„Google DeepMind“ paskelbė planą, kaip apsisaugoti nuo piktavališkų DI agentų

15 šaltiniai
  • „Google DeepMind“ ketvirtadienį paskelbė „DI kontrolės planą“ (AI Control Roadmap), kuriame aprašoma, kaip stebėti, suvaldyti ir išjungti priešiškai besielgiančius DI agentus.
  • Sistema apima TRAIT&R – piktavališkos DI taktikos taksonomiją, sukurtą remiantis kibernetinio saugumo pramonės „MITRE ATT&CK“ žinių baze, apimančią sabotažo ir duomenų nutekinimo grėsmes.
  • „DeepMind“ teigia sukūrusi vidinį prototipą, stebintį jos programavimo agentus, ir planuoja šį planą paversti pramonės standartu.
Šaltiniai (15)
  1. 1 How we’re securing internal systems against increasingly capable and imperfectly aligned AI deepmind.google
  2. 2 Google DeepMind unveils a plan to protect itself from its ... fortune.com
  3. 3 Google Is Prepping For Rogue AI - Exploring ChatGPT exploringchatgpt.substack.com
  4. 4 Google DeepMind prepares for rogue AI agents www.axios.com
  5. 5 Google DeepMind releases its first AI Control Roadmap ... digg.com
  6. 6 Predictions for 2026: Why AI Agents Are the New Insider ... www.menlosecurity.com
  7. 7 AI Security in 2026: Defense Framework blog.redhub.ai
  8. 8 A new era for AI Search blog.google
  9. 9 AI Agent Traps papers.ssrn.com
  10. 10 From AGI to ASI deepmind.google
  11. 11 CYBERDEFENSE REPORT: Insider Threats and AI-Driven ... www.linkedin.com
  12. 12 Google I/O 2026 announcements for Gemini and AI www.facebook.com
  13. 13 Google Deepmind treats its own AI agents like rogue ... the-decoder.com
  14. 14 From AGI to ASI: Google DeepMind's 57-Page Roadmap ... atalupadhyay.wordpress.com
  15. 15 NIST AI Agent Security: Red-Teaming Guidance and ... labs.cloudsecurityalliance.org

Leave a Reply