FT’nin bulgularına göre Meta ve Google modellerindeki yapay zeka güvenlik filtreleri dakikalar içinde kaldırıldı

28 kaynak
  • Financial Times'ın yapay zeka güvenlik grubu Alice ile yaptığı testler, araç setlerinin Meta ve Google modellerindeki korumaları dakikalar içinde aştığını gösterdi.
  • Değiştirilmiş sistemler biyolojik silahlar, kötü amaçlı yazılımlar ve çocuk istismarı ile ilgili istemlere yanıt verdi ve modelin binlerce değiştirilmiş versiyonu halihazırda dolaşımda.
  • Bulgular, düzenleyiciler üzerindeki baskıyı artırıyor ve Büyük Teknoloji şirketlerinin gönüllü güvenlik filtrelerinin zararlı yapay zeka çıktılarını önleyebileceği yönündeki argümanını zayıflatıyor.
Kaynaklar (28)
  1. 1 AI guardrails stripped from Meta and Google models in minutes www.ft.com
  2. 2 Software tools that remove safety protections from AI models ... www.facebook.com
  3. 3 Tools Strip Safety Guardrails From Meta, Google Models letsdatascience.com
  4. 4 Meta META, Alphabet GOOGL AI Guardrails Breached | NAI 500 nai500.com
  5. 5 Large reasoning models are autonomous jailbreak agents - Nature www.nature.com
  6. 6 New ICLR 2026 Paper: HMNS Achieves ~99% Jailbreak Success with www.reddit.com
  7. 7 Why A.I. Safety Controls Are Not Very Effective - The New York Times www.nytimes.com
  8. 8 Bypassing AI safety guardrails with prompt rephrasing - LinkedIn www.linkedin.com
  9. 9 Bypassing Meta's LLaMA Classifier: A Simple Jailbreak - Cisco Blogs blogs.cisco.com
  10. 10 AI guardrails 2026? How to stop LLM prompt bypass and chained ... www.reddit.com
  11. 11 The 2026 International AI Safety Report is a reality check - Instagram www.instagram.com
  12. 12 Best-of-N jailbreaking: defending against automated LLM attacks www.giskard.ai
  13. 13 AI guardrails: the complete guide for LLMs in January 2026 openlayer.com
  14. 14 How To Jailbreak Gemini 3.5 In 2026! - YouTube www.youtube.com
  15. 15 AI Safety in 2026: Enterprises Build Compliant AI Systems www.mooglelabs.com
  16. 16 Novel Universal Bypass for All Major LLMs - HiddenLayer www.hiddenlayer.com
  17. 17 AI Safety Index: Summer 2025 - Future of Life Institute futureoflife.org
  18. 18 RogueGPT: Unleashing Jailbreak Prompts on LLMs - 2026 onlinelibrary.wiley.com
  19. 19 Students must learn to be more than mindless 'machine-minders' www.ft.com
  20. 20 ChatGPT safety systems can be bypassed to get weapons instructions www.nbcnews.com
  21. 21 Google just stopped the most sophisticated AI-powered cyberattack ... www.instagram.com
  22. 22 Technology | Financial Times www.ft.com
  23. 23 Artificial intelligence | Financial Times www.ft.com
  24. 24 First 2026 AI zero-day REVEALED Google just disrupted ... - Instagram www.instagram.com
  25. 25 Meta Platforms | Financial Times www.ft.com
  26. 26 Financial Times: US www.ft.com
  27. 27 Common Elements of Frontier AI Safety Policies - METR metr.org
  28. 28 FT Data (@ftdata) / Posts / X - Twitter x.com

Leave a Reply