OpenAI’s new Astra model is first to hit ‘critical’ cyber risk threshold

5 sources
  • OpenAI said Tuesday that its upcoming Astra model is the first to reach the company's "critical" cybersecurity capability threshold, able to discover and chain zero-day exploits autonomously.
  • The broadly available version will restrict advanced cyber features; select partners including Cisco , Cloudflare , and Palo Alto Networks get early access through a program called Daybreak Blue.
  • OpenAI introduced a "misalignment monitor" that may flag and halt legitimate tasks as false positives, acknowledging the tradeoff between safety and usability.
Sources (5)
  1. 1 OpenAI to limit access to Astra's most powerful cyber capabilities www.axios.com
  2. 2 OpenAI Is About to Release Its First AI Model With ‘Critical' Cyber Abilities www.wired.com
  3. 3 OpenAI says upcoming model is so capable it requires stronger guardrails By Reuters www.investing.com
  4. 4 OpenAI implements stronger safeguards for new Astra Model www.cnbc.com
  5. 5 OpenAI to launch new model with 'stronger safeguards' after hack uk.finance.yahoo.com