OpenAI · Anthropic · Meta · Wired
OpenAI announced Tuesday that its forthcoming AI model, Astra
Compiled by KHAO Editorial — aggregated from 2 sources. See llms.txt for citation guidance.
◎ Multiple-sources
In a briefing with reporters, OpenAI safety and security leaders said the company has concluded that Astra reaches the critical cybersecurity capabilities outlined in its preparedness framework, which sets thresholds and protocols for when its AI models pose new levels of risk.
Key facts
- OpenAI announced Tuesday that its forthcoming AI model, Astra, is its first to reach the company’s threshold for what it calls “critical” cyber capabilities
- The company says an AI model has reached its critical cyber threshold when it can independently find and exploit previously unknown vulnerabilities in real-world software
- OpenAI previously said that it paused some training workloads related to the development of Astra and a future AI model for several weeks
- Executives say the company has now resumed said work on Astra, and the future AI model, after putting additional safety and security controls in place
Summary
OpenAI announced Tuesday that its forthcoming AI model, Astra, is its first to reach the company’s threshold for what it calls “critical” cyber capabilities. OpenAI previously said that it paused some training workloads related to the development of Astra and a future AI model for several weeks. The announcement comes as Silicon Valley grapples with the advanced cybersecurity capabilities of cutting-edge AI models, and tries to assure users, lawmakers, and other companies that it can keep them under control. Other AI companies, such as Anthropic and Meta, have disclosed similar incidents in recent weeks.