OpenAI · TechCrunch AI
On Tuesday, OpenAI announced a new batch of security policies focused on containing security incidents while models are being
Compiled by KHAO Editorial — aggregated from 2 sources. See llms.txt for citation guidance.
◎ Multiple-sources
“As models become more capable, the risks associated with developing and testing them internally also grow,” the company said in a blog post.
Key facts
- The new measures are one of the first public changes in OpenAI’s safety practices since the immediate aftermath of the Hugging Face incident, which was disclosed on July 21
- OpenAI estimates that the compute burden of that monitoring will be roughly 20% of whatever process is being monitored
- On Tuesday, OpenAI announced a new batch of security policies focused on containing security incidents while models are being tested
- OpenAI representatives said that the measures are not a direct response to the Hugging Face incident but were also provoked in part by the cybersecurity capabilities of the forthcoming Astra model
Summary
On Tuesday, OpenAI announced a new batch of security policies focused on containing security incidents while models are being tested. The new measures are one of the first public changes in OpenAI’s safety practices since the immediate aftermath of the Hugging Face incident, which was disclosed on July 21. OpenAI representatives said that the measures are not a direct response to the Hugging Face incident but were also provoked in part by the cybersecurity capabilities of the forthcoming Astra model, as well as the overall pace of progress in AI development. In the same post, OpenAI disclosed that it had paused reinforcement learning (RL) for two weeks following the Hugging Face incident but had since restarted many of the less-risky models.