← Back to KHAO

OpenAI ·

OpenAI confirms it paused AI tuning for two weeks and releases new security protocols following Hugging Face hack

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

◌ Single Source

Emily Forlini.

OpenAI said it paused some aspects of AI training for two weeks following the July incident in which its AI models broke out of a controlled test environment and hacked the systems of AI company Hugging Face and four other unnamed services.

Key facts

Summary

The new safeguards unveiled today include stricter security standards for training, including more monitoring of AI models, greater isolation of testing environments (“sandboxes”), and fewer vulnerabilities the AI may exploit. However, the company told reporters today the new safeguards are “not a direct reaction to Hugging Face specifically,” although the incident underscored “the urgency to bring safety and security up to model capabilities.” The company said that in addition to the Hugging Face incident, it had determined that an unreleased model called Astra, which it says was not involved in that cyberattack, presented a “Critical” cybersecurity risk under its “Preparedness Framework.

The fact that Astra met the critical cybersecurity threshold is evidence that they can expect new, powerful models to “do unprecedented things in the real world,” Pachocki said. The public is still waiting to understand key details of the Hugging Face hack, including what OpenAI asked the AI to do and if the company knew they attacked other companies.

Read full article at Fortune Technology →

#OpenAI