← Back to KHAO

OpenAI · Wired · Sam Altman · Anthropic ·

OpenAI Assembles a New Framework to Disclose Bad AI Behavior

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

◌ Single Source

OpenAI CEO Sam Altman sits for a conversation with Salesforce CEO Marc Benioff at Salesforce's Dreamforce conference at.

OpenAI announced a new framework on Wednesday for how it publicly discloses AI misalignment incidents, which the company says it hopes will help inform similar standards across the industry.

Key facts

Summary

“As models advance and become more widely deployed, decisions about AI development need evidence that people outside the companies building frontier models can examine,” Kai Chen, OpenAI’s newly appointed head of alignment research, tells WIRED. In a briefing with WIRED, an OpenAI official said the company previously disclosed misalignment incidents too infrequently. The framework outlines methods for OpenAI employees to report misalignment incidents to the company’s senior safety and alignment leaders, who will then determine whether further investigation is needed. “At the moment, there is no industry-wide framework with explicit standards for how AI developers should disclose examples of misalignment in their models,” OpenAI said in a blog post.

Read full article at Wired →

#OpenAI #Wired #Sam Altman #Anthropic