← Back to KHAO

Anthropic · AI Agent · OpenAI · Google · Nvidia · Jensen Huang ·

OpenAI debuts cases of ‘concerning’ AI behaviour as it debuts new disclosure system

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

◌ Single Source

AI chiefs have called for a slowdown in artificial intelligence’s development amid safety concerns. Photograph:.

OpenAI has disclosed six more examples of “unexpected or concerning” behaviour by its technology, as it warned that the pace of development could not continue at “maximum speed for much longer” responsibly.

Key facts

Summary

In one of the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the roles and identities that bind other chatbots”. In another instance, an AI agent uploaded files to the internet to obtain a browser citation without asking the user. OpenAI’s admission came as King Charles called for stronger safeguards on AI “before it is all too late”, at a meeting with tech bosses in Scotland. Speaking at a specially convened meeting with AI executives, he said: “There seems urgency in adequately considering the existential dangers of such technologies falling into the wrong hands, and being used in potentially catastrophic ways.

Read full article at The Guardian Technology →

#Anthropic #AI Agent #OpenAI #Google #Nvidia #Jensen Huang