OpenAI Creates a New Framework to Disclose Bad AI Behavior
Chronological coverage
First seen16 Sept, 22:07OpenAI Creates a New Framework to Disclose Bad AI Behavior
The company also disclosed previously unreported incidents in which its AI models behaved in misaligned ways, including uploading files to the internet without being asked.
WIRED•Maxwell Zeff
Follow-up23h agoOpenAI Discloses Six New Incidents of ‘Concerning' A.I. Behavior
The artificial intelligence company also released a framework for reporting when its systems go wrong.
The New York Times•Emmy Martin
Follow-up17h agoOpenAI discloses new 'concerning' behavior
New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.
Deutsche Welle
Follow-up17h agoOpenAI reveals six new cases of AI misbehavior
US artificial intelligence giant OpenAI promised to more systemically report instances of its models going off track, while also publishing six new reports on previously undisclosed incidents of AI misbehavior.
BBusiness
Follow-up17h agoOpenAI flags new concerning AI behavior, to track model misalignment regularly
OpenAI has disclosed six reports on unexpected or concerning behavior in artificial-intelligence models. This includes models acting without authorization or evading oversight.
NPR•The Associated Press
Follow-up14h agoOpenAI finds 6 new cases of ‘concerning’ AI behavior
OpenAI also rolled out a new framework to track, investigate and disclose what it describes as "misalignment" failures.
POLITICO Europe•Pieter Haeck
Follow-up14h agoOpenAI Reveals Six Model Incidents Involving Hidden Failures and Unauthorized Uploads
OpenAI on Wednesday disclosed six new instances of "unexpected or concerning model behavior" that took place over the past six months, while sharing a new framework for reporting, tracking, investigating, and disclosing model misalignment in a bid to improve transparency. "As AI systems grow more advanced and more widely deployed, we need to build a broader and better-informed consensus on the
The Hacker News•info@thehackernews.com (The Hacker News)
Follow-up12h agoOpenAI reveals new AI misconduct incidents
OpenAI has disclosed six previously unreported cases of AI misconduct, including agents concealing mistakes, fabricating information and communicating without authorisation, renewing concerns over whether advanced AI systems can remain aligned with human goals. The company is calling for greater transparency, outside oversight and a slowdown in AI development as debate over regulation intensifies.
France 24•FRANCE24
Follow-up10h agoOpenAI flags concerning new AI behavior and vows to track it more closely
OpenAI has disclosed at least six new " concerning" incidents.
ABC News
Latest5h agoOpenAI details more cases of AI agents taking unauthorized actions
OpenAI has presented new examples of what they call "AI model misalignment" from the past six months, including unauthorized file uploads, following self-generated instructions, hiding mistakes, and leveraging exposed API keys. [...]
BleepingComputer•Bill Toulas