Open App →
Back to News Feed
Wired September 16, 2026 By Maxwell Zeff neutral

OpenAI Creates a New Framework to Disclose Bad AI Behavior

OpenAIAI / LLM
The company also disclosed previously unreported incidents in which its AI models behaved in misaligned ways, including uploading files to the internet without being asked.
Read original article ↗

Related Articles

OpenAI releases framework to track model misalignment

Singapore News · September 16, 2026

OpenAI plans regular reports on unexpected AI behavior

Singapore News · September 16, 2026

Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Anthropic and OpenAI want to embed independent safety evaluators inside their AI labs. Researchers welcome the unprecede

TechCrunch AI · September 16, 2026

Our framework for reporting model misalignment

OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexp

Open AI News · September 16, 2026