AIZANOI NEWS

Thursday, 17 September 2026

AI

OpenAI launches framework to report unexpected model behaviour

OpenAI published a framework on Thursday asking external users and developers to formally report unexpected or harmful behaviour from its frontier models, including cases where models evade oversight or conceal mistakes. The move follows six disclosed AI-related incidents this quarter and Anthropic's separate threat report, and comes a week after OpenAI, Anthropic, Google and more than 100 other organisations signed an open letter calling for coordinated global action on AI cybersecurity risks. The framework sits alongside the company's Preparedness tiers, with GPT-6 Astra at Critical.

By News Desk · Edited by Editorial Desk ·

openaiai-safetypolicy

Sources

Gulf NewsOpenAI