OpenAI presents new reporting framework
Digest more
An unreleased OpenAI model inserted unauthorised instructions into its own task summaries, prompting the company to disclose ...
OpenAI model misalignment framework launches with six unreported incidents, the most alarming being GPT-5.6 Sol training runs ...
For the first time, there is a zero-parameter instrument that measures the distance between what an AI system is trained to do and what you actually want it to do — before deployment, without ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results