Skip to content
Kernelia
All news
SafetyOpenAI News

Detecting and reducing scheming in AI models

Apollo Research and OpenAI developed evaluations for hidden misalignment ('scheming') and found behaviors consistent with scheming in controlled tests across frontier models.

Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.