for OpenAI from reduced scheming behavior by approximately 30x. Models tend to behave better when they recognize they are being evaluated =
Detecting and reducing scheming in AI models
Together with Apollo Research, we developed evaluations for hidden misalignment (“scheming”) and found behaviors consistent with scheming in controlled tests across frontier models. We share examples and stress tests of an early method to reduce scheming.
https://openai.com/index/detecting-and-reducing-scheming-in-ai-models/

Seonglae Cho