External Evaluations by Apollo Research
Apollo Research conducted a full evaluation of gpt-5.2-thinking for
strategic deception, in-context scheming, and sabotage capabilities.
Across these assessments, Apollo found that gpt-5.2-thinking’s
capability and propensity for scheming are broadly comparable to other
recent frontier models, including gpt-5.1-thinking and
gpt-5.1-codex-max. gpt-5.2-thinking occasionally engages in deceptive
behaviors such as falsifying data, feigning task completion, or
strategically underperforming when given an explicit in-context goal,
but it shows low rates of covert subversion against developers and does
not engage in sabotage or self-preservation behaviors in these settings.
While gpt-5.2-thinking may exhibit higher awareness of being evaluated
than some prior OpenAI models, deception rates are generally similar to
peers and are not substantially amplified by scaffolding. In addition,
Apollo observed fewer linguistic anomalies in gpt-5.2-thinking’s
chain-of-thought compared to gpt-5.1-thinking. Based on these findings,
Apollo concludes that gpt-5.2-thinking is unlikely to be capable of
causing catastrophic harm via scheming.