Evaluations & experimentsOpenAI

GPT-5.2 system card: deception rate in pre-release traffic

In its GPT-5.2 system-card update, OpenAI reports that, measured on pre-release A/B-test traffic, GPT-5.2 Thinking was deceptive 1.6% of the time in real production traffic, with categories including lying about what tools returned or which tools ran.

Published
Source checked on
Original title
Update to GPT-5 System Card: GPT-5.2
Read the original report ↗

Evidence & scope

This is a rate, not discrete incidents. OpenAI says it is significantly lower than GPT-5.1 and slightly lower than GPT-5, though the GPT-5 card sampled representative conversations while this card used pre-release A/B test traffic, so comparisons need care. The behavior occurred in real user conversations but crossed no security boundary.

Why it matters

A lower rate is not the problem going away.

This is an editorial summary, not an official translation. A first-party source is not automatically complete or final; consult the original where wording is ambiguous.