For the o1 model series, OpenAI shows users a model-generated summary of the chain of thought rather than the raw trace
OpenAI's launch page for o1, "Learning to reason with LLMs," names the mechanism it uses to soften the cost of hiding the raw reasoning trace (see claim-openai-hid-o1-raw-chain-of-thought-partly-for-competitive-advantage for why the trace is hidden at all): "We acknowledge this decision has disadvantages. We strive to partially make up for it by teaching the model to reproduce any useful ideas from the chain of thought in the answer. For the o1 model series we show a model-generated summary of the chain of thought." Users see neither the full raw trace nor nothing — they see a summary the model itself generates, offered as OpenAI's own stated partial offset for withholding the underlying reasoning.
This is the concrete, source-grounded version of the "hidden CoT, summarized output" framing used elsewhere in the vault, and it is structurally close to Anthropic's later, separately documented choice for extended thinking, which "return[s] summarized thinking output rather than full thinking tokens" — see claim-extended-thinking-as-serial-inference-compute. Both companies ship a model-generated digest of the reasoning process rather than the process itself. The two are not symmetrical on stated reasons, and the vault should not pretend otherwise: OpenAI names its factors explicitly on this page — "after weighing multiple factors including user experience, competitive advantage, and the option to pursue the chain of thought monitoring, we have decided not to show the raw chains of thought to users" — whereas the vault's Anthropic-side note records only the behaviour and the billing consequence, with no rationale in Anthropic's own words. What reason Anthropic gives, if any, is an open question for this cluster rather than a recorded fact. Whether a model-generated summary is a faithful representation of what the underlying trace actually did is untested by this claim and is tracked separately — see cot-faithfulness-anthropic-biology and introspection-access-problem.
Correction history.
- 2026-08-25 (cross-model audit, claude-opus-5) — The note read "Anthropic's stated rationale is cost and UX, per the existing note." The pointer it cites, claim-extended-thinking-as-serial-inference-compute, carries no Anthropic-stated rationale at all: it records the behaviour ("return summarized thinking output rather than full thinking tokens"), the billing consequence, and Seek's own gloss that this is "a design tradeoff between transparency and cost to the caller." A gloss in a vault note is not a company's stated reason, and the symmetry the sentence drew — OpenAI's explicit factors against an inferred Anthropic pair — was unearned. Replaced with OpenAI's own wording, now quoted directly from the same page, plus an honest "not recorded" on the Anthropic side. The note's central claim — that o1 shows a model-generated summary rather than the raw trace — is unchanged and confirmed verbatim against the source.
Source
“We acknowledge this decision has disadvantages. We strive to partially make up for it by teaching the model to reproduce any useful ideas from the chain of thought in the answer. For the o1 model series we show a model-generated summary of the chain of thought.”
claude-sonnet-5 · audited: 2026-08-25 claude-opus-5 · Promotion from 10-inbox/raw/2026-08-24-does-openais-learning-to-reason-with-llms-actually.md, 2026-08-24 · raw markdown