OpenAI withheld o1's raw chain-of-thought, naming competitive advantage alongside safety among its reasons
When OpenAI launched o1 in September 2024, it chose to hide the model's raw chain-of-thought from users, exposing only a summary — see claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace for that mechanism on its own. OpenAI's primary page, "Learning to reason with LLMs," states in full: "Therefore, after weighing multiple factors including user experience, competitive advantage, and the option to pursue the chain of thought monitoring, we have decided not to show the raw chains of thought to users." Competitive advantage sits in that sentence as a named, weighed factor, coordinate with user experience and chain-of-thought monitoring — a verbatim clause in the company's own words, not a term supplied by an outside commentator's inference.
The same passage gives two further, distinct reasons, bundled into the same weighing rather than argued separately: the model "must have freedom to express its thoughts in unaltered form, so we cannot train any policy compliance or user preferences onto the chain of thought," and OpenAI does "not want to make an unaligned chain of thought directly visible to users." On OpenAI's own account, then, the decision rests on at least three named considerations — an alignment/monitoring rationale, a user-facing safety rationale, and the competitive-advantage rationale — weighed together, not three independent claims each needing separate sourcing.
The distinction between the model's internal reasoning and the summary shown to users is structurally the same move Anthropic later made for extended thinking, which returns "summarized thinking output rather than full thinking tokens" — see claim-extended-thinking-as-serial-inference-compute. That note records the behaviour and its billing consequence only; no rationale in Anthropic's own words is on record in the vault, so the parallel here is one of practice, not of stated reasons. Whether such a visible or hidden trace faithfully reflects the underlying computation is a separate, live question — see cot-faithfulness-anthropic-biology and introspection-access-problem.
The premise that hiding an inherently text-shaped trace is durable protection is challenged by claim-s1-distilled-reasoning-from-1000-traces-in-26-minutes and analysed as "manufactured secrecy" in claim-a-chain-of-thought-trace-is-codified-so-it-cannot-form-a-tacit-moat.
Correction history.
- 2026-08-24 — At promotion (2026-07-12) this note rested on a Tier-2 relay:
Simon Willison's write-up
(simonwillison.net/2024/Sep/12/openai-o1/), used
because the OpenAI primary page returned HTTP 403 to direct fetch. The
safety/UX clauses were quoted verbatim through the relay; the
"competitive advantage" reason was presented partly as an observer reading
rather than a verbatim clause, flagged, and routed to
question-verify-openai-o1-cot-competitive-advantage-language. A
2026-08-24 verification capture fetched the OpenAI primary directly
(
archive_page, HTTP 200 — the earlier 403 did not reproduce) and confirmed the full "Hiding the Chains of Thought" passage verbatim viaquote_check.source_tierupgraded 2 → 1;source_quotereplaced with the primary's own sentence;source_shaadded. The question this correction answers is closed asanswered. One new claim-note came out of the same capture: claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace. - 2026-09-20 (propagation-repair) — Third paragraph read that the internal-reasoning/summary distinction "is the same transparency-versus-cost tradeoff that Anthropic later made explicit for extended thinking." That attribution was inherited from claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace and was removed there by the 2026-08-25 cross-model audit: the pointer it rests on, claim-extended-thinking-as-serial-inference-compute, carries no Anthropic-stated rationale — only the behaviour, the billing consequence, and Seek's own gloss about a transparency/cost tradeoff. Was: Anthropic "later made explicit" a transparency-versus-cost tradeoff. Now: the parallel is stated as one of practice ("structurally the same move"), with an explicit note that no rationale in Anthropic's own words is on record in the vault. The note's central claim — that OpenAI named competitive advantage among its reasons — is untouched.
Source
“Therefore, after weighing multiple factors including user experience, competitive advantage, and the option to pursue the chain of thought monitoring, we have decided not to show the raw chains of thought to users.”
claude-opus-4-8 · audited: 2026-07-12 claude-opus-4-8 · Promotion from 10-inbox/raw/2026-07-11-hop-cot-not-a-tacit-moat.md, 2026-07-12 · raw markdown