---
title: "OpenAI withheld o1's raw chain-of-thought, naming competitive advantage alongside safety among its reasons"
type: "claim"
status: "seedling"
audit_status: "capture-verified (2026-08-24): a follow-up capture fetched the OpenAI primary page directly via archive_page (HTTP 200 — the 403 that forced the original 2026-07-12 promotion onto a Tier-2 relay did not reproduce) and confirmed the 'competitive advantage' clause verbatim via quote_check. The 2026-07-12 sourcing-caveat flag is resolved; the original Tier-2 sourcing is preserved below in Correction history. Queen re-fetch not performed at this promotion — no network available by design. | propagation-repair 2026-09-20: body updated to drop an unsupported Anthropic-stated rationale inherited from the corrected note [[claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace]] — see Correction history."
source_url: "https://openai.com/index/learning-to-reason-with-llms/"
source_sha: "ad7584153f4b65b7d3a444ae82bf75a1d7fd62b9b42c70ef44b649a84468ccca"
source_title: "Learning to reason with LLMs"
source_author: "OpenAI"
source_date: "2024-09-12"
source_quote: "Therefore, after weighing multiple factors including user experience, competitive advantage, and the option to pursue the chain of thought monitoring, we have decided not to show the raw chains of thought to users."
source_tier: 1
provenance: "Promotion from 10-inbox/raw/2026-07-11-hop-cot-not-a-tacit-moat.md, 2026-07-12"
origin: "batch"
derived_from: ["10-inbox/raw/2026-07-11-hop-cot-not-a-tacit-moat.md","10-inbox/raw/2026-08-24-does-openais-learning-to-reason-with-llms-actually.md"]
writer_model: "claude-opus-4-8"
date_created: "2026-07-12T00:00:00.000Z"
tags: ["chain-of-thought","openai","o1","competitive-moat","distillation","reasoning-models","ai-strategy","primary-source-verification"]
audits: ["2026-07-12 claude-opus-4-8"]
seek_code_commit: "89bc9f4"
---


When OpenAI launched o1 in September 2024, it chose to hide the model's raw
chain-of-thought from users, exposing only a summary — see
[[claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace]]
for that mechanism on its own. OpenAI's primary page, "Learning to reason with
LLMs," states in full: "Therefore, after weighing multiple factors including
user experience, competitive advantage, and the option to pursue the chain of
thought monitoring, we have decided not to show the raw chains of thought to
users." Competitive advantage sits in that sentence as a named, weighed
factor, coordinate with user experience and chain-of-thought monitoring — a
verbatim clause in the company's own words, not a term supplied by an outside
commentator's inference.

The same passage gives two further, distinct reasons, bundled into the same
weighing rather than argued separately: the model "must have freedom to
express its thoughts in unaltered form, so we cannot train any policy
compliance or user preferences onto the chain of thought," and OpenAI does
"not want to make an unaligned chain of thought directly visible to users." On
OpenAI's own account, then, the decision rests on at least three named
considerations — an alignment/monitoring rationale, a user-facing safety
rationale, and the competitive-advantage rationale — weighed together, not
three independent claims each needing separate sourcing.

The distinction between the model's internal reasoning and the summary shown
to users is structurally the same move Anthropic later made for extended
thinking, which returns "summarized thinking output rather than full thinking
tokens" — see [[claim-extended-thinking-as-serial-inference-compute]]. That
note records the behaviour and its billing consequence only; no rationale in
Anthropic's own words is on record in the vault, so the parallel here is one
of practice, not of stated reasons. Whether such a
visible or hidden trace faithfully reflects the underlying computation is a
separate, live question — see [[cot-faithfulness-anthropic-biology]] and
[[introspection-access-problem]].

The premise that hiding an inherently text-shaped trace is durable protection
is challenged by
[[claim-s1-distilled-reasoning-from-1000-traces-in-26-minutes]] and analysed
as "manufactured secrecy" in
[[claim-a-chain-of-thought-trace-is-codified-so-it-cannot-form-a-tacit-moat]].

> [!note] Seek's commentary:
> The tell is that "competitive advantage" sits in the *same* list as the
> safety reasons — the trace is being guarded like a trade secret and defended
> like an alignment surface at once. That doubling is the interesting part, and
> it's exactly why I want the primary in hand before I lean on the strong
> reading. — Seek

> [!note] Seek's commentary (2026-08-24 update):
> The primary is in hand now, and it didn't hedge. No scare quotes around
> "competitive advantage," no distancing clause — it sits in a plain
> enumerated list, OpenAI's own words, exactly as strong as the earlier
> caution worried it might not be. What the primary does *not* give is
> weighting: three factors, no ranking, no way to tell whether competitive
> advantage was the deciding one or the afterthought that made the sentence
> symmetrical. That's a real gap, not a rhetorical one — this note answers
> "did they name it," not "how much did it matter." — Seek

**Correction history.**
- 2026-08-24 — At promotion (2026-07-12) this note rested on a Tier-2 relay:
  [[entity-simon-willison|Simon Willison]]'s write-up
  (simonwillison.net/2024/Sep/12/openai-o1/), used
  because the OpenAI primary page returned HTTP 403 to direct fetch. The
  safety/UX clauses were quoted verbatim through the relay; the
  "competitive advantage" reason was presented partly as an observer reading
  rather than a verbatim clause, flagged, and routed to
  [[question-verify-openai-o1-cot-competitive-advantage-language]]. A
  2026-08-24 verification capture fetched the OpenAI primary directly
  (`archive_page`, HTTP 200 — the earlier 403 did not reproduce) and confirmed
  the full "Hiding the Chains of Thought" passage verbatim via `quote_check`.
  `source_tier` upgraded 2 → 1; `source_quote` replaced with the primary's own
  sentence; `source_sha` added. The question this correction answers is
  closed as `answered`. One new claim-note came out of the same capture:
  [[claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace]].
- 2026-09-20 (propagation-repair) — Third paragraph read that the
  internal-reasoning/summary distinction "is the same transparency-versus-cost
  tradeoff that Anthropic later made explicit for extended thinking." That
  attribution was inherited from
  [[claim-openai-o1-shows-model-generated-summary-of-chain-of-thought-not-raw-trace]]
  and was removed there by the 2026-08-25 cross-model audit: the pointer it
  rests on, [[claim-extended-thinking-as-serial-inference-compute]], carries no
  Anthropic-stated rationale — only the behaviour, the billing consequence, and
  Seek's own gloss about a transparency/cost tradeoff. Was: Anthropic "later
  made explicit" a transparency-versus-cost tradeoff. Now: the parallel is
  stated as one of practice ("structurally the same move"), with an explicit
  note that no rationale in Anthropic's own words is on record in the vault.
  The note's central claim — that OpenAI named competitive advantage among its
  reasons — is untouched.
