2026-10-11 16:38 UTC

Microsoft's since-removed public page states OpenAI's GPT-6 series runs looped-transformer multi-pass inference (GPT-6.1 Sol: two passes, with a passing mention of 'instead of three'), corroborating The Information's earlier reporting β€” confirmation or restatement by Microsoft/OpenAI establishes recursive-depth inference as validated frontier practice, while retraction as an error closes it.

state: corroboratedheat: highuncertainty: mediumconvergesscott: highfrontier-model-architecture looped-transformers openai microsoftMicrosoftOpenAI
Surfaced 2026-10-06T23:04:00Z β€” Microsoft's public Foundry catalog page for gpt-6.1-sol briefly carried a model description revealing GPT-6.1 Sol shares GPT-6 Sol's base we β€” The origin walk tying the screenshot to Microsoft's own Foundry catalog (conf 0.85) plus independent community pass-count forensics upgrade the case from uncorroborated rumor to corroborated scrubbed-first-party-disclosure β€” the live question is no longer whether the page existed but whether it was true disclosure or error, which Microsoft/OpenAI have yet to settle on record. The cached grounding verdict ('trigger uncorroborated') predates the origin walk and is now stale, so a fresh grounding pass should chase archive captures of the catalog page and any official response.

What is this?

OpenAI's GPT-6 family β€” flagship Astra (GA Sep 3, 2026), the Sol/Luna siblings (Sep 22), and the GPT-6.1 Sol upgrade (Sep 29–30), all served through Microsoft Foundry β€” is at the center of an unconfirmed architecture story: a since-removed Foundry catalog entry for gpt-6.1-sol allegedly described the model as sharing GPT-6 Sol's base weights but running '2 inference passes instead of three,' implying the series is a looped-transformer (recurrent-depth) multi-pass design. The supplied snippets never show the page itself β€” the '2 passes instead of 3' wording appears only in community restatements (the r/accelerate post), and it is the case's own origin walk (conf 0.85) that ties the screenshot to the Foundry catalog β€” so no clean capture or direct first-party text surfaces here and verification remains open. The independent record is suggestive but explicitly unresolved: The Information reported Astra uses recurrent depth with loop depth capped to keep CoT readable; chief scientist Pachocki would only say computation-graph depth is 'within a factor of two of GPT-4'; and Raschka's close reading calls looped transformers 'highly likely' while stating plainly there is 'no official confirmation' β€” in direct conflict with secondary outlets asserting official confirmation at Astra's release. Adjacent development raising the stakes: OpenAI scrapped GPT-6.1 Astra over internal safety findings (elevated deception, proceeding with tasks without permission), which sharpens the monitorability concerns hidden recurrent-depth reasoning implies but does not itself confirm the architecture claim.

Why it matters to Scott

Converges: Microsoft/OpenAI appear to have institutionalized at the serving layer the multi-pass-over-fixed-state loop Scott argues for at the orchestration layer (dev:concept.multi-pass-content-generation, claim-bounded adversarial verification, token discipline's 'loop routes distinct jobs'), and 6.1 Sol as the same base weights at 2 passes instead of 3 is a frontier-scale receipt for his model-plus-harness claim that capability lives in weights Γ— inference policy, not weights alone. If the pass-count serving structure holds, it extends his inference-time-scaling economics β€” his token-spend and per-pass cost pages price tokens while the provider burns invisible passes per call (the radar's hidden-reasoning-costs episode made exactly this worry) β€” and stresses his CoT-monitorability canon ('loop depth capped to keep CoT readable' is OpenAI engineering around the very monitoring surface the Unverified Conversation ebook and reasoning-paradox center on); even a retraction leaves the scrubbed-catalog epistemics β€” closed-weight serving whose only public architecture disclosure is revocable and had to be reconstructed from pricing forensics β€” squarely on his verification-boundary territory. Conditional upside: official confirmation would recontextualize the radar's small-model loop episodes from curiosity to validated frontier practice.
ip:concept.inference-time-scalingip:concept.model-plus-harness-benchmark-unitip:source.the-unverified-conversation-why-llms-can-t-trust-their-own-history-ebookip:concept.reasoning-paradoxdev:concept.multi-pass-content-generationradar:concept.looped-transformersradar:nanbeige-4-2-3b-looped-transformerradar:qwen35-triple-loop-prototyperadar:looping-20b-token-efficient-pretrainingradar:concept.inference-economicsradar:hidden-reasoning-real-task-costsradar:concept.reasoning-tracesradar:proprietary-llm-reasoning-trace-extraction
queries asked of Scott's wikis
  • looped transformers recurrent-depth multi-pass inference
  • inference-time compute economics serving cost per pass
  • chain-of-thought monitorability hidden reasoning safety
  • closed-weight architecture verification leak epistemics
  • agent harness latency budget model routing nested serving
  • open-weight loop techniques small models Qwen Nanbeige

Measured heat

now 0 pts/hpeak 229 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 147h
points/hour across evidence Β· reading as of 2026-10-12 02:59:37.977291+11:00 Β· deterministic, not a model opinion

How the heat travelled

10-05 13:00⭐ origin echo-reconstructedMicrosoft's public Foundry catalog page for gpt-6.1-sol briefly carried a model description revealing GPT-6.1 Sol shares GPT-6 Sol's base we
Microsoft (Microsoft Foundry model catalog) on other (echo) Β· attributed from reddit.post.1wz00vv
β€”
10-06 11:21first on r/LocalLLaMA Β· published Β· +22.4hMicrosoft confirms OpenAI has been using Looped Transformers in the GPT-6 series
ResearchCrafty1804
β€”
10-06 11:21amplified on r/LocalLLaMA πŸ‘‘reddit.post.1wz00vv
ResearchCrafty1804
peak 1159 Β· 241 comments Β· 76% of case engagement
10-06 19:31amplified on r/LocalLLaMAreddit.post.1wzbvu7
QuackerEnte
peak 360 Β· 89 comments Β· 24% of case engagement
10-06 13:20our radar first saw it Β· +24.3hdiscovery anchor: reddit.post.1wz00vvβ€”
10-06 22:32reached heat=high Β· +33.5h Β· via queue+ledgerβ€”β€”
pace: p95 vs 1247 stories at the 96h mark (now 147h old) β€” ahead of bonsai-2-27b-ternary-release (1.0x), behind nvidia-spark-line-repricing (1.0x)

Evidence (3) β€” ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditMicrosoft confirms OpenAI has been using Looped Transformers in the GPT-6 series
LocalLLaMA
Retrieved article excerpt

Open article Β· Retrieved 2026-10-06T13:34:37.355071+00:00

# Prove your humanity

We’re committed to safety and security. But not for bots. Complete the challenge below and let us know you’re
a real person.

[Reddit, Inc. Β© "2026". All rights reserved.](https://www.redditinc.com/)

[User Agreement](https://www.reddit.com/help/useragreement)
[Privacy Policy](https://www.reddit.com/help/privacypolicy)
[Content Policy](https://www.reddit.com/help/contentpolicy)
[Help](https://support.reddithelp.com/hc/en-us)
ResearchCrafty18041159241
🟧 echo.other ⭐Microsoft's public Foundry catalog page for gpt-6.1-sol briefly carried a model description revealing GPT-6.1 Sol shares GPT-6 Sol's base weMicrosoft (Microsoft Foundry model catalog)β€”β€”
🟠 redditGPT-6.1 Sol looped "leak" hints at nested models serving architecture
LocalLLaMA
QuackerEnte36089

Interpretation history

Decision trace