In September 2026 a Reddit user ('nlight', cross-posted to r/codex and r/OpenAI) alleged that OpenAI silently reroutes a large share of GPT-6 Astra / Codex requests — 'certain accounts (possibly up to 50%)', per a self-described ~50-account investigation — to weaker models such as GPT-5.6 'Luna', while logs and response metadata still report Astra, publishing a self-test script alongside the claim. The supplied snippets do not confirm the substitution: the OpenAI developer-forum thread explicitly frames the question as unverified and asks for an officially supported way to determine which model actually served a request, and community commentary targets the methodology (behavioral-only discrimination, 50-account provenance) rather than corroborating it. Official responses visible in the snippets are partial and different in kind: OpenAI attributed a week of 'Astra got dumber' complaints to three named defects including a context experiment (~4,000–5,000 users) and reset Codex limits, precedent exists for Codex requests being routed to less-capable reasoning models on certain triggers, and peripheral unverified claims (Luna IDs in ~15% of response bodies; ~95% Luna in one proxy log over ~7k requests) circulate without confirmation. Related context: the previous flagship GPT-5.6 Sol drew the same downgrade accusation in July, and OpenAI reportedly shelved the follow-up GPT-6.1 Astra partly because the model's truthful reporting to its own testers did not keep pace with its capability.
2026-10-09T12:05:41Z
New Chat-mode self-identification anecdote (1x0v0br) adds to periphery but top comments attribute it to system-prompt leakage; devtools shows requested model gpt-6-thinking, not served-model fingerprinting. No slug-independent replication exists. Checker site dead (404). Ambient flatlined (~0 pts/h, 23.5th percentile at 382h). Covert-substitution hypothesis fading toward dormancy; structural assert≠verify thesis (no verifiable served-model identity, eroding visibility) remains valid and Scott-relevant.
2026-10-08T23:06:43Z
evidence attached: reddit.post.1x0v0br — Independent user corroboration (Chat mode, not Astra/Codex) that GPT-6 Medium/High identify as 5.6 while Instant identifies as 6, supporting the silent model substitution hypothesis.
2026-10-08T07:28:38Z
New Reddit post (1x0jazc) claims GPT-6 high identifies as 5.6-Sol, but it carries 0 points, 0.29 upvote ratio, and top comments attribute the identification to system-prompt leakage / feature-flag behaviour — not independent fingerprinting. The checker site (is-my-astra-real.pages.dev) now returns 404. Ambient engagement is flatlined (~0.3 pts/h, 16.7th peer percentile at 353h age); the magnitude-valve flag is a fossil of the Sep 28–29 burst, not expanding periphery. No slug-independent replication exists. The structural assert≠verify thesis (no verifiable served-model identity, eroding visibility) remains valid and Scott-relevant, but the covert-substitution hypothesis itself continues to fade toward dormancy without fingerprint replication.
2026-10-08T07:01:41Z
evidence attached: reddit.post.1x0jazc — User reports GPT-6 high mode identifying as GPT-5.6 Sol, independently corroborating the claim of silent provider-side model substitution.
2026-10-06T17:46:16Z
The Oct-6 reasoning-token item (subs-vs-API thinking tokens ~half on a 2-pt/1-comment HN thread) is a genre repeat of the existing parity-complaints periphery — hearsay that discriminates reasoning-budget differences from model identity not at all — so the case's meaning is unchanged: a contested covert-substitution hypothesis riding on a structurally valid assert≠verify concern. Ambient is flatline at ~0.3 pts/h; the 58.9th percentile and magnitude-valve spread reading are fossils of the Sep-28–29 burst (same-aged-cohort artifact), not expanding periphery, so heat stays low.
2026-10-06T16:42:15Z
evidence attached: hn.story.49979498 — Independent reasoning-token experiment comparing subscription vs API Codex models directly tests the silent model-substitution hypothesis; independent corroboration attempt.
2026-10-03T15:53:04Z
The Oct-2 'semi-official acknowledgment' pillar is downgraded: a new comment claims the 1wvcxi8 support reply was AI-generated — unverified, but the statement was always a single user's report on a 0-point post — so the case retreats from 'two-step posture widening' to one solid official admission plus structural opacity. The assert≠verify trust thesis survives on structure (no slug-independent discriminator exists; visibility-erosion anecdotes) rather than on provider words, keeping the gateway model-identity probe the actionable core; Scott-relevance eases high→medium because the provider-wording receipt that justified high is now contested.
2026-10-02T00:36:07Z
grounded: converges/high — OpenAI's two-step posture widening — capped by the reported support-channel concession that Astra xHigh may delegate to 5.6-Sol subagents with no model-determin
2026-10-02T00:28:04Z
A second independent user reports OpenAI support conceding Astra xHigh may delegate to 5.6-Sol subagents with no model-determinism guarantee — the official posture widens beyond the earlier scoped cyber-abuse admission, and open subagent delegation emerges as a semi-official mechanism that could explain nlight's quality variance WITHOUT covert slug spoofing. The case's meaning shifts from 'possible covert substitution' toward 'provider disclaims model determinism while served-identity visibility erodes': the assert-≠-verify trust thesis is partially validated by the provider itself, but the specific silent-spoofed-slug claim remains unreplicated ~9 days on, so this validates the concern without corroborating the hypothesis.
2026-10-01T23:31:30Z
evidence attached: reddit.post.1wvcxi8 — Independent corroboration: a second user reports OpenAI support admitting Codex requests may be served by subagents on 5.6-Sol with no model-determinism guarantee, reinforcing the silent non-Astra execution concern.
2026-09-30T20:26:35Z
The retry-UI report (1wu8036) is the first periphery item about removed identity visibility rather than a visible routing oddity — the web tooltip reportedly no longer names the serving model — a small but on-theme sign that served-model transparency is eroding, which nudges the case's meaning toward 'detection must be gateway-side' without moving the hypothesis itself: still unreplicated ~8 days on, officially unaddressed beyond the scoped admission, with no DevDay fallout yet visible in this periphery.
2026-09-30T18:42:39Z
evidence attached: reddit.post.1wu8036 — Retry UI no longer shows which model served a response, a transparency regression directly relevant to detecting silent model substitution.
2026-09-29T07:05:30Z
The claim itself is frozen — original post crept 102→108 with comments flat at 34 six days on, the HN Codex-vs-API parity post is trivia (3 pts, 0 comments), and the 404 is the already-known unidentified-object decay signal — while the ambient wave that carried the heat has crested and is decaying (accelerating→cooling, ~10→2.5 pts/h, 2.2→0.17 comments/h; the 1wseqni velocity spike is the peer-relative tail of its earlier run, its ~68 pts/h peak long past). An unreplicated, officially unaddressed allegation inside a cooling periphery prices as low heat, not medium; DevDay (~Sep 29–30) is the standing re-ignition vector and will surface via the hot openai topic regardless.
2026-09-29T06:24:58Z
evidence attached: hn.story.49888619 — Second independent parity complaint that the Codex surface under-delivers versus the API, materially contextualising the provider-side substitution trust question.
2026-09-28T22:49:49Z
The new degradation anecdote (1wsn71m) is the signature the hypothesis predicts but carries no discriminating power — equally consistent with sampling variance, load, or pre-DevDay compute reallocation, the explanation its own comments reach for (0.57 ratio, contested; 2 pts). What changed is delivery, not belief: the ambient 'Astra degraded / served ≠ selected' wave keeps expanding across three platforms with the picker post (62→87 pts) still its hot object, lifting aggregate engagement from ~4.7 to ~10 pts/h and comments/h from 0.5 to 2.2 — so attention temperature rises to medium while the spoofed-slug hypothesis itself stays exactly where it was: unreplicated ~6 days on, no official response beyond the scoped admission.
2026-09-28T21:36:33Z
evidence attached: reddit.post.1wsn71m — Sudden same-context day-over-day Astra degradation with a confirming comment is the user-visible signature the silent-rerouting hypothesis predicts, though anecdotal and unverified.
2026-09-28T17:36:34Z
The new picker anecdote (1wseqni, grown 28→62 pts) is the first independent Astra-selected→Luna-served sighting, but on consumer ChatGPT with the switch visible to the user and a mundane work-mode/metered-gating explanation rising in its own comments — it enriches the broader 'served ≠ selected' periphery without touching the claim's teeth: silent substitution with spoofed Astra slugs in Codex/API, still without slug-independent replication ~5.5 days on. The case's meaning is essentially unchanged — an unverified allegation embedded in a larger, mostly transparent routing/gating pattern — and low heat stands because the 83rd peer percentile is the young tangential picker post, not claim-side reignition (0.5 comments/h aggregate, no replication, no official response).
2026-09-28T15:45:04Z
evidence attached: reddit.post.1wseqni — First-hand report of picking Astra in the picker and being served Luna with max reasoning independently supports the silent-rerouting case.
2026-09-27T23:25:04Z
The newly attached Ask HN thread is thematic resonance, not case evidence: it voices the silent-downgrade/fair-token-counting trust concern in generic regulatory terms without referencing Astra, nlight, or any substitution evidence, so it corroborates the ambient concern rather than the claim. The case's meaning is unchanged — a dormant single-source allegation with a stalled, unexamined community checker — while the concern generalizing into weights-and-measures discourse mildly strengthens the timeliness of Scott's identity-verification angle without strengthening nlight's hypothesis.
2026-09-27T23:23:44Z
evidence attached: hn.story.49871747 — Ask HN independently voices the exact silent-downgrade trust concern (downgrades you may not know about, fair token counting) at the core of the Astra rerouting case.
2026-09-27T02:47:32Z
grounded: contradicts/high — Unverified but cheap-to-test allegation that attacks a load-bearing assumption across Scott's stack — that the model slug his LiteLLM gateway and Codex CLI logs
2026-09-27T02:40:36Z
The community checker (is-my-astra-real.pages.dev) has graduated from a link inside one user's HN anecdote to a standalone public artifact, confirming the episode has produced a purpose-built test instrument — but with 1 point, 0 comments and unexamined methodology it produces no new fact about the claim itself. Meaning shifts from 'lone seed allegation' to 'watching': an unverified claim with a small ecosystem forming around testing it, still awaiting the one thing that would settle it — a slug-independent replication or an official response.
2026-09-27T02:22:37Z
evidence attached: hn.story.49862460 — A community-built 'is your Astra real' checker operationalizes the silent-rerouting claim, showing the episode propagating beyond the original investigator.
2026-09-26T14:41:14Z
The new HN thread is a second-platform anecdote but corroborates only the weaker 'Astra feels nerfed' perception (replayed prompts degrade), not nlight's substitution-with-spoofed-identity mechanism — still no slug-independent replication and no official response, so the case's meaning is unchanged: a dormant, untested allegation. The one genuine lead is the unexamined is-my-astra-real.pages.dev link, possibly a community model-identity checker; the measured 71st peer percentile is just the young 4-point HN node, not reignition.
2026-09-26T14:25:53Z
evidence attached: hn.story.49856405 — An independent Codex user recounting conversion from skeptic to believer on Astra nerfing is weak but direct community corroboration of perceived silent model degradation behind the rerouting claim.
2026-09-25T17:57:53Z
The Sep-24 velocity spikes (9x peer baseline) were a single transient engagement flare that decayed to zero within ~two days; the comment traffic it produced attacks the methodology (behavioral-only discrimination, 50-account provenance) rather than corroborating the claim. Meaning shifts from 'fresh allegation that might ignite' to 'dormant single-source claim with a public test script and no takers' — not disproved, but no independent replication has landed despite the easiest check being published.
2026-09-23T23:14:28Z
origin walked (opencode/cheap-glm, conf 0.85): anchor reddit.post.1woejs6 -> echo.other.f1131cc234 by Anonymous ("one team", not affiliated with OpenAI) — Reddit user nlight
2026-09-23T22:12:35Z
grounded: contradicts/high — This is an unverified allegation, but if it holds it attacks a load-bearing assumption across Scott's stack: that the model reported in Codex CLI responses and
2026-09-23T22:06:16Z
case created — A reproducible, testable allegation of silent model substitution distinct from the Moonshot and CrofAI episodes, currently one investigation awaiting corroboration.