2026-10-11 16:38 UTC

Griffin's makers claim it is the first Human Interaction Model to pass a video Turing test โ€” 44% of judges took it for a human versus roughly 3% for competing systems โ€” and to rank #1 on NVIDIA's full-duplex video benchmark; identification of the vendor and independent testing resolve whether realtime interactive video crossed perceived-human indistinguishability or this is an inflated demo echo.

state: watchingheat: mediumuncertainty: mediumconvergesscott: highfull-duplex-video-models realtime-interactive-models capability-benchmarksNVIDIA
Surfaced 2026-10-04T13:50:36Z โ€” Tavus's official announcement page (dated "San Francisco, California October 1st, 2026"): "Today we're introducing Griffin, our first Human โ€” The r/OpenAI echo matured into a second front-page thread (12โ†’654 pts, 278 comments), completing cross-community spread โ€” yet day 4 and ~2,000 aggregate points still yield zero independent testing, NVIDIA statement, evaluator access, or debunk; the claim is now hardening into public consensus ('we're cooked') purely through repetition while the safety-gated research preview structurally blocks verification. That widening spread-vs-verification gap is the case's defining fact, but with momentum cooling to 5.8 pts/h off a 241 peak, no new evidence in kind (the magnitude-valve reading counts the same two Reddit threads; the only other object is Tavus's own page), and a 90th-percentile peer rate that just reflects residual votes on aging posts, the episode's attention value has decayed โ€” cooled to low, awaiting external triggers (independent eval, access change, NVIDIA comment, GA timeline).

What is this?

Griffin is a full-duplex video-to-video conversational model announced October 1, 2026 by Tavus, a San Francisco AI research lab previously known for its conversational video avatar stack (Phoenix 4.5, Sparrow-2, Raven-1); it generates full video frames in real time and can listen, watch, and speak simultaneously. Tavus's own live blind study found 48% of participants on a one-minute video call took Griffin for a real human, versus a โ‰ค2โ€“3% pass rate for prior systems including Tavus's own (the case's 44% figure is a weaker echo; the vendor's materials and launch coverage say 48%), and it is credited #1 on the Video Full-Duplex Benchmark that NVIDIA built and scored. It is a research preview, not a public release โ€” Tavus itself says safety review must precede general availability. The Turing-test claim rests on Tavus's own study: the supplied material shows no independent replication, and NVIDIA's documented role is building/scoring the duplex benchmark, not certifying the human-indistinguishability claim.

Why it matters to Scott

Converges with his realtime-interaction and evidence-discipline canon: NVIDIA institutionalizing full-duplex video as a scored benchmark category extends the exact deployment class his Real-Time AI Systems / Fast-Slow Split work and Twilio/Ultravox labs map, while the 44โ€“48% Turing-test number is a vendor-run study with no independent replication โ€” sitting precisely where his Evidence Class Ladder and Positioning Ladder say an unverified claim must be priced, and a live test case for the radar's benchmark-integrity lineage (same actor as the prior Sparrow-2 case). If the claim holds, his Chat-Era-Trust / Cryptographic-Trust position gains its strongest realtime-video forcing function ('your eyes are no longer a verifier'); if it collapses under independent testing, it is a dated receipt for the same argument โ€” and either branch touches what he builds, since realtime duplex video is the next substrate for his voice-agent and talking-head prospecting offers.
ip:concept.evidence-class-ladderip:concept.positioning-ladderip:framework.category-transition-lintip:concept.real-time-ai-systemsip:concept.chat-era-trust-modelip:concept.cryptographic-trustip:concept.synthetic-personasdev:project.twiliodev:technology.ultravoxwork:concept.ai-personalised-outbound-prospectingradar:tavus-sparrow-2-conversational-flowradar:concept.voice-agentsradar:concept.benchmark-integrityradar:concept.model-evaluationradar:realtime-venus-open-av-interactionradar:gemini-38-live-release
queries asked of Scott's wikis
  • vendor-coined AI category launch framing
  • self-reported benchmark claims vs independent evals
  • Turing test marketing capability claims
  • realtime full-duplex video voice agent stack
  • safety-gated research preview release pattern
  • synthetic human indistinguishability trust

Measured heat

now 0 pts/hpeak 241 pts/hcomments 0/hpeers p35momentum: steady2 platformsage 266h
points/hour across evidence ยท reading as of 2026-10-12 02:59:37.977291+11:00 ยท deterministic, not a model opinion

How the heat travelled

09-30 14:00โญ origin echo-reconstructedTavus's official announcement page (dated "San Francisco, California October 1st, 2026"): "Today we're introducing Griffin, our first Human
Tavus (by Hassaan Raza, Co-founder & CEO; Ioannis Patras, Head of Research; Tavus Research Team) on blog (echo) ยท attributed from reddit.post.1wv7q40
โ€”
10-01 18:47first on r/singularity ยท published ยท +28.8hGriffin, the first Human Interaction Model to pass video Turing Test it's already #1 on NVIDIA's benchmark for full-duplex AI video - 44% of people thought it was a real person while other systems are at ~3%
Distinct-Question-16
โ€”
10-02 13:09first on r/OpenAI ยท published ยท +47.2hWe're cooked. This is the first model to pass the Video Turing Test. Half of people who talked to it thought it was a real human
Puzzleheaded-King584
โ€”
10-11 05:00first on r/artificial ยท published ยท +255.0hTavus's Griffin avatar fooled 48% of job-interview subjects, and the fooled were as confident as the un-fooled
lulzxdxdxd
โ€”
10-01 18:47amplified on r/singularity ๐Ÿ‘‘reddit.post.1wv7q40
Distinct-Question-16
peak 1390 ยท 378 comments ยท 62% of case engagement
10-02 13:09amplified on r/OpenAIreddit.post.1wvtedh
Puzzleheaded-King584
peak 779 ยท 312 comments ยท 38% of case engagement
10-11 05:00amplified on r/artificialreddit.post.1x2zfuj
lulzxdxdxd
peak 0 ยท 1 comments ยท 0% of case engagement
10-01 20:21our radar first saw it ยท +30.4hdiscovery anchor: reddit.post.1wv7q40โ€”
10-03 23:41reached heat=high ยท +81.7h ยท via ledgerโ€”โ€”
pace: p97 vs 1188 stories at the 168h mark (now 266h old) โ€” ahead of openai-gpt61-sol-release (1.0x), behind qwen38-flash-dual-3090-speedup (1.0x)

Evidence (4) โ€” โญ canonical anchor

sourceobjectauthorscorecomments
๐ŸŸ  redditGriffin, the first Human Interaction Model to pass video Turing Test it's already #1 on NVIDIA's benchmark for full-duplex AI video - 44% of people thought it was a real person while other systems are at ~3%
singularity
Distinct-Question-161390378
๐ŸŸง echo.blog โญTavus's official announcement page (dated "San Francisco, California October 1st, 2026"): "Today we're introducing Griffin, our first Human Tavus (by Hassaan Raza, Co-founder & CEO; Ioannis Patras, Head of Research; Tavus Research Team)โ€”โ€”
๐ŸŸ  redditWe're cooked. This is the first model to pass the Video Turing Test. Half of people who talked to it thought it was a real human
OpenAI
Puzzleheaded-King584779312
๐ŸŸ  redditTavus's Griffin avatar fooled 48% of job-interview subjects, and the fooled were as confident as the un-fooled
artificial
lulzxdxdxd01

Interpretation history

Decision trace