Gemini 3.8 Live and 3.8 Live Extended Thinking are Google's real-time voice models (announced 2026-09-15 by Tom Ouyang, Malini Jaganathan and the Gemini Audio Team), claiming uninterrupted dialogue with background tool execution and, in Extended Thinking, simultaneous multi-step reasoning plus speech. Per the case's own evidence (retrieved first-party post, API/AI Studio rollout, Artificial Analysis and Sierra τ-Voice benchmarks, integrations across LiveKit, Pipecat, Agora, Vercel, LangChain, Fishjam, Vision Agents), the models are released and developer-accessible; the supplied web snippets, however, do not independently cover 3.8 Live — they confirm the surrounding landscape instead: Gemini 3.1 Flash Live already shipped synchronous function calling during live audio and Gemini 2.5 Native Audio shipped asynchronous tool execution, while xAI's grok-voice-think-fast-1.0 and OpenAI's GPT-Live/GPT-Realtime-2 pursue the same full-duplex 'no awkward pauses' goal. So the specific 3.8 concurrency claims remain documented-but-independently-untested, and the snippets raise a real novelty question versus Gemini 2.5's asynchronous tool use and 3.1's synchronous function calling.
Google has shipped, in a developer-accessible frontier product, exactly the architecture Scott theorised in his Fast-Slow Split / Cognitive Pipelining work — a fast conversational lane with slow reasoning and tool execution running on a different clock — making this a dated-receipts publishing opportunity ('the pattern I named is now a product'), and one he can hands-on test in AI Studio against his Twilio/Ultravox voice-agent lab. It stays medium rather than high because the load-bearing claims (sustained concurrent execution, interruption/cancellation behaviour) remain documented-but-unvalidated with a live novelty question versus Gemini 2.5's async tool execution; a validated hands-on would raise it, and the Voice AI's Fork lens (continuity ≠ verification) is the critical frame for testing whether concurrency actually enables authorised action or just fluent talk over unverified writes.
ip:framework.fast-slow-splitip:concept.cognitive-pipeliningip:source.the-fast-slow-splitip:framework.voice-ais-forkdev:project.twiliodev:technology.geminiradar:concept.voice-agentsradar:openai-gpt-live-voice-architectureradar:openai-gpt-live-1-apiradar:gemini-omni-11-flash-releaseradar:concept.tool-calling
queries asked of Scott's wikis
- fast-slow split voice agents reasoning while speaking
- cognitive pipelining concurrent tool execution dialogue
- async tool execution cancellation interruption behavior agent harness
- Live API hands-on test plan AI Studio voice model evaluation
- continuity vs verification voice agent authorized action
- frontier voice model benchmark tau-voice speech agent comparison notes
| source | object | author | score | comments |
| 🟠 reddit | Gemini 3.8 Live & Gemini 3.8 Live Extended Thinking singularity Retrieved article excerptOpen article · Retrieved 2026-09-15T19:22:03.763085+00:00 # Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Sep 15, 2026
|
- [x.com](https://twitter.com/intent/tweet?text=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking%20%40google&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [Facebook](https://www.facebook.com/sharer/sharer.php?caption=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking&u=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [LinkedIn](https://www.linkedin.com/shareArticle?mini=true&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/&title=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking)
- Mail
- Copy link
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice.
---
Tom Ouyang
Principal Engineer
Malini Jaganathan
Member of Technical Staff, on behalf of the Gemini Audio Team
Share
- [x.com](https://twitter.com/intent/tweet?text=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking%20%40google&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [Facebook](https://www.facebook.com/sharer/sharer.php?caption=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking&u=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [LinkedIn](https://www.linkedin.com/shareArticle?mini=true&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/&title=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking)
- Mail
- Copy link
---
Text "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" with the Gemini Spark, all on a light blue background
Your browser does not support the audio element.
Listen to article
[[duration]] minutes
This content is generated by Google AI. Generative AI is experimental
Voice
Speed
Voice
Speed
0.75X
1X
1.5X
2X
Read AI-generated summary
We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app.
Summaries were generated by Google AI. Generative AI is experimental.
- Check out "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" for smarter voice AI.
- Gemini 3.8 Live offers fast, fluid conversations with real-time visual and language support.
- Use 3.8 Live Extended Thinking to handle complex tasks while keeping the conversation flowing.
- These models work in the background to manage tools while you keep chatting.
- You can try these new features in Google Workspace, Search, and the Gemini app.
Summaries were generated by Google AI. Generative AI is experimental.
Google just launched two new AI models that make talking to your devices feel way more natural. They can handle interruptions, switch between languages, and even explain their thought process while they work. Whether you're solving a complex problem or just chatting, the AI now feels like it's actually listening and thinking along with you. It’s a big step toward making AI feel like a real conversation partner.
Summaries were generated by Google AI. Generative AI is experimental.
#### Explore other styles:
- General summary
- Bullet points
- Basic explainer
Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.
- **Gemini 3.8 Live**: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
- **Gemini 3.8 Live Extended Thinking**: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.
For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.
## Experience more fluid, intelligent conversations
Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on *τ*-Voice and 35.1% on Sierra’s *τ*-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.
Gemini 3.8 Live has shown a high preference among users, securing a second place in the [Speech Agent Arena](https://artificialanalysis.ai/speech-to-speech?api-benchmarks=agentic-performance-vs-cost-to-run#speech-to-speech-arena-results-tabs). In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale.
a chart showing Artificial Analysis Speech to Speech Index
A chart showing Artificial Analysis agentic performance
A chart showing Sierra
A chart showing Artificial Analysis cost per hour of input audio
On ServiceNow’s [EVA-Bench](https://servicenow.github.io/eva/#results), a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality.
Note: This was run on the Live API on Gemini Enterprise Agent Platform.
A chart showing EVA Bench Experience to Task Completion
Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background.
*Gemini 3.8 Live guides employee onboarding in real time, using visual context to answer live questions.*
*Watch Gemini 3.8 Live play chess in near real-time using visual context, reasoning, and natural conversational flow.*
For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like *“Let me check that…”* to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress.
*Watch Gemini 3.8 Live Extended Thinking transform raw sketches and near real-time voice feedback into functional React components.*
*See Gemini 3.8 Live Extended Thinking coordinate multi-step bookings and asynchronous function calls — all without interrupting natural live conversation.*
*Watch Gemini 3.8 Live build complete business plans and custom marketing toolkits on the fly through natural speech.*
Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks.
*Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live.*
## Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live — right inside Search Live.
## Empowering the developer and enterprise voice ecosystem
By using the [Gemini Live API](https://ai.google.dev/gemini-api/docs/live-api), developer platforms such as [Agora](https://docs.agora.io/en/ai/models/mllm/gemini), [Fishjam](https://docs.fishjam.io/tutorials/gemini-live-integration), [LangChain](https://docs.langchain.com/langsmith/trace-gemini-live), [LiveKit](https://docs.livekit.io/agents/models/realtime/plugins/gemini/), [Pipecat](https://docs.pipecat.ai/pipecat/features/gemini-live), [Vercel](https://vercel.com/docs/ai-gateway/modalities/realtime), and [Vision Agents](https://visionagents.ai/integrations/realtime/gemini) enable developers to build and deploy high-performance voice-driven interfaces with ease. These platforms manage complex real-time media streaming infrastructure behind the scenes, allowing developers to focus entirely on crafting the user experience.
We’re also partnering with companies like Salesforce, Genspark, and Lumeris who are excited about 3.8 Live and 3.8 Live Extended Thinking, highlighting its impressive latency, fluidity, and tool-calling capabilities.
Salesforce quote
11Sight Quote
Equal AI Quote
ServiceNow quote
Genspark quote
Lenskart quote
Lumeris quote
Agora quote
Ambr AI quote
LiveKit Quote
Casuu quote
## Ensure transparency with SynthID watermarking
All audio generated by our AI products is watermarked with [SynthID](https://deepmind.google/models/synthid/). This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the [model card](https://deepmind.google/models/model-cards/gemini-3-8-audio/).
## Start using our latest Gemini Audio models:
3.8 Live is rolling out starting today:
- **For developers**: In the [Gemini API](https://ai.google.dev/gemini-api/docs/live-api) and [Google AI Studio](https://aistudio.google.com/live)
- **For enterprises**: In private preview in [Gemini Enterprise](https://console.cloud.google.com/vertex-ai/studio/multimodal-live) and coming soon to [Gemini Enterprise for Customer Experience](https://cloud.google.com/gemini-enterprise-cx)
- **For everyone**: In Search Live
3.8 Live Extended Thinking is rolling out starting today:
- **For developers**: In the [Gemini API](https://ai.google.dev/gemini-api/docs/live-api) and [Google AI Studio](https://aistudio.google.com/live)
- **For enterprises**: In private preview in [Gemini Enterprise](https://console.cloud.google.com/vertex-ai/studio/multimodal-live) and coming soon to [Gemini Enterprise for Customer Experience](https://cloud.google.com/gemini-enterprise-cx) and Google Workspace business customers
- **For everyone**: In Gemini Live and for Google AI Pro and Ultra subscribers in Workspace in Docs, and all Google AI subscribers in Gmail and Keep
## Get the latest news from Google in your inbox
Sign up for our newsletters with product updates, event information, special offers, and more.
Done. Just one step more.
Check your inbox to confirm your subscription.
You can also subscribe with a different email address.
Your information will be used in accordance with [Google's privacy policy.](https://policies.google.com/privacy) You may opt out at any time.
Posted in: | gibbonwalker | 276 | 129 |
| 🟧 hn | Gemini 3.8 Live and 3.8 Live Extended ThinkingRetrieved article excerptOpen article · Retrieved 2026-09-15T19:22:06.137132+00:00 # Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Sep 15, 2026
|
- [x.com](https://twitter.com/intent/tweet?text=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking%20%40google&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [Facebook](https://www.facebook.com/sharer/sharer.php?caption=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking&u=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [LinkedIn](https://www.linkedin.com/shareArticle?mini=true&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/&title=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking)
- Mail
- Copy link
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice.
---
Tom Ouyang
Principal Engineer
Malini Jaganathan
Member of Technical Staff, on behalf of the Gemini Audio Team
Share
- [x.com](https://twitter.com/intent/tweet?text=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking%20%40google&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [Facebook](https://www.facebook.com/sharer/sharer.php?caption=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking&u=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/)
- [LinkedIn](https://www.linkedin.com/shareArticle?mini=true&url=https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/&title=Introducing%20Gemini%203.8%20Live%20and%203.8%20Live%20Extended%20Thinking)
- Mail
- Copy link
---
Text "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" with the Gemini Spark, all on a light blue background
Your browser does not support the audio element.
Listen to article
[[duration]] minutes
This content is generated by Google AI. Generative AI is experimental
Voice
Speed
Voice
Speed
0.75X
1X
1.5X
2X
Read AI-generated summary
We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app.
Summaries were generated by Google AI. Generative AI is experimental.
- Check out "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" for smarter voice AI.
- Gemini 3.8 Live offers fast, fluid conversations with real-time visual and language support.
- Use 3.8 Live Extended Thinking to handle complex tasks while keeping the conversation flowing.
- These models work in the background to manage tools while you keep chatting.
- You can try these new features in Google Workspace, Search, and the Gemini app.
Summaries were generated by Google AI. Generative AI is experimental.
Google just launched two new AI models that make talking to your devices feel way more natural. They can handle interruptions, switch between languages, and even explain their thought process while they work. Whether you're solving a complex problem or just chatting, the AI now feels like it's actually listening and thinking along with you. It’s a big step toward making AI feel like a real conversation partner.
Summaries were generated by Google AI. Generative AI is experimental.
#### Explore other styles:
- General summary
- Bullet points
- Basic explainer
Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.
- **Gemini 3.8 Live**: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
- **Gemini 3.8 Live Extended Thinking**: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.
For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.
## Experience more fluid, intelligent conversations
Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on *τ*-Voice and 35.1% on Sierra’s *τ*-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.
Gemini 3.8 Live has shown a high preference among users, securing a second place in the [Speech Agent Arena](https://artificialanalysis.ai/speech-to-speech?api-benchmarks=agentic-performance-vs-cost-to-run#speech-to-speech-arena-results-tabs). In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale.
a chart showing Artificial Analysis Speech to Speech Index
A chart showing Artificial Analysis agentic performance
A chart showing Sierra
A chart showing Artificial Analysis cost per hour of input audio
On ServiceNow’s [EVA-Bench](https://servicenow.github.io/eva/#results), a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality.
Note: This was run on the Live API on Gemini Enterprise Agent Platform.
A chart showing EVA Bench Experience to Task Completion
Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background.
*Gemini 3.8 Live guides employee onboarding in real time, using visual context to answer live questions.*
*Watch Gemini 3.8 Live play chess in near real-time using visual context, reasoning, and natural conversational flow.*
For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like *“Let me check that…”* to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress.
*Watch Gemini 3.8 Live Extended Thinking transform raw sketches and near real-time voice feedback into functional React components.*
*See Gemini 3.8 Live Extended Thinking coordinate multi-step bookings and asynchronous function calls — all without interrupting natural live conversation.*
*Watch Gemini 3.8 Live build complete business plans and custom marketing toolkits on the fly through natural speech.*
Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks.
*Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live.*
## Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live — right inside Search Live.
## Empowering the developer and enterprise voice ecosystem
By using the [Gemini Live API](https://ai.google.dev/gemini-api/docs/live-api), developer platforms such as [Agora](https://docs.agora.io/en/ai/models/mllm/gemini), [Fishjam](https://docs.fishjam.io/tutorials/gemini-live-integration), [LangChain](https://docs.langchain.com/langsmith/trace-gemini-live), [LiveKit](https://docs.livekit.io/agents/models/realtime/plugins/gemini/), [Pipecat](https://docs.pipecat.ai/pipecat/features/gemini-live), [Vercel](https://vercel.com/docs/ai-gateway/modalities/realtime), and [Vision Agents](https://visionagents.ai/integrations/realtime/gemini) enable developers to build and deploy high-performance voice-driven interfaces with ease. These platforms manage complex real-time media streaming infrastructure behind the scenes, allowing developers to focus entirely on crafting the user experience.
We’re also partnering with companies like Salesforce, Genspark, and Lumeris who are excited about 3.8 Live and 3.8 Live Extended Thinking, highlighting its impressive latency, fluidity, and tool-calling capabilities.
Salesforce quote
11Sight Quote
Equal AI Quote
ServiceNow quote
Genspark quote
Lenskart quote
Lumeris quote
Agora quote
Ambr AI quote
LiveKit Quote
Casuu quote
## Ensure transparency with SynthID watermarking
All audio generated by our AI products is watermarked with [SynthID](https://deepmind.google/models/synthid/). This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the [model card](https://deepmind.google/models/model-cards/gemini-3-8-audio/).
## Start using our latest Gemini Audio models:
3.8 Live is rolling out starting today:
- **For developers**: In the [Gemini API](https://ai.google.dev/gemini-api/docs/live-api) and [Google AI Studio](https://aistudio.google.com/live)
- **For enterprises**: In private preview in [Gemini Enterprise](https://console.cloud.google.com/vertex-ai/studio/multimodal-live) and coming soon to [Gemini Enterprise for Customer Experience](https://cloud.google.com/gemini-enterprise-cx)
- **For everyone**: In Search Live
3.8 Live Extended Thinking is rolling out starting today:
- **For developers**: In the [Gemini API](https://ai.google.dev/gemini-api/docs/live-api) and [Google AI Studio](https://aistudio.google.com/live)
- **For enterprises**: In private preview in [Gemini Enterprise](https://console.cloud.google.com/vertex-ai/studio/multimodal-live) and coming soon to [Gemini Enterprise for Customer Experience](https://cloud.google.com/gemini-enterprise-cx) and Google Workspace business customers
- **For everyone**: In Gemini Live and for Google AI Pro and Ultra subscribers in Workspace in Docs, and all Google AI subscribers in Gmail and Keep
## Get the latest news from Google in your inbox
Sign up for our newsletters with product updates, event information, special offers, and more.
Done. Just one step more.
Check your inbox to confirm your subscription.
You can also subscribe with a different email address.
Your information will be used in accordance with [Google's privacy policy.](https://policies.google.com/privacy) You may opt out at any time.
Posted in: | leumon | 493 | 326 |
| 🟧 echo.blog ⭐ | Google announces both models rolling out through the Gemini API and AI Studio, claiming background tool execution during conversation and si | Tom Ouyang and Malini Jaganathan, Google Gemini Audio Team | — | — |
| 🟧 hn | Gemini 3.8 Flash outperforms SOTA models at agentic CAD coding | Mazer23 | 2 | 0 |
| 🟧 hn | Gemini 3.8 text-to-speech says hello | swolpers | 331 | 138 |
| 🟠 reddit | Gemini 3.8 Flash TTS and Flash-Lite TTS Announced singularity | Recoil42 | 101 | 11 |
| 🟧 hn | Gemini 3.8 TTS Playground | lumpa | 1 | 0 |
| 🟧 hn | Power your agents: Gemini 3.8 Live with Live Avatar is now generally available | planteur | 4 | 0 |
| 🟧 hn | Google is launching a one-stop Gemini agent for your work tasks | mikelgan | 3 | 0 |
| 🟧 hn | Gemini Agent Introduced | coleca | 14 | 2 |
| 🟧 hn | The Gemini Agent | isnotchicago | 16 | 2 |
| 🟧 hn | Welcome to Gemini at Work 2026: Introducing the Gemini Agent | oscarfr | 2 | 1 |
2026-10-09T12:21:45Z
Enterprise GA milestone (2026-10-08) is now fully absorbed; the case is a shipped, developer-accessible product with benchmarks, platform integrations, and enterprise GA confirmed. Measured heat remains at baseline (0.33 pts/h, 45th percentile, steady) despite hot topic neighbourhood (frontier-models, voice-agents). Core concurrency claims — background tool execution during uninterrupted dialogue; Extended Thinking simultaneous reasoning+speech with early verbal cues — remain documented with demos but lack independent hands-on validation of sustained task accuracy, interruption/cancellation behaviour, or novelty versus Gemini 2.5 Native Audio async tools, 3.1 Flash Live sync function calling, and xAI/OpenAI full-duplex peers. The case sits in a holding pattern awaiting production validation.
2026-10-09T09:41:27Z
evidence attached: hn.story.50017609 — shared external link with case evidence
2026-10-09T08:11:18Z
Enterprise GA milestone (2026-10-08) completed the rollout arc; since then no independent hands-on validation of the differentiating concurrency claims (background tool execution during uninterrupted dialogue, Extended Thinking simultaneous reasoning+speech with early verbal cues). Measured engagement has cooled to baseline (0.0 pts/h, 39th percentile, steady across 3 platforms) while the topic neighbourhood (voice-agents, frontier-models) remains hot. The case is now a shipped, developer-accessible product whose core novelty claims versus Gemini 2.5 Native Audio async tools and 3.1 Flash Live sync function calling — and versus xAI/OpenAI full-duplex peers — remain documented but unvalidated in production voice-agent workloads.
2026-10-08T23:06:43Z
evidence attached: hn.story.50012543 — shared external link with case evidence
2026-10-08T20:55:51Z
Enterprise GA milestone landed: Google Cloud's first-party 'Gemini at Work 2026' blog and The Verge coverage confirm the enterprise agent launch built on 3.8 Live, moving the rollout arc from developer launch + private preview to full enterprise GA. Measured heat shows cross-platform spread at 80th percentile (steady, 3 platforms) and topic neighbourhood is hot, so attention lifts from cold despite the core concurrency claims remaining independently unvalidated.
2026-10-08T17:49:20Z
evidence attached: hn.story.50005566 — Official Google Cloud blog announcing 'Gemini at Work 2026', the first-party launch counterpart to the Gemini 3.8 Live release case.
2026-10-08T17:49:20Z
evidence attached: hn.story.50006368 — The Verge coverage of Google's 'Gemini at Work 2026' enterprise agent launch, corroborating the Gemini 3.8 Live rollout for work tasks.
2026-09-25T04:43:45Z
Gemini 3.8 Live with Live Avatar is now reported generally available ~10 days post-launch — the rollout arc (launch → private preview → GA) is complete, so the case's meaning shifts from 'launch to track' to 'shipped product whose specific concurrency claims remain independently unvalidated'. The GA milestone is headline-level (no retrieved page) but first-party and about the exact tracked models, unlike the adjacent TTS/Flash attachments; attention is fully cold (0.17 pts/h vs 280 peak, 30th percentile) and the magnitude-valve spread remains launch echo, not new adoption of concurrent execution. A first independent hands-on of sustained background tool execution / interruption behaviour is the event that would matter next.
2026-09-25T04:22:27Z
evidence attached: hn.story.49840035 — First-party GA milestone for the same models, adding Live Avatar capability that extends the release case.
2026-09-24T03:31:02Z
grounded: converges/medium — Google has shipped, in a developer-accessible frontier product, exactly the architecture Scott theorised in his Fast-Slow Split / Cognitive Pipelining work — a
2026-09-24T03:24:23Z
Re-reading the accumulated evidence rather than reacting to the TTS wave: the retrieved first-party post, live API/AI Studio rollout, third-party benchmarks (Artificial Analysis, ServiceNow EVA-Bench) and seven independent platform integrations with published docs (LiveKit, Pipecat, Agora, Vercel, LangChain, Fishjam, Vision Agents) constitute two-plus independent lines confirming the models are released and developer-accessible — corroborated, not just an announcement. The recent heat is launch echo plus adjacent TTS/Flash releases, not adoption of Live's concurrent execution, so heat stays low despite the magnitude-valve reading.
2026-09-24T00:31:35Z
evidence attached: hn.story.49824314 — Willison's hands-on of Gemini 3.8's TTS stack is third-party practical context on the same voice-agent release wave the open case tracks.
2026-09-23T21:46:35Z
evidence attached: reddit.post.1wobp4x — shared external link with case evidence
2026-09-23T18:05:17Z
The new TTS discussion concerns speech generation and quoted voice-replication features, not evidence that Live sustains dialogue during reasoning and tool execution. Although the spread sensor is loud, its apparent expansion combines the original announcement with different Gemini models; it does not show renewed momentum for this specific capability.
2026-09-23T16:27:31Z
evidence attached: hn.story.49817615 — First-party Google release extending the Gemini 3.8 voice stack the open Live-models case tracks.
2026-09-18T18:41:17Z
The new CAD-coding headline concerns Gemini 3.8 Flash, not either Live model, and supplies no evaluation details; it does not independently corroborate concurrent voice, reasoning or tool execution. The attachment therefore leaves this case a publisher-announced capability awaiting model-specific implementation evidence.
2026-09-18T18:22:40Z
evidence attached: hn.story.49757903 — An external agentic-CAD evaluation provides capability evidence for Gemini 3.8 Flash, though the superiority claim remains publisher-reported.
2026-09-16T15:42:14Z
The refreshed discussion adds no model-specific implementation evidence or credible contradiction; general Gemini opinions and the previously assessed Afrikaans anecdote do not test concurrent reasoning and tool execution. The case remains a concrete publisher announcement awaiting independent validation, not a newly accelerating development.
2026-09-16T03:28:24Z
The new firsthand voice-use report does not identify the model or test concurrent reasoning and tool execution, so it adds no validation of the claimed advance. Discussion remains mostly amplification and general product opinion; the release warrants tracking, not renewed urgency.
2026-09-15T19:28:34Z
grounded: converges/medium — The claimed Google capability would converge directly with Scott’s Fast-Slow Split and Cognitive Pipelining, creating a potential publishing opportunity around
2026-09-15T19:22:35Z
case created — A concrete API model launch with detailed capability claims merits one consolidated case; the two discussions echo the same announcement rather than independently validating performance.