Skip to main content
Back to News
Google ships Gemini 3.8 Live and Extended Thinking for real-time multimodal dialogue
Product
2 min read
US

Google ships Gemini 3.8 Live and Extended Thinking for real-time multimodal dialogue

The AMW Read

Incremental Gemini 3.8-family SKU after Flash, but segment-level because live multimodal dialogue with background tool use ships across API, Workspace, and consumer Gemini at once.
NoveltySignificance
Foundation Models · Player Map

Google ships Gemini 3.8 Live and Extended Thinking for real-time multimodal dialogue

Google introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, calling them its most advanced live dialogue models yet. The release targets near real-time voice agents with upgrades in intelligence and parallel reasoning, real-time visual context, and background task execution that continues without interrupting the conversation. Gemini 3.8 Live is built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Extended Thinking is aimed at high-complexity work with increased intelligence and multi-step reasoning. Availability begins through the Gemini API, Google Workspace, and the Gemini app, with Search also listed among try surfaces.

The move extends the Gemini 3.8 family Google shipped weeks earlier with 3.8 Flash, shifting the same reasoning-first pitch into live multimodal dialogue. That matters because frontier competition is no longer only about static chat quality: hyperscalers are racing to make voice agents that can see context, keep talking, and run tools in the background. Landing the same models across consumer Gemini, Workspace, Search, and the developer API compresses the path from model release to distribution, raising the product bar for rivals that still separate assistant UX from API SKUs.

Builders should treat the pair as a latency-versus-depth tradeoff—Live for fluid, cost-sensitive conversational loops; Extended Thinking when multi-step reasoning must run while the dialogue stays open—and test interruption handling and visual grounding under real workloads. Investors tracking foundation-model platforms should read this as Google packaging agentic voice behavior as a first-class model tier, not merely an app feature.

#Google #Gemini #FoundationModels #VoiceAI #MultimodalAI

#Google#Gemini 3.8 Live#voice agents#multimodal AI#foundation models#Extended Thinking
Read Original

How This Connects

Based on Foundation Models · Player Map

  1. 11h agoDeepSeek reportedly nears RMB 80 billion funding round with Tencent and CATLDeepSeek
  2. 1d agoDeepSeek reportedly nears $12 billion round as investor demand lifts its targetDeepSeek
  3. 3d agoAnthropic infrastructure financing reportedly reaches $60B with Broadcom supportAnthropic
  4. 4d agoAnthropic reportedly files confidentially for a potential October 2026 IPOAnthropic
  5. 6d agoAnthropic reportedly secures up to $42B in Broadcom financing for AI infrastructureAnthropic
  6. 3w agoGoogle ships Gemini 3.8 Live and Extended Thinking for real-time multimodal dialogue · THIS ARTICLE

Related News

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard