LmCast :: Stay tuned in

Gemini 3.8 Live and 3.8 Live Extended Thinking

Recorded: Sept. 15, 2026, 7:08 p.m.

Original Summarized

Gemini 3.8 Live & Gemini 3.8 Live Extended Thinking

Skip to main content

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

Innovation & AI

Products & platforms

Company news

Feed

Newsletter

Back

Innovation & AI


See all in Innovation & AI

Models & Research

Google DeepMind

Google Research

Google Labs

Gemini models

Quantum computing

See all

Products

Developer tools

Gemini app

Gemini Notebook

See all

Infrastructure & cloud

Global network

Google Cloud

See all

Technology

Safety & Security

Health

See all

Learn more:

Google DeepMind blog

Google Research blog

Google Developers blog

Google Cloud blog

Back

Products & platforms


See all in Products & platforms

Products

Search

Maps

Chrome

Google Health

Google Workspace

Learning & Education

Shopping

See all

Platforms

Android

Google Play

Wear OS

See all

Devices

Pixel

Google Nest

Fitbit

Chromebooks

See all

Learn more:

Google Ads & Commerce blog

Waze blog

Back

Company news


See all in Company news

Outreach & initiatives

Creating opportunity

Safety & security

Google.org

Public policy

Sustainability

Health

See all

Leadership

Sundar Pichai, CEO

More authors

See all

Inside Google

Around the globe

Life at Google

See all

Learn more:

Google Security blog

Innovation & AI

Innovation & AI


See all in Innovation & AI

Models & Research

Google DeepMind

Google Research

Google Labs

Gemini models

Quantum computing

See all

Products

Developer tools

Gemini app

Gemini Notebook

See all

Infrastructure & cloud

Global network

Google Cloud

See all

Technology

Safety & Security

Health

See all

Learn more:

Google DeepMind blog

Google Research blog

Google Developers blog

Google Cloud blog

Products & platforms

Products & platforms


See all in Products & platforms

Products

Search

Maps

Chrome

Google Health

Google Workspace

Learning & Education

Shopping

See all

Platforms

Android

Google Play

Wear OS

See all

Devices

Pixel

Google Nest

Fitbit

Chromebooks

See all

Learn more:

Google Ads & Commerce blog

Waze blog

Company news

Company news


See all in Company news

Outreach & initiatives

Creating opportunity

Safety & security

Google.org

Public policy

Sustainability

Health

See all

Leadership

Sundar Pichai, CEO

More authors

See all

Inside Google

Around the globe

Life at Google

See all

Learn more:

Google Security blog

Feed

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

Share

x.com

Facebook

LinkedIn

Mail

Copy link

[]

Preferences

Global (English)

Africa (English)

Australia (English)

Brasil (Português)

Canada (English)

Canada (Français)

Česko (Čeština)

Deutschland (Deutsch)

España (Español)

France (Français)

Greece (Ελληνικά)

India (English)

Indonesia (Bahasa Indonesia)

Ireland (English)

Italia (Italiano)

日本 (日本語)

대한민국 (한국어)

Latinoamérica (Español)

Malaysia (Melayu)

الشرق الأوسط وشمال أفريقيا (اللغة العربية)

MENA (English)

Nederlands (Nederland)

New Zealand (English)

Polska (Polski)

Portugal (Português)

România (Română)

Sverige (Svenska)

ประเทศไทย (ไทย)

Türkiye (Türkçe)

台灣 (中文)

Links

Images

RSS feed

x.com

Facebook

LinkedIn

Mail

Copy link

Newsletter

Breadcrumb

Home

Innovation & AI

Models & research

Gemini Models

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

Sep 15, 2026
|

x.com

Facebook

LinkedIn

Mail

Copy link

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice.

Tom Ouyang
Principal Engineer

Malini Jaganathan
Member of Technical Staff, on behalf of the Gemini Audio Team

Share

x.com

Facebook

LinkedIn

Mail

Copy link

Your browser does not support the audio element.

Listen to article

[[duration]] minutes

This content is generated by Google AI. Generative AI is experimental

Voice

Speed

Voice

Speed
0.75X
1X
1.5X
2X

Read AI-generated summary

We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app.

Summaries were generated by Google AI. Generative AI is experimental.

Check out "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" for smarter voice AI.
Gemini 3.8 Live offers fast, fluid conversations with real-time visual and language support.
Use 3.8 Live Extended Thinking to handle complex tasks while keeping the conversation flowing.
These models work in the background to manage tools while you keep chatting.
You can try these new features in Google Workspace, Search, and the Gemini app.

Summaries were generated by Google AI. Generative AI is experimental.

Google just launched two new AI models that make talking to your devices feel way more natural. They can handle interruptions, switch between languages, and even explain their thought process while they work. Whether you're solving a complex problem or just chatting, the AI now feels like it's actually listening and thinking along with you. It’s a big step toward making AI feel like a real conversation partner.

Summaries were generated by Google AI. Generative AI is experimental.

Explore other styles:

General summary

Bullet points

Basic explainer

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.Experience more fluid, intelligent conversationsGemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena. In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale.

On ServiceNow’s EVA-Bench, a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality.

Note: This was run on the Live API on Gemini Enterprise Agent Platform.

Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background.

Gemini 3.8 Live guides employee onboarding in real time, using visual context to answer live questions.

Watch Gemini 3.8 Live play chess in near real-time using visual context, reasoning, and natural conversational flow.

For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like “Let me check that…” to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress.

Watch Gemini 3.8 Live Extended Thinking transform raw sketches and near real-time voice feedback into functional React components.

See Gemini 3.8 Live Extended Thinking coordinate multi-step bookings and asynchronous function calls — all without interrupting natural live conversation.

Watch Gemini 3.8 Live build complete business plans and custom marketing toolkits on the fly through natural speech.

Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks.

Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live.

Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live — right inside Search Live.

Empowering the developer and enterprise voice ecosystemBy using the Gemini Live API, developer platforms such as Agora, Fishjam, LangChain, LiveKit, Pipecat, Vercel, and Vision Agents enable developers to build and deploy high-performance voice-driven interfaces with ease. These platforms manage complex real-time media streaming infrastructure behind the scenes, allowing developers to focus entirely on crafting the user experience.We’re also partnering with companies like Salesforce, Genspark, and Lumeris who are excited about 3.8 Live and 3.8 Live Extended Thinking, highlighting its impressive latency, fluidity, and tool-calling capabilities.

Ensure transparency with SynthID watermarkingAll audio generated by our AI products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the model card.Start using our latest Gemini Audio models:3.8 Live is rolling out starting today:For developers: In the Gemini API and Google AI StudioFor enterprises: In private preview in Gemini Enterprise and coming soon to Gemini Enterprise for Customer ExperienceFor everyone: In Search Live3.8 Live Extended Thinking is rolling out starting today:For developers: In the Gemini API and Google AI StudioFor enterprises: In private preview in Gemini Enterprise and coming soon to Gemini Enterprise for Customer Experience and Google Workspace business customersFor everyone: In Gemini Live and for Google AI Pro and Ultra subscribers in Workspace in Docs, and all Google AI subscribers in Gmail and Keep

Get the latest news from Google in your inbox

Sign up for our newsletters with product updates, event information, special offers, and more.

Done. Just one step more.

Check your inbox to confirm your subscription.

You can also subscribe with a different email address.

Your information will be used in accordance with Google's privacy policy. You may opt out at any time.

Posted in:

Gemini models

Related stories

Safety & Security

Proactive cyber defense for governments and enterprises

By


Four Flynn

Gemini models

Introducing Gemini 3.8 Flash and 3.8 Flash Cyber

By


Tulsee Doshi

&
Raluca Ada Popa

AI

The latest AI news we announced in August 2026

By


News from Google Team

Gemini models

Introducing agentic video understanding with Gemini

By


Rohan Doshi

&
Mario Lučić

Developer tools

Gemini Omni 1.1 Flash lets you build with more control

By


Anish Nangia

&
Alisa Fortin

Gemini models

Intelligent transcription with Gemini 3.5 Transcribe

By


Diego Melendo Casado

&
Luke Leonhard

Privacy

Terms

Help

More of Google

Google Products

About the Blog

Global (English)

Africa (English)

Australia (English)

Brasil (Português)

Canada (English)

Canada (Français)

Česko (Čeština)

Deutschland (Deutsch)

España (Español)

France (Français)

Greece (Ελληνικά)

India (English)

Indonesia (Bahasa Indonesia)

Ireland (English)

Italia (Italiano)

日本 (日本語)

대한민국 (한국어)

Latinoamérica (Español)

Malaysia (Melayu)

الشرق الأوسط وشمال أفريقيا (اللغة العربية)

MENA (English)

Nederlands (Nederland)

New Zealand (English)

Polska (Polski)

Portugal (Português)

România (Română)

Sverige (Svenska)

ประเทศไทย (ไทย)

Türkiye (Türkçe)

台灣 (中文)

Google has introduced two advanced dialogue models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, designed to enhance voice interactions by making them more natural, fluid, and intelligent, particularly in handling complex reasoning and parallel execution of tasks. Gemini 3.8 Live is characterized by its speed and fluidity, providing fast conversations supported by real-time visual and language context. It is integrated across platforms like the Gemini app, Google Workspace, and Search, enabling more intuitive and collaborative experiences for users when tackling complex tasks via voice commands.

The extended version, Gemini 3.8 Live Extended Thinking, focuses on high-complexity tasks through increased intelligence and sophisticated multi-step reasoning. This model is engineered to operate in the background, managing tools and executing tasks without interrupting the conversational flow, allowing the system to maintain a natural dialogue. This extended thinking capability allows the model to handle complex workflows and simultaneous reasoning while maintaining interactive communication, evidenced by the ability to naturally acknowledge prompts, such as through verbal cues like “Let me check that…”, and narrate the progress of multi-step background operations to the user.

Performance benchmarks highlight the advanced capabilities of these models. Gemini 3.8 Live Extended Thinking demonstrated superior performance in agentic task completion, achieving 68.6 percent success on the tau-Voice benchmark and 35.1 percent on Sierra’s tau-Voice-banking benchmark, while also exhibiting strong reasoning abilities, scoring 97.7 percent on Big Bench Audio. Furthermore, the models achieved a high Speech to Speech Quality Index of 82.6, indicating excellent conversational quality.

The functional integration of these models spans various applications. Gemini 3.8 Live processes visual inputs in near real-time to enrich conversations with contextual information, as demonstrated by real-time visual guidance for employee onboarding questions or visualizing complex processes like playing chess. Extended Thinking facilitates transformations of raw inputs, such as converting sketches and voice feedback into functional elements like React components, and manages complex operational tasks like coordinating asynchronous function calls and multi-step bookings concurrently with live conversation. These capabilities are being leveraged within Google Workspace, Search Live, and the Gemini app to provide real-time troubleshooting and dynamic content generation.

These advancements empower the developer and enterprise voice ecosystem through the Gemini Live API, which connects developers using platforms such as Agora, LangChain, and Vision Agents to build high-performance, voice-driven interfaces. Partnerships with entities like Salesforce and Genspark underscore the industry recognition of the models' impressive latency, fluidity, and tool-calling capabilities. To ensure accountability and safety, all audio generated by the AI products is watermarked with SynthID, an imperceptible watermark woven into the audio output to ensure AI-generated content is detectable and prevents misinformation. The rollout is being implemented across various access points, including the Gemini API, Google AI Studio for developers, private previews within Gemini Enterprise, and general availability in Search Live and the Gemini app for end-users and subscribers.