Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

@GoogleAIStudio
الإنجليزية15 سبتمبر 2026
121K
1.6K
127
61
339

ليرة تركية؛ د

Google introduces Gemini 3.8 Live and Extended Thinking, offering advanced voice AI with parallel reasoning, real-time visual context, and background task execution for developers and enterprises.

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make it more intuitive to collaborate and execute complex tasks using your voice.

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.

  • Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
  • Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.

For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.

Experience more fluid, intelligent conversations

Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.

Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena. In addition to this performance, it remains highly cost-effective—providing developers and enterprises with a capable and efficient model built for scale.

Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image

On ServiceNow’s EVA-Bench, a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality.

Google AI Studio - inline image

Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background.

Google AI Studio - inline image

Gemini 3.8 Live guides employee onboarding in real time, using visual context to answer live questions.

Google AI Studio - inline image

Watch Gemini 3.8 Live play chess in near real-time using visual context, reasoning, and natural conversational flow.

For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like "Let me check that..." to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress.

Google AI Studio - inline image

Watch Gemini 3.8 Live Extended Thinking transform raw sketches and real-time voice feedback into functional React components.

Google AI Studio - inline image

See Gemini 3.8 Live Extended Thinking coordinate multi-step bookings and asynchronous function calls — all without interrupting natural live conversation.

Google AI Studio - inline image

Watch Gemini 3.8 Live build complete business plans and custom marketing toolkits on the fly through natural speech.

Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks.

Google AI Studio - inline image

Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live.

Google AI Studio - inline image

Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live—right inside Search Live.

Empowering the developer and enterprise voice ecosystem

By using the Gemini Live API, developer platforms such as Agora, Fishjam, LiveKit, Pipecat, Vercel, and Vision Agents enable developers to build and deploy high-performance voice-driven interfaces with ease. These platforms manage complex real-time media streaming infrastructure behind the scenes, allowing developers to focus entirely on crafting the user experience.

We’re also partnering with companies like Salesforce, Genspark, and Lumeris who are excited about 3.8 Live and 3.8 Live Extended Thinking, highlighting its impressive latency, fluidity and tool-calling capabilities.

Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image
Google AI Studio - inline image

Ensure transparency with SynthID watermarking

All audio generated by our Al products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio output, ensuring Al-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the model card.

Start using our latest Gemini Audio models:

3.8 Live is rolling out starting today:

3.8 Live Extended Thinking is rolling out starting today:

بنقرة واحدة حفظ

استخدم YouMind للقراءة العميقة للمقالات سريعة الانتشار بتقنية الذكاء الاصطناعي

احفظ المصدر، واطرح أسئلة مركزة، ولخص الحجة، وحوّل المقالة واسعة الانتشار إلى ملاحظات قابلة لإعادة الاستخدام في مساحة عمل واحدة تعمل بالذكاء الاصطناعي.

اكتشف YouMind
للمبدعين

حول Markdown إلى مقالة 𝕏 نظيفة

عندما تنشر كتاباتك الطويلة، فإن الصور والجداول وكتل التعليمات البرمجية تجعل تنسيق 𝕏 مؤلمًا. YouMind يحول مسودة Markdown كاملة إلى مقالة نظيفة وجاهزة للنشر 𝕏.

حاول Markdown إلى 𝕏

المزيد من الأنماط لفك التشفير

المقالات الفيروسية الأخيرة

استكشاف المزيد من المقالات الفيروسية