Skip to this page
THE SUPER INTELLIGENCE TIMESBY OPCELERATE NEURAL RSS
← Back to The Super Intelligence Times
Models Desk / Gemini 3.8 Live / Source-backed briefing / 2026-09-17
← Back to The Super Intelligence Times
The Super Intelligence Times
Source Notes Desk
Sapphire and amber light ribbons linking studio headphones to a glass thinking lattice.
Editorial illustration · AI-generated
Models Desk / Google / Product Release Sep 15 / Briefing Sep 17

Gemini 3.8 Live Is $0.005 Per Minute In. Extended Thinking Speaks While It Reasons.

Google’s Sep 15 launch (updated Sep 17) puts Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking in the Live API. Audio input is $0.005/min; audio output is $0.018/min. Extended Thinking reasons and speaks at the same time.

Quick answerGemini 3.8 Live and Gemini 3.8 Live Extended Thinking shipped September 15, 2026 (Google blog also notes Updated September 17, 2026). Live API audio is $0.005/min input and $0.018/min output (Google’s estimate from $3/1M tokens in and $12/1M tokens out). Extended Thinking reasons and speaks simultaneously, with cues like “Let me check that…”. Google-reported Extended Thinking leads Artificial Analysis Speech to Speech Quality Index at 82.6. For Alberta AI chatbot lead generation cost context, GPT-Live-1 is $0.05/min on our Live-1 briefing—about 10× Gemini’s input audio rate on a same-currency USD compare. USD as published; CAD on the invoice. Foreign API ≠ private Alberta path.

On September 15, 2026, Google introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking—near real-time voice models for voice agents and fluid dialogue. The companion developer post prices Live API audio at $0.005/min input and $0.018/min output. Those figures are USD as published. CAD is on the invoice. This desk is not inventing a Canadian list price.

AI chatbot lead generationAI automation Albertaprivate AI securityGemini 3.8 Live
ModelsGemini 3.8 Live + Gemini 3.8 Live Extended Thinking. Live API + AI Studio.
Audio in / out$0.005/min in · $0.018/min out (token-estimate footnote).
Extended ThinkingReasons and speaks simultaneously; live progress narration.
Languages / watermark97 languages (Live); SynthID on AI audio.

What Google listed for Gemini 3.8 Live

Google frames Gemini 3.8 Live as built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. It automatically detects and transitions between 97 supported languages mid-conversation, and executes tools and API calls in the background while continuing the conversation. The developer post adds asynchronous function calling, visual context, alphanumeric precision, 97+ languages, and incremental content updates.

Gemini 3.8 Live Extended Thinking is framed for high-complexity, multi-step reasoning. Google says it reasons and speaks simultaneously, uses early verbal cues like “Let me check that…” to acknowledge prompts, and uses live progress narration for multi-step background tasks. Configurable thinking is noted on the developer post for complex multi-step work while narrating progress in the main conversation.

Google-reported benchmarks (vendor figures, not this desk’s independent run): Extended Thinking #1 on Artificial Analysis Speech to Speech Quality Index at 82.6; 68.6% on τ-Voice; 35.1% on Sierra’s τ-Voice-banking; 97.7% on Big Bench Audio. Live is described as second place in the Speech Agent Arena. Audio generated by Google’s AI products is watermarked with SynthID.

“For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like “Let me check that…” to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress.”Google blog: Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking, September 15, 2026 (Updated September 17, 2026)

What it costs

These are Google’s published USD Live API audio rates from the developer post. Do not quote them as CAD. CAD is on the invoice. The footnote states the per-minute figures are an estimate based on $3/1M tokens for input and $12/1M tokens for output.

USD Live API audio pricing from Google developer post. CAD on the invoice.
LineRateBillingNotes
Audio input$0.005 / minLive APIEstimate from $3/1M tokens in
Audio output$0.018 / minLive APIEstimate from $12/1M tokens out
Partners——Agora, Fishjam, LangChain, LiveKit, Pipecat, Vercel, Vision Agents
Compare (our desk)$0.05 / minGPT-Live-1OpenAI voice layer from our Live-1 briefing

Rollout Google published

Google’s Sep 15 post (updated Sep 17) lists rollout as follows — exact surface bullets, not invented:

3.8 Live

  • For developers: In the Gemini API and Google AI Studio
  • For enterprises: In private preview in Gemini Enterprise and coming soon to Gemini Enterprise for Customer Experience
  • For everyone: In Search Live

3.8 Live Extended Thinking

  • For developers: In the Gemini API and Google AI Studio
  • For enterprises: In private preview in Gemini Enterprise and coming soon to Gemini Enterprise for Customer Experience and Google Workspace business customers
  • For everyone: In Gemini Live and for Google AI Pro and Ultra subscribers in Workspace in Docs, and all Google AI subscribers in Gmail and Keep

What Alberta operators should do

Alberta work here means AI chatbot lead generation and AI automation Alberta—phone-line receptionist, missed-call capture, and booking flows for shops that lose leads when nobody picks up. Gemini 3.8 Live’s Live API pricing ($0.005/min audio in, $0.018/min audio out) is the operator cost card for a voice front door that can keep chatting while tools run in the background. Extended Thinking is the path when the call needs multi-step reasoning without dropping the conversation.

For same-currency USD cost context only, our GPT-Live-1 briefing lists OpenAI’s voice layer at $0.05/min—about ten times Gemini’s published audio-input per-minute rate. That is not a feature bake-off; it is operator budget math. Foreign hosted API is not a private Alberta path. For regulated data—client health notes, tenders, credentials—prefer a private AI security stack with a human review gate before anything leaves the shop. If the workflow has to stay on the desk with clear ownership, that is AI consulting work, not a public paste into a third-country endpoint. Quote USD as published; expect CAD on the invoice.

The guardrail

Price is not permission. A $0.005-per-minute audio-input card does not put inference in Alberta, and it does not authorize shipping regulated client audio offshore without a written data boundary. For AI chatbot lead generation, use Gemini 3.8 Live when the job fits a real-time voice agent with background tools, use Extended Thinking when the call needs live reasoning narration, keep a human in the loop on bookings and care-adjacent intents, and keep sensitive workloads on a private agents path. SynthID watermarking is Google’s transparency mark on AI audio—not a substitute for your own consent and retention policy.

Opcelerate recommendationTreat Gemini 3.8 Live as a cost-efficient Live API voice front door for Alberta AI chatbot lead generation—$0.005/min audio in, $0.018/min audio out—and Gemini 3.8 Live Extended Thinking when the agent must reason and speak at once. USD as published; CAD on the invoice. Compare budgets to GPT-Live-1 at $0.05/min from our Live-1 briefing, not as a feature winner. Foreign API ≠ private/on-prem; put a human review gate on bookings and regulated intents, and prefer private agents for client data.