Skip to content
-
  • https://www.facebook.com/
  • https://twitter.com/
  • https://t.me/
  • https://www.instagram.com/
  • https://youtube.com/
AI Feed AI Feed AI Feed

AI news, tools, comparisons and practical guides

Subscribe
AI Feed AI Feed AI Feed

AI news, tools, comparisons and practical guides

  • AI News
  • AI Tools Radar
  • AI Comparisons
  • About
  • AI API Prices
  • Local AI
  • AI & Jobs

Sections

  • AI Comparisons
  • AI Features
  • AI Guides
  • AI News
  • Uncategorized

Latest stories

  • Google Releases Gemini 3.8 Live for Real-Time Voice Apps
  • SberTech Describes an MCP Gateway for Enterprise AI Agents
  • T-Bank Publishes Perseus Guide for Four ML Tasks
  • Yandex Opens Alice AI-T5-35B-A0.6B for Fast Search Answers
  • T-Bank Details a VLM-Based AI Agent for Regression Testing
  • AI News
  • AI Tools Radar
  • AI Comparisons
  • About
  • AI API Prices
  • Local AI
  • AI & Jobs
Subscribe
Close

Search

Home/AI Features/Google Releases Gemini 3.8 Live for Real-Time Voice Apps
Иллюстрация к новости: Google выпустила Gemini 3.8 Live для голосовых ИИ-приложений
AI FeaturesAI News

Google Releases Gemini 3.8 Live for Real-Time Voice Apps

Alex
By Alex
15.09.2026 2 Min Read
◉1unique readers

Google released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking for developers on September 15. Available through the Live API in the Gemini API and Google AI Studio, the models are intended for applications that maintain real-time conversations while reasoning and carrying out tasks.

Both models support asynchronous function calling, allowing an agent to make API and tool calls in the background while continuing to stream audio responses. Google also lists live visual context, precise handling of alphanumeric information, incremental merging of audio with structured data, and coverage for more than 97 languages.

Extended Thinking adds a configurable reasoning mode for complex, multi-step requests. According to Google, it can process such work in the background while responding to the user or narrating its progress in the main conversation. Google says the model ranks first on Artificial Analysis’ Speech-to-Speech leaderboard, but its announcement provides no testing details or independent reproduction results.

Google estimates pricing for Gemini 3.8 Live and Extended Thinking at $0.005 per minute of audio input and $0.018 per minute of audio output. A footnote says this estimate is based on rates of $3 per million input tokens and $12 per million output tokens. Access is also offered through integration partners Agora, Fishjam, LiveKit, LangChain, Pipecat, Vercel and Vision Agents.

The audio lineup also includes Gemini 3.5 Transcribe, released a month earlier for converting streamed speech into text. Google reports an average word error rate of 4.0% in streaming mode and 2.6% in non-streaming use. The model supports more than 85 languages, automatically handles language switching and accepts a custom vocabulary list containing up to 1,000 terms.

Smart Transcription mode produces structured, reader-ready text, accounts for speakers’ self-corrections and removes filler words. Through the Interactions API, the model can transcribe audio files up to one hour long with structured timestamps and speaker labels. The error-rate figures and broader performance claims come from Google; the announcement does not identify the evaluation datasets or methodology.

Practical context: In practical terms, the lineup provides developers with two distinct building blocks: Gemini 3.8 Live for two-way voice agents that use actions and context, and Gemini 3.5 Transcribe for captioning, call analysis and other text-only tasks. Asynchronous function calls are particularly relevant when external services are involved because the conversation need not stop while a request runs, although actual latency will also depend on the connected infrastructure.

Gemini 3.8 Live access and stated pricing
Parameter Value
Access channel Live API in the Gemini API and Google AI Studio
Audio input $0.005 per minute
Audio output $0.018 per minute
Integrations Agora, Fishjam, LiveKit, LangChain, Pipecat, Vercel, Vision Agents

Sources

  1. Google Blog

Event date: 2026-09-15. Primary source date: 2026-09-15.

Follow AI Feed on Telegram

New AI stories, practical guides and tool comparisons — in one concise feed.

Open Telegram→

Related reading

  • Google launches Gemini 3.8 Flash and restricted Cyber model
  • Google introduces WeatherNext 3 with hourly forecast updates
  • Google Announces Lyria 3.5 Availability in Gemini and API

Tags:

Editor’s Picks
Alex
Author

Alex

Follow Me
Other Articles
Иллюстрация к новости: СберТех описал MCP Gateway для корпоративных ИИ-агентов
Previous

SberTech Describes an MCP Gateway for Enterprise AI Agents

Recent posts

  • Google Releases Gemini 3.8 Live for Real-Time Voice Apps
  • SberTech Describes an MCP Gateway for Enterprise AI Agents
  • T-Bank Publishes Perseus Guide for Four ML Tasks
  • Yandex Opens Alice AI-T5-35B-A0.6B for Fast Search Answers
  • T-Bank Details a VLM-Based AI Agent for Regression Testing

Recent comments

No comments to show.

Archives

  • September 2026
  • May 2026

Sections

  • AI Comparisons
  • AI Features
  • AI Guides
  • AI News
  • Uncategorized

    © 2026 AI Feed. All rights reserved.
    RUEN
    AboutEditorial PolicySources & methodologyCorrectionsContactPrivacyAnalytics settings
    AI Feed analytics

    Helps us understand which pages are useful. Advertising tracking is disabled.