Google releases Gemini 3.8 Live and 3.5 Transcribe for real-time voice apps
Original titleBuild real-time voice applications with Gemini 3.8 Live and 3.5 Transcribe
Google AI Studio released Gemini 3.8 Live, a native speech-to-speech model with an Extended Thinking variant, and made it available through the Live API. Gemini 3.5 Transcribe, released last month, supports 85+ languages with a reported 4.0% streaming and 2.6% non-streaming Word Error Rate, and accepts a custom vocabulary of up to 1,000 terms.
Live API audio pricing is listed at $0.005/min for input and $0.018/min for output.
The post lists concrete Live API capabilities, per-minute audio pricing, and transcription accuracy figures, helping developers weigh voice agent options against their own cascaded pipelines.
Source: Google AI Studio · x.comPublished · added here