Google's Gemini 3.8 Breaks New Ground with Real-Time Reasoning Narration
Google's latest AI model, Gemini 3.8 Live Extended Thinking, marks a significant shift in transparency by enabling real-time narration of its reasoning process.
This capability allows users to hear the model think through complex workflows without interrupting conversational flow, thanks to early verbal cues and asynchronous function calling.
The model ranks first on the Artificial Analysis Speech-to-Speech Quality Index at 82.6, scores 97.7% on Big Bench Audio, and leads agentic task completion at 68.6% on τ-Voice.
Pricing sits at $0.005 per minute for audio input and $0.018 per minute for audio output, making it competitive enough for developers to build production voice agents at scale.
The shift toward visible inference serves as a new mechanism for error detection and user accountability, moving away from the 'take it or leave it' nature of traditional LLM interactions.