Google Deepmind has introduced two new audio models for developers, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, which currently lead the Artificial Analysis speech-to-speech leaderboard, according to The Decoder. These models showcase advanced capabilities in real-time voice conversation AI.

The Gemini 3.8 Live model is priced at $1.38 per hour of voice conversation, significantly cheaper than OpenAI's competing GPT-Live-1, The Decoder reported. This price point could make advanced speech-to-speech AI more accessible for developers and businesses.

For Japanese markets, where voice recognition and AI-driven communication tools are rapidly evolving, Google Deepmind's new offerings may accelerate adoption in areas such as customer service automation and real-time translation services.