Sections

Search

Google launches Gemini 3.8 Flash TTS and Flash-Lite voice models, claiming top spot on independent audio rankingsAnthropic says Claude found a new enzyme system in bacteriophage DNAOpenAI Introduces MentalHealthBench for Everyday and Crisis AI ConversationsAnthropic says health groups are using Claude in Congo Ebola responseStudy: AI experts lowballed progress speed
All stories

Models·

Google launches Gemini 3.8 Flash TTS and Flash-Lite voice models, claiming top spot on independent audio rankings

Google released Gemini 3.8 Flash TTS and Flash-Lite TTS voice models, with third-party rankings from Hume AI and Voice Arena putting the larger model first.

Google ships Gemini 3.8 Flash TTS and Flash-Lite TTS

Google released two voice models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, according to AI News. The larger model is aimed at interactive entertainment, game development and long-form narration; the Lite version targets automated dubbing, conversational agents and high-throughput translation. AI News reports the pair replaces Google's older fixed catalogue of 30 voices with a directory of more than 2,000 pre-built profiles spanning regional variants such as Quebec French, Scots English and Mexican Spanish across over 100 languages, and that a voice-remixing module for timbre, pitch, pace and accent via text commands is coming. [1]

Benchmarks and evaluations: Hume AI and Voice Arena

Per AI News, independent evaluations put Gemini 3.8 Flash TTS at the top of third-party audio rankings. On the Hume AI Voice Design Benchmark it scored 71.4 overall with a category-leading 60.8 in accent modelling; on the Hume AI Overall Quality Index it placed first, with Flash-Lite TTS second, both ahead of earlier Gemini 3.1 Flash TTS baselines. AI News also cites double-blind human trials through Voice Arena showing preference advantages in Japanese, Brazilian Portuguese, Vietnamese, Modern Standard Arabic, Mexican Spanish and Hindi. [1]

Safety controls

AI News reports that voice cloning requires a 30-second reference recording plus an explicit verbal consent track from the original speaker, with Google validating acoustic alignment before processing. Generated audio carries SynthID watermarks and C2PA provenance metadata, it adds. [1]

Availability and integrations

Both models are available through Google AI Studio and the standard Gemini API, per AI News, connecting to frameworks from Agora, LiveKit, Pipecat and Vercel. AI News lists early commercial integrations including Figma, HeyGen, Linguana, Wondercraft, 99.co and Ollang, with Flash TTS inside Gemini Notebook and Flash-Lite TTS in Google Vids. Administrative API access for Gemini Enterprise customers is described as an upcoming deployment wave. [1]

Sources

  1. AI News · Reporting ·
    Google launches Gemini 3.8 Flash TTS voice models