Close Menu
NCIJ Network NCIJ Network
    What's Hot

    BlackRock sees a new $5 trillion AI trade emerging for stablecoins

    September 23, 2026

    US contractors reportedly purchase tin from company accused of illegal Amazon mining

    September 23, 2026

    Picking a side on Ed Sheeran, Israel and Palestine | Ed Sheeran

    September 23, 2026
    Facebook X (Twitter) Instagram
    Trending
    • BlackRock sees a new $5 trillion AI trade emerging for stablecoins
    • US contractors reportedly purchase tin from company accused of illegal Amazon mining
    • Picking a side on Ed Sheeran, Israel and Palestine | Ed Sheeran
    • Did Denmark play ‘Send in the Clowns’ when Rubio arrived at Greenland deal ceremony?
    • No, this video doesn’t disprove the 9/11 attack on the south tower of the World Trade Center
    • A huit mois de la présidentielle, les comptes de campagne déjà sous surveillance – POLITICO
    • Burnham brands ban on football fans drinking alcohol in stands as ‘discrimination’
    • Tories would prevent long-term jobless spending benefits on alcohol and cigarettes
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Wednesday, September 23
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Artificial Intelligence

    Google Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKSeptember 23, 2026 Artificial Intelligence No Comments6 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Google has released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, 2 new text-to-speech models in its Gemini Audio family. Google calls them its most expressive audio generation models yet. Flash TTS targets creative direction and character voices. Flash-Lite TTS targets high-volume, cost-efficient production. Both let developers direct delivery line by line using natural language.

    Is it deployable? Yes, both models are rolling out now through the Gemini API and Google AI Studio. Access is API-only, with no open weights for self-hosting. Enterprise API access via Gemini Enterprise is listed as coming soon.

    What Google Shipped

    The release splits TTS into 2 tiers with shared direction controls:

    • Gemini 3.8 Flash TTS is built for deep creative direction and character design. Target uses include gaming, immersive audiobooks, podcasts and interactive media. It offers granular control over acting cues, pacing, dialect shifts and backchanneling.
    • Gemini 3.8 Flash-Lite TTS is built for high-volume, cost-efficient scale. Google positions it for dubbing, audio content creation and expressive voice agents. It offers fine-grained control over tone, pacing and expressive nuance.

    In AI Studio, the playground links use the model identifiers gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts.

    Voice Design From a Text Prompt

    Previous Gemini TTS offered 30 original voices. The 3.8 release moves to a much larger voice system.

    • Generative voice design: Flash TTS creates new voices from prompts describing role, accent and voice characteristics. This works across more than 100 languages and dialects. Google’s demos include a Melbourne DJ, a monotone robot and a Japanese dragon.
    • Voice library: Developers get 2,000+ production-ready voices. Coverage includes regional varieties like Mexican Spanish, Quebec French and Scots English.
    • Save and scale: Custom voices can be saved and reused, with minimal drift across projects.
    • Voice remixing (coming soon): Users will adjust a library voice’s timbre, pitch, pace and accent through prompts.

    Directing the Performance

    Both models accept stage directions written in the script. Gemini can also steer delivery from natural script cues.

    • Long-form generation: Voice quality, pacing and timbre hold across hours of continuous audio.
    • Native 2-speaker staging: A single script drives a multi-turn conversation with distinct, separated voices.
    • Vocal bursts: Non-verbal cues like , and add conversational texture.
    • Backchanneling: Active-listening interjections like |mhm| and |yeah| control reaction beats and comedic timing.

    Voice Replication and Safety Controls

    Voice replication builds a consistent vocal profile from a 30-second audio sample. The sample must be your voice or one you have rights to use. Replication requires a verbal consent recording from the voice owner, matched against the reference speaker.

    Every clip from Gemini Audio models carries a SynthID watermark. This imperceptible mark is embedded directly in the audio output. Replicated voices also carry C2PA content credentials. Google points to the Gemini 3.8 Audio model card for its broader safety approach.

    Benchmark Results

    Google reports these results for the new models:

    • Hume AI Voice Design Benchmark: Flash TTS ranks #1 overall with a score of 71.4, per Hume AI.
    • Accent modeling: Flash TTS leads with a score of 60.8.
    • Hume AI Overall Quality Index: Flash TTS ranks #1 and Flash-Lite TTS ranks #2.
    • Voice Arena blind preference: Both models take top positions in Japanese, Brazilian Portuguese, Vietnamese, Modern Standard Arabic, Mexican Spanish and Hindi.

    Key Takeaways

    • Google launched Gemini 3.8 Flash TTS for creative work and Flash-Lite TTS for scale.
    • Flash TTS designs new voices from prompts across 100+ languages and dialects.
    • Developers get 2,000+ production voices, up from 30 originals.
    • Voice replication needs a 30-second sample plus a matching consent recording.
    • Flash TTS ranks #1 on Hume AI’s Voice Design Benchmark with 71.4.

    FAQ

    1. What is Gemini 3.8 Flash TTS? It is Google’s text-to-speech model for creative voice design and line-by-line performance direction. It is available through the Gemini API and Google AI Studio.
    2. How is Flash-Lite TTS different? Flash-Lite TTS is optimized for high-volume, cost-efficient workloads like dubbing and voice agents. It ranks #2 on Hume AI’s Overall Quality Index.
    3. Can I clone my own voice? Yes, with a 30-second sample and a verbal consent recording. It is unavailable in AI Studio in several regions, including the UK, EEA and India.

    Check out the Technical Blog. All credit goes to the researcher of this project. Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

    Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.? Connect with us


    Asif Razzaq is the CEO of Marktechpost AI Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.

    Design Flash FlashLite Gemini Google PromptBased Releases TTS voice
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time

    U.S. TRANSCOM deploys randomised AI to secure military logistics

    SpeakON Ships a MagSafe AI Voice Button With Its Own Microphone

    Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning

    Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model

    AI Agents Are Becoming a New Malware Distribution Channel

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    BlackRock sees a new $5 trillion AI trade emerging for stablecoins

    September 23, 2026

    US contractors reportedly purchase tin from company accused of illegal Amazon mining

    September 23, 2026

    Picking a side on Ed Sheeran, Israel and Palestine | Ed Sheeran

    September 23, 2026

    Did Denmark play ‘Send in the Clowns’ when Rubio arrived at Greenland deal ceremony?

    September 23, 2026
    Latest Posts

    Ransom Cartel ransomware creator sentenced to 16 years in prison

    August 5, 2026

    Uber CEO brushes off reports of a Waymo break-up

    August 5, 2026

    Fauci Faces Contempt Vote. Here Are the Legal Issues Involved.

    August 6, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    BlackRock sees a new $5 trillion AI trade emerging for stablecoins

    September 23, 2026

    US contractors reportedly purchase tin from company accused of illegal Amazon mining

    September 23, 2026

    Picking a side on Ed Sheeran, Israel and Palestine | Ed Sheeran

    September 23, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.