Close Menu
    What's Hot

    MAI-Image-2.6 Text to Image Model Hits #2

    August 11

    Muse Glimmer: 30B Model that Can Run Locally

    August 10

    Prime Agent Self Improving RLM Harness for Coding

    August 6
    Facebook X (Twitter) Instagram
    • AI Robots
    • AI News
    • Text to Video AI Tools
    • ChatGPT
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Rad NeuronsRad Neurons
    • AI Robots
      • AI Coding
    • ChatGPT
    • Text to Video AI
    Subscribe
    Rad NeuronsRad Neurons
    Home » Gemini 2.5 Flash and Pro Text-to-Speech Updates Announced
    AI Audio

    Gemini 2.5 Flash and Pro Text-to-Speech Updates Announced

    AI NinjaBy AI NinjaDecember 113 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Google has already managed to take the lead with Gemini 3 Pro in many areas. It also has incredibly powerful TTS models. The latest updates can now give you more control over style, tone, pace, and accents. These models now offer context-aware speed adjustments and follow instructions better

    We’re launching Gemini 2.5 Flash and Pro Text-to-Speech (TTS) model updates 🚀

    Improvements include:

    – Emotional style and tone versatility
    – Context-aware pacing control
    – Improved multiple-speaker capabilities

    Dive into the blog to learn how these advancements are giving…

    — Google AI Developers (@googleaidevs) December 10, 2025

    Here is a sample prompt you can use to generate your own voice:

    ASMR Pro
    # AUDIO PROFILE: Willow T.
    ## "The ASMR Whisperer"
    
    ## The Scene: Recorded inside a converted Sprinter van parked near Burleigh Heads. The space is small and padded with tapestries and macramé, creating a very "dry" but warm acoustic environment. The microphone is a Neumann KU 100 Dummy Head (binaural), meaning the audio should pan slightly left and right as the character moves, simulating 3D space.
    
    ### DIRECTORS NOTES
    Style: Relaxed Gold Coast bohemian style ASMR content creator.
    Accent: Gold Coast, Australia
    The "Grounding" Breath: Deep, diaphragmatic exhales that sound like ocean waves. Not sharp, but long and audible releases of air.
    Wetness/Mouth Sounds: Essential for ASMR. The listener should hear the sticky, subtle sounds of the tongue moving against the roof of the mouth (the "clicks" and "smacks") between words.
    Prosody & Pacing: The "Drift": The tempo is incredibly slow and liquid. Words bleed into each other. There is zero urgency.
    The "Smile" filter: The voice must sound like the speaker is constantly smiling. This brightens the tone even when whispering.
    High Rising Terminal (Softened): The classic Australian upward inflection at the end of sentences, but slowed down. It shouldn't sound questioning, just open and inviting.
    Tone & Articulation:
    The Gold Coast Vowel Shift: "I" (as in "light") becomes a wide, slow "loit" or "lah-ee-t." "O" (as in "no") drifts into the classic Aussie "naur," but breathy and soft, not harsh. Sibilance: The 'S' sounds should be prominent but crisp, creating a high-frequency "tingle" trigger.
    Vocal Fry (The "Morning Voice"): A rumbly, relaxed texture in the lower register, sounding like they just woke up from a nap on the beach.

    You can try these new models in Google AI Studio. There is also a playground app for playing around with this.

    [HT]

    fTTS
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleCue Chef Cube O1 AI Cooking Gadget with Thermal Imager
    Next Article OpenAI Steps Up with GPT 5.2, Overtakes Claude Opus 4.5 in GDPval-AA
    AI Ninja
    • Website

    Related Posts

    AI Audio

    Lipsync-2-pro: Edit What Anyone Says In Any Video

    September 2
    AI Audio

    ElevenLabs Voice Design v3 Announced

    June 26
    AI Audio

    Nari Labs Dia Outperforms ElevenLabs, Sesame CSM-1B

    April 23
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    50 Sora 2 Prompts You Should Try

    October 1181 Views

    DeepSeek-V3.2-Exp with DeepSeek Sparse Attention(DSA) for Efficient Long-Context Handling

    September 297 Views

    GPT Image 1.5 Model Debuts, Lands at #1?

    December 175 Views
    More
    AI News

    MAI-Image-2.6 Text to Image Model Hits #2

    AI NinjaAugust 11
    AI News

    Muse Glimmer: 30B Model that Can Run Locally

    AI NinjaAugust 10
    AI News

    Prime Agent Self Improving RLM Harness for Coding

    AI NinjaAugust 6
    Most Popular

    Prompt Cannon: Run Prompts Across Multiple Models

    June 243,882 Views

    Dipal D1 2.5K Curved Screen 3D AI Character

    June 231,100 Views

    How to Use Claude in Unity & Unreal Engine with MCP

    March 191,041 Views
    Our Picks

    MAI-Image-2.6 Text to Image Model Hits #2

    August 11

    Muse Glimmer: 30B Model that Can Run Locally

    August 10

    Prime Agent Self Improving RLM Harness for Coding

    August 6
    Tags
    3D agent AI AI model ai video API app avatar ChatGPT Chess Claude Claude Code coding DeepSeek ElevenLabs Gemini glasses GPT Grok Hermes Higgsfield image kling leonardo LLM midjourney Minimax Mini PC model music nano banana o3 offline OpenClaw open source QWEN robot runway sora Veo 2 Veo 3 Vibe coding video video model Voice

    © 2026 Rad Neurons. Inspired by Entropy Grid
    • Home
    • Terms of Use
    • Privacy Policy
    • Disclaimer

    Type above and press Enter to search. Press Esc to cancel.