PercifyAll models
  1. Playground
  2. /
  3. audio models
  4. /
  5. Gemini TTS
AI voice model

Gemini TTS

Natural speech with expressive delivery. Supports two speakers in one take.

audio outputfrom 4 credits2 runsrefunded on failure
Try Gemini TTSExplore more models
audio

What you'll need

  • Script — Supports [sigh], [laughing], [whispering], [short pause].

About Gemini TTS

Gemini TTS is an AI voice model you can run directly in the Percify playground — fill in the inputs and generate right in your browser, or call it programmatically through the Percify API. Each run returns a ready-to-download audio hosted on our CDN, and credits are only charged for successful results.

Run Gemini TTS now

More AI voice models

View the full catalog
XTTS-v2Coqui XTTS-v2: Multilingual Text To Speech Voice CloningChatterbox TurboThe fastest open source TTS model without sacrificing quality.Speech-02-HDText-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Optimized for high-fidelity applications like voiceovers and audiobooks.Speech-02-TurboText-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Designed for real-time applications with low latencyZonos 2Voice cloning + text-to-speech — clone a voice from a short sample and make it say anything, multilingual.ElevenLabs Eleven V3ElevenLabs eleven-v3 is a text-to-speech model available as a hosted endpoint; requests cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.ElevenLabs Multilingual V2ElevenLabs Multilingual V2 is a multilingual text-to-speech model; cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.ElevenLabs MusicElevenLabs Music generates original songs from text descriptions. Create instrumentals or full compositions with customizable duration. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.Lyria 3 ProGoogle Lyria 3 Pro generates high-quality music tracks from text prompts and optional image input.Mirelo SFX 1.6Mirelo SFX1.6 Text To Audio generates sound effects or ambient audio directly from a text prompt, with optional seamless ambience looping.Mureka V9 (Song)mureka ai / mureka v9 / generate song via Mureka official API.Music 2.6Music 2.6 generates complete songs with vocals and instrumentals from text prompts and lyrics.

Ready to generate?

Open Gemini TTS in the playground and run it now.

Try it now