Speech 2.8 Turbo
Fast, natural speech across 40 languages, with expressive sound tags and streaming output.
Paste a script, pick a voice, and generate.
Professional voiceover for every use case
Audiobooks and long-form reading.
Give narrative passages a steady cadence, expressive delivery, and enough vocal warmth for extended listening.E-learning and course modules.
Turn lessons into structured, approachable audio with controlled pacing for definitions, instructions, and key concepts.Explainer and product demos.
Deliver concise product explanations with clear pronunciation, deliberate pauses, and an energetic but controlled tone.IVR and phone prompts.
Generate consistent phone menus and service messages with measured timing and easy-to-follow options.Choose from 300+ voices or use your own
Speech 2.8 Turbo works with more than 300 system voices spanning narrators, presenters, conversational speakers, and character styles. It also supports voice cloning when you need an approved voice to remain consistent across a project.
- •300+ system voices
- •Narration and conversational styles
- •Custom cloned voice support
- •Voice design capability
- •Reusable voice IDs
Speak to audiences in 40 languages
The model supports 40 languages, including English, Spanish, French, German, Arabic, Japanese, Korean, Hindi, and Cantonese. Language enhancement helps the system interpret multilingual scripts and dialect-specific pronunciation more accurately.
- •40 supported languages
- •Automatic language detection
- •Cantonese and Mandarin support
- •Language-specific enhancement
- •Cross-language voice synthesis
Add emotion, breaths, and natural reactions
Seven emotion settings shape the delivery, while Speech 2.8 sound tags add audible details such as laughter, breaths, sighs, gasps, and throat clearing. These controls help dialogue and narration avoid a flat, uninterrupted reading style.
- •Seven emotion settings
- •Laugh and chuckle tags
- •Breaths, sighs, and gasps
- •Cough and throat-clearing tags
- •Natural hesitation and pacing
Fine-tune timing and production settings
Adjust speed, pitch, volume, pronunciation, and custom pauses to fit the destination of the voiceover. Streaming support and multiple audio formats make the model practical for both interactive playback and edited productions.
- •Adjustable speed, pitch, and volume
- •Custom pauses from 0.01 to 99.99 seconds
- •IPA, Pinyin, and Jyutping overrides
- •MP3, PCM, FLAC, and WAV output
- •Sample rates from 8 to 44.1 kHz
- •HTTP and WebSocket streaming
How it works
Paste your script
Enter the exact words you want the model to read. Use paragraph breaks for structure, then add supported pause or sound tags where the delivery needs more nuance.
Pick a voice
Choose a voice that fits your video, lesson, audiobook, advertisement, or phone system. If cloning is available, use only your own voice or one you have explicit permission to reproduce.
Generate and download
Generate the speech, listen for pacing and pronunciation, and adjust the settings if needed. Download the finished audio for editing, publishing, or integration into your project.
Pricing for Speech 2.8 Turbo
Runs on credits — no per-model surcharges, no surprise billing.