Seed Audio 1.0
by ByteDance
ByteDance's all‑in‑one TTS model.
Voice, music, and sound effects in one generation.

Key features
Hear the range
A warm, measured documentary voice-over.
A hushed, tense line read, close and intimate.
A layered open-air market sound bed.
A rolling storm building to a distant thunderclap.
A short rising cue for strings and brass.
A relaxed beat with soft keys and vinyl crackle.
Technical specifications
ByteDance
Developed by ByteDance's Seed research team.
20
English, Chinese, Japanese, Korean, Spanish, French, German and more.
Up to 2 min
Maximum two minutes of generated audio per pass.
3,000 chars
Room for a full scene brief, dialogue included.
Up to 3
Up to three reference clips, each up to 30 seconds.
Non-streaming
Renders the complete track, not a realtime stream.
Use cases
Audiobooks
Narration, character voices, and sound design for a full book. ByteDance puts the cost near a tenth of studio recording.
Video dubbing
Describe the voice or upload a character image, then use timestamps to land each line exactly where the picture needs it.
Game audio
Character barks, scripted performances, and environmental sound effects for immersive scenes, generated from the script.
One-pass video audio
Give a video clip its narration, sound design, and score in one generation, with no separate mixing step afterward.
Ads and promos
A spoken line, sound effects, and music as one ready-to-use track, made for short-form content.
Dialogue and audio drama
Multiple characters, each with a distinct voice and delivery, in one scene with matching ambience and timing.
Prompt examples
Timestamp control
Ryan (warm, breathless): '[5.5s:8.0s] Maya! Wait, you're leaving tonight?'
Edit promptAudiobook scene
Rain on a library window. Narrator, low and unhurried: 'She read it twice.'
Edit promptSports commentary
Packed stadium, roaring crowd. Commentator, exhilarated: 'OH, HE SCORES!'
Edit promptSimple pricing
Get started for free today, with the option to upgrade or cancel anytime.
Basic
900 monthly credits
1 user only
All models
Workflows
Standard
3200 monthly credits
1 user only
All models
Workflows
Pro
6200 shared monthly credits
1 user
All models
Workflows
Pro Max
24000 shared monthly credits
1 user
All models
Workflows
Enterprise
For higher limits
Custom
pricing and billing terms

Free
For playing around
$0
forever free
FAQs
Seedance 2.5
ByteDance
ByteDance's next-generation video model. Up to 30s native clips, 50 references, native audio.
Kling 4.0
Kling
Kuaishou's next Kling flagship, expected soon. Longer clips and sharper detail are on the way.
MiniMax H3
MiniMax
MiniMax's multimodal video model. 2K, native stereo audio, clips up to 15s.
Flux 3
Black Forest Labs
Black Forest Labs' world model for image, video, and audio, with sound that matches the action.