Seed Audio 1.5
de ByteDance
ByteDance's full‑scene audio model.
Dialogue, music and effects, as separate tracks.

Funciones clave
Escucha el rango

Two voices, a corridor hum and a tense pad.

A warning, dripping water, then the gate opens.

A short jingle that ducks under the host.

Japanese narration over a morning market.
Especificaciones técnicas
6 min
Per generation, up from 2 minutes on Seed Audio 1.0
6
Reference clips per generation, up from 3
30
Including regional variants of Spanish and Portuguese
4
Text alone, or text with voice references, video, or both
Casos de uso

Dubbing and localization
Feed in a finished cut and get a dub that follows the picture, with the original effects and music kept.

Short drama and series
Six minutes covers a full scene in one generation, and separate tracks let you revise a line without new renders.

Audiobooks and audio drama
Narration, character voices and room sound together, with the same cast held across a long chapter.

Ads and promos
Voice-over, music swell and sound effects in one pass. When the client wants changes, regenerate only that layer.

Game scenes
Character lines, environment beds and timed effects for cutscenes and trailers, cued to the second.

Podcasts
An intro jingle, two hosts and a room that sounds real, with each voice on its own track for the edit.
Ejemplos de prompts

Rainy café reunion
Rainy café, soft piano. Lena, nervous: 'You kept the ticket?' Theo laughs.

Timed stadium call
[0:00] Crowd builds. [0:03] Commentator: 'He scores!' [0:06] Air horn.

Trailer narration
Low strings, cold wind. Deep narrator: 'Some doors should stay closed.' Thunder.

Coffee ad
Pop beat, coffee pours. Bright voice: 'Fresh coffee, at your door by eight.'

Audiobook opening
Fire crackles, wind outside. Narrator, unhurried: 'One ship was missing.'

Spanish market
Busy market. Vendor, in Spanish: '¡Naranjas frescas, a un euro el kilo!'
Overview
Seed Audio 1.5 is ByteDance's next full-scene audio model and the successor to Seed Audio 1.0. It generates dialogue, music, ambience and sound effects from one prompt, runs up to six minutes, and takes video, voice references and timestamps as direction. ByteDance's partner listings describe it as coming soon.
What Seed Audio 1.5 does differently
The scene comes back in layers. Dialogue, ambience, effects and music arrive as separate tracks, so a wrong line or a loud cue gets fixed on its own. Video input means the sound follows a finished cut, which is what dubbing and translation need.
Seed Audio 1.5 on Morphic
Seed Audio 1.5 is coming to Morphic soon. Seed Audio 1.0 is live now for full scenes, alongside ElevenLabs, Lyria and the rest of the audio models, so you can build the scene on the Canvas today. For prompt structure, see the Seed Audio 1.5 prompt guide, or start from a sample scene in the Seed Audio 1.5 AI audio generator.
Precios simples
Comienza gratis hoy, con la opción de mejorar tu plan o cancelar cuando quieras.
Basic
2400 mensual créditos
1 usuario
Todos los modelos
Workflows
Standard
3625 mensual créditos
1 usuario
Todos los modelos
Workflows
Pro
6350 compartidos mensual créditos
1 usuario
Todos los modelos
Workflows
Max
24650 compartidos mensual créditos
1 usuario
Todos los modelos
Workflows
Enterprise
Para límites más altos
Personalizado
términos de precios y facturación

Free
Para experimentar
$0
gratis para siempre
Preguntas frecuentes
Flux 3 Image
Black Forest Labs
El modelo de imagen de Black Forest Labs que prioriza el control. Ubica cada elemento y edita un cuadro a la vez.
Ideogram 4.5
Ideogram
El modelo de edición precisa de Ideogram. Edita ronda tras ronda sin que la imagen se estropee.
ChatGPT Images 2.5
OpenAI
El modelo de imagen de OpenAI, de septiembre de 2026. Más detalle, ediciones precisas, más rápido.
Lyria 3.5
Google DeepMind
El mejor modelo musical de Google. Canciones con estructura, voz y letra.