Describe the sound scene
Start with characters, emotion, location, timing, dialogue, musical mood, and the effects that should exist in the scene.
Seed Audio 1.0 is ByteDance Seed's all-in-one audio generation model for creating complete sound scenes. Use text, image, or audio context to guide multi-speaker dialogue, emotional delivery, native accents, ambience, background music, and foley-style effects.
Seed Audio 1.0
Scene prompt preview

Prompt concept
Two speakers whisper in a rainy alley, tense strings underneath, distant traffic, footsteps, and a final metallic door slam.
Use the Seed Audio 1.0 workspace below to create sound scenes from a prompt, optional reference audio, or one reference image.
Input
Prompt-first audio generation with optional controls.
Additional Settings
Customize your input with more control.
History
Your recent Seed Audio 1.0 Preview generations.
Sign in to see your generation history.
How to use
To use Seed Audio 1.0 online, write a sound-scene prompt, optionally add an audio or image reference, choose output settings, then generate and review the audio in the SeedAudio.co workspace.

Describe the characters, language, emotion, location, dialogue, ambience, music direction, and sound events you want in the scene.

Use up to three audio references for voice or style direction, or one image reference to guide mood and scene context.

Pick voice behavior, output format, sample rate, speed, volume, pitch, and a max credit cap before running the generation.

Run a short draft first, listen for voice clarity and layer balance, then copy, download, or revise the result.
Prompt formula
Scene + speaker + emotion + language + ambience + BGM + sound effects + timing.
Straightforward workflow
Move from a sound idea to a complete scene direction: define characters, emotion, location, dialogue, music, ambience, and effects in one prompt.
Start with characters, emotion, location, timing, dialogue, musical mood, and the effects that should exist in the scene.
Seed Audio 1.0 is designed to synthesize dialogue, emotional tone, native accents, ambience, BGM, and distinct sound effects together.
Shape audio directions for short films, ads, podcasts, games, learning content, and other projects that need coherent sound scenes quickly.
Seed Audio 1.0 technology
Seed Audio 1.0 is positioned for complete audio scenes: multi-character dialogue, emotion, tone, accents, ambience beds, BGM, and foley in a single creative pass.
Multi-speaker
voice continuity for longer generated scenes
Text / image / audio
multimodal prompting for audio creation
Compose multiple sound layers at once instead of stitching voice, music, ambience, and effects in separate tools.
Guide tone, emotional delivery, dialect, and native-sounding accents while keeping recurring voices recognizable across contexts.
Generate environmental beds, background music, room tone, weather, crowds, or distant city texture alongside the dialogue.
Seed Audio 1.0 is built for longer-form sound scenes, including session-length generation suitable for dialogue, ambience, and music-backed sequences.
Creative possibilities
Seed Audio 1.0 is most interesting when a project needs more than narration: a complete acoustic scene with voices, mood, space, and events.
Draft dialogue, emotional beats, foley, ambience, and music for storyboards or pre-visualization.
Create campaign-ready sound directions for product demos, social clips, and localized ads.
Prototype ambient loops, character barks, UI sounds, and cinematic moments before a final audio pass.
Build scenario-based lessons, character conversations, and immersive explainers with spatial sound cues.
Multi-character delivery with emotional tone
Rain, traffic, rooms, crowds, and natural beds
Footsteps, impacts, doors, texture, and timing
Pricing
Subscribe for the best value, or buy credits when you need a flexible top-up. Every paid plan and credit pack includes Seed Audio 1.0 API access, with one shared credit balance across the web app and API.
Try Seed Audio 1.0 with free signup credits. Perfect for testing prompts and short scenes.
The best annual choice for creators with ongoing audio needs.
The strongest value for teams that expect heavy audio generation.
FAQs
Practical answers for creators who want to understand Seed Audio 1.0 and use it for AI audio generation.
Follow the model's capabilities, access status, and practical use cases for multimodal AI audio generation across dialogue, ambience, music, and sound effects.