Seed Audio today announced support for Doubao Seed-Audio 1.0, the newly released multimodal audio generation model from ByteDance and Volcengine, inside its AI music creation workspace. The integration reflects a broader evolution in AI audio: moving beyond text-to-speech or single-track music generation toward full-scene audio creation, where dialogue, emotion, accents, background music, ambience, and sound effects are generated cohesively.
Doubao Seed-Audio 1.0 is described as a multimodal model that works with text and reference audio, focusing on end-to-end audio creation rather than isolated clips. This distinction matters for creators working on projects like podcast trailers, short dramas, or game teasers, which require multiple audio elements—narration, transition music, room tone, footsteps, and background scores—to be generated as a unified experience. Unlike traditional text-to-speech models that prioritize how words are spoken, Doubao Seed-Audio 1.0 addresses the broader sound of a scene, including voices, music, spatial texture, and timing.
This release has attracted attention beyond musicians. Video creators, marketers, podcast teams, game developers, educators, and social media editors all face the challenge of needing audio that fits a scene, not just a standalone file. Doubao Seed-Audio 1.0 arrives as creators demand more control after generation—refining a chorus, adjusting background music, or extending an intro—which are workflow problems as much as model problems.
Seed Audio positions its workspace to address these needs by placing generation inside an agent-based environment. At the center is the Seed Audio Agent, a guided creation tool that helps users translate plain-language goals—such as a cinematic game loop or podcast intro—into actionable steps. The platform allows users to draft, refine, extend, cover, remix, and organize audio assets without switching between tools.
For creators starting from scratch, the AI Music Generator creates complete songs, instrumental tracks, and hooks from text prompts. Those needing help before generation can use lyric and style assistance to turn a theme or emotion into structured lyrics and production notes. The platform also supports workflows starting from existing material, such as uploaded audio or saved tracks.
Additional tools include AI Cover for creating new vocal or style versions, Extend for lengthening tracks, Add Tracks for adding accompaniment to vocals or instrumentals, and Mashup for combining ideas. Replace Section allows targeted revision of weak parts, while Vocal Remover separates vocals and instrumentals for remixing. Users can explore public tracks via the Explore feature and manage previous generations in My Works.
Seed Audio is available now at https://seedaudio.ai. New users can test Doubao Seed-Audio 1.0-supported workflows, generate sample tracks, and use the platform's editing tools. For visual assets, i2v.ai offers AI image and video generation that pairs naturally with Seed Audio for short videos, social posts, and campaign assets.


