Volcano Engine Launches Audio Creation Model "Doubao Audio Generation Model 1.0"
Published · Jun 24 · Wed Source · 火山引擎 (CN)

Volcano Engine Launches Audio Creation Model "Doubao Audio Generation Model 1.0"

Volcano Engine has launched the Doubao Audio Generation Model 1.0, which supports text or audio reference input for the first time and can generate target audio end-to-end. The model can orchestrate multi-character dialogue, emotional tone, background music, and ambient atmosphere within a single Prompt, directly producing complete audio works while maintaining voice consistency during long-duration generation.

KeywordsVolcanoEngineLaunchesAudioCreationModelDoubaoGeneration

Volcano Engine has launched the Doubao Audio Generation Model 1.0, which supports text or audio reference input for the first time and can generate target audio end-to-end. The model can orchestrate multi-character dialogue, emotional tone, background music, and ambient atmosphere within a single Prompt, directly producing complete audio works while maintaining voice consistency during long-duration generation.

This page provides an editorial summary based on publicly available information. It is not a republished article. Use the source link below for the original report.