ByteDance Launches Seed Audio 1.0, Generates Dialogue and Sound Effects from Single Prompt

According to Beating, ByteDance's Seed Audio 1.0 was launched on July 22, enabling single-prompt generation of dialogue, sound effects, and ambient audio with 100ms timing precision. The model supports text and reference audio inputs, can produce approximately two minutes of audio per generation with extended capability, and maintains consistent character voices across outputs.

The audio demonstrates over 90% usability in most scenarios, with Mean Opinion Score (MOS) exceeding 4 across most supported languages. Seed Audio 1.0 supports 20+ languages with cross-lingual voice transfer and tone customization. The model is now available through Volcano Ark experience center with API integration open.

Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments