ByteDance has launched its audio creation model, Seed Audio 1.0, which is now available for testing at the Volcano Ark Experience Center. This model breaks through the limitations of traditional multi-model stitching and mixing by generating complete audio works end-to-end through a unified framework. Seed Audio 1.0 boasts three core capabilities: precise spatiotemporal arrangement, stable timbre rendering, and authentic multilingual support. It achieves an audio usability rate of over 90% across nine major creative scenarios, with excellent naturalness in multilingual generation. The model lowers the barrier to high-quality audio production, and the team plans to further optimize its capabilities by integrating multimodal inputs and exploring controllable translation technologies in the future.
