ByteDance researchers published a paper on August 3, 2026, describing a generative audio system that handles multi-speaker voice synthesis, environmental soundscapes, local sound effects, and music ...