The authors introduce YuE2, a unified model that combines symbolic and audio music generation.
YuE2 unifies symbolic and audio music generation via explicit score planning
A single mixture-of-transformers model writes readable scores, expands them into semantic tokens, and generates full-song audio, outperforming prior baselines and rivaling proprietary systems.
Top University
Ruibin Yuan · Jiahao Pan · Junyan Jiang · Zhiyue Wu · Ziya Zhou · Jiankai Sun · +29 more
HKUST · M-A-P · Tokenwave.AI · New York University · Stanford University
Research Digest··2 min read
YuE2 unifies symbolic and audio music generation by first writing an explicit score, then expanding it into semantic tokens and synthesizing full-song audio.
Why this paper
From HKUST and 7 others
In one line
YuE2 unifies symbolic and audio music generation by first writing a readable score then producing full song audio.
What we could check
- ·No code link found
- ·No weights link found
- ·No dataset link found
- ·No compute details found
- ·No stated limitations found
- ✓Reports numbers on named benchmarks
Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.
§