← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

YuE2, MERT2, and SheetSage2 models achieve new benchmarks in AI music generation and analysis

🔄 Updated 32m ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • YuE2 scored 6.9632 on SongBench, outperforming Suno v5.
  • MERT2-30s and MERT2-FS achieved SOTA on 14 of 15 MARBLE metrics.
  • SheetSage2 leads on 10 of 13 music transcription benchmarks.
  • These models advance AI capabilities in music creation and analysis.

YuE2 Sets New Music Generation Benchmark

The YuE2 model (best-of-8) achieved a score of 6.9632 on SongBench, a benchmark for music generation. This result represents the highest observed mean among 15 evaluated settings on WildSongBench, surpassing Suno v5, which scored 6.8721 in the same comparison. YuE2's performance indicates progress in generating musical compositions.

MERT2 Achieves State-of-the-Art in Audio Understanding

MERT2-30s and MERT2-FS (full-song) models have established new state-of-the-art results across 14 of 15 MARBLE metrics. These metrics cover various aspects of audio understanding, including tagging, key, genre, and emotion recognition. MERT2-30s achieved 91.72 in genre accuracy on GTZAN, and MERT2-FS scored 67.05 in key refined accuracy on GiantSteps, demonstrating advancements in how AI interprets musical content.

SheetSage2 Leads in Music Transcription

SheetSage2 has achieved state-of-the-art performance on 10 of 13 benchmark metrics for music transcription. This single model handles six transcription tasks: beat, downbeat, key, chord, structure, and melody. Notable scores include 82.51 for vocal melody pitch-class note F1 on RWC-Pop and 90.08 for chord recognition (Maj/min) on osu2017, indicating improved accuracy in converting audio to symbolic musical notation.

Impact on AI Music Development

These benchmark achievements across YuE2, MERT2, and SheetSage2 signify progress in AI's ability to generate, understand, and transcribe music. The results contribute to the development of more sophisticated AI tools for music creation, analysis, and research, potentially influencing future applications in the music technology sector.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~16 min · 14 stories · Sep 10

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

New AI models, YuE2, MERT2, and SheetSage2, have set new performance benchmarks in music generation, audio understanding, and transcription tasks. YuE2 achieved the highest score on SongBench for music generation, while MERT2 models reached state-of-the-art results across 14 of 15 MARBLE metrics for audio recognition, and SheetSage2 leads in 10 of 13 music transcription benchmarks.