Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingNewsWatching0 independent reports0

Qwen-Music Technical Report: Full-Song Generation with Melody-CoT

First seen · 7/20/2026, 12:00 PMLatest activity · 7/20/2026, 12:00 PM

The Qwen-Music technical report presents a music-generation system for text-to-music and cover-song generation with complete vocal singing. Its architecture combines a Qwen-Music-Tokenizer, a Qwen-Music-LLM, and a Qwen-Music-Render module. Audio is compressed into a 25 Hz single-codebook stream of semantic tokens, while Melody-CoT plans melody tokens before full-song generation. Training used more than 5 million hours of multilingual music spanning hundreds of languages, followed by supervised initialization, offline DPO, and online GSPO. On 600 Chinese and English prompts, the report says Qwen-Music led in 13 of 16 objective musicality and audio-quality metrics, with professional evaluators preferring it over leading proprietary systems.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/20, 12:00 PMnot independentRepresentative
    Qwen-Music Technical Report: Full-Song Generation with Melody-CoT