arXiv Machine Learning By Jiawei Yu, Jian Liu

TrueMuse: A Benchmark for Data Attribution in Text-to-Music Models

Read the original on arXiv Machine Learning →

TrueMuse is a new benchmark designed to evaluate data attribution in text-to-music models. It consists of a controlled dataset created by fine‑tuning three diffusion‑based models on curated attribution samples, providing known attribution targets. The benchmark covers four settings—melodic structure, timbral characteristics, artist‑level style, and genre‑level patterns—across 133 attributes, 648 models, and 95,456 generated samples, and is used to assess existing black‑box attribution methods along several dimensions.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 2

On the Human and Computer Alignment of Attribute-Based Music Matches

The paper introduces the MATCHA dataset, comprising 1,105 perceptual assessments from 83 experts on attribute-based music matches across five musical attributes—melody, harmony, rhythm, voice, and timbre. A triplet-based forced-choice experiment with 300 cases, including plagiarism, cover songs, and AI-generated music, was used to gather these judgments. Results show measurable agreement among participants and partial alignment with computational similarity measures, highlighting the need for perceptually grounded evaluation in generative AI for music.

By Roser Batlle-Roca, Woosung Choi, Joan Serr\`a, Fabio Morreale, Wei-Hsiang Liao, Xavier Serra, Emilia G\'omez, Yuki Mitsufuji
arXiv Machine Learning
1d ago

CHORDONOMICON: A Dataset of 666,000 Songs and their Chord Progressions

Chordonomicon is a new dataset of over 666,000 song-level symbolic chord progressions, each annotated with structural parts such as verse, chorus, and bridge, as well as genre and release date. The dataset was compiled by scraping user-generated progressions from multiple sources and shows strong similarity to established prior datasets. The authors also provide a reproducible benchmark suite for next chord prediction, evaluating RNN, GRU, and LSTM models across various context windows and data scales, and find that structural part annotations consistently improve prediction performance.

By Spyridon Kantarelis, Ioannis Liolitsas, Konstantinos Thomas, Vassilis Lyberatos, Edmund Dervakos, Giorgos Stamou