Project Qualia investigates whether experiential similarity between songs can be extracted from listening behavior. Using 1.29 billion scrobbles from 9,396 users, the authors trained a Word2Vec model (Song2Vec) on session data, then applied an artist‑residual procedure to isolate artist‑independent signals. The residual embeddings still contained strong cross‑artist similarity, forming coherent genre and era clusters, demonstrating that experiential structure exists beyond artist identity.
By Nizam Mohammed, Abu B. S. Rahman, Dimuthu D. K. Arachchige
arXiv:2607. 03296v1 Announce Type: cross Abstract: Crossmodal correspondences between sound and taste are well established in psychology and neuroscience, but largely absent from content-based multimedia retrieval.
By Matteo Spanio, Antonio Rod\`a
The paper introduces the MATCHA dataset, comprising 1,105 perceptual assessments from 83 experts on attribute-based music matches across five musical attributes—melody, harmony, rhythm, voice, and timbre. A triplet-based forced-choice experiment with 300 cases, including plagiarism, cover songs, and AI-generated music, was used to gather these judgments. Results show measurable agreement among participants and partial alignment with computational similarity measures, highlighting the need for perceptually grounded evaluation in generative AI for music.
By Roser Batlle-Roca, Woosung Choi, Joan Serr\`a, Fabio Morreale, Wei-Hsiang Liao, Xavier Serra, Emilia G\'omez, Yuki Mitsufuji
MUUNRiver-Bench is a diagnostic benchmark for music retrieval that uses natural‑language instructions to define relevance for reference‑audio queries. It contains 3,440 tracks across 13 genres and 116 sub‑genres and covers seven tasks such as similar‑music, style‑preserving lyric‑rewriting, cover, and segment retrieval. Experiments with six models in eight configurations show that acoustic encoders favor local identity while text‑aligned encoders favor semantic relations, and that instruction‑aware and audio‑text fusion systems do not consistently outperform their backbones.
By Zhancheng Guo, Congren Dai, Shangda Wu, Jianhuai Hu, Danni Zhao, Xiaobing Li, Maosong Sun
The paper evaluates Graph Neural Networks (GNNs) for predicting artist success within collaboration networks, extending prior work on Italian and Danish music scenes by adding a Polish dataset and merging the three into a tri‑national network. Statistical analysis shows the Polish and combined networks share similar clustering properties, while predictive experiments reveal that GNNs match or slightly outperform a Multilayer Perceptron (MLP) in some cases but the MLP generally yields higher success metrics. The findings suggest that internal node attributes such as genre and label affiliation may be more predictive than network topology, and that GNNs may better capture cross‑border relational structures in the merged network.
By Wiktor Dowgia{\l}{\l}o
arXiv:2608. 05153v1 Announce Type: cross Abstract: GraphRAG underperforms vector RAG on citation precision in many reports, but where and why have remained corpus-bound.
By Meftun Akarsu, Burak Ozdemir
arXiv:2605. 03395v2 Announce Type: replace-cross Abstract: Music popularity prediction has attracted growing research interest, with relevance to artists, platforms, and recommendation systems.
By Jaavid Aktar Husain, Dorien Herremans
arXiv:2608.23484v1 Announce Type: new
Abstract: We present Team Semiintelligencn's solution for the ACM RecSys 2026 TalkPlayData Challenge, addressing conversational music recommendation through a mu...
By Naman Garg, Sarika Jain, George Fazekas
We present Team Semiintelligencn's solution for the ACM RecSys 2026 TalkPlayData Challenge, addressing conversational music recommendation through a multi-modal and personalized conversational recomme...
arXiv:2606. 09855v1 Announce Type: cross Abstract: Korean folk painting (minhwa) is built from a small vocabulary of auspicious symbols, a tiger for protection, a pair of birds for marital harmony, a peony for wealth, that recur across many of its painted genres.
By Joonhyung Bae
arXiv:2607. 22413v1 Announce Type: cross Abstract: Sample retrieval tools can help composers find harmonically compatible material, but querying from a fixed reference sample becomes less informative as arrangements evolve and the harmonic context shifts with each musical decision.
By Austin Rockman
arXiv:2606. 07207v1 Announce Type: cross Abstract: Confidence-based loss weighting is usually avoided in generative models because it accelerates errors when the model is confidently wrong, but this intuition breaks down in supervised diffusion training.
By Zixi Li, Youzhen Li