Research Papers research paper arxiv ai artificial-intelligence

Unsupervised Evaluation of Deep Audio Embeddings for Music Structure Analysis

arXivMarch 31, 202610 min read0 views

arXiv:2603.27218v1 Announce Type: cross Abstract: Music Structure Analysis (MSA) aims to uncover the high-level organization of musical pieces. State-of-the-art methods are often based on supervised deep learning, but these methods are bottlenecked by the need for heavily annotated data and inherent structural ambiguities. In this paper, we propose an unsupervised evaluation of nine open-source, generic pre-trained deep audio models, on MSA. For each model, we extract barwise embeddings and segment them using three unsupervised segmentation algorithms (Foote's checkerboard kernels, spectral cl — Axel Marmoret

View PDF HTML (experimental)

Abstract:Music Structure Analysis (MSA) aims to uncover the high-level organization of musical pieces. State-of-the-art methods are often based on supervised deep learning, but these methods are bottlenecked by the need for heavily annotated data and inherent structural ambiguities. In this paper, we propose an unsupervised evaluation of nine open-source, generic pre-trained deep audio models, on MSA. For each model, we extract barwise embeddings and segment them using three unsupervised segmentation algorithms (Foote's checkerboard kernels, spectral clustering, and Correlation Block-Matching (CBM)), focusing exclusively on boundary retrieval. Our results demonstrate that modern, generic deep embeddings generally outperform traditional spectrogram-based baselines, but not systematically. Furthermore, our unsupervised boundary estimation methodology generally yields stronger performance than recent linear probing baselines. Among the evaluated techniques, the CBM algorithm consistently emerges as the most effective downstream segmentation method. Finally, we highlight the artificial inflation of standard evaluation metrics and advocate for the systematic adoption of trimming'', or even double trimming'' annotations to establish more rigorous MSA evaluation standards.

Comments: Submitted to the SMC 2026 conference. 2 figures and 2 tables in the main document, 7 figures in Appendix

Subjects:

Sound (cs.SD); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)

ACM classes: H.5.5

Cite as: arXiv:2603.27218 [cs.SD]

(or arXiv:2603.27218v1 [cs.SD] for this version)

https://doi.org/10.48550/arXiv.2603.27218

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Axel Marmoret [view email] [v1] Sat, 28 Mar 2026 10:18:54 UTC (200 KB)

Original source

arXiv

https://arxiv.org/abs/2603.27218

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

ModelsLive

What Karpathy's Autoresearch Unlocked for Me

I'm not a data scientist. I've trained a few models before — simple classification problems, with AI writing the Python and me running the iterations. It worked. I got confident. Then a friend asked for help with something harder. <h2> Three Weeks at 0.58 </h2> The problem involved predicting an outcome from a mix of CRM data and call recordings. Not trivial, but not exotic either. Quick primer on AUC — the metric I'll use throughout. Imagine your model looks at two random people: one where the answer is yes, one where it's no. AUC measures how often the model correctly ranks the yes above the no. Score of 0.5 means random guessing. Score of 1.0 means always right. I tried everything I knew: XGBoost, feature engineering, extracting features from transcripts u

DEV Community

3m27 minutes ago

Research PapersFresh

Is AI's visual understanding mostly a 'mirage'? New research suggests so - inkl

<a href="https://news.google.com/rss/articles/CBMimwFBVV95cUxNUjItTURkaldjcXhpNjA0b2dSVUhWNGRwaEkzclBMM0FFMk8wNTNsMWFDc1hmVU9jYU5jbzBKSG42WGZpbTI5TUs0R0w5Q2QxX0NCdjgtT2JtUUNsWHkxUFE1MzJjN1RxMmdmQjc5YlBsVTdfVU02Z1FZSlFrLU1hdmxibjBGcUV3STlvc0JnaG9PcXRDZVhQamdmYw?oc=5" target="_blank">Is AI's visual understanding mostly a 'mirage'? New research suggests so</a> inkl

GNews AI multimodal

1mabout 5 hours ago

Self-Evolving AI

Google DeepMind Introduces Aletheia: The AI Agent Moving from Math Competitions to Fully Autonomous Professional Research Discoveries - marktechpost.com

<a href="https://news.google.com/rss/articles/CBMigwJBVV95cUxPOWhkeUl4RkMxb1IzY1NKcUlVeFpYQ3NWc0ZLWVo2OWRESkdsTkZYYlpaamJ6WjE0Z3RaUkFJVEhIQnBFSHdjV2tEd3R1dFVGS28tYXlFNlZwZnVZSnV2TFlNNDFrNDFIdGJ6VzlwbENMZ2x3dHdFNXFWdzlWLWt2OEZQcW1WbTExSUdOVnNjbktiQURwQXRLUFZVeXp5WjZhbTY4dXhpdlphNWl2THNRNGxqbVcxcDlDbW5US3VBLUNvR1owSHNIUE5xMktmcVVDTjl4dXpOMmVPTUdQWkx0aF9yRXowU0NxQ2lHc0VMRzlaNDEyU0lLY0lSdWpLUndFUll30gGIAkFVX3lxTE0tR0JPRll6R09pU1d3NzVSQ0YwSVRJS1Q5YVF6THpfRUhEZ2EyVGJBNi1XX1ZOT19zZkU1WDlqelRzMm5NWjU3VTR2WC1LZ0drMTUyaHZVWFNLa1MwbEJ1OUZYM2R5cWFza0hJbllFSHhPc0tWYTNMbU1PMmw4T0RtLVpZLXBfbERKRUR0LTF6bl94S1FJRDBweVpxRGpTU242a25lMTVHdG5pTXBxUVgtaHp3MG9yX1NEd3p6Z0lKaWxLTVJGNWQxVkxVRFdZbERzSFBJUjUwQkRPNENYWVVrdlM4dTljODRxeWhDMDdLN1czT0tadUM1YmtCcENBcWU4NngwcUI3ZA?oc=5" target="_blank">Google DeepMind Intro

Google News: DeepMind

1m19 days ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 125 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

More in Research Papers

Research PapersFresh

Is AI's visual understanding mostly a 'mirage'? New research suggests so - inkl

GNews AI multimodal

1mabout 5 hours ago

Research PapersLive

Google backs UH Mānoa AI, robotics research - University of Hawaii System

<a href="https://news.google.com/rss/articles/CBMickFVX3lxTE04TGNzcGpVeFNibkdwMzJIOFdrMHYtSS1OUDdFZkR6RFRtUU4yYWN1MlYtRGZQazl5Y1k3SklTYURwYVBycG5NZm5zbzNfRjY0SF9WNVJfM2tTU1BOX0xfb2ZtaWFVVFp3cFg3WXJmNnQwQQ?oc=5" target="_blank">Google backs UH Mānoa AI, robotics research</a> University of Hawaii System

GNews AI Google

1mabout 1 hour ago

Research PapersLive

A Retrospective on the ICLR 2026 Review Process

The selection of papers for ICLR 2026 has fully concluded. We extend our congratulations to the authors whose work will appear at the conference. Creating ICLR’s technical program requires immense effort from the authors, reviewers, and area chairs, and we thank you for your contributions and service. For researchers whose work was rejected, we hope […]

blog.iclr.cc

1mabout 2 hours ago

Research Papers

Vector Researchers present papers at ACL 2024

Vector researchers will be well represented at the 62nd Annual Meeting of the Association for Computational Linguistics in Bangkok, Thailand this year. 14 papers co-authored by Vector-affiliated researchers are being […] The post Vector Researchers present papers at ACL 2024 appeared first on Vector Institute for Artificial Intelligence .

Vector Institute

1mover 1 year ago