Detecting Music Performance Errors with Transformers

Chou, Benjamin Shiue-Hal; Jajal, Purvish; Eliopoulos, Nicholas John; Nadolsky, Tim; Yang, Cheng-Yun; Ravi, Nikita; Davis, James C.; Yun, Kristen Yeon-Ji; Lu, Yung-Hsiang

Computer Science > Sound

arXiv:2501.02030 (cs)

[Submitted on 3 Jan 2025]

Title:Detecting Music Performance Errors with Transformers

Authors:Benjamin Shiue-Hal Chou, Purvish Jajal, Nicholas John Eliopoulos, Tim Nadolsky, Cheng-Yun Yang, Nikita Ravi, James C. Davis, Kristen Yeon-Ji Yun, Yung-Hsiang Lu

View PDF HTML (experimental)

Abstract:Beginner musicians often struggle to identify specific errors in their performances, such as playing incorrect notes or rhythms. There are two limitations in existing tools for music error detection: (1) Existing approaches rely on automatic alignment; therefore, they are prone to errors caused by small deviations between alignment targets.; (2) There is a lack of sufficient data to train music error detection models, resulting in over-reliance on heuristics. To address (1), we propose a novel transformer model, Polytune, that takes audio inputs and outputs annotated music scores. This model can be trained end-to-end to implicitly align and compare performance audio with music scores through latent space representations. To address (2), we present a novel data generation technique capable of creating large-scale synthetic music error datasets. Our approach achieves a 64.1% average Error Detection F1 score, improving upon prior work by 40 percentage points across 14 instruments. Additionally, compared with existing transcription methods repurposed for music error detection, our model can handle multiple instruments. Our source code and datasets are available at this https URL.

Comments:	AAAI 2025
Subjects:	Sound (cs.SD); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2501.02030 [cs.SD]
	(or arXiv:2501.02030v1 [cs.SD] for this version)
	https://doi.org/10.48550/arXiv.2501.02030

Submission history

From: Benjamin Chou [view email]
[v1] Fri, 3 Jan 2025 07:04:20 UTC (2,527 KB)

Computer Science > Sound

Title:Detecting Music Performance Errors with Transformers

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Sound

Title:Detecting Music Performance Errors with Transformers

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators