Detecting Music Performance Errors with Transformers

Chou, Benjamin Shiue_Hal; Jajal, Purvish; Eliopoulos, Nicholas John; Nadolsky, Tim; Yang, Cheng_Yun; Ravi, Nikita; Davis, James C; Yun, Kristen Yeon_Ji; Lu, Yung_Hsiang

doi:10.1609/aaai.v39i22.34539

Citation Details

This content will become publicly available on April 11, 2026

Detecting Music Performance Errors with Transformers

Beginner musicians often struggle to identify specific errors in their performances, such as playing incorrect notes or rhythms. There are two limitations in existing tools for music error detection: (1) Existing approaches rely on automatic alignment; therefore, they are prone to errors caused by small deviations between alignment targets; (2) There is insufficient data to train music error detection models, resulting in over-reliance on heuristics. To address (1), we propose a novel transformer model, Polytune, that takes audio inputs and outputs annotated music scores. This model can be trained end-to-end to implicitly align and compare performance audio with music scores through latent space representations. To address (2), we present a novel data generation technique capable of creating large-scale synthetic music error datasets. Our approach achieves a 64.1% average Error Detection F1 score, improving upon prior work by 40 percentage points across 14 instruments. Additionally, our model can handle multiple instruments compared with existing transcription methods repurposed for music error detection. more »

Award ID(s):: 2326198

PAR ID:: 10626351

Author(s) / Creator(s):: Chou, Benjamin Shiue_Hal; Jajal, Purvish; Eliopoulos, Nicholas John; Nadolsky, Tim; Yang, Cheng_Yun; Ravi, Nikita; Davis, James C; Yun, Kristen Yeon_Ji; Lu, Yung_Hsiang

Publisher / Repository:: PKP Publishing Services Network

Date Published:: 2025-04-11

Journal Name:: Proceedings of the AAAI Conference on Artificial Intelligence

Volume:: 39

Issue:: 22

ISSN:: 2159-5399

Page Range / eLocation ID:: 23687 to 23695

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
This content will become publicly available on April 11, 2026
Journal Article:
https://doi.org/10.1609/aaai.v39i22.34539

More Like this