Exploiting the Two-Dimensional Nature of Agnostic Music Notation for Neural Optical Music Recognition

Alfaro-Contreras, María; Valero-Mas, Jose J.

Exploiting the Two-Dimensional Nature of Agnostic Music Notation for Neural Optical Music Recognition

Por favor, use este identificador para citar o enlazar este ítem: http://hdl.handle.net/10045/114284

Registro completo de metadatos

Registro completo de metadatos
Campo DC	Valor	Idioma
dc.contributor	Reconocimiento de Formas e Inteligencia Artificial	es_ES
dc.contributor.author	Alfaro-Contreras, María	-
dc.contributor.author	Valero-Mas, Jose J.	-
dc.contributor.other	Universidad de Alicante. Departamento de Lenguajes y Sistemas Informáticos	es_ES
dc.date.accessioned	2021-04-20T12:09:04Z	-
dc.date.available	2021-04-20T12:09:04Z	-
dc.date.issued	2021-04-17	-
dc.identifier.citation	Alfaro-Contreras M, Valero-Mas JJ. Exploiting the Two-Dimensional Nature of Agnostic Music Notation for Neural Optical Music Recognition. Applied Sciences. 2021; 11(8):3621. https://doi.org/10.3390/app11083621	es_ES
dc.identifier.issn	2076-3417	-
dc.identifier.uri	http://hdl.handle.net/10045/114284	-
dc.description.abstract	State-of-the-art Optical Music Recognition (OMR) techniques follow an end-to-end or holistic approach, i.e., a sole stage for completely processing a single-staff section image and for retrieving the symbols that appear therein. Such recognition systems are characterized by not requiring an exact alignment between each staff and their corresponding labels, hence facilitating the creation and retrieval of labeled corpora. Most commonly, these approaches consider an agnostic music representation, which characterizes music symbols by their shape and height (vertical position in the staff). However, this double nature is ignored since, in the learning process, these two features are treated as a single symbol. This work aims to exploit this trademark that differentiates music notation from other similar domains, such as text, by introducing a novel end-to-end approach to solve the OMR task at a staff-line level. We consider two Convolutional Recurrent Neural Network (CRNN) schemes trained to simultaneously extract the shape and height information and to propose different policies for eventually merging them at the actual neural level. The results obtained for two corpora of monophonic early music manuscripts prove that our proposal significantly decreases the recognition error in figures ranging between 14.4% and 25.6% in the best-case scenarios when compared to the baseline considered.	es_ES
dc.description.sponsorship	This research work was partially funded by the University of Alicante through project GRE19-04, by the “Programa I+D+i de la Generalitat Valenciana” through grant APOSTD/2020/256, and by the Spanish Ministerio de Universidades through grant FPU19/04957.	es_ES
dc.language	eng	es_ES
dc.publisher	MDPI	es_ES
dc.rights	© 2021 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/licenses/by/4.0/).	es_ES
dc.subject	Optical music recognition	es_ES
dc.subject	Deep learning	es_ES
dc.subject	Connectionist temporal classification	es_ES
dc.subject	Agnostic music notation	es_ES
dc.subject	Sequence labeling	es_ES
dc.subject.other	Lenguajes y Sistemas Informáticos	es_ES
dc.title	Exploiting the Two-Dimensional Nature of Agnostic Music Notation for Neural Optical Music Recognition	es_ES
dc.type	info:eu-repo/semantics/article	es_ES
dc.peerreviewed	si	es_ES
dc.identifier.doi	10.3390/app11083621	-
dc.relation.publisherversion	https://doi.org/10.3390/app11083621	es_ES
dc.rights.accessRights	info:eu-repo/semantics/openAccess	es_ES
Aparece en las colecciones:	INV - GRFIA - Artículos de Revistas

Archivos en este ítem:

Archivos en este ítem:
Archivo	Descripción	Tamaño	Formato
Alfaro-Contreras_Valero-Mas_2021_ApplSci.pdf		3,34 MB	Adobe PDF	Abrir Vista previa Cerrar vista previa

Ver citas en Google Académico

Muestra el registro sencillo

Este ítem está licenciado bajo Licencia Creative Commons