Speech and singing voice classifier based on musical note classification and fundamental frequency estimation.

dc.contributor.advisorRamos Peinado, Germán
dc.contributor.advisorLuzón Álvarez, Claraes_ES
dc.contributor.affiliationDepartamento de Ingeniería Electrónica
dc.contributor.affiliationEscuela Técnica Superior de Ingeniería de Telecomunicación
dc.contributor.affiliationInstituto Universitario de Telecomunicación y Aplicaciones Multimedia
dc.contributor.authorDiana Sánchez, Santiagoes_ES
dc.date.accessioned2022-10-14T17:26:29Z
dc.date.available2022-10-14T17:26:29Z
dc.date.created2022-07-18es_ES
dc.date.issued2022-10-14es_ES
dc.description.abstract[ES] El procesamiento del habla es el estudio de las señales del habla y de los métodos de procesamiento de señal aplicados al análisis y el tratamiento de las mismas. Este tratamiento es actualmente realizado en el dominio discreto, digitalizando previamente la señal del habla. En este ámbito, la distinción entre el voz hablada y voz cantada es una tarea esencial para ayudar a la consecución de varios tipos de algoritmos y de su posterior toma de decisiones. Esta discriminación es a veces difícil incluso para los humanos, dependiendo del tipo de música en el que se basa la voz cantada. Para lograr una buena discriminación entre ambas clases (hablada y cantada), se debe realizar un estudio profundo y amplio debido a que cada tipo de voz puede requerir de un procesado y de algoritmos de decisión diferentes. Este trabajo, presenta un clasificador automático entre voz cantada y hablada o discurso, basado en dos parámetros principales: el tono derivado de la clasificación de las notas musicales y la estima de la frecuencia fundamental. Este método obtiene buenos resultados en la tarea de discriminar entre silencio, habla y voz cantada. Este trabajo tiene una directa aplicación industrial en el campo de la investigación del audio, la aplicación de filtros o algoritmos en función de si hay voz cantada o no, o la discriminación de estilos musicales a partir de los parámetros extraídos de la voz cantada.es_ES
dc.description.abstract[EN] Speech processing is the study of speech signals and signal processing methods applied to their analysis and treatment. This processing is currently performed in the discrete domain, previously digitizing the speech signal. In this area, the distinction between the spoken voice and the sung voice is an essential task to help the achievement of various types of algorithms and their subsequent decision-making. This discrimination is sometimes difficult even for humans, depending on the type of music on which the sung voice is based. To achieve a good discrimination between both classes (spoken and sung), a deep and broad study must be carried out due to each type of voice may require different processing and decision algorithms. This work presents an automatic classifier between sung and spoken voice or speech, based on two main parameters: the tone derived from the classification of musical notes and the estimation of the fundamental frequency. This method obtains good results in the task of discriminating between silence, speech and sung voice. This work has a direct industrial application in the field of audio research, the application of filters or algorithms depending on whether there is sung voice or not, or the discrimination of musical styles from the parameters extracted from the sung voice.en_EN
dc.description.accrualMethodTFGMes_ES
dc.description.bibliographicCitationDiana Sánchez, S. (2022). Speech and singing voice classifier based on musical note classification and fundamental frequency estimation. Universitat Politècnica de València. https://riunet.upv.es/handle/10251/187785es_ES
dc.format.extent61es_ES
dc.identifier.urihttps://riunet.upv.es/handle/10251/187785
dc.languageIngléses_ES
dc.publisherUniversitat Politècnica de Valènciaes_ES
dc.relation.pasarelaTFGM\149786es_ES
dc.rightsReserva de todos los derechoses_ES
dc.rights.accessRightsAbiertoes_ES
dc.subjectDiscursoes_ES
dc.subjectVoz cantadaes_ES
dc.subjectClasificadores_ES
dc.subjectPrototipoes_ES
dc.subjectDetección de notases_ES
dc.subjectTimbrees_ES
dc.subjectSpeechen_EN
dc.subjectSinging voiceen_EN
dc.subjectClassifieren_EN
dc.subjectPrototypeen_EN
dc.subjectNote detectionen_EN
dc.subjectPitchen_EN
dc.subject.classificationTECNOLOGIA ELECTRONICAes_ES
dc.subject.classificationINGENIERIA TELEMATICAes_ES
dc.subject.otherGrado en Ingeniería de Tecnologías y Servicios de Telecomunicación-Grau en Enginyeria de Tecnologies i Serveis de Telecomunicacióes_ES
dc.titleSpeech and singing voice classifier based on musical note classification and fundamental frequency estimation.es_ES
dc.title.alternativeClasificador de voz y canto basado en la clasificación de notas musicales y la estimación de la frecuencia fundamental.es_ES
dc.title.alternativeClassificador de veu i cant basat en la classificació de notes musicals i l'estimació de la freqüència fonamental.es_ES
dc.typeProyecto/Trabajo fin de carrera/gradoes_ES
dspace.entity.typePublication
person.identifier20996
person.identifier.orcid0000-0001-7152-0044
relation.isAdvisorOfPublication180f7656-f3a1-48be-9c39-523e979b2d43
relation.isAdvisorOfPublication.latestForDiscovery180f7656-f3a1-48be-9c39-523e979b2d43
relation.isOrgUnitOfPublicationa3ca7df3-e467-4f69-b673-79a5d6241220
relation.isOrgUnitOfPublicationaa6a0db9-4584-45eb-b7e3-73606ac49444
relation.isOrgUnitOfPublication7eb466f9-a4ba-4215-bab2-52a6a5dd6c63
relation.isOrgUnitOfPublication.latestForDiscoverya3ca7df3-e467-4f69-b673-79a5d6241220
upv.uuid3b4b2053-59b4-42da-b6a3-41fc51e062cdes_ES

Archivos

Bloque original

Mostrando 1 - 1 de 1
Cargando...
Miniatura
Nombre:
Diana - Speech and singing voice classifier based on musical note classification and fundamental ....pdf
Tamaño:
2.92 MB
Formato:
Adobe Portable Document Format