Application of sound source separation methods to advanced spatial audio systems

dc.contributor.advisorLópez Monfort, José Javier
dc.contributor.affiliationEscuela Técnica Superior de Ingeniería de Telecomunicación
dc.contributor.affiliationDepartamento de Comunicaciones
dc.contributor.affiliationInstituto Universitario de Telecomunicación y Aplicaciones Multimedia
dc.contributor.authorCobos Serrano, Máximoes_ES
dc.date.accessioned2010-12-03T07:47:08Z
dc.date.available2010-12-03T07:47:08Z
dc.date.created2009-09-03T08:00:00Zes_ES
dc.date.issued2010-12-03es_ES
dc.description.abstractThis thesis is related to the field of Sound Source Separation (SSS). It addresses the development and evaluation of these techniques for their application in the resynthesis of high-realism sound scenes by means of Wave Field Synthesis (WFS). Because the vast majority of audio recordings are preserved in twochannel stereo format, special up-converters are required to use advanced spatial audio reproduction formats, such as WFS. This is due to the fact that WFS needs the original source signals to be available, in order to accurately synthesize the acoustic field inside an extended listening area. Thus, an object-based mixing is required. Source separation problems in digital signal processing are those in which several signals have been mixed together and the objective is to find out what the original signals were. Therefore, SSS algorithms can be applied to existing two-channel mixtures to extract the different objects that compose the stereo scene. Unfortunately, most stereo mixtures are underdetermined, i.e., there are more sound sources than audio channels. This condition makes the SSS problem especially difficult and stronger assumptions have to be taken, often related to the sparsity of the sources under some signal transformation. This thesis is focused on the application of SSS techniques to the spatial sound reproduction field. As a result, its contributions can be categorized within these two areas. First, two underdetermined SSS methods are proposed to deal efficiently with the separation of stereo sound mixtures. These techniques are based on a multi-level thresholding segmentation approach, which enables to perform a fast and unsupervised separation of sound sources in the time-frequency domain. Although both techniques rely on the same clustering type, the features considered by each of them are related to different localization cues that enable to perform separation of either instantaneous or real mixtures.Additionally, two post-processing techniques aimed at improving the isolation of the separated sources are proposed. The performance achieved by several SSS methods in the resynthesis of WFS sound scenes is afterwards evaluated by means of listening tests, paying special attention to the change observed in the perceived spatial attributes. Although the estimated sources are distorted versions of the original ones, the masking effects involved in their spatial remixing make artifacts less perceptible, which improves the overall assessed quality. Finally, some novel developments related to the application of time-frequency processing to source localization and enhanced sound reproduction are presented.en_EN
dc.description.accrualMethodPalanciaes_ES
dc.description.bibliographicCitationCobos Serrano, M. (2009). Application of sound source separation methods to advanced spatial audio systems [Tesis doctoral]. Universitat Politècnica de València. https://doi.org/10.4995/Thesis/10251/8969es_ES
dc.identifier.doi10.4995/Thesis/10251/8969es_ES
dc.identifier.urihttps://riunet.upv.es/handle/10251/8969
dc.languageIngléses_ES
dc.publisherUniversitat Politècnica de Valènciaes_ES
dc.relation.tesis3130es_ES
dc.rightsReserva de todos los derechoses_ES
dc.rights.accessRightsAbiertoes_ES
dc.sourceRiunet
dc.subjectWave field synthesises_ES
dc.subjectSource separationes_ES
dc.subjectTime frequency processinges_ES
dc.subjectDirection of arrivales_ES
dc.subjectSpatial audioes_ES
dc.subject.classificationTEORIA DE LA SEÑAL Y COMUNICACIONESes_ES
dc.titleApplication of sound source separation methods to advanced spatial audio systems
dc.typeTesis doctorales_ES
dc.type.versioninfo:eu-repo/semantics/acceptedVersiones_ES
dspace.entity.typePublication
person.identifier46
person.identifier.orcid0000-0001-6884-5577
relation.isAdvisorOfPublication2e2fefa1-2108-401b-bacc-956a965b6408
relation.isAdvisorOfPublication.latestForDiscovery2e2fefa1-2108-401b-bacc-956a965b6408
relation.isOrgUnitOfPublicationaa6a0db9-4584-45eb-b7e3-73606ac49444
relation.isOrgUnitOfPublication02a0f2c5-c452-4e1d-a7d9-b731347d078c
relation.isOrgUnitOfPublication7eb466f9-a4ba-4215-bab2-52a6a5dd6c63
relation.isOrgUnitOfPublication.latestForDiscoveryaa6a0db9-4584-45eb-b7e3-73606ac49444
upv.uuid694e0101-de5b-4530-bf02-7990355491dfes_ES

Archivos

Bloque original

Mostrando 1 - 5 de 5
Cargando...
Miniatura
Nombre:
tesisUPV3130.pdf
Tamaño:
5.9 MB
Formato:
Adobe Portable Document Format
Cargando...
Miniatura
Nombre:
tesisUPV3130_ResumenCastellano.txt
Tamaño:
2.95 KB
Formato:
Plain Text
Cargando...
Miniatura
Nombre:
tesisUPV3130_EnglishAbstract.txt
Tamaño:
2.61 KB
Formato:
Plain Text
Cargando...
Miniatura
Nombre:
tesisUPV3130_ResumenValenciano.txt
Tamaño:
2.89 KB
Formato:
Plain Text
Cargando...
Miniatura
Nombre:
tesisUPV3130_Indice.pdf
Tamaño:
174.06 KB
Formato:
Adobe Portable Document Format

Colecciones