Voice quality modelling for expressive speech synthesis

This paper presents the perceptual experiments that were carried out in order to validate the methodology of transforming expressive speech styles using voice quality (VoQ) parameters modelling, along with the well-known prosody (, duration, and energy), from a neutral style into a number of express...

Descripción completa

Detalles Bibliográficos
Autores: Monzo Sánchez, Carlos, Iriondo Sanz, Ignasi, Socoró Carrié, Joan Claudi
Tipo de recurso: artículo
Estado:Versión publicada
Fecha de publicación:2014
País:España
Institución:Varias* (Consorci de Biblioteques Universitáries de Catalunya, Centre de Serveis Científics i Acadèmics de Catalunya)
Repositorio:Recercat. Dipósit de la Recerca de Catalunya
OAI Identifier:oai:recercat.cat:20.500.14342/3439
Acceso en línea:http://hdl.handle.net/20.500.14342/3439
https://doi.org/10.1155/2014/627189
Access Level:acceso abierto
Palabra clave:Parla
81
id ES_8fe0e944bedd46782c39bf5489bd2f6e
oai_identifier_str oai:recercat.cat:20.500.14342/3439
network_acronym_str ES
network_name_str España
repository_id_str
spelling Voice quality modelling for expressive speech synthesisMonzo Sánchez, CarlosIriondo Sanz, IgnasiSocoró Carrié, Joan ClaudiParla81This paper presents the perceptual experiments that were carried out in order to validate the methodology of transforming expressive speech styles using voice quality (VoQ) parameters modelling, along with the well-known prosody (, duration, and energy), from a neutral style into a number of expressive ones. The main goal was to validate the usefulness of VoQ in the enhancement of expressive synthetic speech in terms of speech quality and style identification. A harmonic plus noise model (HNM) was used to modify VoQ and prosodic parameters that were extracted from an expressive speech corpus. Perception test results indicated the improvement of obtained expressive speech styles using VoQ modelling along with prosodic characteristicsHindawi Publishing CorporationUniversitat Ramon Llull. La SalleUniversitat Oberta de Catalunya2014info:eu-repo/semantics/articleinfo:eu-repo/semantics/publishedVersion13 p.http://hdl.handle.net/20.500.14342/3439https://doi.org/10.1155/2014/627189reponame:Recercat. Dipósit de la Recerca de Catalunyainstname:Varias* (Consorci de Biblioteques Universitáries de Catalunya, Centre de Serveis Científics i Acadèmics de Catalunya)InglésThe Scientific World Journal. 2014Attribution 4.0 International© L'autor/ahttp://creativecommons.org/licenses/by/4.0/info:eu-repo/semantics/openAccessoai:recercat.cat:20.500.14342/34392026-05-29T05:05:01Z
dc.title.none.fl_str_mv Voice quality modelling for expressive speech synthesis
title Voice quality modelling for expressive speech synthesis
spellingShingle Voice quality modelling for expressive speech synthesis
Monzo Sánchez, Carlos
Parla
81
title_short Voice quality modelling for expressive speech synthesis
title_full Voice quality modelling for expressive speech synthesis
title_fullStr Voice quality modelling for expressive speech synthesis
title_full_unstemmed Voice quality modelling for expressive speech synthesis
title_sort Voice quality modelling for expressive speech synthesis
dc.creator.none.fl_str_mv Monzo Sánchez, Carlos
Iriondo Sanz, Ignasi
Socoró Carrié, Joan Claudi
author Monzo Sánchez, Carlos
author_facet Monzo Sánchez, Carlos
Iriondo Sanz, Ignasi
Socoró Carrié, Joan Claudi
author_role author
author2 Iriondo Sanz, Ignasi
Socoró Carrié, Joan Claudi
author2_role author
author
dc.contributor.none.fl_str_mv Universitat Ramon Llull. La Salle
Universitat Oberta de Catalunya
dc.subject.none.fl_str_mv Parla
81
topic Parla
81
description This paper presents the perceptual experiments that were carried out in order to validate the methodology of transforming expressive speech styles using voice quality (VoQ) parameters modelling, along with the well-known prosody (, duration, and energy), from a neutral style into a number of expressive ones. The main goal was to validate the usefulness of VoQ in the enhancement of expressive synthetic speech in terms of speech quality and style identification. A harmonic plus noise model (HNM) was used to modify VoQ and prosodic parameters that were extracted from an expressive speech corpus. Perception test results indicated the improvement of obtained expressive speech styles using VoQ modelling along with prosodic characteristics
publishDate 2014
dc.date.none.fl_str_mv 2014
dc.type.none.fl_str_mv info:eu-repo/semantics/article
info:eu-repo/semantics/publishedVersion
format article
status_str publishedVersion
dc.identifier.none.fl_str_mv http://hdl.handle.net/20.500.14342/3439
https://doi.org/10.1155/2014/627189
url http://hdl.handle.net/20.500.14342/3439
https://doi.org/10.1155/2014/627189
dc.language.none.fl_str_mv Inglés
language_invalid_str_mv Inglés
dc.relation.none.fl_str_mv The Scientific World Journal. 2014
dc.rights.none.fl_str_mv Attribution 4.0 International
© L'autor/a
http://creativecommons.org/licenses/by/4.0/
info:eu-repo/semantics/openAccess
rights_invalid_str_mv Attribution 4.0 International
© L'autor/a
http://creativecommons.org/licenses/by/4.0/
eu_rights_str_mv openAccess
dc.format.none.fl_str_mv 13 p.
dc.publisher.none.fl_str_mv Hindawi Publishing Corporation
publisher.none.fl_str_mv Hindawi Publishing Corporation
dc.source.none.fl_str_mv reponame:Recercat. Dipósit de la Recerca de Catalunya
instname:Varias* (Consorci de Biblioteques Universitáries de Catalunya, Centre de Serveis Científics i Acadèmics de Catalunya)
instname_str Varias* (Consorci de Biblioteques Universitáries de Catalunya, Centre de Serveis Científics i Acadèmics de Catalunya)
reponame_str Recercat. Dipósit de la Recerca de Catalunya
collection Recercat. Dipósit de la Recerca de Catalunya
repository.name.fl_str_mv
repository.mail.fl_str_mv
_version_ 1869413248579665920
score 15,198674