Paper published in a journal (Scientific congresses and symposiums)
Towards a Voice Conversion System Based on Frame Selection
Dutoit, Thierry; Holzapfel, A.; Jottrand, Matthieu et al.
2007
 

Files


Full Text
icassp2007_tdahmjamjpys.pdf
Author preprint (182.77 kB)
Request a copy

All documents in ORBi UMONS are protected by a user license.

Send to



Details



Abstract :
[en] The subject of this paper is the conversion of a given speaker's voice (the source speaker) into another identified voice (the target one). We assume we have at our disposal a large amount of speech samples from source and target voice with at least a part of them being parallel. The proposed system is built on a mapping function between source and target spectral envelopes followed by a frame selection algorithm to produce final spectral envelopes. Converted speech is produced by a basic LP analysis of the source and LP synthesis using the converted spectral envelopes. We compared three types of conversion: without mapping, with mapping and using the excitation of the source speaker and finally with mapping using the excitation of the target. Results show that the combination of mapping and frame selection provide the best results, and underline the interest to work on methods to convert the LP excitation.
Disciplines :
Electrical & electronics engineering
Author, co-author :
Dutoit, Thierry ;  Université de Mons > Faculté Polytechnique > Information, Signal et Intelligence artificielle
Holzapfel, A.
Jottrand, Matthieu
Moinet, Alexis ;  Université de Mons > Faculté Polytechnique > Information, Signal et Intelligence artificielle
Perez, Javier
Stylianou, Y.
Language :
English
Title :
Towards a Voice Conversion System Based on Frame Selection
Publication date :
16 April 2007
Event name :
ICASSP 2007 - International Conference on Acoustics, Speech and Signal Processing
Event place :
Honolulu, United States - Hawaii
Event date :
2007
Research unit :
F105 - Information, Signal et Intelligence artificielle
Research institute :
R300 - Institut de Recherche en Technologies de l'Information et Sciences de l'Informatique
R450 - Institut NUMEDIART pour les Technologies des Arts Numériques
Commentary :
see also : proceedings of enterface 2006 : 'Multimodal Speaker Conversion - his master's voice... and face -'
Available on ORBi UMONS :
since 10 December 2010

Statistics


Number of views
1 (0 by UMONS)
Number of downloads
0 (0 by UMONS)

Scopus citations®
 
35
Scopus citations®
without self-citations
35
OpenCitations
 
24

Bibliography


Similar publications



Contact ORBi UMONS