Recherche
Recherche simple
Recherche avancée
Panier électronique
Votre panier ne contient aucune notice
Connexion à la base
Identification
(Identifiez-vous pour accéder aux fonctions de mise à jour. Utilisez votre login-password de courrier électronique)
Entrepôt OAI-PMH
Soumettre une requête
| Consulter la notice détaillée |
| Version complète en ligne |
| Version complète en ligne accessible uniquement depuis l'Ircam |
| Ajouter la notice au panier |
| Retirer la notice du panier |
English version
(full translation not yet available)
Liste complète des articles
|
Consultation des notices
%0 Journal Article
%A Obin, Nicolas
%A Roebel, Axel
%T Similarity Search of Acted Voices for Automatic Voice Casting
%D 2016
%E 1638-1647
%B IEEE/ACM Transactions on Audio, Speech and Language Processing
%V 24
%N 9
%F Obin16a
%K voice casting
%K voice similarity
%K speaker recog-nition
%K speaker traits and states
%K para-linguistics
%K multi-label classification
%X This paper presents a large-scale similarity search of professionally acted voices for computer-aided voice casting. The proposed voice casting system explores GMM-based acoustic models and multi-label recognition of perceived para-linguistic content (speaker states and speaker traits, e.g., age/gender, voice quality, emotion) for the voice casting of professionally acted voices. First, acoustic models (universal background model, super-vector, i-vector) are constructed to model the acoustic space of voices, from which the similarity between voices can be measured directly in the acoustic space. Second, multiple binary classification of speaker traits and states is added to the acoustic models in order to represent the vocal signature of a voice, which is then used to measure the similarity between voices in the para-linguistic space. Finally, a similarity search is processed in order to determine the set of target actors that are the most similar to the voice of a source actor. In a subjective experiment conducted in the real-context of cross-language voice casting, the multi-label scoring system significantly outperforms the acoustic scoring system. This constitutes a proof of concept for the role of perceived para-linguistic categories in the perception of voice similarity.
%1 1
%2 2
%U http://architexte.ircam.fr/textes/Obin16a/
|
|