Emotion conversion based on prosodic unit selection

  • Authors:
  • Daniel Erro;Eva Navas;Inma Hernáez;Ibon Saratxaga

  • Affiliations:
  • AhoLab Signal Processing Laboratory, Electronics and Telecommunications Department, University of the Basque Country, AIda, Bilbao, Spain;AhoLab Signal Processing Laboratory, Electronics and Telecommunications Department, University of the Basque Country, AIda, Bilbao, Spain;AhoLab Signal Processing Laboratory, Electronics and Telecommunications Department, University of the Basque Country, AIda, Bilbao, Spain;AhoLab Signal Processing Laboratory, Electronics and Telecommunications Department, University of the Basque Country, AIda, Bilbao, Spain

  • Venue:
  • IEEE Transactions on Audio, Speech, and Language Processing
  • Year:
  • 2010

Quantified Score

Hi-index 0.00

Visualization

Abstract

Voice conversion has been traditionally focused on spectrum. Current systems lack a solid prosody conversion method suitable for different speaking styles. Recently, the unit selection technique has been applied to transform emotional intonation contours. This paper goes one step beyond: it explores strategies for training and configuring the selection cost function in an emotion conversion application. The proposed system, which uses accent groups as basic intonation units and performs conversion also on phoneme durations and intensity, is evaluated by means of a carefully designed subjective test involving the big six emotions. Although the expressiveness of the converted sentences is still far from that of natural emotional speech, satisfactory results are obtained when different configurations are used for different emotions.