Rhythm Speech Lyrics Input for MIDI-Based Singing Voice Synthesis

  • Authors:
  • Hong-Ru Lee;Chih-Fang Huang;Chih-Hao Hsu;Wen-Nan Wang

  • Affiliations:
  • Department of Mechanical Engineering, National Chiao-Tung University,;Department of Information Communication, Yuan Ze University,;Innovative DigiTech-Enabled Applications & Services Institute, Taipei City, Taiwan 105;Innovative DigiTech-Enabled Applications & Services Institute, Taipei City, Taiwan 105

  • Venue:
  • PCM '09 Proceedings of the 10th Pacific Rim Conference on Multimedia: Advances in Multimedia Information Processing
  • Year:
  • 2009

Quantified Score

Hi-index 0.00

Visualization

Abstract

This paper presents useful techniques and considerations in implementing underlying mandarin singing voice synthesis system using the RSLI unit. The system can receive the continuous speech of the lyrics of a song, and can synthesize the intended song based on the MIDI-based music database. This system is designed based on 3 units.. The first one is the input unit which allows the user specifies a musical score and phonetically-spelled lyrics to system. The second one is the modified unit and it is employed to implement the pitch-shifting function using the PSOLA method. The third one is the mixed unit which has some undesirable artificial-sounding buzzy-effects, including echo and vibrato effects. Moreover, the energy, duration, and spectrum modifications are also implemented in the mixed unit. The synthesized singing voice sounds reasonably good. From the subjective listening test, the MOS (mean opinion score) of 3.3 and 3.2 are obtained for the synthesized singing voices and the similarity of singer's voice, respectively.