Recognition of elderly speech and voice-driven document retrieval

Authors:
S. Anderson;N. Liberman;E. Bernstein;S. Foster;E. Cate;B. Levin;R. Hudson
Affiliations:
Dragon Syst. Inc., Newton, MA, USA;-;-;-;-;-;-
Venue:
ICASSP '99 Proceedings of the Acoustics, Speech, and Signal Processing, 1999. on 1999 IEEE International Conference - Volume 01
Year:
1999

Citing 0
Cited 4

Automatic gender recognition

ICECS'03 Proceedings of the 2nd WSEAS International Conference on Electronics, Control and Signal Processing
Speech Input from Older Users in Smart Environments: Challenges and Perspectives

UAHCI '09 Proceedings of the 5th International on ConferenceUniversal Access in Human-Computer Interaction. Part II: Intelligent and Ubiquitous Interaction Environments
The CARES corpus: a database of older adult actor simulated emergency dialogue for developing a personal emergency response system

International Journal of Speech Technology
Android-based speech processing for eldercare robotics

Proceedings of the companion publication of the 2013 international conference on Intelligent user interfaces companion

Quantified Score

Hi-index	0.00

Visualization

Abstract

We have collected a corpus of 78 hours of speech from 297 elderly speakers, with an average age of 79. We find that acoustic models built from elderly speech provide much better recognition than do non-elderly models (42.1 vs. 54.6% WER). We also find that elderly men have substantially higher word error rates than elderly women (typically 14% absolute). We report on other experiments with this corpus, dividing the speakers by age, by gender, and by regional accent. Using the resulting "elderly acoustic model", we built a document-retrieval program that can be operated by voice or typing. After usability tests with 110 speakers, we tested the final system on 37 elderly speakers. Each retrieved 4 documents from a database of 86,190 Boston Globe articles, 2 by typing and 2 by speech. We measured how quickly they retrieved each article, and how much help they required. We find no difference between spoken and typed queries in either retrieval times or in amount of help required, regardless of age, gender, or computer experience. However, users perceive speech to be substantially faster, and overwhelmingly prefer speech to typing.