Paper: | SLP-L8.5 |
Session: | Efficient Techniques for LVCSR |
Time: | Thursday, May 18, 15:20 - 15:40 |
Presentation: |
Lecture
|
Topic: |
Speech and Spoken Language Processing: Resource constrained ASR for portable/mobile devices |
Title: |
PocketSphinx: A Free, Real-Time Continuous Speech Recognition System for Hand-Held Devices |
Authors: |
David Huggins-Daines, Mohit Kumar, Arthur Chan, Alan W. Black, Mosur Ravishankar, Alexander Rudnicky, Carnegie Mellon University, United States |
Abstract: |
The availability of real-time continuous speech recognition on mobile and embedded devices has opened up a wide range of research opportunities in human-computer interactive applications. Unfortunately, most of the work in this area to date has been confined to proprietary software, or has focused on limited domains with constrained grammars. In this paper, we present a preliminary case study on the porting and optimization of CMU Sphinx-II, a popular open source large vocabulary continuous speech recognition (LVCSR) system, to hand-held devices. The resulting system operates in an average 0.87 times real-time on a 206MHz device, 8.03 times faster than the baseline system. To our knowledge, this is the first hand-held LVCSR system available under an open-source license. |