Publication Details
Title: "Eigenlips" for Robust Speech Recognition
Author: C. Bregler and Y. Konig
Group: ICSI Technical Reports
Date: January 1994
PDF: ftp://ftp.icsi.berkeley.edu/pub/techreports/1994/tr-94-002.pdf
Overview:
In this study we improve the performance of a hybrid connectionist speech recognition system by incorporating visual information about the corresponding lip movements. Specifically, we investigate the benefits of adding visual features in the presence of additive noise and crosstalk (cocktail party effect). Our study extends previous experiments by using a new visual front end, and an alternative architecture for combining the visual and acoustic information. Furthermore, we have extended our recognizer to a multi-speaker, connected letters recognizer. Our results show a significant improvement for the combined architecture (acoustic and visual information) over just the acoustic system in the presence of additive noise and crosstalk.
Bibliographic Information:
ICSI Technical Report TR-94-002
Bibliographic Reference:
C. Bregler and Y. Konig. "Eigenlips" for Robust Speech Recognition. ICSI Technical Report TR-94-002, January 1994
Author: C. Bregler and Y. Konig
Group: ICSI Technical Reports
Date: January 1994
PDF: ftp://ftp.icsi.berkeley.edu/pub/techreports/1994/tr-94-002.pdf
Overview:
In this study we improve the performance of a hybrid connectionist speech recognition system by incorporating visual information about the corresponding lip movements. Specifically, we investigate the benefits of adding visual features in the presence of additive noise and crosstalk (cocktail party effect). Our study extends previous experiments by using a new visual front end, and an alternative architecture for combining the visual and acoustic information. Furthermore, we have extended our recognizer to a multi-speaker, connected letters recognizer. Our results show a significant improvement for the combined architecture (acoustic and visual information) over just the acoustic system in the presence of additive noise and crosstalk.
Bibliographic Information:
ICSI Technical Report TR-94-002
Bibliographic Reference:
C. Bregler and Y. Konig. "Eigenlips" for Robust Speech Recognition. ICSI Technical Report TR-94-002, January 1994
