SuperLectures.com

FRONT-END FEATURE TRANSFORMS WITH CONTEXT FILTERING FOR SPEAKER ADAPTATION

Adaptation for ASR

Full Paper at IEEE Xplore

Presented by: Steven Rennie, Author(s): Jing Huang, IBM T.J. Watson Research Center, United States; Karthik Visweswariah, IBM India Research, India; Peder Olsen, Vaibhava Goel, IBM T.J. Watson Research Center, United States

Feature-space transforms such as feature-space maximum likelihood linear regression (FMLLR) are very effective speaker adaptation technique, especially on mismatched test data. In this study, we extend the full-rank square matrix of FMLLR to a non-square matrix that use neighboring feature vectors in estimating the adapted central feature vector. Through optimizing an appropriate objective function we aim to filter out and transform features through the correlation of the feature context. We compare to FMLLR that just consider the current feature vector only. Our experiments are conducted on the automobile data with different speed conditions. Results show that context filtering improves 23% on word error rate over conventional FMLLR on noisy 60mph data with adapted ML model, and 7%/9% improvement over the discriminatively trained FMMI/BMMI models.


  Speech Transcript

|

  Slides

Enlarge the slide | Show all slides in a pop-up window

0:00:16

  1. slide

0:01:05

  2. slide

0:02:08

  3. slide

0:02:42

  4. slide

0:06:05

  5. slide

0:07:28

  6. slide

0:08:22

  7. slide

0:08:48

  8. slide

0:09:35

  9. slide

0:10:07

 10. slide

0:10:42

 11. slide

0:11:28

 12. slide

0:11:44

 13. slide

0:12:29

 14. slide

0:12:48

 15. slide

0:14:24

    11. slide

0:16:31

    13. slide

  Comments

Please sign in to post your comment!

  Lecture Information

Recorded: 2011-05-24 16:15 - 16:35, Panorama
Added: 15. 6. 2011 14:44
Number of views: 78
Video resolution: 1024x576 px, 512x288 px
Video length: 0:17:10
Audio track: MP3 [5.79 MB], 0:17:10