The AMI system for the transcription of speech in meetings

T Hain, L Burget, J Dines, G Garau… - … , Speech and Signal …, 2007 - ieeexplore.ieee.org
T Hain, L Burget, J Dines, G Garau, V Wan, M Karafiat, J Vepa, M Lincoln
2007 IEEE International Conference on Acoustics, Speech and Signal …, 2007ieeexplore.ieee.org
This paper describes the AMI transcription system for speech in meetings developed in
collaboration by five research groups. The system includes generic techniques such as
discriminative and speaker adaptive training, vocal tract length normalisation,
heteroscedastic linear discriminant analysis, maximum likelihood linear regression, and
phone posterior based features, as well as techniques specifically designed for meeting
data. These include segmentation and cross-talk suppression, beam-forming, domain …
This paper describes the AMI transcription system for speech in meetings developed in collaboration by five research groups. The system includes generic techniques such as discriminative and speaker adaptive training, vocal tract length normalisation, heteroscedastic linear discriminant analysis, maximum likelihood linear regression, and phone posterior based features, as well as techniques specifically designed for meeting data. These include segmentation and cross-talk suppression, beam-forming, domain adaptation, Web-data collection, and channel adaptive training. The system was improved by more than 20% relative in word error rate compared to our previous system and was used in the NIST RT106 evaluations where it was found to yield competitive performance.
ieeexplore.ieee.org
以上显示的是最相近的搜索结果。 查看全部搜索结果