Sciweavers

5942 search results - page 887 / 1189
» On the Processing of Decimated Signals
Sort
View
347
Voted
ICASSP
2011
IEEE
14 years 11 months ago
Deep Belief Networks using discriminative features for phone recognition
Deep Belief Networks (DBNs) are multi-layer generative models. They can be trained to model windows of coefficients extracted from speech and they discover multiple layers of fea...
Abdel-rahman Mohamed, Tara N. Sainath, George Dahl...
198
Voted
ICASSP
2011
IEEE
14 years 11 months ago
Front-end feature transforms with context filtering for speaker adaptation
Feature-space transforms such as feature-space maximum likelihood linear regression (FMLLR) are very effective speaker adaptation technique, especially on mismatched test data. In...
Jing Huang, Karthik Visweswariah, Peder A. Olsen, ...
ICASSP
2011
IEEE
14 years 11 months ago
Video object tracking with differential Structural SIMilarity index
The Structural SIMilarity Measure (SSIM) combined with the sequential Monte Carlo approach has been shown [1] to achieve more reliable video object tracking performance, compared ...
Artur Loza, Fanglin Wang, Jie Yang, Lyudmila Mihay...
ICASSP
2011
IEEE
14 years 11 months ago
Multi-view and multi-objective semi-supervised learning for large vocabulary continuous speech recognition
Current hidden Markov acoustic modeling for large vocabulary continuous speech recognition (LVCSR) relies on the availability of abundant labeled transcriptions. Given that speech...
Xiaodong Cui, Jing Huang, Jen-Tzung Chien
190
Voted
ICASSP
2011
IEEE
14 years 11 months ago
The MIT LL 2010 speaker recognition evaluation system: Scalable language-independent speaker recognition
Research in the speaker recognition community has continued to address methods of mitigating variational nuisances. Telephone and auxiliary-microphone recorded speech emphasize th...
Douglas E. Sturim, William M. Campbell, Najim Deha...