We study key issues related to multilingual acoustic modeling for automatic speech recognition (ASR) through a series of large-scale ASR experiments. Our study explores shared str...
Hui Lin, Li Deng, Dong Yu, Yifan Gong, Alex Acero,...
In this paper, we propose an F0 Frame Error (FFE) metric which combines Gross Pitch Error (GPE) and Voicing Decision Error (VDE) to objectively evaluate the performance of fundame...
Local features are widely used for content-based image retrieval and object recognition. We present an efficient method for encoding digital images suitable for local feature extr...
Mina Makar, Chuo-Ling Chang, David M. Chen, Sam S....
In our previous work, we introduced Spatially Varying Transforms (SVT) for video coding, where the location of the transform block within the macroblock is not fixed but varying. ...
Cixun Zhang, Kemal Ugur, Jani Lainema, Moncef Gabb...
Abstract— This paper describes a new non-invasive brainactuated wheelchair that relies on a P300 neurophysiological protocol and automated navigation. In operation, the subject f...