Many object classes, including human faces, can be modeled as a set of characteristic parts arranged in a variable spatial con guration. We introduce a simpli ed model of a deforma...
This paper investigates the problem of incorporating auxiliary information (e.g. pitch) for speech recognition using dynamic Bayesian networks (DBNs). Previous works usually model...
In this paper, a multilinear formulation of the popular Principal Component Analysis (PCA) is proposed, named as multilinear PCA (MPCA), where the input can be not only vectors, b...
Anastasios N. Venetsanopoulos, Haiping Lu, Konstan...
Speech can be represented as a time/frequency distribution of energy using a multi-band filter bank. A Markov random field model, which takes into account the possible time asynch...
One of the biggest challenges in speaker recognition is dealing with speaker-emotion variability. The basic problem is how to train the emotion GMMs of the speakers from their neu...