Building models of the structure in musical signals raises the question of how to evaluate and compare different modeling approaches. One possibility is to use the model to impute...
Thierry Bertin-Mahieux, Graham Grindlay, Ron J. We...
This paper presents a Bayesian method for temporally aligning a music score and an audio rendition. A critical problem in audio-toscore alignment is in dealing with the wide varie...
Akira Maezawa, Hiroshi G. Okuno, Tetsuya Ogata, Ma...
Stereoscopic video is an important manner for 3-D video applications, and robust stereoscopic video transmission has posed a technical challenge for stereoscopic video coding. In ...
Spoken Language Understanding aims at mapping a natural language spoken sentence into a semantic representation. In the last decade two main approaches have been pursued: generati...
Marco Dinarelli, Alessandro Moschitti, Giuseppe Ri...
We study key issues related to multilingual acoustic modeling for automatic speech recognition (ASR) through a series of large-scale ASR experiments. Our study explores shared str...
Hui Lin, Li Deng, Dong Yu, Yifan Gong, Alex Acero,...