Cargando…
Audiovisual Tracking of Multiple Speakers in Smart Spaces
This paper presents GAVT, a highly accurate audiovisual 3D tracking system based on particle filters and a probabilistic framework, employing a single camera and a microphone array. Our first contribution is a complex visual appearance model that accurately locates the speaker’s mouth. It transforms...
Autores principales: | Sanabria-Macias, Frank, Marron-Romera, Marta, Macias-Guarasa, Javier |
---|---|
Formato: | Online Artículo Texto |
Lenguaje: | English |
Publicado: |
MDPI
2023
|
Materias: | |
Acceso en línea: | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10422319/ https://www.ncbi.nlm.nih.gov/pubmed/37571754 http://dx.doi.org/10.3390/s23156969 |
Ejemplares similares
-
Headgear Accessories Classification Using an Overhead Depth Sensor
por: Luna, Carlos A., et al.
Publicado: (2017) -
Source Localization with Acoustic Sensor Arrays Using Generative Model Based Fitting with Sparse Constraints
por: Velasco, Jose, et al.
Publicado: (2012) -
Nonnative Audiovisual Speech Perception in Noise: Dissociable Effects of the Speaker and Listener
por: Xie, Zilong, et al.
Publicado: (2014) -
Towards End-to-End Acoustic Localization Using Deep Learning: From Audio Signals to Source Position Coordinates
por: Vera-Diaz, Juan Manuel, et al.
Publicado: (2018) -
Stereo Vision Tracking of Multiple Objects in Complex Indoor Environments
por: Marrón-Romera, Marta, et al.
Publicado: (2010)