FaVoA: Face-Voice Association Favours Ambiguous Speaker Detection-Reference-Cited by-同舟云学术

FaVoA: Face-Voice Association Favours Ambiguous Speaker Detection

Published:2021 Issue: Volume: Page:439-450
ISSN:0302-9743
Container-title:Lecture Notes in Computer Science
language:
Short-container-title:

Author:

Carneiro Hugo,Weber Cornelius,Wermter Stefan

Publisher

Springer International Publishing

Link

https://link.springer.com/content/pdf/10.1007/978-3-030-86362-3_36

Reference17 articles.

1. Alcázar, J.L., et al.: Active speakers in context. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2020)

2. Arevalo, J., Solorio, T., Montes-y-Gómez, M., González, F.A.: Gated multimodal units for information fusion. In: 5th International Conference on Learning Representations, ICLR 2017, Workshop Track Proceedings (2017). OpenReview.net

3. Bahrick, L.E., Hernandez-Reif, M., Flom, R.: The development of infant learning about specific face-voice relations. Dev. Psychol. 41(3), 541–552 (2005)

4. Cho, K., et al.: Learning phrase representations using RNN encoder-decoder for statistical machine translation. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), Doha, Qatar, pp. 1724–1734. Association for Computational Linguistics (2014)

5. Choi, H.S., Park, C., Lee, K.: From inference to generation: end-to-end fully self-supervised generation of human face from speech. In: 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, 26–30 April 2020 (2020). OpenReview.net

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. AS-Net: active speaker detection using deep audio-visual attention;Multimedia Tools and Applications;2024-02-05

2. Whose emotion matters? Speaking activity localisation without prior knowledge;Neurocomputing;2023-08

3. A Trained Humanoid Robot can Perform Human-Like Crossmodal Social Attention and Conflict Resolution;International Journal of Social Robotics;2023-04-02

4. End-to-End Active Speaker Detection;Lecture Notes in Computer Science;2022