An Improvement in Audio-Visual Voice Activity Detection for Automatic Speech Recognition-Reference-Cited by-同舟云学术

An Improvement in Audio-Visual Voice Activity Detection for Automatic Speech Recognition

Published:2010 Issue: Volume: Page:51-61
ISSN:0302-9743
Container-title:Trends in Applied Intelligent Systems
language:
Short-container-title:

Author:

Yoshida Takami,Nakadai Kazuhiro,Okuno Hiroshi G.

Publisher

Springer Berlin Heidelberg

Link

http://link.springer.com/content/pdf/10.1007/978-3-642-13022-9_6

Reference16 articles.

1. Nakadai, K., Lourens, T., Okuno, H.G., Kitano, H.: Active audition for humanoid. In: Proceedings of 17th National Conference on Artificial Intelligence, pp. 832–839 (2000)

2. Yamamoto, S., Nakadai, K., Nakano, M., Tsujino, H., Valin, J.M., Komatani, K., Ogata, T., Okuno, H.G.: Real-time robot audition system that recognizes simultaneous speech in the real world. In: Proceedings of IEEE/RSJ International Conference on Intelligent Robots and Systems, pp. 5333–5338 (2006)

3. Potamianos, G., Neti, C., Iyengar, G., Senior, A., Verma, A.: A cascade visual front end for speaker independent automatic speechreading. Speech Technology, Special Issue on Multimedia 4, 193–208 (2001)

4. Tamura, S., Iwano, K., Furui, S.: A stream-weight optimization method for multi-stream hmms based on likelihood value normalization. In: Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing, pp. 469–472 (2005)

5. Fiscus, J.: A post-processing systems to yield reduced word error rates: Recognizer Output Voting Error Reduction (ROVER). In: Proceedings of Workshop on Automatic Speech Recognition and Understanding, pp. 347–354 (1997)

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. An End-to-End Multimodal Voice Activity Detection Using WaveNet Encoder and Residual Networks;IEEE Journal of Selected Topics in Signal Processing;2019-05

2. A deep architecture for audio-visual voice activity detection in the presence of transients;Signal Processing;2018-01

3. Kernel-Based Sensor Fusion With Application to Audio-Visual Voice Activity Detection;IEEE Transactions on Signal Processing;2016-12-15

4. Complementary Sociorobot Issues;Intelligent Systems, Control and Automation: Science and Engineering;2015-07-16