Transformers in Automatic Speech Recognition-Reference-Cited by-同舟云学术

Transformers in Automatic Speech Recognition

Published:2023 Issue: Volume: Page:123-139
ISSN:0302-9743
Container-title:Human-Centered Artificial Intelligence
language:
Short-container-title:

Author:

Evrard Marc^ORCID

Publisher

Springer International Publishing

Link

https://link.springer.com/content/pdf/10.1007/978-3-031-24349-3_8

Reference49 articles.

1. Auli, M.: Wav2vec: self-supervised learning of speech representations. Talk at MIT, CMU, U of Edinburgh, Spring 2021 (2021)

2. Baevski, A., Schneider, S., Auli, M.: VQ-wav2vec: self-supervised learning of discrete speech representations. In: International Conference on Learning Representations (2019)

3. Baevski, A., Zhou, Y., Mohamed, A., Auli, M.: Wav2vec 2.0: a framework for self-supervised learning of speech representations. In: Advances in Neural Information Processing Systems, vol. 33, pp. 12449–12460 (2020)

4. Bello, I., Zoph, B., Vaswani, A., Shlens, J., Le, Q.V.: Attention augmented convolutional networks. In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 3286–3295 (2019)

5. Chan, W., Jaitly, N., Le, Q., Vinyals, O.: Listen, attend and spell: a neural network for large vocabulary conversational speech recognition. In: 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 4960–4964. IEEE (2016)

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Assessing Speech Intelligibility and Severity Level in Parkinson's Disease Using Wav2Vec 2.0;2024 47th International Conference on Telecommunications and Signal Processing (TSP);2024-07-10

2. Time domain speech enhancement with CNN and time-attention transformer;Digital Signal Processing;2024-04

3. Vision and Multi-modal Transformers;Human-Centered Artificial Intelligence;2023