CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP-Reference-Cited by-同舟云学术

CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP

Published:2020-07 Issue: Volume: Page:
ISSN:
Container-title:Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence
language:
Short-container-title:

Author:

Qin Libo¹,Ni Minheng¹,Zhang Yue²³,Che Wanxiang¹

Affiliation:

1. Research Center for Social Computing and Information Retrieval, Harbin Institute of Technology

2. Westlake University

3. Institute of Advanced Technology, Westlake Institute for Advanced Study

Abstract

Multi-lingual contextualized embeddings, such as multilingual-BERT (mBERT), have shown success in a variety of zero-shot cross-lingual tasks. However, these models are limited by having inconsistent contextualized representations of subwords across different languages. Existing work addresses this issue by bilingual projection and fine-tuning technique. We propose a data augmentation framework to generate multi-lingual code-switching data to fine-tune mBERT, which encourages model to align representations from source and multiple target languages once by mixing their context information. Compared with the existing work, our method does not rely on bilingual sentences for training, and requires only one training process for multiple target languages. Experimental results on five tasks with 19 languages show that our method leads to significantly improved performances for all the tasks compared with mBERT.

Publisher

International Joint Conferences on Artificial Intelligence Organization

Cited by 33 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Intent detection and slot filling for Persian: Cross-lingual training for low-resource languages;Natural Language Processing;2024-09-06

2. Use of prompt-based learning for code-mixed and code-switched text classification;World Wide Web;2024-09

3. Mixture-of-languages Routing for Multilingual Dialogues;ACM Transactions on Information Systems;2024-08-05

4. Enhancing Cross-Lingual Sarcasm Detection by a Prompt Learning Framework with Data Augmentation and Contrastive Learning;Electronics;2024-06-01

5. Improving Multilingual and Code-Switching ASR Using Large Language Model Generated Text;2023 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU);2023-12-16