Multilingual Part-of-Speech Tagging: Two Unsupervised Approaches-Reference-Cited by-同舟云学术

Multilingual Part-of-Speech Tagging: Two Unsupervised Approaches

Published:2009-11-17 Issue: Volume:36 Page:341-385
ISSN:1076-9757
Container-title:Journal of Artificial Intelligence Research
language:
Short-container-title:jair

Author:

Naseem T.,Snyder B.,Eisenstein J.,Barzilay R.

Abstract

We demonstrate the effectiveness of multilingual learning for unsupervised part-of-speech tagging. The central assumption of our work is that by combining cues from multiple languages, the structure of each becomes more apparent. We consider two ways of applying this intuition to the problem of unsupervised part-of-speech tagging: a model that directly merges tag structures for a pair of languages into a single sequence and a second model which instead incorporates multilingual context using latent variables. Both approaches are formulated as hierarchical Bayesian models, using Markov Chain Monte Carlo sampling techniques for inference. Our results demonstrate that by incorporating multilingual evidence we can achieve impressive performance gains across a range of scenarios. We also found that performance improves steadily as the number of available languages increases.

Publisher

AI Access Foundation

Subject

Artificial Intelligence

Cited by 14 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. An Automatic POS Tagger System for Code Mixed Indian Social Media Text;Communications in Computer and Information Science;2023-11-30

2. Knowledge Transfer via Word Alignment and Its Application to Vietnamese POS Tagging;Computational Data and Social Networks;2023

3. Bilingual Corpus-based Hybrid POS Tagger for Low Resource Tamil Language: A Statistical approach;Journal of Intelligent & Fuzzy Systems;2022-11-11

4. An Information Theoretic Approach to Symbolic Learning in Synthetic Languages;Entropy;2022-02-10

5. Incorporating Typological Features into Language Selection for Multilingual Neural Machine Translation;Web and Big Data;2021