Approximating Probabilistic Models as Weighted Finite Automata-Reference-Cited by-同舟云学术

Approximating Probabilistic Models as Weighted Finite Automata

Published:2021-05-20 Issue: Volume: Page:1-34
ISSN:0891-2017
Container-title:Computational Linguistics
language:en
Short-container-title:

Author:

Suresh Ananda Theertha¹,Roark Brian²,Riley Michael³,Schogol Vlad⁴

Affiliation:

1. Google Research. theertha@google.com

2. Google Research. roark@google.com

3. Google Research. riley@google.com

4. Google Research. vlads@google.com

Abstract

Abstract Weighted finite automata (WFAs) are often used to represent probabilistic models, such as ngram language models, because among other things, they are efficient for recognition tasks in time and space. The probabilistic source to be represented as a WFA, however, may come in many forms. Given a generic probabilistic model over sequences, we propose an algorithm to approximate it as a WFA such that the Kullback-Leibler divergence between the source model and the WFA target model is minimized. The proposed algorithm involves a counting step and a difference of convex optimization step, both of which can be performed efficiently.We demonstrate the usefulness of our approach on various tasks, including distilling n-gram models from neural models, building compact language models, and building open-vocabulary character models. The algorithms used for these experiments are available in an open-source software library.

Publisher

MIT Press - Journals

Subject

Artificial Intelligence,Computer Science Applications,Linguistics and Language,Language and Linguistics

Link

http://direct.mit.edu/coli/article-pdf/doi/10.1162/coli_a_00401/1919755/coli_a_00401.pdf

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. VQ-T: RNN Transducers using Vector-Quantized Prediction Network States;Interspeech 2022;2022-09-18