Character-Aware Neural Language Models-Reference-Cited by-同舟云学术

Character-Aware Neural Language Models

Published:2016-03-05 Issue:1 Volume:30 Page:
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Kim Yoon,Jernite Yacine,Sontag David,Rush Alexander

Abstract

We describe a simple neural language model that relies only on character-level inputs. Predictions are still made at the word-level. Our model employs a convolutional neural network (CNN) and a highway net work over characters, whose output is given to a long short-term memory (LSTM) recurrent neural network language model (RNN-LM). On the English Penn Treebank the model is on par with the existing state-of-the-art despite having 60% fewer parameters. On languages with rich morphology (Arabic, Czech, French, German, Spanish, Russian), the model outperforms word-level/morpheme-level LSTM baselines, again with fewer parameters. The results suggest that on many languages, character inputs are sufficient for language modeling. Analysis of word representations obtained from the character composition part of the model reveals that the model is able to encode, from characters only, both semantic and orthographic information.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 109 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Geometry-Aware Weight Perturbation for Adversarial Training;Electronics;2024-09-04

2. A multi-model attention based CNN-BiLSTM model for personality traits prediction based on user behavior on social media;Knowledge-Based Systems;2024-09

3. FedRoLA: Robust Federated Learning Against Model Poisoning via Layer-based Aggregation;Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining;2024-08-24

4. Normalization of Web of Science Institution Names Based on Deep Learning;Algorithms;2024-07-14

5. Custom Natural Language Understanding for Healthcare Chatbots and A Case Study;2024 IEEE International Conference on Digital Health (ICDH);2024-07-07