LET: Linguistic Knowledge Enhanced Graph Transformer for Chinese Short Text Matching-Reference-Cited by-同舟云学术

LET: Linguistic Knowledge Enhanced Graph Transformer for Chinese Short Text Matching

Published:2021-05-18 Issue:15 Volume:35 Page:13498-13506
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Lyu Boer,Chen Lu,Zhu Su,Yu Kai

Abstract

Chinese short text matching is a fundamental task in natural language processing. Existing approaches usually take Chinese characters or words as input tokens. They have two limitations: 1) Some Chinese words are polysemous, and semantic information is not fully utilized. 2) Some models suffer potential issues caused by word segmentation. Here we introduce HowNet as an external knowledge base and propose a Linguistic knowledge Enhanced graph Transformer (LET) to deal with word ambiguity. Additionally, we adopt the word lattice graph as input to maintain multi-granularity information. Our model is also complementary to pre-trained language models. Experimental results on two Chinese datasets show that our models outperform various typical text matching approaches. Ablation study also indicates that both semantic information and multi-granularity information are important for text matching modeling.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 23 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Second-Order Text Matching Algorithm for Agricultural Text;Applied Sciences;2024-08-09

2. bjEnet: a fast and accurate software bug localization method in natural language semantic space;Software Quality Journal;2024-07-22

3. A Sentence-Matching Model Based on Multi-Granularity Contextual Key Semantic Interaction;Applied Sciences;2024-06-14

4. JMS-QA: A Joint Hierarchical Architecture for Mental Health Question Answering;IEEE/ACM Transactions on Audio, Speech, and Language Processing;2024

5. Exploring into the Unseen: Enhancing Language-Conditioned Policy Generalization with Behavioral Information;Cyborg and Bionic Systems;2024-01