Author:
Wei Jianxiang,Hu Tianling,Dai Jimin,Wang Ziren,Han Pu,Huang Weidong
Abstract
Introduction: Adverse drug reactions (ADR) are directly related to public health and become the focus of public and media attention. At present, a large number of ADR events have been reported on the Internet, but the mining and utilization of such information resources is insufficient. Named entity recognition (NER) is the basic work of many natural language processing (NLP) tasks, which aims to identify entities with specific meanings from natural language texts.Methods: In order to identify entities from ADR event data resources more effectively, so as to provide valuable health knowledge for people, this paper introduces ALBERT in the input presentation layer on the basis of the classic BiLSTM-CRF model, and proposes a method of ADR named entity recognition based on the ALBERT-BiLSTM-CRF model. The textual information about ADR on the website “Chinese medical information query platform” (https://www.dayi.org.cn) was collected by the crawler and used as research data, and the BIO method was used to label three types of entities: drug name (DRN), drug component (COM), and adverse drug reactions (ADR) to build a corpus. Then, the words were mapped to the word vector by using the ALBERT module to obtain the character level semantic information, the context coding was performed by the BiLSTM module, and the label decoding was using the CRF module to predict the real label.Results: Based on the constructed corpus, experimental comparisons were made with two classical models, namely, BiLSTM-CRF and BERT-BiLSTM-CRF. The experimental results show that the F1 of our method is 91.19% on the whole, which is 1.5% and 1.37% higher than the other two models respectively, and the performance of recognition of three types of entities is significantly improved, which proves the superiority of this method.Discussion: The method proposed can be used effectively in NER from ADR information on the Internet, which provides a basis for the extraction of drug-related entity relationships and the construction of knowledge graph, thus playing a role in practical health systems such as intelligent diagnosis, risk reasoning and automatic question answering.
Funder
Major Project of Philosophy and Social Science Research in Colleges and Universities of Jiangsu Province
National Social Science Fund of China
National Natural Science Foundation of China
Subject
Pharmacology (medical),Pharmacology
Reference25 articles.
1. Long short-term memory neural networks for Chinese word segmentation;Chen,2015
2. Named entity recognition with bidirectional LSTM-CNNs;Chiu;Trans. Assoc. Comput. linguistics,2016
3. Language modeling with gated convolutional networks;Dauphin,2017
4. A survey of the usages of deep learning for natural language processing;Dw OtterMedina;IEEE Trans. neural Netw. Learn. Syst.,2020
5. Adverse drug reactions: Definitions, diagnosis, and management;Edwards;Lancet,2000
Cited by
3 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献