Joint Pre-Trained Chinese Named Entity Recognition Based on Bi-Directional Language Model

Changxia Ma,Chen Zhang

doi:10.1142/s0218001421530037

Abstract

The current named entity recognition (NER) is mainly based on joint convolution or recurrent neural network. In order to achieve high performance, these networks need to provide a large amount of training data in the form of feature engineering corpus and lexicons. Chinese NER is very challenging because of the high contextual relevance of Chinese characters, that is, Chinese characters and phrases may have many possible meanings in different contexts. To this end, we propose a model that leverages a pre-trained and bi-directional encoder representations-from-transformers language model and a joint bi-directional long short-term memory (Bi-LSTM) and conditional random fields (CRF) model for Chinese NER. The underlying network layer embeds Chinese characters and outputs character-level representations. The output is then fed into a bidirectional long short-term memory to capture contextual sequence information. The top layer of the proposed model is CRF, which is used to take into account the dependencies of adjacent tags and jointly decode the optimal chain of tags. A series of extensive experiments were conducted to research the useful improvements of the proposed neural network architecture on different datasets without relying heavily on handcrafted features and domain-specific knowledge. Experimental results show that the proposed model is effective, and character-level representation is of great significance for Chinese NER tasks. In addition, through this work, we have composed a new informal conversation message corpus called the autonomous bus information inquiry dataset, and compared to the advanced baseline, our method has been significantly improved.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Joint Pre-Trained Chinese Named Entity Recognition Based on Bi-Directional Language Model

Abstract

Talk to us

Similar Papers

More From: International Journal of Pattern Recognition and Artificial Intelligence

Lead the way for us

Journal: International Journal of Pattern Recognition and Artificial Intelligence	Publication Date: Apr 5, 2021
Citations: 3

Similar Papers

Named entity recognition of local adverse drug reactions in Xinjiang based on transfer learning
Keming Kang ... Long Yu
Journal of Intelligent & Fuzzy Systems | VOL. 40
Keming Kang, et. al.Keming Kang ... Long Yu
01 Jan 2020
Journal of Intelligent & Fuzzy Systems | VOL. 40

A Joint Learning Model to Extract Entities and Relations for Chinese Literature Based on Self-Attention
Li-Xin Liang ... Wu-Shao Wen
Mathematics | VOL. 10
Li-Xin Liang, et. al.Li-Xin Liang ... Wu-Shao Wen
24 Jun 2022
Mathematics | VOL. 10

A Chinese Named Entity Recognition Method Fusing Word and Radical Features
Shan Deng ... Ping Lu
-
Shan Deng, et. al.Shan Deng ... Ping Lu
23 Sep 2022
23 Sep 2022

A fine-grained Chinese word segmentation and part-of-speech tagging corpus for clinical text
Ying Xiong ... Zhongmin Wang
BMC Medical Informatics and Decision Making | VOL. 19
Ying Xiong, et. al.Ying Xiong ... Zhongmin Wang
01 Apr 2019
BMC Medical Informatics and Decision Making | VOL. 19

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Joint Pre-Trained Chinese Named Entity Recognition Based on Bi-Directional Language Model

Abstract

Talk to us

Similar Papers

More From: International Journal of Pattern Recognition and Artificial Intelligence