Extracting Chinese events with a joint label space model.

Wenzhi Huang,Donghong Ji,Junchi Zhang,Fu Lee Wang

doi:10.1371/journal.pone.0272353

Wenzhi Huang, Donghong Ji + Show 2 more

Open Access

https://doi.org/10.1371/journal.pone.0272353

Copy DOI

Journal: PloS one	Publication Date: Sep 27, 2022
Citations: 1	License type: CC BY 4.0

Affiliation: Wuhan University, Wuhan Institute of Technology

Abstract

The task of event extraction consists of three subtasks namely entity recognition, trigger identification and argument role classification. Recent work tackles these subtasks jointly with the method of multi-task learning for better extraction performance. Despite being effective, existing attempts typically treat labels of event subtasks as uninformative and independent one-hot vectors, ignoring the potential loss of useful label information, thereby making it difficult for these models to incorporate interactive features on the label level. In this paper, we propose a joint label space framework to improve Chinese event extraction. Specifically, the model converts labels of all subtasks into a dense matrix, giving each Chinese character a shared label distribution via an incrementally refined attention mechanism. Then the learned label embeddings are also used as the weight of the output layer for each subtask, hence adjusted along with model training. In addition, we incorporate the word lexicon into the character representation in a soft probabilistic manner, hence alleviating the impact of word segmentation errors. Extensive experiments on Chinese and English benchmarks demonstrate that our model outperforms state-of-the-art methods.

Full Text