Hybrid Neural Network for Automatic Recovery of Elliptical Chinese Quantity Noun Phrases.

Autor: Hanyu Shi, Weiguang Qu, Tingxin Wei, Junsheng Zhou, Yunfei Long, Yanhui Gu, Bin Li
Předmět:
Zdroj: Computers, Materials & Continua; 2021, Vol. 69 Issue 3, p414-4127, 15p
Abstrakt: In Mandarin Chinese, when the noun head appears in the context, a quantity noun phrase can be reduced to a quantity phrase with the noun head omitted. This phrase structure is called elliptical quantity noun phrase. The automatic recovery of elliptical quantity noun phrase is crucial in syntactic parsing, semantic representation and other downstream tasks. In this paper, we propose a hybrid neural network model to identify the semantic category for elliptical quantity noun phrases and realize the recovery of omitted semantics by supplementing concept categories. Firstly, we use BERT to generate character-level vectors. Secondly, Bi-LSTM is applied to capture the context information of each character and compress the input into the context memory history. Then CNN is utilized to capture the local semantics of n-gramswith various granularities. Based on the ChineseAbstractMeaning Representation (CAMR) corpus and Xinhua News Agency corpus, we construct a hand-labeled elliptical quantity noun phrase dataset and carry out the semantic recovery of elliptical quantity noun phrase on this dataset. The experimental results show that our hybrid neural network model can effectively improve the performance of the semantic complement for the elliptical quantity noun phrases. [ABSTRACT FROM AUTHOR]
Databáze: Complementary Index