In Mandarin Chinese,when the noun head appears in the context,a quantity noun phrase can be reduced to a quantity phrase with the noun head omitted.This phrase structure is called elliptical quantity noun phrase.The a...In Mandarin Chinese,when the noun head appears in the context,a quantity noun phrase can be reduced to a quantity phrase with the noun head omitted.This phrase structure is called elliptical quantity noun phrase.The automatic recovery of elliptical quantity noun phrase is crucial in syntactic parsing,semantic representation and other downstream tasks.In this paper,we propose a hybrid neural network model to identify the semantic category for elliptical quantity noun phrases and realize the recovery of omitted semantics by supplementing concept categories.Firstly,we use BERT to generate character-level vectors.Secondly,Bi-LSTM is applied to capture the context information of each character and compress the input into the context memory history.Then CNN is utilized to capture the local semantics of n-grams with various granularities.Based on the Chinese Abstract Meaning Representation(CAMR)corpus and Xinhua News Agency corpus,we construct a hand-labeled elliptical quantity noun phrase dataset and carry out the semantic recovery of elliptical quantity noun phrase on this dataset.The experimental results show that our hybrid neural network model can effectively improve the performance of the semantic complement for the elliptical quantity noun phrases.展开更多
基金This research is supported by the National Science Foundation of China(Grant 61772278,author:Qu,W.Grant Number:61472191,author:Zhou,J.http://www.nsfc.gov.cn/),the National Social Science Foundation of China(Grant Number:18BYY127,author:Li B.http://www.cssn.cn),the Philosophy and Social Science Foundation of Jiangsu Higher Institution(Grant Number:2019SJA0220,author:Wei,T.https://jyt.jiangsu.gov.cn)and Jiangsu Higher Institutions’Excellent Innovative Team for Philosophy and Social Science(Grant Number:2017STD006,author:Qu,W.https://jyt.jiangsu.gov.cn)。
文摘In Mandarin Chinese,when the noun head appears in the context,a quantity noun phrase can be reduced to a quantity phrase with the noun head omitted.This phrase structure is called elliptical quantity noun phrase.The automatic recovery of elliptical quantity noun phrase is crucial in syntactic parsing,semantic representation and other downstream tasks.In this paper,we propose a hybrid neural network model to identify the semantic category for elliptical quantity noun phrases and realize the recovery of omitted semantics by supplementing concept categories.Firstly,we use BERT to generate character-level vectors.Secondly,Bi-LSTM is applied to capture the context information of each character and compress the input into the context memory history.Then CNN is utilized to capture the local semantics of n-grams with various granularities.Based on the Chinese Abstract Meaning Representation(CAMR)corpus and Xinhua News Agency corpus,we construct a hand-labeled elliptical quantity noun phrase dataset and carry out the semantic recovery of elliptical quantity noun phrase on this dataset.The experimental results show that our hybrid neural network model can effectively improve the performance of the semantic complement for the elliptical quantity noun phrases.