首页 | 官方网站   微博 | 高级检索  
     

语料对中文名词短语指代消解影响研究
引用本文:高俊伟,孔芳,朱巧明,李培峰.语料对中文名词短语指代消解影响研究[J].中文信息学报,2013,27(3):61-69.
作者姓名:高俊伟  孔芳  朱巧明  李培峰
作者单位:苏州大学 计算机科学与技术学院, 江苏省计算机信息处理技术重点实验室,江苏 苏州 215006
基金项目:国家自然科学基金资助项目,江苏省高校自然科学重大基础研究资助项目
摘    要:指代是自然语言中一种常见的语言现象,对简化语言,减少冗余有很大的作用。指代消解是用计算机找出这些指代现象的一个过程。近几年英文指代消解研究取得了很大的成就,然而,中文指代消解研究目前还较少,一方面是由于中文自然语言处理的研究起步较晚,相关的知识较少,另外一方面就是中文相关的语料库较少,目前已知的仅有ACE2005, OntoNotes等。为了探讨语料库对中文名词短语指代消解的影响,该文实现了一个基于有监督学习方法的中文名词短语指代消解平台和一个基于无监督聚类方法的中文名词短语指代消解平台,在此平台的基础上从语料库的数量和质量两个方面来探讨语料对中文名词短语指代消解的影响。

关 键 词:指代消解  名词短语  无监督  聚类  语料  

Research on the Corpus Effect to the Chinese Noun Phrase Anaphora Resolution
GAO Junwei , KONG Fang , ZHU Qiaoming , LI Peifeng.Research on the Corpus Effect to the Chinese Noun Phrase Anaphora Resolution[J].Journal of Chinese Information Processing,2013,27(3):61-69.
Authors:GAO Junwei  KONG Fang  ZHU Qiaoming  LI Peifeng
Affiliation:School of Computer Science & Technology, Soochow University,
Key Lab of Computer Information Processing Technology of Jiangsu Province, Suzhou, Jiangsu 215006,China
Abstract:Coreference is a common phenomenon in natural language, with a great effect in making the natural language clear and explicit illusions. Coreference resolution is the process to detect these phenomena by the computer. A great deal of research has been conducted on this task in English with substantial achievements in recent years. However, much less work has been done in this area in Chinese. One problem is the lack of public Chinese corpus for this research in except for ACE2005, OntoNotes and so on. To discuss the effect of the corpus to the Chinese Noun Phrase Anaphora Resolution, we present a Chinese noun phrase coreference resolution system that based on supervised learning approach and another system that based on unsupervised clustering approach. We discussed the effect of the corpus to the Chinese noun phrase coreference resolution based on the two platforms from the quantity and the quality of the corpus.
Key wordscoreference resolution; noun phrase; unsupervised; clustering; corpus
Keywords:coreference resolution  noun phrase  unsupervised  clustering  corpus  
本文献已被 万方数据 等数据库收录!
点击此处可从《中文信息学报》浏览原始摘要信息
点击此处可从《中文信息学报》下载全文
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司    京ICP备09084417号-23

京公网安备 11010802026262号