引用本文:王志国,宗成庆.基于高阶词汇依存的短语结构树重排序模型.软件学报,2012,23(10):2628-2642
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览 4132次   下载 6957 本文二维码信息
码上扫一扫!
分享到: 微信 更多
基于高阶词汇依存的短语结构树重排序模型
王志国, 宗成庆
模式识别国家重点实验室(中国科学院自动化研究所), 北京 100190
摘要:
在句法分析中,已有研究工作表明,词汇依存信息对短语结构句法分析是有帮助的,但是已有的研究工作都仅局限于使用一阶的词汇依存信息.提出了一种使用高阶词汇依存信息对短语结构树进行重排序的模型,该模型首先为输入句子生成有约束的搜索空间(例如,N-best 句法分析树列表或者句法分析森林),然后在约束空间内获取高阶词汇依存特征,并利用这些特征对短语结构候选树进行重排序,最终选择出最优短语结构分析树.在宾州中文树库上的实验结果表明,该模型的最高 F1 值达到了 85.74%,超过了目前在宾州中文树库上的最好结果.另外,在短语结构分析树的基础上生成的依存结构树的准确率也有了大幅提升.
关键词:  短语结构  依存结构  句法重排序  高阶词汇依存关系  句法森林
DOI:10.3724/SP.J.1001.2012.04192
分类号:
基金项目:国家自然科学基金(60975053, 61003160); 中国科学院对外合作交流项目
Phrase Parses Reranking Based on Higher-Order Lexical Dependencies
WANG Zhi-Guo, ZONG Cheng-Qing
National Laboratory of Pattern Recognition (Institute of Automation, The Chinese Academy of Sciences), Beijing 100190, China
Abstract:
The existing works on parsing show that lexical dependencies are helpful for phrase tree parsing.However, only first-order lexical dependencies have been employed and investigated in previous research. Thispaper proposes a novel method for employing higher-order lexical dependencies for phrase tree evaluation. Themethod is based on a parse reranking framework, which provides a constrained search space (via N-best lists orparse forests) and enables the parser to employ relatively complicated lexical dependency features. The models areevaluated on the UPenn Chinese Treebank. The highest F1 score reaches 85.74% and has outperformed allpreviously reported state-of-the-art systems. The dependency accuracy of phrase trees generated by the parser hasbeen significantly improved as well.
Key words:  phrase structure  dependency structure  parse reranking  higher-order lexical dependencies  parseforest

引用本文:
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览次   下载  
分享到: 微信 更多
摘要:
关键词:  
DOI:
分类号:
基金项目:
Abstract:
Key words: