首页>
外国专利>
Method and system for simplifying implicit rhetorical relation prediction in large scale annotated corpus
Method and system for simplifying implicit rhetorical relation prediction in large scale annotated corpus
展开▼
机译:简化大规模带注释语料中隐式修辞关系预测的方法和系统
展开▼
页面导航
摘要
著录项
相似文献
摘要
The present invention provides a method and system directed to predicting implicit rhetorical relations between two spans of text, e.g., in a large annotated corpus, such as the Penn Discourse Treebank ("PDTB"), Rhetorical Structure Theory corpus, and the Discourse Graph Bank, and particularly directed to determining a rhetorical relation in the absence of an explicit discourse marker. Surface level features may be used to capture pragmatic information encoded in the absent marker. In one manner a simplified feature set based only on raw text and semantic dependencies is used to improve performance for all relations. By using surface level features to predict implicit rhetorical relations for the large annotated corpus the invention approaches a theoretical maximum performance, suggesting that more data will not necessarily improve performance based on these and similarly situated features.
展开▼