咨询与建议

看过本文的还看了

相关文献

该作者的其他文献

文献详情 >Attributed rhetorical structur... 收藏
arXiv

Attributed rhetorical structure grammar for domain text summarization

作     者:Lu, Ruqian Hou, Shengluan Wang, Chuanqing Huang, Yu Fei, Chaoqun Zhang, Songmao 

作者机构:Key Laboratory of MADIS Academy of Mathematics and Systems Science Chinese Academy of Sciences Beijing100190 China Key Laboratory of Intelligent Information Processing Institute of Computing Technology Chinese Academy of Sciences Beijing100190 China University of Chinese Academy of Sciences Beijing100049 China 

出 版 物:《arXiv》 (arXiv)

年 卷 期:2019年

核心收录:

主  题:Text processing 

摘      要:This paper presents a new approach of automatic text summarization which combines domain oriented text analysis (DoTA) and rhetorical structure theory (RST) in a grammar form: the attributed rhetorical structure grammar (ARSG), where the non-terminal symbols are domain keywords, called domain relations, while the rhetorical relations serve as attributes. We developed machine learning algorithms for learning such a grammar from a corpus of sample domain texts, as well as parsing algorithms for the learned grammar, together with adjustable text summarization algorithms for generating domain specific summaries. Our practical experiments have shown that with support of domain knowledge the drawback of missing very large training data set can be effectively compensated. We have also shown that the knowledge based approach may be made more powerful by introducing grammar parsing and RST as inference engine. For checking the feasibility of model transfer, we introduced a technique for mapping a grammar from one domain to others with acceptable cost. We have also made a comprehensive comparison of our approach with some others. Copyright © 2019, The Authors. All rights reserved.

读者评论 与其他读者分享你的观点

用户名:未登录
我的评分