期刊论文详细信息
Information
A Framework for Word Embedding Based Automatic Text Summarization and Evaluation
TuluTilahun Hailu1  Junqing Yu1  TessfuGeteye Fantaye1 
[1] School of Computer Science and Technology, Huazhong University of Science and Technology, Wuhan 430074, China;
关键词: automatic evaluation metrics;    extrinsic evaluation;    intrinsic evaluation;    natural language processing;    text summarization;    word embedding;   
DOI  :  10.3390/info11020078
来源: DOAJ
【 摘 要 】

Text summarization is a process of producing a concise version of text (summary) from one or more information sources. If the generated summary preserves meaning of the original text, it will help the users to make fast and effective decision. However, how much meaning of the source text can be preserved is becoming harder to evaluate. The most commonly used automatic evaluation metrics like Recall-Oriented Understudy for Gisting Evaluation (ROUGE) strictly rely on the overlapping n-gram units between reference and candidate summaries, which are not suitable to measure the quality of abstractive summaries. Another major challenge to evaluate text summarization systems is lack of consistent ideal reference summaries. Studies show that human summarizers can produce variable reference summaries of the same source that can significantly affect automatic evaluation metrics scores of summarization systems. Humans are biased to certain situation while producing summary, even the same person perhaps produces substantially different summaries of the same source at different time. This paper proposes a word embedding based automatic text summarization and evaluation framework, which can successfully determine salient top-n sentences of a source text as a reference summary, and evaluate the quality of systems summaries against it. Extensive experimental results demonstrate that the proposed framework is effective and able to outperform several baseline methods with regard to both text summarization systems and automatic evaluation metrics when tested on a publicly available dataset.

【 授权许可】

Unknown   

  文献评价指标  
  下载次数:0次 浏览次数:0次