错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Investigating a Semantic Similarity Loss Function for the Parallel Training of Abstractive and Extractive Scientific Document Summarizers

  • Sudipta Singha Roy,
  • Robert E. Mercer

摘要

Scientific document summarization focusses on condensing scientific literature, research papers, and technical documents into concise summaries while preserving crucial scientific concepts, findings, and conclusions. In this work, we present a novel loss function that incorporates semantic similarity, and use it in the parallel training of extractive and abstractive summarizers, thereby improving the performance of the individual summarizer units. The new loss function is a union of the summarizer cross-entropy losses and the semantic similarity losses among the generated and reference summaries. To validate the effectiveness of the proposed loss function joint with the parallel training, the experiments use a combination of four recently state-of-the-art extractive summarizers and four recently state-of-the-art abstractive summarizers. Results indicate that for all combinations, the extractive and abstractive summarizers both gain significant performance boosts. It is conjectured that the new semantic similarity-induced cross-entropy loss combined with the parallel training will improve any combination of quality extractive and abstractive summarizers.