Less is More for Long Document Summary Evaluation by LLMs (2309.07382v2)

Published 14 Sep 2023 in cs.CL

Abstract: LLMs have shown promising performance in summary evaluation tasks, yet they face challenges such as high computational costs and the Lost-in-the-Middle problem where important information in the middle of long documents is often overlooked. To address these issues, this paper introduces a novel approach, Extract-then-Evaluate, which involves extracting key sentences from a long source document and then evaluating the summary by prompting LLMs. The results reveal that the proposed method not only significantly reduces evaluation costs but also exhibits a higher correlation with human evaluations. Furthermore, we provide practical recommendations for optimal document length and sentence extraction methods, contributing to the development of cost-effective yet more accurate methods for LLM-based text generation evaluation.

References (29)

Authors (5)

Yunshu Wu (4 papers)
Hayate Iso (19 papers)
Pouya Pezeshkpour (25 papers)
Nikita Bhutani (20 papers)
Estevam Hruschka (23 papers)

Citations (27)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

GitHub

GitHub - megagonlabs/llm-longeval: 💵 Code for Less is More for Long Document Summary Evaluation by LLMs (Wu, Iso et al; EACL 2024) (9 stars)

Less is More for Long Document Summary Evaluation by LLMs (2309.07382v2)

Summary

Related Papers

GitHub