Pith. sign in

REVIEW 1 cited by

ROUGE-K: Do Your Summaries Have Keywords?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.05186 v1 pith:ZVJLOUOL submitted 2024-03-08 cs.CL

classification cs.CL
keywords summarieskeywordsevaluationfindimportantincludeinformationmetric
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Keywords, that is, content-relevant words in summaries play an important role in efficient information conveyance, making it critical to assess if system-generated summaries contain such informative words during evaluation. However, existing evaluation metrics for extreme summarization models do not pay explicit attention to keywords in summaries, leaving developers ignorant of their presence. To address this issue, we present a keyword-oriented evaluation metric, dubbed ROUGE-K, which provides a quantitative answer to the question of -- \textit{How well do summaries include keywords?} Through the lens of this keyword-aware metric, we surprisingly find that a current strong baseline model often misses essential information in their summaries. Our analysis reveals that human annotators indeed find the summaries with more keywords to be more relevant to the source documents. This is an important yet previously overlooked aspect in evaluating summarization systems. Finally, to enhance keyword inclusion, we propose four approaches for incorporating word importance into a transformer-based model and experimentally show that it enables guiding models to include more keywords while keeping the overall quality. Our code is released at https://github.com/sobamchan/rougek.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory

    cs.LG 2025-01 conditional novelty 6.0 of 10

    RevertLM reconstructs private text from intermediate LLM representations in split learning using a projection into embedding space plus a generative decoder, outperforming prior embedding-only attacks.

Pith tools