Pith. sign in

Paper Citation Record · LEDGER

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization

As of 12 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2604.25130.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.25130 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T16:43:03.770431Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-09T03:36:57.168246Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T03:45:55.345415Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact5
  • verified fuzzy42
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b9716f39-905d-423a-8545-11e30bb5a7af · outbound

This paper cites A comprehensive survey for automatic text summarization: Techniques, approaches and perspectives.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization A comprehensive survey for automatic text summarization: Techniques, approaches and perspectives

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.389064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:4ea3c93ca9e0b979887fdd740ab009e87dd61309977aa9caca26cba4cf2a0a58

Observation ae242ad5-16ba-4e17-a0e0-ccfa1f2dcebc · outbound

This paper cites A comprehensive survey on automatic text summarization with exploration of llm-based methods.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization A comprehensive survey on automatic text summarization with exploration of llm-based methods

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.349914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:0ab4d0f1700d1510b73b7c8a1d5f79b0ecc126b7533f1c700634fcd0842276cf

Observation 61d86dbf-f32a-4f0b-90b7-f14ab3c6434a · outbound

This paper cites How far are we from robust long abstractive summarization?.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization How far are we from robust long abstractive summarization?

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.370982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:5e4581d81244ec770ebf0a4915880fa7937dd679763334d7b3b8019bd99be72a

Observation 655e998f-666f-4401-a684-de797c739b3d · outbound

This paper cites Current and future state of evaluation of large language models for medical summarization tasks.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Current and future state of evaluation of large language models for medical summarization tasks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.366674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:4c6aeb7847d7c4fe79c288a77125c6669991d2a31a1a4700dd25d5ef8126f143

Observation a346ce2e-bf56-4383-ba7b-0db3dda35e17 · outbound

This paper cites Effectiveness in retrieving legal precedents: exploring text summarization and cutting-edge lan- guage models toward a cost-efficient approach.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Effectiveness in retrieving legal precedents: exploring text summarization and cutting-edge lan- guage models toward a cost-efficient approach

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.393385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:44cba4264d98eefbcd367695c66688c52409cb43b28961d6e0b2cf8a982f95e1

Observation cbdcefff-1dcb-4065-830f-6301bafe704b · outbound

This paper cites The power of graphs in medicine: Introducing biographsum for effective text summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization The power of graphs in medicine: Introducing biographsum for effective text summarization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.382071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:cf31a38ab2c1e25c35e11566af0f7adddd5c865ff9a455fbb9221d905e62cf54

Observation 2cdead4b-aada-4d92-ae4c-0b21de368553 · outbound

This paper cites A Comparative Study of Quality Evaluation Methods for Text Summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization A Comparative Study of Quality Evaluation Methods for Text Summarization

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:18.137617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:6c8a368adcaf9d1b3f655e037a8a987e5cb6f81beb727e1a5ccdf6f7fa452428

Observation 3f28253e-00da-4116-8cd3-68f8f79a34d8 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Rouge: A package for automatic evaluation of summaries

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.354456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:0651f88be44a550840d4f0f3c81818d8b2075515684911a753f98a9137b409bf

Observation 7e40da3b-8f7f-45d9-afa0-2038ac8a8617 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Bleu: a method for automatic evaluation of machine translation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.378227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:e074363371c1a7978b858bdb826095e0909b44ad6b64c89b2d7fc91dd7c7d9dc

Observation 512e7586-5381-42d8-a6f7-35ec0860aef5 · outbound

This paper cites Bertscore: Evaluating text generation with bert.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Bertscore: Evaluating text generation with bert

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.397108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:6d331b19de5053d0074c82fe67b48e08c22ff5fbcad0b32c5343109ed533492d

Observation 88d97bd6-594b-4e8b-8aeb-4078bbaddd2a · outbound

This paper cites Summeval: Re-evaluating summarization evaluation.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Summeval: Re-evaluating summarization evaluation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.326605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:cc5ca342ce50d013890ec894d8103eb7767d0a63c85707d8f1ff86d12aad8a2e

Observation b6f9fec6-8967-4485-ad3b-da686748b09b · outbound

This paper cites Towards question- answering as an automatic metric for evaluating the content quality of a summary.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Towards question- answering as an automatic metric for evaluating the content quality of a summary

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.361712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:77947f056fdab2564ec8332e3739a6f149433f8d3aae411eccfb64db9bc47178

Observation 6051f8ae-68ec-4407-ab2d-2f439db69512 · outbound

This paper cites News summarization and evaluation in the era of gpt-3.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization News summarization and evaluation in the era of gpt-3

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.357719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:ec6473b0d5681850a75972f3e52e19d7f3392d0dc5dfcd917332c068886fdfe7

Observation 2de151eb-6e13-4726-9b1a-eb8b04dde426 · outbound

This paper cites FABLES: Evaluating faithfulness and content selection in book-length summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization FABLES: Evaluating faithfulness and content selection in book-length summarization

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:18.044669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:dcaace3f99e2140dd25d0f4d03bf4d7b9bb573bc03c563a1903593df69918532

Observation 3985e453-96f7-43d7-b8c0-cfc841bbe814 · outbound

This paper cites Quality evaluation of summarization models for patent documents.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Quality evaluation of summarization models for patent documents

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.210103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:3219684f88d915f70f012eade9ef9bb8c8b3d56b064a96a4baddaa1c40ca76f2

Observation e8989c75-4737-4600-8f53-1bf686198a21 · outbound

This paper cites Questeval: Summarization asks for fact-based evalu- ation.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Questeval: Summarization asks for fact-based evalu- ation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.213160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:f0565e2463b4f36d44727bee1f49586a2ca7820e68ea2d9da964eecc2f88dde2

Observation ef74e3dd-7a3f-451e-9774-85cf45086168 · outbound

This paper cites Asking and answering questions to evaluate the factual consistency of summaries.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Asking and answering questions to evaluate the factual consistency of summaries

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.291310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:9145be0f086e0c83a7bb86d2cb3841f11d9557875a46158ba6972d394a40d26b

Observation d2163ec1-1f22-4b08-859f-c7bbec4f107f · outbound

This paper cites Evaluation of question-answering based text summarization using llm invited paper.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Evaluation of question-answering based text summarization using llm invited paper

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.230422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:5ceca3974adae6b12603ef560cdd485adc2c10b9f2a42e10649a953f835f97cd

Observation 76051203-de85-463a-b27e-5ec813469dce · outbound

This paper cites q 2: Evaluating factual consistency in knowledge-grounded dialogues via question generation and question answering.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization q 2: Evaluating factual consistency in knowledge-grounded dialogues via question generation and question answering

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.323364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:87a2f01d985aa599731610ecd8e8da9ec48a18b54d0735bda3fc2a8db1746602

Observation 72554925-45bc-4377-8938-7aee10d87dfd · outbound

This paper cites Qapyramid: Fine-grained evaluation of content selection for text sum- marization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Qapyramid: Fine-grained evaluation of content selection for text sum- marization

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.312751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:1716ae963112267691e7c48e54c0aa6c5013c3bc0ea3fdcdd11582b060ee5655

Observation 2f83d3b3-8d87-4729-b81f-4d6493085a39 · outbound

This paper cites Feqa: A question answering evaluation framework for faithfulness assessment in abstractive summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Feqa: A question answering evaluation framework for faithfulness assessment in abstractive summarization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.216455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:b177c69a78f81c00ccdbe1be705ce40eb2eb6dc30846f1be01e3b70b33ef11c5

Observation 3a7223f0-3be5-4f98-a4d0-f2b950043991 · outbound

This paper cites On faithfulness and factuality in abstractive summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization On faithfulness and factuality in abstractive summarization

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.315995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:7aab9f4295f56709f0d4e711ce898a1ce2b3c79b839033195a17a869c965ed80

Observation 7d09c09b-4f66-4384-a753-45140489581d · outbound

This paper cites Faithful to the original: Fact aware neural abstractive summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Faithful to the original: Fact aware neural abstractive summarization

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.302131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:ea7edca39c6e33c0a1a7f2dc2404d1954e2dd482e34e4e2263d141d041864078

Observation 02930183-7536-4ef9-b6ed-e86f7fe22442 · outbound

This paper cites Evaluating the factual consistency of abstractive text summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Evaluating the factual consistency of abstractive text summarization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.306045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:109db5781c5d6f26b3556e33519029e35a2eee1f38fd23290c5ada6b54191766

Observation 189fc783-3aaa-4dd4-9591-248f65da7514 · outbound

This paper cites Scaling Up Summarization: Leveraging Large Language Models for Long Text Extractive Summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Scaling Up Summarization: Leveraging Large Language Models for Long Text Extractive Summarization

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:17.859134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:5efbd9ee6d0be0d1e0b0faeb28a334bd637dac39e38a6f7ed729a4f980732909

Observation b6cdc995-4553-4092-914b-73263947a9a6 · outbound

This paper cites Self-refine: Iter- ative refinement with self-feedback.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Self-refine: Iter- ative refinement with self-feedback

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.297920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:1a79db9afd6c032d74d62a450eead8db1ca241d1dd387665997779dace80d364

Observation 9832d6d9-a002-4cbb-b189-a2bd94b6513a · outbound

This paper cites Teaching large language models to self-debug.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Teaching large language models to self-debug

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.309158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:501299f8cdf60b7dc163c5053e77cc8e716c8539917f8bacdc02d3170a53dc6b

Observation 36653993-8dfd-4ed4-b48d-6231df3a6290 · outbound

This paper cites Prometheus: Inducing fine-grained evaluation capability in language models.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Prometheus: Inducing fine-grained evaluation capability in language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.226993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:a332d63ef10869ff6353c4e924d34c51635b7deb3956105c04cfe7af140465e1

Observation 49952e71-ed43-48ec-8404-05f2e4b794e3 · outbound

This paper cites Summit: Iterative text summarization via chatgpt.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Summit: Iterative text summarization via chatgpt

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.223387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:492750cce974fcab9368871d8cf8f79875839d709685f7e02a9c18d196129b29

Observation 6654a19e-2e52-42d1-8c40-f3a6bda3b2a8 · outbound

This paper cites Multi-fact correction in abstractive text summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Multi-fact correction in abstractive text summarization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.265946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:b1104ec7aa9f7d491cdc9adc862366610ec1315485787051d1a5fc5d6d9e8fb3

Observation 4a71d314-1902-4911-a8d1-6760dca816e3 · outbound

This paper cites Meta-rewarding language models: Self-improving alignment with llm-as-a-meta-judge.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Meta-rewarding language models: Self-improving alignment with llm-as-a-meta-judge

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.255292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:f27d88552440a35e2e55ad495d6b33ff21b194af183f0381a586b834a646adda

Observation 8139034c-05c3-4d31-a2c6-4a16d4101952 · outbound

This paper cites The lighthouse of language: Enhancing llm agents via critique-guided improvement.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization The lighthouse of language: Enhancing llm agents via critique-guided improvement

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:18.132029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:7e9be1d96fbf6e289346b89c8ca5429ef0baea89b45a81732791214794bda1a9

Observation 181e2903-8da0-4d6e-9530-48ee170c12e7 · outbound

This paper cites Large language models improve clinical decision making of medical students through patient simulation and structured feedback: a randomized controlled trial.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Large language models improve clinical decision making of medical students through patient simulation and structured feedback: a randomized controlled trial

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.259363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:0ef88bec86566e147cae2afa2437c2f83c23fbdc440b065d16e8302c5e30c540

Observation f6aee5f8-bea8-4e9a-b8b5-3623ba839c82 · outbound

This paper cites CRITIC: Large language models can self-correct with tool-interactive critiquing.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization CRITIC: Large language models can self-correct with tool-interactive critiquing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.262765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:693ff4d8a04b54d638a9c8dbfb42f89864d5781e9ed916629d0b5bbf1ab71f0a

Observation c4b44ddd-799b-4954-8126-e3b6890ab6a3 · outbound

This paper cites Answers unite! unsupervised metrics for reinforced summarization models.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Answers unite! unsupervised metrics for reinforced summarization models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.269080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:24c45a15078f73019f9c44e84d17dc738e8102eab36820be2a0dbcb99376d8aa

Observation da892bd2-4d46-48cc-bec4-f0557a26f40c · outbound

This paper cites Re3: Generating longer stories with recursive reprompting and revision.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Re3: Generating longer stories with recursive reprompting and revision

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.287844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:572852242a6a5ddb74ec3d1442f8cbc151c53630bda669021fbb636c8b333037

Observation a098760d-d380-4b23-8d8e-d75cc5872d2c · outbound

This paper cites Summac: Re- visiting nli-based models for inconsistency detection in summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Summac: Re- visiting nli-based models for inconsistency detection in summarization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.248440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:dd2fa591a8208730866101b8854693d97a0e52779d4cceece1e85aed74422965

Observation 50b0fe57-c975-4416-8410-033118cfe47a · outbound

This paper cites Iragkr: Iterative retrieval augmented generation with fine- grained knowledge refinement.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Iragkr: Iterative retrieval augmented generation with fine- grained knowledge refinement

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.240889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:f0deaf60d8a6df166fcab6f585a799e6d6179dc421b7a24c9ddd15bec1692f45

Observation 9e8ccdae-1bde-4b4d-9637-56d199c1a1a4 · outbound

This paper cites Adaptive iterative retrieval for enhanced retrieval-augmented genera- tion.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Adaptive iterative retrieval for enhanced retrieval-augmented genera- tion

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.233759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:a1bf958362a4a82e08558ebd4053957e63eedfbe4d5275351bd8158c2d81c4da

Observation 28348270-6986-4a95-87b7-c1979728975d · outbound

This paper cites Learning to summarize from human feedback.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Learning to summarize from human feedback

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.237169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:db195583b9947d49f87258a11069a1b4860eb54fbf81379f52d660a773476410

Observation 4100afb7-ce99-46e5-a16a-a5cfac36c97e · outbound

This paper cites Bigpatent: A large-scale dataset for abstractive and coherent summarization.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Bigpatent: A large-scale dataset for abstractive and coherent summarization

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.245023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:759c5121ea7c6b57cad08db1cfc06db4b8e93fd622b8928cd49bfed21bbaf098

Observation 32324a1b-a8bd-4292-acc6-908bba0412ee · outbound

This paper cites The Llama 3 Herd of Models.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization The Llama 3 Herd of Models

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:36:17.756357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:709083d358e01f01a4c4d4f2e05f356237c5dcdf901b1e48994b7e06553f2ea6

Observation 13198aee-a608-4961-988f-d2aceef41ffd · outbound

This paper cites Llmrefine: Pinpointing and refining large language models via fine-grained actionable feedback.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Llmrefine: Pinpointing and refining large language models via fine-grained actionable feedback

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.251862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:fbaf14c4d6fba92115bed3c8d7a611ef2b4005f4354461a2f7745f4bd44d3bae

Observation 6a057d42-e887-4d02-a1aa-1d2d86f132a9 · outbound

This paper cites Learning to refine with fine-grained natural language feedback.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Learning to refine with fine-grained natural language feedback

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.272413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:34115277db42ec376a0c340d754a72e8972809c6f4a5e2ec9541e1501497c099

Observation fa3bb994-b66a-40d6-9ce9-be6c7b5de490 · outbound

This paper cites Frame: Feedback- refined agent methodology for enhancing medical research insights.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Frame: Feedback- refined agent methodology for enhancing medical research insights

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.294499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:2ee4598753f81f1740c5ea12e4655c6638ab7ae5ee1047baddeab805825eeaad

Observation 1c5a6bb7-5c5a-4410-8b14-891572eaadba · outbound

This paper cites Beyond factual accuracy: Evaluating coverage of diverse factual information in long-form text generation.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Beyond factual accuracy: Evaluating coverage of diverse factual information in long-form text generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.219718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:4b1ca33ab2cfb0c14f919703e0258077a8233d6f984c55a9a720979edd73a2c1

Observation bf4703eb-4cce-46dc-a88a-fada6c744337 · outbound

This paper cites Confidence vs critique: A decomposition of self-correction capability for llms.

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization Confidence vs critique: A decomposition of self-correction capability for llms

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:53:23.319677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-07T16:43:03.770431Z digest=sha256:6e5b90883ed4f577134204147828c4c3a1c096f9a8b682d772e7e1558b5ee6cb

Pith citing papers

Observation a3fa82e2-1ada-4cf6-bc33-458cd539e9d4 · inbound

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops cites this paper.

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-09T03:45:55.347538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-09T03:36:57.168246Z digest=sha256:927f9834da8ad9fae957a05c54abbb713b357ad4130a41a89f5786b5beeb44ec