Pith. sign in

Paper Citation Record · LEDGER

Human-like Summarization Evaluation with ChatGPT

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2304.02554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.02554 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:17:21.048674Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

38
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 67ef8c81-3b94-4d19-84a7-3575f5e7b04d · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate Human-like Summarization Evaluation with ChatGPT

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.810777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:74ab85f85a30f5e679c2d3e3defffa265a900890efda36613794b61491e2729f

Observation 3ef92e9e-ed7a-4170-8e6b-ff61fae66aea · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Human-like Summarization Evaluation with ChatGPT

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:46:27.555802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:f00c3bc59cfa43ff156f58afb05f807151bcf9833529c4df0236731b4ced9034

Observation 902dbcb4-60da-4398-829b-1baadea826d7 · inbound

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cites this paper.

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions Human-like Summarization Evaluation with ChatGPT

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:55:49.997989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T21:54:26.670284Z digest=sha256:7be408a8c844d9b5d040c1a3d79503c1b1cc452f91b2fe8fa88754773f91dd69

Observation 06000e1c-39a1-4252-8da1-202daa152937 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Human-like Summarization Evaluation with ChatGPT

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.287787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:ffcae1a6bdfac79fe3d72c2a94cf25c105cb83a7d3004c31eeb0b16c0d9407d9

Observation 3e78450a-e50d-46a7-ab9b-7121ba863c7e · inbound

Evaluating Small Language Models for News Summarization: Implications and Factors Influencing Performance cites this paper.

Evaluating Small Language Models for News Summarization: Implications and Factors Influencing Performance Human-like Summarization Evaluation with ChatGPT

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T18:17:21.048674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:17:21.048674Z digest=sha256:d9db249af2c57db60a8d4f51855e35910588bb950796ef86540519a25ee39705

Observation 32a5e30b-a6fc-44c1-842d-092c8b0d8f73 · inbound

LLMs to Support a Domain Specific Knowledge Assistant cites this paper.

LLMs to Support a Domain Specific Knowledge Assistant Human-like Summarization Evaluation with ChatGPT

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T23:35:38.362684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:35:38.362684Z digest=sha256:b35530d65c156ea23b92e9c703ab79ed1bc0a1df23c0ea607f366fecb122c697

Observation 7ed344d8-cc99-4521-8a1f-236530c45dda · inbound

Efficient Online RFT with Plug-and-Play LLM Judges: Unlocking State-of-the-Art Performance cites this paper.

Efficient Online RFT with Plug-and-Play LLM Judges: Unlocking State-of-the-Art Performance Human-like Summarization Evaluation with ChatGPT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:34.827221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:34.827221Z digest=sha256:6a90482731ef6f149296274f581b74ceb4571fee3782e2048d9e2f15488e9752

Observation 6d4cbf06-a589-4f36-8b55-fe83b1168291 · inbound

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation cites this paper.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Human-like Summarization Evaluation with ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.575596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.575596Z digest=sha256:656c9f83ae9c3395a170fc745cc655d2066b8371b04a9db4bacda22613c4b70a

Observation 741f0943-23d7-4e97-a828-35f987c12e2a · inbound

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications cites this paper.

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications Human-like Summarization Evaluation with ChatGPT

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:31.638854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:31.638854Z digest=sha256:5a2f40c11f8c45fbb6dd5f49da1ea5f3e43e5fe2b9b5db87d9c611b6821939fe

Observation 94ea1d46-cb64-4f75-8c1d-946df320e2c7 · inbound

Byzantine-Robust Decentralized Coordination of LLM Agents cites this paper.

Byzantine-Robust Decentralized Coordination of LLM Agents Human-like Summarization Evaluation with ChatGPT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:01.188665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:01.188665Z digest=sha256:0a9757fd446e3c45befb7eddc12d6400089be07b1f32551b9ac1c7e9f372885d

Observation 4fe04d12-60a9-4794-9cab-1de6c3c9ada0 · inbound

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge cites this paper.

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge Human-like Summarization Evaluation with ChatGPT

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-05T22:40:42.192331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:40:42.192331Z digest=sha256:8e2ce5f50a21d41457fcbcbd3fc46eebeee80f51bd4d82ce30fffd127a55cba5

Observation 136b11e9-e3d3-4a0a-81a7-45a6755596d3 · inbound

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization cites this paper.

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization Human-like Summarization Evaluation with ChatGPT

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:50.428652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:59:50.428652Z digest=sha256:a433319ee792a5b75a278fe2fadaf4d3063115cd6ce9b99fbc6dae0e796224a5

Observation 9cf5c829-93f6-42de-be24-672595c7a65e · inbound

AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs cites this paper.

AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs Human-like Summarization Evaluation with ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:05.221826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T06:00:05.221826Z digest=sha256:f64610c28cb37e6fe1cb4c95a2c289600c68b51a36c6727d12730c9f2a345290

Observation 5e33f5bb-19b6-4662-ab56-48f6373323aa · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning Human-like Summarization Evaluation with ChatGPT

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.562656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.562656Z digest=sha256:2af2635372a0a7f202d71f71087835dd29ac96fbe634c2d70415f68e4d60dcbd

Observation fc0e4375-8325-4003-854b-0831ba72c975 · inbound

From Moderation to Mediation: Can LLMs Serve as Mediators in Online Flame Wars? cites this paper.

From Moderation to Mediation: Can LLMs Serve as Mediators in Online Flame Wars? Human-like Summarization Evaluation with ChatGPT

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T18:57:00.265923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:57:00.265923Z digest=sha256:725e1520162fa1021d1d7089402e2556976a070068426d44d1721c3a35f694c9

Observation 98df71dd-262c-4bc6-a17e-7cb5936d0a21 · inbound

User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study cites this paper.

User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study Human-like Summarization Evaluation with ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T13:05:28.935067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:05:28.935067Z digest=sha256:5e9ee4b68aef59ca540f8243bb5ff9e9e72f7c300c930a941560b88847f07525

Observation 704a3560-05c3-4bcc-87d5-054ee716e1fc · inbound

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory cites this paper.

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory Human-like Summarization Evaluation with ChatGPT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T06:08:26.400539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:08:26.400539Z digest=sha256:aa6b4e72df82be2a4811db23bfdbede01f582fd746c30482a81bb3affd66a116

Observation 0dfd8e45-4b92-4d8b-b3ce-17fc499899f5 · inbound

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation cites this paper.

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation Human-like Summarization Evaluation with ChatGPT

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:37.680608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T16:20:50.111108Z digest=sha256:fd1a5decaa15c3b5965eb4e176a4e4508323f0018b9de38f854dfb68d082eacd

Observation 7c59b5e1-23bc-45bf-9f81-85e651965119 · inbound

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance cites this paper.

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance Human-like Summarization Evaluation with ChatGPT

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:37.284099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T10:34:52.474916Z digest=sha256:fc7b2e56c4e14dd202e925745c0f9a3eb4e0a95e6a76d4236ee1db53d060fd5f

Observation 21bfc490-1d5f-464a-8b0a-c7a871e1b691 · inbound

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation cites this paper.

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation Human-like Summarization Evaluation with ChatGPT

Reference 114

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:37:22.701785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T20:20:08.996005Z digest=sha256:5539920b7c0da5e1d36edce4c8a06f0d8b3059e1854a69c4a2ff1a2d582d75e8