Pith. sign in

Paper Citation Record · LEDGER

Towards a Unified Multi-Dimensional Evaluator for Text Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2210.07197.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.07197 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:44:59.082443Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0b59baa5-9b17-4a33-89a9-bb4f6aa1a4dc · inbound

G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment cites this paper.

G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T22:55:50.890163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T22:55:50.014540Z digest=sha256:ffda91ec8d1ffea427a6ff68281396607951418f2356e01d78bf129b46f8076d

Observation 6c7fe812-66f3-46b8-a252-445d4f05d543 · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.793778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:a9848bf61e027f928fd12c131b4bf5b7d322e7b5b74c9804896548c2f882bd67

Observation 32a81edf-6fe9-4a5d-86c7-f4fa59c6ae85 · inbound

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries cites this paper.

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:54:19.733847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T13:54:19.264893Z digest=sha256:a934a9546e8556e5cd0d3e920bf97fd486db4055914d1700706d88c4735c5568

Observation 12a7c8c7-c809-4a00-b9d0-4ef9ec49ac77 · inbound

Benchmark Data Contamination of Large Language Models: A Survey cites this paper.

Benchmark Data Contamination of Large Language Models: A Survey Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 186

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:10:41.155762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:10:40.420241Z digest=sha256:3b16b5db5d9a71667a05a0c846f19ecc9b95ec204fd58996d99b860e43cd4652

Observation 7d324e0b-8baa-4f26-9137-688bea2d9641 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 230

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.307524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:5ba655fa1adfd6cf65c68fd9fd2c42e319ac1bf121a06466c24c086b8e7c5a3a

Observation 2fee3956-a6eb-42d0-bf04-d02c694340a4 · inbound

CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models cites this paper.

CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:44:59.082443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:44:59.082443Z digest=sha256:100d3e85906c7c859b45baf49ecdbf98141a583ef6ceae8d871829ba874c5f4e

Observation f98f101b-0ad5-4e3b-9867-9a0a144ccef5 · inbound

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation cites this paper.

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:04.315586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:03:04.315586Z digest=sha256:4f51c3ea36be11837a93276f2db8a3577bd17fb8b8327fcf3dc06bf6925b7dbd

Observation 233c6e28-a4f0-47d1-a3ca-b202903a2b38 · inbound

AllSummedUp: un framework open-source pour comparer les metriques d'evaluation de resume cites this paper.

AllSummedUp: un framework open-source pour comparer les metriques d'evaluation de resume Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:38.290508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:23:38.290508Z digest=sha256:a072c0abaa6b6c9638fc03b3dd11dcb684e65b4a9653a8d230f87f8251f9bb7a

Observation b043c2ea-cc58-4209-9186-bf500f00ea96 · inbound

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization cites this paper.

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:50.743278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:59:50.743278Z digest=sha256:77a34e311cfaa8a55ab37094f507fbc8880bc6b396005dd9700ec610d359aa3b

Observation c9334041-cf23-4940-8acb-41b65d3db52f · inbound

Evalet: Evaluating Large Language Models through Functional Fragmentation cites this paper.

Evalet: Evaluating Large Language Models through Functional Fragmentation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:01:39.958945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T16:57:25.259866Z digest=sha256:3816fd9e4997aae2d711f92e7bbb57ecf18541e0c83d5deb582ac3fa3dd1e509

Observation 6fe3e5d8-7bea-4abe-9409-3ff10dad1166 · inbound

Calibrating Model-Based Evaluation Metrics for Summarization cites this paper.

Calibrating Model-Based Evaluation Metrics for Summarization Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:41:36.994674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T06:36:55.334742Z digest=sha256:8357293ee92dba504c253d68b01f7516e53f48db761f66ca6a43b24bb6d6e5e0

Observation 7c25fa95-6a51-4ad2-8f45-98c4a037682a · inbound

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives cites this paper.

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 159

Resolution
verified exact
arxiv_id, observed 2026-05-10T01:04:50.328006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T01:00:41.543394Z digest=sha256:f8259895fbfc35e66c73520eca7be6ebc1d2052439b22f11e4bb751a1e6e599b

Observation 98c69d14-4e51-454f-a729-333fa5dfbd14 · inbound

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering cites this paper.

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:50:40.215960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T18:05:38.705480Z digest=sha256:c78808dea5fa2794a6b8e9d69ebd46dcfb823ec20909b658370211f9ee68d3b2

Observation 10abaf87-5b51-4118-9055-2e0ba8eaef8c · inbound

A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation cites this paper.

A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:58:05.094673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T05:57:25.521881Z digest=sha256:0ec949c9f980b1a20f589d17998aa6186024369ea4add77e2312c643ae2a9856

Observation 9a24dbf6-560e-403a-b992-0f051920b6be · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:50:11.202748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:d3ec68c027a2922475d3c1c0c8afef41173256d5c8cd9138214c91a89ed53261

Observation 96fc4647-b9a6-4cb3-aa22-04d7b24d75d6 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:43.611908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:43.611908Z digest=sha256:6cb0061b6415464bb427d7c4bf11b565ae2627f947ee4cd1bc9f52984650b004

Observation af975a08-464c-4f98-ad01-ee223761082c · inbound

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank cites this paper.

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-26T03:58:57.037544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T03:56:28.271760Z digest=sha256:1ff12ff6eae56ce30e6b1448993d3af4917cf0a7f1937a33700e4c12267c8bee

Observation ba466b6a-e344-47e1-b8b0-97fc6236b753 · inbound

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 cites this paper.

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T02:23:00.951912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T02:17:26.872854Z digest=sha256:46666dd0d06ee1b9c1eeb98723108d92b9624feec4fbf5ed6f2ef4530dccccc3