Pith. sign in

Paper Citation Record · LEDGER

Towards a Unified Multi-Dimensional Evaluator for Text Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2210.07197.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.07197 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:44:59.082443Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0b59baa5-9b17-4a33-89a9-bb4f6aa1a4dc · inbound

G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment cites this paper.

G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T22:55:50.890163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T22:55:50.014540Z digest=sha256:e515efbe8fba9e30427f96a548f31017befa2a78eaa5ecbac9b730eccabfeb14

Observation 6c7fe812-66f3-46b8-a252-445d4f05d543 · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.793778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:5a7ac3130afe06f321a7ea9b226cb56c97abe70d703349cbed9756d021563129

Observation 32a81edf-6fe9-4a5d-86c7-f4fa59c6ae85 · inbound

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries cites this paper.

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:54:19.733847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T13:54:19.264893Z digest=sha256:a1576c6ea7da9890dc9594b65c62eee616d207e6e077497ec0aa54138cec348a

Observation 12a7c8c7-c809-4a00-b9d0-4ef9ec49ac77 · inbound

Benchmark Data Contamination of Large Language Models: A Survey cites this paper.

Benchmark Data Contamination of Large Language Models: A Survey Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 186

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:10:41.155762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T23:10:40.420241Z digest=sha256:6fca6382117b266a823ef1c89655020b277a9342b539aaae7ea1490a560f7a39

Observation 7d324e0b-8baa-4f26-9137-688bea2d9641 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 230

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.307524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:aa35d5595023f70b62d7457ebbcb84dbad35f9a3fe4346e9e4abd14a3fae657e

Observation 2fee3956-a6eb-42d0-bf04-d02c694340a4 · inbound

CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models cites this paper.

CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:44:59.082443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:44:59.082443Z digest=sha256:ef0db6dcee5ad0458f3b78dddeb6a53235cf97e87919225d545bb9d413c6c1a8

Observation f98f101b-0ad5-4e3b-9867-9a0a144ccef5 · inbound

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation cites this paper.

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:04.315586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:03:04.315586Z digest=sha256:9c1e2132ca18417a612ae32f1697da9d18e99bd5910df2785fae29035e54d6e6

Observation 233c6e28-a4f0-47d1-a3ca-b202903a2b38 · inbound

AllSummedUp: un framework open-source pour comparer les metriques d'evaluation de resume cites this paper.

AllSummedUp: un framework open-source pour comparer les metriques d'evaluation de resume Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:38.290508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:23:38.290508Z digest=sha256:3bad1d51a1ca0474a5ecd166c098b111b7097b7ebb7181a7a3812d7960f0d228

Observation b043c2ea-cc58-4209-9186-bf500f00ea96 · inbound

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization cites this paper.

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:50.743278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:59:50.743278Z digest=sha256:448ff76cf07c779c9741e9ad2481ca47b92e9520fa3ccb5fb007724c6a2193b9

Observation c9334041-cf23-4940-8acb-41b65d3db52f · inbound

Evalet: Evaluating Large Language Models through Functional Fragmentation cites this paper.

Evalet: Evaluating Large Language Models through Functional Fragmentation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:01:39.958945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T16:57:25.259866Z digest=sha256:fedf833c86c46c5ac72aa1e0657e2e1b3626101606d73a5e657470142dae1fe8

Observation 6fe3e5d8-7bea-4abe-9409-3ff10dad1166 · inbound

Calibrating Model-Based Evaluation Metrics for Summarization cites this paper.

Calibrating Model-Based Evaluation Metrics for Summarization Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:41:36.994674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T06:36:55.334742Z digest=sha256:588ff0194e251e5b2978f4d1ce5258cad31688a88b40a0aca9bde097ec5c6154

Observation 7c25fa95-6a51-4ad2-8f45-98c4a037682a · inbound

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives cites this paper.

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 159

Resolution
verified exact
arxiv_id, observed 2026-05-10T01:04:50.328006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T01:00:41.543394Z digest=sha256:182bd67f38aefb149c19741fa36a3ddd89c69c8065ae829b1ffee381a568c564

Observation 98c69d14-4e51-454f-a729-333fa5dfbd14 · inbound

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering cites this paper.

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:50:40.215960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T18:05:38.705480Z digest=sha256:8e580b6ec3b175761a535801c7690817afd4c01c5656e13118b1337737fcf646

Observation 10abaf87-5b51-4118-9055-2e0ba8eaef8c · inbound

A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation cites this paper.

A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:58:05.094673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T05:57:25.521881Z digest=sha256:ed0080edbe259d9fed050c093e1606d374dc1e9c150cb12e46de12c04d5f15d0

Observation 9a24dbf6-560e-403a-b992-0f051920b6be · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:50:11.202748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:86d1976d44e6fb32213e81203aba301dd55f20d285694282daa807ad0f6d9196

Observation 96fc4647-b9a6-4cb3-aa22-04d7b24d75d6 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:43.611908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:43.611908Z digest=sha256:6cb0061b6415464bb427d7c4bf11b565ae2627f947ee4cd1bc9f52984650b004

Observation af975a08-464c-4f98-ad01-ee223761082c · inbound

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank cites this paper.

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-26T03:58:57.037544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T03:56:28.271760Z digest=sha256:2968edd609daca83fde3818f6c4207af7c8b774173e749449646e40606e21ce3

Observation ba466b6a-e344-47e1-b8b0-97fc6236b753 · inbound

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 cites this paper.

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T02:23:00.951912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T02:17:26.872854Z digest=sha256:587e45c092ba480013ded054d3f75daadd704f71ed948be7221f8626f4486e15