Pith. sign in

Paper Citation Record · LEDGER

Human-Centered Design Recommendations for LLM-as-a-Judge

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2407.03479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.03479 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:56:37.483408Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 698c9c48-3d7c-4663-9190-028b2013d533 · inbound

Engineering AI Judge Systems cites this paper.

Engineering AI Judge Systems Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T11:59:18.356623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:59:18.356623Z digest=sha256:d6505a0786b9a0ac505ea5bdaa5dec162d1afaeb83982d8acadd4fecc2a47d3d

Observation 2c4cc155-e7f7-4117-98d5-dd7338a6cb85 · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.317267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:8385efdbf846e1697820753c267b6433a07a70944e4f6de4c5398673afe4d6d0

Observation fd190ed0-e42c-4e1b-94df-fab570a7591c · inbound

tAIfa: Enhancing Team Effectiveness and Cohesion with AI-Generated Automated Feedback cites this paper.

tAIfa: Enhancing Team Effectiveness and Cohesion with AI-Generated Automated Feedback Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T11:56:37.483408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:56:37.483408Z digest=sha256:fb5090a0af01fd5f9995600af5ed582fab9c11de4eba4a25a940bfe3b04d8198

Observation 70fa1381-9ced-4a64-ac84-4ef91d15ef08 · inbound

A Design Space for the Critical Validation of LLM-Generated Tabular Data cites this paper.

A Design Space for the Critical Validation of LLM-Generated Tabular Data Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:30:29.502034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:30:29.502034Z digest=sha256:c4eb9c487e3ff0d309ba60e5bc134c2014fb0f6949bad87cd5850e69e03058f4

Observation 6095d48a-b2e7-45c7-9b7e-a8654c37d8f5 · inbound

A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development cites this paper.

A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T22:13:04.805330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:13:04.805330Z digest=sha256:03efca65d24869df6907993b798167226011775048e6299cef2b5a72e05c293e

Observation 487ff4ce-77a9-460c-ae37-99a31a85c4a9 · inbound

Arbiters of Ambivalence: Challenges of Using LLMs in No-Consensus Tasks cites this paper.

Arbiters of Ambivalence: Challenges of Using LLMs in No-Consensus Tasks Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:51.588167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:23:51.588167Z digest=sha256:d63f4f40bb5bff3144910f4f605419e9bce2f9cd9946356b008214d5e881db9d

Observation 5d4767b0-ebb5-4a5c-a42b-402c88401fa7 · inbound

Measurement as Bricolage: Examining How Data Scientists Construct Target Variables for Predictive Modeling Tasks cites this paper.

Measurement as Bricolage: Examining How Data Scientists Construct Target Variables for Predictive Modeling Tasks Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:08.262727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:08.262727Z digest=sha256:e8ad24f833eceada1e1d7eea8004e574cc476943aafb5e583c6d0713c9a323c6

Observation 22b28149-dc76-4615-91ab-df4e7e495a21 · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:02:52.282207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T22:02:36.307598Z digest=sha256:da3e56cba2c90aa15f30b7e2ad15711dd6d6056afc8b8906f3d530c814a1b07b

Observation cdb0bbe9-dc64-4ecc-83aa-5144c59b0d2a · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:20:31.357401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T08:18:18.448122Z digest=sha256:66e1be1bf3d94638eb38c7e777c7640fb5980145b79ccb5b7271eba61a6c5c26

Observation c45983fe-e20d-4776-843b-6212170bc539 · inbound

LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection cites this paper.

LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:45:48.078704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T19:37:19.360378Z digest=sha256:777576099ef3158ad0ba037f1e45340b0c9d6f36e414f4a1b204b78f97018182

Observation 75230a94-7698-4b96-a6b1-4f25dd1cc06a · inbound

Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild cites this paper.

Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:27:48.184935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T11:26:38.634540Z digest=sha256:8b7854c12c56b9a1b319f87588216dff6510555a2845f7a2f38466e06add057a

Observation c8a6e547-ba27-42f7-bd36-33f7d81ad6ab · inbound

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation cites this paper.

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation Human-Centered Design Recommendations for LLM-as-a-Judge

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:30.820148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T16:46:34.169875Z digest=sha256:de8d4f13f62394f8bf9cc1c738f6e1702b0e9177e609946591ca4ef5485bcc5e