Pith. sign in

Paper Citation Record · LEDGER

Black-box Uncertainty Quantification Method for LLM-as-a-Judge

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2410.11594.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.11594 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:13:04.887831Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T03:35:50.526763Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation af9711ed-856e-4e79-95df-f95c0051aa06 · inbound

A Survey of Calibration Process for Black-Box LLMs cites this paper.

A Survey of Calibration Process for Black-Box LLMs Black-box Uncertainty Quantification Method for LLM-as-a-Judge

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T13:48:40.102166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:48:40.102166Z digest=sha256:7c3f39eb09f12b36294fe25f6c14dd20567399503e5562d1bb79c3d4201e435c

Observation f4cad87c-1e69-4c3e-baf0-c9f4ffad7ec2 · inbound

A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development cites this paper.

A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development Black-box Uncertainty Quantification Method for LLM-as-a-Judge

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T22:13:04.887831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:13:04.887831Z digest=sha256:ea8f3640c9429a818b066de792e38fa5176dd913f3d2a7bb203f004dce87c3d3

Observation 01c1984b-176b-4dd4-8e64-63fba250c49f · inbound

Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks cites this paper.

Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks Black-box Uncertainty Quantification Method for LLM-as-a-Judge

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:35:50.530335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T03:32:53.109030Z digest=sha256:614fa1b87b0db677a7b52fc92ee8ded1cccb4cb7a71a4e54d26b506f81f67514

Observation 00d28e57-a6e8-4d3f-9c50-454595b1038e · inbound

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory cites this paper.

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory Black-box Uncertainty Quantification Method for LLM-as-a-Judge

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T06:08:26.735504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:08:26.735504Z digest=sha256:cd718fe353c1fa3b1caa3891c736bd5f961e34b6f526464b08acb5a393cfeab3

Observation 73d6154b-b99d-41ee-8d98-2e2dcef72e02 · inbound

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation cites this paper.

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation Black-box Uncertainty Quantification Method for LLM-as-a-Judge

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:17.371041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-07T16:46:01.749101Z digest=sha256:42e3b0c39663a9b69ca273daa489b163137dfcab2d7982136c897eaa0f1b0396