Pith. sign in

Paper Citation Record · LEDGER

Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2402.14016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.14016 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:09:54.139224Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T17:35:43.987419Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04588e05-18b4-49ea-860c-c307ab898808 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 119

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:43.990648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:57915b31e176be9835f569b20da0ce4d0664bbb0187119dd7f6db16ed40d964f

Observation 304c9cd5-a2d1-453d-9854-a2a083b98a4e · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:34.639350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:fbc520bbfebb55a645582f6a683f766356f7debcec664008096b5c298c2897dd

Observation e39f4052-00bc-4d4b-9d89-f8c258d4d13e · inbound

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations cites this paper.

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T17:09:54.139224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:09:54.139224Z digest=sha256:eb6df62e9fbd7717026e4b472775ed44142c290e94c7bd26483d833ea5d71239

Observation a5657ff9-64fe-44bd-9873-1096621b70c6 · inbound

Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-box Neural Ranking Models cites this paper.

Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-box Neural Ranking Models Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T04:32:36.186677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:32:36.186677Z digest=sha256:7618010b88b62e5cd966e7fd59c53006fa9a6f83f9ff90ce1dfa41094a14d534

Observation 00788e6f-822a-407f-92e2-abf6df22ed5b · inbound

One Token to Fool LLM-as-a-Judge cites this paper.

One Token to Fool LLM-as-a-Judge Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:38.612180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:15:38.612180Z digest=sha256:fede49b9a6737811fbc2929f12a9355aaeebdef754dc94e91fcc4da5c6eced84

Observation 251f3ac3-81a6-405d-b40c-8702098bca6a · inbound

TripTailor: A Real-World Benchmark for Personalized Travel Planning cites this paper.

TripTailor: A Real-World Benchmark for Personalized Travel Planning Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T05:40:31.814090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:40:31.814090Z digest=sha256:4e3b888e72b98ba34694d28f4f3d47d3296cf5a9e52ef722428a4bd897b44e61

Observation ac8e5f32-794d-40f0-ab27-6d0b0dd180b1 · inbound

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization cites this paper.

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.605293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-18T12:53:45.767341Z digest=sha256:61c94cd8cbe712fd037983b49463d6ddfb70e18ac30d8c659381389a004dea76

Observation 73696ae7-9264-4465-9555-69c1a26a9c9f · inbound

When AI reviews science: Can we trust the referee? cites this paper.

When AI reviews science: Can we trust the referee? Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:16:11.606972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-08T06:19:54.727724Z digest=sha256:6b696b15931c627a575f0ac38657b2b3228424e83458f97c954f424c33cb03ea

Observation a4435910-bfd7-4897-827d-2203faf2ba23 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:20.567329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:20.567329Z digest=sha256:7e2c3a1a20aa5ac4fbbbe2c2143d6ff9f503e893f6172697c1bb482b3f0e0d3e