Pith. sign in

Paper Citation Record · LEDGER

LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2311.09766.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.09766 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:09:03.743155Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T12:53:50.204969Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d2160a96-009f-4a07-81c8-52bb9ac2f147 · inbound

LLM Evaluators Recognize and Favor Their Own Generations cites this paper.

LLM Evaluators Recognize and Favor Their Own Generations LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T18:44:28.834486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T18:44:28.766639Z digest=sha256:d9e718cfbaf7303b096d4eaf91ad291373f67805795795086c8b260d2101260f

Observation 9adf1b1a-4ebb-4f81-87af-babbbc0f498f · inbound

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback cites this paper.

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:38:32.513854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T22:37:43.230753Z digest=sha256:efeb664ce248169b62d097a1dbe7fe634341ffab9d400803c356ae63d1960934

Observation 32ddb30e-97e4-434d-b92a-a0e9df2442ba · inbound

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model cites this paper.

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:13:57.439123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:13:57.368874Z digest=sha256:da98b526e190306b55d4f8cddf37dac97f76534413a00edc5f0f780fb2f9970d

Observation e106cbb7-b709-4a03-aa4b-d339d548924e · inbound

Adaptive-VP: A Framework for LLM-Based Virtual Patients that Adapts to Trainees' Dialogue to Facilitate Nurse Communication Training cites this paper.

Adaptive-VP: A Framework for LLM-Based Virtual Patients that Adapts to Trainees' Dialogue to Facilitate Nurse Communication Training LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:03.743155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:03.743155Z digest=sha256:2a1ee5ae0b272afcc5aeb146455a0f1af8e3d330c8bfe8db6b7d37d4cfeefa15

Observation c5a216dd-79c7-4ed1-89f9-9c84f203cf38 · inbound

Beyond the Surface: Measuring Self-Preference in LLM Judgments cites this paper.

Beyond the Surface: Measuring Self-Preference in LLM Judgments LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:04.946179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:04.946179Z digest=sha256:e5577ea0951cc64a8dd15bf88b93047f8ccfbc85f64a8bf6b43885fcf09cdb2e

Observation d2c512fc-a62d-4465-8867-b396368ee3b4 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.175033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.175033Z digest=sha256:91c911b4d313eb5a050baaf457247d37eba84a10fb5e64c327dd5f25f4e7ba28

Observation 8e4873f1-8fb5-4d9f-b160-79cf59d49357 · inbound

TripTailor: A Real-World Benchmark for Personalized Travel Planning cites this paper.

TripTailor: A Real-World Benchmark for Personalized Travel Planning LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T05:40:31.799364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:40:31.799364Z digest=sha256:f9a5de430c71c7d1e38b0cca13d2da3c6bfe26fd0ac7bce85f0460849a7c127e

Observation 577d1d4b-13f4-41ea-839a-c6fcf9d46fab · inbound

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge cites this paper.

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-05T22:40:42.215645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:40:42.215645Z digest=sha256:82bda7c4f83581aaddfd2417a9542df46c21ef557564b975f223966aeac5dc0f

Observation df1eccca-07c1-48a8-b0c5-70f8f8d8a90d · inbound

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators cites this paper.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.492755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.492755Z digest=sha256:b093cd681497692e8642bd0436dcc8e3e3332b676b2c6795997fe9d7bf3c6e3d

Observation a2461b85-a270-4cc7-8fe6-478af047a7a1 · inbound

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities cites this paper.

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:22:41.940048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T17:20:39.969447Z digest=sha256:08d2de2541890af7872f9620fa031371513b463fe5a36e9cebcc21d4de14c047

Observation 23576d6b-c1f9-4de1-b490-af927d97c239 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 208

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.090816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:13a89a8e062ba8c152b372c4f03b25cb60a0eab3127771be0d19d7da6307e90f

Observation ddf56c74-3e75-4cc1-8072-cbea4f5d93e1 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 207

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:55:59.658761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:f58575b0ae6ad62b2717b5529a4c9c94091681ae9091a77e470964e99f64e58c

Observation a3b551bd-1e0c-495e-8b75-80d806efdb70 · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:53:50.207304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:ec5dfe3d46ea53b72ec28c269b2c906a5c3e8c6e46f184ce354a04aae61d8cac

Observation 9d903bc7-1671-45df-82a5-b0383e87779d · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:1bab4c316927d7f8a5c3cea8f4631bb5131f5be63c557859e27bd18f746e0f48