Pith. sign in

Paper Citation Record · LEDGER

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

As of 19 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 3 inbound Pith citation observations for arXiv:2509.03647.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03647 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:51:44.513163Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:36:43.356416Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T00:50:50.435265Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 884fc43c-5944-4f2e-a110-dc1f6eefbd75 · outbound

This paper cites Phi-4 Technical Report.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.470003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.470003Z digest=sha256:6649fe5356ecb78ab10db68fed76d0bc7df091ace45467c6e4289af8808eec5d

Observation 796afacc-1d83-4d1f-8b57-db5f3b926950 · outbound

This paper cites DeepSeek-V3 Technical Report.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators DeepSeek-V3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.481798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.481798Z digest=sha256:8c9f926bf3808e27d0c90d19f653fd8bffd7020305e6a80b5aae4a736ce89411

Observation 377b2df0-811b-4180-8ce9-86e3676dd4e3 · outbound

This paper cites One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.484491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.484491Z digest=sha256:966e33d5b51899084996e33431161c0a477c416851020e02ee9e65dfd20a6510

Observation 61946c37-67fa-43da-8bd9-ec4f82121088 · outbound

This paper cites The Llama 3 Herd of Models.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators The Llama 3 Herd of Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.487461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.487461Z digest=sha256:835f5da152d1c312fede81e9c76337c9400d0b7b1db1bc0daa151c203a21c7dc

Observation 96025470-ecf9-4e35-a854-5c51dd027bba · outbound

This paper cites A Survey on LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators A Survey on LLM-as-a-Judge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.490091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.490091Z digest=sha256:735aaf7dd6c13bd88c3998bb2e9eb5584c03cbaee3a8f6533e9eef2567e0253e

Observation df1eccca-07c1-48a8-b0c5-70f8f8d8a90d · outbound

This paper cites LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.492755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.492755Z digest=sha256:7f49d814af36def123ed2ea1944dbd1bd9218dcda311b024b95254114e72d83a

Observation d3be7bdb-19ae-4b70-853f-6e0d66ffbc5c · outbound

This paper cites LLM Evaluators Recognize and Favor Their Own Generations.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators LLM Evaluators Recognize and Favor Their Own Generations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.497879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.497879Z digest=sha256:da18ef6ff878f5d6d12fe5d2a7f1e84205a03764082c36d0154aab0c9e536435

Observation 2a05c428-d11d-4870-9ba0-98eebba3cf36 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.500620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.500620Z digest=sha256:3b9e4f155953f2ba21cc0aa4c5f87ab01fceb31c591b39bf592f54c5cd9020d8

Observation 6bdeb55c-ce83-427d-969d-5e8c8a117827 · outbound

This paper cites Rerouting LLM Routers.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Rerouting LLM Routers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.503048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.503048Z digest=sha256:b11cf4f1d168cb01af06a9d7a259d4c58ca54aa2446e51f3050938ba0947baad

Observation c33fa0e4-a6f3-491d-9d2b-c883088b9adb · outbound

This paper cites Self-Preference Bias in LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Self-Preference Bias in LLM-as-a-Judge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.505631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.505631Z digest=sha256:ba9e6a8e5b0d81a3ef8eb7398dbcb265fa13397f9311df87c323cd169d0d5348

Observation abb5264d-5f38-4023-999b-72fc205319bd · outbound

This paper cites Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.508182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.508182Z digest=sha256:9945bdfb1432be103062dec751b501cb86e48575d85af357255f311733222224

Observation 7bd8e9d5-44b3-46e2-9d24-04caeee55b22 · outbound

This paper cites Leveraging Uncertainty Estimation for Efficient LLM Routing.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Leveraging Uncertainty Estimation for Efficient LLM Routing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.510706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.510706Z digest=sha256:64d75560df7a348a85a3425b1bce3791250419e5c301cef7e128fb0c88adf6b5

Observation 95f663a8-0f6f-434c-a24f-bfc367ab9160 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.513163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.513163Z digest=sha256:9124526001a458bc573ec88854a6b75aba45eb045586102550d89fe58c01b3b9

Observation 47a93026-2b4c-4a59-b9cc-8bc98faef0b3 · outbound

This paper cites Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.495367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.495367Z digest=sha256:da5afbbc0bc393579929511adc570a229def433a66a16b57dc4565eba937cfde

Observation 8a537809-b3ad-4d46-a0ed-39b94ad498ba · outbound

This paper cites VisIT-Bench: A Benchmark for Vision-Language Instruction Following Inspired by Real-World Use.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators VisIT-Bench: A Benchmark for Vision-Language Instruction Following Inspired by Real-World Use

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.475907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.475907Z digest=sha256:f221d0c9004077d430c64c59e818ec17fd2f4d4e4d5529a9e4edb2e2cc1b1795

Observation d673015e-3920-4df5-beac-72cb4138ac6a · outbound

This paper cites Do LLM Evaluators Pre- fer Themselves for a Reason?, April 2025a.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Do LLM Evaluators Pre- fer Themselves for a Reason?, April 2025a

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.478823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.478823Z digest=sha256:f8a45bf01f35d647b86e436a2288c307b5ae731a91c50d24a8f1b64a261d71c6

Observation 81a2150a-7c62-44be-a267-e3c553b4f0de · outbound

This paper cites Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.473173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.473173Z digest=sha256:3d4318b0dcdcb90702967cbaef1d751afa37858d48f6ea4242413bbcb9857a03

Pith citing papers

Observation 7cb81bf0-2e4d-4052-936b-77c570e966e6 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-24T01:14:21.924280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:18:19.955943Z digest=sha256:b866b56ea00be1a9de897ea054ac65088259fd5711e770f352171284face5d67

Observation 02ac22cc-c851-49b1-8ac1-3267d1deaeb0 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T16:40:55.442545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:40:55.442545Z digest=sha256:7dacaca739fb60505bbfac8609a61911694a9b6f59d6c47c702cb460e7ce7330

Observation 6ab5c2fe-4aef-4e52-8ce4-19efc153bb24 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:36:43.356416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:36:43.356416Z digest=sha256:e4e6775faca6c339655dbc1a845ce8523cd5199f2b33d223149d3fdd3afc0c06