Pith. sign in

Paper Citation Record · LEDGER

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

As of 8 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 3 inbound Pith citation observations for arXiv:2509.03647.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03647 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:51:44.513163Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:36:43.356416Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T00:50:50.435265Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 884fc43c-5944-4f2e-a110-dc1f6eefbd75 · outbound

This paper cites Phi-4 Technical Report.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.470003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.470003Z digest=sha256:424dc96cbefd44110f0384897ae2c42140932a65d90f33df4fa2229ba5ee6266

Observation 796afacc-1d83-4d1f-8b57-db5f3b926950 · outbound

This paper cites DeepSeek-V3 Technical Report.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators DeepSeek-V3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.481798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.481798Z digest=sha256:f9af128a1eeb7f87f0410edd0747bee6da24a554529ee0f329eed3d05fac57c1

Observation 377b2df0-811b-4180-8ce9-86e3676dd4e3 · outbound

This paper cites One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.484491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.484491Z digest=sha256:a23ffd9096a1d2d95d24c49cdff4642efd7f7804d9f2e679ac5502c0d6525e62

Observation 61946c37-67fa-43da-8bd9-ec4f82121088 · outbound

This paper cites The Llama 3 Herd of Models.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators The Llama 3 Herd of Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.487461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.487461Z digest=sha256:7492e2de0843cdaafaacf9d06c4cf6f14a83577a598a43de0b94dd29ef7cdaed

Observation 96025470-ecf9-4e35-a854-5c51dd027bba · outbound

This paper cites A Survey on LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators A Survey on LLM-as-a-Judge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.490091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.490091Z digest=sha256:765874660930fd67c5e448654578dc6b04ec40685bf3583a1640942a2a097b07

Observation df1eccca-07c1-48a8-b0c5-70f8f8d8a90d · outbound

This paper cites LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.492755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.492755Z digest=sha256:b093cd681497692e8642bd0436dcc8e3e3332b676b2c6795997fe9d7bf3c6e3d

Observation d3be7bdb-19ae-4b70-853f-6e0d66ffbc5c · outbound

This paper cites LLM Evaluators Recognize and Favor Their Own Generations.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators LLM Evaluators Recognize and Favor Their Own Generations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.497879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.497879Z digest=sha256:e596e888ed9e38fe0a595990cfb95bf8b3e18d0a01cf930a846ea295ed1d5f84

Observation 2a05c428-d11d-4870-9ba0-98eebba3cf36 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.500620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.500620Z digest=sha256:f55eaf61514a58842a9a53b11773d9c0c93d00d73e56d424381bb7c104e4e418

Observation 6bdeb55c-ce83-427d-969d-5e8c8a117827 · outbound

This paper cites Rerouting LLM Routers.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Rerouting LLM Routers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.503048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.503048Z digest=sha256:ca5bf1ffacb679a198d9582fa0b51705aafb7362383b27b8e2611ed7afd449b7

Observation c33fa0e4-a6f3-491d-9d2b-c883088b9adb · outbound

This paper cites Self-Preference Bias in LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Self-Preference Bias in LLM-as-a-Judge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.505631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.505631Z digest=sha256:7cbe50dfcdea5b993a672790272e09f7871dc6d42107877828e60ce427c66e25

Observation abb5264d-5f38-4023-999b-72fc205319bd · outbound

This paper cites Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.508182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.508182Z digest=sha256:cc8fd727fbc5afe44291a41c3d8c43dadada878cb5febb6d3ad042d0daf27ab9

Observation 7bd8e9d5-44b3-46e2-9d24-04caeee55b22 · outbound

This paper cites Leveraging Uncertainty Estimation for Efficient LLM Routing.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Leveraging Uncertainty Estimation for Efficient LLM Routing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.510706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.510706Z digest=sha256:5f126f9ed1c6248734970ddd99f9f414a68b5ac3261964bb78bb15a396005984

Observation 95f663a8-0f6f-434c-a24f-bfc367ab9160 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.513163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.513163Z digest=sha256:1cd3fc6cdf071ab39704872c970c3e50df5905819b76e846c4dfb882662cda75

Observation 47a93026-2b4c-4a59-b9cc-8bc98faef0b3 · outbound

This paper cites Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.495367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.495367Z digest=sha256:ce742eeb7e061ddda8abcacc592841ec57ad9c6b1ceabd9d2e1d3e441b2ae7a9

Observation 8a537809-b3ad-4d46-a0ed-39b94ad498ba · outbound

This paper cites VisIT-Bench: A Benchmark for Vision-Language Instruction Following Inspired by Real-World Use.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators VisIT-Bench: A Benchmark for Vision-Language Instruction Following Inspired by Real-World Use

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.475907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.475907Z digest=sha256:bbb0944263c85da4bab3ccac74e24e1fabbfbe3f4b3da55d52754f23dd6b0537

Observation d673015e-3920-4df5-beac-72cb4138ac6a · outbound

This paper cites Do LLM Evaluators Pre- fer Themselves for a Reason?, April 2025a.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Do LLM Evaluators Pre- fer Themselves for a Reason?, April 2025a

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.478823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.478823Z digest=sha256:d00425732098b6001cb43f9fa5878b1918d0810882f7cd8bc3c81f943c944673

Observation 81a2150a-7c62-44be-a267-e3c553b4f0de · outbound

This paper cites Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct.

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T10:51:44.473173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:51:44.473173Z digest=sha256:1d348fdc3c17d6bdb05adf3ad03ee6fef38eb2fb8ea3f1d5220297086b9fc127

Pith citing papers

Observation 7cb81bf0-2e4d-4052-936b-77c570e966e6 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-24T01:14:21.924280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:18:19.955943Z digest=sha256:1f7d276beaa787a1db0d35bb682590b22aed3d0e1becf6de31faeff100d01ae1

Observation 02ac22cc-c851-49b1-8ac1-3267d1deaeb0 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T16:40:55.442545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:40:55.442545Z digest=sha256:e4ef98242800a7c2919595b92e2caf0751fa60e4bf0048d07b4ba5f8a5584029

Observation 6ab5c2fe-4aef-4e52-8ce4-19efc153bb24 · inbound

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models cites this paper.

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:36:43.356416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:36:43.356416Z digest=sha256:d7cd61809382d0a19bdf53b5e14830258d76bc7a37b84f0c411bc99cd83d5e5a