Pith. sign in

Paper Citation Record · LEDGER

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 2 inbound Pith citation observations for arXiv:2411.14483.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14483 v2

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:14:07.204201Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:51:45.801126Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:12:29.896082Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0faa153b-8cc9-45c1-9e35-cacb099a0f87 · outbound

This paper cites online" 'onlinestring :=.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.113437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.113437Z digest=sha256:7284c86ea860c452ed4d595369fe0e9991515bcf42436a55c7cf1a48c911cbe4

Observation 5ea6b7f8-37d7-4f15-af7e-794d9dd9675a · outbound

This paper cites write newline.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.118103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.118103Z digest=sha256:57f872577f7cd2d47dd3906f5b2a60d5d89d1051d48c2fb43185fc6421628b79

Observation d1df92e1-a4fc-4fb0-8af5-f59bff7d72a2 · outbound

This paper cites Elo Uncovered: Robustness and Best Practices in Language Model Evaluation.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Elo Uncovered: Robustness and Best Practices in Language Model Evaluation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.122126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.122126Z digest=sha256:85307b398cfc39af114bf71c64e3cdd20b004c4d25a8ce8f59337158837b5db7

Observation 0fc82042-0240-439b-83a4-0c6d43bbdfec · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.126918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.126918Z digest=sha256:4c8515d661a516d5c0256e5a17db74d255af98ec21d90b08d181bfbb7239dd18

Observation 2e2c5c88-ce1b-433f-8b35-56f1a4afb37c · outbound

This paper cites Random Walker Ranking for NCAA Division I-A Football.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Random Walker Ranking for NCAA Division I-A Football

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:14:07.280389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.130593Z digest=sha256:f2f1ba67c2fc638bbbc18e1f88e165f0321dba84b94540668445f8f0164cf46c

Observation 8745f801-9476-442d-a55e-36dc78b4e8f5 · outbound

This paper cites The Bowl Championship Series: A Mathematical Review.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat The Bowl Championship Series: A Mathematical Review

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:14:07.262409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.135694Z digest=sha256:789692d2fe3c10c450a94a793afff3bf598a8a3439d677b150205b56495f99b0

Observation 7b4ccd4e-50ab-4fc4-a132-2de46853ee93 · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Gonzalez, Ion Stoica, and Eric P

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.141008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.141008Z digest=sha256:37efb8edef44a38ccc46076ff53bcbe9a043c37d8feb1a3921f2b149042b845d

Observation f82815bd-a9a5-4175-9ab3-d4f106ab3ba9 · outbound

This paper cites Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.145202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.145202Z digest=sha256:972db4cef142ffadb09926340a5f507b7dd165a667b9daac6b37b2a8be557588

Observation 77e2b03e-967c-4894-b92e-4007a3548ecd · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.434268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.149523Z digest=sha256:775d4b1f3df884c2b62203bdd93c7681f8de63922608236fdb58e304a7413938

Observation 88da050e-756d-4203-bc16-765af77b21f9 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.153215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.153215Z digest=sha256:cda9851da2ce1132807a296dda4b9911e15e1d8650b7d2aa97426b06a086e6e9

Observation 901f8317-a6d8-4037-aa23-4b3158a60b7a · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.410615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.157474Z digest=sha256:6b1960262df3849ff1ebd90df8143bbde7410c79dc4721bb0a1b48dbdf15198a

Observation 31a274a7-8fc9-4333-a1ee-795c28588d9f · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.161722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.161722Z digest=sha256:f9a3a28fc4b4292f908ad9383ab34a2b9f1a364c30b77613635a38a331b6ca8e

Observation 69201b8c-3014-4676-add3-3fd8a804dba2 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.398399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.165801Z digest=sha256:3825bb6ec540632948670b7498e75fd94c602396de53d6b35e51ba0a7bb23082

Observation 86b0d2e0-f74a-4db8-be1e-aa9a95e20b1f · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.386661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.170154Z digest=sha256:6e9f1bb276348c7c0df14c1b53f2a10152e5bdf09deca5f26a509d345a8d6b22

Observation 97604a30-f602-487b-84c3-c88f0e8f2cff · outbound

This paper cites Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.173862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.173862Z digest=sha256:94d59b79a9066771b96bbc78e5f29cee5df27972f44c004e692a9ff1e4a78fe3

Observation dfc8f062-b4b0-49e1-8db4-a2e7280e0138 · outbound

This paper cites LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.177986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.177986Z digest=sha256:155e32c2bfa7fd638493fc32c691e1049c71dd4a7745f6c7ebda30000622936a

Observation 5d87537c-57e1-4a35-b2e2-ea76946e0077 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.374231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.183153Z digest=sha256:ff88a617536d8e803e9816066f7a8f0087e6378bb8df3f4804ed4a5e4ef3ff3a

Observation 48450322-31bb-4ba7-8003-20c09d4383e5 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.361647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.187372Z digest=sha256:acd5ac2918086b07bd9fa6ca544b9b24b33454360c4b0e15d18552a3ee5d6f70

Observation 8502a7b6-ba61-47de-ba23-4cf32a818f38 · outbound

This paper cites SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.191359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.191359Z digest=sha256:3a321d245aeb798ad1b86df01135835abf376d5f2288a3f294fbf88ba233f3d3

Observation eadf0244-935b-4c52-93b3-eb4560620dd3 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.195683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.195683Z digest=sha256:1df35393654270b5908e1c0da8f0a13dee7d82744e19efadcc33fa4d61d9ccce

Observation 22f145d0-e36f-48f7-999d-d95152ca568a · outbound

This paper cites Style Over Substance: Evaluation Biases for Large Language Models.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Style Over Substance: Evaluation Biases for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.199865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.199865Z digest=sha256:c45bd83a3a18d221f87d1a29efb67fb64e10eaefb552fa1569e4fbe101f51c81

Observation 32e177b6-3461-4d0b-8a85-195111fd4784 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.204201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.204201Z digest=sha256:21004ca7a4836400591b034ed86be7ac83d1a5720987a67dad4f5335fdba6a5d

Pith citing papers

Observation 50be9836-8e8d-4c41-a469-89c5026fa4b9 · inbound

Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference cites this paper.

Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

Reference 1952

Resolution
unresolved
no resolver link, observed 2026-08-10T22:51:45.801126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:51:45.801126Z digest=sha256:8fc5e3ecd923afe3eb4063395ae6e577b0a815bee8237e76f7d2f95b88ff9b51

Observation c7de3eed-8795-47f5-ba78-8065e605404f · inbound

SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models cites this paper.

SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:12:29.990097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:12:26.391373Z digest=sha256:232c87b5d601b0db30cb01cb1689bf1b20ff8177cc4292473e2a6928ea85ad21