Pith. sign in

Paper Citation Record · LEDGER

Behavioral Fingerprinting of Large Language Models

As of 15 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 4 inbound Pith citation observations for arXiv:2509.04504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04504 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:04:09.577795Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:52:24.750109Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 27894d4b-329d-4aa4-af84-5d61342cbb8f · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Behavioral Fingerprinting of Large Language Models OPT: Open Pre-trained Transformer Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:07.725926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:07.725926Z digest=sha256:6dc9cbc7f7aa91eac122a5422e42bad21ba3d225523636da0f595381c63075b4

Observation 527be7cc-77e9-4e5c-8d70-41edac7f6399 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Behavioral Fingerprinting of Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:07.805287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:07.805287Z digest=sha256:e316e9673d8bad80e4cab0074180148e7b66f31fc9b6be60c39bffb1adbfdb31

Observation 3bf5cc7b-bdb9-4cd4-86d9-08da44344ef5 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Behavioral Fingerprinting of Large Language Models Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:07.879716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:07.879716Z digest=sha256:942510a71fc5671f4cd1a759208a4abf90c96e13ef93d95e6c7da04fa99f5398

Observation 07948307-46af-436b-911a-77dddad52066 · outbound

This paper cites DeepSeek-V3 Technical Report.

Behavioral Fingerprinting of Large Language Models DeepSeek-V3 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:07.951912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:07.951912Z digest=sha256:f1f61acc12fa9294a017ad8df25fca6e3fe60e98deb5533eecd2df7ab0efa56c

Observation 1adc9ca0-d81c-4c94-bb3a-cdc3a6cded33 · outbound

This paper cites Language Models are Few-Shot Learners.

Behavioral Fingerprinting of Large Language Models Language Models are Few-Shot Learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.033770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.033770Z digest=sha256:f42a4bac3bb2a605bdcf03a0b1705cd847a7d31c3bd5f7732dd37af8c40c44ca

Observation 5a63c54f-6fe2-4f30-ae94-b591328f53f7 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Behavioral Fingerprinting of Large Language Models GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.126591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.126591Z digest=sha256:08692fb0537496d32816838efd5bf9686897ed57179e50ea145c61a5b0b2aa3d

Observation eb6b9a97-fff6-4b56-ad08-98fbaea5a516 · outbound

This paper cites Superglue: A stickier benchmark for general-purpose language understanding systems.

Behavioral Fingerprinting of Large Language Models Superglue: A stickier benchmark for general-purpose language understanding systems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.228418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.228418Z digest=sha256:c37f70c59c97fb93b0a3ea12e71a595ec5fe4efc5531a7cc086a3bc62f132760

Observation db3ddec0-ee3a-4b27-9d80-7eb0fe6fae23 · outbound

This paper cites Holistic Evaluation of Language Models.

Behavioral Fingerprinting of Large Language Models Holistic Evaluation of Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.369556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.369556Z digest=sha256:41d96ce90b9bcc4dee610e17ee4836ec9609a10e86f96d3b6caced7ae9206d9f

Observation 4fa7b725-b52d-45e0-8b72-29cddf72971e · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Behavioral Fingerprinting of Large Language Models Measuring Massive Multitask Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.460978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.460978Z digest=sha256:9b1381638983740a6ce4c8483cb8e41341847c8c48412c28c5c516b70f33bba7

Observation 089921d1-5279-466d-ba26-809c720e9e08 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Behavioral Fingerprinting of Large Language Models Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.489296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.489296Z digest=sha256:a74ef08a40b0180b72f013121e036830284ad2239773050154e5b1b016f0dfb2

Observation 5bb923c7-cf60-4dc3-9200-bdfc63772f7d · outbound

This paper cites Gifts differing: Understanding personality type.

Behavioral Fingerprinting of Large Language Models Gifts differing: Understanding personality type

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:04:10.678169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:04:08.605020Z digest=sha256:ebd57a83ee5e7ac51f23da1ad445c1c713d2f84f16babffbd82b167cbf121e48

Observation 4555e996-8f42-4382-8d3e-559df4a39d47 · outbound

This paper cites Checkeval: A reliable llm-as-a-judge framework for evaluating text generation using checklists.

Behavioral Fingerprinting of Large Language Models Checkeval: A reliable llm-as-a-judge framework for evaluating text generation using checklists

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.734783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.734783Z digest=sha256:8e720d17c62a213dc55f218d7cbcaa097999cf49dc6760aea52626844733ea28

Observation 0c2c9326-3176-47d3-bd94-fa9397f04101 · outbound

This paper cites FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models.

Behavioral Fingerprinting of Large Language Models FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:08.940093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:08.940093Z digest=sha256:261098d9d4673c7314988aa04f3c3bed509d9c59fd12072bbc4cc04734553773

Observation 6f138fb8-9bf2-42e6-b126-684a9c0efdcf · outbound

This paper cites UltraEval: A Lightweight Platform for Flexible and Comprehensive Evaluation for LLMs.

Behavioral Fingerprinting of Large Language Models UltraEval: A Lightweight Platform for Flexible and Comprehensive Evaluation for LLMs

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-05T12:04:10.150544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:04:09.105197Z digest=sha256:c6f368aecd328037e66d2997380af29703d65273c8c05415e88d61b970faf1ca

Observation 9a928db3-6004-4dec-9b89-606bdd2a7e46 · outbound

This paper cites The waluigi effect.

Behavioral Fingerprinting of Large Language Models The waluigi effect

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:04:10.632395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:04:09.266832Z digest=sha256:91840d6ea584b2c17a9393e586a02eb8195b178093d9db3a3e461500a71fe8cf

Observation 849c67b7-9266-4733-b9aa-fbd0f4c6c464 · outbound

This paper cites A Computational Framework for Behavioral Assessment of LLM Therapists.

Behavioral Fingerprinting of Large Language Models A Computational Framework for Behavioral Assessment of LLM Therapists

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:09.417454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:09.417454Z digest=sha256:90cef971df2247937c70f8385e474ec3586f5b2cf51c6f8be26b08bab57de0e4

Observation 3b0ec61f-ec61-41ac-863a-513826aa83fd · outbound

This paper cites Learning on llm output signatures for gray-box behavior analysis.

Behavioral Fingerprinting of Large Language Models Learning on llm output signatures for gray-box behavior analysis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:09.448435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:09.448435Z digest=sha256:3e404873563abfe03e0d4f8f3d78304dfdf2f90ebef4137a6e95e7ebd302bde8

Observation 8de0819b-1298-4f83-8711-5f0370f25770 · outbound

This paper cites Identifying Multiple Personalities in Large Language Models with External Evaluation.

Behavioral Fingerprinting of Large Language Models Identifying Multiple Personalities in Large Language Models with External Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:09.501889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:09.501889Z digest=sha256:8c60ff9965a0a20e88a89f4b9c8f902dd7d2ea41500baa9179fb67184fb8b3f3

Observation 930de71d-a330-491a-97db-56932d7a91e0 · outbound

This paper cites behavioral fingerprint.

Behavioral Fingerprinting of Large Language Models behavioral fingerprint

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:04:10.609246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:04:09.577795Z digest=sha256:ef7d78424ca49a1fb57afc91af3856d2b33f6d7d50c65142ae8dab1317262e95

Pith citing papers

Observation 124d1761-fbd5-416d-b0d3-f73abc5eda1f · inbound

The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs cites this paper.

The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs Behavioral Fingerprinting of Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T23:30:07.574017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T23:30:07.574017Z digest=sha256:762b64ffb67488004296d578d60c5865e74d3c20e51e197fadec90d065753271

Observation 7351c9bf-9528-4d79-88ab-9236f5bf24dc · inbound

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms cites this paper.

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms Behavioral Fingerprinting of Large Language Models

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:32:53.253917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T00:26:54.019256Z digest=sha256:c7c36342160da75ad4e30b7fdb511f78512430054e18987388f1d4a712335798

Observation 72afae11-08a2-48b9-b27a-eb434ad90a19 · inbound

Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement cites this paper.

Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement Behavioral Fingerprinting of Large Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T12:15:36.476876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:15:36.476876Z digest=sha256:9ce54ec7f09837b9ba02ecf561a35f7ac167376b53fe5d9644fa2f19997b5ce7

Observation 117ac520-201a-4050-bc00-7f7a033673a2 · inbound

Mapping and Measuring the Behavioral Evolution of Large Language Models cites this paper.

Mapping and Measuring the Behavioral Evolution of Large Language Models Behavioral Fingerprinting of Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:52:24.750109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T11:52:24.750109Z digest=sha256:2f91d4e6ed132ec0b9d10e911cabdbe4d73e11b0971630c96336980659d75805