Pith. sign in

Paper Citation Record · LEDGER

Tell me about yourself: LLMs are aware of their learned behaviors

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2501.11120.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11120 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:42:43.585155Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.989193Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1cc15828-65e6-45bb-9b97-2908bc6a33fc · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Tell me about yourself: LLMs are aware of their learned behaviors

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.585155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.585155Z digest=sha256:ac820773fda834bf53953538ee9b9bbb8d99c5c8b8ce316a6e48e89dfe67e8fb

Observation 511f238a-e3aa-4478-8dce-1dd08fed1d9a · inbound

Towards eliciting latent knowledge from LLMs with mechanistic interpretability cites this paper.

Towards eliciting latent knowledge from LLMs with mechanistic interpretability Tell me about yourself: LLMs are aware of their learned behaviors

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:40:05.176888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:40:05.176888Z digest=sha256:1b7c7d59dad0f60d1a2c8c032307e3d526156b2946b0b54b72762a2767232379

Observation d0d69705-22b9-400f-9c5b-10aa758c6c34 · inbound

Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling cites this paper.

Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling Tell me about yourself: LLMs are aware of their learned behaviors

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:41:56.079078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:41:56.079078Z digest=sha256:f481563e1a3ed774d17692b7ffc8770c30471dc997b443b6c08315c5e7117ad5

Observation 9ad14d7a-bed1-4286-be12-d33bd9948047 · inbound

VLMs Can Aggregate Scattered Training Patches cites this paper.

VLMs Can Aggregate Scattered Training Patches Tell me about yourself: LLMs are aware of their learned behaviors

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:05.788018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:05.788018Z digest=sha256:cf80ee05320fb74264a1654a85702250874441e5b82e804bbc04c188de35a733

Observation 20fc7a9c-5247-45fd-9664-9ca132f4b0d4 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Tell me about yourself: LLMs are aware of their learned behaviors

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.142377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.142377Z digest=sha256:5db702211a3f25f6eb77b692aeecc54feb2e42620527af4f602e16fd118fb319

Observation cea16080-f244-401d-9a0b-8ffdfad45e35 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Tell me about yourself: LLMs are aware of their learned behaviors

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:24.818491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:24.818491Z digest=sha256:3e4a4ccfe862945111b9a800a163caa4331c4fa27b3abf2154bc3845c072863a

Observation 13fcd288-f8df-4cc1-9fe7-642e61e93595 · inbound

Emergent misalignment as prompt sensitivity: A research note cites this paper.

Emergent misalignment as prompt sensitivity: A research note Tell me about yourself: LLMs are aware of their learned behaviors

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:58.060361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:53:58.060361Z digest=sha256:47ef9ce008d66c7708f8bfcdb077ad7939fcc1348970d6cbeec0250862613c7d

Observation e0f3776a-7cb8-4bf8-b343-a45db62e6e41 · inbound

Simple Mechanistic Explanations for Out-Of-Context Reasoning cites this paper.

Simple Mechanistic Explanations for Out-Of-Context Reasoning Tell me about yourself: LLMs are aware of their learned behaviors

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:43.029366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:28:43.029366Z digest=sha256:caf82797c2f29a5fc4bcb9c4880f6bf9f1cb8498b6f7c5220f8846de9da7c0b0

Observation e33121b4-fdb1-4485-a060-897e7dc12f0c · inbound

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs cites this paper.

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs Tell me about yourself: LLMs are aware of their learned behaviors

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T16:57:07.514379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:57:07.514379Z digest=sha256:1f74863ac0e41046ab8bd34fc3d67e12861877dbc786be16030683a37e283b95

Observation bfb3b4a6-70cc-42e3-bd3a-a73a2b25a421 · inbound

Artificial Phantasia: Emergent Mental Imagery in Large Language Models cites this paper.

Artificial Phantasia: Emergent Mental Imagery in Large Language Models Tell me about yourself: LLMs are aware of their learned behaviors

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:44:22.783938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T21:41:39.111769Z digest=sha256:015d525cf7960d65a7e1531d6b7f0099f5f0c49207994f9a655f8dae4fb1b00d

Observation 7e1e79b9-6930-4740-95d6-05578307b2af · inbound

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models cites this paper.

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models Tell me about yourself: LLMs are aware of their learned behaviors

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:51:00.523876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T06:50:13.312970Z digest=sha256:6e287bf82ec2ea652ac5a4d95b0d29f3f62e421a31bd023467dbf456e908f7a2

Observation 1f0856e6-598c-45e3-a773-2ed5005c40bf · inbound

BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling cites this paper.

BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling Tell me about yourself: LLMs are aware of their learned behaviors

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T19:14:46.482128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:14:46.482128Z digest=sha256:8c2180b41b8be408d4720114e7d47f61f794f6d9e013bc286126d298dc819029

Observation eeda695c-7298-426f-bc5f-6a99f20c9cf9 · inbound

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models cites this paper.

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models Tell me about yourself: LLMs are aware of their learned behaviors

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:23:26.935064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T23:18:34.388467Z digest=sha256:0507d8945efd4f9b3583b1bf38e3c20178b0add4f24fb357fd0c16dd93b87dba

Observation 504b10a1-38c8-43d3-b5fd-ce79e5baee30 · inbound

Characterizing the Consistency of the Emergent Misalignment Persona cites this paper.

Characterizing the Consistency of the Emergent Misalignment Persona Tell me about yourself: LLMs are aware of their learned behaviors

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:29.637018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T08:06:58.635103Z digest=sha256:7aad0e8c9e92d88e9148c8d65219254129da3e52b5eeb6e62efa42f29b796d8a

Observation 8dc0b65d-d0de-4af9-ba87-11e2a8371864 · inbound

The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences cites this paper.

The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences Tell me about yourself: LLMs are aware of their learned behaviors

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:16.069847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T17:08:29.027351Z digest=sha256:3f6adb7f42e1e195714aa64a4dc586a0b687b4a6d04dba866d5c7706a46a6ac0

Observation 828add19-7a57-4b93-b49d-d11d1beb5b0e · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions Tell me about yourself: LLMs are aware of their learned behaviors

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:34:48.503379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:fb0868e9d3ca6ceecaa901ebb87bbb2e7c700e245dfaa81826e03f04069c5e2d

Observation ab053d5b-a1fd-4579-b7ca-8f749530f028 · inbound

The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition cites this paper.

The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition Tell me about yourself: LLMs are aware of their learned behaviors

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:42:35.858382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T19:34:27.009061Z digest=sha256:316bef1b618e4f7c69b9d7393ee6d530bea08ed59bff7ce2a0a2e7f07896fcff

Observation bd11fd80-7374-40e7-9982-88e021787f30 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Tell me about yourself: LLMs are aware of their learned behaviors

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:44.990464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:f99a816fedb6adf11402bd1d1e62224bf933f5c85de1966a585872665d6d0c0c

Observation 053a97c1-4b83-4f52-b93a-b49ea9377533 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Tell me about yourself: LLMs are aware of their learned behaviors

Reference 224

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:55:59.611181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:85039c6fd495a93a7257f6edf9bd5d441ce3f971ef46767b4d2505f6b24924c2

Observation e2b060b1-c3ee-45f7-80fc-197cfd40c59f · inbound

Out-of-Distribution Generalization of Risk Aversion in Language Models cites this paper.

Out-of-Distribution Generalization of Risk Aversion in Language Models Tell me about yourself: LLMs are aware of their learned behaviors

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T07:15:51.883780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T07:15:51.883780Z digest=sha256:4213a9fbb56c54068896e60150c35eebdae0670f2b824e81d2d0848dc77874ed

Observation 57bc6d85-c289-495b-ad03-e8e8bc3323b7 · inbound

Verbalizable Representations Form a Global Workspace in Language Models cites this paper.

Verbalizable Representations Form a Global Workspace in Language Models Tell me about yourself: LLMs are aware of their learned behaviors

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T23:15:18.170826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:15:18.170826Z digest=sha256:afc45b6a34105b57d336adc0c71430655962c3ad6631ab4e13b4478652c95286

Observation 53abd113-7977-453a-8b82-67a813624b44 · inbound

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary cites this paper.

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary Tell me about yourself: LLMs are aware of their learned behaviors

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T15:08:36.489413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:08:36.489413Z digest=sha256:80b0959229060c96682e488ac65a288a9bff94c65ba55254bef558c49f003f38

Observation bea06045-4976-4100-b131-86ee60aeda0e · inbound

Emergent Misalignment Recruits a Pre-existing Persona Subspace cites this paper.

Emergent Misalignment Recruits a Pre-existing Persona Subspace Tell me about yourself: LLMs are aware of their learned behaviors

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:07.723958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:46:07.723958Z digest=sha256:40dc556e0cc51c0cf4d2ac79c21fc01a1253cee87b9119a3ad302c5fc3482667

Observation 612afea2-fffd-490a-9f3f-0aa6930f6634 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Tell me about yourself: LLMs are aware of their learned behaviors

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.801901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.801901Z digest=sha256:4ade6d2c496dff51ea8340c964c13af46bb66e060e690ac79dbffd60f4c4c32f