Pith. sign in

Paper Citation Record · LEDGER

Interpretability Illusions in the Generalization of Simplified Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2312.03656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03656 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:59:11.526314Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:47:27.728455Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6116d3e7-b668-47a1-b5ab-62ef0ec4e3bb · inbound

Obfuscated Activations Bypass LLM Latent-Space Defenses cites this paper.

Obfuscated Activations Bypass LLM Latent-Space Defenses Interpretability Illusions in the Generalization of Simplified Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T16:59:11.526314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:59:11.526314Z digest=sha256:5e8cd6e547f8d19a1e84df10923ac0f8f802c920a5625d19a7051e12bec0f801

Observation f16e372b-3f59-4f61-b756-93891421f381 · inbound

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning cites this paper.

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning Interpretability Illusions in the Generalization of Simplified Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:34.186793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:34.186793Z digest=sha256:5bf520e67c1dc31e6e19615fa27b57a23c55d9152aa7507ea07de010ce6536e2

Observation f5c9e6aa-3ad4-4688-b966-1105b9e50060 · inbound

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims cites this paper.

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims Interpretability Illusions in the Generalization of Simplified Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:47:27.729853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T19:33:19.875024Z digest=sha256:c585d0479b6c32f66c1087ce530dba667cefbc98e8f18ad0c817cf53170201f6

Observation 1f2b9ba0-0c10-4259-807e-88dc0b332516 · inbound

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders cites this paper.

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders Interpretability Illusions in the Generalization of Simplified Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:15:29.651363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-01T07:15:16.674714Z digest=sha256:d3c8498167926875035d384935c26eddf029bcc327a1113f81d9dd5575d8c98b

Observation f3a52cfc-c8dd-4f5c-b080-4e2c37d4933d · inbound

LAWFUL: Law-Aligned Witness for Faithful Use of Latents cites this paper.

LAWFUL: Law-Aligned Witness for Faithful Use of Latents Interpretability Illusions in the Generalization of Simplified Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T00:47:37.317996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:47:37.317996Z digest=sha256:62107c3db826270729f5cd8013def5b7dd56af12f14fdcf63318848c8a4d9f36