Pith. sign in

Paper Citation Record · LEDGER

SelfIE: Self-Interpretation of Large Language Model Embeddings

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2403.10949.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.10949 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:06:55.812069Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ae97825f-34b0-4041-aa53-5e11bac06f2d · inbound

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring cites this paper.

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T21:06:55.812069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:06:55.812069Z digest=sha256:868be87f044a731c485df8a0a17e44133cbe263c7b692df561647c8254b328c5

Observation 00cfff6c-03a5-4672-a484-d602f0ba15bd · inbound

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race cites this paper.

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:14.424151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:13:14.424151Z digest=sha256:23bcb6dcbef9f1dd19d0da32448272370bf86bd9ca28f4d39f842c128c1d99c9

Observation 989dcab7-d626-48ea-9f7b-930feae4c211 · inbound

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models cites this paper.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.821849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.821849Z digest=sha256:441ec34c9f5577684702f9c4f8b7e7edabfe159e8ec40ffb83e133e84b7d3d7e

Observation e8686af4-31dd-459b-bdb8-9e77202f550c · inbound

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models cites this paper.

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:58.441959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:18:58.441959Z digest=sha256:45521970d1cfd57c54932dabf50e7606bcba188d9466f47f66df61d6443c34cb

Observation 1a78986c-6ee0-4bd4-bb99-2d35c1a3dd8d · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:56.622659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:7514c1e5eb03f79256cce94dbc041fd29d205ec467147b4792dc82040f006dd1

Observation 0869d5f4-9fb2-4c0a-b8af-4b47da688e1d · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:56.217131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:a243552679c53c5c2ba8a57ae6821ce717d8eb469bf5d40f02e95add94ad2612

Observation 5f2e92f1-f4b4-4221-ad2e-846f195f435f · inbound

From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Models cites this paper.

From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:01:11.987581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T06:56:18.332754Z digest=sha256:30445154a807e69c1e2365a8780ec1f7a250af1ffe6739c7e3d3d7313f9763a4

Observation 8b2aabb9-eacd-46fd-8753-56276e0c2046 · inbound

PRISM: Recovering Instruction Sets from Language Model Activations cites this paper.

PRISM: Recovering Instruction Sets from Language Model Activations SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-27T17:01:08.103275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T16:52:02.948457Z digest=sha256:f54112b304fdd642c90c68ec755ac6594b0e0a26db75758c7c4ddcf7bdae5edd

Observation 1973f5a3-a744-4e2d-8728-63df377402b6 · inbound

Verbalizable Representations Form a Global Workspace in Language Models cites this paper.

Verbalizable Representations Form a Global Workspace in Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T23:15:19.338403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:15:19.338403Z digest=sha256:04075e735149eee4771b1b654b908cf7ae4e9acf8b19961d18f6d9dfb66b2a98