Pith. sign in

Paper Citation Record · LEDGER

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

As of 10 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2607.07903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07903 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T15:42:46.392593Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:59:10.702144Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-08T16:59:12.165147Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact12
  • verified fuzzy48
  • unresolved2
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 680e0b50-e187-4317-a51e-20bdf9bb804d · outbound

This paper cites Language models are few-shot learners.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Language models are few-shot learners

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.566944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:06a212a9d3ec538efac5fe0ee14e25384be79ae21c0857d27b33eebf9986817f

Observation 8383f39d-3cce-4125-ab39-1861964c3d2c · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Attention is all you need.Advances in neural information processing systems, 30

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.568568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fd4cb218dc4b7b0a0911b5063a43d2e846f585d31ccdab8fb4a42946f6157b0d

Observation 9c69b763-8d3c-48a5-b556-67661ba8f832 · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explaining and Harnessing Adversarial Examples

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.187453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b65b20482a7cf9b78fa1b4b63654603659760ec74e22da25439cc7fd1b0c0d40

Observation f1c3ed2e-5048-4696-a2b7-60fac5d7d455 · outbound

This paper cites Towards deep learning models resistant to adversarial attacks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards deep learning models resistant to adversarial attacks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.563775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:747d541b5b165ca4a4f81a2bb354393c09a39b458805eab79bf65ed910cdb2d7

Observation 59b98507-bd75-4512-8aa5-5b54239ac5ff · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.185136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b968b2ca4bf8fd5d329b43593a78b0af9cfb351af2a7a3db460a999a1a4b69a7

Observation c27c8f31-64b1-46b0-9b8a-559bb29ca91a · outbound

This paper cites Training language models to follow instructions with human feedback.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Training language models to follow instructions with human feedback

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.570237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:da748abda8c99502ed01bbce0a83aff417ba079809dc132aebe021484af44d0e

Observation 1b58cf5e-ea7b-4e69-8c94-62bf1c701a03 · outbound

This paper cites Bridging Interpretability and Robustness Using LIME-Guided Model Refinement.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Bridging Interpretability and Robustness Using LIME-Guided Model Refinement

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.205521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e2c824d1572a916973912c3a61c5712c13d8ffcfd23cdc30187edbeee558cfd0

Observation 169b105d-5578-4346-8fd4-1a8ad474fdfd · outbound

This paper cites Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.571948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d04b131157d9bfe8b24e01ea54913b3d69208e584677f3a497946d87b9e1788a

Observation 37d3f899-fd79-4665-9b35-54a34dd9b37f · outbound

This paper cites Explainability-guided defense: Attribution-aware model refinement against adversarial data attacks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explainability-guided defense: Attribution-aware model refinement against adversarial data attacks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.607238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4e66725cb4ccbffc20ad8098866ea5028ddce6223a7ee8f7095e587184c66c3b

Observation 601bde07-b71b-497b-ae28-47824b59c8f6 · outbound

This paper cites Representation learning and nature encoded fusion for heterogeneous sensor networks.IEEE Access, 7:39227–39235, 2019.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Representation learning and nature encoded fusion for heterogeneous sensor networks.IEEE Access, 7:39227–39235, 2019

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.622563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:abe6691b8ef4c27e0179890c9fe501472a84ed45c7216f756bb2a7bc1f6d48c7

Observation 08ace148-a560-4c8f-befa-881010b32548 · outbound

This paper cites Congestion aware dynamic user association in heteroge- neous cellular network: A stochastic decision approach.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Congestion aware dynamic user association in heteroge- neous cellular network: A stochastic decision approach

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.648376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:cf5157ad4237a8f40573f4739ce2ad49b0dbecc70dad878da8ba79c83ebeb8ce

Observation 38a14a71-814c-4758-9711-2cb9b5ec9305 · outbound

This paper cites Explaining the behavior of neuron activations in deep neural networks.Ad Hoc Networks, 111:102346, 2021.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explaining the behavior of neuron activations in deep neural networks.Ad Hoc Networks, 111:102346, 2021

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.581786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:387bc1c76bdb879f4d2df116d5e53ca5f9c7de1a30bf30952ac0bd699308a063

Observation 2e3ff8f1-6604-4e63-820e-64a19a80671b · outbound

This paper cites Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.596843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:bf965a079ddb4bf16297799dcb5cec70d8abc63971363b9b3ad42d81e7ed66e8

Observation 2a0bb3cb-2c49-49a7-b1a4-cd8b05c3e35b · outbound

This paper cites Deep reinforcement learning based computation offloading for mobility-aware edge computing.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Deep reinforcement learning based computation offloading for mobility-aware edge computing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.573528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:1b6a2fe2f63770f6d86832433e76c48c487a95b7370c4519ee99da89aa2b687a

Observation 01fb565e-efb7-4e11-966d-9208c0af634f · outbound

This paper cites Improving robustness of deep neural networks via large-difference transformation.Neurocomputing, 450:411–419, 2021.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Improving robustness of deep neural networks via large-difference transformation.Neurocomputing, 450:411–419, 2021

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.577671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e539bbe24a6cddf53c202d495e6c7538da84522c554064c18086089a82d349e4

Observation be1049dd-cf8c-461f-b0b2-d8dcf7021290 · outbound

This paper cites Looking beyond content: Modeling and detection of fake news from a social context perspective.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Looking beyond content: Modeling and detection of fake news from a social context perspective

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.646719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:5d50fdbec1b7df6efbd1b54c289c4c26d5012e25856bbe421023e2912346e346

Observation 409e11a4-216b-4dfc-845d-1bb5de8c45be · outbound

This paper cites Layer-wise entropy analysis and visualization of neurons activation.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Layer-wise entropy analysis and visualization of neurons activation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.579372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fdebf3581f8a0f6721952ea08e89b4bd083f7734fe72249da39098e8e3139e29

Observation 65310eba-0524-4ea7-98df-0a2f9476d6ff · outbound

This paper cites Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.189844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d955349c7c3f49d2ea25e1d0377a84719977249809fd394d220ce4c11145729d

Observation 47b283e7-d437-4037-ad18-9d223467b674 · outbound

This paper cites Explainability- driven defense: grad-cam-guided model refinement against adversarial threats.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explainability- driven defense: grad-cam-guided model refinement against adversarial threats

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.650203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:11eccd8c43a84dd7147ca60aa67fac3329e55e7d59f61c7212375e234d5b2e27

Observation b2927e6b-2bb0-4846-8c83-fea33baa0375 · outbound

This paper cites Expert-guided explainable few-shot learning for medical image diagnosis.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Expert-guided explainable few-shot learning for medical image diagnosis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.565360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:da70ca3eabecb71969eb0c5dc2e145003732523dfe7af1b04d76cd8aa39414de

Observation 44002a84-b16f-4c4b-b3a4-c475c6d416dc · outbound

This paper cites GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.196757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:06405cddf21c94b7b450d4968d093d4e62cb16913255bb55b3ca6ebc0c052b03

Observation 7d378a65-33d7-4f85-832f-c5ff4b39d6ba · outbound

This paper cites Toward carbon-neutral human ai: Rethinking data, computation, and learning paradigms for sustainable intelligence.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Toward carbon-neutral human ai: Rethinking data, computation, and learning paradigms for sustainable intelligence

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.575486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:06fcf19e47140dfdbb8e684f25b9cc3149c10a208fc3402ce7428a5635c12c20

Observation 8a5350e3-9bc5-4180-a294-53ad400d5a1f · outbound

This paper cites Expert-guided explainable few-shot learning with active sample selection for medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2026.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Expert-guided explainable few-shot learning with active sample selection for medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2026

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.643323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:61ed83b4f125340ceaa1dbff461555cfd925e2d2b2d3474328da82f7ae5a3515

Observation cd9bbc7a-77e1-4460-9b12-369abd128e75 · outbound

This paper cites Acting flatterers via llms sycophancy: Combating clickbait with llms opposing-stance reasoning.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Acting flatterers via llms sycophancy: Combating clickbait with llms opposing-stance reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.645010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ba3f669038a2a16db31295309966b9e4d464bb992e5dda66bcdab951cace2472

Observation 56c2b89e-d054-4708-9420-f080785bbee7 · outbound

This paper cites Bridging symmetry and robustness: On the role of equivariance in enhancing adversarial robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Bridging symmetry and robustness: On the role of equivariance in enhancing adversarial robustness

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.651911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:c42f9d05368cb69b7b53601d26c1013cb4eff667ccc0635bcc19a4475bd88b00

Observation 1ef30688-2eb1-4f7a-9580-3f2f42d3d60f · outbound

This paper cites Channel- selected stratified nested cross-validation for clinically relevant eeg-based parkinson’s disease detection.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Channel- selected stratified nested cross-validation for clinically relevant eeg-based parkinson’s disease detection

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.638178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:82b810c416533b78afd7d18a6cfd0ab2926a6db310f1bc77cf7f7b659487a9db

Observation 21ce97ec-117e-4049-b15b-71a8d95945cc · outbound

This paper cites Winsor-cam: Human-tunable visual explanations from deep networks via layer-wise winsorization.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Winsor-cam: Human-tunable visual explanations from deep networks via layer-wise winsorization.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.636392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0740b2d91cafccbebbb8e3eb636d0c4132ca0231158851e9780f7bbefca16230

Observation 4d088faa-8f00-4701-a457-ed8ef819dd97 · outbound

This paper cites Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.639983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fabb9a562adbbe9d0b2acc3d849439f7058427811b642c3441e4721c3741bbaf

Observation 9826756c-7d04-4d67-abcb-0816e0db9894 · outbound

This paper cites CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.210188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:7554694d16b8fd89dab420d3cb98e30a30836a9e0d5b7901f9ec7a8b83c8ef72

Observation 125e7890-7f5b-484f-9f52-1cd7f5b3d5e7 · outbound

This paper cites Zoom in: An introduction to circuits.Distill, 5(3):e00024–001.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Zoom in: An introduction to circuits.Distill, 5(3):e00024–001

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.632952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0b1936184a38ee04efcbd2810934025602327bf25ff07fceccbf9f0371e187e7

Observation 0e02460d-b20c-4e45-877c-28c239cd849f · outbound

This paper cites A mathematical framework for transformer circuits.Transformer Circuits Thread.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs A mathematical framework for transformer circuits.Transformer Circuits Thread

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.634679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fc97bc096fafbbb92f24f2a7be373e395f05301a579895d2b8a9ff0e3c6a48c6

Observation 5d3ef84b-4c08-426c-bd4a-f016d587c3ad · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-07-10T15:47:23.641528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:c44b635a2b393d3c1d76e557d34d1854ca78dc6cd0965d09988583ad5c19e2f6

Observation f64d98e3-b05d-4a52-ac51-21ffc217c163 · outbound

This paper cites Axiomatic attribution for deep networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Axiomatic attribution for deep networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.627734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:2a6a8ef92e088f9a66d817e9564eab21b153ce6962e27036a575fa5b281267bd

Observation 19c3e22f-0b0f-4cb3-bd67-a265097642f9 · outbound

This paper cites Intriguing properties of neural networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Intriguing properties of neural networks

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.207919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b565cb42a4be4bf92ff0e498f90a2b1ee6b8f4cea8b7f49a658ce7013b1b8835

Observation 9d1f8654-46f5-4d79-b9f4-af89aeedf039 · outbound

This paper cites Towards evaluating the robustness of neural networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards evaluating the robustness of neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.629384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e9dc9a663cefe0946a54b3494dc64a5375ff8589544e7363e088c7e29e9353fa

Observation 696d66d3-169e-44dc-b7eb-64a09802ec10 · outbound

This paper cites Visual adversarial examples jailbreak aligned large language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Visual adversarial examples jailbreak aligned large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.625900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:cb7dc78a8ca5acafe2b63000b9c8a1eb6f6467fb218509b7afcb89eba92f28e5

Observation 418d9386-7a86-4bdb-ac2a-3bacb33122ee · outbound

This paper cites AutoDAN: Interpretable gradient-based adversarial attacks on large language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs AutoDAN: Interpretable gradient-based adversarial attacks on large language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.631154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ea49dcc2ec3b4145a330df8f218a43f951f55f5cb7b2a98cb6b503d925b8713c

Observation 2f013e77-14d3-4bad-8f75-077096bd4231 · outbound

This paper cites DOI:10.23915/distill.00010.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs DOI:10.23915/distill.00010

Reference 38

Resolution
metadata mismatch
doi, observed 2026-07-10T15:47:22.952574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:69fcb8174c9ea73e21f799f3a558295149c4086bf37d959987dc250a7416b036

Observation 834c07d4-92ff-4961-8321-d67b623b6f49 · outbound

This paper cites In-context learning and induction heads.Transformer Circuits Thread.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs In-context learning and induction heads.Transformer Circuits Thread

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.653540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:5030ea897d711d8d41367c6ffcc0ef898c76768a9b8aa4a372c514154db30a77

Observation e9fd6d66-7bb3-4da6-90b3-89fb35b85ab4 · outbound

This paper cites Sparse autoencoders find highly interpretable features in language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Sparse autoencoders find highly interpretable features in language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.619321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4e38bda3294e402cbaa23bfee9be2a34bc4f90a218bdafa5acb672c1181d03e6

Observation 009f8fb8-8db9-4acb-8ff1-5628c34fbb89 · outbound

This paper cites Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.203361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0c1ba40dd4f719b98a4cc493ca3da3b09852054f51c3e3d031c97dc7c52bf64a

Observation 5905a272-3c72-45d8-b89e-9b09ca947f98 · outbound

This paper cites A unified approach to interpreting model predictions.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs A unified approach to interpreting model predictions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.620873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:9201a29ea6f88dce5c5bce1dcb0078451c7c6977ae81f869821cc260ec02957c

Observation b3b1e879-d152-4107-99f6-b041cf7d6430 · outbound

This paper cites why should i trust you?.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs why should i trust you?

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.624246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d673131a6ece4c1a7872b929fb7ec8ec41cdfa8c14600d7e760027c42bd827c1

Observation c1ac42e5-49e7-4172-b397-bfc69a7a492f · outbound

This paper cites Attribution patching: Activation patching at industrial scale.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Attribution patching: Activation patching at industrial scale

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.614430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ff14ad724fd8bc7dc828aec9d40e66c15b8612abdd5ee5067ae1fb9d642956c8

Observation 9e490a52-f32a-4541-ab21-a0c902bcf5e2 · outbound

This paper cites Causal abstraction: A theoretical foundation for mechanistic interpretability.Journal of Machine Learning Research, 26(83):1–64, 2025.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Causal abstraction: A theoretical foundation for mechanistic interpretability.Journal of Machine Learning Research, 26(83):1–64, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.616097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:6e37c21d9525bc0e7cec82e457de0c96dc7cd0f06c9ee8c6229be25b122843d2

Observation 50fd132c-5ad3-4a3e-9f9f-c77909e573fc · outbound

This paper cites Investigating gender bias in language models using causal mediation analysis.Advances in neural information processing systems, 33:12388–12401.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Investigating gender bias in language models using causal mediation analysis.Advances in neural information processing systems, 33:12388–12401

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.612655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b4e4e4b401cabac7af489f39ac1212077e06f00b5c60bf05b3c1769adcf424ca

Observation ce825ff0-56c2-4612-98cd-990467af4351 · outbound

This paper cites Locating and editing factual associations in gpt.Advances in neural information processing systems, 35:17359–17372.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Locating and editing factual associations in gpt.Advances in neural information processing systems, 35:17359–17372

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.617762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:06c243386a35c0b7ba6688d6c8590295f236e687175df005f894161a18015296

Observation 3185777b-bdb9-46c8-8c98-75c128dab86c · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs LLaMA: Open and Efficient Foundation Language Models

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.201157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:5a6b69508a6ee046b03966957e82d530a61f5bff1830914a9c19dcfbc89ab584

Observation e8b718b0-36eb-476e-9798-19ee400e2889 · outbound

This paper cites Jailbroken: How does llm safety training fail?Advances in neural information processing systems, 36:80079–80110, 2023.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Jailbroken: How does llm safety training fail?Advances in neural information processing systems, 36:80079–80110, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.605525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:32b3e8817bf8a549a63031bc686db64b19408016bf68e2046724cc8ff5148027

Observation 167d61a9-581f-43ff-a259-c3e74f7c7326 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.198943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:943a0ebd3cb306bcef7b446eb7a2a3cd1ab858471bf28f2a9356756641bc8992

Observation da099723-6c95-48f6-8c9a-fb37b679a2fb · outbound

This paper cites Adversarial examples are not bugs, they are features.Advances in neural information processing systems, 32, 2019.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Adversarial examples are not bugs, they are features.Advances in neural information processing systems, 32, 2019

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.609025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:dee80ba4defc223b425f6d40e4c2e3fcc6212fae9b879c6c4953c2806f2ce933

Observation f6dd829d-a2b0-4d9b-8998-b4a26c6e62d7 · outbound

This paper cites Learning important features through propagating activation differences.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Learning important features through propagating activation differences

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.600092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:a673b00e6ec2a75a97d8dc98efe080b956b117f7e984a41a09861def937bf2fa

Observation 979d68bc-b92e-48b9-8b2d-8381f37b570d · outbound

This paper cites Many-shot jailbreaking.Advances in Neural Information Processing Systems, 37:129696–129742.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Many-shot jailbreaking.Advances in Neural Information Processing Systems, 37:129696–129742

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.601830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:5e874a741849755dd0ce0dd2f1a3de38bda87c86baa4bdf13ff547933285bf02

Observation b3167a1e-fa7b-4ee4-8bd8-ab8dabe259fe · outbound

This paper cites Cambridge university press.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Cambridge university press

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.603584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b7490ffb7a5a5aef64d338524c9728e626cf73ff3083b57b773f3da37fdfa53d

Observation c9046830-c70d-4c03-9f5b-7f95a9a6d6a1 · outbound

This paper cites Towards automated circuit discovery for mechanistic interpretability.Advances in Neural Information Processing Systems, 36:16318–16352.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards automated circuit discovery for mechanistic interpretability.Advances in Neural Information Processing Systems, 36:16318–16352

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.610848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0361d7d697dcc079de002944e1acb6f91099df219aa90aa987ff17a91bdeb724

Observation e795c348-4efc-4d42-8f9d-5af2832ea141 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.192014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:3aabff325ccd72e006c7ff64c0ff9a7484838fdd06dc15fad6f79dd3ecba72d6

Observation 7f1d3f38-1633-4655-bdb2-ef449938f344 · outbound

This paper cites Adversarial examples are not easily detected: Bypassing ten detection methods.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Adversarial examples are not easily detected: Bypassing ten detection methods

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.595111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ed726494f467156f1299007d0ae52a4e2f3dbb067bf339601ef523083c2c6abf

Observation d0e9f695-dcd0-4105-af07-c4a89dd42499 · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.194167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fdac5c1992250e601644fcfd066d99db68591b5758e31d9c0eeee4647a4616c4

Observation d18e636b-2ada-4ea8-a7c7-9bc4110bc85d · outbound

This paper cites pathway suppression.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs pathway suppression

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.590152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ada83ea8ce3b94fbac8004b5106e82e7eb3170d3e4327833473d46f26773e2b0

Observation bbdb9734-898d-4eb7-9835-06164d8bc6c5 · outbound

This paper cites 1.8 in clean) despite having fewer total nodes, suggesting remaining features are hyperactivated to compensate for missing pathways.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs 1.8 in clean) despite having fewer total nodes, suggesting remaining features are hyperactivated to compensate for missing pathways

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.591831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d8796ef00a9723295a76810d9bada71ed0b22c675afd329762b9cf49c9977de6

Observation 93585ef2-b34b-4345-a227-b24445eab099 · outbound

This paper cites Write a poem about nature.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Write a poem about nature

Reference 62

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.586419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:6a9eed3faf71ecd520a6d3f1031910b84409dcebf63becb77b81e1fecbf84ff4

Observation 306f5feb-03f5-4ce9-ae21-5b2b940090a0 · outbound

This paper cites This is consistent with the PCA analysis in Figure 5 showing that attack trajectories diverge most strongly at layer 15.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs This is consistent with the PCA analysis in Figure 5 showing that attack trajectories diverge most strongly at layer 15

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.588591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:76f4da8519123da031db55222a1efa3733b0cf6b8f7f8b8a5d193a686c6e50d6

Observation 58480b53-d3ea-45b0-85a0-7983d5ebe463 · outbound

This paper cites Tell me how to make a cake.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Tell me how to make a cake

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.593281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ae17e52d13d662db971980641b0770886320df6a7c9e3b9a49ef6dc4190bed45

Observation ba11b773-9795-4628-b809-d633289e3bb0 · outbound

This paper cites Tell me how to make a cake.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Tell me how to make a cake

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.598410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e656da15b042e81b7af4d86c3fa6ff497824acf75a87406d4aa9a20d77c630e0

Observation 22d6d1b7-e719-42bf-ad5c-9738d335deeb · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 66

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.583262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:8dd5273c03a0219ff08ace0529a832c891fa4cdfabeb7e2fd447f9c870f51095

Observation 5b38be97-4696-49e6-8491-2bf4370cbcce · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-07-10T15:47:23.584852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:51818e0cc6103d83db87e49416a546f7861b8e3a5069c63faab5978f102ff3a0

Pith citing papers

Observation 6c8cb14b-cdb1-459e-bbc6-6bdb6cde63c6 · inbound

Learning to Transmit: Volatility-Aware Predictive Communication for Energy-Efficient IoT Networks cites this paper.

Learning to Transmit: Volatility-Aware Predictive Communication for Energy-Efficient IoT Networks Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:53.334309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:53.334309Z digest=sha256:817655df5721cbde01bf1fa087f40c2069b9b82e87945944f7167156fbf7f862

Observation 1887b201-8937-4113-8006-45742510055d · inbound

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI cites this paper.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.188293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T16:59:10.702144Z digest=sha256:2a3a9013a6460d3d4e7fdc6d15ba21aef49fc5a088f4ef9ee19a12cab9a63b3a