Pith. sign in

Paper Citation Record · LEDGER

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

As of 10 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 3 inbound Pith citation observations for arXiv:2506.05774.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05774 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:30.371390Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T20:04:33.986438Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 31d3c6a6-e3f1-4bd7-9ff0-7e42d0e1b6c2 · outbound

This paper cites write newline.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:23.946527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:23.946527Z digest=sha256:b2c443df8a8e0bb572d652a767065491d9430b793c6514d568b9bf5dea874c56

Observation fb599d49-8b63-4db8-9560-558bb6d3e2b7 · outbound

This paper cites Sanity checks for saliency maps.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sanity checks for saliency maps

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.647224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:23.999650Z digest=sha256:c5edce59c4947e660fae4825e5d767e6f041601d9d71f7d7448b43d75b363988

Observation 443e0fb3-0cce-46c0-9582-da2cd15f5a31 · outbound

This paper cites A., Oikarinen, T., Kulkarni, A., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., Oikarinen, T., Kulkarni, A., and Weng, T.-W

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.512906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.118920Z digest=sha256:351362248026febbdccaaff71e8ea043e78abfe34242ea4fe72741f9996d8b44

Observation 982694b4-a84b-4674-8521-9cfcb4e94bfe · outbound

This paper cites Network dissection: Quantifying interpretability of deep visual representations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Network dissection: Quantifying interpretability of deep visual representations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.330669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.230402Z digest=sha256:b577d1698e9691526e4d634468a5e41488fb60b2a3fde30a60bc0e3596930d95

Observation 65cbf3bd-035c-42ee-8d6e-fdc3aaf4b87b · outbound

This paper cites Understanding the role of individual units in a deep neural network.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Understanding the role of individual units in a deep neural network

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.177225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.346189Z digest=sha256:b7e10df367eb176451e79898ec1d91c86c36ccf8b20fa5ffd66ee430e03da266

Observation d776b68d-e338-43ad-b376-b63aa2eae2f2 · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T10:18:30.669419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.449973Z digest=sha256:9d2526090d3ca9a4c67237b2d44175a0a04dfab51da691290b3a5a89b770edc0

Observation fda4dac8-6469-4a7a-be1a-5b7707196bff · outbound

This paper cites Language models can explain neurons in language models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Language models can explain neurons in language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.579347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.579347Z digest=sha256:705289af8900bd12de34d4023002c7ad1b721febadbcc5f03905249bc4409192

Observation e1b8db61-ae1a-4685-a0a9-8957651cb965 · outbound

This paper cites An Interpretability Illusion for BERT.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks An Interpretability Illusion for BERT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.712297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.712297Z digest=sha256:d14000913eb79a5b4386ea23199a85dd0b59cf398f8cb8f085325d1441980fbd

Observation ddeecd5a-83d1-45e4-8693-9053e03656b0 · outbound

This paper cites V., Lombrozo, T., Smith-Renner, A., and Tan, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks V., Lombrozo, T., Smith-Renner, A., and Tan, C

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.941318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.846246Z digest=sha256:29cf87b8a8aa90f96183f712a7e32c877351214146a2fc72f37171c7c374b146

Observation 741ff664-356a-44cf-b11d-757a07c91353 · outbound

This paper cites E., Hume, T., Carter, S., Henighan, T., and Olah, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks E., Hume, T., Carter, S., Henighan, T., and Olah, C

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.943620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.943620Z digest=sha256:0b0345df413ba2a7891128dfe4d859b0ea1344302132fd7ce5321fed4f5b6b9c

Observation a3cbfcbc-bfb7-4c27-afb5-a39cd0604421 · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:18:39.725040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.038821Z digest=sha256:662a955514572a5b6d82dd90e71178914119a68acbafb3c096235bd3da5b5f76

Observation ea0d4756-2973-42a2-8d8d-5bf9dbe68217 · outbound

This paper cites Sparse autoencoders find highly interpretable features in language models, 2023.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sparse autoencoders find highly interpretable features in language models, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.505824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.148810Z digest=sha256:ae65ee8526733688b1e5485fd5ad9c0a9f7087642ce0d7cec887d5d247dae3c8

Observation 20f8f412-4049-4bba-baf7-18cdfec4f235 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Imagenet: A large-scale hierarchical image database

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.255706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.255706Z digest=sha256:e63b8ff48f716eac27849d2f7b80a38aae11f4bd0b8456d1faf2cc48b79ec197

Observation 2d68b32a-d330-4664-b841-abf25556bc45 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks An image is worth 16x16 words: Transformers for image recognition at scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.376864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.376864Z digest=sha256:01e3ab1517b0a9bc93cdde2ed4cd704bc1d1d2317e4db96df570d218ac924f58

Observation 5df7d047-7536-4294-83ea-7abdaede6a7c · outbound

This paper cites Look at the variance! efficient black-box explanations with sobol-based sensitivity analysis.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Look at the variance! efficient black-box explanations with sobol-based sensitivity analysis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.295999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.486603Z digest=sha256:232efa71c1351fb8f5c96f88eec0306a2b103ce90d4cd395f25579f35de924e7

Observation 64acfe81-fbe2-48ff-aacd-e06e8a7517a1 · outbound

This paper cites Craft: Concept recursive activation factorization for explainability.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Craft: Concept recursive activation factorization for explainability

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.034285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.637005Z digest=sha256:7c445cd8f6a97bbfc4febded81a720f65d2751c4a1d59f661d0b8024fdfd070f

Observation c5e7ff17-afbc-4612-b615-9157c2f6c1b6 · outbound

This paper cites A holistic approach to unifying automatic concept extraction and concept importance estimation.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A holistic approach to unifying automatic concept extraction and concept importance estimation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.646488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.734826Z digest=sha256:4811f5b76ce00227bae4d41ec9c71e76d888e7291ae1d9c2c667420b20cfd31b

Observation 41e663dd-1009-4e1a-b7e1-ebe6fd15587f · outbound

This paper cites Interpreting the Second-Order Effects of Neurons in CLIP.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpreting the Second-Order Effects of Neurons in CLIP

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.835979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.835979Z digest=sha256:a5945636c993921ea6cc1263d2ec4de32aef538beda7d8651c4a9200602c76c6

Observation 0aebe1b9-cfd9-461c-a228-58d5bbbee6ad · outbound

This paper cites Y., and Kim, B.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Y., and Kim, B

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.405896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.969375Z digest=sha256:46acccd519e774fd53c27ca1a885381d4c32758c37cc8d40e28246384aa8f554

Observation b27ef534-15fd-476d-b34d-89fcc521190d · outbound

This paper cites Openwebtext corpus.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Openwebtext corpus

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:26.115122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:26.115122Z digest=sha256:e6c25373b20f79bf60e362407db06ecfd0b992e1b5ba4271d75c25d2c21af937

Observation 4fc2ab76-3405-44b3-9b96-221fec5afd3c · outbound

This paper cites Enhancing Automated Interpretability with Output-Centric Feature Descriptions.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Enhancing Automated Interpretability with Output-Centric Feature Descriptions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:26.276323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:26.276323Z digest=sha256:963bf1a09dfa012fb3a5d769dc11843a9ba63d00514bf528f6bf02c7da428e66

Observation 3252810e-e471-489f-9341-f7692e92b815 · outbound

This paper cites Finding neurons in a haystack: Case studies with sparse probing, 2023.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Finding neurons in a haystack: Case studies with sparse probing, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.078827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.408910Z digest=sha256:57ec7303a7dc5be214b3d2d1503898466b6c0c8d79a700f343e42a15f88e1049

Observation 20680be1-6e5c-47aa-bf25-62a04c2f6557 · outbound

This paper cites Natural language descriptions of deep visual features.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Natural language descriptions of deep visual features

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.729555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.579774Z digest=sha256:8bea7eb9314e89d5669f0f7134d8bbadb2019bf407409181519eed0f33f21456

Observation 7b86f752-d9d4-4d9b-ad8b-30dcd1c8e742 · outbound

This paper cites A benchmark for interpretability methods in deep neural networks.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A benchmark for interpretability methods in deep neural networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.410088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.742962Z digest=sha256:c7b853437a1c23c64df463f7a1192a66bc4f25fe7fd00df10dcb84a663ce57d1

Observation 841b2854-eeea-4cc2-b99f-f2dd6db3bbac · outbound

This paper cites Rigorously assessing natural language explanations of neurons.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Rigorously assessing natural language explanations of neurons

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.116707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.966423Z digest=sha256:91725bcfe02746d4baaccf86f98eef8cbf12644a337e91761562d5e06b148ccf

Observation cb117414-1cd1-4d38-b5a6-2f6190c41053 · outbound

This paper cites Benchmarking XAI Explanations with Human-Aligned Evaluations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Benchmarking XAI Explanations with Human-Aligned Evaluations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:27.145795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:27.145795Z digest=sha256:e912dbd01c783c41584e709f994c25e0676f6c162d9454e9f19de57672746e7c

Observation e4aef4ab-ebad-4e17-923b-b71a11d58b10 · outbound

This paper cites Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav).

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.907734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.294767Z digest=sha256:49c74f9268a6ccf6737498d665182007417eb8811fe4624cd4857ef9112095ea

Observation fc42fa35-a658-4dd4-9ee0-095ba060c2cb · outbound

This paper cites Human-centered evaluation of explainable ai applications: a systematic review.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Human-centered evaluation of explainable ai applications: a systematic review

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.667551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.412784Z digest=sha256:15e3c8537ff8f498fb0b51fb08d72678f64610957c0e5743e5a497866d34b271

Observation 39c7f83c-830f-4767-9cbe-6c83fdb00837 · outbound

This paper cites W., Nguyen, T., Tang, Y.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks W., Nguyen, T., Tang, Y

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:27.508320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:27.508320Z digest=sha256:d4277b116e5e4f85532012fbcfaf5656f976994b6767cc77bf9ce5fed03e3a34

Observation e36413ea-ccf0-4381-bfa7-87ec1fcdfe82 · outbound

This paper cites CoSy: Evaluating Textual Explanations of Neurons.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks CoSy: Evaluating Textual Explanations of Neurons

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T10:18:30.887986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.600701Z digest=sha256:a1be8768ca2e2a745b15eeac275ae91d77fe0164d19373976a1d7baa7aec783c

Observation 6cff607f-48dd-45df-851a-1755f53ef1fe · outbound

This paper cites Towards a fuller understanding of neurons with clustered compositional explanations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Towards a fuller understanding of neurons with clustered compositional explanations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.353211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.699706Z digest=sha256:0efefc3d4fb84fffe786af2624d18eca01c60ff317c92f150e8983021d45ed0a

Observation f43aedbd-0774-43d3-a5fc-743eace1b7be · outbound

This paper cites The importance of prompt tuning for automated neuron explanations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks The importance of prompt tuning for automated neuron explanations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.050252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.810365Z digest=sha256:39c678740c22eaf074f9ca02de762a3123cb94d8ded5609e74cbabc6300c5b3b

Observation 8a646d95-57a6-453c-b5a3-ccb7740d337f · outbound

This paper cites S., Iofinova, E., Frantar, E., and Alistarh, D.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks S., Iofinova, E., Frantar, E., and Alistarh, D

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.753595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.953068Z digest=sha256:ce32639fd1926c6b33f683a56ab0aabb4026739a43e6d1696b9d462f7e6e8f39

Observation 8120d656-9a1f-43d3-8b4b-cd633db81700 · outbound

This paper cites and Andreas, J.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Andreas, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.486095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.084315Z digest=sha256:455060c7e301e534feb0252c8eeba3c9e93e5d20c118cef8221bc5badadf4bfb

Observation 687ee0ec-43b6-4f35-a124-6eeea4ae6420 · outbound

This paper cites and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Weng, T.-W

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.201778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.224188Z digest=sha256:0959413686d1081fc7c0f9a480f11f03e47545d75a440a26e722d0116e545f4e

Observation 1bc06d1c-d5e7-4cad-9de9-7f234684040d · outbound

This paper cites and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Weng, T.-W

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.960193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.335730Z digest=sha256:46a786c5cd370599d1e4b7aabbe2cd3e4d28ee0fa3178e866d03f89ca9515505

Observation 1ec7d962-1cbc-4929-9bd8-75614f12f906 · outbound

This paper cites M., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks M., and Weng, T.-W

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.687582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.423556Z digest=sha256:e9099db9483345120a7ae0cafb0b1f2c20a34b0ae1e41e1a0fa15fff53558ee4

Observation 13083056-2851-4025-92d4-b346b0948116 · outbound

This paper cites Rise: Randomized input sampling for explanation of black-box models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Rise: Randomized input sampling for explanation of black-box models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.468718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.549606Z digest=sha256:7e1da65d35061a78dd2f46dd8319d1138d897f961d39c356e35012bd4cbfa9a5

Observation 83da099d-7541-40e3-b642-cc6b43a5f3cb · outbound

This paper cites Language models are unsupervised multitask learners.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Language models are unsupervised multitask learners

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:28.657772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:28.657772Z digest=sha256:9fd7aff2fab56ca218cf3f9e4a530844faaa6a6b5292261a90f92760c4373b4a

Observation 26844bd0-7900-4b8c-baa8-5bc0f78ceebf · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:28.761564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:28.761564Z digest=sha256:fbadf19699b23d2fbcd0a2dcfaa5826236dee9e19f13e61f039156def08a9140

Observation 8e4a5fa3-f1be-429c-baae-d2ffa26d94af · outbound

This paper cites Interpretable machine learning: Fundamental principles and 10 grand challenges.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpretable machine learning: Fundamental principles and 10 grand challenges

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.227782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.876081Z digest=sha256:6068094c3a4513ae8234c6153b97e097eda5601f792c3c4bfc2161e3d1c042e2

Observation 7c2c7182-d8ca-4e47-a8c6-c2d8dedcb1ad · outbound

This paper cites Find: A function description benchmark for evaluating interpretability methods.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Find: A function description benchmark for evaluating interpretability methods

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.962085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.007046Z digest=sha256:fbbf2164dac3341921d2a2f9380688de0f1cd99047a4946b11c0e5e31bb8b8af

Observation 50ae9110-4f71-4973-b64d-3276fc2a1793 · outbound

This paper cites R., Schwettmann, S., Wang, F., Rajaram, A., Hernandez, E., Andreas, J., and Torralba, A.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks R., Schwettmann, S., Wang, F., Rajaram, A., Hernandez, E., Andreas, J., and Torralba, A

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.061562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.061562Z digest=sha256:02774da372f9dffde89497d7b83da9dcbcc7dd1e46980f6fcb6bacc1a09d02af

Observation a87b55cb-db2e-4314-97a0-c589c4ac4e98 · outbound

This paper cites R., Antonello, R., Jain, S., Huth, A.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks R., Antonello, R., Jain, S., Huth, A

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.625224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.106445Z digest=sha256:ed903c538f7b242fb9aa8c1c8fb97cbb8ef2487b452fa2619aa44dc836b63f88

Observation 44c11d88-ed41-4905-8f7e-63e9cb2597b8 · outbound

This paper cites A., Oikarinen, T., Srivastava, D., Weng, W.-H., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., Oikarinen, T., Srivastava, D., Weng, W.-H., and Weng, T.-W

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.349513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.109993Z digest=sha256:26421f967ec686ea1b24db27de847db363cb68d46579faf4ef5db1fc1056cc19

Observation e5324f97-f8aa-42dc-ac97-eaa57593a494 · outbound

This paper cites Vlg-cbm: Training concept bottleneck models with vision-language guidance.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Vlg-cbm: Training concept bottleneck models with vision-language guidance

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.071418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.155437Z digest=sha256:7ee234af1e2912d7d744809094175f0b6c13b066e5c3f0e576ecb58c9001ec1b

Observation 6493ee0b-bdb3-4aa9-a640-f880e170eb7d · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:18:32.827490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.257892Z digest=sha256:a1eeb6e89a6c016c33d8ae1f7d6e9cc54ea5e9a352bf1de014cd99983e2d73bb

Observation 237ba960-b0fd-4e58-9e09-69d91c45405d · outbound

This paper cites L., McDougall, C., MacDiarmid, M., Freeman, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks L., McDougall, C., MacDiarmid, M., Freeman, C

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.421782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.421782Z digest=sha256:daf2522d3a13e5924b8d46fac3e269621b2dd7aa3cd7ca61508e60196ecc956f

Observation 20b18915-b1d5-4346-a85f-01f0f8b1c9fe · outbound

This paper cites The caltech-ucsd birds-200-2011 dataset.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks The caltech-ucsd birds-200-2011 dataset

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.533625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.533625Z digest=sha256:be9f384b1c35b267ac70a267c3b47514727d5c44bad831492a91a092874ed5c3

Observation f4be5411-3829-426d-8b45-a359c8a2c570 · outbound

This paper cites Post-hoc concept bottleneck models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Post-hoc concept bottleneck models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:32.500409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.687107Z digest=sha256:fcf3c4925645d8863fcb39a0d28b4ab26b29d86947651564bacbac797b45349c

Observation 97f6663c-32bf-4ab5-9aff-3689da9c74e3 · outbound

This paper cites Sigmoid loss for language image pre-training.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sigmoid loss for language image pre-training

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:32.246966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.810712Z digest=sha256:2e897ea2048204437e3b16072d62e1fe4ce0b797fd636e83da6653e08ae2d778

Observation 66377515-a453-4c18-b9a7-7234a9b035be · outbound

This paper cites A., and Rubinstein, B.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., and Rubinstein, B

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.972108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.969311Z digest=sha256:32c4736d763dd02b972159accab0d71baecf8d66cfec2fb8d3499a27baf7c41c

Observation 1af7c1b8-9f38-43f6-a1e7-2005a117face · outbound

This paper cites Object detectors emerge in deep scene cnns.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Object detectors emerge in deep scene cnns

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.754931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.103614Z digest=sha256:0bba11495dfdcca076556b204af4a83eed57a6a025e2d35dbbc35a50c7c70912

Observation 397a2133-f133-404c-848c-2fab1ee83488 · outbound

This paper cites Places: A 10 million image database for scene recognition.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Places: A 10 million image database for scene recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.460894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.211798Z digest=sha256:7481e87894a24c0cbe5fedc75f673c1d22bfbb0fac0c9436d9e86a719d849161

Observation 14cf9674-d9d7-4d6f-bb13-556dc7ded694 · outbound

This paper cites S., Klein, T., and Brendel, W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks S., Klein, T., and Brendel, W

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.203304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.371390Z digest=sha256:b079970b65e78364426be0dbd8664e451f28a9a4349e77a0ffdd1e5e264c00c7

Pith citing papers

Observation 737a2ae0-66cd-4cf6-bf96-d5d6b26088a5 · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 260

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:51:29.531017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T18:50:52.313363Z digest=sha256:f671cbc35c145172e979c5d882ee27346682f75c347ca4f2764a4a1e5a6318f9

Observation 04967e2f-35bd-4bd0-a0d9-2d8d9d96237b · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 252

Resolution
unresolved
no resolver link, observed 2026-08-02T20:04:33.986438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:04:33.986438Z digest=sha256:c3fe85dc230f445dc7f1264b603b1ac132b0a9c305a072a1f77d39f27e386f19

Observation 84620f7e-13be-40f8-9b64-7c857542c718 · inbound

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders cites this paper.

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.916851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T22:22:50.474397Z digest=sha256:a757e95312eed3f66267a63cb564215dcd6ef7bfd72be7a35f9b6c8ffa7d8227