Pith. sign in

Paper Citation Record · LEDGER

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

As of 18 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 3 inbound Pith citation observations for arXiv:2506.05774.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05774 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:30.371390Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T20:04:33.986438Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 31d3c6a6-e3f1-4bd7-9ff0-7e42d0e1b6c2 · outbound

This paper cites write newline.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:23.946527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:23.946527Z digest=sha256:159a0243edec71669bfe7e44c487d3feac0608e13eb873370c996a2514bf3105

Observation fb599d49-8b63-4db8-9560-558bb6d3e2b7 · outbound

This paper cites Sanity checks for saliency maps.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sanity checks for saliency maps

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.647224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:23.999650Z digest=sha256:64e6cb4f18291a9cfebe6c39ed71cda9317cfa7f19e19c4f8476eb068f52f35b

Observation 443e0fb3-0cce-46c0-9582-da2cd15f5a31 · outbound

This paper cites A., Oikarinen, T., Kulkarni, A., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., Oikarinen, T., Kulkarni, A., and Weng, T.-W

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.512906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.118920Z digest=sha256:d2dfc148d73fea361356d33553e5dc5ab7240070156cc072fc83eca21a97ee31

Observation 982694b4-a84b-4674-8521-9cfcb4e94bfe · outbound

This paper cites Network dissection: Quantifying interpretability of deep visual representations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Network dissection: Quantifying interpretability of deep visual representations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.330669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.230402Z digest=sha256:426907e8b4708c0f98dd6ba37cff8d62d1c77e9d3ea239ff3b9ee0bf6ec8b1b7

Observation 65cbf3bd-035c-42ee-8d6e-fdc3aaf4b87b · outbound

This paper cites Understanding the role of individual units in a deep neural network.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Understanding the role of individual units in a deep neural network

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:40.177225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.346189Z digest=sha256:4f95d40557152809c94529f20dd0f0191ce45bc3e00021c0e34d30c3ee9488f0

Observation d776b68d-e338-43ad-b376-b63aa2eae2f2 · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T10:18:30.669419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.449973Z digest=sha256:93a01af03e7f0aeb5a9c20cc3afde83a75a9fdd2022c185a20a78bc349218f8c

Observation fda4dac8-6469-4a7a-be1a-5b7707196bff · outbound

This paper cites Language models can explain neurons in language models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Language models can explain neurons in language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.579347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.579347Z digest=sha256:c722bf8cab15794a1a4cf0aa6ad6f90c81c0dfe55c1c3214d7e8c77dff02c1da

Observation e1b8db61-ae1a-4685-a0a9-8957651cb965 · outbound

This paper cites An Interpretability Illusion for BERT.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks An Interpretability Illusion for BERT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.712297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.712297Z digest=sha256:a29e55ff2aa01b95743bf6b2468e754767e1dd37c329b590605341c651f221a7

Observation ddeecd5a-83d1-45e4-8693-9053e03656b0 · outbound

This paper cites V., Lombrozo, T., Smith-Renner, A., and Tan, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks V., Lombrozo, T., Smith-Renner, A., and Tan, C

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.941318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:24.846246Z digest=sha256:fd7573e1bdce4b49f724804fceabfeddb91b1a6ca820ba10d1a4f06953d1db0a

Observation 741ff664-356a-44cf-b11d-757a07c91353 · outbound

This paper cites E., Hume, T., Carter, S., Henighan, T., and Olah, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks E., Hume, T., Carter, S., Henighan, T., and Olah, C

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.943620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:24.943620Z digest=sha256:e5491ebf5e4d2b475fcb632a45ebdb3823429239f2166c423809cede1d5f725f

Observation a3cbfcbc-bfb7-4c27-afb5-a39cd0604421 · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:18:39.725040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.038821Z digest=sha256:c43aee059478aa932f15c094a25b89743bf3dec5526ca6d523acfedb35947812

Observation ea0d4756-2973-42a2-8d8d-5bf9dbe68217 · outbound

This paper cites Sparse autoencoders find highly interpretable features in language models, 2023.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sparse autoencoders find highly interpretable features in language models, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.505824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.148810Z digest=sha256:8e7dfa742a2cce0baf846c8895d009a3314dc8d0516c160ff13bc282467e8603

Observation 20f8f412-4049-4bba-baf7-18cdfec4f235 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Imagenet: A large-scale hierarchical image database

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.255706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.255706Z digest=sha256:beb871b51486fc32d692e40bd76b9cfa318a5ad1332033efcfd1049874e1c810

Observation 2d68b32a-d330-4664-b841-abf25556bc45 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks An image is worth 16x16 words: Transformers for image recognition at scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.376864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.376864Z digest=sha256:d4d873267e7aa7490236fbdd3cab6539f4156289cbc7c6c8c0e2a422e87f6fc9

Observation 5df7d047-7536-4294-83ea-7abdaede6a7c · outbound

This paper cites Look at the variance! efficient black-box explanations with sobol-based sensitivity analysis.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Look at the variance! efficient black-box explanations with sobol-based sensitivity analysis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.295999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.486603Z digest=sha256:668119ef78d58a259d549922552ed469e147ea5f69acd757425410834481e9e1

Observation 64acfe81-fbe2-48ff-aacd-e06e8a7517a1 · outbound

This paper cites Craft: Concept recursive activation factorization for explainability.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Craft: Concept recursive activation factorization for explainability

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:39.034285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.637005Z digest=sha256:6f66e3f0610024b2ef8a93fe7d82173ec36b576e1746dfa274eb68cad56970c5

Observation c5e7ff17-afbc-4612-b615-9157c2f6c1b6 · outbound

This paper cites A holistic approach to unifying automatic concept extraction and concept importance estimation.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A holistic approach to unifying automatic concept extraction and concept importance estimation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.646488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.734826Z digest=sha256:f550e62409dd7f9605ed0b40f74085dfc1d28cd3aa1a6c508219e7a204f6dac2

Observation 41e663dd-1009-4e1a-b7e1-ebe6fd15587f · outbound

This paper cites Interpreting the Second-Order Effects of Neurons in CLIP.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpreting the Second-Order Effects of Neurons in CLIP

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:25.835979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:25.835979Z digest=sha256:01841daaa6b17d8edb70cdc0d78314c4d90283ba4128d20267acfaa9ed3dbe5f

Observation 0aebe1b9-cfd9-461c-a228-58d5bbbee6ad · outbound

This paper cites Y., and Kim, B.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Y., and Kim, B

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.405896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:25.969375Z digest=sha256:dba29a31d5b02f9900ec0801d1cbbddcb67b803487183fbcef789fecf6cc77dc

Observation b27ef534-15fd-476d-b34d-89fcc521190d · outbound

This paper cites Openwebtext corpus.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Openwebtext corpus

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:26.115122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:26.115122Z digest=sha256:3779001da47baced809ed6555710c29d7387d53b948ede8bf5438af6303d621a

Observation 4fc2ab76-3405-44b3-9b96-221fec5afd3c · outbound

This paper cites Enhancing Automated Interpretability with Output-Centric Feature Descriptions.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Enhancing Automated Interpretability with Output-Centric Feature Descriptions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:26.276323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:26.276323Z digest=sha256:a8545fd76d8d457f3954a5924a2732c34c34337f8ea64b798148e027bd827197

Observation 3252810e-e471-489f-9341-f7692e92b815 · outbound

This paper cites Finding neurons in a haystack: Case studies with sparse probing, 2023.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Finding neurons in a haystack: Case studies with sparse probing, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:38.078827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.408910Z digest=sha256:11b6bc8477f9b42b6480403814350e9cc78e5960a911a9828448cd85be59d363

Observation 20680be1-6e5c-47aa-bf25-62a04c2f6557 · outbound

This paper cites Natural language descriptions of deep visual features.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Natural language descriptions of deep visual features

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.729555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.579774Z digest=sha256:aed932e43efcc3439c85bfd288fb6aa7120ad6c7e7d7674fa4945d069392cfea

Observation 7b86f752-d9d4-4d9b-ad8b-30dcd1c8e742 · outbound

This paper cites A benchmark for interpretability methods in deep neural networks.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A benchmark for interpretability methods in deep neural networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.410088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.742962Z digest=sha256:b798c1cc15d3529651d74c9cfde47364db51bd8c63530118c8df25a03c9ff2c0

Observation 841b2854-eeea-4cc2-b99f-f2dd6db3bbac · outbound

This paper cites Rigorously assessing natural language explanations of neurons.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Rigorously assessing natural language explanations of neurons

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:37.116707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:26.966423Z digest=sha256:2e269f159c3ffb651abf25ea8bc14ef74f1bcdd9704ef8d2a8eba84ae469fa76

Observation cb117414-1cd1-4d38-b5a6-2f6190c41053 · outbound

This paper cites Benchmarking XAI Explanations with Human-Aligned Evaluations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Benchmarking XAI Explanations with Human-Aligned Evaluations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:27.145795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:27.145795Z digest=sha256:67717385b7a9999015907bf6e735e1da6b7f4492abd429770e72e813eac5d9db

Observation e4aef4ab-ebad-4e17-923b-b71a11d58b10 · outbound

This paper cites Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav).

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.907734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.294767Z digest=sha256:d32322f2b22ec978884c6157522fc7fa808887d6647a4bfa6b19d43563419db6

Observation fc42fa35-a658-4dd4-9ee0-095ba060c2cb · outbound

This paper cites Human-centered evaluation of explainable ai applications: a systematic review.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Human-centered evaluation of explainable ai applications: a systematic review

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.667551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.412784Z digest=sha256:e4fe05b49aa4d804ca514332297366c3be3af2bf859c5d195a0a181e5835d705

Observation 39c7f83c-830f-4767-9cbe-6c83fdb00837 · outbound

This paper cites W., Nguyen, T., Tang, Y.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks W., Nguyen, T., Tang, Y

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:27.508320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:27.508320Z digest=sha256:6690918f7e3008fc4a548a1d32641262ca6ef09c4e6724a91c12e2083a6b902d

Observation e36413ea-ccf0-4381-bfa7-87ec1fcdfe82 · outbound

This paper cites CoSy: Evaluating Textual Explanations of Neurons.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks CoSy: Evaluating Textual Explanations of Neurons

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T10:18:30.887986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.600701Z digest=sha256:43db8917df21c2d01ef64c623d9ea94e58075acbfa9f5054cf720b618dbd01b7

Observation 6cff607f-48dd-45df-851a-1755f53ef1fe · outbound

This paper cites Towards a fuller understanding of neurons with clustered compositional explanations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Towards a fuller understanding of neurons with clustered compositional explanations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.353211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.699706Z digest=sha256:92a9c16bb4d81be608237e68b08756a812bc1174ad6b885ff638c416d22e98f6

Observation f43aedbd-0774-43d3-a5fc-743eace1b7be · outbound

This paper cites The importance of prompt tuning for automated neuron explanations.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks The importance of prompt tuning for automated neuron explanations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:36.050252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.810365Z digest=sha256:decf3522c39110998e86cc54e729fc1be5e432744c4f60f1d47ac321a9dcb69c

Observation 8a646d95-57a6-453c-b5a3-ccb7740d337f · outbound

This paper cites S., Iofinova, E., Frantar, E., and Alistarh, D.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks S., Iofinova, E., Frantar, E., and Alistarh, D

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.753595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:27.953068Z digest=sha256:164437263a810885739a04ed623fdfec54e6d8338a7141197e99f933b871d177

Observation 8120d656-9a1f-43d3-8b4b-cd633db81700 · outbound

This paper cites and Andreas, J.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Andreas, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.486095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.084315Z digest=sha256:402970ed9e9b94d2f4d0e1da15d3d5686f979c71a87407e7d8b55fb080528398

Observation 687ee0ec-43b6-4f35-a124-6eeea4ae6420 · outbound

This paper cites and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Weng, T.-W

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:35.201778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.224188Z digest=sha256:3714a014519dbb7d51a8d68c92817a9b7d5960d9be80f5558b1644986aa6dc7c

Observation 1bc06d1c-d5e7-4cad-9de9-7f234684040d · outbound

This paper cites and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks and Weng, T.-W

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.960193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.335730Z digest=sha256:ad85105c4e2e9c06280b20b5dd2b662a6a1dc89eb99cdbb454b73f1968a65cd4

Observation 1ec7d962-1cbc-4929-9bd8-75614f12f906 · outbound

This paper cites M., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks M., and Weng, T.-W

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.687582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.423556Z digest=sha256:09525df0dad0d16d711bccfbd202d953d1c6770dc93e4a2ebec937f9841580e3

Observation 13083056-2851-4025-92d4-b346b0948116 · outbound

This paper cites Rise: Randomized input sampling for explanation of black-box models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Rise: Randomized input sampling for explanation of black-box models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.468718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.549606Z digest=sha256:73118787c0d6e6c40cdd207b17500574c68a94b7fb01ed7027eaeee37f7f76b7

Observation 83da099d-7541-40e3-b642-cc6b43a5f3cb · outbound

This paper cites Language models are unsupervised multitask learners.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Language models are unsupervised multitask learners

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:28.657772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:28.657772Z digest=sha256:86804b76397f895f75a53a3c1c41be9bb2dd026e1f997b90509f278a57af9aab

Observation 26844bd0-7900-4b8c-baa8-5bc0f78ceebf · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:28.761564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:28.761564Z digest=sha256:1d94f775a80ff5d050447baba4426bc92ad47beb6ac718a851afd83078f1d12e

Observation 8e4a5fa3-f1be-429c-baae-d2ffa26d94af · outbound

This paper cites Interpretable machine learning: Fundamental principles and 10 grand challenges.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Interpretable machine learning: Fundamental principles and 10 grand challenges

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:34.227782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:28.876081Z digest=sha256:ddbaf7ab6cf30dbc8fe2d9c59ee55fdd6b8ef3b067418a74558defdf62adc437

Observation 7c2c7182-d8ca-4e47-a8c6-c2d8dedcb1ad · outbound

This paper cites Find: A function description benchmark for evaluating interpretability methods.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Find: A function description benchmark for evaluating interpretability methods

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.962085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.007046Z digest=sha256:c5b3a328844bf312edf5a81ea4ba07657d9959ea6c85b87811e91c3c474fc55d

Observation 50ae9110-4f71-4973-b64d-3276fc2a1793 · outbound

This paper cites R., Schwettmann, S., Wang, F., Rajaram, A., Hernandez, E., Andreas, J., and Torralba, A.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks R., Schwettmann, S., Wang, F., Rajaram, A., Hernandez, E., Andreas, J., and Torralba, A

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.061562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.061562Z digest=sha256:565de0cda9eac271fb0d0dbe96c0ba81b5c8f0675759edae26a8e0b9782c728f

Observation a87b55cb-db2e-4314-97a0-c589c4ac4e98 · outbound

This paper cites R., Antonello, R., Jain, S., Huth, A.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks R., Antonello, R., Jain, S., Huth, A

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.625224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.106445Z digest=sha256:5f37241d2980e7a2d159334009201aee84c9435c083d1e7f21639125aba0a3d8

Observation 44c11d88-ed41-4905-8f7e-63e9cb2597b8 · outbound

This paper cites A., Oikarinen, T., Srivastava, D., Weng, W.-H., and Weng, T.-W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., Oikarinen, T., Srivastava, D., Weng, W.-H., and Weng, T.-W

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.349513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.109993Z digest=sha256:d4aba31d998ce5db937b501bdf2ccb8dba201ae54ccaeaf4f8c7a1842f680834

Observation e5324f97-f8aa-42dc-ac97-eaa57593a494 · outbound

This paper cites Vlg-cbm: Training concept bottleneck models with vision-language guidance.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Vlg-cbm: Training concept bottleneck models with vision-language guidance

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:33.071418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.155437Z digest=sha256:4cd7af54f2c39f0b7f545dd1a68b04cdbebadb2311479d7263605eb45a7ef90b

Observation 6493ee0b-bdb3-4aa9-a640-f880e170eb7d · outbound

This paper cites an unresolved cited work.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:18:32.827490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.257892Z digest=sha256:122c26635618a0bfbdf9da7dd6f4894744fb88c5bcfa8a0442438701838c54c1

Observation 237ba960-b0fd-4e58-9e09-69d91c45405d · outbound

This paper cites L., McDougall, C., MacDiarmid, M., Freeman, C.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks L., McDougall, C., MacDiarmid, M., Freeman, C

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.421782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.421782Z digest=sha256:27ec36c72a6b82873eefbf2dae59c8f40c083ad3d227168fa4853d3d9acaf67f

Observation 20b18915-b1d5-4346-a85f-01f0f8b1c9fe · outbound

This paper cites The caltech-ucsd birds-200-2011 dataset.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks The caltech-ucsd birds-200-2011 dataset

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:29.533625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:29.533625Z digest=sha256:5c62eedde68a2cc05e48d9be16a28ab32066b22e9676cc71658a4e35c18b6488

Observation f4be5411-3829-426d-8b45-a359c8a2c570 · outbound

This paper cites Post-hoc concept bottleneck models.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Post-hoc concept bottleneck models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:32.500409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.687107Z digest=sha256:be829f6b7ae9a5de16af8b32492823a317d14085b48bf541d6ee11253ed2cf61

Observation 97f6663c-32bf-4ab5-9aff-3689da9c74e3 · outbound

This paper cites Sigmoid loss for language image pre-training.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Sigmoid loss for language image pre-training

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:32.246966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.810712Z digest=sha256:bf8e6c518425ab4b7ae141a892e4afb9a850347a607475ba33998b0b7276b3bc

Observation 66377515-a453-4c18-b9a7-7234a9b035be · outbound

This paper cites A., and Rubinstein, B.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks A., and Rubinstein, B

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.972108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:29.969311Z digest=sha256:217c7a9664e5c47a70f634f32bb31d4c995c866755468e9aaab98d267fc75ca2

Observation 1af7c1b8-9f38-43f6-a1e7-2005a117face · outbound

This paper cites Object detectors emerge in deep scene cnns.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Object detectors emerge in deep scene cnns

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.754931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.103614Z digest=sha256:23be6db552a06bed7def3c5f558dc8f18d96d93fc197c4e70340a7d1015b95c6

Observation 397a2133-f133-404c-848c-2fab1ee83488 · outbound

This paper cites Places: A 10 million image database for scene recognition.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks Places: A 10 million image database for scene recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.460894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.211798Z digest=sha256:30c83b071724545d58f4ed5e6583c03686321797c9e3e929d55e96284b56ec6a

Observation 14cf9674-d9d7-4d6f-bb13-556dc7ded694 · outbound

This paper cites S., Klein, T., and Brendel, W.

Evaluating Neuron Explanations: A Unified Framework with Sanity Checks S., Klein, T., and Brendel, W

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:31.203304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:18:30.371390Z digest=sha256:b99f68123f8692113e4559f5a9595335734f5f113477ba817d70449f4d54553b

Pith citing papers

Observation 737a2ae0-66cd-4cf6-bf96-d5d6b26088a5 · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 260

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:51:29.531017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T18:50:52.313363Z digest=sha256:bc14513c3e653a73a814a010ecc8a0f185baf2af829c2d4cba1d9141506e72bf

Observation 04967e2f-35bd-4bd0-a0d9-2d8d9d96237b · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 252

Resolution
unresolved
no resolver link, observed 2026-08-02T20:04:33.986438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:04:33.986438Z digest=sha256:20eff8ac5aa43ae3a0869d0a9bcbd736e857236b02934875db7828d67df6b179

Observation 84620f7e-13be-40f8-9b64-7c857542c718 · inbound

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders cites this paper.

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders Evaluating Neuron Explanations: A Unified Framework with Sanity Checks

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.916851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T22:22:50.474397Z digest=sha256:4264da60e99707f640ec9445778eaf67d7e9aaed55eed1471c10776386fd208c