Pith. sign in

Paper Citation Record · LEDGER

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

As of 18 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 3 inbound Pith citation observations for arXiv:2505.11189.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11189 v3

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:01:51.951512Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T00:06:52.820343Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy32
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation f2d08f5e-406a-4ab3-9a1f-05a4432812ed · outbound

This paper cites Sustainable development goals.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Sustainable development goals

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.756977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.748437Z digest=sha256:c8d6f03cddc91fbb30ae1b3fd52da48dfa159c1bc971a9b347631af4f2b0af1f

Observation acd52332-5f94-4dce-b2eb-b10280581c38 · outbound

This paper cites The role of artificial intelligence in achieving the sustainable development goals.Nature communications, 11(1):1–10, 2020.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP The role of artificial intelligence in achieving the sustainable development goals.Nature communications, 11(1):1–10, 2020

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.746451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.753006Z digest=sha256:3620f6a183df86bd8c40701140d20f3376d02d31ca14e1148d864f3c4172f5d8

Observation a62b6b06-f480-4853-a107-afb8097baabd · outbound

This paper cites Schäfer, Afra Amini, Heidi Lam, Massimiliano Ciaramita, Ben Gaiarin, Michelle Chen Huebscher, Christian Buck, Niels Mede, Markus Leippold, and Nadine Strauß.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Schäfer, Afra Amini, Heidi Lam, Massimiliano Ciaramita, Ben Gaiarin, Michelle Chen Huebscher, Christian Buck, Niels Mede, Markus Leippold, and Nadine Strauß

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.735874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.757077Z digest=sha256:12222a7894146cc0fecc6690c1189196254bb08b88f76f6e9d1d30120efd93ea

Observation cac9303b-8325-47f4-a759-f6b903803bcd · outbound

This paper cites Cognitive biases and artificial intelligence.NEJM AI, 1(12):AIcs2400639, 2024.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Cognitive biases and artificial intelligence.NEJM AI, 1(12):AIcs2400639, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.724824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.760968Z digest=sha256:8c44cb3bb72d4e24885bb737803fbdca7532e75982ea680d23f61d36db407fcf

Observation 0cad5b62-29bd-4f33-b738-7128a4974cd5 · outbound

This paper cites Misinformation spreading on facebook.Complex spreading phenomena in social systems: Influence and contagion in real-world social networks, pages 177–196, 2018.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Misinformation spreading on facebook.Complex spreading phenomena in social systems: Influence and contagion in real-world social networks, pages 177–196, 2018

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.714352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.765027Z digest=sha256:a1bd1a1f552c0de573e7bb6bf106b13ad74ef0013810610fa64d416750d57938

Observation 840b270b-7490-4f38-9d2b-24cf17c2e809 · outbound

This paper cites Information overload, multi-tasking, and the socially networked jury: Why prosecutors should approach the media gingerly.J.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Information overload, multi-tasking, and the socially networked jury: Why prosecutors should approach the media gingerly.J

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.704122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.768773Z digest=sha256:69a510e1060574065713aa855347aae573200717092498a820776aa7df8ced13

Observation ffcc856b-5c5a-4ae6-b05a-1f8ea4dd0f89 · outbound

This paper cites How do expectations shape perception? Trends in cognitive sciences, 22(9):764–779, 2018.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP How do expectations shape perception? Trends in cognitive sciences, 22(9):764–779, 2018

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.692978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.772644Z digest=sha256:321cc4a4aa947c3437df02b5e1c6c859c92e00a7f25665e666831a312edfb921

Observation 4d3a8a42-7237-4a38-8164-7590a95aa7b3 · outbound

This paper cites Default beliefs as a basis of social decision-making.Trends in Cognitive Sciences, 26(12):1026–1028, 2022.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Default beliefs as a basis of social decision-making.Trends in Cognitive Sciences, 26(12):1026–1028, 2022

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.682636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.776304Z digest=sha256:d60b08d5e9311b117b54957dc1d773d308bf84c8b948a548d11fe4b18d2ded81

Observation 4859a1eb-828f-4370-b691-10d6e06c6e51 · outbound

This paper cites Evaluating the moral beliefs encoded in llms.Advances in Neural Information Processing Systems, 36:51778–51809, 2023.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Evaluating the moral beliefs encoded in llms.Advances in Neural Information Processing Systems, 36:51778–51809, 2023

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.780029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.780029Z digest=sha256:e60fc40d92defcc5769af24ef0852c1cbb2edadf427d29c589474583e68e49c6

Observation b725c165-ba54-4b29-92b5-a80d2a714f56 · outbound

This paper cites Arithmetic Without Algorithms: Language Models Solve Math With a Bag of Heuristics.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Arithmetic Without Algorithms: Language Models Solve Math With a Bag of Heuristics

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.783972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.783972Z digest=sha256:373d4302c57f521e7370572a0964b1bbe510cce53674ed7b315bf7a30939f0f9

Observation ae766a60-7b3c-4571-8333-5c26cafe370a · outbound

This paper cites Responsible generative ai: A comprehensive study to explain llms.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Responsible generative ai: A comprehensive study to explain llms

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.665560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.788114Z digest=sha256:8fff970112eaba5375f5fd43ae3bc316dff3f018bf3ffddf4b11696ec720d0e8

Observation be10ec29-f389-435d-a5d7-69e84387a585 · outbound

This paper cites A Unified Approach to Interpreting Model Predictions.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP A Unified Approach to Interpreting Model Predictions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.791608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.791608Z digest=sha256:7cdea990839b61b0b559cb09cba20766c3a73708618a690a1489761263b19231

Observation 9bd4faf2-e405-48e9-ae5a-15684d81c0ef · outbound

This paper cites Impossibility theorems for feature attribution.Proceedings of the National Academy of Sciences, 121(2):e2304406120, 2024.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Impossibility theorems for feature attribution.Proceedings of the National Academy of Sciences, 121(2):e2304406120, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.795747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.795747Z digest=sha256:7e0ffb9c87576b44aee5ba37b8df89c0a4a8259afe014895398eddbe9b8e790c

Observation 2856ca25-a81e-437b-9d06-542b49be727d · outbound

This paper cites Predictive learning via rule ensembles.The Annals of Applied Statistics, pages 916–954, 2008.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Predictive learning via rule ensembles.The Annals of Applied Statistics, pages 916–954, 2008

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.647966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.799072Z digest=sha256:d29ed276588807a91f039551848f2aca717d41cdb2f40aa616381d54f782ee46

Observation 49b8b619-f702-433c-88ca-73f9ed751db9 · outbound

This paper cites ShapG: new feature importance method based on the Shapley value.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP ShapG: new feature importance method based on the Shapley value

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:01:52.242814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.802744Z digest=sha256:2e3b1b56551215318b668ea513fbcdb60e2c0720dbe6bd1d671c53415af7c97c

Observation 4c46a628-5721-48a9-b436-f39baf06f101 · outbound

This paper cites Rational shapley values.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Rational shapley values

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.637693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.806640Z digest=sha256:d84b6498078118751baaba4910f1672c342b31a6ca96123abc9a69a0d36b5ec1

Observation 8bc23dec-6543-4df1-ab43-bcf4bb5b4778 · outbound

This paper cites Explainable AI for Trees: From Local Explanations to Global Understanding.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Explainable AI for Trees: From Local Explanations to Global Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.810258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.810258Z digest=sha256:e48ee08b70a03dd24a1afa838ac5bc330849806fc31bfd9ab2d08636abf037c6

Observation 6469d35b-d8d3-440a-9ed1-35c88592f77b · outbound

This paper cites TokenSHAP: Interpreting Large Language Models with Monte Carlo Shapley Value Estimation.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP TokenSHAP: Interpreting Large Language Models with Monte Carlo Shapley Value Estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.814254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.814254Z digest=sha256:b16c6e7ad43452f6d44fe961a63687a136881a63e8d0649b13c6cf2d5963d381

Observation fb72a985-5ae2-4caf-bdc6-5473b80e50bf · outbound

This paper cites Concept-Level Explainability for Auditing & Steering LLM Responses.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Concept-Level Explainability for Auditing & Steering LLM Responses

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.818562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.818562Z digest=sha256:be753e7b2f5cef649146b2d54f5df2f4e34df9e62a0907e28b2eaba9225d9948

Observation 53afd7d0-5d41-440a-a7b1-efdccb91bfd9 · outbound

This paper cites Transcoders find interpretable llm feature circuits.Advances in Neural Information Processing Systems, 37:24375–24410, 2024.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Transcoders find interpretable llm feature circuits.Advances in Neural Information Processing Systems, 37:24375–24410, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.822373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.822373Z digest=sha256:9fb38498e6ec444f06af6d709926756c2b99fbf8b9985c072bf3fbd280b80c08

Observation 9241c8cb-7cf3-41f2-9b87-ccf855317fd7 · outbound

This paper cites Chatgpt: Literate or intelligent about un sustainable development goals?Plos one, 19(4):e0297521, 2024.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Chatgpt: Literate or intelligent about un sustainable development goals?Plos one, 19(4):e0297521, 2024

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.621843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.826033Z digest=sha256:0b835e0e50ad3667c3887f7e6d52b5d870f726eb4dc132de7b7bb669a9cc6091

Observation 67050bfb-2253-43a5-a617-96ce272e0872 · outbound

This paper cites Surveying Attitudinal Alignment Between Large Language Models Vs. Humans Towards 17 Sustainable Development Goals.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Surveying Attitudinal Alignment Between Large Language Models Vs. Humans Towards 17 Sustainable Development Goals

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.829457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.829457Z digest=sha256:605f5120936a473241d440e7580095e388d241bacdb2172966312315f6933242

Observation 7d33bee7-eef1-4346-a940-69603f2df734 · outbound

This paper cites Decoding Biases: Automated Methods and LLM Judges for Gender Bias Detection in Language Models.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Decoding Biases: Automated Methods and LLM Judges for Gender Bias Detection in Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.832936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.832936Z digest=sha256:100246b5006d7b2966b63da92e208e002b85d8948a8c36ec6771a6bb2630b1c0

Observation 505f286c-531a-419f-b90c-58386211ef30 · outbound

This paper cites Benchmarking Cognitive Biases in Large Language Models as Evaluators.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Benchmarking Cognitive Biases in Large Language Models as Evaluators

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.836687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.836687Z digest=sha256:3ac07518fc23b4bed4a0b9a3738b2cd700a8bf28f49c0b0b7720b7feaebb039f

Observation 6649fbb7-79cb-4c8a-816f-d30c6dac5bac · outbound

This paper cites Model-Agnostic Interpretability of Machine Learning.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Model-Agnostic Interpretability of Machine Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.840330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.840330Z digest=sha256:e80a00f24168e3f3cc77df1fa6c015410bb0ed6ca79711640b0a67bd78a90b72

Observation 57f0dc9d-352c-416f-8377-7732d12bc0b8 · outbound

This paper cites Addressing cognitive bias in medical language models.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Addressing cognitive bias in medical language models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.843848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.843848Z digest=sha256:27ce758ec9b7c3e99e565c2db819a1bff97385f819cb76bbd98e6f72ac2ba1a9

Observation 51b2bbe2-ca24-4514-81a9-64d1ab99f495 · outbound

This paper cites Is general-purpose ai reasoning sensitive to data-induced cognitive biases? dynamic benchmarking on typical software engineering dilemmas.arXiv preprint arXiv:2508.11278, 2025.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Is general-purpose ai reasoning sensitive to data-induced cognitive biases? dynamic benchmarking on typical software engineering dilemmas.arXiv preprint arXiv:2508.11278, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.847760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.847760Z digest=sha256:36053e74b71f2cba2a2a773f7e7c725f67db076a3fece8854ab11d558e159993

Observation 6c930cd7-8bd8-460b-a0f5-99896935fc7a · outbound

This paper cites A semantic embedding space based on large language models for modelling human beliefs.Nature Human Behaviour, pages 1–13, 2025.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP A semantic embedding space based on large language models for modelling human beliefs.Nature Human Behaviour, pages 1–13, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.611143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.851347Z digest=sha256:089a4ce78faaaf8d346691bf65456afe57119d33a388ae8715883665dd38be2a

Observation 3a18d70f-3b97-4032-89ce-f9633c77be91 · outbound

This paper cites Brownian distance covariance.The Annals of Applied Statistics, pages 1236–1265, 2009.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Brownian distance covariance.The Annals of Applied Statistics, pages 1236–1265, 2009

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.599360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.855029Z digest=sha256:399cda9fb4ebe1b8bd908b4b959edf85aa7308719afaead6b9a49a2882a411d4

Observation dd1b08ab-273e-42aa-bc9b-464a73eca82a · outbound

This paper cites Dealing with information overload: a comprehensive review.Frontiers in psychology, 14:1122200, 2023.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Dealing with information overload: a comprehensive review.Frontiers in psychology, 14:1122200, 2023

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.588415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.858879Z digest=sha256:096dd04155157e1f497b7796637519298ec2875da85a3942895b999016558fc8

Observation fba8df43-2d47-49b4-999e-7999bea9c7b1 · outbound

This paper cites Why information overload damages decisions? an explanation based on limited cognitive resources.Advances in Psychological Science, 27(10):1758, 2019.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Why information overload damages decisions? an explanation based on limited cognitive resources.Advances in Psychological Science, 27(10):1758, 2019

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.578452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.862088Z digest=sha256:69e40c8dd237efb914225d67cb47128b687b51e5dc0fa05c372d1ed9cef1abeb

Observation d36d7188-7226-427b-8497-96b86ad340d4 · outbound

This paper cites The power of moral words: Loaded language generates framing effects in the extreme dictator game.Judgment and Decision Making, 14(3):309–317, 2019.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP The power of moral words: Loaded language generates framing effects in the extreme dictator game.Judgment and Decision Making, 14(3):309–317, 2019

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.567931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.865586Z digest=sha256:5a5a77c42aaf93e4725c08c13a8d761e65fe34afb060a892043024271685fe01

Observation 014a9778-f42a-4091-968d-8a5fa664b39f · outbound

This paper cites The fog index after twenty years.Journal of Business Communication, 6(2): 3–13, 1969.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP The fog index after twenty years.Journal of Business Communication, 6(2): 3–13, 1969

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.557646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.868797Z digest=sha256:e54ad2b9229868eb6e72ac1d2d921ae611ddbc59dd1203a4b415bb3d38130d55

Observation 5573a6ad-20ac-4323-8c99-43b2f7255ab2 · outbound

This paper cites Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.872250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.872250Z digest=sha256:1437e7e409694dadafd1253a075eadc3e8c953c19245de83d3e06182006d2501

Observation b7189856-fc07-49a5-9ce0-2ba10a371fda · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.876051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.876051Z digest=sha256:614faf9576c4f173ce5a6657ce15d79e04290e6035d8dbd46a3eb03b765ee5c1

Observation 21f41034-7d34-451c-8bf5-411d07f3d5a4 · outbound

This paper cites Shap for actuaries: Explain any model.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Shap for actuaries: Explain any model

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.541238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.879307Z digest=sha256:8b3e4a2692f29da94d10baa4702865774f1338a22f171f9f5db10d34385b7893

Observation 60ee1f59-d908-4e79-ae70-d8bc92d49156 · outbound

This paper cites shap.explainers.partition — shap documentation.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP shap.explainers.partition — shap documentation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.530297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.882779Z digest=sha256:703d9a8b5894ee6ddf2205324f89d364ca7b7001fb2971909986c364e5715207

Observation c0b60cc8-8205-4462-8ce7-36c045ad33f8 · outbound

This paper cites Gradient boosting machines, a tutorial.Frontiers in neuro- robotics, 7:21, 2013.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Gradient boosting machines, a tutorial.Frontiers in neuro- robotics, 7:21, 2013

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.520156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.886332Z digest=sha256:31567aa66321a51d9644a2596eb1210707cafe17ca8313deca88b7900aa08523

Observation 7f867c7f-bf2a-43d2-b3e3-c99ed96c5f4b · outbound

This paper cites Lasso regression.Journal of British Surgery, 105(10): 1348–1348, 2018.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Lasso regression.Journal of British Surgery, 105(10): 1348–1348, 2018

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.509973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.890236Z digest=sha256:6bdf172f470b3449d9db3b2f1461bfa2f3685f515baa956b2ed55851291ce4ea

Observation 810403e8-7293-48b5-a9a0-d4c52ac06dc2 · outbound

This paper cites Xgboost: A scalable tree boosting system.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Xgboost: A scalable tree boosting system

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.893531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.893531Z digest=sha256:f2f2777f410046bde8dae42372e373b3a213bc2eec14c7e45bc8ca9114d85e62

Observation f214d0e1-60b2-41f9-b5c8-b484dc365998 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.896774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.896774Z digest=sha256:5fc12627b37426a4fdda409b0e1d6e28865dcb78ee605374759d5f198f9fb144

Observation 9d9cf436-a0aa-4865-a884-acdbe4a0a70f · outbound

This paper cites Glocalx-from local to global explanations of black box ai models.Artificial Intelligence, 294:103457, 2021.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Glocalx-from local to global explanations of black box ai models.Artificial Intelligence, 294:103457, 2021

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.491676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.900143Z digest=sha256:fc4c8064e25ef20b31169969bcc6ecbcd12276cf1828ecb2aa4e0a8551fa0f1b

Observation d3623743-947d-45c9-b1db-f5073f46709f · outbound

This paper cites imodels: a python package for fitting interpretable models.Journal of open source software, 6(61):3192, 2021.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP imodels: a python package for fitting interpretable models.Journal of open source software, 6(61):3192, 2021

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.377173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.903373Z digest=sha256:3065863e5c21a6fe9698a44640fc4bd9e6f73e4a368466756d697b885e2a2ef9

Observation 4f88f50b-cd82-461a-a43b-6596ae33e7a2 · outbound

This paper cites Bayesian rule sets for interpretable classification.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Bayesian rule sets for interpretable classification

Reference 44

Resolution
verified exact
raw_fallback, observed 2026-08-15T21:01:52.072826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.906628Z digest=sha256:6e011605525e958348f96456a73680c1f0ad688e109300e6dad857fd1b714cbb

Observation 6e79db4c-d360-4632-9512-c09a7466b7c2 · outbound

This paper cites Fast interpretable greedy-tree sums (figs).

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Fast interpretable greedy-tree sums (figs)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.366610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.910063Z digest=sha256:9cec5f381b276b57e614a3c13561abdadeaefb711162890217419f0ab3dea2ab

Observation 61835fce-7881-4913-b0f1-741aa68e9050 · outbound

This paper cites Post-hoc explanation using a mimic rule for numerical data.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Post-hoc explanation using a mimic rule for numerical data

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.355409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.913247Z digest=sha256:eeb2deb52ee28fc6395c3e613e787b80450144649d54bab23bdef5526dcff39b

Observation 03f5a0a9-8438-4381-a980-093c63e61c34 · outbound

This paper cites Palm: Machine learning explanations for iterative debugging.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Palm: Machine learning explanations for iterative debugging

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.344762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.916772Z digest=sha256:c4c5a3a24ee63c038b1394a5e249e634e31ac6182913d7e3bc5e7e58ff4a0444

Observation c01d6249-3db3-4df7-9c20-f00dfa736c7d · outbound

This paper cites V oorhees.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP V oorhees

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.334651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.919929Z digest=sha256:b35aa53d9808b2b8b969f711853907448d3a044675ccd5670e31abc43325bef3

Observation 09b0c17c-a771-40b6-a2bc-13201b5aed04 · outbound

This paper cites Geni: A framework for the generation of explanations and insights of knowledge graph embedding predic- tions.Neurocomputing, 521:199–212, 2023.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Geni: A framework for the generation of explanations and insights of knowledge graph embedding predic- tions.Neurocomputing, 521:199–212, 2023

Reference 49

Resolution
verified exact
doi, observed 2026-08-15T21:01:51.999112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.923351Z digest=sha256:991cdc4011115b20935614d5da148a616b6bf943ea574a472e104f7656145d50

Observation 97b34abb-da29-4cd6-93a2-c35c3dbf18c3 · outbound

This paper cites Vera Liao, Yunfeng Zhang, Ronny Luss, Finale Doshi-Velez, and Amit Dhurandhar.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Vera Liao, Yunfeng Zhang, Ronny Luss, Finale Doshi-Velez, and Amit Dhurandhar

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.926754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.926754Z digest=sha256:047f096eb061659f94477d96d993cb852fb9c1d4c022293bfa5c76636c49e3ef

Observation dddd1c30-fda1-4072-9c6a-9c940e7e35af · outbound

This paper cites From anecdotal evidence to quantitative evaluation methods: A systematic review on evaluating explainable AI.ACM Comput.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP From anecdotal evidence to quantitative evaluation methods: A systematic review on evaluating explainable AI.ACM Comput

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.930229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.930229Z digest=sha256:10d5d085e4e12c1dde79cd86a34810735f4158811c65a5c9fc945ef158e05cbc

Observation 6ef35fd7-55f0-4ecd-baba-b2311ef7cd10 · outbound

This paper cites Notions of explainability and evaluation approaches for explainable artificial intelligence.Information Fusion, 76:89–106, 2021.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Notions of explainability and evaluation approaches for explainable artificial intelligence.Information Fusion, 76:89–106, 2021

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.323871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.933736Z digest=sha256:e5e8e5de61cb77e73f44ae312b80832b3a83cec0f7b795c0adef027ea2d280cf

Observation f25509dc-f30e-4d15-b214-aafc184cdf4f · outbound

This paper cites When do neural nets outperform boosted trees on tabular data?Advances in Neural Information Processing Systems, 36:76336–76369, 2023.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP When do neural nets outperform boosted trees on tabular data?Advances in Neural Information Processing Systems, 36:76336–76369, 2023

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:01:51.937316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:01:51.937316Z digest=sha256:d32cbf08e08948057ff6fc62bf336eca691846113bb7982d25c8f3dedd339eaa

Observation 704a30b4-2abb-47ef-b473-45ce910177aa · outbound

This paper cites textstat: Python library for readability statistics.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP textstat: Python library for readability statistics

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.306308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.940684Z digest=sha256:def7b863158715e186db6c38195c56aa5920c5e45c7e0617ad73f1c2a9ef2c82

Observation 24648db5-586d-4a05-b26f-c35182c01dbd · outbound

This paper cites Evaluate the conceptual density of the texts in the whole web about {topic}. Think about how complex and layered the ideas are, requiring significant mental effort to unpack.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Evaluate the conceptual density of the texts in the whole web about {topic}. Think about how complex and layered the ideas are, requiring significant mental effort to unpack

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.295730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.943984Z digest=sha256:3d333212112723034fe1befbcc4d414cd8234091a2f5616d098cb978245ccbd4

Observation 2e863372-6a5d-48fa-9050-f41aed66e89f · outbound

This paper cites an unresolved cited work.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:01:52.284090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.948006Z digest=sha256:c9d9ad1b4c2e79c5b2780c61cc053edb0001869f4a3ff47b71b5e2c802228074

Observation 0065c3cd-0353-45cf-a90d-515adca17284 · outbound

This paper cites Write␣one␣short␣sentence.

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP Write␣one␣short␣sentence

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:01:52.273031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:01:51.951512Z digest=sha256:48627c61b826890ede07a4ab29ef301599bf231ff8028a24f7765c172dcb2e52

Pith citing papers

Observation 233fe333-708c-47a0-8a9d-42f6a6f2788a · inbound

Assessing Model-Agnostic XAI Methods against EU AI Act Explainability Requirements cites this paper.

Assessing Model-Agnostic XAI Methods against EU AI Act Explainability Requirements Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-09T02:06:03.418835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T08:55:42.238398Z digest=sha256:3477f93ce2473f5f2a2b780807c87ea40e7f342686be60feb14e08d832476af4

Observation 028fcf2d-83cd-4b98-9901-eccc0fe81f8c · inbound

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation cites this paper.

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-09T02:06:03.418835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T18:55:24.214794Z digest=sha256:039d22c6239151cfaae1d2da704a323ecda0946e78ecc36b685731a8b3499599

Observation 9702af26-3437-441a-b840-677723c88003 · inbound

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation cites this paper.

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-07-01T00:15:09.520927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T00:06:52.820343Z digest=sha256:6e1e9ced4fad1e6f944bb66e974244e4361c4daa8d5cf9f8452342d5b58627e1