Pith. sign in

Paper Citation Record · LEDGER

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 4 inbound Pith citation observations for arXiv:2505.19532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19532 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:13.861239Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:44:01.319756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T02:11:15.649761Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact6
  • verified fuzzy37
  • unresolved12
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 24f0970d-618e-4102-8706-4439c8b32b7c · outbound

This paper cites Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:25.052191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:08.745552Z digest=sha256:a01ba9fbdcde0e7d456eb5a98df7de37305ab66becfebc6381d1758a45daa4b7

Observation ee685072-8c34-402e-b082-b7a4187c5205 · outbound

This paper cites Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.767945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:08.815703Z digest=sha256:9b674eeb80ce3a7c2f2c12df4b3e3013341c08fc2448125960e0ddc7e03fabf1

Observation 0c1b0ddf-32dc-4637-bcba-72c79c6bb8a8 · outbound

This paper cites The Theory of Dynamic Programming.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The Theory of Dynamic Programming

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.402866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:08.913367Z digest=sha256:c0236ab43da98b028cad2a4e7de94cd567f10e4e365bf9962d1c618d7e8f79ad

Observation 4d8d3aa4-062a-4087-a14e-4432c50246b2 · outbound

This paper cites Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.118281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:08.991357Z digest=sha256:881ead507bec5db9a343612e297b7581c6598002b60da46ed8f2573f3b57c868

Observation 60c86d88-0001-4dbb-9dfe-207899ff98b0 · outbound

This paper cites Evasion Attacks Against Machine Learning at Test Time.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Evasion Attacks Against Machine Learning at Test Time

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.863576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.056165Z digest=sha256:6b3793d9273a5669857d3c0969604104557e52e75bd139a8330129257b8e00b8

Observation 1f7c4ae6-b121-41a8-a1f8-4e3e5a839b5a · outbound

This paper cites Poisoning Attacks against Support Vector Machines.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Attacks against Support Vector Machines

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.120007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.120007Z digest=sha256:81819f6bbf382c50cf1072a9cd7fa0164315659533a15a3a48a0aa43f78ecbc2

Observation b98bfa1e-704b-4477-b2c3-a5bea58a8677 · outbound

This paper cites A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.564667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.219209Z digest=sha256:e800c07a339252e85be4973a67b2d23ffdb0f51c35c9dee9dd1d311e42f84846

Observation cea9ad1c-3035-43c9-a257-4727e7a5eaf2 · outbound

This paper cites Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.296878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.314319Z digest=sha256:52cc825e62f2979b669dbbf28ce0c1796313ffe6c1034ae7d70f8f061163273d

Observation d5950921-2832-4264-a2d5-f3a85efe7f24 · outbound

This paper cites ABIDES: Towards High-Fidelity Market Simulation for AI Research.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ABIDES: Towards High-Fidelity Market Simulation for AI Research

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.371334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.371334Z digest=sha256:401e29109c88a5cb7bbe8f15fcf93904ec1b51effb6a078f382c70fe9990899e

Observation 7316298d-0a5b-4fea-89ad-d3a01200ca4f · outbound

This paper cites Backdoor Attacks on Multiagent Collaborative Systems.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Backdoor Attacks on Multiagent Collaborative Systems

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.524825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.462751Z digest=sha256:c5d1e9b94ac54a5a691908ff5d39acc219a1340d7628bc5c55ed520ff0d28b34

Observation 228bd688-ce3d-4e83-bebd-9c51564df737 · outbound

This paper cites MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.979794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.542591Z digest=sha256:7a884f56f21a459237266efd02cb6ae97939544dedca88db9991ec579bbceae3

Observation 84e91ac8-425c-4771-b0da-0d076d7ed0aa · outbound

This paper cites Certified Adversarial Robustness via Random- ized Smoothing.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Certified Adversarial Robustness via Random- ized Smoothing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.687430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.602938Z digest=sha256:03872c4ec3dbaf5b6ccc2c443fcc071901c54f407503db6e0c768bf1588dfa59

Observation 10ca9940-2439-4c6d-80e2-9620d89070f8 · outbound

This paper cites Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.416479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.655891Z digest=sha256:5154b7cea27d3aaad0a0e3a5f50d55a678030aa65286a8387064a2a3d058ef72

Observation 30f1801b-c281-4a89-8e84-da716c6ad6f4 · outbound

This paper cites Cullen, Shijie Liu, Paul Montague, Sarah M.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cullen, Shijie Liu, Paul Montague, Sarah M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.235882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.746018Z digest=sha256:87fd5e18b73e049c965ccd75eb0b9979ef1553c0925d758ef16198b892c2ad3d

Observation c57c3dea-9816-4319-acd6-78e9897e0bb7 · outbound

This paper cites It’s Simplex! Disaggregating Measures to Improve Certified Robustness.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning It’s Simplex! Disaggregating Measures to Improve Certified Robustness

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.028532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.797497Z digest=sha256:22eb99b1cc76b9d978614cb60cd27e1f3dcfaf1bf33aa49891b5b0c0df32253d

Observation 2fbd68f7-da95-443a-94e7-81af0edc5fe0 · outbound

This paper cites An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.801447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.873759Z digest=sha256:6da90932277df373ff2786f899e036d63fa3d3be017adbe3c42fb7fe55b60511

Observation 90a40535-83b9-480c-9d43-9e9baa9bbf9a · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning CARLA: An Open Urban Driving Simulator

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.534425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:09.939953Z digest=sha256:bc667a77fa6b45fd6d9967114aeee5b771d7e7bb3402bc8da546f4f3486f819b

Observation 3fcd4da5-b74e-4869-8304-82d192365c6f · outbound

This paper cites Guiding pretraining in reinforcement learning with large language models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Guiding pretraining in reinforcement learning with large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.315128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.019452Z digest=sha256:3d8a275d248bd519e5bca0411dc1150ef5a6c1cac5c9b035a269fc8e9d9fc666

Observation 0e7fa66a-4ce9-48a8-8290-8f3d704957fa · outbound

This paper cites Language Guided Exploration for RL Agents in Text Environments.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Language Guided Exploration for RL Agents in Text Environments

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.282384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.180714Z digest=sha256:60606fb876173bbe2848bebbc4b4f5327229dec18d7ecde4aae2a941a9401c7e

Observation 8c463425-7772-4b7f-a911-6cdc02f0fcfa · outbound

This paper cites Planting Undetectable Backdoors in Machine Learning Models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Planting Undetectable Backdoors in Machine Learning Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.237019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.237019Z digest=sha256:cf49fca77b3484f0b4b16ac8545012dfa32183e3fdeed80ac6f1c1015ce18597

Observation 71cfbef9-06ea-4ae8-abae-6f8b66386102 · outbound

This paper cites BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.927351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.324278Z digest=sha256:5ba5697a7a22dfafde3d76b5d8202e640949c2a8f64bfcf8b841959f61ce08c0

Observation ebb433b1-9ce8-4b8b-9d65-76e9f3449a2f · outbound

This paper cites BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.384895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.384895Z digest=sha256:cfe1edd643d568d36886fd8b318eab71bf28f847c735f4723db345a31ec4a6fa

Observation b8c93030-60ed-4cdc-9598-1a16f925bc06 · outbound

This paper cites PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.965276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.446404Z digest=sha256:e2016a7b42f4deed9918db269a46b8a6d75afd329ac231a9deee5bebe7d262ba

Observation bfb968c8-6543-42d1-be6d-6446a4d9efd6 · outbound

This paper cites Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.707748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.454275Z digest=sha256:28c0e6a46734d62012c5c3d1091f6333a63d01437ab9ebf4fd1f6585f6b5c57e

Observation 07f1cc15-035a-44ff-b755-668f87d8a211 · outbound

This paper cites TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.511519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.582816Z digest=sha256:bf2d9fee43211eeb005877519f70172ccd1ce9f9f59ac69d85318f913cc930dd

Observation 1b805143-6fd5-4b8d-a1c3-54e9f60583d3 · outbound

This paper cites Policy Smoothing for Provably Robust Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Smoothing for Provably Robust Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.671575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.671575Z digest=sha256:7d3009f90884a903dbc68ff3202ea261f466a6cf8ad3c0dffa07619fa30ee3b8

Observation 97c9257d-11e0-4009-b752-278fae881646 · outbound

This paper cites Architectural Neural Backdoors from First Principles.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Architectural Neural Backdoors from First Principles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.746820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.746820Z digest=sha256:20d9a3eeca9a1af3c98f2f287daaa468be46ef57e504ad36d16791b8116a4797

Observation cf09a675-8643-4606-bef5-eb1fe31efd53 · outbound

This paper cites Markov Games as a Framework for Multi-Agent Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Markov Games as a Framework for Multi-Agent Reinforcement Learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.254614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.832696Z digest=sha256:2d18e22bdd8bef97b656ed201b7427eac56d4ab4cd5d9be8aa2f55d637af8937

Observation 0aa6899e-3fcd-4b29-a60d-34cdb5a316ca · outbound

This paper cites Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.059640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.941533Z digest=sha256:f0dbfef9be7a35c83e6d7807a04301707260fbf4123122085281c9d1e7867fcd

Observation 1cd0789d-5bd6-49e1-8c89-d91d9c566f1c · outbound

This paper cites Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.850917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.053113Z digest=sha256:c71e9c117abba0e2bed27af16d227e8b8379b35e146479b35620dc90d9730c55

Observation c32f9f8e-6dfe-433d-b824-e736751c2eb5 · outbound

This paper cites Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.667735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.147442Z digest=sha256:10f6b1e8c1585f2abc48189a93316251c21de7e43339af3df406cd1f8fa847d6

Observation f0a1f8c2-923b-44fa-80ea-cd194c8970be · outbound

This paper cites Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.529582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.221298Z digest=sha256:148e5410d8ca1992480eff99bef5a8feb152804cf2d4b9b989f0462fc5f3a9c2

Observation 27edbecb-bf88-439f-a9f5-8a8150c26499 · outbound

This paper cites Neural Trojans.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Neural Trojans

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.305670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.298570Z digest=sha256:ee8644080a96086e88c7f6def1cccddd5918059391cd5ced82bc4b1992d027b8

Observation 566ccd0f-48cf-4514-83bd-6f35319822f3 · outbound

This paper cites Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.078769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.401800Z digest=sha256:6cd505508fa5cf8b9349af74701237ad9737eee15dc149d3f4d924c2adbbd306

Observation b6c1dd5b-9f31-483b-9d63-c915d53d8349 · outbound

This paper cites Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.616466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.458479Z digest=sha256:fa965e90e64c463eb85d93ae2765d1655da621e444286ca0e5b70962af93b6f1

Observation ac2a61ab-3ea7-4b80-bbf6-a8d1491e6787 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning, December.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning, December

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.899224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.554843Z digest=sha256:39b43ae64f8fa5ed98476d0e165ea9bc6d82b612a42d8bddc617f4b0edc8ba83

Observation e85d0210-ac6e-4afa-80c7-406526022785 · outbound

This paper cites Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.724484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.785271Z digest=sha256:cf0c9fceba3adbfbb55fa75ce4e67313431fb7a8a1319dee59e5391f4992f9cb

Observation 312ce2f5-0951-4555-b69f-4fe3c0baf33c · outbound

This paper cites Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.548415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:11.951247Z digest=sha256:926949bf5a6a5581d814c6221889c7f09299eb11754c3c30a9ae3a672fcd0e9d

Observation 63a6a1f2-4391-48e2-a735-baa6b864a371 · outbound

This paper cites JPMorgan Develops Robot to Execute Trades, July 2017.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning JPMorgan Develops Robot to Execute Trades, July 2017

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.354926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:12.018846Z digest=sha256:2d6cdda6194f2c9ed9c6ec6c5b286606aecd774d0389d373b72507a7b70b54fb

Observation 5fd7afc1-43d6-47fa-b9b4-8887be6340c4 · outbound

This paper cites Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.180560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:12.096873Z digest=sha256:903dddaf789c959760f65c5546d43acfdd6a0b332c65519d185f1a79ae749775

Observation 8f8095b2-b532-429f-a7a3-492acb00e3f5 · outbound

This paper cites Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.973244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:12.219825Z digest=sha256:781c53486f9fb47035968c4cea88718e4b15b41b6d506d35b830193d521ea3fd

Observation cebaed85-ccee-44ee-994d-8c12717c64a3 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.347692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.347692Z digest=sha256:addfa2d5dc79f634bfd236db4ec9739f312a815fe862072489c38b8c21f723ba

Observation ec7b3903-b31b-4ac9-ab90-b4119d1d39ec · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.417191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.417191Z digest=sha256:31f23109953fb0774d8400c3adb3aafe2e5592f7ccbab4e8347ae6076b6d00f4

Observation 8264c65a-711a-47d2-9eac-6f488a616ac8 · outbound

This paper cites Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.239670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:12.543771Z digest=sha256:de0cb921dbbda5350ff1500c08c4e206755c688d7af3594d5bab0d9bead30b29

Observation 5634e41f-5f95-4688-b95c-55bc2b167cff · outbound

This paper cites PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.716099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:12.676060Z digest=sha256:fa3284ba98d05cd182e1ba9466d0d6b4a4f65d214c52985c9566b9d9ffaa57af

Observation e88ff9e9-ce87-47a3-b961-3e01d43c85ba · outbound

This paper cites Terry, Ariel Kwiatkowski, John U.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Terry, Ariel Kwiatkowski, John U

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.837650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.837650Z digest=sha256:05ac64b925fb9093ef9371525c783b08e2d99fb6942f74ec2683594e8a72a76b

Observation 88a1cf82-57b3-4079-9032-e64371323bc7 · outbound

This paper cites BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.520817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.019993Z digest=sha256:377718dc80168919f6b7bc29c4dc13d56481cdd93efdcbcdc9a3204be383d8ae

Observation 9b059332-876c-4306-b361-e1bfa42b28de · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.328253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.140080Z digest=sha256:06621b2d6100212ed6e7ee91840343acece9b01a928f20d8dc7c7468a298b152

Observation 62011085-d68e-4778-b113-93c53cb405d2 · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.170905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.258616Z digest=sha256:a97a70ac504dc50d0b08718292136d2824e49c981669da5c33cb38aa5c33cb7f

Observation 13b09dc5-92ce-43d8-a4d0-5c1e9a631854 · outbound

This paper cites Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.949314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.418890Z digest=sha256:d828b79a09ee5a481443ceebf0fa69191bd5fa23464bd9c3dec53df1da07f91d

Observation 8a8452bd-4632-4c30-81ed-369fcda75eea · outbound

This paper cites Design of intentional backdoors in sequential models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Design of intentional backdoors in sequential models

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.073814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.532135Z digest=sha256:c3e34bf644e3ebfcdf7bda1881b96449da169768f8377f00aa0780690266057b

Observation cfb2976a-90cc-4cd2-a21a-7e79fd02024d · outbound

This paper cites A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.721920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.626270Z digest=sha256:8301604887f5482de287ff9f5e325c2821e393549950457e147fc3a187d074d1

Observation 3187cb81-6d16-4ab9-a3c2-96b1066f48b0 · outbound

This paper cites ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.447981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.695320Z digest=sha256:79f391e4784b39f3e0c5ca8d86d2ec08a3d3f4e2672965e3999e114db8c9a1c1

Observation 1d7e9e0c-9f3f-4dbb-9b36-00235c7d7530 · outbound

This paper cites The trigger action and backdoor action remain consistent with those outlined in Section 4.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The trigger action and backdoor action remain consistent with those outlined in Section 4

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.111055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.771099Z digest=sha256:13112a6227b8cfbbf71368b4522ef8053c8036a829a32b532f69d2de3248140c

Observation 23cb00e4-16b6-4331-8592-0c6c4634803b · outbound

This paper cites move-up,.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning move-up,

Reference 57

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:16:15.835427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:13.861239Z digest=sha256:2727550c931225a3ad8fdaf874e338206c9060801758c6548143bf22d459e182

Observation e701deac-8449-41e7-bbfe-6a7ef1d6d8ec · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:11.645830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:11.645830Z digest=sha256:bbc6bb1cd0d6adf8661018bacb9e8967baed6dbf8da37c1ac2b39466d5cc4fb0

Observation 8a09f742-749b-4f0d-86e0-0bb66a3f1c2f · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 8677

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T14:16:21.105155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:16:10.094719Z digest=sha256:ac16ce230feb57a2f53531bf127a7d1b88399cf652f7644b16b5d23f28eed9f8

Pith citing papers

Observation 26814004-f1fd-40ac-8dfe-eaa6b2887a79 · inbound

Position: Certified Robustness Does Not (Yet) Imply Model Security cites this paper.

Position: Certified Robustness Does Not (Yet) Imply Model Security Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:44:01.319756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:44:01.319756Z digest=sha256:43edf12e8c6a4a8b7cace1cad93b7484344c18a16523a5b852732980165b1240

Observation 3d8ff5a5-ea96-4e0b-91cb-22d95099f4d2 · inbound

Agent Safety Alignment via Reinforcement Learning cites this paper.

Agent Safety Alignment via Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:29:36.704014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:29:36.704014Z digest=sha256:1b911f9f74c156981820f3f97ba6a51e477bbf7a7aad0d4e55642c510d229772

Observation 71bedfe3-d47f-4a69-a628-94828df796af · inbound

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning cites this paper.

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T10:45:47.028500Z digest=sha256:1762025e68b3d8d3184b448c037fd7e4a18e659fde3579791171cb7d0a609d9a

Observation 34f9838f-1e5f-413e-ba7a-4122c125163c · inbound

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning cites this paper.

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T02:10:02.988190Z digest=sha256:3884924929987aa1e9d83e298dcfce9fcb07ecd3e50346b7b4aa93c2d1c2f012