Pith. sign in

Paper Citation Record · LEDGER

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 4 inbound Pith citation observations for arXiv:2505.19532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19532 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:13.861239Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:44:01.319756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T02:11:15.649761Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact6
  • verified fuzzy37
  • unresolved12
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 24f0970d-618e-4102-8706-4439c8b32b7c · outbound

This paper cites Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:25.052191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:08.745552Z digest=sha256:c0634066f4e045adc0dc375a622741471d7460dc8897e23ea3f83ebdbb233978

Observation ee685072-8c34-402e-b082-b7a4187c5205 · outbound

This paper cites Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.767945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:08.815703Z digest=sha256:00cf3158d73eeacbd564bcb0a4e451759c062df66a227be4d2f4e9e7c9207b53

Observation 0c1b0ddf-32dc-4637-bcba-72c79c6bb8a8 · outbound

This paper cites The Theory of Dynamic Programming.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The Theory of Dynamic Programming

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.402866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:08.913367Z digest=sha256:7f9d1af4617a05c2bc55b5cd7d98d9b8a1f274fb5642123b6086340ee79922d6

Observation 4d8d3aa4-062a-4087-a14e-4432c50246b2 · outbound

This paper cites Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.118281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:08.991357Z digest=sha256:3a7a3fbc9c95f87740d181bca27ccc89bd03286fa0ea2de4a992918f878592d8

Observation 60c86d88-0001-4dbb-9dfe-207899ff98b0 · outbound

This paper cites Evasion Attacks Against Machine Learning at Test Time.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Evasion Attacks Against Machine Learning at Test Time

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.863576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.056165Z digest=sha256:26537b0b64b53cecc72146df0596a010d112dbbaee8b54f6ddebc87adfd3d10a

Observation 1f7c4ae6-b121-41a8-a1f8-4e3e5a839b5a · outbound

This paper cites Poisoning Attacks against Support Vector Machines.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Attacks against Support Vector Machines

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.120007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.120007Z digest=sha256:68915abde3bc0921c3915fb73a75a731a5dfcb007c3da073c1da7fbead7f308e

Observation b98bfa1e-704b-4477-b2c3-a5bea58a8677 · outbound

This paper cites A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.564667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.219209Z digest=sha256:2f4a26d11135f44ed93cae7196a836dd0580b7242dcbf0731e6aa24ce8f1b44e

Observation cea9ad1c-3035-43c9-a257-4727e7a5eaf2 · outbound

This paper cites Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.296878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.314319Z digest=sha256:eafcb037308366b9c35ba4e4a1a0f953c463d0c87029a4aff36dd18c79d59fee

Observation d5950921-2832-4264-a2d5-f3a85efe7f24 · outbound

This paper cites ABIDES: Towards High-Fidelity Market Simulation for AI Research.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ABIDES: Towards High-Fidelity Market Simulation for AI Research

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.371334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.371334Z digest=sha256:403f6201c751a0eae71c972b46432a0ff191b20fb65544e575276b151f956665

Observation 7316298d-0a5b-4fea-89ad-d3a01200ca4f · outbound

This paper cites Backdoor Attacks on Multiagent Collaborative Systems.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Backdoor Attacks on Multiagent Collaborative Systems

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.524825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.462751Z digest=sha256:4335751495ba996b5c5a6d2b3cfb1f488fb97f84dcff89f1619d71d773d28d58

Observation 228bd688-ce3d-4e83-bebd-9c51564df737 · outbound

This paper cites MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.979794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.542591Z digest=sha256:4b0c986b5c52a08568cefb4977359d5e436ee251421144ef61eaef68f53dbeb5

Observation 84e91ac8-425c-4771-b0da-0d076d7ed0aa · outbound

This paper cites Certified Adversarial Robustness via Random- ized Smoothing.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Certified Adversarial Robustness via Random- ized Smoothing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.687430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.602938Z digest=sha256:58bb7b069bcca24e16693863fd08c5b29bf7b5359b2b0e97542728e8f4309902

Observation 10ca9940-2439-4c6d-80e2-9620d89070f8 · outbound

This paper cites Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.416479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.655891Z digest=sha256:54b3ca6f15f0c399376695dfb5da77b1f819e87b4caf6120ea9788f6ee8e6a34

Observation 30f1801b-c281-4a89-8e84-da716c6ad6f4 · outbound

This paper cites Cullen, Shijie Liu, Paul Montague, Sarah M.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cullen, Shijie Liu, Paul Montague, Sarah M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.235882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.746018Z digest=sha256:e6e83762fc21e870fb806c1e09625d3885c77a2b63ec956db2e2d32003e25939

Observation c57c3dea-9816-4319-acd6-78e9897e0bb7 · outbound

This paper cites It’s Simplex! Disaggregating Measures to Improve Certified Robustness.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning It’s Simplex! Disaggregating Measures to Improve Certified Robustness

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.028532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.797497Z digest=sha256:ec2428fdd150262f2954e2d7d061b2b079e5b98bb43c96045cc2cc5bf42a3611

Observation 2fbd68f7-da95-443a-94e7-81af0edc5fe0 · outbound

This paper cites An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.801447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.873759Z digest=sha256:b9b9a8c37247f18880cb2567355e3380d2b2bce2b481973d9a3561a580f95952

Observation 90a40535-83b9-480c-9d43-9e9baa9bbf9a · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning CARLA: An Open Urban Driving Simulator

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.534425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:09.939953Z digest=sha256:ed61b75c46de381cfc44a14b105f43d701048784845cc09d90e0cafe014cad2c

Observation 3fcd4da5-b74e-4869-8304-82d192365c6f · outbound

This paper cites Guiding pretraining in reinforcement learning with large language models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Guiding pretraining in reinforcement learning with large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.315128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.019452Z digest=sha256:70a20646cd9045d84f8d44d2807525dc53706d309e940b00836ef246eed97f53

Observation 0e7fa66a-4ce9-48a8-8290-8f3d704957fa · outbound

This paper cites Language Guided Exploration for RL Agents in Text Environments.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Language Guided Exploration for RL Agents in Text Environments

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.282384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.180714Z digest=sha256:14c509a8e863176b948e888737dd0b93846b24889a69de05838b50736c69ccac

Observation 8c463425-7772-4b7f-a911-6cdc02f0fcfa · outbound

This paper cites Planting Undetectable Backdoors in Machine Learning Models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Planting Undetectable Backdoors in Machine Learning Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.237019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.237019Z digest=sha256:2638af2fcffbb8a867bedcb1f08f9b58659a95f5f24a5cb5949b6986537b69a9

Observation 71cfbef9-06ea-4ae8-abae-6f8b66386102 · outbound

This paper cites BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.927351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.324278Z digest=sha256:9748201a7741d214e78009516403bd4d4bb83fb0c3bfbdd3afed7d8050b8294a

Observation ebb433b1-9ce8-4b8b-9d65-76e9f3449a2f · outbound

This paper cites BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.384895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.384895Z digest=sha256:582187caddc77ed4f5d94ef5f641de50d0acc19ee70ff31b6665fd51479cdbf9

Observation b8c93030-60ed-4cdc-9598-1a16f925bc06 · outbound

This paper cites PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.965276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.446404Z digest=sha256:d6fef4c3ba88e64a22a5b5e1e2d29c05542f362bfae8cc4f3cc9eb797509f9b5

Observation bfb968c8-6543-42d1-be6d-6446a4d9efd6 · outbound

This paper cites Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.707748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.454275Z digest=sha256:496f21adc3b488f694c946d45d460f9bc219b057b5d73b6ad53e4851b74a72dc

Observation 07f1cc15-035a-44ff-b755-668f87d8a211 · outbound

This paper cites TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.511519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.582816Z digest=sha256:3807ee263e227668148306795b67cb9d6953cfcdd678573d612a2d74f551921c

Observation 1b805143-6fd5-4b8d-a1c3-54e9f60583d3 · outbound

This paper cites Policy Smoothing for Provably Robust Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Smoothing for Provably Robust Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.671575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.671575Z digest=sha256:b4d4b00ab268fb37f9da1e5747ea657b31aba06f04982bec0fc4c2b756beb8cd

Observation 97c9257d-11e0-4009-b752-278fae881646 · outbound

This paper cites Architectural Neural Backdoors from First Principles.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Architectural Neural Backdoors from First Principles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.746820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.746820Z digest=sha256:87239470d9615f0b5a0f1fd8a8b58aa11c444125b71993ad78b9bc09ee3635cb

Observation cf09a675-8643-4606-bef5-eb1fe31efd53 · outbound

This paper cites Markov Games as a Framework for Multi-Agent Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Markov Games as a Framework for Multi-Agent Reinforcement Learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.254614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.832696Z digest=sha256:f575d79d1c803a609ef5325c9ad9675dec535ac1e2d974e461d9febdbc0ae0da

Observation 0aa6899e-3fcd-4b29-a60d-34cdb5a316ca · outbound

This paper cites Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.059640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.941533Z digest=sha256:7b98f6630f22f7feb71f3fa8365c575e642b34c6807e48dc10325747a792f385

Observation 1cd0789d-5bd6-49e1-8c89-d91d9c566f1c · outbound

This paper cites Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.850917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.053113Z digest=sha256:48e7475ac9ffb63ddd9b3f7c31c885fd564becece6d5676f2abe3b4c61b870bc

Observation c32f9f8e-6dfe-433d-b824-e736751c2eb5 · outbound

This paper cites Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.667735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.147442Z digest=sha256:c45b2c7bc003ec47478ae29b89883d41945a267920f86d4219c4ca7bfeaead6f

Observation f0a1f8c2-923b-44fa-80ea-cd194c8970be · outbound

This paper cites Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.529582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.221298Z digest=sha256:6893a01d85b97709e82f7738905088317f628737e1dbb7761f5f28b83ff633a2

Observation 27edbecb-bf88-439f-a9f5-8a8150c26499 · outbound

This paper cites Neural Trojans.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Neural Trojans

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.305670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.298570Z digest=sha256:5827396761ffcbb1dae768027ff51899d97e27b8db57e65e2a20986b7f5a62f7

Observation 566ccd0f-48cf-4514-83bd-6f35319822f3 · outbound

This paper cites Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.078769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.401800Z digest=sha256:58237e45cba7ccd5e5b720edbd3791d20972328755e98ef6580f1d06d3f6ced9

Observation b6c1dd5b-9f31-483b-9d63-c915d53d8349 · outbound

This paper cites Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.616466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.458479Z digest=sha256:582cd50d211e7dc16235e8be04474166805ad93e55f12e8e775a0d4961d77445

Observation ac2a61ab-3ea7-4b80-bbf6-a8d1491e6787 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning, December.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning, December

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.899224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.554843Z digest=sha256:e97bd812e7cbba1966160978a3e24559b224342221322af358c6671560625fb0

Observation e85d0210-ac6e-4afa-80c7-406526022785 · outbound

This paper cites Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.724484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.785271Z digest=sha256:ca213a42be2c7c09cfe1913139c6055684266663777418bfafeec7bbb3d9457f

Observation 312ce2f5-0951-4555-b69f-4fe3c0baf33c · outbound

This paper cites Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.548415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:11.951247Z digest=sha256:5dfe319112d0d04653fbd74ca836ef8a2f0caf097af630b5de17c5568f423276

Observation 63a6a1f2-4391-48e2-a735-baa6b864a371 · outbound

This paper cites JPMorgan Develops Robot to Execute Trades, July 2017.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning JPMorgan Develops Robot to Execute Trades, July 2017

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.354926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:12.018846Z digest=sha256:d131d044046c7eb3dab326653ad5ab3a7715219401abe4e896f139afb101348e

Observation 5fd7afc1-43d6-47fa-b9b4-8887be6340c4 · outbound

This paper cites Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.180560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:12.096873Z digest=sha256:f79bf360f5a85ef06a2d27753ab2af6a0e3775eaca9d1142a6d827e824278e70

Observation 8f8095b2-b532-429f-a7a3-492acb00e3f5 · outbound

This paper cites Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.973244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:12.219825Z digest=sha256:7314c0917a51b52ac385defcaebc86361a8822ce9b6e8408390fdc6d69658ac0

Observation cebaed85-ccee-44ee-994d-8c12717c64a3 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.347692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.347692Z digest=sha256:e56ed7404b023367e68ca7cab845ce28e552c5cb32aad972152c26b7696b00e1

Observation ec7b3903-b31b-4ac9-ab90-b4119d1d39ec · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.417191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.417191Z digest=sha256:c807aa9fa87e6c9f4cc3af73d30511cf9369c2a79d1bfbc5b8c417b48839a103

Observation 8264c65a-711a-47d2-9eac-6f488a616ac8 · outbound

This paper cites Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.239670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:12.543771Z digest=sha256:8589c1503c95d1a1e0930553e18c5d25e8c9966d23ee33cb20aa71335bf6ceb4

Observation 5634e41f-5f95-4688-b95c-55bc2b167cff · outbound

This paper cites PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.716099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:12.676060Z digest=sha256:d4e468a97e04c08ba894ca8bd9e1a3f40ab79259de3bc0380b7cdb379630507b

Observation e88ff9e9-ce87-47a3-b961-3e01d43c85ba · outbound

This paper cites Terry, Ariel Kwiatkowski, John U.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Terry, Ariel Kwiatkowski, John U

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.837650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.837650Z digest=sha256:a010a735a2f347b9aa935988cc30d6ac5ba5180142c0db2ad8f4bf26c3735ffa

Observation 88a1cf82-57b3-4079-9032-e64371323bc7 · outbound

This paper cites BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.520817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.019993Z digest=sha256:05544c379b5525156be09321ca56a2358971b14b3bbb1dcfc75b320b8b7d3666

Observation 9b059332-876c-4306-b361-e1bfa42b28de · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.328253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.140080Z digest=sha256:6b0661143cf9fe66a70cd9bd9717edde9e2539316886fa80c720b7eea1d09845

Observation 62011085-d68e-4778-b113-93c53cb405d2 · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.170905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.258616Z digest=sha256:38ac9780a253cc042b6bb42a5fa0e7731ba96df88e904a4fe230e816b5beeb0b

Observation 13b09dc5-92ce-43d8-a4d0-5c1e9a631854 · outbound

This paper cites Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.949314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.418890Z digest=sha256:020431ff750a41c41add84d5d9e173956a9d706ee4eb39ba2e1dbcd547a90274

Observation 8a8452bd-4632-4c30-81ed-369fcda75eea · outbound

This paper cites Design of intentional backdoors in sequential models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Design of intentional backdoors in sequential models

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.073814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.532135Z digest=sha256:b47bdea3891cbb4389d3a079891226ef9a9a57d76c1c9242f5b536dffc275547

Observation cfb2976a-90cc-4cd2-a21a-7e79fd02024d · outbound

This paper cites A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.721920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.626270Z digest=sha256:11bd8f1980e24ea6a88f6eddf2cd48eac976f6d61e90211613fb7bc331163e8d

Observation 3187cb81-6d16-4ab9-a3c2-96b1066f48b0 · outbound

This paper cites ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.447981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.695320Z digest=sha256:4c56cf66fab63126594e92e7344748be2cf21fa5729c5b701838a0d48f2c7c08

Observation 1d7e9e0c-9f3f-4dbb-9b36-00235c7d7530 · outbound

This paper cites The trigger action and backdoor action remain consistent with those outlined in Section 4.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The trigger action and backdoor action remain consistent with those outlined in Section 4

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.111055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.771099Z digest=sha256:86233d667af104b6137fdf26e9d56cfbaf8242b59600a8bf4837ea8d6098e885

Observation 23cb00e4-16b6-4331-8592-0c6c4634803b · outbound

This paper cites move-up,.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning move-up,

Reference 57

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:16:15.835427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:13.861239Z digest=sha256:44e6167cacd5919c4b1ef24d8db17ebe8db0c605807fddad2247d1bce1ae7d29

Observation e701deac-8449-41e7-bbfe-6a7ef1d6d8ec · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:11.645830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:11.645830Z digest=sha256:6b91e83e7be682dff4705de83fb59dfa3bfde6291efcfa6e1c546fbdd76cd2ed

Observation 8a09f742-749b-4f0d-86e0-0bb66a3f1c2f · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 8677

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T14:16:21.105155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:16:10.094719Z digest=sha256:2b66e91823a324245cd22ec49174a1e1e3f065683ed0253089d2dfa2668c2d04

Pith citing papers

Observation 26814004-f1fd-40ac-8dfe-eaa6b2887a79 · inbound

Position: Certified Robustness Does Not (Yet) Imply Model Security cites this paper.

Position: Certified Robustness Does Not (Yet) Imply Model Security Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:44:01.319756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:44:01.319756Z digest=sha256:51add822c4966cf46cce5f12c2c9d82a4b45c86e2825a3d31b39fbc4167aaa5f

Observation 3d8ff5a5-ea96-4e0b-91cb-22d95099f4d2 · inbound

Agent Safety Alignment via Reinforcement Learning cites this paper.

Agent Safety Alignment via Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:29:36.704014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:29:36.704014Z digest=sha256:fd44add5143d929f00dae7b898b3606a63e138cede2dc1c2be6a70ed2e4a6096

Observation 71bedfe3-d47f-4a69-a628-94828df796af · inbound

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning cites this paper.

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T10:45:47.028500Z digest=sha256:dff2e1787516500ed3bf9b59bcea691080660282c330adb8879f2ac4d83d70c0

Observation 34f9838f-1e5f-413e-ba7a-4122c125163c · inbound

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning cites this paper.

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:10:02.988190Z digest=sha256:78c768e138e0b77649e0de5326621e92629e39d77918d052422552875ffc93ec