Pith. sign in

Paper Citation Record · LEDGER

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 4 inbound Pith citation observations for arXiv:2505.19532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19532 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:13.861239Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:44:01.319756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T02:11:15.649761Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact6
  • verified fuzzy37
  • unresolved12
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 24f0970d-618e-4102-8706-4439c8b32b7c · outbound

This paper cites Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Deep Reinforcement Learning Agents with In- Distribution Triggers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:25.052191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:08.745552Z digest=sha256:6b6067d5a853348f1e783ff050003e86cacd2616e5fc08620ec34cc4972ad0fc

Observation ee685072-8c34-402e-b082-b7a4187c5205 · outbound

This paper cites Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Can Machine Learning be Secure? InProceedings of the 2006 ACM Symposium on Information, Computer and Communications Security, pages 16–25, 2006

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.767945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:08.815703Z digest=sha256:110e8ea28296b027575a2c87637c1d0279055ec3a1f7939b30bcc07c1b9bee98

Observation 0c1b0ddf-32dc-4637-bcba-72c79c6bb8a8 · outbound

This paper cites The Theory of Dynamic Programming.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The Theory of Dynamic Programming

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.402866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:08.913367Z digest=sha256:7ecf51d05d8c0fa67152ca0a138f512744ab6dcd93027fe654b98b21e1602b1d

Observation 4d8d3aa4-062a-4087-a14e-4432c50246b2 · outbound

This paper cites Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provable Defense against Backdoor Policies in Reinforcement Learning.Advances in Neural Information Processing Systems, 35:14704–14714, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:24.118281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:08.991357Z digest=sha256:fe7c5c0ec0a9e851e28b3bb78d70482ca2423d51390685455ea568d37db23f29

Observation 60c86d88-0001-4dbb-9dfe-207899ff98b0 · outbound

This paper cites Evasion Attacks Against Machine Learning at Test Time.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Evasion Attacks Against Machine Learning at Test Time

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.863576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.056165Z digest=sha256:972ac64411befc2b20a491b99cbfcc5ca6cad93b4faf514d69cb05186daa1f12

Observation 1f7c4ae6-b121-41a8-a1f8-4e3e5a839b5a · outbound

This paper cites Poisoning Attacks against Support Vector Machines.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Poisoning Attacks against Support Vector Machines

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.120007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.120007Z digest=sha256:68915abde3bc0921c3915fb73a75a731a5dfcb007c3da073c1da7fbead7f308e

Observation b98bfa1e-704b-4477-b2c3-a5bea58a8677 · outbound

This paper cites A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A conceptual framework for externally-influenced agents: An assisted reinforcement learning review.Journal of Ambient Intelligence and Humanized Computing, 14(4):3621–3644, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.564667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.219209Z digest=sha256:7a19f67adef4e18e033fb178c6e678be2578796f27a9dfec5e76cab1e34ddfa1

Observation cea9ad1c-3035-43c9-a257-4727e7a5eaf2 · outbound

This paper cites Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Strong Data Augmentation Sanitizes Poisoning and Backdoor Attacks Without an Accuracy Tradeoff

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:23.296878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.314319Z digest=sha256:304e5a45d518b6cbab8d258ca230eaec7086354bc285d29ecceb674916dabfa7

Observation d5950921-2832-4264-a2d5-f3a85efe7f24 · outbound

This paper cites ABIDES: Towards High-Fidelity Market Simulation for AI Research.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ABIDES: Towards High-Fidelity Market Simulation for AI Research

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:09.371334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:09.371334Z digest=sha256:403f6201c751a0eae71c972b46432a0ff191b20fb65544e575276b151f956665

Observation 7316298d-0a5b-4fea-89ad-d3a01200ca4f · outbound

This paper cites Backdoor Attacks on Multiagent Collaborative Systems.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Backdoor Attacks on Multiagent Collaborative Systems

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.524825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.462751Z digest=sha256:7c7bb26fd0190e38e0bf5c528563360c29ad5592571a59127a048c21f0d86387

Observation 228bd688-ce3d-4e83-bebd-9c51564df737 · outbound

This paper cites MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning MARNet: Backdoor Attacks against Cooperative Multi-Agent Reinforcement Learning.IEEE Transactions on Dependable and Secure Computing, 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.979794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.542591Z digest=sha256:6d3e7bf8d4bec5bcb28b3725330d4ecf6909a171a3ce59d0dbb3a02f9f5c1dcf

Observation 84e91ac8-425c-4771-b0da-0d076d7ed0aa · outbound

This paper cites Certified Adversarial Robustness via Random- ized Smoothing.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Certified Adversarial Robustness via Random- ized Smoothing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.687430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.602938Z digest=sha256:f81f0413dae57a9e69d3d3d3ce3771ef3a169e9ca23559108b3d7557474aeaa4

Observation 10ca9940-2439-4c6d-80e2-9620d89070f8 · outbound

This paper cites Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Double bubble, toil and trouble: Enhancing certified robustness through transitivity.Advances in Neural Information Processing Systems, 35:19099–19112, 2022

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.416479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.655891Z digest=sha256:bbb64610dfc402582197dc73b02651fe949b941f2fdaa062ae76af9945415bc4

Observation 30f1801b-c281-4a89-8e84-da716c6ad6f4 · outbound

This paper cites Cullen, Shijie Liu, Paul Montague, Sarah M.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cullen, Shijie Liu, Paul Montague, Sarah M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.235882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.746018Z digest=sha256:1feaeb11aed79aa0aa0fdad338ebec5e870542755bdf8e5ac4cb693c03e8b63e

Observation c57c3dea-9816-4319-acd6-78e9897e0bb7 · outbound

This paper cites It’s Simplex! Disaggregating Measures to Improve Certified Robustness.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning It’s Simplex! Disaggregating Measures to Improve Certified Robustness

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:22.028532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.797497Z digest=sha256:5cd1af0ff4452f2fb7083e2baa136bf04a6236c6c4f45a7ee193619ed563293a

Observation 2fbd68f7-da95-443a-94e7-81af0edc5fe0 · outbound

This paper cites An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning An Automated FX Trading System using Adaptive Reinforcement Learning.Expert Systems with Applications, 30(3):543–552, 2006

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.801447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.873759Z digest=sha256:4eca61ba799aa83c938d344aa980486c436d98a01bde34cf76bd063a0a905b7f

Observation 90a40535-83b9-480c-9d43-9e9baa9bbf9a · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning CARLA: An Open Urban Driving Simulator

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.534425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:09.939953Z digest=sha256:f32179412a7cc974b03e45e8d3ee6dc97bf41c6857cfa66fc764eef5bdc30c50

Observation 3fcd4da5-b74e-4869-8304-82d192365c6f · outbound

This paper cites Guiding pretraining in reinforcement learning with large language models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Guiding pretraining in reinforcement learning with large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:21.315128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.019452Z digest=sha256:be000b7ae2380e35f7ed96c31d719a1e9433e3642ad849743dc8ad1d22d3dfd6

Observation 0e7fa66a-4ce9-48a8-8290-8f3d704957fa · outbound

This paper cites Language Guided Exploration for RL Agents in Text Environments.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Language Guided Exploration for RL Agents in Text Environments

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:15.282384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.180714Z digest=sha256:4b8bf7d0639b4cc4b1a6f9bf71e785f0258a5f12eef5b972ee6def62629f884b

Observation 8c463425-7772-4b7f-a911-6cdc02f0fcfa · outbound

This paper cites Planting Undetectable Backdoors in Machine Learning Models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Planting Undetectable Backdoors in Machine Learning Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.237019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.237019Z digest=sha256:19564a3b4b6b65a520acf2cf6cc60a04e03a14adeffcb14d11acd091cd942106

Observation 71cfbef9-06ea-4ae8-abae-6f8b66386102 · outbound

This paper cites BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BAFFLE: Hiding Backdoors in Offline Reinforcement Learning datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.927351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.324278Z digest=sha256:4da9bda9a9a8c79fc586e9813deaf5a228d75d7eb0a0e630a73eef77fbbd21f7

Observation ebb433b1-9ce8-4b8b-9d65-76e9f3449a2f · outbound

This paper cites BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.384895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.384895Z digest=sha256:582187caddc77ed4f5d94ef5f641de50d0acc19ee70ff31b6665fd51479cdbf9

Observation b8c93030-60ed-4cdc-9598-1a16f925bc06 · outbound

This paper cites PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.965276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.446404Z digest=sha256:76a22666751459e98cda47c9bfc6e1f79f8505445932cec317da695a1719d6dd

Observation bfb968c8-6543-42d1-be6d-6446a4d9efd6 · outbound

This paper cites Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.707748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.454275Z digest=sha256:e25c409c35c339ed823ff62c28e1239f7f0fc72f19a7a574021bd7e8c460250b

Observation 07f1cc15-035a-44ff-b755-668f87d8a211 · outbound

This paper cites TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning TrojDRL: Evaluation of Backdoor Attacks on Deep Reinforcement Learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.511519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.582816Z digest=sha256:44de6d81af4dff16a9eff462d286f4578e20394781883c7984c3df419ccfb02c

Observation 1b805143-6fd5-4b8d-a1c3-54e9f60583d3 · outbound

This paper cites Policy Smoothing for Provably Robust Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Smoothing for Provably Robust Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.671575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.671575Z digest=sha256:b4d4b00ab268fb37f9da1e5747ea657b31aba06f04982bec0fc4c2b756beb8cd

Observation 97c9257d-11e0-4009-b752-278fae881646 · outbound

This paper cites Architectural Neural Backdoors from First Principles.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Architectural Neural Backdoors from First Principles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:10.746820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:10.746820Z digest=sha256:87239470d9615f0b5a0f1fd8a8b58aa11c444125b71993ad78b9bc09ee3635cb

Observation cf09a675-8643-4606-bef5-eb1fe31efd53 · outbound

This paper cites Markov Games as a Framework for Multi-Agent Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Markov Games as a Framework for Multi-Agent Reinforcement Learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.254614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.832696Z digest=sha256:5c98459cac76b52397bb23016f23769ab463983fb640096fb610764d573da483

Observation 0aa6899e-3fcd-4b29-a60d-34cdb5a316ca · outbound

This paper cites Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Provably Efficient Black-Box Action Poisoning Attacks against Reinforcement Learning.Advances in Neural Information Processing Systems, 34:12400–12410, 2021

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:20.059640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.941533Z digest=sha256:a0a9a91565879fa2d9e6d6de7e84b57c57c13bf5a336555fb7e6ba70d7103133

Observation 1cd0789d-5bd6-49e1-8c89-d91d9c566f1c · outbound

This paper cites Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Fine-Pruning: Defending against Dackdooring Attacks on Deep Neural Networks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.850917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.053113Z digest=sha256:2e6407d8dd2d40049f73a93e12fa0672042377d84b792a3ec0ad0b795109a593

Observation c32f9f8e-6dfe-433d-b824-e736751c2eb5 · outbound

This paper cites Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.667735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.147442Z digest=sha256:48820222fe8f5a620d90c2aa578b39bfa1025f9b0486f8dd433d73d4888a64af

Observation f0a1f8c2-923b-44fa-80ea-cd194c8970be · outbound

This paper cites Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.529582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.221298Z digest=sha256:990f36107d7138db97bc1738becc3bc4d423e38cbaa52f0fdea13a5c35e60a61

Observation 27edbecb-bf88-439f-a9f5-8a8150c26499 · outbound

This paper cites Neural Trojans.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Neural Trojans

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.305670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.298570Z digest=sha256:6c88a40c86cabd4d86ca30124ae1da0ac17d3993ef7006ce03b068d9bbc4a752

Observation 566ccd0f-48cf-4514-83bd-6f35319822f3 · outbound

This paper cites Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Poisoning in Batch Reinforcement Learning and Control.Advances in Neural Information Processing Systems, 32, 2019

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:19.078769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.401800Z digest=sha256:dcbd165fd415098081a71f50b3229a11a012b0c093b99297b121e276c2960874

Observation b6c1dd5b-9f31-483b-9d63-c915d53d8349 · outbound

This paper cites Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.616466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.458479Z digest=sha256:47ccd810426ecf3463c4e108903b56b96ed3e5939a304bcf596d75770dc6a3ba

Observation ac2a61ab-3ea7-4b80-bbf6-a8d1491e6787 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning, December.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning, December

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.899224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.554843Z digest=sha256:35c8b5a0d755bb385a1e60ebf4d0b2758f012ed81607e39d061dd1b0f0200a17

Observation e85d0210-ac6e-4afa-80c7-406526022785 · outbound

This paper cites Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Implicit Poisoning Attacks in Two-Agent Reinforcement Learning: Adversarial Policies for Training-Time Attacks

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.724484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.785271Z digest=sha256:2a828a575050ec4caf3ea624549b0aa262f563b796a17297ae01fe1606c270ac

Observation 312ce2f5-0951-4555-b69f-4fe3c0baf33c · outbound

This paper cites Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Cal-QL: Calibrated offline RL pre-training for efficient Online Fine-Tuning.Advances in Neural Information Processing Systems, 36, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.548415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:11.951247Z digest=sha256:37d2bcf75c4f207f121340ddc477a5384292aaca2993c390787dbd918a1c8aa2

Observation 63a6a1f2-4391-48e2-a735-baa6b864a371 · outbound

This paper cites JPMorgan Develops Robot to Execute Trades, July 2017.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning JPMorgan Develops Robot to Execute Trades, July 2017

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.354926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:12.018846Z digest=sha256:04d89b776bbc37c19ed5c9320cc55dd453bfe38244fcc1de489f38cb8eca36f2

Observation 5fd7afc1-43d6-47fa-b9b4-8887be6340c4 · outbound

This paper cites Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement Learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:18.180560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:12.096873Z digest=sha256:59379d3118558589324728aa6211279fa897c26fd5bf3f05cabcdb4992da05f3

Observation 8f8095b2-b532-429f-a7a3-492acb00e3f5 · outbound

This paper cites Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Understanding the Limits of Poisoning Attacks in Episodic Reinforcement Learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.973244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:12.219825Z digest=sha256:be2870123241b305763d106e5dc25959ff96bc601b10396abc80456e0bcdd5de

Observation cebaed85-ccee-44ee-994d-8c12717c64a3 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.347692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.347692Z digest=sha256:e56ed7404b023367e68ca7cab845ce28e552c5cb32aad972152c26b7696b00e1

Observation ec7b3903-b31b-4ac9-ab90-b4119d1d39ec · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.417191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.417191Z digest=sha256:c807aa9fa87e6c9f4cc3af73d30511cf9369c2a79d1bfbc5b8c417b48839a103

Observation 8264c65a-711a-47d2-9eac-6f488a616ac8 · outbound

This paper cites Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Vulnerability-Aware Poisoning Mechanism for Online RL with Unknown Dynamics

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.239670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:12.543771Z digest=sha256:7ca2d8c14022b61b2d4c2bae87ace4569b68ab9615beeacb03de06aebe5ef746

Observation 5634e41f-5f95-4688-b95c-55bc2b167cff · outbound

This paper cites PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning.Advances in Neural Information Processing Systems, 34:15032–15043, 2021

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.716099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:12.676060Z digest=sha256:f75bb7b013b44c8af02d8c98f016d27b48d6fdfebdc65e46d2ab121914aab2de

Observation e88ff9e9-ce87-47a3-b961-3e01d43c85ba · outbound

This paper cites Terry, Ariel Kwiatkowski, John U.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Terry, Ariel Kwiatkowski, John U

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:12.837650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:12.837650Z digest=sha256:a010a735a2f347b9aa935988cc30d6ac5ba5180142c0db2ad8f4bf26c3735ffa

Observation 88a1cf82-57b3-4079-9032-e64371323bc7 · outbound

This paper cites BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:17.520817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.019993Z digest=sha256:0516595a6183b64b799242628a37aca6b32d312dcaca95dba56fd7c7d6a4d861

Observation 9b059332-876c-4306-b361-e1bfa42b28de · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.328253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.140080Z digest=sha256:96a34fe0f86a9bd158932c2b644cc746ebf3205b5ef00daa470304a12d1f3483

Observation 62011085-d68e-4778-b113-93c53cb405d2 · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:17.170905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.258616Z digest=sha256:274f93fcf2a46546d18795f12ec5e2625fc77f772ee350f0b89334b887c60a0e

Observation 13b09dc5-92ce-43d8-a4d0-5c1e9a631854 · outbound

This paper cites Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Transferable Environment Poisoning: Training-time Attack on Reinforcement Learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.949314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.418890Z digest=sha256:c434dcee5eecc9aea9245c93cd95f4261662a1bd8b9973e06b67332faec97a18

Observation 8a8452bd-4632-4c30-81ed-369fcda75eea · outbound

This paper cites Design of intentional backdoors in sequential models.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Design of intentional backdoors in sequential models

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:16:14.073814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.532135Z digest=sha256:5d1e056bb90f935d1eb6fe88e42d4bd31b04301cd88412beb76465bf9fec9db4

Observation cfb2976a-90cc-4cd2-a21a-7e79fd02024d · outbound

This paper cites A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning A Temporal-Pattern Backdoor Attack to Deep Reinforcement Learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.721920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.626270Z digest=sha256:c68b090ef57d59f9fb89a8bdd320dbc560d1f1e5192bbc175b3ea9b28d377b59

Observation 3187cb81-6d16-4ab9-a3c2-96b1066f48b0 · outbound

This paper cites ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning ∞X k=0 γkRopp(si+k+1, aopp i+k+1, aatt i+k+1) aopp i+k+1 ∼π opp, aatt i+k+1 ∼π # , (7) πopp = argminπ∈Π E

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.447981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.695320Z digest=sha256:cfc0558909242ece6163bb562b9344bc493de761c3bbdd719715ac231e8c92e8

Observation 1d7e9e0c-9f3f-4dbb-9b36-00235c7d7530 · outbound

This paper cites The trigger action and backdoor action remain consistent with those outlined in Section 4.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning The trigger action and backdoor action remain consistent with those outlined in Section 4

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:16.111055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.771099Z digest=sha256:3e7329e3def32257da0b0d95d8088b30123cfd759349f27152d7941bb340fb27

Observation 23cb00e4-16b6-4331-8592-0c6c4634803b · outbound

This paper cites move-up,.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning move-up,

Reference 57

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:16:15.835427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:13.861239Z digest=sha256:801708e2fb4ea7e7bca5d9eef9be6774f2f497170a385881245ce11c0bcfa176

Observation e701deac-8449-41e7-bbfe-6a7ef1d6d8ec · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:11.645830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:11.645830Z digest=sha256:6b91e83e7be682dff4705de83fb59dfa3bfde6291efcfa6e1c546fbdd76cd2ed

Observation 8a09f742-749b-4f0d-86e0-0bb66a3f1c2f · outbound

This paper cites an unresolved cited work.

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning Unresolved cited work

Reference 8677

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T14:16:21.105155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:16:10.094719Z digest=sha256:a1415481de0d572a48d0c9349d1ff2fcd6dc9daa227dff877db4f20eaef1e0d4

Pith citing papers

Observation 26814004-f1fd-40ac-8dfe-eaa6b2887a79 · inbound

Position: Certified Robustness Does Not (Yet) Imply Model Security cites this paper.

Position: Certified Robustness Does Not (Yet) Imply Model Security Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:44:01.319756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:44:01.319756Z digest=sha256:51add822c4966cf46cce5f12c2c9d82a4b45c86e2825a3d31b39fbc4167aaa5f

Observation 3d8ff5a5-ea96-4e0b-91cb-22d95099f4d2 · inbound

Agent Safety Alignment via Reinforcement Learning cites this paper.

Agent Safety Alignment via Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:29:36.704014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:29:36.704014Z digest=sha256:fd44add5143d929f00dae7b898b3606a63e138cede2dc1c2be6a70ed2e4a6096

Observation 71bedfe3-d47f-4a69-a628-94828df796af · inbound

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning cites this paper.

BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T10:45:47.028500Z digest=sha256:eda9bdbf2e5e3d42e04e0a2baaf0d89515256e07e4d15c7ca248893540815c03

Observation 34f9838f-1e5f-413e-ba7a-4122c125163c · inbound

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning cites this paper.

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-25T02:17:29.965443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:10:02.988190Z digest=sha256:48332132440e7e23b0b540f9fe21a59b70a7aa0e1fca2d9f4c08b4b3c211cada