Pith. sign in

Paper Citation Record · LEDGER

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents

As of 10 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2507.07848.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07848 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:36:43.205768Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:01:52.583181Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T11:16:11.205349Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact3
  • verified fuzzy7
  • unresolved22
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95f26998-1e1a-4156-a40e-5a588965f3ad · outbound

This paper cites Exploratory Not Explanatory: Counterfactual Analysis of Saliency Maps for Deep Reinforcement Learning.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Exploratory Not Explanatory: Counterfactual Analysis of Saliency Maps for Deep Reinforcement Learning

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:36:44.309466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:40.623694Z digest=sha256:9b97e8a3a4f71a5ce7098d75ac42a489bc9bedc7dcb7e8eed97edf0f3e73ae36

Observation f6f4d27e-7671-4d63-9e53-ab4d21e4f298 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:46.273991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:40.660920Z digest=sha256:66a5770ebd4b818a9508bf74e72ba9951dacbf6b31dc248efe2d71e8183c4e09

Observation 104725c9-b729-4cfc-8cd3-4f41033a0f4a · outbound

This paper cites Bastani, Y.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Bastani, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:46.087533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:40.718851Z digest=sha256:45694933a2a50a18138a5d99bcba38fb0f8651122baa93c8b39f1e69d181b306

Observation 27e70747-2d56-420a-baef-1cbe2d3e9fd6 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:40.793229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:40.793229Z digest=sha256:3cf7ea91a6b2c2247a84d8aa11d041ceaff182cf33c9447f97020c3d566e935a

Observation c0e22369-6a5d-41d8-b013-eff8da762ef3 · outbound

This paper cites XGBoost: A Scalable Tree Boosting System.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents XGBoost: A Scalable Tree Boosting System

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:40.967804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:40.967804Z digest=sha256:7ac567d8c19c2544bb8b438116f9dcd9752f23752cb0a422ed6b40609f62f20c

Observation 23ecd545-926a-4970-8040-937706c6b044 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:45.889984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:41.043766Z digest=sha256:efdd0a0b01c8266f75621888f64a0e12146e2f3d087daf50a9c811428e090b86

Observation 489c0ef0-f354-42e1-a11a-235f990ce760 · outbound

This paper cites Opportunities and Challenges in Explainable Artificial Intelligence (XAI): A Survey.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Opportunities and Challenges in Explainable Artificial Intelligence (XAI): A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.113929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.113929Z digest=sha256:d8caac9e0f6832fbf0b48aff4e0ae22acd5089cc09b40e5b0511bc16772a278f

Observation 16701200-0bf9-4c6f-9939-32dcf01d684c · outbound

This paper cites Ernst, P.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Ernst, P

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:45.687095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:41.216849Z digest=sha256:4ab9821a1945577286ad9204951f74b2c589a0c1fdc6dd1c494115d52919f587

Observation 656f1846-0d53-4fbf-813d-e5f41bd200be · outbound

This paper cites Limit Order Books.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Limit Order Books

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:36:44.062766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:41.329870Z digest=sha256:9435e136ea88a0110d89251b02c3e54466e2febcc2f3bb11da6b3904fe395017

Observation 6ba3920d-a826-461d-bcb6-cbd1bffb587c · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.411462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.411462Z digest=sha256:2b0d60ba7c722376e2d5732626d033ea30a1d3c651dcd662080bb11d2af0b42e

Observation e656d7bd-221f-4f68-a2ca-59e5cc2fcdf0 · outbound

This paper cites Generative Adversarial Imitation Learning.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Generative Adversarial Imitation Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.470280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.470280Z digest=sha256:73712b507fa35c66e73a8e46f90002bd6f4007589b54c2fdec23ea2c9b532777

Observation 77dba5d7-122a-4f18-b8ce-95e62973adcf · outbound

This paper cites Scaling Laws for Neural Language Models.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Scaling Laws for Neural Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.556842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.556842Z digest=sha256:1bd2205a212bb9b06fdf4902216982019c830a55da89540a1450dc21f634e1f3

Observation f7b17c13-5645-4e65-b605-f3a140014281 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Adam: A Method for Stochastic Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.595975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.595975Z digest=sha256:9abe054f70b2a9cfcf4fc06735b9a1045f852875fb9be36e94d3a62771d23101

Observation 47703804-323e-4b65-ac94-1eba4ec94ead · outbound

This paper cites Likmeta, A.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Likmeta, A

Reference 15

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:36:41.687392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.687392Z digest=sha256:75039a970cb37b8fa73eb5a93c756081755e943f6e583f8a67c901e1a2f7fea1

Observation 70293423-5513-4844-baa7-0a4d62ec54de · outbound

This paper cites A Unified Approach to Interpreting Model Predictions.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents A Unified Approach to Interpreting Model Predictions

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:41.924890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:41.924890Z digest=sha256:a331237f25262b3d5e816f0bfbdc69d509a5fa60932ab1dd85df4c301e23a414

Observation bacba575-7fc9-4467-99da-ea6302ffcf9b · outbound

This paper cites Madumal, T.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Madumal, T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:45.257069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.021385Z digest=sha256:7edc24dd780c6a894a74985a802218f20fa1156d6d10b6bc179401001092efa9

Observation c154b661-19aa-4164-b508-bee8ded4520d · outbound

This paper cites Milani, N.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Milani, N

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.087560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.087560Z digest=sha256:5bc4981bc5c9b359a545e9e51f804c88ba69f32470d69a61fc937410caaf6092

Observation 6c056548-b9df-4bf5-807d-18451ced58ad · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Playing Atari with Deep Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.181918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.181918Z digest=sha256:754a6e56bc4499b5fb0a782c1d79fc245a842d14b3709fba6cafadd215130910

Observation b4690ba5-7521-406e-aae4-ab9fa632357d · outbound

This paper cites Pirotta, M.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Pirotta, M

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:45.047639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.277595Z digest=sha256:465c85e413691e3dc79564953dd05a8649968bfbb3a290bea49ddf1550913ba1

Observation f085db74-bd2b-42a6-a196-1ac52e376e34 · outbound

This paper cites Raffin, A.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Raffin, A

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:44.908894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.358506Z digest=sha256:441973b3efc870f748e16603154fd942a2b2d6cc426c194d7cd54d944c30fb36

Observation cb27aa3c-4d3d-4bda-9c18-0f555ee3d69e · outbound

This paper cites Ross and D.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Ross and D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:44.783928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.413581Z digest=sha256:1db097d7b03e2ae0467301833cd95e857f7e575657c9128c7d9dd8fb6acb6a39

Observation 486bbe3a-042f-4a19-9e8f-de5fcefde6d3 · outbound

This paper cites A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.485427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.485427Z digest=sha256:46fd2953adbacfc29ea401526fd9460cce1952f0fab1267124966c7c43163c71

Observation c55ec9d5-e171-4c49-9c7f-1b1f518d4003 · outbound

This paper cites Policy Distillation.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Policy Distillation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.537434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.537434Z digest=sha256:00cbc453d3413932bdb32aa81d930f900779e80a7e3174e5824d00034b2694e3

Observation d18e3e25-e6a1-490b-8443-d6213004d448 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.598639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.598639Z digest=sha256:a68ecc608ccddf3972095477c05505a3993780db00ea41a345bceba84cdcce52

Observation 388688a3-fd09-4b5b-949f-687de9beb5a0 · outbound

This paper cites Silver, A.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Silver, A

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:44.645786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.677150Z digest=sha256:12cec75d5cb40e28473462dfee07bc111ebad718a6cb7a4fa7ccc5fb11c453f3

Observation e95d96f0-0bdb-4331-b3d0-b4e332e00e93 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.779161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.779161Z digest=sha256:2edc960637af50e4ee1a12803e06a6b6650234269685cc45bb36d195ef180297

Observation f1eb5a81-083f-4518-aecf-1bafe7294808 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:44.445279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:42.846940Z digest=sha256:d1deab2a5de4b7c950e9c4bc91bbb5e9bdc06af0bd1009843d13313b7d647f84

Observation 089f3cf2-aea5-43ee-9f2f-e3b3a05e0edb · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:42.916690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:42.916690Z digest=sha256:d0620d6dfeabc3424b0465ccc03efadcf36a007b2decfedee55a3dc304626ef7

Observation e26d1c88-95fd-4fcf-8e06-47c21d96bc1e · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:43.159262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:43.159262Z digest=sha256:c3cf977ad440ae07a7004e56c31bd7ed4202f1faaa0cf0a3fc3ce1a0a8088615

Observation b86bce2e-9fc8-48d8-aea1-7379b5bca79b · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 33

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:36:43.205768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:43.205768Z digest=sha256:c1cab15d64a1e278de83e238ec906edf9bc88c9e7db2803b73b9820775286019

Observation 87dc3488-3b3b-4306-bec0-a90966db9e2b · outbound

This paper cites Programmatically Interpretable Reinforcement Learning.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Programmatically Interpretable Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:43.074284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:43.074284Z digest=sha256:c7dfe44a012119e810408600d84afe81e40fcd68605fa9131a19b2e3dee261f0

Observation 69b28478-8241-419b-8742-e5a319229a82 · outbound

This paper cites an unresolved cited work.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Unresolved cited work

Reference 2005

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:45.437240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:41.267103Z digest=sha256:c0ba41ccb2ac28e00b12c3b06bc37d499511068ef693959551f4a737c1779d0d

Observation 84856176-ead1-4ab0-afb5-719450b71aec · outbound

This paper cites Toward Interpretable Deep Reinforcement Learning with Linear Model U-Trees.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Toward Interpretable Deep Reinforcement Learning with Linear Model U-Trees

Reference 2018

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:36:43.680643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:36:41.844751Z digest=sha256:9c399c952a4246ce699175c779647030ccc8bf986d139e742de28041590f30b5

Observation e18f7011-50c9-4c5c-a035-9cb088ecb4d3 · outbound

This paper cites Deep Reinforcement Learning for Active High Frequency Trading.

"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents Deep Reinforcement Learning for Active High Frequency Trading

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:40.897742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:36:40.897742Z digest=sha256:4fda0d3b44f066e7798f640ab5cdd338411e7f632730ec9b7e231170cc7c0809

Pith citing papers

Observation 7bde5198-5a12-4239-a8ec-c2bba0cf3aba · inbound

Selecting Feature Interactions for Generalized Additive Models by Distilling Foundation Models cites this paper.

Selecting Feature Interactions for Generalized Additive Models by Distilling Foundation Models "So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:16:11.210837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:01:52.583181Z digest=sha256:e4faf3cd83d2d05a3b47dc56e0ba3a5e64032aaa3ff5e23824dd382cc017216e