Pith. sign in

Paper Citation Record · LEDGER

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies

As of 10 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2507.14901.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14901 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:51:37.495684Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T01:13:11.483599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:38:55.871103Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved6
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c996a0ed-63c1-43bc-acea-b116e46d4fae · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Mastering the game of Go with deep neural networks and tree search

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.350707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.266533Z digest=sha256:fe618325b5af80892ef86e3c848965769518805354766fc8fd2441a124652881

Observation d3a92e27-364d-4403-8015-737a8d3a7b7d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Playing Atari with Deep Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.272028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.272028Z digest=sha256:0fc4ee214d4815c9e75a5bca28d0f559a99b904faf0af6533a00769c7d59350b

Observation 13efe497-c48a-4eb8-b7f2-aa9cd992a0a9 · outbound

This paper cites Reinforcement learning in robotics: A survey.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning in robotics: A survey

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.330692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.278064Z digest=sha256:cabeb8944e9c28ea3bfa21aa654ab9f549df251c0c8f0c7bd75576794cd8226b

Observation 9c5ddcdc-76b8-4b30-b285-98c9dcfb7fc1 · outbound

This paper cites Resource management with deep reinforcement learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Resource management with deep reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.302120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.283822Z digest=sha256:fbb5b6206c0fdb66a23d8c4f5e59701254296253e6a9182aee5db624b29945f9

Observation f6bc40c8-296c-41b7-831f-8d63d4b72b9c · outbound

This paper cites End to End Learning for Self-Driving Cars.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies End to End Learning for Self-Driving Cars

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.289123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.289123Z digest=sha256:ad9e061003d4994002f92fe6e543dfb383d1ac6dbd06358136eaab83cea25ee3

Observation 5f3ee64d-9fcd-4e88-934d-8000b51c2f17 · outbound

This paper cites Reinforcement learning based recommender systems: A survey.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning based recommender systems: A survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.277061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.294637Z digest=sha256:c0cd7e97d2c69527d4a8b4ab7e5b5ec4139bf2bb6c014ae4d438e7824ef6bb9b

Observation b8c337f6-bd16-4313-b1f4-08e289837d3c · outbound

This paper cites A review on reinforcement learning: Introduction and applications in industrial process control.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies A review on reinforcement learning: Introduction and applications in industrial process control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.259262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.301545Z digest=sha256:af1343f85e0f393c3e23f3db7b3fe6308bab621ed60caba268688cb6fb754643

Observation d5e13382-5942-4d95-a0a0-e8ae69e58dcd · outbound

This paper cites Targeted Reduction of Causal Models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Targeted Reduction of Causal Models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.237252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.306817Z digest=sha256:2f9fbf73ef10261e370049a728830010e748c34ed2312306b0f798b8e4138836

Observation dc779504-5c2d-4899-9bac-a3eba4581158 · outbound

This paper cites Causality.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causality

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.192843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.316979Z digest=sha256:f3b607130453501cea4739b79a72df6137d6cc602c4133813e6f6b8f9c6a08ce

Observation 3d7f21a3-a9c1-4a92-96a6-5a3cba8a0666 · outbound

This paper cites Elements of Causal Inference – Foundations and Learning Algorithms.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Elements of Causal Inference – Foundations and Learning Algorithms

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.171955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.321785Z digest=sha256:d263c9d8652615bf534a049d7b161d5a78d31dcc41dc6e3f23a8c3c6e74bb159

Observation 1842fbfd-2e17-4c27-baa7-6860a213115d · outbound

This paper cites Abstracting Causal Models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Abstracting Causal Models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.151157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.327967Z digest=sha256:377b33ddc10ee121daf3a43c94c948dae5d0be6f6002e281839b54e0e6774204

Observation 39bc8457-3cd2-4e20-bcb0-5b1bd2e8210a · outbound

This paper cites Approximate Causal Abstractions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Approximate Causal Abstractions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.128006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.333042Z digest=sha256:fbd0499f9703afda3719da1b60c802a30ebbcdaa11516de396d1bda0acbc4837

Observation 7694d6ea-c01d-4faa-aaf3-59cc967fae51 · outbound

This paper cites Causal Abstraction with Soft Interventions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction with Soft Interventions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.108359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.337908Z digest=sha256:eb097b59650f1a46173fa36e8a3134d65be0955c8d382f0b9032e42ef008f9ef

Observation e58c0533-2ce7-4186-8d35-9ec41c874277 · outbound

This paper cites Compositional abstraction error and a category of causal models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Compositional abstraction error and a category of causal models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.086804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.343403Z digest=sha256:d1bfe8c3ef389c8e0f321a4c4152234d0fc8cfcad6830feadc13036e8fe6ce6c

Observation 4ecdeea4-ca1e-4230-8825-1855bc0ca1db · outbound

This paper cites Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.348803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.348803Z digest=sha256:68920afb7762bc761bb81ec18f14270eea4a3ff9f4718205509c725164405ebf

Observation ef440e19-8954-40df-8a91-650817e81f34 · outbound

This paper cites Visual causal feature learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visual causal feature learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.064441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.354099Z digest=sha256:1e695c3420b6aaa5ecb4d6a5acc33f8d51a4bab23829c55806c22de516583e08

Observation 442984d5-a9f9-4ade-b3e0-cd46dd24725e · outbound

This paper cites Causal consistency of structural equation models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal consistency of structural equation models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.042834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.358642Z digest=sha256:06431a187f15b96f0b01464e57e8ba038df5167226a02a17e4c19723830522c0

Observation 78cf3f97-fb04-412d-ad77-fc8b2fef3ff5 · outbound

This paper cites Homomor- phism Autoencoder–Learning Group Structured Representations from Observed Transitions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Homomor- phism Autoencoder–Learning Group Structured Representations from Observed Transitions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.024586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.363208Z digest=sha256:782ec61fa77b1aca869de822fea1bdad9276cb52277cfcf38182c390ee0d31b5

Observation cfe8954d-212a-4e89-a73b-2bbd02d0cc79 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.368467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.368467Z digest=sha256:a84ae5967dcddc72102db4692cf4fb943232984114d2d1ff787ee5c6f79accb3

Observation 8191d131-a17a-4d61-8fad-91056cfc21e6 · outbound

This paper cites Learning to play table tennis from scratch using muscular robots.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning to play table tennis from scratch using muscular robots

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.000926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.373740Z digest=sha256:d4253be24e857434bc00756b3f2797228bf5bd414f84562f9e97d9405890308a

Observation 94589efd-bd16-4145-8b33-daccf0f594cf · outbound

This paper cites Safe & accurate at speed with tendons: A robot arm for exploring dynamic motion.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Safe & accurate at speed with tendons: A robot arm for exploring dynamic motion

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.983185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.378670Z digest=sha256:1d89c5da162755df0de2f17a420cf62487635102168bc977a0cc03759c28e119

Observation 9cee4fc6-7cd1-419b-ab02-b0d7cbe61986 · outbound

This paper cites Explainable reinforcement learning: A survey and comparative review.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning: A survey and comparative review

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.967569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.383546Z digest=sha256:631a13f1594536d6d71dff4921bff31865cc3c1cb75cda4f78a3ac8bf04950bd

Observation 073134fe-be44-444b-a248-1261899fbdd7 · outbound

This paper cites Visualizing and understanding Atari agents.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visualizing and understanding Atari agents

Reference 23

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.950030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.389396Z digest=sha256:b06968744b638c5bc00ff14f070bc3e531b40ef947c1910c29c230eff10c8b13

Observation 422b1cd2-0d11-49dc-ae39-b79406e8d39f · outbound

This paper cites Transparency and explanation in deep reinforcement learning neural networks.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Transparency and explanation in deep reinforcement learning neural networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.933941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.394262Z digest=sha256:9ca152cdf5a810f79711e65c15159d13bd80122860e89e1f8064174c2cb09601

Observation 0d0c8e15-c119-4931-85b4-07b1195709af · outbound

This paper cites Towards interpretable reinforcement learning using attention augmented agents.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Towards interpretable reinforcement learning using attention augmented agents

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.918084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.400049Z digest=sha256:4bac401ef5759615f7df55e45c36ab1c19a48ce9f6ee8e7c215dcd6fc8d6f9e5

Observation 2f817359-2cf9-4f5b-b3ed-8c2149af47a4 · outbound

This paper cites Explainable robotic systems: Under- standing goal-driven actions in a reinforcement learning scenario.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable robotic systems: Under- standing goal-driven actions in a reinforcement learning scenario

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.900836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.404901Z digest=sha256:a0c341556e4b5b19d2b2c240c4f2073da14ada164b2fa072de669a98dcc0b976

Observation 0731a134-e957-444a-bc9f-6ead1652a7bf · outbound

This paper cites Explaining reinforcement learning to mere mortals: An empirical study.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explaining reinforcement learning to mere mortals: An empirical study

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.883975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.410288Z digest=sha256:da139e9355db012bf171b707624171b749368e64169dcdd856c589f1d52bf1e5

Observation 72e5ddf2-be85-4494-929f-e2850e358765 · outbound

This paper cites Learning "what-if" explanations for sequential decision-making.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning "what-if" explanations for sequential decision-making

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.867244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.415048Z digest=sha256:411617d8063893f9e1ecdc023665dd95f5a139fb24721d0a5537c1c10b852a05

Observation a54b121b-5235-4e93-8eee-1d2dfbd287c8 · outbound

This paper cites Graying the black box: Understanding DQNs.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Graying the black box: Understanding DQNs

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.849813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.419818Z digest=sha256:11e68607ea7c9457a0906058cde5fc5ca9c796fb30b7d7318e2af811b86f7a8c

Observation ee84d02c-791a-4b0f-a01d-1ece076cd3c5 · outbound

This paper cites Generation of policy-level explanations for reinforcement learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Generation of policy-level explanations for reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.833618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.425178Z digest=sha256:093ae529691699dfad6f42d95cf7645163b048acb78df474eb38363f43662e1d

Observation 9267c37f-385a-4c34-a9c2-cec05a521d51 · outbound

This paper cites TLdR: Policy summarization for factored SSP problems using temporal abstractions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies TLdR: Policy summarization for factored SSP problems using temporal abstractions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.812888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.430381Z digest=sha256:b8435ffff4a28635ff660d56f6efc6009918047db5026c84da570191a63e8c40

Observation f290b596-ced0-4c63-b473-c29dd69035af · outbound

This paper cites Explainable reinforcement learning through a causal lens.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning through a causal lens

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.795008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.436342Z digest=sha256:2195925620363de595488fb549b054080d46cba145661f4548796324bd538c32

Observation d0371ab7-c37f-4c11-9fb1-b4cd23cf7617 · outbound

This paper cites Causal abstractions of neural networks.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal abstractions of neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.777692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.441148Z digest=sha256:f2f662f96a606a0ad5931d92188fd53d501f2c7a24ae21d004e4b3348328bff9

Observation 32c7e9ef-373c-46ad-a735-46ac73f5df31 · outbound

This paper cites Segment anything.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Segment anything

Reference 34

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.760111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.446132Z digest=sha256:b7721761485d8e0e620f7f3d07b42cce65bcf10ef9370a1010b7b8dd90008b42

Observation c6d279c7-9ac4-4e9c-96c7-7fdf2563759d · outbound

This paper cites Grad-CAM: Visual explanations from deep networks via gradient-based localization.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Grad-CAM: Visual explanations from deep networks via gradient-based localization

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.744252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.453979Z digest=sha256:a5519f2b27c240a1bd55a667bbaf3b0c50c495a3afc432d87dc14c7b81485681

Observation 506935a4-89b3-441b-8d66-ae28f18e397e · outbound

This paper cites Foundations of structural causal models with cycles and latent variables.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Foundations of structural causal models with cycles and latent variables

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.727134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.459902Z digest=sha256:4aeac99c5fa2eb7cc535444aa0882694beaed25f10a2974ab8d415ecb4d1b8c1

Observation 0f986416-cc18-4aaf-b1d5-2d6a3fa1cbeb · outbound

This paper cites Dependence, correlation and gaussianity in independent component analysis.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Dependence, correlation and gaussianity in independent component analysis

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.709254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.466360Z digest=sha256:c11dbaadcdc4c8d2cee096dd1e966c065d1336deb3edb50b9c454b234bccd426

Observation 1fcdcd2d-dcc3-4c41-9ec7-25fe42b0c189 · outbound

This paper cites Stable-Baselines 3: Reliable reinforcement learning implementations.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Stable-Baselines 3: Reliable reinforcement learning implementations

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.692501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.471536Z digest=sha256:ff6f59b3361edb93d23b0ae74053363c510579683d522f03ebce5891e02f3723

Observation 814cd2cb-92ff-4d17-a93c-72f28dd5e93e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Proximal Policy Optimization Algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.477154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.477154Z digest=sha256:7a8579c3a96789ed3280df2cac9c578639374779d90b41ab9082ce0b93a791b8

Observation bf472ef0-6772-40db-b29b-1bce7b2d9b99 · outbound

This paper cites The high-level model has (n+1) endogenous variables {Y, Z1,.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The high-level model has (n+1) endogenous variables {Y, Z1,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.673668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.482934Z digest=sha256:faf8ee855bbf49da8efbf8e85f82cd305fdf63c845faeae664b7624858cfd78b

Observation c7b768c2-5b5a-4aed-b5e3-e492691afd4f · outbound

This paper cites The exogenous variables {W0, W1,.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The exogenous variables {W0, W1,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.653791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.489483Z digest=sha256:14924dd55080c147d281b6b2ce2a1452cfbf33fcc94a654f86ec02f497303f08

Observation b4eb86eb-1987-477e-bc8b-154b844b785d · outbound

This paper cites (id − f1)−1.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies (id − f1)−1

Reference 43

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.634033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.495684Z digest=sha256:3ab5274c4300d9e7ba01f9e114d9543283da01fb8170d1fa23f296fdca34d99f

Observation 824cb148-9f02-49f3-90c1-c643ed51f9a8 · outbound

This paper cites an unresolved cited work.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:51:38.211855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:51:37.312001Z digest=sha256:c8f396e3053c57848091b82672c49bb31336e77457c9bbaa4548a37713e50164

Pith citing papers

Observation 750427c3-6156-4f82-80be-4ce14ec2e6ed · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:55.872495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:5303c1058c9ca98d68006217f936e42cb1f6d124b8ca7f8332509f1603db7ec0