Pith. sign in

Paper Citation Record · LEDGER

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic

As of 11 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.05445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05445 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:01.072056Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9d48139d-8836-41e3-ae6b-2f8ff8f667bd · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.279110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.605339Z digest=sha256:89a0cc5f011182994b75615abffffb0aa2362a61a6ffeb49c60ac9fca0ab8fc0

Observation ebfad1ab-7e77-4f31-908c-554a0d59c0ca · outbound

This paper cites and Pearl, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Pearl, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.227641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.665123Z digest=sha256:b57605e6d6755669c29f727f20c6673b02abd6eadcb47979243e1c68683d9fa9

Observation d45b2319-f964-4499-bb6c-e92a8eae836f · outbound

This paper cites OpenAI Gym.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:56.756738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:56.756738Z digest=sha256:7289e4918666652769da782e5eaaf237cadcdfc0f7de993c3edbbd0b7d7ebac2

Observation c9389d4d-dee5-477b-bc34-d3fbf3a36cdc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.172607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.823275Z digest=sha256:0d5823a560621b2a756eeb15f7d12ab94f1ba900505cfcd28c4051f0e7d7f738

Observation 269ca6ea-c4e5-459a-b44b-bdd72d06a58d · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.126760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.896354Z digest=sha256:a506b65e1b66f0b6c668cb24ad2fbe0883776a1e5dea710ab6a2fac8e50e6092

Observation 0246fd85-67da-4d61-be80-6ef066d20f43 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.066815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.970646Z digest=sha256:586a346c6310b6aab7eb148ab2a685dffd93a61c6e0cc3889e9a051dfcbe26f3

Observation 54e0fd50-c7ff-48b2-92bc-d12fdeea3cd7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.004765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.045401Z digest=sha256:8b9df9e4fc828552c532a7173a3911b18295de25fe842d1d7b7c45691ac36d77

Observation bcaa6341-0c28-4e2d-b9e9-124faad92207 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.944649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.100420Z digest=sha256:67be5417b8f5fc1ba5deb44ca91428495d23417de2033e011dd2d8b5eafd0fdd

Observation 0db23e0e-4f18-4ba0-9a9a-f05a486d13d7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.877110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.151008Z digest=sha256:35a49a997b3465a010a1c10501f1f8fdf1ab92b92dc30ae3325dad3d4ca8ef37

Observation 19c2d0ed-9252-4fb4-a7f4-1523034c1e75 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.794972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.207458Z digest=sha256:d9c6e26eb2d0b3ff2c10d9b8a7707f1c80c093d44afe5562dc56bc04e40e0da8

Observation 93cededb-216f-4353-96b7-e98981c165ab · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.744172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.263934Z digest=sha256:425767b1337d769fca5ab822c1239fbe23199b71d3873a03b45d843727993bea

Observation a3c08d5c-115f-451d-aa63-d8a87d64f85c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.686487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.336074Z digest=sha256:ff3f7be563712e772e0a6976163cb13ee5e0f61fb498c97c392f7786a51a1111

Observation e26ac464-1600-40d0-a5c7-8eac7cf39a4a · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.606814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.402866Z digest=sha256:60a5dcb8390cfa234d35f58b41eddd9846b50df02346fc020ae27914bc34d433

Observation 1bb3ecff-3de4-4781-a77d-9f9581352569 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.529392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.466460Z digest=sha256:17a05adf1d296011cbb19bfcffa0048eafe36ed6b7e40f5e28bdb80b77810cfe

Observation 1c900d36-b428-43cd-bbd6-7f6eb859d864 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.519671Z digest=sha256:c49c2f26da6b9f1e3f37c2c7bdb32b2f4caa536b4da17fc9f06715b71ef6ce69

Observation a5195288-9cf4-4d74-8860-25e085d20bd6 · outbound

This paper cites P., Hunt, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic P., Hunt, J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.392073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.574314Z digest=sha256:8574ccc8f021c892339acb486ec0f11700f853b07c92fabd8efa6d49f553a921

Observation 1f25f871-a3d6-4f61-b15e-c6f07e9ebbbd · outbound

This paper cites Continuous control with deep reinforcement learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Continuous control with deep reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:57.644181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:57.644181Z digest=sha256:a783f98068cc684452a952ca55f8a97b18c67bd1e2b343754804712dcf032e3d

Observation 94d44d67-9b6d-4e78-b265-4a61b205d29e · outbound

This paper cites and Krishnamurthy, A.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Krishnamurthy, A

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.330213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.711279Z digest=sha256:879f13f3f4184eb0067620282bbbf06120a1077f3f6df2bff75c5979a7eb0443

Observation 254e7504-a210-4e33-9413-903c850d747a · outbound

This paper cites Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:36:01.307595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.781252Z digest=sha256:a856a3b4caf1f9b8157258c8fbafe5c8b2c194c891bcec75c7ffb9f522b9c810

Observation 100ddbe6-78c6-4a76-95b9-2d421e05afc1 · outbound

This paper cites V., Sima, K., and Leong, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Sima, K., and Leong, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.277446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.852453Z digest=sha256:4bb6ee41b86e462767118d7293b789a861ec932c01ac0833e59d9a097b6d5d2d

Observation cd705998-f34d-428b-b67f-219cefb0ea64 · outbound

This paper cites V., Fu, D., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Fu, D., and Leong, T.-Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.223838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.918663Z digest=sha256:f8d2e8d400d8eb51e4b8e8c31b2f3fa6f53d7577e6f160005e86cb3419f19ccf

Observation 169f980f-2cc9-41a2-b6cc-4eb0f12180f6 · outbound

This paper cites V., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., and Leong, T.-Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.201758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.008523Z digest=sha256:a55c3a01074f722c640a331b225c12c653d76a797b9bc4f76524f2819009b537

Observation 08ee4ad2-093f-4500-914e-d302f1fdeed7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.173753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.116971Z digest=sha256:989e6838824cd2fc41b6fb6398ecd7ac7f3fb8d8ecff3acf661ae1194d968a3a

Observation cf3d0505-6724-4442-8426-b9eb59d5a31c · outbound

This paper cites and Sontag, D.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Sontag, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.138332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.205073Z digest=sha256:0b46785625c0aea7cae01e442eb34f19e67a35652dbf3d0fc76414613f94947a

Observation 14a94d8c-10b4-4183-af31-0c2e57d9b490 · outbound

This paper cites o lkopf, B., R \.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic o lkopf, B., R \

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.099722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.290756Z digest=sha256:0f6f8f8fcbe07cc8dbc7f0ee523fcd3289fa3971b91b929ea0e13c644dfd5bf4

Observation 581578b9-8bdf-4b87-9b1e-d95a612d4adc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.072534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.377118Z digest=sha256:de9711016af6715fd0f0a3769c6bba3089b8e47344bb4ec3c85b53dc73e9950f

Observation e968b5eb-b467-4199-881c-78b8a88897fa · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.045301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.488977Z digest=sha256:9a4d1033dd31df1e8eba552205f8e4b6143482f058461ad3606c1b73773e4e7b

Observation 217fafc5-d441-4cde-a5b1-50b4c0ab6376 · outbound

This paper cites Robust Policy Optimization in Deep Reinforcement Learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Robust Policy Optimization in Deep Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.559435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.559435Z digest=sha256:e98f317ef3de27ab019f03819e2dc95c8f13aef811f7105cf5680027f9728f32

Observation 4fc2e2ad-4e12-49c6-800c-3f5284aa57b3 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.023254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.624485Z digest=sha256:579bd643e1f4a427b43aa130582581e389a3d4dc0c5d25bb6703b496a7100272

Observation 9d0de39e-ea15-431c-b06e-bd4542495984 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.000514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.710725Z digest=sha256:9ee1e572cf515d59289616b27ff3e710cd252b23dab6d1c39cc27fe23c50218c

Observation 402f2e2e-4070-4651-aad8-3c837732275a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.897292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.897292Z digest=sha256:2323a5380c8a361d028cc5573a177c3da5728e1f50a2eaa42eef7f3eb3fcd833

Observation 48daab15-c73f-4d61-a90a-335b1c953714 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.979816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.021496Z digest=sha256:04969027221ee87076688081c782caee77c9f45d44b6f0a7c5c3e912c279d709

Observation 8daf1743-76ff-4305-9b2c-f5d5373ddb91 · outbound

This paper cites A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.958176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.121778Z digest=sha256:b3f5fd2087758c9fa7c4c7591bb267147797856d60793130a21f48f0f67a395d

Observation c5b5fe2e-7771-4e7e-9717-807af026f026 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.244632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.244632Z digest=sha256:1ae082d11a205e8c817862d43daedfac44bb7c531ea6cc49da4af31e24cac947

Observation 3352881e-78dc-4e8d-93d7-ad8bbee579a2 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.864149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.407242Z digest=sha256:67b9bdbdd00d3eb4eecae43d627caa447ce1ebad8eb10b50aead1029640c971b

Observation 4fac4cf2-e0b3-47da-87f2-e820812baf33 · outbound

This paper cites V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.723703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.524679Z digest=sha256:2842a77ba230f9f99ac504368853a7fc32523acf28c85119dcdf2459d4f664bc

Observation d8dd8711-3e83-4881-b14d-2ab66946ff93 · outbound

This paper cites V., Lee, Y., Hoang, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Lee, Y., Hoang, T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.388402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.653506Z digest=sha256:694a4bbc88a78f7e1b447e7c249ff033e09a5692bd90e1349c91be99525829e2

Observation bfa7acdc-9448-4e86-9ce1-1a13d2e9f8a9 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.121133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.776828Z digest=sha256:92be0e762665c24aa03db2964581e0ca688e90f38d0fb70aa14cbe385472a0c0

Observation 1af0e510-f2e3-4913-8113-8e41f75d71a1 · outbound

This paper cites T., and Athey, S.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic T., and Athey, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.919306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.919306Z digest=sha256:17bf4ca25334a0e1bdf0d389f819b7808b0254827cf00cd757ecfb7fdbf7b766

Observation db3c2665-14ee-4a57-9ddb-ebd1843af036 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.970054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.076368Z digest=sha256:fbf200ee28cb0bf6bb20af3db7b41f95528c3584ad2d13b4c44630d2fedfcf7e

Observation c0da3255-67e7-4351-81f7-c722802bac01 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.787674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.191439Z digest=sha256:8b705226da6271e8f0a800e5499c853d82814a7b333fdfc887c91f9944773315

Observation 32c8258f-d0a5-4952-8da4-b8b519314caa · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.590333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.346427Z digest=sha256:4c3d014558b8b066fd08dd8af10285e5f5b7abd14b13815d096bf61ad68bc49d

Observation c21a2795-aca7-4916-b332-982f024c8b3f · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.404749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.501633Z digest=sha256:9a23e56c5c8ede185f87a398eedf8f25fc45f72bd8bd9d699b97841e6fd092d9

Observation 68da7043-9bbd-49a3-8a32-de12c8a35dec · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.249365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.671471Z digest=sha256:3ef7b0cbd4699606b34b0148efe5df00ff962db9974816113035e9a62799be81

Observation 23686f90-8263-4ecb-a308-ad80d9cc31a7 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.052297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.792691Z digest=sha256:02653eeb2f35f6aa95b547b55fd34165d2ea92a627a6abfc16d98b59d3c04218

Observation 0adad830-802a-444e-aef3-221fd80f7c2c · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:01.892224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.884611Z digest=sha256:2880d726ad02e09e3bc1d85874b261372e62d53e02bb52f15c2bada21f4ef9ec

Observation 30a6359d-8f35-45dc-9155-a15bfcea025c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.747435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.981793Z digest=sha256:80f6e382a25af276c08fa2fdfebed4c8d44df0bffc4bb752f0e540b6fed8e75f

Observation d1c23697-be22-4220-a151-691c6b834a31 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.551087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T10:36:01.072056Z digest=sha256:60f11eadea8fd9ca93b80ce7d20378d2f00eb8a4636c4e16c1d8f669aba672d7

Pith citing papers

No inbound Pith citation observations are available.