Pith. sign in

Paper Citation Record · LEDGER

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic

As of 12 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.05445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05445 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:01.072056Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9d48139d-8836-41e3-ae6b-2f8ff8f667bd · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.279110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.605339Z digest=sha256:feed9c5b8b3df7d246762b43ebe9109e028916a311bb078f03def2c7248cd5a0

Observation ebfad1ab-7e77-4f31-908c-554a0d59c0ca · outbound

This paper cites and Pearl, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Pearl, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.227641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.665123Z digest=sha256:45f04456b4ff047cbcccba79b8ad7b1a950ee70a8c06d051d2d5161ae74ab61e

Observation d45b2319-f964-4499-bb6c-e92a8eae836f · outbound

This paper cites OpenAI Gym.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:56.756738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:56.756738Z digest=sha256:d244b19f93378ed7160875e937859efc46442cceabf27cdf66a88f70a76f4ea9

Observation c9389d4d-dee5-477b-bc34-d3fbf3a36cdc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.172607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.823275Z digest=sha256:9d2db10ef6010b665a156c94c17ff24393ff6c3257e91bcaa8f12e1bb033248e

Observation 269ca6ea-c4e5-459a-b44b-bdd72d06a58d · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.126760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.896354Z digest=sha256:bdb638c19fc8b51060b41db70cbce71259e429362200b25a55d7801204fada4e

Observation 0246fd85-67da-4d61-be80-6ef066d20f43 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.066815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.970646Z digest=sha256:bea951b7ec62faa32d0524e77875dab38bb6273bcd36e509bf56f9f224c3c7ee

Observation 54e0fd50-c7ff-48b2-92bc-d12fdeea3cd7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.004765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.045401Z digest=sha256:55c26cd3129c1e801a20f1cdee784829126adf673f766366e759f3d122ace7f7

Observation bcaa6341-0c28-4e2d-b9e9-124faad92207 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.944649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.100420Z digest=sha256:1c67fae29d4101070d7b57a1be1c8651bde2ab75c3e503f9afe07e2811d5acec

Observation 0db23e0e-4f18-4ba0-9a9a-f05a486d13d7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.877110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.151008Z digest=sha256:4ef2ae91e4883d8a3396771bbf2d6b04bfbba4de1ea1691ecd8836f17ba98f3a

Observation 19c2d0ed-9252-4fb4-a7f4-1523034c1e75 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.794972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.207458Z digest=sha256:012cc8a796f63cb5c187272967d4ae374bbfdeb58dfeb5528be5612c3b2661da

Observation 93cededb-216f-4353-96b7-e98981c165ab · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.744172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.263934Z digest=sha256:8c1a6d90b3aedd4da54849733cb4d1ad4001ea5c5e18615d1f74eeb93cb97ad2

Observation a3c08d5c-115f-451d-aa63-d8a87d64f85c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.686487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.336074Z digest=sha256:cdc6bb5ab739d295e9890bbba73610a80baadc3d76be16d2e63336006caf595e

Observation e26ac464-1600-40d0-a5c7-8eac7cf39a4a · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.606814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.402866Z digest=sha256:7417e85010524da1bac54e7d833725dde625bfdf4dcc3cc6ce789ed8bf36b964

Observation 1bb3ecff-3de4-4781-a77d-9f9581352569 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.529392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.466460Z digest=sha256:07391624f0864f2d4ac70b060e56c32fd1fc939b159514760b927a2dcb402353

Observation 1c900d36-b428-43cd-bbd6-7f6eb859d864 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.519671Z digest=sha256:4c80ccfa6380ba02ee16237806d6ed10c9bed516d09801c85c87efddc896cf3c

Observation a5195288-9cf4-4d74-8860-25e085d20bd6 · outbound

This paper cites P., Hunt, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic P., Hunt, J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.392073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.574314Z digest=sha256:139b7a3ac3dfc3a7f8c044b6fdd0f71826a100a2fd1fa05d234f780f7644eef3

Observation 1f25f871-a3d6-4f61-b15e-c6f07e9ebbbd · outbound

This paper cites Continuous control with deep reinforcement learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Continuous control with deep reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:57.644181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:57.644181Z digest=sha256:a783f98068cc684452a952ca55f8a97b18c67bd1e2b343754804712dcf032e3d

Observation 94d44d67-9b6d-4e78-b265-4a61b205d29e · outbound

This paper cites and Krishnamurthy, A.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Krishnamurthy, A

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.330213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.711279Z digest=sha256:92b93433e641ff50c4019dc22cd2e1026f910941af18e9baca2934ac3402e7b9

Observation 254e7504-a210-4e33-9413-903c850d747a · outbound

This paper cites Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:36:01.307595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.781252Z digest=sha256:9be931469610f3326e9ee5223bb619ce518c96ca961f0b5c62238cf80240d9d3

Observation 100ddbe6-78c6-4a76-95b9-2d421e05afc1 · outbound

This paper cites V., Sima, K., and Leong, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Sima, K., and Leong, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.277446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.852453Z digest=sha256:c6c87653c01577797a11a5833a935631fda7f17b3af1daaf2aba20ef81c79ff8

Observation cd705998-f34d-428b-b67f-219cefb0ea64 · outbound

This paper cites V., Fu, D., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Fu, D., and Leong, T.-Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.223838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.918663Z digest=sha256:010804a123587fd077fe56ed4ee03b23c3bda0f2c3024272fdda4de8d2caee7e

Observation 169f980f-2cc9-41a2-b6cc-4eb0f12180f6 · outbound

This paper cites V., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., and Leong, T.-Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.201758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.008523Z digest=sha256:6a697d0e3ce97adc13d30843cf855de860a0427aa2003e849a6f09041e7311de

Observation 08ee4ad2-093f-4500-914e-d302f1fdeed7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.173753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.116971Z digest=sha256:fea2ef3ff561fedb3dff110f2d47dc7beab0b874f59018c54772dfe2432b8a19

Observation cf3d0505-6724-4442-8426-b9eb59d5a31c · outbound

This paper cites and Sontag, D.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Sontag, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.138332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.205073Z digest=sha256:3f2ec3873bbe37aaae34c999b03b776a16e706da4c9028abb1e7d61fdd7a2f39

Observation 14a94d8c-10b4-4183-af31-0c2e57d9b490 · outbound

This paper cites o lkopf, B., R \.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic o lkopf, B., R \

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.099722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.290756Z digest=sha256:abf877564427a18430f52420518f68dec98fe152d3d7566f117e78d415babc8b

Observation 581578b9-8bdf-4b87-9b1e-d95a612d4adc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.072534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.377118Z digest=sha256:100a530f78c013c04e2a095ac64213591c04d218355d4978ea7cae9a5ff530ce

Observation e968b5eb-b467-4199-881c-78b8a88897fa · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.045301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.488977Z digest=sha256:4c704b3ece1dc47a68f340366490a03e3a329babdfbf79250deaa4a911becd45

Observation 217fafc5-d441-4cde-a5b1-50b4c0ab6376 · outbound

This paper cites Robust Policy Optimization in Deep Reinforcement Learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Robust Policy Optimization in Deep Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.559435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.559435Z digest=sha256:e98f317ef3de27ab019f03819e2dc95c8f13aef811f7105cf5680027f9728f32

Observation 4fc2e2ad-4e12-49c6-800c-3f5284aa57b3 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.023254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.624485Z digest=sha256:5016aba135e1ab4b66c0e0b594e47e5e0a5d882b67cef76ac73afd97d8bf02d6

Observation 9d0de39e-ea15-431c-b06e-bd4542495984 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.000514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.710725Z digest=sha256:ed2bd2069f50e4d0c9416975432b92b9e5eb140be3f5409e64f8b1cac9ecfe7f

Observation 402f2e2e-4070-4651-aad8-3c837732275a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.897292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.897292Z digest=sha256:2323a5380c8a361d028cc5573a177c3da5728e1f50a2eaa42eef7f3eb3fcd833

Observation 48daab15-c73f-4d61-a90a-335b1c953714 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.979816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.021496Z digest=sha256:ba22a802663da45eb09ea09cbe814c4a9564e831e1e7e83bb0e26d3c37ae81d0

Observation 8daf1743-76ff-4305-9b2c-f5d5373ddb91 · outbound

This paper cites A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.958176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.121778Z digest=sha256:1b96a6a8348d37d77b708659cf74e57b2e9246647ade688fd7f60acf224ae939

Observation c5b5fe2e-7771-4e7e-9717-807af026f026 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.244632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.244632Z digest=sha256:1ae082d11a205e8c817862d43daedfac44bb7c531ea6cc49da4af31e24cac947

Observation 3352881e-78dc-4e8d-93d7-ad8bbee579a2 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.864149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.407242Z digest=sha256:26a2849d3450017ad8998b7448d8d8fbd3902922f40d825be0c863b30a33027e

Observation 4fac4cf2-e0b3-47da-87f2-e820812baf33 · outbound

This paper cites V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.723703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.524679Z digest=sha256:de0c7e23b06715769174c211ada086b781c8f904493e663e1981dca515040882

Observation d8dd8711-3e83-4881-b14d-2ab66946ff93 · outbound

This paper cites V., Lee, Y., Hoang, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Lee, Y., Hoang, T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.388402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.653506Z digest=sha256:b3485dba075d2065d50c48effcb2da998cf94d4d5d30c84634113478d743e1a1

Observation bfa7acdc-9448-4e86-9ce1-1a13d2e9f8a9 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.121133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.776828Z digest=sha256:2b2be9d6734f5e15b11cd3dfe0def3db3a61c4a0ddf7a438c0d6c273cdf5ad2d

Observation 1af0e510-f2e3-4913-8113-8e41f75d71a1 · outbound

This paper cites T., and Athey, S.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic T., and Athey, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.919306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.919306Z digest=sha256:17bf4ca25334a0e1bdf0d389f819b7808b0254827cf00cd757ecfb7fdbf7b766

Observation db3c2665-14ee-4a57-9ddb-ebd1843af036 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.970054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.076368Z digest=sha256:c70008851d423fe96902842f3b1dc4693808bfd328aa7f9286ee6ae56d5265eb

Observation c0da3255-67e7-4351-81f7-c722802bac01 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.787674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.191439Z digest=sha256:2168f25832d67358d87b071e4f5ae15067e7f4c1a4dcc42b7ef766568c4c6943

Observation 32c8258f-d0a5-4952-8da4-b8b519314caa · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.590333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.346427Z digest=sha256:805dbed9bef71b1e6084fff726d63752659e9e799bfc3c89d2d052fa3f7d0441

Observation c21a2795-aca7-4916-b332-982f024c8b3f · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.404749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.501633Z digest=sha256:f6685f44904786dccfdc5548cf1d8fd7ccdfef5d6d38e4886dd6ccab5f3cbc49

Observation 68da7043-9bbd-49a3-8a32-de12c8a35dec · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.249365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.671471Z digest=sha256:265702d783e45afe29cdbf2ec235a929e6735c403a3da7ad9c99426d04e9d18b

Observation 23686f90-8263-4ecb-a308-ad80d9cc31a7 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.052297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.792691Z digest=sha256:7e969882a350a74ac61709343d6961ccb3af963e4fa118a2ef8bc158e29fe5af

Observation 0adad830-802a-444e-aef3-221fd80f7c2c · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:01.892224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.884611Z digest=sha256:1f9ab89cf0d34e2bbb484f27dbfeb5779c849d6b4e456fcc8188e70ba93467da

Observation 30a6359d-8f35-45dc-9155-a15bfcea025c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.747435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.981793Z digest=sha256:eb3decefcf4d58a0dbf3aaf3f0cd88f75e1dcf286ebe5413e1dc92c735a1380f

Observation d1c23697-be22-4220-a151-691c6b834a31 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.551087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T10:36:01.072056Z digest=sha256:6218b159ab1850475e36e976a89452e2daf093e06d2f837015ae122050f3975d

Pith citing papers

No inbound Pith citation observations are available.