Pith. sign in

Paper Citation Record · LEDGER

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation

As of 21 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2411.09891.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09891 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:17:36.456856Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact2
  • verified fuzzy34
  • unresolved16
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5613138e-a2b7-4f09-b075-fa23f26b74ca · outbound

This paper cites Deep rein- forcement learning for dynamic treatment regimes on medical registry data.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Deep rein- forcement learning for dynamic treatment regimes on medical registry data

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.243880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.247911Z digest=sha256:98dbc70713013d37413ae6677a0884403301358248bc13a38f3fa1f77f132f9d

Observation 3ff031b4-4c2d-4c0a-9c54-4fa6922eb086 · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Deep reinforcement learning for autonomous driving: A survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.252742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.252742Z digest=sha256:dd55daff21a9cc6f0e0693db595e95cfcaf833e648fd4ec8512ab235e75a276e

Observation 4baa2b41-6a6a-4dce-8b3a-f3a1fef53fbd · outbound

This paper cites Off-Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Off-Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.257541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.257541Z digest=sha256:ab7c6b7d5e2ae727a179414116e8ead986eebd52fa88fecaf9098cd498b619a6

Observation fa5f8d35-7bd6-4766-924c-660ddfb74b70 · outbound

This paper cites Sim-to-real interactive recommendation via off-dynamics reinforcement learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Sim-to-real interactive recommendation via off-dynamics reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.222949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.261957Z digest=sha256:63458c7ffdabf8cec509b919216e9feacf65d3f9fce3eac7ba580401f58d79cc

Observation dd32523b-2d57-48a7-a4f0-b2c20d772568 · outbound

This paper cites DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.266052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.266052Z digest=sha256:68f7da17e2940f31e9d4c7182ebd6404e5932172faa696340a87fe9f62fc67df

Observation 3a913f97-f674-4ba4-9dce-f911e3608368 · outbound

This paper cites Unsupervised domain adaptation with dynamics-aware rewards in reinforcement learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Unsupervised domain adaptation with dynamics-aware rewards in reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.211161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.269733Z digest=sha256:985488bfd1e92260510e632baa01ff888b6b0d0aad53c546a06b3b0255fcc175

Observation 7499ddd5-1a1b-4f04-9a6f-ce7f3de53280 · outbound

This paper cites Generative adversarial imitation learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative adversarial imitation learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.273364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.273364Z digest=sha256:fcd52f2d3a14d0e1055d852abc201261c9d5d5eb355beab3762704e38e5e4aa8

Observation fd30c84c-8a34-4fd0-8d82-19e9a3177f5a · outbound

This paper cites Generative Adversarial Imitation from Observation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative Adversarial Imitation from Observation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.277182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.277182Z digest=sha256:95df490b58aeb54d55a367feb3e873905f7157664dfe118e36b030832ffec657

Observation 67525ab4-c36a-43f7-ba45-e7263eda8d37 · outbound

This paper cites Offline imitation learning with a misspecified simulator.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Offline imitation learning with a misspecified simulator

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.192290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.280800Z digest=sha256:026881b386c5d47d80d46071505b2aa8103989b2d4d8bca6027fb86668ac6b4a

Observation 5fb528e4-6cfa-4290-a0b0-cef63400ccd0 · outbound

This paper cites An imitation from observation approach to transfer learning with dynamics mismatch.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation An imitation from observation approach to transfer learning with dynamics mismatch

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.179110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.284153Z digest=sha256:c432823777bfffd1dd47a4907bc0c0759b2a0c42243bd0f6508f2a6386abe044

Observation 2673544c-0942-4e83-b05d-3a617f6e9d54 · outbound

This paper cites State-only Imitation with Transition Dynamics Mismatch.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation State-only Imitation with Transition Dynamics Mismatch

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.288161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.288161Z digest=sha256:373d32097678c8cc50d8ccaf57bc394a9d7ae94f8b32e0bfdeed084abb18e2d8

Observation 81241666-5130-4ff8-9e30-5624202aeeea · outbound

This paper cites Doubly Robust Policy Evaluation and Learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly Robust Policy Evaluation and Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.292286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.292286Z digest=sha256:18cf124a633b04c2bd8c203ba7b9e7418c953aa86bbf5d1e14f3f076a2f0a21c

Observation 5ff84bb0-e52e-42df-8718-052153b1fb00 · outbound

This paper cites Doubly robust off-policy value evaluation for reinforcement learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy value evaluation for reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.164640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.296742Z digest=sha256:51428f3378ba9d1dfa61c96bfaa88020d5413172fe0747fc4edec232ab0b17d1

Observation 3fa4eb3f-6143-4c6c-b60b-5c3ca1ec149c · outbound

This paper cites Doubly robust off-policy evaluation with shrinkage.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy evaluation with shrinkage

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.153092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.301670Z digest=sha256:07d67f0a4ced833ebe1d5a481bce291337f50dc909667710c43df0f92aa4afe8

Observation 93013354-fb64-4f17-a28a-0693ab6bfc36 · outbound

This paper cites Doubly robust off-policy actor-critic: Convergence and optimality.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy actor-critic: Convergence and optimality

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.139372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.305848Z digest=sha256:36af319efd165813211a384ddd9f728602ec2fdfa086cfdf4431403ee77b4102

Observation a67cd4fa-e58b-40a3-a2ee-0b12438c2009 · outbound

This paper cites Doubly robust distribu- tionally robust off-policy evaluation and learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust distribu- tionally robust off-policy evaluation and learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.127030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.309412Z digest=sha256:bbea36724ec64b13c742b9251cd2c9893e404c5537741d66e64c80e7e7d7b1b5

Observation 838e9822-b528-4ded-88ef-3632000b9f7e · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.313407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.313407Z digest=sha256:aff64c99673064b8f5705d163d77ff266e504793be2324c58b51d4cdf0ce9296

Observation b521f84e-07f0-4a70-bd00-a9afb35559db · outbound

This paper cites On the off-dynamics approach to reinforcement learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation On the off-dynamics approach to reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.106346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.316877Z digest=sha256:ba108a1a1954fbab7f6563ee85777457534ef8bcf88f114a4078188a9adaeddf

Observation 983a1df7-5f2d-4b03-b13b-48f96b33386d · outbound

This paper cites When to trust your model: Model-based policy optimization.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation When to trust your model: Model-based policy optimization

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.093026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.320568Z digest=sha256:6fdcafaa5e550ad14242d162f1d26fd8c57d2d3c13031a91363cbac297bea8f3

Observation 2322f4d4-362a-4411-b663-48c7abd51685 · outbound

This paper cites Mutual alignment transfer learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Mutual alignment transfer learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.080478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.323875Z digest=sha256:eb666d4e008fec9830bcce8082d448e9246e0ae51443aff1b19071bc7dd53150

Observation 859041a9-8267-4a5a-b3cc-e5099afa6451 · outbound

This paper cites Domain Adaptation for Reinforcement Learning on the Atari.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain Adaptation for Reinforcement Learning on the Atari

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-12T20:17:36.578965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.327207Z digest=sha256:5dd707afa2b1ee97aecf0614c0bdaecfc8e239da151be483c63514b572b911bd

Observation c9e7c01a-2133-406d-a598-39bc065876a7 · outbound

This paper cites Domain adaptation in reinforcement learning via latent unified state representation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain adaptation in reinforcement learning via latent unified state representation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.066843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.331292Z digest=sha256:fcf2bcdef28c6f4617dc93f5e1c9a513b0265e1afe90586f001f30cb5631e62f

Observation 8df7979f-e2dd-4bd4-bc39-b81e5579f731 · outbound

This paper cites Transfer learning in deep reinforce- ment learning: A survey.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Transfer learning in deep reinforce- ment learning: A survey

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.055530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.335280Z digest=sha256:7ca0eb4ce58f36b5d1c838a5c44ba3adb74d2da47a29c3c74f2a76aedfaf1855

Observation 40b228c3-039a-44d7-b42e-dd51a0522d74 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.339277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.339277Z digest=sha256:e82574e0d5763b34c1ed3ed0cf17852ec798755c100a5749eee2013cd1162fbe

Observation 1a0e4323-7e16-45d5-a78e-8ab3d648c68b · outbound

This paper cites State regularized policy optimization on data with dynamics shift.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation State regularized policy optimization on data with dynamics shift

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.043346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.343376Z digest=sha256:34b3ece5ad245fdf5463e86d832efdd398a4dae048c3adfbdb791d333acacc85

Observation 3cc6b026-d4f9-4812-b364-efe5a0e14d5b · outbound

This paper cites Generative adversarial nets.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative adversarial nets

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.347569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.347569Z digest=sha256:a129a46a2e9827888d2b74a0b66513f1079c4567a9184d86978b10fbf5e4110f

Observation 68384332-9525-40be-9946-9bd08d26d87c · outbound

This paper cites Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.021172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.351975Z digest=sha256:75217cf80d2b03396adde1ee6659104700f8e8968462e3faf95e197fbda25a93

Observation 41eaee32-c7f3-45e9-b0c1-aaa0782819e2 · outbound

This paper cites Learning Robust Rewards with Adversarial Inverse Reinforcement Learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.355560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.355560Z digest=sha256:ca33f5755d23258d573a9d46bdcf4e60bcc15ee6c17a03ab49d8f2262d0349c0

Observation a118e89a-86dd-401d-bf96-580d44b86363 · outbound

This paper cites Imitation learning via kernel mean embedding.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation learning via kernel mean embedding

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:37.007973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.360166Z digest=sha256:9063cb0a6a7106b9fc9b1948cb2426ef2c5579967420a5f59456b8a66e0b19be

Observation 8eb1bf54-3c9e-42ae-a0e2-dbb200a85835 · outbound

This paper cites Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.364226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.364226Z digest=sha256:9aae2a57179816f180347d36319a27d3cbec76d9d61fbf0d13c0dba39fcecda4

Observation 6f4a63f3-1e08-4d69-9b0b-6e5b06b5f65b · outbound

This paper cites Task transfer by preference-based cost learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Task transfer by preference-based cost learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.994200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.368737Z digest=sha256:9f74e61894d6536d88f1e750bb4852d721473296a834e1c621c278cc1a42542e

Observation c26186d0-e667-49f6-8204-4dc139411b68 · outbound

This paper cites Imitation Learning from Video by Leveraging Proprioception.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation Learning from Video by Leveraging Proprioception

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-12T20:17:36.524377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.372938Z digest=sha256:bb802cada41a571a0b58bb4ae98cb8e2ab715dd368f4d515c4dfefdd46271254

Observation 16e2bad1-a861-4f73-a727-725b0cec41eb · outbound

This paper cites Imitation from observation: Learning to imitate behaviors from raw video via context translation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation from observation: Learning to imitate behaviors from raw video via context translation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.981041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.376193Z digest=sha256:5bc574e087baebcbb87d948506065a09ddd1290cc5923138c4677d03b7b3766e

Observation e3e99ab7-cf1f-401e-ade8-3e44214fb1a7 · outbound

This paper cites Behavioral Cloning from Observation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Behavioral Cloning from Observation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.379282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.379282Z digest=sha256:12dfc2df4646f315e9e70be53bec8bb7fda78d6946109aa0b58f4747891591b6

Observation cd31525b-a218-4e47-b5c5-196c1eb8eb77 · outbound

This paper cites Recent Advances in Imitation Learning from Observation.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Recent Advances in Imitation Learning from Observation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.383384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.383384Z digest=sha256:f690ceaf593c7d93ae9f704ca9e4d7d836e45be828327c77f8be3fee08d5aaa1

Observation b555e714-d3e7-4ea0-92ec-8fa8f3e60266 · outbound

This paper cites Domain adaptive imitation learning.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain adaptive imitation learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.387268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.387268Z digest=sha256:d4fc06f76aa208a3dc1a57174c205cb1370f942a9525b6ea199598f746d51831

Observation 99553806-6256-4bc9-89a8-cd82d7a4e646 · outbound

This paper cites Generalization and equilibrium in generative adversarial nets (gans).

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generalization and equilibrium in generative adversarial nets (gans)

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T20:17:36.390932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:17:36.390932Z digest=sha256:27be5fe447c40c29451d9b1b452df95102752c109b7508c6fa2b0103685da86e

Observation 961c3012-d6c7-4f41-ab91-36a32d94b33b · outbound

This paper cites vf+MOWCXlXkD7CB/Zj9tqm3hyT0=.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation vf+MOWCXlXkD7CB/Zj9tqm3hyT0=

Reference 38

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T20:17:36.954194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.394593Z digest=sha256:9f304607bc14c561c007194cb229b537b721bbf412334456924cb66591efe7f9

Observation d54b067f-39d8-4ab1-bdba-831510ef0bcf · outbound

This paper cites And in the introduc- tion section, we have a contribution list.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation And in the introduc- tion section, we have a contribution list

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.941654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.399572Z digest=sha256:416da103bba29f35d78af70613a19cdaf2e730cd4cd367e3008da4bf7a4277ff

Observation a1a4d7ba-c253-4094-ab40-5eb181ccc063 · outbound

This paper cites Limitations.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Limitations

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.930341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.403904Z digest=sha256:417cb7877f70243916e2b7bb219f148ef20acd139b7ddca004b3a3e56a6fb574

Observation 0155ba99-1d8a-4638-aa00-a973529a342e · outbound

This paper cites We present our theoretical result in Section 4 and the proof is in Appendix B.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation We present our theoretical result in Section 4 and the proof is in Appendix B

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.919164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.407613Z digest=sha256:4af813e0c209b0cce5c618ae3f5841ac640e7ccd8fe57557cfa16ff7140cb284

Observation b0cdfb74-91ad-4fcd-9254-ad5b9abc7f1f · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.907775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.412052Z digest=sha256:a0fbc928fd9b662ebfd341acd027e9974405779eeb08ce610c1dd776e9d5ec38

Observation 33486d2d-81cc-46f8-985c-89e0480a87b9 · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.895880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.416059Z digest=sha256:788eb0f5551647c3f17f955517eb5ab00a1f9d30c5818725af866ccfeaae9de2

Observation 7601504f-19b7-48d6-abc1-8400b07369ed · outbound

This paper cites We also describe the hyperparameter tuning in the Appendix D.4.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation We also describe the hyperparameter tuning in the Appendix D.4

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.884247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.420974Z digest=sha256:ac236b942391c1549679186e3325dbc6d1f92beb7b5171a2f735c8e6a70f1f67

Observation e3f0739b-d912-422c-815e-64f036c38ccd · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.872996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.425518Z digest=sha256:a78a9e6327da436df7e82b725c288a456691fddf9ce172abc7a768c8ad437c17

Observation 6c8fae34-bc3a-4d87-8a50-d06a4c9b1a9b · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.861005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.429220Z digest=sha256:476aa59dabd1b05c0079c8edfd5632a381d2e2cc9f5046b4dd187fe51d355a2e

Observation 03c7ca5f-168a-4789-8b6d-8373cf47c48e · outbound

This paper cites Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.848492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.433498Z digest=sha256:0997f8a1d3f4edae6345d42db9b241608b115643c243d844b12037d7046ecb6f

Observation 9398912b-84f5-473b-8888-9b4764310394 · outbound

This paper cites Guidelines: • The answer NA means that there is no societal impact of the work performed.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that there is no societal impact of the work performed

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.836178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.437415Z digest=sha256:70ed7bf8edb8205e9ea4a22c2f966a09fab5a44ad11132c5763deb048244c2bc

Observation 9b412b4b-d424-486a-b262-bf11491823fe · outbound

This paper cites Guidelines: • The answer NA means that the paper poses no such risks.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper poses no such risks

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.822673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.441482Z digest=sha256:d21f611c0c47c779e1122e59a4cbe94864a1bccd832de0f4c341a9d38abf84b5

Observation eccf6cce-6e09-4549-8514-24227f4aff22 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not use existing assets

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.810650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.445554Z digest=sha256:c8f9aff3b458c21db29878cab36b3f4bf212509dfc41829bdee35194901e49fe

Observation aadab1a8-e210-4843-bc9d-b1dcb1ae1500 · outbound

This paper cites Also, details about the implementation are included in the paper.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Also, details about the implementation are included in the paper

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.798768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.449898Z digest=sha256:b1336e848100aa60e0eb8c6b92f876a0fef7ab45242e742d070a3cd21b7f5523

Observation 44774b8e-2c5d-4bd7-b370-8cea819e8347 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.786580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.453463Z digest=sha256:ab2d10095f6e790737e614e07291b9ea09508c9cd1137b98f387b9f8703ea447

Observation c3afa81f-16bb-4e33-b2b4-90db56f1ed3d · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:17:36.773313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T20:17:36.456856Z digest=sha256:b9dcf5be8d57f58c8d070b085e44d948f60b976d320fcf0fdfb973752ec618a8

Pith citing papers

No inbound Pith citation observations are available.