Pith. sign in

Paper Citation Record · LEDGER

Hybrid Cross-domain Robust Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2505.23003.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23003 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:03:56.760781Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy54
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87394822-1158-4e2a-87a6-348b9e263bd9 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Human-level control through deep reinforcement learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:06.986317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:51.710833Z digest=sha256:4e5d82ee073fd1605cb383d344d1ba2738c84519548ec4022d9f39b8f1bece8d

Observation 8a273763-11b3-472c-952a-ba77d63aa8ee · outbound

This paper cites Mastering atari, go, chess and shogi by planning with a learned model,.

Hybrid Cross-domain Robust Reinforcement Learning Mastering atari, go, chess and shogi by planning with a learned model,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:06.678932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:51.743162Z digest=sha256:22f302cc85692c1b679be879d4ee89af4e73f280124cc65fe07e719830ecd96f

Observation 2fd1546c-ea6e-4d82-8626-dde48f253588 · outbound

This paper cites The limits and potentials of deep learning for robotics,.

Hybrid Cross-domain Robust Reinforcement Learning The limits and potentials of deep learning for robotics,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:06.349343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:51.816316Z digest=sha256:7d58df18714f73858006005346205dc26f71b9c84deb03e8070524beb58176ff

Observation 002a37b1-83d8-4b53-81ad-d11512796553 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Hybrid Cross-domain Robust Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:51.926044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:51.926044Z digest=sha256:006c7898d2aee9c858cc2ef101f58f62b50c46489598cd43b44714ce60cab048

Observation 48e32230-1d57-44db-b425-10a9c96f24d9 · outbound

This paper cites Mildly conservative q-learning for offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Mildly conservative q-learning for offline reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:06.072621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:51.980942Z digest=sha256:47d867b3a1e8b4e28d88dc9e3f8ed6bda1d9a4809f653ad6ff173f1ff62774d6

Observation 367462c2-33ff-4703-b988-82980a9bc3ba · outbound

This paper cites Morel: Model-based offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Morel: Model-based offline reinforcement learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.968363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.046563Z digest=sha256:0e2fd373a47ad8c564b53c05ecd9f71660996ddfb9a2a9a398991b2a07ba0db3

Observation affdc510-2133-4651-982d-fd29ce0e93f1 · outbound

This paper cites A Conservative Approach for Few-Shot Transfer in Off- Dynamics Reinforcement Learning,.

Hybrid Cross-domain Robust Reinforcement Learning A Conservative Approach for Few-Shot Transfer in Off- Dynamics Reinforcement Learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.834644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.133369Z digest=sha256:83a8afa0d3decda00bd51452d61be88a596bfd71c26934dc1388ce3e55251b5d

Observation 09fd12e4-4552-42df-8036-c29b447f1380 · outbound

This paper cites A Comprehensive Survey of Cross-Domain Policy Transfer for Embodied Agents,.

Hybrid Cross-domain Robust Reinforcement Learning A Comprehensive Survey of Cross-Domain Policy Transfer for Embodied Agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.656151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.254160Z digest=sha256:88d89a971da7af96dcf240dbb414b00c66e922b7b2ade09e8bb241001498f166

Observation e7d05e23-4de3-44f3-900e-bc388edabed5 · outbound

This paper cites OCEAN-MBRL: Offline Conservative Exploration for Model-Based Offline Reinforcement Learning,.

Hybrid Cross-domain Robust Reinforcement Learning OCEAN-MBRL: Offline Conservative Exploration for Model-Based Offline Reinforcement Learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.463189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.315486Z digest=sha256:6b9ac3ba2110d9cae6bfd25e6ff49f59e3f851612548599e57836c56e13578ff

Observation c4838456-3f8f-40b5-8859-c6eb7ffa252a · outbound

This paper cites When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.277973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.439479Z digest=sha256:9506ce55a46e826f6ff6b148437a5b5d80ce00d9862c46a91c3bea0ee6899ab6

Observation 5fde591e-9809-4b8e-ad12-2313ea00fd07 · outbound

This paper cites H2O+: An Improved Framework for Hybrid Offline-and-Online RL with Dynamics Gaps.

Hybrid Cross-domain Robust Reinforcement Learning H2O+: An Improved Framework for Hybrid Offline-and-Online RL with Dynamics Gaps

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:52.498739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:52.498739Z digest=sha256:e9f0b7e570152ab370d4b6fe033b91b55dffc258372c085f6348365ca89b604f

Observation 30a34d6f-bd9f-48d5-9f3d-c40f5d21773b · outbound

This paper cites DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning,.

Hybrid Cross-domain Robust Reinforcement Learning DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:05.143507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.536485Z digest=sha256:3f1ec2f6b9817175d16ad1e315f67844f8f9553a4931e40cb1ad65e2c26a69c6

Observation 9863a988-28be-49ba-85bb-7bd9587d266e · outbound

This paper cites Beyond ood state actions: Supported cross-domain offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Beyond ood state actions: Supported cross-domain offline reinforcement learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.992162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.636075Z digest=sha256:e7ce172e8c0c7c0f32eea4c31ff214e4b7a74ded8b30f466381528936a1c084e

Observation 148637c7-3901-48ea-a4e7-b42781ba9324 · outbound

This paper cites Contrastive Rep- resentation for Data Filtering in Cross-Domain Offline Reinforcement Learning,.

Hybrid Cross-domain Robust Reinforcement Learning Contrastive Rep- resentation for Data Filtering in Cross-Domain Offline Reinforcement Learning,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.872205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.683004Z digest=sha256:54faefd74f215b624dd2f44b8e6c5b5fce1f90629cc950c2a5e2651dcf0ce4af

Observation dd89e5db-4dc4-4870-8fa0-e7c3a389cd44 · outbound

This paper cites Off- Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers,.

Hybrid Cross-domain Robust Reinforcement Learning Off- Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.684468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.755958Z digest=sha256:821e3cc2ca462fc4f53cacce6e707b599b072d59424d3774bdf4506ff05bbf8f

Observation a76b4118-69e9-4131-85c0-7b2e7cad8f5f · outbound

This paper cites Policy Learning for Off-Dynamics RL with Deficient Support,.

Hybrid Cross-domain Robust Reinforcement Learning Policy Learning for Off-Dynamics RL with Deficient Support,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.555221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.822622Z digest=sha256:7ae4c30e7cb06bcef82619fadf4f8a11994262a910f001822702e564c780e121

Observation 3ad239e2-1cd4-45ff-9689-fa26d57387e1 · outbound

This paper cites Cross- domain policy adaptation via value-guided data filtering,.

Hybrid Cross-domain Robust Reinforcement Learning Cross- domain policy adaptation via value-guided data filtering,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.311896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.875747Z digest=sha256:aba3cfdcc9d6858a267ef3d528fd4db6d3423e2cee9c4ca80236fa3e027a7f52

Observation 02d113a7-65a0-41ba-ae77-7c7499ad273f · outbound

This paper cites Cross-Domain Policy Adaptation by Capturing Representation Mismatch,.

Hybrid Cross-domain Robust Reinforcement Learning Cross-Domain Policy Adaptation by Capturing Representation Mismatch,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:04.044759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:52.993127Z digest=sha256:cfc1ca36733937fcf508ab254a960e04ceaaf1fe314b3998d7f070cb54e9985e

Observation 7da39e0c-304f-4272-994f-29a4ef906acb · outbound

This paper cites Sim-to-real transfer of robotic control with dynamics randomization,.

Hybrid Cross-domain Robust Reinforcement Learning Sim-to-real transfer of robotic control with dynamics randomization,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:03.882049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.042357Z digest=sha256:44719840a685d4322dcdc1c6c16f732eb7b5bf6f1bf38ef155db63e05541d235

Observation 145cbc1a-6d94-4d4b-b850-8b08b5f3167e · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world,.

Hybrid Cross-domain Robust Reinforcement Learning Domain randomization for transferring deep neural networks from simulation to the real world,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:03.661930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.111229Z digest=sha256:3d698f634af09ff48d380dc30e1351c89d71119966a35421da365075865d5a95

Observation 32e376c8-f2e8-44c5-b9de-45dcffe9276f · outbound

This paper cites CAD2RL: Real Single-Image Flight Without a Single Real Image,.

Hybrid Cross-domain Robust Reinforcement Learning CAD2RL: Real Single-Image Flight Without a Single Real Image,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:03.454026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.189497Z digest=sha256:0d216765a60a9d46b49361e659848b74eab89a32365f2e02777c343ee7d1c685

Observation c9fb2cf2-adc1-4e99-9760-fa4483dbd27a · outbound

This paper cites Neural networks for control and system identification,.

Hybrid Cross-domain Robust Reinforcement Learning Neural networks for control and system identification,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:03.193874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.285895Z digest=sha256:a4dbe4e8ea5dc89c402b36d10ec3f488f1fda29f4b552a17d4545f8177e4f072

Observation 9c911f14-a743-456c-a728-989bd090d3d9 · outbound

This paper cites Fast model identification via physics engines for data-efficient policy search,.

Hybrid Cross-domain Robust Reinforcement Learning Fast model identification via physics engines for data-efficient policy search,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.922280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.372933Z digest=sha256:72da910995ce9fe2206017bd3bbaa48c43717859068ad28a312e49da1cedb167

Observation 1a20c7f0-0000-405a-9f1a-227ddeb4d42f · outbound

This paper cites Closing the sim-to-real loop: Adapting simulation randomization with real world experience,.

Hybrid Cross-domain Robust Reinforcement Learning Closing the sim-to-real loop: Adapting simulation randomization with real world experience,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.690470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.422518Z digest=sha256:09c47285705fceb44560d693f704d2ef20a63e535b7fe4f36ecd9ccee2e43ee0

Observation 6eea4444-5e06-4c27-92c9-c004d784e2b6 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

Hybrid Cross-domain Robust Reinforcement Learning Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.419331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.516978Z digest=sha256:a5d3b76a3263ce8e50dcbdaa3b52bd367c15928e68519e7ea93c09bde5560c3e

Observation 580266bb-4c62-4cdf-846c-8794ae4b81a2 · outbound

This paper cites Learning to Adapt in Dynamic, Real-World Environments through Meta- Reinforcement Learning,.

Hybrid Cross-domain Robust Reinforcement Learning Learning to Adapt in Dynamic, Real-World Environments through Meta- Reinforcement Learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.249818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.596714Z digest=sha256:e9937faf8a265ca3ad67164559c5bd1120e1e11bc9d79a7da130db0e5145697e

Observation 5ce9a842-9c91-4969-b833-f743813ed8bd · outbound

This paper cites Zero-shot policy transfer with disentangled task representation of meta- reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Zero-shot policy transfer with disentangled task representation of meta- reinforcement learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.040841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.634400Z digest=sha256:d2958cdf81ddbe7959101ee2283728d475abe760035082b8ca9c8c536b0d8cd8

Observation c68b46a9-de9d-43a5-ac8d-94f6ba9bd43d · outbound

This paper cites Provably good batch off- policy reinforcement learning without great exploration,.

Hybrid Cross-domain Robust Reinforcement Learning Provably good batch off- policy reinforcement learning without great exploration,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.814857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.694525Z digest=sha256:807e76ca511c80a82968058aa59c7f14ac5542575ed3a003cacb749f53fe1d9f

Observation 59de49a2-ba08-4b3b-9e75-fa038ee6f5aa · outbound

This paper cites Conservative q-learning for offline re- inforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Conservative q-learning for offline re- inforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.591693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.799782Z digest=sha256:1aa83763df5efed1d1e118a2a5e8d1c41d90e031832273455a5c4e313ed9a11c

Observation 3c564a64-660a-468b-84fb-8d5875d12b02 · outbound

This paper cites Rambo-rl: Robust adversarial model-based offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Rambo-rl: Robust adversarial model-based offline reinforcement learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.376539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.855773Z digest=sha256:6abca1aa20e11bae8812f6451479d533db87edfe77ee3619581e197e1ab0ec16

Observation b13760c2-771e-4d62-a5a2-2f8a37e9708c · outbound

This paper cites Robust dynamic programming,.

Hybrid Cross-domain Robust Reinforcement Learning Robust dynamic programming,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.119170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:53.936953Z digest=sha256:6066d2be675b79ca99e788d7df377689a5814fbd5b849d98eda12a7ac71e054e

Observation 26094e38-8cbf-4fdd-b247-a7c846d25e51 · outbound

This paper cites Robust control of Markov decision processes with uncer- tain transition matrices,.

Hybrid Cross-domain Robust Reinforcement Learning Robust control of Markov decision processes with uncer- tain transition matrices,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.898945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.010480Z digest=sha256:c8b508f71cb58379fb212616a9ec2acdcdbd2b400686dbb4ce2987ec9705664b

Observation e86b3880-c6eb-41b1-95b6-9a4b3ac32919 · outbound

This paper cites Distributionally robust Markov decision processes,.

Hybrid Cross-domain Robust Reinforcement Learning Distributionally robust Markov decision processes,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.673085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.110255Z digest=sha256:930a2b5f4e64fbb7d8f10669bdd106da9a3316001648ec8bd7610c7e502402d4

Observation 873468c7-91db-4ae5-9636-29fed1e512db · outbound

This paper cites Policy gradient method for robust reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Policy gradient method for robust reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.460610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.242757Z digest=sha256:005e7ed49d93883af070cea484447813c7107cc6822921f577f3910b6a25f48e

Observation 6f1da6b5-b889-49a8-8390-6e3e00541838 · outbound

This paper cites Policy Gradient in Robust MDPs with Global Convergence Guarantee.

Hybrid Cross-domain Robust Reinforcement Learning Policy Gradient in Robust MDPs with Global Convergence Guarantee

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:54.312568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:54.312568Z digest=sha256:cd81c0da794e57a4c32c86e758ac4717d72f2f457d797264d2adb4ff5e543d43

Observation 3495fd71-931c-4174-9809-1c43c14856cf · outbound

This paper cites Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics,.

Hybrid Cross-domain Robust Reinforcement Learning Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.292564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.394415Z digest=sha256:3c71785c421637cb51d4a2e12823b599b6b5ee1398a6216d5f64c37ca04e45d3

Observation a645f45b-6780-4488-b6bc-314e8be9efa1 · outbound

This paper cites Improved sample complexity bounds for distri- butionally robust reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Improved sample complexity bounds for distri- butionally robust reinforcement learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.108759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.478229Z digest=sha256:b62dca1728cc70959744d023af755ce92cd0b7dc6679598f3992891bc12aaaca

Observation 369eb0e5-810c-41d8-9b61-7a9e48e8c098 · outbound

This paper cites Online robust reinforcement learning with model uncertainty,.

Hybrid Cross-domain Robust Reinforcement Learning Online robust reinforcement learning with model uncertainty,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.946253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.549558Z digest=sha256:ff285022cd64fc3b2d0d47d89c0dc7aa1447c300483a81152035988e47bb3393

Observation 49f95e52-ae19-4839-b84e-00120c6e8c93 · outbound

This paper cites Online Policy Optimization for Robust MDP.

Hybrid Cross-domain Robust Reinforcement Learning Online Policy Optimization for Robust MDP

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:54.632287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:54.632287Z digest=sha256:e83764fb52191a7c4ef4909c6dad1a4e1bf43aca6bfb475cd4196f6dbd857567

Observation 1d4a5a95-50c4-4cc0-9296-6ae259515ffb · outbound

This paper cites Finite-sample re- gret bound for distributionally robust offline tabular reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Finite-sample re- gret bound for distributionally robust offline tabular reinforcement learning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.744758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.736585Z digest=sha256:745dd6d3c67a9c7bafa6611ee91071e971d22c3aa479bad0c344b126d40b53bf

Observation 03022763-18cb-4a4d-9113-1f771d41a8e7 · outbound

This paper cites Robust reinforcement learning using offline data,.

Hybrid Cross-domain Robust Reinforcement Learning Robust reinforcement learning using offline data,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.635353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.815160Z digest=sha256:5affef8d02dd58ff1e4b7f659e12db1af385aef9e8dc85711b17a6305585958c

Observation 56dad26c-385e-40dd-802d-5074168bff4a · outbound

This paper cites Learning models with uniform performance via distri- butionally robust optimization,.

Hybrid Cross-domain Robust Reinforcement Learning Learning models with uniform performance via distri- butionally robust optimization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.465546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:54.881160Z digest=sha256:f203f4d909146910873dad99c81a28ade2ca60afac1dd343ab6dc00b7559261d

Observation c8cd894a-0807-4521-b6a3-3d3e495534f1 · outbound

This paper cites Distributionally Robust Model-Based Offline Reinforcement Learning with Near-Optimal Sample Complexity.

Hybrid Cross-domain Robust Reinforcement Learning Distributionally Robust Model-Based Offline Reinforcement Learning with Near-Optimal Sample Complexity

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:54.976956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:54.976956Z digest=sha256:27a62263056477f4240f67c57e827e39b5d34f02c8c7820b64e00647fe8d68e9

Observation b34cba95-4416-4216-b7b1-bb5a0b928fc9 · outbound

This paper cites Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation.

Hybrid Cross-domain Robust Reinforcement Learning Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:57.022169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.069175Z digest=sha256:729f584427af58dc2d96e2dbb157306527d8392f09bbe9271f09a4ea75924f95

Observation 41e46250-17b2-4f09-b02c-d4b8c80b38e7 · outbound

This paper cites Double pessimism is provably effi- cient for distributionally robust offline reinforcement learning: Generic algorithm and robust partial coverage,.

Hybrid Cross-domain Robust Reinforcement Learning Double pessimism is provably effi- cient for distributionally robust offline reinforcement learning: Generic algorithm and robust partial coverage,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.332299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.166807Z digest=sha256:fa31d1c9c32b6b4c12bba6efd60a1a9c3203fc1867669c7d350ab00dd9d8e676

Observation 48ea1df1-0783-4823-a819-bd7e56b38e09 · outbound

This paper cites Using simulation and domain adaptation to improve efficiency of deep robotic grasping,.

Hybrid Cross-domain Robust Reinforcement Learning Using simulation and domain adaptation to improve efficiency of deep robotic grasping,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.209500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.260086Z digest=sha256:ac9979b9f0b1d0a090aa4939d60b9a050b70594b4b4b4f183a156c964cdeebc3

Observation a4b886d6-de1f-4d63-bf24-368515875f3c · outbound

This paper cites Darla: Improving zero-shot transfer in reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Darla: Improving zero-shot transfer in reinforcement learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.063394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.384899Z digest=sha256:d2d51dd54475872fe2f100fc3ddb2add74d5adb4eb8368d80fe8e62bca371cd0

Observation 8e15cec3-3307-469b-9806-3a2877d65008 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Hybrid Cross-domain Robust Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.466436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:55.466436Z digest=sha256:c045a588e3eb723652d5c49939c119a827cd0f03bd0d9feea742af0527c9c585

Observation c59f9d8e-47e3-4e64-afad-b54722738100 · outbound

This paper cites Mopo: Model-based offline policy optimization,.

Hybrid Cross-domain Robust Reinforcement Learning Mopo: Model-based offline policy optimization,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.973425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.540687Z digest=sha256:044c244c224d70a10fc880483bb9bc283a137308fb82959d133e51944e96a625

Observation afe29559-3ad9-4b6b-a139-6719ce0d4608 · outbound

This paper cites Robust adversarial reinforce- ment learning,.

Hybrid Cross-domain Robust Reinforcement Learning Robust adversarial reinforce- ment learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.835115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.636939Z digest=sha256:25b1705f87bd3d7a7afb24e0b6eef30a2b2c821a20aca671cc81c8004c4f6666

Observation 81d67441-b454-4d96-a0e1-162ab450fdb3 · outbound

This paper cites Exponential bellman equation and im- proved regret bounds for risk-sensitive reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Exponential bellman equation and im- proved regret bounds for risk-sensitive reinforcement learning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.699419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.672126Z digest=sha256:be9caebe9045a1dc079107cb7e06cc5f48a6a1a91af4079fe46953bed56c5182

Observation 50448d80-034c-4d4f-b16c-face20210058 · outbound

This paper cites One risk to rule them all: A risk-sensitive perspective on model-based offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning One risk to rule them all: A risk-sensitive perspective on model-based offline reinforcement learning,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.610582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.758678Z digest=sha256:0cba38bd9464f7c4bc6a5024d45c46a708969ea9bbe157420ec0a3fe52641f75

Observation 435bff98-ffe5-417d-90ea-eb3073ccf41d · outbound

This paper cites Corruption-robust offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Corruption-robust offline reinforcement learning,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.451377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.913745Z digest=sha256:1d7094f45e7a126c0ee9106fe0eb62247f8b16103dd668b17e284e0ed98d76db

Observation c21a02b3-c07a-4a0d-9b38-4d1252472888 · outbound

This paper cites Corruption-robust offline reinforcement learning with general function approximation,.

Hybrid Cross-domain Robust Reinforcement Learning Corruption-robust offline reinforcement learning with general function approximation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.317081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:55.982637Z digest=sha256:8363aa8cb894fea4330342d6f56498e28d4914b759be2b628ca9f1df955bfda9

Observation a2d6c9af-417d-42bc-b9ff-29439b77f3f2 · outbound

This paper cites Distributionally robust stochastic programming,.

Hybrid Cross-domain Robust Reinforcement Learning Distributionally robust stochastic programming,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.215785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.038538Z digest=sha256:302f4fa53ff8a0cc966556bd012dcbbd2c739f07464c3fc79335ac333f4b9e8f

Observation 0af4f2ac-0178-4c0c-ba3b-b66a8bc1d4f7 · outbound

This paper cites Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees,.

Hybrid Cross-domain Robust Reinforcement Learning Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.100122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.144266Z digest=sha256:c1d6e842870d234d54bd34a12ba1be5516f6e71bc8d17b8e09f0f463302dfb4d

Observation f8fd7989-8a91-4d0f-8422-8aeaca0fd5e8 · outbound

This paper cites Deep reinforcement learning in a handful of trials using probabilistic dynamics models,.

Hybrid Cross-domain Robust Reinforcement Learning Deep reinforcement learning in a handful of trials using probabilistic dynamics models,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:57.940029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.272869Z digest=sha256:401cbf4a38f7ade716648ddf1ae0ffa44aad492f701571d6f0d2d9700a12f8f8

Observation 71c50236-d4b2-4e64-9896-da802303d253 · outbound

This paper cites Model-Bellman inconsistency for model-based offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Model-Bellman inconsistency for model-based offline reinforcement learning,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:57.745498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.341894Z digest=sha256:89ceb88b45b4613d64272f93895bee853af3a1e4672902b3e29ade495e87bf8f

Observation 698b2d49-2b42-478f-a274-f2f6dd91adf3 · outbound

This paper cites Uncertainty-driven trajectory truncation for data augmentation in offline reinforcement learning,.

Hybrid Cross-domain Robust Reinforcement Learning Uncertainty-driven trajectory truncation for data augmentation in offline reinforcement learning,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:57.606895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.426774Z digest=sha256:d91b5bb39bf0f95c704b681f7b369100497dac5473585b08f8656737531e53ee

Observation ac45104f-9f24-472f-b7e2-db601b000a2b · outbound

This paper cites Prioritized Experience Replay.

Hybrid Cross-domain Robust Reinforcement Learning Prioritized Experience Replay

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:56.555963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:56.555963Z digest=sha256:27c92301e58203935848625c67ecc82893d7160b1ace9053d9cb307b013f669e

Observation e63780ca-5c5c-49d7-966c-9d136d33e92a · outbound

This paper cites labml.ai Annotated Paper Implementations,.

Hybrid Cross-domain Robust Reinforcement Learning labml.ai Annotated Paper Implementations,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:57.446148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.620245Z digest=sha256:87e9d0f01008b2655e1fe6da8bf402dfee49fc128ebbd1693a843424559b9fec

Observation 37130817-48e1-4fa4-8085-b305024b6914 · outbound

This paper cites REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy Transfer.

Hybrid Cross-domain Robust Reinforcement Learning REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy Transfer

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:56.703511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:56.703511Z digest=sha256:5d8b27c687cbcd80e82271936fe0d5537cba8ed7190f42d35459056804696e57

Observation 0a77678a-1738-4910-8d0f-efec2c3f7d5a · outbound

This paper cites -m": multi comp, “-s.

Hybrid Cross-domain Robust Reinforcement Learning -m": multi comp, “-s

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:57.237850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:03:56.760781Z digest=sha256:c0b833c57f1ccf11dc1eebec9135843b2671c80319424e60f5438a37ad4bb189

Pith citing papers

No inbound Pith citation observations are available.