Pith. sign in

Paper Citation Record · LEDGER

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

As of 14 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 5 inbound Pith citation observations for arXiv:2502.02705.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02705 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T11:31:10.810713Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:53:09.773129Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T05:34:40.281222Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f200afa1-4088-42b5-a626-598fb834fd70 · outbound

This paper cites Further suppose ∆r = max s r(s) − mins r(s) and ∆V = max s Vs(s) − mins Vs(s) are finite.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Further suppose ∆r = max s r(s) − mins r(s) and ∆V = max s Vs(s) − mins Vs(s) are finite

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.366765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.791377Z digest=sha256:6813b8c1aa4c53dd0e2d2ae660281e8fc105d7f6c0e1178151a07e1c6d4ab4dc

Observation ba8189ae-c835-4de2-a693-30fa28d15dae · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.639789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.639789Z digest=sha256:6de99a4def8310c358c7040b4df8e6bded8ac39a9ea440f193bd2d9e81bfd98a

Observation 1f0ad521-36c7-459a-9c43-a803c24533e0 · outbound

This paper cites ∞X t=0 γt¯r(st) # = V π real(s) − Vs(s0) ≥ Eρπ real(s).

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning ∞X t=0 γt¯r(st) # = V π real(s) − Vs(s0) ≥ Eρπ real(s)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.334021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.800593Z digest=sha256:c0e60b9c6f401b7c0161e65cb7cecb5fe517c3a351546743c713da3b4daaf4da

Observation a6374873-e67e-449d-b2f5-f380bb0fd588 · outbound

This paper cites Visual Foresight: Model-Based Deep Reinforcement Learning for Vision-Based Robotic Control.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Visual Foresight: Model-Based Deep Reinforcement Learning for Vision-Based Robotic Control

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.660450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.660450Z digest=sha256:7a41cf7d3eb2adb51f2627f56143ed8223d11668d346347c5883a758740da62c

Observation 2d05657e-a2fc-4cb6-a6fe-7a3f97185d8c · outbound

This paper cites FurnitureBench: Reproducible Real-World Benchmark for Long-Horizon Complex Manipulation.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning FurnitureBench: Reproducible Real-World Benchmark for Long-Horizon Complex Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.671456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.671456Z digest=sha256:28f8b60c91a03ac9da4352c3d27ba66779f1b9078c99d37a5988bd0a78042233

Observation 4b4e8856-c57c-45dc-9851-1db3dc103445 · outbound

This paper cites What went wrong? closing the sim-to-real gap via differentiable causal discovery.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning What went wrong? closing the sim-to-real gap via differentiable causal discovery

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.482741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.682334Z digest=sha256:efbf0474a4ad4f071305cdb69c8aeb0aab864758e4e9a738adfadb9ecbc03687

Observation 31e1ca70-602d-456e-b6d9-d70ba6bdd539 · outbound

This paper cites RMA: rapid motor adaptation for legged robots.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning RMA: rapid motor adaptation for legged robots

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.468213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.687275Z digest=sha256:e653d5b60661c27cc6fff64466a6eb7bccb5da4035bce599e02b9000a492296a

Observation b420593d-7593-4074-bc39-2452b91062a2 · outbound

This paper cites DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.691667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.691667Z digest=sha256:b65f615375743c85ce7297567b49420b2a878a59b412a7e9393c1c4c57cf8146

Observation 5819d1c4-b456-4357-947c-e70806da3f95 · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.696945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.696945Z digest=sha256:3f08880f0eac32e9a44f8abfd27c9537be0593e61c12458dda56276c8136c61a

Observation 5d417802-c356-4b0b-a75e-2116b30f597a · outbound

This paper cites Isaac gym: High performance gpu-based physics simulation for robot learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Isaac gym: High performance gpu-based physics simulation for robot learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.452871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.702724Z digest=sha256:96a915f231edf8bb738f97f77db6934a7615a0eb5b6e5e54705d619f493e8c11

Observation 6d3a13f7-8526-4076-96b7-fd3175598aab · outbound

This paper cites ASID: Active Exploration for System Identification in Robotic Manipulation.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning ASID: Active Exploration for System Identification in Robotic Manipulation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.707879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.707879Z digest=sha256:2e74270ec4e3e78ad9c3108e0bffdb9bb43c96442cd9bcd328fe274feab29c9a

Observation f607b61d-4dd9-4cce-bd01-df8e5bd3d53b · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.713116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.713116Z digest=sha256:8b35deac9acd5e571eca5e5d7639a074362fb21c7f5accc90b9829f023f837f9

Observation 9d664a4c-e269-4b2c-852f-a80d26a62428 · outbound

This paper cites Sim-to-real transfer of robotic control with dynamics randomization.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Sim-to-real transfer of robotic control with dynamics randomization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.436844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.718238Z digest=sha256:11831e037f4281897861438c23f5b9d730e78d854ac81a178fc94272a4446606

Observation c997893d-ca5d-4f12-8ff1-c0110fae2c31 · outbound

This paper cites 14 Published as a conference paper at ICLR 2025 Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning 14 Published as a conference paper at ICLR 2025 Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.421532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.723265Z digest=sha256:baaf7662ecccfda25fce4553c3b5b3fde05ddd0d4e62f30b7d7f40bf59fc9405

Observation 691abc2e-98b5-4e5c-a791-9e7d5ee60c29 · outbound

This paper cites A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.727780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.727780Z digest=sha256:858b04ea4df583c024abfa6c6aea3b9af407f4ba0644fa263cf4a0c3df44ccf0

Observation 5172ce9e-65eb-4d46-80d5-5abb480c8306 · outbound

This paper cites Integrated architectures for learning, planning, and reacting based on approxi- mating dynamic programming.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Integrated architectures for learning, planning, and reacting based on approxi- mating dynamic programming

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.403276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.733142Z digest=sha256:4a254891a631b4ad4dd4ecc4d3545c02fa252c663fcbe5fc3811a2f9e53cd3e1

Observation a7d09495-6b85-43a4-b46d-889e013f183d · outbound

This paper cites Domain Randomization via Entropy Maximization.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Domain Randomization via Entropy Maximization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.742814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.742814Z digest=sha256:da1db923532c3569edb634eb0f29059cac54f6a66c04921644d7159e28877474

Observation c4fb5b0d-48b6-4735-b0cf-67483e337c62 · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.753680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.753680Z digest=sha256:970b2b48116c21d2e274e6a005864e42f2e9c274b69f530167104eaa48e7040f

Observation 2b78bf66-7275-4f0c-94a6-2e1ccaa5869e · outbound

This paper cites Exploring Model-based Planning with Policy Networks.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Exploring Model-based Planning with Policy Networks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.759214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.759214Z digest=sha256:9a0b97054c28a8d720a0e718b83aaf6df1ca21c673da2433f78fc692bc25cf02

Observation 27ab88ee-2b3d-42f3-8844-7d79ac2833d4 · outbound

This paper cites Lyapunov Design for Robust and Efficient Robotic Reinforcement Learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Lyapunov Design for Robust and Efficient Robotic Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.765781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.765781Z digest=sha256:23834e9919189065d22aa134cee20e11a3c32f5993a3a1320e05810a6317e8dd

Observation 2bda9692-72c5-4e35-b572-8005643dc710 · outbound

This paper cites Language to Rewards for Robotic Skill Synthesis.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Language to Rewards for Robotic Skill Synthesis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.776502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.776502Z digest=sha256:96aa84bc4505fe9a2640ed9b02d2c74dfe50a6909e691d97b541f481739e7b48

Observation 1ddfece0-6ab0-4aea-b6bf-7c44da3370bb · outbound

This paper cites Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.781558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.781558Z digest=sha256:3de7023092d1a9d20c8dab42e823a1e74b12f3bb4c38756bab91bd297e2f5e98

Observation 71ae0c97-51cd-4358-a01e-b5d603cbb963 · outbound

This paper cites γH Vsim(sH ) + H−1X t=1 γtr(st) − Vsim(s0) # = E γH Vsim(sH ) − γH−1Vsim(sH−1) + γH r(sH−1) + E.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning γH Vsim(sH ) + H−1X t=1 γtr(st) − Vsim(s0) # = E γH Vsim(sH ) − γH−1Vsim(sH−1) + γH r(sH−1) + E

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.350095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.795959Z digest=sha256:6e419858a2013ae34f9a0790db5e57088c1071f5262cf954cf8b5ee9babbe78e

Observation b4f302e0-4a95-4672-83ea-5809449075c5 · outbound

This paper cites an unresolved cited work.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Unresolved cited work

Reference 512

Resolution
unresolved
raw_fallback, observed 2026-08-09T11:31:11.317789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.805828Z digest=sha256:6efd0b6875a3b0ec4e33a05200519cf835544e29706df0fb0d0bc463600b1103

Observation 4c9efcfc-060b-4a32-aa62-4a803ab023ed · outbound

This paper cites We don’t train on any simulation data during real-world fine-tuning because we empirically found it didn’t help fine-tuning performance in our settings.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning We don’t train on any simulation data during real-world fine-tuning because we empirically found it didn’t help fine-tuning performance in our settings

Reference 1536

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.297912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.810713Z digest=sha256:fb2e6c0128a6ba681dd12d9fa54646b325000458ad704c1efd31bdfa6ff51cf6

Observation ffad0629-a6c6-4468-bb3f-66938311de73 · outbound

This paper cites doi: 10.1145/122344.122377.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning doi: 10.1145/122344.122377

Reference 1991

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.738072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.738072Z digest=sha256:0efe80610b5525f2271298e5bdfd3bf153ba03e0cf6d7090bfaeed5c17db1393

Observation 9fba0ee2-6ab2-40b1-a112-85f1dd30b7f9 · outbound

This paper cites 16 Published as a conference paper at ICLR 2025 A P ROOFS Notation Recap.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning 16 Published as a conference paper at ICLR 2025 A P ROOFS Notation Recap

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.384997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.786477Z digest=sha256:20c7b931c69bfc417168a653b30d94feff77a2d5e508cc2a6386b33c67695280

Observation 990175b0-5743-4621-98f1-df4ffdb23c40 · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.747797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.747797Z digest=sha256:f2cd54a126a4b43c30d14ab2eda0254c1900e5a4bcb56e6b075f0141b9b2b422

Observation abd4174c-d514-4e8b-9ee8-4ab1c9174b25 · outbound

This paper cites Imitation Bootstrapped Reinforcement Learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Imitation Bootstrapped Reinforcement Learning

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.676629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.676629Z digest=sha256:e9fab26364b6a118115f1e43c59b0f91c9187203b70f0fcd2cd5ebc7080cef3d

Observation 80f01451-24cf-4fe0-bae2-1284d03b69fb · outbound

This paper cites Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.771362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.771362Z digest=sha256:c9019dd3e6fe48359d4546350fbb99f03676f9f84cfcb74d7de935d05f542895

Observation fe82fe4d-9f49-434f-9614-cfe244d6b4e6 · outbound

This paper cites Off-Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Off-Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.665872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.665872Z digest=sha256:40f638f41031000b51007e0472f482f4c4de46f93cf39668571f2650a75e2179

Observation 8c2261ac-41e5-444e-8a78-21e3ab34c1f8 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Dota 2 with Large Scale Deep Reinforcement Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.635032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.635032Z digest=sha256:db23d4b31bd9ae5eb0b30f137866d8d7ef101674bfafcc3c331270066e6d63bc

Observation c79dad8c-042a-4cc1-bbbc-84b1a4888474 · outbound

This paper cites URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-09T11:31:10.645494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:31:10.645494Z digest=sha256:c2be8ecfcd34229f6fca20d288188d91783775e8b51012d1f1422528e7e77ee7

Observation d004ffad-3089-484b-9a69-bcc78b15ca00 · outbound

This paper cites Yuqing Du, Olivia Watkins, Trevor Darrell, Pieter Abbeel, and Deepak Pathak.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning Yuqing Du, Olivia Watkins, Trevor Darrell, Pieter Abbeel, and Deepak Pathak

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.497464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.655412Z digest=sha256:a017b0a9a82f342c100e11310238108a10925d4a8331889287b7710202e51176

Observation f868a3a1-a76a-4d96-848b-d441a5776a05 · outbound

This paper cites ProcTHOR: Large-scale embodied AI using procedural generation.

Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning ProcTHOR: Large-scale embodied AI using procedural generation

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T11:31:11.513809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-09T11:31:10.650564Z digest=sha256:28327b323993adb61964340361ded2cd758c31da8af0947226bcc123e83fa5ef

Pith citing papers

Observation 9b64537e-37e9-42e4-8dd3-a89bd3ce4825 · inbound

SLAC: Safe and Efficient Real-Robot Reinforcement Learning via Unsupervised Simulation Pre-Training cites this paper.

SLAC: Safe and Efficient Real-Robot Reinforcement Learning via Unsupervised Simulation Pre-Training Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T10:53:09.773129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:53:09.773129Z digest=sha256:2ee9f923df1284477d148ed9ed8d55df1cf489ce9c109589257ed6481b3c2bd8

Observation c8bcfd32-6511-4a9f-ba0f-67ebe3691da3 · inbound

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training cites this paper.

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:06.350159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:06.350159Z digest=sha256:ba964d5a4ea63f7f68930dca8145f17379c31710742315cc49b7672dabf3dde5

Observation 067fe033-1ea1-4e0c-b807-4ff601eb191c · inbound

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation cites this paper.

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.475351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T09:49:03.333757Z digest=sha256:ac4e88f071602a69f0b2b74d93f85d3c7db80d4b272217b38ebadcd4a14b0a56

Observation bbef72da-9f9e-42e2-be14-678f7680644e · inbound

Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations cites this paper.

Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:34:40.284030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T05:32:17.780012Z digest=sha256:e2fb6b46eefb8896157f9ae1fc7278afc1c7cbbb20436a51c4468a643ef084f9

Observation b285372f-d2df-4ee9-8cb0-5417e19f5528 · inbound

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller cites this paper.

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T01:41:50.557922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:41:50.557922Z digest=sha256:3f4c9f3e283f9b8f16099a4401afebedc2c66050b548c71c85c2f870edfc3768