Pith. sign in

Paper Citation Record · LEDGER

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

As of 16 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 3 inbound Pith citation observations for arXiv:2508.12252.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.12252 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:30:47.443824Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T12:40:11.283716Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:48:11.092966Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact3
  • verified fuzzy16
  • unresolved36
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d37e4d1d-4dba-42c5-95e7-7f7825a7d8d7 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.514637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.238609Z digest=sha256:00dae39c727cfb70619759e3e802dca5dd563c82968ed0bc58088c1f4f30b90d

Observation f41eb466-3514-4d85-8663-48e083fba3fc · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.504633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.242784Z digest=sha256:ec6be0d3f6411209ac06d8b6fe84ebdee252cd5783ddab465dd6a372c4f73484

Observation 968334f4-9e7c-498c-a6dd-cb1679b83093 · outbound

This paper cites Radosavovic, T.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Radosavovic, T

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.246922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.246922Z digest=sha256:4e4a77afe5b3e09165a3f366700e772e2aac1cc313eab1f57c8300c30d9cc69a

Observation b5a0e30d-2e2b-4382-b56c-bc80d1f746e9 · outbound

This paper cites Radosavovic, S.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Radosavovic, S

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.495368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.250920Z digest=sha256:9fc2d544fcc941f57fd8defc1ba1fc78cd87456b873b83dc0d10a5cb380568a3

Observation f9939da6-5c34-4410-98d7-b0dfbd6d781f · outbound

This paper cites Zhuang, S.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Zhuang, S

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.485495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.254384Z digest=sha256:22940c908e868813aa4555b63ec876d09f39ec91d7d00baabe00406fef3b84f0

Observation c5324bb9-50b3-499e-bf79-31b215d2b273 · outbound

This paper cites Tobin, R.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Tobin, R

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.258035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.258035Z digest=sha256:aac8ca90633c15226111b7c8de75209a542ac758488d64bdd2309dcbb4ad2b2e

Observation 83c56bdb-f31f-4c6e-830d-8f7d13086e37 · outbound

This paper cites FiLM: Visual Reasoning with a General Conditioning Layer.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids FiLM: Visual Reasoning with a General Conditioning Layer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.261747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.261747Z digest=sha256:61c9f17b9564ffa0a750b7fa2f3be1483e72da9479fc841370ceaefcf220f7ff

Observation 1e0c12f6-a0f2-4a30-959a-e2be77825452 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Proximal Policy Optimization Algorithms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.265375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.265375Z digest=sha256:cbe89e984017ac8e0e4024d2519ec1f63e1b179919bc7e01edd44f9ea82acbb6

Observation 3c9840ed-ebd8-4794-809d-75667a2a2342 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.475851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.268569Z digest=sha256:d75cdb010dfa6e1b7fb2f4023a8b394cd4dd9c79bd6d382cbf231fba175f2299

Observation ea45d224-4cf6-45b7-bb85-64ad43ea7a7d · outbound

This paper cites Industrial robotic arm market size, share, growth trends, regional share, competitive intelligence, forecast report 2025–2037, March 2025.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Industrial robotic arm market size, share, growth trends, regional share, competitive intelligence, forecast report 2025–2037, March 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.466195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.271862Z digest=sha256:7435f2dc989901c44dcef9bbbf69826ba0d0c7aa77a892002b5fadb1bb5f1cff

Observation 239d04ba-3511-4340-9346-18f28822b316 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.275933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.275933Z digest=sha256:a35fd36919edbbc470d316684cd957b343488ea70907758e8042bcbf9d7af4f1

Observation 40e18464-1ba9-4096-b405-e9a456367d0e · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.455860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.279354Z digest=sha256:ebcd38f37b098f6b5c05b65fd410d2a65a7381d013f600106cdfc47e57e7037c

Observation 35570d82-456e-49db-80cc-85dcdd010055 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.446096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.282369Z digest=sha256:b84ed37670734222568113f3f5b771d358a5121791ea45c03acf023b2cf1d37c

Observation 5ad0fd6f-d9a5-47a8-93ad-f99f45ef128c · outbound

This paper cites Rudin, D.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Rudin, D

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.436445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.285541Z digest=sha256:688b82642a3a408c34a5c1fcbf962951b295224ac52f8a4b4bef34f1043a78b4

Observation b5f78d16-e46f-4749-83da-e5a3919d9d08 · outbound

This paper cites Cheng, K.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Cheng, K

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.425997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.288867Z digest=sha256:c066e2229bbd8ca584d010d41ab83bb8f4e136df686d67faa10f317caceb27e4

Observation c6f08d01-01e6-473d-8f28-4c72fbe77855 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.415947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.295225Z digest=sha256:cbf716de784ad6bb92064a416b6d4aed1efeab6df017485bd2c1e2819aac625c

Observation 9c331c7e-e278-4c79-ae1b-02bcee479537 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.405847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.298623Z digest=sha256:04af27c7d495341f0f1b07a38f4e5ddb3982c1d7adf1fe2e960951f28cc95c31

Observation 3fbec8d8-bf90-4bf7-a7a2-bb8316c618bb · outbound

This paper cites Akkaya, M.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Akkaya, M

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.395164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.301961Z digest=sha256:92d370b3cd875babfe95f45c6258f73bf92482cddbcc9f764118c479a485a45f

Observation f3640d43-d959-40a4-bd69-85da64deaf76 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.384823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.305225Z digest=sha256:8e09a2ecb305a9a6059124518b41e371006682bbf954e73718da1db7e14a1103

Observation 362b0df5-864b-49d6-929f-f2860f573c0a · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.375052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.312269Z digest=sha256:62a06cc92a932984cef716f2769907a326216a0942c0e548f1b62f931644b6ce

Observation 005b2106-4642-4f28-917e-27e4bf15fb12 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.365376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.315617Z digest=sha256:34f40dc14e0dae70a33673a7885799c2b5793287e93a8b34998f89405e2902f0

Observation 3d15ee66-7e34-41ef-8bd6-c59f2774d61a · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.355605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.318996Z digest=sha256:957dfed0b167321610f53e7c377c382bd1ebd58d46269bc37aba16ee2362ee40

Observation 6c90e035-a773-44ff-bc52-23bcde027bf7 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.345347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.322394Z digest=sha256:1b42dada1ade6718a51d480df469edbe8e3bd84a8314512adb76b72c50d137c2

Observation 71aab0ef-ca1a-4c6c-b2c8-216f244ef453 · outbound

This paper cites Gu, Y .-J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Gu, Y .-J

Reference 24

Resolution
verified exact
doi, observed 2026-08-15T17:30:47.486718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.325760Z digest=sha256:b55f65ce2ebf1f9e561705efb332e31d167bbe1dd8e317e03c6eb4cc1a2e16f4

Observation 75044096-4563-4aec-8718-ebc873d26e13 · outbound

This paper cites Haarnoja, B.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Haarnoja, B

Reference 25

Resolution
malformed identifier
doi_truncated, observed 2026-08-15T17:30:47.475757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.329298Z digest=sha256:b4298e597daab0f115019fcd3289ae8db473183c9b9771eb7bf373259ae2e50f

Observation 7effc648-d84c-4594-94aa-cad0434d7e01 · outbound

This paper cites Chebotar, A.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Chebotar, A

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.332964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.332964Z digest=sha256:1e50f71261fb5ee727c8f75d7843f17e126d4b84d44ce99d992127c0f191f1eb

Observation 4594d7a6-2660-4202-8a24-36a184897cc3 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.329914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.340494Z digest=sha256:15c327a6d787dc6ed6fcf2a6df8192e938047d15d6e410e87963580958ddc2fe

Observation 2c988812-ad24-46eb-a598-4e21d3ba1634 · outbound

This paper cites Huang, X.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Huang, X

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.320569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.343765Z digest=sha256:1017a090e000ce158682297c73a4d01a33b82db98e74fd97dde267cdb84583ea

Observation 76a9b329-a1b5-4a6a-921c-86b6c23e5ef7 · outbound

This paper cites Schoettler, A.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Schoettler, A

Reference 29

Resolution
metadata mismatch
raw_fallback, observed 2026-08-15T17:30:47.861898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.347182Z digest=sha256:59e586eaff7b293d1b40a882068d78f306dbfb865fa21f8aaa8ab4503df16119

Observation b706cc61-59a9-4e04-8bbb-782f2aa96d53 · outbound

This paper cites Kumar, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Kumar, Z

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.311130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.350785Z digest=sha256:a01a6929568e8e5fde9769cdf3aaa316ed3bd33caf94857ed3d10e47bc3614cf

Observation 873c65de-f2f3-4303-9c32-f320f8768a07 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.301902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.354172Z digest=sha256:1c506631293b17daba8248c81261968f58942d095f2bb74ccef7d0ace9bc54de

Observation a12e9ed7-7076-4192-bdfc-1cadc6ae3705 · outbound

This paper cites Kumar, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Kumar, Z

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.357450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.357450Z digest=sha256:c1117371e3044161951d8b68673d3032c8ee5a243977e9b771a0961013f42327

Observation 8cc4ef6e-2177-47f1-a950-e0c6c7822e64 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.361177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.361177Z digest=sha256:c3fb96d1edb3340dff5df69a4c9b30c32df66465d479f0cd94d4d3604b68cc57

Observation 9ccaed51-65be-4678-a6c4-26680fa46111 · outbound

This paper cites Zhang, L.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Zhang, L

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.292248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.364767Z digest=sha256:cd9c39b6fb87d554a41e4925860283d455fd8b4e4b5c988c0aa98f10cda910a5

Observation 2f915a04-24f3-4785-9a51-c8c520b94092 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.282832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.368283Z digest=sha256:1b8360cf1c1653c1cb4df67a179a49f57bad1c68f103ae4f15d74d7701e07338

Observation 12025c06-e099-4625-a20e-9b147c563d68 · outbound

This paper cites Xiong, R.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Xiong, R

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.273004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.371895Z digest=sha256:641a8edee584e6a6dcb39958cb563d778d89600c04ed6b0776ca7b0f55af10a2

Observation e8e5dce2-2f4a-498b-8fe0-1cc134e84f1a · outbound

This paper cites Jiang, C.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Jiang, C

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.263242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.375113Z digest=sha256:d58088a634a327588f615b68bb563264291fef0190b2e2376e2592afd693a764

Observation 6a661b1a-84bc-4ede-b854-34eda987ecc4 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.252509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.378516Z digest=sha256:ea02e20fa81592f441ef63f39679e643edaf0746056ba0670e30e21d7dbb823a

Observation b0f8de91-dd9a-4916-8aa5-9e57315a01c2 · outbound

This paper cites Meta-Reinforcement Learning of Structured Exploration Strategies.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Meta-Reinforcement Learning of Structured Exploration Strategies

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.381761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.381761Z digest=sha256:fed7e9855d833e46533a43bb374f88747dfc5e02738ebf296218ffca32fc93d2

Observation e6582a89-3d1e-4e84-a0fa-295e8966074a · outbound

This paper cites Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.386032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.386032Z digest=sha256:56cf0a7db456ba51481cb0ca8a0d6b0208f18190a7e6363db9b8baf560b1235a

Observation e22770a7-cd73-483e-84f4-8e9d1d459e4c · outbound

This paper cites Learning Fast Adaptation with Meta Strategy Optimization.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Learning Fast Adaptation with Meta Strategy Optimization

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:30:47.678993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.390195Z digest=sha256:f9928721998681e419c20193533699ef7906499eadf482f625e975c97e0ab011

Observation 4d432c3d-ef69-4828-bbb6-7b0732dcdc0f · outbound

This paper cites Learning Agile Robotic Locomotion Skills by Imitating Animals.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Learning Agile Robotic Locomotion Skills by Imitating Animals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.393858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.393858Z digest=sha256:7a3c16a22343e67daf934f63714032ba91f00e335d089a84ffcd2ff45da5d1b6

Observation ebd9c970-5feb-4211-a91c-8b71ed150709 · outbound

This paper cites Sim-to-Real Transfer for Biped Locomotion.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Sim-to-Real Transfer for Biped Locomotion

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:30:47.655543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.397602Z digest=sha256:58938aa7f5b6fe59fe6636952f39ac6e44d6e8e6a873565d8995641347ef3c26

Observation 010b27c3-4c1a-47dc-8d4b-5812403aae1b · outbound

This paper cites Huang, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Huang, Z

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.242605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.401207Z digest=sha256:6b54568b3368036368f6e39557d15ec76dcf857c91c8f148b50163a7f10f9119

Observation 392ba3d7-f554-4dcf-95b1-e562534909f5 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.404534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.404534Z digest=sha256:5fee7569f39d2447923b2ba1b74405fe0964f9ca89a6f4a9704ac56a109e2a7f

Observation 538ccfa4-410e-4e7a-b367-5de25c014983 · outbound

This paper cites Mendonca, E.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Mendonca, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.232719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.407769Z digest=sha256:c0ce3489bede835ddc2bf5b4ee168184f6018e2d714e33cf87502f36aa9cca79

Observation 306647f7-cc4e-4172-bcc1-aff4a81347ad · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.222731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.410932Z digest=sha256:15cf256b5eaa9c3941face682f8f96de5e1d982c90ed66473290e0d2a7180276

Observation f05fe392-6314-41d1-a490-5b2dc110742c · outbound

This paper cites Smith, I.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Smith, I

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.212596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.414738Z digest=sha256:9ebeb0b53682196902e87a52752592eded0211ed3b1016be2c0e98ea98cc0e42

Observation 6271e74c-cf73-4d76-833f-58c66ac519ed · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.201870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.418660Z digest=sha256:c94c3db41874159e5c6dee7c12baebaa5d1928879d336bb1d88403b7c98700b8

Observation 247a93df-ba07-4820-b7e1-182055958d70 · outbound

This paper cites Grow Your Limits: Continuous Improvement with Real-World RL for Robotic Locomotion.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Grow Your Limits: Continuous Improvement with Real-World RL for Robotic Locomotion

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.422280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.422280Z digest=sha256:a7ea00c9df579410eddc7e86688718583300a2473e2f5a17e17542c1d9776095

Observation 0b4aca1f-ca57-4229-9fa7-d45a7f598f58 · outbound

This paper cites Bloesch, J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Bloesch, J

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.191518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.425779Z digest=sha256:450dfc5f5ccb21a198c624ccd9856c317163952c313e5c9ed5e33e6ce95c2532

Observation 3f6a04bf-98b0-48ad-95b6-d8bea14a7926 · outbound

This paper cites ROBOTIS OP3 e-Manual.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids ROBOTIS OP3 e-Manual

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.181109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.429123Z digest=sha256:bf4ddde5873fdcca3eea013ba827234e6e9ee9d693d84867866715091557fe06

Observation da8b53df-09e7-472e-b669-921d70efa3c7 · outbound

This paper cites Maples and J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Maples and J

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.436554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.436554Z digest=sha256:b015a9ede9cacbbc6b078c9f893f76781bbf778bde234515e5c2add582dabf7b

Observation babeb17d-7682-41c5-b5d2-6449d5601f2e · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.166595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.440117Z digest=sha256:02df936a3469b311580572db1899d2d61752622c27301d91faa9ecb52c600f29

Observation 366b1d46-dcd3-41d9-9605-e3460b3704aa · outbound

This paper cites Ground Truth.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Ground Truth

Reference 55

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T17:30:48.156115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T17:30:47.443824Z digest=sha256:f66e6ea5c2e0d2b1cfc36882011e3b9144fa2a06e83118488a1327e9a65f54c0

Observation e1526a1b-6b16-46b8-b251-92cb4582d7b0 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.337310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.337310Z digest=sha256:6039c7647347188cf88c68c3019e3725aebb2a081a0b4aa0cfd35ddb15ace141

Observation e4dd0581-0ade-4c68-aa84-8ff18796156b · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.308759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.308759Z digest=sha256:b2d70fd61f8b15b9b88b60180cd2b5a64e8cf000d14fede0d54efb0901823075

Observation ebca5cb5-7f21-4995-9d3f-98db6954476b · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.292052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.292052Z digest=sha256:141eca2da21814c11532f859686e1e6633509299329084da1a5cc55dbc9a1d78

Pith citing papers

Observation ea9dacbd-33d9-400b-a4dd-ea8d845c60b7 · inbound

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation cites this paper.

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:06:21.288640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T14:37:01.237169Z digest=sha256:e0219e725478e855bb4fed2c5c9a852d4bea44d45b9127156ccf376e201bc573

Observation d1a1e7ef-d564-4433-a598-c3b7b16ba780 · inbound

FADA: Few-Shot Domain Adaptation via Dynamics Alignment for Humanoid Control cites this paper.

FADA: Few-Shot Domain Adaptation via Dynamics Alignment for Humanoid Control Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:25:48.479963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-30T01:28:29.441778Z digest=sha256:fc604043235afefc805a0fa9fb4c1e4c8dc8e75e76a21bfbec9474b10533e45e

Observation d5dff37c-69d6-4fb0-bded-f9c65388aa07 · inbound

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning cites this paper.

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:48:11.094715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-03T12:40:11.283716Z digest=sha256:d358749f1593a950f9622362acd08f0ab29da6d700280f5af7f657bde1668fad