Pith. sign in

Paper Citation Record · LEDGER

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

As of 19 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 3 inbound Pith citation observations for arXiv:2508.12252.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.12252 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:30:47.443824Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T12:40:11.283716Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:48:11.092966Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact3
  • verified fuzzy16
  • unresolved36
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d37e4d1d-4dba-42c5-95e7-7f7825a7d8d7 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.514637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.238609Z digest=sha256:c98f510100b84142056fe26eb0b25400ee62e7bd2aaad35ccff36824c6b9857e

Observation f41eb466-3514-4d85-8663-48e083fba3fc · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.504633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.242784Z digest=sha256:14db610536c024f54d2a2976e41b7ec1f768e100186ced93032b053a90a6302f

Observation 968334f4-9e7c-498c-a6dd-cb1679b83093 · outbound

This paper cites Radosavovic, T.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Radosavovic, T

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.246922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.246922Z digest=sha256:4e4a77afe5b3e09165a3f366700e772e2aac1cc313eab1f57c8300c30d9cc69a

Observation b5a0e30d-2e2b-4382-b56c-bc80d1f746e9 · outbound

This paper cites Radosavovic, S.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Radosavovic, S

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.495368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.250920Z digest=sha256:dcff21d97a63477171dba5e372c1a783daaee149bcf18ddcbc3c3cb535431d53

Observation f9939da6-5c34-4410-98d7-b0dfbd6d781f · outbound

This paper cites Zhuang, S.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Zhuang, S

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.485495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.254384Z digest=sha256:ed50666d47ae5963f8d8bf2816665e450b36fe102b94c7dd0ff217f5a352a113

Observation c5324bb9-50b3-499e-bf79-31b215d2b273 · outbound

This paper cites Tobin, R.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Tobin, R

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.258035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.258035Z digest=sha256:aac8ca90633c15226111b7c8de75209a542ac758488d64bdd2309dcbb4ad2b2e

Observation 83c56bdb-f31f-4c6e-830d-8f7d13086e37 · outbound

This paper cites FiLM: Visual Reasoning with a General Conditioning Layer.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids FiLM: Visual Reasoning with a General Conditioning Layer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.261747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.261747Z digest=sha256:61c9f17b9564ffa0a750b7fa2f3be1483e72da9479fc841370ceaefcf220f7ff

Observation 1e0c12f6-a0f2-4a30-959a-e2be77825452 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Proximal Policy Optimization Algorithms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.265375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.265375Z digest=sha256:cbe89e984017ac8e0e4024d2519ec1f63e1b179919bc7e01edd44f9ea82acbb6

Observation 3c9840ed-ebd8-4794-809d-75667a2a2342 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.475851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.268569Z digest=sha256:adfd415ab6647e2c07c5bc87f0b15d721c54049a9d887b649ee4b57d0886266c

Observation ea45d224-4cf6-45b7-bb85-64ad43ea7a7d · outbound

This paper cites Industrial robotic arm market size, share, growth trends, regional share, competitive intelligence, forecast report 2025–2037, March 2025.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Industrial robotic arm market size, share, growth trends, regional share, competitive intelligence, forecast report 2025–2037, March 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.466195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.271862Z digest=sha256:b8990f5cda3c3a99e98840052b227a6de416c357ef08bea4702f95b0b49c6fe4

Observation 239d04ba-3511-4340-9346-18f28822b316 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.275933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.275933Z digest=sha256:a35fd36919edbbc470d316684cd957b343488ea70907758e8042bcbf9d7af4f1

Observation 40e18464-1ba9-4096-b405-e9a456367d0e · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.455860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.279354Z digest=sha256:54e6d176c37e210154a38e21b95bcc25957c04d49585b80ac419a7b31dd1e6d6

Observation 35570d82-456e-49db-80cc-85dcdd010055 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.446096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.282369Z digest=sha256:21e6aa44556a687a23c083167700c07ecae52d9765b89f1e504e5b0811af9eba

Observation 5ad0fd6f-d9a5-47a8-93ad-f99f45ef128c · outbound

This paper cites Rudin, D.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Rudin, D

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.436445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.285541Z digest=sha256:0b6c8a2ca842724b1939d7a6711b03099523cb481b71603430fcea63fa4b436d

Observation b5f78d16-e46f-4749-83da-e5a3919d9d08 · outbound

This paper cites Cheng, K.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Cheng, K

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.425997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.288867Z digest=sha256:193f7fe7006968d603adf74d421e23c69015561c20104f41d66f60462ff61813

Observation c6f08d01-01e6-473d-8f28-4c72fbe77855 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.415947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.295225Z digest=sha256:e0547ce7b4ccb914c8c4ac52b031866e7576327a2bca24646c45a548f1ab2915

Observation 9c331c7e-e278-4c79-ae1b-02bcee479537 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.405847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.298623Z digest=sha256:51d57ebebc34a7a60977dd9ef64b4968e61371a34d6af3820ccae582314fc54f

Observation 3fbec8d8-bf90-4bf7-a7a2-bb8316c618bb · outbound

This paper cites Akkaya, M.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Akkaya, M

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.395164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.301961Z digest=sha256:dc07e687a9975827aa238d96b2582865fb4cd09576f2f76158126f8839b76538

Observation f3640d43-d959-40a4-bd69-85da64deaf76 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.384823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.305225Z digest=sha256:f49a49c064a0db7c659aa3f19f891b5946acc2ffcd03563fd1cfb5d69f41625e

Observation 362b0df5-864b-49d6-929f-f2860f573c0a · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.375052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.312269Z digest=sha256:40234557ceaa530edb058a99e26f9f4dfd68fa7e649f16b639d57bb486da4a4e

Observation 005b2106-4642-4f28-917e-27e4bf15fb12 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.365376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.315617Z digest=sha256:d2505b20268ecb6f9e3948aeae0640312dff22b2f10918908bc3ae14eb6b9d86

Observation 3d15ee66-7e34-41ef-8bd6-c59f2774d61a · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.355605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.318996Z digest=sha256:428143ef9403688d0dd1e65134668e41ddc64a273d023766f0ae67cc7f9c1d51

Observation 6c90e035-a773-44ff-bc52-23bcde027bf7 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.345347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.322394Z digest=sha256:40e3c257412482b184495151cc5b0a380551f478fba74bb6ab3e6e8249033d7e

Observation 71aab0ef-ca1a-4c6c-b2c8-216f244ef453 · outbound

This paper cites Gu, Y .-J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Gu, Y .-J

Reference 24

Resolution
verified exact
doi, observed 2026-08-15T17:30:47.486718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.325760Z digest=sha256:dab4d1fb0df3be9140408a1de6c23dbd99f76410125f15d2314a12385c304c2d

Observation 75044096-4563-4aec-8718-ebc873d26e13 · outbound

This paper cites Haarnoja, B.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Haarnoja, B

Reference 25

Resolution
malformed identifier
doi_truncated, observed 2026-08-15T17:30:47.475757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.329298Z digest=sha256:4377196dc5e550cd02949b58479876d9ad6b639594a199dacec4fa9e474cd606

Observation 7effc648-d84c-4594-94aa-cad0434d7e01 · outbound

This paper cites Chebotar, A.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Chebotar, A

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.332964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.332964Z digest=sha256:1e50f71261fb5ee727c8f75d7843f17e126d4b84d44ce99d992127c0f191f1eb

Observation 4594d7a6-2660-4202-8a24-36a184897cc3 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.329914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.340494Z digest=sha256:5dbf1c8f95bd61d47663d05c1f9882c7ba642920ad6cbaad0ce47f9012f0ae25

Observation 2c988812-ad24-46eb-a598-4e21d3ba1634 · outbound

This paper cites Huang, X.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Huang, X

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.320569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.343765Z digest=sha256:d99a4debdc5d507f27119950137b3c3dcd5a83ac94e1b2b26bc38b2888047c6e

Observation 76a9b329-a1b5-4a6a-921c-86b6c23e5ef7 · outbound

This paper cites Schoettler, A.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Schoettler, A

Reference 29

Resolution
metadata mismatch
raw_fallback, observed 2026-08-15T17:30:47.861898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.347182Z digest=sha256:2db28dfd112d3d5656455c807002cf552ec6258b2491511f26bb48d68002065c

Observation b706cc61-59a9-4e04-8bbb-782f2aa96d53 · outbound

This paper cites Kumar, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Kumar, Z

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.311130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.350785Z digest=sha256:5a0e6a3bf708920e943417dec169592e57bfb74c69b8890e7de912dd66d9e9d3

Observation 873c65de-f2f3-4303-9c32-f320f8768a07 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.301902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.354172Z digest=sha256:1397e0b721d6db207dca76a4a39212d3c08ee54bdd61ab6dde9f00de76f28422

Observation a12e9ed7-7076-4192-bdfc-1cadc6ae3705 · outbound

This paper cites Kumar, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Kumar, Z

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.357450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.357450Z digest=sha256:c1117371e3044161951d8b68673d3032c8ee5a243977e9b771a0961013f42327

Observation 8cc4ef6e-2177-47f1-a950-e0c6c7822e64 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.361177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.361177Z digest=sha256:c3fb96d1edb3340dff5df69a4c9b30c32df66465d479f0cd94d4d3604b68cc57

Observation 9ccaed51-65be-4678-a6c4-26680fa46111 · outbound

This paper cites Zhang, L.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Zhang, L

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.292248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.364767Z digest=sha256:af2976fadfeecddcea21c7d8d51a9f9b22086801f367a7346f9ab94b34384446

Observation 2f915a04-24f3-4785-9a51-c8c520b94092 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.282832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.368283Z digest=sha256:2712a27840b0e5c07b09cf7284134914cc0b80818ebb48389ed3ec8f7c39d81d

Observation 12025c06-e099-4625-a20e-9b147c563d68 · outbound

This paper cites Xiong, R.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Xiong, R

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.273004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.371895Z digest=sha256:00c8ed504a7b10222ea7d21ef00c43a1cc360a21527761d10ac563bb3bb79a08

Observation e8e5dce2-2f4a-498b-8fe0-1cc134e84f1a · outbound

This paper cites Jiang, C.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Jiang, C

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.263242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.375113Z digest=sha256:1ab61005f3b740bde7c928a52f5b54f5839eb4ab122c5e6e982fc2e0c8f1922a

Observation 6a661b1a-84bc-4ede-b854-34eda987ecc4 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.252509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.378516Z digest=sha256:81fe82434499acf17c953918640e14aa7304083559bb30851fb13d6d583363c1

Observation b0f8de91-dd9a-4916-8aa5-9e57315a01c2 · outbound

This paper cites Meta-Reinforcement Learning of Structured Exploration Strategies.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Meta-Reinforcement Learning of Structured Exploration Strategies

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.381761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.381761Z digest=sha256:fed7e9855d833e46533a43bb374f88747dfc5e02738ebf296218ffca32fc93d2

Observation e6582a89-3d1e-4e84-a0fa-295e8966074a · outbound

This paper cites Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.386032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.386032Z digest=sha256:56cf0a7db456ba51481cb0ca8a0d6b0208f18190a7e6363db9b8baf560b1235a

Observation e22770a7-cd73-483e-84f4-8e9d1d459e4c · outbound

This paper cites Learning Fast Adaptation with Meta Strategy Optimization.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Learning Fast Adaptation with Meta Strategy Optimization

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:30:47.678993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.390195Z digest=sha256:229a3a9239ae9a5dc3cded4337fe73133e286e532be8377a66aef9be50bbb602

Observation 4d432c3d-ef69-4828-bbb6-7b0732dcdc0f · outbound

This paper cites Learning Agile Robotic Locomotion Skills by Imitating Animals.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Learning Agile Robotic Locomotion Skills by Imitating Animals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.393858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.393858Z digest=sha256:7a3c16a22343e67daf934f63714032ba91f00e335d089a84ffcd2ff45da5d1b6

Observation ebd9c970-5feb-4211-a91c-8b71ed150709 · outbound

This paper cites Sim-to-Real Transfer for Biped Locomotion.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Sim-to-Real Transfer for Biped Locomotion

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:30:47.655543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.397602Z digest=sha256:83a0f4627347ab6c2e609ed68189f0c6b4571dcfd596bb0b75b995032fed1b82

Observation 010b27c3-4c1a-47dc-8d4b-5812403aae1b · outbound

This paper cites Huang, Z.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Huang, Z

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.242605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.401207Z digest=sha256:86de8b016417ec7c55c21ca08e5d1855000b8ff0772d6af5b579cc10276768d1

Observation 392ba3d7-f554-4dcf-95b1-e562534909f5 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.404534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.404534Z digest=sha256:5fee7569f39d2447923b2ba1b74405fe0964f9ca89a6f4a9704ac56a109e2a7f

Observation 538ccfa4-410e-4e7a-b367-5de25c014983 · outbound

This paper cites Mendonca, E.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Mendonca, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.232719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.407769Z digest=sha256:32eedfee84b2f78a8b3eb7cb1c5384363459606768aae414e20b0ed5197d424d

Observation 306647f7-cc4e-4172-bcc1-aff4a81347ad · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.222731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.410932Z digest=sha256:eafa1a29bd91002b6cae0deca4904601658391a60f976f71112ebbe545cad40b

Observation f05fe392-6314-41d1-a490-5b2dc110742c · outbound

This paper cites Smith, I.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Smith, I

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.212596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.414738Z digest=sha256:9695301ca21150477563ba359d80f2eef153a09bab497cff93b3ea282cc7e500

Observation 6271e74c-cf73-4d76-833f-58c66ac519ed · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.201870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.418660Z digest=sha256:9238b929a97f1eb14e5f40d735f93711e99ecc35ebec6b3c4f54a6e5169ce571

Observation 247a93df-ba07-4820-b7e1-182055958d70 · outbound

This paper cites Grow Your Limits: Continuous Improvement with Real-World RL for Robotic Locomotion.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Grow Your Limits: Continuous Improvement with Real-World RL for Robotic Locomotion

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.422280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.422280Z digest=sha256:4d354a70ac8227590a44b4c8d2d6adabfdebce082da03ddbab5bbe7ab3c3a34f

Observation 0b4aca1f-ca57-4229-9fa7-d45a7f598f58 · outbound

This paper cites Bloesch, J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Bloesch, J

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.191518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.425779Z digest=sha256:7a758e9b5e34c1840136890428ff11fea0c8d7b7a0e36e249da9c6415dc71ef9

Observation 3f6a04bf-98b0-48ad-95b6-d8bea14a7926 · outbound

This paper cites ROBOTIS OP3 e-Manual.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids ROBOTIS OP3 e-Manual

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:30:48.181109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.429123Z digest=sha256:bd34e04c5db72421db4cc8a31bebdfcae95efe10e4cf9929eae26a6693ffd2e6

Observation da8b53df-09e7-472e-b669-921d70efa3c7 · outbound

This paper cites Maples and J.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Maples and J

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.436554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.436554Z digest=sha256:b015a9ede9cacbbc6b078c9f893f76781bbf778bde234515e5c2add582dabf7b

Observation babeb17d-7682-41c5-b5d2-6449d5601f2e · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:48.166595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.440117Z digest=sha256:c4ebfb9023a16a028370d8cd14cce0d27fb4e8936fb95f60389e4a1ffaca0019

Observation 366b1d46-dcd3-41d9-9605-e3460b3704aa · outbound

This paper cites Ground Truth.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Ground Truth

Reference 55

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T17:30:48.156115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:30:47.443824Z digest=sha256:7c4ca9995bddfca05691271ef55242593bb30dd958bf175979905d7204308bce

Observation e1526a1b-6b16-46b8-b251-92cb4582d7b0 · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.337310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.337310Z digest=sha256:6039c7647347188cf88c68c3019e3725aebb2a081a0b4aa0cfd35ddb15ace141

Observation e4dd0581-0ade-4c68-aa84-8ff18796156b · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.308759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.308759Z digest=sha256:b2d70fd61f8b15b9b88b60180cd2b5a64e8cf000d14fede0d54efb0901823075

Observation ebca5cb5-7f21-4995-9d3f-98db6954476b · outbound

This paper cites an unresolved cited work.

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:47.292052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:47.292052Z digest=sha256:141eca2da21814c11532f859686e1e6633509299329084da1a5cc55dbc9a1d78

Pith citing papers

Observation ea9dacbd-33d9-400b-a4dd-ea8d845c60b7 · inbound

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation cites this paper.

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:06:21.288640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T14:37:01.237169Z digest=sha256:ff2c8e1cbf0895f3cfa8d0eae4cbea68b165d34fdea38d8e47b6892bb58bfcaf

Observation d1a1e7ef-d564-4433-a598-c3b7b16ba780 · inbound

FADA: Few-Shot Domain Adaptation via Dynamics Alignment for Humanoid Control cites this paper.

FADA: Few-Shot Domain Adaptation via Dynamics Alignment for Humanoid Control Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:25:48.479963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-30T01:28:29.441778Z digest=sha256:7d4da9f5837617918dea44a7894f57b9cc359d66eb7e5200406d75ecb01b6cac

Observation d5dff37c-69d6-4fb0-bded-f9c65388aa07 · inbound

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning cites this paper.

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:48:11.094715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-03T12:40:11.283716Z digest=sha256:32492d57d0c943a28d625812011617d6ec711f6760f9a2b427b784598dcccff4