Pith. sign in

Paper Citation Record · LEDGER

4D Visual Pre-training for Robot Learning

As of 7 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2508.17230.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.17230 v2

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:03:21.467027Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T04:18:02.341742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:05:50.616851Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a635510-11fe-4aeb-9a56-f8e3070eb9b9 · outbound

This paper cites Diffusion-Based Representation Learning.

4D Visual Pre-training for Robot Learning Diffusion-Based Representation Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.017683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.017683Z digest=sha256:d1a2cb82418ccd9051b7cf3341083f7b5f9c8e5ff775dd565ada69f0e794a34d

Observation 966afc96-74aa-45a7-8078-86a83e867bc3 · outbound

This paper cites Cobot magic: An open-source robotic system.https://global.agilex.ai/products/ cobot-magic, 2025.

4D Visual Pre-training for Robot Learning Cobot magic: An open-source robotic system.https://global.agilex.ai/products/ cobot-magic, 2025

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.338006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.139352Z digest=sha256:6f8f7d4d3fd8ef7684a72e69a075e9cedb6155bfdd3ade265e033bf2b2e53bf4

Observation eb778560-faca-4d28-99e0-00c7c75dabbd · outbound

This paper cites Is conditional gen- erative modeling all you need for decision-making?, 2023.

4D Visual Pre-training for Robot Learning Is conditional gen- erative modeling all you need for decision-making?, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.324751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.225443Z digest=sha256:999ffe244acd3d1ea6c4090cf939e933e43800c1fedb2aef2cce54576fec32aa

Observation 342b70fd-7065-47be-b0f9-bb647dbf006c · outbound

This paper cites Diffusion Policy: Visuomotor Policy Learning via Action Diffusion.

4D Visual Pre-training for Robot Learning Diffusion Policy: Visuomotor Policy Learning via Action Diffusion

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.230022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.230022Z digest=sha256:e64d487227614d32c6038f8bb73338639251494eb3d2a5a2659da6eaff8412bd

Observation 080d14c6-c09b-47f0-a02b-db9d260702ed · outbound

This paper cites An unbiased look at datasets for visuo-motor pre-training.

4D Visual Pre-training for Robot Learning An unbiased look at datasets for visuo-motor pre-training

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.311444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.235078Z digest=sha256:5596e0ee03b54102cb9ee4cb69cfdc5aa20e9a82b5d4753f2adb900ff5ef3270

Observation a5c3e32a-eacd-4969-9d67-60d8dc9eb804 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

4D Visual Pre-training for Robot Learning Imagenet: A large-scale hierarchical image database

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.239383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.239383Z digest=sha256:01f4a3f9b36e512a8830ded5d3070ae9774a45d03286dd81d7a28a308c0e2924

Observation fe6f73e4-40b5-457f-b7bc-28a6fa23caca · outbound

This paper cites Implicit behavioral cloning, 2021.

4D Visual Pre-training for Robot Learning Implicit behavioral cloning, 2021

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.285073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.243345Z digest=sha256:75b8b62ccfea9396d67fafe421162d822a69a71f16bc515c2a29988956926bc8

Observation c6797594-bb4b-411a-bd3a-acd7e4450313 · outbound

This paper cites Mobile aloha: Learning bimanual mobile manipulation using low-cost whole-body teleoperation.

4D Visual Pre-training for Robot Learning Mobile aloha: Learning bimanual mobile manipulation using low-cost whole-body teleoperation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.270882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.247675Z digest=sha256:c3cf5c4c8f5b29c92fcd753090d6765d60042bd0b2fbf943b741452961bbeb84

Observation 46de06a7-c92b-41d2-a328-70e47d45189a · outbound

This paper cites Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation.

4D Visual Pre-training for Robot Learning Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.251510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.251510Z digest=sha256:c5fb1681736b0529dc95f168f4dc6a87353ad85bca8da73c4c87430708b7a258

Observation 2b504652-2d46-471b-9b87-7fbefc7645a0 · outbound

This paper cites Rvt: Robotic view transformer for 3d object manipulation.

4D Visual Pre-training for Robot Learning Rvt: Robotic view transformer for 3d object manipulation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.256034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.256034Z digest=sha256:f872b786415b608e2829718ba606e1e445366aa8e87fe509969e910725688b32

Observation 1870c50b-f686-4f1c-9515-cb016b6942d0 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

4D Visual Pre-training for Robot Learning Ego4d: Around the world in 3,000 hours of egocentric video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.260015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.260015Z digest=sha256:23058e48c1d8b75f8b9ac3ac8d978768f3dfabf9895f8f199fce979ecc318a1a

Observation dd8634bf-ca7c-421b-a163-a6e0de0fb295 · outbound

This paper cites Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

4D Visual Pre-training for Robot Learning Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.263906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.263906Z digest=sha256:6a06b159903927f08ab1036812ff599a58d74097801e8c90b7c66d8902197e3d

Observation ccfc23a4-d579-4898-8f7e-a97b12171a8f · outbound

This paper cites Spatio-temporal self-supervised representation learning for 3d point clouds.

4D Visual Pre-training for Robot Learning Spatio-temporal self-supervised representation learning for 3d point clouds

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.231230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.267884Z digest=sha256:d1babf7a51f5222be95c039e65d1f0dc4f7683631df4a23605bc1bcd78f32a61

Observation bd3e5533-a953-4d6f-8a46-80ed849f5f90 · outbound

This paper cites Diffusion Reward: Learning Rewards via Conditional Video Diffusion.

4D Visual Pre-training for Robot Learning Diffusion Reward: Learning Rewards via Conditional Video Diffusion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.271906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.271906Z digest=sha256:fd840de9b23a621f1829190f370e204241b8d61ae2dd38e58477e4e9d059a0d4

Observation 6beb5b35-e348-4134-bc37-0d3edd306821 · outbound

This paper cites SODA: Bottleneck Diffusion Models for Representation Learning.

4D Visual Pre-training for Robot Learning SODA: Bottleneck Diffusion Models for Representation Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.276099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.276099Z digest=sha256:bf3f1536874f3c1a89e64287ff7ea0708767aa7653b3ec940b94e5a2b13000ad

Observation 2386e89f-27f0-4fd7-82d2-4ae15162be01 · outbound

This paper cites Equilibrium free-energy differences from nonequilibrium measurements: A master-equation ap- proach.Physical Review E, 56(5):5018, 1997.

4D Visual Pre-training for Robot Learning Equilibrium free-energy differences from nonequilibrium measurements: A master-equation ap- proach.Physical Review E, 56(5):5018, 1997

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.217786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.280367Z digest=sha256:b81d45f2716a110bd5284c6c2273015c13b72ef743a1495d57647e20a2366d9c

Observation e70ad54d-784e-41ca-9509-eec2899144b2 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

4D Visual Pre-training for Robot Learning RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.284052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.284052Z digest=sha256:c5f26098601f0f3c0cd16b4f709c130baebed87542962602d34886aa9bb31f68

Observation 3c498518-9db0-48d2-87c9-381072874f5e · outbound

This paper cites Point- voxel cnn for efficient 3d deep learning.Advances in neural information processing systems, 32, 2019.

4D Visual Pre-training for Robot Learning Point- voxel cnn for efficient 3d deep learning.Advances in neural information processing systems, 32, 2019

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.204285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.288110Z digest=sha256:95d571d9238409b7062abce3dcf04c0a421f728264c4fb834c8c632b95e677db

Observation c2315448-5b1d-40e2-b8e8-bb1a31c695e6 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelli- gence?Advances in Neural Information Processing Systems, 36, 2024.

4D Visual Pre-training for Robot Learning Where are we in the search for an artificial visual cortex for embodied intelli- gence?Advances in Neural Information Processing Systems, 36, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.191842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.291729Z digest=sha256:ebc7c577db79a1cac6988a434311e13783553a791c3251c4cecbd1baa8b35020

Observation 0f1c4598-5ea5-478c-865a-46dbad63b42e · outbound

This paper cites What mat- ters in learning from offline human demonstrations for robot manipulation, 2021.

4D Visual Pre-training for Robot Learning What mat- ters in learning from offline human demonstrations for robot manipulation, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.178483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.295581Z digest=sha256:e4e688607c7c4d508116770ddbc4d983bbbd8bf1cb6913756de54d918fb3d9a2

Observation 15af83ac-f03b-4365-b84c-63b1c1091ae3 · outbound

This paper cites Self-supervised point cloud prediction using 3d spatio-temporal convolutional networks.

4D Visual Pre-training for Robot Learning Self-supervised point cloud prediction using 3d spatio-temporal convolutional networks

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.165293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.299132Z digest=sha256:ceb796fd6c535a6de9c9908df25ae99e50e65e448d9e8fa9249ac7ce055aef37

Observation f22873ba-802a-4dab-8b1d-4e9684a4ecff · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

4D Visual Pre-training for Robot Learning R3M: A Universal Visual Representation for Robot Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.302768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.302768Z digest=sha256:2fa0e561e7d88edfcd2e82c41c859e55c1a30728e35837b2ee984884bfccd0c6

Observation 0fe27fdf-95fd-4700-b6cf-bc3b8bad835c · outbound

This paper cites Henriques.

4D Visual Pre-training for Robot Learning Henriques

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.151794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.306611Z digest=sha256:a8b79492c8ec7610b1d1f2c40776bb7e1c8a08e71734b6526f8a9d17b180fccc

Observation 6b24ad63-b50d-469b-a89c-1e25287e20ed · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

4D Visual Pre-training for Robot Learning Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.310338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.310338Z digest=sha256:236d082db7a9e993e158286818b26db0473f4d421507ce5da6d4d602afd85669

Observation a95946da-2f04-4e23-be52-f1fb9e04c6aa · outbound

This paper cites Masked autoencoders for point cloud self-supervised learning.

4D Visual Pre-training for Robot Learning Masked autoencoders for point cloud self-supervised learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.138718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.314258Z digest=sha256:00d42019990e621b04d43ecf36c1600b11512e1496489d051fcdc04493337057

Observation ea8bc1c0-84f5-4536-b868-67f7ae53481d · outbound

This paper cites Recon- structing hands in 3d with transformers.

4D Visual Pre-training for Robot Learning Recon- structing hands in 3d with transformers

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.124476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.317885Z digest=sha256:d73ab5751edb3d5d5ae498f017153884b4d0dc6dc5b055e9acd034e7ad0bd9c6

Observation 873be75f-5acb-4648-bb70-3c580c09c8d2 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30, 2017.

4D Visual Pre-training for Robot Learning Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30, 2017

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.110432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.321493Z digest=sha256:37cdef42c97c2c4c5d0213b25fa5cf08707be9e0f5a78e39f9556109660f2930

Observation 8df40319-180d-46a0-85e5-20a73191b6b3 · outbound

This paper cites Dexmv: Imita- tion learning for dexterous manipulation from human videos,.

4D Visual Pre-training for Robot Learning Dexmv: Imita- tion learning for dexterous manipulation from human videos,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.096738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.325644Z digest=sha256:f6c61ac3eda5d73eaa0356c0c5af5f1f128307da98865c56bf933b7e74198743

Observation 587b6cd7-41b6-42ad-b69b-bbbaf3b0cb7b · outbound

This paper cites AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System.

4D Visual Pre-training for Robot Learning AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.329618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.329618Z digest=sha256:15d70cdf273b52fb487a2f2491f1c8eeda9340953ef95934b0a8a5a75ab4a876

Observation 6ff6bb74-1dcb-46bc-b59d-99192dcf5c15 · outbound

This paper cites Robot learning with sen- sorimotor pre-training.

4D Visual Pre-training for Robot Learning Robot learning with sen- sorimotor pre-training

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.083894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.333713Z digest=sha256:f21d83824e54240edc0324a2e79652d78819ab7b8c6fe00a9f940525e1862bf6

Observation 01aea381-36a0-446c-ae4a-6c80dead3cb8 · outbound

This paper cites Real-world robot learn- ing with masked visual pre-training.

4D Visual Pre-training for Robot Learning Real-world robot learn- ing with masked visual pre-training

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.070451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.337505Z digest=sha256:df5aa1ccd79bae8a6cb22fea3b97b5c3f4cc01dae3c59371ab0196e0ead8ba75

Observation 7819deb8-c349-417d-aa66-fb36aa512f19 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

4D Visual Pre-training for Robot Learning Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.341137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.341137Z digest=sha256:b066fefeec5e6c2ae3aa65c19c9f46fce5db03311c762cad8983c5c7eaef4279

Observation 7219eb5c-e32d-48ec-b163-99877d0db233 · outbound

This paper cites Learning complex dexterous manipulation with deep reinforcement learning and demonstrations, 2018.

4D Visual Pre-training for Robot Learning Learning complex dexterous manipulation with deep reinforcement learning and demonstrations, 2018

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.057737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.345166Z digest=sha256:a8b0b5b38778e0634adb68f4a26a67eced2beb255ee64b810d707906a9bb45b6

Observation 83e69080-267e-4218-a39e-0ae7159eb3d6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

4D Visual Pre-training for Robot Learning High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.349065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.349065Z digest=sha256:f623fb80f5995ab17154538dc591b79a37d1ba417ea7e7491102b5c31be815f4

Observation 6d495e07-46de-4594-870b-be071a7cf748 · outbound

This paper cites On bringing robots home, 2023.

4D Visual Pre-training for Robot Learning On bringing robots home, 2023

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.035067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.352885Z digest=sha256:f1215f4f5d34af94afd4a3c69b8a648cff1e25be0efb4ad8aec350a446a00208

Observation cf8d3867-1c0e-4240-9737-d6ee8115d10d · outbound

This paper cites RRL: Resnet as representation for Reinforcement Learning.

4D Visual Pre-training for Robot Learning RRL: Resnet as representation for Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.357019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.357019Z digest=sha256:c512861079d7c761cf0bf0f21ef2d19380600b20be430ed332c66a04ad53f56b

Observation 95454da0-576b-4b4e-b3a7-4e83b7279563 · outbound

This paper cites Perceiver- actor: A multi-task transformer for robotic manipulation.

4D Visual Pre-training for Robot Learning Perceiver- actor: A multi-task transformer for robotic manipulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.361117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.361117Z digest=sha256:a3c292d0ab1b335dfa8d32afb9cb2454bf20f6e1d78bbb966ef72f0ac7dc0906

Observation 49078f5f-c4b8-47b6-a274-86a6d7c86745 · outbound

This paper cites Shelving, stacking, hanging: Relational pose diffusion for multi-modal rearrangement, 2023.

4D Visual Pre-training for Robot Learning Shelving, stacking, hanging: Relational pose diffusion for multi-modal rearrangement, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.009913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.364875Z digest=sha256:1f317496553afdd429c2feb0b6017b6c3bd9806b38c1f50bbc5943076a6762b3

Observation 9f46b6ad-ee52-44f0-80a0-87d9a34d2370 · outbound

This paper cites Denoising Diffusion Implicit Models.

4D Visual Pre-training for Robot Learning Denoising Diffusion Implicit Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.368501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.368501Z digest=sha256:f1162ac9b5b731fc021fc1bcbd133788a5de39b9a48756cabb3bf89268350144

Observation 96721581-04c8-4a0b-a2f2-303837a39311 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

4D Visual Pre-training for Robot Learning Score-Based Generative Modeling through Stochastic Differential Equations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.372448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.372448Z digest=sha256:7acc4f839aea2baab3d5507fc216050a108c2872e3d080a2fed037f24e5eb361

Observation 1444e2da-ffe5-44fb-9b2e-0c47c5c6a110 · outbound

This paper cites Memory-consistent neural networks for imitation learning, 2024.

4D Visual Pre-training for Robot Learning Memory-consistent neural networks for imitation learning, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.995667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.376270Z digest=sha256:85a50791f6d1d8236c84107e4aa10b15a877d29fdb2ff0fe7d7c4d167999447b

Observation bc7977ff-7c9b-41fa-9420-907c6db02869 · outbound

This paper cites Mujoco: A physics engine for model-based control.

4D Visual Pre-training for Robot Learning Mujoco: A physics engine for model-based control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.981393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.380078Z digest=sha256:78d47e7136fa7bfc205c11c655d47f51164e1da1fb2e9cbd030d7ecfd38ea5bd

Observation 585ba0d6-7156-49ef-b8b0-cab3c9a8a156 · outbound

This paper cites Se(3)-diffusionfields: Learning smooth cost func- tions for joint grasp and motion optimization through diffu- sion, 2023.

4D Visual Pre-training for Robot Learning Se(3)-diffusionfields: Learning smooth cost func- tions for joint grasp and motion optimization through diffu- sion, 2023

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.967856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.384022Z digest=sha256:de094dc89400c8f82cd4badaec2efc3cb6a8cf43dece989aa50ebbe91754b5f0

Observation 036d5778-04d8-49ac-bae7-4f680b674757 · outbound

This paper cites RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective.

4D Visual Pre-training for Robot Learning RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.388478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.388478Z digest=sha256:5f237da3518d1eda84ac4a3c9703d6348a9fcbe6e09fbe0cc96e5c2039c24bb7

Observation a11434cf-28ad-42a3-a12e-98d85db0e938 · outbound

This paper cites Equivariant Diffusion Policy.

4D Visual Pre-training for Robot Learning Equivariant Diffusion Policy

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.392455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.392455Z digest=sha256:8f7feeab0d8e81f61852103175021507b6148754579457a680b1dca1fa8fa1a4

Observation 3814d33a-265c-484a-948c-05b98b483286 · outbound

This paper cites Diffusion models as masked autoencoders.

4D Visual Pre-training for Robot Learning Diffusion models as masked autoencoders

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.954631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.396448Z digest=sha256:0e3f5b76cbb38d5fb6b5d8c623f1fcffc015454bc0393bb05227afbff8078fe7

Observation b26cf4c9-f623-4209-9d3e-215b8295e897 · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

4D Visual Pre-training for Robot Learning RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.400105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.400105Z digest=sha256:7eb7c3c0f91b524e738af628b2f02ef77bd96160399de7bb475dde17f53fa481

Observation f979dbea-2f6b-4fb0-a28e-5b27753ca864 · outbound

This paper cites Tiangong.https://x-humanoid.com/ bt.html, 2025.

4D Visual Pre-training for Robot Learning Tiangong.https://x-humanoid.com/ bt.html, 2025

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.941597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.404049Z digest=sha256:0c7f87b200cabae81038bc5ed875871d75a5eefa53702ccfb2e419d438d24b2f

Observation 81782620-d8d6-45c8-8234-d84116b6eef8 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

4D Visual Pre-training for Robot Learning Masked Visual Pre-training for Motor Control

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.407790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.407790Z digest=sha256:2473a7966e1711283440aa49b27f54e610702c298b8139c86d876da16e6ec93b

Observation d885646e-f68e-477c-8b69-0b544d592100 · outbound

This paper cites NeRFuser: Dif- fusion guided multi-task 3d policy learning, 2024.

4D Visual Pre-training for Robot Learning NeRFuser: Dif- fusion guided multi-task 3d policy learning, 2024

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.927021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.411478Z digest=sha256:055fd744277c225ecc0c6927f0dfe256d6ad9aa6ecbc8600ac8c23edb614f17c

Observation 2e69c9c8-8ab2-4008-977c-30668a545c17 · outbound

This paper cites EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning.

4D Visual Pre-training for Robot Learning EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.415481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.415481Z digest=sha256:e6547e919f314221c46dfe7eb1ac782cfc8d077e9cdc2682d3f00eb6dd874f87

Observation 9b44c538-5c30-499b-bfec-bc31f1d38e88 · outbound

This paper cites Visual point cloud forecasting enables scalable autonomous driving.

4D Visual Pre-training for Robot Learning Visual point cloud forecasting enables scalable autonomous driving

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.419705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.419705Z digest=sha256:6000a117133a9a2a7ed66da7c52f7ee1a1f5ede9d5248606579ed21ef4dddedc

Observation d0832534-c200-4442-8c82-819a061c9a55 · outbound

This paper cites Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning.

4D Visual Pre-training for Robot Learning Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.904281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.423343Z digest=sha256:d489977f1b19a60383b8cbb15acf41ba8edf6f44a55a5dbd45e9e8d8eb0e3c15

Observation 417a6389-f1d0-460f-a1d6-f24e13ea1190 · outbound

This paper cites Visual reinforcement learning with self- supervised 3d representations.IEEE Robotics and Automa- tion Letters, 8(5):2890–2897, 2023.

4D Visual Pre-training for Robot Learning Visual reinforcement learning with self- supervised 3d representations.IEEE Robotics and Automa- tion Letters, 8(5):2890–2897, 2023

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.890339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.427088Z digest=sha256:7715ba8c454d127450a05f8a6bc2855f6bb0b4f19376ab49fb20cec277f02d1b

Observation 7ea36b4e-021c-4441-a4fb-60df61569095 · outbound

This paper cites Gnfactor: Multi-task real robot learning with generalizable neural feature fields.

4D Visual Pre-training for Robot Learning Gnfactor: Multi-task real robot learning with generalizable neural feature fields

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.430762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.430762Z digest=sha256:f356b7dcdcf8990420b33d7d458db354b7674aacda38d46a6e6c3c95c39f4e5d

Observation 1b55d1fb-3853-4c82-b410-47d256aa4e45 · outbound

This paper cites Generalizable Humanoid Manipulation with 3D Diffusion Policies.

4D Visual Pre-training for Robot Learning Generalizable Humanoid Manipulation with 3D Diffusion Policies

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.434778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.434778Z digest=sha256:370883d935286fce65f3f872fc45350c930b01c0e3b4945e88b25416dbe6eb64

Observation 4771448d-7443-4c28-ac4e-6864a87f1567 · outbound

This paper cites 3d diffusion policy: Gen- eralizable visuomotor policy learning via simple 3d repre- sentations.

4D Visual Pre-training for Robot Learning 3d diffusion policy: Gen- eralizable visuomotor policy learning via simple 3d repre- sentations

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.867924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.439233Z digest=sha256:baafea2f0780bc7eb8860813e0a11b40625b256b7d5dc6d7f4f8ef0e147629d9

Observation 4260ea29-7672-4d90-8b1c-6304b1696c41 · outbound

This paper cites Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing sys- tems, 35:27061–27074, 2022.

4D Visual Pre-training for Robot Learning Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing sys- tems, 35:27061–27074, 2022

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.442925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.442925Z digest=sha256:b67ec817a54bb39580163001e31d5c1d78e9b926dbacdcaa83ffe8d62f576bdc

Observation 2d85b8e6-4ffb-4fe8-823f-1a8e01e82eb8 · outbound

This paper cites VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks.

4D Visual Pre-training for Robot Learning VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.446755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.446755Z digest=sha256:08a86cb706040f174c5d017598f9f76354bfc1567c5ffaa5301167e5370a1d80

Observation f6840b21-1ffe-46aa-8b0f-5fd9e948e0b5 · outbound

This paper cites Complete-to-partial 4d distillation for self-supervised point cloud sequence representation learning.

4D Visual Pre-training for Robot Learning Complete-to-partial 4d distillation for self-supervised point cloud sequence representation learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.846631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.451231Z digest=sha256:95bda91cf88bad9f40567b385a78111961ec3c9d41ca1bc3c4b86eb1f22f87d9

Observation b6a15598-3cd3-4011-9959-ac5f20d1bb43 · outbound

This paper cites Point transformer.

4D Visual Pre-training for Robot Learning Point transformer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.833350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.454842Z digest=sha256:0cb78171a705798a0b44d712f5c2b7e1518574396c3a947bd269a311bc7ab2c2

Observation e334a8fc-fadb-4334-90f8-de9692eb52d2 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

4D Visual Pre-training for Robot Learning Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.458999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.458999Z digest=sha256:c3353afa90d1a12c2d07735921aff84b886c86d0ae0210c80fd820bb000c3619

Observation 07fd99e0-d749-4779-88ba-ef697ffc9aea · outbound

This paper cites Point Cloud Pre-training with Diffusion Models.

4D Visual Pre-training for Robot Learning Point Cloud Pre-training with Diffusion Models

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:03:21.509768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.462860Z digest=sha256:960305e1ca89bb81df7a597fe8d190c99404c6cff848c74b24fee7fb086b76d6

Observation 6d5802aa-b5b4-439d-9860-80114dee874f · outbound

This paper cites 3d shape generation and completion through point-voxel diffusion.

4D Visual Pre-training for Robot Learning 3d shape generation and completion through point-voxel diffusion

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.820003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T17:03:21.467027Z digest=sha256:c726bbe51d42757accf94e6a375dbd2f71c8b823c39c44f068bd7f1f252b1a00

Pith citing papers

Observation 5a7e9431-ae2c-4ab2-8f13-ae8a4b1ad604 · inbound

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration cites this paper.

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration 4D Visual Pre-training for Robot Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:05:50.618485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T04:18:02.341742Z digest=sha256:5db578d4a5de10a2154e52ba959ca315be5f7aa55504a60f1d0c23ab04e98326