Pith. sign in

Paper Citation Record · LEDGER

4D Visual Pre-training for Robot Learning

As of 21 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2508.17230.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.17230 v2

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:03:21.467027Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T04:18:02.341742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:05:50.616851Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a635510-11fe-4aeb-9a56-f8e3070eb9b9 · outbound

This paper cites Diffusion-Based Representation Learning.

4D Visual Pre-training for Robot Learning Diffusion-Based Representation Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.017683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.017683Z digest=sha256:4fdf2d5cbadee44f74a2853e030a849356266996f3722a74cc485266db2cbf89

Observation 966afc96-74aa-45a7-8078-86a83e867bc3 · outbound

This paper cites Cobot magic: An open-source robotic system.https://global.agilex.ai/products/ cobot-magic, 2025.

4D Visual Pre-training for Robot Learning Cobot magic: An open-source robotic system.https://global.agilex.ai/products/ cobot-magic, 2025

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.338006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.139352Z digest=sha256:ace8c1c3c936c7219501085429128a4dd93da1a10092cd1cdaea689fb8b6dcb6

Observation eb778560-faca-4d28-99e0-00c7c75dabbd · outbound

This paper cites Is conditional gen- erative modeling all you need for decision-making?, 2023.

4D Visual Pre-training for Robot Learning Is conditional gen- erative modeling all you need for decision-making?, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.324751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.225443Z digest=sha256:f0a7edb5ef3b16848f9ff7863384f9a03b9e02942d10b54671eb43fb54842f56

Observation 342b70fd-7065-47be-b0f9-bb647dbf006c · outbound

This paper cites Diffusion Policy: Visuomotor Policy Learning via Action Diffusion.

4D Visual Pre-training for Robot Learning Diffusion Policy: Visuomotor Policy Learning via Action Diffusion

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.230022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.230022Z digest=sha256:6acfbcb3e89cb6fe1a4ff41c7903fbd9a1d17c557b1f1c1441570c962b3b3e58

Observation 080d14c6-c09b-47f0-a02b-db9d260702ed · outbound

This paper cites An unbiased look at datasets for visuo-motor pre-training.

4D Visual Pre-training for Robot Learning An unbiased look at datasets for visuo-motor pre-training

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.311444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.235078Z digest=sha256:59dfcb137bc28a4258d155044108ba4c28bb53fd44e1873ab0a9ee330ab20e0c

Observation a5c3e32a-eacd-4969-9d67-60d8dc9eb804 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

4D Visual Pre-training for Robot Learning Imagenet: A large-scale hierarchical image database

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.239383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.239383Z digest=sha256:81ad4111307264f402020ead9266c5ff3b05ea071a3346cb698ad76247e17306

Observation fe6f73e4-40b5-457f-b7bc-28a6fa23caca · outbound

This paper cites Implicit behavioral cloning, 2021.

4D Visual Pre-training for Robot Learning Implicit behavioral cloning, 2021

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.285073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.243345Z digest=sha256:da9513a1e1a362b83324aeedb3a82ce856ac3774b3fb2f6531aaf2629ccb358f

Observation c6797594-bb4b-411a-bd3a-acd7e4450313 · outbound

This paper cites Mobile aloha: Learning bimanual mobile manipulation using low-cost whole-body teleoperation.

4D Visual Pre-training for Robot Learning Mobile aloha: Learning bimanual mobile manipulation using low-cost whole-body teleoperation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.270882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.247675Z digest=sha256:c5f5f145516827936ee1ec6a16f1063efb1e1aeb680b5c3ef78f1fdbc807f6c9

Observation 46de06a7-c92b-41d2-a328-70e47d45189a · outbound

This paper cites Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation.

4D Visual Pre-training for Robot Learning Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.251510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.251510Z digest=sha256:2a3bbb4752a898984b29a911ef82b0e5092a67057937ff4d1f8ca1082025fd3a

Observation 2b504652-2d46-471b-9b87-7fbefc7645a0 · outbound

This paper cites Rvt: Robotic view transformer for 3d object manipulation.

4D Visual Pre-training for Robot Learning Rvt: Robotic view transformer for 3d object manipulation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.256034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.256034Z digest=sha256:1abce3825dccb3fe9127c2a438821f6feda39132278417e99dbf996cfc2f5725

Observation 1870c50b-f686-4f1c-9515-cb016b6942d0 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

4D Visual Pre-training for Robot Learning Ego4d: Around the world in 3,000 hours of egocentric video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.260015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.260015Z digest=sha256:32026ca6cf8e1176069a7fc2fcf389b7a9d6aa2a5fb07ef84c3d6e47af8754a4

Observation dd8634bf-ca7c-421b-a163-a6e0de0fb295 · outbound

This paper cites Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

4D Visual Pre-training for Robot Learning Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.263906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.263906Z digest=sha256:0a6587b53d7ce824fbfd5e15cb9e50f34a37f6c95b4d6cd0626f038592396aa7

Observation ccfc23a4-d579-4898-8f7e-a97b12171a8f · outbound

This paper cites Spatio-temporal self-supervised representation learning for 3d point clouds.

4D Visual Pre-training for Robot Learning Spatio-temporal self-supervised representation learning for 3d point clouds

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.231230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.267884Z digest=sha256:672e5796b0c3434192318949ed67a3e079aac8c7180b343a42fde46e6b432da0

Observation bd3e5533-a953-4d6f-8a46-80ed849f5f90 · outbound

This paper cites Diffusion Reward: Learning Rewards via Conditional Video Diffusion.

4D Visual Pre-training for Robot Learning Diffusion Reward: Learning Rewards via Conditional Video Diffusion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.271906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.271906Z digest=sha256:0d9c9a14f930ca5742afca5e93102f1d30c2271cad39fb8126b45cda82984b1a

Observation 6beb5b35-e348-4134-bc37-0d3edd306821 · outbound

This paper cites SODA: Bottleneck Diffusion Models for Representation Learning.

4D Visual Pre-training for Robot Learning SODA: Bottleneck Diffusion Models for Representation Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.276099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.276099Z digest=sha256:81912c1b34fcf2218e755f17a8799915fd090c39b402ca9a8118f78745d82039

Observation 2386e89f-27f0-4fd7-82d2-4ae15162be01 · outbound

This paper cites Equilibrium free-energy differences from nonequilibrium measurements: A master-equation ap- proach.Physical Review E, 56(5):5018, 1997.

4D Visual Pre-training for Robot Learning Equilibrium free-energy differences from nonequilibrium measurements: A master-equation ap- proach.Physical Review E, 56(5):5018, 1997

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.217786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.280367Z digest=sha256:da93ff9783a48346bad7a5e2e22a91320f02559cdd88c9925066ceb27e2e95c7

Observation e70ad54d-784e-41ca-9509-eec2899144b2 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

4D Visual Pre-training for Robot Learning RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.284052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.284052Z digest=sha256:cb1736de458bc69ce2e919d85a6b5b833e6716ece6ce2c918cc098b7deb3c667

Observation 3c498518-9db0-48d2-87c9-381072874f5e · outbound

This paper cites Point- voxel cnn for efficient 3d deep learning.Advances in neural information processing systems, 32, 2019.

4D Visual Pre-training for Robot Learning Point- voxel cnn for efficient 3d deep learning.Advances in neural information processing systems, 32, 2019

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.204285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.288110Z digest=sha256:4345f2e663246c052e2391c9c204ab57c9345ad0e8dc7b5af5238dfed63b8f59

Observation c2315448-5b1d-40e2-b8e8-bb1a31c695e6 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelli- gence?Advances in Neural Information Processing Systems, 36, 2024.

4D Visual Pre-training for Robot Learning Where are we in the search for an artificial visual cortex for embodied intelli- gence?Advances in Neural Information Processing Systems, 36, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.191842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.291729Z digest=sha256:8d787b6923a34fb3c46d2e4c43c24724cfe29e4540265286913cce6915edf27e

Observation 0f1c4598-5ea5-478c-865a-46dbad63b42e · outbound

This paper cites What mat- ters in learning from offline human demonstrations for robot manipulation, 2021.

4D Visual Pre-training for Robot Learning What mat- ters in learning from offline human demonstrations for robot manipulation, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.178483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.295581Z digest=sha256:86f2e407d42278fb6b2a52b81bac9248ec5396fb2d1b4c1b7d04423d0f36738d

Observation 15af83ac-f03b-4365-b84c-63b1c1091ae3 · outbound

This paper cites Self-supervised point cloud prediction using 3d spatio-temporal convolutional networks.

4D Visual Pre-training for Robot Learning Self-supervised point cloud prediction using 3d spatio-temporal convolutional networks

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.165293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.299132Z digest=sha256:839629a4845b78fed60b94f72e7c9fa3a5c1c51b71813161cb35c91236835ca0

Observation f22873ba-802a-4dab-8b1d-4e9684a4ecff · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

4D Visual Pre-training for Robot Learning R3M: A Universal Visual Representation for Robot Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.302768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.302768Z digest=sha256:0d5893bc78e3aaf8d8034eecaffadc066842c847d17f4ac3bf1e6631603e3335

Observation 0fe27fdf-95fd-4700-b6cf-bc3b8bad835c · outbound

This paper cites Henriques.

4D Visual Pre-training for Robot Learning Henriques

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.151794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.306611Z digest=sha256:bfa6c1473adc6b1b981b6b2b5804be5dc8ecb690ebe10b3a8574903093bf47d5

Observation 6b24ad63-b50d-469b-a89c-1e25287e20ed · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

4D Visual Pre-training for Robot Learning Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.310338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.310338Z digest=sha256:5b214b47687d3f8d726526919516031b9041f2c61bdfc6deb811aba4af87faf2

Observation a95946da-2f04-4e23-be52-f1fb9e04c6aa · outbound

This paper cites Masked autoencoders for point cloud self-supervised learning.

4D Visual Pre-training for Robot Learning Masked autoencoders for point cloud self-supervised learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.138718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.314258Z digest=sha256:f77789cfb9c673eee588589dfb3cd7129d9095f4eb11359b7405a73ad68efec7

Observation ea8bc1c0-84f5-4536-b868-67f7ae53481d · outbound

This paper cites Recon- structing hands in 3d with transformers.

4D Visual Pre-training for Robot Learning Recon- structing hands in 3d with transformers

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.124476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.317885Z digest=sha256:4e8f496405af38d3615259bf006751d050a5c6e414181996328ebae2de48ac95

Observation 873be75f-5acb-4648-bb70-3c580c09c8d2 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30, 2017.

4D Visual Pre-training for Robot Learning Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30, 2017

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.110432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.321493Z digest=sha256:189b77c8cdd43886a684e0433fca664e17df14481b9168df7310940ce2ebdb6d

Observation 8df40319-180d-46a0-85e5-20a73191b6b3 · outbound

This paper cites Dexmv: Imita- tion learning for dexterous manipulation from human videos,.

4D Visual Pre-training for Robot Learning Dexmv: Imita- tion learning for dexterous manipulation from human videos,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.096738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.325644Z digest=sha256:7b82f5dae0cee18929a92c1728dc538f9978bf68326dcf87a436a392fae01b7a

Observation 587b6cd7-41b6-42ad-b69b-bbbaf3b0cb7b · outbound

This paper cites AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System.

4D Visual Pre-training for Robot Learning AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.329618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.329618Z digest=sha256:4061f2011916ba10e01cbb0cebfed5f3a9821a782d14ac0f31e26b685ce68ef5

Observation 6ff6bb74-1dcb-46bc-b59d-99192dcf5c15 · outbound

This paper cites Robot learning with sen- sorimotor pre-training.

4D Visual Pre-training for Robot Learning Robot learning with sen- sorimotor pre-training

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.083894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.333713Z digest=sha256:f45737de67b00a23c6b2c1f79ab8ebfb588c608597b29971f8d8c3cd6215ce78

Observation 01aea381-36a0-446c-ae4a-6c80dead3cb8 · outbound

This paper cites Real-world robot learn- ing with masked visual pre-training.

4D Visual Pre-training for Robot Learning Real-world robot learn- ing with masked visual pre-training

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.070451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.337505Z digest=sha256:507aa011d8adac6e458d969c3d01b0f47191bf5fe49557c98ce9187ed0e0092e

Observation 7819deb8-c349-417d-aa66-fb36aa512f19 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

4D Visual Pre-training for Robot Learning Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.341137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.341137Z digest=sha256:bd3770225541ad5fb8ebc6a08664b83c216990f2b774cdf9965d8699ed8b5aa1

Observation 7219eb5c-e32d-48ec-b163-99877d0db233 · outbound

This paper cites Learning complex dexterous manipulation with deep reinforcement learning and demonstrations, 2018.

4D Visual Pre-training for Robot Learning Learning complex dexterous manipulation with deep reinforcement learning and demonstrations, 2018

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.057737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.345166Z digest=sha256:d29760702fc823c49fb5c692c8c6b4f31da39db0e58493f9837c8d2a4f9e6098

Observation 83e69080-267e-4218-a39e-0ae7159eb3d6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

4D Visual Pre-training for Robot Learning High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.349065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.349065Z digest=sha256:b4385c1608ee75fffb313ca7ca70086428754d42c7d0d40a313d43d2f2633801

Observation 6d495e07-46de-4594-870b-be071a7cf748 · outbound

This paper cites On bringing robots home, 2023.

4D Visual Pre-training for Robot Learning On bringing robots home, 2023

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.035067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.352885Z digest=sha256:695d9293de5c880427c94493067af4378ba79e29cd339454f1b04c1ad0805d4f

Observation cf8d3867-1c0e-4240-9737-d6ee8115d10d · outbound

This paper cites RRL: Resnet as representation for Reinforcement Learning.

4D Visual Pre-training for Robot Learning RRL: Resnet as representation for Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.357019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.357019Z digest=sha256:25af6281f66b88749d283426d5098f98d7b44844da2a11823893ba3fc3ab3e5c

Observation 95454da0-576b-4b4e-b3a7-4e83b7279563 · outbound

This paper cites Perceiver- actor: A multi-task transformer for robotic manipulation.

4D Visual Pre-training for Robot Learning Perceiver- actor: A multi-task transformer for robotic manipulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.361117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.361117Z digest=sha256:85c905e9f260e6d29b30fa88985e5e38c07a06e02f6c34c2bdae80b3db2bd0db

Observation 49078f5f-c4b8-47b6-a274-86a6d7c86745 · outbound

This paper cites Shelving, stacking, hanging: Relational pose diffusion for multi-modal rearrangement, 2023.

4D Visual Pre-training for Robot Learning Shelving, stacking, hanging: Relational pose diffusion for multi-modal rearrangement, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:22.009913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.364875Z digest=sha256:0d6a12cfaf6e51ca4c2c9393c4dc0871da5fba5055534afbe716fbb9a6454966

Observation 9f46b6ad-ee52-44f0-80a0-87d9a34d2370 · outbound

This paper cites Denoising Diffusion Implicit Models.

4D Visual Pre-training for Robot Learning Denoising Diffusion Implicit Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.368501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.368501Z digest=sha256:e89cc33d97ecede4311acddb7c2720a681f4513ba3be6408ab81562db7911275

Observation 96721581-04c8-4a0b-a2f2-303837a39311 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

4D Visual Pre-training for Robot Learning Score-Based Generative Modeling through Stochastic Differential Equations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.372448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.372448Z digest=sha256:1e8c279d1188591637c8427f00245178fe2e7324521ee1e17171bdfc919974bd

Observation 1444e2da-ffe5-44fb-9b2e-0c47c5c6a110 · outbound

This paper cites Memory-consistent neural networks for imitation learning, 2024.

4D Visual Pre-training for Robot Learning Memory-consistent neural networks for imitation learning, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.995667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.376270Z digest=sha256:fe5fa30a60a6595b20ffe3033007e435c717a10a3d6f68cfed78c2b0d61353b9

Observation bc7977ff-7c9b-41fa-9420-907c6db02869 · outbound

This paper cites Mujoco: A physics engine for model-based control.

4D Visual Pre-training for Robot Learning Mujoco: A physics engine for model-based control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.981393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.380078Z digest=sha256:99fb128f84d6e64b4bec40a390035f7132248771f67579eea4b3e62416bd16c8

Observation 585ba0d6-7156-49ef-b8b0-cab3c9a8a156 · outbound

This paper cites Se(3)-diffusionfields: Learning smooth cost func- tions for joint grasp and motion optimization through diffu- sion, 2023.

4D Visual Pre-training for Robot Learning Se(3)-diffusionfields: Learning smooth cost func- tions for joint grasp and motion optimization through diffu- sion, 2023

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.967856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.384022Z digest=sha256:c0cfd0e7d747a4d671239310768308901ec8b6dd2b7b3a81ea15031f08746c8d

Observation 036d5778-04d8-49ac-bae7-4f680b674757 · outbound

This paper cites RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective.

4D Visual Pre-training for Robot Learning RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.388478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.388478Z digest=sha256:138cfa13d911b316c0a294509b572a2828a3633e8ea0ba8f83fea649cf5c11cc

Observation a11434cf-28ad-42a3-a12e-98d85db0e938 · outbound

This paper cites Equivariant Diffusion Policy.

4D Visual Pre-training for Robot Learning Equivariant Diffusion Policy

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.392455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.392455Z digest=sha256:4e9940de20dc5883393327d983619d5906328a9ef6fd73bf9c1d61f3d4d63e2f

Observation 3814d33a-265c-484a-948c-05b98b483286 · outbound

This paper cites Diffusion models as masked autoencoders.

4D Visual Pre-training for Robot Learning Diffusion models as masked autoencoders

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.954631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.396448Z digest=sha256:2370388a6edd2d819113fb82ceabc16b796227eb689c70205b1aff5bef9db80f

Observation b26cf4c9-f623-4209-9d3e-215b8295e897 · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

4D Visual Pre-training for Robot Learning RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.400105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.400105Z digest=sha256:4cb42713204bdba490c1c8376c347f3a10254e338f06ff0fb2f5a372d3c2bd3f

Observation f979dbea-2f6b-4fb0-a28e-5b27753ca864 · outbound

This paper cites Tiangong.https://x-humanoid.com/ bt.html, 2025.

4D Visual Pre-training for Robot Learning Tiangong.https://x-humanoid.com/ bt.html, 2025

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.941597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.404049Z digest=sha256:29ef361806067f95d2d57870751abe72a5fb2d15fec3611739e81bef3f4d66b9

Observation 81782620-d8d6-45c8-8234-d84116b6eef8 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

4D Visual Pre-training for Robot Learning Masked Visual Pre-training for Motor Control

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.407790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.407790Z digest=sha256:d75eb47290704aefe759b6e9299be044e55838436dff11293c1261ff0598ec34

Observation d885646e-f68e-477c-8b69-0b544d592100 · outbound

This paper cites NeRFuser: Dif- fusion guided multi-task 3d policy learning, 2024.

4D Visual Pre-training for Robot Learning NeRFuser: Dif- fusion guided multi-task 3d policy learning, 2024

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.927021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.411478Z digest=sha256:c2d8a0ca0abb01180d8d51f7caa70e96a2173da3343c09452d85b69eaea446cd

Observation 2e69c9c8-8ab2-4008-977c-30668a545c17 · outbound

This paper cites EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning.

4D Visual Pre-training for Robot Learning EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.415481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.415481Z digest=sha256:112f448798d52f0765de69de87ba54214c548e8815f784d81c3ed5315b7676c6

Observation 9b44c538-5c30-499b-bfec-bc31f1d38e88 · outbound

This paper cites Visual point cloud forecasting enables scalable autonomous driving.

4D Visual Pre-training for Robot Learning Visual point cloud forecasting enables scalable autonomous driving

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.419705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.419705Z digest=sha256:742b2e0874f8ac23317031e14eec1143ac948a15a71f32d9fae08382b1079a2d

Observation d0832534-c200-4442-8c82-819a061c9a55 · outbound

This paper cites Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning.

4D Visual Pre-training for Robot Learning Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.904281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.423343Z digest=sha256:3ec0e0712ba923745f8e5fe7889ceb2a3b86d2ee90f8c55574633d6ebb15d08b

Observation 417a6389-f1d0-460f-a1d6-f24e13ea1190 · outbound

This paper cites Visual reinforcement learning with self- supervised 3d representations.IEEE Robotics and Automa- tion Letters, 8(5):2890–2897, 2023.

4D Visual Pre-training for Robot Learning Visual reinforcement learning with self- supervised 3d representations.IEEE Robotics and Automa- tion Letters, 8(5):2890–2897, 2023

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.890339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.427088Z digest=sha256:0c3ac3368c03c2e3d2624cee7982f1186d895f4d93fa3b427c1420c9f83dd011

Observation 7ea36b4e-021c-4441-a4fb-60df61569095 · outbound

This paper cites Gnfactor: Multi-task real robot learning with generalizable neural feature fields.

4D Visual Pre-training for Robot Learning Gnfactor: Multi-task real robot learning with generalizable neural feature fields

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.430762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.430762Z digest=sha256:4ffcb0f730bb96c173e0a959cdf41e9f50d9fb60291269d345289a1e6356e1c2

Observation 1b55d1fb-3853-4c82-b410-47d256aa4e45 · outbound

This paper cites Generalizable Humanoid Manipulation with 3D Diffusion Policies.

4D Visual Pre-training for Robot Learning Generalizable Humanoid Manipulation with 3D Diffusion Policies

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.434778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.434778Z digest=sha256:dd7c7f47e2e24d1d78f183dd28392542da01dca34ec64d67a94eb894c7346ec4

Observation 4771448d-7443-4c28-ac4e-6864a87f1567 · outbound

This paper cites 3d diffusion policy: Gen- eralizable visuomotor policy learning via simple 3d repre- sentations.

4D Visual Pre-training for Robot Learning 3d diffusion policy: Gen- eralizable visuomotor policy learning via simple 3d repre- sentations

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.867924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.439233Z digest=sha256:fc9c0f875e3f82acca28e6189892fb8381eef0e507d607433561be55c82586f2

Observation 4260ea29-7672-4d90-8b1c-6304b1696c41 · outbound

This paper cites Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing sys- tems, 35:27061–27074, 2022.

4D Visual Pre-training for Robot Learning Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing sys- tems, 35:27061–27074, 2022

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.442925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.442925Z digest=sha256:0e97721adb2a0d0f98e34022bb0b7fc38ee79c6424da4fd843c1b87b4535e2cb

Observation 2d85b8e6-4ffb-4fe8-823f-1a8e01e82eb8 · outbound

This paper cites VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks.

4D Visual Pre-training for Robot Learning VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.446755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.446755Z digest=sha256:002a58da495c15966fbd780f137e7c43d56ad47ce1e7b841ba6eb9c30a21357c

Observation f6840b21-1ffe-46aa-8b0f-5fd9e948e0b5 · outbound

This paper cites Complete-to-partial 4d distillation for self-supervised point cloud sequence representation learning.

4D Visual Pre-training for Robot Learning Complete-to-partial 4d distillation for self-supervised point cloud sequence representation learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.846631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.451231Z digest=sha256:4843a68524beb4e1ec8a9ce8a2cbcda8303e60642ba7683976dfae73894ae4de

Observation b6a15598-3cd3-4011-9959-ac5f20d1bb43 · outbound

This paper cites Point transformer.

4D Visual Pre-training for Robot Learning Point transformer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.833350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.454842Z digest=sha256:7d0950afddcd66f4c659599b55ef74e3d2e6012726cacc8285dd7c66f8ad522e

Observation e334a8fc-fadb-4334-90f8-de9692eb52d2 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

4D Visual Pre-training for Robot Learning Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T17:03:21.458999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:03:21.458999Z digest=sha256:217279bd8fa5c87149e06f10de9ff5e39a1ea55d6a82de855091449017ad3de3

Observation 07fd99e0-d749-4779-88ba-ef697ffc9aea · outbound

This paper cites Point Cloud Pre-training with Diffusion Models.

4D Visual Pre-training for Robot Learning Point Cloud Pre-training with Diffusion Models

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:03:21.509768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.462860Z digest=sha256:8f4d05d4a696a893610716c7dcb5830e6c32d7abee345c943b79a9006f800ba9

Observation 6d5802aa-b5b4-439d-9860-80114dee874f · outbound

This paper cites 3d shape generation and completion through point-voxel diffusion.

4D Visual Pre-training for Robot Learning 3d shape generation and completion through point-voxel diffusion

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:03:21.820003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T17:03:21.467027Z digest=sha256:994a1ffb67bac02e3ad6c246df872bd31db23bf80a2d0d8fc53c9883eee25c82

Pith citing papers

Observation 5a7e9431-ae2c-4ab2-8f13-ae8a4b1ad604 · inbound

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration cites this paper.

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration 4D Visual Pre-training for Robot Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:05:50.618485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T04:18:02.341742Z digest=sha256:b3a4c9187186f20ef3e45cc58f9e4f4c023ad775480ef93cbf5b6f4d646e6f12