Pith. sign in

Paper Citation Record · LEDGER

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

As of 16 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2604.17887.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.17887 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T04:41:45.903419Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T08:04:48.963890Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact42
  • verified fuzzy13
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fa88f65d-e5da-4722-a0db-200d273d0e41 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:65c191187641db7013ad6f86973d91d32996c9fc6f4abe9b60d2ea5023041d2e

Observation 9e4f991c-eeba-48da-ad34-4bff349b8cd4 · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:47:10.348602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:64e92f883c5ea18241a4f91d01cb89743ff8707d58e7fa71da72ae12c8a3863b

Observation 73c73cbf-a0dc-43f8-92b5-adff13846f15 · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.877393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:8d055320f47be4581da1cfdd67a14a763f1fff49a109c6dd7a08d82aab51aac6

Observation 9ed6ea99-5520-4157-b64b-6cbb56d431dd · outbound

This paper cites Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:17:01.597427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:280f89eebc9d4e4f3991b835aa3d849484dbf50e0b3dd8ac0093eb2bba33622b

Observation fc8bdcd0-3449-41b2-8cb8-fee98bfacc32 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:54:59.300665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:cb38b6c9f84c90d8bc31639a17bc3be842ef24969bbe14ba8fb3342ebd4c5949

Observation 821736e2-2818-4591-8e0d-b9bfc7c4210f · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:38:24.665044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:01d16e940a7b6497fbc34acb28f52be68a654d26720ddd2c4f84db1a09805ad2

Observation 3a804e4c-5014-49da-8114-0b0d8ef5fc33 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:58:52.491420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:092abffa1ad02b4a8dd95b735418d9965328f08901a56c174b337c43e8b69503

Observation 3dfb972a-aeee-4f43-a7ec-a1976ba90d32 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Align your latents: High-resolution video synthesis with latent diffusion models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.406758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:2f4a68ebb4c22abc5e66edbec11f8bf6e4f022ec9fb3ae9bad8d040dd5edc1fc

Observation dcc8ed7e-2b1b-468d-8a2b-0064fc5af883 · outbound

This paper cites Genie: Generative interactive environments.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Genie: Generative interactive environments

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.415849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:52f68a951f1aae062efd3e84a095d404249fe90fb3789a4105ceac5052b85b81

Observation 3fcd4e95-7c82-4814-b484-9a353b5db8be · outbound

This paper cites Closed-loop visuomotor control with generative expectation for robotic manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Closed-loop visuomotor control with generative expectation for robotic manipulation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.403187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:93a5f1894a65bdcb332de88738ff7a62d700bbb3d05f46e2499a0f593f737d1c

Observation d37c29b2-f0d6-4f21-bfc8-effd2f864b41 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:09:24.919963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:282900b3cedea132d83cb17ebd91af378801f90cdab7e39e8fbc18ab68470953

Observation 0c50f298-35b7-4fca-bdd2-a363c430bc6a · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:09:34.215368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:98014aecdf3f465ea74a34a3ef6c1a9ab551844d1cb25ce9263e0f63d230860b

Observation d4e3056f-8206-4972-808e-c140606e80bf · outbound

This paper cites GR-3 Technical Report.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement GR-3 Technical Report

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:04:12.844097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:0dbd425fbd4691f20a7073acb05b6e99d3da79203ff789ecfcf13f6ebb26048d

Observation a5c12e48-6202-46dd-a236-2f166bbfd2ca · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:40:27.997500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:d254dcb71402dda654e559e340be6d0757db2c892e7a723d882d1c6be734e9d7

Observation abb8c667-2a6b-4cdd-a281-00ff4a257105 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.The International Journal of Robotics Research.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Diffusion policy: Visuomotor policy learning via action diffusion.The International Journal of Robotics Research

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.412071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:45629ad651610b6fe9674cbc3e831429c88f0b5c91b5936af32220f1ec7864fa

Observation c265f48e-4422-4ac2-96b7-e11db43ab1be · outbound

This paper cites Wow: Towards a world omniscient world model through embodied interaction.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Wow: Towards a world omniscient world model through embodied interaction

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.836000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:a054a93fce80bf2b3733ce6ce800373b2073a7b96631e31ca3d4e62ff75d4140

Observation cbaf94d8-4ba7-4fa9-bf82-34a2b90eebd7 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.770851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:81b8cda1ed92fa89672855700cb53ccf4134f9fac36dd13190416d4d14872273

Observation 71dc6d70-daab-48cd-bc12-d04b9b5ce411 · outbound

This paper cites Learning universal policies via text-guided video generation.Advances in neural information processing systems.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Learning universal policies via text-guided video generation.Advances in neural information processing systems

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.419408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:315f550a20a4226b3ec7f20caa8818c7686129188ceaa6abb87cc99513f0a9b5

Observation cc25a280-e697-4226-9d46-5393a83d8e22 · outbound

This paper cites Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:55:55.512010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:1cd8daeae7ecce7ce75019304ae81d7a2fa6e160485cc12fd1cd2b7e7fe4a51a

Observation 9dd91235-30d7-431b-8aeb-7a1fe8212171 · outbound

This paper cites Vidar: Embodied Video Diffusion Model for Generalist Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Vidar: Embodied Video Diffusion Model for Generalist Manipulation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:54:28.431425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:5b3e94d132177c5523a4aa40068ed42c7f3c559176800d855a007e76b21fb911

Observation 2f4847b4-1252-48dd-a8a0-acc37bd80b41 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:02:55.841967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:3fcd4a98f757df245c4c70a272e7b44fe3e04b02cb2287fd9d97e4a72ed96889

Observation 6d3869e6-7723-4117-9a94-35a574c16a6d · outbound

This paper cites Deep residual learning for image recognition.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Deep residual learning for image recognition

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.437563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:4cf390c64a72ebdda581b84143e03ffbe48139d3ce062cd93f7e1e9cc9ec5dff

Observation c66e1695-5600-4905-b553-2f80defd84d2 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:24:30.435137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:dcdeccb469b2ef4e1bcb7bc70457e1cc9af68fe907e8f033a70261ae5843d11d

Observation 54155f2f-18f4-4ecc-a2bd-cdf4518d2692 · outbound

This paper cites EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:40:29.471447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:37987e5d985f94bbfd9a485ec1c0406f0c8c6c6d352265dfc7eceda5a7e463b2

Observation 8877372a-be52-408e-9426-4921e0447b0f · outbound

This paper cites Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:38:11.644840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:2cb903a01183c914830bedadac4c633cdfeb4b3ef2da432aec1732ff60057eff

Observation 50d2bea1-49e5-401a-8121-2928f7995180 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.908259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:ddcefed55c78174e41244c0d2af9d1744f1fc39f741447a02616a357e79180fc

Observation 8c29e990-54d7-44e0-a5ae-39dd08c7fd69 · outbound

This paper cites Dreamgen: Unlocking generalization in robot learning through neural trajectories.arXiv e-prints.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Dreamgen: Unlocking generalization in robot learning through neural trajectories.arXiv e-prints

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.403495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:db424b0c02fd847af3d8e212a968bfb5f0ab0646cc57cf7f7a74efd011942533

Observation 9c720f23-7ab7-411a-8c73-cd934f44f14f · outbound

This paper cites Galaxea Open-World Dataset and G0 Dual-System VLA Model.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Galaxea Open-World Dataset and G0 Dual-System VLA Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.724021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:5f84209ba97b4ff14b37e72dc5279cfe41aadac36aca5f30e3f1d139a60acb88

Observation d1840719-130b-42b4-85fc-879af2671e77 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement OpenVLA: An Open-Source Vision-Language-Action Model

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:37.168344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:7722ff9eba762c866c41b2f9c84c5c09d1c0219a32d7f16d58afa0203324e667

Observation 0d7f32bd-9e26-4d04-8615-7ce3ba22ee09 · outbound

This paper cites Segment anything.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Segment anything

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.434016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:9f276fd65389b8ada2d6d6790266722be71bf7341e1c15f9b018a57a51a14f78

Observation 2bbe9494-217d-4a7e-b8c0-9d71f8916cef · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.934902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:54a43beb0a1f3ae6f3ac7a3f3841b14d718af39a68f70c3a8e43f1930c7daaec

Observation 19cb5f75-7462-4198-9e0a-d3da92de4bac · outbound

This paper cites Unified Video Action Model.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Unified Video Action Model

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:50:29.857677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:119be965e0ef93eb1c682a78dd23a1c3a5cf07400d6d9045785a93434d2a0fd3

Observation df326d75-1924-406a-82a2-369ce6fab3b4 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:46:30.715521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:5b59c2614d3922a8056a6b396499d08e32a53797d273609e9cbd48dbf22a2fab

Observation 1b0f78e0-49a2-4e82-8f6d-0d856bd5a075 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:43:11.511714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:c2c6df211399a522361dc4a6f0002d5ecddb2030a54638b84d58cb95ba2bbf63

Observation 5696192e-7a96-434b-9377-1760ccbfde1e · outbound

This paper cites Octo: An open-source generalist robot policy.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Octo: An open-source generalist robot policy

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.430203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:eefe315a3fb993bcc52a4b7c67b1a1622c90b5a4193fe59e67d52ddcce53b8fc

Observation 2d399f3e-75ff-4a5e-980c-7a1d4fa34f26 · outbound

This paper cites Robotwin: Dual-arm robot benchmark with generative digital twins (early version).

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Robotwin: Dual-arm robot benchmark with generative digital twins (early version)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.399898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:b2b475cb36f06a17c7d2386bdc86bf6938a625716fd3241f75fda78a781741ac

Observation 82778423-18b2-43aa-a972-dec0c26bb939 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement DINOv2: Learning Robust Visual Features without Supervision

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.754249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:2e20b096ed5ff839239ded1f829e87933c9bae8cb242f0f60bdc247e380ca637

Observation 7a487226-a11b-4683-970e-186fbbca8a4c · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.408052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:664f395960d263ff5c84482774b42c098da9e94f54803ef6ca2e881f3df6be68

Observation c7199917-fcdf-46e4-a5e2-1629a42e72cb · outbound

This paper cites Diffusion Policy Policy Optimization.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Diffusion Policy Policy Optimization

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:48:15.161771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:7fb7a90d3b3ee0d0fc7244d551c89d1b0abf5a49641f50e9a0dfcb5a06b5a086

Observation e701aafc-3872-420f-be51-75fc843c0727 · outbound

This paper cites Manibox: Enhancing spatial grasping generalization via scalable simulation data generation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Manibox: Enhancing spatial grasping generalization via scalable simulation data generation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.830360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:6990a81f20f7024e156f444beaebfa55fc20990fc96c68cddd227e5df04ac431

Observation 08998509-6321-4fef-912f-16272aa421ed · outbound

This paper cites AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.903288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:1338f6a39ed62cd2348883eab721e27c88361fe2d09e390ce071bcaecf06acaa

Observation c705e43a-64bb-4a5a-ad4a-01d91fa8760f · outbound

This paper cites Gigabrain-0: A world model-powered vision-language- action model.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Gigabrain-0: A world model-powered vision-language- action model

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.924219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:fb056a571a8aa6945c5ff92fbae6275af6fa92845290e18d34d43d7a17278142

Observation 040efece-7673-4417-a2ce-f59eadbcdfc0 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:38:25.165602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:2299f9a476358ddf516359c2bf3ea958631965411071e7bbd4546ec5b2addaec

Observation 8da45d2d-cf39-4407-83b1-24ffa6bcfae0 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Attention is all you need.Advances in neural information processing systems

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.423158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:bffa625a4906e50b34fcaf68f5071398cb1feaf83f83df441ea1c3c41caa14d8

Observation ac53f542-d3da-43b8-a37a-640cf0c5d62c · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Wan: Open and Advanced Large-Scale Video Generative Models

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:05:22.803097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:e43ec40c662416fef547da00d13b4c7384d97d17214c9bd103cbedd0b5a9ce06

Observation 8177bee3-1b04-4aef-9ba4-a6f8f1fc1b32 · outbound

This paper cites Vidman: Exploiting implicit dynamics from video diffusion model for effective robot manipulation.Advances in Neural Information Processing Systems.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Vidman: Exploiting implicit dynamics from video diffusion model for effective robot manipulation.Advances in Neural Information Processing Systems

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T02:44:33.426852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:aefc328de8c9d7b7b668b5ce2cf9c8ee8f868758c253fd74b81e4bc872df4280

Observation 86bf127f-5215-4149-b398-a220d3935e29 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:32:05.974462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:bf2d12992e882ed3d8314b36dc54073a30ea58a64489ee39cc52721125c6c26d

Observation a6c79d11-ffbc-409a-a8ca-b92b45b37a97 · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:14:18.432154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:b4f50336d3fd7163e8b018a1d270cb4f751b9ae549c38f8bd437840037ba2ab5

Observation 3e63988f-31b2-4368-9fd6-bd08e65e0be3 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:26:22.534732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:d588245a6d5817443ed542da22c05fa2858d87d080c4f13d146118e0d9973405

Observation 674ae899-f34f-4ce6-9d4a-fe1c6d310a06 · outbound

This paper cites RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.919554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:a022afaca34854937fef176e85b9a2dc402a041229493e39a4eb5a4b5fed6ee9

Observation 21cbcb39-2689-493d-ac96-e1f57d587f4a · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:49:37.451338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:ff98934b32952d6c406fcef14a41fbfe32dd0239c6657afb998fb71dc9e10e49

Observation dfaf73e6-78d8-4c43-8c32-0967cecd0041 · outbound

This paper cites Igniting vlms toward the embodied space.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Igniting vlms toward the embodied space

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.808094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:5279e52b519e97c6346ddec64ea5dac33146c22bb6a1bc235951370e8bf19e53

Observation 277e00de-e952-4092-b036-ada21e968f83 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:36.384383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:360ae5230122cf9a1c43705756aebc66b1778a69b7d86a651a748ba606cea2b2

Observation aab213db-3f04-42d6-a506-b59277502132 · outbound

This paper cites ALOHA Unleashed: A Simple Recipe for Robot Dexterity.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement ALOHA Unleashed: A Simple Recipe for Robot Dexterity

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.819178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:dcb3d1bac5f48b71f8bd40de4533c30495594d1297e56c72caa8d31ef1e32186

Observation 8fd15fd1-09df-448c-97d9-70c53ddd0368 · outbound

This paper cites RoboDreamer: Learning Compositional World Models for Robot Imagination.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement RoboDreamer: Learning Compositional World Models for Robot Imagination

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:47:30.400020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:44b2989f162248440e91e74705e9fceeae21a95671b5a9a827e0efb207bfa674

Pith citing papers

Observation 02f3c703-9065-450e-977f-66ffcecae504 · inbound

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation cites this paper.

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-12T08:04:48.963890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:04:48.963890Z digest=sha256:7db872a6f144eb99d6580ca49c412e01be83fa1008c4709066897ac3fc6b77f0

Observation 5d2855cf-0c46-4967-8be0-ed314527d0a9 · inbound

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training cites this paper.

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T06:41:57.276146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:41:57.276146Z digest=sha256:0eb89adcfa69ccfeaa66d0cfae06e1229dc4f1a5a3e18a20b4404165ea677a0b