Pith. sign in

Paper Citation Record · LEDGER

Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2503.14492.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.14492 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:35:01.921193Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:09:35.034771Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 893384aa-73d2-42de-98cc-9ac124a31722 · inbound

Generative Physical AI in Vision: A Survey cites this paper.

Generative Physical AI in Vision: A Survey Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 192

Resolution
unresolved
no resolver link, observed 2026-08-10T18:53:00.815168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:53:00.815168Z digest=sha256:8f35e87b6ed6ca90b8e3d5b82d8b43b4afd69066eb693fbf08bfc1e5bc7ee965

Observation 5e282bed-5e98-430f-8d20-22b08266d0e8 · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:50:45.507330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:5c0474c3f9b9bc36b7c2fdc232843a0cc014dd00c1538ad2fa7691ab4b72676c

Observation 45a7ca65-cb9b-4658-a9d7-987c00a473dd · inbound

Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen cites this paper.

Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 125

Resolution
unresolved
no resolver link, observed 2026-08-07T13:28:32.774748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:28:32.774748Z digest=sha256:d2892d2fa02af881788e45e65afdae69714fcf1ff341c48cea055be2b808df9b

Observation 6310431a-f881-4d67-a5a2-7c6e51ccaa36 · inbound

A Survey: Learning Embodied Intelligence from Physical Simulators and World Models cites this paper.

A Survey: Learning Embodied Intelligence from Physical Simulators and World Models Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 295

Resolution
unresolved
no resolver link, observed 2026-08-06T21:09:18.911635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:09:18.911635Z digest=sha256:52d89bbf3fb4b49313b200686b61fa29cd2586983af6b58965f0dcbcba4fc405

Observation 2da991c7-4eda-40fe-8528-f2e42769e2ed · inbound

LLM-based Realistic Safety-Critical Driving Video Generation cites this paper.

LLM-based Realistic Safety-Critical Driving Video Generation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T07:17:08.402206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T07:13:57.747558Z digest=sha256:c81aba78a75110ad439db28bdbc5e656070e0ee32f7b7b2a7c6389c006af063e

Observation f5ad7f50-a70a-4fb8-8496-ec4a2b63727e · inbound

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels cites this paper.

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:26:02.032190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:26:02.032190Z digest=sha256:81cc2d0bf7c7b1318a8891bc040db5a51e5a4f67066634097a25954386a6d87a

Observation 3ba50032-e5d0-456e-979a-531c29d79df5 · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.660028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:5edc6e71ae060a6b4f946c5085268eb9574f9fcd9c329a0e5ab28a49174e7865

Observation ba600a2d-7b05-4c19-865c-c04250791d22 · inbound

Non-invasive Assessment of Pancreatic Duct Hypertension Using Computational Flow Modeling cites this paper.

Non-invasive Assessment of Pancreatic Duct Hypertension Using Computational Flow Modeling Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T18:05:42.520975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:05:42.520975Z digest=sha256:18b3c50db5d95b0f23aca3a20c8079bd3286e58ac3f005b2f202eab6e03b62bc

Observation 5b8e6b00-7cd2-48f1-8537-3a7cf435e0a7 · inbound

3D and 4D World Modeling: A Survey cites this paper.

3D and 4D World Modeling: A Survey Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T06:04:03.292717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:04:03.292717Z digest=sha256:9b8350528843e754735ee521c386481ac72127d849764cfe530abbf86e0ed5d0

Observation ad860091-6241-488a-adf2-3e0c71ae3199 · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:32.995933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:32.995933Z digest=sha256:890bf2c7f816830755015929fd453af7abb41edf5c932070e249c0b8b89b5d17

Observation f344ecf1-a33d-49c0-8634-5443f93647f8 · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.712342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:287e7128cde782de5d3c9822af785b2fe73def82a1ed1dde776221bc6386ace4

Observation 413fc347-8f25-4d88-b12e-6edb1181e763 · inbound

SimScale: Learning to Drive via Real-World Simulation at Scale cites this paper.

SimScale: Learning to Drive via Real-World Simulation at Scale Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:34:01.784781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T04:33:03.629533Z digest=sha256:3e1bdce7aab2ea69b6409c39dad2eaba08a4d4bae70b35af6132250e5b1dc573

Observation 71b2a9e8-8232-40ee-af7a-485e338f3e93 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:31:22.067393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:cd76bd6ab0bf853d270268767c4c17dd6300c108b9d2103887e76c5c97fdc2bb

Observation 3dccb086-3cfa-411f-b109-557b3f4459f1 · inbound

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis cites this paper.

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T16:49:01.716396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:49:01.716396Z digest=sha256:81c10ac3b7a5ebab323b87653210da7c82be5e6866835130bb4aa7e779ffc72c

Observation 6fb857da-9b42-483e-9399-8a3d8b802345 · inbound

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation cites this paper.

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T13:24:42.566290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:24:42.566290Z digest=sha256:f14eec4f034f07bd85ff7429b48dda850c0a4c2015e329fd5bbb33e6fc7f4277

Observation 0272825e-e743-4e18-a674-283a3ff737fa · inbound

HyPER-GAN: Hybrid Patch-Based Image-to-Image Translation for Real-Time Photorealism Enhancement in Game Engines cites this paper.

HyPER-GAN: Hybrid Patch-Based Image-to-Image Translation for Real-Time Photorealism Enhancement in Game Engines Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T23:29:30.577918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:29:30.577918Z digest=sha256:363a5217af52c023adde64556b424a0135f3d025368321d647ec2ac7fba0e691

Observation 4a4a10b6-253f-42e4-8f14-bd9ddd52bf8b · inbound

InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model cites this paper.

InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:09:59.863054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T12:07:17.341611Z digest=sha256:8eeb0c7a5f883504e31a6336ea2eb55b2b73e99189aef88b94394ef4c701e658

Observation ae7c6457-8b68-451b-ba46-09e1bab78836 · inbound

From Virtual Environments to Real-World Trials: Emerging Trends in Autonomous Driving cites this paper.

From Virtual Environments to Real-World Trials: Emerging Trends in Autonomous Driving Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:05:20.323081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T09:02:21.114658Z digest=sha256:581c438cd276aa1d03db74d7c833a480ef9d204afbc4ce4a861f3c6d6519a875

Observation 3f0f9fb4-ec28-4450-9ba4-9227d43d7231 · inbound

HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes cites this paper.

HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:50.102371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T19:39:31.357223Z digest=sha256:f4a0a6d8be4e49f4c3cfddfe1d4886a8e6422a285ed170ff189c018bf99f04d6

Observation 86f88068-b48d-4c81-b1b6-7cfe6e2af4ab · inbound

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction cites this paper.

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:45:58.697450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:30:53.578491Z digest=sha256:8dc1260faa3dccd6ad9358f44d006add8f6691208aacb537e1804551bb881e3b

Observation 7789962c-23b6-48a7-9e69-19b41cc8b409 · inbound

ShapeGen: Robotic Data Generation for Category-Level Manipulation cites this paper.

ShapeGen: Robotic Data Generation for Category-Level Manipulation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:19:20.659201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T10:17:05.939312Z digest=sha256:7be8d3537739e2f1f45c8fb5eeace17e10508263a0de19e8f6ac5a79df8338be

Observation 51215b4c-d1d4-4180-b51c-808bb123b4b6 · inbound

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving cites this paper.

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:19:47.220532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T00:09:18.068337Z digest=sha256:da15f0ff9a0d866ded5ae6a7889154637bd0348c4c0da1288a73f598843c4f32

Observation eaa97439-eff9-4e1b-a624-221d3ae74aff · inbound

Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation cites this paper.

Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:35:38.590103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T18:21:00.089755Z digest=sha256:0d6fd50f80567f0d73efc8309f9276087d5875bda9d2490c1e15d16313207d85

Observation 7dca4217-d4eb-4efb-86bb-ebcd2e9dbffa · inbound

GEM: Generating LiDAR World Model via Deformable Mamba cites this paper.

GEM: Generating LiDAR World Model via Deformable Mamba Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.336546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:31:09.604703Z digest=sha256:e39ad00aee5114564cf738e309bf02fb1b5aede1b0a3f4b37ee61d71d733809a

Observation e9f829c1-89a9-4b72-a575-39e509e4ddfc · inbound

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation cites this paper.

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.826892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T03:15:40.944411Z digest=sha256:e93eaccaa31b89089892486cc473da5bb8ff6e20a7e2b35f7c8cf1973b7911ee

Observation 9f5332e1-14a8-41c8-8b9d-e055560af3c8 · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 296

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:07:18.079258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:4bd603efa07a8f15b32c5c78b4b5dd0ebafb116cf7f465a4654430bc2e004014

Observation bce5491f-574c-42f6-9951-6721c7c2c7f4 · inbound

EgoExo-WM: Unlocking Exo Video for Ego World Models cites this paper.

EgoExo-WM: Unlocking Exo Video for Ego World Models Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:32:36.109827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T14:32:27.920029Z digest=sha256:84cfd34e488cca4512efcdcbc8c000e2144b61ee84aca552a87b81842ca29d0b

Observation ecc64b75-8f5b-49e3-bcc0-f48ca5f9812c · inbound

EgoExo-WM: Unlocking Exo Video for Ego World Models cites this paper.

EgoExo-WM: Unlocking Exo Video for Ego World Models Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:35:47.482613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T20:26:11.104536Z digest=sha256:416291744b67d9725710850a3182dfab6a61247f0b2ce6c521beb7c11f120cb2

Observation 968dc632-0e55-4fa1-8cef-c33b988c931d · inbound

RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations cites this paper.

RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:08:04.874886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T05:07:39.815615Z digest=sha256:39e9410cddaeb5ca1b7fd910e8db77dc60f5f8375e1c61913a23cd296c0eedb2

Observation efb92402-e075-48d6-8dad-750b10bbf3f3 · inbound

CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration cites this paper.

CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.526235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T05:44:58.627525Z digest=sha256:188b59ce2f409632b6c699aa204fda150cdbdd34b0bc1e9b6748ea3d841cedc5

Observation c0fa389b-a2db-4b08-84b2-460798eaa30d · inbound

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications cites this paper.

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 275

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:43:15.745971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T08:36:23.776293Z digest=sha256:4bfb3a02333327c2bf180c00e0f1995528dd77122971b3dfc0e78d3be211a0e2

Observation b22c48c3-17b1-4e22-950d-5d80f993366f · inbound

TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing cites this paper.

TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:06:30.059920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T10:19:39.120104Z digest=sha256:101061924dbb1901f0270d4a61f6cbc1a952754173a10d70fdb232401d7c1b49

Observation 264d94bf-b860-4798-82bf-ef815477d434 · inbound

JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid cites this paper.

JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:34:36.277338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T10:31:54.292897Z digest=sha256:7fbd12a8cbca79fba403ef4e1bfb9d8ea3e8d02e548ae93696041ea3707f86df

Observation 5ed1139f-9600-4d64-8b52-90d39ef99734 · inbound

DriveJudge: Rethinking Autonomous Driving Evaluation with Vision-Language Models cites this paper.

DriveJudge: Rethinking Autonomous Driving Evaluation with Vision-Language Models Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:28:48.900716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T03:04:13.554098Z digest=sha256:c9f972fd88562919eb3934d82ffe60d0ba0456af7aafedb2113261d7b0e69930

Observation fc251de0-052a-49f8-ad33-759b3d1cfd83 · inbound

World Action Models: A Survey cites this paper.

World Action Models: A Survey Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 129

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T04:09:35.037927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T17:11:12.686936Z digest=sha256:841b5811273c78c36c73ba0ae8127ec8ca7959774d9fe7971d99dd7db9972ff8

Observation bb79cc3d-066d-403d-9d7c-fb93c367e906 · inbound

Semantic-Aware, Physics-Informed, Geometry-Grounded Weather Video Synthesis cites this paper.

Semantic-Aware, Physics-Informed, Geometry-Grounded Weather Video Synthesis Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:24:32.484310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T09:19:50.748415Z digest=sha256:0804458a1cc927f08fc2256d39bd70d0128b704b05225879b189a4d932315c55

Observation 998a84b1-3e48-42a4-b8a7-977d2f7f7f24 · inbound

HorizonRelight: Relighting Long-horizon Videos Consistently via Diffusion Transformers cites this paper.

HorizonRelight: Relighting Long-horizon Videos Consistently via Diffusion Transformers Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:24:32.107963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T09:24:04.427234Z digest=sha256:9b35f88eb3df347ce599bffbb0874d95b1b0aa311100a5752e9f53b7fe3cefdd

Observation cb3eaec8-6f38-4d73-9081-531eae02fd22 · inbound

DANTE-W: Diffuse Albedo Neural Texturing in the Wild cites this paper.

DANTE-W: Diffuse Albedo Neural Texturing in the Wild Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:05:28.523396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T07:02:23.273554Z digest=sha256:e022979f6d2fc068413598867d4dda6691281787bcebd7ee8351c27657058264

Observation f7466bbb-15a8-4171-be02-cd866fe194bb · inbound

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation cites this paper.

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T08:04:48.963890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:04:48.963890Z digest=sha256:7479bf7c08e93bbc4a2640b110577ed56d1f45060667a632e6af48e976429924

Observation 1452d2f7-5f44-4469-9bff-9ac11d4998d7 · inbound

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control cites this paper.

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T22:42:31.529313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:42:31.529313Z digest=sha256:3f2660ba867af16887a2ff98bd0ce21e24e6a5d698f9bdd90ed6da38498a51c3

Observation babce225-27ae-4451-b350-a6b723bbe9cf · inbound

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch cites this paper.

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T03:16:51.693644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:16:51.693644Z digest=sha256:427a31425cfb6b67806c049ad6f5a351190b6abba4eb390c2a228d3993ae3434

Observation c315cee6-6f04-4a6e-a37a-b747b6da00b7 · inbound

Adaptive Model-Based Transfer Learning for Dynamic HVAC Control cites this paper.

Adaptive Model-Based Transfer Learning for Dynamic HVAC Control Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T22:44:38.587764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:44:38.587764Z digest=sha256:ee406422fe3c9a146f8b16fedf0406d992c656e242c107b7e28d843137f128d8

Observation a01029ff-4844-4745-b95b-7a609e529c82 · inbound

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering cites this paper.

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T06:35:27.692918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:35:27.692918Z digest=sha256:195a186eb2a1496bc3450d7b84d88eceb5f3da4110b33c9c76dfdfc385c5ebb6

Observation 7fd41cc2-0ce4-428b-abf4-4df544bd91cf · inbound

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh cites this paper.

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T00:40:18.395572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:40:18.395572Z digest=sha256:4fb44a47fb097589b1aca80f4ca56415c70878a4b4f7992d9a18eabca95e3222

Observation b9585c3c-98f4-4271-baed-acc6429fb9b6 · inbound

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh cites this paper.

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T04:26:19.738561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:26:19.738561Z digest=sha256:fba33ae370194a94fbe52918a8defa670c1a5986c61790e67091008eea0f1b54

Observation 1c2dc313-8bca-4936-a7fc-585266d48572 · inbound

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards cites this paper.

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:12.254829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:34:12.254829Z digest=sha256:6fd820fcc41ed15da4e239b6668119c0fc9462d8af748daa192e1c18e9e3e897

Observation a03fb650-5bff-4705-b7c1-4543145f788a · inbound

GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions cites this paper.

GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:11:17.907762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:11:17.907762Z digest=sha256:b3752e648ded78a53a621581c026fa0237d4eabf528df979b580f8fb92eb5b43

Observation 224c9b46-35bb-4a1b-a8e6-dc9440f9fa36 · inbound

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts cites this paper.

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:35:01.921193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:35:01.921193Z digest=sha256:8e7d1121f52da5e2b0c51c7681aef947e4135370091142b1a8b52dc769720c11