Pith. sign in

Paper Citation Record · LEDGER

Cosmos World Foundation Model Platform for Physical AI

As of 9 August 2026, this Paper Citation Record lists 100 of 253 outbound references and 100 inbound Pith citation observations for arXiv:2501.03575.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03575 v3

Coverage vector

measured 100 of 253 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T23:38:44.933410Z

measured 200 of 200 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 100 of 423 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:42:56.918869Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 253 outbound references displayed

  • verified exact34
  • verified fuzzy63
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

9
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b9eddad2-76cb-4889-99c8-a678007a23f1 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

Cosmos World Foundation Model Platform for Physical AI SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:43:31.136179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f68d6020b1efcf1005cd8067839c019891ac7065d892fa5970300e531da0d92d

Observation d1a7fd91-439a-4324-839e-3360cddc9da9 · outbound

This paper cites Nemotron-4 340B Technical Report.

Cosmos World Foundation Model Platform for Physical AI Nemotron-4 340B Technical Report

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:38:45.226532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e641c6c9f50e69da19019d6f282aa7e34d90941b5550db57fe816d9d2e576e8c

Observation 76d9eb47-b380-48e9-8d31-8ffadc314883 · outbound

This paper cites Pixtral 12B.

Cosmos World Foundation Model Platform for Physical AI Pixtral 12B

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:53:29.932651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ee5fd3c00b8a5d21c50476fa18d7a3804f79ec8833e802152b28b87e90b2bac7

Observation b8248dc6-b877-4311-91ff-ad33d19df6b9 · outbound

This paper cites Bbc planet earth dataset.

Cosmos World Foundation Model Platform for Physical AI Bbc planet earth dataset

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.954091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:42560402449a8469799fd59ed234455e4ee73dd044ee7ea32f473e54660ee229

Observation 6efd1fdc-cf2f-4d05-a6c2-ec65ee35ac70 · outbound

This paper cites Diffusion for world modeling: Visual details matter in atari.

Cosmos World Foundation Model Platform for Physical AI Diffusion for world modeling: Visual details matter in atari

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.960688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9e379ed9b9f4741ba2349d6d0a90151c1431a6b4df38deedc970455bbfdb935d

Observation f0494f37-92a4-4ff7-a3d6-915ed3006afd · outbound

This paper cites Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation.

Cosmos World Foundation Model Platform for Physical AI Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.256610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:37e39e64e71429abfacc225015bcf120816e1cca0ab05b59ca715180fd81b2dc

Observation 9fa0fb77-f2be-4623-9000-9911990835d7 · outbound

This paper cites Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models.

Cosmos World Foundation Model Platform for Physical AI Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.273741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ed81ce2cf1f14816b5eea301eb64484e47933c30712f37c6ac5365b93c5ff3e6

Observation cf7f9cde-43d7-4a40-921f-1846801b588a · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Cosmos World Foundation Model Platform for Physical AI eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:44:22.991611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:383c0a67ab3032969cf5a8b7a0ecae18fcda7c00110c2d8a96d79247cb66d3f4

Observation 8d2efc41-4c23-47e5-89e2-3780ca54d986 · outbound

This paper cites Navigation World Models.

Cosmos World Foundation Model Platform for Physical AI Navigation World Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.293773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c1321110fc835d609b5e563965612958442c7caeca788579a15d9e505b0b5c5e

Observation 44af0301-d311-4d46-a3f7-1583e5c2ae40 · outbound

This paper cites Improving image generation with better captions.Computer Science.

Cosmos World Foundation Model Platform for Physical AI Improving image generation with better captions.Computer Science

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.032345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:43445a86108a73140bf526b642e02679fb758787b4e5ee560edcd2074fc54b20

Observation 5d9fd3bb-f132-439c-a9af-8c5c96f2ee6e · outbound

This paper cites Zero-shot robotic manipulation with pre-trained image-editing diffusion models.

Cosmos World Foundation Model Platform for Physical AI Zero-shot robotic manipulation with pre-trained image-editing diffusion models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.050351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2252fb589145f960cbd045c22e1b3f3feadaa790185e8622104b3726b5faf37b

Observation 82540ea0-d2ca-4dd0-b9b6-1a4ebb59f932 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Cosmos World Foundation Model Platform for Physical AI Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.307012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:57ca724c7a5652612245f97a0bd54c6581b8c320793d15c9e40660765ecfa53d

Observation 8907886e-a7aa-47a4-95d0-fdc8a09b4772 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Cosmos World Foundation Model Platform for Physical AI Align your latents: High-resolution video synthesis with latent diffusion models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.077346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:38a22d552c3700a2d21bea2f84c33d8973432d5eefc30b74a1987939e916d0b4

Observation 97be606c-2830-4c56-81e2-a1217a4a21c8 · outbound

This paper cites an unresolved cited work.

Cosmos World Foundation Model Platform for Physical AI Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-10T23:38:47.103360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e119a9b216052eedcfd1d3666c3d6d13cdd2879f7220cd45fedd14d7aebd2bf3

Observation 24ff8b8c-8423-454a-abab-5bbac740306e · outbound

This paper cites Video generation models as world simulators.

Cosmos World Foundation Model Platform for Physical AI Video generation models as world simulators

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.112379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a52e3eaa049cc998755b6cc81fa117014aa1d681476a9a6592c5c287fecc7874

Observation 7b001c47-7138-46fc-b888-302c1c553f15 · outbound

This paper cites Language models are few-shot learners.

Cosmos World Foundation Model Platform for Physical AI Language models are few-shot learners

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.125359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d6885ed91c9e0f2eb5e738b9b2b65988594fc9f1ba036e4c7037d905e4dc45f5

Observation c510d33b-5c7b-4c81-918e-6bbdc0952019 · outbound

This paper cites Genie: Generative interactive environments.

Cosmos World Foundation Model Platform for Physical AI Genie: Generative interactive environments

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.138349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:4f41c18f395b39320a10a2bab9353f527319af5a7e8190b118a452f270d4284d

Observation db82a035-dd8e-4187-ba25-023dbe0e05d9 · outbound

This paper cites Lee, Deming Chen, and Tri Dao.

Cosmos World Foundation Model Platform for Physical AI Lee, Deming Chen, and Tri Dao

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.146855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c9275a5888a6092173b024d4aec03cca6f6e7149981aedeb2feaee6cbfeb20fc

Observation f45fb0e2-4feb-4837-9304-8db4035cc461 · outbound

This paper cites Pyscenedetect.

Cosmos World Foundation Model Platform for Physical AI Pyscenedetect

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.156354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:dedd56d1573aa66db70997dd10aebb2cd1503b885c37e3f18281747a09f7bd01

Observation 082efe4e-f27d-4cdd-902e-c9c115d9211a · outbound

This paper cites an unresolved cited work.

Cosmos World Foundation Model Platform for Physical AI Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-10T23:38:47.160400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b9f4a1ff8e400ad28ca09cfffb6894a44f656acfb4482a34df6b91950979dbac

Observation 3a08caa9-f224-41a1-a58b-ea867f4fce90 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

Cosmos World Foundation Model Platform for Physical AI pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.166349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e911b3b6748301c864f9a33cd651a4869cc28f0dd1300454fa41a9612ab549d2

Observation 0d1d4298-c3d0-4f38-b305-c9920570d619 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

Cosmos World Foundation Model Platform for Physical AI GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:09:34.215368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:fedd5e2371f71353a7b67adee3601522930e9c7db4b93bc917adbf9edf6c82bf

Observation df32755f-f906-4059-9229-11da98eeece8 · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

Cosmos World Foundation Model Platform for Physical AI Training Deep Nets with Sublinear Memory Cost

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:44:07.405726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9e43537207a359150a5e68c4663c9a33389e8e074564bf50acb5a6cb95184a6a

Observation 610c0e48-b34e-48b6-be5f-9aab5e2955e9 · outbound

This paper cites On the Importance of Noise Scheduling for Diffusion Models.

Cosmos World Foundation Model Platform for Physical AI On the Importance of Noise Scheduling for Diffusion Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.368038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d74e67777005f1a7fe007285a1eb8d1891313cc9230727d8e4114974f9922c42

Observation 68e5fbf5-4dff-4f0f-a559-1e7c26f17cb3 · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

Cosmos World Foundation Model Platform for Physical AI Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.194838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:865fe93f2a06d3c3de91f1515d62d5c0f32778b6f4a9348d62b845eef13cb22c

Observation fd0ccab1-fe39-4eb6-9969-9a718150bb5a · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.RSS.

Cosmos World Foundation Model Platform for Physical AI Diffusion policy: Visuomotor policy learning via action diffusion.RSS

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.201403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a07bc9106034f898f6fc41ea275d59e3bf40322dad91eabfd16cda2ab9a38f85

Observation 74b91a0b-aac7-4316-90e4-3e26a37af673 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

Cosmos World Foundation Model Platform for Physical AI Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.377347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:3c1abce7e65e443fd18b3e11aec58442116be3dca3a4120a34fa31f7c6b5b54d

Observation 9a568bc0-fb85-4164-9823-5d8aa4abe361 · outbound

This paper cites The Z-loss: a shift and scale invariant classification loss belonging to the Spherical Family.

Cosmos World Foundation Model Platform for Physical AI The Z-loss: a shift and scale invariant classification loss belonging to the Spherical Family

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.394809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5e2d90a6fdb2227e4af3389599a70ad8ab63ae6e84f1a05cdfc6a1863221646d

Observation 3e634f34-4f4c-4401-822e-87f9a4f806fc · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

Cosmos World Foundation Model Platform for Physical AI Scaling vision transformers to 22 billion parameters

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.218259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cef21158fc8ba013254e488c78a5e8b063b49f5f8af8a583eb2a5a5899ef1c56

Observation 351ca4d4-d0e0-43e3-b8ea-cdfbfb372e8c · outbound

This paper cites Autoregressive Video Generation without Vector Quantization.

Cosmos World Foundation Model Platform for Physical AI Autoregressive Video Generation without Vector Quantization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:07:39.939813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:70672f4b3ab7bd3e9a2c0a9584b039ffac464248ce87091d1c935625ceed90d9

Observation f0919126-1aa5-4eea-be12-28b9dc668c4b · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Cosmos World Foundation Model Platform for Physical AI Imagenet: A large-scale hierarchical image database

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.226707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:09683de6a98a5a27afa54fc04e0f017d26baf3ccaaa362d1b3551cdd2759261d

Observation 42020881-2bce-4ab7-86f2-b28455c74624 · outbound

This paper cites Retinaface: Single-stage dense face localisation in the wild.CVPR.

Cosmos World Foundation Model Platform for Physical AI Retinaface: Single-stage dense face localisation in the wild.CVPR

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.231215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:53bbd27347e354c080a8111f3e7e86bbe8ec0843372446c6dd5fcef8c59bf97b

Observation 131fd15d-901f-48ce-8280-e7f4fbc4bcd0 · outbound

This paper cites Superpoint: Self-supervised interest point detection and description.

Cosmos World Foundation Model Platform for Physical AI Superpoint: Self-supervised interest point detection and description

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.240577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d81e71a5ce22623aa8a87714de1d47cf021d4041bddd1f83be06bc6719087f88

Observation 5481f156-04b8-48e8-a82e-f270de850db1 · outbound

This paper cites Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning.

Cosmos World Foundation Model Platform for Physical AI Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.413511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f18fade01f656bd78db8b856af0064426257927d65e8e4ba3a9921bbf18b33fe

Observation a97cd7aa-e6e6-4248-a7f3-ae9150f3d901 · outbound

This paper cites An image is worth 16x16 words: transformers for image recognition at scale.

Cosmos World Foundation Model Platform for Physical AI An image is worth 16x16 words: transformers for image recognition at scale

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.251179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a3cc8f71d4a6921b6e8adaa90ebb67f4bb4310149b755d3fedcaa6ae27881be8

Observation f6fe4fe4-14b2-4d62-b9d3-f0e7f5c79cd4 · outbound

This paper cites Learning universal policies via text-guided video generation.

Cosmos World Foundation Model Platform for Physical AI Learning universal policies via text-guided video generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.447628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:1d108f88b79b25d8427aeba1b983f7a7357f1e8c945e21477ee033f3abf13da5

Observation b403a2cb-1c99-478b-b0bf-44738389745c · outbound

This paper cites The Llama 3 Herd of Models.

Cosmos World Foundation Model Platform for Physical AI The Llama 3 Herd of Models

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.422498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:497f6fd9e211be8103272cc56203683bbccba53e2d85c823e1ff0d2f0509ded5

Observation 585782af-4b38-427a-a647-45a3e5edf30c · outbound

This paper cites Bridge data: Boosting generalization of robotic skills with cross-domain datasets.

Cosmos World Foundation Model Platform for Physical AI Bridge data: Boosting generalization of robotic skills with cross-domain datasets

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.479120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9aa18919eb94285682b3717f61eb2777e0bc8f29f50262657aaf019369c63433

Observation 7692dfdc-dafe-4b03-967c-a37691e03551 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Cosmos World Foundation Model Platform for Physical AI Taming transformers for high-resolution image synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.531141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f426209fc486992f560ef98eee08b9ca71242d883b2435e886333292fa5c72e1

Observation 3791fcb7-9592-4788-9496-2d62610ab126 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Cosmos World Foundation Model Platform for Physical AI Scaling rectified flow transformers for high-resolution image synthesis

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.362308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:65ca045be49bbe979a144aa2d78b4988b898fbd0dfad517784f62c90d7dfe529

Observation 1ac09ec0-84e0-4eb0-bcd2-b33870d039a4 · outbound

This paper cites Two-frame motion estimation based on polynomial expansion.

Cosmos World Foundation Model Platform for Physical AI Two-frame motion estimation based on polynomial expansion

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.474704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cdccfe24dcb97aa8281f79511a288ce3abb210adeaf44fbbffbc281ee66cd6ea

Observation 327fb231-5d79-474e-b826-a73383ac6036 · outbound

This paper cites Deep visual foresight for planning robot motion.

Cosmos World Foundation Model Platform for Physical AI Deep visual foresight for planning robot motion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.475477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e609a1afb26752f83c6638fe28e98799af86d0828f971453b92bc8d7c4a0f027

Observation 3d79f949-f051-4164-89cf-ba8b3f6329f6 · outbound

This paper cites FLUX.1: Image generation.

Cosmos World Foundation Model Platform for Physical AI FLUX.1: Image generation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.521517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:1fab903ace9f5be67acf92d2ca7a26d504b562f43014164fd087363c8aefe107

Observation d9c09137-44a6-4317-a1a7-eab79132a116 · outbound

This paper cites Dreamsim: Learning new dimensions of human visual similarity using synthetic data.

Cosmos World Foundation Model Platform for Physical AI Dreamsim: Learning new dimensions of human visual similarity using synthetic data

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.366304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5c33629374b8f39d5d9df10a89991f41380e052b54c206a1e8154498bb8e0c69

Observation 41cc40ae-4f80-4429-98bb-b4268818dce9 · outbound

This paper cites Datacomp: In search of the next generation of multimodal datasets.

Cosmos World Foundation Model Platform for Physical AI Datacomp: In search of the next generation of multimodal datasets

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.517500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:3eded53cacacf4a9bf6badd5d9907e931e809ee774404d0527288120f6f1497b

Observation 3e58b3af-452c-48f5-a3fe-33af67801cad · outbound

This paper cites Make-a-scene: Scene-based text-to-image generation with human priors.

Cosmos World Foundation Model Platform for Physical AI Make-a-scene: Scene-based text-to-image generation with human priors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.548648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:aab5fdc22cd5ff61af605db51d34131e10c7fbdebfab567560096053a273c219

Observation 7e0c643b-a5a5-4b66-b1a7-b77ca8068e58 · outbound

This paper cites A new algorithm for data compression.The C Users Journal.

Cosmos World Foundation Model Platform for Physical AI A new algorithm for data compression.The C Users Journal

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.427828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:4c9226b9555b15f4a2bdc96e0a349e64ad5acaf3ff914035c78d82ce684daa7a

Observation a061fda2-7525-411b-b07d-d61826ecd72f · outbound

This paper cites Murphy, and Tim Salimans.

Cosmos World Foundation Model Platform for Physical AI Murphy, and Tim Salimans

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.503457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ef2d5eec212f962e4f0e8eaa4521af5f974839f047304059ff18c070d70a7e5f

Observation 09e0f4a1-c4a4-45c4-b6c3-c6c62c124c8b · outbound

This paper cites MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control.

Cosmos World Foundation Model Platform for Physical AI MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.431062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2f42ac94536a7d2bcc95b7b32c8cdc99817cb6ed89ad3a6d0155d031c583877c

Observation 052f054b-77b5-40f2-b58d-192145f5dc0e · outbound

This paper cites Magicdrive: Street view generation with diverse 3d geometry control.

Cosmos World Foundation Model Platform for Physical AI Magicdrive: Street view generation with diverse 3d geometry control

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.499767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e6593627bdaed8c160293504f6e045114a0f7aa24e4c5c249941dcc1c73bcd8e

Observation 917ffa15-80ca-4f3e-945d-7ca6ada99ebb · outbound

This paper cites Vista: A generalizable driving world model with high fidelity and versatile controllability.

Cosmos World Foundation Model Platform for Physical AI Vista: A generalizable driving world model with high fidelity and versatile controllability

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.508533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b7f42517f1bbd8b2f7636e63beb09baeb021c2534db79995cb40d38f5852e714

Observation 7efcc68e-0777-4923-90dc-da561c561e9a · outbound

This paper cites Image style transfer using convolutional neural networks.

Cosmos World Foundation Model Platform for Physical AI Image style transfer using convolutional neural networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.491965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ef6096ffb949209447f4742a275c82b57f103726a77de6f93b77456136c64f44

Observation 8419c0a4-cbfe-435e-b835-8f2cc198a026 · outbound

This paper cites Long video generation with time-agnostic vqgan and time-sensitive transformer.

Cosmos World Foundation Model Platform for Physical AI Long video generation with time-agnostic vqgan and time-sensitive transformer

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.441584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:623aa21745c5db1eff965a44198a3a1050e3a378668da5b168eef0177761848e

Observation a7b15757-ef00-4560-b4a7-59310c3b1dd4 · outbound

This paper cites Preserve your own correlation: A noise prior for video diffusion models.

Cosmos World Foundation Model Platform for Physical AI Preserve your own correlation: A noise prior for video diffusion models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.471705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:bfc4de2b22dfa7f1bdae043f1da7e086d0e7bd903291f8931b2a67872fa2ffff

Observation bb242d0a-a2b3-4b03-9913-3f5b3f4abce5 · outbound

This paper cites Visual fact checker: Enabling high-fidelity detailed caption generation.

Cosmos World Foundation Model Platform for Physical AI Visual fact checker: Enabling high-fidelity detailed caption generation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.420285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:4f654d42893bb31108910a660bac7732151e777c7beb50e9643c46e5e5d8b3e4

Observation c38e0513-05f2-4a38-8b44-ff848401c888 · outbound

This paper cites AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts.

Cosmos World Foundation Model Platform for Physical AI AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.444469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5d0fa478e0e028601026eecb74a18afdaa76823238e00635fe6aef5b22f24975

Observation f6763656-85ec-4941-81a1-d5b0e763cb35 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Cosmos World Foundation Model Platform for Physical AI Imagebind: One embedding space to bind them all

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.556492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6d47ebed3a537f14d4dc95d03f8087797568fc251d369889db85d1acbc2d053e

Observation 20df0eaf-79b7-4907-8ccb-70d10a214247 · outbound

This paper cites Emu video: Factorizing text-to-video generation by explicit image conditioning.

Cosmos World Foundation Model Platform for Physical AI Emu video: Factorizing text-to-video generation by explicit image conditioning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.540196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:7b38bee2a98fb55626823ff6c548a4c8d74178e87c2bec8ebad1ed461dfcc317

Observation eecd8030-5738-4296-9b79-c6da8741e23a · outbound

This paper cites Ego-exo4d: Understanding skilled human activity from first-and third-person perspectives.

Cosmos World Foundation Model Platform for Physical AI Ego-exo4d: Understanding skilled human activity from first-and third-person perspectives

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.495940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:3ba728fcd3d49251502fbfcbb77f00f8530e4fff53de0fd09f3a6e83f3837840

Observation d636bb7a-7ddd-4903-a7ee-df722c13e080 · outbound

This paper cites Photorealistic video generation with diffusion models.

Cosmos World Foundation Model Platform for Physical AI Photorealistic video generation with diffusion models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.544202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5f928dc915ff35f3e765770a73a793c543429d0437d81c1c7524f069b8e2734c

Observation 71ac6de8-988a-4710-837c-de74d6a31447 · outbound

This paper cites Pre-trained text-to-image diffusion models are versatile representation learners for control.

Cosmos World Foundation Model Platform for Physical AI Pre-trained text-to-image diffusion models are versatile representation learners for control

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.483224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:962fb7abb711292dd112cfd42cfbf2708f7d307ee15553752179d860b96a397f

Observation 01b8d6ca-8679-4bb8-bc57-9d7a95f16fa0 · outbound

This paper cites SPACE:Speech-driven Portrait Animation with Controllable Expression.

Cosmos World Foundation Model Platform for Physical AI SPACE:Speech-driven Portrait Animation with Controllable Expression

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.401003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:fac37742fc1fbdaf1ae5876b7512253dbc4e762e58d035b1aa96957edd3aca05

Observation 1cd3b051-b359-4c21-991a-c774a606a1f2 · outbound

This paper cites World Models.

Cosmos World Foundation Model Platform for Physical AI World Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:08:36.665942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a852e836d639ff365a140f8fe25babaa144148adbaf9c64d21649c2a6aa7eb23

Observation 4bf1f6d6-0a88-45e8-8516-61591075103a · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Cosmos World Foundation Model Platform for Physical AI Dream to Control: Learning Behaviors by Latent Imagination

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:16:36.618198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d7a5f31248da741383fae2f7e35a2e87b46ed4b3bd8211e85e64cac89f7f2c49

Observation 1b88f61f-2bcd-4d2a-a1ac-a3dad34dff1d · outbound

This paper cites Mastering atari with discrete world models.

Cosmos World Foundation Model Platform for Physical AI Mastering atari with discrete world models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.430428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:529beeebc6d2c096bcbc12a0e6c980d71ac8306a9cc85dce73e0c853325118f9

Observation 122bdbad-d14d-4b14-9a82-23007c0ebf3b · outbound

This paper cites Mastering Diverse Domains through World Models.

Cosmos World Foundation Model Platform for Physical AI Mastering Diverse Domains through World Models

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:08:22.800194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:4bab8d6717d5d9aa839c1e4de5592b1da080d0040baffe18c48329490c67ee23

Observation edfa3c1e-a956-4442-8e01-695c33dcb755 · outbound

This paper cites Td-mpc2: Scalable, robust world models for continuous control.

Cosmos World Foundation Model Platform for Physical AI Td-mpc2: Scalable, robust world models for continuous control

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.453439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b74251fbebe66f9ff47e4c5a62790c3986548f6699bbe6381b63478ddda7fc58

Observation c884df7d-420d-43b2-b2dc-e9f2dbe0ed98 · outbound

This paper cites Cambridge university press.

Cosmos World Foundation Model Platform for Physical AI Cambridge university press

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.525849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a0ef29198b4d6f48729c09ab20486b9d77bedbc3a6ad00acf8b40c150e815d2e

Observation 11282f87-2687-45a5-a525-cb7e44d96d25 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

Cosmos World Foundation Model Platform for Physical AI CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:06:24.067006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:682c39ad2e73e0542e917ae11eeb4f9fea5789f34efada53888efc73ca4ad2bb

Observation bba45182-89b5-48e0-932b-838544dfd566 · outbound

This paper cites Learning an actionable discrete diffusion policy via large-scale actionless video pre-training.

Cosmos World Foundation Model Platform for Physical AI Learning an actionable discrete diffusion policy via large-scale actionless video pre-training

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.391013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:301698ad349c9d8450e0b86e95564ec60a74f5c17c125d89d1c0428d07478a32

Observation 0ac13db5-4cf7-4b4e-987b-26417e0e75c1 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Cosmos World Foundation Model Platform for Physical AI Masked autoencoders are scalable vision learners

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.484329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:398abcc79024ab8ee08b05cc69a1ef6782c32f5138a5f8f162b405ba8cd9ce2d

Observation 00f28e15-78dc-4604-acd8-d2a850d7f5b5 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Cosmos World Foundation Model Platform for Physical AI Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.535933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:bf4d6a7b85907148a0b806bcd71f6b884b556df16776b4aba8389228bbc8af13

Observation b959d73d-d7c9-4306-a626-96b50251e9d9 · outbound

This paper cites wake-sleep.

Cosmos World Foundation Model Platform for Physical AI wake-sleep

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.487460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ca711df31836eacf00ca3ebe7e575a8d5a6bd788202f9c40a86f8eb29fe18f73

Observation e9f131af-e71d-402b-8206-6ec61072c4a5 · outbound

This paper cites Classifier-Free Diffusion Guidance.

Cosmos World Foundation Model Platform for Physical AI Classifier-Free Diffusion Guidance

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.494382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a82fb45cad00d088a153590de05ef23050b92b77dd46ce9aaea5cb7dba2c3378

Observation c7a204ae-f586-430e-aff6-27a1dac0a726 · outbound

This paper cites Denoising diffusion probabilistic models.

Cosmos World Foundation Model Platform for Physical AI Denoising diffusion probabilistic models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.436107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:689069cd5b82266e197ce43d249994be6c459f6427ef22263ce305f357bd2524

Observation e694479b-d138-4d0d-98db-aeadf1da4a02 · outbound

This paper cites Video diffusion models.

Cosmos World Foundation Model Platform for Physical AI Video diffusion models

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.462831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6697e4d58812f90f7c308f7d3e0b29ff47e869a7f86f73fbd8177acae7c708c1

Observation 8fb25a0f-7222-4691-9deb-742c3600272c · outbound

This paper cites Cogvideo: Large-scale pretraining for text-to-video generation via transformers.

Cosmos World Foundation Model Platform for Physical AI Cogvideo: Large-scale pretraining for text-to-video generation via transformers

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.560558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:43aaa3cd51d12e154ccad16b9844f0eab27ffc7a91420749f63b5a4c99c99165

Observation 162e0e9c-1f84-464a-bd31-c3ede4134fca · outbound

This paper cites Simple diffusion: End-to-end diffusion for high resolution images.

Cosmos World Foundation Model Platform for Physical AI Simple diffusion: End-to-end diffusion for high resolution images

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.471142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d046ec6016b3cc31884cf83c3f3730305222c34e12e2adf967cd512bad81d4da

Observation 8ffd52d9-526e-4394-99e8-712bbd521c8e · outbound

This paper cites Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion.

Cosmos World Foundation Model Platform for Physical AI Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.501344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0396b993755f5b677aaf44153c7564dd6072a6c23d1cdc7205b8f37b3f6d4310

Observation eb161c84-38d1-4100-bda3-dab8034346ed · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Cosmos World Foundation Model Platform for Physical AI GAIA-1: A Generative World Model for Autonomous Driving

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:15:10.779550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:603b5f453a285b109a989205f12ac015b60b1f36595990309fec83e121c6f26a

Observation 9a07989a-0508-4b42-86d2-5b744b87a199 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Cosmos World Foundation Model Platform for Physical AI LoRA: Low-rank adaptation of large language models

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.451531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:149ed7e182015666e2e3c9ac17dc7889bf8590f632936bd000a1c72099691b2d

Observation 6412333a-37af-4624-9afa-a6bef513e673 · outbound

This paper cites GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs.

Cosmos World Foundation Model Platform for Physical AI GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.524063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2cd291a99d3f761bedd1d241534b27ce18f66bed9fbda09d87e62f9130e7da9a

Observation c126c200-4f63-41a8-9780-7af978d91c6d · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

Cosmos World Foundation Model Platform for Physical AI Vbench: Comprehensive benchmark suite for video generative models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.564785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cfe5fed1e73436e963f579ca3415f0ddc3afe2c938c4016b3ea7e25314b49feb

Observation c077c955-8436-4525-b0ca-a0fd153a664d · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Cosmos World Foundation Model Platform for Physical AI Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 84

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.530796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:eafc4754387e2c04696ec51b7a6262f58c77f9b3a12afcf8578c43b39f6de6b3

Observation 5cb2648e-9301-4a93-8528-17243aacb2eb · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Cosmos World Foundation Model Platform for Physical AI DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:22.650116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ab7bc33443ee04e58442a27e0d3eacafd5274db649f169e898c335a82656e8a1

Observation 747c2d4e-dcdc-4926-b9f2-09752f00729a · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

Cosmos World Foundation Model Platform for Physical AI ADriver-I: A General World Model for Autonomous Driving

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.553921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d47aa94ad6bddbb85dff25f43d914c034a940217f1d76bcf5786b01bac8ccd04

Observation 8cccda4b-da3c-42a0-a2f1-5dba18b811e7 · outbound

This paper cites Mistral 7B.

Cosmos World Foundation Model Platform for Physical AI Mistral 7B

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.559658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f6a64fdd635d2fa122d77961da9ceb272bc8cbeafe87d260398c95598c1113ce

Observation bf0fded8-bb87-4995-8ca2-da5b34d34525 · outbound

This paper cites How Far is Video Generation from World Model: A Physical Law Perspective.

Cosmos World Foundation Model Platform for Physical AI How Far is Video Generation from World Model: A Physical Law Perspective

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:13:41.491170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9fa77e782f21f2820b3adb9278ece0ff3d6063792ec6ed037f7f5ecf33e28936

Observation 77c7af6d-1fa0-45e2-a2eb-b79dd46bf9f3 · outbound

This paper cites Scaling Laws for Neural Language Models.

Cosmos World Foundation Model Platform for Physical AI Scaling Laws for Neural Language Models

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.577077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:555662a4bdc0e1b391c6d63f1ee37bf9b0f188baee551caf9040a84cdedacaf2

Observation cb7d58f3-d100-4925-9604-bc3dbdd0df0c · outbound

This paper cites Analyzing and improving the image quality of stylegan.

Cosmos World Foundation Model Platform for Physical AI Analyzing and improving the image quality of stylegan

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.219507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6ef7c19e6c57f3cff1c3e6bcf9a2d5cef57121ef121252e6afe0450f3b8c4773

Observation eaebe424-dd94-4215-af7e-f4b7ad315ba7 · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

Cosmos World Foundation Model Platform for Physical AI Elucidating the design space of diffusion-based generative models

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.236004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:15b1d5474068c49d4577e37d96beae70bd00032dc99cb6ed887ffe5d5cb608af

Observation f48f9da6-31ba-4f65-bf5e-34f43a025811 · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

Cosmos World Foundation Model Platform for Physical AI Analyzing and improving the training dynamics of diffusion models

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.241534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:d75ea85e88c5682cec48c446829c97016d243a070331e6e0fa7051479d1c144c

Observation d0944257-c6c4-4007-ac27-6d8db6044f2f · outbound

This paper cites 3d diffuser actor: Policy diffusion with 3d scene representations.

Cosmos World Foundation Model Platform for Physical AI 3d diffuser actor: Policy diffusion with 3d scene representations

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.253128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c3800ea3be3a769db4281530d2a0f9a5649de66015980deda3a1d4dbef7ad546

Observation 1b849c64-4670-4591-82a1-601b4c16b9ae · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Transactions on Graphics (TOG).

Cosmos World Foundation Model Platform for Physical AI 3d gaussian splatting for real-time radiance field rendering.ACM Transactions on Graphics (TOG)

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.261429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:93af72873069976b0c8580156ef072c71e560b7bb2a262e6185064438c36bc00

Observation 2e255add-5a32-4427-a3be-bab393a1f6fa · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

Cosmos World Foundation Model Platform for Physical AI YOLOv11: An Overview of the Key Architectural Enhancements

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-12T13:29:16.418786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:98deae32578b36d08b80095fe4b3ce375e2271a0f4d46a91a0948437162d49f0

Observation 752045ee-252c-4ac9-8516-4dccea3b3dc9 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Cosmos World Foundation Model Platform for Physical AI OpenVLA: An Open-Source Vision-Language-Action Model

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.595747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:03f8ac7ba7a5bcf7ae5820bf4a7b8db371f44e0406c15005938bd59ad1c75160

Observation 8ee927dc-3969-497a-ae2e-cdfc08a66b50 · outbound

This paper cites Learning to Simulate Dynamic Environments with GameGAN.

Cosmos World Foundation Model Platform for Physical AI Learning to Simulate Dynamic Environments with GameGAN

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.295756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:008ae30c0a17c813c8ec55a98efffab404673d0b2532a8704e61530ada95c4e4

Observation e36484a6-c7e1-4902-b27b-112c6787ae3c · outbound

This paper cites DriveGAN: Towards a Controllable High-Quality Neural Simulation.

Cosmos World Foundation Model Platform for Physical AI DriveGAN: Towards a Controllable High-Quality Neural Simulation

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.308459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:8e1104732066fd5fb02cf872671d1ab2894bfcce94f8d72925f09baaea6e2b0e

Observation c8ab99c4-d22a-4efc-8137-a9b8fda50ff8 · outbound

This paper cites Auto-Encoding Variational Bayes.

Cosmos World Foundation Model Platform for Physical AI Auto-Encoding Variational Bayes

Reference 99

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.606393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e6a5ea00d3c4dcaeff6a91613eb83cad50f059308287e6e9b010264248225574

Observation ba229324-e554-46c4-8eea-fb3ec7296d31 · outbound

This paper cites Learning to act from actionless videos through dense correspondences.

Cosmos World Foundation Model Platform for Physical AI Learning to act from actionless videos through dense correspondences

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.337589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:7de25f114f6419666c8c9e4c7108d058746a1e9b1f7c9387ff2c05793770ca3a

Pith citing papers

Observation 87e26b48-3ac2-407e-b763-1e55d40d7074 · inbound

Latte: Latent Diffusion Transformer for Video Generation cites this paper.

Latte: Latent Diffusion Transformer for Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-13T21:45:35.796766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T21:45:35.754742Z digest=sha256:66e7b67bc6bb2ca5dd4b2450b5b33da9fbbf4ca2c262e94bed862c55af729302

Observation 4097a27e-822c-4e46-b132-57a05af61dd6 · inbound

Do generative video models understand physical principles? cites this paper.

Do generative video models understand physical principles? Cosmos World Foundation Model Platform for Physical AI

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-20T12:47:05.854160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:47:05.825659Z digest=sha256:525975e5d777c18a340e159c561ddf07ee5994e37eb7338eaf0e5a6976a514ce

Observation 07d9e05a-eb3a-484f-b201-1502b89bf29a · inbound

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling cites this paper.

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-17T02:52:20.670876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T02:52:20.643070Z digest=sha256:5c7cd4a3fb8c40601ecf4445187cf5fb3a5612433d25f122bf0f94356d4ae6d0

Observation dc8db966-b688-4d8c-82b4-cdffaee7983e · inbound

Multimodal Medical Code Tokenizer cites this paper.

Multimodal Medical Code Tokenizer Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T00:42:56.918869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:42:56.918869Z digest=sha256:9e70e86e88843f193b4d146d29d9b5a67bdf370d9a39270b9121535257e2908a

Observation 85928938-7044-4a89-a653-5151f2f030ab · inbound

Goku: Flow Based Video Generative Foundation Models cites this paper.

Goku: Flow Based Video Generative Foundation Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T21:07:32.156095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:07:32.156095Z digest=sha256:a91b6bf1a55f3d978ca4e32f6fa38181ea8f67d5e4c6f86914ce8c8d8bd392be

Observation 6f1d706e-5be6-48ce-a7eb-88a2acd01955 · inbound

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model cites this paper.

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Cosmos World Foundation Model Platform for Physical AI

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:02:23.962171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T08:02:23.002090Z digest=sha256:bb9a8f3a1b071a0f0b429adde978670dea11af1bb3a32e12f270b6c43d20139d

Observation eaad51e5-1286-4666-9fd2-e6c708b049e3 · inbound

Simulus: Combining Improvements in Sample-Efficient World Model Agents cites this paper.

Simulus: Combining Improvements in Sample-Efficient World Model Agents Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-23T02:32:26.169585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T02:28:23.113408Z digest=sha256:e9b53a61d1e6bd40d43a961246dcd68799f562d8086e4a924e24d8a1ed542552

Observation f4bafd84-306a-4ebd-80cf-94f7275ee836 · inbound

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots cites this paper.

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:09:10.112304Z digest=sha256:b487cd17939aa48cc3ce36cc37156db25a1a422233eea7eaab7d3cbde0879ecc

Observation 0086e66b-c74d-4771-86c5-fc8b87d6633c · inbound

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning cites this paper.

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning Cosmos World Foundation Model Platform for Physical AI

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:47:10.244526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:47:10.146795Z digest=sha256:e9fc7e302ef1e2efc50a0d00073cdf8efeccf5fcd131c40280f5fe5bdbcfe82d

Observation a3d5a6f2-c170-4668-8183-428a9b4e96b2 · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:05:17.269584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:10972b9c7dfec752ebf420255a0acd4108b77b5dae7bfd29a3705264f0020b28

Observation 6f53b649-f81f-4bcd-8611-eb03a817272e · inbound

GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving cites this paper.

GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-15T13:48:22.318514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T13:48:22.279405Z digest=sha256:9bc071fa6f32ca25e404cc106cc7c3ece2f59ac74caaf8d9f0564d4e2536db80

Observation d8035884-cc36-4141-b377-9d2473332ee3 · inbound

AccidentSim: Generating Vehicle Collision Videos with Physically Realistic Collision Trajectories from Real-World Accident Reports cites this paper.

AccidentSim: Generating Vehicle Collision Videos with Physically Realistic Collision Trajectories from Real-World Accident Reports Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-22T22:27:12.531213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T22:25:57.796922Z digest=sha256:26062b771a8b6ecb852e8a397bfdcbb0bfd418bc39dfdad02a6b0ee615ff649a

Observation b2a5a481-821d-47ca-b442-25c02053e973 · inbound

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness cites this paper.

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness Cosmos World Foundation Model Platform for Physical AI

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-14T18:42:03.311760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T18:42:02.940250Z digest=sha256:40289bf03d59482105c3beaa878ab6348bf3fcdf7d18bb82456f8f85a09c6b2f

Observation faa18e6f-0cb2-4d91-a77a-f29599f73889 · inbound

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets cites this paper.

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets Cosmos World Foundation Model Platform for Physical AI

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-13T16:25:00.461445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T16:25:00.365534Z digest=sha256:ddb487acf145def2d887817455f349d3a312c189cef37e53483ea673ceb039ed

Observation 727c359b-f3fa-4b70-802e-74e0fe353586 · inbound

SkyReels-V2: Infinite-length Film Generative Model cites this paper.

SkyReels-V2: Infinite-length Film Generative Model Cosmos World Foundation Model Platform for Physical AI

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-14T20:23:04.195745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:23:04.022599Z digest=sha256:fbb07cc1e54f8b3a7ec37aeefc032728bac7bbd666639e6d3910ce09a02c0aa2

Observation fda9cd74-497c-42ec-ba96-b2fb99f30de6 · inbound

EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video cites this paper.

EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video Cosmos World Foundation Model Platform for Physical AI

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T15:40:29.404704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T15:40:29.361187Z digest=sha256:7510879c2c8fe4d8a2b289ede1c8ec30eb8f077124b34b65a7e5c3bffaabac5b

Observation 3be5cab3-7d3b-43a4-96a2-d5a64357f8ab · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Cosmos World Foundation Model Platform for Physical AI

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-15T23:50:45.427504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:7aa123ce2308ac914259d33afdc9447e9a849fdc644c33f5f49d4699859bd75f

Observation 3d2a7284-190a-4ad7-a50a-dfcb170daa82 · inbound

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model cites this paper.

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:24.006927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:43:24.006927Z digest=sha256:2a573a09f4813862dcace8d9a2aa20c8dba1b2d29cde6777bf75ae602924b82e

Observation e75c1a73-b4f3-4a38-878b-ce2f6d7f4a94 · inbound

Interspatial Attention for Efficient 4D Human Video Generation cites this paper.

Interspatial Attention for Efficient 4D Human Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:14:20.861002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:14:20.861002Z digest=sha256:e2d0d66046606ac8d20a58dc9a6b3721c249e55f86b1b6525fe0605a7e24a050

Observation 4e2ee009-0eea-42d5-80d2-b1e18af85000 · inbound

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design cites this paper.

QuickVideo: Real-Time Long Video Understanding with System Algorithm Co-Design Cosmos World Foundation Model Platform for Physical AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:00.103049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:00.103049Z digest=sha256:c1d376038f2067b1a5d94a32bd2e4e521dc1df85ed56ef273d53a34da495f8fc

Observation 589dae3b-092e-4076-be5d-cc5563d90ebf · inbound

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence cites this paper.

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-22T13:11:35.653194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:07:11.548885Z digest=sha256:87670cc986600ff527c46ee2b8a96e873bdbaec0c39941563d35f8961b121d53

Observation b40e162f-bb38-4acf-a8b4-4ee07e587fdb · inbound

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild cites this paper.

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T12:57:17.713523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T12:56:30.634284Z digest=sha256:45c2916011c6d8d6f76046e04ad88da0707eaab9d9b7604930d65ff68883f88f

Observation 694019ef-d3d2-49e6-b448-33ea55633d84 · inbound

SCIZOR: A Self-Supervised Approach to Data Curation for Large-Scale Imitation Learning cites this paper.

SCIZOR: A Self-Supervised Approach to Data Curation for Large-Scale Imitation Learning Cosmos World Foundation Model Platform for Physical AI

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:33.421703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:33.421703Z digest=sha256:5e4fd58b96c0be60d1c2099b8b96f724b86646cc23e00c55ca9fdc2f25a929a8

Observation f9a3f170-251e-44e0-964c-d8d060f8bb88 · inbound

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models cites this paper.

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:04.895413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:04.895413Z digest=sha256:358cbf6c29410c02fe4a48e43c1a2199baf7947f319e9947ec12c381b60dce4a

Observation e4f598ac-8b86-4fea-86e9-ae2af88e63ad · inbound

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis cites this paper.

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T12:02:16.767815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T12:00:37.025335Z digest=sha256:7f2392ae3968e33f3888fbfc932de4dff15858d8e137b576541646cf1b712e8f

Observation 13f4c28b-51d2-409c-a6ab-5e75404bc1cc · inbound

FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation cites this paper.

FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:54:16.397022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:54:16.397022Z digest=sha256:91c45f8d482dd255bb39a5c5668bf7e8f943da332de59c153a8de2fd5c4b4a6e

Observation d1063c36-9b65-4a8f-a1eb-08ceecfdc30b · inbound

Humanoid World Models: Open World Foundation Models for Humanoid Robotics cites this paper.

Humanoid World Models: Open World Foundation Models for Humanoid Robotics Cosmos World Foundation Model Platform for Physical AI

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:54:56.598202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:54:56.598202Z digest=sha256:08c7732b2272212f35a4769c6d655ed396300e660d961ce4a41edd862df3a94f

Observation 244b70f9-ebf8-4ce1-8757-fca7a6da178c · inbound

Playing with Transformer at 30+ FPS via Next-Frame Diffusion cites this paper.

Playing with Transformer at 30+ FPS via Next-Frame Diffusion Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:49:41.669315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:49:41.669315Z digest=sha256:5ea5c03042086f3d57cf71a0a604239ba2beec28d6338235b975f03af83e18b1

Observation 95e7e55d-1297-4f69-be1c-775a2973b5ea · inbound

WorldPrediction: A Benchmark for High-level World Modeling and Long-horizon Procedural Planning cites this paper.

WorldPrediction: A Benchmark for High-level World Modeling and Long-horizon Procedural Planning Cosmos World Foundation Model Platform for Physical AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:50.945500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:50.945500Z digest=sha256:e745ec140493dbe6da4c49b44adb96a4ca28885169c4ef39d10594c67adac5e2

Observation 2182539b-505d-4ca3-aaaf-b2657f45872e · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Cosmos World Foundation Model Platform for Physical AI

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:40.968572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:40.968572Z digest=sha256:2dd019508c3861d5dc0df6b9b24bd509338045cdea1ab8505aa3de0aeac9e21b

Observation 4286e6d7-37ae-4216-b2cc-2fbec26dba99 · inbound

Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics cites this paper.

Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:23.995257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:23.995257Z digest=sha256:3653ecd7e03a760708cb9cacc80806f8b8e142d92d792bb99c4561de4ed080f4

Observation bf189677-c2e6-4593-8b7f-a41be0088752 · inbound

Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion cites this paper.

Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:47:14.332355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:47:14.332355Z digest=sha256:6dd4c130f43818814b2498c2e4e2b981959c7537911cdaa78e5807cdd5beb6a3

Observation 1194eee9-5f64-4a62-b271-043a3751e3ce · inbound

From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models cites this paper.

From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:30.982950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:30.982950Z digest=sha256:9d79c1fcaa345c9028ec52c0b6f7a0bab2e92db7f4e569239c1bce3b935b5eea

Observation cf5137c5-86e3-4333-9e18-23dcab4e0c56 · inbound

EgoM2P: Egocentric Multimodal Multitask Pretraining cites this paper.

EgoM2P: Egocentric Multimodal Multitask Pretraining Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:31:42.352925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:31:42.352925Z digest=sha256:fca87add970997bcfa427fb246cbf1937e021ad0b6cc02a02bbfb094f38ce8b9

Observation e296f37f-7262-4e36-9ab5-699e0e348ad8 · inbound

MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation cites this paper.

MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation Cosmos World Foundation Model Platform for Physical AI

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:02.174857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:26:02.174857Z digest=sha256:e95984d0e475cb31576c5fb689c5f0cebfbf65cf9798d1811ca180e4623b6794

Observation a969e97e-b74a-4849-88ad-4430658d805b · inbound

Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models cites this paper.

Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.961334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.961334Z digest=sha256:65bbb7e17e0673e1cb20a083934ee233bb506de43e3c95e31d87fb55efe5ebe2

Observation 32cabdd5-345d-416c-ad84-abcc019dd647 · inbound

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models cites this paper.

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:54.075933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:54.075933Z digest=sha256:2e3452520822b82da8973821baaf2f232b14432b56963fdc04b132eba02cbd75

Observation 64075647-6577-40f5-bbab-de06c42ae1a8 · inbound

VideoMat: Extracting PBR Materials from Video Diffusion Models cites this paper.

VideoMat: Extracting PBR Materials from Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:48:24.773347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:48:24.773347Z digest=sha256:0b0ffe8751a97a35e558ff7bec5c2fc7d058a9a4e3f5122f3f610e6ceb39f12b

Observation f7a93f73-ff0f-464d-b438-24e6dbb99af4 · inbound

IntPhys 2: Benchmarking Intuitive Physics Understanding In Complex Synthetic Environments cites this paper.

IntPhys 2: Benchmarking Intuitive Physics Understanding In Complex Synthetic Environments Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:43:58.401203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:43:58.401203Z digest=sha256:292b98c0520a99e079375991b1c60dc03869628d79ee792af9f24e2171e637c7

Observation 201bc43f-a55d-4a40-abd1-5a4ab1bd16a4 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:33:50.706975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:7c118b6d7494682583095aa4fc230869d974cf1acd7f31dfaa46a4c2d72738c3

Observation 2c9d0627-ec58-446f-bd33-3bbd218ec092 · inbound

GenWorld: Towards Detecting AI-generated Real-world Simulation Videos cites this paper.

GenWorld: Towards Detecting AI-generated Real-world Simulation Videos Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:41.910019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:41.910019Z digest=sha256:060809fddf05401fd394376181f95f10bdbff2e1d1e3cb5133084b8ff92f98e9

Observation 4d51a2ba-e790-4222-a9af-b695382d504e · inbound

Perspective on Utilizing Foundation Models for Laboratory Automation in Materials Research cites this paper.

Perspective on Utilizing Foundation Models for Laboratory Automation in Materials Research Cosmos World Foundation Model Platform for Physical AI

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T00:55:16.991821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:55:16.991821Z digest=sha256:7b9183ad51c172a0ae13e9e09bda490a7e77ad633406b819047a2be7659098e2

Observation 2b23abd8-3f55-401d-bffb-a97364c2626b · inbound

VideoMAR: Autoregressive Video Generatio with Continuous Tokens cites this paper.

VideoMAR: Autoregressive Video Generatio with Continuous Tokens Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:26.225981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:26.225981Z digest=sha256:6818e9f6a5dfd4218c2b22285f6d5972d89a11e9fdf67363d21bd93659b910a4

Observation 8a349e2f-cf44-4191-9e88-63c8540fcb83 · inbound

AMPLIFY: Actionless Motion Priors for Robot Learning from Videos cites this paper.

AMPLIFY: Actionless Motion Priors for Robot Learning from Videos Cosmos World Foundation Model Platform for Physical AI

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:25:28.121151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:25:28.121151Z digest=sha256:15ab485f065d9b75dd0af7b86f6503e73597bcd2cb4e48c5f4a9c05ac5b8f1d9

Observation e084a9d0-7372-4584-8387-a3e17518e62a · inbound

UniRelight: Learning Joint Decomposition and Synthesis for Video Relighting cites this paper.

UniRelight: Learning Joint Decomposition and Synthesis for Video Relighting Cosmos World Foundation Model Platform for Physical AI

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:37.011999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:37.011999Z digest=sha256:c2911c9e927e40735861279be414dee93a054dce3b454d33240444fcd691575d

Observation 807b8b19-2ca7-4da7-8abe-da57faf68049 · inbound

Matrix-Game: Interactive World Foundation Model cites this paper.

Matrix-Game: Interactive World Foundation Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:19:59.423764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:19:59.423764Z digest=sha256:c905293c8ee29dc77f33a704e159fa669a02dd186f465a00d32734064ac76619

Observation ec3b4777-ccef-40fa-bbb0-fd13f8f9a177 · inbound

WorldVLA: Towards Autoregressive Action World Model cites this paper.

WorldVLA: Towards Autoregressive Action World Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T22:57:07.947528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T22:57:07.883617Z digest=sha256:47354feb99fbb5840622b07f44b6916b7576234f0e28dc5ae41a61987ff93845

Observation 2e4c45f8-31e5-4b5a-975f-31b039fcfc9a · inbound

RoboScape: Physics-informed Embodied World Model cites this paper.

RoboScape: Physics-informed Embodied World Model Cosmos World Foundation Model Platform for Physical AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:56:07.501402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:56:07.501402Z digest=sha256:60ba44777793c35cc733f10fc4a2de7a95ebcc6c126d9bf02410374b79059da5

Observation 10c8e198-bb2a-4f56-a575-d78b59a44d09 · inbound

Epona: Autoregressive Diffusion World Model for Autonomous Driving cites this paper.

Epona: Autoregressive Diffusion World Model for Autonomous Driving Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:29:46.636813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:29:46.636813Z digest=sha256:21fb573c1d02651d7a640dfc0c75dc133353b9e2598395db47ec91dfa9b93129

Observation 04808656-9896-405a-b698-c4a2c9708aa9 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion Cosmos World Foundation Model Platform for Physical AI

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:16.402862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:16.402862Z digest=sha256:384b61f70c4eb1a55276f821f4ddab16b6506bba84e9c3140fb2fccde26cb441

Observation 5672c289-a6aa-4547-8b0d-01ff6d2e0014 · inbound

A Survey: Learning Embodied Intelligence from Physical Simulators and World Models cites this paper.

A Survey: Learning Embodied Intelligence from Physical Simulators and World Models Cosmos World Foundation Model Platform for Physical AI

Reference 294

Resolution
unresolved
no resolver link, observed 2026-08-06T21:09:18.844124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:09:18.844124Z digest=sha256:6d13a705455a2b8acbe2a6a5e01861fdb93620d6b50aa4dfd40371ba8d491a0a

Observation a6c0ac49-d018-4331-8fcb-3f91c30212e7 · inbound

Are Synthetic Videos Useful? A Benchmark for Retrieval-Centric Evaluation of Synthetic Videos cites this paper.

Are Synthetic Videos Useful? A Benchmark for Retrieval-Centric Evaluation of Synthetic Videos Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:37:04.316465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:37:04.316465Z digest=sha256:31cc3afe5c3e35151c6287a34585d42e763bfa4fda0dc6f2991d9bca488d1844

Observation cbea5d6c-2af4-48c4-a9e3-0f92c51cfc1f · inbound

Grounding Intelligence in Movement cites this paper.

Grounding Intelligence in Movement Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:23.659772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:23.659772Z digest=sha256:907303da799ac01dee9b82e94a04d52a18d82d63ae34de8739e5249fe70ee55d

Observation ed99d9b5-29e1-4a9f-861f-f8e8401cd53f · inbound

Dyn-O: Building Structured World Models with Object-Centric Representations cites this paper.

Dyn-O: Building Structured World Models with Object-Centric Representations Cosmos World Foundation Model Platform for Physical AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:19:03.343161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:19:03.343161Z digest=sha256:8529ff1640d5156976a9afbb03457fc9367d6e229ff31a8084d6b94398c6e391

Observation b0d8c5d9-569b-4c5d-85fd-171432c7653b · inbound

Critique of World Model cites this paper.

Critique of World Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:42:24.639870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:42:24.639870Z digest=sha256:eea23691dacdb2d2aa3ca7a865128ede6b5cdde93ed6cef032f67932587073bf

Observation f8339469-0154-4c3f-b8f1-fac204824ed7 · inbound

Physics-Grounded Motion Forecasting via Equation Discovery for Trajectory-Guided Image-to-Video Generation cites this paper.

Physics-Grounded Motion Forecasting via Equation Discovery for Trajectory-Guided Image-to-Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T18:58:02.048778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:58:02.048778Z digest=sha256:242b7ecf46dfc874e94d85d1c964e4896946727ff5495ed780921686d3eaeddd

Observation f4a82a89-61d6-4790-b488-fad07dea0890 · inbound

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling cites this paper.

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-19T05:17:06.758003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T05:13:28.767788Z digest=sha256:0ed0e153086f88cc325a5fd791569a9bd666c6243c162dd5db0801fa0912cea6

Observation fa25afe9-7edb-429d-bea0-1266cf94a24a · inbound

MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization cites this paper.

MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:32:54.656073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:32:54.656073Z digest=sha256:7e47b71d8129f28543bd3e9370f5328bc588a871f606a73501fe0278efe47a4a

Observation 2f0afddc-1342-4958-8d1e-b4ec0e7088fc · inbound

Learning human-to-robot handovers through 3D scene reconstruction cites this paper.

Learning human-to-robot handovers through 3D scene reconstruction Cosmos World Foundation Model Platform for Physical AI

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T18:18:28.491430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:18:28.491430Z digest=sha256:b73d42b73d1dd100c9e1469072154211938c6759441fa70f0656c3a3e48dfe49

Observation 01d0693d-1a66-4237-85d4-664f63fa52b0 · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding Cosmos World Foundation Model Platform for Physical AI

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:12.942821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:12.942821Z digest=sha256:3baa04cf61c97413bd60c9a4ee8a2281829a38295ecb6ab43a6f93141295c9b3

Observation 276ef8b4-8d1d-4c8b-8999-1e05e8846ebf · inbound

MobiWorld: World Models for Mobile Wireless Network cites this paper.

MobiWorld: World Models for Mobile Wireless Network Cosmos World Foundation Model Platform for Physical AI

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T17:59:41.742363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:59:41.742363Z digest=sha256:ee1c6cbc5841dda28bc41cbb8155b5f31e001b0bed76bce8de225a1cabbf16c6

Observation 6d13776c-15ee-4bca-b638-36f7d06c1d33 · inbound

Quantize-then-Rectify: Efficient VQ-VAE Training cites this paper.

Quantize-then-Rectify: Efficient VQ-VAE Training Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:34:53.324850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:34:53.324850Z digest=sha256:6d05358c42595cc99b45998157fbe757bbbc80d06981640a78d0318a0a73ff82

Observation 38d58f02-6377-4c88-8266-ca01c57b2ee7 · inbound

Bridging Brains and Machines: A Unified Frontier in Neuroscience, Artificial Intelligence, and Neuromorphic Systems cites this paper.

Bridging Brains and Machines: A Unified Frontier in Neuroscience, Artificial Intelligence, and Neuromorphic Systems Cosmos World Foundation Model Platform for Physical AI

Reference 184

Resolution
verified exact
local_arxiv, observed 2026-05-19T04:42:04.946578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T04:37:33.928616Z digest=sha256:125fc3d7b0335144e1278188d5b7fc7388393ab1dd8a6d73682d084171cf9c43

Observation f724913f-e660-45f2-b610-9c76b74a65e6 · inbound

Diffusion-Based Imaginative Coordination for Bimanual Manipulation cites this paper.

Diffusion-Based Imaginative Coordination for Bimanual Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:18:50.317023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:18:50.317023Z digest=sha256:55279d8466b3b9c161d213f7e1663a1ab8357ae8ee33610637301b012c035238

Observation 6108ab6a-a227-464e-ae20-e8c38503606b · inbound

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models cites this paper.

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:50.279897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:50.279897Z digest=sha256:c38f542dc95712044bfcf3258e7114b5924d9efdc9e2a5ef779f24967e640810

Observation dd05d835-036a-4d73-a28f-484ef0c0855e · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.219696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.219696Z digest=sha256:27a7ca7427107ccffc5324b02733c2c288f0a622c28f926a545857ccc2b548eb

Observation 1b330f06-99a8-433f-a996-18e374455eca · inbound

Yume: An Interactive World Generation Model cites this paper.

Yume: An Interactive World Generation Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:31.709925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:31.709925Z digest=sha256:74f64038a69ca91381d6d20046f182b0dbe27a715474548928f87479becce54e

Observation ffa81ed9-aa70-491f-b8cb-3ff58afef1fb · inbound

Back to the Features: DINO as a Foundation for Video World Models cites this paper.

Back to the Features: DINO as a Foundation for Video World Models Cosmos World Foundation Model Platform for Physical AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:24:19.930972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:24:19.930972Z digest=sha256:9cf4aa4b7e56972f502d43def189e0b736ec00bfbbae35b3c148fad0f81e2daf

Observation 1d93675a-5d02-4460-a2d7-560cce27e24a · inbound

PROVCREATOR: Synthesizing Complex Heterogenous Graphs with Node and Edge Attributes cites this paper.

PROVCREATOR: Synthesizing Complex Heterogenous Graphs with Node and Edge Attributes Cosmos World Foundation Model Platform for Physical AI

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T13:08:47.960825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:08:47.960825Z digest=sha256:f5a2df6302698531d304908e4814f6c1924949d7f9fbf3fa5359eac49761cdf8

Observation 8e83fc43-eece-4238-9e7d-4277fd04c5fb · inbound

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels cites this paper.

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T12:26:02.027355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:26:02.027355Z digest=sha256:6be4ed03a5f072967d61bb2eaaeca0450caf58b8e2de6db3560c23a36696adbb

Observation be4d09c0-3d2e-44f3-adea-cc784d8148f5 · inbound

"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth cites this paper.

"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth Cosmos World Foundation Model Platform for Physical AI

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T05:15:41.372062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:15:41.372062Z digest=sha256:f14bba89020f5ecd3ea7bfd0a6cb370e59e2108414ebd93b6db64b8929b6affb

Observation 023a61aa-dea6-453f-911a-67a3005c28a2 · inbound

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models cites this paper.

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models Cosmos World Foundation Model Platform for Physical AI

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T05:12:32.027479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:12:32.027479Z digest=sha256:8124308a0597385def11b0b0b61b1255680e727b8c187d1de39b72b4600bf160

Observation 88c3cd79-cdf0-4f59-a40d-b06cf58dcd00 · inbound

Qwen-Image Technical Report cites this paper.

Qwen-Image Technical Report Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:29:06.883874Z digest=sha256:197581a8012b432554a0f06f72cf413d949050008a83fa7e9a144e98e6fbffd4

Observation 782a905c-a10e-465f-b814-4353d40c42e0 · inbound

DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning cites this paper.

DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning Cosmos World Foundation Model Platform for Physical AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T23:27:34.279803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:27:34.279803Z digest=sha256:ef732a950992d66d11f6207251eea82e4d3800c07f94c64f737dfa18934057d9

Observation a726a1b0-1197-4cf7-b00a-1ea8bc50e8fb · inbound

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation cites this paper.

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:28:41.941408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T21:28:41.904725Z digest=sha256:3c3f6c91b0d829ccca4065947cd210b8fd5fbfdda4c79f216786e5fc4066a007

Observation 1d82efa8-cf79-4816-a462-52bf55b1e559 · inbound

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing cites this paper.

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing Cosmos World Foundation Model Platform for Physical AI

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:29.081553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:29.081553Z digest=sha256:f0f25f0c0f1f59160bdb07adb7b98bda08b205f447375fb62c7304d9dcc657ba

Observation e558eb34-43f5-4d67-99a8-757ea8f54df1 · inbound

Towards Theoretical Understanding of Transformer Test-Time Computing: Investigation on In-Context Linear Regression cites this paper.

Towards Theoretical Understanding of Transformer Test-Time Computing: Investigation on In-Context Linear Regression Cosmos World Foundation Model Platform for Physical AI

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T22:10:17.411409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:10:17.411409Z digest=sha256:b06b1e501bdcb792a683c179a2337c77500fe5b9516253db5b38d769e5d5fad4

Observation 3103a322-c00a-4cac-8189-0e99aea15310 · inbound

Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges cites this paper.

Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges Cosmos World Foundation Model Platform for Physical AI

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-05T21:02:04.810077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:02:04.810077Z digest=sha256:08a79ea5fd9dc471e09d5b391c57acb33107f12feea75e3917e5cba10415e54c

Observation d64476b5-49d1-42ef-9c3f-9cfe5d576912 · inbound

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation cites this paper.

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:20:18.913406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:20:18.913406Z digest=sha256:328d841580ff2a46301aab3f4a6b8efdc0b36a9e2c1f0b560980aa6de70310d7

Observation 2c4296a2-004f-4801-bd74-854b8d2cd7f1 · inbound

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms cites this paper.

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:43.268081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:17:43.268081Z digest=sha256:6776bfa0f49e78bb874b57df4d5d829a936ba8066b73ba0e7bb8d0fa5307e39a

Observation 9183893f-1457-4eb2-9d8f-cea04e2bebd3 · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.695393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:62ee85f2e91b0efd5a67067f896027eedddd9155683b66600856b545d5395ffb

Observation a8af67dd-e818-4905-9946-171650e69f48 · inbound

Matrix-game 2.0: An open-source real-time and streaming interactive world model cites this paper.

Matrix-game 2.0: An open-source real-time and streaming interactive world model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:36:53.350371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:36:31.044743Z digest=sha256:2e429889b79fe3a4651d8d0dfc3524a169d326bb034d2f02c4cce497a5036449

Observation 1daa2b1b-c9ed-47ff-9302-d7b07c4255a8 · inbound

Precise Action-to-Video Generation Through Visual Action Prompts cites this paper.

Precise Action-to-Video Generation Through Visual Action Prompts Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T19:14:34.361533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:14:34.361533Z digest=sha256:9487afdfa37cfaeec33c92963803cff67a12878927cf8a92ee822e04c80683d4

Observation 0f405814-3c69-4ed3-8836-3670c0dcd93b · inbound

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models cites this paper.

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:11:52.772377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:10:16.645510Z digest=sha256:0b8f5bdb9efe45d40c76927c56c173467801d92a5a6fee5e0045ff2fdd93c4e5

Observation 51a8f960-5ebe-4255-a418-2d0284ce4141 · inbound

LuxDiT: Lighting Estimation with Video Diffusion Transformer cites this paper.

LuxDiT: Lighting Estimation with Video Diffusion Transformer Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:52:02.100197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:52:02.100197Z digest=sha256:fad83e0bac6c3934f812f3ff227a87ba3e016561e6d09fe29bdf7db370147faf

Observation 8d07c304-da9f-4247-b043-8013e0921c41 · inbound

Exploring Autoregressive Vision Foundation Models for Image Compression cites this paper.

Exploring Autoregressive Vision Foundation Models for Image Compression Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:52.435440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:36:52.435440Z digest=sha256:922dbc126f291ccd2433943f06a759a14e4762a9c75b079ed4e7310f8fadffc5

Observation eda46ae9-ced9-464a-b77b-f65d436b6b31 · inbound

3D and 4D World Modeling: A Survey cites this paper.

3D and 4D World Modeling: A Survey Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T06:04:02.866781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:04:02.866781Z digest=sha256:c2f77dee17446297c1f3b47e5ed12c73e2d0380b4fa174e05be6bc47fa3665bc

Observation 5a788f97-59ff-4b64-86ba-6431ab55bc65 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:56:24.406818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:281d345f7eda56865e3f8ae18b43280885715859767f7e0a92aa6df3d11efc31

Observation 64df5fca-a99b-4165-9c84-2164f926d364 · inbound

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation cites this paper.

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation Cosmos World Foundation Model Platform for Physical AI

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:52.135555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:52.135555Z digest=sha256:771494ede5531c68854bec73dee7e5af17f25b2e7c38087fe704bc4ef36c17ca

Observation 6453827b-e935-4b90-b87e-1381952a8a0e · inbound

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models cites this paper.

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:21:23.781604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T13:21:02.738561Z digest=sha256:9e36821fd454c6f2a9f443bfccc6d80e62131971f0c906a921293d651ff4660c

Observation afd23de3-b2cc-4395-9129-b0f4404e7310 · inbound

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization cites this paper.

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T11:25:33.514720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:25:33.514720Z digest=sha256:c4e6f50099a374a43f0926a6d472409a8ca71ed8285d7545dbf1a9d89be7dd8a

Observation cad30db5-7e0d-47a5-aa2c-b4e6a4c5b74d · inbound

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation cites this paper.

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T13:21:35.661138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:19:50.252345Z digest=sha256:b03b9dfe5548263cea208a20ffb386d8eb195027c7b412c216ded46252bec46d

Observation 62134817-4749-42ed-a948-0a647c7ae0ff · inbound

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency cites this paper.

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency Cosmos World Foundation Model Platform for Physical AI

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T08:51:09.155679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:46:16.541104Z digest=sha256:bce6971ff4e331a63a3e95d17c729d95890cd43e470f16a939f7c9e7dee780c6

Observation cb8ef6e5-5063-4c47-b854-74fbad2a2bca · inbound

Ctrl-World: A Controllable Generative World Model for Robot Manipulation cites this paper.

Ctrl-World: A Controllable Generative World Model for Robot Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T01:14:10.550114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T01:14:10.174044Z digest=sha256:ce03ab392415b9f61635bd2e851337d43da771ddfe41a4cf0d46c3bb5101016f

Observation d6c2b98c-7420-43ca-ab21-8a9e0bd5f380 · inbound

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer cites this paper.

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:30.449145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:30.449145Z digest=sha256:63523d3a44b7fbf7e5680a6b1c9010226bdba6294fa2366688084bd180beda77

Observation 83f7ae15-d0ae-425b-a7b1-22c1e2a35cef · inbound

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation cites this paper.

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T03:20:48.732645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T03:20:35.021120Z digest=sha256:70f8aacf3f8e5fd0e7124831026360e5e0c60912f04c4065cc5a09e7bc0c1343

Observation 6f8d0196-3608-45c3-a9f9-cdf0d070daea · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI Cosmos World Foundation Model Platform for Physical AI

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-12T23:01:13.716640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:9134ac06f47d7f3ce6349fe112eb17ab8404f2f9cd96c46c477d162df5fc9f20

Observation 756e2c06-caf1-4485-8a01-b0dcfa310336 · inbound

Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction? cites this paper.

Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction? Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:20:29.138865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T18:17:53.581119Z digest=sha256:935abf287f724f13615506c46ed66c8d3c451d594470c8d826b369ec9f87a776

Observation 5c57004e-ad8f-4ff7-91ab-003a995d78e4 · inbound

Saving Foundation Flow-Matching Priors for Inverse Problems cites this paper.

Saving Foundation Flow-Matching Priors for Inverse Problems Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:40:15.123284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:36:13.408128Z digest=sha256:a3d4a141f249cb71f1bae02e0e36f94f1a879c4ab71e4f1be80d4306e1c46d35

Observation a5fcd7b0-61c9-4d43-856c-684c59195325 · inbound

RynnVLA-002: A Unified Vision-Language-Action and World Model cites this paper.

RynnVLA-002: A Unified Vision-Language-Action and World Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T20:59:49.860461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:59:49.860461Z digest=sha256:101fe648d22aa6cb6ed51afbf0ded80b5d6adb89f001fcf249344632be8b2dc7