Pith. sign in

Paper Citation Record · LEDGER

Physical Informed Driving World Model

As of 16 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 6 inbound Pith citation observations for arXiv:2412.08410.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08410 v2

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:58:58.492878Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:31:52.984933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T17:51:54.796433Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b85c0c5-43be-41c7-b032-495d5ca1f6dd · outbound

This paper cites Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation.

Physical Informed Driving World Model Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.286202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.286202Z digest=sha256:9ecbaf4dfdbf189783be0e0fcd024afcb533d7bde8ff810bcea62c86492f0c6c

Observation 75343fff-0023-4f32-94c7-0de5bcc45895 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Physical Informed Driving World Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.291656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.291656Z digest=sha256:754a48ab9355d704c56145d9c90100691d6a233f1b27de1f31b4ce528efd1445

Observation 5ea56432-d879-4eff-9285-8c966f4918e2 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

Physical Informed Driving World Model Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.124634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.296448Z digest=sha256:2967f143f7d46831ff4b37b887a80fc4baa254956d0338b596b6d8aeea872581

Observation 987239e9-29b5-4c09-826a-43275ec99950 · outbound

This paper cites Virtual KITTI 2.

Physical Informed Driving World Model Virtual KITTI 2

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.301381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.301381Z digest=sha256:abdbb6d5e6d4fda48745a9c5f297104db69db5cb32f3baf5ad0603ffcca2190e

Observation a0521e8f-14e0-49c9-8f01-6c981d7c9f8d · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

Physical Informed Driving World Model nuscenes: A multi- modal dataset for autonomous driving

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.111242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.306570Z digest=sha256:6aae04824e5277b11bf93350122cd4a5302abcb579fd86a12046c3566bd4793d

Observation 1d4b8f78-ef90-42aa-bc7f-c1477707f121 · outbound

This paper cites CARLA: An open urban driving simulator.

Physical Informed Driving World Model CARLA: An open urban driving simulator

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.097512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.311038Z digest=sha256:e4c485f09644b84809a87332249d44ddbebd7eab622b4e17092f59a7583e949c

Observation b98e65e1-a64d-4f98-b560-91243918a682 · outbound

This paper cites Magicdrive: Street view generation with diverse 3d geometry control.

Physical Informed Driving World Model Magicdrive: Street view generation with diverse 3d geometry control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.082815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.315655Z digest=sha256:3dadb84930c85e2c6083b1b08b551dfd7a9dc12170ac8811eadbaf6c182c3aa0

Observation 03fda73b-3a78-451d-aa1d-89884b150742 · outbound

This paper cites Gan-based virtual- to-real image translation for urban scene semantic segmen- tation.

Physical Informed Driving World Model Gan-based virtual- to-real image translation for urban scene semantic segmen- tation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.069599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.319643Z digest=sha256:b716a51690741cb1cb76210d79b31543bcb1a7939d19c08cb8bfbaff4fc17569

Observation f02bec2d-0df7-47cb-93b1-e0cb677f3d8b · outbound

This paper cites Learning video rep- resentations of human motion from synthetic data.

Physical Informed Driving World Model Learning video rep- resentations of human motion from synthetic data

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.055054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.324204Z digest=sha256:ea4a107b51f2d235f7535e55fa62138fa96e4bdcd55171c08978df6c87f28c79

Observation 18ab6450-a9f5-4acb-9073-c03814939a8e · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Physical Informed Driving World Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.328661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.328661Z digest=sha256:18dffd06a30a3a90331845ebe77cf447dbda730c72d797856f5c45237d31cc66

Observation ab10fe61-4f49-4b5e-bdc9-be30527bf3b3 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Physical Informed Driving World Model Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.332946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.332946Z digest=sha256:c1f439c585c1bf189fe12bd5ebb43c687d5719632fddaeec2de43d365c5e6556

Observation 570188ec-28e4-48cd-a558-2dc92374c107 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

Physical Informed Driving World Model Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.042253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.337261Z digest=sha256:fb930618d8c3560e981c1d6e103f00743c6698f306b8cc5e3f5ad34c01781e3b

Observation f08eac2d-9bd4-47e3-9412-e15a8629f442 · outbound

This paper cites Classifier-free diffusion guidance.

Physical Informed Driving World Model Classifier-free diffusion guidance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.028635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.341410Z digest=sha256:6f8246980ab2960e5eb7c2e649fb256938c4c1ee01998de6739723eb4ff49eef

Observation 1fd1a30d-f4a4-4914-a1ab-7d9d60b7dff3 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Physical Informed Driving World Model Denoising diffu- sion probabilistic models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.345291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.345291Z digest=sha256:84ef4cbfea5cf1684c2bc664fafabe0e2fd30314130571234c930f0cdb215c3c

Observation cf78aa91-1a4e-4d1e-9d16-020a886fe525 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Physical Informed Driving World Model Imagen Video: High Definition Video Generation with Diffusion Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.349163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.349163Z digest=sha256:af6b6aa6924e4b181607f192397c123372529d53f466977f1e569a4836424435

Observation e682f2e9-29f4-4a09-b41d-33420f9fc021 · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

Physical Informed Driving World Model ADriver-I: A General World Model for Autonomous Driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.353543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.353543Z digest=sha256:8d2ed162209a08c272ab03028a0bdc33b5bdf9a4c81a53a63e26dfc6d47a7ea9

Observation 3732d6b9-4e17-47fa-b897-4dabdc7a8d15 · outbound

This paper cites Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023.

Physical Informed Driving World Model Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.007453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.357846Z digest=sha256:1cbdcb09c85de999efc612c4e61890944b89866197cc525b5e44c589af7ef53c

Observation 756cf46a-6078-4185-9a99-6265064ede36 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Physical Informed Driving World Model Gligen: Open-set grounded text-to-image generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.994439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.361557Z digest=sha256:ea08883dd401babcda2d749a77f5e5e60b9505372abff0ffb03750c97355e5ed

Observation 935c05c4-bc57-4ca6-ae59-9fc8f3096eb2 · outbound

This paper cites Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024.

Physical Informed Driving World Model Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.980598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.365871Z digest=sha256:7f5d6f00388bb127f7dfff0835b77f7c7eb438ad7e548ca92bea4635204f4315

Observation fc67a0b2-7b4b-4fc6-8331-72b88f1c2bf3 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

Physical Informed Driving World Model Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.967463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.369874Z digest=sha256:efc154b7fcab10bc54c9f7db9812f45f4aebf13c071e46f509582cdb902e37be

Observation fe44afa6-d804-46f2-adda-0e2d06dfbf58 · outbound

This paper cites Scalable diffusion models with transformers.

Physical Informed Driving World Model Scalable diffusion models with transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.373643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.373643Z digest=sha256:c9d8fe3ae9507b824d3652ae4dcdb923515b34892fae183649c37a3aa8bfddd8

Observation 7c6fbc32-7cfd-4b7d-9b09-51467ffa4ebb · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Physical Informed Driving World Model SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.377429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.377429Z digest=sha256:b360d7b810b066ee463aa1c992dd16a8ccfd4400dcf983ccbde54b301a672424

Observation 000271a3-7cdf-48c8-8fd4-d4982ba29049 · outbound

This paper cites Exploring the limits of transfer learning with a uni- fied text-to-text transformer.

Physical Informed Driving World Model Exploring the limits of transfer learning with a uni- fied text-to-text transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.945335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.381250Z digest=sha256:b71614e963f3aede254df27e34006cf83bb99b87cc4a7ba398305c8c84feb5d8

Observation 91ff9ac6-cfd6-4cd7-b9cd-09546e831195 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Physical Informed Driving World Model Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.384644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.384644Z digest=sha256:2aebf2dc9bb82913860b7d4c52a9e2917d83bc776cd0dcce518343598a17426b

Observation a36f5579-6f58-457f-a3e9-6b94fb7e4cb1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Physical Informed Driving World Model High-resolution image synthesis with latent diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.388465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.388465Z digest=sha256:a17d9291609f45c3c82a7cafdb68ee27d154c92fec0f600ceef78c869f87c68f

Observation 148fed5e-2a00-4fef-994d-477f64264158 · outbound

This paper cites The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes.

Physical Informed Driving World Model The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.924702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.391900Z digest=sha256:0b4515ed1a68f2b1644e867e6fff39eb2602672b7de3b38449a35c2262d3cc52

Observation 4482455a-6548-4333-9044-ee7eb77ee734 · outbound

This paper cites Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016.

Physical Informed Driving World Model Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.911860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.396122Z digest=sha256:b1a5b73e2ebbb16345580a910febd0bab0bef6fd5bebe9489f9fa474cd34e653

Observation ba0116da-ad7c-4838-918a-474525c11b6a · outbound

This paper cites Learning 12 from simulated and unsupervised images through adversarial training.

Physical Informed Driving World Model Learning 12 from simulated and unsupervised images through adversarial training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.899709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.399608Z digest=sha256:52d89ee1396c237f02f9b7e2860ffbebcf085857999d2b8a0001a49f1d12afb9

Observation 29c9a3a0-48ed-415d-b8bb-71af4b71e4b5 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Physical Informed Driving World Model Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.403342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.403342Z digest=sha256:17712a64658fd33447b428371f827e725f4261842473ca4a961e1463948b6f0e

Observation f4eafede-498a-4c98-ab1a-c614d33a7628 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Physical Informed Driving World Model Score-based generative modeling through stochastic differential equa- tions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.407998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.407998Z digest=sha256:ac8a10f7fc847e0dda289369a8fd3625edd1696e9dc71f76b562b49ed7400c25

Observation 76aa1697-7fa0-4fbc-a1c7-4b5ac5bd28c3 · outbound

This paper cites Street-View Image Generation from a Bird's-Eye View Layout.

Physical Informed Driving World Model Street-View Image Generation from a Bird's-Eye View Layout

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.411831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.411831Z digest=sha256:5247f1919ef7e081e4b6fc9194cc9bca0e7107eab018f398c3e3b1876a762254

Observation 043a437b-8ece-492d-9f88-e875ab79b733 · outbound

This paper cites MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion.

Physical Informed Driving World Model MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.416036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.416036Z digest=sha256:dee01119f3d7e5cfbdb29289d26b859117c9aba55014bdf35f2aaa6d3b76fb94

Observation 90dfeb38-14e8-42b5-8ec6-0ab6ae84e8cf · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Physical Informed Driving World Model Domain randomization for transferring deep neural networks from simulation to the real world

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.419800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.419800Z digest=sha256:18588b5018a9c51077442170968d0f4540193183ca5af77f5c9764e249991851

Observation 00f8a4eb-a254-474f-847c-7390af8597c4 · outbound

This paper cites Consistent view synthesis with pose-guided diffusion models.

Physical Informed Driving World Model Consistent view synthesis with pose-guided diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.871909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.423768Z digest=sha256:2ef1932de3ec24baecd88f040484098681572030760e7839913d63c13d6335df

Observation 9fd47125-8eb4-4303-9e66-2585f4993afe · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Physical Informed Driving World Model Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.427881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.427881Z digest=sha256:8c8e0a74cee0b3a79d72471bba8717c4f59f511119520fa8362618cda6123a73

Observation 49bfcf36-bfed-4090-9e16-ab534fe517e0 · outbound

This paper cites Attention is all you need.

Physical Informed Driving World Model Attention is all you need

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.859969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.432062Z digest=sha256:6a2e3a65dddbb67a4ea729ff0cc4e85573fee9b2a502d5f090b716d552722fa3

Observation 4e491b20-097c-4473-b511-1ea4fa851230 · outbound

This paper cites Exploring object-centric temporal modeling for efficient multi-view 3d object detection.

Physical Informed Driving World Model Exploring object-centric temporal modeling for efficient multi-view 3d object detection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.847896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.436030Z digest=sha256:458f2a4c1cd8ff852a3fe4cea9f7ad92119e429ddddb352bedbe62d799972011

Observation 430b276b-b838-4d2e-85ad-6320e44c9131 · outbound

This paper cites Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023.

Physical Informed Driving World Model Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.834623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.439999Z digest=sha256:ef65878a955095e22ba13d2dd8146b4f80a1c840d8b25c73047f4797cc639928

Observation 892263d8-6b64-47c9-a70a-6e8de63377f7 · outbound

This paper cites DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving.

Physical Informed Driving World Model DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.444090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.444090Z digest=sha256:7701adb3ca639b456f607537932ddf67d9cf5967a746be833022c10ec43049f7

Observation d2a94574-c5cc-4af1-8eea-8bde469010f2 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023.

Physical Informed Driving World Model Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.821545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.447595Z digest=sha256:822bcdafa3b6d9303d3616701d1ad2325684c9e699ae6165fd12ece8a3387569

Observation b72f2ca1-c71f-4fd6-923e-f1fbfb03b7b2 · outbound

This paper cites Panacea: Panoramic and controllable video generation for autonomous driving.

Physical Informed Driving World Model Panacea: Panoramic and controllable video generation for autonomous driving

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.808701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.451136Z digest=sha256:20c61ee9c989c300a03ce8d17aee19cecd4003887173d4915ae02ec5038fcf3a

Observation 9d3a015a-af50-402d-b044-4a7af83a78b3 · outbound

This paper cites BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout.

Physical Informed Driving World Model BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.455339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.455339Z digest=sha256:0a90221f8f22ce09377213c128d7479680f004bea74de4d48b00bc69446063ce

Observation eb4378b4-cb85-42ff-b4b1-7f2dbfaecf36 · outbound

This paper cites Magvit: Masked generative video transformer.

Physical Informed Driving World Model Magvit: Masked generative video transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.796348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.459666Z digest=sha256:19a29e4d1b783926edf96890efcbc1392f6649945e52d6048d6c9ba969a53c07

Observation 88a19d93-3bd4-4807-9d47-24f40d1aadb2 · outbound

This paper cites Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024.

Physical Informed Driving World Model Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.783763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.463513Z digest=sha256:99700f14081e519783b3d56ebfb6bf29c53b20d2fade3890a1a2d8a9acbd01d9

Observation 85472ba3-cd49-4454-86d6-6f450f65df25 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Physical Informed Driving World Model Adding conditional control to text-to-image diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.467551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.467551Z digest=sha256:5f99b521d1eefc577dd6b20cc18ac7387fc6ccc44d0caa31635c642e6bdfeb4b

Observation d760a2ca-8a8a-463e-999a-4977e19d6fc3 · outbound

This paper cites Collaborative and adversarial network for unsupervised do- main adaptation.

Physical Informed Driving World Model Collaborative and adversarial network for unsupervised do- main adaptation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.764198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.471434Z digest=sha256:7cbf26027553cabb477710d10b787a6ae39087f4cb7f72c53b1d5349cf6fd0e7

Observation 4b2863e2-328f-4295-9fad-e79343f57c3b · outbound

This paper cites DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation.

Physical Informed Driving World Model DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.475852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.475852Z digest=sha256:7fdb786514d63d295ff75ac9d29c868c005cf4e2400066206a2e2d8657a1b0e2

Observation 3768ceed-6242-4fd5-851f-519b592fb326 · outbound

This paper cites GenAD: Generative End-to-End Autonomous Driving.

Physical Informed Driving World Model GenAD: Generative End-to-End Autonomous Driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.480102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.480102Z digest=sha256:4ce113d0d5fe620aa0f71ae3b286906e7b85dc8575cb337fc2dd86b8f3a76137

Observation bd638ff6-3a72-44aa-a7ac-bd090241bf8c · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

Physical Informed Driving World Model Open-sora: Democratizing efficient video production for all, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.752042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.484128Z digest=sha256:d2b140859a5b036e834aaa344597f527d829f9a5687051ea64b32f07a7bfc249

Observation 16d0aac3-1ead-4981-b815-6823e7a80c92 · outbound

This paper cites patchified.

Physical Informed Driving World Model patchified

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.738519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.488155Z digest=sha256:460097273f5b8655ed307f595e4694063df237034abaaec86c99d5ccc81c5356

Observation 35f79062-b189-47a1-960b-f48e0bd19b09 · outbound

This paper cites Simulation-to-Real Visual Translation.

Physical Informed Driving World Model Simulation-to-Real Visual Translation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.724458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T17:58:58.492878Z digest=sha256:65df67e77b7462e1acfaffc0fb9d1893f5d7b5c7bc576b2281df66ce5ad1b733

Pith citing papers

Observation 41ebb73c-9c33-49a7-98e5-dd8f28167298 · inbound

A Survey of World Models for Autonomous Driving cites this paper.

A Survey of World Models for Autonomous Driving Physical Informed Driving World Model

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-10T18:31:52.984933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:31:52.984933Z digest=sha256:aefbf9c8a5838c9198e571a7cdef8d30750e070392433f09c4bb9e1a0589a578

Observation 3afa36f0-f4de-441e-aba1-ea4e330d1edf · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment Physical Informed Driving World Model

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.799714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:88a7e0e2a17738f3594285567c7721636861c480bc13648b440af8372976e8bb

Observation 24f0798c-92af-4e58-bc33-13b694187851 · inbound

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities cites this paper.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Physical Informed Driving World Model

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.587203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.587203Z digest=sha256:278713961c6c488e63fb39e957f8a1fa02a3fedb4d31450cae5f70cc202ecdfe

Observation 2c5cbbd6-3523-43cc-a13a-25dbb91d6b98 · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI Physical Informed Driving World Model

Reference 214

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:54.374022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:54.374022Z digest=sha256:d25fdf6f90e950ab9a6d28f66e468c8cee3c9175d1924e235b9948545a3f8c56

Observation a03e8934-99c5-43f2-bf1e-d4066f579245 · inbound

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World cites this paper.

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World Physical Informed Driving World Model

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-03T17:02:42.473873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:02:42.473873Z digest=sha256:92313933f906b5df9520869a920a49415f87010262a8e991b3b1ecbb232d3c8d

Observation dabfd1d9-faf5-4c98-aa6b-3507fe97cfca · inbound

MultiWorld: Scalable Multi-Agent Multi-View Video World Models cites this paper.

MultiWorld: Scalable Multi-Agent Multi-View Video World Models Physical Informed Driving World Model

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:09:08.093011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T05:06:11.514186Z digest=sha256:736faee4459e3c3b58f74d316e4b3dfb1af9f6ecdab419f5f7f25c472d07166e