Pith. sign in

Paper Citation Record · LEDGER

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

As of 15 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 29 inbound Pith citation observations for arXiv:2412.19505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.19505 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:21:49.227847Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:23:02.078100Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:46:57.095080Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved28
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 992cf9e4-efa8-4255-b75c-ddaa1a364e19 · outbound

This paper cites Model-Based Offline Planning.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Model-Based Offline Planning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.862449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.862449Z digest=sha256:f09c0f94b9bb8e9b6135de9d0395c4c06930055ab8362d53d48606d5f793d8ed

Observation 5ef92d33-62fc-40cb-ae00-cc8a1633112b · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.872608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.872608Z digest=sha256:b7f58cfd6c5e5d691fbe6c5319a96ac6590855fd121b906be86e4c29d7507a58

Observation 5e2d042e-35ff-48ed-9fbe-5effc3a6053a · outbound

This paper cites Language Models are Few-Shot Learners.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Language Models are Few-Shot Learners

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.885106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.885106Z digest=sha256:46dc36df74d8de97bbb48c360447d925cae84f4ba05120cb19d71ab7763f6abe

Observation 7f6c6b40-c619-41ee-a230-40ab15309a8c · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT nuscenes: A multi- modal dataset for autonomous driving

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.294271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:48.892714Z digest=sha256:061391646d8fd30753e3618aa9cf4d65b27d2e3e001ae391ed43a380480cfd76

Observation 56125511-3d54-4d01-a35c-ee61dcd8f350 · outbound

This paper cites NuPlan: A closed-loop ML-based planning benchmark for autonomous vehicles.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT NuPlan: A closed-loop ML-based planning benchmark for autonomous vehicles

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.904049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.904049Z digest=sha256:dec65cac93c19e6af53fd23a1a35fbbadbf268d51debed0cf1baadead245fb87

Observation f84c48d1-dd00-426a-8600-923881c8408d · outbound

This paper cites UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.920295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.920295Z digest=sha256:e2c604aec8bbf6ee26145bd683d5258edee6ae8e6465b97ff4702e436f24b77b

Observation 5b727d2d-5693-4a1d-b599-5065ae919f25 · outbound

This paper cites Uncertainty-aware model-based offline reinforcement learning for automated driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Uncertainty-aware model-based offline reinforcement learning for automated driving

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.272580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:48.927406Z digest=sha256:9cb75c4f7d73db0d391ce02494909b95bf43fad86bec839610c25a4668c9cdd1

Observation 60b72b98-76cb-4212-9fe7-4542f11bd8b8 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Taming transformers for high-resolution image synthesis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.253430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:48.934941Z digest=sha256:c68bfa8e854a6846cfff5f4cdc3171504eec8350e6d8575b665bbf872d83f75c

Observation 4347f416-1b71-4c11-ad51-97f1d229fca8 · outbound

This paper cites Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.940157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.940157Z digest=sha256:e2181d2dbb78865dbd41b08d7159ae5cd39df0a6f07fda97e6ca847f12e45a3c

Observation 5b5eaa11-74a3-4209-a7a1-809a6a59c8fe · outbound

This paper cites World models for autonomous driving: An initial survey.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT World models for autonomous driving: An initial survey

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.235250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:48.950257Z digest=sha256:9e18830eec9c7b2a21e4c79f0b5acfc6d3ee0b5ec959052465f8cf55b90fcf1e

Observation fb62c909-4854-4e03-9d28-c94ccdd7b717 · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Dream to Control: Learning Behaviors by Latent Imagination

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.960535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.960535Z digest=sha256:79c24f4cc5402acad20daa83631d43440e3fb54532d37903043369506322b45b

Observation 3b129139-b1e6-4793-8022-cd87baa51d96 · outbound

This paper cites Mastering Atari with Discrete World Models.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Mastering Atari with Discrete World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.971614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.971614Z digest=sha256:d088f56e396e4a45187a3334403be40005e8d430614dc1204747f9013840c6b8

Observation 47cb3455-5fae-47df-87a5-c5cf99c30e47 · outbound

This paper cites Mastering Diverse Domains through World Models.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Mastering Diverse Domains through World Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.977956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.977956Z digest=sha256:c2505cba17533cfc4b7310c89668fa9b81f46151967f453bdacc787779518ebb

Observation 4f597b95-f6cd-4f37-a728-ac2544093686 · outbound

This paper cites Show me what and tell me how: Video synthesis via multimodal conditioning.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Show me what and tell me how: Video synthesis via multimodal conditioning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.218513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:48.985048Z digest=sha256:14e95b105d0d6685a9836ab5e9a80e8a044c67b5cdf193b9dc0e9d1693bce2a5

Observation af2682c2-60df-4882-a81e-e464c58594dc · outbound

This paper cites Model-Predictive Policy Learning with Uncertainty Regularization for Driving in Dense Traffic.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Model-Predictive Policy Learning with Uncertainty Regularization for Driving in Dense Traffic

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.990647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.990647Z digest=sha256:599760d5e2656889fbc908516ef69f4470b1e69c297451d82909d4b2c23b6599

Observation fdc8f408-ed22-440d-b429-40bccd627f93 · outbound

This paper cites Query-Key Normalization for Transformers.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Query-Key Normalization for Transformers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:48.995839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:48.995839Z digest=sha256:c651e9756d8049286482ba503a5df0cc017c3e41d3ee4fb017dcfc95ab0a7419

Observation f2f0a016-8d43-46a0-a247-a30dfe0ec8c8 · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT GAIA-1: A Generative World Model for Autonomous Driving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.001732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.001732Z digest=sha256:b0d7eb223544dc72aeb72fdd0a11c7a6a4e9e7c5a050aa56a49deeebe17c5d2a

Observation 9c164218-d3bf-464f-8bd1-e5af50137c5b · outbound

This paper cites Image-to-image translation with conditional adversarial net- works.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Image-to-image translation with conditional adversarial net- works

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.007117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.007117Z digest=sha256:c803161fb384e2cfd4e0a3ba5b88292c1bff748dd998718c91b6a7e126ba8c75

Observation f3e2005e-9778-4319-a4de-7751335f7636 · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT ADriver-I: A General World Model for Autonomous Driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.011945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.011945Z digest=sha256:e0f5afdf5a55f43012e2ad6c7106e6b876e16affa0861734bfdf87c599ca7ced

Observation 94589463-b9cc-4b15-a39d-b98587ee2f30 · outbound

This paper cites Drivegan: Towards a controllable high-quality neural simulation.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Drivegan: Towards a controllable high-quality neural simulation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.186676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.016523Z digest=sha256:94382b73c6b487f89667ea899886671830edab12950b0a30486695146f10d149

Observation 0dd486fe-edca-41be-b031-b7417d8a22e4 · outbound

This paper cites Adam: A method for stochastic optimization.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Adam: A method for stochastic optimization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.169142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.025369Z digest=sha256:3a8f62675d6a112989da1aa4ef9b7e2cfb94e7aa76238dcb6b549aa0d26855eb

Observation eb052aa0-1ab7-44bb-8395-385fc26a2b3f · outbound

This paper cites The open im- ages dataset v4: Unified image classification, object detection, and visual relationship detection at scale.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT The open im- ages dataset v4: Unified image classification, object detection, and visual relationship detection at scale

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.151917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.031890Z digest=sha256:e67b51a9a96442e68481f55a54a6c80068c28f6bad4358caccf6d1ba2891fed0

Observation 0203f2ee-0557-4ca3-a464-a9e94489fbe7 · outbound

This paper cites Fast and accurate image super-resolution with deep laplacian pyramid networks.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Fast and accurate image super-resolution with deep laplacian pyramid networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.132811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.040380Z digest=sha256:932fa432218a979ab36b1adc2a035d1d954487d0bb1583a81a540be5e7393f6e

Observation e70f41e2-101b-4fa9-8ade-c5c4feecafbc · outbound

This paper cites A path towards autonomous machine intelli- gence version 0.9.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT A path towards autonomous machine intelli- gence version 0.9

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.047118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.047118Z digest=sha256:5b1d3fa58099ff26935b75136ad5b776535251892bedff121ac5a43ac26c1a09

Observation d7f4f25e-64b5-42e8-8667-a858c733d8fc · outbound

This paper cites Microsoft coco: Common objects in context.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Microsoft coco: Common objects in context

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.101344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.056740Z digest=sha256:c743d3cfd6ff72582211b463e7ae18935b6f2fe9c8b40ca3ac1b33e92db5d66f

Observation b872b970-e03f-48ea-89c4-be0692140a87 · outbound

This paper cites Decoupled Weight Decay Regularization.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Decoupled Weight Decay Regularization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.064106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.064106Z digest=sha256:c5ff1a79dc9f25762b72b716e24b76c3e18a846640a107cf4724afe858ef3fa9

Observation 97c13c3f-72e9-4ca0-aff5-7f3e74e8eb5c · outbound

This paper cites WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.071699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.071699Z digest=sha256:7d9c78bca27b459d31afde0d8470697c86702287f63f9ef2415a143a2289e8d3

Observation f6acfe43-03fe-419f-8fc6-d8550ea53f1b · outbound

This paper cites Improving language understanding by genera- tive pre-training.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Improving language understanding by genera- tive pre-training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.077752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.077711Z digest=sha256:88e4ef1fa27057613e14b602bc7b8422d022c7877c81903b3a0bf5b44bcd7fbe

Observation 245b9cb9-d6be-4a48-821e-3647adc481eb · outbound

This paper cites Language models are unsuper- vised multitask learners.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Language models are unsuper- vised multitask learners

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.057463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.084622Z digest=sha256:b81ddc9dcb440d6f107d1c65fa8ea0366cf198525eec11bd1344e1c202c84721

Observation 0d410798-ec2a-4296-b5dd-a849c578c43a · outbound

This paper cites Learning a Driving Simulator.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Learning a Driving Simulator

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.092154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.092154Z digest=sha256:4e7de618e2d323c6a4dabfc3f4426321f8c455e4769027227adca1bc9ebe52c4

Observation 5813e40e-caf5-4355-9cf2-b440bdb9dbe5 · outbound

This paper cites Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.097879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.097879Z digest=sha256:6ae40a03108d44c274c5c22c27b3417e2e13eefa7920a3ac898bbaed2694de37

Observation 2c9ee85b-d917-4ed7-b298-236e1382c618 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Roformer: Enhanced transformer with rotary position embedding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.105194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.105194Z digest=sha256:4518837979470931a779cbd793b8c477f7d14b008303a2b586e6c59aaa547942

Observation af69369a-118b-42f5-a1ee-7ba031ac8e4c · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.111282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.111282Z digest=sha256:71a2b013e791cfb4dc0e3dabf88234ba37ab7418b9115673e73f828b73ca1ac4

Observation 441621c2-8bf1-4738-869a-8cfc14e27081 · outbound

This paper cites Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.117719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.117719Z digest=sha256:7b9b8901beae3bc99ce8c1142de2b1edbc44e1e3c9fc4834e1c734aefacbee02

Observation 7392c75f-c6b8-447b-8ece-d7b0f381cfb1 · outbound

This paper cites A Good Image Generator Is What You Need for High-Resolution Video Synthesis.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT A Good Image Generator Is What You Need for High-Resolution Video Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.125257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.125257Z digest=sha256:d6a6c72038e60356d5175618c9d64364d9b6b313328811588892089f24b4e7ba

Observation 3119c715-45ee-44a2-89e1-01de51f3409d · outbound

This paper cites Neural discrete representation learning.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Neural discrete representation learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:50.011147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.132585Z digest=sha256:b386c46c3b5c861a16280387ee49fb9316290bb782177c3b9faffd2ea986cb85

Observation 05547562-d343-4383-9d3b-f677ca994290 · outbound

This paper cites DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.139374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.139374Z digest=sha256:d2e0ec3740001b8f2196c4fce55c84b5e3b8c2d8c8204ede989a568f915e324e

Observation 5fd2c5a9-005a-47dd-8634-0aa9cfbc59d6 · outbound

This paper cites Driving into the Future: Multiview Visual Forecasting and Planning with World Model for Au- tonomous Driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Driving into the Future: Multiview Visual Forecasting and Planning with World Model for Au- tonomous Driving

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.995483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.147436Z digest=sha256:a4046d3da14100f618b8b498f2ad4ebd584264b76b4153870b1e38c94d3a6c51

Observation 82deb782-5ead-4184-aaed-f3c1598552e7 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.153030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.153030Z digest=sha256:0db5b731a09e8591e6563644d9db76b53f77eaecc1a6f09c52e47ae1f27ad651

Observation fc217fea-bcd6-4692-8bcc-9955cd889d94 · outbound

This paper cites Daydreamer: World models for physical robot learning.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Daydreamer: World models for physical robot learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.964335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.159691Z digest=sha256:ea61838eed0cc9e57ddf4328d6d925f21af80aa8086105e2bf1a5abc978462ee

Observation fa749b3e-ccb2-47aa-8860-449b51f9b65a · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.164470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.164470Z digest=sha256:30a5ca2315f510514a7de86042f033fb3f750393c41839c70b27b7d516e74183

Observation 22927ece-7dd2-4871-a19c-16689a281071 · outbound

This paper cites Generalized Predictive Model for Autonomous Driving.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Generalized Predictive Model for Autonomous Driving

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.938902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.171293Z digest=sha256:9c67c22c3566b8268b305805824de80319afa9ffed5e0b39ae35e6894227a5d5

Observation 09a8872d-7e12-4e2e-961a-2a8860dfd52b · outbound

This paper cites Vector-quantized Image Modeling with Improved VQGAN.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Vector-quantized Image Modeling with Improved VQGAN

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.178281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.178281Z digest=sha256:bab906da157e36f501fa259f502786eba4d8f357477100e99910a3a60471bea1

Observation ac5acc64-fbbf-492f-9ba1-5114455b2c36 · outbound

This paper cites Generating Videos with Dynamics-aware Implicit Generative Adversarial Networks.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Generating Videos with Dynamics-aware Implicit Generative Adversarial Networks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T00:21:49.184250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:21:49.184250Z digest=sha256:61bdec7ffebb1ffae0565e38a408a95f8817a1251d7bd7ca2cb040de6798d0ac

Observation 9aa6987b-fbd7-41ac-a37a-59132c2f0ad7 · outbound

This paper cites Learning to drive by watching youtube videos: Action-conditioned contrastive policy pretraining.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Learning to drive by watching youtube videos: Action-conditioned contrastive policy pretraining

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.923303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.189846Z digest=sha256:a14e47a045afb4f19136732441529e0c11019094fc6e79c5834559c969e5ab15

Observation 26c8db5e-159f-4ec0-9d42-a0b041bb72bc · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT The unreasonable effectiveness of deep features as a perceptual metric

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.905925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.194430Z digest=sha256:fd8f00dc4a73c19931b64d8ba5c61ffca6d9c03b890f38d2cf90115473ee22c3

Observation 3bfe43c7-265e-4d44-9d85-9eed6368f5af · outbound

This paper cites Movq: Modulating quantized vectors for high-fidelity image generation.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Movq: Modulating quantized vectors for high-fidelity image generation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.888778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.199512Z digest=sha256:1cf425d684561d60b8a75ee195bdd315d72892dfb957c0fd4650753ada1b9378

Observation 0162864f-c5ca-467e-96df-1293cdbce82c · outbound

This paper cites The authors believe that this work has small potential negative impacts.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT The authors believe that this work has small potential negative impacts

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.866768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.204381Z digest=sha256:d74f8465cef857dfec25ab7b00fb91396262d990d38185a5ec631f6c8a135da9

Observation 7f7d23a9-99bb-4d9c-9e3c-c66df082f550 · outbound

This paper cites Due to limited GPU memory, each image is restricted to a resolution of 256 × 512.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT Due to limited GPU memory, each image is restricted to a resolution of 256 × 512

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.844782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.209863Z digest=sha256:8b1cddd2050e5315e81650acb43e831070eac32776c314f826386cd37c2adb2b

Observation d4d847be-07b0-4687-90bb-977aef0c73ec · outbound

This paper cites This dataset is automatically annotated using a state-of-the-art offline perception system.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT This dataset is automatically annotated using a state-of-the-art offline perception system

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.825507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.216091Z digest=sha256:41bc0644ba8a05c8fa2c6679f081a1f993ad1d2caf854a6e1b694f2e69f1e626

Observation a824a3a7-acd9-46c3-98db-f8f99313c4d5 · outbound

This paper cites The images are with size of 256 × 512 and tokenized into 512 tokens.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT The images are with size of 256 × 512 and tokenized into 512 tokens

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:21:49.806904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.222557Z digest=sha256:b8aec53b51d2b9a24142184e0327a38da39bae5a6643862317a6e228c005d237

Observation 6f9985d5-488a-4f52-885a-7de8d1069b53 · outbound

This paper cites w/o AR” and “Ours.

DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT w/o AR” and “Ours

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T00:21:49.784312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T00:21:49.227847Z digest=sha256:cdcde6920aa8c3d93bd409cbf80e98eb28057e55b32889972140b070c1be9ecf

Pith citing papers

Observation 9acad089-1f21-497d-9a75-b60439efa4ba · inbound

ARCON: Advancing Auto-Regressive Continuation for Driving Videos cites this paper.

ARCON: Advancing Auto-Regressive Continuation for Driving Videos DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T22:11:52.549073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:11:52.549073Z digest=sha256:956952dc17ddb09e571957bd1033182cde6dd5c1dd4b36c1b926e5be53b4292a

Observation 7f7ba31e-8639-496b-8631-e77f823c5b90 · inbound

DriveGPT: Scaling Autoregressive Behavior Models for Driving cites this paper.

DriveGPT: Scaling Autoregressive Behavior Models for Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:48.098811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:48.098811Z digest=sha256:3e7cbace86ad950d5612c21398f8be0f01f78d312cc142560c8411bbb561847a

Observation 4727ccc5-6884-4c1b-97e9-7ea6f12e52d5 · inbound

A Survey of World Models for Autonomous Driving cites this paper.

A Survey of World Models for Autonomous Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T18:31:52.474058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:31:52.474058Z digest=sha256:83a1945159f56d9395bc2ce15f2d4d75ab9969312a6229db45c466e998410e1f

Observation 0db9d819-b5bb-41d5-8218-c317dea44932 · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.689364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:531ee00004af1590fffd327df026238a4f857ff49797b738e3bb454f7cf9db25

Observation b78d4b06-e9f4-4de9-acbc-71041aa5bb28 · inbound

GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control cites this paper.

GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:13:43.254703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:13:43.254703Z digest=sha256:8d11d190b4afe5bc6a85d9be1f320fb7d4853fd6c2441804bb23a009506267f4

Observation 82e4a2a7-0850-40ef-9200-49d35d6dd197 · inbound

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model cites this paper.

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:12.935432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:12.935432Z digest=sha256:725ce8e9c6728c685844300cb9d2c3176bc8175edad59a41923b35cab784a49b

Observation bb55e9a0-0669-4394-a68e-ee0fb00a8013 · inbound

Epona: Autoregressive Diffusion World Model for Autonomous Driving cites this paper.

Epona: Autoregressive Diffusion World Model for Autonomous Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:29:49.144575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:29:49.144575Z digest=sha256:f1e14e84053d222e6e38aa3613abe1c9e74022ed54283ccf627975cd563689b2

Observation 65b2649b-0e03-4020-8598-f513f30aee3d · inbound

Depth Anything at Any Condition cites this paper.

Depth Anything at Any Condition DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:52:02.480908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:52:02.480908Z digest=sha256:a6b4d6dca2babfdaea1a21884e23a85b7e4047a25bfbd942d9e2c6555072b98e

Observation b2409850-6265-41af-87aa-044ac4955e59 · inbound

MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization cites this paper.

MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:32:54.718781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:32:54.718781Z digest=sha256:db25aa6c6aca88a7facab9561fda5ad85d16e62ca39ecaf415361a379f3ce241

Observation c4fa05f2-f4f2-4f68-b7ba-c6ffff482753 · inbound

3D and 4D World Modeling: A Survey cites this paper.

3D and 4D World Modeling: A Survey DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-05T06:04:15.535672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:04:15.535672Z digest=sha256:cab003e35b51a4282025338f47e7ad9b840c200f1bb154888b9431f128a735d5

Observation db6ed7d7-8fec-4f55-87cc-dcd438111853 · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:45.347912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:45.347912Z digest=sha256:cfdd098218acdf33a1f2f583e696a45e2f2289d729254b3ae36cde1d5578cd49

Observation 4be31751-4d72-4bd0-8886-6c587a079584 · inbound

OmniNWM: Omniscient Driving Navigation World Models cites this paper.

OmniNWM: Omniscient Driving Navigation World Models DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T08:57:09.462192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:57:09.462192Z digest=sha256:50746aefdddd9d9df87412dcb3a16ef4b7d8008e2451e8d7b8397db5f5531121

Observation d054fc5c-8512-4b6f-8410-dd03d7a4fa39 · inbound

Thinking Ahead: Foresight Intelligence in MLLMs and World Model cites this paper.

Thinking Ahead: Foresight Intelligence in MLLMs and World Model DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T20:41:57.783447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:41:57.783447Z digest=sha256:647a467f9e1b7cadb43ae7a460e5ff6d092eb0fb77d867373846a762d278fc54

Observation 6e6cbc5b-8e79-4619-8c5a-1bd740c0b85f · inbound

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World cites this paper.

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:38:21.152627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T19:34:39.518649Z digest=sha256:7ddf5e6cac4a1e1b117cd1c81c3dc540a4d259982e6e357065478a0b7d29162f

Observation fccf9409-c6f8-4df7-8646-f78bce711583 · inbound

Learning Vision-Language-Action World Models for Autonomous Driving cites this paper.

Learning Vision-Language-Action World Models for Autonomous Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:31:00.739946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:08:10.442655Z digest=sha256:1a9e43b29aaa600b213a31d1a9687bbf6c4e9fc1262b44cffd5bce3e4e4f3f1f

Observation 88dd64db-3028-4efc-94c8-2045588a8501 · inbound

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic cites this paper.

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:26:02.515766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:00:59.662003Z digest=sha256:50c726302eb968d4619c783a40160064f7212dee9ab1576dba8f5c3cccedf9d1

Observation 1523f59a-96a8-4f5a-bc00-c1ea18730cce · inbound

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation cites this paper.

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:29.886059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T05:31:59.676725Z digest=sha256:a6054e41a4835cd253de140e1d567ee1ef32ee524324f90f570fc31d7bcc0596

Observation 23b6a125-2a6d-4deb-a730-0d96c104108d · inbound

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving cites this paper.

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:56:28.345061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T03:48:36.717026Z digest=sha256:12527b82808c5647368bee9f1ee8ef7749eeae47604291c70ac5519bd91e912e

Observation 200f7029-45a7-4d53-903c-1e4a480634f1 · inbound

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving cites this paper.

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:38:00.700467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T21:36:52.396245Z digest=sha256:e46b1aa8539e876a49c4badb5291a12fa7fbda4fed69e0742a86ffa00ecd2a72

Observation 63fe9d34-9637-40c3-81c9-f3479298f508 · inbound

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond cites this paper.

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:04:00.902020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T21:15:12.465970Z digest=sha256:d6e0f5f24c0433dd89c2a695abd01f3d4f71cfa7085be0de5d3ea2e2c12021de

Observation 724b52d1-eaac-4d39-9e19-98173b2f592c · inbound

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends cites this paper.

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 213

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.867181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T17:29:18.513507Z digest=sha256:9958668eed5f8c130fdc557ebda1d6b85d5f7975079c783c599f3013a0ca9f5f

Observation ef710abf-7490-4f43-9c2a-61978eb80f36 · inbound

Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning cites this paper.

Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:46:57.096758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T01:49:25.510681Z digest=sha256:7c12b42b494b1320dcaf3229437fa38fc45f9944dedd32b07272129cfaf363b5

Observation 68ca223f-0f8f-4af5-9abd-6e3415aa9a5d · inbound

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation cites this paper.

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T08:19:04.131379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:19:04.131379Z digest=sha256:ac33e8074ce4d7d25c273fb9783121bb58a62a09ba35b6828f24d16bfcfbb31d

Observation 18800ef9-5faf-4b61-897d-cf4217d7db68 · inbound

OpenLongTail: Generative Scaling of Long-Tail Driving Data cites this paper.

OpenLongTail: Generative Scaling of Long-Tail Driving Data DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T01:26:27.220907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T01:26:27.220907Z digest=sha256:c0d272ee7cca74e16ff1fd96187ff1e1af77719aef5fc393665fc7f4b4794b9c

Observation 7d925026-4feb-4bd7-a003-b0f70351c1f7 · inbound

Instant NuRec: Feed-Forward 3D Gaussian Reconstruction for Driving Scene Simulation cites this paper.

Instant NuRec: Feed-Forward 3D Gaussian Reconstruction for Driving Scene Simulation DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T02:55:46.103411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:55:46.103411Z digest=sha256:e45cf25cfdd6cba569a572d9903c8df757a05372db45428471617c45e969b6f0

Observation 8bd34162-89f3-46e2-a849-b45a45eed1cb · inbound

Orbis 2: A Hierarchical World Model for Driving cites this paper.

Orbis 2: A Hierarchical World Model for Driving DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T22:05:02.737118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:05:02.737118Z digest=sha256:87ddaa7724af8c10e5823136962b83a4ed5ffdb90d6e161006b3f7d8a42d2aeb

Observation a95fcd14-a47f-445b-b5dc-410b0a3670d1 · inbound

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features cites this paper.

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T18:25:51.323654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:25:51.323654Z digest=sha256:db7cd525c8586de6c19521e6226bfc401000251c7b343c579748cf64a04af10a

Observation 90a528d6-e066-47c5-a81c-0a10ad410c84 · inbound

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting cites this paper.

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T00:28:52.210464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:28:52.210464Z digest=sha256:b665b201d165d7bfbc98ddf80370ba0b4df353fc231003ed0d423b889c30718b

Observation 671d015b-a27f-4d65-9d33-36c446c29423 · inbound

PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives cites this paper.

PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-14T04:23:02.078100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:23:02.078100Z digest=sha256:0d08425b9bd53035bb32cd6947544daca829c1475f6c9968355226b90f3e6486