Pith. sign in

Paper Citation Record · LEDGER

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth

As of 17 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2505.01729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01729 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:18:26.321884Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T06:04:18.694737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d0b9f937-8cb9-4772-a7ea-d4bf2c897223 · outbound

This paper cites Planning-oriented autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Planning-oriented autonomous driving,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:24.875269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:24.875269Z digest=sha256:0449a605994c1393643ccfb9b20e01adaa641b28c02f71b19603c9ee66a1fbde

Observation 2220247f-e894-4f63-bb95-8a44be236b32 · outbound

This paper cites Unipad: A uni- versal pre-training paradigm for autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unipad: A uni- versal pre-training paradigm for autonomous driving,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.544765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.928459Z digest=sha256:12f7b93365c35522a130d5a783f174f4839e57c425b5b666968be1b8c8c032f4

Observation 68c80e70-6c3f-429a-b138-9880249c42dd · outbound

This paper cites P-mapnet: Far-seeing map gen- erator enhanced by both sdmap and hdmap priors,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth P-mapnet: Far-seeing map gen- erator enhanced by both sdmap and hdmap priors,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.528086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.966548Z digest=sha256:240f6d08e2ae86f15f4f83552f64d23b9a3da7b96f0b849337600451c85365ef

Observation 6c20a22a-232d-4147-b1ee-2abfeb375ec6 · outbound

This paper cites Lode: Locally conditioned eikonal implicit scene completion from sparse lidar,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Lode: Locally conditioned eikonal implicit scene completion from sparse lidar,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.509244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.971670Z digest=sha256:c3b1c508483e5aba3537bf08b67d76f8beaea026a4aa45568a688ae484edcd63

Observation b2e106d1-d8c0-430a-9c64-81dd0f82538a · outbound

This paper cites Monoocc: Digging into monocular semantic occupancy prediction,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Monoocc: Digging into monocular semantic occupancy prediction,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.491967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.977088Z digest=sha256:19de8335357befa3c2120f822baef5f313da491fba7457648d059108794a0563

Observation d36a9073-2cbf-41c5-bc39-06c20e1dab11 · outbound

This paper cites Unsupervised road anomaly detection with language anchors,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsupervised road anomaly detection with language anchors,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.263748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.983137Z digest=sha256:488ecf8a4b8e1767d78bf8b2d98bac19db845bd4a4fe630ff8ee65da1a8ed392

Observation c7a091f1-d52c-444e-94bb-a7e9b2dcdc88 · outbound

This paper cites Tod3cap: Towards 3d dense captioning in outdoor scenes,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Tod3cap: Towards 3d dense captioning in outdoor scenes,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:28.054631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.991133Z digest=sha256:fc753e7885569ed378e9d48fda98e6ce66bbf2606ea8f6fd86081132d09a6399

Observation 52dbd232-e09f-494f-98e0-8af719dc224e · outbound

This paper cites Uniscene: Unified occupancy-centric driving scene generation,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Uniscene: Unified occupancy-centric driving scene generation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.889091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:24.996600Z digest=sha256:847a6913ef43712121fa4642914ed540f7a3f9e9ec52848d5a80ef2e7548e608

Observation 68e2975e-89a0-4b18-95e9-964da30eaf6a · outbound

This paper cites Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.001865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.001865Z digest=sha256:1c31fa437ac53b923b8040fe41f2c8268ad243719c8f457f6cbdcf446617319f

Observation abb38e74-21a6-4d08-be40-84386559d30e · outbound

This paper cites Model- based imitation learning for urban driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Model- based imitation learning for urban driving,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.008003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.008003Z digest=sha256:f9104ae2d6730a2f3f1c3eed391f8b01ee0583a956c341fc56e98779782f1813

Observation f18702c0-e0e4-4db0-9b9f-8ce6000e6ea5 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Driving into the future: Multiview visual forecasting and planning with world model for autonomous driving,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.861457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.025704Z digest=sha256:13130006727e3a63b5d2c83c384db9fcd8a2c80e59f71857dbf1d1e0d3d76cc8

Observation cdb84a7b-4e80-45a9-84fa-a30c26b82172 · outbound

This paper cites DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.082648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.082648Z digest=sha256:adcdb425a3bda922b1a1576328d1551a78709ef56438ddcb6f89ed02befe509b

Observation c1fa35cb-ef5a-4ff6-a98c-0193e8f87431 · outbound

This paper cites BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.142175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.142175Z digest=sha256:d59cbf67f0ff5022c045477e0adf5f1569253caa7ec3e25303f1a5bfb9764bf3

Observation 097fa55c-425b-46f8-9928-803a641d404c · outbound

This paper cites MagicDrive: Street View Generation with Diverse 3D Geometry Control.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth MagicDrive: Street View Generation with Diverse 3D Geometry Control

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.242114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.242114Z digest=sha256:f66da2dfd3cfa0610e16419ee6c7b4306159b71c428242b18bce72bf107fb320

Observation abc94007-f54a-4d9e-8401-ada4d12d25a7 · outbound

This paper cites Panacea: Panoramic and controllable video generation for au- tonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Panacea: Panoramic and controllable video generation for au- tonomous driving,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.845330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.247901Z digest=sha256:3aaca21a5528fbf1e7dbf265638ee0d8eaa041c4d025285f47dac1b218776003

Observation 5609a65b-0af6-4c4f-8509-4b033e71728a · outbound

This paper cites SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.252453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.252453Z digest=sha256:a61aa3f8599b66cad5854d97ac2a30645781ae37a28bcf87edb230e5ccf03d40

Observation c1be9b42-d98b-44bf-9e88-dc8ebc4b74d7 · outbound

This paper cites Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.257948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.257948Z digest=sha256:62980de7faa1dc4be9054de419a4deb21b7a9ae73e06cbd3d75457fcaf173637

Observation c3842a00-66fb-4417-9f58-e3985250eddd · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Motionctrl: A unified and flexible motion controller for video generation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.774760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.263395Z digest=sha256:261c17e093195bed57cef7d5adb1b61048a5edc96129c129dd0441750f876d65

Observation 07c9566a-3f23-4ff1-a529-a4b70dad7da3 · outbound

This paper cites Direct-a-video: Customized video generation with user-directed camera movement and object motion,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Direct-a-video: Customized video generation with user-directed camera movement and object motion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.557248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.268270Z digest=sha256:e4ce354203066a8cca457f6ee0a16194a8132705d6b7485eb48dd7bb9dee15ef

Observation 0de58162-7554-4a7d-a295-0e5c4f8b9961 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.273255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.273255Z digest=sha256:10f3bedcfe498401ce55d115b774a4d49b4bf6969ec4d1e404ff52803aa9d222

Observation f02c3dce-ee33-42b4-a942-a49f8d1bb818 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.280565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.280565Z digest=sha256:e40aec14b7ba3e148ad2c6cf4499a95a610ba1d0dcd206669959f4b2012598de

Observation a5994da4-ed3e-4729-b4d7-7b189547e168 · outbound

This paper cites Unsu- pervised learning of depth and ego-motion from video,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsu- pervised learning of depth and ego-motion from video,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.539900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.341750Z digest=sha256:8c12987a4c4b88bf81a9f0bc5bb84cdee075cbf1fdfbe5730e406a7eff741135

Observation 24d5a49e-0a34-4f66-91e4-818ce0c28a73 · outbound

This paper cites Unsuper- vised learning of depth and ego-motion from monocular video using 3d geometric constraints,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsuper- vised learning of depth and ego-motion from monocular video using 3d geometric constraints,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.523519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.460878Z digest=sha256:7b6e5f47de5faf1db06de1acb8ac6ef65d93581530a12561e8b36f6d81cf6ec4

Observation 40106678-e772-42ff-aee6-bb2aef04f8b4 · outbound

This paper cites Unsupervised scale-consistent depth and ego-motion learning from monocular video,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsupervised scale-consistent depth and ego-motion learning from monocular video,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.506661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.535748Z digest=sha256:c147445ccf1afc89071e774b1b4951aad0dd4f064b3e36fcf007160924e6e05f

Observation f133852c-53b1-464e-bed8-a6bbb148e17f · outbound

This paper cites Unsuper- vised cnn for single view depth estimation: Geometry to the rescue,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsuper- vised cnn for single view depth estimation: Geometry to the rescue,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.490127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.541618Z digest=sha256:ad6d01c15831a79397ef6218c77a837b14e05db87de9c5b0a866e616593c5be7

Observation 63136b1b-e3a6-43c7-9b26-a05fe4c8e6d7 · outbound

This paper cites Unsupervised scale-consistent depth learning from video,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Unsupervised scale-consistent depth learning from video,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.419828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.547835Z digest=sha256:7611256fca7df8b5dff178ff4d872b001e525930d7864e395671d34a2a359c26

Observation 7ac30406-5f83-47ee-880c-ff72da263438 · outbound

This paper cites A Survey of World Models for Autonomous Driving.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth A Survey of World Models for Autonomous Driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.552710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.552710Z digest=sha256:2465b46a9530cdd1bb36ba1c8fc2d363806dca831b0f06d17ec8c7edbf39b704

Observation 93bcaf3d-bd91-4105-a262-c5368cd4ee2e · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth End-to-end autonomous driving: Challenges and frontiers,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.648613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.648613Z digest=sha256:2fcbb5441a92a30ef68d311eae87cdd53a0bc760f6a862d4b322363af8ac3c80

Observation 4c292dd5-dc3a-479b-9e32-bd51206cdc9f · outbound

This paper cites DOME: Taming Diffusion Model into High-Fidelity Controllable Occupancy World Model.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth DOME: Taming Diffusion Model into High-Fidelity Controllable Occupancy World Model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.692226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.692226Z digest=sha256:b1682654e69f9b12fe1fc63c214ec3e48a9450319f026caf2df708933180a8d5

Observation 85990901-eb2a-46ef-ae95-b0b67864392a · outbound

This paper cites World4Drive: End-to-End Autonomous Driving via Intention-aware Physical Latent World Model.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth World4Drive: End-to-End Autonomous Driving via Intention-aware Physical Latent World Model

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.774194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.774194Z digest=sha256:b34082e6adebd1a9ce0e82dab88535902a5e2dd83c4cab172666e379d494bc86

Observation c0d0452f-72f8-4faf-ba04-4a4c4bdba3df · outbound

This paper cites Int2: Interactive trajectory pre- diction at intersections,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Int2: Interactive trajectory pre- diction at intersections,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.101830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.789180Z digest=sha256:6c33b1819f97cbecc4df046c2bd86907c57c1624465fd131e3ab4d0b143a5f07

Observation 33e6ab5d-973c-488e-b263-3129b6f3d63b · outbound

This paper cites Transfusion: Robust lidar-camera fusion for 3d object detection with transformers,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Transfusion: Robust lidar-camera fusion for 3d object detection with transformers,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.086036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.795222Z digest=sha256:6d9ff6d24fed99a8b03425c677079957582ad8667e7f6ad951d857664bdd0386

Observation 082817f6-a619-4e8f-8449-246d7c3dbb2e · outbound

This paper cites Pivotnet: Vectorized pivot learning for end-to-end hd map con- struction,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Pivotnet: Vectorized pivot learning for end-to-end hd map con- struction,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:27.067809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.799577Z digest=sha256:7bfc2862b70414718ea90490050b473c631188f55952e3a577d03c5fcf7ab74d

Observation 51bf9803-1309-42c7-b4d9-7533474f45ea · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth GAIA-1: A Generative World Model for Autonomous Driving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.805171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.805171Z digest=sha256:1e2e28a19cee52ee65642f1134e88a6dc51c3b6834aed209f24436d29fcecec5

Observation acc0e849-302a-48e8-b94a-b8ea27adac97 · outbound

This paper cites Drivedreamer: Towards real-world-drive world models for autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Drivedreamer: Towards real-world-drive world models for autonomous driving,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:26.880196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.810356Z digest=sha256:162da0a87f898f302ca82a456f0bda3ed697089e95c3e0002e43bae48c7c7528

Observation ae15b249-dd66-4f6e-8c6c-842c24b81e53 · outbound

This paper cites WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.814646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.814646Z digest=sha256:c01587aaa26a5dd57b155a00b983d91a97d04167526ea778aec8ae496a13437b

Observation f22adbba-0cc8-4771-983e-2568c74624ec · outbound

This paper cites MUVO: A Multimodal Generative World Model for Autonomous Driving with Geometric Representations.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth MUVO: A Multimodal Generative World Model for Autonomous Driving with Geometric Representations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:25.831067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:25.831067Z digest=sha256:bacbca9a9cdc5dc36ae0a959ecaffea8ee98a6567586bb3c773099e29504626b

Observation 119e7ef6-ebb0-4b3a-aeb3-9a109fedaca5 · outbound

This paper cites Occworld: Learning a 3d occupancy world model for autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Occworld: Learning a 3d occupancy world model for autonomous driving,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:26.816905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:25.931424Z digest=sha256:ae218c014a3a2879ce6c9510ed30d7f5e5fcfeefc49d928bd1b82657d74e9e22

Observation 3b6152e7-6f08-4166-8575-878239903a5c · outbound

This paper cites Motionbooth: Motion-aware customized text-to-video generation,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Motionbooth: Motion-aware customized text-to-video generation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:26.798749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:26.063521Z digest=sha256:44f03d9ae8ee812fe59d125529e0581304a26ff58a0eb164443e30b57f9dc5d4

Observation eefe5b7f-468a-47d6-8702-1b66136b57d2 · outbound

This paper cites Training-free Camera Control for Video Generation.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Training-free Camera Control for Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.068471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.068471Z digest=sha256:cda7ec93a5f346ab501d7d22671e626590aa621f922e32963effa366c15062a0

Observation 8d131752-e731-4bce-997a-5b3167178583 · outbound

This paper cites CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.073739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.073739Z digest=sha256:35f656b60e68bf93519e7efe1bf12ac6e34ec8077891f6b3f2f0bc3a057596b5

Observation 86c30a36-b022-4746-ac38-02d390094fb1 · outbound

This paper cites CamI2V: Camera-Controlled Image-to-Video Diffusion Model.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth CamI2V: Camera-Controlled Image-to-Video Diffusion Model

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.078737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.078737Z digest=sha256:4b45d996d2dbab6a6ef3b934e4c92f654745c9862fd60a2930d82eade7824769

Observation 6cb9c08b-6887-43d7-b1ed-6cef6b5a4e2b · outbound

This paper cites I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.084748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.084748Z digest=sha256:fd2d07b97ef7b26c62a9bbb5afe284fb6167906104524cd1a1910e170f481325

Observation dbc67455-d2c9-480b-a162-616f7fe3619a · outbound

This paper cites DiVE: DiT-based Video Generation with Enhanced Control.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth DiVE: DiT-based Video Generation with Enhanced Control

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.091366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.091366Z digest=sha256:eafb94d767d9bb0a84b3c9c708afef0a3f770ecb85a25b9386127a1a30ab2b68

Observation c3c39100-9998-42e8-96aa-a59ba53efb26 · outbound

This paper cites DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.096340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.096340Z digest=sha256:ad87318b737d1eb8daee258b91324bfb4d75b804bd8c76cf37b40537ad056320

Observation 61b0939e-94aa-42ce-b30f-6314d97e3429 · outbound

This paper cites Image quality assessment: from error visibility to struc- tural similarity,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Image quality assessment: from error visibility to struc- tural similarity,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:18:26.780118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:18:26.150773Z digest=sha256:b6ed656bdd8989ed854566a9183509a35b4b7c9c42ad0816169ea262606017d6

Observation 67a12022-cc92-46b4-add0-aff028746541 · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth nuscenes: A multimodal dataset for autonomous driving,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.269401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.269401Z digest=sha256:e5a0d169c6171eeef4e71464a3b29a826c591a7c221cc9147822027fb386c540

Observation e9bf8e4d-a2b6-4058-8652-c04c31f50c82 · outbound

This paper cites Stereo magnification: Learning view synthesis using multiplane images,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Stereo magnification: Learning view synthesis using multiplane images,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.315681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.315681Z digest=sha256:cc6ea069f2c6bf23c5a20b317709cb3c62610a58e1b8467f24b7d648a733324b

Observation 9db2c7cf-78ee-4789-8bd6-c4ebb664d404 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium,.

PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth Gans trained by a two time-scale update rule converge to a local nash equilibrium,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:26.321884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:26.321884Z digest=sha256:12d72fb34d62f390d4fc52032c853fa9465d89bfa1b2a6b8b88024648957a61f

Pith citing papers

Observation 7be77090-e6f2-46f3-9685-261b3488ecd5 · inbound

3D and 4D World Modeling: A Survey cites this paper.

3D and 4D World Modeling: A Survey PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-05T06:04:18.694737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:04:18.694737Z digest=sha256:cb208b173d7269fff244b8aa0458efa87512ea7cecc05d562decb9db843967cb