Pith. sign in

Paper Citation Record · LEDGER

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving

As of 14 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2411.14716.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14716 v2

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:27.362145Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-17T01:59:16.251640Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T02:01:25.681466Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6aa18881-9c91-4198-9ba5-ec4a4570edb5 · outbound

This paper cites ALSO: automotive lidar self- supervision by occupancy estimation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving ALSO: automotive lidar self- supervision by occupancy estimation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.990546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.162166Z digest=sha256:61e0d6ec4d676902ad6e3996299e7ad37d83e63f37376f043d9bf257767287fe

Observation 44b446bc-cc5f-4bb7-9ba8-5a0dba7f7946 · outbound

This paper cites Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.979309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.166625Z digest=sha256:a565e4ccee5d0080d8fb63239d21dc820fee7f959ad9ae8d6124404479e1ebeb

Observation 32523c7e-23e3-449b-9568-9bbac70847ef · outbound

This paper cites GaussianBeV: 3D Gaussian Representation meets Perception Models for BeV Segmentation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianBeV: 3D Gaussian Representation meets Perception Models for BeV Segmentation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.170569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.170569Z digest=sha256:a2b6500cb2187c421471c095c48db72da5ef8859f7e8690b6133b991ff95d0f1

Observation 364d9dda-a1c3-405d-b6e4-669cea8f2498 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.967870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.175223Z digest=sha256:86591e21a00f244506c773f203a2af069fcdd7a7c8ba58db3adf840fced6af1f

Observation 8eea7ba1-caaa-41f6-9fae-5252d194ba42 · outbound

This paper cites Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.179157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.179157Z digest=sha256:64e197ff36a065a45d4161fab4d2fc232b9e9af930e1bbd07dc70b1d15be513e

Observation b38e2d77-f02b-4646-9620-213bc74216f0 · outbound

This paper cites Gaussianpro: 3d gaussian splatting with progressive propagation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Gaussianpro: 3d gaussian splatting with progressive propagation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.183375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.183375Z digest=sha256:c7ebce472746bfc5eff4a0e6ffffe32c80a3992784f364359e196d550d372831

Observation b5b72d35-31ca-4ea2-869d-422cddf16f66 · outbound

This paper cites MMDetection3D: Open- MMLab next-generation platform for general 3D object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving MMDetection3D: Open- MMLab next-generation platform for general 3D object detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.187557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.187557Z digest=sha256:5ddf1da860285da11e323e793f8f10b0b3db99c3259811bc1bce694b7be09bb4

Observation fedf35c9-25ba-4600-b167-292b5e65e2b1 · outbound

This paper cites an unresolved cited work.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:05:27.942940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.191103Z digest=sha256:b38861a9907bbc8a2bca68b80215816ce71fd53305e2faa62ca0d978366d50f6

Observation b3ff89b9-ad5a-455f-9514-d4d324f85172 · outbound

This paper cites GaussianOcc: Fully Self-supervised and Efficient 3D Occupancy Estimation with Gaussian Splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianOcc: Fully Self-supervised and Efficient 3D Occupancy Estimation with Gaussian Splatting

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.194774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.194774Z digest=sha256:80457ac41db0f6feac6fdae03a47d376330d272e4fb4567047edb561a077f146

Observation 9b0b4069-e556-435b-b94c-37a0b2dac1f7 · outbound

This paper cites Digging into self-supervised monocular depth estimation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Digging into self-supervised monocular depth estimation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.198601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.198601Z digest=sha256:594a1c17fd99319e12afc0eea432286c4e407895c5e2ae3448772ba1c875474e

Observation 916b2e7e-76e6-4e23-be54-de7b1ffacc25 · outbound

This paper cites Planning-oriented autonomous driv- ing.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Planning-oriented autonomous driv- ing

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.925559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.202410Z digest=sha256:0e65ada3319ea57e4cd4f2d97f97c49462076dcd15f019dbbda3709b0542659a

Observation 0eddf914-42da-4150-ab07-76c9ac3658fe · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.206211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.206211Z digest=sha256:9cf0222b3642762b6c13323be7b6336b27d8caab5f4e4175cf2d43493cf6a799

Observation c335db95-c393-495e-8087-40c28c78c88a · outbound

This paper cites Tri-perspective view for vision- based 3d semantic occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Tri-perspective view for vision- based 3d semantic occupancy prediction

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.913850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.210238Z digest=sha256:3f60b42bcae561261f9c0ba1e7a9dad2c08de41cfc250bc47e7e218deb2bebfd

Observation beac3a8e-e369-458b-96cf-09dd9cd43e12 · outbound

This paper cites GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.213896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.213896Z digest=sha256:46c3e2b77035c7ebaea87d3c0bc0f93bef036a5a8d022ebafe8002e7ab422901

Observation 67d713cc-7e4f-4c58-9197-877d90925caf · outbound

This paper cites Nerf-mae: Masked autoencoders for self-supervised 3d representation learning for neural radiance fields.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Nerf-mae: Masked autoencoders for self-supervised 3d representation learning for neural radiance fields

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.901390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.218037Z digest=sha256:7380d5ca405cfe68bbf1edcc743d79e1c014d9a14954b5223b2649c03d781bff

Observation f3c37000-0a54-4416-8f8d-99939aeb0ad4 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving 3d gaussian splatting for real-time radiance field rendering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.221978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.221978Z digest=sha256:3f1a357d1f0218396ff215bd524af01611101c5d16a8df4f9b035f85c795e309

Observation 5e45081c-65f6-4f5a-9663-f4f21306d076 · outbound

This paper cites Maeli: Masked autoencoder for large-scale lidar point clouds.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Maeli: Masked autoencoder for large-scale lidar point clouds

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.883157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.226043Z digest=sha256:e9700054a0a1cf366a9923ae55aacf8c2c4bdee2a57d3b7ec624fafef1f46e7f

Observation 95bb179d-5273-4705-83c1-5e3fe8d5889a · outbound

This paper cites Unifying voxel-based representation with transformer for 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unifying voxel-based representation with transformer for 3d object detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.872005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.229982Z digest=sha256:f6768e1a82891faf392fe33f87e9422f714121bc87c4ed2d190d0d0ba470add3

Observation eac0eb53-c7c9-4304-ba0b-0ac1bb817413 · outbound

This paper cites Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.233686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.233686Z digest=sha256:5c289ab57cda3f653351c20979dd3999999deaf07189f0f8858cfc077e79dc4c

Observation 4e9b9075-1a1d-4b0c-a118-522fdd1133b3 · outbound

This paper cites Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.855036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.237676Z digest=sha256:245d258a89127786cfd853299936275f3693a56dcbee2a2406fb06891ea60e51

Observation 43ea415c-9bc2-49c8-878f-f7f5bfa6db9b · outbound

This paper cites Bevformer: 9 Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevformer: 9 Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.844433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.241236Z digest=sha256:9088c0223070dfd99c0fc7268a51f9b41523782c04a88c8bc8588a6142c1bd1f

Observation 098e1ebf-13ae-4cd8-8efe-c414cfbc80be · outbound

This paper cites Fb-occ: 3d occupancy prediction based on forward-backward view transformation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fb-occ: 3d occupancy prediction based on forward-backward view transformation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.833377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.245115Z digest=sha256:7a462aaccfb447ba5fbfc6f77c49218cb28de2b8b2980b56b0ff0690dec642a5

Observation 377dcf41-c935-4ebc-857b-33e47c637523 · outbound

This paper cites Fully sparse 3d occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fully sparse 3d occupancy prediction

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.821803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.248649Z digest=sha256:5aa2e0a70435629e2f82df625fd6cebc66274f442e6f9495e6923f4d379d3177

Observation 86fd28be-248a-4547-a518-7b5ac43312eb · outbound

This paper cites PETR: position embedding transformation for multi-view 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving PETR: position embedding transformation for multi-view 3d object detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.809763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.252268Z digest=sha256:b378cde87b7aa02681b0b6dff4ab485e45b1925902435e295bf86d3d6b5e1c4f

Observation 11d79bf9-fe4e-4747-ae37-6c204346d9d6 · outbound

This paper cites A convnet for the 2020s.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving A convnet for the 2020s

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.797683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.255728Z digest=sha256:a3da78d7dfd9721b65642470c36c76b9dc74f7c8bcd58c2cbcbc1f8a98e78f61

Observation 1968de2e-ded2-4e2e-b29e-bb6204b7e931 · outbound

This paper cites Learning ego 3d representation as ray tracing.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Learning ego 3d representation as ray tracing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.785592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.259402Z digest=sha256:9133924db3f70b5f7567077536fa02310cfe44923ec918c441225f771e3213c0

Observation 0a75d406-cd44-413e-9288-0057978490ba · outbound

This paper cites Srinivasan, Matthew Tancik, Jonathan T.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Srinivasan, Matthew Tancik, Jonathan T

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.772732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.262909Z digest=sha256:1a6e7a7c7a543eb69920964b25a9aa05b117690d2123174f23905821351eacb0

Observation d3440214-7ec2-433a-9352-7229afe9559a · outbound

This paper cites Occupancy-MAE: Self-supervised Pre-training Large-scale LiDAR Point Clouds with Masked Occupancy Autoencoders.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occupancy-MAE: Self-supervised Pre-training Large-scale LiDAR Point Clouds with Masked Occupancy Autoencoders

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.266517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.266517Z digest=sha256:5d2977f711fac38e145d29693d69d13756f3e3574c9615cb0801edf7a003b399

Observation e9304f3a-0c36-415f-ac8e-3e88144a545e · outbound

This paper cites Occupancy-mae: Self-supervised pre-training large- scale lidar point clouds with masked occupancy autoen- coders.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occupancy-mae: Self-supervised pre-training large- scale lidar point clouds with masked occupancy autoen- coders

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.759678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.269898Z digest=sha256:9d92a9eb369324d7a48d496552fdc9aa313d2ae53de70a83c22ae44902656ae0

Observation e62b69de-5654-4eb0-a3f0-1b139c2b025b · outbound

This paper cites Segcontrast: 3d point cloud feature representation learning through self-supervised seg- ment discrimination.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Segcontrast: 3d point cloud feature representation learning through self-supervised seg- ment discrimination

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.745396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.273588Z digest=sha256:fa72a663581f932e58cd26ebe1f1c1ef5f5e4d499f359296e0cc2ed5deb269b7

Observation 2cb62c01-bad2-4606-a181-191bf402a020 · outbound

This paper cites Renderocc: Vision-centric 3d occupancy pre- diction with 2d rendering supervision.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Renderocc: Vision-centric 3d occupancy pre- diction with 2d rendering supervision

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.733778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.277305Z digest=sha256:f93acbb6ad4f34e4fd8e9906f52a19b8376c2d326ffa99c7f418a0e02d96b772

Observation bf5ff595-6865-4b10-a764-9f8e63014d23 · outbound

This paper cites Is pseudo-lidar needed for monocular 3d object detection? In Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pages 3142–3152,.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Is pseudo-lidar needed for monocular 3d object detection? In Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pages 3142–3152,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.280812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.280812Z digest=sha256:1e312f9069f335489d6c92cc05ce36b430bb4cefdf684d7c4df12e95b63ac54a

Observation 65fbab39-a80b-4f6d-86b5-af7c3848ba3b · outbound

This paper cites BEVContrast: Self-supervision in bev space for automotive lidar point clouds.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVContrast: Self-supervision in bev space for automotive lidar point clouds

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.713543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.284207Z digest=sha256:fd9ce929c464b6cac250714c3e478706c2f3ac4844e1899801d911773dacdacf

Observation d216e97f-d406-4719-89f6-905f7aed8f78 · outbound

This paper cites 3dppe: 3d point positional encoding for multi-camera 3d ob- ject detection transformers.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving 3dppe: 3d point positional encoding for multi-camera 3d ob- ject detection transformers

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.700033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.287586Z digest=sha256:cf5b7a235ea3526c8faaf4de242e857b6da40965e5b06e5a6414a8d88701f728

Observation 5b95f62a-1da0-4832-b2e3-5134d95c2f1f · outbound

This paper cites Sparseocc: Re- thinking sparse latent representation for vision-based seman- tic occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Sparseocc: Re- thinking sparse latent representation for vision-based seman- tic occupancy prediction

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.687493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.291176Z digest=sha256:74373f0e979e2b3e6d87f9c31d30bab5470e5b1ba74c5dbeea526b1b0f488809

Observation 72f9ee61-f12f-4a46-9bf9-4ece5f3170d5 · outbound

This paper cites Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.675621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.294803Z digest=sha256:54f7123fcf16d68283c9797276d86044553781de93e6735876637f63495c8c35

Observation f5f41528-31e7-4f6e-b302-11c24e52979f · outbound

This paper cites Scene as occupancy.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Scene as occupancy

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.298136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.298136Z digest=sha256:4469a88ad88c067f79f9b040412e64fbbb0371dab61b7670b11c5bf85bb5fbbd

Observation 5198f398-2c1f-45a7-88e7-04ec9d3df602 · outbound

This paper cites Opus: occupancy prediction using a sparse set.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Opus: occupancy prediction using a sparse set

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.655341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.301497Z digest=sha256:8838010c654570f068f10f282954ecc2b0334b4fb609da7a90d2f4e1293b8b11

Observation bd4eaec4-fa88-449a-9cb0-7cc9b50c78b0 · outbound

This paper cites Fcos3d: Fully convolutional one-stage monocular 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fcos3d: Fully convolutional one-stage monocular 3d object detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.643571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.305041Z digest=sha256:701681fb61f56efd0da528652c9385ef92cc0ab9b0cee67dd3494d8d84df3757

Observation 6f9e0367-fec3-43f0-a83f-8fb675511538 · outbound

This paper cites Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.629682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.308616Z digest=sha256:ff313a6a38919c47fa55182000aab04130ca69b0ebac63a242d797e46812aa0c

Observation befcdced-425b-4aa2-8bbc-c27033a9dd55 · outbound

This paper cites Cross modal transformer via coordinates encoding for 3d object dectection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Cross modal transformer via coordinates encoding for 3d object dectection

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.616552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.312339Z digest=sha256:58f8bbf48b331e7b797465e3c054d4303435046925a11860698a1895c45028b1

Observation 656bbc4b-a1d4-4b23-bd45-445b060157d9 · outbound

This paper cites SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-12T15:05:27.459509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.316012Z digest=sha256:3f9bfef4ad408e8dd3e5b8268237b93f9ebe1812f53167c2d467abfa0f96e1e8

Observation 8d289ecb-128c-4847-b39b-9a2608a3034e · outbound

This paper cites Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.319583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.319583Z digest=sha256:f487b41ee7509e6a8b183c85856d9905935c62f7af1ef8da9b4fd9f8a046c12f

Observation 2e8f9ebd-155d-4b75-bfd3-bd65ba0b3416 · outbound

This paper cites Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.323607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.323607Z digest=sha256:ee4a9b0ab2fc6ec647f33c69b38b3142cff422041673b0a106bf922f57a2ace9

Observation 4707a4ac-6400-442b-990b-eb6ce7df8164 · outbound

This paper cites Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective supervision.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective supervision

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.603727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.328573Z digest=sha256:8af213908862fd91e65ac67c3acfd660211d430f778d52850fd7c8260787ab95

Observation c9c3ce52-b2ec-4e16-8d1d-0c078af174ff · outbound

This paper cites Unipad: A universal pre-training paradigm for autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unipad: A universal pre-training paradigm for autonomous driving

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.589686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.332025Z digest=sha256:f508fb39fb1aa38d9e4d9232e7a11adaeda7396a7b60a863dc9c228f0a8d8880

Observation 7fc07850-f963-4f17-8002-fa91e2da01d7 · outbound

This paper cites Visual point cloud forecasting enables scalable autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Visual point cloud forecasting enables scalable autonomous driving

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.577009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.335585Z digest=sha256:1d63568c72adb0e6f56066a82413ae9479c782905fd418da96ed0047ae957b2d

Observation 94a9b3b0-098f-4769-9ab1-63ec187a6663 · outbound

This paper cites Ad-pt: Autonomous driving pre-training with large-scale point cloud dataset.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Ad-pt: Autonomous driving pre-training with large-scale point cloud dataset

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.565893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.338784Z digest=sha256:c0e2c73bad4648c3ff9eba3e05992fe90544283979b0d4a5f35799ac5caeb0f0

Observation 5a445193-deb4-44c1-8a0f-1fded900b4ce · outbound

This paper cites BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.342139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.342139Z digest=sha256:8ae0f838460e57a6143667584b1cadb8eebe6521e51248ac7bfeed1df11ef4d2

Observation 1c66db66-0e73-4b5f-899c-dee36e580ca5 · outbound

This paper cites Radocc: Learning cross-modality occupancy knowledge through ren- dering assisted distillation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Radocc: Learning cross-modality occupancy knowledge through ren- dering assisted distillation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.554208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.346075Z digest=sha256:dcc69e298cf99a384bf9ebf27f278661c86b1cf8b3ee7c0a9c66bc01b3a7c5dc

Observation 7a88f6a2-edf0-49e9-83e9-c4a5585f7d03 · outbound

This paper cites Hugs: Holistic urban 3d scene understanding via gaus- sian splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Hugs: Holistic urban 3d scene understanding via gaus- sian splatting

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.541582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.350412Z digest=sha256:cdf126407a7ff0a713b2b2f188e56afcfa58de43bf40eaf161eae7cf5b760b22

Observation 02d46c52-78ae-4cf2-8904-252f418a856d · outbound

This paper cites Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.354297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.354297Z digest=sha256:a91a1663ce7210f37606dd9e908b353276453b24df2d2b4d64b5e8e55c188a68

Observation d8316332-66dd-46a4-8b7b-6ca159295ba6 · outbound

This paper cites Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.358066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.358066Z digest=sha256:b0683c70d308227c172f0a9c82f2ed1a29455a96edde61aad36f886dbcb1b774

Observation 92d8d6a5-fac6-4d1e-96b3-b4d0aea5bc2c · outbound

This paper cites MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-12T15:05:27.399830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T15:05:27.362145Z digest=sha256:0f6b83edace468225e8d496b259b282d6e41c3e8da2da1eda37beb71ed5b9274

Pith citing papers

Observation 42612af9-ea0c-4ba2-b543-5b219ee4f4d0 · inbound

Flux4D: Flow-based Unsupervised 4D Reconstruction cites this paper.

Flux4D: Flow-based Unsupervised 4D Reconstruction VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:01:25.683602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T01:59:16.251640Z digest=sha256:4e6db0570e8743cca17daf4d4b04a84d9fe4a704fb57d713f796ffde4642acf4