Pith. sign in

Paper Citation Record · LEDGER

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities

As of 8 August 2026, this Paper Citation Record lists 100 of 144 outbound references and 0 inbound Pith citation observations for arXiv:2509.08302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08302 v1

Coverage vector

measured 100 of 144 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:52:06.559453Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 144 outbound references displayed

  • verified exact4
  • verified fuzzy13
  • unresolved83
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f9257e52-cd33-479b-abe0-31b69599653d · outbound

This paper cites 3d object detection for autonomous driving: A survey,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities 3d object detection for autonomous driving: A survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.955693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.955693Z digest=sha256:05e0bcd0dc57373aa433d94ec73cc32db48583a986ae76dc250b8e56aebf298a

Observation c764dff8-139f-46ed-bdb8-23c5de0b39fb · outbound

This paper cites Vision-based semantic segmentation in scene understanding for autonomous driving: Recent achievements, challenges, and out- looks,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Vision-based semantic segmentation in scene understanding for autonomous driving: Recent achievements, challenges, and out- looks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.961813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.961813Z digest=sha256:b6ea0cbfd8840cab2aae391bf73774ff513f526631fe3a5591233047062cce7e

Observation 5f403c37-4c36-41ed-8fb0-76bce0f8ac70 · outbound

This paper cites A review of deep learning-based visual multi-object tracking algorithms for autonomous driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities A review of deep learning-based visual multi-object tracking algorithms for autonomous driving,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.966653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.966653Z digest=sha256:f0daa8caf1d417c461fe1a80eff7cedb736eb96709127a08445c41d4c64731ab

Observation d65b8950-a850-4da0-acd1-0ccb76d5e3ae · outbound

This paper cites Towards long-tailed 3d detection,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Towards long-tailed 3d detection,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.972339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.972339Z digest=sha256:bec68dd11e5620802e80a4c78d1736c4c7d35731e093d18fe3f8eab4807da379

Observation e465f32f-7cec-4df0-ad99-fe71a4c90dd1 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities On the Opportunities and Risks of Foundation Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.979955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.979955Z digest=sha256:e6fa78725f26ffcd626e113abd60cc562c1982104d341f381d97d1ee71be980e

Observation fae8c30c-c100-4f15-9b78-df95426c161b · outbound

This paper cites Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.987906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.987906Z digest=sha256:ff59ed647fbaf2e9f375e90e94db423b8e648472731d7e6ec5007bf5d83bea89

Observation 376c9556-8aff-430c-8e30-75316db34504 · outbound

This paper cites Applications of Large Scale Foundation Models for Autonomous Driving.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Applications of Large Scale Foundation Models for Autonomous Driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.993544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.993544Z digest=sha256:bc4f6f4b6468ab403cd1242d04b00b5a28802ba9b0f4218b8b9a0a95ae4318fa

Observation d0e341af-95d9-4c15-8f84-6db07cbdc06b · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities A Survey for Foundation Models in Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:05.999643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:05.999643Z digest=sha256:1e172697f5fb706b92e56786ae36668802672cfecd632c4d1f169ae3fb4ac28e

Observation 1888f426-0ae8-4a6a-a30a-965628e0aa72 · outbound

This paper cites Prospective role of foundation models in advancing autonomous vehicles,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Prospective role of foundation models in advancing autonomous vehicles,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.006340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.006340Z digest=sha256:80d7e402d17508b9dcb2ca425a011f7cb47211526a4734912e2918d0457e70e4

Observation 044ac690-0940-4526-a418-a6c219682e59 · outbound

This paper cites LLM4Drive: A Survey of Large Language Models for Autonomous Driving.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities LLM4Drive: A Survey of Large Language Models for Autonomous Driving

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.018250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.018250Z digest=sha256:d9d18b01438f38242493d96565df8757234cf2455f5e3b2ad2c68c960491724a

Observation d9f5dca6-921b-453c-bc8c-5a1e01fa2b39 · outbound

This paper cites Vision language models in autonomous driving: A survey and outlook,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Vision language models in autonomous driving: A survey and outlook,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.023527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.023527Z digest=sha256:11da0d155724d0524b92f7609967be10f7f35d698b4e55d38e432a83b359d26b

Observation de83e643-f1ad-4c54-9db8-23edaf01ae9f · outbound

This paper cites A simple framework for contrastive learning of vi- sual representations,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities A simple framework for contrastive learning of vi- sual representations,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.032288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.032288Z digest=sha256:521b7078adfd00cad9ac16442d7f08be9c333d6798aeede01bf69acfa2c2064e

Observation 5bad0149-16d2-4d99-a209-330436757d0d · outbound

This paper cites Momentum contrast for unsupervised visual repre- sentation learning,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Momentum contrast for unsupervised visual repre- sentation learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.037588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.037588Z digest=sha256:ad0d12496777d5885e36a4ecbac1903852b0e3330100ca30d7cd51568ce08970

Observation a1f52ee6-a4a5-4f23-a82e-2645a82f60fb · outbound

This paper cites Improved Baselines with Momentum Contrastive Learning.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Improved Baselines with Momentum Contrastive Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.043634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.043634Z digest=sha256:f09556033d0ec6dcda21af6042712162d6cdfa225f76844470e61c0d092d134d

Observation 2925c295-deca-4236-ad6e-2b8327b9cd3c · outbound

This paper cites Masked autoencoders are scalable vision learners,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Masked autoencoders are scalable vision learners,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.049114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.049114Z digest=sha256:e41abbb05936429ef7f5c2be11af0f11abcaea1421918609e2f13dcbbd241f71

Observation d69169fb-6421-4219-bfed-0898ef033262 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.054078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.054078Z digest=sha256:e2ad10016a55df1ff8f157f86fe2432e1ae5034b3eb24d300a58d2c2d98920f2

Observation b399a9ea-240a-4d5c-a12f-56223cef3f85 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Distilling the Knowledge in a Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.060395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.060395Z digest=sha256:7daacd28aeade2715531835cfe2f0cd18fd22ad5ac025d81b4081b245c42c8dd

Observation 7a598d00-ea11-4715-a1d1-607a76e0c35d · outbound

This paper cites Knowledge distillation: A survey,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Knowledge distillation: A survey,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.070561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.070561Z digest=sha256:c99e12d3eedc58f953e58dc4b02e7b6fc8b5728aa8522655c71deca8d462d9ad

Observation 75a8cb68-f62b-4daa-a0b1-342804835a49 · outbound

This paper cites Self- training with noisy student improves imagenet classi- fication,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Self- training with noisy student improves imagenet classi- fication,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.078479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.078479Z digest=sha256:6b3a5c1fdb645b03dc3eab5c8d4aa4069da4c8b5b8423d37b1a6afbfc535dd7f

Observation 4ba07694-93b0-44ed-9bf7-c19fb1310825 · outbound

This paper cites Bootstrap your own latent-a new approach to self- supervised learning,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Bootstrap your own latent-a new approach to self- supervised learning,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.087477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.087477Z digest=sha256:bdcf03c84f69bb9870a9825b2d78ed3ddcc683beafbb457436076c888c4c9b31

Observation 29de20c4-8203-44df-8485-5d5e16fd24b2 · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Emerging properties in self-supervised vision transformers,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.093522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.093522Z digest=sha256:30f32eca66ce6b744fc986c5df7acd99f1982f12a478144f04ef6adc573105b6

Observation 997aae34-57c4-4592-92e7-b5ab8cd71ea4 · outbound

This paper cites Structure-from- motion revisited,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Structure-from- motion revisited,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.098456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.098456Z digest=sha256:f32e2e3c5780dd82d1c064618e7b6eba8392a698900549d0596dcd829bf6a0b5

Observation d9c70e9f-8a34-4c5f-b79f-5ae34ecf8564 · outbound

This paper cites Multi-view stereo: A tutorial,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Multi-view stereo: A tutorial,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.103056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.103056Z digest=sha256:def745f73219997d524f2179785ed2d3320709656f9e05ef34569b54af155ae7

Observation 4546bf8d-89ba-441f-a2e0-d6ccdc165bf5 · outbound

This paper cites Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.108496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.108496Z digest=sha256:d8c5acd52f7eb42185a5da43c3b90e5e309911b478e470bd1337f8f23151e4c8

Observation 96ab3df6-150f-4ccf-a4c3-f3435abbd036 · outbound

This paper cites 3d-r2n2: A unified approach for single and multi- view 3d object reconstruction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities 3d-r2n2: A unified approach for single and multi- view 3d object reconstruction,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.114614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.114614Z digest=sha256:244dd11ea4d07b9dbfd44e156250c749a7398af8966994f9c0b5169ff1a97ff8

Observation 0e257a96-f1ae-48f4-8d41-bbbc38e21073 · outbound

This paper cites Predicting depth, surface normals and semantic labels with a common multi- scale convolutional architecture,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Predicting depth, surface normals and semantic labels with a common multi- scale convolutional architecture,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.121630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.121630Z digest=sha256:4875982cc73c7c0d7b22a7a81c1116b2d716df3b8a0a4a22acc8c71a019b814b

Observation bbb8a972-b932-4b27-96bd-8e63b568c99d · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Nerf: Representing scenes as neural radiance fields for view synthesis,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.126104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.126104Z digest=sha256:d9b62e3be69017723be8b6402fe626ce44fa5cae1717e3811caa2bb15600358c

Observation 1dfc9792-243d-4bc1-8da5-8ebb0307a6e2 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities 3d gaussian splatting for real-time radiance field rendering

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.131796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.131796Z digest=sha256:0aa9cb4b8345af91affe0b66f4190a188aca3c60a7b0e1d2a42d14a229548086

Observation 7465bef0-9337-46c8-912c-65294b4f72fb · outbound

This paper cites YOLOv3: An Incremental Improvement.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities YOLOv3: An Incremental Improvement

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.136278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.136278Z digest=sha256:e27db83bb9d90e24d9e69f96f8dd8f76a45bd94109077bc08364bb1ef1a678fd

Observation 8d63dc12-8850-4364-8ce0-a4ae4ccea62f · outbound

This paper cites Deep residual learning for image recognition,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Deep residual learning for image recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.141344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.141344Z digest=sha256:7d5cd0849acea9db0c48460ee6c51bb0675a3dd9a71d0fd18f6221a768da2639

Observation 6bbaf5a6-75be-4b30-b814-8e055348648e · outbound

This paper cites GPT-4 Technical Report.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities GPT-4 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.146791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.146791Z digest=sha256:71b326cecb3013c59a06778dee5cca81722f3db5bde524aed34a19622eadbe88

Observation 306ef2ef-2bec-429d-880a-6fdb354c8660 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.151255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.151255Z digest=sha256:d5248027153f2d0ec77221da0aa6ebfefd7e58c12997f4bd890a63d28afd0ec1

Observation 808ca6a7-f1cc-4b29-828d-6a2e237583af · outbound

This paper cites Segment anything,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Segment anything,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.156549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.156549Z digest=sha256:bb6ec8812024365702c88934ad9e38ca2771d79a09de43398afc5ce384fb47d1

Observation 3b690062-15d8-4053-9c37-476f70e7730b · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities DINOv2: Learning Robust Visual Features without Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.161060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.161060Z digest=sha256:5a6faf9104ad6081e952e7b9fb171f911f851ccbfda513e8a9c97857fc7d4fb0

Observation ed7d736d-bf60-42b9-b47b-a6ce60c45e49 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Learning transferable visual models from natural language supervision,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.165537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.165537Z digest=sha256:927fe59146c1069b8c3efd49fc1a0787bf2dc0123d74a47e2db8608f22371087

Observation 491988f0-11c5-474c-9803-79e69ab57e2d · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for VOLUME 00, 2024 27 Authoret al.: Preparation of Papers for IEEE OPEN JOURNALS open-set object detection,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Grounding dino: Marrying dino with grounded pre-training for VOLUME 00, 2024 27 Authoret al.: Preparation of Papers for IEEE OPEN JOURNALS open-set object detection,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.170279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.170279Z digest=sha256:906c46bc3c3f8d126988b2fa62a7c352db174f0bd7428d220182ad25096a95f3

Observation c087e1f2-c3d8-45da-b1b0-21f0d6b1d44b · outbound

This paper cites Denoising diffusion probabilistic models,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Denoising diffusion probabilistic models,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.174614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.174614Z digest=sha256:68bed5501860583fe6c949513fcf44f77c572069c9a447cbe14281fce15714cc

Observation 45984016-7677-47d3-96dd-fbc7e6d7a997 · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities High-resolution image synthesis with latent diffusion models,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.179142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.179142Z digest=sha256:e6d5441b65c499c37442e39da7745cf6420b59f9e57dfc4aaa8fe979fc46606a

Observation 06588573-edd6-4f41-82f2-8f80df7780c8 · outbound

This paper cites Video diffusion models,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Video diffusion models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.183470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.183470Z digest=sha256:6ec78efd2421e16a21117fd414c78fe2ef47b6b6eb39d4aeacf9dca4dbb7de23

Observation cb6f5e0e-29a8-4af9-88d8-6d20ad38a38b · outbound

This paper cites 3d shape generation and completion through point-voxel diffusion,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities 3d shape generation and completion through point-voxel diffusion,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.188424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.188424Z digest=sha256:3041e4642141f90a52ceb5318a5ba72ed1b947b5b695dbd70079d5a406363834

Observation 254311b5-9d0d-4818-b96f-c978d422a22f · outbound

This paper cites Diffusion-based signed distance fields for 3d shape generation,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Diffusion-based signed distance fields for 3d shape generation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.193245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.193245Z digest=sha256:470bb576ac59a67ad239c74bf04673c0f06605069cea7dc2859fd37b745878ce

Observation 1c0950a9-d402-41ed-875f-898878a7fcf2 · outbound

This paper cites Language models are few-shot learners,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Language models are few-shot learners,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.198082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.198082Z digest=sha256:039830f8a9d3ee7533ff18b7359590eb7010e569c4bab5864c23b6f4eacb99e5

Observation a6b6b936-605d-4ed9-a1ac-5a8190f8a493 · outbound

This paper cites Image-to-lidar self-supervised distilla- tion for autonomous driving data,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Image-to-lidar self-supervised distilla- tion for autonomous driving data,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.204799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.204799Z digest=sha256:6cd0d903637f7c9371d396f8a891b98908726000ab51ec7af649d5eb6563ea72

Observation 48a87cf9-5322-4f28-b184-8470b2895fb4 · outbound

This paper cites Segment any point cloud sequences by distilling vision foundation models,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Segment any point cloud sequences by distilling vision foundation models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.209914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.209914Z digest=sha256:ae3b13f8138b8233cf6b76bad04a5d9401b6ffecf2285ad870d5caac73a39507

Observation 6b775d6d-9da4-45eb-9ceb-f1a6ea681d32 · outbound

This paper cites Better call sal: Towards learning to segment anything in lidar,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Better call sal: Towards learning to segment anything in lidar,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.214452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.214452Z digest=sha256:6f4f9b1cb765f40767fe25590c23e877ad6463666478ad187f79487252178a7e

Observation 6faf96f5-d14a-42f4-b8c5-6b7743b1559d · outbound

This paper cites Sam4udass: When sam meets unsupervised domain adaptive semantic segmentation in intelligent ve- hicles,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Sam4udass: When sam meets unsupervised domain adaptive semantic segmentation in intelligent ve- hicles,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.219199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.219199Z digest=sha256:5ddcd16999a14ae204603995476ca6fa6c5ed5c59c1b0a0b20610956d6e01c72

Observation e9f52271-fced-471d-8acd-cf8a1584d203 · outbound

This paper cites Occnerf: Self-supervised multi- camera occupancy prediction with neural radiance fields,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Occnerf: Self-supervised multi- camera occupancy prediction with neural radiance fields,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.226674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.226674Z digest=sha256:0ae595c7baa3defa7e9cf8bb0b7079a11841bb73912a25de0d1cac21eaaaa178

Observation a46554fc-03ac-48d2-9dd0-7cc169d2c97d · outbound

This paper cites Open 3D World in Autonomous Driving.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Open 3D World in Autonomous Driving

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:52:07.504246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.234374Z digest=sha256:4bd727211db27e9491b5245944ad2a697c8fe1c2efffc2449504a219ade5b7f5

Observation f9188c8d-3d07-49f5-bccb-eeb23a512c27 · outbound

This paper cites OVO: Open-Vocabulary Occupancy.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities OVO: Open-Vocabulary Occupancy

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.241136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.241136Z digest=sha256:77928de0f158622a4648f87a92a0bf51b45b5dcc3ff87fac891d446a45afad96

Observation 3efc75fb-904e-4f13-b4a2-75e4c755c603 · outbound

This paper cites Clip2scene: Towards label-efficient 3d scene understanding by clip,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Clip2scene: Towards label-efficient 3d scene understanding by clip,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.246751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.246751Z digest=sha256:7e0d25cc868c47f46d50f7f706eb043c5751c99d968165b750d474079ce0a091

Observation 1ed4d6d4-5cfb-4759-88dd-7b6a889fd293 · outbound

This paper cites Vlm2scene: Self- supervised image-text-lidar learning with foundation models for autonomous driving scene understanding,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Vlm2scene: Self- supervised image-text-lidar learning with foundation models for autonomous driving scene understanding,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.251773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.251773Z digest=sha256:1eb8450020d990b794bccfe16bc7ac2a13bbc4f0e0bbd66b724a5227951a6ec3

Observation 04225145-4aa3-4e2d-b62a-7ca4a398a769 · outbound

This paper cites Unsupervised 3d perception with 2d vision-language distillation for autonomous driv- ing,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Unsupervised 3d perception with 2d vision-language distillation for autonomous driv- ing,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.256601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.256601Z digest=sha256:9052a032b9a4ffcf7e981c9d72e07d68cfe346daef03102a52d9abf29077fc51

Observation 27d822af-a0f7-4591-9dcb-8509cecbd4bf · outbound

This paper cites Opensight: A simple open-vocabulary framework for lidar-based object detection,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Opensight: A simple open-vocabulary framework for lidar-based object detection,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.261543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.261543Z digest=sha256:841cf8f5037b57518af0d42186c4f95915caa06cf6ff6eec1d4442376a47b93c

Observation f0edd2f8-565c-4683-b2e1-15f3c2bc8e0a · outbound

This paper cites SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:52:07.465345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.267283Z digest=sha256:7568927bc9c742e84bc42120ead001c9d16f0ff520c4298f5916bd494cdbc98e

Observation 6545c1ab-899f-40ff-971a-1f30917478df · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities GPT-Driver: Learning to Drive with GPT

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.272252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.272252Z digest=sha256:ffcb4c1b91031e340b0e0867a69d69b489e604d0f3c32c30c6b9528f32aa9ee3

Observation ab1bbacd-b53e-48fb-bb5b-3d99c6f92b34 · outbound

This paper cites OmniDrive: A Holistic Vision-Language Dataset for Autonomous Driving with Counterfactual Reasoning.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities OmniDrive: A Holistic Vision-Language Dataset for Autonomous Driving with Counterfactual Reasoning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.279971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.279971Z digest=sha256:02652f79bb8eeeace194e398b96687e79362da29477feb261d318f5b13678000

Observation 2c0a3865-ec48-4065-8669-aaee8b2ef20e · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Drivegpt4: Interpretable end-to-end autonomous driving via large language model,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.285794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.285794Z digest=sha256:e2c39874fe0ebe1dfe8d0c23b0b3c1f399dbfbe68a71468f3ce6e97ed327f282

Observation e5bb838d-6f68-4829-a109-84ab3eec627f · outbound

This paper cites Dol- phins: Multimodal language model for driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Dol- phins: Multimodal language model for driving,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.290926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.290926Z digest=sha256:29c53a83273ce6d9d64aa104f3fb6a8c1e0697067f899e9106ca543af0c38fba

Observation 3598c952-d75e-4441-ae7e-319b1ece4b4f · outbound

This paper cites EMMA: End-to-End Multimodal Model for Autonomous Driving.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities EMMA: End-to-End Multimodal Model for Autonomous Driving

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.295417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.295417Z digest=sha256:b7db481ae553376a743eb92760222f13dcef5311fe8c199ea458625021356641

Observation 5e4f1d86-3b82-4191-9958-3027096e23f7 · outbound

This paper cites Lidar-llm: Exploring the potential of large language models for 3d lidar understanding,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Lidar-llm: Exploring the potential of large language models for 3d lidar understanding,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.299745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.299745Z digest=sha256:16d731f885c67776e8615cd221bf78d8bcadec274b5fdbea1ec0dd9a2b377363

Observation 1c1f6882-564a-4c43-bfa7-9e01b4c0f605 · outbound

This paper cites A survey on hallucination in large language mod- els: Principles, taxonomy, challenges, and open ques- tions,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities A survey on hallucination in large language mod- els: Principles, taxonomy, challenges, and open ques- tions,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.304116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.304116Z digest=sha256:41c4aa8502d945dbeb38cd279d79f5eb940cf86268c3152e3b63c8990853c8ca

Observation 977244c2-3572-457e-ba6a-ed1b2b120fdc · outbound

This paper cites Retrieval-augmented genera- tion for knowledge-intensive nlp tasks,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Retrieval-augmented genera- tion for knowledge-intensive nlp tasks,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.308606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.308606Z digest=sha256:14a7dddcff9a6c7cca9e7d6f5c017f5db898556d59a4ba78fa4c13132ad843bd

Observation 703757cb-f79b-48ae-a467-4b4bd68e965e · outbound

This paper cites Self-rag: Learning to retrieve, generate, and critique through self-reflection,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Self-rag: Learning to retrieve, generate, and critique through self-reflection,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.313814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.313814Z digest=sha256:d8ba6f41bcef45722f51e95930353b9640d125916314229db8ab90b4603f7d41

Observation 73b7e1da-1908-430c-a002-6ffd354230d4 · outbound

This paper cites Driving with llms: Fusing object-level vector modal- ity for explainable autonomous driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Driving with llms: Fusing object-level vector modal- ity for explainable autonomous driving,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.320327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.320327Z digest=sha256:b4d621dc53f77a883f5747e99ecf34ce49ded5eb4be99ef1fcec236088936cb3

Observation 001be390-273c-42c5-9086-c894ec33fd87 · outbound

This paper cites A survey on occupancy perception for autonomous driv- ing: The information fusion perspective,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities A survey on occupancy perception for autonomous driv- ing: The information fusion perspective,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.324944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.324944Z digest=sha256:92ff8744e72695f59684ed63ad55463899fdf731b70d8f4d1f3e260ea4a353af

Observation 1479719f-f425-4895-b9df-12cec0af5abb · outbound

This paper cites Neural vol- umetric world models for autonomous driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Neural vol- umetric world models for autonomous driving,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.331586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.331586Z digest=sha256:4bc663e341622b8e6e804373b7369061a8367471f8b6f12d62ae2a14b886aba0

Observation 34593224-8340-4434-88b9-dcf1ba67067f · outbound

This paper cites Tri-perspective view for vision-based 3d semantic oc- cupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Tri-perspective view for vision-based 3d semantic oc- cupancy prediction,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.336689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.336689Z digest=sha256:868e776d57d4c6e2202546ad930cddba6a83071fcfa018f157ba2b100bb784f9

Observation 82f575c2-6de9-438b-9b6d-0f9a26cb95d1 · outbound

This paper cites V oxformer: Sparse voxel transformer for camera-based 3d se- mantic scene completion,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities V oxformer: Sparse voxel transformer for camera-based 3d se- mantic scene completion,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.341604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.341604Z digest=sha256:9c6ef3b0217f0752c03cd7ac103e80607e792e7f64cdae8b0d528725a1556322

Observation 65acb063-59a9-4f83-be6d-2983ad8596a0 · outbound

This paper cites Occformer: Dual-path transformer for vision-based 3d semantic occupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Occformer: Dual-path transformer for vision-based 3d semantic occupancy prediction,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.348253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.348253Z digest=sha256:eae9ada0299692438daababfbe5f6f3a24fdb8cddf3ad32b981f5f1b13f2a2df

Observation 30f16fba-7139-4035-b7a9-0c136bee422f · outbound

This paper cites Fully sparse 3d occupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Fully sparse 3d occupancy prediction,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.353724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.353724Z digest=sha256:8dbd3f5e4c3697a79efc829030b36ef900b2e291a54aeab71ff999ad4fdd3696

Observation 9ffaaa56-33af-47de-b459-9b0f51844400 · outbound

This paper cites Hybridocc: Nerf enhanced transformer-based multi- camera 3d occupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Hybridocc: Nerf enhanced transformer-based multi- camera 3d occupancy prediction,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.363094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.363094Z digest=sha256:9065c969feadf27dc2fa39da6a463fe3ed416b1e0cd3ca50d46e8f953f1b7ae7

Observation 4615e532-1b5b-4770-9edd-3de223996b30 · outbound

This paper cites Renderocc: Vision- centric 3d occupancy prediction with 2d rendering supervision,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Renderocc: Vision- centric 3d occupancy prediction with 2d rendering supervision,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.368464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.368464Z digest=sha256:4226b44396bd37d6e301a250328c7ef7a940de85e37b8dc3d1b7626c21e48a5b

Observation be6922b5-27bf-4b10-a3f6-a4516abf7fb0 · outbound

This paper cites S-nerf++: Autonomous driving simu- lation via neural reconstruction and generation,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities S-nerf++: Autonomous driving simu- lation via neural reconstruction and generation,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.375736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.375736Z digest=sha256:4feb26ec0be40f26eb4c6e4a63f097e99490b62d999bd0e212f4f818cb34201a

Observation d5c805ad-40d8-4c26-9740-77783c7ae86f · outbound

This paper cites Selfocc: Self-supervised vision-based 3d occupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Selfocc: Self-supervised vision-based 3d occupancy prediction,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.387522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.387522Z digest=sha256:dae4262ac4347b847d553bad8744f723d5fafd90863628973b5f35d6bc986cf1

Observation a473cd25-a1ab-4d80-8798-767ed4d65a06 · outbound

This paper cites RenderWorld: World Model with Self-Supervised 3D Label.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities RenderWorld: World Model with Self-Supervised 3D Label

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.391983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.391983Z digest=sha256:8f301de88e41b6f4a7e93c6ec5afb6eefb73172864ad010ef629db657018dd70

Observation 1d99020e-7f55-40ca-b8d5-8c61fb58add6 · outbound

This paper cites GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.396798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.396798Z digest=sha256:84c10a4df92b130c74d778eed54f4858953a10ed07fd9055bf2305787cf7db7b

Observation 24b1a303-cb16-4c99-a1b9-33942cd6e303 · outbound

This paper cites Street gaussians: Modeling dynamic urban scenes with gaussian splat- ting,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Street gaussians: Modeling dynamic urban scenes with gaussian splat- ting,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.404842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.404842Z digest=sha256:7c29bbbe0ac05e38879edb10f016e011bc751164e08bae6cc0debfd706b12982

Observation d704555e-22fb-4d26-89bc-67fec162f7df · outbound

This paper cites Gaussianformer: Scene as gaussians for vision-based 3d semantic occupancy prediction,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Gaussianformer: Scene as gaussians for vision-based 3d semantic occupancy prediction,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.412278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.412278Z digest=sha256:189ac289a9cf2d942dbd791a3e1715a2cf55515ef3b278af7822c892fd42abb9

Observation 2e9902a1-753f-4e91-9b0d-b471be8fa350 · outbound

This paper cites Surroundocc: Multi-camera 3d occupancy pre- diction for autonomous driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Surroundocc: Multi-camera 3d occupancy pre- diction for autonomous driving,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.419424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.419424Z digest=sha256:33922cc6b87850ecdd610df423c5f0d8fa809f885c98f58ce15073c557e04eef

Observation 1c398fed-632f-4b0f-9b3b-a027bdf1a06d · outbound

This paper cites Uno: Unsupervised occupancy fields for per- ception and forecasting,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Uno: Unsupervised occupancy fields for per- ception and forecasting,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.430373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.430373Z digest=sha256:7cd76a04452bf26a65cf50a569cace5bdad52152a963405a22745b1177b356d3

Observation b2d0919c-459d-47d2-9266-e11f47851c21 · outbound

This paper cites Maeli: Masked au- toencoder for large-scale lidar point clouds,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Maeli: Masked au- toencoder for large-scale lidar point clouds,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.440273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.440273Z digest=sha256:8b056eb05aa1eff11d4565878143ad9e2a4a60cf0d37079b08eab336c0fb6e7f

Observation a71e55e2-ee72-4f0e-9534-fc3aeab6833b · outbound

This paper cites Masked autoencoder for self-supervised pre-training on lidar point clouds,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Masked autoencoder for self-supervised pre-training on lidar point clouds,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.454199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.454199Z digest=sha256:b65055dc6be1ce184e5c2553a0ba23d9326b82ec2be54b6c778efc5288af1a9d

Observation ad181b6e-cb92-4571-a1b0-7d6603404a52 · outbound

This paper cites Bev-mae: Bird’s eye view masked autoencoders for point cloud pre-training in autonomous driving sce- narios,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Bev-mae: Bird’s eye view masked autoencoders for point cloud pre-training in autonomous driving sce- narios,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.461440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.461440Z digest=sha256:c1dfc37a0ebc36945ac75781370652cb5c4189c2836f25c8d371d0f5ab357f1f

Observation 717ed5c5-dc35-4283-81b9-1c546ed7bbf2 · outbound

This paper cites Geo- mae: Masked geometric target prediction for self- supervised point cloud pre-training,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Geo- mae: Masked geometric target prediction for self- supervised point cloud pre-training,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.958701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.466253Z digest=sha256:3adad9d59306774cde08d018e022681c2020149adb34ab4ccd85d84d92236767

Observation 2cb009e3-6fad-4b92-af25-a0e5f5d355db · outbound

This paper cites Openoccupancy: A large scale benchmark for surrounding semantic oc- cupancy perception,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Openoccupancy: A large scale benchmark for surrounding semantic oc- cupancy perception,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.853318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.473340Z digest=sha256:dbd45602e7d548cf782ee78e20f30302a1ab583c6c40316a2dde01857fd9109d

Observation f6a9f70f-24ed-4655-8ec9-92d3e3846131 · outbound

This paper cites Occ3d: A large-scale 3d occu- pancy prediction benchmark for autonomous driving,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Occ3d: A large-scale 3d occu- pancy prediction benchmark for autonomous driving,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.741657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.479214Z digest=sha256:ebfead1504ad4e4cb5686bbe74352fd8bbd6049f96fdcba9595d109b6a578caf

Observation f207edd4-de38-4471-a912-3265676c83f5 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Representation Learning with Contrastive Predictive Coding

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.484270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.484270Z digest=sha256:8fff68434bfc6f56c0c1204e09272cfa7a8e434379d5e8aadeb27d1c5393c433

Observation 64004ea8-13af-4853-be77-bf72b6e0d902 · outbound

This paper cites Cross-modal contrastive learning for domain adap- tation in 3d semantic segmentation,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Cross-modal contrastive learning for domain adap- tation in 3d semantic segmentation,

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.602627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.490510Z digest=sha256:4dd390e447f124b229984a82eabcc32742d597730c79816c1b8a26fee967e454

Observation 17f96b34-aa93-4c4c-bd53-8196ce44d4eb · outbound

This paper cites 4d contrastive superflows are dense 3d representation learners,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities 4d contrastive superflows are dense 3d representation learners,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.493586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.499824Z digest=sha256:0a9a40f7b3f77c07a148cef3ec7ef7a791106596f7d05039615c833319b7beb2

Observation e5ae1d23-a4cc-42f0-96be-6a0d24730c45 · outbound

This paper cites Superflow++: Enhanced spatiotemporal con- sistency for cross-modal data pretraining,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Superflow++: Enhanced spatiotemporal con- sistency for cross-modal data pretraining,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.409874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.505917Z digest=sha256:3300a96799b54e780ca925b4ce456ee5bee5eb1037f0829b2391de26832aec3d

Observation e51c080a-0dbe-4bfd-a2bc-371a2c3211e6 · outbound

This paper cites ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.510402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.510402Z digest=sha256:29fcde5ac74eacd4de5393f516c6fdcbbb40c089c88f3b0c5bf6be07b4f42d6b

Observation f7d3ad64-de83-4a49-9c0b-b076dc108545 · outbound

This paper cites BEVDistill: Cross-Modal BEV Distillation for Multi-View 3D Object Detection.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities BEVDistill: Cross-Modal BEV Distillation for Multi-View 3D Object Detection

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:52:07.308999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.514900Z digest=sha256:81fcd6d077f17474095171055d754ade0ca32518fa2f318a1561a0ad40d97e66

Observation 1fe045da-f0f8-46b0-8cf4-87fc20ed479b · outbound

This paper cites Distill- bev: Boosting multi-camera 3d object detection with cross-modal knowledge distillation,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Distill- bev: Boosting multi-camera 3d object detection with cross-modal knowledge distillation,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.268784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.519795Z digest=sha256:c34fe71b625d21d2313b3ef985a8fc95fb305a90c5be5566fb454c0e49a533f5

Observation e5b5c084-ceb6-4252-9a09-6123e0307ead · outbound

This paper cites Geometric-aware Pretraining for Vision-centric 3D Object Detection.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Geometric-aware Pretraining for Vision-centric 3D Object Detection

Reference 94

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:52:07.279648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.524843Z digest=sha256:36d78c3babdde62cbf507c1f1999aec88da0b51d32114f6acf607d2d7654402c

Observation 6b67b156-6f08-49d4-8398-ee7bd8d43878 · outbound

This paper cites Unidis- till: A universal cross-modality knowledge distillation framework for 3d object detection in bird’s-eye view,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Unidis- till: A universal cross-modality knowledge distillation framework for 3d object detection in bird’s-eye view,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:11.129176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.530339Z digest=sha256:ac71cec627bd12c92e48c2f08cc5c4a5ec1fa667909a06dd7d828da4691f645e

Observation 2c0bd976-ee51-4a9c-8764-b69dcaafee5c · outbound

This paper cites Revisiting domain generalized stereo match- ing networks from a feature consistency perspec- tive,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Revisiting domain generalized stereo match- ing networks from a feature consistency perspec- tive,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:10.988761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.535252Z digest=sha256:81773a5a55da40583f555acc35f75ee0875d7a7f0a2ab1d2c07cd2b426865ad6

Observation aa4d1f90-2d62-4d39-93f3-c3222569a5c0 · outbound

This paper cites Weakly supervised monocular 3d object detection using multi-view projection and direction consis- tency,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Weakly supervised monocular 3d object detection using multi-view projection and direction consis- tency,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:10.852859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.539772Z digest=sha256:75de9fc85604ee10d5cd8c0936e6378e34a735998b7af9b3d2aabd6d3d161103

Observation 60efd1e8-6ef1-4470-9b09-ffaf730b17a8 · outbound

This paper cites Bevformer: learning bird’s-eye- view representation from lidar-camera via spatiotem- poral transformers,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Bevformer: learning bird’s-eye- view representation from lidar-camera via spatiotem- poral transformers,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:10.737724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.548022Z digest=sha256:3218508cc6d7be2c9f2e8b6a359211b781a859523b87ee0424b1515bea66038d

Observation 13f62e17-857c-4cc5-8769-e33ca9238a43 · outbound

This paper cites Ega- depth: Efficient guided attention for self-supervised multi-camera depth estimation,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Ega- depth: Efficient guided attention for self-supervised multi-camera depth estimation,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:10.597236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.554368Z digest=sha256:f3dd9b9e0da3eb8a992cee8eb6f636931d4d7b88b480ee6244df103bbc091c84

Observation e79ebaf7-1f1e-4b72-991a-0aa03f6a2c33 · outbound

This paper cites Multimae: Multi-modal multi-task masked autoen- coders,.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Multimae: Multi-modal multi-task masked autoen- coders,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:52:10.430743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T20:52:06.559453Z digest=sha256:9f943eceefe31d72c7918a2cf9dc7b69b39896855b78f7cdd50f51492d92b518

Pith citing papers

No inbound Pith citation observations are available.