Pith. sign in

Paper Citation Record · LEDGER

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving

As of 23 August 2026, this Paper Citation Record lists 88 of 88 outbound references and 0 inbound Pith citation observations for arXiv:2504.12709.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12709 v1

Coverage vector

measured 88 of 88 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:29:04.998023Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

88 of 88 outbound references displayed

  • verified exact0
  • verified fuzzy60
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e72e2a1e-c9e3-4114-8066-9b2e0fb63bd4 · outbound

This paper cites Qwen Technical Report.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Qwen Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.449609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.449609Z digest=sha256:86314c87cb540df9b27d2005d507840f5d2bbecda054317cca03240e41d1e462

Observation d66fcc00-127b-4914-8c64-4c9edd6da90a · outbound

This paper cites Transfusion: Ro- bust lidar-camera fusion for 3d object detection with trans- formers.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Transfusion: Ro- bust lidar-camera fusion for 3d object detection with trans- formers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.455259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.455259Z digest=sha256:ec0600cb8bf182c40b6396e9412ce5129a50f96edefd7e685474c1af53cca851

Observation 63b4f935-7f43-4d97-a21d-8636b7f5ac9e · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving nuscenes: A multi- modal dataset for autonomous driving

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.460448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.460448Z digest=sha256:1d005a602f0a0ebe41f5077684b584c603b9c8095183ce63b58316acb4e886ad

Observation 4d0d5267-deb9-42d4-b02b-d7054dc782d0 · outbound

This paper cites Nuplan: A closed-loop ml-based plan- ning benchmark for autonomous vehicles, 2022.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Nuplan: A closed-loop ml-based plan- ning benchmark for autonomous vehicles, 2022

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.465599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.465599Z digest=sha256:2c1869b25997d3019c2df48c38cafe972afab46f35a4549537a877b3e526a272

Observation 07e58b06-8d32-41f8-8cd1-f897bee6c9ce · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Emerg- ing properties in self-supervised vision transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.470630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.470630Z digest=sha256:faa19f1b230f74468f5d15f6b0d52b36d093d9449a414615a5dff8684d785325

Observation c72d8efa-0a94-4c10-84fd-9b4d62fd71a8 · outbound

This paper cites A simple framework for contrastive learning of visual representations.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving A simple framework for contrastive learning of visual representations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.476510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.476510Z digest=sha256:33c9d064eb548c4e5bb432a0293dd06fcdb25a66071833245772f67852b072c7

Observation f7d1ff38-0f05-4940-9f3f-e48aa94a35b0 · outbound

This paper cites Multi-view 3d object detection network for autonomous driving.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Multi-view 3d object detection network for autonomous driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.481756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.481756Z digest=sha256:b226312f0e447bc6f38ddbc2080c5beeaa4b2cbe132ec5701b1503da6dc4725e

Observation 607f9cf9-42a5-46a3-a452-eb8cb80a8206 · outbound

This paper cites Improved Baselines with Momentum Contrastive Learning.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Improved Baselines with Momentum Contrastive Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.487054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.487054Z digest=sha256:56acf81830ca215b83a8d7ec1726f202b16b4b78b2a76a60ed8cd5ff2a56e509

Observation b6a1d6a7-bf0b-4690-8610-700ad7391b4e · outbound

This paper cites BEVDistill: Cross-Modal BEV Distillation for Multi-View 3D Object Detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving BEVDistill: Cross-Modal BEV Distillation for Multi-View 3D Object Detection

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.492171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.492171Z digest=sha256:5f16b19d607d9273ab8296c5bab6d499435a3119c75f45e6b946bb4a7f001831

Observation d2ed17b4-7f48-4004-9557-5f5304daa1e9 · outbound

This paper cites Back-tracing representative points for voting- based 3d object detection in point clouds.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Back-tracing representative points for voting- based 3d object detection in point clouds

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.497205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.497205Z digest=sha256:2c498a3237c97ee0631a5e1b08d23371eaedfa11228d1933a1be53352d1e6230

Observation d0c19000-001c-4c47-89e8-26aa6b069643 · outbound

This paper cites Lyft 3d object detection for autonomous vehicles.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Lyft 3d object detection for autonomous vehicles

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.501994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.501994Z digest=sha256:e1ef9223bf9b1e5352e0b4396a218e9d27b284f2de41aa6e75f87e380edabaa3

Observation 5218d50f-fe7a-4ec5-a78c-08f5947374c8 · outbound

This paper cites Scaling vision transformers to 22 billion pa- rameters.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Scaling vision transformers to 22 billion pa- rameters

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.506366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.506366Z digest=sha256:01a768a20f801a579d1f63d761758a2c0ded11003300caf36172db6575b24de0

Observation 92757011-8770-453a-9af1-5733e097c343 · outbound

This paper cites V oxel r-cnn: To- wards high performance voxel-based 3d object detec- tion.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving V oxel r-cnn: To- wards high performance voxel-based 3d object detec- tion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.333426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.511138Z digest=sha256:8f9fe5d2cb4d777940111fc19f6ae834768cba589896fec2483665c5e3a0644b

Observation b0fd37f0-c2e6-45b8-95b2-72c1116ad499 · outbound

This paper cites Embracing single stride 3d object detector with sparse trans- former.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Embracing single stride 3d object detector with sparse trans- former

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.515500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.515500Z digest=sha256:eafe4cf245a4139b6865732a9438c06e0fd4bbe76ecbf426636ad2679c0358de

Observation a680525e-cc87-4285-b76a-bb8b4c30ed88 · outbound

This paper cites Sparse dense fusion for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Sparse dense fusion for 3d object detection

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.307496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.520656Z digest=sha256:8f0a3182d07f0739fe5e4a1dee6109ed7e6d5df6e1b3e106be7490adcf0b087f

Observation 98a104f9-f67f-4936-ab26-be96a0c0fb96 · outbound

This paper cites Momentum contrast for unsupervised visual rep- resentation learning.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Momentum contrast for unsupervised visual rep- resentation learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.292320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.525685Z digest=sha256:2ee9c784f1d145d962512dd61e9a5614718435acceabbcad68502d1df4c1fa21

Observation 66d0ca62-ca8f-416a-bac3-2370bec36515 · outbound

This paper cites Masked autoencoders are scalable vision learners, 2021.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Masked autoencoders are scalable vision learners, 2021

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.276923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.530641Z digest=sha256:ac89595fa4c11b31b92bc262ba6e31e8d41eac830f333a3e7e40dcef82160f08

Observation 089bfaa8-7625-40f0-8fe2-7b840912a737 · outbound

This paper cites Data-efficient image recognition with con- trastive predictive coding.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Data-efficient image recognition with con- trastive predictive coding

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.262228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.535260Z digest=sha256:990f3483a7cb814910bfbd082561b3c6b7c8d11ddf4487ecc1f401ff1a3dde2b

Observation 73d38042-0a3f-4c3f-a310-8657655e4257 · outbound

This paper cites Masked autoencoder for self-supervised pre-training on lidar point clouds.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Masked autoencoder for self-supervised pre-training on lidar point clouds

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.247260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.540202Z digest=sha256:8031444e8a13045140026af62bf147cc37a8f1e1fe6f8ce55833567d1b18bb5d

Observation 6f9926f3-f595-4ed4-a807-149d362b0cb7 · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.544790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.544790Z digest=sha256:75bfbe2f38adbe3ab7920ddf54e87724d6c6d46b83a337edc6197b02a6ad5933

Observation d9eaeec5-7be3-472c-9334-e772ad785941 · outbound

This paper cites Ep- net: Enhancing point features with image semantics for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Ep- net: Enhancing point features with image semantics for 3d object detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.229691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.550010Z digest=sha256:6aaf06a4bd45d77860579d54f0f181b6bb0ff9aad4b8f54f8526611dae9ef00c

Observation 7fd0c8c0-7fc3-4343-8305-72f792b8985a · outbound

This paper cites Tri-perspective view for vision- based 3d semantic occupancy prediction.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Tri-perspective view for vision- based 3d semantic occupancy prediction

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.210127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.555109Z digest=sha256:bfafa667fa8bebe941c2eacb7e7f4c52e045614f894a3695cf405655f28d423f

Observation 97b86348-507e-4bb5-80a9-61c8353fcab0 · outbound

This paper cites Learning semantic segmentation from multiple datasets with label shifts.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Learning semantic segmentation from multiple datasets with label shifts

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.192703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.559702Z digest=sha256:c264244420ef5996470758060b91def6e683ed346bdaad754365f69cc12aeafb

Observation 71585212-1172-44f1-acdd-260fce955433 · outbound

This paper cites OccMamba: Semantic Occupancy Prediction with State Space Models.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving OccMamba: Semantic Occupancy Prediction with State Space Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.564292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.564292Z digest=sha256:153b073b4d6606e8563cef9fb494171d92e757f1eff683edf236b9f41ab650fe

Observation ec34b3f8-a8a7-49d0-bc9c-a944839062a6 · outbound

This paper cites Deepfusion: Lidar-camera deep fusion for multi-modal 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Deepfusion: Lidar-camera deep fusion for multi-modal 3d object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.172644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.569341Z digest=sha256:d678aeecdf8a7e21cb4192e213b0ade47a495341dc6ddcb7d204bc9914544f0a

Observation 65131e67-1e96-450d-8aa8-8904315e788b · outbound

This paper cites Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.152448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.573629Z digest=sha256:7686c51f7aecdd31aae792b5e3d3ea9578ceea233b37b5921dc14bb518c53852

Observation b4e4d9f2-a8d4-4bf5-b411-d50e69850c26 · outbound

This paper cites Bev- former: Learning bird’s-eye-view representation from multi- camera images via spatiotemporal transformers.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Bev- former: Learning bird’s-eye-view representation from multi- camera images via spatiotemporal transformers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.136244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.696083Z digest=sha256:5496144785eeabbb53911870a761689255414a3051f48379d74aa03dcc12bdd8

Observation 0a9ca2cf-a20f-4435-9e2d-2699ae2da981 · outbound

This paper cites Bevfusion: A simple and robust lidar-camera fu- sion framework.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Bevfusion: A simple and robust lidar-camera fu- sion framework

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.119937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.700912Z digest=sha256:c18c1a0e9dfff611618d19d6644d064dc6afe831c58d3d040bb2544d3871c715

Observation 28fca571-aead-4149-bba9-21e4922bc6ec · outbound

This paper cites Geomim: Towards better 3d knowledge transfer via masked image modeling for multi-view 3d un- derstanding.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Geomim: Towards better 3d knowledge transfer via masked image modeling for multi-view 3d un- derstanding

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.103232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.705414Z digest=sha256:ba63140f37b05c086ab0ce7df16d587c9cea576610598512048e4167d2ef1813

Observation 0fc43de0-54a0-4fc7-b8ba-4f6b99ac018a · outbound

This paper cites P4Contrast: Contrastive Learning with Pairs of Point-Pixel Pairs for RGB-D Scene Understanding.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving P4Contrast: Contrastive Learning with Pairs of Point-Pixel Pairs for RGB-D Scene Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.709889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.709889Z digest=sha256:b55903236067399af43702b6ee176c9920c06ec58bd9225aa5bcc1e75b3ad3eb

Observation fe8faa54-45f3-4beb-81ee-2ef73f05460d · outbound

This paper cites Petr: Position embedding transformation for multi-view 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Petr: Position embedding transformation for multi-view 3d object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.083322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.715081Z digest=sha256:1ebcacbb461d5ca03944affd6d6b16ce51c52ed42dffe4d25bc5978b735f250a

Observation 46289bd7-e323-4e71-a664-46562fdec579 · outbound

This paper cites Swin trans- former: Hierarchical vision transformer using shifted win- dows, 2021.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Swin trans- former: Hierarchical vision transformer using shifted win- dows, 2021

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.720385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.720385Z digest=sha256:30b368b35bc57f3dc4a80838f0170a66b988701143d4f975b479f8e47f7c7682

Observation 82dc1b25-46a9-4e80-908c-86673467a312 · outbound

This paper cites Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.055953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.725189Z digest=sha256:93aaa1c9518c44b16e7305f7ea09a65b00fbf52f4fbd45cfd96bcbee921639bd

Observation de6c53ea-46f2-45c9-b663-ca8702143c84 · outbound

This paper cites One Million Scenes for Autonomous Driving: ONCE Dataset.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving One Million Scenes for Autonomous Driving: ONCE Dataset

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.730095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.730095Z digest=sha256:37782f003122126d49a40396b79e57a5367cee6f77a00a9b118d3fb9d9a8da4b

Observation 8e771814-ec7a-49f8-a268-5fe4ee774820 · outbound

This paper cites V oxel transformer for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving V oxel transformer for 3d object detection

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.037906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.735732Z digest=sha256:4390666962f9d02c6a969c4e2a3c73717b227a7761888e843cd90152dd02b0a9

Observation a2acd7c4-c104-4838-a26a-348a22eea51c · outbound

This paper cites Multi-camera unified pre-training via 3d scene recon- struction.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Multi-camera unified pre-training via 3d scene recon- struction

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.020710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.740615Z digest=sha256:7fc209410dcd65ac2acbc0970bff0822eb7f0534e57e8b0e3bf9efad9f6bfedd

Observation 382dff7f-c24b-4d35-bf0f-3a5c349e3da2 · outbound

This paper cites Uniscene: Multi-camera unified pre-training via 3d scene reconstruction for autonomous driving, 2024.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Uniscene: Multi-camera unified pre-training via 3d scene reconstruction for autonomous driving, 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:06.003658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.746036Z digest=sha256:fd68716a8fbd20445b2aeb8b7707a16440609bd4f739328543a6c10bd5ec94cc

Observation b3a4e140-91d8-41d1-8c22-e8b7066e602e · outbound

This paper cites Self-supervised learning of pretext-invariant representations.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Self-supervised learning of pretext-invariant representations

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.985516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.750966Z digest=sha256:993a85a53b6895cbbd4aed6ab0d6d1580e6ba16ff993de665d24453238634a8e

Observation a17bb197-64d9-4fa7-92fd-d4802ecd0248 · outbound

This paper cites Masked autoencoders for point cloud self-supervised learning.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Masked autoencoders for point cloud self-supervised learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.969142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.756015Z digest=sha256:5e1131ef5ce51f187d2138deca382f622533f1190e8c68a52fd8203fa789784f

Observation 8d4b43a7-29c5-46a6-8169-1568233ce3da · outbound

This paper cites Simpletrack: Understanding and rethinking 3d multi-object tracking.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Simpletrack: Understanding and rethinking 3d multi-object tracking

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.952969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.760693Z digest=sha256:b599855e4dd4453567d845877685713100a8621cece2c7cdbf0fed5eded6f02c

Observation b089837e-43fb-49fa-a2a7-b3171c396649 · outbound

This paper cites Standing between past and future: Spatio-temporal modeling for multi-camera 3d multi- object tracking.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Standing between past and future: Spatio-temporal modeling for multi-camera 3d multi- object tracking

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.935375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.765204Z digest=sha256:0e725ec8911bb3fb00e52c5c2c84948c597bbb773eecceca0f719e0e69cc871a

Observation a58a8716-1fab-4a4d-90eb-31829caef806 · outbound

This paper cites What Do Self-Supervised Vision Transformers Learn?.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.769709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.769709Z digest=sha256:7830a7c429bd74448149262b8a2b3ee8823769c3825b6524979851a403b96346

Observation 22d1af9b-6df6-4419-a516-9cb00d7e5871 · outbound

This paper cites Lift, splat, shoot: En- coding images from arbitrary camera rigs by implicitly un- projecting to 3d.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Lift, splat, shoot: En- coding images from arbitrary camera rigs by implicitly un- projecting to 3d

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.918876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.774663Z digest=sha256:2c44fd0651f9115e2346305b6ebb2ac026289da7d4367ffef11484b283549210

Observation bd90abcb-f341-43ca-8e03-8a06373428f1 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Deep hough voting for 3d object detection in point clouds

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.901850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.779336Z digest=sha256:8cf623c90bafb4eba6148822efe4a3b79792abcb2757f4f9cfec9d0ef6b94b57

Observation 12d9a611-b081-4544-b5a0-2dbb6954f258 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Learning transferable visual models from natural language supervision, 2021

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.883422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.783829Z digest=sha256:4afb22fb914c0da50753ee812bdef35bca652381c8a3d6ac2e153828777dc3be

Observation 067ba32f-7414-46fa-a9d4-3e7961cf9b8d · outbound

This paper cites Image-to-lidar self-supervised distillation for autonomous driving data,.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Image-to-lidar self-supervised distillation for autonomous driving data,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.867407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.788501Z digest=sha256:e927f57ab8388630f090bcac2c09de84cc24e3c4330b993776d63d1b122bdea9

Observation 441789eb-f4c8-4084-9466-02b10e4434ce · outbound

This paper cites Pointr- cnn: 3d object proposal generation and detection from point cloud.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pointr- cnn: 3d object proposal generation and detection from point cloud

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.793465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.793465Z digest=sha256:e2e6eae56d24e4db2625db804e00aaf0c11956ca33661068d0485ac752de87de

Observation bc8433c6-2d9a-466a-901e-e83bb8d9437d · outbound

This paper cites Pv-rcnn: Point-voxel feature set abstraction for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pv-rcnn: Point-voxel feature set abstraction for 3d object detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.840845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.798848Z digest=sha256:5620e817f21a4dd18d22402c00687973b28c38d6ba8ec212e37cfac8d50cd619

Observation 711ba46e-d296-4241-895d-c68d1081852f · outbound

This paper cites From points to parts: 3d object detec- tion from point cloud with part-aware and part-aggregation network.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving From points to parts: 3d object detec- tion from point cloud with part-aware and part-aggregation network

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.824006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.803561Z digest=sha256:67026f126d7a7d2119611e1bba97e17064883c990cd8f411a9ed3813a83d0ccf

Observation 759d9ceb-e6c0-44f1-a6e2-4e4c2f273872 · outbound

This paper cites Mvx- net: Multimodal voxelnet for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Mvx- net: Multimodal voxelnet for 3d object detection

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.808012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.807789Z digest=sha256:f702deb38f6abbc96beab1c38ff33993d7c43e846e4fb9f228e4f4855035dc20

Observation e0bd6fcd-9720-42f1-96b4-cec0a5d0e4d8 · outbound

This paper cites CALICO: Self-Supervised Camera-LiDAR Contrastive Pre-training for BEV Perception.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving CALICO: Self-Supervised Camera-LiDAR Contrastive Pre-training for BEV Perception

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.812438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.812438Z digest=sha256:bd63a9b1d4f06309b503fc16c406403e5ca849686e2be8b1b45717ec4272eb2e

Observation c774d46b-34e7-42cb-a9b4-ac89e00497aa · outbound

This paper cites Scalability in perception for autonomous driving: Waymo open dataset.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Scalability in perception for autonomous driving: Waymo open dataset

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.791952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.817238Z digest=sha256:1c15074d3c8f9843d3c07db1df26cee77b6664de155e045e97e00bf59cee14e2

Observation 53296ce4-7430-4028-99b0-6954e4a48747 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving LLaMA: Open and Efficient Foundation Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.821696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.821696Z digest=sha256:847311fadbe72f98e3517d62341fc71b885d7f4bc5d7d9a86a080dc7ac7185f1

Observation 9013c791-794b-4faa-a182-5b20bf3d9467 · outbound

This paper cites Pointpainting: Sequential fusion for 3d object de- tection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pointpainting: Sequential fusion for 3d object de- tection

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.826720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.826720Z digest=sha256:5bb93468a68ccc23caa484ba86ef4637bcd8266173380c1993a43c53c871dc91

Observation 7e97f82f-3563-41fd-b699-934a97a97fec · outbound

This paper cites Pointaugmenting: Cross-modal augmentation for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pointaugmenting: Cross-modal augmentation for 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.766198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.831294Z digest=sha256:73cdde501fc44a5a4b4dd0e1da1cb380a056561ebfc449d586e7b7c4876d5665

Observation 6a35e5f2-8da3-4e26-b80b-d5ba1351f7cb · outbound

This paper cites Rbgnet: Ray-based grouping for 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Rbgnet: Ray-based grouping for 3d object detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.750442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.835839Z digest=sha256:bd11981055b94e3b1ef7cc33a3f7ed04a4df076df4114891b8d67d664af1b8f4

Observation f2524166-d622-4b6a-af02-e075a62f82fa · outbound

This paper cites Dsvt: Dy- namic sparse voxel transformer with rotated sets, 2023.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Dsvt: Dy- namic sparse voxel transformer with rotated sets, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.735440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.840626Z digest=sha256:f4044aa1cc59dde6513f3232c9d2619db78d004706422c95fede6e117c6f7943

Observation 2d4dbffb-8fbe-4a31-a2d8-ad18531ee729 · outbound

This paper cites Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.720762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.845494Z digest=sha256:a50003c8113d104ab5e51ac3d19b21684dc944be2d64a28cbebaf2f347d897e6

Observation 007d9845-ff13-4cfa-b028-321adec6f841 · outbound

This paper cites Cross-dataset collaborative learning for seman- tic segmentation in autonomous driving.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Cross-dataset collaborative learning for seman- tic segmentation in autonomous driving

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.704281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.850433Z digest=sha256:652ab9c56afaf0ef1ec2f698f3afb61a143d32fbfad2424dc3033edb54130a01

Observation 16d8a9ce-5bd2-4f14-816d-73aa3931097d · outbound

This paper cites Mv- contrast: Unsupervised pretraining for multi-view 3d object recognition.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Mv- contrast: Unsupervised pretraining for multi-view 3d object recognition

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.687952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.854929Z digest=sha256:39a1f17ba07cd5e96b85a2f525b07b8df8f70dde88739e0e7e46cf3a44f1ab94

Observation 795e196f-bedf-441e-8bf3-6abc2b608bdf · outbound

This paper cites Tracking everything everywhere all at once.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Tracking everything everywhere all at once

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.669575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.859569Z digest=sha256:f834b67db8e2d8dae41143eb7edb6bb730fe18c75bc7a1f3ef2dba3b05d1fef9

Observation e6942a76-e74d-4def-a3ed-e016705bbd7d · outbound

This paper cites Fcos3d: Fully convolutional one-stage monocular 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Fcos3d: Fully convolutional one-stage monocular 3d object detection

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.651697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.864755Z digest=sha256:16799549a9ffba7e0cbff2f212e787e9d7ec4bc3262a59ca97649b240ebea0b7

Observation 815aa0ba-092b-4f3a-abc7-051cfe9ceefa · outbound

This paper cites Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.630559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.869666Z digest=sha256:747f41b60c719e9e0e9984adb322cd97f5d138004bbb0b84e7559eaa61e3438d

Observation 2e1e4659-1248-4419-b092-7480dd02e2fb · outbound

This paper cites Detr3d: 3d object detection from multi-view images via 3d-to-2d queries.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Detr3d: 3d object detection from multi-view images via 3d-to-2d queries

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.609928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.875066Z digest=sha256:fd0d381a597409b6d60ffa16a6f8193a82e7fc52bc0e5d5cdd32ca08d0426301

Observation 93aaf10b-9847-40ba-ad20-e5dcc37610dc · outbound

This paper cites Masked feature predic- tion for self-supervised visual pre-training.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Masked feature predic- tion for self-supervised visual pre-training

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.589009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.879815Z digest=sha256:747c72db5a6d83e512de35e6e727d389e1287a42ce64407f90ad7a08c71f8c94

Observation c53a3cda-4791-4f94-8b56-9554a4035a2a · outbound

This paper cites Virtual sparse convolution for multimodal 3d ob- ject detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Virtual sparse convolution for multimodal 3d ob- ject detection

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.571702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.884523Z digest=sha256:d6e382ea2a1b0f494d46e0c125b3da849d869378579af5717f31f96779ee0fe1

Observation 7bd373c8-3107-4676-a4a4-f7d5f3890473 · outbound

This paper cites Towards large-scale 3d representation learning with multi-dataset point prompt training.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Towards large-scale 3d representation learning with multi-dataset point prompt training

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.554412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.890415Z digest=sha256:af172aa06d46db329f438c9006abb6eb3e55595906082307058aae1a02432392

Observation faa3470d-ce8e-44c9-ae3c-ae33d20e704e · outbound

This paper cites M$^2$BEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving M$^2$BEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.895032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.895032Z digest=sha256:1df5f939093d4c28378c6f08c9240571e5853359a28a755895046b4cafa3470a

Observation 1c8a4479-cd35-4eac-b096-376af9139f9f · outbound

This paper cites Pointcontrast: Unsupervised pre- 11 training for 3d point cloud understanding.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pointcontrast: Unsupervised pre- 11 training for 3d point cloud understanding

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.537360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.899905Z digest=sha256:e4f61b36ca57112996025d94de8f1c6eab2c4fe670ff12e1f9c6a9bbd6ccf03c

Observation 1d3b27e7-539a-42f5-af28-bcbac3e648d8 · outbound

This paper cites Qi, Leonidas J.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Qi, Leonidas J

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.520405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.904654Z digest=sha256:7f5f8c39b5f93fc3d252493f6912622681a0e40cecab2a7b03894909bcccf8f9

Observation 9568c564-2863-49ec-9f9c-f74133bfdaef · outbound

This paper cites Simmim: A simple framework for masked image modeling.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Simmim: A simple framework for masked image modeling

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.503265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.909297Z digest=sha256:9b50463f3863310fbc079acede900b8035a96943b91ba51f1d9f6ddf9184abb5

Observation 0158ed4e-f3c0-4d85-bd18-25eb96396fa4 · outbound

This paper cites Cross modal trans- former: Towards fast and robust 3d object detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Cross modal trans- former: Towards fast and robust 3d object detection

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.484923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.914256Z digest=sha256:d1de15c4d4b3e3adf9d64d042e7e69a4defeeee1447d3bce8a253fdc30046ab7

Observation e8a26a85-bb4c-4141-9708-ec40c172cd99 · outbound

This paper cites SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.918936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.918936Z digest=sha256:6a55c23f12cb6b03406b5bc1b34321cfdcfd10a792a1c734055eb31f3de90085

Observation e66dd811-29ac-4201-9bd2-437a68ce9aae · outbound

This paper cites Second: Sparsely embed- ded convolutional detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Second: Sparsely embed- ded convolutional detection

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.469494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.923624Z digest=sha256:5ed12addf825427ac023f45df363bcfb7b495a8f83c720cc098c3c35ec2fb393

Observation c5cddb47-7c53-4fce-89f6-2b76cab1f634 · outbound

This paper cites Boosting 3D Object Detection via Object-Focused Image Fusion.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Boosting 3D Object Detection via Object-Focused Image Fusion

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.929624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.929624Z digest=sha256:6a60156cef22a723fa8626c88617098db03f17a92d0951fd0119737333d3de7f

Observation 8bae362b-5d29-441d-972b-3ce993682baf · outbound

This paper cites Gd-mae: Gen- erative decoder for mae pre-training on lidar point clouds,.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Gd-mae: Gen- erative decoder for mae pre-training on lidar point clouds,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.453699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.935221Z digest=sha256:c81802e45f9545fe03bb81a44a6358408aa8650f0462ab8f62ec834cdb91a25d

Observation 13aefb75-e5ab-4689-9634-7608ad071a3a · outbound

This paper cites Pred: pre-training via semantic rendering on lidar point clouds.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Pred: pre-training via semantic rendering on lidar point clouds

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.436633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.941512Z digest=sha256:860b188a1bd4e57fe92addde53d1ca414a43fc306fc5ae5031326592cec18dd2

Observation 27384206-382a-4dfe-9272-eb97fc8ee122 · outbound

This paper cites 3dssd: Point-based 3d single stage object detector.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving 3dssd: Point-based 3d single stage object detector

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.419325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.946603Z digest=sha256:ba278434ccf973eea63fc92a52fd90462140c974664ce30f4e31e76e2aa90c7c

Observation 74519015-6f4d-48fc-b2c5-18df8ee25ba6 · outbound

This paper cites Center- based 3d object detection and tracking.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Center- based 3d object detection and tracking

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.403184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.952697Z digest=sha256:40f260c4a65dd776433735ed133e403b6f2734e2a74cd6ff26eca886b5576a9a

Observation f39a6e44-707e-43a2-89ef-9604fd80d027 · outbound

This paper cites 3d-cvf: Generating joint camera and lidar fea- tures using cross-view spatial feature fusion for 3d ob- ject detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving 3d-cvf: Generating joint camera and lidar fea- tures using cross-view spatial feature fusion for 3d ob- ject detection

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.386205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.958493Z digest=sha256:d36c87a6bf85d5961a755255df9f86f804601ffa5d09c752dd5740aa6a91bc08

Observation e9b68d7a-11af-47ae-8541-f2bd9cbdf0d8 · outbound

This paper cites SparseLIF: High-Performance Sparse LiDAR-Camera Fusion for 3D Object Detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving SparseLIF: High-Performance Sparse LiDAR-Camera Fusion for 3D Object Detection

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.963199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.963199Z digest=sha256:af061abf0d894de010232fc39700de59852e5299f4eeb8d08579f2a7219598f1

Observation 661e0502-73ce-4c45-a5e8-8c80045eeb71 · outbound

This paper cites Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Point-m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.368023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.968253Z digest=sha256:0dbe7d257f2bcf858f0ae0a6034166efd8dbd8a0365aac12a83f3cccdee6c18c

Observation d82f25d1-9b79-4b79-a14d-21f7a3d85aaf · outbound

This paper cites Hvdistill: Transferring knowl- edge from images to point clouds via unsupervised hybrid- view distillation.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Hvdistill: Transferring knowl- edge from images to point clouds via unsupervised hybrid- view distillation

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.351182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.973466Z digest=sha256:d3cc7c8d8101453f2c0ad3bf2eedf9f196a69b455ace04f3e9b60aa4d60abd7e

Observation 499a02cd-1d0b-4ce9-85f6-5bf4882c3fe0 · outbound

This paper cites Self-supervised pretraining of 3d features on any point-cloud.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Self-supervised pretraining of 3d features on any point-cloud

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.335464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.978031Z digest=sha256:d38298f259c9d08448d62533aca6e7101f85a6c45296fc97c059826256dfebd7

Observation 0ff301f7-a203-4bae-a331-d710e3274e89 · outbound

This paper cites Unidistill: A universal cross-modality knowl- edge distillation framework for 3d object detection in bird’s- eye view, 2023.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Unidistill: A universal cross-modality knowl- edge distillation framework for 3d object detection in bird’s- eye view, 2023

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.319944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.982624Z digest=sha256:ca49e815370b15396f8b0ca965e562b35b6fbf0ec89d140ea835879043b629c3

Observation 931d9d47-bb01-4354-938e-3d9b8cae00f7 · outbound

This paper cites Sim- ple multi-dataset detection.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Sim- ple multi-dataset detection

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.302684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.987292Z digest=sha256:827b7672726c42cbccf7a9bccdb364c5807e8b1012f640a695db96d3a9c2899a

Observation 621502f4-3981-4bbe-bba0-f7bbd8d943bf · outbound

This paper cites Cylindrical and asymmetrical 3d convolution networks for lidar seg- mentation.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving Cylindrical and asymmetrical 3d convolution networks for lidar seg- mentation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.992486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.992486Z digest=sha256:13665fa51971b2e708d9a25002b9c5c0fdf6587c171fa88e2c6d575041b385ac

Observation 938485b7-09dd-447a-851d-1b49ebbb6fdf · outbound

This paper cites NuScenes.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving NuScenes

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:29:05.273987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T12:29:04.998023Z digest=sha256:0ffa7fe889035090e7041eedfa519f99e1c94eec968fc1dbb01348a8dd367a6f

Pith citing papers

No inbound Pith citation observations are available.