Pith. sign in

Paper Citation Record · LEDGER

BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2203.17270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.17270 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:47:13.263203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:08:22.583557Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e62581fb-6d06-48b9-b7a1-b77bca286d76 · inbound

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning cites this paper.

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:18:49.535690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T03:16:03.058129Z digest=sha256:ff93d30735b564cecbc596cdeb36eb093e56a717fb37555da4ce0a7ea541e6fc

Observation b43c0cbb-c37b-456c-8b1c-44108f63848e · inbound

Enhancing End-to-End Autonomous Driving with Latent World Model cites this paper.

Enhancing End-to-End Autonomous Driving with Latent World Model BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:38:51.943862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T07:38:51.885383Z digest=sha256:0584d7ac270b3361a7981f293c92c7c2b6416f575de3024c326aa17d5389ab6e

Observation 0cd24a7c-0bac-45db-8965-9174ce6315cf · inbound

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving cites this paper.

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:25:51.484526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T21:25:20.390835Z digest=sha256:1c6e70f2dd7bbafb31a3f40ceea35838d624161f68de3cebd444d135897a1988

Observation 80bad068-16b6-43ad-80b8-bf7d4bd7d0d3 · inbound

Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving cites this paper.

Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:24:23.914142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T15:24:23.756052Z digest=sha256:f6fe9843c4f10a98205316e88d3236837da91f167e03e6c5439da740043e6311

Observation f69f2ee9-8c18-4e30-9e69-12438284e762 · inbound

CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving cites this paper.

CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:47:13.263203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:47:13.263203Z digest=sha256:9a4bcd78823ea67a6e9a3b71ef07e3b72a610accbdd9c66fc26a7360ecbad3ee

Observation c4908c9d-5c01-49b5-9620-2d62f93274a1 · inbound

SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels cites this paper.

SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:11:45.426681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:11:45.426681Z digest=sha256:c1866d42a1056fb34bab418191bed45e2267de571b8ae0a0b41372026ef86d86

Observation 096a0178-c3d0-405d-aea7-f3aec25945d9 · inbound

BEVCALIB: LiDAR-Camera Calibration via Geometry-Guided Bird's-Eye View Representations cites this paper.

BEVCALIB: LiDAR-Camera Calibration via Geometry-Guided Bird's-Eye View Representations BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:12:15.488257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:10:31.571968Z digest=sha256:c839668cad8b92cea2d3519698439bb02839cc85660d1e2b15dcd61a2116d7a1

Observation a6edca6f-c2c7-4d14-8e66-efec5e422ac1 · inbound

S2GO: Streaming Sparse Gaussian Occupancy Prediction cites this paper.

S2GO: Streaming Sparse Gaussian Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:09.658408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:09.658408Z digest=sha256:a95e82e75855544247ba1e102cce1d208c865f60de7790e6ddba7a0046af97a7

Observation 717ab366-15ab-45aa-8d82-ac83504eb1f4 · inbound

FocalAD: Local Motion Planning for End-to-End Autonomous Driving cites this paper.

FocalAD: Local Motion Planning for End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:17:15.800775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:15:53.988156Z digest=sha256:6c6a50b599cafc4c43eab5dee2c42d88c1f8886b825573122bc193b7fc8048f3

Observation bd4ee98a-b979-4876-a424-03869bb94fc0 · inbound

Learning to Generate Vectorized Maps at Intersections with Multiple Roadside Cameras cites this paper.

Learning to Generate Vectorized Maps at Intersections with Multiple Roadside Cameras BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:22.884613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:22.884613Z digest=sha256:31572d6bb9fb90425a827f59fbbbc71b78bf7e82d56690477a6072f0fef4acec

Observation aa3b2e12-822b-4e27-a50c-77f6bfcf1f42 · inbound

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection cites this paper.

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:42:00.957398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:42:00.957398Z digest=sha256:3e748206561e018d76de84ad329da2f2df75339cb07350216198f0c39636188f

Observation d418e1cc-9c3e-4074-99fe-07da8eb6ef48 · inbound

Progressive Bird's Eye View Perception for Safety-Critical Autonomous Driving: A Comprehensive Survey cites this paper.

Progressive Bird's Eye View Perception for Safety-Critical Autonomous Driving: A Comprehensive Survey BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-05T22:06:55.337304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:06:55.337304Z digest=sha256:49758119d82cf531012a1158398beee03730d85ac26970bce0fe037590a10662

Observation 45ada445-1397-4301-9130-0e54f13ea8f8 · inbound

SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer cites this paper.

SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T14:55:48.356558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:55:48.356558Z digest=sha256:82d1353bef624b12fcbb2ed28f152b78cb68ee66649fee385471642a61f42e73

Observation 0b72d4ed-8679-4e5f-a2b2-d9c4406ad130 · inbound

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models cites this paper.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.959381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.959381Z digest=sha256:10226a9bf9315b1e6d77d2eef0fc596a1bac960f45bfeb0a51223f4fb5353e21

Observation 6ded0fc4-1f91-46a2-8582-bb5106ccb31b · inbound

Semantic Causality-Aware Vision-Based 3D Occupancy Prediction cites this paper.

Semantic Causality-Aware Vision-Based 3D Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:48:57.622231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:48:57.622231Z digest=sha256:c8c9ed7dadfe901876423a27e739f62e59fa3798b473be28c72a58abbb17833b

Observation 73dae037-6dbb-4b3e-83b8-84eb82da462a · inbound

DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion cites this paper.

DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T09:29:20.708420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:29:20.708420Z digest=sha256:bc6b170c3b542161ce264314da593c128148bc910d0ea8efde6af103cdb43eab

Observation f2ee8ce3-f025-44bc-a2cf-8780d356ba4e · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.632850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:096034804072ad89bcfed1d48bb58a8e6888c5227948dfb5aa94efd2482ac9ac

Observation 4637504e-1921-45fc-ad75-5ca91801d746 · inbound

Fast-BEV++: Fast by Algorithm, Deployable by Design cites this paper.

Fast-BEV++: Fast by Algorithm, Deployable by Design BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T18:44:18.886433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T18:42:37.182403Z digest=sha256:7424b7b2103a0e23e5dbc530a8fd83009735f250927795f159c3576f8442a976

Observation 7fddc862-931f-4209-9130-5905bd0d3039 · inbound

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving cites this paper.

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T12:06:27.585778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:06:27.585778Z digest=sha256:8749f31888d365bf673f0daa32eba41a5a0a40f26cb0a8b4b28f3943a3c62068

Observation f73c6746-4829-4ace-8259-f46256c95782 · inbound

More than the Sum: Panorama-Language Models for Adverse Omni-Scenes cites this paper.

More than the Sum: Panorama-Language Models for Adverse Omni-Scenes BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:00:02.865442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T13:58:18.635091Z digest=sha256:a458067c340c344ef12d00c6db9247debc153b46820213a7ddaf5d85a9bdc555

Observation b4e21461-f0d2-488e-9dbd-eb891e87f78c · inbound

BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving cites this paper.

BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:03:12.414025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:00:21.130075Z digest=sha256:10afa4b41587394d3555031894b3b671dc7df454729dd14ca9d41aa5a7c2b7d2

Observation 442fffd0-908d-4b5e-b539-94bf4942ff64 · inbound

Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving cites this paper.

Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:00:56.801261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:43:52.542973Z digest=sha256:3dfcd2bc26172cb272c82830b8149b4247979647f658e2b12ca953c41af804d6

Observation 5366bd63-cc7f-4830-be9a-7fa8f7bb297f · inbound

Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion cites this paper.

Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:47.954582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T20:22:04.663359Z digest=sha256:0a18a5f2c0e0ebf1bfef72eebfe0bfbe677c9c970b629ba731ac7c407d464966

Observation 5fa18e0b-93cc-49c5-ac98-343690cc4eb2 · inbound

ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution cites this paper.

ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:56:14.121985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T15:58:59.744568Z digest=sha256:1d59d571fdb3a480c41b25febded51731bab20206efe7de9fee04b23802bd6c8

Observation c8d8e482-4ed2-4c66-a9c4-9054d4a82e92 · inbound

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making cites this paper.

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.053667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T17:01:31.011188Z digest=sha256:e7aa6b51713dbf166a72b12af040f889e03085a8d269bd2f7ba7c069aad6f5ef

Observation 2d7ede07-3691-4f8c-8848-21f70c4366bd · inbound

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding cites this paper.

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:26.590225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T00:53:35.188721Z digest=sha256:7576397e3716fd03435563e6ee40333b44e7d95900239bf9a2532fc0e7e28205

Observation ef25aee8-786a-4d1c-80fe-cf3bee87bdf5 · inbound

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers cites this paper.

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:49:44.370206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:47:32.476614Z digest=sha256:7d95e2c61289edb353397098f6a46f47f1492b7b5a57bad14778d1ac3e4246fd

Observation 21f924e8-64f1-450b-98a1-7a4e0013d1ea · inbound

Deformba: Vision State Space Model with Adaptive State Fusion cites this paper.

Deformba: Vision State Space Model with Adaptive State Fusion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:03:57.771091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:03:12.736238Z digest=sha256:9b9f6088512ba1f3cf43ceb499b9fe57839d3bb69ca6887193b8b76487f726c0

Observation 025a8355-0606-4108-856f-3dbb2e19596a · inbound

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation cites this paper.

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.452299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T11:11:46.428727Z digest=sha256:75f19d01ce4b92a543d29a1481531cfa4ce87d96407a8dd6f9f8997124b67084

Observation ac39401a-4851-47df-b270-f419b7319e87 · inbound

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation cites this paper.

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T12:34:16.033188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:34:16.033188Z digest=sha256:5b902233b84ee59e2c3b8a21f342b7aba1c8de02423b8d25a9d64faaae7a43a0

Observation 11892c46-f107-46b5-9eda-0a28e3b81e72 · inbound

Isolation-aware Scheduling Framework for DNN-based End-to-End Autonomous Driving System on Tile-based Accelerators cites this paper.

Isolation-aware Scheduling Framework for DNN-based End-to-End Autonomous Driving System on Tile-based Accelerators BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:47:44.612320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T11:49:46.697087Z digest=sha256:08580277c398c365aaa8ffd9650210986f40b1d21c2f20aef4844a19089a3033

Observation 6f9966dc-0deb-4187-9a50-bc27b622f84c · inbound

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models cites this paper.

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:22.585056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T07:08:41.071757Z digest=sha256:dffc4a395246371272478331637f8911e42aaa8b59e3ab6e9f094333e0c91ad3

Observation 140ef37a-73d5-423c-b6f5-1f951345f183 · inbound

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models cites this paper.

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-15T10:49:28.959330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:49:28.959330Z digest=sha256:10175eeba6c7774b89ac207f533743791d5534046d6ff678e563f00d2a6e77b6

Observation b4c6a4db-587d-4420-9252-6586c38d9d2d · inbound

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling cites this paper.

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.659347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:35:58.372801Z digest=sha256:0cb772956df383819e2cef2e9b9e33d241609914d525065df97a127b8f23b24a

Observation 01b31408-dafb-43ed-838e-739b059f9d83 · inbound

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling cites this paper.

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T09:58:51.929181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:58:51.929181Z digest=sha256:c95cdfa8bae2223cd3ecb65f15bb164313b8e06c74015a8f1559e0cc8a547481

Observation 2657712e-93b6-4f30-94bc-52cbdd4fbd41 · inbound

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction cites this paper.

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T23:42:57.473601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:42:57.473601Z digest=sha256:ccca8a7b61ee2ebfc471f8c5414f2650d502f6d8e68483ba48eab7fcf883e46b

Observation b06769f7-8c62-460c-94c6-aad4a59852fa · inbound

RayOcc: Occlusion-Aware Ray Occupancy Estimation via Gaussian Mixture Intensity cites this paper.

RayOcc: Occlusion-Aware Ray Occupancy Estimation via Gaussian Mixture Intensity BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T17:26:12.201838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:26:12.201838Z digest=sha256:7b7b6ee9daa6d9fe296fd54992dd2dc6e0b7e764d171d1bd744c43f52b67e34b