Pith. sign in

Paper Citation Record · LEDGER

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.10715.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.10715 v4

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:28:32.866993Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cf34b8cf-97eb-463f-9cd9-2f91a96750f6 · outbound

This paper cites Transfusion: Robust lidar-camera fusion for 3d object detection with transform- ers.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Transfusion: Robust lidar-camera fusion for 3d object detection with transform- ers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.356826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.869339Z digest=sha256:0b06ba619b0e9c7c381db53b860212ac8ac4a67cc5137467f6e186a28cb21454

Observation 831a497f-5d72-44a3-8a01-048fa561c219 · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection nuscenes: A multi- modal dataset for autonomous driving

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.902701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.902701Z digest=sha256:568f4eb0d1704fb194a9594cf6962f2144e6f3d5b11a8d4ea209107116284895

Observation 9cdc39d3-77ac-4cd9-837f-fe535a365661 · outbound

This paper cites BEVFusion4D: Learning LiDAR-Camera Fusion Under Bird's-Eye-View via Cross-Modality Guidance and Temporal Aggregation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVFusion4D: Learning LiDAR-Camera Fusion Under Bird's-Eye-View via Cross-Modality Guidance and Temporal Aggregation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.945210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.945210Z digest=sha256:234316ab7cbfbf0530c1870ec67886a4a05ab11a42f870d648e67ae38cc89617

Observation ba345c00-23e6-4c50-84dc-0c33205e9177 · outbound

This paper cites Objectfusion: Multi-modal 3d object detection with object-centric fusion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Objectfusion: Multi-modal 3d object detection with object-centric fusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.324462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.952456Z digest=sha256:40e5270392c9f3212f7dd425f566473e04b9657701df259293690170b967ff41

Observation df852ae9-8fd4-47a0-8677-3426972f01fd · outbound

This paper cites Futr3d: A unified sensor fusion framework for 3d detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Futr3d: A unified sensor fusion framework for 3d detection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.960051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.960051Z digest=sha256:0045b5d283de65184848f84659546ed34a7e0c1e74f8ba1229096dab9c7ad9b5

Observation d326804a-495d-447c-91f9-b19b9c4e3102 · outbound

This paper cites Focal- former3d: focusing on hard instance for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Focal- former3d: focusing on hard instance for 3d object detection

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.281120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.966404Z digest=sha256:cc16cf32c96882c402133e01d6d4e29dbc193f6a773de271459fed3b440ae7ac

Observation 1cc38717-a2b4-40fb-91d2-b16553222e95 · outbound

This paper cites Deformable feature aggregation for dynamic multi-modal 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deformable feature aggregation for dynamic multi-modal 3d object detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.973386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.973386Z digest=sha256:0aa7271c67f78709d64f688da24d1e68994dddd904399979e4a3550f8dba6df5

Observation 000040bf-6d9c-4652-a0e4-886a086bbae2 · outbound

This paper cites AutoAlign: Pixel-Instance Feature Aggregation for Multi-Modal 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection AutoAlign: Pixel-Instance Feature Aggregation for Multi-Modal 3D Object Detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.979698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.979698Z digest=sha256:b07e695289c4af0950f17038ad900093de53c4ff2c42f8c804a87210d8549d19

Observation 80d41623-02eb-4ff9-ab5e-091d93f53cba · outbound

This paper cites Li3detr: A li- dar based 3d detection transformer.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Li3detr: A li- dar based 3d detection transformer

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.247997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.986071Z digest=sha256:7e142bbc321841a7453cbfbb91b7598dfa85fe3000d79d241f0ffb9beb035b57

Observation 1b2aa19d-a697-4fc9-bee0-4efde8581b5a · outbound

This paper cites Adamixer: A fast-converging query-based object detector.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Adamixer: A fast-converging query-based object detector

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.221067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.992416Z digest=sha256:483b0b93df73d467ee64cefea7944543336b4523abd79aeeb6ad513dc8780ab7

Observation cc1897e8-9e8f-4d78-b1fc-36d58d4269e4 · outbound

This paper cites Deep residual learning for image recognition.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deep residual learning for image recognition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.196184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:31.997847Z digest=sha256:036e02aaab5b5d97b00f54b425cd9c3abe3061752c176f5d511834794fc82a54

Observation b5f09690-b9c6-4449-851b-f0a9da5c29b9 · outbound

This paper cites FusionFormer: A Multi-sensory Fusion in Bird's-Eye-View and Temporal Consistent Transformer for 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection FusionFormer: A Multi-sensory Fusion in Bird's-Eye-View and Temporal Consistent Transformer for 3D Object Detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.026565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.026565Z digest=sha256:23bbe8dfce1079276f087c2f23af52e9108d759b08598a0d8427fa35cdd6067b

Observation c57c0f88-cab8-4ab1-bac7-078e928635af · outbound

This paper cites EA-LSS: Edge-aware Lift-splat-shot Framework for 3D BEV Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection EA-LSS: Edge-aware Lift-splat-shot Framework for 3D BEV Object Detection

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.073565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.073565Z digest=sha256:46a6d6cdb00933fd4a78f34084c4b2c982b08d9ea71409266f1adae325d0f422

Observation 8624b5a5-52ff-43fc-b83e-7b8d73ff62ab · outbound

This paper cites BEVDet4D: Exploit Temporal Cues in Multi-camera 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVDet4D: Exploit Temporal Cues in Multi-camera 3D Object Detection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.100561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.100561Z digest=sha256:6c2a2e2d619b68e72a6acb9a97b00f63aa65d3d62d1de0f89fd291426c2d3acc

Observation 529056fc-d4a2-4640-9363-78d345cb6783 · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.110177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.110177Z digest=sha256:222690e1a2122753270c4e239dd5b404eb0bf1a603c34898e5ab08be1d43d4a8

Observation 90c89b00-7150-4bd5-9216-75d36552350c · outbound

This paper cites Far3d: Expanding the horizon for surround-view 3d object detec- tion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Far3d: Expanding the horizon for surround-view 3d object detec- tion

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.177090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.117315Z digest=sha256:337e6b9d306bf475db1793bf9e2311fe28447d495fae9752859dd5e045686c30

Observation 41f62e9f-b0e6-4a1a-8721-3ba7d64d5c6b · outbound

This paper cites Msmdfusion: Fusing lidar and camera at multiple scales with multi-depth seeds for 3d ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Msmdfusion: Fusing lidar and camera at multiple scales with multi-depth seeds for 3d ob- ject detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.152651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.123896Z digest=sha256:3f822ca34fa25cc3e404b9711f1aaaecf7147380f38775479b1f3fefb7dcbb63

Observation dd3d40dd-fbd1-4db6-bd2a-a50017551b09 · outbound

This paper cites Centermask: Real-time anchor-free instance segmentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Centermask: Real-time anchor-free instance segmentation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.129462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.132234Z digest=sha256:9b13ff7f119240aabe2b4712f271846c20966ad8338e6844cb4d846eb904381a

Observation d540e775-3042-4f19-8dd5-9d4307478b73 · outbound

This paper cites Dn-detr: Accelerate detr training by intro- ducing query denoising.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Dn-detr: Accelerate detr training by intro- ducing query denoising

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.107666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.139659Z digest=sha256:f9f680110656dbe79ed6762d0eb428cdf70a41f0833035d059f7ff9a5dd0aa6c

Observation eda1fe5b-6c09-458c-8853-54738a770d88 · outbound

This paper cites Gafusion: Adaptive fusing lidar and camera with multi- ple guidance for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Gafusion: Adaptive fusing lidar and camera with multi- ple guidance for 3d object detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.087414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.147120Z digest=sha256:c5f8984d56ccbb0895695164b262569cb567e399be3b71a585384a185e5aaede

Observation 633861bd-206c-4d43-9e82-cab98f7642cc · outbound

This paper cites Unifying voxel-based representation with transformer for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unifying voxel-based representation with transformer for 3d object detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.062138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.187755Z digest=sha256:133dc01fd9f2b027fb5f2a774bf90f03051a9cbbeec7db3dba14e2b99a74c9df

Observation d26534f8-9184-474a-a38b-09bb6f801027 · outbound

This paper cites Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.040921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.252018Z digest=sha256:d33e25d13ecbf6cd3c6d2b08476790030a9f405c4f61d756b2f4883ede614d78

Observation 195c2a61-2a56-487a-a9d3-efa5a5848897 · outbound

This paper cites Fast-bev: A fast and strong bird’s- eye view perception baseline.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Fast-bev: A fast and strong bird’s- eye view perception baseline

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.020290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.258150Z digest=sha256:e962eae1210132e656eb81346a3aef8745ead056af51346829ae8b0fe65eb36c

Observation 9ef83d42-a8b3-4692-ac9a-4dd59905c3be · outbound

This paper cites Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.264382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.264382Z digest=sha256:657b73384ab08ea185ad8358824e60bffbb74f6e20b82217929ef86cd513a70b

Observation e6f6b9fe-c088-4b14-879b-a74451b72c3d · outbound

This paper cites Bevfusion: A simple and robust lidar-camera fusion framework.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevfusion: A simple and robust lidar-camera fusion framework

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.972462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.270371Z digest=sha256:e5a24b58ec686898ceef72df9fe22819bf692acad974913751e4a5f28147f3c6

Observation 923fb104-c073-41ce-a7c0-94a155be2e4b · outbound

This paper cites Feature pyra- mid networks for object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Feature pyra- mid networks for object detection

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.276545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.276545Z digest=sha256:f4e2ba48d917c7889707dab26eab7bb69f1c58872c3901cefc3fe7449ecf038e

Observation 6331c775-a5b6-4b73-a6d9-ff63cd19df36 · outbound

This paper cites Sparsebev: High-performance sparse 3d object de- tection from multi-camera videos.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparsebev: High-performance sparse 3d object de- tection from multi-camera videos

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.933417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.288586Z digest=sha256:86b466eade95f09800896cff20bc51684cee2e5242ca9ca1f62335116f0b045b

Observation d4560ec2-925c-499d-9383-79c3165e1b88 · outbound

This paper cites DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.349200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.349200Z digest=sha256:1236774fde717950b215f32e98df72ce83a01a18ce4ea6975323c58c69bb1e09

Observation e4e42334-4b28-4ebf-a877-f52f50e8dead · outbound

This paper cites Petr: Position embedding transformation for multi-view 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Petr: Position embedding transformation for multi-view 3d object detection

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.916200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.389505Z digest=sha256:74bdf485fbcfcc2e5dbb6f07bf27c9ece9c02530602d94939155d730ee3fbfa0

Observation d9c64888-8943-46c3-b6b3-26b7948b1dd2 · outbound

This paper cites Petrv2: A unified framework for 3d perception from multi-camera images.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Petrv2: A unified framework for 3d perception from multi-camera images

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.898541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.396866Z digest=sha256:28d3242ccf15a84d53c6df7da4f6ecd954de523d7200fe09865dadd443b9d5f0

Observation 9a4c979a-4e49-4e10-82c4-c4b8b8e122b3 · outbound

This paper cites Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.880251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.403024Z digest=sha256:fb869465c1d8b3ab6fc3c16c7a7adf08bc4596c7c7b7aa3b90f98cd0f6335325

Observation 3df95c01-9555-4966-a908-6f2004176533 · outbound

This paper cites Decoupled weight decay regularization.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Decoupled weight decay regularization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.409589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.409589Z digest=sha256:fad486b0a93d6ead9b1c3da1c675afa8e765dd8054ad15d912157b208409d892

Observation 6e2743bd-466a-4abd-ade4-bc393fa33d7b · outbound

This paper cites DETR4D: Direct Multi-View 3D Object Detection with Sparse Attention.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DETR4D: Direct Multi-View 3D Object Detection with Sparse Attention

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.417165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.417165Z digest=sha256:88674b5bb4b9e339e90d5e7fc6e8dac94ee483863dea9a7c2d324023f48e748e

Observation a6aa273d-edaa-4cac-a3e3-dc0251ef65e8 · outbound

This paper cites Lift, splat, shoot: Encod- ing images from arbitrary camera rigs by implicitly unpro- jecting to 3d.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Lift, splat, shoot: Encod- ing images from arbitrary camera rigs by implicitly unpro- jecting to 3d

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.473306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.473306Z digest=sha256:695145bcdb99d219cfa35fa3128e45c6eb2911a31349cd5a81038c5250ac7000

Observation 713536b3-5403-487e-b5ab-765cdf112c45 · outbound

This paper cites Categorical depth distribution network for monocular 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Categorical depth distribution network for monocular 3d object detection

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.526518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.526518Z digest=sha256:9a523b3d9b4ea54b0065a6e017818e5d506e56ee8ed3266dac7f1004e94d04c7

Observation 39402c44-1e5e-4d69-aa7a-e5655a14d85f · outbound

This paper cites Focal loss for dense ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Focal loss for dense ob- ject detection

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.533375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.533375Z digest=sha256:fbb660dd4d7920726c1a7beb057a93d464d5140365dbc536657067ba22b71af8

Observation 3c7976a6-4263-4681-9bca-286c31ecf2d5 · outbound

This paper cites Cyclical learning rates for training neural networks.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Cyclical learning rates for training neural networks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.540886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.540886Z digest=sha256:5dc36ddfec51526048e9bb1a80537507f34709908ab1e947234479fa7852bfda

Observation 01a03ecf-58fd-43a4-91ce-6c44abcfb3fb · outbound

This paper cites Attention is all you need.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Attention is all you need

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.782414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.547697Z digest=sha256:0fcd6bc3bf559e1edcac6e681c989aefab6052ea22cd557d4ef860e8f7b65779

Observation ca45c5d4-86e2-4fc5-ba7a-b8b2a9ef9295 · outbound

This paper cites Pointpainting: Sequential fusion for 3d object de- tection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Pointpainting: Sequential fusion for 3d object de- tection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.553384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.553384Z digest=sha256:f9db4aa27f235f9485f28b41706f03ae2655b6a93b7e9c2dbd4317fb6de794bd

Observation 38f4380e-b09d-4b7c-949b-cd9e077f00db · outbound

This paper cites Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.738673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.559169Z digest=sha256:18d83b2e2560e93445bd77335f98bc3cfc4f9c00814b9b7dcfeb39e53b4fafae

Observation 9433cc5e-d441-4f00-a1bc-487f80013518 · outbound

This paper cites an unresolved cited work.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:28:33.716692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.564250Z digest=sha256:e0187b1996dffc1dfe1a4c4b31b677cb801ebc6e5f65318128328efd431dacd1

Observation 31193498-e793-4c76-b2cc-cdb91510b183 · outbound

This paper cites Exploring object-centric temporal modeling for efficient multi-view 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Exploring object-centric temporal modeling for efficient multi-view 3d object detection

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.694585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.570535Z digest=sha256:ba2d76e8af5dcd448a753ae9c11af954c1d62a6f8b7809cb97486290e451b40a

Observation 5da9b462-c8e6-47cd-a085-3d26e0abc278 · outbound

This paper cites Detr3d: 3d object detection from multi-view images via 3d-to-2d queries.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Detr3d: 3d object detection from multi-view images via 3d-to-2d queries

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.675157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.576703Z digest=sha256:591636529a4c8874cee02e7949ae0da4cda42948eb4467f1c7a79ce80255c651

Observation dbd26f5a-ede3-45d9-9a24-5a8c516ac9d6 · outbound

This paper cites MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.581814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.581814Z digest=sha256:8a4ae7dd74878e4d75d34b068205577a9459559dc8fe05d5a8fc235a32aa3b6f

Observation 34b09d61-d230-4752-9ab5-acd1aceee89c · outbound

This paper cites Sparsefusion: Fusing multi-modal sparse rep- resentations for multi-sensor 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparsefusion: Fusing multi-modal sparse rep- resentations for multi-sensor 3d object detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.651370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.636500Z digest=sha256:a3815ebd82a68252854901758d909adb634a5d78560e49107135a749178ff545

Observation 545362d3-1372-4ecd-94ec-8cea2e99e7ef · outbound

This paper cites Cross modal trans- former: Towards fast and robust 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Cross modal trans- former: Towards fast and robust 3d object detection

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.624715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.681023Z digest=sha256:264b5c2ca1f90deed8c0aa0389c9499a74b336cd1ef33e521b05285f1eee2e42

Observation e0ed78f5-633e-4b80-9302-1e313eb20f62 · outbound

This paper cites Second: Sparsely embed- ded convolutional detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Second: Sparsely embed- ded convolutional detection

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.601168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.686676Z digest=sha256:568de36090880023436befe8a3ee50edfc5b4e05f7e71f86132eafe263740c78

Observation acfb5e71-4fb1-4d32-800e-5829fb21f471 · outbound

This paper cites Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective su- pervision.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective su- pervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.694155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.694155Z digest=sha256:f8323f1e8faea992a7886c465d24ff2b38695d930f31f3e371d32a718eb9b12d

Observation dfef9795-ebcb-4d7d-8300-72ca824feeb4 · outbound

This paper cites Deepinteraction: 3d object detection via modality interaction.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deepinteraction: 3d object detection via modality interaction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.555150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.700956Z digest=sha256:465a34b4e162ca815f131698526fb4e70f9fd1d72d6b747a7c5e18620398c9c8

Observation 88b54127-fd66-471d-9912-1d16bbf095de · outbound

This paper cites DeepInteraction++: Multi-Modality Interaction for Autonomous Driving.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DeepInteraction++: Multi-Modality Interaction for Autonomous Driving

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.705803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.705803Z digest=sha256:f6669023703ac64e2d192446328680c99c3ce238ad476bad6cf31b828807b8be

Observation 2b76b5ee-3b94-4e1e-a582-7b4dd311ea82 · outbound

This paper cites Is-fusion: Instance-scene collaborative fusion for multimodal 3d ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Is-fusion: Instance-scene collaborative fusion for multimodal 3d ob- ject detection

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.530527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.714103Z digest=sha256:4e5bbab93cf92f107fccecc752177fd1173097699b6a49be18ee39d33c85b38e

Observation 6781b3c6-4158-41da-a7cd-142cc06a06fb · outbound

This paper cites Center- based 3d object detection and tracking.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Center- based 3d object detection and tracking

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.509180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.721909Z digest=sha256:ed348ff104a2ed95fd8d367f0238cd887dcb832d0a26e6f7b52f55423c5a29c7

Observation 58607b24-2237-4a45-a873-7aacf3030b23 · outbound

This paper cites Multi- modal virtual point 3d detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Multi- modal virtual point 3d detection

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.482646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.783175Z digest=sha256:350491484302f77b5a6dc23e8acb9b2a18402aebfddd55a5e0a5549015291df6

Observation 431aa5a5-6523-46d0-869d-d4cf80606bb9 · outbound

This paper cites DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.828663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.828663Z digest=sha256:36506e98d01a80ec16642a4de325cb2e435c88efa1a76f4406e2a961cb5a07da

Observation 1b0c9fab-649f-4a49-96ac-7b9ebb21924b · outbound

This paper cites Sparselif: High-performance sparse lidar- camera fusion for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparselif: High-performance sparse lidar- camera fusion for 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.459457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.833803Z digest=sha256:e086a0e1acf4bc25ce5d54e1a53781fb3da4a771380680e08c15d4ec6e6bd43a

Observation d4b743e8-a98a-4ffe-90d4-43c4a94fd101 · outbound

This paper cites SimpleBEV: Improved LiDAR-Camera Fusion Architecture for 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection SimpleBEV: Improved LiDAR-Camera Fusion Architecture for 3D Object Detection

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.841383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.841383Z digest=sha256:65e11b4e6f51cea49b0f5a6768631541ac0693a9b3f8591f30402cc1e8d883ff

Observation 6e9daa43-c357-4cbc-a104-ca6d2336b3f7 · outbound

This paper cites V oxelnet: End-to-end learning for point cloud based 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection V oxelnet: End-to-end learning for point cloud based 3d object detection

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.333466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.847930Z digest=sha256:10e2a450268dbdea19f29f9a7a81452b15a53e850691a610aaad91061ed57423

Observation f77c4e70-c88f-41a6-8587-eb0658698e35 · outbound

This paper cites Centerformer: Center-based transformer for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Centerformer: Center-based transformer for 3d object detection

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.275324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:28:32.854555Z digest=sha256:af7c56a766d3e462b4174c4ba8611f16e45494f89a7c79c7694a1987d02e3d2f

Observation 44122285-e031-4b7f-9445-9c6d9fc1437d · outbound

This paper cites Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.860314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.860314Z digest=sha256:9704b611c7c74f466726ebb4c7467ddddd3fbc73603b97e48cf1da122e570d54

Observation c2c59f5a-8caa-48e3-8951-fb90bccff40d · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.866993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.866993Z digest=sha256:1897514fc97daa473ff5002b50a3d02829a0a542a293e5bb084ccd8a53961a4e

Pith citing papers

No inbound Pith citation observations are available.