Pith. sign in

Paper Citation Record · LEDGER

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection

As of 16 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.10715.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.10715 v4

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:28:32.866993Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cf34b8cf-97eb-463f-9cd9-2f91a96750f6 · outbound

This paper cites Transfusion: Robust lidar-camera fusion for 3d object detection with transform- ers.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Transfusion: Robust lidar-camera fusion for 3d object detection with transform- ers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.356826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.869339Z digest=sha256:a8692a6997f7d1aceabbfc15d5c6fde6b8b744ec194b1a351d6e595295259b5e

Observation 831a497f-5d72-44a3-8a01-048fa561c219 · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection nuscenes: A multi- modal dataset for autonomous driving

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.902701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.902701Z digest=sha256:45e2ba5a6d26898bf4a6274476cd73c112e68327509d12953cb574453632faa6

Observation 9cdc39d3-77ac-4cd9-837f-fe535a365661 · outbound

This paper cites BEVFusion4D: Learning LiDAR-Camera Fusion Under Bird's-Eye-View via Cross-Modality Guidance and Temporal Aggregation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVFusion4D: Learning LiDAR-Camera Fusion Under Bird's-Eye-View via Cross-Modality Guidance and Temporal Aggregation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.945210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.945210Z digest=sha256:3a7d4df8e0c3c4c486b40f5ee47e16b9647a259ca9490092bd8e05a214cca044

Observation ba345c00-23e6-4c50-84dc-0c33205e9177 · outbound

This paper cites Objectfusion: Multi-modal 3d object detection with object-centric fusion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Objectfusion: Multi-modal 3d object detection with object-centric fusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.324462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.952456Z digest=sha256:9be34f6b015eb4e1afd8f114f769c7b6c0709146f392d8a262593951311621fd

Observation df852ae9-8fd4-47a0-8677-3426972f01fd · outbound

This paper cites Futr3d: A unified sensor fusion framework for 3d detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Futr3d: A unified sensor fusion framework for 3d detection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.960051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.960051Z digest=sha256:8918192897170bd85115259230c92457a2fdc6674e74c1041348ccc3c89925b6

Observation d326804a-495d-447c-91f9-b19b9c4e3102 · outbound

This paper cites Focal- former3d: focusing on hard instance for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Focal- former3d: focusing on hard instance for 3d object detection

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.281120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.966404Z digest=sha256:55500175ed602028f486b12b542b7087866637209f7653ed979d2815fce59d48

Observation 1cc38717-a2b4-40fb-91d2-b16553222e95 · outbound

This paper cites Deformable feature aggregation for dynamic multi-modal 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deformable feature aggregation for dynamic multi-modal 3d object detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.973386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.973386Z digest=sha256:1c6209307395a9b4bf94def341503bdea48f80bdacf21f5fb417f4b6d0f3596f

Observation 000040bf-6d9c-4652-a0e4-886a086bbae2 · outbound

This paper cites AutoAlign: Pixel-Instance Feature Aggregation for Multi-Modal 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection AutoAlign: Pixel-Instance Feature Aggregation for Multi-Modal 3D Object Detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:31.979698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:31.979698Z digest=sha256:ba3b533175853864885d714b7de7f42884f8dc704c7c2df4d49500f36691a4b4

Observation 80d41623-02eb-4ff9-ab5e-091d93f53cba · outbound

This paper cites Li3detr: A li- dar based 3d detection transformer.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Li3detr: A li- dar based 3d detection transformer

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.247997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.986071Z digest=sha256:cffccb9366c5ce8a1a6da33faed5c6d77fdf59e94183bf4265fbbfef8e3a1735

Observation 1b2aa19d-a697-4fc9-bee0-4efde8581b5a · outbound

This paper cites Adamixer: A fast-converging query-based object detector.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Adamixer: A fast-converging query-based object detector

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.221067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.992416Z digest=sha256:db1074e36d85e4db038c593a0a9038de75ce85de8ee207d8197c16a44088d6ff

Observation cc1897e8-9e8f-4d78-b1fc-36d58d4269e4 · outbound

This paper cites Deep residual learning for image recognition.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deep residual learning for image recognition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.196184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:31.997847Z digest=sha256:7beba834d13b43dec97d6b2fbbae98895fa46f2d39da7b55d6ac10a348fb6036

Observation b5f09690-b9c6-4449-851b-f0a9da5c29b9 · outbound

This paper cites FusionFormer: A Multi-sensory Fusion in Bird's-Eye-View and Temporal Consistent Transformer for 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection FusionFormer: A Multi-sensory Fusion in Bird's-Eye-View and Temporal Consistent Transformer for 3D Object Detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.026565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.026565Z digest=sha256:f78ff85ef25b213ffda4e36672ad5f8df9d8a49e5081dcee70139c10c1b14803

Observation c57c0f88-cab8-4ab1-bac7-078e928635af · outbound

This paper cites EA-LSS: Edge-aware Lift-splat-shot Framework for 3D BEV Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection EA-LSS: Edge-aware Lift-splat-shot Framework for 3D BEV Object Detection

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.073565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.073565Z digest=sha256:a8a3da94b0ecb1423458c875f1b6221d333b8dda8aa0d9b5cbf235857fa69c0d

Observation 8624b5a5-52ff-43fc-b83e-7b8d73ff62ab · outbound

This paper cites BEVDet4D: Exploit Temporal Cues in Multi-camera 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVDet4D: Exploit Temporal Cues in Multi-camera 3D Object Detection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.100561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.100561Z digest=sha256:70661531ede38fcda249fedfc8efb674537866529e6f31efc547fa73c708fb9b

Observation 529056fc-d4a2-4640-9363-78d345cb6783 · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.110177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.110177Z digest=sha256:733da68c17219afbbacd1aa47cd542b4361ad6aa3a0f81e7393c52a754ff037a

Observation 90c89b00-7150-4bd5-9216-75d36552350c · outbound

This paper cites Far3d: Expanding the horizon for surround-view 3d object detec- tion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Far3d: Expanding the horizon for surround-view 3d object detec- tion

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.177090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.117315Z digest=sha256:59d64264acc0d975b3e394b1de5a9ff9fdb0ddaaf4df910474b1bf3383fed119

Observation 41f62e9f-b0e6-4a1a-8721-3ba7d64d5c6b · outbound

This paper cites Msmdfusion: Fusing lidar and camera at multiple scales with multi-depth seeds for 3d ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Msmdfusion: Fusing lidar and camera at multiple scales with multi-depth seeds for 3d ob- ject detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.152651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.123896Z digest=sha256:ae4446aa617c5d519b9c8ae4ff401bb8917e609d29b38992131667a4f5d48cda

Observation dd3d40dd-fbd1-4db6-bd2a-a50017551b09 · outbound

This paper cites Centermask: Real-time anchor-free instance segmentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Centermask: Real-time anchor-free instance segmentation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.129462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.132234Z digest=sha256:af2758d92acf2d8d2cf8586cab6391ed2100c9e5414ee86ff36bca45615129e3

Observation d540e775-3042-4f19-8dd5-9d4307478b73 · outbound

This paper cites Dn-detr: Accelerate detr training by intro- ducing query denoising.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Dn-detr: Accelerate detr training by intro- ducing query denoising

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.107666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.139659Z digest=sha256:8c502a2f3269082b11cf7f771f868b70959af0cc2cc5163924d9e3035d8f8768

Observation eda1fe5b-6c09-458c-8853-54738a770d88 · outbound

This paper cites Gafusion: Adaptive fusing lidar and camera with multi- ple guidance for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Gafusion: Adaptive fusing lidar and camera with multi- ple guidance for 3d object detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.087414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.147120Z digest=sha256:43c0549dcb796e8a690a703e15f085e924d1217588009ecae48c77904bf2ef0e

Observation 633861bd-206c-4d43-9e82-cab98f7642cc · outbound

This paper cites Unifying voxel-based representation with transformer for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unifying voxel-based representation with transformer for 3d object detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.062138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.187755Z digest=sha256:3f8aa249800f89bf0d8acd3f7df18c26e2741143589efb0806eab70051ac0938

Observation d26534f8-9184-474a-a38b-09bb6f801027 · outbound

This paper cites Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.040921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.252018Z digest=sha256:05a2a63fe4e9ad759e23b1c6036336aeb76969ecbb1439ad65b483ba5c88de0d

Observation 195c2a61-2a56-487a-a9d3-efa5a5848897 · outbound

This paper cites Fast-bev: A fast and strong bird’s- eye view perception baseline.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Fast-bev: A fast and strong bird’s- eye view perception baseline

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:34.020290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.258150Z digest=sha256:8a68b67ede32f41c1f7e0c267377f18bec8b5146f30a445f0add152117600ea6

Observation 9ef83d42-a8b3-4692-ac9a-4dd59905c3be · outbound

This paper cites Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.264382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.264382Z digest=sha256:af8bf40ea66f775822e556fbf77d63c260640782fe1fa3615f4b308a4b15cc2a

Observation e6f6b9fe-c088-4b14-879b-a74451b72c3d · outbound

This paper cites Bevfusion: A simple and robust lidar-camera fusion framework.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevfusion: A simple and robust lidar-camera fusion framework

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.972462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.270371Z digest=sha256:d558e6368ad3b9c5948155da089bcc95537580ea8e07fc7e87e418c8a24802b6

Observation 923fb104-c073-41ce-a7c0-94a155be2e4b · outbound

This paper cites Feature pyra- mid networks for object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Feature pyra- mid networks for object detection

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.276545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.276545Z digest=sha256:ed2d681b3ec307a88c698325ea4321013beb9827db975e351833089d7ac6e06b

Observation 6331c775-a5b6-4b73-a6d9-ff63cd19df36 · outbound

This paper cites Sparsebev: High-performance sparse 3d object de- tection from multi-camera videos.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparsebev: High-performance sparse 3d object de- tection from multi-camera videos

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.933417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.288586Z digest=sha256:2b5b509289b26fa475fec3c85fd321eeab68541ce8c474e132df729e8eab9f6b

Observation d4560ec2-925c-499d-9383-79c3165e1b88 · outbound

This paper cites DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.349200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.349200Z digest=sha256:991749bf82353671902905b2a7f8b7e12c9b5cb75164137493688ce7700f954d

Observation e4e42334-4b28-4ebf-a877-f52f50e8dead · outbound

This paper cites Petr: Position embedding transformation for multi-view 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Petr: Position embedding transformation for multi-view 3d object detection

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.916200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.389505Z digest=sha256:b05e3f6f47ab7cf5bd925945b67a205a4a65bada85e5855d7437fea2da892204

Observation d9c64888-8943-46c3-b6b3-26b7948b1dd2 · outbound

This paper cites Petrv2: A unified framework for 3d perception from multi-camera images.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Petrv2: A unified framework for 3d perception from multi-camera images

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.898541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.396866Z digest=sha256:f45bb940b7b1d08fe4e2294fce093774fcebdb89ce23c9d4325c6d8b44b48268

Observation 9a4c979a-4e49-4e10-82c4-c4b8b8e122b3 · outbound

This paper cites Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevfusion: Multi- task multi-sensor fusion with unified bird’s-eye view repre- sentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.880251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.403024Z digest=sha256:98f2433999e25430e0469c400d660c5c0dbb251728f759d30ec9a7d899e1811e

Observation 3df95c01-9555-4966-a908-6f2004176533 · outbound

This paper cites Decoupled weight decay regularization.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Decoupled weight decay regularization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.409589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.409589Z digest=sha256:84ffb62a0f881caded7f5b2d9838706ea95862380c4b5decbbbb422d16e88857

Observation 6e2743bd-466a-4abd-ade4-bc393fa33d7b · outbound

This paper cites DETR4D: Direct Multi-View 3D Object Detection with Sparse Attention.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DETR4D: Direct Multi-View 3D Object Detection with Sparse Attention

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.417165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.417165Z digest=sha256:a05e16b65f36bdead66ff411fe99deb7b6c79422a0a51797084a040c5d4974b9

Observation a6aa273d-edaa-4cac-a3e3-dc0251ef65e8 · outbound

This paper cites Lift, splat, shoot: Encod- ing images from arbitrary camera rigs by implicitly unpro- jecting to 3d.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Lift, splat, shoot: Encod- ing images from arbitrary camera rigs by implicitly unpro- jecting to 3d

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.473306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.473306Z digest=sha256:7f9f996aa13731c887a0cb14f9f326d269be6681e86d4d31765648a8d05a8b2a

Observation 713536b3-5403-487e-b5ab-765cdf112c45 · outbound

This paper cites Categorical depth distribution network for monocular 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Categorical depth distribution network for monocular 3d object detection

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.526518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.526518Z digest=sha256:e2679bdce4bb7f1c6c1f8775821e04d32faed62a44095b772e4e54f015697ac9

Observation 39402c44-1e5e-4d69-aa7a-e5655a14d85f · outbound

This paper cites Focal loss for dense ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Focal loss for dense ob- ject detection

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.533375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.533375Z digest=sha256:2c9d8a22d4163a07eb968c40432051410b2abccf49ec7c63a06ee644507d6a73

Observation 3c7976a6-4263-4681-9bca-286c31ecf2d5 · outbound

This paper cites Cyclical learning rates for training neural networks.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Cyclical learning rates for training neural networks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.540886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.540886Z digest=sha256:5896f08018e311de63bb04b1f4e914d61c1206a51aa289c0ee42126252ae7c74

Observation 01a03ecf-58fd-43a4-91ce-6c44abcfb3fb · outbound

This paper cites Attention is all you need.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Attention is all you need

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.782414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.547697Z digest=sha256:163daf0bf33aa10d08bcb9c60902f9ea140a4ec68dcc51662d02669555c607d6

Observation ca45c5d4-86e2-4fc5-ba7a-b8b2a9ef9295 · outbound

This paper cites Pointpainting: Sequential fusion for 3d object de- tection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Pointpainting: Sequential fusion for 3d object de- tection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.553384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.553384Z digest=sha256:eb0a53f98c67d40e4eadb9199bb371ba979555ecc0ca1db03e39a38660374d0c

Observation 38f4380e-b09d-4b7c-949b-cd9e077f00db · outbound

This paper cites Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unitr: A unified and efficient multi-modal transformer for bird’s-eye-view repre- sentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.738673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.559169Z digest=sha256:1f7e98786581fee53495c2cad150ecdf90d43f7b5a454aeff2da66e14c685747

Observation 9433cc5e-d441-4f00-a1bc-487f80013518 · outbound

This paper cites an unresolved cited work.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:28:33.716692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.564250Z digest=sha256:3019acbed6150c6a058dd02fb8604682bea5c51673fe48c1d2f8172e140cd578

Observation 31193498-e793-4c76-b2cc-cdb91510b183 · outbound

This paper cites Exploring object-centric temporal modeling for efficient multi-view 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Exploring object-centric temporal modeling for efficient multi-view 3d object detection

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.694585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.570535Z digest=sha256:5c33e8e4dba09cf020214be6196dcaacf11f4ec4833c87ec00f4dca416c67f9a

Observation 5da9b462-c8e6-47cd-a085-3d26e0abc278 · outbound

This paper cites Detr3d: 3d object detection from multi-view images via 3d-to-2d queries.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Detr3d: 3d object detection from multi-view images via 3d-to-2d queries

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.675157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.576703Z digest=sha256:b05fef66cacf41e4d66dffc5a8bce590682d3fa329987421f7d88af676da52cf

Observation dbd26f5a-ede3-45d9-9a24-5a8c516ac9d6 · outbound

This paper cites MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.581814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.581814Z digest=sha256:9b60cf9fe1f54f0c033a5444f87e4172088f87f36d8ddd2fe42760b9edcde364

Observation 34b09d61-d230-4752-9ab5-acd1aceee89c · outbound

This paper cites Sparsefusion: Fusing multi-modal sparse rep- resentations for multi-sensor 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparsefusion: Fusing multi-modal sparse rep- resentations for multi-sensor 3d object detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.651370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.636500Z digest=sha256:3c78026f25854270a121e2878ed18703cac963ecc9fab0f5de8bef9b5dd5320c

Observation 545362d3-1372-4ecd-94ec-8cea2e99e7ef · outbound

This paper cites Cross modal trans- former: Towards fast and robust 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Cross modal trans- former: Towards fast and robust 3d object detection

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.624715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.681023Z digest=sha256:36a2c414d3eaf79c240376de2b4285598ab44f0ee5f140a244636864bcc39567

Observation e0ed78f5-633e-4b80-9302-1e313eb20f62 · outbound

This paper cites Second: Sparsely embed- ded convolutional detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Second: Sparsely embed- ded convolutional detection

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.601168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.686676Z digest=sha256:dd0d566c70843672e55aa16026a9f5b3d09a1f88d829da0d4ef3782d58f5a5a2

Observation acfb5e71-4fb1-4d32-800e-5829fb21f471 · outbound

This paper cites Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective su- pervision.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective su- pervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.694155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.694155Z digest=sha256:1a80260bf166fbebc286a5e42f26112a665a6f42adef5470c2f681ad97f1189c

Observation dfef9795-ebcb-4d7d-8300-72ca824feeb4 · outbound

This paper cites Deepinteraction: 3d object detection via modality interaction.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deepinteraction: 3d object detection via modality interaction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.555150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.700956Z digest=sha256:a133ba79b432ae848fe89160d04c8ccc9dd92efe6bf97d33c0623bc163d15919

Observation 88b54127-fd66-471d-9912-1d16bbf095de · outbound

This paper cites DeepInteraction++: Multi-Modality Interaction for Autonomous Driving.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DeepInteraction++: Multi-Modality Interaction for Autonomous Driving

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.705803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.705803Z digest=sha256:c2eef7781acbe4ed566d2514d7839a442e12724d93685bceacf272b8d7904ac2

Observation 2b76b5ee-3b94-4e1e-a582-7b4dd311ea82 · outbound

This paper cites Is-fusion: Instance-scene collaborative fusion for multimodal 3d ob- ject detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Is-fusion: Instance-scene collaborative fusion for multimodal 3d ob- ject detection

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.530527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.714103Z digest=sha256:04006872951d7b636b64d458f2f25451e929ee61253f7b5b5cbd9fd256bc7084

Observation 6781b3c6-4158-41da-a7cd-142cc06a06fb · outbound

This paper cites Center- based 3d object detection and tracking.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Center- based 3d object detection and tracking

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.509180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.721909Z digest=sha256:e5119e3cd8b14e3ef49f18d7e9e1b85bb1bb8716224967c3b2a6a2f2358c80f9

Observation 58607b24-2237-4a45-a873-7aacf3030b23 · outbound

This paper cites Multi- modal virtual point 3d detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Multi- modal virtual point 3d detection

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.482646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.783175Z digest=sha256:cd06a410bafa6cb898c1a30f53bd2da12798a2788225f0cf495be7b213739832

Observation 431aa5a5-6523-46d0-869d-d4cf80606bb9 · outbound

This paper cites DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.828663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.828663Z digest=sha256:b53b4fc80633a6d2a8c757546cd37d56fed6b5fc047e1ebb9427f4170fe5727b

Observation 1b0c9fab-649f-4a49-96ac-7b9ebb21924b · outbound

This paper cites Sparselif: High-performance sparse lidar- camera fusion for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Sparselif: High-performance sparse lidar- camera fusion for 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.459457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.833803Z digest=sha256:6665dd5dce639c4ef97cca867167f1791fe9064a4e2de1d8ffad94f6b053d6ab

Observation d4b743e8-a98a-4ffe-90d4-43c4a94fd101 · outbound

This paper cites SimpleBEV: Improved LiDAR-Camera Fusion Architecture for 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection SimpleBEV: Improved LiDAR-Camera Fusion Architecture for 3D Object Detection

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.841383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.841383Z digest=sha256:3b52f6a5dbd8d76a09c02d2c42fbe17542a54ad25b24c9c539692faf24e42613

Observation 6e9daa43-c357-4cbc-a104-ca6d2336b3f7 · outbound

This paper cites V oxelnet: End-to-end learning for point cloud based 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection V oxelnet: End-to-end learning for point cloud based 3d object detection

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.333466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.847930Z digest=sha256:d729416e8f0838491c59f407d5f1707ce7b127e9d56da63309c64b75182c7b67

Observation f77c4e70-c88f-41a6-8587-eb0658698e35 · outbound

This paper cites Centerformer: Center-based transformer for 3d object detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Centerformer: Center-based transformer for 3d object detection

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:28:33.275324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T19:28:32.854555Z digest=sha256:b2e468a336e68e7eabe7739307f8605864277e43689a55d1ec9e9d25446b158b

Observation 44122285-e031-4b7f-9445-9c6d9fc1437d · outbound

This paper cites Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.860314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.860314Z digest=sha256:a851dce3302019584afafddbadce085b909401c8805662e78910dbbbf419eeec

Observation c2c59f5a-8caa-48e3-8951-fb90bccff40d · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:32.866993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:28:32.866993Z digest=sha256:a723090f23995a2b82f6ff4826606656afcc4c8de271764087aa188d742dfe77

Pith citing papers

No inbound Pith citation observations are available.