Pith. sign in

Paper Citation Record · LEDGER

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion

As of 19 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2506.15610.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15610 v3

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:57:23.112112Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved16
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1810e1c3-0f9a-48e4-825b-965a552770ad · outbound

This paper cites Omni3d: A large benchmark and model for 3d object detection in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Omni3d: A large benchmark and model for 3d object detection in the wild

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:29.098750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:19.710060Z digest=sha256:e6142ed172f29f44b3134036ab621fdbf20e30ddb395ba4dfb9b09d5e011431c

Observation 2c0833e2-f403-4540-b463-5e4af2c0c3db · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.953346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:19.773113Z digest=sha256:12783481e0c8c0775214857023fc7338c6a6ab8e978ffd265332f5708f503b2b

Observation 0b8c31d0-66db-44f9-9ef4-7319f9e8b31b · outbound

This paper cites CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:19.841794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:19.841794Z digest=sha256:26f2b29f2fdd24f48329a7fb05a1dbadc759a493ff3cb1a0646ad696b4d5d3ea

Observation 388626aa-3df2-4a49-9063-f2f5c4043019 · outbound

This paper cites Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.757700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:19.882914Z digest=sha256:39ac832115f91bf92f86f03f88e7b2fd9fcf8427ee16e81285c7c4cb5f543e9d

Observation 74be73ef-6495-406f-9033-f653c0842a43 · outbound

This paper cites A hierarchical graph network for 3d object detection on point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion A hierarchical graph network for 3d object detection on point clouds

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.627155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:19.973335Z digest=sha256:95f8429ac0d2450f7bcf51515547903083eb6d5708bd95b20d80d875e4e692e4

Observation 075ef2eb-c0fd-4efd-ad99-59efe427d5b2 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.052396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.052396Z digest=sha256:383af7601ae90446a07f1b64603621eb5823024de7a677dcf0a9d167d24856dc

Observation 38ade343-5ea4-4662-bfd7-b99dbd518673 · outbound

This paper cites Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.354338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.088537Z digest=sha256:fd744895544f9908274aeac6415625470ecc9dbe79aad9a9bfc5dd65bea5fcf5

Observation d1050601-59d1-432a-8b23-85b38cba320d · outbound

This paper cites Disarm: Displacement aware relation mod- ule for 3d detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Disarm: Displacement aware relation mod- ule for 3d detection

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.112593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.164335Z digest=sha256:baaa6da736122f5e5aacaa0ea17268e569383b6e2aecadfbf9df7d1089a22d15

Observation d8b0e6e4-1366-422a-ad17-2a5387f33182 · outbound

This paper cites 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.230116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.230116Z digest=sha256:c88e90ec6a73cd61acd134fd5eeca9246f0ac2bf86996037a12f0cb83388ee58

Observation 31c53f9e-0276-4e75-b58f-926c262a9e92 · outbound

This paper cites Generic objects as pose probes for few- shot view synthesis.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Generic objects as pose probes for few- shot view synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.266963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.266963Z digest=sha256:bb94007a6e6302b297b472fdee9a2a8e0b8406d1304ca99679159ede8a8671cd

Observation 0d60c4cf-3e4f-4aa1-be89-2bee002f3ca6 · outbound

This paper cites Training an open-vocabulary monocular 3d detection model without 3d data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Training an open-vocabulary monocular 3d detection model without 3d data

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.903048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.342896Z digest=sha256:a4a8a3396860c5e595ad48bb20e302099086e166c0acf7445d248fef0306538c

Observation f90b80f4-833b-4283-954a-ebe59cff9416 · outbound

This paper cites Particle filter with swarm move for optimiza- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Particle filter with swarm move for optimiza- tion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.648559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.428794Z digest=sha256:ceb74db815abbbfdeb1907d792d3172dcb72cd705361cc2c626ea4ed29ac01a3

Observation deac8d6b-e5d5-4cd7-bbc0-6472fae19635 · outbound

This paper cites Open-vocabulary 3d semantic segmentation with foundation models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary 3d semantic segmentation with foundation models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.458507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.468149Z digest=sha256:9f1ab5280b0050cd6f7b831c81f2709b592068bdb59adc3f4fd593342d1ed1c0

Observation c84d6efc-034c-45c1-9efa-fc92731b8523 · outbound

This paper cites Segment any- thing.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Segment any- thing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.519699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.519699Z digest=sha256:817ba26daba4f99e6019eb7c0abe729c712cbb3136e01e9288f01f4e11d1366b

Observation b7fbde06-5851-4147-b246-6cbd89a7e6b2 · outbound

This paper cites Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.208995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.613522Z digest=sha256:fe255d4b489f5fdd45545f5a38df8b340bab44b500248fb8425b54bc84273378

Observation d59def1e-fea6-4475-af13-29d4b993c969 · outbound

This paper cites Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.039717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.681004Z digest=sha256:40423333b1d015280a744c5285d0ca4c42ee2a1496b0fd2e03d4f644192250a3

Observation 3227a5b0-11eb-43fa-9c32-8ba76137b8f6 · outbound

This paper cites Arm3d: Attention-based re- lation module for indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Arm3d: Attention-based re- lation module for indoor 3d object detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.883501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.729580Z digest=sha256:0d887c9b59f3660289c6968a5006eefa8f0463b0037ab04e21231bda63080339

Observation dee36243-c969-4f07-99cd-a68810ab5ec5 · outbound

This paper cites Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction

Reference 18

Resolution
verified exact
raw_fallback, observed 2026-08-06T23:57:23.622737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.762678Z digest=sha256:ad743c188caac8df3d624378471d7fb6385372d49a8f0a8821856f264ab3f4ec

Observation c3d0f184-e88b-462e-b7ed-749b937e064d · outbound

This paper cites Cubify Anything: Scaling Indoor 3D Object Detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Cubify Anything: Scaling Indoor 3D Object Detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.836056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.836056Z digest=sha256:fac4f57fbd33791676f7d268cdf992445b523cbee8054f8c8afd95c5bbe445fd

Observation ec19e30d-dcdb-4eec-bffa-9f5e596b0550 · outbound

This paper cites Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.706827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.923788Z digest=sha256:4baff195af5a562be7399f6a83635e5430580e1053a0584cb52d07ef36e34fb9

Observation 910dc8e6-aeb9-4fc5-a6a5-3799fbe6260d · outbound

This paper cites Ground- ing image matching in 3d with mast3r.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ground- ing image matching in 3d with mast3r

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.588898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:20.959777Z digest=sha256:7e3d46af18764ec1aaeac2342030bb148fbdf47c4946cda33ae2631fa5562f70

Observation 59b8dedd-56c5-4851-8b47-19d2c55907c0 · outbound

This paper cites Grass: Generative recursive autoencoders for shape structures.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grass: Generative recursive autoencoders for shape structures

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.435530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.029161Z digest=sha256:6a958abe95f365078b06c69acc048224e510fde6de4a534ff6b11b871badb07e

Observation 54325d4d-4b34-4b92-a097-d8f71b7b38e0 · outbound

This paper cites Prompting depth anything for 4k resolution accurate metric depth estimation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Prompting depth anything for 4k resolution accurate metric depth estimation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.106713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.106713Z digest=sha256:4cc3acdda0bb0202a7265ee232f44e3ba9f5dad70de9dbb88fad991abe2dcc5b

Observation 9575c0d0-f9d0-40b1-b110-fd3ae562ec36 · outbound

This paper cites Microsoft coco: Common objects in context.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Microsoft coco: Common objects in context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.156947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.156947Z digest=sha256:298e71dce8012f8f19baad5d55309fbb531512a8f18168573dfc7bdd7fb8efc3

Observation 780674c6-9d6d-4cff-8789-faf511f73e8a · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.283502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.186493Z digest=sha256:55e7c39c624e4df62a0dec4c462fbc9d6c1e0a09150f9d5648f1ff73650163d8

Observation dd589e9a-127b-47bc-907e-1b2799c5740b · outbound

This paper cites Group-free 3d object detection via transformers.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Group-free 3d object detection via transformers

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.161915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.272610Z digest=sha256:965aac2f5ae0aa26ba761cfd74ea6a82380809d2c4e6c77c227eef9c101c97ea

Observation 91b030f3-9a99-4eec-a10d-918adbd6017f · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary point-cloud object detection without 3d an- notation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.006363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.345759Z digest=sha256:49867e9e11c96d4e7d4a7a6186f4c601015d30973d993181df348195d7c23924

Observation 2a316920-c355-4120-94c4-cefac417f91f · outbound

This paper cites Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.853536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.395450Z digest=sha256:fa1c55a9fc1f8701afb377b2b8d35f4551a2670118b2c915e4f429a143a0ae1d

Observation a546f684-a63e-49ea-95b2-b8102b33067a · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.482945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.482945Z digest=sha256:dbd669a200d91d82004f42c6b4231a010132bb68de778e054417b3495872e389

Observation 4839aa72-f402-41b5-ae08-ff2ffea84526 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Deep hough voting for 3d object detection in point clouds

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.752058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.554743Z digest=sha256:2fa0483ca34e3d6cc139542f47fb3c719039384c9c6a06ee2fc222d6e5a4738c

Observation b82ca3dc-b805-459e-a51a-130f286f6df0 · outbound

This paper cites Imvotenet: Boosting 3d object detection in point clouds with image votes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Imvotenet: Boosting 3d object detection in point clouds with image votes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.631090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.596405Z digest=sha256:7d61d4af91bde4cde09d811dc571d2b86f60c0f561ada97f9d393606a9737de2

Observation ef7a074c-2554-4528-ad20-99f58a209b56 · outbound

This paper cites High quality entity segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion High quality entity segmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.480306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.698243Z digest=sha256:a65857e3d04bcea3df667103c7cc7c94cdd6660f3405db8069d8c58f391d38d1

Observation 6b2d430e-9b97-48aa-96ac-120d0f782967 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.761942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.761942Z digest=sha256:88dce6c287d61f112a871e42b46652b56abcc34dc4ba4916a8bda87f63369fe6

Observation df1d47f0-249e-4836-be9c-390e09e8170c · outbound

This paper cites Fcaf3d: Fully convolutional anchor-free 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Fcaf3d: Fully convolutional anchor-free 3d object detection

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.372401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.814853Z digest=sha256:f14b3e92f44f0e2a8b84207aa547a6a9a37a0fa9ea508a0298e16666d9562312

Observation 1c1b4d4e-87f8-4e03-8a56-88ab71b2c584 · outbound

This paper cites Tr3d: Towards real-time indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Tr3d: Towards real-time indoor 3d object detection

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.175588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.930156Z digest=sha256:cc736bf3bea54b047a15d0e2714c5d446b041569c292225b89644e10fd0b943b

Observation f2a6349b-9088-4ba2-8369-76754884eeb7 · outbound

This paper cites CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.984804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.984804Z digest=sha256:dab46ac6aacbad69786913fe7bf8771e2bc00e6cdfcd88b07215f4a0a32d9fc0

Observation 71bfd0e5-8d45-4aee-b09c-70204a46798a · outbound

This paper cites Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.058640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.095652Z digest=sha256:d7d0047de7da65e387e57326dc3f5e354bd6c4417580b2119d5e8e37f97f97e4

Observation afcaf3b6-a840-4370-ad8d-9f1204a89f5b · outbound

This paper cites OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:57:23.380699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.137587Z digest=sha256:ad157edf92a1e709d806092112dbff5de6b6478dd8814c789cb7dd57a8010bda

Observation b51d861a-0292-4570-92da-364bf9ed4fc3 · outbound

This paper cites Spatiallm: Large language model for spatial understanding.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Spatiallm: Large language model for spatial understanding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.873173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.277083Z digest=sha256:9969042afcccd14d8a5b6df8830d2065fa2528cb4c03d5008a2e627aac559858

Observation 80edae4e-0db0-48cd-ac43-a86dd91cdb9f · outbound

This paper cites Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.755358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.318089Z digest=sha256:8a614f73a7d0d1046d7a901c0e140663630ddaede80e6fb29e541243dd3c7801

Observation 2c411c6a-c3fb-4ff7-ad39-c8b88103abb5 · outbound

This paper cites Dust3r: Geometric 3d vi- sion made easy.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Dust3r: Geometric 3d vi- sion made easy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.649214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.370207Z digest=sha256:687328d2f7a32d6fa570edcf0d31178b8cac9a21e667513e4720b923efa93e7d

Observation f1c669fc-2e85-4f4a-9ec6-ca30275c7c6c · outbound

This paper cites Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.553960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.452928Z digest=sha256:1c6c265f42fde687983b27b4584e3d05fd2d4009d9807b51a705e8071564cc12

Observation 5a1a0568-8def-4c35-b0eb-3b748886b4fb · outbound

This paper cites Mlcvnet: Multi-level con- text votenet for 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mlcvnet: Multi-level con- text votenet for 3d object detection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.451083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.497715Z digest=sha256:5554cb5c82189fa68234f97b07cd497e87cffde8a5c24d4898b5356b4b4cbc28

Observation 1ac36b7a-40cb-4bff-b7a9-79acb6ce0495 · outbound

This paper cites EmbodiedSAM: Online Segment Any 3D Thing in Real Time.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion EmbodiedSAM: Online Segment Any 3D Thing in Real Time

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.527722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.527722Z digest=sha256:36cf7fc7f800d6d12021a7bf313977c6b7c7eff47675e13dc74bf6174a177403

Observation ff68512c-d9b6-40c6-b356-a34588984168 · outbound

This paper cites Memory-based adapters for online 3d scene perception.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Memory-based adapters for online 3d scene perception

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.343165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.557788Z digest=sha256:1bd0cbaa8e16ab697cdec9e18454ecc4e782f1fb5795a27334898b6e2a6d102a

Observation 6cf57c01-6a2e-455c-8279-9dd6305f69c8 · outbound

This paper cites M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.240521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.628208Z digest=sha256:928c104c10deaf730ee882a8d73856c594dad5f34c201086c176fc373724e3ad

Observation 74bc0535-9c10-4fa6-9e84-dc79c263eb7d · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.146333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.685438Z digest=sha256:83aab0114988a1fe14115bf3662d1a255bdbd27302014482aa37a63981e321f5

Observation 015464ae-5d32-4ab9-9b2c-29a2ac39231c · outbound

This paper cites SAM3D: Segment Anything in 3D Scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion SAM3D: Segment Anything in 3D Scenes

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.728460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.728460Z digest=sha256:fe83329d3cb96b6fedb9c9d05a35b640b3965727db15f1d94b475b42e2ec03bc

Observation 552dd05e-4ae1-4825-bcb0-e26422f24d7e · outbound

This paper cites Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.026578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.763335Z digest=sha256:48f2c277853f877c133427989a9f17fed1821bb3bd4aabaf497174775e88bf06

Observation 924d9691-d87c-4949-824a-7fdcc3e8ba7a · outbound

This paper cites Detect anything 3d in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Detect anything 3d in the wild

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.826709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.826709Z digest=sha256:43a40c4be070af6f8a57ddf2a7dad929016c395058e721d145d9ee7c003e4f2b

Observation 4e9fc568-6ea5-47c4-b2a4-90ae0ef3af85 · outbound

This paper cites Rosefusion: random optimization for online dense recon- struction under fast camera motion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Rosefusion: random optimization for online dense recon- struction under fast camera motion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.943526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.878121Z digest=sha256:b7967b83b148ab97aaa36076e40b7b74c8e3c53d7a3c1c369cc3d28a619f4e30

Observation 55298ec4-c903-45d7-a8f6-00eb2c24a56b · outbound

This paper cites Asro- dio: Active subspace random optimization based depth iner- tial odometry.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Asro- dio: Active subspace random optimization based depth iner- tial odometry

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.835827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.925803Z digest=sha256:dd218140297c3849e7bbf9230d1c6015110e5607ef63c65aacae813fc9081ded

Observation 9ecbafac-1468-433c-84d3-d9f093d7f8cc · outbound

This paper cites Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.010201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.010201Z digest=sha256:b0e6436c1c5d30dc556aa31c6a7dc361568a1c6a8041c331319b14d2852325ea

Observation a182b4fd-c592-49a9-9db0-a63b974bf72d · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.071247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.071247Z digest=sha256:3de64b2503366d6459e7443a720f5199ce0119e9ded2b6b276fbd878dd8893d8

Observation a8379172-f2ab-4ea4-856e-f298b4209031 · outbound

This paper cites V oxelnet: End-to-end learning for point cloud based 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion V oxelnet: End-to-end learning for point cloud based 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.706522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:23.112112Z digest=sha256:11590b16889b73b806edcfee71c17555856958c0b0683ef63a27bccd4f95660b

Observation 01a24053-0f17-464d-ac33-dffac7eb64fb · outbound

This paper cites 1, 2, 6, 7, 8.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 1, 2, 6, 7, 8

Reference 493

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.260929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:21.877808Z digest=sha256:0196813396a6a1aeb33ee5e29585a5d95b435acb360ffd9c9d23a36bcaec43f4

Observation e659c0d4-049c-4405-9e6f-930cac7aa061 · outbound

This paper cites an unresolved cited work.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Unresolved cited work

Reference 2025

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T23:57:24.976651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:57:22.189342Z digest=sha256:7c6d878da59fd55b5e920f6bd65be2152db709aec050b3aa1f03070e74b336a3

Pith citing papers

No inbound Pith citation observations are available.