Pith. sign in

Paper Citation Record · LEDGER

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations

As of 21 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2412.11412.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11412 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:01:33.101578Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 88d7e4c3-a819-40c6-89bc-7e14d3ffb88a · outbound

This paper cites ARK- itscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations ARK- itscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.130075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.819069Z digest=sha256:fa0e349ed39b311b131dfc1a522488deb811dc690dfe886ea536a5097cfd8c8b

Observation 5deb3153-9e0e-48cd-a832-a81b6b5bce56 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.825902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.825902Z digest=sha256:9d5eed9d45df54ab347ffd974b9f57abf7d00c7e33b531a7a2aa5122a1798601

Observation 45870842-9e13-4590-b8ba-13c3663a0c1d · outbound

This paper cites Omni3d: A large benchmark and model for 3d object detection in the wild.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Omni3d: A large benchmark and model for 3d object detection in the wild

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.107460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.832395Z digest=sha256:9a68cd082e6f5d2d1f090de791a4e21c93feb66694a6435600f7cec728b0ea0c

Observation 5441bd91-2ee2-4822-9941-6f2a2edf5ac8 · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.087970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.837973Z digest=sha256:dcc487c0a2d4c5bd025c8ca7c0767ebe11bdb2936a1f44e6fb2c870a691c14cf

Observation 1e58a82f-6466-4a47-ad05-da3854ba49e2 · outbound

This paper cites End-to- end object detection with transformers.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations End-to- end object detection with transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.843953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.843953Z digest=sha256:827cf11d7a855bd49453dd2bd5ebd3e045037034241fda4212cd858cc994d6c3

Observation f0eac324-f7a8-4a70-8c99-c8d8b793f9b4 · outbound

This paper cites Exploring classification equilibrium in long-tailed object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Exploring classification equilibrium in long-tailed object detection

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.052947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.850680Z digest=sha256:06f34167bcb1efa83e23fb93eb021ce5a85c8556825093fa006a03b0f6afbb19

Observation aeb6ac87-6f10-461a-b8e0-22c7034c164b · outbound

This paper cites Dqs3d: Densely-matched quantization- aware semi-supervised 3d detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Dqs3d: Densely-matched quantization- aware semi-supervised 3d detection

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.033969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.856107Z digest=sha256:99a44434b536b4476c1b91c4bd9c48de14e788a75e26c1af01de6f9fa972d057

Observation 990c192f-eb67-4516-807f-431d14fa9c31 · outbound

This paper cites Fast r-cnn.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Fast r-cnn

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.015333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.862565Z digest=sha256:1d3863bd94275da04fc74c4ba84a03af8e8ab800eac1ebf5d2797aecb77723e0

Observation c5d034e0-83f5-4a34-bd6b-d9030947d806 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Rich feature hierarchies for accurate object detection and semantic segmentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.998956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.868468Z digest=sha256:5dd9f5da472903e3b610c96606cf72509fcd6dff8d3be37449cea2aefe00f0c0

Observation 710d132a-d2d7-49fb-a757-7b7180520ace · outbound

This paper cites Open-vocabulary object detection via vision and language knowledge distillation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary object detection via vision and language knowledge distillation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.980561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.873751Z digest=sha256:ec60e9568b6ce61fa64288ae35b4d87992e534ead6c1e69b06b3517aaad87eda

Observation 99a5487d-c825-4040-9ecd-e184b426a038 · outbound

This paper cites LVIS: A dataset for large vocabulary instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations LVIS: A dataset for large vocabulary instance segmentation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.962656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.879582Z digest=sha256:bf45cb6a4d6db2ceb0a4f77850ee1074c30006ec52fbd786d285374935e7dd3b

Observation 59b1d90f-f408-4140-b337-526a3e9fe003 · outbound

This paper cites Mask r-cnn.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Mask r-cnn

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.944279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.885508Z digest=sha256:b66bb41741beaeda8cfe12787af7d66ed1139c20cf91039b0c72ac30cce3d492

Observation dd0d3eb3-9740-45a9-90e9-c545304fbf58 · outbound

This paper cites Cooperative holistic scene understanding: Unifying 3d object, layout and camera pose estimation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Cooperative holistic scene understanding: Unifying 3d object, layout and camera pose estimation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.926961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.891321Z digest=sha256:65f63d5060a1248550c5e1cdfd944406be9d17762253f4d126452a8be8fa8646

Observation 166f0d29-6c2e-441e-a923-87f8aa162fd8 · outbound

This paper cites 3d-relnet: Joint object and relational network for 3d prediction.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations 3d-relnet: Joint object and relational network for 3d prediction

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.909519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.897240Z digest=sha256:e65a217a3ade7d157367a1208c02b776e03913616111778bc5cf61a60eb0a7c7

Observation 654a91fe-b3c2-4320-810f-cba6662aa775 · outbound

This paper cites Cornernet: Detecting objects as paired keypoints.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Cornernet: Detecting objects as paired keypoints

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.891078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.904736Z digest=sha256:ba1f87de7f273b2bb629b39633db55fae66ae0ce09e18d19f147a7b7be47f38c

Observation b1519e68-ff5f-4d98-a294-c8e93c924c1c · outbound

This paper cites Overcoming classifier im- balance for long-tail object detection with balanced group softmax.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Overcoming classifier im- balance for long-tail object detection with balanced group softmax

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.873986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.910845Z digest=sha256:828e71f592047e70417f341cf05f44dbbe11c4749a287bc444586178c5b68ae0

Observation 53370ef3-ee1f-4b19-8a72-1da4c33ad235 · outbound

This paper cites Towards Unified 3D Object Detection via Algorithm and Data Unification.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Towards Unified 3D Object Detection via Algorithm and Data Unification

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-11T15:01:33.248430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.917088Z digest=sha256:036a36a222aa18091eb5dd7450e52ec5d308acf637c8cf4c6a9e75bb9e485fa2

Observation 97e0f6ff-fb58-40ea-ab56-4712a01e5ce1 · outbound

This paper cites Ssd: Single shot multibox detector.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Ssd: Single shot multibox detector

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.856016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.926872Z digest=sha256:b33a38114cc0e4d0dbf779183a8d6de747d84e03b2510dd0285edc310ea6078c

Observation 36342aa1-3951-4467-8802-58fbcdd5f140 · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary point-cloud object detection without 3d an- notation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.838797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.931772Z digest=sha256:aa422a62548bcd17c74c2d08ee90ffc1d7ff238f7aab0fac9d369ee21f112812

Observation 25943ee2-4262-4070-9764-530e0ac5ee3c · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary point-cloud object detection without 3d an- notation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.821828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.936810Z digest=sha256:07df8115709073da803ba163493a453e13ca532ac89ea3648bbad71a8169c30a

Observation 491d8e27-e166-4fa6-8764-fc4de63ac762 · outbound

This paper cites Total3dunderstanding: Joint lay- out, object pose and mesh reconstruction for indoor scenes from a single image.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Total3dunderstanding: Joint lay- out, object pose and mesh reconstruction for indoor scenes from a single image

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.804084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.941913Z digest=sha256:127a97ce3c64e50f9181a4d25724ac6302daf5319786cb6dfc8f8d725902b48a

Observation 5e8ce9cd-2d90-46ca-b45f-0e1d706b16d8 · outbound

This paper cites On model calibration for long-tailed object detection and instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations On model calibration for long-tailed object detection and instance segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.785036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.946996Z digest=sha256:92cb606ca8aaaf19440e3aa83d1f6e23956791e4ebf0d798881c3cf81f929a2c

Observation 6188f051-6e9c-4e9c-8eb0-636d3780fc1c · outbound

This paper cites High quality entity segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations High quality entity segmentation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.762241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.952133Z digest=sha256:79b99d2ab153ba6cd7147d233e8c83b5ee76eec4054959cb88dc9c2b20371793

Observation 440c4190-259e-4f95-bcff-a94ccf8a616a · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Learning Transferable Visual Models From Natural Language Supervision

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.956830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.956830Z digest=sha256:9883c8d5f5118718f08e1464a5dbb4cf6afe90be2ac27b8005a61d09a9b3dc28

Observation 486a7f24-64bb-4348-8eee-799f93b9bbab · outbound

This paper cites Improved visual-semantic alignment for zero-shot object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Improved visual-semantic alignment for zero-shot object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.744752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.962414Z digest=sha256:f5a7f2fb0e463b2d5784ec2d292723c74e85cf9762e5f8bc2e21021fad0ac8a4

Observation 424b6811-85c8-4ad4-a860-5c4da54398ad · outbound

This paper cites You Only Look Once: Unified, Real-Time Object Detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations You Only Look Once: Unified, Real-Time Object Detection

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.967400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.967400Z digest=sha256:378009d98deaa6cb77d88335a24a03d45d3d736f00517d84140d68c14cd4434b

Observation dc4f12af-3a8d-479a-a7a8-266e3a327367 · outbound

This paper cites Yolo9000: Better, faster, stronger.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Yolo9000: Better, faster, stronger

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.724336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.972620Z digest=sha256:f0277a060bb18f964ca2dfc9ac29cd713accb3c1be0b0220920c020f82af0ecf

Observation 68565193-9ad0-4579-b4b8-11e148445e53 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.705623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.977879Z digest=sha256:c9100121bdbe9ab57d94d424e1ff2f401694f5bf7c62b92b4766577dffab0401

Observation 65d8cf1b-13cd-4010-8595-fdfdc80b6c0c · outbound

This paper cites Susskind.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Susskind

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.686990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.983136Z digest=sha256:42bfccc91fa9ceb8fd010f663a8dd2004538882a13012c8daf97574b8803689d

Observation 237754b2-08d2-4433-8b2c-fc09950fcb95 · outbound

This paper cites Language- grounded indoor 3d semantic segmentation in the wild.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Language- grounded indoor 3d semantic segmentation in the wild

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.667928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.989741Z digest=sha256:3bd775a328ed9754ddd5e80336f079e265d0c508308138548a59a255aff2b8ea

Observation 9a20f1d5-44cb-400c-96d4-f15d7f3359de · outbound

This paper cites Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.648685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:32.997903Z digest=sha256:13cd932a9afbc4a21a0e8ce94b363b6eda8643f64996c72efd99f7d344b3ee6e

Observation e88424da-79a8-4aaf-80db-bd918550f311 · outbound

This paper cites Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.629442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.005910Z digest=sha256:3086d62004c3eb552128187fff27d9a97ded54088954277cf3e27e0162adcb04

Observation 09201718-0f36-4011-a98a-dc2f9dccb921 · outbound

This paper cites Object detection with trans- formers: A review, 2023.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Object detection with trans- formers: A review, 2023

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.608207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.012374Z digest=sha256:5075a211339a7111cfc91a0a31d07173e3180aa40f2e47794ac5139c4c957f57

Observation 9eac7d22-94d9-46bb-af41-dce3d90e9998 · outbound

This paper cites an unresolved cited work.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:33.587612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.018628Z digest=sha256:0187bf639760e87cee947505d00afca146db2e9cd979e10c57a0076c07468105

Observation 8c365951-a0a7-46b0-af0b-821fb756bf85 · outbound

This paper cites Lichtenberg, and Jianxiong Xiao.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Lichtenberg, and Jianxiong Xiao

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.566083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.024300Z digest=sha256:d067ee3d2b8a2cb302ad14e3e643bdbcc05fa8927d8ef616fad8be15cbfb6518

Observation 95431c26-5f2f-42de-927c-354b01eabecf · outbound

This paper cites Equalization loss v2: A new gradient balance ap- proach for long-tailed object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Equalization loss v2: A new gradient balance ap- proach for long-tailed object detection

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.547484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.030727Z digest=sha256:ddd58f6dc4b6afee37753ed96ed20ca61b0fb37c59c4fe6222f6082b7fcf7576

Observation e303a519-1db5-482f-8a30-314b9f5829a1 · outbound

This paper cites Equalization loss for long-tailed object recognition.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Equalization loss for long-tailed object recognition

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.529520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.036167Z digest=sha256:5f8032e63f11962843c137a52991e2de0a574903fb564fe4c18ed03a4cc0c1ca

Observation 73c663c1-342d-4235-a9d8-3fb7f4a982ee · outbound

This paper cites Im- geonet: Image-induced geometry-aware voxel representation for multi-view 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Im- geonet: Image-induced geometry-aware voxel representation for multi-view 3d object detection

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.507325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.042021Z digest=sha256:3a5180462d3246e0d53d5fe0fdf90a1b9f78139679da31596d2d4db952b3eec4

Observation 444d46ca-d067-48d7-9b5b-4df0c17251fd · outbound

This paper cites Efros, and Jitendra Malik.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Efros, and Jitendra Malik

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.485450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.047593Z digest=sha256:910a53c4451a0e8320c94d3ab712d8ee5c697ed644a2ab71d61273cc934b50ab

Observation 7fae7716-30e6-4cc2-9446-230a528ea6be · outbound

This paper cites Seesaw loss for long- tailed instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Seesaw loss for long- tailed instance segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.461296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.053003Z digest=sha256:bd1699fc29721b7b4bc9a305f9c7779f58e4d56b825a0dc0fd6261f0cb38893e

Observation f5e5559b-19c4-400c-a8a5-70fcd0e49400 · outbound

This paper cites Detecting 11k classes: Large scale object detection without fine-grained bounding boxes.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Detecting 11k classes: Large scale object detection without fine-grained bounding boxes

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.442801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.060006Z digest=sha256:7ce22414a1cb77b8a5b36af331e2c092c21562ff14f841959369fb63c3583a8c

Observation 32bc1dbe-0e64-45bc-90fe-d0747c3c6652 · outbound

This paper cites Open-vocabulary object detection using captions.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary object detection using captions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.422681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.065394Z digest=sha256:81a4409e096bdafbbe87134225ed58b5314c6b79c730fd76be206174ff25fd74

Observation 8c77f4fb-a743-4bea-a29f-651a1107745d · outbound

This paper cites Distribution alignment: A unified framework for long-tail visual recognition.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Distribution alignment: A unified framework for long-tail visual recognition

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.398101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.072574Z digest=sha256:ed36dcd7f8dedf37ec381bdb51069754e17363213c3a59e8dc8ab227852ba7fb

Observation 244f5ff2-1e3e-4507-a658-852a106662d4 · outbound

This paper cites Detecting twenty-thousand classes using image-level supervision.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Detecting twenty-thousand classes using image-level supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.377779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.078707Z digest=sha256:4fbef99ba1bb10384936c8f137c0e19d37ab3504ce7b66d820366b37aeb9cc6c

Observation 97d2a028-2783-4762-933f-9bf4dad93dd6 · outbound

This paper cites Objects as Points.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Objects as Points

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.085113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:33.085113Z digest=sha256:5f9fd2361e1a4ef3d204a718283dd68d72d165bbdba107dc6d498b5db7ac0e13

Observation 45c93069-dc93-4b80-a892-1c2003893245 · outbound

This paper cites On the continuity of rotation representations in neural networks.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations On the continuity of rotation representations in neural networks

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.355539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.091840Z digest=sha256:9b433c1b7db8290830a9cba56712c2285d101c4f103d6347acd7bdce7244bb5a

Observation 3eca40a0-15ec-4b56-8658-495af94c19a5 · outbound

This paper cites Tame a wild camera: in-the-wild monocular camera calibra- tion.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Tame a wild camera: in-the-wild monocular camera calibra- tion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.332911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.097085Z digest=sha256:b724e78b6f68eeb53499a38f4711567d0fad06cf49c8f1636c80508fb4f2d699

Observation 5578a354-832c-4560-a367-3464fc881645 · outbound

This paper cites Deformable {detr}: Deformable transform- ers for end-to-end object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Deformable {detr}: Deformable transform- ers for end-to-end object detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.307728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T15:01:33.101578Z digest=sha256:d33cb42075daf645774f1cd6dc603fb5228c4358bab64d1e5a887106c72da224

Pith citing papers

No inbound Pith citation observations are available.