Pith. sign in

Paper Citation Record · LEDGER

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations

As of 21 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2508.20063.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20063 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:19:09.753843Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact1
  • verified fuzzy59
  • unresolved4
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 077c1ffc-6d71-43c9-9649-db5388c34799 · outbound

This paper cites Multi-view depth estimation by fusing single-view depth probability with multi-view geometry.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Multi-view depth estimation by fusing single-view depth probability with multi-view geometry

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.280789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.573384Z digest=sha256:2470782bd10da78407cdf4e3e58fda1585f844e9e92a4f0dde56ee487ad5e6e9

Observation 22b74cbf-1921-4152-bd47-b8bee8e2c6a3 · outbound

This paper cites Arkitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile rgb-d data.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Arkitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile rgb-d data

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.272912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.576911Z digest=sha256:4edba4e65a7af7f2ad39f90f1d21ce7f5e4dda8e52c7fb424fb8f04e3d3e38b6

Observation 882ae802-c3c6-47d7-9a2c-92ca3c281dbf · outbound

This paper cites Coda: Collaborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Coda: Collaborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.265469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.579508Z digest=sha256:4d34c64bdc56c7cc915787977b47ff8aa81a52a38b7960e4b7a9eb2769ddd65b

Observation cb95a57f-e18e-46e2-96b6-d01db232c90c · outbound

This paper cites End-to-end object detection with trans- formers.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations End-to-end object detection with trans- formers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.257575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.582665Z digest=sha256:38e8dfbc1e425e8d52f0c643c577791482f9b15fe975d40dbb98863e5ee114dd

Observation 270c4124-e4fc-4704-a788-7633f1b16b91 · outbound

This paper cites Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.250879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.585132Z digest=sha256:824c1180d3fda7fd154cfb4994915dd0e5f55cb4937aa809f5dfc33d27f169eb

Observation bdb0772f-5294-4399-bd60-1104ea802948 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Imagenet: A large-scale hierarchical image database

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.243791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.587681Z digest=sha256:341ec223ab6caeaf56a21a411a5fd64bf7fd7324faf0204f78aa888b6d6f1f90

Observation b58466ef-07b8-4495-bd4d-faa22de70fcd · outbound

This paper cites V oxel r-cnn: To- wards high performance voxel-based 3d object detec- tion.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations V oxel r-cnn: To- wards high performance voxel-based 3d object detec- tion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.236035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.590269Z digest=sha256:bb006e1f8506a7d4f622f21607b3d850e35dd1a2341598eb010032ec10fa89a6

Observation 7ff973e2-cbb0-4b7d-927f-83310d3bdfc0 · outbound

This paper cites Effi- cient graph-based image segmentation.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Effi- cient graph-based image segmentation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.229064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.592735Z digest=sha256:e1fd16857143655f7482549c26c42838dd0f9052207396e0f35130bde9d31233

Observation 8d73ed2e-4250-4252-b941-ec29e3247cd6 · outbound

This paper cites Scaling open-vocabulary image segmentation with image-level labels.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Scaling open-vocabulary image segmentation with image-level labels

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.222335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.598223Z digest=sha256:14d681e0c4d0e84d2f68786b593f1c1a65ccf4beaa5b62a57aca6ee16d4d38ba

Observation daff6758-417e-43d3-95d2-8c7c9e8f90dc · outbound

This paper cites Open-vocabulary object detection via vision and lan- guage knowledge distillation.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Open-vocabulary object detection via vision and lan- guage knowledge distillation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.214888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.600529Z digest=sha256:16b4d0a03ee1b253525370018ef92938b9b001d7439a8d6e17cb2847ad1a9717

Observation 2420e6ef-5ea1-437d-a7c8-2e1060ee3fb2 · outbound

This paper cites Generative sparse detection networks for 3d single-shot object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Generative sparse detection networks for 3d single-shot object detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.207221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.603918Z digest=sha256:266e5c0f992be80549e84df6fa60ee25292f4b4847d54e22e4244c7a338edb27

Observation 798f0c41-e90b-46f7-95d5-17f3e6bff649 · outbound

This paper cites Semantic abstraction: Open- world 3d scene understanding from 2d vision-language models.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Semantic abstraction: Open- world 3d scene understanding from 2d vision-language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.199809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.606380Z digest=sha256:7fe623c4db804cb72ce36f2279a97790ffcf13e94063f91d5645b6b6dff3bb48

Observation c28e0875-5a9f-45ce-9a47-e7ae54e5716b · outbound

This paper cites Deep residual learning for image recognition.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Deep residual learning for image recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.193013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.609756Z digest=sha256:bb1d2f1635af3134098bc2d463faf73016708d34dc2e4d9a6bf8eec29fc6c4cf

Observation 95268950-d0b7-455e-baf6-91da83f2ff88 · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:09.612241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:19:09.612241Z digest=sha256:9147ea0c9dfa2b4194ea49b21a637e69c5fb3886c6767bf53343ce6421edb246

Observation f4478643-8281-4203-a03f-9525539691b8 · outbound

This paper cites Tenenbaum, Celso Miguel de Melo, Madhava Krishna, Liam Paull, Florian Shkurti, and An- tonio Torralba.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Tenenbaum, Celso Miguel de Melo, Madhava Krishna, Liam Paull, Florian Shkurti, and An- tonio Torralba

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.185098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.615293Z digest=sha256:f43b2618b35fc0b6f759e4cfd8db609e5f50e5a2f6fa0fc03b7f5b91cb84b586

Observation adc71f5f-b766-494f-a8db-533a7f8d1781 · outbound

This paper cites Scaling up visual and vision- language representation learning with noisy text su- pervision.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Scaling up visual and vision- language representation learning with noisy text su- pervision

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.178138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.618108Z digest=sha256:27963a30be8fac05e48fc8273d99cefa996d70a597c8ef599ba5593efe5f3b57

Observation 9b3c69a2-6a6e-4c95-a41d-74707cfafd33 · outbound

This paper cites Lerf: Language embedded radiance fields.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Lerf: Language embedded radiance fields

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.170673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.621252Z digest=sha256:645ed8a82d558c54dceb05f3e4f9e29a7ba6fec2c52f72f5843c4bc03aa24d72

Observation 00cc5d74-303a-4a4c-b80e-d16a985ba014 · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Dollar, and Ross Girshick.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Berg, Wan-Yen Lo, Piotr Dollar, and Ross Girshick

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.162967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.623423Z digest=sha256:673fdc61998f176866dae364e071d53c6198649a2b729fe66babf8c3312f8985

Observation e33bcafd-2c20-4030-beda-cdf1e3d396b6 · outbound

This paper cites Learning multiple layers of features from tiny images.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Learning multiple layers of features from tiny images

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.154687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.625792Z digest=sha256:b774fbb1d7e8eda1f87218dfcd0d7e393a014b2c226a0933194c74d669e86372

Observation 783dffc3-db45-4cdd-8106-b36035d3542a · outbound

This paper cites Pointpillars: Fast encoders for object detection from point clouds.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Pointpillars: Fast encoders for object detection from point clouds

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.147914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.628366Z digest=sha256:f094b96a68e65a58e77b2f23d24d6800da49b69a59236f9173d7555b1dc95839

Observation 18f42ccb-3a63-4a59-9fee-791806916830 · outbound

This paper cites Mask dino: Towards a unified transformer-based framework for object detection and segmentation.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Mask dino: Towards a unified transformer-based framework for object detection and segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.140796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.630663Z digest=sha256:b9cc7025dcca215586f95b2ec8cef1e40d383f9babfe1649ffb0b96131c89227

Observation 0d063751-48bd-491d-90a0-a3ac7f76fc32 · outbound

This paper cites Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal trans- formers.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal trans- formers

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.133002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.633072Z digest=sha256:b08992ff446401c079e1f4f8fb0d67837ca5c1508a0881daf3946773f7d1d81d

Observation ce358653-a3c1-46fa-a842-53eee2ace87d · outbound

This paper cites Focal loss for dense object detec- tion.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Focal loss for dense object detec- tion

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.124883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.635980Z digest=sha256:182177e9eb8b18f5985dcd589d7846d74719d362d260789c13cc72e37ab32a8f

Observation d8beca92-0ca9-4714-af61-cf0b5e699d5f · outbound

This paper cites Petr: Position embedding transformation for multi-view 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Petr: Position embedding transformation for multi-view 3d object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.118157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.638980Z digest=sha256:4d3fae9e03fe67ac1626af1e2223d04bb6a3bdf7d8698175042fbc516647964a

Observation cfa2cdf6-9989-48af-99b3-10e0853c39ad · outbound

This paper cites Decoupled weight decay regularization.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Decoupled weight decay regularization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.110615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.641404Z digest=sha256:b32e5b3df32d022840867706b185c63f045b3a45f00a124cdb0786591e37bec0

Observation 85a70d69-4463-4285-971a-b7c09b9771c6 · outbound

This paper cites High-quality entity segmentation.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations High-quality entity segmentation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.101863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.644399Z digest=sha256:24a227bcdb7a0363fc4c53a66f78d573cb1f60bfa848e95e4a40aeeb127765a6

Observation f031b041-aa10-486a-ac24-3001f9323481 · outbound

This paper cites Ovir-3d: Open- vocabulary 3d instance retrieval without training on 3d data.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Ovir-3d: Open- vocabulary 3d instance retrieval without training on 3d data

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.093413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.647868Z digest=sha256:4856fa91265c3da5e6c7f895637ce1f75e94d6b309d233952b7489c31cf2ebce

Observation e058cdba-2b51-4ba7-afee-83e1a40a901d · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d annotation.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Open-vocabulary point-cloud object detection without 3d annotation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.085125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.651682Z digest=sha256:4e2eb612f9cb042eb86a0485bd7c9ae4339cb67a5ada2123f30bcdfe404bf4c8

Observation 8d5bb906-311f-488d-b221-f404e712a6f7 · outbound

This paper cites V oxel transformer for 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations V oxel transformer for 3d object detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.078312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.655058Z digest=sha256:a52b416c43ef328ea4fc69066eea3d2a9d84d6913636c1ebf4b2954bf67ef8d0

Observation c339d6cb-86f0-436b-91d1-49ad59abeb57 · outbound

This paper cites Atlas: End-to-end 3d scene reconstruction from posed images.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Atlas: End-to-end 3d scene reconstruction from posed images

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.070690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.657539Z digest=sha256:63a78a239495f1fb2eb573ec1a35d1eb0f2ca1af486ab2962e66fced2c3f88ec

Observation b496e5b2-8849-435b-8787-e0dc7790baa9 · outbound

This paper cites 3d object detection with pointformer.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations 3d object detection with pointformer

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.060710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.660946Z digest=sha256:04c62cbfde9c1ecbd061847d2b612f1f9fc2561c72d9da41fce5c2ea6196632d

Observation 8597b81e-6556-4126-bdd6-d432cf1a2c87 · outbound

This paper cites Openscene: 3d scene understanding with open vocabularies.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Openscene: 3d scene understanding with open vocabularies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.051927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.663167Z digest=sha256:bdd6c4d2773bf8d515eed66ebc7dba8930e098456868c470686fe0225438947e

Observation b4ad6772-67e5-43d8-bc52-a4ea34f58f41 · outbound

This paper cites Deepwalk: Online learning of social representations.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Deepwalk: Online learning of social representations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.042087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.666187Z digest=sha256:5861374d0eacd40869d733bd803f7ccad7edc0a1727c03f516abfcc0e4af6ea8

Observation 323c479a-43e8-4226-9dcc-0c290600d893 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Pointnet++: Deep hierarchical feature learning on point sets in a metric space

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.032391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.668439Z digest=sha256:bc9040b9d5115e1f74d96380bc1988397d5b2c53f92cd643efe4c10c76dd5182

Observation 05f009a5-5cf0-4fee-8450-8658bffcdb7d · outbound

This paper cites Frustum pointnets for 3d object detection from rgb-d data.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Frustum pointnets for 3d object detection from rgb-d data

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.023746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.670689Z digest=sha256:b7f554b68652dca5cf0d54bb4d31cc2b4fea99840e7e796576db9bdab180708a

Observation ac84dc69-80ab-4c47-9d08-864b0c77e165 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Deep hough voting for 3d object detection in point clouds

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.015685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.672976Z digest=sha256:d49165cd99b053172909d46ffa0f8186317b00130bf5796f561718b2e36e488a

Observation 7c15497b-4b8c-4e21-9cd5-2c620e772dda · outbound

This paper cites Learning transferable vi- sual models from natural language supervision.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Learning transferable vi- sual models from natural language supervision

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:10.007815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.675793Z digest=sha256:46a0bc7d49e9b0ecb51f2ccb7a7b1facc867bfb971180b2c7f5485fe7ef516d7

Observation ace5f1ff-a5df-4dc8-b741-2bdc1df0ed7f · outbound

This paper cites Im- proved visual-semantic alignment for zero-shot object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Im- proved visual-semantic alignment for zero-shot object detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.998309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.678634Z digest=sha256:3750fb5a94b621a11d24ac6f10fd7df02352c8b9177a1798c89db60a2ff8aaca

Observation 755178b6-c092-4a4b-9ff0-eec728197989 · outbound

This paper cites Language-grounded indoor 3d semantic segmentation in the wild.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Language-grounded indoor 3d semantic segmentation in the wild

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.989388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.681007Z digest=sha256:4a54e57197ef166946854b9f24a32668ab078e44f9cbaa1ee063dfb74a2d5630

Observation 40c6b85d-315c-4f0f-9bf4-b0295c0a9ae7 · outbound

This paper cites Fcaf3d: fully convolutional anchor-free 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Fcaf3d: fully convolutional anchor-free 3d object detection

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.981221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.683296Z digest=sha256:82682cfbe0e43ff7f0cc73649e8de544b66d6fe2edbbef6e6aa9dec3f04770af

Observation 43570cb9-34b1-4f0d-9608-e7af431e7a8e · outbound

This paper cites Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d ob- ject detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d ob- ject detection

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.973557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.686051Z digest=sha256:7687dcae668c7c1e1dc9167a4cd75c3124d3c42fedef7be8ebc2ce8e8d7efb88

Observation 2d80a827-31b7-4c6e-87cd-df3c13b3080f · outbound

This paper cites Pointrcnn: 3d object proposal generation and detection from point cloud.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Pointrcnn: 3d object proposal generation and detection from point cloud

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.965066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.688331Z digest=sha256:03ab628f73860f2f7a99418340b201732d55d8dfe56cca1d4f2c1517b4da04fe

Observation dbfafd01-9bd5-4108-9b96-e45fda437a90 · outbound

This paper cites From points to parts: 3d object detection from point cloud with part-aware and part-aggregation network.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations From points to parts: 3d object detection from point cloud with part-aware and part-aggregation network

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.958075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.690752Z digest=sha256:f6cb9a307138c8b7446d9970e625ec669ec93c7658d6530fb0a1e9cd8e203024

Observation cafedb21-c360-48c9-8698-503dd9305b08 · outbound

This paper cites Point-gnn: Graph neural network for 3d object detection in a point cloud.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Point-gnn: Graph neural network for 3d object detection in a point cloud

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.950653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.694038Z digest=sha256:b8a13b30046089548286e2c877ef843472ef97f3fdb61b81c492386a522d5499

Observation 1a7d2f19-3e14-4139-a986-b434eabaf0e8 · outbound

This paper cites Sumner, Marc Pollefeys, Federico Tombari, and Francis Engel- mann.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Sumner, Marc Pollefeys, Federico Tombari, and Francis Engel- mann

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.942245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.696948Z digest=sha256:2f8937407f3c0601debdcd4540e40e64619947c2967e59aab59952a57b989816

Observation bb8b220d-d61f-40e6-bdd7-90a55a34b9d2 · outbound

This paper cites MV-DUSt3R+: Single-Stage Scene Reconstruction from Sparse Views In 2 Seconds.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations MV-DUSt3R+: Single-Stage Scene Reconstruction from Sparse Views In 2 Seconds

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:09.699878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:19:09.699878Z digest=sha256:05162f8b5c295f1954e24f267f8ec740527be6f2a49c2eb096e317ba492b96f4

Observation b564b683-fd54-4e2c-91a0-7cc8aff25215 · outbound

This paper cites Fcos: A simple and strong anchor-free object detector.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Fcos: A simple and strong anchor-free object detector

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.935278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.703125Z digest=sha256:9374b8efc8a556f60d6d0fb9fe3f4f90044be10a3f23ca9315ee2903c555d8dd

Observation 4dc3b85a-9940-4033-b8d1-079a29aa8caf · outbound

This paper cites CrossDTR: Cross-view and Depth-guided Transformers for 3D Object Detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations CrossDTR: Cross-view and Depth-guided Transformers for 3D Object Detection

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:19:09.794914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.705999Z digest=sha256:8e5f899cc4187d9ba4a25f5f9f4699fab5c13ce11d1fbb8af8496cfd7900da03

Observation 2397532f-7423-47e7-8b52-86162117f48a · outbound

This paper cites Imgeonet: Image-induced geometry-aware voxel repre- sentation for multi-view 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Imgeonet: Image-induced geometry-aware voxel repre- sentation for multi-view 3d object detection

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.928081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.708428Z digest=sha256:103e85810efcc009f3ad9fcf095abaf1abd204b6d512e3ddb1fc71174a7be776

Observation 52284338-5e2b-4ff1-95c2-148c809914fe · outbound

This paper cites Vggt: Visual geometry grounded transformer.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Vggt: Visual geometry grounded transformer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.920605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.711643Z digest=sha256:cf1efb771334ab24b26facf6c05c56251335d5b7f66b3aee61ba385f6041cb7c

Observation 7b430f11-994c-4ca5-a138-29360910565a · outbound

This paper cites Detr3d: 3d object detection from multi-view images via 3d-to-2d queries.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Detr3d: 3d object detection from multi-view images via 3d-to-2d queries

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.913364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.714126Z digest=sha256:d7fc51b3e15f8aa0daf51a95453186bae842d5423de5286508667dfbb17bf238

Observation c5cef031-9e9e-45aa-a976-28d7687270a3 · outbound

This paper cites Nerf-det: Learning geometry-aware volumetric representation for multi-view 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Nerf-det: Learning geometry-aware volumetric representation for multi-view 3d object detection

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.906158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.716630Z digest=sha256:cb43a27052c25f7609534d0e37bd1a133c297e88b103a89ffc07a6a9e6e5039f

Observation b147ef40-9f35-48a7-b85d-983108732769 · outbound

This paper cites A simple base- line for open-vocabulary semantic segmentation with pre-trained vision-language model.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations A simple base- line for open-vocabulary semantic segmentation with pre-trained vision-language model

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.898606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.719017Z digest=sha256:4f6f2f985d0df896916ca8512c9252588bbc9a5f7d1f46853feafbd609a1286e

Observation 102b8f82-94d2-4e17-9b55-313523c65ac5 · outbound

This paper cites Second: Sparsely embedded convolutional detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Second: Sparsely embedded convolutional detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.890092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.722088Z digest=sha256:14a27b1d80d6ea4e3f28a88f8c9bb8f4d03c43549d0f80d2431d968386490363

Observation 36153ef6-dade-4e31-82c8-54ea889d24f2 · outbound

This paper cites Pixor: Real-time 3d object detection from point clouds.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Pixor: Real-time 3d object detection from point clouds

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.882006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.724732Z digest=sha256:c8aa0e0a08fa51f46171a6d850080aaaa547b31b6f9486f45c1a7de59d11972a

Observation 5342bf0c-eace-418e-a571-3c98ae7398b1 · outbound

This paper cites Imov3d: Learn- ing open-vocabulary point clouds 3d object detection from only 2d images.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Imov3d: Learn- ing open-vocabulary point clouds 3d object detection from only 2d images

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.874193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.727324Z digest=sha256:2045a1a6a040c9bd9efa23e5c19e40f3689f11f84e877328baa80aae2091bf44

Observation ba465949-8078-4741-821e-54727f47dbad · outbound

This paper cites SAM3D: Segment Anything in 3D Scenes.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations SAM3D: Segment Anything in 3D Scenes

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:09.730952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:19:09.730952Z digest=sha256:34987162b742565f21ab670ad2e3fc5fa68f94bc75bc089b423c89fc88012aa5

Observation 337a8f46-03c8-4416-96a7-b4801890d295 · outbound

This paper cites Std: Sparse-to-dense 3d object detector for point cloud.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Std: Sparse-to-dense 3d object detector for point cloud

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.866010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.734232Z digest=sha256:0ee39dafe014172cf0750fd81f914111c435759382bcf646e4c1d3159d274073

Observation 7ea55eaa-c3fc-46cc-8c45-a578ddf0175b · outbound

This paper cites 3dssd: Point-based 3d single stage object detector.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations 3dssd: Point-based 3d single stage object detector

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.857859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.737314Z digest=sha256:5f61143f65313f1e4619bb070fd3f6a96c324a975192ad8a1271f5eab475c897

Observation 08ec8f88-8d18-49be-88bd-985b3fa24936 · outbound

This paper cites Open-vocabulary object detection us- ing captions.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Open-vocabulary object detection us- ing captions

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.849389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.739611Z digest=sha256:5f1717152ee45cbc44030fea01b2e4c3b3b7e0c73d0e1577c3f59e4e4a631fc1

Observation ae7fb3c8-a23c-4b70-9d6f-6909f3678a76 · outbound

This paper cites PointCLIP: Point Cloud Understanding by CLIP.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations PointCLIP: Point Cloud Understanding by CLIP

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:09.742405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:19:09.742405Z digest=sha256:eaea49956567dc27d03e92df4ba2026d2ce2859bd53807be201893397b6ae406

Observation c04332f9-e844-4e41-bc75-1ac44ab1bb6d · outbound

This paper cites H3dnet: 3d object detection using hybrid ge- ometric primitives.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations H3dnet: 3d object detection using hybrid ge- ometric primitives

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.841065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.745353Z digest=sha256:69f8b2722bce0240f082434f90efd5ea64e37cba3cc312f138bc3b8f2b6015a7

Observation 4d5f1441-966f-4805-87a8-7869b4877861 · outbound

This paper cites Iou loss for 2d/3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Iou loss for 2d/3d object detection

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.833223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.749159Z digest=sha256:f18bec5afca2592cdec52205c20fb2c66e1cb29021ea5abd130eff88a337f07a

Observation 7b7e2875-1e17-47db-b47c-f5abd6ee60df · outbound

This paper cites Detecting twenty- thousand classes using image-level supervision.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations Detecting twenty- thousand classes using image-level supervision

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:09.825173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.751547Z digest=sha256:9f070fb4411b475d577314f6be567af90123b40193fbf014e8134e8dbee6b96e

Observation 8f19b685-158a-4d14-9c8f-ab7425337d27 · outbound

This paper cites V oxelnet: End-to-end learn- ing for point cloud based 3d object detection.

OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations V oxelnet: End-to-end learn- ing for point cloud based 3d object detection

Reference 66

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T15:19:09.817590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T15:19:09.753843Z digest=sha256:fb6cdc812ac9daed95dab030a9609102ec9a6089b2bfe7baeafcfea41c98bc7e

Pith citing papers

No inbound Pith citation observations are available.