Pith. sign in

Paper Citation Record · LEDGER

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

As of 20 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 5 inbound Pith citation observations for arXiv:2507.07781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07781 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:37:21.425460Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T13:51:30.008232Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:37:29.656433Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact5
  • verified fuzzy27
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9376cfc3-aed9-4cf8-83f9-69393bdf2cb9 · outbound

This paper cites Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.624177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.114837Z digest=sha256:3886ec5b01a462790720af88158e36895f14bc122a64a45b4380fa3e845372da

Observation d8b9af87-1faa-4dd8-b234-3bcb510b0aef · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.491641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.120880Z digest=sha256:9913e6f39e43354dda8041359961be079acd60544b3f1f7c6b973e188038c2f6

Observation 1b40e861-a928-4ed0-997f-f53026a46edb · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanqa: 3d question answering for spatial scene understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.126873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.126873Z digest=sha256:ae8a524fc8620da5f16a6ea0351d10572a2303e52f17c74cbeaa28c4a610c7de

Observation b7e840de-bc4e-4b8c-901a-a6f8dc94ff4b · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.133258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.133258Z digest=sha256:3fb445d4bbdd7a0332efb2b582006d72dae3cdd3b6bfda3ede1d0beea302f7ad

Observation 543dca29-9e36-4ef3-9fa9-ac2c3d0652e9 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.323272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.139122Z digest=sha256:d4d4b3992c4d74ad2e1414e1d84da111c32797dbb8068c82f1fbe2b3bee1b7e4

Observation 7aa50b6c-d6c4-4909-84da-bdd7a04867a1 · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:26.181776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.145147Z digest=sha256:366a50cab399e31426c7027b4fcbd6e9a7d0ea06d78bb6da741a0ac17e56ab66

Observation a49de642-bb8a-4a85-ab87-769d11004151 · outbound

This paper cites Towards label-free scene understanding by vision foundation models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Towards label-free scene understanding by vision foundation models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.042482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.151006Z digest=sha256:2457a3a0e1790a5f67a03ea1fc8f0068f30ba5a7d6d6892f11a45ff281287615

Observation bf533800-795b-499b-937b-7d9eb46bb329 · outbound

This paper cites Clip2scene: Towards label-efficient 3d scene understanding by clip.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Clip2scene: Towards label-efficient 3d scene understanding by clip

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.903944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.156005Z digest=sha256:aca619e39a0c3564366deeeb5775d8baca5006d70e1068bdb30ea3c4bd1b499e

Observation 4148b6cc-5b0e-4a17-aed6-75c5f3173d47 · outbound

This paper cites OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.160848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.160848Z digest=sha256:8a2804ce3f8da12a9ec241ff78ff564c085d2c3a6d7ae1e79f80e3f30cf271d9

Observation c6b1ba35-5ff8-42c0-bb3c-8e907b4ee5c9 · outbound

This paper cites Zero-shot point cloud segmentation by transferring geometric primitives.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Zero-shot point cloud segmentation by transferring geometric primitives

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.363226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.166231Z digest=sha256:36feae021d4e0ee9f9cda22a048d7bec3bf83ea54c38e45d9955443449fe897c

Observation 241bf559-752a-4323-ac7b-128ab4e7f507 · outbound

This paper cites Bridging language and geometric primitives for zero-shot point cloud segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Bridging language and geometric primitives for zero-shot point cloud segmentation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.784832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.171663Z digest=sha256:b81a8af652b7b8e6d1c8001b42fd5cc9683acd7c7ce71d6abb81bde7ab794873

Observation b452f0b8-dea8-4990-ad3f-496a7d70925b · outbound

This paper cites Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.684857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.176775Z digest=sha256:e555ee7b015ba3be38658cb1d3bc039e5e814641a85f943187c4d2e420757480

Observation b292152c-adad-48cc-b2e8-98a6979bbdb0 · outbound

This paper cites Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.181823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.181823Z digest=sha256:cc0027be6f9a17948169f84a3498694e4f7807f0b251a62ef26a08a17241db14

Observation 93c3ecf0-d3c8-466b-af6a-67e62075dabd · outbound

This paper cites Grounded 3D-LLM with Referent Tokens.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Grounded 3D-LLM with Referent Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.187570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.187570Z digest=sha256:1c7bbe8cdcafb0b4edd4afb7119be72cb7c6e1fe0d6ddc7aca7a63c12781ea0e

Observation 3d085221-29a1-4e6c-a68a-63ec232906cd · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:25.545431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.193512Z digest=sha256:c2ac3469e5c8fa91127750d59d40680274195d32d2dbb330eabf67795e5c96db

Observation b5a4291c-4900-4b9b-9bad-50c7e71bef18 · outbound

This paper cites Segpoint: Segment any point cloud via large language model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Segpoint: Segment any point cloud via large language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.369542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.199440Z digest=sha256:e8284bfdde9e3384f7fc3912e19fe443e8486bd2b4e14b3662fab8671a2b3220

Observation 1c0ae59c-f287-4a9d-b403-2c10bf660c0b · outbound

This paper cites 3d concept learning and reasoning from multi-view images.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d concept learning and reasoning from multi-view images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.223082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.205237Z digest=sha256:d1019156c04f5c6cbd1e59013c72bc2e5690b85c15e4354c71ad6f83cd925354

Observation 88a24fde-1d8c-443a-8169-e0f0c819eac2 · outbound

This paper cites 3d-llm: injecting the 3d world into large language models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-llm: injecting the 3d world into large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.076517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.211899Z digest=sha256:c9ef39daf1589d659900ac22c020ec81e8c7390baeb467972dbd8e6bec74045c

Observation 38fc496f-7a4c-4614-af2a-f2a6eef7d609 · outbound

This paper cites Chat-scene: Bridging 3d scene and large language models with object identifiers.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-scene: Bridging 3d scene and large language models with object identifiers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.930825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.217983Z digest=sha256:ae6421e855bb7bf7d2d9904177879bd5d113fd1f2da52a6526f5f9ff250cbc5c

Observation f7ee941a-0953-47fe-822d-78787c2cc680 · outbound

This paper cites An embodied generalist agent in 3d world.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes An embodied generalist agent in 3d world

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.808362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.224453Z digest=sha256:f2ee31fa0e3ceb4b97356922f87ec420ce88afca1a59c83a28a4f0f365935bda

Observation 66773a7b-6149-40a1-9c5a-b97c8fc6ddeb · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.231214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.231214Z digest=sha256:47879f81a9cc2d616f053cb3ff0f60741135abb7ad4f403107db4d67358f6669

Observation 7600b532-ee8d-4d1b-8f5d-8ae1bade0159 · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.674946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.237038Z digest=sha256:a45bd2ce0e9b23f20a668d7f0aa41547993b7ea8b41f20c24bb7a71ea4b06665

Observation 042d3286-fb21-4c1f-ab19-3971b06480f7 · outbound

This paper cites Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.046731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.243530Z digest=sha256:8a7c093770d76439024f446bb102d2b775e48324eb301128bbadf9136d14836d

Observation 2d6570b7-2cda-4a48-95c0-ecf1afc66378 · outbound

This paper cites Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.529985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.255021Z digest=sha256:84992040728cd920d8538316c7236918a489e2862f3dbb07759698856d8fb8a3

Observation d9ac1b62-ec57-42e0-8d96-4246995d0ff7 · outbound

This paper cites Dense object grounding in 3d scenes.ACM Multimedia, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Dense object grounding in 3d scenes.ACM Multimedia, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.383244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.261876Z digest=sha256:50cd7d3f6860585b46ece5aa8f8c75124785be7a9433b90e8a827dd60b19531f

Observation 4471163f-cb6b-4db4-89ab-5737e3b369ce · outbound

This paper cites Sceneverse: Scaling 3d vision-language learning for grounded scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sceneverse: Scaling 3d vision-language learning for grounded scene understanding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.255373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.267560Z digest=sha256:39dab6d87984047cd3c966ccba177e3296dc71fb7f4c3d55b9457be704429433

Observation c9d89ce4-7ef4-476e-a7ca-79ec368e6539 · outbound

This paper cites Multimodal 3D Reasoning Segmentation with Complex Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multimodal 3D Reasoning Segmentation with Complex Scenes

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.860973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.273496Z digest=sha256:c043509377559c5deb32d6f614dc8782c00d478170d5078c0cce518ab3a6cbaf

Observation 6c0c30db-6c45-4dbf-bf87-63ee7a7f99dd · outbound

This paper cites Intent3d: 3d object detection in rgb-d scans based on human intention.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Intent3d: 3d object detection in rgb-d scans based on human intention

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.091965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.280081Z digest=sha256:83cda51e9e9d4975187df397eb407ce1c025fe0263417e0a9344a686196e209a

Observation 83c48785-f887-4b67-be35-c9b7c3c962ab · outbound

This paper cites M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.955399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.288173Z digest=sha256:3acbf87683afe07ede549aa0f59b25b650d5d2723121e9e30ac5eda46eb147b2

Observation 362f372f-b1d3-4e5f-a3aa-0d1015d2618b · outbound

This paper cites 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.293763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.293763Z digest=sha256:243ca14aa9c96fd56d2a528326a39ab2bc50fb801cb7996397c13985db7bcbf0

Observation b948b5bd-2019-432c-820e-060b18784fe6 · outbound

This paper cites Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.821728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.299608Z digest=sha256:b4fbf890d4dc84b0fcaef1b044d17dce53cf995cd9055a7fe9e2282968384aef

Observation e141879a-7cc9-41f9-bcd1-4302cbbbcea4 · outbound

This paper cites See more and know more: Zero-shot point cloud segmentation via multi-modal visual data.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes See more and know more: Zero-shot point cloud segmentation via multi-modal visual data

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.645328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.307224Z digest=sha256:3e72a781cb4d1beeb8884202504537f00c5dc5e1db28b4b702126bc35d979b1d

Observation 346a79c3-8c57-4270-93cb-b5eba08d6163 · outbound

This paper cites 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.314108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.314108Z digest=sha256:d947639fa2738f8be33b36171e44fd3f0d7219b9376e26d3f7e001ea16f97b17

Observation 99404787-2c76-42ee-98e0-c7d6561e17b2 · outbound

This paper cites Sqa3d: Situated question answering in 3d scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sqa3d: Situated question answering in 3d scenes

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.319192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.319192Z digest=sha256:0db20603bfae3ff3d1b89b94323d6b7ecb00681ff4d7ff8f71f149af8ba5d6bb

Observation fd250474-66e0-431b-bd7d-d4b83e2ddde7 · outbound

This paper cites X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.395769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.324456Z digest=sha256:a3eff8b80f694500a7700d7f7722c32161ff2e79ded4266a438e798c04c8e973

Observation 181c2de8-7ad0-41ba-9f83-b112b78c4865 · outbound

This paper cites Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.282816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.330745Z digest=sha256:5001cfa02607cb936489aa104bf01ae67cccaf2b11b7ebeb3f8a821c1fd9e341

Observation 644d44f1-f96e-423b-b35c-a0403c355991 · outbound

This paper cites Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.182958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.337102Z digest=sha256:4a12000274aad1080a220fe0a9edae110daab4231f00b3942895bded8ffdd3f8

Observation 3770243f-952a-4a6c-94dd-05f355ed0d06 · outbound

This paper cites Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.344604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.344604Z digest=sha256:9f3672f2c2c6b74dfbb457c754476c6ad9d35d335ad2d842854b5a618facceb0

Observation b23ebd52-1067-4957-95b2-552b4f7e0a31 · outbound

This paper cites 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.052030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.352680Z digest=sha256:1b3ee5c0005e17ce5b897fc82409c62b392dd3b6a55026c93eeccfe5048bcbae

Observation eac10f64-5bb7-472e-a662-ba3269557217 · outbound

This paper cites Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.908862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.360014Z digest=sha256:8a9d34b1031082264db7a1982dad7e7b84941fdf23f77db44c745d3f7813489d

Observation c2c71b5d-d935-4927-bd90-1dc2685e2feb · outbound

This paper cites Fouhey, and Joyce Chai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fouhey, and Joyce Chai

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.814955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.367267Z digest=sha256:d3a7b993ac2356dd27f131b1d58b8a7f9ee8d9144820df86bbf1f46b978ea200

Observation e9854bb2-90b6-42a3-b153-64c907f4714c · outbound

This paper cites Scannet++: A high-fidelity dataset of 3d indoor scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scannet++: A high-fidelity dataset of 3d indoor scenes

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.375086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.375086Z digest=sha256:97495c677379cb7d1b37200e5574505bbb1f2f4db10946954ba7b0eb9f0dbdc2

Observation 30d36a3c-73ef-4ec8-8854-6d3fbf77c3c3 · outbound

This paper cites ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.594527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.381545Z digest=sha256:264f02a4f0aa2182c92ec6cab57cc8232fb54b2f0fb196ff09cb64a46b507add

Observation 6a899754-8dde-43f0-912c-07f2f9f8c2bb · outbound

This paper cites Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.542408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.389400Z digest=sha256:c04131df00d5744d4966d30ead769a819ea2813a8748d5bd5af8dbba14341fdb

Observation 4253dfeb-ddb6-4e97-9581-e8379f3a1fde · outbound

This paper cites VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.399482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.399482Z digest=sha256:a17cc65fe1ecd7bbfd0860503ffb3ae119a45bd7d86b7de008a494bf7061acfb

Observation 869ee334-947e-452a-a4cf-0c9bcf0a93f8 · outbound

This paper cites ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.413309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.413309Z digest=sha256:c3ef9158a13d8971cd1a481a4d0f4a0db5f9cc0310ed913f7f3a23dbc3d7449a

Observation 773e9acc-2f00-4620-961d-89669b1f0225 · outbound

This paper cites 3d-vista: Pre-trained transformer for 3d vision and text alignment.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-vista: Pre-trained transformer for 3d vision and text alignment

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.621098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T18:37:21.425460Z digest=sha256:d744288ffc1d6992bbbc2fcaeda4468b0e6943c42294b6ef3f84992abb6d0895

Pith citing papers

Observation 6a07e2e5-fa18-4df6-a88c-1e53ebdf9428 · inbound

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing cites this paper.

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:33:44.661273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T00:33:29.282349Z digest=sha256:a3dc69e02cde1ac6d7b60e66ae649b7232cb07c351d14e1bb8b07a48dea1c848

Observation f83cbea5-9f4e-44d1-bdd0-89ab94a8b873 · inbound

What if? Emulative Simulation with World Models for Situated Reasoning cites this paper.

What if? Emulative Simulation with World Models for Situated Reasoning SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T13:51:30.008232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:51:30.008232Z digest=sha256:5bc6e5faa5acbd2dd799b9dd2731e083535137b47271cd1112a2a230946f0548

Observation bba1b6d4-3609-4ea9-a515-fc025b6adf68 · inbound

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models cites this paper.

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:29.657958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T17:08:22.609095Z digest=sha256:04af22d4698680a05f9455c32ddee1190a4b6ed9cbfcc4099286667062238401

Observation e147e5dc-eb09-4b09-b902-16ed734977d4 · inbound

Holo-Captioning: Toward the Text Equivalent of 3D Scenes cites this paper.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:18f44ec7b55dee27f075b5112d6cac953317e924e8ec260a5e2ba47d74670a2e

Observation 01684f2e-885d-4528-b5d0-182573a6ee2e · inbound

G$^2$TAM: Geometry Grounded Track Anything Model cites this paper.

G$^2$TAM: Geometry Grounded Track Anything Model SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T23:56:52.009530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:56:52.009530Z digest=sha256:3558810f7e737eeadfd999ec110100e2779c2f5916033867d9c8b9eb0e4d5a7b