Pith. sign in

Paper Citation Record · LEDGER

MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.09401.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.09401 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:29:55.209758Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:14:17.058826Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 995eef7d-b09c-4c69-b567-da4c74d59daf · inbound

EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding cites this paper.

EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T21:29:55.209758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:29:55.209758Z digest=sha256:aadf5a3e06325fe1528292a14841b37b4615fc0573c4c6697c4b9c339f5efb08

Observation 54e54abd-5c8a-41e1-a8aa-5dc31cbc0da5 · inbound

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding cites this paper.

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:32:38.694683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:32:38.694683Z digest=sha256:69b00ef67e2977a79c7175dcebb8b083d83b95abb2d34214d59a97deb8bc94ed

Observation b7d85442-4198-47da-8d4a-051b4a91c961 · inbound

RoboBrain 2.0 Technical Report cites this paper.

RoboBrain 2.0 Technical Report MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:47:29.800018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:47:29.800018Z digest=sha256:d74660315061ec57064671fe3e3a1d144363136f3bf7bf46e1afa70c21da95c5

Observation b5898b26-4b7f-4614-b97e-825b081c6a13 · inbound

Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models cites this paper.

Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:14:17.062781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T15:14:16.922877Z digest=sha256:c7fbf92492c603c1c326ff363e1d375ab0aab7ffe38e28bc1900edc69b747e09

Observation be15af40-6c2c-48c9-9f03-acbbf5123251 · inbound

VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs cites this paper.

VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T13:56:48.156476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:56:48.156476Z digest=sha256:35e1f09ef9b9047c7536ea22bbc027e5c50731c4451edcdcd7f980fc38192948

Observation 17f506ea-f715-4c0e-857e-5abb8e42b3ab · inbound

Data Pyramid for Embodied Manipulation: A Survey cites this paper.

Data Pyramid for Embodied Manipulation: A Survey MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations

Reference 264

Resolution
unresolved
no resolver link, observed 2026-07-31T06:18:55.870335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:18:55.870335Z digest=sha256:51990f0e704f90fd47faaef3abc9f1df0e99cfd144276871fd75bbdb4db7e747