Pith. sign in

Paper Citation Record · LEDGER

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling

As of 13 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 1 inbound Pith citation observation for arXiv:2411.19492.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19492 v2

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:11:49.782513Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:45:17.233257Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact5
  • verified fuzzy50
  • unresolved44
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c10ed52d-ecbc-492b-a1d2-bede40549c6a · outbound

This paper cites SATR: Zero-shot semantic segmentation of 3D shapes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SATR: Zero-shot semantic segmentation of 3D shapes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.493918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.493918Z digest=sha256:b8e19ceab09b0419517c919a6df29b7880d18e98bc760a99ea11adcf748f2dbe

Observation b0fc3940-b544-4ef7-b7e4-e7d03103bcb2 · outbound

This paper cites SceneCom- plete: Open-world 3D scene completion in complex real world environments for robot manipulation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneCom- plete: Open-world 3D scene completion in complex real world environments for robot manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.498994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.498994Z digest=sha256:599c9dab88e9f23cf18956c47688b9d415d3bfd3f158d1080aaa089be4bca185

Observation 1f4f9780-9735-48f0-b581-eaa8adef76f4 · outbound

This paper cites Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.502353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.502353Z digest=sha256:75f21bf5b076e21666de4f2af4a9b6febe52f51646adca6f69eb212b5c1e8f48

Observation 4569fd82-0eb5-4ccf-8408-1bcf034bcbb4 · outbound

This paper cites Scan2CAD: Learning CAD model alignment in RGB-D scans.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scan2CAD: Learning CAD model alignment in RGB-D scans

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.505334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.505334Z digest=sha256:65c775ac9e1ed203f2a86797f556e352b8e9f3f2832d4bb911903e5c60d418d7

Observation e2900653-dc28-4d3e-bff9-dcf7312b8ac8 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Emerg- ing properties in self-supervised vision transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.507805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.507805Z digest=sha256:dd7bb85a3394fa3d46b3e8912bdfb1b8edf52af643758d49f6a2812d307124d3

Observation 7da11930-16d1-4549-9077-4548384cca77 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ShapeNet: An Information-Rich 3D Model Repository

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.511213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.511213Z digest=sha256:6fe0303a5afc29f6e7fda14f451eca10aaba6c8fe413eaf24846da9b8c2c81b7

Observation 6fa30b9e-676e-459b-9dad-b82ad1ec18d5 · outbound

This paper cites CLIP2Scene: Towards label-efficient 3D scene un- derstanding by CLIP.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP2Scene: Towards label-efficient 3D scene un- derstanding by CLIP

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.514058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.514058Z digest=sha256:5ae595f2e8dc7005ef5125f7a9017977f665d93bc169a99b62379c3fcdfdc986

Observation 3461e84d-5a87-42d0-8f08-242c6964026d · outbound

This paper cites Single-view 3D scene reconstruc- tion with high-fidelity shape and texture.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Single-view 3D scene reconstruc- tion with high-fidelity shape and texture

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.516584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.516584Z digest=sha256:360ed8a8fe07f581c6cd8c8be846c3c4fff88eaa0cba5867c8478d296183a2bb

Observation d1a26e7b-6693-4eaf-a69e-a43ad1dc52d4 · outbound

This paper cites URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.518832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.518832Z digest=sha256:15a737ceecae29ac0532b5e1bdf2797dfccac74c5e8ab358ec8fcfe3a66eef8c

Observation c152bf9f-fd1c-4877-9db7-f4ef625b6bb6 · outbound

This paper cites ScanNet: Richly-annotated 3D reconstructions of indoor scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ScanNet: Richly-annotated 3D reconstructions of indoor scenes

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.521898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.521898Z digest=sha256:47a9456f20d791f7f66d6f0637d9b056a5da2c14e870d7145e935edd5b76aefc

Observation 6b6e1817-d7fd-48a5-805d-f8c3b396e0e0 · outbound

This paper cites Automated Creation of Digital Cousins for Robust Policy Learning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Automated Creation of Digital Cousins for Robust Policy Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.524207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.524207Z digest=sha256:99de98c4b75f5df829f664eeb33a0f822ea2e17912ef524284fd31ac3f9d8294

Observation 4b986f3b-c3e1-4058-93f5-6e6687f99228 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Objaverse: A universe of annotated 3d objects

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.527276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.527276Z digest=sha256:3d06a9e005e465e1c0a2a73d6d575c1375f255702a8b72f90819070285ef558a

Observation 90fe9c9d-96ce-4944-81a5-ffad12978faa · outbound

This paper cites SceneFun3D: Fine-grained functionality and affordance un- derstanding in 3D scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneFun3D: Fine-grained functionality and affordance un- derstanding in 3D scenes

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.530240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.530240Z digest=sha256:0ae3c363f9a1455dad8851c69252595321ff8a095af588f97ab413b3d7d0c9c1

Observation f034f4e7-ed66-4761-82f8-d719e5bb0389 · outbound

This paper cites PLA: Language-driven open- vocabulary 3D scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PLA: Language-driven open- vocabulary 3D scene understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.532957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.532957Z digest=sha256:d2d480f2af08e27016c895bb6eca48dc8f7321891540464d82e89d44c1b345b5

Observation 2ac572d8-4fd1-4b72-b3ac-7d1c6df8ec1c · outbound

This paper cites PanoContext-Former: Panoramic total scene understanding with a transformer.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PanoContext-Former: Panoramic total scene understanding with a transformer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.535892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.535892Z digest=sha256:afec7893aa859d94640890822590e0dba9981a809ad9298bc4ba188464ee2a49

Observation eaaf686a-693b-4f76-8ab8-5e89260c4f29 · outbound

This paper cites CLIP- Away: Harmonizing focused embeddings for removing ob- jects via diffusion models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP- Away: Harmonizing focused embeddings for removing ob- jects via diffusion models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.538456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.538456Z digest=sha256:7316faa12bcf29d0435486716d8812fcb1a36dafd5c2e13816bbae40f85a47b9

Observation a4b66764-a5ea-4d09-8c3f-6eb2e3eed339 · outbound

This paper cites Prob- ing the 3D awareness of visual foundation models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Prob- ing the 3D awareness of visual foundation models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.540729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.540729Z digest=sha256:9165ca054f37fb1786f7fb7d5dcacaf3c20d31604070faa27c01afe6ef4077eb

Observation 4307cfc7-8e60-4dc0-bb28-44d4669cff56 · outbound

This paper cites A density-based algorithm for discovering clusters in large spatial databases with noise.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling A density-based algorithm for discovering clusters in large spatial databases with noise

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.543346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.543346Z digest=sha256:5aef630549d41df8d6f0412e48646abb90eed2d6c1194eb632129e8a69dd3958

Observation 4b65af23-3c8f-4cbc-94a4-b502f79ca123 · outbound

This paper cites Fischler and Robert C.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Fischler and Robert C

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.546454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.546454Z digest=sha256:4dacdf952b5ad75a79864805107dfaf3a9c3f6aea84bd34d51f3b5f8f1c3c6db

Observation 36b5064f-c43d-45c7-b707-0bb01a285626 · outbound

This paper cites Example-based synthesis of 3d object arrangements.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Example-based synthesis of 3d object arrangements

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.549432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.549432Z digest=sha256:00c9b7613d213324861538b802a474c1ccc012d4d3712ce39cb2b8beb3dec056

Observation 7f63adc0-3367-483c-8376-984dc9be9b26 · outbound

This paper cites 3d-front: 3d furnished rooms with layouts and semantics.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling 3d-front: 3d furnished rooms with layouts and semantics

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.551727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.551727Z digest=sha256:2c9bdf53965f28ba1a128fc20b5e24ed66bfd83367ab873d4a40446cd0679f2f

Observation a4cc9940-54dc-4f87-85a2-32d1c84734fe · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.554635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.554635Z digest=sha256:2a2ef854a80af5dab8a88d3b2cc8930f22e39ef5a6527a8e0517caf16a1a5790

Observation 4cc51535-0c6e-4a92-87c1-4c6e61dea9bb · outbound

This paper cites Any- home: Open-vocabulary generation of structured and tex- tured 3d homes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Any- home: Open-vocabulary generation of structured and tex- tured 3d homes

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.558068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.558068Z digest=sha256:5a43121d16c6e3653744a6b19dd0d4d983037207f732a04c5989677c455a335d

Observation 11432c41-d44e-4321-9968-19ef40bb787c · outbound

This paper cites DiffCAD: Weakly-supervised probabilistic CAD model retrieval and alignment from an RGB image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DiffCAD: Weakly-supervised probabilistic CAD model retrieval and alignment from an RGB image

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.560446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.560446Z digest=sha256:e6dc0dd5b50cc4402466fdb38bea73badaa8e58096c03dd646f513ccf72d1d6f

Observation b99148ee-bb8a-4528-840f-8e326469b597 · outbound

This paper cites GraphDreamer: Compositional 3D scene synthesis from scene graphs.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GraphDreamer: Compositional 3D scene synthesis from scene graphs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.563920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.563920Z digest=sha256:e05a8b6aa723c5b6b60d46ce599e3cb84a146268b37541e52caa3cfb7bbacc43

Observation aa3bb3ee-95c7-442f-a61e-50183eddfc04 · outbound

This paper cites Zero-shot category-level object pose estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Zero-shot category-level object pose estimation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.568274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.568274Z digest=sha256:1a9037aacbcfcd1dfc30da58ed900f40a6100e0a709cb0b68882aa786314dd95

Observation 44c9488f-8f54-4e50-a6e5-e4f699ea6f0d · outbound

This paper cites ROCA: Robust CAD model retrieval and alignment from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ROCA: Robust CAD model retrieval and alignment from a single image

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.509037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.571479Z digest=sha256:a935205c44a80ed895c8a186173c259430a98fc0231799fe865048e4eb5f57c3

Observation 1389a93c-938b-429c-93de-ea4f091f3eb5 · outbound

This paper cites 3D-LLM: In- jecting the 3D world into large language models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling 3D-LLM: In- jecting the 3D world into large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.501591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.573680Z digest=sha256:deae2c869769772a576cab89fbd4744c5ba08a39da55c0d7597901e83bc32ad8

Observation eadedfa4-e170-4b51-94f0-c1486f70a6ed · outbound

This paper cites Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.576177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.576177Z digest=sha256:d031f8585473bdcfdeeebe4308232913c4a420001ea2d2710f3346b0e78fdc40

Observation 3914b068-dd94-433a-9877-e4b1564c8537 · outbound

This paper cites Aladdin: Zero-Shot Hallucination of Stylized 3D Assets from Abstract Scene Descriptions.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Aladdin: Zero-Shot Hallucination of Stylized 3D Assets from Abstract Scene Descriptions

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:50.011876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.579347Z digest=sha256:83db54aa7e7a30433fd7299aac95942203b2a66c955455af372cd645f65ea167

Observation b1682dc6-5c43-417d-963c-0c57922a5317 · outbound

This paper cites Holistic 3D scene parsing and re- construction from a single RGB image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Holistic 3D scene parsing and re- construction from a single RGB image

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.493482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.582343Z digest=sha256:38e13e5bd1dfd1d39a36373750b59e9bb3ef240a3f9accd9a7f46cdfb5f22d40

Observation f40c009a-a11c-4b74-b560-43cc7a141fc7 · outbound

This paper cites OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.584612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.584612Z digest=sha256:9997086a6567f05cf4baa363e4933e826604339df081b7943ca80b0ae208bfad

Observation e0d13309-f173-4a11-aa4c-20ffb32777b9 · outbound

This paper cites CenterSnap: Single-shot multi-object 3D shape reconstruction and categorical 6D pose and size estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CenterSnap: Single-shot multi-object 3D shape reconstruction and categorical 6D pose and size estimation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.485608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.587024Z digest=sha256:60f9987c3f0341a30d252080b264aeacc1f83b5c72759db6f5a6d2677f14d02d

Observation 4029bd80-e635-4220-9332-c665f271c800 · outbound

This paper cites an unresolved cited work.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.590378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.590378Z digest=sha256:0a096b47d4720540fa0ed73733299e938c31fc830f1585dccece7e56523b67a1

Observation c4df1580-0934-4c3e-ab63-e5cae16d210a · outbound

This paper cites ConceptFusion: Open-set Multimodal 3D Mapping.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ConceptFusion: Open-set Multimodal 3D Mapping

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.592608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.592608Z digest=sha256:0a8d2998b8d77c532e1f9dfe2c3b04c192b092448211c289abdaec0a69231154

Observation 8d8df2bf-ae13-4266-b896-c28c4e3f334a · outbound

This paper cites SceneVerse: Scaling 3D vision-language learning for grounded scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneVerse: Scaling 3D vision-language learning for grounded scene understanding

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.473781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.595695Z digest=sha256:0811b134470d46bb6774bb20d92180191d6cb42bb3e0fb5bc49f7e83842f3520

Observation 9fd86233-c5dd-4801-b59e-9f150ccc453d · outbound

This paper cites LERF: Language embed- ded radiance fields.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LERF: Language embed- ded radiance fields

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.466992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.598600Z digest=sha256:40efc35dcd213569f3c9e53b499dc08f6f49d4cb0571ab418c742bd29bb38bb0

Observation fbf60058-fe52-4163-9c27-40d3938ce0a2 · outbound

This paper cites Habitat synthetic scenes dataset (HSSD-200): An analysis of 3D scene scale and realism tradeoffs for objectgoal naviga- tion.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Habitat synthetic scenes dataset (HSSD-200): An analysis of 3D scene scale and realism tradeoffs for objectgoal naviga- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.460093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.600942Z digest=sha256:be3a6766550cdb56fcb114968671680e8b4bc25b3b0adb40d443fd8762e79ce1

Observation 78cb8e7f-801f-44dc-aa8b-676e4c882697 · outbound

This paper cites Segment Anything.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Segment Anything

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.604227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.604227Z digest=sha256:b6fd76b25ba0ebf79b5fbf4da84e910ab7d31cb0a60923335189777c573a6eb9

Observation 63ef7ff9-186c-42a7-95a2-83510ec31c31 · outbound

This paper cites Mask2CAD: 3D shape prediction by learning to seg- ment and retrieve.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Mask2CAD: 3D shape prediction by learning to seg- ment and retrieve

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.452832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.607531Z digest=sha256:34502edc0f305120f8534372b70fea3ae4c5a7b3fe86b430f9cc40be3e8e37bc

Observation b19deefe-7550-41d4-8e8a-3d8bba338e28 · outbound

This paper cites Patch2cad: Patchwise embedding learning for in-the- wild shape retrieval from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Patch2cad: Patchwise embedding learning for in-the- wild shape retrieval from a single image

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.445461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.609993Z digest=sha256:d85e653a97ef38c94d37772a23292af51945dc922c5b33940ee1533166c5ea71

Observation 49a26a17-6fdb-4f7a-9957-55c070ed88bc · outbound

This paper cites Langer, G.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Langer, G

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.438477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.613122Z digest=sha256:89c2295a4bae142677f6853882a820079817e65702748b86ff8f21ced1056ab6

Observation ad0699e8-db43-43c7-8977-55b4a505aef9 · outbound

This paper cites FastCAD: Real-Time CAD Retrieval and Alignment from Scans and Videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling FastCAD: Real-Time CAD Retrieval and Alignment from Scans and Videos

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.617182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.617182Z digest=sha256:eab3f1e685929315162930a253d935a60fb49b0d3c3e38336cc6a7ef7b07fee0

Observation 141f1074-31b5-42c3-ad5e-6632d1020d50 · outbound

This paper cites Duoduo CLIP: Efficient 3D Understanding with Multi-View Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Duoduo CLIP: Efficient 3D Understanding with Multi-View Images

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.619909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.619909Z digest=sha256:d9064c910291dd3e2cff5d4403fcb21cbc673a218a13027582301717451a9c30

Observation 51b90bbf-2d31-4901-abb9-0b501532ab37 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.622478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.622478Z digest=sha256:7e43a6edf4e685f824c995905d75136247508f305cda023badea7dc66293ad7c

Observation 08c3a02c-a967-425e-889e-8e411437f904 · outbound

This paper cites InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.625257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.625257Z digest=sha256:2e24da13540134c3c97c90426b6473c7a45b061ec329c68d5e258090922e94cc

Observation 33e30b14-43ea-407b-af39-8b59d8bfdcd7 · outbound

This paper cites Towards high-fidelity single-view holistic reconstruction of indoor scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Towards high-fidelity single-view holistic reconstruction of indoor scenes

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.430827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.628147Z digest=sha256:a2833a4fe10bb44f892fece85b3c0997d0ae88e6a5bbb2411b3766722c313e68

Observation 5afda80b-60e8-4cef-a612-0b847c5398af · outbound

This paper cites LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.953020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.630597Z digest=sha256:151ead302bbd2de79991e8acccb2f2408b1c44a94054c8d68871eec6698a3892

Observation 05e95c9d-3b13-421d-b51b-dfc76d29ae6b · outbound

This paper cites PartSLIP: Low-shot part segmentation for 3D point clouds via pretrained image- language models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PartSLIP: Low-shot part segmentation for 3D point clouds via pretrained image- language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.422596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.633616Z digest=sha256:51c6fb1ceb4cace43a164208a19d7a5b4a294a000c3d095aeca37501de02ed9c

Observation a4aa060b-a51c-4e72-ae46-90abf372a5c4 · outbound

This paper cites OpenShape: Scaling up 3D shape representation towards open-world understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenShape: Scaling up 3D shape representation towards open-world understanding

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.415948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.637708Z digest=sha256:e445281d2e4d33f63730d44e7b022b001ec8894cfc37da574e3a48f1d0daa2c2

Observation 899abf58-de0f-4be0-92c4-72564e622cab · outbound

This paper cites Open-vocabulary point-cloud object detection without 3D annotation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open-vocabulary point-cloud object detection without 3D annotation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.408637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.640419Z digest=sha256:53ff65cb11c2c32bbae9e726eb55fd1b5bbd9d484b378ce984621435913676c4

Observation 299a2f2d-ca6e-4ec6-a677-6b5b8c355a71 · outbound

This paper cites Vid2CAD: CAD model alignment using multi-view constraints from videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Vid2CAD: CAD model alignment using multi-view constraints from videos

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.400234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.642820Z digest=sha256:101372443cb552643dee7a7f19cb85d5afd6fd2cd5e4f3d7ab3fcd34f04b1fb3

Observation 5c630e24-16f8-4a04-8cc6-96b3964b67d9 · outbound

This paper cites Cad-estate: Large-scale cad model annota- tion in rgb videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Cad-estate: Large-scale cad model annota- tion in rgb videos

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.393160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.645585Z digest=sha256:f421d4cc0a256f0e9ff84b9b64a5d9dbff5a6ba21cade4e468a5b2a3534b37e9

Observation 01ecffa7-6669-4a6e-9d06-1f38217748c9 · outbound

This paper cites Scaling open-vocabulary object detection.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scaling open-vocabulary object detection

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.386021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.648858Z digest=sha256:7883ce01f4f26215e9bc0597a277ab1f9cb910bcfa89b84ce8025c4bd86adafc

Observation aef7d29d-6dd1-4f0a-ad72-4fc34ac6a4bb · outbound

This paper cites Open3DIS: Open-vocabulary 3D instance segmentation with 2D mask guidance.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open3DIS: Open-vocabulary 3D instance segmentation with 2D mask guidance

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.378614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.651350Z digest=sha256:0ce3769949c894f14e2ad6a6d87bdc841ae61a4f2c9999213213b789d67a52dc

Observation 2bf060e5-9f48-423b-aa8b-4477f6de6072 · outbound

This paper cites GigaPose: Fast and Robust Novel Ob- ject Pose Estimation via One Correspondence.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GigaPose: Fast and Robust Novel Ob- ject Pose Estimation via One Correspondence

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.369090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.653661Z digest=sha256:5569bc77d7937c2e15335ffc2f7f46a372db757dc35ca5ef8a338958b6328c4f

Observation dea198db-3ebb-431c-89e1-4664d7dd54ea · outbound

This paper cites Total3DUnderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Total3DUnderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.360211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.656294Z digest=sha256:5aefdacb0d1b577173632dbdcabc1fba88cdc50deac453aea64be9ddbcc96cf6

Observation 8e93a9eb-b3e8-4bb6-85ec-f0c01fb757b6 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DINOv2: Learning Robust Visual Features without Supervision

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.658751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.658751Z digest=sha256:4fb85db15e7e6d280821a256429e0ea79fa1e137bf2bb35b87b27c97b73588b1

Observation 13524e66-3cc2-4a6e-92c3-956d93cbd301 · outbound

This paper cites ATISS: Autore- gressive transformers for indoor scene synthesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ATISS: Autore- gressive transformers for indoor scene synthesis

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.351186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.661845Z digest=sha256:53ca21b609014112f39b9e3764509a4fbeac1aba32a8fee74e8a3946a3a3bb06

Observation 264b8f0a-3f84-4350-b2bf-25d1fbaa5973 · outbound

This paper cites OpenScene: 3D scene understanding with open vocabular- ies.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenScene: 3D scene understanding with open vocabular- ies

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.344077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.665002Z digest=sha256:8d4271d946a4c41111046b8f238a1e4c9c22a0c5204747b62db103f6f525964d

Observation d82d9222-8027-40e7-98db-82eb806064aa · outbound

This paper cites LangSplat: 3D language gaussian splat- ting.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LangSplat: 3D language gaussian splat- ting

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.336354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.668569Z digest=sha256:89f3118d212f2445783a5ba3b18f91061f70d3085c433665799d9f9a0aa7cb0d

Observation 8d9bafd8-7cf4-4c13-8d8a-7564cbd35969 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Learning transferable visual models from natural language supervi- sion

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.671433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.671433Z digest=sha256:819652a4124e4f93f234ff1e0dae476da83657fd908fb556bd3e4adc7b2282ad

Observation e3a9a5f5-90cc-4ae7-b497-2e2fe0b355e9 · outbound

This paper cites Fast and flex- ible indoor scene synthesis via deep convolutional genera- tive models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Fast and flex- ible indoor scene synthesis via deep convolutional genera- tive models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.325043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.674137Z digest=sha256:f33856d92f4ef5edd937a2422f336e09123dc0526cda97441a942e90fe0348cf

Observation 787cabb3-1af9-40d0-81f9-3dbcd1c2acb4 · outbound

This paper cites Hypersim: A photorealistic syn- thetic dataset for holistic indoor scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Hypersim: A photorealistic syn- thetic dataset for holistic indoor scene understanding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.318328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.677118Z digest=sha256:5f7061eb256a0ffeed2558645f32cef84242e878a21e2d890ebf3388b6a70fa9

Observation 880af085-46f0-4a31-b380-5f96591c318d · outbound

This paper cites Estimating generic 3D room structures from 2D annotations.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Estimating generic 3D room structures from 2D annotations

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.309562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.680491Z digest=sha256:f26ef91bb098a4dfcfcec41feae42efddfc863cf484328cd9c3b0f5ff9d3531e

Observation 2676c924-eef1-447c-8041-2becce7e76e2 · outbound

This paper cites Computational geometry.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Computational geometry

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.301546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.683425Z digest=sha256:6b1e3e47c2f2a4d5315b9047d69dbf507d2b908aa3e56d57efe2048348584902

Observation ae21a2f4-160a-4e51-9794-8c91025c928d · outbound

This paper cites PlaneRecTR: Uni- fied query learning for 3D plane recovery from a single view.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PlaneRecTR: Uni- fied query learning for 3D plane recovery from a single view

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.295397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.685824Z digest=sha256:e4b436a19b16d0b2ec863abf4119c0b0486ab87fc08e4dd72171acba063c23bd

Observation ff8537de-ce15-41b0-82eb-07afa53accbf · outbound

This paper cites General 3D room layout from a single view by render-and-compare.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling General 3D room layout from a single view by render-and-compare

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.287652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.688793Z digest=sha256:7e579b121ff942341650b6683469c528caa0b767149341ebc030d9f997039bfa

Observation 7d498c06-3dfd-46a6-b560-7a03cd10f95b · outbound

This paper cites Lempitsky.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lempitsky

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.280447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.691376Z digest=sha256:b8d18f847a69a545adaf1b7d1fd613e68421508342f3610fbfb2ec31605a4fc8

Observation c8d8126c-03d9-4792-a825-8a3a4f38bf44 · outbound

This paper cites Habitat 2.0: Training home assistants to rearrange their habitat.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Habitat 2.0: Training home assistants to rearrange their habitat

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.273410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.693691Z digest=sha256:50a6f0238a794ea869ca5f74df1c72eecc057d9644cccf6189012d2dfc2c7d1d

Observation 54f70c6a-f0bf-47ed-b740-f7e59808c83c · outbound

This paper cites OpenMask3D: Open-Vocabulary 3D Instance Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenMask3D: Open-Vocabulary 3D Instance Segmentation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.697493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.697493Z digest=sha256:f013a0afded431dfa2f695624cf2ae0b0ec539913704e77289b0b24d84a17686

Observation cabec119-1f50-43de-a3f9-688cb1002400 · outbound

This paper cites SceneMotifCoder: Example-driven Visual Program Learning for Generating 3D Object Arrangements.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneMotifCoder: Example-driven Visual Program Learning for Generating 3D Object Arrangements

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.700931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.700931Z digest=sha256:bee1037ccd68134c4dd277451162bf4da178d0055de929d89867e4c2191b3645

Observation 7d6a11eb-53e0-4c2a-a996-ecd22a6e70e4 · outbound

This paper cites DiffuScene: Denoising diffu- sion models for generative indoor scene synthesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DiffuScene: Denoising diffu- sion models for generative indoor scene synthesis

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.266035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.703907Z digest=sha256:ecd4da5f93be78f8f0f01e01c1bc3f7753768b0dc338385d32d46f1dd54d2975

Observation 41f804fa-f2cc-4271-812a-cb9dc1345bda · outbound

This paper cites Least-squares estimation of transforma- tion parameters between two point patterns.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Least-squares estimation of transforma- tion parameters between two point patterns

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.257927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.707342Z digest=sha256:92b79274f1caf5ccb90edb78897a043bbc3d5d90555826922a2212f112941375

Observation aa30b035-e05d-4931-b51e-827bac432aa7 · outbound

This paper cites Deep convolutional priors for indoor scene syn- thesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Deep convolutional priors for indoor scene syn- thesis

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.251260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.709763Z digest=sha256:d070445a2aaa82140c9f83b60f779f78f90fb1fcc62eb6aacdfa6db5f935a5f8

Observation 0815e5db-bc2b-483d-97d8-b42e5d68612d · outbound

This paper cites PlanIT: Planning and in- stantiating indoor scenes with relation graph and spatial prior networks.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PlanIT: Planning and in- stantiating indoor scenes with relation graph and spatial prior networks

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.244285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.712598Z digest=sha256:ff269082876e9212a8ffa3884c39bb902cabd77686964c8367a1de28a27e9097

Observation 3a365394-dbd9-407f-8cd0-cdf9e83bc041 · outbound

This paper cites Lift3D: Zero-shot lifting of any 2D vi- sion model to 3D.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lift3D: Zero-shot lifting of any 2D vi- sion model to 3D

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.237503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.715682Z digest=sha256:c5407e98fc9e1d331346686c6d23719f2db9b6ad5c607097217213b7d052e5dc

Observation a4d9314b-7d34-466e-b97b-048b3bb45935 · outbound

This paper cites SceneFormer: Indoor scene generation with transformers.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneFormer: Indoor scene generation with transformers

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.230954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.718055Z digest=sha256:eeae9ce4d9788f49bbb967b54772d53d7d45f19d97f43b454cf0b65b409f541f

Observation 5d969a9a-5799-4b47-ad3b-779e1785e78a · outbound

This paper cites Lego-net: Learning regular rearrangements of ob- jects in rooms.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lego-net: Learning regular rearrangements of ob- jects in rooms

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.224002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.720297Z digest=sha256:f83e51c88e70893b4d7715d5805bd684e40b608c4776c6b2278e3d648e0b94a7

Observation 85a496ba-cd1d-4b62-9b6f-182dd497fc87 · outbound

This paper cites R3ds: Reality-linked 3d scenes for panoramic scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling R3ds: Reality-linked 3d scenes for panoramic scene understanding

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.217254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.722620Z digest=sha256:c1168eeb7717d164dd2c86f50686626fbda25208b92c1ef1e9c55189a217a107

Observation a7760160-f15f-44c5-b27c-5c3f8fe3d692 · outbound

This paper cites Generalizing single-view 3D shape retrieval to occlu- sions and unseen objects.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Generalizing single-view 3D shape retrieval to occlu- sions and unseen objects

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.209991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.725620Z digest=sha256:88c4c746bc2e4f27301e0e3ded77d690c9b15bc8afd0a48ef481ae49735f00a6

Observation e18b82c3-98e4-4f11-89d0-3f068ae12108 · outbound

This paper cites ULIP-2: Towards scal- able multimodal pre-training for 3D understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ULIP-2: Towards scal- able multimodal pre-training for 3D understanding

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.201800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.728700Z digest=sha256:ea3e02b2fe1183ec2fd8e052e4e47850b6e99e5a3aa48492bfd923b8538e2fe1

Observation 02d3f13e-c7db-4f0a-82b0-1d96e5be249b · outbound

This paper cites Learning to reconstruct 3d non-cuboid room layout from a single rgb image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Learning to reconstruct 3d non-cuboid room layout from a single rgb image

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.194470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.731710Z digest=sha256:ae18d8714d8d67f6992ed91fc9b647dba011c2314021bbae4d638a620e326f30

Observation d06723d3-9339-40d2-ad98-66f40cf7dde5 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.734860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.734860Z digest=sha256:5dba8d4537bd0b5bc5acf4417fc5b2cbd1df36822b39225a7f7bad8cef857cb8

Observation 11ce0f39-aab4-4e95-a6de-e8952184a690 · outbound

This paper cites Depth Anything V2.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Depth Anything V2

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.737583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.737583Z digest=sha256:bc6199628e2e1d67a443c3bc552a84fe2ff8d0fc86ec2c3e49107ab986595a94

Observation dd72d8e8-af42-4cde-aa27-6e12c9c63ac3 · outbound

This paper cites ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images

Reference 86

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.916053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.740308Z digest=sha256:11b08755c3890900e54f8dd128e267157762321b3e193d74933d68c269f04a67

Observation ae042157-ac9d-4eac-a7dd-24e250697efa · outbound

This paper cites Holodeck: Language guided gen- eration of 3D embodied AI environments.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Holodeck: Language guided gen- eration of 3D embodied AI environments

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.186658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.743249Z digest=sha256:20094f15f85db9752261f9a55f5af6dae4508e32ef6e8ed1086cf481a3c8484f

Observation cf604c72-7298-4844-9d0f-897e80fac004 · outbound

This paper cites Multi-view Aggregation Network for Dichotomous Image Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Multi-view Aggregation Network for Dichotomous Image Segmentation

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.905489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.745506Z digest=sha256:daa608acd0828d180f826b7b568eaf2ee05521ea9c58f6152c3f9a10eabe2184

Observation d35df8f3-c6fb-426b-86db-10310aa4bad0 · outbound

This paper cites Inpaint Anything: Segment Anything Meets Image Inpainting.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Inpaint Anything: Segment Anything Meets Image Inpainting

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.748866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.748866Z digest=sha256:347e38884f3953064cbaa3460156c379ab4cd62e02c7826827e569a45d99c69b

Observation 096e0f4b-243d-4d72-a29f-042671238d5e · outbound

This paper cites Improving 2D feature representations by 3D-aware fine-tuning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Improving 2D feature representations by 3D-aware fine-tuning

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.178196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.751681Z digest=sha256:f814cc7933da74a840aa397ce8d524535d9b5e6446d46ef7504f4011fb41213b

Observation c9c5a416-f891-4c26-9f76-9ce0382ff57d · outbound

This paper cites DeepPanoCon- text: Panoramic 3D scene understanding with holistic scene context graph and relation-based optimization.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DeepPanoCon- text: Panoramic 3D scene understanding with holistic scene context graph and relation-based optimization

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.170172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.754700Z digest=sha256:4d573ad4555443c4b557b58a5d63386ba48e63fea52aeeaafd2ba75b361abda9

Observation 8bb5fafc-8cf4-4e2e-939d-5b266b049bb5 · outbound

This paper cites CLIP-FO3D: Learning free open-world 3D scene representations from 2D dense CLIP.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP-FO3D: Learning free open-world 3D scene representations from 2D dense CLIP

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.162667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.757458Z digest=sha256:edd39e28045276b4c921992a922e3ccd3d742def79811e591128d6433ec771b6

Observation 5a632b2e-c84f-44a5-b1b9-8c367523e5e8 · outbound

This paper cites Structured3D: A large photo-realistic dataset for structured 3D modeling.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Structured3D: A large photo-realistic dataset for structured 3D modeling

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.155290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.759672Z digest=sha256:5935c7c2cb5433ffee07c6a1254476daebd4bab1f771e079881f5ae2c6e88aee

Observation 8dc8a302-6b17-4eb6-a682-4dddcdc1ae71 · outbound

This paper cites Bilateral refer- 12 ence for high-resolution dichotomous image segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Bilateral refer- 12 ence for high-resolution dichotomous image segmentation

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.147406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.761875Z digest=sha256:90f1aab12875d67446dd0e6c8a5545c029f9983c366a9fa49f03d28112dce949

Observation ff514cac-978c-4aae-a9e8-ff68e3efb5e8 · outbound

This paper cites Zero-Shot Scene Reconstruction from Single Images with Deep Prior Assembly.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Zero-Shot Scene Reconstruction from Single Images with Deep Prior Assembly

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.764963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.764963Z digest=sha256:add6a0abb466dc93b15102a9a3ab5fefacc79774cd848b8d44cdc26c1ff2a948

Observation 601a6c67-5a14-4e0a-b27a-c02bae60cec9 · outbound

This paper cites Open3D: A Modern Library for 3D Data Processing.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open3D: A Modern Library for 3D Data Processing

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.769541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.769541Z digest=sha256:a01842411313f3cf5a10d01960507c74bdc447049d8b8a85d7154f1a0a011f8d

Observation 0326cd5a-3ce0-4e48-8e58-c24586f365e0 · outbound

This paper cites Point- CLIP v2: Prompting CLIP and GPT for powerful 3D open- world learning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Point- CLIP v2: Prompting CLIP and GPT for powerful 3D open- world learning

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.138459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.772565Z digest=sha256:35ddf5cfca7c1a9856272a5ca2cad0f2e88698619115262cc63adf280dfb4173

Observation 843593dc-c620-4229-a2c7-860359cbced8 · outbound

This paper cites GRS: Generating robotic simulation tasks from real-world images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GRS: Generating robotic simulation tasks from real-world images

Reference 98

Resolution
verified exact
raw_fallback, observed 2026-08-12T10:11:49.871719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.775705Z digest=sha256:281674b6440085ed100a8883e9b32bb06211a4b588bb356fd65258de83835fd4

Observation ae5943dc-7d99-4022-ab7e-5ce3bce352b7 · outbound

This paper cites gpt-4o-2024-08-06.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling gpt-4o-2024-08-06

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.128905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.778819Z digest=sha256:17c1c1f8ff2e0790596988857863c1b90578e368360c5cbe7f29580aa52c64f2

Observation 775dd22d-dda3-4b03-8033-b0e597ed2b6c · outbound

This paper cites a photo of CLASS.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling a photo of CLASS

Reference 100

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T10:11:50.120553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T10:11:49.782513Z digest=sha256:49c9c616c95ddb7390ca4757f5aade2375beb73af609063f9f9ac144b2e453c6

Pith citing papers

Observation 946354f0-d8de-47df-80b8-e47a1d45a800 · inbound

Advances in 4D Representation: Geometry, Motion, and Interaction cites this paper.

Advances in 4D Representation: Geometry, Motion, and Interaction Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling

Reference 299

Resolution
unresolved
no resolver link, observed 2026-08-04T08:45:17.233257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:45:17.233257Z digest=sha256:3bf5b5b564a5fb08a31ce6fc408a4129d7884662279804172212819a97c36aa9