Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:48:15.371509Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 15 inbound Pith citation observations for arXiv:2412.02193.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:48:15.371509Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:38:09.713884Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T12:53:50.400526Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6d6ddcd9-9b63-45a6-85dd-07989b7eff82 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Lego-net: Learning regular rearrangements of ob- jects in rooms
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7dbb07f-49ec-48d8-a1f2-ea6eaffcefb6 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models House- gan++: Generative adversarial layout refinement network towards intelligent computational agent for professional archi- tects
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 51d7a2b4-2741-4856-befc-0f7b2269f093 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Atiss: Autoregres- sive transformers for indoor scene synthesis
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e29642a8-ecc7-4fca-ab68-285e83185fd0 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Layoutgpt: Compositional visual plan- ning and generation with large language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 23ce5c91-3468-467e-9c27-03cc0f999f95 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Holodeck: Language guided gen- eration of 3d embodied ai environments
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e10ca49f-b47f-4b15-b57c-353f0ca7dde2 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Controlroom3d: Room generation using semantic proxy rooms
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a4222826-c1cd-4b82-a1b5-6bb00f437a43 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models GALA3D: Towards Text-to-3D Complex Scene Generation via Layout-guided Generative Gaussian Splatting
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23be55f6-ed98-4e59-a814-f3d6829a55c0 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Compositional 3d scene generation using locally conditioned diffusion
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4bf646e7-10c9-475c-978b-25146c521a2e · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Disentangled 3D Scene Generation with Layout Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4431e7e0-7cdd-4423-a6f0-20d5e5831367 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Lay-A-Scene: Personalized 3D Object Arrangement Using Text-to-Image Priors
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b42060f2-8b4d-4e2d-8dd6-4860706157bb · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Any- home: Open-vocabulary generation of structured and textured 3d homes
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2932abca-4467-43b8-a6ba-b08badf07fe0 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 395a8679-064f-417b-8626-ab0ea59b5ff0 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models I-design: Personal- ized llm interior designer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8136846-15eb-4e25-9dde-ac0a732f0a55 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d29277ef-75ff-4214-a4f4-f443bc873c54 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models 3d-llm: Injecting the 3d world into large language models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c3d2d28-f487-4537-897c-3bcef10e31c9 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models 3D-GRAND: A Million-Scale Dataset for 3D-LLMs with Better Grounding and Less Hallucination
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c733c3f-09c1-494e-80a9-27acd6f49e68 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c09669-ef9d-45bd-86b4-c2f77c9c7239 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Ll3da: Visual interactive instruction tuning for omni-3d understand- ing reasoning and planning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5b0d31eb-8328-4b88-b4a8-6a0da23fdc75 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Gpt4point: A unified framework for point-language understanding and generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9711f590-b682-4a98-b062-490b481bb930 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Pointllm: Empowering large language models to understand point clouds
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 49a8f54d-1cfb-4ab0-b06e-dcdfd3087ceb · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Chat-scene: Bridging 3d scene and large language models with object identifiers
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8e343bae-1874-4a32-af99-75971b360f05 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Spatialvlm: Endowing vision-language models with spatial reasoning capabilities
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9ada32df-b4a2-4556-8cc0-ffa8e2d7a687 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 967c1500-952a-46ca-adf0-f0d3bb7486a7 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models SceneScript: Reconstructing Scenes With An Autoregressive Structured Language Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d0ad26c-a010-45e9-a9ae-d7b3fb6c9efa · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Agent3D-Zero: An Agent for Zero-shot 3D Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf573247-a95a-4d89-ac37-18f59cc6a25d · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e90c9a-ebdc-4051-8844-98c479e8168d · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Making large multimodal models understand arbitrary visual prompts
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b32eace2-2241-4156-86fb-7e8cccf9698e · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Scaffolding Coordinates to Promote Vision-Language Coordination in Large Multi-Modal Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20dbff0f-dfc9-401b-9b77-0939d86b169f · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models GPT-4 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16fffcb9-f2e8-4239-938d-8391968d01c7 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Self-consistency improves chain of thought reasoning in lan- guage models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8c31f780-a1b3-44e7-853b-f45d3f2912de · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Distance-iou loss: Faster and better learning for bounding box regression
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d6380aa2-385d-4c96-9914-232c4ce9d792 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Iou loss for 2d/3d object detection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ab50028a-55f3-4933-87fd-7fcfd6f0284b · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e1d099-1f22-4d88-8566-e30a3ea59b58 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Objaverse: A universe of annotated 3d objects
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7a177637-c3e3-4ea3-b83e-cf8df995c60d · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Gpt-4v (ision) is a human-aligned evaluator for text-to-3d generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cce554ee-f073-4002-b4da-95908ac86b30 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Learning transferable visual models from natural language supervi- sion
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0c85c859-f66e-42fb-840a-faffcd21986e · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Based on the assets, you should describe the general layout of the scene, the types of assets present, and any notable features
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 01ec0038-57dc-476d-a99c-436b2f49f40a · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models You should consider the functional, semantic, and geometric relationships between the assets
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3bc13331-ae60-4c8d-b6cd-5174fea093cd · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models You should explain the rationale behind each group and how the assets within each group are related to each other
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e28ee188-53b3-423c-bc43-563a25c7fc72 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models You should consider the significance of each group and the logical flow of the scene layout
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e8af38f1-ae42-4329-8ad0-6c7548387578 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models **Example:** Suppose you are examining a bedroom scene
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae6cd132-b61e-4fe6-870e-d02da1724fcd · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 234dce07-c011-46b0-832e-6b02a61c0193 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models The bedside table should be close to the bed for easy access
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 497c45bf-ecbb-4cc2-963c-fcb7a1575a90 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models The rational is that the bed is the central piece, the nightstand is next to the bed, and the lamp is on the nightstand
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 91ae464a-d310-48a2-bbb7-cfdb217b3253 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models They should be placed first to establish the sleeping area
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 303474b0-4a4c-4de8-b529-547fcca635d5 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models list": [ {
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7fc0ccc-1438-47ba-80c7-3a9a6b3bdb6e · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 67b3da58-7adc-41ed-9337-cf696fd92c3e · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9c617651-d7a3-4cc2-a35c-1117fa7d1286 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cd4739c3-f68d-4cf0-b111-1f10cc718f05 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Walls are also labeled with orientation arrows
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8adbdb27-7f26-461e-9aee-000406a4345c · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9d80e8ce-67e5-40d8-8b4e-38b5c7b44180 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Your task is to write a program that:
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a2546d91-7b76-450a-8d9e-dac5a9b5281f · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d0891190-bf54-41b4-9b5c-75344f43bdf4 · outbound
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models point towards
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fc7645bf-da88-479f-80c3-f5f96de379e8 · inbound
ScanEdit: Hierarchically-Guided Functional 3D Scan Editing LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eda677e5-5435-4db7-9840-a55a897ab99a · inbound
Scenethesis: A Language and Vision Agentic Framework for 3D Scene Generation LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86fe945b-ef27-42b0-aff0-8ed64a02c7fc · inbound
ReSpace: Text-Driven Autoregressive 3D Indoor Scene Synthesis and Editing LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 254ee44b-87fe-44e0-9831-d03ce9f537f4 · inbound
Handle-based Mesh Deformation Guided By Vision Language Model LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffaa7004-b558-46a1-8dd9-ec650a2b8439 · inbound
Video Perception Models for 3D Scene Synthesis LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56d960d6-db08-4b58-8524-19e5526477d4 · inbound
IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5629b80-0571-4fe0-ac2f-9bd72f80e7a7 · inbound
CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 189c25ce-7354-44c2-8df8-900b39caf607 · inbound
3D-Generalist: Self-Improving Vision-Language-Action Models for Crafting 3D Worlds LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3ffe8f1-c6df-4b91-9845-311658c049dc · inbound
Canvas3D: Empowering Precise Spatial Control for Image Generation with Constraints from a 3D Virtual Canvas LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0968b885-b33c-4917-b21f-d5e43dd16988 · inbound
SemLayoutDiff: Semantic Layout Generation with Diffusion Model for Indoor Scene Synthesis LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e96530ee-f361-49dc-81d0-8573a7ba476e · inbound
ProcFunc: Function-Oriented Abstractions for Procedural 3D Generation in Python LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4eb1d6a7-34ae-40f8-b2ad-58f42882f7d5 · inbound
STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 24a7f4a2-ea31-47fa-8bca-4bae8d2f9f32 · inbound
Perceive-then-Plan: Layout-as-Policy for Monocular 3D Scene Layout Estimation LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0f9fbde0-f7cd-43d1-afdb-1e1ed5342152 · inbound
Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 69399460-b668-4f77-a536-3f960aa7ea98 · inbound
SynCity 3000: Bootstrapping Scene-Scale 3D Diffusion LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.