Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:37:28.194043Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 11 inbound Pith citation observations for arXiv:2506.01853.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:37:28.194043Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:33:58.615197Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
100 of 108 outbound references displayed
External citation measurements
1
pith, observed 2026-08-05T02:28:24.338817Z
Observation ba8c87c7-db74-4930-87fd-38483b2ac911 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e67a45-0d95-49d2-ad57-f6878b600c9b · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98ae7833-6501-422f-982d-d66f178fedaf · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 431a6cff-ab11-49c0-a0af-cbcc2a26e144 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen2.5-VL Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ff1d3cc-2991-4aaa-8022-55e05ba1df37 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Demystifying MMD GANs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15a240ce-734a-4bff-a78c-99f5c6ffb750 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Piqa: Reasoning about physical common- sense in natural language
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4c94635-96c4-4813-9bcb-98c590cdd290 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ShapeNet: An Information-Rich 3D Model Repository
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df554099-ae9d-42df-8865-559aacaa6905 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointgpt: Auto-regressively generative pre-training from point clouds.Advances in Neural Information Processing Systems, 36: 29667–29679, 2023
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fac3dbdc-9738-4617-a4b5-69380ddb9b01 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Microdreamer: Zero- shot 3d generation in 20 seconds by score-based iterative reconstruction.arXiv e-prints, pages arXiv–2404, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d73be3a2-3ba7-4eed-98d7-a5312da0811e · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e4238e-3e26-437d-86a1-3efcac9a38d8 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshxl: Neural coordinate field for generative 3d foundation models.Advances in Neural Information Processing Systems, 37:97141–97166, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 896097ac-e20f-4d87-b711-5026c9c62f5f · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a6bc40-8c86-42a6-9dcf-d5c03422d69d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f24bea8-e00b-49f2-9e4e-4107916a9cfe · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sar3d: Autoregressive 3d object generation and understanding via multi-scale 3d vqvae
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcf52c9d-a488-486a-b3a3-e457db965fcf · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c88e12a0-259d-48eb-8b96-c7327300499a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d961ad31-b56e-46e5-b363-1a4cdeb8319d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Text-to-3d using gaussian splatting
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48d4bde8-b383-4712-ab79-1700c8d3854a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding V3D: Video Diffusion Models are Effective 3D Generators
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2572e5a1-cca0-42b4-a625-31633d5ce4c8 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ac3aefb-c295-4021-91ca-5abc2230011a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Training Verifiers to Solve Math Word Problems
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b60ef697-a219-4f90-bfa1-d5bc904e71e4 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Abo: Dataset and benchmarks for real-world 3d object understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 639067a0-0b5d-49b2-8d36-6e581760bd8c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813, 2023
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ad52bd0-0fca-4016-a5be-4de7f1d9c053 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54825719-4198-4040-90f9-a9b38724e13e · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4f9af82-9a08-4a1f-9e1a-9e4809b1ea4a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 103e5f92-6dfd-4379-9e05-f70539c4e620 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9147454a-076a-417b-9476-b1b865b8706c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Measuring Massive Multitask Language Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdbbf27a-5a48-497c-90c3-e5c75d8961ff · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fd4d6fc-20cc-4b7e-97b1-1f8de06e6372 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 942049f0-3b07-40fa-97f4-6a14fc2b7c8c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LRM: Large Reconstruction Model for Single Image to 3D
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fe629d8-23a1-4183-9b5d-0ee7fa08a603 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d-llm: Injecting the 3d world into large language models.Advances in Neural Information Processing Systems, 36:20482–20494, 2023
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b15196e8-a502-435d-90cc-8b359ff2eb5a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9076281c-26af-4d9f-8c9c-8c3da0f7ccf9 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4o System Card
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a76dd746-bd58-48b0-906a-ebe47ea4b7bf · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d gaussian splatting for real-time radiance field rendering.ACM Trans
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7d9d9cc-4934-4b58-a423-07e87fe83476 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3efcc19-5340-44bd-804b-4b8b5892457c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a9fc0b-18f2-4d97-92e7-b9b3b34eae23 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03f2517b-494e-4742-9d66-8ec00652d90c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Llama-vid: An image is worth 2 tokens in large language models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c7c7950-2d39-4754-8b51-66a2a9e8d4bf · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Magic3d: High-resolution text-to-3d content creation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13da5421-ef5a-463f-b0d1-d0baabc52e3b · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DeepSeek-V3 Technical Report
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74095d9c-1ffe-4bd6-b19c-4e50f8bf22b6 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2116498f-7835-4ec4-aaca-8038acba5261 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 257ffe6e-744b-4afd-8b31-b8d531ce6a18 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f9a76d7-4f34-4e0d-b007-b63ab93000e0 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding One-2- 3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.Advances in Neural Information Processing Systems, 36:22226–22246, 2023
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546aa401-5af1-4c68-8e73-6e704e7a80b6 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero-1-to-3: Zero-shot one image to 3d object
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4e66ff1-2c0d-4203-a7bb-7eeab6e9ddf9 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 797de06d-30b3-4924-9935-285baa29eb8d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Wonder3d: Single image to 3d using cross-domain diffusion
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 062289b7-18e4-485a-b8c9-7f645c9fd044 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Marching cubes: A high resolution 3d surface construction algorithm
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5bd8967-5b5b-4e3c-a510-0b481d94d934 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9f363de-9fd6-48ae-ac56-afed71cacb67 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106, 2021
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1ece4fd-c453-4518-829b-b590bc230090 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hierarchical Transformers Are More Efficient Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb870eb-21a7-4369-9200-2d15c9bd6605 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Barron, and Ben Mildenhall
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b290a26a-c611-4aeb-917e-a03990121824 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Shapellm: Universal 3d object understanding for embodied interaction
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3e9253d-e0d0-4a72-a5c2-ad447b622c28 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1af2ac5e-da91-4d85-917b-fa425a63a366 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Learning transferable visual models from natural language supervision
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3505aa5-dc7a-42e5-b561-7b8c755a4dfb · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreambooth3d: Subject-driven text-to- 3d generation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 730b779a-3d70-4221-b63b-758239a60fc9 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding High-resolution image synthesis with latent diffusion models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ba98690-c867-4fe3-85c5-c59f27da9e74 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SocialIQA: Commonsense Reasoning about Social Interactions
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e85a51f8-1b74-4af5-9364-79b437d50b68 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9b6bf24-129d-4c69-bcb3-d12693b51df2 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MVDream: Multi-view Diffusion for 3D Generation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d8f952e-a96c-4781-97be-45bc86f02bcc · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshgpt: Generating triangle meshes with decoder-only transformers
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 98441879-1280-47a4-bcb3-dd9ed5f42fd4 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ad398fcf-12eb-46a6-8c73-02d32654f82d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Using shape to categorize: Low-shot learning with an explicit shape bias
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7b342ea8-2d4e-45ea-856d-8beb961b26ca · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfa94b3-8baf-440e-86eb-08db26dc9238 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rethinking the inception architecture for computer vision
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69950d54-7dc1-499d-a15c-a6f8348e83e6 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab3c4712-7db3-4395-a871-b1b142448e63 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24096128-49eb-46d1-b74e-bc34832ef2c2 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 023147e7-e977-4526-9975-2b55d0234824 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0918483a-fa5d-4cf2-99d5-fe69eb2e92af · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ea2661-bbee-4ba1-977c-2e8fcb81a3c5 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Neural discrete representation learning.Advances in neural information processing systems, 30, 2017
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae417b46-ae3d-485a-8264-6912330d817c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4584e442-efc3-4f67-a60d-f7f128ddc773 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Score Jacobian Chaining: Lifting Pretrained 2D Diffusion Models for 3D Generation
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ac3ca65-a9ac-48c4-83a5-05eac76d239e · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc377646-8cf4-46df-998f-cb695723e7c8 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d35027-5f05-495c-8e8b-feb8f919aa34 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rodin: A generative model for sculpting 3d digital avatars using diffusion
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 24cdf760-9d93-4149-b3ce-c53b97e3e2bc · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Emu3: Next-Token Prediction is All You Need
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b2bb92d-a1ce-4c3a-9e50-3f5acc78d6ae · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5491dbc-1144-45c2-8e25-78dfb6672c7d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7c388385-c785-46f1-8863-7b6063936576 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0000ef5-a4f6-44b7-b7e4-d8ae783ebf64 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44bd630b-97db-4f4b-9c5d-e36a6ac0bb3c · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshLRM: Large Reconstruction Model for High-Quality Meshes
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07fba7ca-3dda-4c42-b7dc-25eeac45eda7 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Consistent123: Improve Consistency for One Image to 3D Object Synthesis
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55c8ad8b-06b0-453b-8a8e-43df60ec9c9a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PivotMesh: Generic 3D Mesh Generation via Pivot Vertices Guidance
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3263c51f-e5c3-4123-b0de-d29da82cab29 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Scaling Mesh Generation via Compressive Tokenization
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ef0d01b-adfe-4c8e-b14a-e045329109c9 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instructblip: Towards general-purpose vision-language models with instruction tuning [c]
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5b00c8d3-f7c8-4d21-b935-8441a953c2ea · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Unique3d: High-quality and efficient 3d mesh generation from a single image
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 42565f75-fbb4-4cd2-9e10-26813377ff33 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79cc8c9e-cd5c-4b99-8491-f56d2fcf006a · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Structured 3D Latents for Scalable and Versatile 3D Generation
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7bd91bf-fc85-464e-b05d-50badb484302 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0a569b7-a257-4d1a-9ce9-4b18ca190283 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f68bb29-47c1-4b6f-9b10-3749e0f69e84 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointllm: Empow- ering large language models to understand point clouds
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 248e081b-02e5-43b8-93d4-c0415496802f · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35c12590-3fb0-410a-8e52-c71e84238aa7 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7341f3c5-bcf7-41d6-8fb0-f87f3d62177f · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e7b8e8c-facb-4bb0-a876-ad901c364d41 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Stablenormal: Reducing diffusion variance for stable and sharp normal.ACM Transactions on Graphics (TOG), 2024
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 58f0e69d-0225-4bb6-b596-1ff0dadb5337 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hi3DGen: High-fidelity 3D Geometry Generation from Images via Normal Bridging
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aa9e4a7-b61b-45d7-9a7e-dcc7c9c80f0d · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreamreward: Text-to-3d generation with human preference
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 519efb95-2708-446d-8153-02e02ec65380 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 033307e0-27f1-49de-92b0-ea735b1e2f90 · outbound
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49058870-ee52-4be4-a8da-900f9056dd8e · inbound
Motus: A Unified Latent Action World Model ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6e979359-e9ca-4a99-8336-e3b74fc31a21 · inbound
CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1205b8f5-5fd6-45e7-9e5a-66869ba40c08 · inbound
LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b674a98d-3a17-4bfc-bbc8-a7776f48ac9e · inbound
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation be4efb64-eb67-4096-88c9-579956d8c3fc · inbound
EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f772d536-8266-4d3e-a899-6807867d5ea2 · inbound
PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 34de4d10-d593-4454-b3e4-34752fbae90a · inbound
GEM: Generative Supervision Helps Embodied Intelligence ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 286806b0-ee16-4238-944e-efafdf6ea434 · inbound
PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2b0a997f-e45c-4d49-a429-70802dad9b4d · inbound
ELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ee48ee27-38cd-4b2b-bd55-0b9d0bbe1f9d · inbound
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffd247c5-ee90-44c9-a736-bfad8d98e0fd · inbound
PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.