Pith. sign in

Paper Citation Record · LEDGER

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

As of 15 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 11 inbound Pith citation observations for arXiv:2506.01853.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01853 v1

Coverage vector

measured 100 of 108 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:37:28.194043Z

measured 111 of 111 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:33:58.615197Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 108 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved89
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ba8c87c7-db74-4930-87fd-38483b2ac911 · outbound

This paper cites GPT-4 Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.654734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.654734Z digest=sha256:2af0fd2debd369b7eb0ff3507fcee24644984efb8ce4493f30531ed0fdfb7236

Observation 26e67a45-0d95-49d2-ad57-f6878b600c9b · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.660428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.660428Z digest=sha256:f3c3be3b5eb34614c786005ac466cac631b010a1b063a895b3fc41abc8b8a408

Observation 98ae7833-6501-422f-982d-d66f178fedaf · outbound

This paper cites Qwen Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.664880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.664880Z digest=sha256:ded13fc04a9f329ba180776707ea1916bb3e61dc45a86cf93fe4bfa2df5ad560

Observation 431a6cff-ab11-49c0-a0af-cbcc2a26e144 · outbound

This paper cites Qwen2.5-VL Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.669361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.669361Z digest=sha256:4d62ee4b9d8ec18ebcba138f75dbe8b7c1d07595ce92faef97cb4691d1069c34

Observation 7ff1d3cc-2991-4aaa-8022-55e05ba1df37 · outbound

This paper cites Demystifying MMD GANs.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Demystifying MMD GANs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.673540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.673540Z digest=sha256:a110ed79b050ea42d67892debe3f865eb14ab9bfdcb0af0258702aa613e10db7

Observation 15a240ce-734a-4bff-a78c-99f5c6ffb750 · outbound

This paper cites Piqa: Reasoning about physical common- sense in natural language.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Piqa: Reasoning about physical common- sense in natural language

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.678469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.678469Z digest=sha256:5c0542e810e7ba3ebe0fd8c74b5d26d1a2ddeb2549256a9d8dd7e5714836f814

Observation e4c94635-96c4-4813-9bcb-98c590cdd290 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ShapeNet: An Information-Rich 3D Model Repository

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.683058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.683058Z digest=sha256:3a84f7e4f3545042690ab183d557be0e13bbbd0936e22be2562f39ae7d045cf9

Observation df554099-ae9d-42df-8865-559aacaa6905 · outbound

This paper cites Pointgpt: Auto-regressively generative pre-training from point clouds.Advances in Neural Information Processing Systems, 36: 29667–29679, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointgpt: Auto-regressively generative pre-training from point clouds.Advances in Neural Information Processing Systems, 36: 29667–29679, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.687548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.687548Z digest=sha256:2db38c755adfe161d66d37bfac04faee20156875071cdc7d0c2e2dae71082f5d

Observation fac3dbdc-9738-4617-a4b5-69380ddb9b01 · outbound

This paper cites Microdreamer: Zero- shot 3d generation in 20 seconds by score-based iterative reconstruction.arXiv e-prints, pages arXiv–2404, 2024.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Microdreamer: Zero- shot 3d generation in 20 seconds by score-based iterative reconstruction.arXiv e-prints, pages arXiv–2404, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.691522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.691522Z digest=sha256:a9abd18f9f55e591bbf1aa30b89b8265882684c7c9ca0c284c02c43ffb8491c1

Observation d73be3a2-3ba7-4eed-98d7-a5312da0811e · outbound

This paper cites Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.696648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.696648Z digest=sha256:48fa378b0a84c29eaaeb272c79b8ab900e029cf069a12c3fc9514aa0ff03a2be

Observation 82e4238e-3e26-437d-86a1-3efcac9a38d8 · outbound

This paper cites Meshxl: Neural coordinate field for generative 3d foundation models.Advances in Neural Information Processing Systems, 37:97141–97166, 2025.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshxl: Neural coordinate field for generative 3d foundation models.Advances in Neural Information Processing Systems, 37:97141–97166, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.701108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.701108Z digest=sha256:27b28cd114e537a8b59809f52e3a9aa9ee6c6384fa150588d59d4d34f6ff9768

Observation 896097ac-e20f-4d87-b711-5026c9c62f5f · outbound

This paper cites MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.705114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.705114Z digest=sha256:fc5cf41eb1a2bb8f5cc5d1eeb96f8ac8f081fefe46fe64435a21aa17e47b850a

Observation 44a6bc40-8c86-42a6-9dcf-d5c03422d69d · outbound

This paper cites MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.709245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.709245Z digest=sha256:3f634965da34312d0386a9e9a7e3facf4b7144a4306800a21bf3581e2d45173f

Observation 0f24bea8-e00b-49f2-9e4e-4107916a9cfe · outbound

This paper cites Sar3d: Autoregressive 3d object generation and understanding via multi-scale 3d vqvae.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sar3d: Autoregressive 3d object generation and understanding via multi-scale 3d vqvae

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.713269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.713269Z digest=sha256:0edbde74b26521b009e026790b9a1461a491b3cde13cc123438426274e90f750

Observation fcf52c9d-a488-486a-b3a3-e457db965fcf · outbound

This paper cites 3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.717242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.717242Z digest=sha256:d0db3af83021b5297ecbc1ed3ca8d1eb37842beb993eb7ed5b72c33ca4888265

Observation c88e12a0-259d-48eb-8b96-c7327300499a · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.721752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.721752Z digest=sha256:a09ad047b1f23f12287658ed397dbe3112e96526989483841e6c1e6cda63e9e2

Observation d961ad31-b56e-46e5-b363-1a4cdeb8319d · outbound

This paper cites Text-to-3d using gaussian splatting.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Text-to-3d using gaussian splatting

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.726222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.726222Z digest=sha256:0493af8918fa652403546b59fea8b8ccf56861dfbbbd55eb20431f52395657a2

Observation 48d4bde8-b383-4712-ab79-1700c8d3854a · outbound

This paper cites V3D: Video Diffusion Models are Effective 3D Generators.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding V3D: Video Diffusion Models are Effective 3D Generators

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.730398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.730398Z digest=sha256:61dc0936908d1055491e88df1a5dba60b925afe61de59fd9106757fed4691a5c

Observation 2572e5a1-cca0-42b4-a625-31633d5ce4c8 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.734517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.734517Z digest=sha256:3d92c52eb41f474c01db91018c172ab43d487ef181b1fe6f3a738e9f4af57bba

Observation 9ac3aefb-c295-4021-91ca-5abc2230011a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.738665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.738665Z digest=sha256:d8d1a3fa4ac141c1df0e408af54e1ecab75908eefb5ee8ba10a7b46add09285f

Observation b60ef697-a219-4f90-bfa1-d5bc904e71e4 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Abo: Dataset and benchmarks for real-world 3d object understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.742712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.742712Z digest=sha256:d749776f7b71e551a5ee54b53eae8c57d3bbbe00e3de11ba528946bc82f8083f

Observation 639067a0-0b5d-49b2-8d36-6e581760bd8c · outbound

This paper cites Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.750433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.750433Z digest=sha256:cb2a00e70a68d8c5f781d9f887d9836bf2adfd1c255d738e94c7cfeae7398f23

Observation 3ad52bd0-0fca-4016-a5be-4de7f1d9c053 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.755217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.755217Z digest=sha256:1e34ef76153e0306a105c4860214efdf2cec05c1e13f73fdcd6f4639d9a11063

Observation 54825719-4198-4040-90f9-a9b38724e13e · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.760327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.760327Z digest=sha256:5224235004f1b61a1ed6b67347685600b9aad59ce41c945cf26b8cdb82d4f74e

Observation e4f9af82-9a08-4a1f-9e1a-9e4809b1ea4a · outbound

This paper cites FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.764649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.764649Z digest=sha256:8477610009bbcd699dfe59c5948a1db8f35da30f17e64a726ed4404e19b21f66

Observation 103e5f92-6dfd-4379-9e05-f70539c4e620 · outbound

This paper cites Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.769279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.769279Z digest=sha256:4a7ac54041fdd3338f33e1f0ad2dd6e68b30a7ba6d38e43de2ebc8e6072674f1

Observation 9147454a-076a-417b-9476-b1b865b8706c · outbound

This paper cites Measuring Massive Multitask Language Understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Measuring Massive Multitask Language Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.773382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.773382Z digest=sha256:bc5a7b2d359c1fde9427b424152b5025358391e4a42cf8f828ea4461efd8703c

Observation cdbbf27a-5a48-497c-90c3-e5c75d8961ff · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.777736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.777736Z digest=sha256:8ea0a35a88cc82ef40e8795436e6e82fc7b6d5f0c78ec698b4e4e54bd375eb91

Observation 6fd4d6fc-20cc-4b7e-97b1-1f8de06e6372 · outbound

This paper cites Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.781209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.781209Z digest=sha256:392fb0283d8370b2031fe9aef37080c0dd5bea5bda06072f1ec0aa70571231db

Observation 942049f0-3b07-40fa-97f4-6a14fc2b7c8c · outbound

This paper cites LRM: Large Reconstruction Model for Single Image to 3D.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LRM: Large Reconstruction Model for Single Image to 3D

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.784805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.784805Z digest=sha256:025f05c66788d3b30a64537daa86372f2cfeec508a3b5d305d1cdcebdc514e74

Observation 0fe629d8-23a1-4183-9b5d-0ee7fa08a603 · outbound

This paper cites 3d-llm: Injecting the 3d world into large language models.Advances in Neural Information Processing Systems, 36:20482–20494, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d-llm: Injecting the 3d world into large language models.Advances in Neural Information Processing Systems, 36:20482–20494, 2023

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.789046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.789046Z digest=sha256:554cc8e64379788cf87908715a72f57140e3df99b21225eb7e5cb105d372bbdd

Observation b15196e8-a502-435d-90cc-8b359ff2eb5a · outbound

This paper cites SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.792792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.792792Z digest=sha256:f7e4fa23f8233558e74a0bb76c03e49b6f4591cca6e20e732f890398e78fecbc

Observation 9076281c-26af-4d9f-8c9c-8c3da0f7ccf9 · outbound

This paper cites GPT-4o System Card.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4o System Card

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.800204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.800204Z digest=sha256:b66b582fb5303a50976be7d1a46580ad953768d18ed780847f6d331eec163667

Observation a76dd746-bd58-48b0-906a-ebe47ea4b7bf · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Trans.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d gaussian splatting for real-time radiance field rendering.ACM Trans

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.803830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.803830Z digest=sha256:e2c2caf649f9dd29accdf18b705eb04ef5f7d650eab6dd290298e0f7d77026df

Observation a7d9d9cc-4934-4b58-a423-07e87fe83476 · outbound

This paper cites Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.807634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.807634Z digest=sha256:6310d8c141779edb68e9258af659c927bb95d005d3df4e4239353dbf89b9c6d9

Observation b3efcc19-5340-44bd-804b-4b8b5892457c · outbound

This paper cites SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.811750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.811750Z digest=sha256:d866f785e6e9156304c9293a87a56559f91b2b652738cd70c705c812c6ea7fdb

Observation c0a9fc0b-18f2-4d97-92e7-b9b3b34eae23 · outbound

This paper cites CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.815679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.815679Z digest=sha256:b60c1db9374680d2bfe8dd92fb953bf1fa57a62b544fbd68cc31af2ec4eb0323

Observation 03f2517b-494e-4742-9d66-8ec00652d90c · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Llama-vid: An image is worth 2 tokens in large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.819869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.819869Z digest=sha256:5d4516be3d838da745a68c339a421546d323afad0c99047aaeab097487a6cb5a

Observation 8c7c7950-2d39-4754-8b51-66a2a9e8d4bf · outbound

This paper cites Magic3d: High-resolution text-to-3d content creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Magic3d: High-resolution text-to-3d content creation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.823750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.823750Z digest=sha256:6139402ba6f894d9fdcefb382454ceb3310896b5af9b60f68d525a75efb71db0

Observation 13da5421-ef5a-463f-b0d1-d0baabc52e3b · outbound

This paper cites DeepSeek-V3 Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DeepSeek-V3 Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.827900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.827900Z digest=sha256:f93b3a4f7177c7732d3b0b85b7f6eae693a356f4745f31f0721948546fa3ec5e

Observation 74095d9c-1ffe-4bd6-b19c-4e50f8bf22b6 · outbound

This paper cites ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.831955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.831955Z digest=sha256:d1550538db2c9ba5b76035726fc8cd2ef6af11498442cf40b1c75d73bb87a6a7

Observation 2116498f-7835-4ec4-aaca-8038acba5261 · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.835988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.835988Z digest=sha256:58766df6ee1b2d53d288122d3f18a342774c32c23d10f76c285053d53644a890

Observation 257ffe6e-744b-4afd-8b31-b8d531ce6a18 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.840516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.840516Z digest=sha256:4cab628c5f04938b37e2c49149e6d6a2f203d020c5903165d82659575e2ac652

Observation 3f9a76d7-4f34-4e0d-b007-b63ab93000e0 · outbound

This paper cites One-2- 3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.Advances in Neural Information Processing Systems, 36:22226–22246, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding One-2- 3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.Advances in Neural Information Processing Systems, 36:22226–22246, 2023

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.844476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.844476Z digest=sha256:3e777acad48d997c4cffeed42e42e71aa77724661dbb8a356bc6ad153b21390c

Observation 546aa401-5af1-4c68-8e73-6e704e7a80b6 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero-1-to-3: Zero-shot one image to 3d object

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.849128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.849128Z digest=sha256:34f61a241db201870cb27aec4e3bbea5362092f8fb687b2375a7accd52e78336

Observation b4e66ff1-2c0d-4203-a7bb-7eeab6e9ddf9 · outbound

This paper cites SyncDreamer: Generating Multiview-consistent Images from a Single-view Image.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.854089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.854089Z digest=sha256:171b6154063d3dc8cd968fb3f3396719119d791074ee040a062b7293d4093947

Observation 797de06d-30b3-4924-9935-285baa29eb8d · outbound

This paper cites Wonder3d: Single image to 3d using cross-domain diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Wonder3d: Single image to 3d using cross-domain diffusion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.858378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.858378Z digest=sha256:48fcafeb725e8f912ccd20724210bf1f1b46e84e08dcb832dd51beb6a5c37275

Observation 062289b7-18e4-485a-b8c9-7f645c9fd044 · outbound

This paper cites Marching cubes: A high resolution 3d surface construction algorithm.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Marching cubes: A high resolution 3d surface construction algorithm

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.862454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.862454Z digest=sha256:73072d3e7037b92424191c4228661a22b32e63e108de5a319d10d83f37f112a4

Observation c5bd8967-5b5b-4e3c-a510-0b481d94d934 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.866814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.866814Z digest=sha256:de4b6b5fe0cd566658d65ee15580b2ba672c4e21ab8542267dc04132e8510c16

Observation b9f363de-9fd6-48ae-ac56-afed71cacb67 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106, 2021.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106, 2021

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.871119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.871119Z digest=sha256:25d7d9d0ad808a8aba22503b4f1d12579e6297006e1ceb786f8466668e4ce233

Observation c1ece4fd-c453-4518-829b-b590bc230090 · outbound

This paper cites Hierarchical Transformers Are More Efficient Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hierarchical Transformers Are More Efficient Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.875673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.875673Z digest=sha256:4252d257f751865f4e4e0446f6f59b706f6bbb99620aa2370600a26a9cc19409

Observation ffb870eb-21a7-4369-9200-2d15c9bd6605 · outbound

This paper cites Barron, and Ben Mildenhall.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Barron, and Ben Mildenhall

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.879675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.879675Z digest=sha256:0be2dd4244c1a3382e21bb7b141e9714539d379a1c7ba852a36a3fac92aaa9d1

Observation b290a26a-c611-4aeb-917e-a03990121824 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Shapellm: Universal 3d object understanding for embodied interaction

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.883351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.883351Z digest=sha256:8075bbef4ea0818f4b651416e9ec82ce9e5dee832bef59d16c19200f6bf4a205

Observation a3e9253d-e0d0-4a72-a5c2-ad447b622c28 · outbound

This paper cites Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.887417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.887417Z digest=sha256:0775730da363001391d944e558de1fc32ef51be920acd846c02d5d1e242957c9

Observation 1af2ac5e-da91-4d85-917b-fa425a63a366 · outbound

This paper cites Learning transferable visual models from natural language supervision.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Learning transferable visual models from natural language supervision

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.891434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.891434Z digest=sha256:bbc70a81fb6fdac43b161e0fdc6c3d6d3364836be60960ab9ea080e6f4eef94f

Observation e3505aa5-dc7a-42e5-b561-7b8c755a4dfb · outbound

This paper cites Dreambooth3d: Subject-driven text-to- 3d generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreambooth3d: Subject-driven text-to- 3d generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.895233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.895233Z digest=sha256:1b6dd6f71eb7bffb72251dcbcaa238972e2f08df953a7b01e793113b36e404a1

Observation 730b779a-3d70-4221-b63b-758239a60fc9 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding High-resolution image synthesis with latent diffusion models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.898979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.898979Z digest=sha256:53892c23abda4f208e58515a108d835a94282781889135e2197132d8236dc9eb

Observation 3ba98690-c867-4fe3-85c5-c59f27da9e74 · outbound

This paper cites SocialIQA: Commonsense Reasoning about Social Interactions.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SocialIQA: Commonsense Reasoning about Social Interactions

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.902878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.902878Z digest=sha256:49f27da17e1af2b73828abcd7f1b14198549772d06deb31fe8ec1c1a18cd6230

Observation e85a51f8-1b74-4af5-9364-79b437d50b68 · outbound

This paper cites Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.906927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.906927Z digest=sha256:05eef4e322e41a58014e8403474229e9f2d24f6ab6dd8da81dc533347c3eb00e

Observation e9b6bf24-129d-4c69-bcb3-d12693b51df2 · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MVDream: Multi-view Diffusion for 3D Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.910859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.910859Z digest=sha256:49f05baceb5cffd8826c9f4d0fa77b8b686dec08f07545a7bb7629d0aa206a3d

Observation 8d8f952e-a96c-4781-97be-45bc86f02bcc · outbound

This paper cites Meshgpt: Generating triangle meshes with decoder-only transformers.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshgpt: Generating triangle meshes with decoder-only transformers

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.325646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.914925Z digest=sha256:4457ffa6fec7fbdf84305def2538a50ee5d4aec9debc13e1c486799fd8adca96

Observation 98441879-1280-47a4-bcb3-dd9ed5f42fd4 · outbound

This paper cites Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.310511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.919555Z digest=sha256:192e553ef1d891a33d46ce9fef4fd72b892c07d8a91e478194259b8c34ed80ab

Observation ad398fcf-12eb-46a6-8c73-02d32654f82d · outbound

This paper cites Using shape to categorize: Low-shot learning with an explicit shape bias.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Using shape to categorize: Low-shot learning with an explicit shape bias

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.294237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.923782Z digest=sha256:d5574638406473eef683bb1094e46cf8f857459c8db0b3c1c7386867644e9400

Observation 7b342ea8-2d4e-45ea-856d-8beb961b26ca · outbound

This paper cites DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.927781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.927781Z digest=sha256:bef1eef1c644cf64b9399fefc6628865241f7bc510312b1e886fafa2ddb39bf7

Observation 0bfa94b3-8baf-440e-86eb-08db26dc9238 · outbound

This paper cites Rethinking the inception architecture for computer vision.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rethinking the inception architecture for computer vision

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.931926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.931926Z digest=sha256:a31e567e16bb472514a20c065f4d6cf528c183c15f1c900f27321c5cc31fdc8d

Observation 69950d54-7dc1-499d-a15c-a6f8348e83e6 · outbound

This paper cites DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.935654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.935654Z digest=sha256:99b405cec8e0107e2250a13a4f93d2872c6d741188aafd68296d3bd13d2ba707

Observation ab3c4712-7db3-4395-a871-b1b142448e63 · outbound

This paper cites LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.939831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.939831Z digest=sha256:89f1b95c476e49f204f87544e950f56d5dbed20af1a0d51ba83aca5d298a5105

Observation 24096128-49eb-46d1-b74e-bc34832ef2c2 · outbound

This paper cites EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.943894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.943894Z digest=sha256:88a8136fcd8e62e99fa1807ea68f30c67447c324364309a0bdb1d3bb979e0fe3

Observation 023147e7-e977-4526-9975-2b55d0234824 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.947973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.947973Z digest=sha256:4a39dae2918c1f838e60cdba9c00deaf151bce3e0f8823dfdd283d8247485ae2

Observation 0918483a-fa5d-4cf2-99d5-fe69eb2e92af · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA: Open and Efficient Foundation Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.952820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.952820Z digest=sha256:d2ca801324b292ca4426676fa47c720a9ec82c0a0fea8c2a3c1db3a1723f112f

Observation 59ea2661-bbee-4ba1-977c-2e8fcb81a3c5 · outbound

This paper cites Neural discrete representation learning.Advances in neural information processing systems, 30, 2017.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Neural discrete representation learning.Advances in neural information processing systems, 30, 2017

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.957223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.957223Z digest=sha256:4e9412a3feee8132f2bd3d262afdb7efc224924a97537d83039278a6b09a6044

Observation ae417b46-ae3d-485a-8264-6912330d817c · outbound

This paper cites Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.961196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.961196Z digest=sha256:8edadf3857443a53dfc4d422c6b742fbc35c5033efb39c37eb05c975e6f7bba1

Observation 4584e442-efc3-4f67-a60d-f7f128ddc773 · outbound

This paper cites Score Jacobian Chaining: Lifting Pretrained 2D Diffusion Models for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Score Jacobian Chaining: Lifting Pretrained 2D Diffusion Models for 3D Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.965143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.965143Z digest=sha256:523066d448304e775b69c776ee994bf0be8a5c48a84fadfb2d206b2ad0bcb1cd

Observation 9ac3ca65-a9ac-48c4-83a5-05eac76d239e · outbound

This paper cites ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.969205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.969205Z digest=sha256:dfdcd56475655354f6a117e500a2fc43bcfe541656e8374a19459b12beada6c8

Observation cc377646-8cf4-46df-998f-cb695723e7c8 · outbound

This paper cites PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.973250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.973250Z digest=sha256:7204445274e61ab99de432a40d764b10ba726e1a0c39cfe31074655c96b3a142

Observation f0d35027-5f05-495c-8e8b-feb8f919aa34 · outbound

This paper cites Rodin: A generative model for sculpting 3d digital avatars using diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rodin: A generative model for sculpting 3d digital avatars using diffusion

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.254652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.977321Z digest=sha256:0da837f392ed47c44cdc9f23ff4f934d8655d5e1b3839c13e3e62a9cb406cfd7

Observation 24cdf760-9d93-4149-b3ce-c53b97e3e2bc · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Emu3: Next-Token Prediction is All You Need

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.981864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.981864Z digest=sha256:5ea84dd40b930f2e89d4e820a46879db56e30b89e6f137b07595c642b9d1a194

Observation 1b2bb92d-a1ce-4c3a-9e50-3f5acc78d6ae · outbound

This paper cites AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.099457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.099457Z digest=sha256:2601b62cb9024e6b896c2027748f8877e56faa2ef21e69640e165289d7e64bd0

Observation f5491dbc-1144-45c2-8e25-78dfb6672c7d · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.241422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.104174Z digest=sha256:710cb4351fca096fe572c66e324a5fd05531dc51e6cb7a2259d072fb4c6d9cd7

Observation 7c388385-c785-46f1-8863-7b6063936576 · outbound

This paper cites LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.108498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.108498Z digest=sha256:badab216fe564db51c437d78b8984f151e166f6b251bf0e67fb12e3cb46daf0f

Observation b0000ef5-a4f6-44b7-b7e4-d8ae783ebf64 · outbound

This paper cites CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.112744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.112744Z digest=sha256:5c43ae8631783ac9b5a473d40ea853f3260cb27d582585b334d60eb79ef30290

Observation 44bd630b-97db-4f4b-9c5d-e36a6ac0bb3c · outbound

This paper cites MeshLRM: Large Reconstruction Model for High-Quality Meshes.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshLRM: Large Reconstruction Model for High-Quality Meshes

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.117281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.117281Z digest=sha256:3e7c5d4553c5089dda8cce24ce396124fa435c3e2ca59c40c555e0758e23de5b

Observation 07fba7ca-3dda-4c42-b7dc-25eeac45eda7 · outbound

This paper cites Consistent123: Improve Consistency for One Image to 3D Object Synthesis.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Consistent123: Improve Consistency for One Image to 3D Object Synthesis

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.121774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.121774Z digest=sha256:00d21ca73e5674e1b9775fc27d1fb0c8fccd12ed4db2d3f182b83d037d7aee22

Observation 55c8ad8b-06b0-453b-8a8e-43df60ec9c9a · outbound

This paper cites PivotMesh: Generic 3D Mesh Generation via Pivot Vertices Guidance.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PivotMesh: Generic 3D Mesh Generation via Pivot Vertices Guidance

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.126162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.126162Z digest=sha256:8a6e8406680051aa16ffc991c58b8305c1703207f450686c938266e504a122ab

Observation 3263c51f-e5c3-4123-b0de-d29da82cab29 · outbound

This paper cites Scaling Mesh Generation via Compressive Tokenization.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Scaling Mesh Generation via Compressive Tokenization

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.130354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.130354Z digest=sha256:d1ef06e143942889b234f2e0c7441a6e096ce233ecd44c28bbd77bf4882c46e8

Observation 5ef0d01b-adfe-4c8e-b14a-e045329109c9 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning [c].

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instructblip: Towards general-purpose vision-language models with instruction tuning [c]

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.226115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.134585Z digest=sha256:500917ad0c7acb93fdb3a704fede81d63ab85bb31740a866eebfdda39b395ad9

Observation 5b00c8d3-f7c8-4d21-b935-8441a953c2ea · outbound

This paper cites Unique3d: High-quality and efficient 3d mesh generation from a single image.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Unique3d: High-quality and efficient 3d mesh generation from a single image

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.210607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.138377Z digest=sha256:30aae2afdb4ca67e6e4d7495d30f51112b35bc3e428942a7beb3487341f6a8dc

Observation 42565f75-fbb4-4cd2-9e10-26813377ff33 · outbound

This paper cites Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.142230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.142230Z digest=sha256:e8f9a0a80532b7b8c4fa2838ce3f05e0c8aed811a316fad625b9889627212ae6

Observation 79cc8c9e-cd5c-4b99-8491-f56d2fcf006a · outbound

This paper cites Structured 3D Latents for Scalable and Versatile 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Structured 3D Latents for Scalable and Versatile 3D Generation

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.146343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.146343Z digest=sha256:f105ab5fd51665e1fb8d4f9947c85bb013dc652e908c5cf6fe04b5b25351ebec

Observation d7bd91bf-fc85-464e-b05d-50badb484302 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.150610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.150610Z digest=sha256:d2c195beea5969d70d007473de8254cfb5a552f3d3a512e95b756c28c6d09022

Observation a0a569b7-a257-4d1a-9ce9-4b18ca190283 · outbound

This paper cites InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.155632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.155632Z digest=sha256:490b59b71a411a1e8399ad420ef4616d6369839fd5c38a5a348a0d3487c73e34

Observation 4f68bb29-47c1-4b6f-9b10-3749e0f69e84 · outbound

This paper cites Pointllm: Empow- ering large language models to understand point clouds.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointllm: Empow- ering large language models to understand point clouds

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.195514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.159995Z digest=sha256:092f212bf2e6994ae012e80c7612faa31b4d1fec8e31bbc416db32979d6f88f7

Observation 248e081b-02e5-43b8-93d4-c0415496802f · outbound

This paper cites DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.163979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.163979Z digest=sha256:fab29348f742a6bb035b46c178c50d76e6b6f9083b8085b51f84835d3d4aaac5

Observation 35c12590-3fb0-410a-8e52-c71e84238aa7 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.181138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.168417Z digest=sha256:88899088f749a3bd40914bfe3d440bef9c63364ca5974bf1e5087784c0e96225

Observation 7341f3c5-bcf7-41d6-8fb0-f87f3d62177f · outbound

This paper cites Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.173079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.173079Z digest=sha256:e6d0dbcd7cbbcc95b8599698c9634a3f39fae91a6efa8cbc7c77d45a0fcde796

Observation 9e7b8e8c-facb-4bb0-a876-ad901c364d41 · outbound

This paper cites Stablenormal: Reducing diffusion variance for stable and sharp normal.ACM Transactions on Graphics (TOG), 2024.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Stablenormal: Reducing diffusion variance for stable and sharp normal.ACM Transactions on Graphics (TOG), 2024

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.166583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.178056Z digest=sha256:5f28b843ee77285b9014bf2c7f6904255b2c221612e55d01f4edeb8d1c6a3776

Observation 58f0e69d-0225-4bb6-b596-1ff0dadb5337 · outbound

This paper cites Hi3DGen: High-fidelity 3D Geometry Generation from Images via Normal Bridging.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hi3DGen: High-fidelity 3D Geometry Generation from Images via Normal Bridging

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.182157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.182157Z digest=sha256:2d34064ff49b6210c8e96dc2605ca55aa88a26fe6e8c571a846eff9ed402cf64

Observation 9aa9e4a7-b61b-45d7-9a7e-dcc7c9c80f0d · outbound

This paper cites Dreamreward: Text-to-3d generation with human preference.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreamreward: Text-to-3d generation with human preference

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.152175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.186239Z digest=sha256:5335de4a0882d70d2c4ec730eb6bde8077f7e5bc6b125fa70cc0babce2f41534

Observation 519efb95-2708-446d-8153-02e02ec65380 · outbound

This paper cites Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.190244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.190244Z digest=sha256:b347a3618f4e59f218f6a6b1735696e2b12423e0a92bcffaceb21e2a77fdf440

Observation 033307e0-27f1-49de-92b0-ea735b1e2f90 · outbound

This paper cites GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.194043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.194043Z digest=sha256:4896c843b9bb6d78473d3b5ab23b45b9ac1db70f9bb431c8168522b7beacbe96

Pith citing papers

Observation 49058870-ee52-4be4-a8da-900f9056dd8e · inbound

Motus: A Unified Latent Action World Model cites this paper.

Motus: A Unified Latent Action World Model ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:44:36.865476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T18:44:36.636455Z digest=sha256:59fec49823c301b123ac91051fa079953cf872a4f9c7446df2666efaec8b280d

Observation 6e979359-e9ca-4a99-8336-e3b74fc31a21 · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:50:14.832521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:f83b51e34680f2f1d8d2bc600813aa838825e0dd469aaeb98b787b54bf7f8c7e

Observation 1205b8f5-5fd6-45e7-9e5a-66869ba40c08 · inbound

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation cites this paper.

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 206

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:04.533037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T15:35:37.095627Z digest=sha256:90fc424653a59fbb097ad9c33eda583706dde65b4e80fb6714f893ba6863a09d

Observation b674a98d-3a17-4bfc-bbc8-a7776f48ac9e · inbound

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection cites this paper.

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:41:12.157141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-09T15:26:50.289626Z digest=sha256:d2695d1cf25e8deb646dc5dfa212deee08596ae1c03985e3208ce2101be2323f

Observation be4efb64-eb67-4096-88c9-579956d8c3fc · inbound

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers cites this paper.

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:27:47.816722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-19T21:26:36.707554Z digest=sha256:bb2b0e58e4a280c7f2808a977874928b5ba71de6b90db2fb57de9b49faa703bf

Observation f772d536-8266-4d3e-a899-6807867d5ea2 · inbound

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects cites this paper.

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:41:21.940604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T09:39:48.573379Z digest=sha256:efaa41be9daead94126ef947da0e35eaa848821af1469098e186cfa790f053b1

Observation 34de4d10-d593-4454-b3e4-34752fbae90a · inbound

GEM: Generative Supervision Helps Embodied Intelligence cites this paper.

GEM: Generative Supervision Helps Embodied Intelligence ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.911677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-29T13:38:27.263726Z digest=sha256:c3b81d06ec01db9aa2cfd6516124a7ebbd06e5022f6a2ee458ab5b1c30880b49

Observation 286806b0-ee16-4238-944e-efafdf6ea434 · inbound

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation cites this paper.

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:55:28.946818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-01T06:54:22.727115Z digest=sha256:b1993cda7450b36bbe11609b97200214c29de8788892d640769161c6b238d975

Observation 2b0a997f-e45c-4d49-a429-70802dad9b4d · inbound

ELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and Generation cites this paper.

ELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T01:44:26.235368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-08T01:40:40.690020Z digest=sha256:d353eccb31498dd8741de14ba3d6a1760a6be3ec5e10423b7ed04ffd80635856

Observation ee48ee27-38cd-4b2b-bd55-0b9d0bbe1f9d · inbound

Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing cites this paper.

Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:47.730082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:47.730082Z digest=sha256:757fe14c50b5c8e0a97f66333ffe6a22e9c59ced4f123179749023f421efd1a4

Observation ffd247c5-ee90-44c9-a736-bfad8d98e0fd · inbound

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets cites this paper.

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T00:33:58.615197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:33:58.615197Z digest=sha256:0d14dd21b4e7a94c58a0a1ddf6d36ab0b2863d087c621e559e4f90c1e90855cb