Pith. sign in

Paper Citation Record · LEDGER

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

As of 14 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 11 inbound Pith citation observations for arXiv:2506.01853.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01853 v1

Coverage vector

measured 100 of 108 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:37:28.194043Z

measured 111 of 111 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:33:58.615197Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 108 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved89
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ba8c87c7-db74-4930-87fd-38483b2ac911 · outbound

This paper cites GPT-4 Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.654734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.654734Z digest=sha256:8e14393c1eea86b484f21a1b267e73873f67ae9933f4b8d3c0334fbcc908b38b

Observation 26e67a45-0d95-49d2-ad57-f6878b600c9b · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.660428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.660428Z digest=sha256:af88ecef4ad1ddc0049382e928c77067fa09d0566af51f725b2b4e19800111e7

Observation 98ae7833-6501-422f-982d-d66f178fedaf · outbound

This paper cites Qwen Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.664880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.664880Z digest=sha256:2a039db2f2bc0c4cb62459d87e9529ef6fe651ade731263a84797a66c6d8716d

Observation 431a6cff-ab11-49c0-a0af-cbcc2a26e144 · outbound

This paper cites Qwen2.5-VL Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.669361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.669361Z digest=sha256:f401f530d4825fe784dc40876fe3734e181d1c8f89c1817ece7b61f0a116a8bf

Observation 7ff1d3cc-2991-4aaa-8022-55e05ba1df37 · outbound

This paper cites Demystifying MMD GANs.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Demystifying MMD GANs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.673540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.673540Z digest=sha256:4d1969bcc97b04a4f48077cbea17bf38a8ac3675afadc5c146558bec0204a84f

Observation 15a240ce-734a-4bff-a78c-99f5c6ffb750 · outbound

This paper cites Piqa: Reasoning about physical common- sense in natural language.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Piqa: Reasoning about physical common- sense in natural language

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.678469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.678469Z digest=sha256:64173db45cb4e28adba6fddd934f3c699e6e807532f86b9451fcdaeafe49f7c0

Observation e4c94635-96c4-4813-9bcb-98c590cdd290 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ShapeNet: An Information-Rich 3D Model Repository

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.683058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.683058Z digest=sha256:1cfa1c2ecd3331e770386bcfe31849c3768fa62556dfd9e6b56a816228dd4525

Observation df554099-ae9d-42df-8865-559aacaa6905 · outbound

This paper cites Pointgpt: Auto-regressively generative pre-training from point clouds.Advances in Neural Information Processing Systems, 36: 29667–29679, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointgpt: Auto-regressively generative pre-training from point clouds.Advances in Neural Information Processing Systems, 36: 29667–29679, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.687548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.687548Z digest=sha256:3420b600d5919ecf50a8d15da2d7322fa789f811e5abf98f1626b07c87c03070

Observation fac3dbdc-9738-4617-a4b5-69380ddb9b01 · outbound

This paper cites Microdreamer: Zero- shot 3d generation in 20 seconds by score-based iterative reconstruction.arXiv e-prints, pages arXiv–2404, 2024.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Microdreamer: Zero- shot 3d generation in 20 seconds by score-based iterative reconstruction.arXiv e-prints, pages arXiv–2404, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.691522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.691522Z digest=sha256:afff1fce568640011c09c0fc323f889bbed2cc9b01a118266c3f847f627e5c09

Observation d73be3a2-3ba7-4eed-98d7-a5312da0811e · outbound

This paper cites Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.696648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.696648Z digest=sha256:16143b60a7847241cfe4ce06298c8eee2801df3ba46001b25d9f51906ef62616

Observation 82e4238e-3e26-437d-86a1-3efcac9a38d8 · outbound

This paper cites Meshxl: Neural coordinate field for generative 3d foundation models.Advances in Neural Information Processing Systems, 37:97141–97166, 2025.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshxl: Neural coordinate field for generative 3d foundation models.Advances in Neural Information Processing Systems, 37:97141–97166, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.701108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.701108Z digest=sha256:49dd245c6f62c1d951108684ddf79058ad74f385639ac86e8fc29a5f379a6b59

Observation 896097ac-e20f-4d87-b711-5026c9c62f5f · outbound

This paper cites MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.705114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.705114Z digest=sha256:3351bfab4f7fc9784aa237d05a3b8d982740d1164ec004db889cb7ecfec8565a

Observation 44a6bc40-8c86-42a6-9dcf-d5c03422d69d · outbound

This paper cites MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.709245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.709245Z digest=sha256:438e980093d6e9c9a32ce97d1131727c18956ff75db625b1c1a1c716f87b8788

Observation 0f24bea8-e00b-49f2-9e4e-4107916a9cfe · outbound

This paper cites Sar3d: Autoregressive 3d object generation and understanding via multi-scale 3d vqvae.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sar3d: Autoregressive 3d object generation and understanding via multi-scale 3d vqvae

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.713269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.713269Z digest=sha256:95ab32415b6b9955f64df6dfdca8eecca432e3bf7a1630156fd425de0ad2480e

Observation fcf52c9d-a488-486a-b3a3-e457db965fcf · outbound

This paper cites 3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.717242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.717242Z digest=sha256:293beba167bb494a44724e135aad92ac4e558d8a7aa0fa7a7a9a3e77c04b451c

Observation c88e12a0-259d-48eb-8b96-c7327300499a · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.721752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.721752Z digest=sha256:d8b8ea60b4c23627978f3b15b41a1dc3763dec8fe2cf0a1116d69a9f2a79d06f

Observation d961ad31-b56e-46e5-b363-1a4cdeb8319d · outbound

This paper cites Text-to-3d using gaussian splatting.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Text-to-3d using gaussian splatting

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.726222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.726222Z digest=sha256:b6ee2213651c3efeff5f035b2fd291626329202489cbf9bb96d54c46f2765363

Observation 48d4bde8-b383-4712-ab79-1700c8d3854a · outbound

This paper cites V3D: Video Diffusion Models are Effective 3D Generators.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding V3D: Video Diffusion Models are Effective 3D Generators

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.730398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.730398Z digest=sha256:3accdd07e212347cdaec1f620d729e39f8347094b38cc2ee9ad4c614feea2bbb

Observation 2572e5a1-cca0-42b4-a625-31633d5ce4c8 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.734517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.734517Z digest=sha256:1d5f4e29086bcd9e5c59645b9710d996fb0d883aa746ab407a6420afceafc7b2

Observation 9ac3aefb-c295-4021-91ca-5abc2230011a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.738665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.738665Z digest=sha256:635a37577ba9a25c90cdb9dc812bcfe853d5306eb6321e29fe9243d9c57133ae

Observation b60ef697-a219-4f90-bfa1-d5bc904e71e4 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Abo: Dataset and benchmarks for real-world 3d object understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.742712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.742712Z digest=sha256:352b704b99850a7581e86bbb13479412759af0feb3a3079b2612401bfd32be0d

Observation 639067a0-0b5d-49b2-8d36-6e581760bd8c · outbound

This paper cites Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799–35813, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.750433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.750433Z digest=sha256:9aaf3a378c435cadf21aa83046d515089d62a27a41f5d8975d67bc9c8c4bb1d7

Observation 3ad52bd0-0fca-4016-a5be-4de7f1d9c053 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.755217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.755217Z digest=sha256:54a15f2c425399a6965b183b75bdebf4cbbf603f926bd995ddd85d2e426165d1

Observation 54825719-4198-4040-90f9-a9b38724e13e · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.760327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.760327Z digest=sha256:3c7a510800b404776e81ea36d38ff9f4dba57ffa70e1cea0a84e00dd02c967e1

Observation e4f9af82-9a08-4a1f-9e1a-9e4809b1ea4a · outbound

This paper cites FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.764649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.764649Z digest=sha256:4e21b972279d0808702c45ba108267eb241f7765f9912710160b3dc6a18ac5c7

Observation 103e5f92-6dfd-4379-9e05-f70539c4e620 · outbound

This paper cites Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.769279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.769279Z digest=sha256:82592ffaf51645d6cbbccb835fff61fea1ab4b6a879479abd96efd9c899c8d34

Observation 9147454a-076a-417b-9476-b1b865b8706c · outbound

This paper cites Measuring Massive Multitask Language Understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Measuring Massive Multitask Language Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.773382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.773382Z digest=sha256:cf003349b294a85eb105cd3fc9389cd641817be33f17db86456100d09dff4949

Observation cdbbf27a-5a48-497c-90c3-e5c75d8961ff · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.777736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.777736Z digest=sha256:6fb6cd222099403d604b654320cf2f2db242e3727e9e9c18f97cdb4162c45cbc

Observation 6fd4d6fc-20cc-4b7e-97b1-1f8de06e6372 · outbound

This paper cites Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.781209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.781209Z digest=sha256:81e3c9c332552ffc5bfefbaaed5ff3c1474e35d79810d213be82748a6824ab26

Observation 942049f0-3b07-40fa-97f4-6a14fc2b7c8c · outbound

This paper cites LRM: Large Reconstruction Model for Single Image to 3D.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LRM: Large Reconstruction Model for Single Image to 3D

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.784805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.784805Z digest=sha256:e26a0bfde68f0b04b8a257ffc19149d08108455222df345db592d4e6a0857695

Observation 0fe629d8-23a1-4183-9b5d-0ee7fa08a603 · outbound

This paper cites 3d-llm: Injecting the 3d world into large language models.Advances in Neural Information Processing Systems, 36:20482–20494, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d-llm: Injecting the 3d world into large language models.Advances in Neural Information Processing Systems, 36:20482–20494, 2023

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.789046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.789046Z digest=sha256:d2680da094abef1dfe33fa50f6167efc67d9b3b2677ae97439ae0ecdf9f9c951

Observation b15196e8-a502-435d-90cc-8b359ff2eb5a · outbound

This paper cites SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.792792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.792792Z digest=sha256:84b28ee617683e1592119764fc0109fd8dd186e6fb8563e4a57bdb0c03319347

Observation 9076281c-26af-4d9f-8c9c-8c3da0f7ccf9 · outbound

This paper cites GPT-4o System Card.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GPT-4o System Card

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.800204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.800204Z digest=sha256:3e7f1a70cddbc33f87ab744e3797d04c17d90034c28c5dc411cc819c0dd030e6

Observation a76dd746-bd58-48b0-906a-ebe47ea4b7bf · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Trans.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding 3d gaussian splatting for real-time radiance field rendering.ACM Trans

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.803830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.803830Z digest=sha256:7378f812fba2e47721077e2107877bf26c888f8f357297e8f5a9442de8a34bb4

Observation a7d9d9cc-4934-4b58-a423-07e87fe83476 · outbound

This paper cites Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.807634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.807634Z digest=sha256:1948de52bd1928e9f3bc55e173ec202f29d95e640e0eb95e4521370ad23af661

Observation b3efcc19-5340-44bd-804b-4b8b5892457c · outbound

This paper cites SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SweetDreamer: Aligning Geometric Priors in 2D Diffusion for Consistent Text-to-3D

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.811750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.811750Z digest=sha256:9402a87a9c500ee61771813babdac0ae30f92401bbcbbb17b5cb2c39c50f8cab

Observation c0a9fc0b-18f2-4d97-92e7-b9b3b34eae23 · outbound

This paper cites CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.815679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.815679Z digest=sha256:561da122900708743ab1f2c17d12379513f62a0e7665b8f29cbdb1ec86df3ca9

Observation 03f2517b-494e-4742-9d66-8ec00652d90c · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Llama-vid: An image is worth 2 tokens in large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.819869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.819869Z digest=sha256:e544c51e95f8c5c1a75b57eb16814ea5d6f966d488c073f82ec6f15826af79d3

Observation 8c7c7950-2d39-4754-8b51-66a2a9e8d4bf · outbound

This paper cites Magic3d: High-resolution text-to-3d content creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Magic3d: High-resolution text-to-3d content creation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.823750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.823750Z digest=sha256:53405edadc53538e7cea46eaa1c8beeca0e9f62a0d22961acb719234f9f832b4

Observation 13da5421-ef5a-463f-b0d1-d0baabc52e3b · outbound

This paper cites DeepSeek-V3 Technical Report.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DeepSeek-V3 Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.827900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.827900Z digest=sha256:f4608edd155dcccf86c914ced2b9648544d5e74fefa42b6091eec20cd5e2acb3

Observation 74095d9c-1ffe-4bd6-b19c-4e50f8bf22b6 · outbound

This paper cites ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.831955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.831955Z digest=sha256:f64d365b0f7bf5c86970ff829ae2098605ef066c066402a5032f5c4c4058a8fa

Observation 2116498f-7835-4ec4-aaca-8038acba5261 · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.835988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.835988Z digest=sha256:2d5aa222c683408450a508f6939563438895092b38e8a14273100ca2e668d0a2

Observation 257ffe6e-744b-4afd-8b31-b8d531ce6a18 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.840516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.840516Z digest=sha256:76bba84416c66df5a266dc11f3bd97aacc15805e8be45d38fee02b92d71b7621

Observation 3f9a76d7-4f34-4e0d-b007-b63ab93000e0 · outbound

This paper cites One-2- 3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.Advances in Neural Information Processing Systems, 36:22226–22246, 2023.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding One-2- 3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.Advances in Neural Information Processing Systems, 36:22226–22246, 2023

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.844476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.844476Z digest=sha256:6f36f4ba1d2b673d807ad5b48493b807cf24233eef76b68f42582ad102af00a5

Observation 546aa401-5af1-4c68-8e73-6e704e7a80b6 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero-1-to-3: Zero-shot one image to 3d object

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.849128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.849128Z digest=sha256:188a34b158c727566d308feed4ffb69253d24ac80977dbbc013f6be49ecd28ca

Observation b4e66ff1-2c0d-4203-a7bb-7eeab6e9ddf9 · outbound

This paper cites SyncDreamer: Generating Multiview-consistent Images from a Single-view Image.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.854089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.854089Z digest=sha256:779d3d3044fec1d3e93fbcbf6ac4593e4daf3bbe24fb8477da70a1659f964e02

Observation 797de06d-30b3-4924-9935-285baa29eb8d · outbound

This paper cites Wonder3d: Single image to 3d using cross-domain diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Wonder3d: Single image to 3d using cross-domain diffusion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.858378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.858378Z digest=sha256:cb78c9c4e576c42bd3324e4f644027a5cecbeb9e82f5db783b6069696030e493

Observation 062289b7-18e4-485a-b8c9-7f645c9fd044 · outbound

This paper cites Marching cubes: A high resolution 3d surface construction algorithm.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Marching cubes: A high resolution 3d surface construction algorithm

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.862454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.862454Z digest=sha256:2dc4d0259147f64475f7e10766052bfa3aad12fc6b0f263c2a4deb9486e93b9b

Observation c5bd8967-5b5b-4e3c-a510-0b481d94d934 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.866814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.866814Z digest=sha256:55be490b88526a088511828634dc1a26f59278aa853b0c4d89108c38d963f5c0

Observation b9f363de-9fd6-48ae-ac56-afed71cacb67 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106, 2021.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106, 2021

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.871119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.871119Z digest=sha256:0b16596def2d3b620609188eab0166540e0fbcf953f698f1e1a2bae881facb4c

Observation c1ece4fd-c453-4518-829b-b590bc230090 · outbound

This paper cites Hierarchical Transformers Are More Efficient Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hierarchical Transformers Are More Efficient Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.875673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.875673Z digest=sha256:fb341ece6f82d82a57eee3c652242d5d5bbadd46761f668acf8474540da85383

Observation ffb870eb-21a7-4369-9200-2d15c9bd6605 · outbound

This paper cites Barron, and Ben Mildenhall.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Barron, and Ben Mildenhall

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.879675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.879675Z digest=sha256:5c4f8f385ac08a93604a24a9c06a7d75fff5f883e8fc64307e0942ad76945315

Observation b290a26a-c611-4aeb-917e-a03990121824 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Shapellm: Universal 3d object understanding for embodied interaction

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.883351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.883351Z digest=sha256:2a152f4eb8659c752b076a7429b8a080b9a91a282207d921e79fc7f0cd950ed0

Observation a3e9253d-e0d0-4a72-a5c2-ad447b622c28 · outbound

This paper cites Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.887417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.887417Z digest=sha256:3df0355f558cec413f541551d298bb48e7c9580eaa1262fee13f9d10dbf169fe

Observation 1af2ac5e-da91-4d85-917b-fa425a63a366 · outbound

This paper cites Learning transferable visual models from natural language supervision.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Learning transferable visual models from natural language supervision

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.891434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.891434Z digest=sha256:758fe16312f113491438bb824af0f9e4dfab1348ccb11103a95df63e4a24a36d

Observation e3505aa5-dc7a-42e5-b561-7b8c755a4dfb · outbound

This paper cites Dreambooth3d: Subject-driven text-to- 3d generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreambooth3d: Subject-driven text-to- 3d generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.895233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.895233Z digest=sha256:23085c0353bd57050d0c829c2c3e37cd0bf2f9add791940558aef182d4710e7b

Observation 730b779a-3d70-4221-b63b-758239a60fc9 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding High-resolution image synthesis with latent diffusion models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.898979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.898979Z digest=sha256:583605ff83c9b40f4f7ca62208751f1168761be2bd60229df1a9d8fa0e63d201

Observation 3ba98690-c867-4fe3-85c5-c59f27da9e74 · outbound

This paper cites SocialIQA: Commonsense Reasoning about Social Interactions.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding SocialIQA: Commonsense Reasoning about Social Interactions

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.902878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.902878Z digest=sha256:4aac3767883d44d100bb956c0b5f964912c47bf0fa3c90e364a6ef64c9773abb

Observation e85a51f8-1b74-4af5-9364-79b437d50b68 · outbound

This paper cites Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.906927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.906927Z digest=sha256:24fbd2d831049113750b471310483818799ea8aab58be75e3dcbd2bfd0ef8f6f

Observation e9b6bf24-129d-4c69-bcb3-d12693b51df2 · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MVDream: Multi-view Diffusion for 3D Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.910859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.910859Z digest=sha256:b78d5b6db7a7dd40925e7f3cfa0317c8c69adfdc39103616e6324ca2b879b4b8

Observation 8d8f952e-a96c-4781-97be-45bc86f02bcc · outbound

This paper cites Meshgpt: Generating triangle meshes with decoder-only transformers.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meshgpt: Generating triangle meshes with decoder-only transformers

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.325646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.914925Z digest=sha256:48e9aea07739153c94cd5685c4eca628abbccb1a73b8f369f585db08edadd763

Observation 98441879-1280-47a4-bcb3-dd9ed5f42fd4 · outbound

This paper cites Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.310511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.919555Z digest=sha256:8a079aa055a1b7b5f8afd189b7c3f1b9981c971b5925eb9471216b42f1ba6017

Observation ad398fcf-12eb-46a6-8c73-02d32654f82d · outbound

This paper cites Using shape to categorize: Low-shot learning with an explicit shape bias.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Using shape to categorize: Low-shot learning with an explicit shape bias

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.294237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.923782Z digest=sha256:db165b02037ed074df846040e96132797271b0e526c0d8bbdecca4d81f2c69b5

Observation 7b342ea8-2d4e-45ea-856d-8beb961b26ca · outbound

This paper cites DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.927781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.927781Z digest=sha256:27244798eea49f03ed9abf0944a2ac7f089fb7b6527b7f841d4d8bd34618d9fb

Observation 0bfa94b3-8baf-440e-86eb-08db26dc9238 · outbound

This paper cites Rethinking the inception architecture for computer vision.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rethinking the inception architecture for computer vision

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.931926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.931926Z digest=sha256:b7eb980e07473dde978d3608cd3e93f6cd9c8ee366702df71a11a56a8b983f88

Observation 69950d54-7dc1-499d-a15c-a6f8348e83e6 · outbound

This paper cites DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.935654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.935654Z digest=sha256:9ce81ac4b555e6ddf66685a076244e46479745cd73eda2e845afdaa51584bf90

Observation ab3c4712-7db3-4395-a871-b1b142448e63 · outbound

This paper cites LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.939831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.939831Z digest=sha256:7c8e6a5d235e7981c30648dbf036c614b3776e0fc278e10f5b9099b7912d397f

Observation 24096128-49eb-46d1-b74e-bc34832ef2c2 · outbound

This paper cites EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.943894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.943894Z digest=sha256:bcc261eba6f34e152a1ebde3d64b56db9352f21c3a4135beb42d3b7bff6f6ba9

Observation 023147e7-e977-4526-9975-2b55d0234824 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.947973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.947973Z digest=sha256:2be55989daa55a0b0d5e88b7518c2301661ec1fb4c1204d0dee051d80565c8a4

Observation 0918483a-fa5d-4cf2-99d5-fe69eb2e92af · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA: Open and Efficient Foundation Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.952820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.952820Z digest=sha256:5d4ea38be08a907c8cfe1d04e36659438427be03fc55982c4bb70abae24137da

Observation 59ea2661-bbee-4ba1-977c-2e8fcb81a3c5 · outbound

This paper cites Neural discrete representation learning.Advances in neural information processing systems, 30, 2017.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Neural discrete representation learning.Advances in neural information processing systems, 30, 2017

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.957223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.957223Z digest=sha256:15655e84274f59489ae58504edffaddacc6322d718af621790f551c0c68784c1

Observation ae417b46-ae3d-485a-8264-6912330d817c · outbound

This paper cites Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.961196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.961196Z digest=sha256:bcbf207ca07e2bd6ce888ea74ded88a275d74caa9e7dca2123ac985fe8f123ee

Observation 4584e442-efc3-4f67-a60d-f7f128ddc773 · outbound

This paper cites Score Jacobian Chaining: Lifting Pretrained 2D Diffusion Models for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Score Jacobian Chaining: Lifting Pretrained 2D Diffusion Models for 3D Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.965143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.965143Z digest=sha256:506d04bdf10cecfdb8e593a70caae7e4bba95f8103fcc590982c549ff232b8e7

Observation 9ac3ca65-a9ac-48c4-83a5-05eac76d239e · outbound

This paper cites ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.969205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.969205Z digest=sha256:165a51be63e42580fbb8d783e22811d64405b6e9d3d143b095d6e5e34e543581

Observation cc377646-8cf4-46df-998f-cb695723e7c8 · outbound

This paper cites PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.973250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.973250Z digest=sha256:8df4d3b8eab169614115f3b461deb77aa9c08619f24128aac07600273f20cc1d

Observation f0d35027-5f05-495c-8e8b-feb8f919aa34 · outbound

This paper cites Rodin: A generative model for sculpting 3d digital avatars using diffusion.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Rodin: A generative model for sculpting 3d digital avatars using diffusion

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.254652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:27.977321Z digest=sha256:d2b463f6e5d8e58b1811719690a1ac8f04210920b9e29cbfb2c95cf4d76dce54

Observation 24cdf760-9d93-4149-b3ce-c53b97e3e2bc · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Emu3: Next-Token Prediction is All You Need

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:27.981864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:27.981864Z digest=sha256:71339827f50cc017eef729084eba93ee83af8c424c24792c7a21e34079b14ea2

Observation 1b2bb92d-a1ce-4c3a-9e50-3f5acc78d6ae · outbound

This paper cites AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.099457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.099457Z digest=sha256:07287097e576129cac28bc77cfeec0ea9cb7769b83fea374a704119e4f766858

Observation f5491dbc-1144-45c2-8e25-78dfb6672c7d · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.241422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.104174Z digest=sha256:ee3ebfdef295961ea379d48278bafed357801bbb3972cb5920f2cf37c6bb4b18

Observation 7c388385-c785-46f1-8863-7b6063936576 · outbound

This paper cites LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.108498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.108498Z digest=sha256:d55a8e8fc0834a1a93676f84eb36a6a9a536ab49dd6adede266dcda3effcd8d9

Observation b0000ef5-a4f6-44b7-b7e4-d8ae783ebf64 · outbound

This paper cites CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.112744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.112744Z digest=sha256:b3a9624761c60e4aa1cc0e215e939940d1e50662199fac3b142e4fd232e5c5b2

Observation 44bd630b-97db-4f4b-9c5d-e36a6ac0bb3c · outbound

This paper cites MeshLRM: Large Reconstruction Model for High-Quality Meshes.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding MeshLRM: Large Reconstruction Model for High-Quality Meshes

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.117281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.117281Z digest=sha256:50fe7f5845a924b9c6ba0d159bcbe30d5dfa09fc0adb69c90d8334bff6dd2bee

Observation 07fba7ca-3dda-4c42-b7dc-25eeac45eda7 · outbound

This paper cites Consistent123: Improve Consistency for One Image to 3D Object Synthesis.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Consistent123: Improve Consistency for One Image to 3D Object Synthesis

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.121774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.121774Z digest=sha256:e3b7279069d971fc4e5cd6a8d2dba26f449bdbe9eaea0089df118769804bc3ed

Observation 55c8ad8b-06b0-453b-8a8e-43df60ec9c9a · outbound

This paper cites PivotMesh: Generic 3D Mesh Generation via Pivot Vertices Guidance.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding PivotMesh: Generic 3D Mesh Generation via Pivot Vertices Guidance

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.126162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.126162Z digest=sha256:4657ee18abe07b625df895cd077f4e34b380c29c9f5bcff964af1168f2ae7dd8

Observation 3263c51f-e5c3-4123-b0de-d29da82cab29 · outbound

This paper cites Scaling Mesh Generation via Compressive Tokenization.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Scaling Mesh Generation via Compressive Tokenization

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.130354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.130354Z digest=sha256:137335d8c4fef3f5bbfaf82f44590751ff428ea06954e7043dcc9597fc179b2a

Observation 5ef0d01b-adfe-4c8e-b14a-e045329109c9 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning [c].

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Instructblip: Towards general-purpose vision-language models with instruction tuning [c]

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.226115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.134585Z digest=sha256:5475634503162c4ffd878bc1fecd5c12f5c5a20cf389b9af99b673fc835ddbc4

Observation 5b00c8d3-f7c8-4d21-b935-8441a953c2ea · outbound

This paper cites Unique3d: High-quality and efficient 3d mesh generation from a single image.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Unique3d: High-quality and efficient 3d mesh generation from a single image

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.210607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.138377Z digest=sha256:3525b99a8ced6c803509110215c5ef2b9d164ad8d1ade7fba849047615a89fe3

Observation 42565f75-fbb4-4cd2-9e10-26813377ff33 · outbound

This paper cites Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.142230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.142230Z digest=sha256:88c99723effbc52930b86c282410d03f0286571e33dbb664979858a97f5b52b7

Observation 79cc8c9e-cd5c-4b99-8491-f56d2fcf006a · outbound

This paper cites Structured 3D Latents for Scalable and Versatile 3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Structured 3D Latents for Scalable and Versatile 3D Generation

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.146343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.146343Z digest=sha256:4ad833a4988c26d61e5b8a83866e12f2fe8736a28e47eee91d4cdcd31601db16

Observation d7bd91bf-fc85-464e-b05d-50badb484302 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.150610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.150610Z digest=sha256:042819296dd93a081e8c26c185a0428dd9975549753b1185545bb478a45de266

Observation a0a569b7-a257-4d1a-9ce9-4b18ca190283 · outbound

This paper cites InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.155632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.155632Z digest=sha256:b0326fdc8d17d2a11465987a34b0a83d37d610a3304b6fb9ad137ce24bf5a1cc

Observation 4f68bb29-47c1-4b6f-9b10-3749e0f69e84 · outbound

This paper cites Pointllm: Empow- ering large language models to understand point clouds.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Pointllm: Empow- ering large language models to understand point clouds

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.195514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.159995Z digest=sha256:fe3bc288773ddf5c89a276fe3706842bc553c5ee81eef2fc725308470870975e

Observation 248e081b-02e5-43b8-93d4-c0415496802f · outbound

This paper cites DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.163979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.163979Z digest=sha256:22eef1f2e03b1970bf71ed3473b4b8e80369be086ff671858c2cd49cb0e3b0b9

Observation 35c12590-3fb0-410a-8e52-c71e84238aa7 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.181138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.168417Z digest=sha256:f332e1b6428f4f2e8a73645fb93a52717bc6019fc3b9d84691f5707a2b1d1ce0

Observation 7341f3c5-bcf7-41d6-8fb0-f87f3d62177f · outbound

This paper cites Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.173079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.173079Z digest=sha256:68d01b73d78f17eade5f5c4944db6b614dd012c2c490448e8d0dbb9cbb0bdaad

Observation 9e7b8e8c-facb-4bb0-a876-ad901c364d41 · outbound

This paper cites Stablenormal: Reducing diffusion variance for stable and sharp normal.ACM Transactions on Graphics (TOG), 2024.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Stablenormal: Reducing diffusion variance for stable and sharp normal.ACM Transactions on Graphics (TOG), 2024

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.166583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.178056Z digest=sha256:f122ed32598836855963f64f5365765220417f862400991437404519ac80e691

Observation 58f0e69d-0225-4bb6-b596-1ff0dadb5337 · outbound

This paper cites Hi3DGen: High-fidelity 3D Geometry Generation from Images via Normal Bridging.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Hi3DGen: High-fidelity 3D Geometry Generation from Images via Normal Bridging

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.182157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.182157Z digest=sha256:6172468e6cd79d3d3f3f53bff888342fd3ef7fb80ac72f56de2294b9a3de4081

Observation 9aa9e4a7-b61b-45d7-9a7e-dcc7c9c80f0d · outbound

This paper cites Dreamreward: Text-to-3d generation with human preference.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Dreamreward: Text-to-3d generation with human preference

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:37:29.152175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:37:28.186239Z digest=sha256:4d785eb65adf6b04f7023667b108db2aa2947b62346a3ef3c38db94a4fdbfefa

Observation 519efb95-2708-446d-8153-02e02ec65380 · outbound

This paper cites Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.190244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.190244Z digest=sha256:7e17d61a2b9aae946ca89ea437b6daaa58ec8ba89a9146e60a21e716e3ae42de

Observation 033307e0-27f1-49de-92b0-ea735b1e2f90 · outbound

This paper cites GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.194043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.194043Z digest=sha256:67512b2639cd52f278191c87b3d7a5ecb76ecf89b162becae45562027f5606a4

Pith citing papers

Observation 49058870-ee52-4be4-a8da-900f9056dd8e · inbound

Motus: A Unified Latent Action World Model cites this paper.

Motus: A Unified Latent Action World Model ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:44:36.865476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T18:44:36.636455Z digest=sha256:eeaf9e2abc4fd61c875dc58e5857ac26291c73fa9e7ca987a1abfc633a05fd7c

Observation 6e979359-e9ca-4a99-8336-e3b74fc31a21 · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:50:14.832521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:9597664ef760369c1afcc06d0d905fe99a3f7ec398b640cdbc9375180d332be6

Observation 1205b8f5-5fd6-45e7-9e5a-66869ba40c08 · inbound

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation cites this paper.

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 206

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:04.533037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T15:35:37.095627Z digest=sha256:270f23d52eebc00996a22438727cb73665fb672070a7f7da42083baa491878ca

Observation b674a98d-3a17-4bfc-bbc8-a7776f48ac9e · inbound

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection cites this paper.

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:41:12.157141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-09T15:26:50.289626Z digest=sha256:5341629ec207402058db6e8af43d744088f2eae0fdea022146f60d0f0d74b670

Observation be4efb64-eb67-4096-88c9-579956d8c3fc · inbound

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers cites this paper.

EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:27:47.816722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-19T21:26:36.707554Z digest=sha256:7e38843d806eb897c169390e36b019b4d195a90fab9e8a6e04a04ae70e0ada8f

Observation f772d536-8266-4d3e-a899-6807867d5ea2 · inbound

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects cites this paper.

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:41:21.940604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T09:39:48.573379Z digest=sha256:bbc20b17d3d5a1719a2f83c6d559ac865d3474c22dfd3e5603d4e613a2d9a109

Observation 34de4d10-d593-4454-b3e4-34752fbae90a · inbound

GEM: Generative Supervision Helps Embodied Intelligence cites this paper.

GEM: Generative Supervision Helps Embodied Intelligence ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.911677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-29T13:38:27.263726Z digest=sha256:45c096c90c3b4aa82aa380a2df5de19a9d22fafe96a3d6d0f08f62376c78efe9

Observation 286806b0-ee16-4238-944e-efafdf6ea434 · inbound

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation cites this paper.

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:55:28.946818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-01T06:54:22.727115Z digest=sha256:5aa5e90aae0bd4d5df35627fb9bf1bcbd67980f655127860fe6e89b582864b16

Observation 2b0a997f-e45c-4d49-a429-70802dad9b4d · inbound

ELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and Generation cites this paper.

ELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and Generation ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T01:44:26.235368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-08T01:40:40.690020Z digest=sha256:797bdf13d35b09a24acdb405a353e008b2d731a413c3fb8487493c22db864b08

Observation ee48ee27-38cd-4b2b-bd55-0b9d0bbe1f9d · inbound

Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing cites this paper.

Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:47.730082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:47.730082Z digest=sha256:3e253abfc04e57f80294bcc0f53d0bcabb0f35a91f41db8c324055bf7e025888

Observation ffd247c5-ee90-44c9-a736-bfad8d98e0fd · inbound

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets cites this paper.

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T00:33:58.615197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:33:58.615197Z digest=sha256:9eb367fa57c47054f5e3e505ba9c47c119bd2b5b1ab93c8a8bbc778eebe1cef9