Pith. sign in

Paper Citation Record · LEDGER

Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2405.14832.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.14832 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:37:28.142230Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:57:09.705005Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d447d58-241c-4c30-9221-e525abf9fbfd · inbound

Structured 3D Latents for Scalable and Versatile 3D Generation cites this paper.

Structured 3D Latents for Scalable and Versatile 3D Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:10:39.660054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T15:10:39.554217Z digest=sha256:52a19632fba60358a1081f8e5c2b7c3cda6370464327fc06c3a2028d59492094

Observation 987fcea7-2c09-4728-9dc4-5f8a43c22838 · inbound

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation cites this paper.

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:55:24.942855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T04:53:58.465445Z digest=sha256:0283f69b7349b5bf35ebd6e455f2100d6b7e7f70004ee4ae3b9213ec2d63a54a

Observation 568f5956-5cf8-4455-9092-04f9e2a65db0 · inbound

TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models cites this paper.

TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 109

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T21:51:18.424465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T21:51:18.323840Z digest=sha256:f940f23fb170c2cc2bab71014f3628b9ec8eb7651e287a9bd74b26f203f9f2cf

Observation 42565f75-fbb4-4cd2-9e10-26813377ff33 · inbound

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding cites this paper.

ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T11:37:28.142230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:37:28.142230Z digest=sha256:a7de47bd3d693fa491fdf4b4c9773d1b1500f40556db7f78ab02c112e2da1898

Observation 65176052-c786-458a-8748-996a8c8d7fb6 · inbound

Squeeze3D: Your 3D Generation Model is Secretly an Extreme Neural Compressor cites this paper.

Squeeze3D: Your 3D Generation Model is Secretly an Extreme Neural Compressor Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-07T05:33:39.664262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:33:39.664262Z digest=sha256:098d42184d24de7c6099755c483db97daa70a8d94260fc560fcdabd4732c9570

Observation 77849f1e-d34f-4eb6-a147-8f6e31bc0355 · inbound

Efficient Part-level 3D Object Generation via Dual Volume Packing cites this paper.

Efficient Part-level 3D Object Generation via Dual Volume Packing Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:18.301179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:40:18.301179Z digest=sha256:f85b9d565888245578fe849d9fdd7f758a6c99dfbdc6973526985f99b076c1fe

Observation c57557a1-20fb-4d9d-a998-09f228de034f · inbound

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation cites this paper.

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:39:02.324366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:39:02.324366Z digest=sha256:645724ee08c3707590d88e46e8f195a81c9e42c48f74c521a3c8f1d616f16617

Observation d698e675-1372-4324-8105-78e9d3e16d4c · inbound

XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration via Multi-Head Speculative Decoding cites this paper.

XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration via Multi-Head Speculative Decoding Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:29:56.216416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:29:56.216416Z digest=sha256:8befda39f478c643abcb0c8db5bfde6f9ec2039c1b0532d1ca8820c508d5ad87

Observation c4e5e105-5f30-43a7-8e1b-d6fad8f9c9a9 · inbound

Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis cites this paper.

Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T10:30:15.918137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:30:15.918137Z digest=sha256:b167fa7d4084727e0d02a917d4a71273f304a29490a25b34c1a43008987873fc

Observation 297074a3-aad0-4a91-b23c-8c8206ac578b · inbound

VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space cites this paper.

VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:44.552196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:52:44.552196Z digest=sha256:9b728679f9ba560d2b6792a94052298c5296a617778fe56dd686aaf75040cee6

Observation 3f1b884d-fbdc-45c6-bb8a-0c8c289edb72 · inbound

Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework cites this paper.

Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:25.690899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:40:25.690899Z digest=sha256:580aab5c43900caf3ddd0f83dc0a6fe5f3e5dbcb5a834ca60dada9c38cb5325b

Observation 54789b1a-a093-4e36-9027-ff1aa3217edf · inbound

Few-step Flow for 3D Generation via Marginal-Data Transport Distillation cites this paper.

Few-step Flow for 3D Generation via Marginal-Data Transport Distillation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T10:19:33.859731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:19:33.859731Z digest=sha256:42f465196a135018a2f8d8103314918f211ad28b6351dd223d94d2a4b74c8a78

Observation 59917d2a-1886-43b7-bad5-a6fcb92013c5 · inbound

Native and Compact Structured Latents for 3D Generation cites this paper.

Native and Compact Structured Latents for 3D Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:20:42.378116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:20:42.300247Z digest=sha256:5edfdaadbfe8fc5474868a37421abfb8c9fbacf954f9fcbd77aad640ecf6d31c

Observation a11cdd7c-e046-417f-b44b-7ac12314e33a · inbound

MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation cites this paper.

MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T13:00:00.560729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T12:59:47.236252Z digest=sha256:4355d971fa16880991ba525844f96c28136537a478b7b1bb0f7f012dc9647b22

Observation c1f79350-3414-43af-9408-0420b5654a6c · inbound

SegviGen: Repurposing 3D Generative Model for Part Segmentation cites this paper.

SegviGen: Repurposing 3D Generative Model for Part Segmentation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:29:53.475852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T09:29:08.513396Z digest=sha256:acb9a8becd7474e474f25251a39f15e75b2ccc0910b887947db67ab2b460fc28

Observation 39dcc95e-8e14-43b4-961e-0c71fe283296 · inbound

TInR: Exploring Tool-Internalized Reasoning in Large Language Models cites this paper.

TInR: Exploring Tool-Internalized Reasoning in Large Language Models Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T22:20:17.569353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:20:17.569353Z digest=sha256:22e7e001bd887a8bb30f4775c10ad763b868c49276d6bc99f9d02ff61f9900cd

Observation 9959ba70-6cc1-454b-92fc-e618508bcf8e · inbound

ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment cites this paper.

ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:06:05.657587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:38:03.377613Z digest=sha256:40bbcd0b32b3c1372d57c69b67cb59be9d39c7ce1c02450658ea379037f3b5a4

Observation bf6e7051-0093-4fbd-a1bb-839ef7f072f8 · inbound

Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation cites this paper.

Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:06.163432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T04:18:35.668119Z digest=sha256:15147bd97f38b56645d5aa53e936615670cb490837465a06d5cf5bb084432cb9

Observation 16533ece-7949-4a59-98e2-413a3b350a09 · inbound

Velox: Learning Representations of 4D Geometry and Appearance cites this paper.

Velox: Learning Representations of 4D Geometry and Appearance Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:41:07.769821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T17:19:58.259824Z digest=sha256:1a43e0211fd88504a99e2e92220cd7e4e5b4dec992ed57261aa07755c35d80a0

Observation 2aba649f-c8de-4334-846e-ad7809df8985 · inbound

ROAR-3D: Routing Arbitrary Views for High-Fidelity 3D Generation cites this paper.

ROAR-3D: Routing Arbitrary Views for High-Fidelity 3D Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T05:04:37.277546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:04:07.790893Z digest=sha256:5f2c2ca2c32d04203552102af7eb1d2d3e6a2b6995e92dc66b02af92663133bd

Observation 13828f95-c54c-4cff-ab83-e5bd81913f82 · inbound

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment cites this paper.

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:57:09.711869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T22:17:37.135761Z digest=sha256:d46513a777f53b4434646a130ef048209cf2900ffa43911a56a37c3f8cbf70ad

Observation 577172a5-468a-47e7-8a24-972925c62977 · inbound

ReScene: Structured Indoor Scene Reconstruction from Multi-View Captures cites this paper.

ReScene: Structured Indoor Scene Reconstruction from Multi-View Captures Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:57.251059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:30:56.107507Z digest=sha256:610ec73f222764a0aa87be3706b5c5127ef7a2ca69f51ce2ad3516d325693bde

Observation f8cf2917-be86-448a-8bd4-390ae49080a5 · inbound

PointSplat: Compact Gaussian Splatting via Human-Centric Prediction cites this paper.

PointSplat: Compact Gaussian Splatting via Human-Centric Prediction Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:35:41.397854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:25:22.114369Z digest=sha256:cab87ad4a24a6b20d6e220caad77abe0c1804dd30f1ef6eafadf69757a52bba5

Observation ecc3e900-c1d1-4b6a-bfbd-2159cff5e437 · inbound

Restore3D: Breathing Life into Broken Objects with Shape and Texture Restoration cites this paper.

Restore3D: Breathing Life into Broken Objects with Shape and Texture Restoration Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-07-02T14:47:03.245117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T14:42:16.438956Z digest=sha256:2c6b85fef1bd443d8302009f61dd4d84b9c277fd1aed96ae2a7a4cf7ba2392e2

Observation 9f8aeaa6-6313-404f-a3f9-9d9d337fc285 · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 211

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:48.512499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:48.512499Z digest=sha256:f1f18487c321ee605c3b77d2c6270b78e557ce2f8cf7f2b732b94530fcf4a31e

Observation aa0b44c1-f36e-4ddc-aa3d-4dfbf4db2ba9 · inbound

UMI3D: Robust 3D Generation on Unconstrained Multi-Image Inputs via Simultaneous Focus Cross-Attention Routing cites this paper.

UMI3D: Robust 3D Generation on Unconstrained Multi-Image Inputs via Simultaneous Focus Cross-Attention Routing Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T18:55:32.025725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T18:55:32.025725Z digest=sha256:3217204471e69ba29eb8540084a473fd7944e025ce17ff88a998aa5d5427f9a7

Observation 03d71310-bb40-463f-8603-f881a8e3c458 · inbound

EditFlow3D: Automated Local Editing of 3D Assets with Trajectory Preservation cites this paper.

EditFlow3D: Automated Local Editing of 3D Assets with Trajectory Preservation Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T00:13:57.257536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:13:57.257536Z digest=sha256:f5d8f48dfd5281766666396508667e6c8dedd67cd408feecb817c33564a9113e