Pith. sign in

REVIEW 10 cited by

Advances in 3D Generation: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.17807 v1 pith:ZAE7O3K5 submitted 2024-01-31 cs.CV cs.GR

classification cs.CVcs.GR
keywords generationfieldmethodsmodelssurveyapplicationscontentdatasets
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Generating 3D models lies at the core of computer graphics and has been the focus of decades of research. With the emergence of advanced neural representations and generative models, the field of 3D content generation is developing rapidly, enabling the creation of increasingly high-quality and diverse 3D models. The rapid growth of this field makes it difficult to stay abreast of all recent developments. In this survey, we aim to introduce the fundamental methodologies of 3D generation methods and establish a structured roadmap, encompassing 3D representation, generation methods, datasets, and corresponding applications. Specifically, we introduce the 3D representations that serve as the backbone for 3D generation. Furthermore, we provide a comprehensive overview of the rapidly growing literature on generation methods, categorized by the type of algorithmic paradigms, including feedforward generation, optimization-based generation, procedural generation, and generative novel view synthesis. Lastly, we discuss available datasets, applications, and open challenges. We hope this survey will help readers explore this exciting topic and foster further advancements in the field of 3D content generation.

Discussion (0). Sign in to comment.

Forward citations

Cited by 10 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. IDEAL-Bench: Indoor Dataset and Evaluation suite for Analyzing 3D Layout reasoning

    cs.CV 2026-07 accept novelty 7.0 of 10

    Current VLMs top out at 62.1/100 on holistic single-image 3D indoor layout prediction, with strong recognition but weak geometric regression, and mid-tier rankings that shift relative to QA and primitive-reconstructio...

  2. Objectness Similarity: Capturing Object-Level Fidelity in 3D Scene Evaluation

    cs.CV 2025-09 conditional novelty 7.0 of 10

    OSIM aligns more closely with human perception than existing 3D scene metrics by measuring object-level feature similarity within detected objects and weighting scores by saliency.

  3. Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Hallo4D uses vision-language models to detect and correct spatial and temporal mistakes in AI-generated 3D and 4D content, improving consistency without retraining the base generators.

  4. Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework

    cs.GR 2025-09 conditional novelty 6.0 of 10

    SDF grids reconstruct best, Dual Octrees score best on automatic generation metrics, but users prefer SDF output, and reconstruction plus compression errors make up a large share of generation error.

  5. Detecting Visual Information Manipulation Attacks in Augmented Reality: A Multimodal Semantic Reasoning Approach

    cs.CV 2025-07 conditional novelty 6.0 of 10

    A new taxonomy and dataset of semantic AR manipulation attacks, plus a VLM+OCR detector that achieves 88.94% accuracy and roughly 7-second latency.

  6. Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling

    cs.CV 2025-05 conditional novelty 6.0 of 10

    A sparse, deformable marching-cubes representation plus a sparse-convolution VAE convert rough meshes into watertight 1024^3 surfaces with less detail loss, faster training, and better downstream 3D generation than pr...

  7. SplatPainter: Interactive Authoring of 3D Gaussians from 2D Edits via Test-Time Training

    cs.CV 2025-12 conditional novelty 5.0 of 10

    A test-time-trained feedforward model that propagates 2D edits onto 3D Gaussian attributes at interactive speeds.

  8. Drive Any Mesh: 4D Latent Diffusion for Mesh Deformation from Video

    cs.CV 2025-06 conditional novelty 5.0 of 10

    A video-conditioned latent diffusion model generates mesh vertex trajectories that deform an input 3D asset into render-ready 4D animations.

  9. Reconstructing 4D Spatial Intelligence: A Survey

    cs.CV 2025-07 accept novelty 4.0 of 10

    A review that classifies 4D scene reconstruction methods into five progressive levels: low-level cues, scene components, dynamic scenes, interactions, and physics.

  10. LLM-to-Phy3D: Physically Conform Online 3D Object Generation with LLMs

    cs.CV 2025-06 conditional novelty 4.0 of 10

    An iterative prompt-refinement wrapper around LLM-to-3D generation, using CFD drag, vision-language domain scores, and visual novelty, reports 4.5% to 106.7% DPAR gains over non-refined baselines in car design.

Pith tools