Pith. sign in

REVIEW 3 cited by

Disentangled Clothed Avatar Generation from Text Descriptions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2312.05295 v2 pith:5GA4DRAR submitted 2023-12-08 cs.CV

classification cs.CV
keywords avatarso-smpltextanimationbodyclothesgenerationhuman
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this paper, we introduce a novel text-to-avatar generation method that separately generates the human body and the clothes and allows high-quality animation on the generated avatar. While recent advancements in text-to-avatar generation have yielded diverse human avatars from text prompts, these methods typically combine all elements-clothes, hair, and body-into a single 3D representation. Such an entangled approach poses challenges for downstream tasks like editing or animation. To overcome these limitations, we propose a novel disentangled 3D avatar representation named Sequentially Offset-SMPL (SO-SMPL), building upon the SMPL model. SO-SMPL represents the human body and clothes with two separate meshes but associates them with offsets to ensure the physical alignment between the body and the clothes. Then, we design a Score Distillation Sampling (SDS)-based distillation framework to generate the proposed SO-SMPL representation from text prompts. Our approach not only achieves higher texture and geometry quality and better semantic alignment with text prompts, but also significantly improves the visual quality of character animation, virtual try-on, and avatar editing. Project page: https://shanemankiw.github.io/SO-SMPL/.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation

    cs.CV 2024-11 conditional novelty 6.0 of 10

    SIMS couples retrieval-augmented LLM scripts with a text-conditioned, physics-based control policy to generate stylized human-scene interactions.

  2. PhyCAGE: Physically Plausible Compositional 3D Asset Generation from a Single Image

    cs.CV 2024-11 conditional novelty 6.0 of 10

    A single-image pipeline that generates physically plausible compositional 3D Gaussian Splatting assets by using a physics simulator as a gradient-driven optimizer.

  3. SimAvatar: Simulation-Ready Avatars with Layered Hair and Clothing

    cs.CV 2024-12 conditional novelty 5.0 of 10

    SimAvatar generates text-described 3D avatars with separate body, garment, and hair layers that can be driven by off-the-shelf physics simulators.

Pith tools