REVIEW 1 cited by
HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure Priors
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Despite recent advancements in high-fidelity human reconstruction techniques, the requirements for densely captured images or time-consuming per-instance optimization significantly hinder their applications in broader scenarios. To tackle these issues, we present HumanSplat which predicts the 3D Gaussian Splatting properties of any human from a single input image in a generalizable manner. In particular, HumanSplat comprises a 2D multi-view diffusion model and a latent reconstruction transformer with human structure priors that adeptly integrate geometric priors and semantic features within a unified framework. A hierarchical loss that incorporates human semantic information is further designed to achieve high-fidelity texture modeling and better constrain the estimated multiple views. Comprehensive experiments on standard benchmarks and in-the-wild images demonstrate that HumanSplat surpasses existing state-of-the-art methods in achieving photorealistic novel-view synthesis.
Forward citations
Cited by 1 Pith paper
-
Snap-Snap: Taking Two Images to Reconstruct 3D Human Gaussians in Milliseconds
A feed-forward pipeline predicts 3D human Gaussian splats from two input images (front and back) in 190 ms, using a DUSt3R-style point cloud predictor with extra side-view heads, nearest-neighbor color warping, and a ...
Discussion (0). Continue with ORCID to comment.