Learning 3D-Gaussian Simulators from RGB Videos

Andreas Ren\'e Geist; Georg Martius; Mikel Zhobro

arxiv: 2503.24009 · v3 · pith:VCPUBTEOnew · submitted 2025-03-31 · 💻 cs.GR · cs.AI· cs.CV· cs.LG· cs.RO

Learning 3D-Gaussian Simulators from RGB Videos

Mikel Zhobro , Andreas Ren\'e Geist , Georg Martius This is my paper

classification 💻 cs.GR cs.AIcs.CVcs.LGcs.RO

keywords dynamicsdatadgsiminformationparticlephysicaltemporalcapture

0 comments

read the original abstract

Realistic simulation is critical for applications ranging from robotics to animation. Learned simulators have emerged as a possibility to capture real world physics directly from video data, but very often require privileged information such as depth information, particle tracks and hand-engineered features to maintain spatial and temporal consistency. These strong inductive biases or ground truth 3D information help in domains where data is sparse but limit scalability and generalization in data rich regimes. To overcome the key limitations, we propose 3DGSim, a learned 3D simulator that directly learns physical interactions from multi-view RGB videos. 3DGSim unifies 3D scene reconstruction, particle dynamics prediction and video synthesis into an end-to-end trained framework. It adopts MVSplat to learn a latent particle-based representation of 3D scenes, a Point Transformer for particle dynamics, a Temporal Merging module for consistent temporal aggregation and Gaussian Splatting to produce novel view renderings. By jointly training inverse rendering and dynamics forecasting, 3DGSim embeds the physical properties into point-wise latent features. This enables the model to capture diverse physical behaviors, from rigid to elastic, cloth-like dynamics, and boundary conditions (e.g. fixed cloth corner), along with realistic lighting effects that also generalize to unseen multibody interactions and novel scene edits.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy
cs.LG 2026-05 unverdicted novelty 7.0

MoSA learns residual stress operators on an isotropic backbone using a physics-informed cascaded network and motion constraints to capture mild anisotropy and heterogeneity for improved real-to-sim dynamics.
Learning Action-Conditional and Object-Centric Gaussian Splatting World Models for Rigid Objects
cs.RO 2026-06 unverdicted novelty 5.0

MRO-GWM learns action-conditional 3D dynamics of rigid objects via object-centric Gaussians and a transformer, evaluated on synthetic multi-object scenes and in simulation for non-prehensile manipulation.