DAG uses frozen text-to-image diffusion features plus a learned decoder to predict manipulable regions on 3D object point clouds from a reference human-object interaction image and an affordance verb.
In 2015 IEEE International Conference on Robotics and Automation (ICRA), 1374–1381
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning
DAG uses frozen text-to-image diffusion features plus a learned decoder to predict manipulable regions on 3D object point clouds from a reference human-object interaction image and an affordance verb.