
Gerard Pons-Moll
@GerardPonsMoll1 • 6,359 subscribers
Professor of Computer Science at the University of Tübingen. Machine Learning, Graphics and 3D Computer Vision enthusiast.
Shorts
Videos

Capturing every combination of human-object interaction is not feasible. We need compositionally. Most interactions are local. A hand holds a cup, a chair supports the pelvis, while much of the body remains free. COSMI, by Daniel Escandar Ilya Petrov builds on this observation to compose existing single-object datasets into 222k multi-object sequences, totaling 275 hours, with up to 5 objects per sequence. LLM reasoning and geometric checks ensure the compositions remain plausible. On this data, we train a single diffusion transformer for interactions with 1 to 5 objects, predicting each object relative to its interacting body part and generalizing to unseen objects and combinations. Data and code will be made available: 🌐
Gerard Pons-Moll20,393 görüntüleme • 2 gün önce

Check out PhysHead: Simulation-Ready Gaussian Head Avatars (CVPR 2026)! We introduce a layered head & hair model using strands + Gaussian splats, enabling physics-based animation from multi-view video. PS: We didn’t dare compute our bald versions… maybe you’re braver 🙂
Gerard Pons-Moll10,232 görüntüleme • 5 ay önce

Nice demo from Unitree. I shaked hands with it but didn't dare to box against it :-)
Gerard Pons-Moll13,979 görüntüleme • 1 yıl önce
Daha fazla içerik yok.