
Grace Luo @ ICML 2026
@graceluo_ • 2,283 subscribers
phd student @berkeley_ai, vision + language
Videos

✨New preprint: Dual-Process Image Generation! We distill *feedback from a VLM* into *feed-forward image generation*, at inference time. The result is flexible control: parameterize tasks as multimodal inputs, visually inspect the images with the VLM, and update the generator.🧵
Grace Luo @ ICML 2026133,504 görüntüleme • 1 yıl önce

Guidance on top of diffusion models can now be used to drag and manipulate images, create pose-conditioned images, and so much more! Check out Readout Guidance: Work w/ trevordarrell, Oliver Wang, Dan Goldman, Aleksander Holynski. More in thread 🧵.
Grace Luo @ ICML 202642,717 görüntüleme • 2 yıl önce
Daha fazla içerik yok.