Loading video...
Video Failed to Load
MVDiffusion: How to take a pre-trained text2image model for a perspective view (e.g., Stable Diffusion) and retrain to generate multiple consistent views (e.g., a panorama). Project site: Hugging Face demo: Code out in a month.
33,456 views • 3 years ago •via X (Twitter)
8 Comments

Yasutaka Furukawa3 years ago
I am very sorry. Typo fixing. Had to delete the old one and retweet.

Yasutaka Furukawa3 years ago
I had to delete and repost. Highly appreciate it if you could reshare/retweet this post

Ziyu Wan3 years ago
Awesome work!!!!! I wonder where we could find the paper😊

Yasutaka Furukawa3 years ago
Thank you Ziyu. Did not realize that we haven't uploaded to arxiv yet... Asking students to upload in a day.

Jake Harrison3 years ago
I would add to my newsletter

Philipp Tsipman3 years ago
🔥

Asriel H3 years ago
does it support image prompt as input or only text prompt?

Yasutaka Furukawa3 years ago
Only text for this work. But it seems trivial to add image-prompt capability.
