Загрузка видео...
Не удалось загрузить видео
Scaling up GANs for Text-to-Image Synthesis present our 1B-parameter GigaGAN, achieving lower FID than Stable Diffusion v1.5, DALL·E 2, and Parti-750M. It generates 512px outputs at 0.13s, orders of magnitude faster than diffusion and autoregressive models, and inherits the disentangled, continuous, and controllable latent space of GANs abs: project page:
278,115 просмотров • 3 лет назад •via X (Twitter)
Комментарии: 10

Daniel Losey 🔀3 лет назад
amazing

David Marx (@digthatdata.bsky.social)3 лет назад
GANs are back baybee

Nicolay Mausz3 лет назад
Adobe research - I guess this will be part of CC

Draz ⚛️3 лет назад
The upscaling is quite insane on how it accurately fills in details

Nerdy Rodent 🐀🤓💻3 лет назад
It’s been hours now, why isn’t it showing up? 😉

Asriel H3 лет назад
It has the same schema of injecting latent vector into every scaling layer as StyleGAN has

okaris3 лет назад
The examples provided don’t look as good as diffusion models. Some details obscured or looking weird.

Adhik Joshi3 лет назад
Weights aren't open-source

Julien Genoud3 лет назад
The 4k upsampler 🤯

Clarence Hu3 лет назад
paging @gwern
