Switti -- a new scale-wise transformer for text-to-image generation...

Gradio's profile picture

Gradio

29,314 Aufrufe • vor 1 Jahr

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 Aufrufe • vor 1 Jahr

Snap presents MoA Mixture-of-Attention for Subject-Context Disentanglement in Personalized...

AK's profile picture

AK

47,488 Aufrufe • vor 2 Jahren

Text-to-image diffusion transformer models learn to align text and...

Alec Helbling's profile picture

Alec Helbling

94,095 Aufrufe • vor 7 Monaten

🤯 𝐒𝐥𝐢𝐜𝐞𝐝𝐢𝐭 : A revolutionary approach to zero-shot video...

Gradio's profile picture

Gradio

19,586 Aufrufe • vor 2 Jahren

Google presents MobileDiffusion Subsecond Text-to-Image Generation on Mobile Devices...

AK's profile picture

AK

150,554 Aufrufe • vor 2 Jahren

🚨Announcing our #ICLR2025 Oral! 🔥Diffusion LMs are on the...

Marianne Arriola @ICML's profile picture

Marianne Arriola @ICML

150,079 Aufrufe • vor 1 Jahr

We are excited to introduce Mercury, the first commercial-grade...

Inception's profile picture

Inception

1,915,287 Aufrufe • vor 1 Jahr

Image generation with Gemini just got a bananas upgrade...

Google DeepMind's profile picture

Google DeepMind

1,482,123 Aufrufe • vor 11 Monaten

The v0.14 release of BERTopic is here 🥳 Fine-tune...

Maarten Grootendorst's profile picture

Maarten Grootendorst

124,476 Aufrufe • vor 3 Jahren

LongWriter Unleashing 10,000+ Word Generation from Long Context LLMs...

AK's profile picture

AK

50,995 Aufrufe • vor 1 Jahr

Tired of writing image generation prompts? Introducing "Describe with...

AI Waifus's profile picture

AI Waifus

4,350,661 Aufrufe • vor 1 Jahr

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,746 Aufrufe • vor 3 Jahren

NVIDIA just released a very impressive text-to-video paper. Video...

Lior Alexander's profile picture

Lior Alexander

158,565 Aufrufe • vor 3 Jahren

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,801 Aufrufe • vor 1 Jahr

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 Aufrufe • vor 2 Jahren

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,123 Aufrufe • vor 3 Jahren

Today we're sharing details on AudioCraft, a new family...

AI at Meta's profile picture

AI at Meta

677,780 Aufrufe • vor 3 Jahren

Big news from Boltz - our biggest update yet!...

Gabriele Corso's profile picture

Gabriele Corso

136,389 Aufrufe • vor 1 Monat