Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 просмотров • 1 год назад

New Generation Model! 🚨 We're introducing the Mistral model...

ALCHEMIST AI 🔮's profile picture

ALCHEMIST AI 🔮

16,432 просмотров • 1 год назад

Switti -- a new scale-wise transformer for text-to-image generation...

Gradio's profile picture

Gradio

29,314 просмотров • 1 год назад

NVIDIA AI Released DiffusionRenderer: An AI Model for Editable,...

Marktechpost AI Dev News ⚡'s profile picture

Marktechpost AI Dev News ⚡

104,741 просмотров • 1 год назад

LongWriter Unleashing 10,000+ Word Generation from Long Context LLMs...

AK's profile picture

AK

50,995 просмотров • 1 год назад

With Hunyuan3D World Model 1.0 now released and open-sourced,...

Tencent Hy's profile picture

Tencent Hy

23,150 просмотров • 1 год назад

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,047 просмотров • 1 год назад

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,746 просмотров • 3 лет назад

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 просмотров • 2 лет назад

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 просмотров • 1 год назад

Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering discuss: The...

AK's profile picture

AK

19,101 просмотров • 1 год назад

🇨🇳 Another great Chinese Model, OmniHuman-1.5 from ByteDance Turns...

Rohan Paul's profile picture

Rohan Paul

63,859 просмотров • 10 месяцев назад

Microsoft presents Windows Agent Arena Evaluating Multi-Modal OS Agents...

AK's profile picture

AK

19,684 просмотров • 1 год назад

A New Era with V3🪄 V3's new engine introduces...

ALCHEMIST AI 🔮's profile picture

ALCHEMIST AI 🔮

48,721 просмотров • 1 год назад

STEVE-1: A Generative Model for Text-to-Behavior in Minecraft paper...

AK's profile picture

AK

144,783 просмотров • 3 лет назад

We’re excited to introduce Text-to-LoRA: a Hypernetwork that generates...

Sakana AI's profile picture

Sakana AI

403,159 просмотров • 1 год назад

AI Is Moving Beyond “Generating Videos” — Toward “Generating...

雪踏乌云's profile picture

雪踏乌云

112,114 просмотров • 8 дней назад

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,123 просмотров • 3 лет назад

Introducing VL-JEPA: Vision-Language Joint Embedding Predictive Architecture for streaming,...

Pascale Fung's profile picture

Pascale Fung

90,144 просмотров • 7 месяцев назад

Diffusion models are an amazing tool for cofolding, they...

Sergey Edunov's profile picture

Sergey Edunov

36,985 просмотров • 1 месяц назад

We are excited to introduce Stable Fast 3D, Stability...

Stability AI's profile picture

Stability AI

438,350 просмотров • 2 лет назад