Gradio's banner
Gradio's profile picture

Gradio

@Gradio57,025 subscribers

Build and share machine learning apps in 3 lines of Python. Part of the @Huggingface family 🤗. DMs are open for sharing your gradio app with us for promotion!

Shorts

😀 ToonCrafter: Generative Cartoon Interpolation 🔆 Can interpolate two cartoon images by leveraging the pre-trained image-to-video diffusion priors. Model on 🤗👇

😀 ToonCrafter: Generative Cartoon Interpolation 🔆 Can interpolate two cartoon images by leveraging the pre-trained image-to-video diffusion priors. Model on 🤗👇

93,074 views

Microsoft's OmniParser V2 just dropped. Now AI can : *understands every button* *identifies all interactive elements* *structures the entire GUI* Pure vision-based parsing. GUI automation just got a lot easier 🎯

Microsoft's OmniParser V2 just dropped. Now AI can : *understands every button* *identifies all interactive elements* *structures the entire GUI* Pure vision-based parsing. GUI automation just got a lot easier 🎯

61,978 views

🆕 Image super-resolution model just dropped! Superior results even with a single sampling step. 🔥InvSR: Arbitrary-steps Image Super-resolution via Diffusion Inversion.

🆕 Image super-resolution model just dropped! Superior results even with a single sampling step. 🔥InvSR: Arbitrary-steps Image Super-resolution via Diffusion Inversion.

53,331 views

Leffa from Meta is a 🆕 unified framework for controllable person image generation that enables precise manipulation of both appearance (i.e., Virtual try-on) and pose (i.e., Pose transfer). Learn more about 📚 Code,🔥 Demo, and 🤗 Model ⬇️

Leffa from Meta is a 🆕 unified framework for controllable person image generation that enables precise manipulation of both appearance (i.e., Virtual try-on) and pose (i.e., Pose transfer). Learn more about 📚 Code,🔥 Demo, and 🤗 Model ⬇️

52,913 views

Trending: Step1X-3D for generating high-fidelity 3D assets with versatile textures! Apache-2.0 with exceptional results in terms of geometry and texture mapping🔥🔥

Trending: Step1X-3D for generating high-fidelity 3D assets with versatile textures! Apache-2.0 with exceptional results in terms of geometry and texture mapping🔥🔥

36,405 views

🔥Introducing VILA: The next-gen VLM for Video and inter-leaved image understanding with Apache-2.0 license! Is VILA the best in-class small VLM for edge-deployment? Keep reading🧵 👀

🔥Introducing VILA: The next-gen VLM for Video and inter-leaved image understanding with Apache-2.0 license! Is VILA the best in-class small VLM for edge-deployment? Keep reading🧵 👀

51,694 views

😍 🆕 DeepSeek Mixture-of-Experts VLMs are out on Hugging Face Spaces! Witness advanced multimodal understanding in the app linked. Available in 3 sizes: DeepSeek-VL2-Tiny, DeepSeek-VL2-Small & DeepSeek-VL2, with 1.0B, 2.8B & 4.5B activated params resp. out of available 16B.

😍 🆕 DeepSeek Mixture-of-Experts VLMs are out on Hugging Face Spaces! Witness advanced multimodal understanding in the app linked. Available in 3 sizes: DeepSeek-VL2-Tiny, DeepSeek-VL2-Small & DeepSeek-VL2, with 1.0B, 2.8B & 4.5B activated params resp. out of available 16B.

36,491 views

This simple HTML app has a drag-and-drop canvas, 20+ text effects, 3D extrusion, and perspective transforms for your 'text behind image'. 😎 Crazy part: The ML backend is ~50 lines of Python Built with gradio.Server ➡️ bring any frontend you want!

This simple HTML app has a drag-and-drop canvas, 20+ text effects, 3D extrusion, and perspective transforms for your 'text behind image'. 😎 Crazy part: The ML backend is ~50 lines of Python Built with gradio.Server ➡️ bring any frontend you want!

11,744 views

This is built using Gradio 6's super easy new HTML component! You can play with the demo here:

This is built using Gradio 6's super easy new HTML component! You can play with the demo here:

15,652 views

Switti -- a new scale-wise transformer for text-to-image generation 🦾 🔥 Improved generation of fine-grained details. Outperforms existing T2I AR models and competes with state-of-the-art T2I diffusion models while being up to 7x faster.

Switti -- a new scale-wise transformer for text-to-image generation 🦾 🔥 Improved generation of fine-grained details. Outperforms existing T2I AR models and competes with state-of-the-art T2I diffusion models while being up to 7x faster.

29,314 views

Videos