Hugging Apps's banner
Hugging Apps's profile picture

Hugging Apps

@HuggingApps3,710 subscribers

Tweeting the coolest apps and demos on @huggingface Spaces curation by @multimodalart

Shorts

ABot World 0.5B is out! we have now world models at home, running in real-time on consumer GPU feed it an initial image, steer it with with your keyboard 🌎 you can run it locally or play now on spaces ▶️

ABot World 0.5B is out! we have now world models at home, running in real-time on consumer GPU feed it an initial image, steer it with with your keyboard 🌎 you can run it locally or play now on spaces ▶️

141,113 görüntüleme

MiniMax H3 inpainting allows for precise video editing with incredible control ✏️🎞️ supports clickable or promptable masks, or user-drown ones. works with comfyui and 🧨 diffusers and of course, there's an app by Linoy Tsaban ▶️

MiniMax H3 inpainting allows for precise video editing with incredible control ✏️🎞️ supports clickable or promptable masks, or user-drown ones. works with comfyui and 🧨 diffusers and of course, there's an app by Linoy Tsaban ▶️

39,701 görüntüleme

LTX Ripple is a new efficient IC LoRA approach for video editing ✏️🎞️ edit just the first frame and have the edits ripple into the rest of the video. lands perfect edits, incredibly fast! LTX 2.5 based ▶️

LTX Ripple is a new efficient IC LoRA approach for video editing ✏️🎞️ edit just the first frame and have the edits ripple into the rest of the video. lands perfect edits, incredibly fast! LTX 2.5 based ▶️

25,532 görüntüleme

AntResearch just dropped 4DAnyone on Hugging Face! Convert a video of any person into a 4D gaussian splat for that person ▶️

AntResearch just dropped 4DAnyone on Hugging Face! Convert a video of any person into a 4D gaussian splat for that person ▶️

41,036 görüntüleme

Canter 2B is a photography-oriented text-to-image model being trained on just a single GPU by data-archetype 🐎 are we entering the indie ai models era? still training, the v0002 checkpoint is available, more to come ▶️ play on spaces

Canter 2B is a photography-oriented text-to-image model being trained on just a single GPU by data-archetype 🐎 are we entering the indie ai models era? still training, the v0002 checkpoint is available, more to come ▶️ play on spaces

24,291 görüntüleme

Krea2-SDA LoRA fixes the biggest issue with Krea2 Turbo: the lack of variety on the same prompt, even with different seeds by applying semantic directional alignment in 2 of the 8 generation steps, we get way more diversity of outcomes ▶️

Krea2-SDA LoRA fixes the biggest issue with Krea2 Turbo: the lack of variety on the same prompt, even with different seeds by applying semantic directional alignment in 2 of the 8 generation steps, we get way more diversity of outcomes ▶️

12,864 görüntüleme

Microsoft Asia just dropped Mage-Flow on Hugging Face a smol 4B model for image generation and editing that matches much larger models in quality and it's fast! 4 steps in < 1s at 1024x1024 (model goes up to 4K) try out on spaces

Microsoft Asia just dropped Mage-Flow on Hugging Face a smol 4B model for image generation and editing that matches much larger models in quality and it's fast! 4 steps in < 1s at 1024x1024 (model goes up to 4K) try out on spaces

48,598 görüntüleme

NVIDIA’s ARDY is a glimpse of where AI animation is heading: real time, open source and you can play with it right now on Hugging Face Spaces Type what the character should do and generate a motion sequence in seconds. Give a sentence a body.

NVIDIA’s ARDY is a glimpse of where AI animation is heading: real time, open source and you can play with it right now on Hugging Face Spaces Type what the character should do and generate a motion sequence in seconds. Give a sentence a body.

50,367 görüntüleme

every 3D mesh is kind of a compiled binary: printable, usable, but frozen. you can't just "make this hole 2mm wider" Cadena is just out to fix this, a decompiler of 3D meshes - converting it into an editable CAD program, step by step 🧊 app on spaces ▶️

every 3D mesh is kind of a compiled binary: printable, usable, but frozen. you can't just "make this hole 2mm wider" Cadena is just out to fix this, a decompiler of 3D meshes - converting it into an editable CAD program, step by step 🧊 app on spaces ▶️

30,354 görüntüleme

Alibaba finally dropped Wan 2.2 Animate 14B on Hugging Face. Fully open source, video-to-video smooth and high quality motion transfer ▶️ locally or on Spaces

Alibaba finally dropped Wan 2.2 Animate 14B on Hugging Face. Fully open source, video-to-video smooth and high quality motion transfer ▶️ locally or on Spaces

24,138 görüntüleme

Microsoft just dropped Fara 1.5 computer use on Hugging Face a collection (4B, 9B and 27B) computer-use models that drive a real browser, looping screenshot → click or keystroke → repeat - highly performant, even the 4B one, awesome for local use ▶️ on spaces

Microsoft just dropped Fara 1.5 computer use on Hugging Face a collection (4B, 9B and 27B) computer-use models that drive a real browser, looping screenshot → click or keystroke → repeat - highly performant, even the 4B one, awesome for local use ▶️ on spaces

27,845 görüntüleme

ScenA is a new omni-TTS model out on Hugging Face, it takes a text description + two voice references and produces a full mixed scene with sound effects 🎧 does pineapple belong on pizza? 🍕 🍍 ▶️ on Spaces

ScenA is a new omni-TTS model out on Hugging Face, it takes a text description + two voice references and produces a full mixed scene with sound effects 🎧 does pineapple belong on pizza? 🍕 🍍 ▶️ on Spaces

22,301 görüntüleme

your band's jam session, but now it's editable MIDI Muscriptor (Mirelo x Kyutai) is the first model that transcribes audio into per-instrument MIDI note tracks really well ▶️ on Spaces

your band's jam session, but now it's editable MIDI Muscriptor (Mirelo x Kyutai) is the first model that transcribes audio into per-instrument MIDI note tracks really well ▶️ on Spaces

21,046 görüntüleme

SenseNova U1 is the best open image model you maybe didn't heard of yet Unified language-image (like Nano Banana 2 / GPT-Image-2) that does precise generations and edits It has a general version and a specialized infographics one, where it shines ▶️

SenseNova U1 is the best open image model you maybe didn't heard of yet Unified language-image (like Nano Banana 2 / GPT-Image-2) that does precise generations and edits It has a general version and a specialized infographics one, where it shines ▶️

26,858 görüntüleme

UniSE is here to remove the background noise from any audio - finally open source caught up here your voice memos recorded inside of a blender are salvageable now 🔊 try it for yourself on Spaces ▶️

UniSE is here to remove the background noise from any audio - finally open source caught up here your voice memos recorded inside of a blender are salvageable now 🔊 try it for yourself on Spaces ▶️

22,224 görüntüleme

CharacterSheet LoRA by Alisson Pereira just dropped on HF and is so good 🎨 feed it one image of a character, photo or illustration, and this FLUX.2-klein LoRA generates a full model sheet: portrait, front, side and back, identity intact ▶️ on Spaces

CharacterSheet LoRA by Alisson Pereira just dropped on HF and is so good 🎨 feed it one image of a character, photo or illustration, and this FLUX.2-klein LoRA generates a full model sheet: portrait, front, side and back, identity intact ▶️ on Spaces

16,295 görüntüleme

BS-Roformer Leap is here! separating vocal 🗣️ from instrument 🎸 in an absolutely clean and precise manner ▶️ come karaoke on spaces

BS-Roformer Leap is here! separating vocal 🗣️ from instrument 🎸 in an absolutely clean and precise manner ▶️ come karaoke on spaces

17,641 görüntüleme

WordVoice TTS ships what I've always wanted: a TTS system with per word control you can let the system auto-pilot (based on CosyVoice3) or control every word with duration, loudness, pitch or tone works with cloned or pre-set voices ▶️ on spaces

WordVoice TTS ships what I've always wanted: a TTS system with per word control you can let the system auto-pilot (based on CosyVoice3) or control every word with duration, loudness, pitch or tone works with cloned or pre-set voices ▶️ on spaces

18,725 görüntüleme

reshooting a video from a camera that was never there 📹 This LTX-2.3 IC-LoRA re-renders real footage from a NEW angle - same scene, same action, different camera. Pick "far to the left, higher, further" and it just… moves Cooked by Cseti 🧑‍🍳 try it live on Hugging Face: ▶️

reshooting a video from a camera that was never there 📹 This LTX-2.3 IC-LoRA re-renders real footage from a NEW angle - same scene, same action, different camera. Pick "far to the left, higher, further" and it just… moves Cooked by Cseti 🧑‍🍳 try it live on Hugging Face: ▶️

17,074 görüntüleme

CrisperWhisper 2 is out and it adds an "intended" mode to transcription Removing spoken speech markers from the transcript, keeping accurate word-level timings ▶️ on Spaces

CrisperWhisper 2 is out and it adds an "intended" mode to transcription Removing spoken speech markers from the transcript, keeping accurate word-level timings ▶️ on Spaces

12,232 görüntüleme

Videos

Daha fazla içerik yok.