正在加载视频...

视频加载失败

🎥 Frustrated by Sora's credit limits? Still waiting for Veo 2? 🚀 Open-source video DiTs are actually on par. We introduce FastVideo, an open-source stack to support fast video generation for SoTA open models. We have supported Mochi and Hunyuan, 8x faster inference, 720P 5-second video in 62 seconds.

69,735 次观看 • 1 年前 •via X (Twitter)

11 条评论

Hao AI Lab 的头像
Hao AI Lab1 年前

(2/6) 🎬Compared to the original Hunyuan Video, FastVideo reduces the diffusion time from 232 seconds to 27 seconds, and the end-to-end time from 267 seconds to 62 seconds. (all measured on 8XH100)

Hao AI Lab 的头像
Hao AI Lab1 年前

(3/6) 🎥 Compared to the original Mochi, FastMochi reduces the diffusion time from 63 seconds to 26 seconds, and the end-to-end time from 123 seconds to 81 seconds. (all measured on 8XH100)

Hao AI Lab 的头像
Hao AI Lab1 年前

(4/6) Behind the scenes, FastVidep uses consistency distillation (CD). CD was proposed to accelerate image diffusion models, but its application to video Diffusion Transformers (DiT) has been scattered—until now. We’re sharing the first open recipe for CD on video DiTs with open data, checkpoints, and codebase. You can follow our recipe to distill your own model! 🚀 HF link:

Hao AI Lab 的头像
Hao AI Lab1 年前

(5/6) Beyond CD, FastVideo is lightweight yet powerful, packed with many useful features: 🏆 Support for distilling, finetuning, and inferencing SoTA video DiTs: Mochi and Hunyuan ⚡ Scalable training with FSDP, sequence parallelism, and selective activation checkpointing—achieving near-linear scaling to 64 GPUs. 🛠️ Memory-efficient fine-tuning with LoRA.

Hao AI Lab 的头像
Hao AI Lab1 年前

(6/6) Github: Shout out to @PY_Z001 @RunlongSu @_foreverpiano @BrianChen112900 @DachengLi177 and @haozhang for their hard work! We thank @anyscalecompute @richliaw and @mbzuai for their support throughout this project.

destream official 的头像
destream official1 年前

Creators, receiving payments shouldn't be a hassle. With destream, it's fast and seamless. Boost your income today!

Esteban 的头像
Esteban1 年前

but can you add image to video, instead of just text to video?

Pip 的头像
Pip1 年前

How many steps?

Hao AI Lab 的头像
Hao AI Lab1 年前

FastMochi-8 FastHunayuan-6

neb 的头像
neb1 年前

Firstly thanks that really cool ! Been trying the model since this afternoon but can’t figure proper setting they often come « overcooked » :| (tried the 6 steps / 6 guidance 17 flow shift)

Rohit Bharadwaj 的头像
Rohit Bharadwaj1 年前

@DrYangSong What are the min VRAM/GPU requirements for inference?

相关视频

Here are 10 AI video editor GitHub repos worth bookmarking: 1. Shotcut Most actively maintained open source video editor in 2026. 14K stars. Cross-platform with AI-assisted features. Just shipped a new release April 30, 2026. 2. Kdenlive The closest open source alternative to Adobe Premiere Pro. Multi-track editing, proxy editing, VST audio, and customizable workspace. Best for professional workflows. 3. OpenShot The easiest entry point for beginners. Drag and drop, 400+ transitions, 3D titles, and AI-assisted trimming. 5,700 stars. 4. Blender Not just 3D. Blender's video sequence editor and compositing pipeline is used in professional film production. 18,300 stars. Unmatched for VFX. 5. Recordly Screen recorder with auto-zoom, cursor polish, webcam overlays, and styled frames built in. Built for demo videos and walkthroughs. 6. Wan2.1 Alibaba's open source text-to-video model. Cinema-grade 1080p generation. Apache 2.0. The gold standard for open source video generation in 2026. 7. HunyuanVideo Tencent's 13B parameter open source video model. 11.9K stars. Handles 720p and 1080p with high temporal coherence. 8. CogVideoX Apache 2.0 licensed. Loads natively via Hugging Face Diffusers. Strong prompt following and smooth frame transitions. Needs 16GB VRAM minimum. 12.5K stars. 9. Open-Sora Most starred open source video generation project at 24K stars. Full training pipeline for $200K. Production-level output quality. 10. Mochi 1 Focused entirely on motion quality. The most natural-looking physics of any open source video model. Water, fabric, and human gestures without AI jitter. Apache 2.0.

Kanika

17,309 次观看 • 1 个月前