Enze Xie's banner
Enze Xie's profile picture

Enze Xie

@xieenze_jr1,965 subscribers

Staff Research Scientist at NVIDIA, doing GenAI, CS PhD from HKU MMLab, interned at NVIDIA.

Shorts

🚀 Sol Video Inference Engine is here! An agent-native, training-free full-stack accelerator for video diffusion. It auto-tunes cache + sparse attn + token pruning + quant + kernel fusion for any model/hardware/config. >2× end-to-end speedup on 64B Cosmos3-Super, 22B LTX-2.3 and 2B SANA-Video — near-lossless VBench quality, minimal human effort. Practical acceleration for real video gen deployment. 📄 Paper: 🌐 Project: 💻 Code: Proud of the team! 🎉

🚀 Sol Video Inference Engine is here! An agent-native, training-free full-stack accelerator for video diffusion. It auto-tunes cache + sparse attn + token pruning + quant + kernel fusion for any model/hardware/config. >2× end-to-end speedup on 64B Cosmos3-Super, 22B LTX-2.3 and 2B SANA-Video — near-lossless VBench quality, minimal human effort. Practical acceleration for real video gen deployment. 📄 Paper: 🌐 Project: 💻 Code: Proud of the team! 🎉

35,295 просмотров

Videos

xieenze_jr's profile picture

Fast-dLLM accelerates multi-modal diffusion VLM LLaDA-V 10 times! 🚀

Enze Xie

11,314 просмотров • 11 месяцев назад

Больше нет контента для загрузки