Video yükleniyor...
Video Yüklenemedi
FlexAttention is a novel compiler-driven programming model that allows implementing the majority of attention variants in a few lines of idiomatic PyTorch code Boyuan Feng & Avik Chaudhuri show how many existing attention variants can be implemented via FlexAttention & that we achieve competitive performance compared to handwritten kernels.... show more
17,929 görüntüleme • 1 yıl önce •via X (Twitter)
3 Yorum

Lakshya1 yıl önce
@Boyuan_Feng @__avik How does it compare to MagiAttention

Jilong | We provide AI marketer - 24/7 marketing1 yıl önce
@Boyuan_Feng @__avik FlexAttention sounds promising! Compilers beat brute force coding.

Vinayak N Baddi1 yıl önce
@Boyuan_Feng @__avik Possible to share the slides @__avik Thanks
