Загрузка видео...
Не удалось загрузить видео
FlexAttention is a novel compiler-driven programming model that allows implementing the majority of attention variants in a few lines of idiomatic PyTorch code Boyuan Feng & Avik Chaudhuri show how many existing attention variants can be implemented via FlexAttention & that we achieve competitive performance compared to handwritten kernels.... show more
17,929 просмотров • 1 год назад •via X (Twitter)
Комментарии: 3

Lakshya1 год назад
@Boyuan_Feng @__avik How does it compare to MagiAttention

Jilong | We provide AI marketer - 24/7 marketing1 год назад
@Boyuan_Feng @__avik FlexAttention sounds promising! Compilers beat brute force coding.

Vinayak N Baddi1 год назад
@Boyuan_Feng @__avik Possible to share the slides @__avik Thanks
