Video wird geladen...
Video konnte nicht geladen werden
FlexAttention is a novel compiler-driven programming model that allows implementing the majority of attention variants in a few lines of idiomatic PyTorch code Boyuan Feng & Avik Chaudhuri show how many existing attention variants can be implemented via FlexAttention & that we achieve competitive performance compared to handwritten kernels.... show more
17,929 Aufrufe • vor 1 Jahr •via X (Twitter)
3 Kommentare

Lakshyavor 1 Jahr
@Boyuan_Feng @__avik How does it compare to MagiAttention

Jilong | We provide AI marketer - 24/7 marketingvor 1 Jahr
@Boyuan_Feng @__avik FlexAttention sounds promising! Compilers beat brute force coding.

Vinayak N Baddivor 1 Jahr
@Boyuan_Feng @__avik Possible to share the slides @__avik Thanks
