Loading video...
Video Failed to Load
I wrote GPU kernels for GPT2 124M inference engine from scratch in ~1500 lines of CUDA. Here I explain it simply.
157,478 views • 6 days ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
