
Atharva Ingle
@AtharvaIngle7 • 2,304 subscribers
AI @e2enetworks (building @jarvislabsai) , @kaggle Competition Expert, Dev Expert @weights_biases
Shorts
Videos

GLM 5.2 in pi running at almost 100 tokens/sec on 8x H200. I really like this model. It feels a frontier-level model you can actually self-host. Next step: squeeze more out of it by benchmarking different quants, tuning the stack, and figuring out the minimum GPU setup people can realistically run it on. Expect more on this soon.
Atharva Ingle26,751 görüntüleme • 2 ay önce
Daha fazla içerik yok.