CEO & Co-Founder @river_ai_inc. Previously @xAI, Research & Engineering
Shorts
Grok 2 mini is now 2x faster than it was yesterday. In the last three days Lianmin Zheng and Saeed Maleki rewrote our inference stack from scratch using SGLang ( This has also allowed us to serve the big Grok 2 model, which requires multi-host inference, at a reasonable speed. Both models didn’t just get faster, but also slightly more accurate. Stay tuned for further speed improvements!