正在加载视频...

视频加载失败

Try now Hyperfast LLM running on custom built GPUs Answers in miliseconds, not seconds How? 🤯

602,378 次观看 • 2 年前 •via X (Twitter)

10 条评论

@levelsio 的头像
@levelsio2 年前

Via @mattshumer_

Beff – e/acc 的头像
Beff – e/acc2 年前

Hyper-optimizing a particular program for a TPU can yield 100×+ speedup. Speaking from experience

Suhail 的头像
Suhail2 年前

"An LPU Inference Engine, with LPU standing for Language Processing Unit™, is a new type of end-to-end processing unit system that provides the fastest inference for computationally intensive applications with a sequential component to them, such as AI language applications (LLMs)."

ben sima 的头像
ben sima2 年前

Answer: Custom hardware and great software (I work at Groq)

Tony Dinh 🎯 的头像
Tony Dinh 🎯2 年前

Seriously impressed. At this speed I think it's possible to build a human-like conversation experience where the AI can even interrupt you while you're speaking.

@levelsio 的头像
@levelsio2 年前

Not affiliated btw, it's made by @JonathanRoss321

Justin / Get Impeached 2! 的头像
Justin / Get Impeached 2!2 年前

waiting for someone to get confused between grok and groq

Groq Inc 的头像
Groq Inc2 年前

@levelsio No confusion, we trademarked our name in 2016.

Jay Scambler 的头像
Jay Scambler2 年前

They developed their own hardware that are built using LPUs instead of GPUs!

Krishiv 的头像
Krishiv2 年前

This is insane. What are some novel use-cases now possible because of this speed?

相关视频