Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Try now Hyperfast LLM running on custom built GPUs Answers in miliseconds, not seconds How? 🤯

602,378 görüntüleme • 2 yıl önce •via X (Twitter)

10 Yorum

@levelsio profil fotoğrafı
@levelsio2 yıl önce

Via @mattshumer_

Beff – e/acc profil fotoğrafı
Beff – e/acc2 yıl önce

Hyper-optimizing a particular program for a TPU can yield 100×+ speedup. Speaking from experience

Suhail profil fotoğrafı
Suhail2 yıl önce

"An LPU Inference Engine, with LPU standing for Language Processing Unit™, is a new type of end-to-end processing unit system that provides the fastest inference for computationally intensive applications with a sequential component to them, such as AI language applications (LLMs)."

ben sima profil fotoğrafı
ben sima2 yıl önce

Answer: Custom hardware and great software (I work at Groq)

Tony Dinh 🎯 profil fotoğrafı
Tony Dinh 🎯2 yıl önce

Seriously impressed. At this speed I think it's possible to build a human-like conversation experience where the AI can even interrupt you while you're speaking.

@levelsio profil fotoğrafı
@levelsio2 yıl önce

Not affiliated btw, it's made by @JonathanRoss321

Justin / Get Impeached 2! profil fotoğrafı
Justin / Get Impeached 2!2 yıl önce

waiting for someone to get confused between grok and groq

Groq Inc profil fotoğrafı
Groq Inc2 yıl önce

@levelsio No confusion, we trademarked our name in 2016.

Jay Scambler profil fotoğrafı
Jay Scambler2 yıl önce

They developed their own hardware that are built using LPUs instead of GPUs!

Krishiv profil fotoğrafı
Krishiv2 yıl önce

This is insane. What are some novel use-cases now possible because of this speed?

Benzer Videolar