Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

17,255 Aufrufe • vor 3 Tagen •via X (Twitter)

36 Kommentare

Profilbild von Sufyan
Sufyanvor 3 Tagen

500

Profilbild von Miyuru
Miyuruvor 3 Tagen

locking in 900. now tell me the model and how many cards are coughing up smoke behind it

Profilbild von Meet Limbani
Meet Limbanivor 3 Tagen

i know that scrolling speed in PI, it's around 350

Profilbild von Gil
Gilvor 3 Tagen

750

Profilbild von Neo
Neovor 3 Tagen

420

Profilbild von 0xSero
0xSerovor 3 Tagen

420.69

Profilbild von Decatalyst 🌱
Decatalyst 🌱vor 3 Tagen

OH MY TOMATOES I FEEL LIKE A CAVEMAN USING OPUS 5 NOW 😭

Profilbild von ruleryak
ruleryakvor 3 Tagen

went from 24k output at 7s to 44k output at 34s so end to end throughput including tool calls, prefill, and decode is around 750 tps there. So the actual decode is probably 1k+

Profilbild von Maciej
Maciejvor 3 Tagen

No tool call streaming? They spend bilions on hardware and lack money for API software.

Profilbild von Valentin Yanakiev
Valentin Yanakievvor 3 Tagen

384

Profilbild von D
Dvor 3 Tagen

450

Profilbild von pythongiant
pythongiantvor 3 Tagen

a majillion

Profilbild von PaulNL
PaulNLvor 3 Tagen

400ish ?

Profilbild von Mike Reese
Mike Reesevor 3 Tagen

How does cerebras compare to your local setup with deepseek 4.1 flash in terms of tok/s?

Profilbild von AArchimedes64
AArchimedes64vor 3 Tagen

1024

Profilbild von cherki
cherkivor 3 Tagen

How?

Profilbild von Arman
Armanvor 3 Tagen

1337

Profilbild von Fraser Price
Fraser Pricevor 3 Tagen

Bout tree fiddy

Profilbild von Crispybits || Parroty account
Crispybits || Parroty accountvor 3 Tagen

3 - you've sped up the AI writing sections of the video 1000x 😹

Profilbild von Redmix
Redmixvor 3 Tagen

~1850

Profilbild von Wayne
Waynevor 3 Tagen

I bet that's at least 1t/s

Profilbild von Serhii Y.
Serhii Y.vor 3 Tagen

you're typing 1-2tts, model 100+

Profilbild von mo
movor 3 Tagen

320

Profilbild von Kutluk
Kutlukvor 3 Tagen

~300-400?

Profilbild von Jason Frisch
Jason Frischvor 3 Tagen

What do I need to run GLM-5.3-flash? Can a Mac Studio manage it for regular coding tasks?

Profilbild von Loco Legend
Loco Legendvor 3 Tagen

170-250

Profilbild von Sektur Toilet
Sektur Toiletvor 3 Tagen

3-400?

Profilbild von Jeff Steve
Jeff Stevevor 3 Tagen

700

Profilbild von Simao Soares 🇪🇺🇵🇹🇳🇱
Simao Soares 🇪🇺🇵🇹🇳🇱vor 3 Tagen

I've seen a tok/s, is there a metric for token quality? I'm honestly curious. So many tools to reduce token usage like context compression, context understanding, prompt engineering for summarised high quality replies, MTP and non-MTP...

Profilbild von saltyclaw707
saltyclaw707vor 3 Tagen

300 ish

Profilbild von vveerrgg
vveerrggvor 3 Tagen

280 T/Sec. … how are you configuring your cards … that’s the question I have

Profilbild von Petey struggles
Petey strugglesvor 3 Tagen

1500

Profilbild von Matt
Mattvor 3 Tagen

700+

Profilbild von Asif Sheikh 💸
Asif Sheikh 💸vor 3 Tagen

~1,850 tokens/s you're welcome.

Profilbild von byovd
byovdvor 3 Tagen

300

Profilbild von Kachow Towmater
Kachow Towmatervor 3 Tagen

280

Ähnliche Videos