Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

DeepSeek v4.1 Flash via API is almost TOO fast 🤯

137,788 Aufrufe • vor 6 Tagen •via X (Twitter)

46 Kommentare

Profilbild von keys 🧪
keys 🧪vor 6 Tagen

Actually don’t want the model to run too fast then you have keep scrolling up to see what happened it’s annoying after a while

Profilbild von Mia
Miavor 6 Tagen

Good problem to have imo

Profilbild von hkdom
hkdomvor 6 Tagen

You will wow more if you run it with DSH

Profilbild von Mia
Miavor 6 Tagen

You're probably right

Profilbild von Santanu Sinha
Santanu Sinhavor 6 Tagen

Genuine question.. what harness is that and why is it flickering so much? Also .. yes amazing speed

Profilbild von Mia
Miavor 6 Tagen

Hermes agent CLI

Profilbild von Santanu Sinha
Santanu Sinhavor 6 Tagen

Ok thanks

Profilbild von kaith
kaithvor 6 Tagen

How’s it’s performance?

Profilbild von Mia
Miavor 6 Tagen

Looking better than v4 flash

Profilbild von Consigliere
Consiglierevor 6 Tagen

via Cloud API?

Profilbild von Mia
Miavor 6 Tagen

Yes

Profilbild von Gaurav Bhatia
Gaurav Bhatiavor 6 Tagen

Imagine after thinking for that long it doesn’t complete the work.

Profilbild von Mia
Miavor 6 Tagen

But it did 😀

Profilbild von Wassollichhier
Wassollichhiervor 6 Tagen

is vision onboard?

Profilbild von Mia
Miavor 6 Tagen

Not on this one, it's beta

Profilbild von Samuel McHargue
Samuel McHarguevor 6 Tagen

@grok how much is v4.1 flash compared to the older one.

Profilbild von Bear
Bearvor 6 Tagen

Vision support?

Profilbild von Daniel Carneiro
Daniel Carneirovor 6 Tagen

WTF is that ugly TUI?

Profilbild von nah
nahvor 6 Tagen

Imagine every model being this fast. That'd be amazing.

Profilbild von Steven Cheng
Steven Chengvor 6 Tagen

Latency that low feels like cheating.

Profilbild von base
basevor 6 Tagen

get in my spark … lol

Profilbild von CV.YH
CV.YHvor 6 Tagen

Wow

Profilbild von Chris Winslow
Chris Winslowvor 6 Tagen

Speed ≠ Quality?

Profilbild von Saurav
Sauravvor 6 Tagen

actually useful?

Profilbild von Mia
Miavor 6 Tagen

Seems like it's better than v4 flash

Profilbild von bitflipgremlin
bitflipgremlinvor 6 Tagen

this isn't quite mercury-2 speed but it's damn fucking close, and quality leads me to suspect it's a different mechanism. i wonder what kind of bullshit the whale came up with now

Profilbild von Ian Hailey
Ian Haileyvor 6 Tagen

It can never be too fast!

Profilbild von dan0mad
dan0madvor 6 Tagen

There’s 4.1???!? Since wen?

Profilbild von forreal
forrealvor 6 Tagen

better than glm 5.3 flash? it only matters if it reaches performance of kimi k3 or glm 5.3.

Profilbild von Devin Oldenburg
Devin Oldenburgvor 6 Tagen

~176 tokens/sec

Profilbild von Redmix
Redmixvor 6 Tagen

It's a shame we can't enjoy it for very long.

Profilbild von Mia
Miavor 6 Tagen

Squeezing what I can !

Profilbild von Redmix
Redmixvor 6 Tagen

✌🏻

Profilbild von Cato Nooka
Cato Nookavor 6 Tagen

if there is another model break through for this kind of speed, I don't know what to say, not only so fast at 400 TPS, but also more token efficient, the time for each task is shink down to minutes

Profilbild von Yume_X
Yume_Xvor 6 Tagen

Yeah I was shocked too when I saw it , makes a good argument for 300 tok/s builds when you see it lol

Profilbild von CryptoYeti
CryptoYetivor 6 Tagen

@grok wie teuer ist die API von der neuste Model von Deepseek? Und vergleiche sie mit Muse spark corbitor .

Profilbild von Jayden
Jaydenvor 6 Tagen

The latency is fun. The bill is the actual benchmark.

Profilbild von Mike
Mikevor 6 Tagen

I really wish Nous would fix the CLI redraw flashing in Hermes. That’s mainly why I stick to the TUI.

Profilbild von remakw
remakwvor 6 Tagen

fastest model i gave ever used switched the trace off because my eyes were hurting lol

Profilbild von rapidTools / Mike
rapidTools / Mikevor 6 Tagen

o_0... Dafuq... I hope we get this open as well. It would speed up local flash model az well I guess. (not to this speed obviously, but making it even faster wpuld be cool)

Profilbild von Tom
Tomvor 6 Tagen

Why I only see v40730? Where is 4.1?

Profilbild von Steven Cheng
Steven Chengvor 6 Tagen

Almost" is doing a lot of heavy lifting there. Latency spikes still kill the vibe when you are chaining calls.

Profilbild von Reelix
Reelixvor 6 Tagen

Do a speed comparison :p

Profilbild von Cave
Cavevor 6 Tagen

Woo.

Profilbild von Jarno
Jarnovor 6 Tagen

Fast until you hit the rate limit, then it's back to staring at a spinner like every other model.

Profilbild von Akshay Joshi
Akshay Joshivor 6 Tagen

Thought it’s free and went rampant and burned 96million token from Deekseek platform hopefully 97% cache hit hence 1.30$ only

Ähnliche Videos