正在加载视频...

视频加载失败

DeepSeek v4.1 Flash via API is almost TOO fast 🤯

137,788 次观看 • 6 天前 •via X (Twitter)

46 条评论

keys 🧪 的头像
keys 🧪6 天前

Actually don’t want the model to run too fast then you have keep scrolling up to see what happened it’s annoying after a while

Mia 的头像
Mia6 天前

Good problem to have imo

hkdom 的头像
hkdom6 天前

You will wow more if you run it with DSH

Mia 的头像
Mia6 天前

You're probably right

Santanu Sinha 的头像
Santanu Sinha6 天前

Genuine question.. what harness is that and why is it flickering so much? Also .. yes amazing speed

Mia 的头像
Mia6 天前

Hermes agent CLI

Santanu Sinha 的头像
Santanu Sinha6 天前

Ok thanks

kaith 的头像
kaith6 天前

How’s it’s performance?

Mia 的头像
Mia6 天前

Looking better than v4 flash

Consigliere 的头像
Consigliere6 天前

via Cloud API?

Mia 的头像
Mia6 天前

Yes

Gaurav Bhatia 的头像
Gaurav Bhatia6 天前

Imagine after thinking for that long it doesn’t complete the work.

Mia 的头像
Mia6 天前

But it did 😀

Wassollichhier 的头像
Wassollichhier6 天前

is vision onboard?

Mia 的头像
Mia6 天前

Not on this one, it's beta

Samuel McHargue 的头像
Samuel McHargue6 天前

@grok how much is v4.1 flash compared to the older one.

Bear 的头像
Bear6 天前

Vision support?

Daniel Carneiro 的头像
Daniel Carneiro6 天前

WTF is that ugly TUI?

nah 的头像
nah6 天前

Imagine every model being this fast. That'd be amazing.

Steven Cheng 的头像
Steven Cheng6 天前

Latency that low feels like cheating.

base 的头像
base6 天前

get in my spark … lol

CV.YH 的头像
CV.YH6 天前

Wow

Chris Winslow 的头像
Chris Winslow6 天前

Speed ≠ Quality?

Saurav 的头像
Saurav6 天前

actually useful?

Mia 的头像
Mia6 天前

Seems like it's better than v4 flash

bitflipgremlin 的头像
bitflipgremlin6 天前

this isn't quite mercury-2 speed but it's damn fucking close, and quality leads me to suspect it's a different mechanism. i wonder what kind of bullshit the whale came up with now

Ian Hailey 的头像
Ian Hailey6 天前

It can never be too fast!

dan0mad 的头像
dan0mad6 天前

There’s 4.1???!? Since wen?

forreal 的头像
forreal6 天前

better than glm 5.3 flash? it only matters if it reaches performance of kimi k3 or glm 5.3.

Devin Oldenburg 的头像
Devin Oldenburg6 天前

~176 tokens/sec

Redmix 的头像
Redmix6 天前

It's a shame we can't enjoy it for very long.

Mia 的头像
Mia6 天前

Squeezing what I can !

Redmix 的头像
Redmix6 天前

✌🏻

Cato Nooka 的头像
Cato Nooka6 天前

if there is another model break through for this kind of speed, I don't know what to say, not only so fast at 400 TPS, but also more token efficient, the time for each task is shink down to minutes

Yume_X 的头像
Yume_X6 天前

Yeah I was shocked too when I saw it , makes a good argument for 300 tok/s builds when you see it lol

CryptoYeti 的头像
CryptoYeti6 天前

@grok wie teuer ist die API von der neuste Model von Deepseek? Und vergleiche sie mit Muse spark corbitor .

Jayden 的头像
Jayden6 天前

The latency is fun. The bill is the actual benchmark.

Mike 的头像
Mike6 天前

I really wish Nous would fix the CLI redraw flashing in Hermes. That’s mainly why I stick to the TUI.

remakw 的头像
remakw6 天前

fastest model i gave ever used switched the trace off because my eyes were hurting lol

rapidTools / Mike 的头像
rapidTools / Mike6 天前

o_0... Dafuq... I hope we get this open as well. It would speed up local flash model az well I guess. (not to this speed obviously, but making it even faster wpuld be cool)

Tom 的头像
Tom6 天前

Why I only see v40730? Where is 4.1?

Steven Cheng 的头像
Steven Cheng6 天前

Almost" is doing a lot of heavy lifting there. Latency spikes still kill the vibe when you are chaining calls.

Reelix 的头像
Reelix6 天前

Do a speed comparison :p

Cave 的头像
Cave6 天前

Woo.

Jarno 的头像
Jarno6 天前

Fast until you hit the rate limit, then it's back to staring at a spinner like every other model.

Akshay Joshi 的头像
Akshay Joshi6 天前

Thought it’s free and went rampant and burned 96million token from Deekseek platform hopefully 97% cache hit hence 1.30$ only

相关视频