Loading video...

Video Failed to Load

Go Home

DeepSeek V4.1 Flash is now in Codex!! The pricing is insane. $10 on OpenCode Go gets you 26,000 DeepSeek requests every 5 hours (they just 4x'd it) for 72 hours. If you ran the same usage on the DeepSeek API it would cost you about $12, so one afternoon...

164,733 views • 17 hours ago •via X (Twitter)

23 Comments

Ziwen's profile picture
Ziwen17 hours ago

here is the repo if anyone wants to try it.

Ewgeni's profile picture
Ewgeni17 hours ago

imho not a bad model but when it comes to token use and image recognition, i cannot recommend it. Tried a few things on my Chessboard Bench but was not able achieve any good results. It overthinks too much, produces a lot of tokens and delivers bad.

Joe's profile picture
Joe16 hours ago

Sol level for coding I find that hard to believe

Bk1man's profile picture
Bk1man15 hours ago

Great math — and it holds after the promo too: OpenRouter lists deepseek-v4-flash at $0.089/M in, $0.177/M out with 1M ctx. Even the 26k-requests-per-5h habit stays cheap on the API after the 72h window ends. Receipts:

Ziwen's profile picture
Ziwen14 hours ago

Using open code is worth it ain't gonna lie.

Rahul's profile picture
Rahul16 hours ago

is this Mac app?

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack17 hours ago

@ziwenxu_ that's crazy value. I remember paying way more for something similar last year.

Noah Muller's profile picture
Noah Muller15 hours ago

Is the model picker also accessible with remote connection?

Ziwen's profile picture
Ziwen14 hours ago

Yes

Argona's profile picture
Argona17 hours ago

deepseek keeps making these models cheaper than they should be i switched my subagents over already

Asim Khan's profile picture
Asim Khan15 hours ago

Subagent support is what makes this actually useful, not just cheap filler when the main model hits a wall.

Ziwen's profile picture
Ziwen14 hours ago

That's the part I use the most

AdiiX's profile picture
AdiiX17 hours ago

swapping to flash mid loop when the main model hits a limit means handing off state to a different model with different habits mid task, not just continuing the same reasoning.

J A Z I I's profile picture
J A Z I I17 hours ago

bro da goat, everything is in codex now

Umar Saeed's profile picture
Umar Saeed16 hours ago

The API is great.

Herb Page's profile picture
Herb Page14 hours ago

quantized to hell

frenetic | golden era's profile picture
frenetic | golden era14 hours ago

what counts as long runs? how good would it be at orchestration? as the manager and subagent?

Andy Scott's profile picture
Andy Scott16 hours ago

what is this UI for codex router??

catman's profile picture
catman14 hours ago

the 72-hour window is the weird part here—does the subscription still make sense once that promo ends?

Md Fahim's profile picture
Md Fahim17 hours ago

That's a pretty sweet deal! Seems like a no-brainer if

jogn's profile picture
jogn14 hours ago

Tf is codex router

David Hampl's profile picture
David Hampl14 hours ago

This is why sub complainers confuse me. Heavy users get 10x vs API. Just test if it codes well.

Zenkun's profile picture
Zenkun15 hours ago

Why doesn't that interface appear on my Codex?

Related Videos