Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

MiMo-X-Pro-Preview vs DeepSeek V4.1 Flash the difference in quality is honestly insane, MiMo X Pro Preview, especially when paired with MiMo-Flash-Preview, feels noticeably ahead You can already tell the kind of benchmark scores MiMo must be hitting right now. DeepSeek V4.1 Flash is good and fast, but MiMo feels...

18,359 Aufrufe • vor 1 Tag •via X (Twitter)

15 Kommentare

Profilbild von NOTICIAS IA | Amaliometria
NOTICIAS IA | Amaliometriavor 1 Tag

Perhaps the tru comparison would be against Mimo Flash, not pro. Anyway it is a very useful insight, thanks!!

Profilbild von n0geegee
n0geegeevor 1 Tag

mimo with ds pricing would be a blast

Profilbild von SK
SKvor 1 Tag

DeepSeek is crazy good, but when it comes to aesthetics and visual quality they still have some work to do. But again this is still a flash model.

Profilbild von Iam_
Iam_vor 1 Tag

You're absolutely right ,really fast and efficient.

Profilbild von EDDY VU
EDDY VUvor 1 Tag

What specific use cases or prompts made the quality difference feel that obvious to you?

Profilbild von Hunter Bown
Hunter Bownvor 1 Tag

whoa

Profilbild von Abdelrahman Alkerdawy
Abdelrahman Alkerdawyvor 1 Tag

bro you do the most unique posts fr🔥🔥

Profilbild von Iam_
Iam_vor 1 Tag

Thanks, brother! Waiting for Grok 4.7 to wrap up the week; we've had about one new model a day since September started, haha.

Profilbild von Abdelrahman Alkerdawy
Abdelrahman Alkerdawyvor 1 Tag

yes!!!

Profilbild von jabdbdb djjdj
jabdbdb djjdjvor 1 Tag

这两不是一个赛道的模型

Profilbild von 魏文博
魏文博vor 1 Tag

我不相信,在中国,一致的评价是 MIMO ��别垃圾,但在中国审核过的人很少而且还用不了 Computer use,难道模型不一样?

Profilbild von yang zhou
yang zhouvor 1 Tag

how tps about xiaomi flash?

Profilbild von Ibesh
Ibeshvor 1 Tag

feel is fair but drop the prompts that separated them. mimo ahead on vibes is one thing, same clip rendered on both tells me more than a score prediction

Profilbild von J A Z I I
J A Z I Ivor 1 Tag

wow, super excited for mimo next model

Profilbild von Ruben Garcia Jr
Ruben Garcia Jrvor 1 Tag

You’re the reason the model is slow !!!! lol

Ähnliche Videos

hy3 vs mimo-v2.5 vs deepseek v4 flash vs minimax m3 the four models on top of the openrouter leaderboard by tokens this week: #1 hy3 (Tencent Hy) – 7.5t #2 mimo-v2.5 (Xiaomi MiMo) – 6.56t #3 deepseek v4 flash (DeepSeek) – 5.24t #4 minimax m3 (MiniMax (official)) – 4.21t so we tested them. 3 prompts, single-file html, Three.js from a cdn, fully procedural, no external assets. all run via AI/ML API each prompt is a transparent cutaway machine that has to be mechanically correct, not decorative: • 4-stroke engine with full oil circulation – slider-crank kinematics, cam at 2:1, valve lift driven by lobes, oil loop from sump to gallery to big-end • watt walking-beam steam engine – four-bar vector-loop closure, eccentric-driven slide valve, steam events synced to real port position • francis reaction water turbine – 20 guide vanes on a regulating ring, 17 lofted runner blades, gpu particle advection, precessing vortex rope at part load the takeaway up front: none of the four cleared all three scenes on the first attempt. but the price spread between them is roughly 70x – hy3 fixed included costs less than two cents overall results (summed across all 3 scenes): cost #1 hy3 – $0.016 #2 deepseek v4 flash – $0.025 #3 mimo-v2.5 – $0.97 #4 minimax m3 – $1.17 tokens #1 hy3 – 19,326 #2 deepseek v4 flash – 63,126 #3 mimo-v2.5 – 322,523 #4 minimax m3 – 702,900 lines of code #1 hy3 – 1,047 #2 mimo-v2.5 – 2,759 #3 deepseek v4 flash – 3,273 #4 minimax m3 – 3,354 scenes needing a second attempt #1 hy3 – 1 (engine) #1 mimo-v2.5 – 1 (turbine) #1 minimax m3 – 1 (turbine) #4 deepseek v4 flash – 2 (steam engine, turbine) observations: 1. the token spread is the real story – minimax burns 36x hy3's tokens and lands in the same place, one retry, ~3.3k lines 2. hy3 is the outlier on density: 1,047 lines total, fewest tokens, cheapest run, and only one scene needed a second pass. deepseek is the opposite trade – near-hy3 pricing but the most retries 3. mimo and minimax seem to overthink instead of writing the code. minimax spent 359.1k tokens on the steam engine and produced 1,346 lines – the tokens are going somewhere other than the file 4. the francis turbine broke three of the four. the spec that separates them is the one with 20 linked guide vanes and gpu particle advection, not the one with the most parts overall impression: none of these models excelled at any of the tasks we gave them. but they were close, and they were extremely cheap. the gap that matters isn't quality anymore – it's that hy3 ran all three scenes for less than two cents while the frontier labs charge dollars for the same work right now you pick these because they're good for the zero price you pay. soon that's something openai and anthropic will have to think about follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

17,145 Aufrufe • vor 1 Monat