
OmedTheVibeCoder
@OmedVibeCodes • 1,229 subscribers
🇩🇪 Software Developer | Extreme AI model testing, honest results, no hype | Building something that could change game development 🎮
Shorts
Videos

This is probably my most important post about DeepSeek V4 Flash. Do NOT judge this model after testing only one harness. Pi, Claude Code, and OpenCode produced dramatically different results—and OpenCode completely changed my verdict. You really need to read the details below.
OmedTheVibeCoder66,795 görüntüleme • 6 gün önce

Okay… DeepSeek V4 Flash actually COOKED 😭 With OpenCode, I’d place it around GPT-5.6 Terra Max—and definitely above Luna Max in this test. Almost unlimited usage for around €9, by the way. This model is a MASSIVE success. Just don’t use the wrong harness.
OmedTheVibeCoder35,417 görüntüleme • 6 gün önce

DeepSeek V4 Flash completely failed this test. Compared to Opus 4.8, GPT-5.6 Sol, Kimi K3, and the other current top models, it’s honestly not good. The gap is massive. Against Luna, however, it performs similarly well. Both have different strengths and weaknesses, so I still need to test them more. The weird part: the Pi harness performed worse than OpenCode for DeepSeek. More tests are coming, guys. Don’t worry
OmedTheVibeCoder25,844 görüntüleme • 7 gün önce

Guys, stop underestimating the harness. GPT-5.6 Sol through Claude Code vs Codex produced a GIGANTIC difference in my benchmark. But Opus 5 is the real innovation: High, xHigh, and Max don’t just feel like different thinking levels—they feel like COMPLETELY DIFFERENT MODELS. It’s basically like getting three or four Opus models in one.
OmedTheVibeCoder35,970 görüntüleme • 13 gün önce
Daha fazla içerik yok.