Loading video...

Video Failed to Load

Go Home

This completely changes how I look at AI benchmarks. Qwen 3.8 Max Preview felt nearly unusable in OpenCode and Qwen Code. With the Claude Code harness, it suddenly came surprisingly close to Kimi K3 and Fable 5. Same model. Same task. Different harness. Wildly different result. How is this...

30,223 views • 1 month ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos