正在加载视频...

视频加载失败

This completely changes how I look at AI benchmarks. Qwen 3.8 Max Preview felt nearly unusable in OpenCode and Qwen Code. With the Claude Code harness, it suddenly came surprisingly close to Kimi K3 and Fable 5. Same model. Same task. Different harness. Wildly different result. How is this...

30,223 次观看 • 1 个月前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频