Loading video...
Video Failed to Load
Big win for open-source LLMs! DeepSeek V4 Pro holds the top open-weights score on SWE-bench Verified, in the GPT-5.5 range. GLM 5.2 leads the open-weight intelligence index and sits near the closed frontier on long-horizon coding. But this leaderboard number is a weak proxy for real performance. It comes... show more
44,124 views • 1 month ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
