Loading video...
Video Failed to Load
Holy moly: GLM-5.3 got much better in cybersecurity since our pre-release evaluation with Z.ai. It now matches GPT-5.6-Sol on our cybersecurity benchmark at 0.4x the cost 🤯 - At pass@1: it went from 60.4% to 65.6% CVEs rediscovered, crushing every other open model on one-shot tasks - At pass@3:... show more
34,641 views • 1 month ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
