
GLM-5.1: Towards Long-Horizon Tasks
GLM-5.1 is a model with enhanced coding capabilities that achieves state-of-the-art 58.4% on SWE-Bench Pro, outperforming GLM-5, GPT-5.4, Opus 4.6, an...
A collection of bookmarks filtered by the tag "Performance Benchmarks".
Showing 1-2 of 2 bookmarks tagged with "Performance Benchmarks"
All Stories