Chinese AI company Z.AI has released GLM-5.1, a 754-billion parameter open-source model designed for autonomous software engineering tasks, achieving a score of 58.4 on SWE-Bench Pro and outperforming OpenAI's GPT-5.4, Anthropic's Opus 4.6, and Google's Gemini 3.1 Pro according to Z.AI's benchmarks. The model can run autonomously for up to 8 hours and sustained performance over 600 iterations on a vector database optimization task, reaching 21,500 queries per second—six times better than single-session results. Released under the MIT License with full model weights available, GLM-5.1 offers enterprises cost control and data sovereignty advantages, though analyst Pareekh Jain notes its Chinese origins may raise compliance concerns for some US companies. The release signals an industry shift from prompt-based AI tools toward systems capable of handling extended, multi-step assignments with minimal supervision, though critics like Lian Jye Su caution that benchmark performance may not reflect real-world enterprise environments with legacy systems and complex workflows.
Story comments
Loading comments…