GLM-5.1 Tops Open-Source Models in Coding Agent Benchmark

GLM-5.1 has emerged as the leading open-source model in the Artificial Analysis Coding Agent Benchmark, according to a report by Artificial Analysis. The benchmark evaluates model performance on three key tests: SWE-Bench-Pro-Hard-AA, Terminal-Bench v2, and SWE-Atlas-QnA, which simulate real-world programming and technical tasks. While the proprietary Opus 4.7 model secured the top global position, GLM-5.1, operating on Claude Code, led among open-source models, showcasing its advanced capabilities in programming agent scenarios.

Source: Show Original

Disclaimer: The content provided on Phemex News is for informational purposes only. We do not guarantee the quality, accuracy, or completeness of the information sourced from third-party articles. The content on this page does not constitute financial or investment advice. We strongly encourage you to conduct you own research and consult with a qualified financial advisor before making any investment decisions.

You may also like