Z.ai's flagship model and the most capable open-weights model for coding, with major gains on long-horizon agentic tasks.
| Benchmark | GLM-5.3 | GLM-5.2 |
|---|---|---|
| Terminal Bench 2.1 | 88.2 | 81.0 |
| Terminal Bench 3.0 | 28.3 | 4.6 |
| DeepSWE (v1.1) | 66.9 | 46.2 |
| NL2Repo | 58.0 | 48.9 |
| ProgramBench (Almost Solved) | 19.0 | 9.5 |
| FrontierSWE | 78.1 | 67.5 |
| SWE-Marathon (v1.1) | 42.5 | 19.4 |
| PostTrainBench | 39.8 | 31.7 |