Benchmark
PostTrainBench
Benchmark by PostTrainBench team
PostTrainBench v1.1 ยท September 2026 results
31.70%
GLM 5.2, Max reasoning / Claude Code
frontier lab
Organization attribution for GLM 5.2; benchmark agent used Claude Code.
Benchmark by PostTrainBench team
PostTrainBench v1.1 ยท September 2026 results
31.70%
GLM 5.2, Max reasoning / Claude Code