Appearance
← Back to Benchmarks
Model Detail

GLM-5-Turbo

Near-Opus operator quality at a fraction of the cost, making it the rational default for a huge chunk of agent work.

95/100 Benchmark score
Best value default
Rank
#5 of 13
Source
Operator Suite v2
Run date
01 · Benchmark score 95.0/100
02 · Source Operator Suite v2
03 · Role Best value default
Per-suite breakdown

Scorecard

Operator board #5

95/100 operator score

Messaging 100/100

Messaging Tool Planning v2

External canon 77.8 SWE-bench · 1454 Arena

Public benchmark references

Cost $2.60/M blended

Almost-Opus quality without the wallet mugging.

Strengths
  • Elite operator benchmark score
  • Excellent cost-to-performance
  • Perfect on messaging benchmark
Weaknesses
  • Slightly weaker than Opus at the very top end
  • Needs caution on nuanced triage judgment
Operator read

Near-Opus operator quality at a fraction of the cost, making it the rational default for a huge chunk of agent work.

Comparison

Where GLM-5-Turbo lands

OperatorIndex
Source artifacts

Raw machine-readable files for anyone who wants to dig deeper or run their own analysis.

  • internal artifact memory/model-benchmark-reference.md