Appearance
← Back to Benchmarks
Model Detail

Qwen 3.6 27B NVFP4

Best current 27B result of the three, but still too soft to trust as a default local operator model.

40/100 Benchmark score
Best of the weak pass
Rank
#11 of 13
Source
Enterprise oMLX
Run date
Avg latency
17.44s
01 · Benchmark score 40/100
02 · Average latency 17.44s
03 · Role Best of the weak pass
Per-suite breakdown

Scorecard

Operator board #13

40/100 operator score

Messaging Quick execution pack

Messaging Tool Planning v2

External canon First internal 27B oMLX pass

Public benchmark references

Cost Enterprise local

Best current 27B result of the three, but still too soft to trust as a default local operator model.

Strengths
  • Tied top score in this 27B run
  • Live on oMLX/OpenClaw
  • Clean no-thinking API path
Weaknesses
  • Only 40/100 on the quick pack
  • Missed multiple operator-judgment tasks
  • Needs rerun/tuning before routing decisions
Operator read

Best current 27B result of the three, but still too soft to trust as a default local operator model.

Comparison

Where Qwen 3.6 27B NVFP4 lands

OperatorIndex
Source artifacts

Raw machine-readable files for anyone who wants to dig deeper or run their own analysis.

  • internal artifact output/benchmarks/2026-04-25-qwen36-27b-api/README.md
  • internal artifact output/benchmarks/2026-04-25-qwen36-27b-api/results-raw.json
  • internal artifact output/benchmarks/2026-04-25-qwen36-27b-api/results-scored.json