Compare / head-to-head

GLM-4.1V-9B-ThinkingvsGLM-4.5V

GLM-4.5V leads 13 of 13 shared benchmarks. Both offer a 64K-token context window.

Benchmarks from cited public sources; pricing from official pages; status from official provider feeds.

Shared benchmarks
0 – 13
GLM-4.5V leads
Cheaper per token
—
list price, input + output
Larger context
Tie
both 64K tokens
Providers
Z.ai
same provider
GLM-4.1V-9B-Thinking
Z.ai · released 2025-07-01
textvisionvideo
Context
64K
Max out
—
Input /1M
—
Output /1M
—
Cached /1M
—
Scores
13 · 0 core
GLM-4.5V
Z.ai · released 2025-08-11
textvisionvideo
Context
64K
Max out
16K
Input /1M
$0.6 ↗
Output /1M
$1.8 ↗
Cached /1M
$0.11
Scores
17 · 2 core
Quality

Benchmark matrix

BenchmarkGLM-4.1V-9B-ThinkingGLM-4.5VΔEdge
Reported by both · 13
AndroidWorld41.7% ↗57% ↗-15.3 ptGLM-4.5V
ChartQAPro59.5% ↗64% ↗-4.5 ptGLM-4.5V
Design2Code64.7% ↗82.2% ↗-17.5 ptGLM-4.5V
MathVision54.4% ↗65.6% ↗-11.2 ptGLM-4.5V
MathVista80.7% ↗84.6% ↗-3.9 ptGLM-4.5V
MMBench v1.185.8% ↗88.2% ↗-2.4 ptGLM-4.5V
MMMU (val)68% ↗75.4% ↗-7.4 ptGLM-4.5V
MMMU Pro57.1% ↗65.2% ↗-8.1 ptGLM-4.5V
MMStar72.9% ↗75.3% ↗-2.4 ptGLM-4.5V
OCRBench84.2% ↗86.5% ↗-2.3 ptGLM-4.5V
OSWorld14.9% ↗35.8% ↗-20.9 ptGLM-4.5V
VideoMME (w/o sub)68.2% ↗74.6% ↗-6.4 ptGLM-4.5V
VideoMMMU61% ↗72.4% ↗-11.4 ptGLM-4.5V
Only GLM-4.1V-9B-Thinking reports · 0
GLM-4.1V-9B-Thinking reports nothing GLM-4.5V does not.
Only GLM-4.5V reports · 4
LMArena Elonot reported1352.4 ↗——
LMArena Vision (LMArena)not reported1152.7 ↗——
MMLongBench-Docnot reported44.7% ↗——
WebVoyagerSomnot reported84.4% ↗——
Scores tagged "epoch" or "matharena" are independent runs, used only where the lab has not published its own; ⚠ marks rows MathArena flags as released after the competition. Higher is better on every row. Δ is GLM-4.1V-9B-Thinking minus GLM-4.5V in the benchmark's own unit. "Not reported" means the lab has not published that figure; it is not a zero. ↗ opens the source.
Specs & pricing

Side by side

SpecGLM-4.1V-9B-ThinkingGLM-4.5VEdge
Context window64K tokens64K tokensTie
Max output—16K tokens—
Input price / 1M—$0.6 ↗—
Output price / 1M—$1.8 ↗—
Cached input / 1M—$0.11 ↗—
Input + output / 1M
Lower is cheaper. List prices; batch, tool and regional fees excluded.
—$2.4—
Modalitiestext · vision · videotext · vision · videoTie
Released2025-07-012025-08-11—
Cited benchmark scores1317—
Reliability

Z.ai status

All providers →
Built by Respan
Which one wins on your data?

Public benchmarks are a starting point. Run GLM-4.1V-9B-Thinking and GLM-4.5V on your own prompts with Respan evals, or route to either through one gateway key with automatic failover.