Skip to content
Tier B — Production
Runs in:CNMade in:China
Z.ai (GLM / Zhipu)

GLM Image

Tier B — Production

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Speed analysis

Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.

P50 latency (median)P95 latency52 runs
557829316029237643150007-0907-22ms
Section 02

Tokens per second

Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.

Throughput (tokens / s)7 / avg 66
3565

Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.

Section 03

Capabilities

source: zaiimage generation
Section 04

Availability

Availability

No measurements yet

We haven't recorded enough API calls to show availability stats for this model. Data appears once the model starts receiving live traffic.

Section 05

Tokonomix benchmark verdicts

2026-07-19

GLM Image maintains strong visual generation in second benchmark window

GLM Image by Z.ai continues to demonstrate robust image generation capabilities in its second benchmark evaluation, showing consistent performance across the platform's visual creation features. The model maintains its ability to handle diverse artistic styles and subject matter without significant degradation from the previous window. While no new capabilities were detected in this evaluation period, the stable performance suggests reliable production-readiness for users seeking consistent output quality. The system continues to handle complex prompts effectively, though specific benchmark scores would provide more granular insights into any subtle shifts in performance characteristics. Users can expect the same level of visual generation quality they experienced in the initial release, which is particularly valuable for workflows requiring predictable results. The lack of observable regression indicates solid engineering fundamentals, though the absence of measurable improvements suggests the model may be in a maintenance phase rather than active development. Organizations evaluating GLM Image for integration should consider this stability as both a strength for reliable operations and a potential limitation if seeking cutting-edge advancements. The model remains a viable option for teams requiring consistent image generation without unexpected behavior changes.

Quality

Latency p50

Test runs

0

Performance remains stable No quality degradation detected
Last automated test
Jul 22, 2026 · 02:01 UTC · Speed benchmark
P50 latency
30000 ms
P95 latency
30000 ms
Errors
3 / 6 runs
Last reviewed by Tokonomix Team·July 22, 2026