Speed analysis
Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.
Quality scores
Evaluation results from judge-model scoring across diverse task categories. Scores reflect coherence, accuracy and instruction-following.
Tokens per second
Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.
Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.
Capabilities
Availability
Availability
No measurements yet
We haven't recorded enough API calls to show availability stats for this model. Data appears once the model starts receiving live traffic.
Tokonomix benchmark verdicts
CogView-4 maintains steady performance across image generation benchmarks
CogView-4 continues to demonstrate consistent performance in the current benchmark window, maintaining the same capability profile as its debut. The model shows no significant changes across key image generation metrics, suggesting a stable release without major updates during this period. Users can expect reliable image generation functionality that remains competitive within its tier. The model's performance characteristics have remained unchanged, indicating a focus on stability rather than rapid iteration. This consistency may benefit users seeking predictable outputs for production workflows. For those evaluating CogView-4, the current benchmarks confirm that its strengths and limitations from the previous window remain intact. The lack of benchmark fluctuations suggests either model stability or minimal real-world usage variation during this measurement period. Organizations already using CogView-4 should not expect different performance characteristics from earlier deployments. The steady state of capabilities means CogView-4 continues to serve its original use cases without expansion into new modalities or significant performance improvements. Users seeking cutting-edge advances may want to monitor future benchmark windows for potential updates, while those prioritizing consistency will find the current stability reassuring.
Quality
—
Latency p50
—
Test runs
0
CogView-4
by Z.ai (GLM / Zhipu)
- Context window
- — tokens
- Input price
- — / 1M
- Output price
- — / 1M
- Tier
- Tier B — Production
- Modality
- Text
- API type
- REST · streaming
- Benchmark runs
- 56
More from Z.ai (GLM / Zhipu)