Benchmarks

Compare AI models

Pick two models to see release dates, benchmark scores, and community ratings side by side.

Open page →
Provider
Z.ai
-
Out yet?
Yes
-
Status
Available
-
Release date
Jun 13, 2026
-
LM Arena?LM Arena's Elo-style rating from blind head-to-head votes: people compare two anonymous model answers and pick the better one. Higher is better.
1470.9620912904825Verify at LM ArenaAug 8, 2026
-
Agentic?BenchLM.ai's score for multi-step agentic work - planning, tool use, and acting autonomously, normalized 0–100 across multiple benchmarks. Higher is better.
58.16Verify at BenchLM.aiAug 8, 2026
-
Coding?BenchLM.ai's score for code generation and software-engineering tasks, normalized 0–100 across multiple benchmarks. Higher is better.
64.31Verify at BenchLM.aiAug 8, 2026
-
InstructionFollowing?BenchLM.ai's score for following precise, detailed instructions, normalized 0–100 across multiple benchmarks. Higher is better.
81Verify at BenchLM.aiJul 21, 2026
-
Knowledge?BenchLM.ai's score for factual knowledge and question answering, normalized 0–100 across multiple benchmarks. Higher is better.
82.5Verify at BenchLM.aiAug 8, 2026
-
SWE-Bench Pro?SWE-Bench Pro: resolving real GitHub issues in large codebases end to end. Score is the percentage of issues resolved.Hugging Face
62.1Verify at Hugging FaceAug 8, 2026
-
Vibe rating?OutYet's community rating: signed-in users score the model 1–10. Shown as the average and the number of votes.
8.0 / 10 · 2
-
Predecessor
-
-
Successor
-
-

Benchmark scores are mirrored from third-party sources and captured on the dates shown. Numbers from different benchmarks, sources, or capture dates are not directly comparable.

Data from BenchLM.ai.