<!-- LLM_VERSION_INFO
FORMAT: text/markdown
CONTENT_TYPE: article
ORIGINAL_URL: https://www.vals.ai/models/zai_glm-5.1-thinking
ALTERNATE_VERSION: models/zai_glm-5-1-thinking.html (text/html)
EXTRACTION_DATE: 2026-04-17T00:43:36.928Z

This is the markdown version with text-only content (images converted to alt-text).
For rich formatting with images, request the HTML version at: models/zai_glm-5-1-thinking.html
-->

# Open Weights & Proprietary

## All Companies

### Models

#### Release date

| Model                              | Release Date |  
|------------------------------------|--------------|  
|  Claude Opus 4.7 | 4/16/2026 |  
|  Muse Spark | 4/8/2026 |  
|  Gemma 4 31B IT | 4/2/2026 |  
|  Qwen 3.6 Plus | 4/1/2026 |  
|  GLM 5.1 | 4/1/2026 |  
|  Trinity Large Thinking | 3/17/2026 |  
|  GPT 5.4 Mini | 3/17/2026 |  
|  GPT 5.4 Nano | 3/17/2026 |  
|  MiniMax-M2.7 | 3/9/2026 |  
|  Grok 4.20 (Reasoning) | 3/5/2026 |  
|  GPT 5.4 | 3/3/2026 |  
|  Gemini 3.1 Flash Lite Preview | 2/24/2026 |  
|  GPT 5.3 Codex | 2/23/2026 |  
|  Qwen 3.5 Flash | 2/19/2026 |  
|  Gemini 3.1 Pro Preview (02/26) | 2/17/2026 |  
|  Claude Sonnet 4.6 | 2/16/2026 |  
|  Qwen 3.5 Plus | 2/12/2026 |  
|  MiniMax-M2.5 | 2/12/2026 |  
|  MiniMax-M2.5 | 2/11/2026 |  
|  GLM 5 | 2/5/2026 |  
|  Claude Opus 4.6 (Nonthinking) | 2/5/2026 |  
|  Claude Opus 4.6 (Thinking) | 1/26/2026 |  
|  Kimi K2.5 | 1/23/2026 |  
|  Qwen 3 Max Thinking | 12/23/2025 |  
|  MiniMax-M2.1 | 12/22/2025 |  
|  GLM 4.7 | 12/17/2025 |  
|  Gemini 3 Flash (12/25) | 12/17/2025 |  
|  MiMo V2 Flash | 12/11/2025 |  
|  GPT 5.2 | 12/11/2025 |  
|  GPT 5.2 Codex | 12/11/2025 |

### View All Models

#### GLM 5.1

| Model Name         | Details            |
|--------------------|--------------------|
|  | [GLM 5.1](https://docs.z.ai/) |

**Release Date**: 4/1/2026

**Accuracy (Vals Index)**: 63.17% ± 1.95

**Latency (Vals Index)**: 457.94s

**Cost/Test (Vals Index)**: $0.22

**Context Window**: 200k

**Max Output Tokens**: 131k

**Input Modality**: Hyperparameter settings

**Default Provider**: Zhipu AI

**Temperature**: 1

**Top P**: 0.95

**Top K**: Default

**Max Output Tokens**: 131,072

**Show rankings only among open weight models**

### Benchmarks

#### Accuracy Rankings

* [Vals Index](/content/benchmarks/vals_index/index.html) -394.32% ± 1.95 
* [CaseLaw (v2)](/content/benchmarks/case_law_v2/index.html) -352.94% ± 1.13 
* [CorpFin](/content/benchmarks/corp_fin_v2/index.html) -490.59% ± 0.94 
* [Finance Agent (v1.1)](/content/benchmarks/finance_agent/index.html) -485.73% ± 2.80 
* [MedCode](/content/benchmarks/medcode/index.html) -386.28% ± 2.12 
* [MedScribe](/content/benchmarks/medscribe/index.html) -737.02% ± 2.06 
* [ProofBench](/content/benchmarks/proof_bench/index.html) -245.59% ± 4.16 
* [TaxEval (v2)](/content/benchmarks/tax_eval_v2/index.html) -867.23% ± 0.90 
* [Vibe Code Bench](/content/benchmarks/vibe-code/index.html) -417.04% ± 4.55 
* [AIME](/content/benchmarks/aime/index.html) -1321.89% ± 0.44 
* [GPQA](/content/benchmarks/gpqa/index.html) -1316.53% ± 1.82 
* [LiveCodeBench](/content/benchmarks/lcb/index.html) -1369.27% ± 1.06 
* [LegalBench](/content/benchmarks/legal_bench/index.html) -1530.72% ± 0.43 
* [MMLU Pro](/content/benchmarks/mmlu_pro/index.html) -1695.33% ± 0.33 
* [SWE-bench](/content/benchmarks/swebench/index.html) -1600.28% ± 1.90 
* [Terminal-Bench 2.0](/content/benchmarks/terminal-bench-2/index.html) -1210.67% ± 5.31

---

##### Contact us

Or send us an email at [contact@vals.ai](mailto:contact@vals.ai)

**Proprietary Benchmarks** (contact us to get access)

**Academic Benchmarks**
Read about our [methodology](/content/methodology/index.html).
