<!-- LLM_VERSION_INFO
FORMAT: text/markdown
CONTENT_TYPE: article
ORIGINAL_URL: https://www.vals.ai/models/anthropic_claude-opus-4-5-20251101-thinking
ALTERNATE_VERSION: models/anthropic_claude-opus-4-5-20251101-thinking/index.html (text/html)
EXTRACTION_DATE: 2026-04-17T00:42:45.104Z

This is the markdown version with text-only content (images converted to alt-text).
For rich formatting with images, request the HTML version at: models/anthropic_claude-opus-4-5-20251101-thinking/index.html
-->

# Open Weights & Proprietary

## Release date

| Models | Release Date |
|--------|--------------|
| Claude Opus 4.7 | 4/16/2026 |
| Muse Spark | 4/8/2026 |
| Gemma 4 31B IT | 4/2/2026 |
| Qwen 3.6 Plus | 4/2/2026 |
| GLM 5.1 | 4/1/2026 |
| Trinity Large Thinking | 4/1/2026 |
| GPT 5.4 Mini | 3/17/2026 |
| GPT 5.4 Nano | 3/17/2026 |
| MiniMax-M2.7 | 3/17/2026 |
| Grok 4.20 (Reasoning) | 3/9/2026 |
| GPT 5.4 | 3/5/2026 |
| Gemini 3.1 Flash Lite Preview | 2/24/2026 |
| GPT 5.3 Codex | 2/23/2026 |
| Qwen 3.5 Flash | 2/19/2026 |
| Gemini 3.1 Pro Preview (02/26) | 2/17/2026 |
| Claude Sonnet 4.6 | 2/16/2026 |
| Qwen 3.5 Plus | 2/12/2026 |
| MiniMax-M2.5 | 2/12/2026 |
| GLM 5 | 2/11/2026 |
| Claude Opus 4.6 (Nonthinking) | 2/5/2026 |
| Claude Opus 4.6 (Thinking) | 2/5/2026 |
| Kimi K2.5 | 1/26/2026 |
| Qwen 3 Max Thinking | 1/23/2026 |
| MiniMax-M2.1 | 12/23/2025 |
| GLM 4.7 | 12/22/2025 |
| Gemini 3 Flash (12/25) | 12/17/2025 |
| MiMo V2 Flash | 12/17/2025 |
| GPT 5.2 | 12/11/2025 |
| GPT 5.2 Codex | 12/11/2025 |

### Claude Opus 4.5 (Thinking)  
  
[Compare Models](/content/comparison?modelA=anthropic%2Fclaude-opus-4-5-20251101-thinking/index.html)  
[Claude Opus 4.5 (Thinking)](https://docs.claude.com/en/docs/about-claude/models/overview)  
**Release Date:** 11/24/2025

### Accuracy (Vals Index)
- **Accuracy:** 62.93% ± 1.98
- **Latency:** 245.14s
- **Cost/Test:** $0.91
- **Context Window:** 200k
- **Max Output Tokens:** 64k
- **Input Modality:** Hyperparameter settings
  
**Default Provider:** Anthropic

### Benchmarks
#### Accuracy Rankings
- **[Vals Index](/content/benchmarks/vals_index/index.html)**
  - -96.37% ± 1.98  
  - 87/ 40  
- **[Vals Multimodal Index](/content/benchmarks/vals_multimodal_index/index.html)**  
  - -118.98% ± 1.55  
  - 68/ 28  
- **[CaseLaw (v2)](/content/benchmarks/case_law_v2/index.html)**  
  - -143.31% ± 0.11  
  - 118/ 47  
- **[CorpFin](/content/benchmarks/corp_fin_v2/index.html)**  
  - -177.18% ± 0.94  
  - 323/ 97  
- **[Finance Agent (v1.1)](/content/benchmarks/finance_agent/index.html)**  
  - -187.71% ± 2.81  
  - 170/ 45  
- **[MedCode](/content/benchmarks/medcode/index.html)**  
  - -181.81% ± 2.01  
  - 203/ 51  
- **[MedScribe](/content/benchmarks/medscribe/index.html)**  
  - -362.02% ± 1.90  
  - 246/ 51  
- **[MortgageTax](/content/benchmarks/mortgage_tax/index.html)**  
  - -327.00% ± 0.92  
  - 344/ 69  
- **[ProofBench](/content/benchmarks/proof_bench/index.html)**  
  - -196.54% ± 4.82  
  - 128/ 24  
- **[SAGE](/content/benchmarks/sage/index.html)**  
  - -319.41% ± 3.40  
  - 331/ 49  
- **[TaxEval (v2)](/content/benchmarks/tax_eval_v2/index.html)**  
  - -512.64% ± 0.85  
  - 741/ 104  
- **[AIME](/content/benchmarks/aime/index.html)**  
  - -726.48% ± 0.37  
  - 736/ 96  
- **[GPQA](/content/benchmarks/gpqa/index.html)**  
  - -723.57% ± 2.48  
  - 807/ 99  
- **[IOI](/content/benchmarks/ioi/index.html)**  
  - -188.03% ± 5.16  
  - 403/ 50  
- **[LiveCodeBench](/content/benchmarks/lcb/index.html)**  
  - -853.43% ± 1.04  
  - 950/ 103  
- **[LegalBench](/content/benchmarks/legal_bench/index.html)**  
  - -944.56% ± 0.39  
  - 1311/ 116  
- **[MedQA](/content/benchmarks/medqa/index.html)**  
  - -1168.00% ± 0.18  
  - 1131/ 95  
- **[MMLU Pro](/content/benchmarks/mmlu_pro/index.html)**  
  - -1156.92% ± 0.38  
  - 1224/ 97  
- **[MMMU Pro](/content/benchmarks/mmmu/index.html)**  
  - -1193.61% ± 0.90  
  - 814/ 66  
- **[SWE-bench](/content/benchmarks/swebench/index.html)**  
  - -1190.04% ± 1.90  
  - 571/ 41  
- **[Terminal-Bench 2.0](/content/benchmarks/terminal-bench-2/index.html)**  
  - -907.40% ± 5.31  
  - 759/ 52

### Contact us
Proprietary Benchmarks ( [contact us](/content/models/anthropic_claude-opus-4-5-20251101-thinking#contact-form/index.html) to get access)
Academic Benchmarks
Read about our [methodology](/content/methodology/index.html).
