<!-- LLM_VERSION_INFO
FORMAT: text/markdown
CONTENT_TYPE: article
ORIGINAL_URL: https://www.vals.ai/models/cohere_command-a-03-2025
ALTERNATE_VERSION: models/cohere_command-a-03-2025/index.html (text/html)
EXTRACTION_DATE: 2026-04-17T00:43:42.728Z

This is the markdown version with text-only content (images converted to alt-text).
For rich formatting with images, request the HTML version at: models/cohere_command-a-03-2025/index.html
-->

# Open Weights & Proprietary

## All Companies

### Release date

| Models                                | Release Date  | Image                                                      |
|---------------------------------------|----------------|-----------------------------------------------------------|
| Claude Opus 4.7                       | 4/16/2026     |  |
| Muse Spark                            | 4/8/2026      |  |
| Gemma 4 31B IT                        | 4/2/2026      |  |
| Qwen 3.6 Plus                        | 4/2/2026      |  |
| GLM 5.1                              | 4/1/2026      |  |
| Trinity Large Thinking                | 4/1/2026      |  |
| GPT 5.4 Mini                          | 3/17/2026     |  |
| GPT 5.4 Nano                          | 3/17/2026     |  |
| MiniMax-M2.7                         | 3/17/2026     |  |
| Grok 4.20 (Reasoning)                | 3/9/2026      |  |
| GPT 5.4                              | 3/5/2026      |  |
| Gemini 3.1 Flash Lite Preview         | 2/24/2026     |  |
| GPT 5.3 Codex                        | 2/23/2026     |  |
| Qwen 3.5 Flash                       | 2/19/2026     |  |
| Gemini 3.1 Pro Preview (02/26)       | 2/17/2026     |  |
| Claude Sonnet 4.6                    | 2/16/2026     |  |
| Qwen 3.5 Plus                        | 2/12/2026     |  |
| MiniMax-M2.5                         | 2/12/2026     |  |
| MiniMax-M2.5                         | 2/11/2026     |  |
| GLM 5                                | 2/5/2026      |  |
| Claude Opus 4.6 (Nonthinking)        | 2/5/2026      |  |
| Claude Opus 4.6 (Thinking)           | 1/26/2026     |  |
| Kimi K2.5                            | 1/23/2026     |  |
| Qwen 3 Max Thinking                   | 12/23/2025    |  |
| MiniMax-M2.1                         | 12/22/2025    |  |
| GLM 4.7                              | 12/17/2025    |  |
| Gemini 3 Flash (12/25)               | 12/17/2025    |  |
| MiMo V2 Flash                        | 12/11/2025    |  |
| GPT 5.2                              | 12/11/2025    |  |
| GPT 5.2 Codex                        | 12/11/2025    |  |

### Command A

[Compare Models](/content/comparison?modelA=cohere%2Fcommand-a-03-2025/index.html)  
[Command A](https://docs.cohere.com/v2/docs/models)

**Release Date:** 3/13/2025

#### Accuracy (Vals Index)
**19.20%** ± 0.98

#### Latency (Vals Index)
**305.05s**

#### Cost/Test (Vals Index)
**$1.00**

#### Context Window
**256k**

#### Max Output Tokens
**8k**

#### Input Modality
**Hyperparameter settings**

**Default Provider:** Cohere  
Some benchmarks may use different provider and parameters. Please refer to the benchmark page for more information.

#### Benchmark Settings
- **Temperature:** 0.3
- **Top P:** Default
- **Top K:** Default
- **Max Output Tokens:** 8,000

### Benchmarks

#### Accuracy Rankings

| Benchmark Name                     | Performance     | Ranking               |
|-------------------------------------|------------------|-----------------------|
| [CaseLaw (v2)](/content/benchmarks/case_law_v2/index.html)               | -205.56% ± 1.09 | 162/47                |
| [CorpFin](/content/benchmarks/corp_fin_v2/index.html)                   | -169.69% ± 0.97 | 163/97                |
| [Finance Agent (v1.1)](/content/benchmarks/finance_agent/index.html)    | -17.92% ± 0.97  | 62/45                 |
| [TaxEval (v2)](/content/benchmarks/tax_eval_v2/index.html)             | -296.02% ± 0.95 | 191/104               |
| [AIME](/content/benchmarks/aime/index.html)                             | -72.67% ± 0.94  | 151/96                |
| [GPQA](/content/benchmarks/gpqa/index.html)                             | -296.87% ± 2.28 | 166/99                |
| [LiveCodeBench](/content/benchmarks/lcb/index.html)                    | -239.90% ± 1.09 | 144/103               |
| [LegalBench](/content/benchmarks/legal_bench/index.html)               | -606.17% ± 0.42 | 588/116               |
| [MedQA](/content/benchmarks/medqa/index.html)                           | -677.88% ± 0.36 | 289/95                |
| [MMLU Pro](/content/benchmarks/mmlu_pro/index.html)                     | -641.79% ± 0.46 | 199/97                |
| [SWE-bench](/content/benchmarks/swebench/index.html)                   | -79.49% ± 1.20  | 41/41                 |
| [Terminal-Bench 2.0](/content/benchmarks/terminal-bench-2/index.html)   | -25.11% ± 1.58  | 63/52                 |
