Mistral AIopen-weight
Released: 2025-01-14Codestral 25.01
Mistral dedicated 256k context window coding model tailored for fill-in-the-middle completion and polyglot development.
Intelligence Index
78.5/ 100
Calibrated multi-domain compositeThroughput Speed
90tok/sec
Output streaming throughputTime To First Token
320ms
Initial chunk server latencyLiveBench Contam-Free
—%
September 2026 suite scoreSourced Benchmark Evaluations (2)
Verified performance across authoritative benchmarks with provenance tracking.| Benchmark | Status | Score | Trust | Date | Source Type | Provenance |
|---|---|---|---|---|---|---|
| ARC-AGI-2 | Nearing Saturation | 57.7% | 75 | 2025-02-11 | independent | Source ↗ |
| SWE-bench Verified | Saturated | 54.7% | 25 | 2025-01-23 | independent | Source ↗ |
Local Hardware Execution (Ollama)Hardware Compatible
# 1. Pull and execute model locally with Ollama
ollama run codestral-25-01
# 2. Or invoke via local OpenAI-compatible endpoint
curl http://localhost:11434/v1/chat/completions \
-H "Content-Type: application/json" \
-d '{
"model": "codestral-25-01",
"messages": [{"role": "user", "content": "Analyze reasoning chains on AIME 2026"}]
}'Evaluation & Sourcing Notes
Scores listed for Codestral 25.01 represent verified evaluations extracted from official research papers, independent evaluation suites (HELM, LMSYS, OpenCompass, LiveBench), and verified audit reports.
All benchmarks marked as saturated or deprecated reflect historical performance where the benchmark no longer provides active discriminative power.