All Comparisons

Codestral-22B vs GPT-4o

Codestral-22B
Mistral AI πŸ‡«πŸ‡·
GPT-4o
OpenAI πŸ‡ΊπŸ‡Έ
58.3
#136β€”
high
$2.50/M in Β· $10.00/M out
128K context
Most Affordable
GPT-4o
$2.50/M input
Largest Context
GPT-4o
128K tokens
Best Benchmark Score
GPT-4o
58.3/100

Model Specifications

SpecCodestral-22BGPT-4o
ProviderMistral AIOpenAI
Release DateMay 2024May 2024
Knowledge Cutoffβ€”Oct 2023
Parameters22200000000Bunknown
Context Windowβ€”128K
Max Outputβ€”β€”
Open SourceYesNo
Licensemnpl_0_1proprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Codestral-22BGPT-4o
Inputβ€”$2.50
Outputβ€”$10.00

API Provider Pricing

Capabilities

FeatureCodestral-22BGPT-4o
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkCodestral-22BGPT-4o
MMLU-Proβ€”
74.5
IFEvalβ€”
83.6
Chatbot Arena Eloβ€”
1285

coding

BenchmarkCodestral-22BGPT-4o
SWE-bench Verifiedβ€”
33.2
HumanEvalβ€”
90.2

math

BenchmarkCodestral-22BGPT-4o
AIME 2024β€”
13.4
MATHβ€”
76.6
GSM8Kβ€”
95.8

reasoning

BenchmarkCodestral-22BGPT-4o
BigBench-Hardβ€”
87.3
ARC-Challengeβ€”
96.4
GPQA Diamondβ€”
53.6

language

BenchmarkCodestral-22BGPT-4o
HellaSwagβ€”
96.4
WinoGrandeβ€”
85.7

multimodal

BenchmarkCodestral-22BGPT-4o
MMMUβ€”
69.1

safety

BenchmarkCodestral-22BGPT-4o
TruthfulQAβ€”
73.5

Category Winners

coding
GPT-4o
38.2
math
GPT-4o
54.5
reasoning
GPT-4o
54.4
general
GPT-4o
70.5
language
GPT-4o
79.9
multimodal
GPT-4o
54.0
safety
GPT-4o
94.3

Compare More Models

Frequently Asked Questions