All Comparisons

GPT-4o vs Llama 4 Scout

GPT-4o
OpenAI 🇺🇸
58.3
#136
high
$2.50/M in · $10.00/M out
128K context
M
Llama 4 Scout
Meta 🇺🇸
38.6
#214
low
Most Affordable
GPT-4o
$2.50/M input
Largest Context
GPT-4o
128K tokens
Best Benchmark Score
GPT-4o
58.3/100

Model Specifications

SpecGPT-4oLlama 4 Scout
ProviderOpenAIMeta
Release DateMay 2024Apr 2025
Knowledge CutoffOct 2023
Parametersunknown109000000000B
Context Window128K
Max Output
Open SourceNoYes
Licenseproprietaryllama_4_community_license_agreement
Tokenizer
Modality
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)GPT-4oLlama 4 Scout
Input$2.50
Output$10.00

API Provider Pricing

Capabilities

FeatureGPT-4oLlama 4 Scout
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkGPT-4oLlama 4 Scout
MMLU-Pro
74.5
IFEval
83.6
Chatbot Arena Elo
1285

coding

BenchmarkGPT-4oLlama 4 Scout
SWE-bench Verified
33.2
HumanEval
90.2

math

BenchmarkGPT-4oLlama 4 Scout
AIME 2024
13.4
MATH
76.6
GSM8K
95.8

reasoning

BenchmarkGPT-4oLlama 4 Scout
BigBench-Hard
87.3
ARC-Challenge
96.4
GPQA Diamond
53.6
57.2

language

BenchmarkGPT-4oLlama 4 Scout
HellaSwag
96.4
WinoGrande
85.7

multimodal

BenchmarkGPT-4oLlama 4 Scout
MMMU
69.1
69.4

safety

BenchmarkGPT-4oLlama 4 Scout
TruthfulQA
73.5

Category Winners

coding
GPT-4o
38.2
math
GPT-4o
54.5
reasoning
GPT-4o
54.4
general
GPT-4o
70.5
language
GPT-4o
79.9
multimodal
M
Llama 4 Scout
54.8
safety
GPT-4o
94.3

Compare More Models

Frequently Asked Questions