All Comparisons

Command R+ vs GPT-4o

Command R+
Cohere πŸ‡¨πŸ‡¦
8.5
#316β€”
high
$2.50/M in Β· $10.00/M out
128K context
GPT-4o
OpenAI πŸ‡ΊπŸ‡Έ
58.3
#136β€”
high
$2.50/M in Β· $10.00/M out
128K context
Most Affordable
Command R+
$2.50/M input
Largest Context
Command R+
128K tokens
Best Benchmark Score
GPT-4o
58.3/100

Model Specifications

SpecCommand R+GPT-4o
ProviderCohereOpenAI
Release DateApr 2024May 2024
Knowledge Cutoffβ€”Oct 2023
Parameters104Bunknown
Context Window128K128K
Max Outputβ€”β€”
Open SourceOpen WeightsNo
Licensecc-by-nc-4.0proprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Command R+GPT-4o
Input$2.50$2.50
Output$10.00$10.00

API Provider Pricing

Capabilities

FeatureCommand R+GPT-4o
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkCommand R+GPT-4o
IFEval
73.8
83.6
Chatbot Arena Elo
1168
1285
MMLU-Pro
55.8
74.5

coding

BenchmarkCommand R+GPT-4o
HumanEval
73.2
90.2
SWE-bench Verifiedβ€”
33.2

math

BenchmarkCommand R+GPT-4o
MATH
47.2
76.6
GSM8K
87.5
95.8
AIME 2024β€”
13.4

reasoning

BenchmarkCommand R+GPT-4o
ARC-Challenge
88.4
96.4
BigBench-Hardβ€”
87.3
GPQA Diamondβ€”
53.6

language

BenchmarkCommand R+GPT-4o
WinoGrande
80.2
85.7
HellaSwag
85.3
96.4

safety

BenchmarkCommand R+GPT-4o
TruthfulQA
58.8
73.5

multimodal

BenchmarkCommand R+GPT-4o
MMMUβ€”
69.1

Category Winners

coding
GPT-4o
38.2
math
GPT-4o
54.5
reasoning
GPT-4o
54.4
general
GPT-4o
70.5
language
GPT-4o
79.9
multimodal
GPT-4o
54.0
safety
GPT-4o
94.3

Compare More Models

Frequently Asked Questions