All Comparisons

Command R+ vs GPT-4.1

Command R+
Cohere 🇨🇦
8.5
#316
high
$2.50/M in · $10.00/M out
128K context
GPT-4.1
OpenAI 🇺🇸
32.8
#238
high
$2.00/M in · $8.00/M out
1.0M context
Most Affordable
GPT-4.1
$2.00/M input
Largest Context
GPT-4.1
1.0M tokens
Best Benchmark Score
GPT-4.1
32.8/100

Model Specifications

SpecCommand R+GPT-4.1
ProviderCohereOpenAI
Release DateApr 2024Apr 2025
Knowledge Cutoff2024-06-01
Parameters104BUndisclosed
Context Window128K1.0M
Max Output
Open SourceOpen WeightsNo
Licensecc-by-nc-4.0proprietary
Tokenizer
Modality
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Command R+GPT-4.1
Input$2.50$2.00
Output$10.00$8.00

API Provider Pricing

Capabilities

FeatureCommand R+GPT-4.1
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkCommand R+GPT-4.1
IFEval
73.8
Chatbot Arena Elo
1168
MMLU-Pro
55.8
MMMLU
87.3

coding

BenchmarkCommand R+GPT-4.1
HumanEval
73.2
SWE-bench Verified
54.6

math

BenchmarkCommand R+GPT-4.1
MATH
47.2
GSM8K
87.5
AIME 2025
46.4

reasoning

BenchmarkCommand R+GPT-4.1
ARC-Challenge
88.4
GPQA Diamond
66.3
Humanity's Last Exam
5.4

language

BenchmarkCommand R+GPT-4.1
WinoGrande
80.2
HellaSwag
85.3

safety

BenchmarkCommand R+GPT-4.1
TruthfulQA
58.8

multimodal

BenchmarkCommand R+GPT-4.1
MMMU
74.8
CharXiv Reasoning
56.7

agent

BenchmarkCommand R+GPT-4.1
TAU-Bench Retail
68

Category Winners

coding
GPT-4.1
25.1
math
GPT-4.1
6.0
reasoning
GPT-4.1
27.5
general
GPT-4.1
66.0
language
Command R+
6.8
multimodal
GPT-4.1
40.2
safety
Command R+
21.0
agent
GPT-4.1
50.7

Compare More Models

Frequently Asked Questions