AI Models

Opus 4.7 Burns More Tokens Than 4.6 at Same Price

Early measurements show Anthropic's Opus 4.7 model consumes significantly more tokens per request than Opus 4.6, despite identical pricing. Developer Abhishek Ray reported averages like 1.325x for code content and up to 1.47x for technical docs. Community data indicates a 37.4 percent rise in token use and costs across 483 tests, with modest gains in instruction following.

Neura News

Neura News

Neura Market Editorial

April 19, 20262 min read
Opus 4.7 Burns More Tokens Than 4.6 at Same Price

Opus 4.7 Burns More Tokens Than 4.6 at Same Price

Anthropic released Opus 4.7 with the same listed price as Opus 4.6. Initial token counts indicate it uses substantially more tokens for each request. Developer Abhishek Ray shared these findings on Claude Code Camp.

Anthropic, a company focused on safe AI systems, built the Claude family of models. Former OpenAI staff founded it in 2021. Opus serves as the top-tier model in this lineup, handling complex tasks like coding and analysis. Pricing follows a token-based system, where costs depend on input and output tokens processed.

Token Increase Matches and Exceeds Guide

Anthropic's migration guide mentions a token rise of 1.0 to 1.35 times. Ray's tests align closely with this for most cases. Real Claude Code content averaged 1.325 times more tokens. A CLAUDE.md file reached 1.445 times. Technical documentation hit 1.47 times.

Code content suffered the largest increase, according to Ray. Prose experienced a smaller rise. Chinese and Japanese texts showed almost no change.

Community Tests Show Even Higher Usage

A community evaluation on tokens.billchambers.me reported stronger results. It covered 483 submissions. Token usage jumped 37.4 percent. Per-request costs followed the same pattern.

These figures come from diverse user tests. They highlight real-world differences beyond official estimates.

The #1 Newsletter in AI

Stay ahead of the AI curve

The most important updates, news, and content — delivered weekly.

No spam. Unsubscribe anytime.

Cost Estimates for Typical Sessions

Ray calculated costs for a sample session with 80 turns. Switching to Opus 4.7 added 20 to 30 percent to expenses. The original bill stood at $6.65. New totals ranged from $7.86 to $8.76.

Such increases matter for heavy users. Developers and businesses track token counts to manage budgets. Flat pricing masked these shifts until hands-on tests appeared.

Gains in Instruction Following

Users receive some benefits. Opus 4.7 follows instructions better in certain tests. On the IFEval benchmark with 20 prompts, it succeeded five percentage points more often than Opus 4.6.

IFEval measures adherence to rules in prompts. This small improvement suits tasks needing precision, like code generation or strict formatting.

Anthropic continues to update Claude models. Opus versions evolve with better capabilities, though efficiency varies. Early adopters now adjust workflows based on these token realities.

Related on Neura Market

More from Neura News

General

Open-weight AI mirrors Kubernetes ecosystem shift

Tobi Knaup, co-founder of Mesosphere, draws parallels between the rise of Kubernetes and the current trajectory of open-weight AI models. He argues that open-weight models are becoming a neutral substrate for innovation, attracting a global ecosystem of developers, startups, and enterprises. The piece warns against US restrictions on Chinese open-weight models, advocating instead for American leadership through open releases, procurement strategies, and standards.

Jul 25·7 min read
General

Open-weight AI mirrors Kubernetes rise, US warned on bans

The author, a Mesosphere co-founder, draws parallels between the rise of Kubernetes and the current open-weight AI ecosystem. He argues that open-weight models are becoming a neutral platform for innovation, and warns that US restrictions on Chinese open-weight models could isolate American developers from a global ecosystem. The piece urges the US to compete by releasing frontier models, using procurement to create demand, building the stack, and setting standards rather than imposing bans.

Jul 25·7 min read
Industry

Power line failure reveals AI data center grid risks and solutions

A fallen power line near Washington, DC caused over 3 gigawatts of data center load to vanish from the PJM grid in seconds, spiking voltage across the region. The event, which made lights flicker from Northern Virginia to Chicago, highlights a growing problem as AI data centers become larger and more concentrated. Experts warn that without better coordination or technology like ON.Energy's battery-backed uninterruptible power supply, such disruptions will become more frequent and severe.

Jul 25·5 min read