AI Models

OpenAI ChatGPT Images 2.0 Adds Reasoning and Web Search

OpenAI has released ChatGPT Images 2.0, powered by GPT Image 2, which reasons before generating images and incorporates web search. It produces up to eight consistent images per prompt and improves text handling, especially non-Latin scripts. All users see better quality, while paid tiers access advanced thinking modes.

Neura News

Neura News

Neura Market Editorial

April 21, 20264 min read
OpenAI ChatGPT Images 2.0 Adds Reasoning and Web Search

OpenAI ChatGPT Images 2.0 Adds Reasoning and Web Search

OpenAI rolled out ChatGPT Images 2.0, its updated image generator based on the GPT Image 2 model. This version includes reasoning steps before creation and web search capabilities. Users can now generate as many as eight consistent images from one prompt. The tool also manages text much better overall, with strong results for non-Latin scripts.

The company confirmed the release in a blog post. GPT Image 2 mirrors features in Google's Nano Banana Pro by taking time to think based on the chosen mode. It adjusts reasoning duration and pulls web data when needed. OpenAI expects this to boost image variety and precision. Paid ChatGPT Plus, Pro, and Business subscribers alone get the full thinking outputs.

Thinking Mode Enables Multi-Image Consistency

When users turn on thinking mode, ChatGPT Images 2.0 produces up to eight images simultaneously from a single instruction. Characters, objects, and styles remain uniform across outputs. OpenAI cites examples like full-page mangas from one image and text prompt, sets of social media visuals, and house room design series.

Quality Upgrades Reach Every User

All ChatGPT accounts benefit from enhanced image quality, no matter the mode. The generator now grasps key photo traits more accurately. It excels in pixel art, manga, film stills, and similar styles. Previous models faltered on details like tiny text, icons, UI components, crowded scenes, and nuanced style cues. This one handles them well.

Supported aspect ratios stretch from 3:1 ultra-wide to 1:3 ultra-tall. These fit banners, slides, and phone displays. API access reaches 2K resolution.

API Details and Token-Based Costs

The #1 Newsletter in AI

Stay ahead of the AI curve

The most important updates, news, and content — delivered weekly.

No spam. Unsubscribe anytime.

Developers access the model as gpt-image-2 through OpenAI's API. Charges follow a token system: $8 per million image input tokens, $30 per million image output tokens. Text inputs run $5 per million, outputs $10 per million. Cached inputs cost less.

Actual per-image prices depend on quality and size. OpenAI's pricing page lists a 1024 x 1024 low-quality image at $0.006, medium at $0.053, high at $0.211. A 1024 x 1536 version costs $0.005 low, $0.041 medium, $0.165 high. Other sizes like 1536 x 1024 match those rates.

| Model | Quality | 1024 x 1024 | 1024 x 1536 | 1536 x 1024 | |, , , , , , -|, , , , -|, , , , , , -|, , , , , , -|, , , , , , -| | GPT Image 2 | Low | $0.006 | $0.005 | $0.005 | | | Medium | $0.053 | $0.041 | $0.041 | | | High | $0.211 | $0.165 | $0.165 | | GPT Image 1.5 | Low | $0.009 | $0.013 | $0.013 | | | Medium | $0.034 | $0.05 | $0.05 | | | High | $0.133 | $0.2 | $0.2 |

GPT Image 2 undercuts prior versions at bigger sizes. High-quality 1024 x 1536 runs $0.165 against $0.20 for GPT Image 1.5 and $0.25 for GPT Image 1.5 at other large formats. Standard 1024 x 1024 high quality costs more at $0.211 versus $0.133. Outputs over 2K stay in beta with possible inconsistencies.

OpenAI points to uses in localized ads, infographics, school materials, design apps, and creative sites. Codex integrates image generation right in the workspace, no extra API key required.

Real-World Tests and Pre-Launch Buzz

Tests show ChatGPT Images 2.0 performs well on tough prompts. One benchmark asked for a hyper-realistic DSLR photo: a monkey holding a pink banana sits on a tiger upfront. Behind, a horse rides an astronaut, who serves as a living spacesuit saddle. The horse controls as rider, fully clear, no reversal. High-res, sharp focus, real lighting.

Instant mode yields a somewhat artificial result. Thinking mode captures true DSLR quality better.

Before launch, codenamed gpt-image-2 reached select US testers. Samples on X and Reddit looked photo-real. It shines on complex scenes, diagrams, text-heavy screenshots, suiting ads and education like infographics. OpenAI fixed the smooth skin and ideal lighting flaws from GPT Image 1.5, where Nano Banana Pro led. A livestream at 12 pm PT revealed it. Examples included a fake Nadella image touting Chrome downloads via Edge and a teased AI screenshot.

Related on Neura Market

More from Neura News

AI Models

42 Mathematicians Urge Royal Society to Warn Government and Media About AI Existential Risk

Forty-two mathematical fellows, including Fields Medal winners Martin Hairer, Peter Scholze, and Wendelin Werner, have signed an open letter urging the Royal Society to warn the UK government and media about existential risks from advanced AI. The letter follows recent breakthroughs in which leading models solved open research problems, including a Millennium Problem. None of the signatories are affiliated with AI companies. The group warns that AI labs' estimates of existential risk above ten percent must not be dismissed as hype, and that by the time the situation becomes obvious to the public, it may be too late to act.

Sep 18·2 min read
Developer

Steve Yegge Shuts Down Gas Town After Failing to Build Anything Else With It

Steve Yegge shut down Gas Town, his ultra-vibed coding agent orchestrator, after admitting he never built anything else with it despite heavy subscription spend. Databricks reported a 60% coding spend increase after rolling out GPT-6 Astra to 3,500 engineers, OpenAI published a misalignment disclosure framework with six case reports, and Xiaomi ran MiMo-V2.6 RL training in public with live telemetry.

Sep 18·21 min read