惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

云风的 BLOG
云风的 BLOG
M
MIT News - Artificial intelligence
博客园 - Franky
J
Java Code Geeks
V
Visual Studio Blog
G
Google Developers Blog
罗磊的独立博客
MongoDB | Blog
MongoDB | Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Recent Announcements
Recent Announcements
Last Week in AI
Last Week in AI
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Stack Overflow Blog
Stack Overflow Blog
博客园 - 司徒正美
The GitHub Blog
The GitHub Blog
腾讯CDC
阮一峰的网络日志
阮一峰的网络日志
V
V2EX
博客园 - 【当耐特】
IT之家
IT之家
I
InfoQ
U
Unit 42
C
Check Point Blog
Martin Fowler
Martin Fowler

PYMNTS.com

Google Accelerates Agentic AI Shift With New Enterprise Platform DeFi Security Suffers New Blow With $3 Million Volo Exploit Uninvited Users Access Anthropic’s Mythos AI Model Block and Uber Expand Partnership Across Several Global Markets OpenAI Pledges $1.5 Billion to PE Enterprise AI Project Podcast: Inside the $9 Billion DeFi Hack That’s Shaking Crypto’s Foundations Synchrony CFO Flags Momentum in Spending and Credit Banks Risk Slowing the Emerging Middle Market Firms Driving Growth Paysafe Expands Digital Wallet Availability Across 18 European Markets Bad Data Can Break Good AI in Payments 50% More Digital Shopping Days Put Parents at the Center of Retail’s Shift 65% Call Insurance Essential. Why Most Spending Isn’t So Clear-Cut Amazon Recasts Marketplace Fraud as a Broader Trust Problem Capital One’s Q1 Shifts Attention From Spending to Strategy Lawmakers Question JetBlue About Surveillance Pricing Allegations Small Businesses Stop Chasing Amazon on Delivery Speed Google Embeds AI Into Chrome for 3.5 Billion Users Adobe Plans Outcome-Based Pricing for New AI Product Suite UnitedHealth Spends $1.5 Billion on AI and Wants Double Back MiCA Forces Crypto Firms to Get Licensed or Get Out Prediction Market Kalshi Targets Crypto Perpetuals New York Sues Coinbase and Gemini Over Prediction Markets Amazon and Anthropic Deepen Ties With Investment and Hardware Pact Commercial Loans Show US Economy Defies Sluggish Forecasts The Web Is Gaslighting AI Agents and Nobody Can Tell OCC Enters the Interchange Fight and Raises the Stakes Amazon Dismisses New Evidence in California Antitrust Suit AI Finds Its Best Customer on Main Street Coinbase Opens Services Marketplace for Agentic Commerce Feds Start Processing $127 Billion in Tariff Refunds for Importers
OpenAI Images 2.0 Is a Real Leap With a Real Price Tag
PYMNTS · 2026-04-23 · via PYMNTS.com

By  |  April 22, 2026

 | 

OpenAI

Two years ago, asking an artificial intelligence (AI) image model for a software dashboard mockup meant getting back something that looked like a dashboard had melted, with corrupted labels and drifted columns. A designer would have to spend an hour cleaning it up.

PYMNTS tested ChatGPT Images 2.0 on the same prompt. The layout held. Text rendered cleanly across both the dashboard and a set of product-style images. Outputs came back as strong drafts. Only minor corrections were needed.

OpenAI released the model on Tuesday (April 21). While the quality gap is real, the business case still needs work.

What the New Model Does Differently

OpenAI said Images 2.0 “brings an unprecedented level of specificity and fidelity to image creation,” describing it as able to follow instructions, preserve requested details and render fine-grained elements including small text, iconography, UI elements and dense compositions at up to 2K resolution.

The model includes a thinking mode that reasons before generating, spending more or less time depending on the complexity of the prompt, and can search the web during that process, according to Open AI. The output is built from a plan rather than reconstructed from noise. That shift is what fixes text. Diffusion models treated letters as pixels. The new model treats them as instructions.

With thinking mode active, the model generates up to eight images at once from a single prompt, with characters, objects and styles held consistent across all outputs, according to The Decoder. Extended thinking is restricted to Plus, Pro and Business subscribers. Free users get the base quality improvements. Developers can access the model via the application programming interface (API) under the name gpt-image-2.

Advertisement: Scroll to Continue

Text rendering improvements extend to Japanese, Korean, Hindi and Bengali, expanding addressable use cases for global commerce and localized product content.

Where the Business Case Holds

The use cases that work are the ones where output quality directly cuts labor. Marketing teams producing ad variants, eCommerce operators generating product imagery at scale and design teams building UI mockups are the clearest examples. The previous problem wasn’t the idea. It was that images requiring human correction on every pass was slower than images made by hand.

A model that returns a strong draft on the first pass changes that math. The correction loop shortens. Per-output hours drop. At volume, that’s where savings appear.

OpenAI lists localized advertising, infographics, educational content and design tools among its target enterprise use cases. TechRadar noted the model’s reasoning step makes it better suited to multi-part design requests where elements need to stay coherent across a composition, which maps onto real production workflows in marketing and product teams.

What’s Limiting Adoption

Image generation doesn’t fit the same cost model as text. Text models run at high frequency across coding, customer support and finance operations. Image generation is episodic. It doesn’t sit inside a daily workflow the way a language model does. Lower volume means fewer opportunities to amortize per-image API costs against measurable output.

The pricing reflects that tension. At the standard 1024×1024 resolution in high quality, the new model costs $0.211 per image via the API, up from $0.133 for its predecessor GPT Image 1.5, The Decoder reported. At larger resolutions, the new model is cheaper than prior versions. The structure rewards scale and penalizes low-frequency use.

Latency is a separate constraint. The thinking step takes time. Generating a complex multi-element output takes minutes rather than seconds, as TechRadar noted, which matters in workflows where speed is the point.

There’s also a measurement problem that text models don’t share. Text model return on investment (ROI) maps cleanly onto time saved per query or tickets resolved. Image generation ROI is harder to isolate. Design cycles are longer. Creative review adds variability. The line between a model that saved time and one that shifted where the work happens isn’t always clear.

For all PYMNTS AI and digital transformation coverage, subscribe to the daily AI and Digital Transformation Newsletters.