Pricing Comparison: OpenAI vs Google Gemini pricing - a comparison for AI product builders
OpenAI and Google run the two most-watched rate cards in AI token pricing. Both repriced six times in the last 12 months, which makes this comparison a moving target. Here is where the numbers stand as of July 13, 2026.
OpenAI vs Google at a glance
OpenAI | ||
|---|---|---|
Cheapest flagship | gpt-5.6-sol at $5 / $30 | Gemini 3.1 Pro at $2 / $12 |
Cheapest small model | gpt-5.4-nano at $0.20 / $1.25 | Gemini 3.1 Flash-Lite at $0.125 / $0.75 |
Context window | Standard rates to 200k, then $10 / $45 on gpt-5.6-sol | 1M on Pro, $4 / $18 above 200k |
Pricing stability (12mo) | High. 6 price events in 12 months, flagship input up 4x | High. 6 price events in 12 months, Flash tier up 5x |
Takeaway: Gemini is cheaper at every tier today. Neither price list survives a quarter unchanged.
Model overview
OpenAI's current lineup is the gpt-5.6 family, launched July 9, 2026. gpt-5.6-sol is the flagship at $5 / $30, gpt-5.6-terra the mid tier at $2.50 / $15, and gpt-5.6-luna the value pick within the family at $1 / $6. The gpt-5.5 and gpt-5.4 generations stay on the card, with gpt-5.5-pro as the premium reasoning option and gpt-5.4-nano as the cheapest model at $0.20 / $1.25.
Google's flagship is Gemini 3.1 Pro, in preview at $2 / $12 for prompts up to 200k. Gemini 3.5 Flash, launched May 19, 2026, is the newest speed tier at $1.50 / $9. Gemini 3.1 Flash-Lite at $0.125 / $0.75 is the value pick, and Gemini 2.5 Pro remains on the card as a legacy model being displaced.
Key features
OpenAI | ||
|---|---|---|
Context window | Standard rates up to 200k tokens, long-context tier above (gpt-5.6-sol: $10 / $45) | 1M on Pro, long-context rates above 200k ($4 / $18 on 3.1 Pro) |
Caching discount | Cached input 10x cheaper than base input | Supported, rates vary per model |
Batch discount | Roughly 50% off | 50% off |
Multimodal | Text, image, audio | Text, image, audio, video |
Tool pricing | Web search $10 per 1k calls | Grounded search $14 per 1k queries after 5,000 free per month |
Pricing snapshot
Prices verified July 13, 2026. Providers reprice often, see the update timeline below.
OpenAI API pricing
Standard tier, per 1M tokens, prompts up to 200k. Long-context rates where noted.
Model | Input | Output | Notes |
|---|---|---|---|
gpt-5.6-sol | $5.00 | $30.00 | Flagship. $10 / $45 above 200k context |
gpt-5.6-terra | $2.50 | $15.00 | Mid tier |
gpt-5.6-luna | $1.00 | $6.00 | Value tier |
gpt-5.5 | $5.00 | $30.00 | Prior flagship |
gpt-5.5-pro | $30.00 | $180.00 | Premium reasoning |
gpt-5.4 | $2.50 | $15.00 | |
gpt-5.4-mini | $0.75 | $4.50 | |
gpt-5.4-nano | $0.20 | $1.25 | Cheapest |
Batch API: roughly 50% off. Cached input: 10x cheaper than base input. Web search tool: $10 per 1k calls.
Google API pricing
Per 1M tokens, paid tier, prompts up to 200k where tiered.
Model | Input | Output | Notes |
|---|---|---|---|
Gemini 3.1 Pro (preview) | $2.00 | $12.00 | Flagship. $4 / $18 above 200k |
Gemini 3.5 Flash | $1.50 | $9.00 | Newest speed tier |
Gemini 3 Flash (preview) | $0.50 | $3.00 | |
Gemini 3.1 Flash-Lite | $0.125 | $0.75 | Cheapest |
Gemini 2.5 Pro (legacy) | $1.25 | $10.00 | Being displaced |
Batch: 50% off. Grounded search: $14 per 1k queries after 5,000 free per month. A free tier exists for most Flash models.
Chat plans
Provider | Plan | Price | Notes |
|---|---|---|---|
OpenAI | ChatGPT Free | $0 | |
OpenAI | ChatGPT Go | $8/mo | |
OpenAI | ChatGPT Plus | $20/mo | |
OpenAI | ChatGPT Pro | $100/mo or $200/mo | 5x or 20x usage |
OpenAI | ChatGPT Business | $25/user/mo | $20 annual, min 2 users. Renamed from Team in April 2026 |
OpenAI | ChatGPT Enterprise | Custom | |
Free | $0 | ||
AI Pro | $19.99/mo | ||
AI Ultra | $99.99/mo or $199.99/mo | 5x or 20x usage. 20x tier cut from $249.99 at I/O 2026 |
What this costs at a real workload
The standardized scenario: 1,000 requests per day at roughly 1,000 input and 1,000 output tokens each, which works out to about 30M input and 30M output tokens per month. Monthly cost = 30 x input rate + 30 x output rate.
Model | Monthly cost |
|---|---|
Claude Fable 5 (reference) | $1,800 |
gpt-5.6-sol / gpt-5.5 | $1,050 |
gpt-5.6-terra / gpt-5.4 | $525 |
Gemini 3.1 Pro | $420 |
Gemini 3.5 Flash | $315 |
gpt-5.6-luna | $210 |
DeepSeek V4 Flash (reference) | $12.60 |
Tokenizer caveat for the Claude rows: Fable 5, Sonnet 5, and Opus 4.7+ use a new tokenizer that produces roughly 30% more tokens for the same text. Comparing per-token rates across providers understates Anthropic's effective cost per request by about that margin. Rate cards don't tell you what the invoice will say. The tokenizer does.
At the flagship level, Gemini 3.1 Pro runs $420 per month against $1,050 for gpt-5.6-sol, a 2.5x gap. At the value end, gpt-5.6-luna at $210 sits between Gemini 3.5 Flash at $315 and the much cheaper Flash-Lite.
The spread across all six providers is 143x between the most and least expensive model for the same workload. The full 15-row table is on the AI Pricing Index. For how per-token rates turn into unit economics, see token economics.
Timeline of past updates
OpenAI
July 9, 2026: gpt-5.6 family launch. Sol at $5 / $30, terra at $2.50 / $15, luna at $1 / $6
April 24, 2026: gpt-5.5 launch at $5 / $30, doubling flagship input over gpt-5.4
April 9, 2026: ChatGPT Pro $100 tier added below the $200 tier
April 2, 2026: Team plan renamed Business, $25/user/mo
March 5, 2026: gpt-5.4 family launch at $2.50 / $15, doubling over gpt-5
August 2025: GPT-5 launch at $1.25 / $10
Trajectory: Flagship input went $1.25 to $2.50 to $5.00 in under a year. 4x.
May 2026 (I/O): AI Ultra restructured, $99.99 5x tier added, 20x tier cut from $249.99 to $199.99
May 19, 2026: Gemini 3.5 Flash at $1.50 / $9. 3x over Gemini 3 Flash
February 19, 2026: Gemini 3.1 Pro preview at $2 / $12, 3.1 Flash-Lite at $0.125 / $0.75
December 17, 2025: Gemini 3 Flash preview at $0.50 / $3, up 67% from 2.5 Flash
November 18, 2025: Gemini 3 Pro preview at $2 / $12, up from 2.5 Pro's $1.25 / $10
June 2025: Gemini 2.5 Flash GA at $0.30 / $2.50
Trajectory: The Flash tier is 5x more expensive than 13 months ago. The era of the ever-cheaper budget model ended here first.
The AI Pricing Index
Provider | Price events (12mo) | Flagship price direction | Reprice risk for builders |
|---|---|---|---|
OpenAI | 6 | Up 4x | High |
6 | Up 5x (Flash tier) | High |
Both providers sit at the high end of the index. OpenAI's flagship input went from $1.25 to $5.00 in under a year. Google's Flash tier costs 5x what it did 13 months ago. If you build on either, budget for a price event roughly every two months, in either direction.
The index counts documented list-price events per provider over the trailing 12 months, recounted monthly. Full methodology and all six provider timelines are on the AI Pricing Index.
What this means for your own pricing
If you build on OpenAI, Google, or both, this pricing complexity becomes your pricing complexity. Four problems show up regardless of which side you pick.
The margin problem. Every user interaction has variable cost. A power user generating long responses on a flagship model costs 10-50x more than a casual user on a value model. Per-seat pricing doesn't see this. You need usage-aware billing to protect margins.
The model mix problem. Products route different requests to different models, and the table above shows the cost range that creates inside a single product. Your billing system needs to know which model served which request and price accordingly.
The credit translation problem. Many AI products abstract provider costs with credit-based pricing. That means maintaining a conversion layer from usage to credits to dollars. Every provider price event forces you to re-derive the conversion rates.
The visibility problem. Finance needs margin by customer, by feature, by model, by month. If your metering system lives apart from your billing system, that reconciliation happens in spreadsheets.
The index above adds the repricing angle: these two providers logged 12 price events in 12 months between them. Each one forces you to re-derive your own margins, credit conversion rates, and FX exposure. Solvimon runs that layer: metering per provider, credits and wallets as first-class primitives, rate cards you update without an engineering sprint. See Solvimon for AI.
Related
OpenAI vs Anthropic pricing. Two flagships priced within dollars, and a tokenizer that changes the math.
Gemini vs DeepSeek pricing. The two price-aggressive providers, compared on the same workload.
Ready to Solve Monetization?
Solvimon monetizes small and large companies alike to drive more revenue through effective pricing and billing.
Why Solvimon
AI monetization that drives innovation
The Solvimon platform is extremely flexible allowing us to bill the most tailored enterprise deals automatically.
Ciaran O'Kane
Head of Finance
Solvimon is not only building the most flexible billing platform in the space but also a truly global platform.
Juan Pablo Ortega
CEO
I was skeptical if there was any solution out there that could relieve the team from an eternity of manual billing. Solvimon impressed me with their flexibility and user-friendliness.
János Mátyásfalvi
CFO
Working with Solvimon is a different experience than working with other vendors. Not only because of the product they offer, but also because of their very senior team that knows what they are talking about.
Steven Burgemeister
Product Lead, Billing



