Choosing a cheap API model can be daunting, especially with the August 2026 Gemini plan ladder overhaul, recent renames, and price cuts. If you're looking to get the lowest API cost while balancing features like Deep Research, Flow credits, and storage limits, this guide will help you pick smartly. We'll focus on key pricing points such as gemini-3.1-flash-lite cheapest, versions offering batch or flex at half price, and why you should avoid Pro for budget.
August 2026 Gemini Plan Ladder and Pricing Overview
As of August 2026, the Gemini API plans saw a strategic restructuring. The new tier ladder reflects the industry's push for more flexible and affordable pricing while maintaining premium capabilities at the top end. Here's the high-level breakdown:
Plan Price Key Features Usage Limits Free $0 Basic API access, 100K tokens/month 100,000 tokens/month, 10 GB storage Flash-Lite $15 / month gemini-3.1-flash-lite cheapest model, batch & flex support 1 million tokens/month, 50 GB storage Batch & Flex $0.005 / 1K tokens Half price API calls compared to Flash-Lite, no storage Pay-as-you-go, no storage included Pro $49 / month Full Gemini-3.1 features, Deep Research credits 5 million tokens/month, 200 GB storage Ultra (5x) $99 / month Ultra speed, 5x faster token throughput 10 million tokens/month, 500 GB storage Ultra (20x) $249 / month Top tier, 20x faster token throughput Unlimited tokens, 1 TB storageKey Takeaway:
Free is obviously $0, but limited in tokens and features. Flash-Lite (gemini-3.1-flash-lite) hits the sweet spot for lowest-cost casual API use with 1 million tokens/month and modest storage—all for $15. The Batch & Flex versions offer a separate flexible, pay-per-use pricing at half the cost of Flash-Lite's token rate but come without storage or advanced features.
Recent Renames and Price Cuts: What Changed?
Since the last quarter, OpenAI’s pricing saw several important changes:
- Ultra split into two tiers: Previously, Ultra was a single $249 plan, which confused many budget-driven users. Now, there’s Ultra (5x) at $99 and Ultra (20x) at $249, providing clear speed and cost choices. Batch and Flex half-price token rate: The token costs for batch or flex usage were halved, making them the cheapest options once you factor in high-volume token consumption—but note they lack storage and Deep Research credits. Pro renamed and repriced: What was once called “Premium” or “Standard Pro” is now ‘Pro' at $49. Many budget-conscious API users report better cost efficiency elsewhere. Storage bundles clarified: Storage tiers now align tightly with usage limits, preventing hidden overage fees.
Price & Tier Renames Summary
Old Name New Name Price Before Price Now Ultra Ultra (5x) / Ultra (20x) $249 Single Tier $99 / $249 Premium Pro $59 / Month $49 / Month Batch & Flex (Token Rate) Batch & Flex (Token Rate) $0.01 / 1K tokens $0.005 / 1K tokensUsage Limits vs. Features: Deep Research, Flow Credits, and Storage
Understanding what you get for the price is key to avoiding overspending on features you don’t need:
- Deep Research credits: Exclusive to Pro and above tiers. Useful if your app needs semantic search and extended context windows but costly if used in volume. Flow credits: These are a form of API credit used to unlock enhanced streaming and multitasking flows. Typically bundled only with Pro and Ultra plans. Storage: Starting from Free’s 10 GB, storage increases with tier—50 GB at Flash-Lite, 200 GB at Pro, and up to 1 TB on Ultra 20x.
In general, if your use case is primarily token-driven API calls with minimal advanced features or storage, Flash-Lite is best. But if you need Deep Research or Flow credits, you’ll have to weigh if a Pro plan fits your budget or if you can limit those features.
Why Avoid Pro for Budget?
While Pro at $49/month offers great feature breadth, it’s too expensive for simple token usage when compared to Flash-Lite ($15) or pay-as-you-go Batch & Flex. Pro bundles in Deep Research credits and big storage, which inflate the cost. For startups or hobby projects aiming to optimize costs, it’s better to pick Flash-Lite or the Batch & Flex model, then layer on storage or Deep Research only when necessary.
Ultra Split into 5x and 20x Tiers - What Does This Mean?
The Ultra plan was the premium $249 plan providing ultra-high throughput and unlimited tokens. Now it’s divided into two:
Ultra (5x) at $99/month: A fast 5x speed tier with 10 million tokens/month and moderate storage. Great for medium workloads requiring accelerated responses without max budget. Ultra (20x) at $249/month: The original full-power tier, offering max throughput and unlimited tokens, plus 1 TB storage—for enterprise-grade workloads.This split means smaller teams or heavier users can now pick a more tailored cost-speed combo. It also encourages users who only need moderate speeds to save by not paying for full Ultra 20x capacity.
Example Price Comparison: Picking the Cheapest Lowest-Cost Model
Model Price Token Rate Storage Features Free $0 Included 100K tokens 10 GB Basic API access gemini-3.1-flash-lite cheapest $15 / month 1 million tokens/month 50 GB Batch & Flex enabled Batch or Flex (half price) Pay-as-you-go $0.005 / 1K tokens Pay per token None included Cheap tokens, no features Pro (avoid for budget) $49 / month 5 million tokens/month 200 GB Deep Research, Flow creditsFinal Recommendations
- If your goal is lowest monthly cost with modest usage and storage: Pick gemini-3.1-flash-lite at $15/month. It’s the cheapest full-features model with decent storage and batch/flex support. If you want truly minimal fixed cost and pay only for token usage: Use Batch or Flex half-price pay-as-you-go API calls at $0.005/1K tokens. But add external storage if needed. Avoid Pro at $49/month if you don’t need Deep Research or Flow credits to save money. Ultra tiers: Choose only if you need high throughput and generous token limits. Ultra (5x) offers a more affordable middle-ground.
Picking an API plan is about matching your workload and budget flexibility to the right tier. August 2026’s Gemini pricing changes introduced clarity, halved token price for batch/flex, and split Ultra for granularity. Keep your usage profile in mind and monitor your storage needs to avoid surprise costs.
Summary
If API cost suprmind optimization is your #1 priority, start with the free tier and upgrade to gemini-3.1-flash-lite cheapest at $15/month. If your token volume is high but storage & features simple, Batch or Flex at half price per token beats Pro for budget. Avoid Pro if you don’t need bundled research and flow credits. Ultra tier splits now let you scale throughput with cost, but it’s overkill for low-cost scenarios.


Stay tuned for updates—pricing policies evolve rapidly, and the best deal today may be reshaped come next quarter.