◆ AI Article Pricer Compare AI article generation costs
Google · 10 models tracked

Gemini API Pricing Calculator

Price an article on Gemini Pro, Flash and Flash-Lite, including the promotional rates and the 200K context threshold.

Catalog rates verified . See verification and correction history.

Calculate here

Your Gemini article budget

Estimated API cost / article$0.00082

Monthly API budget$0.0819

The cheapest Gemini model at these settings — Gemini 3.1 Flash-Lite costs $0.2203/mo more.

Normal synchronous pricing.

Every tier this model offers, at these settings
  • Standard selected $0.00082 $0.0819 / month · reference rate
  • Flex $0.00041 $0.0410 / month · 50% off · saves $0.0410/mo
  • Batch $0.00041 $0.0410 / month · 50% off · saves $0.0410/mo
Brief, drafts and token assumptions

Base text-token estimate, not an invoice. Excludes reasoning beyond the word estimate, output overshoot, cache writes/storage, tools, regional surcharges and editing. Unsupported tiers use Standard and say so. Check the calculation assumptions and budget for human editing.

Compare this workload across all providers →

The short answer

Lowest estimated API cost: Gemini 2.5 Flash-Lite

One 1,500-word article costs $0.00082 on Gemini 2.5 Flash-Lite, Google's cheapest tracked model, and $0.0242 on Gemini 3.1 Pro, its most expensive. On Gemini 2.5 Flash-Lite, Batch costs $0.00041 at these assumptions.

This is a price ranking, not a writing-quality recommendation. Check the model-selection guide and real AI-written article examples before deciding what fits your editorial workflow.

Cheapest article
$0.00082
Dearest article
$0.0242
Models
10
Avg out/in
6.2×
Bar chart of the API cost of one 1,500-word article on the 10 tracked Google models, from $0.00082 on Gemini 2.5 Flash-Lite to $0.0242 on Gemini 3.1 Pro, at rates verified 2026-09-24.
Every tracked Google model at 1,500 words and one draft, against Google's own published rates.
Download chart
Cost by article length

What Gemini charges for a real draft

One draft, a 250-word brief, 20% prompt overhead. Change any of it in the calculator.

Model500w1,000w1,500w2,000w3,000w
Gemini 2.5 Flash-Lite$0.00030$0.00056$0.00082$0.00108$0.00160
Gemini 3.1 Flash-Lite$0.00107$0.00205$0.00302$0.00400$0.00595
Gemini 3.5 Flash-Lite$0.00174$0.00337$0.00499$0.00662$0.00987
Gemini 2.5 Flash$0.00174$0.00337$0.00499$0.00662$0.00987
Gemini 3.8 Flash$0.00273$0.00517$0.00760$0.0100$0.0149
Gemini 3.7 Flash$0.00273$0.00517$0.00760$0.0100$0.0149
Gemini 3.6 Flash$0.00273$0.00517$0.00760$0.0100$0.0149
Gemini 3.5 Flash$0.00643$0.0123$0.0181$0.0240$0.0357
Gemini 2.5 Pro$0.00699$0.0135$0.0200$0.0265$0.0395
Gemini 3.1 Pro$0.00858$0.0164$0.0242$0.0320$0.0476

Cheapest first. These are list rates applied to a token model — see the methodology for exactly where that diverges from an invoice.

The 1,000w column is the per-1,000-words figure people quote: $0.00056 on Gemini 2.5 Flash-Lite at these settings.

How Google prices

What is specific to this provider

Providers do not price the same way, and the differences change which model wins.

Rate structure

The shape of the lineup

Google's lineup is the most aggressive at the cheap end: the Flash and Flash-Lite tiers are priced for volume, and Flash-Lite is among the least expensive ways to generate a full-length draft from any provider. The trade is a wider output premium — output typically costs eight times input, against five times on Claude.

Processing tiers

What discounts exist

Google offers Standard, Flex and Batch. Flex and Batch are both half price — Flex runs synchronously with slower scheduling, Batch asynchronously within 24 hours — so the decision is simply whether the job can wait. Some Gemini models keep the Standard cache-read rate on both. Google's 1.8× Priority tier is not modeled here.

Context pricing

Where the rate changes

The Pro models change rate above 200K input tokens — well under the 1M window they advertise. The rates below are the under-threshold rates, which is what an article workflow pays; the footnotes give the higher figures. This catches people who assume a 1M window means one price for 1M tokens.

Watch out

What to check before you budget

Gemini 3.8 Flash and Gemini 3.7 Flash and Gemini 3.6 Flash are on promotional rates that end 2026-12-31, after which they roughly double. They are attractive precisely because of that discount, so treat any annual budget built on them as provisional and re-check the footnotes below.

Reference

Google list prices

USD per million tokens, cheapest output first.

ModelInputCached in OutputContext
Gemini 2.5 Flash-Lite $0.10 $0.010 $0.40 1,000K
Gemini 3.1 Flash-Lite Text, image and video input rate; audio input is $0.50. $0.25 $0.025 $1.50 1,000K
Gemini 3.5 Flash-Lite $0.30 $0.030 $2.50 1,000K
Gemini 2.5 Flash $0.30 $0.030 $2.50 1,000K
Gemini 3.8 Flash Promotional rate through 2026-12-31; $1.50 / $7.50 from 2027-01-01. $0.75 $0.075 $3.75 1,000K
Gemini 3.7 Flash Promotional rate through 2026-12-31; $1.50 / $7.50 from 2027-01-01. $0.75 $0.075 $3.75 1,000K
Gemini 3.6 Flash Promotional rate through 2026-12-31; $1.50 / $7.50 from 2027-01-01. $0.75 $0.075 $3.75 1,000K
Gemini 3.5 Flash $1.50 $0.150 $9.00 1,000K
Gemini 2.5 Pro Rate shown is for prompts up to 200K tokens; $2.50 / $15.00 above that. $1.25 $0.125 $10.00 1,000K
Gemini 3.1 Pro Rate shown is for prompts up to 200K tokens; $4.00 / $18.00 above that. $2.00 $0.200 $12.00 1,000K

Source: Google's published pricing. Compare against the other providers on the full pricing table.

Common questions

Google pricing questions

Answers specific to this provider's rate structure.

Which Gemini model is cheapest for article writing?

Gemini 2.5 Flash-Lite is the cheapest tracked Gemini model for a full draft, and among the cheapest from any provider. The Flash and Flash-Lite tiers are Google’s volume play and are priced accordingly.

What is the 200K context threshold on Gemini Pro?

The Pro models change rate above 200K input tokens, despite advertising a 1M window. The rates on this page are the under-threshold rates, which is what an article workflow pays. The advertised context and the single-rate context are not the same number, which catches a lot of people budgeting from the headline figure.

Are the Gemini Flash prices permanent?

No. Three Flash models are on promotional rates that expire, after which they double. They are attractive largely because of that discount, so an annual budget built on them should be treated as provisional. The footnotes on this page give the post-promotional figures.

Does Gemini have a batch tier?

Yes. Batch input and output are 50% off on every tracked Gemini model, and Google also sells Flex — slower synchronous processing — at the same half price. Some models keep the Standard cache-read rate on both. Google’s 1.8× Priority tier is not modeled here.

Is Gemini cheaper than GPT for content?

At the volume end they are close, and which one wins moves with promotional rates. Gemini carries a wider output premium — output averages over six times input — which matters more for article writing than for chat, because a draft is almost entirely output.

The bottom line

Gemini is one column in a five-column decision

Price your workflow here, then check it against the other providers before committing. At article volumes the difference between two reasonable models is usually smaller than a single hour of editing. When the work can run asynchronously, compare Gemini's batch rate — not its standard rate — against DeepSeek, where no Batch tier is published. DeepSeek instead offers a separate off-peak discount.