Base text-token estimate, not an invoice. Excludes reasoning beyond the word estimate, output overshoot, cache writes/storage, tools, regional surcharges and editing. Unsupported tiers use Standard and say so. Check the calculation assumptions and budget for human editing.
Promotional rate through 2026-12-31; $1.50 / $7.50 from 2027-01-01.
One 1,500-word article costs $0.00082 on Gemini 2.5 Flash-Lite, Google's cheapest tracked model, and $0.0242 on Gemini 3.1 Pro, its most expensive. On Gemini 2.5 Flash-Lite, Batch costs $0.00041 at these assumptions.
Cheapest first. These are list rates applied to a token model — see
the methodology for exactly where that diverges from an invoice.
The 1,000w column is the per-1,000-words figure people quote: $0.00056 on Gemini 2.5 Flash-Lite at these settings.
How Google prices
What is specific to this provider
Providers do not price the same way, and the differences change which model wins.
Rate structure
The shape of the lineup
Google's lineup is the most aggressive at the cheap end: the Flash and
Flash-Lite tiers are priced for volume, and Flash-Lite is among the least expensive
ways to generate a full-length draft from any provider. The trade is a wider output
premium — output typically costs eight times input, against five times on Claude.
Processing tiers
What discounts exist
Google offers Standard, Flex and Batch. Flex and Batch are both
half price — Flex runs synchronously with slower scheduling, Batch asynchronously within
24 hours — so the decision is simply whether the job can wait. Some Gemini models keep
the Standard cache-read rate on both. Google's 1.8× Priority tier is not modeled here.
Context pricing
Where the rate changes
The Pro models change rate above 200K input tokens — well
under the 1M window they advertise. The rates below are the under-threshold rates,
which is what an article workflow pays; the footnotes give the higher figures. This
catches people who assume a 1M window means one price for 1M tokens.
Watch out
What to check before you budget
Gemini 3.8 Flash and Gemini 3.7 Flash and Gemini 3.6 Flash are on promotional rates that end
2026-12-31, after which they roughly double. They are attractive
precisely because of that discount, so treat any annual budget built on them as
provisional and re-check the footnotes below.
What these rates buy
Read an unedited Gemini draft before you budget
Every gallery below includes an unedited Gemini 2.5 Flash draft, published
blind with a current-rate cost comparison. The model and the price stay hidden until you
reveal them, so this page can name who took part without spoiling which draft is which.
Answers specific to this provider's rate structure.
Which Gemini model is cheapest for article writing?
Gemini 2.5 Flash-Lite is the cheapest tracked Gemini model for a full draft, and among the cheapest from any provider. The Flash and Flash-Lite tiers are Google’s volume play and are priced accordingly.
What is the 200K context threshold on Gemini Pro?
The Pro models change rate above 200K input tokens, despite advertising a 1M window. The rates on this page are the under-threshold rates, which is what an article workflow pays. The advertised context and the single-rate context are not the same number, which catches a lot of people budgeting from the headline figure.
Are the Gemini Flash prices permanent?
No. Three Flash models are on promotional rates that expire, after which they double. They are attractive largely because of that discount, so an annual budget built on them should be treated as provisional. The footnotes on this page give the post-promotional figures.
Does Gemini have a batch tier?
Yes. Batch input and output are 50% off on every tracked Gemini model, and Google also sells Flex — slower synchronous processing — at the same half price. Some models keep the Standard cache-read rate on both. Google’s 1.8× Priority tier is not modeled here.
Is Gemini cheaper than GPT for content?
At the volume end they are close, and which one wins moves with promotional rates. Gemini carries a wider output premium — output averages over six times input — which matters more for article writing than for chat, because a draft is almost entirely output.
The bottom line
Gemini is one column in a five-column decision
Price your workflow here, then check it against the other providers before committing.
At article volumes the difference between two reasonable models is usually smaller than a
single hour of editing. When the work can run asynchronously, compare Gemini's batch rate — not its standard rate — against DeepSeek, where no Batch tier is published. DeepSeek instead offers a separate off-peak discount.