<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>CostMon — Google (Gemini) Price Changes</title><description>Every Google (Gemini) price change CostMon tracks that actually moved a bill, cuts and increases both, with the source and what to do about it.</description><link>https://costmon.com</link><item><title>Gemini 3 Flash turned on context caching by default, cutting repeated-token cost up to 90%.</title><link>https://costmon.com/price-changes#gemini-3-flash-caching-batch-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-3-flash-caching-batch-pricing-2025</guid><description>Gemini 3 Flash launched at $0.50 per million input tokens and $3 per million output tokens, with context caching standard rather than opt-in, plus a 50% discount for asynchronous Batch API jobs. (up to −90% w/ caching) — Confirm caching is active on your repeated prompts, and route non-interactive jobs through the Batch API to capture the 50% discount.</description><pubDate>Wed, 17 Dec 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Cut</category></item><item><title>Gemini 3 Pro launched as Google&apos;s new flagship, at a higher output rate than 2.5 Pro.</title><link>https://costmon.com/price-changes#gemini-3-pro-launch-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-3-pro-launch-pricing-2025</guid><description>Google&apos;s new flagship reasoning and agentic model launched at $2 per million input tokens and $12 per million output tokens for prompts up to 200K tokens. ($2 / $12 per M tokens) — Benchmark real task cost against your current model before switching. The higher output rate needs to be earned back in fewer tokens per task, not assumed.</description><pubDate>Tue, 18 Nov 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Restructure</category></item><item><title>Gemini 2.5 Flash&apos;s stable release raised input price but cut output price and merged two rate tiers into one.</title><link>https://costmon.com/price-changes#gemini-2-5-flash-stable-repricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-2-5-flash-stable-repricing-2025</guid><description>Going stable, Google removed 2.5 Flash&apos;s separate thinking vs. non-thinking price tiers, raising input from $0.15 to $0.30 per million tokens while cutting output from $3.50 to $2.50. (in +100%, out −29%) — Recompute cost-per-request at the new $0.30/$2.50 blended rate before migrating off the preview model. Input-heavy workloads got pricier even though output-heavy ones got cheaper.</description><pubDate>Tue, 17 Jun 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Restructure</category></item></channel></rss>