<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>CostMon — The Cloud &amp; AI Price Change Log</title><description>Every cloud and AI price change that actually moved a bill, sourced from the provider that made it — cuts and increases both, with the source and what to do about it.</description><link>https://costmon.com</link><item><title>OpenAI cut GPT-5.6 Luna&apos;s price 80% and Terra&apos;s 20%, leaving the flagship tier untouched.</title><link>https://costmon.com/price-changes#gpt-5-6-luna-terra-price-cut</link><guid isPermaLink="true">https://costmon.com/price-changes#gpt-5-6-luna-terra-price-cut</guid><description>Citing inference-stack efficiency gains, OpenAI dropped Terra to $2/$12 per million tokens and Luna to $0.20/$1.20, while the top-tier Sol model&apos;s price stayed flat. (Luna −80%, Terra −20%) — Confirm the lower rate appears on your next invoice, then check whether tasks still pinned to Sol can safely move down to the now much cheaper Luna or Terra.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><category>OpenAI</category><category>Cut</category></item><item><title>GitHub replaced Copilot&apos;s flat premium-request billing with usage-based AI credits.</title><link>https://costmon.com/price-changes#github-copilot-ai-credits-2026</link><guid isPermaLink="true">https://costmon.com/price-changes#github-copilot-ai-credits-2026</guid><description>Every Copilot plan now gets a monthly allotment of GitHub AI Credits consumed by actual token usage, instead of a flat per-request charge that billed a quick question the same as a multi-hour agent session. ($10-$39 credits/mo) — Track per-seat credit burn from June 1, 2026, and budget for the promotional credit boost on Business/Enterprise expiring after August 2026.</description><pubDate>Mon, 27 Apr 2026 00:00:00 GMT</pubDate><category>GitHub</category><category>Restructure</category></item><item><title>GPT-5.5 launched priced higher than GPT-5.4, betting token efficiency would offset it.</title><link>https://costmon.com/price-changes#gpt-5-5-price-increase</link><guid isPermaLink="true">https://costmon.com/price-changes#gpt-5-5-price-increase</guid><description>GPT-5.5 launched at roughly double GPT-5.4&apos;s per-token rate. OpenAI&apos;s own announcement frames this as a deliberate increase, offset by claimed gains in token efficiency per task. ($5 / $30 per M tokens) — Run a side-by-side cost comparison on real workloads before moving production traffic. Confirm the token savings offset the higher headline price for you.</description><pubDate>Fri, 24 Apr 2026 00:00:00 GMT</pubDate><category>OpenAI</category><category>Increase</category></item><item><title>Vercel cut Turbo build machine pricing 16% and unified every tier to one per-CPU rate.</title><link>https://costmon.com/price-changes#vercel-turbo-build-machines-price-cut-2026</link><guid isPermaLink="true">https://costmon.com/price-changes#vercel-turbo-build-machines-price-cut-2026</guid><description>All build machine tiers now bill at a flat $0.0035 per CPU per minute, dropping the 30-CPU Turbo machine from $0.126 to $0.105 a minute. (−16% ($0.126→$0.105/min)) — Confirm the lower rate applied on your next invoice, and re-check whether Turbo is still worth it over Standard or Enhanced now that the price gap has narrowed.</description><pubDate>Wed, 15 Apr 2026 00:00:00 GMT</pubDate><category>Vercel</category><category>Cut</category></item><item><title>VPC Encryption Controls stopped being a free preview and started billing per VPC-hour.</title><link>https://costmon.com/price-changes#aws-vpc-encryption-controls-new-charge</link><guid isPermaLink="true">https://costmon.com/price-changes#aws-vpc-encryption-controls-new-charge</guid><description>After a free preview that ran from November 2025, AWS began charging a fixed hourly rate for every non-empty VPC running Encryption Controls in monitor or enforce mode. Empty VPCs stay free. The announcement itself doesn&apos;t print a rate. AWS&apos;s VPC pricing page puts it at $0.15 per hour in US East (N. Virginia), rising to $0.31 in the priciest regions. ($0.15 / VPC-hour (us-east-1)) — Audit which VPCs need enforce mode versus monitor mode, and turn it off on non-empty VPCs where it isn&apos;t earning its per-hour charge. Check your own region&apos;s rate too; it&apos;s up to twice the US East price.</description><pubDate>Sun, 01 Mar 2026 00:00:00 GMT</pubDate><category>AWS</category><category>New charge</category></item><item><title>Gemini 3 Flash turned on context caching by default, cutting repeated-token cost up to 90%.</title><link>https://costmon.com/price-changes#gemini-3-flash-caching-batch-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-3-flash-caching-batch-pricing-2025</guid><description>Gemini 3 Flash launched at $0.50 per million input tokens and $3 per million output tokens, with context caching standard rather than opt-in, plus a 50% discount for asynchronous Batch API jobs. (up to −90% w/ caching) — Confirm caching is active on your repeated prompts, and route non-interactive jobs through the Batch API to capture the 50% discount.</description><pubDate>Wed, 17 Dec 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Cut</category></item><item><title>Snowflake moved Snowpipe to one flat per-GB rate, cutting ingestion cost for some teams over 50%.</title><link>https://costmon.com/price-changes#snowflake-snowpipe-unified-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#snowflake-snowpipe-unified-pricing-2025</guid><description>Snowpipe pricing, previously based on compute consumed and file count, moved to a single 0.0037-credits-per-GB rate across file ingestion and streaming. (flat 0.0037 credits/GB) — Check ingestion cost after the automatic switchover, and stop over-batching files purely to dodge per-file charges. Billing is now per-GB regardless of file count.</description><pubDate>Mon, 08 Dec 2025 00:00:00 GMT</pubDate><category>Snowflake</category><category>Cut</category></item><item><title>Claude Opus 4.5 cut Opus-tier pricing to $5/$25 per million tokens.</title><link>https://costmon.com/price-changes#claude-opus-4-5-pricing-cut</link><guid isPermaLink="true">https://costmon.com/price-changes#claude-opus-4-5-pricing-cut</guid><description>Opus 4.5 launched at $5 input / $25 output per million tokens, a cut Anthropic frames as putting frontier-tier capability within reach of more teams. ($5 / $25 per M tokens) — Re-evaluate workloads kept on Sonnet purely for cost. At the new rate, Opus may be affordable for tasks where the extra capability wasn&apos;t worth the old premium.</description><pubDate>Mon, 24 Nov 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Cut</category></item><item><title>Gemini 3 Pro launched as Google&apos;s new flagship, at a higher output rate than 2.5 Pro.</title><link>https://costmon.com/price-changes#gemini-3-pro-launch-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-3-pro-launch-pricing-2025</guid><description>Google&apos;s new flagship reasoning and agentic model launched at $2 per million input tokens and $12 per million output tokens for prompts up to 200K tokens. ($2 / $12 per M tokens) — Benchmark real task cost against your current model before switching. The higher output rate needs to be earned back in fewer tokens per task, not assumed.</description><pubDate>Tue, 18 Nov 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Restructure</category></item><item><title>Cloud Build repriced its e2 machines by region, raising build minutes in some and trimming others.</title><link>https://costmon.com/price-changes#cloud-build-e2-regional-price-update-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#cloud-build-e2-regional-price-update-2025</guid><description>Per-build-minute pricing for e2 machine types moved from one flat rate to a region-by-region table: us-east1 (South Carolina) went up across e2-medium, e2-highcpu-8, and e2-highcpu-32, while us-central1 (Iowa) went down slightly on the same machine types. (e.g. $0.064→$0.0705/build-min (us-east1)) — Check your build pool&apos;s region against the new per-region rate table, and shift latency-insensitive builds to a cheaper region if yours got more expensive.</description><pubDate>Wed, 15 Oct 2025 00:00:00 GMT</pubDate><category>Google Cloud</category><category>Increase</category></item><item><title>Claude Haiku 4.5 launched at a third of Sonnet 4&apos;s price for comparable coding performance.</title><link>https://costmon.com/price-changes#claude-haiku-4-5-pricing</link><guid isPermaLink="true">https://costmon.com/price-changes#claude-haiku-4-5-pricing</guid><description>The new small model launched at a flat $1 input / $5 output per million tokens, positioned by Anthropic as Sonnet-4-level coding at a third of the cost. ($1 / $5 per M tokens) — Benchmark current Sonnet-4-class workloads against Haiku 4.5 and shift eligible traffic down a tier, checking quality on your own tasks first.</description><pubDate>Wed, 15 Oct 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Restructure</category></item><item><title>Azure cut Ultra Disk IOPS and throughput pricing in UK South by up to 80%.</title><link>https://costmon.com/price-changes#azure-ultra-disk-price-cut-uk-south-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#azure-ultra-disk-price-cut-uk-south-2025</guid><description>In UK South, provisioned IOPS on Ultra Disk dropped from $0.06205 to $0.02482 per month, and provisioned throughput from $0.40588 to $0.08103 per MBPS per month. Per-GiB capacity pricing didn&apos;t move. (−60% IOPS / −80% throughput) — Re-run the disk cost estimate for UK South workloads at the new per-unit rates before assuming last quarter&apos;s numbers still hold.</description><pubDate>Tue, 02 Sep 2025 00:00:00 GMT</pubDate><category>Microsoft Azure</category><category>Cut</category></item><item><title>GKE folded paid multi-cluster management into the Standard tier at no extra charge.</title><link>https://costmon.com/price-changes#gke-single-tier-multicluster-free-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gke-single-tier-multicluster-free-2025</guid><description>As part of moving to a single paid GKE tier, Fleets, Teams, Config Management, and Policy Controller are now included with GKE Standard. All four were previously gated behind a pricier tier or add-on. (fleet mgmt now $0) — Check your current GKE edition. If you upgraded purely for fleet management or Policy Controller, you can likely drop back to Standard and keep the feature.</description><pubDate>Tue, 26 Aug 2025 00:00:00 GMT</pubDate><category>Google Cloud</category><category>Free tier</category></item><item><title>GPT-5 replaced the GPT-4o/o-series lineup with one model family sold at three price points.</title><link>https://costmon.com/price-changes#gpt-5-launch-pricing</link><guid isPermaLink="true">https://costmon.com/price-changes#gpt-5-launch-pricing</guid><description>GPT-5 launched at $1.25 per million input tokens and $10 per million output tokens, with mini and nano tiers underneath it: one family instead of separate reasoning and non-reasoning models. ($1.25 / $10 per M tokens) — Map existing GPT-4o and o3 call sites to gpt-5, gpt-5-mini, or gpt-5-nano by task difficulty, and re-benchmark spend rather than assuming a like-for-like swap.</description><pubDate>Thu, 07 Aug 2025 00:00:00 GMT</pubDate><category>OpenAI</category><category>Restructure</category></item><item><title>Vercel switched Functions billing to active CPU time instead of full wall-clock duration.</title><link>https://costmon.com/price-changes#vercel-active-cpu-pricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#vercel-active-cpu-pricing-2025</guid><description>A Standard-size function running at 100% active CPU now costs about $0.149 an hour instead of $0.318, because idle time waiting on a database call or an LLM response no longer bills as CPU time. (−53% ($0.318→$0.149/hr)) — Re-run cost estimates for Functions-heavy AI workloads under the new model rather than assuming automatic savings, and confirm your plan has switched over.</description><pubDate>Wed, 25 Jun 2025 00:00:00 GMT</pubDate><category>Vercel</category><category>Cut</category></item><item><title>Gemini 2.5 Flash&apos;s stable release raised input price but cut output price and merged two rate tiers into one.</title><link>https://costmon.com/price-changes#gemini-2-5-flash-stable-repricing-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#gemini-2-5-flash-stable-repricing-2025</guid><description>Going stable, Google removed 2.5 Flash&apos;s separate thinking vs. non-thinking price tiers, raising input from $0.15 to $0.30 per million tokens while cutting output from $3.50 to $2.50. (in +100%, out −29%) — Recompute cost-per-request at the new $0.30/$2.50 blended rate before migrating off the preview model. Input-heavy workloads got pricier even though output-heavy ones got cheaper.</description><pubDate>Tue, 17 Jun 2025 00:00:00 GMT</pubDate><category>Google (Gemini)</category><category>Restructure</category></item><item><title>AWS cut On-Demand GPU instance prices by up to 45%.</title><link>https://costmon.com/price-changes#aws-ec2-gpu-instances-price-cut-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#aws-ec2-gpu-instances-price-cut-2025</guid><description>On-Demand pricing dropped across the P5, P5en, P4d, and P4de instance families: P5 by up to 45%, P5en by up to 26%, P4d/P4de by up to 33%. (up to −45% (P5)) — Compare the new On-Demand rate against your existing Savings Plan commitment. It may no longer be worth the lock-in.</description><pubDate>Thu, 05 Jun 2025 00:00:00 GMT</pubDate><category>AWS</category><category>Cut</category></item><item><title>Anthropic cut tool-use output tokens on Claude 3.7 Sonnet by up to 70%.</title><link>https://costmon.com/price-changes#anthropic-token-saving-updates-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#anthropic-token-saving-updates-2025</guid><description>A new token-efficient tool-use mode cut output tokens consumed on tool-calling workloads, and prompt-cache reads stopped counting against input-token-per-minute rate limits. (up to −70% output tokens) — Turn on token-efficient tool use for tool-calling workloads, and check whether old throttling logic built around the rate limit is still needed.</description><pubDate>Thu, 13 Mar 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Cut</category></item><item><title>AWS cut GuardDuty&apos;s S3 malware-scanning price by 85%.</title><link>https://costmon.com/price-changes#aws-guardduty-malware-protection-s3-cut</link><guid isPermaLink="true">https://costmon.com/price-changes#aws-guardduty-malware-protection-s3-cut</guid><description>The per-GB price for the data-scanned dimension of Malware Protection for S3 dropped from $0.60 to $0.09 in US East (N. Virginia), applied automatically to existing customers. (−85% ($0.60→$0.09/GB)) — Re-run the cost projection for buckets you skipped scanning because it was too expensive. Coverage that didn&apos;t pencil out before likely does now.</description><pubDate>Thu, 06 Feb 2025 00:00:00 GMT</pubDate><category>AWS</category><category>Cut</category></item><item><title>MongoDB merged its Shared and Serverless tiers into one Flex tier with a capped bill.</title><link>https://costmon.com/price-changes#mongodb-atlas-flex-tier-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#mongodb-atlas-flex-tier-2025</guid><description>Atlas Flex bills an $8 base fee plus usage up to 500 ops/sec, capped at $30 a month. It replaces both the old Shared clusters and standalone Serverless instances. ($8 base, capped $30/mo) — Plan the migration off Shared or Serverless ahead of MongoDB&apos;s cutover, and model your peak burst usage against the $30 cap before assuming it still fits.</description><pubDate>Thu, 06 Feb 2025 00:00:00 GMT</pubDate><category>MongoDB Atlas</category><category>Restructure</category></item><item><title>GitHub made Copilot free in VS Code, capped at 2,000 completions a month.</title><link>https://costmon.com/price-changes#github-copilot-free-tier-2024</link><guid isPermaLink="true">https://costmon.com/price-changes#github-copilot-free-tier-2024</guid><description>Any GitHub account gets capped code completions and chat in VS Code: 2,000 completions and 50 chat messages a month, with no trial or subscription required. (2,000 completions free/mo) — Point casual or evaluation users at the free tier instead of provisioning a paid seat, and watch for contributors who lean on it in place of a managed seat.</description><pubDate>Wed, 18 Dec 2024 00:00:00 GMT</pubDate><category>GitHub</category><category>Free tier</category></item><item><title>Anthropic launched a batch endpoint priced at half the standard API rate.</title><link>https://costmon.com/price-changes#anthropic-message-batches-api-50-off</link><guid isPermaLink="true">https://costmon.com/price-changes#anthropic-message-batches-api-50-off</guid><description>The Message Batches API processes up to 10,000 queries asynchronously within 24 hours, at half the price of a standard synchronous call. (−50% vs standard API) — Audit API traffic for anything without a real-time latency requirement and move it to the Batches API to cut that portion of the bill in half.</description><pubDate>Tue, 08 Oct 2024 00:00:00 GMT</pubDate><category>Anthropic</category><category>Restructure</category></item><item><title>Cloud SQL began billing instances left running on end-of-life MySQL or PostgreSQL versions.</title><link>https://costmon.com/price-changes#cloud-sql-extended-support-new-charge-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#cloud-sql-extended-support-new-charge-2025</guid><description>Instances still running a MySQL or PostgreSQL major version past its community end-of-life get auto-enrolled in a paid &quot;extended support&quot; add-on, billed per vCPU-hour (or per instance-hour on shared-core). Google waived the charge through April 2025, then started billing it. (new per-vCPU/hr fee) — Check instance versions against Cloud SQL&apos;s end-of-life schedule and upgrade off any EOL major version instead of paying the ongoing surcharge.</description><pubDate>Wed, 29 May 2024 00:00:00 GMT</pubDate><category>Google Cloud</category><category>New charge</category></item><item><title>Azure zeroed out the bandwidth charge for traffic between Availability Zones in the same region.</title><link>https://costmon.com/price-changes#azure-inter-az-bandwidth-free-2024</link><guid isPermaLink="true">https://costmon.com/price-changes#azure-inter-az-bandwidth-free-2024</guid><description>Microsoft eliminated the per-GB charge for data moving across Availability Zones within a region, for both private and public IP traffic. (→ $0/GB) — Revisit any old cost-optimization advice that steered you away from zone-redundant architecture to dodge bandwidth charges. That tradeoff is gone.</description><pubDate>Tue, 21 May 2024 00:00:00 GMT</pubDate><category>Microsoft Azure</category><category>Cut</category></item><item><title>GPT-4o launched at half the price of GPT-4 Turbo, with 5x the rate limit.</title><link>https://costmon.com/price-changes#gpt-4o-half-price-vs-gpt-4-turbo</link><guid isPermaLink="true">https://costmon.com/price-changes#gpt-4o-half-price-vs-gpt-4-turbo</guid><description>OpenAI&apos;s new flagship model replaced GPT-4 Turbo as the default, matching or beating its quality at half the per-token price. (−50% vs GPT-4 Turbo) — Re-point integrations at gpt-4o and re-run your eval suite before rolling it out. A cheaper model at the same price tier can still shift output style.</description><pubDate>Mon, 13 May 2024 00:00:00 GMT</pubDate><category>OpenAI</category><category>Cut</category></item><item><title>AWS paired the new IPv4 charge with 750 free IPv4-hours a month, for new accounts only.</title><link>https://costmon.com/price-changes#aws-ipv4-free-tier-750-hours</link><guid isPermaLink="true">https://costmon.com/price-changes#aws-ipv4-free-tier-750-hours</guid><description>New AWS accounts get 750 hours a month of public IPv4 usage at no charge when they launch an EC2 instance with a public address. That&apos;s enough to cover one always-on instance. (750 IPv4-hrs/mo free) — Don&apos;t assume this offsets your bill. Check whether your account qualifies as &quot;new&quot; before crediting it against the IPv4 charge.</description><pubDate>Thu, 01 Feb 2024 00:00:00 GMT</pubDate><category>AWS</category><category>Free tier</category></item><item><title>Google Cloud raised the hourly charge on in-use external IPv4 addresses.</title><link>https://costmon.com/price-changes#gce-external-ipv4-price-increase-2024</link><guid isPermaLink="true">https://costmon.com/price-changes#gce-external-ipv4-price-increase-2024</guid><description>The External IP Charge for a standard VM rose from $0.004 to $0.005 an hour, and for a Spot VM from $0.002 to $0.0025, for every in-use external IPv4 address, on the same schedule as AWS&apos;s own IPv4 pricing move. ($0.004→$0.005/hr) — Move workloads to internal-only addressing behind Cloud NAT or a load balancer, and keep external IPs reserved for the resources that genuinely need one.</description><pubDate>Thu, 28 Sep 2023 00:00:00 GMT</pubDate><category>Google Cloud</category><category>Increase</category></item><item><title>AWS started charging for every public IPv4 address, attached or not.</title><link>https://costmon.com/price-changes#aws-public-ipv4-charge</link><guid isPermaLink="true">https://costmon.com/price-changes#aws-public-ipv4-charge</guid><description>Every public IPv4 address across AWS started accruing an hourly charge, whether it&apos;s attached to a running instance or just sitting on an Elastic IP nobody released. (+$0.005 / IP-hour) — Run VPC IP Address Manager&apos;s Public IP Insights, move workloads behind a shared NAT gateway or load balancer, and release unattached Elastic IPs.</description><pubDate>Fri, 28 Jul 2023 00:00:00 GMT</pubDate><category>AWS</category><category>New charge</category></item><item><title>BigQuery raised on-demand query pricing 25% the same day it retired flat-rate slots.</title><link>https://costmon.com/price-changes#bigquery-editions-on-demand-increase-2023</link><guid isPermaLink="true">https://costmon.com/price-changes#bigquery-editions-on-demand-increase-2023</guid><description>Google retired flat-rate annual, flat-rate monthly, and flex-slot commitments in favor of three new editions billed per-second, and raised the on-demand analysis price 25% across all regions on the same date. (+25% on-demand) — Cut bytes scanned with partitioning and clustering before the 25% increase compounds, and evaluate a slot-based edition if your on-demand spend is now significant.</description><pubDate>Wed, 29 Mar 2023 00:00:00 GMT</pubDate><category>Google Cloud</category><category>Increase</category></item><item><title>Azure raised local-currency prices in the UK, EU, and Nordics to track the USD exchange rate.</title><link>https://costmon.com/price-changes#azure-currency-price-alignment-2023</link><guid isPermaLink="true">https://costmon.com/price-changes#azure-currency-price-alignment-2023</guid><description>Microsoft moved to a twice-yearly currency-alignment cadence and raised local-currency Azure pricing to track USD: GBP up 9%, DKK/EUR/NOK up 11%, SEK up 15%. (+9% to +15% (non-USD)) — Check your billing currency and agreement type, and rebuild your forecast around the new rate instead of the one you budgeted with.</description><pubDate>Tue, 31 Jan 2023 00:00:00 GMT</pubDate><category>Microsoft Azure</category><category>Increase</category></item></channel></rss>