<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>CostMon — Anthropic Price Changes</title><description>Every Anthropic price change CostMon tracks that actually moved a bill, cuts and increases both, with the source and what to do about it.</description><link>https://costmon.com</link><item><title>Claude Opus 4.5 cut Opus-tier pricing to $5/$25 per million tokens.</title><link>https://costmon.com/price-changes#claude-opus-4-5-pricing-cut</link><guid isPermaLink="true">https://costmon.com/price-changes#claude-opus-4-5-pricing-cut</guid><description>Opus 4.5 launched at $5 input / $25 output per million tokens, a cut Anthropic frames as putting frontier-tier capability within reach of more teams. ($5 / $25 per M tokens) — Re-evaluate workloads kept on Sonnet purely for cost. At the new rate, Opus may be affordable for tasks where the extra capability wasn&apos;t worth the old premium.</description><pubDate>Mon, 24 Nov 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Cut</category></item><item><title>Claude Haiku 4.5 launched at a third of Sonnet 4&apos;s price for comparable coding performance.</title><link>https://costmon.com/price-changes#claude-haiku-4-5-pricing</link><guid isPermaLink="true">https://costmon.com/price-changes#claude-haiku-4-5-pricing</guid><description>The new small model launched at a flat $1 input / $5 output per million tokens, positioned by Anthropic as Sonnet-4-level coding at a third of the cost. ($1 / $5 per M tokens) — Benchmark current Sonnet-4-class workloads against Haiku 4.5 and shift eligible traffic down a tier, checking quality on your own tasks first.</description><pubDate>Wed, 15 Oct 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Restructure</category></item><item><title>Anthropic cut tool-use output tokens on Claude 3.7 Sonnet by up to 70%.</title><link>https://costmon.com/price-changes#anthropic-token-saving-updates-2025</link><guid isPermaLink="true">https://costmon.com/price-changes#anthropic-token-saving-updates-2025</guid><description>A new token-efficient tool-use mode cut output tokens consumed on tool-calling workloads, and prompt-cache reads stopped counting against input-token-per-minute rate limits. (up to −70% output tokens) — Turn on token-efficient tool use for tool-calling workloads, and check whether old throttling logic built around the rate limit is still needed.</description><pubDate>Thu, 13 Mar 2025 00:00:00 GMT</pubDate><category>Anthropic</category><category>Cut</category></item><item><title>Anthropic launched a batch endpoint priced at half the standard API rate.</title><link>https://costmon.com/price-changes#anthropic-message-batches-api-50-off</link><guid isPermaLink="true">https://costmon.com/price-changes#anthropic-message-batches-api-50-off</guid><description>The Message Batches API processes up to 10,000 queries asynchronously within 24 hours, at half the price of a standard synchronous call. (−50% vs standard API) — Audit API traffic for anything without a real-time latency requirement and move it to the Batches API to cut that portion of the bill in half.</description><pubDate>Tue, 08 Oct 2024 00:00:00 GMT</pubDate><category>Anthropic</category><category>Restructure</category></item></channel></rss>