DeepSeek Slashes API Cache-Hit Pricing to One-Tenth, Extends V4 Pro Discount
Chinese AI company DeepSeek has long challenged OpenAI and Anthropic with low-cost APIs. Input caching allows existing content to be reused, directly affecting the operating costs of services with high volumes of repetitive requests, including RAG and agents. Developers are therefore watching the price changes closely.
DeepSeek most recently announced that input cache-hit API rates across its entire model lineup would fall to one-tenth of their previous levels. It also extended the flagship V4 Pro model’s limited-time 75% discount through May 5, 2026, bringing the discounted API output price to less than NT$30 per million tokens.
All Coverage
5 original reportsThe Backstory
The history behind this eventNo historical echoes for this signal
Subscribe to Mark Radar Weekly
Every Friday, the week's strongest signals in your inbox. Unsubscribe anytime.