DeepSeek V4
3 published articles
API economics
DeepSeek V4 drops in July with a pricing shift that will hurt if you code when the sun is up
DeepSeek V4 arrives mid-July with peak-load pricing. The API costs double during business hours, up to $1.74 per million output tokens, but off-peak rates can be a steal for cache-hit workloads.
2026-07-09
Model Release
DeepSeek-V4 preview lands, and the open-weight math gets harder for everyone
DeepSeek unveils DeepSeek-V4 preview with top-tier reasoning and stronger agent abilities, now available across platforms. The release signals a major escalation in the open-weight AI race, putting pressure on OpenAI and Google to justify their closed-source strategies.
2026-07-05
Open-source AI research
Million-token context on a tenth of the KV cache: DeepSeek-V4's efficiency bet
DeepSeek's V4 preview pairs a 1.6T-parameter Pro model with a 284B Flash variant, both at one million tokens of context. The paper claims 27% of the inference FLOPs and 10% of the KV cache of V3.2, the numbers that make the context cost-effective to serve.
2026-06-22