
DeepSeek V4 Flash: The $0.28 Open-Weight Model That Beats Its Own Pro at Agentic Coding — No GPU Required
DeepSeek re-post-trained V4-Flash and it now beats the larger V4-Pro on every published agentic benchmark at roughly a third of the output price. Same 284B/13B architecture, 1M-token context, MIT-licensed weights, $0.28 per million output tokens. Here are the real numbers, the cache-pricing trick that changes agent economics, what self-hosting actually costs, and how to decide whether to migrate.
Read the full analysis →
