DeepSeek announced the general availability release of V4 Pro 0813, a large-scale mixture-of-experts model designed for production workloads.
The model features a 1-million-token context window and is available through OpenRouter's API platform. Pricing is set at $0.435 per million input tokens and $0.87 per million output tokens, reflecting competitive positioning in the market for efficient inference.
V4 Pro 0813 launched on August 12, 2026. The model supports OpenAI-compatible API calls, enabling developers to integrate it using standard SDKs with only a base URL change. OpenRouter provides monitoring infrastructure, performance metrics (throughput, latency, time-to-first-token), and benchmark comparisons against other hosted models.
The release includes standard capabilities such as tool calling and structured output support. The mixture-of-experts architecture allows the model to optimize computation by activating only relevant expert networks for each task, a design pattern increasingly adopted across large language models for efficiency gains.
Comments
No comments yet — be the first.
Open the discussion
No account or password needed — just enter your e-mail and we’ll send you a one-time sign-in link. First time here? You’re set up automatically.
Your rating will be applied automatically after you sign in.
Check your inbox
We’ve sent a sign-in link to …. Open it on this device — this tab will sign you in automatically.
Nothing arrived? Check your spam folder — and mark the mail as "Not spam" so it lands in your inbox next time.