DeepSeek announced the general availability release of V4 Pro 0813, a large-scale mixture-of-experts model designed for production workloads.

The model features a 1-million-token context window and is available through OpenRouter's API platform. Pricing is set at $0.435 per million input tokens and $0.87 per million output tokens, reflecting competitive positioning in the market for efficient inference.

V4 Pro 0813 launched on August 12, 2026. The model supports OpenAI-compatible API calls, enabling developers to integrate it using standard SDKs with only a base URL change. OpenRouter provides monitoring infrastructure, performance metrics (throughput, latency, time-to-first-token), and benchmark comparisons against other hosted models.

The release includes standard capabilities such as tool calling and structured output support. The mixture-of-experts architecture allows the model to optimize computation by activating only relevant expert networks for each task, a design pattern increasingly adopted across large language models for efficiency gains.