DeepSeek V4 Pro 0813 is the general availability release of a large-scale mixture-of-experts model, priced at $0.435 per million input tokens and $0.87 per million output tokens, with a 1M token context window and up to 384,000 output tokens. It is accessible via OpenRouter, hosted by a single provider, with 100% average uptime over the past three days. Top applications sending traffic to the model include SillyTavern, Janitor AI, and Open WebUI. Integration is possible through OpenRouter's OpenAI-compatible API, Anthropic Messages format, or the Responses API, with support for reasoning tokens and streaming.
Questions this post answers
What is the API pricing for DeepSeek V4 Pro 0813 on OpenRouter?
DeepSeek V4 Pro 0813 costs $0.435 per million input tokens and $0.87 per million output tokens at the listed provider rate. The weighted average price customers actually pay is lower, around $0.04999 per million input tokens and $0.8698 per million output tokens, due to caching and discounts. It supports a 1,048,576 token context window and up to 384,000 output tokens. Developers comparing LLM API costs can track pricing shifts like this via daily.dev.
What is the context window and max output for DeepSeek V4 Pro 0813?
DeepSeek V4 Pro 0813 has a context window of 1,048,576 tokens (1M) and a maximum output of 384,000 tokens. It is a large-scale mixture-of-experts model released as the general availability version of DeepSeek V4 Pro on August 12, 2026, and is accessible through OpenRouter's API. Keep up with new model context and output limits relevant to your builds on daily.dev.