<!-- mobian-agent-page publisher="dailydev" canonical="https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb" -->

---
title: DeepSeek V4 Pro 0813 | daily.dev
description: DeepSeek V4 Pro 0813 is the general availability release of a large-scale mixture-of-experts model, priced at $0.435 per million input tokens and $0.87 per...
canonical: https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb
twitter:card: summary_large_image
twitter:site: @dailydotdev
og:type: website
og:site_name: daily.dev
og:title: DeepSeek V4 Pro 0813 | daily.dev
og:description: DeepSeek V4 Pro 0813 is the general availability release of a large-scale mixture-of-experts model, priced at $0.435 per million input tokens and $0.87 per...
og:url: https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb
og:image: https://api.daily.dev/og/posts/b1MmDMaJb.png
og:image:alt: DeepSeek V4 Pro 0813
og:image:width: 1200
og:image:height: 630
og:locale: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# DeepSeek V4 Pro 0813

**[Hacker News](https://daily.dev/sources/hn)** · 5 min read · 1 upvotes · 0 comments

## Summary

DeepSeek V4 Pro 0813 is the general availability release of a large-scale mixture-of-experts model, priced at $0.435 per million input tokens and $0.87 per million output tokens, with a 1M token context window and up to 384,000 output tokens. It is accessible via OpenRouter, hosted by a single provider, with 100% average uptime over the past three days. Top applications sending traffic to the model include SillyTavern, Janitor AI, and Open WebUI. Integration is possible through OpenRouter's OpenAI-compatible API, Anthropic Messages format, or the Responses API, with support for reasoning tokens and streaming.

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://openrouter.ai/deepseek/deepseek-v4-pro-0813>

## Questions this post answers

### What is the API pricing for DeepSeek V4 Pro 0813 on OpenRouter?

DeepSeek V4 Pro 0813 costs $0.435 per million input tokens and $0.87 per million output tokens at the listed provider rate. The weighted average price customers actually pay is lower, around $0.04999 per million input tokens and $0.8698 per million output tokens, due to caching and discounts. It supports a 1,048,576 token context window and up to 384,000 output tokens.

_Developers comparing LLM API costs can track pricing shifts like this via daily.dev._

### What is the context window and max output for DeepSeek V4 Pro 0813?

DeepSeek V4 Pro 0813 has a context window of 1,048,576 tokens (1M) and a maximum output of 384,000 tokens. It is a large-scale mixture-of-experts model released as the general availability version of DeepSeek V4 Pro on August 12, 2026, and is accessible through OpenRouter's API.

_Keep up with new model context and output limits relevant to your builds on daily.dev._

## Community take

How the wider developer community reacted, aggregated from 1 discussion and 91 comments across hackernews (as of 2026-08-13).

**TL;DR:** Reception is largely a value-for-money debate: many find V4 Pro 0813 impressively cheap and competitive with pricier frontier models, but plenty argue the cheaper Flash variant already delivers similar quality, making Pro's premium hard to justify.

**Sentiment:** 40% positive · 45% mixed · 15% skeptical

**The case for**

- Extremely cheap per-task cost compared to Opus-class models, especially factoring in cache discounts.
- Benchmarks show clear gains over the previous Flash checkpoint on coding/agentic tasks.
- Some report it plans tasks well when paired with Flash for implementation.
- Its excellent caching and low pricing make it feel 'too cheap to meter' for heavy usage.

**The pushback**

- Many argue the cheaper Flash model gets 80-90% of the way there, making Pro's price premium not worth it.
- Benchmarks put it behind Kimi-K3 and other frontier models on several tasks.
- Burns a very high number of tokens to complete tasks despite low per-token cost.
- No vision support, unlike some competitors.
- Concerns about data privacy/training policy and geopolitical risk of using a Chinese-hosted model.

**By community**

- hackernews (mixed): Extensive discussion weighing V4 Pro's benchmark gains and cost efficiency against whether it's actually worth the price jump over Flash, alongside side debates on China-hosting risk, pricing transparency, and comparisons to Opus, Kimi, and other competitors.

**Hottest debate:** Whether the more expensive Pro model is actually worth it over the much cheaper Flash model, which many feel already delivers comparable real-world results.

**Open questions**

- Will DeepSeek actually raise prices soon, and by how much?
- How does V4 Pro really compare in practice (not just benchmarks) to Kimi-K3, GLM-5.2, and Opus 5?
- Will additional third-party hosting providers reduce reliance on DeepSeek's own data-training policy?

**Highlights**

> So not worth it over flash? Even at ~7x the size it isn't worth the price hike. Flash may be a monster of a model due to all the RL it received from free usage everywhere.
> — [Gecko4072 on hackernews · 5 comments](https://news.ycombinator.com/item?id=49275256)

> If that wasn't impressive enough, it's actually ~60x cheaper if you take into account the typical cache-read/input/output split in agentic coding, and the deep discount for cache reads offered by DeepSeek. Opencode has some public data on the typical split [1]: For DeepSeek V4 Pro the typical split is 750 in, 290 out, 82k cached. Cost per request for V4 Pro: $0.000875 per request. Equivalent Opus cost (w/o taking into account cache write costs): $0.052 per request. [1] https://opencode.ai/docs/go/#usage-limits
> — [xynelius on hackernews](https://news.ycombinator.com/item?id=49275775)

> Benchmarks:     | Benchmark                | DS-V4-Pro | DS-V4-Flash | DS-V4-Pro | DS-V4-Flash | GLM-5.2   | Kimi-K3   | Opus-4.8  | Fable 5       |     |                          | 0813      | 0731        | Preview   | Preview     |           |           |           | (w/ fallback) |     |--------------------------|-----------|-------------|-----------|-------------|-----------|-----------|-----------|---------------|     | HLE (wo/w tools)         | 42.7/60.0 | 37.8/51.5   | 37.7/48.2 | 34.8/45.1   | 40.5/54.7 | 43.5/56.0 | 49.8/57.9 | 53.3/63.0     |     | Terminal Bench 2.1       | 87.9      | 82.7        | 72.1      | 61.8        | 81.0      | 88.3      | 85.0      | 88.0          |     | NL2Repo                  | 61.5      | 54.2        | 38.5      | 39.4        | 48.9      | -         | 69.7      | -             |     | Cybergym                 | 83.3      | 76.7        | 52.7      | 38.7        | -         | 80.0      | 78.3      | 83.1          |     | DeepSWE                  | 62.7      | 54.4        | 12.8      | 7.3         | 46.2      | 67.5      | 58.0      | 70.0          |     | Toolathlon-Verified      | 74.1      | 70.3        | 55.9      | 49.7        | 59.9      | 76.5      | 76.2      | 77.9          |     | Agents' Last Exam        | 25.7      | 25.2        | 16.5      | 15.8        | 23.8      | 27.6      | 25.7      | -             |     | AutomationBench (Public) | 31.8      | 25.1        | 12.8      | 10.8        | 12.9      | 30.8      | 27.2      | 29.1          |     | DSBench-FullStack        | 71.1      | 68.7        | 41.8      | 37.0        | 61.8      | 73.7      | 71.6      | 77.2          |     | DSBench-Hard             | 67.2      | 59.6        | 31.1      | 25.8        | 54.5      | 63.0      | 71.7      | 68.3          | Source: https://reddit.com/r/LocalLLaMA/comments/1vmi0fg/deepseek_v4...
> — [scrlk on hackernews · 4 comments](https://news.ycombinator.com/item?id=49275180)

> Even though cost-per-token is low, Deepseek v4 tends to burn an immense number of tokens to accomplish tasks.
> — [nullbyte on hackernews](https://news.ycombinator.com/item?id=49276127)

> Geometric mean of all these benchmarks : * GPT-5.6 Sol: 65.5 * Fable 5 (w/ fallback): 64.5 * Opus 5: 64.0 * DS-V4-Pro 0813: 62.5 * Kimi-K3: 62.3 * DS-V4-Flash 0731: 55.8 * GLM-5.2: 47.3
> — [goldenarm on hackernews · 1 comments](https://news.ycombinator.com/item?id=49275754)

**Source threads**

- [hackernews](https://news.ycombinator.com/item?id=49274600) · 249 points · 91 comments

## Similar posts on daily.dev

- [Models & Pricing](https://daily.dev/posts/models-pricing-6vw4gyejf) · Hacker News · 2 upvotes · 0 comments
- [DeepSeek’s steep V4-Pro price cut escalates AI pricing war](https://daily.dev/posts/deepseek-s-steep-v4-pro-price-cut-escalates-ai-pricing-war-ssyvrmhnt) · InfoWorld · 0 upvotes · 0 comments

---

Tags: [#llm](https://daily.dev/tags/llm), [#deepseek](https://daily.dev/tags/deepseek), [#mixture-of-experts](https://daily.dev/tags/mixture-of-experts), [#openrouter](https://daily.dev/tags/openrouter)

[View this post on daily.dev](https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb)

```json
{"@context":"https://schema.org","@graph":[{"@type":"Organization","@id":"https://daily.dev/#organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180},"sameAs":["https://twitter.com/dailydotdev","https://github.com/dailydotdev","https://www.linkedin.com/company/daily-dev-ltd"]},{"@type":"WebSite","@id":"https://daily.dev/#website","url":"https://daily.dev","name":"daily.dev","publisher":{"@id":"https://daily.dev/#organization"},"potentialAction":{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https://daily.dev/search?q={search_term_string}"},"query-input":"required name=search_term_string"}}]}
{"@context":"https://schema.org","@type":"TechArticle","headline":"DeepSeek V4 Pro 0813","url":"https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb","mainEntityOfPage":{"@type":"WebPage","@id":"https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb"},"datePublished":"2026-08-12T16:52:41.749Z","dateModified":"2026-08-13T15:45:27.773Z","description":"DeepSeek V4 Pro 0813 is the general availability release of a large-scale mixture-of-experts model, priced at $0.435 per million input tokens and $0.87 per...","image":"https://media.daily.dev/image/upload/s--ZrL_HSsR--/f_auto/v1722860399/public/Placeholder%2006","thumbnailUrl":"https://media.daily.dev/image/upload/s--ZrL_HSsR--/f_auto/v1722860399/public/Placeholder%2006","isAccessibleForFree":true,"articleSection":"Hacker News","inLanguage":"en","publisher":{"@type":"Organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180}},"author":{"@type":"Organization","name":"Hacker News","logo":"https://media.daily.dev/image/upload/t_logo,f_auto/v1/logos/hn","url":"https://daily.dev/sources/hn"},"commentCount":0,"discussionUrl":"https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb","interactionStatistic":[{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":1},{"@type":"InteractionCounter","interactionType":{"@type":"CommentAction"},"userInteractionCount":0}],"keywords":"llm,deepseek,mixture-of-experts,openrouter","timeRequired":"PT5M"}
{"@context":"https://schema.org","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://daily.dev"},{"@type":"ListItem","position":2,"name":"Hacker News","item":"https://daily.dev/sources/hn"},{"@type":"ListItem","position":3,"name":"DeepSeek V4 Pro 0813"}]}
{"@context":"https://schema.org","@type":"FAQPage","@id":"https://daily.dev/posts/deepseek-v4-pro-0813-b1mmdmajb#faq","mainEntity":[{"@type":"Question","name":"What is the API pricing for DeepSeek V4 Pro 0813 on OpenRouter?","acceptedAnswer":{"@type":"Answer","text":"DeepSeek V4 Pro 0813 costs $0.435 per million input tokens and $0.87 per million output tokens at the listed provider rate. The weighted average price customers actually pay is lower, around $0.04999 per million input tokens and $0.8698 per million output tokens, due to caching and discounts. It supports a 1,048,576 token context window and up to 384,000 output tokens. Developers comparing LLM API costs can track pricing shifts like this via daily.dev."}},{"@type":"Question","name":"What is the context window and max output for DeepSeek V4 Pro 0813?","acceptedAnswer":{"@type":"Answer","text":"DeepSeek V4 Pro 0813 has a context window of 1,048,576 tokens (1M) and a maximum output of 384,000 tokens. It is a large-scale mixture-of-experts model released as the general availability version of DeepSeek V4 Pro on August 12, 2026, and is accessible through OpenRouter's API. Keep up with new model context and output limits relevant to your builds on daily.dev."}}]}
```

