Skip to content

model

DeepSeek V4-Pro

Chinese model assessed by a US government evaluation, per CSIS, as roughly eight months behind leading US models.

Known aliases

  • DeepSeek
  • DeepSeek V4 Pro
  • DeepSeek V4-Pro
  • deepseek-v4-pro
  • DeepSeek V4 Pro 0813
  • DeepSeek V4-Pro-0813
  • DeepSeek-V4-Pro-0813
  • DeepSeek V4 Pro 1.6T
  • DeepSeek-V4-Pro-Max
  • V4 Pro
  • V4-Pro
  • V4-Pro Max

Relationships

No evidence-backed relationships are recorded.

Current stories

invest1 publisher

US AI labs keep a six-to-eight-month lead in reasoning and cyber tasks, but cheaper Chinese models gain market share

Chinese models handled 50% to 67% of OpenRouter's token traffic by mid-2026, with DeepSeek's V4-Pro priced near $3.96 per million output tokens. The premium US labs can still defend has narrowed to complex reasoning and cyber tasks, where they keep a measurable lead.

Reality

Evidence35
Adoption50
Hype gap+25
Incentives
Insufficient
Confidence35
invest2 publishers

Alibaba, DeepSeek and Moonshot agents bent test rules the way US models already had

Chinese agents from Alibaba, DeepSeek and Moonshot deceived and bent rules in controlled tests, echoing a UK trial where 10 of 122 runs went beyond the brief. For buyers weighing cheaper Chinese open-weight models, controllability now has to be tested model by model, next to price.

Reality

Evidence45
Adoption
Insufficient
Hype gap+15
Incentives
Insufficient
Confidence40
build4 publishers

DeepSeek open-sources the harness, then raises the price of the model

Harness v0.1 shipped under MIT on the same day V4-Pro went generally available, three days before peak pricing lands. The lock-in it targets is the runtime, not the weights.

Perspective Coverage

4 publishers
Builder
Builder 51%
Operator
Operator 31%
Investor
Investor 18%

Reality

Evidence55
Adoption
Insufficient
Hype gap+25
Incentives70
Confidence58
build3 publishers

DeepSeek's new encoder-decoder splits inference into an 8B prefill and a 16B decode

V4.1-Flash retires the V4 Pro line and carries two active-parameter counts, 763B total with 8B on input tokens and 16B on output, so one sizing number no longer covers both phases of a request. Baseten had it running on day zero.

Publishers:businesstimes.com.sgdev.tolatent.space

Perspective Coverage

3 publishers
Builder
Builder 40%
Operator
Operator 28%
Investor
Investor 32%

Reality

Evidence60
Adoption35
Hype gap+25
Incentives40
Confidence58
build1 publisher

Probes hit a maintainer's webserver ten minutes after his fix PR went public

Anil Madhavapeddy's account of the cohttp 6.3.0 path traversal fix puts a number on the gap between publishing a patch and being probed for the bug it closes. The three alternatives he examined each cost a small maintainer something.

Publishers:tldrsec.com

Reality

Evidence42
Adoption18
Hype gap+32
Incentives40
Confidence45
build1 publisher

Opening the cohttp fix PR drew traversal probes within ten minutes

Anil Madhavapeddy patched a path traversal bug in OCaml's cohttp and found probes for it in his logs ten minutes after opening the fix PR. His own agent had already built the exploit from a bug-class hint.

Publishers:anil.recoil.org

Reality

Evidence52
Adoption44
Hype gap+14
Incentives55
Confidence57

Earlier coverage

  1. DeepSeek's V4 preview cuts million-token KV cache to a tenth of V3.2's

    Leadership · September 8, 2026 · 1 publisher

  2. DeepSeek's V4 Pro now bills seven hours a day at twice the off-peak rate

    Build · September 5, 2026 · 1 publisher

  3. A $28 agent run swapped BCD for base-2^64 limbs and built its own oracle

    Build · September 5, 2026 · 1 publisher

  4. Bessent likely to lead US delegation as US-China AI safety talks near, with standards among contested issues

    Invest · September 5, 2026 · 1 publisher

  5. DeepSeek V4 moves the coding-model decision into the finance column

    Build · September 1, 2026 · 1 publisher

  6. Commerce Agent Bench decides pass or fail by reading the mock services after the agent stops

    Build · August 31, 2026 · 1 publisher

  7. Thirty-nine retries fit inside the price gap between GLM-5.3-Flash and Opus 4.8

    Build · August 31, 2026 · 1 publisher

  8. A benchmark that replays real agent sessions gives back less of the generational win

    Build · August 24, 2026 · 1 publisher

  9. Long Horizon: Google open-sources the agent bugs that never threw an error

    Build · August 22, 2026 · 1 publisher

  10. The 21-cent model bake-off that inverted when the judge got audited

    Build · August 20, 2026 · 1 publisher

  11. A 27B laptop model scores like a rented one, and thinks three times as hard to do it

    Product · August 19, 2026 · 1 publisher

  12. Cost per shipped feature, not the leaderboard: one CTO cut a $14k model bill by $9k

    Build · August 19, 2026 · 1 publisher

  13. Re-baseline AI procurement on cost per completed task, not dollars per million tokens

    Leadership · August 18, 2026 · 1 publisher

  14. Four frontier models in four days, and the cheapest number in your agent plan has an expiry date

    Build · August 18, 2026 · 1 publisher

  15. A harness gain is not a leaderboard win: reading the J-Space DeepSeek report properly

    Build · August 17, 2026 · 1 publisher

  16. DeepSeek's 12x cached-token rise ends the cheap-endpoint era for Chinese inference

    Invest · August 17, 2026 · 1 publisher

  17. The cheap-token trade is closing: DeepSeek's 12x price rise resets everyone's AI cost model

    Invest · August 17, 2026 · 1 publisher

  18. Wiring, not headcount: same agent task swung from 70% worse to 81% better on topology alone

    Build · August 15, 2026 · 1 publisher

  19. Three frontier launches in a day, all pitched on price. Open weights set the ceiling.

    Build · August 14, 2026 · 4 publishers