MODEL PROFILE / EDITION 2026.09
DeepSeek V4 Pro 0813
Cost-controlled agentic backend and full-stack work, and teams that need MIT-licensed weights for compliance
Balanced model evidence score; rank 8. No rank delta is shown because there is no prior numeric snapshot.
Best for
Cost-controlled agentic backend and full-stack work, and teams that need MIT-licensed weights for compliance
Price & access
DeepSeek direct peak: $1.32 in / $3.96 out per 1M; off-peak exactly half at $0.66/$1.98; cache-hit input $0.044 peak / $0.022 off-peak. Peak hours 01:00-04:00 and 06:00-10:00 UTC Mon-Fri. OpenRouter blended: $0.5795 in / $1.738 out, cache read $0.01932.
Evidence summary
First-party pricing and version table, first-party Hugging Face card with a full benchmark table and stated harness, two independent preference boards, two third-party benchmark leaderboards and two independent speed measurements.
Edition status
No prior snapshot. Tracked in the September 2026 baseline.
Strengths
- Terminal-Bench 2.1 87.9% is within 0.9 pts of the best model in the set while costing a fraction of it
- MIT-licensed 1.6-1.7T MoE weights, so the exact model can be self-hosted
- Sub-2-second TTFT and 1M context on the hosted API
- Off-peak pricing halves an already low rate ($0.66/$1.98)
Tradeoffs
- Design Arena Website 1258 (rank 42) shows visual taste well behind frontier models
- Text-only: screenshot-to-code needs the separate vision-exp model
- SWE-bench Pro 55.4% is far below the Anthropic tier
- Serving 1.6T+ weights locally is out of reach for most teams
Score profile
visual 71 · code 84 · agentic 82 · debug 80 · context 82 · speed 74 · value 92
Linked evidence
Evidence source 1 ↗Evidence source 2 ↗Evidence source 3 ↗Evidence source 4 ↗Evidence source 5 ↗Evidence source 6 ↗Evidence source 7 ↗Evidence source 8 ↗Evidence source 9 ↗Use the same published data, filters, and side-by-side metrics.