AI ScoreboardFrontend & Coding Index← Back to rankings

MODEL PROFILE / EDITION 2026.09

DeepSeek V4 Pro 0813

Cost-controlled agentic backend and full-stack work, and teams that need MIT-licensed weights for compliance

DeepSeekRank 8Confidence ANo prior snapshot
79.7

Balanced model evidence score; rank 8. No rank delta is shown because there is no prior numeric snapshot.

Best for

Cost-controlled agentic backend and full-stack work, and teams that need MIT-licensed weights for compliance

Price & access

DeepSeek direct peak: $1.32 in / $3.96 out per 1M; off-peak exactly half at $0.66/$1.98; cache-hit input $0.044 peak / $0.022 off-peak. Peak hours 01:00-04:00 and 06:00-10:00 UTC Mon-Fri. OpenRouter blended: $0.5795 in / $1.738 out, cache read $0.01932.

Evidence summary

First-party pricing and version table, first-party Hugging Face card with a full benchmark table and stated harness, two independent preference boards, two third-party benchmark leaderboards and two independent speed measurements.

Edition status

No prior snapshot. Tracked in the September 2026 baseline.

Strengths

  • Terminal-Bench 2.1 87.9% is within 0.9 pts of the best model in the set while costing a fraction of it
  • MIT-licensed 1.6-1.7T MoE weights, so the exact model can be self-hosted
  • Sub-2-second TTFT and 1M context on the hosted API
  • Off-peak pricing halves an already low rate ($0.66/$1.98)

Tradeoffs

  • Design Arena Website 1258 (rank 42) shows visual taste well behind frontier models
  • Text-only: screenshot-to-code needs the separate vision-exp model
  • SWE-bench Pro 55.4% is far below the Anthropic tier
  • Serving 1.6T+ weights locally is out of reach for most teams

Score profile

visual 71 · code 84 · agentic 82 · debug 80 · context 82 · speed 74 · value 92

Linked evidence

Evidence source 1 ↗Evidence source 2 ↗Evidence source 3 ↗Evidence source 4 ↗Evidence source 5 ↗Evidence source 6 ↗Evidence source 7 ↗Evidence source 8 ↗Evidence source 9 ↗
Compare this model on the full scoreboard

Use the same published data, filters, and side-by-side metrics.

Open comparison view →