AI ScoreboardFrontend & Coding Index← Back to rankings

MODEL PROFILE / EDITION 2026.09

GPT-5.5

Teams already pinned to GPT-5.5 snapshots who need a stable, known quantity

OpenAIRank 28Confidence BNo prior snapshot
66.5

Balanced model evidence score; rank 28. No rank delta is shown because there is no prior numeric snapshot.

Best for

Teams already pinned to GPT-5.5 snapshots who need a stable, known quantity

Price & access

$5 in / $30 out per 1M (AA); GPT-5.6 Sol now undercuts it at $4/$20

Evidence summary

Deeply sampled Arena and Design Arena rows plus third-party SWE-bench Pro, but the docs vs AA context-window conflict (1.05M vs 922k) is unresolved.

Edition status

No prior snapshot. Tracked in the September 2026 baseline.

Strengths

  • Very well-sampled Arena row (13,258 votes, +/-6)
  • SWE-Atlas debugging 44.7% is the best number in ByteDance's cross-vendor table
  • 1M-class context with image input
  • Mature, widely integrated model

Tradeoffs

  • Costs more than GPT-5.6 Sol while scoring lower on SWE-bench Pro
  • Arena frontend Elo 1513 is bottom-quartile in this cohort
  • No `max` effort tier
  • Superseded by GPT-5.6, so investment is short-lived

Score profile

visual 63 · code 72 · agentic 68 · debug 68 · context 84 · speed 60 · value 50

Linked evidence

Evidence source 1 ↗Evidence source 2 ↗Evidence source 3 ↗Evidence source 4 ↗Evidence source 5 ↗Evidence source 6 ↗
Compare this model on the full scoreboard

Use the same published data, filters, and side-by-side metrics.

Open comparison view →