AI ScoreboardFrontend & Coding Index← Back to rankings

MODEL PROFILE / EDITION 2026.09

GPT-5.6 Terra

High-volume production coding where SWE-bench-style correctness matters more than visual flair

OpenAIRank 26Confidence BNo prior snapshot
72.0

Balanced model evidence score; rank 26. No rank delta is shown because there is no prior numeric snapshot.

Best for

High-volume production coding where SWE-bench-style correctness matters more than visual flair

Price & access

$2 in / $12 out per 1M; cached input $0.20; rates quoted for context below 272K

Evidence summary

Price/context/speed are first-party or AA-measured, and SWE-bench Pro placement is third-party; agentic and debugging evidence is family-level rather than Terra-specific.

Edition status

No prior snapshot. Tracked in the September 2026 baseline.

Strengths

  • SWE-bench Pro 63.4% is within ~1 pt of Sol at half the price
  • 111.4 tok/s is the fastest OpenAI throughput in this set
  • 1.05M context with image input
  • Cheap cached input ($0.20)

Tradeoffs

  • Arena frontend Elo 1520 is mid-pack
  • TTFT 168.78 s at max effort is very high
  • No Design Arena or visual-web benchmark evidence
  • No first-party frontend benchmark published

Score profile

visual 63 · code 78 · agentic 74 · debug 72 · context 86 · speed 62 · value 76

Linked evidence

Evidence source 1 ↗Evidence source 2 ↗Evidence source 3 ↗Evidence source 4 ↗Evidence source 5 ↗
Compare this model on the full scoreboard

Use the same published data, filters, and side-by-side metrics.

Open comparison view →