MODEL PROFILE / EDITION 2026.09
GPT-5.6 Terra
High-volume production coding where SWE-bench-style correctness matters more than visual flair
Balanced model evidence score; rank 26. No rank delta is shown because there is no prior numeric snapshot.
Best for
High-volume production coding where SWE-bench-style correctness matters more than visual flair
Price & access
$2 in / $12 out per 1M; cached input $0.20; rates quoted for context below 272K
Evidence summary
Price/context/speed are first-party or AA-measured, and SWE-bench Pro placement is third-party; agentic and debugging evidence is family-level rather than Terra-specific.
Edition status
No prior snapshot. Tracked in the September 2026 baseline.
Strengths
- SWE-bench Pro 63.4% is within ~1 pt of Sol at half the price
- 111.4 tok/s is the fastest OpenAI throughput in this set
- 1.05M context with image input
- Cheap cached input ($0.20)
Tradeoffs
- Arena frontend Elo 1520 is mid-pack
- TTFT 168.78 s at max effort is very high
- No Design Arena or visual-web benchmark evidence
- No first-party frontend benchmark published
Score profile
visual 63 · code 78 · agentic 74 · debug 72 · context 86 · speed 62 · value 76
Linked evidence
Evidence source 1 ↗Evidence source 2 ↗Evidence source 3 ↗Evidence source 4 ↗Evidence source 5 ↗Use the same published data, filters, and side-by-side metrics.