OpenAI reports 82% lower successful coding-agent task cost for GPT-5.6 Terra in Kiro
OpenAI and AWS testing found GPT-5.6 Terra completed successful Terminal-Bench 2.1 coding-agent tasks at roughly 82% lower cost in Kiro. The deployment also exposes Sol, Terra, and Luna for specification-driven implementation, multi-step coding, review checkpoints, and property-based testing.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
The primary report adds practical evidence that a frontier coding agent can complete benchmark tasks at sharply lower cost while operating across long-running, tool-using software workflows. Lower cost and deeper integration increase the feasible scale of consequential autonomous coding, although the evidence is provider-run and limited to one benchmark and product environment.
Assessment history
-
R1
Toward 48 · confidence 93
New primary deployment evidence reports a measured 82% reduction in successful Terminal-Bench task cost and names the exact GPT-5.6 models integrated into Kiro.
25 Aug 2026
Share this page
-
DoomBench assesses “OpenAI reports 82% lower successful coding-agent task cost for GPT-5.6 Terra in Kiro” as evidence moving toward doom, with magnitude 48 and confidence 93 out of 100 in the deployment reach category.
-
The DoomBench assessment of “OpenAI reports 82% lower successful coding-agent task cost for GPT-5.6 Terra in Kiro” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “OpenAI reports 82% lower successful coding-agent task cost for GPT-5.6 Terra in Kiro” as follows: OpenAI and AWS testing found GPT-5.6 Terra completed successful Terminal-Bench 2.1 coding-agent tasks at roughly 82%...
https://www.doombench.com/news/openai-reports-82-lower-successful-coding-agent-task-cost-for-gpt-5-6-terra-in-kiro-2026-08-24