OpenAI & AWS Tests Cut GPT-5.6 Agent Task Costs by 82%

新智元 · wechat · 2026-08-28

OpenAI and AWS released joint test results showing that by optimizing the environment and model, the cost for GPT-5.6 Terra to complete a task on the Kiro platform (Terminal-Bench 2.1) dropped by approximately 82%.

Optimization Logic

The 62% cost difference did not come from official price cuts (Terra dropped only 20%), but from efficiency gains:

Framework & Model Division

Tests show that billing costs are heavily influenced by the combination of "Agent Framework + Model". Kiro uses a Spec-driven process, breaking down requirements before handing them to the model, significantly saving tokens previously wasted on trial and error. Additionally, OpenAI suggests using Sol for planning and Luna for implementation in the pipeline to optimize cost-performance.

Competitive Landscape

The GPT-5.6 series (Sol/Terra/Luna) is now available on AWS Kiro, listed alongside Claude Fable. Kiro employs a hidden Chain of Thought strategy. This move allows OpenAI to compete with Anthropic on AWS's turf, shifting the evaluation standard from pure "accuracy" to a "balance of score and cost".

Original post →

More from coding & agent

coding & agent channel →