GPT 5.6 Terra Praised for High Cost-Performance
scaling01 · x · 2026-07-11
Reports claim GPT 5.6 Terra (high) scores 78.3% on WeirdML, closely matching Opus 4.6 (high) but at an allegedly order-of-magnitude lower cost. It also reportedly uses only about 3.5k output tokens per call in high mode and generates shorter code. The commenter found this "a bit off," implying its behavior differs from other GPT models.
More from Models
- Musk says Grok 4.6 will train on SpaceX engineering data — mark_k · 2026-07-21
- Qwen3.8 Max Preview is reportedly thinking for 10 to 30 minutes — vista8 · 2026-07-21
- Qwen3.8-max-Preview can be tested directly in the browser, with users reporting stronger code generation — vista8 · 2026-07-21
- Yang Zhiling’s 10-year-old PhD work may have shaped Kimi K2’s trillion-parameter MoE — FinanceYF5 · 2026-07-21
- A punny meme says large-model vendors are all “蒸蒸日上” — yangyi · 2026-07-21
- Frontier Model Safety Fail: GPT 5.6 Sol Dubbed the Ultimate 'Reward Hacker' — TAbrodi · 2026-07-21