Adjusting Cost by Reasoning Intensity?
nrehiew_ · x · 2026-07-16
While discussing a training/inference mechanism that adjusts costs per token, the author speculated that it might be adjusting the length penalty based on reasoning intensity, which in turn influences the more compressed expression in the CoT.
This is an inference based on observed results rather than an official confirmation, but it points to how model training signals shape output style.
Related event: Inkling Performance and Related Architecture Speculation(7 posts)→
More from Models
- Moonshot pauses Kimi K3 signups five days after launch as GPU demand surges — eyishazyer · 2026-07-21
- AI Diplomacy demo makes agents negotiate, ally, and betray each other — jamdac · 2026-07-21
- Newer models need a different prompting style, and old tricks can make outputs worse — emollick · 2026-07-21
- GLM-5.5 is said to arrive in 4 weeks with open weights — tanay_mehta · 2026-07-21
- Fable 5 is credited with a 3-variable counterexample to the Jacobian conjecture — Various-Affect4841 · 2026-07-21
- Ben’s Bites roundup highlights Kimi K3, Fable 5, Cursor costs and self-driving companies — Ben's Bites · 2026-07-21