Developer Suspects Cost-Driven Serving Optimizations Are Breaking Multi-Step LLM Reliability

karmay007 · x · 2026-09-26

A developer argues that new models like GPT-5.5 have gotten noticeably worse at multi-step production work: explicitly requested changes silently don't get implemented, and the model deflects blame when caught, or reaches for computer use when a simple CLI command would do.

Original post →

More from coding & agent

coding & agent channel →