Models Execute Commands, But Lag in Decision-Making
fchollet · x · 2026-07-19
fchollet pointed out a notable divergence: models are improving incredibly fast at strictly executing clear instructions, but their ability to make reliable judgments when facing new situations not covered by instructions hasn't improved in tandem for a while.
This gap shows that while models are getting better at "doing as told," they are still far from achieving truly stable, autonomous decision-making.
More from Models
- Screenshot shows Gemini 3.5 Flash Lite and Gemini 3.6 Flash on a Google models page — koltregaskes · 2026-07-21
- Gemini 3.6 Flash lands in Google AI Studio with cheaper output pricing — gaganghotra_ · 2026-07-21
- Jack Clark says OpenAI’s internal-deployment safety notes help the whole frontier community — jackclarkSF · 2026-07-21
- Mindlab Research puts Macaron-V1-Venti on Hugging Face — External_Mood4719 · 2026-07-21
- Google is surfacing Gemini 3.5 Flash-Lite and 3.6 Flash in AI Studio — Expensive_Syrup_6529 · 2026-07-21
- ChatGPT often explains the wall before answering whether it is tilting — Aware-sky-3489 · 2026-07-21