Community tests new model: impressive but brand bias persists
Early user tests of a new model show strong Fermi estimation and strict-format poetry rewriting, while other users caution that subjective model judgments are heavily influenced by brand names, with one test revealing a plausible-sounding but flawed argument.
2026-08-22 ~ 2026-08-22 · 2 related posts
- Opinion: People Rely on Brand Names for Subjective Model Evaluations — jd_pressman · 2026-08-22
- User Tests New Model: Strong at Fermi Estimates and Strict Poetry Rewriting — jd_pressman · 2026-08-22