GPT-5.6 Sol nails a one-shot answer in a new FDR-BH one-sided test result
lihua_lei_stat · x · 2026-07-22
- A thread on progress around controlling the false discovery rate with the Benjamini–Hochberg procedure for one-sided Gaussian tests.
- The author says GPT-5.6 Sol got the correct answer in one shot, and attached the exact prompt.
- The post is framed as Episode 3 of the FDR(BHq) series, with the author noting the verification and editing took several hours.
More from Research
- Itshi shows a full embodied AI model driving an automotive wiring line at WAIC 2026 — 量子位 · 2026-07-22
- A new RL finding says the training harness may induce generalization — inductionheads · 2026-07-22
- XBOW paper breaks insecure LLM code generation into fixable failure categories — moyix · 2026-07-22
- Gemini 3.6 Flash “disproves” a 43-year-old conjecture in a joke post — georgemillo · 2026-07-22
- AI Security Institute tests lie detectors across 31 open-weight models — geoffreyirving · 2026-07-22
- DeltaNet notes unpack Kimi Delta Attention with equations and a state-update diagram — nrehiew_ · 2026-07-22