astra's 'Wild' Off-Beaten-Path Solutions Impress in Auto-Research Benchmark Re-Runs
generativist · x · 2026-09-13
AI researcher generativist adds detail on his astra auto-research benchmark re-runs: his harness specifically selects for off-beaten-path solutions, and astra's outputs were "wild," reinforcing his earlier claim that frontier models may have already done deep RL on such tasks and self-improvement is underway.
Related event: Researcher Reruns Auto-Research Benchmark, Finds Stunning Results(2 posts)→
More from AGI Musings
- Beff Jezos vows to 'overthrow the token harvester cartel' in anti-big-lab manifesto — beffjezos · 2026-09-13
- Comment: If OpenAI and Anthropic can't control the risks, they should stop releasing models — AlexTensor · 2026-09-13
- AI czar David Sacks backs frontier labs slowing down — but slams cartel and METR independence claims — kevinnbass · 2026-09-13
- Compute to shift from RL maxxing to interpretability until reward hacking is solved — zephyr_z9 · 2026-09-13
- Reddit users speculate Musk, Amodei and Altman know of an undisclosed AI incident behind slowdown calls — Traditional-Chip8339 · 2026-09-13
- tszzl predicts open-source AI will be banned after a major disaster, wants monitored APIs — mimi10v3 · 2026-09-13