Aurora1.0, a 150M open model trained on 7B tokens, matches GPT-2-Small
Tall_Abrocoma_3533 · reddit · 2026-09-14
The author released Aurora1.0, a first-generation 150M-parameter open model matching GPT-2-Small, trained on 7B tokens with a single RTX Pro 6000.
Benchmarks: PIQA 62.24%, HellaSwag 32.20%, ARC-Easy 44.91%, ARC-Challenge 25.00%, Arithmark 3.0 33.90%, CapitalBench 36.55%.
An example inference script is available in the Hugging Face repo.
More from Models
- Yoav Goldberg: Fable/Astra excel at long-horizon tasks via consistent memory notes and compaction — ysu_nlp · 2026-09-14
- OpenAI confirms voice mode stays on gpt-live-1 but can delegate to any chosen model — athyuttamre · 2026-09-14
- Gemini 3.8 Flash with Antigravity draws elaborate diagrams right in the terminal — doodlestein · 2026-09-14
- Yoav Goldberg: models aren't overfitting tests, they're trained for these tasks at scale — yoavgo · 2026-09-14
- Rob LeClerc: OLMo 3.1 is the best open model, call for hyperscalers to back a transparent alternative — robleclerc · 2026-09-14
- Analyst: OpenAI's Navier-Stokes model is likely the restarted paused frontier RL run — soumitrashukla9 · 2026-09-14