Comparing Math Proofs: GPT-4o vs. Astra
danshipper · x · 2026-08-03
A developer shared the analysis generated by GPT-4o attempting a mathematical conjecture and compared it directly against the results from OpenAI's frontier model, Astra. This provides further insight into the reasoning capabilities of weaker models versus frontier models.
Related event: Prompt Distance: Can Weaker Models Reproduce Frontier Proofs?(6 posts)→
More from Models
- ChatGPT co-inventor launches Jev: claims 20-200x faster, 40-400x cheaper than LLMs — GabGarrett · 2026-09-18
- Dev: Jev could run cheaply in browsers — classification tasks were shoehorned into LLMs — GabGarrett · 2026-09-18
- Ran Jev across 10 services for hours, still couldn't spend $1 — multiply_matrix · 2026-09-18
- ChatGPT co-inventor launches Jev, claiming 200x faster, 400x cheaper frontier model — multiply_matrix · 2026-09-18
- Tencent's Hy4 Preview ranks #4 among open-weight models, cheapest in top ten — mariofilhoml · 2026-09-18
- Simple letter-counting test exposes huge gap: GPT-6-Astra hits 93%, Fable 5.1 flounders — scaling01 · 2026-09-18