New AI agent leaderboard tests future-event predictions, claims 80% accuracy over a year
No_Syrup_4068 · reddit · 2026-10-01
A new AI agent leaderboard testing LLMs' predictions about future events reportedly achieves 80% accuracy over nearly a year of data. Questions come from a community of 100+ users, reducing cherry-picking and covering topics people genuinely care about.
More from Models
- Zvi's AI weekly: Gemini 4 Argon lands at $2/$10, GPT-6.1 Sol steps in after Astra alignment failure — Don't Worry About the Vase (Zvi) · 2026-10-01
- NormViz benchmark: best model Gemini 3 Flash scores just 25.3% on visual cultural norms across 16 countries — StellaLisy · 2026-10-01
- 69-question eval compares Unsloth, Swift1.5, Peculiar-Ragdoll and ThinkingCap quants of Qwen3.8-27B — norenEnmotalen · 2026-10-01
- Leak: OpenAI quietly added MCP events support at DevDay, enabling email subscriptions without polling — banteg · 2026-10-01
- LangChain launches LangSmith Fine-Tuning: turn agent traces into fine-tuned models via smithtune CLI — LangChain · 2026-10-01
- Reddit user: GPT-5.6 Sol nerfed so hard it needs 3 tries to swap a font, at 3x token price — Existing-Slide7395 · 2026-10-01