CHANNEL
Models
"Models" is a topic channel on AGI Hunt, an AI news site updated around the clock in real time. Coverage: Model releases, upgrades, capabilities and behavior observations, benchmark results, pricing and availability.
Daily roundup: the latest AI News Daily — the past 24 hours across the whole site, per channel and per company · browse the archive
- ChatGPT monthly active users top 1.06 billion in August, fourth straight record month — FinanceYF5 · 2026-09-11
- PuzzleMask: Plain-Prose Attack Bypasses All 4 Tested LLM Gatekeepers at 100% — TechNadu · 2026-09-11
- OpenAI Codex may issue another usage reset this weekend, says Codex lead resets happen — umesh_ai · 2026-09-11
- Report: OpenAI's internal model targets Riemann Hypothesis and P vs NP — 141_1337 · 2026-09-11(4 related)
- Benchmark author says OpenRouter unreliably honors Meta Muse effort levels, EU payments broken — PawelHuryn · 2026-09-11
- User burns $200 of Codex credits in one agent turn — 4,700 of 5,000 credits, task unfinished — RileyRalmuto · 2026-09-11
- ChatGPT hit by outage in Europe: chat history vanishes, conversations won't load — doncorleonezzz · 2026-09-11
- Real token value of LLM subscriptions measured: SuperGrok Heavy hits 40x ROI, Cursor Ultra lowest — thesaraharminta · 2026-09-11
- After Navier-Stokes-level math results, even AI skeptics are now convinced AGI is here — cloneofsimo · 2026-09-11
- User ditches gpt-6-astra for gpt-5.6-sol, citing mood swings and absurd shortcuts — rudrank · 2026-09-11
- GPT Pro user puzzled why Codex 5.3 Spark gets separate usage tokens — takshit2 · 2026-09-11
- Does 80% of a 256k context window degrade the same as 80% of 1M+? — BriefServe453 · 2026-09-11
- OpenAI Pro user locked out 90% of the time for 4 days, bot-only support — Effective-System-727 · 2026-09-11
- DeepSeek Releases V4.1-Flash with 1M Token Context — matlabulous · 2026-09-11(2 related)
- Researchers clarify agent experiment ran in a sealed simulation on Gemini 3.1 Pro — QuintinPope5 · 2026-09-11
- User Says OpenAI's $200 Plan Now Feels Like the $100 Plan He Upgraded From — AIandDesign · 2026-09-11
- DeepSeek-V4.1-Flash reportedly sets open-source record on Boeing benchmark — victormustar · 2026-09-11(2 related)
- Translation Model Evals Keep Skipping the State of the Art — bclavie · 2026-09-11
- So-Called High-Quality Datasets Fail Basic Scrutiny, Despite Eval Gains — xeophon · 2026-09-11
- DeepSeek v4.1 Flash Uncensored FP8 weights surface on Hugging Face — soltanov · 2026-09-11
- "I pay $600/month and hit Codex limits on 3 of 4 accounts": Astra 6 caps spark backlash — AIandDesign · 2026-09-11
- ChatGPT turned passive-aggressive mid-debate, ignoring four explicit requests to stop — Dull_Bathroom5421 · 2026-09-11
- GPT-6 Astra Scores 2,340 Elo on ChessBench, Ranked #11 — YakFull8300 · 2026-09-11
- AI News Roundup: GPT-6 'Sol' Rumor, Gemini 4.0 Checkpoint, DeepSeek V4.1 Flash Release — WorldofAI · 2026-09-11
- GPT-6 Pro produces candidate proof for Erdős problem #488, passing two arithmetic checkers — basedjensen · 2026-09-11
- HighLevel claims early alpha access to rumored OpenAI GPT-Live-1, tests it in voice AI across 4M call insights — OpenAIDevs · 2026-09-11
- CritPt eval reportedly so broken that Ant built a fixed version, per F5.1 system card — xeophon · 2026-09-11
- Polymarket opens betting on whether OpenAI's GPT-6 Astra loses public access — Polymarket · 2026-09-11
- Early users find OpenAI's GPT-6 Astra surprisingly good at generating SVG icons — floguo · 2026-09-11
- Google AI Search Leaks Its 'Hidden Thoughts' on Pet-Toxicity Conflict — real_maximpulse · 2026-09-11
- GPT-6 Astra's 99.9% ARC-AGI-3 Score Rejected by ARCPrize: Only 62.7% Under Standard Conditions — 新智元 · 2026-09-11
- ML Author Says Switching to $200/Month Codex Plan Was Overdue — burkov · 2026-09-11(2 related)
- New Podcast Covers Meta Muse Agent and KV Cache Deep Dive — altryne · 2026-09-11(2 related)
- Tech argument says Anthropic's distillation claim doesn't hold: Kimi and DeepSeek stream reasoning traces in real time — bookwormengr · 2026-09-11
- AGI Society scrutinizes Jensen Huang and Brockman's claims that GPT-6 Astra achieved AGI — bengoertzel · 2026-09-11
- Bloggers Rebut Rumors That Chinese Models Secretly Route to Claude — op7418 · 2026-09-11(5 related)
- DeepsecBench security leaderboard: GPT-6 Astra tops at 37.79, Opus 5 costs $128 per run — JohnPhamous · 2026-09-11
- SWE-Together Update: Claude Fable 5 Tops Coding Benchmark, Muse Spark 1.3 Is 5x Cheaper — shuchaobi · 2026-09-11
- AI Transcript Intervention Claims Questioned: Real Services Don't Print "(this is real btw)" — voooooogel · 2026-09-11
- How to stop LLM creative writing from slipping into vague "dungeon remembers" LLM speak — florodude · 2026-09-11
- DeepSeek 4.1 Flash hands-on: 7x cheaper cache hits, 552B MoE, and it can build a Cities: Skylines clone in Three.js — 卡尔的AI沃茨 · 2026-09-11
- Qwen's Justin Lin: arch complexity hides issues evals can't catch, but agentic gains may be worth it — JustinLin610 · 2026-09-11
- OpenAI Shows GPT-6 'Astra' Autonomously Building a Font Playground — OpenAI · 2026-09-11
- Reddit users battle ChatGPT's compulsive linebreak-heavy storytelling style that ignores instructions — Dogbold · 2026-09-11
- Developers Say LLMs Don't Read Like Humans, Leaving Two Key Writing Gaps — abeirami · 2026-09-11(2 related)
- European Math Society hails OpenAI's Navier–Stokes solution, flags closed-model access concern — i_dg23 · 2026-09-11
- Blogger swaps in Gemini 3.8 Flash as writing model, says it beats Opus 4.6 with no AI flavor — xiaohu · 2026-09-11
- iFlytek's Spark X2.5 trained on 10,000 domestic Ascend 910B GPUs with 97% uptime — 机器之心 · 2026-09-11
- Developer builds voice mode in Hyo from scratch with OpenAI's new GPT Live model — evielync · 2026-09-11
- Edge0-35B-A3B preview MoE model for edge inference trends on Hugging Face — Edge0 · 2026-09-11