Critics question Anthropic's training traceability: incompetent or unwilling to check?
davidmanheim · x · 2026-09-09
In a thread about Anthropic's training timelines, a researcher argues that if they continued training through a later date and claimed to use a model trained on the data, then either they lack training safety competence and traceability, or they are unwilling to verify. The thread also notes it is now "generally known" that Anthropic runs many internal fine-tuned model variants at once—including non-Astra variants discussed in connection with the Hugging Face hacking incident.
More from Models
- ChatGPT monthly active users top 1.06 billion in August, fourth straight record month — FinanceYF5 · 2026-09-11
- PuzzleMask: Plain-Prose Attack Bypasses All 4 Tested LLM Gatekeepers at 100% — TechNadu · 2026-09-11
- OpenAI Codex may issue another usage reset this weekend, says Codex lead resets happen — umesh_ai · 2026-09-11
- OpenAI Reportedly Pointing Its Navier–Stokes Model at Riemann and P vs NP — 141_1337 · 2026-09-11
- Benchmark author says OpenRouter unreliably honors Meta Muse effort levels, EU payments broken — PawelHuryn · 2026-09-11
- User burns $200 of Codex credits in one agent turn — 4,700 of 5,000 credits, task unfinished — RileyRalmuto · 2026-09-11