Sean Taylor claims fast progress eradicating hallucinations; Andrew Ng: capabilities and safety can align
irinarish · x · 2026-09-04
Sean Taylor weighs in on the Astra discussion: the team is making fast progress on eradicating hallucinations, using evals that realistically capture user experience rather than academic benchmarks. Andrew Ng reshares to argue the perceived capabilities-vs-safety tradeoff is not inevitable — reducing hallucinations is a case where the two align.
More from Models
- Sam Altman hints full Astra access may arrive this weekend, plus three usage resets — iruletheworldmo · 2026-09-04
- Analyst: no model trained without NVIDIA has ever beaten one trained on NVIDIA — BenBajarin · 2026-09-04
- Mostik's 12-PhD team tops ARC-AGI leaderboard by passing hidden states from frontier to small models — VoidAsuka · 2026-09-04
- OpenAI Astra team member hypes launch after months of work — oyhsu · 2026-09-04
- Fable 5.1 Names Agents in Same Thematic Style as Fable 5, Noting Personality Consistency — repligate · 2026-09-04
- OpenAI Commits $1B to Subsidize Cybersecurity Defenders Alongside GPT-6 Astra Launch — EricBuess · 2026-09-04