Codex Hits 5M Weekly Users; AI Model Eval & Transparency Discussed
altryne · x · 2026-07-24
This episode of the ThursdAI podcast focuses on the latest in AI coding tools and model evaluation:
- Codex Growth & Features: OpenAI's Romain Huet reveals Codex has hit over 5 million weekly users, marking an inflection point. The show discusses new features like /goal, AppShots, autonomous thread management, and the future GPT-5.6.
- Daily AI Workflow: The host shares 15 minutes of his actual daily AI usage, including fixing papercuts with Computer Use and tools built for family vacations.
- Model Eval & Security: A guest appearance on Insecure Agents discusses the Wolfbench evaluation framework, emphasizing the critical need to get identity and authorization right as models grow more capable.
Related event: Codex tops 5M weekly active users, teases new features(2 posts)→
More from coding & agent
- Chaining dependent MCP tool calls: no rollback, duplicate risk — agentrsdg · 2026-09-11
- DeepMind-led paper makes design docs the source of truth, code disposable — SMART regenerates in 1.5-3h for ~$100 — Roger_M_Taylor · 2026-09-11
- Agent-built classifier labels 192k docs for $0.70 vs $13-26 with frontier LLMs — vanstriendaniel · 2026-09-11
- MathModelAgent gains traction: auto-solves math modeling and writes a submission-ready paper — jihe520 · 2026-09-11
- alphaXiv open-sources OpenResearch to run parallel research agents with any model — alphaXiv · 2026-09-11
- DeskcommCRM: open-source AI sales CRM with native agents and WhatsApp hits 1k stars — melgarafael · 2026-09-11