Security Expert Mocks AI Eval Labs Over Basic Cybersecurity Flaws
nptacek · x · 2026-08-05
A cybersecurity expert sarcastically calls out AI evaluation labs, noting the irony that organizations failing at basic cybersecurity are positioning themselves as authorities on frontier model capabilities.
More from Models
- GPT 5.6 Sol Successfully Formalizes Complex Math Proof for Nonsofic Existence — Sauers_ · 2026-08-05
- GPT 5.6 Successfully Formalizes Complex Math Proofs Without Internal Lean Code — Sauers_ · 2026-08-05
- Chinese Models Dominate Forecasting Leaderboard with Advanced AI Agents — teortaxesTex · 2026-08-05
- Developer Take: Free Gemini 2.0 Flash Offers Better Value Than Kimi — mertdumenci · 2026-08-05
- Quantized DeepSeek V3 hits 247 tokens/s decode in just 162GB — teortaxesTex · 2026-08-05
- Nous Research Co-founder: Model Sycophancy is a 'Reward Hack', Not Loyalty — petergyang · 2026-08-05