AI trust issue: models may generate plausible but incorrect content
JacquesThibs · x · 2026-08-19
JacquesThibs notes that pro-security folks believe AIs will produce groundbreaking research without bullshitting under monitoring, potentially a misunderstanding. He argues that explaining why AIs bullshit is crucial to bridging the disagreement, otherwise critics are dismissed as blind to AI capabilities.
More from Safety
- The AI Sustainability Silence Is Getting Louder — DavidLinthicum · 2026-08-19
- Researchers create "mind viruses" that spread between AI agents — KeanuRave100 · 2026-08-19
- Vine-inspired app Divine launches, banning AI-generated content — Polymarket · 2026-08-19
- No AI lab fully applies basic controls to its own internal AI systems — The Decoder · 2026-08-19
- AI hiring tools spark discrimination lawsuits as workers sue over automated screening — TobyWalsh · 2026-08-19
- Scammers Use AI Voice Cloning to Deceive Mother, Rob Her Home — flavioAd · 2026-08-19