DeepTeam: Open-source framework for red teaming LLMs locally
tom_doerr · x · 2026-09-19
DeepTeam is an open-source red teaming framework for LLM systems (2.8k stars on GitHub) — think penetration testing, but for LLMs.
It runs locally and simulates attacks including jailbreaking, prompt injection, and multi-turn exploitation to uncover vulnerabilities like bias, PII leakage, and SQL injection in AI agents, RAG pipelines, and chatbots. It also offers guardrails to prevent these issues in production.
Built on DeepEval (the open-source LLM evaluation framework) and maintained by Confident AI, with an accompanying platform for managing risk assessments and sharing reports.
More from coding & agent
- Split coding and testing across two agents to save weekly quota — ___Patrice___ · 2026-09-19
- onPanda: open-source token-level LLM editor that can harness Claude Code and Codex — Fancy_Fanqi77 · 2026-09-19
- Muse use-case directory adds 90 new entries, now 678 total with ready-to-run prompts — armand_ruiz · 2026-09-19
- Skip the waitlist: running Jev's evaluate API via Vercel AI Gateway for $0.00074 — AlchainHust · 2026-09-19
- Simulated Worlds for Sonnet 3.6: Bureaucracy, Commutes and TONS OF KINDNESS — repligate · 2026-09-19
- Anthropic's Fable 5.1 lands in Kiro, built for long-running context-heavy agentic sessions — DigitalColmer · 2026-09-19