Open-source InfernoSIM v4.0: failure simulator for AI agents, inject faults and verify behavior
pranaysparihar · reddit · 2026-08-24
The author released InfernoSIM v4.0, a local reliability-testing tool for tool-using AI agents. It records sanitized model/tool traffic, replays deterministically, injects failures, and verifies actual agent behavior. Supports OpenAI, Anthropic, Ollama, MCP HTTP/stdio, etc. Simulates lost side effects, malformed args, schema drift, delayed responses. Outputs JSON, JUnit, SARIF, HTML for CI. Seeking engineers to test on real stacks, especially parallel calls, custom MCP, streaming. MIT licensed, runs locally.
Related event: InfernoSIM: Open Source AI Agent Failure Simulator(2 posts)→
More from coding & agent
- Repo includes Claude and Codex implementations — tekbog · 2026-08-25
- Developer Argues Codex Remains the Best AI Coding Product — nickbaumann_ · 2026-08-25
- Apodex 1.1 mini: open 35B local model, harness swap adds up to 10 points — SimonShaoleiDu · 2026-08-25
- How AI Agents Understand Design System Languages — round · 2026-08-25
- From Scripts to Loops: Engineering Challenges and Solutions for Production Agentic Systems — Pavan_Belagatti · 2026-08-25
- AWS launches Agent Registry and backs ARD, an open spec for cross-environment agent discovery — AWS ML Blog · 2026-08-25