Open-source InfernoSIM v4.0: failure simulator for AI agents, inject faults and verify behavior

pranaysparihar · reddit · 2026-08-24

The author released InfernoSIM v4.0, a local reliability-testing tool for tool-using AI agents. It records sanitized model/tool traffic, replays deterministically, injects failures, and verifies actual agent behavior. Supports OpenAI, Anthropic, Ollama, MCP HTTP/stdio, etc. Simulates lost side effects, malformed args, schema drift, delayed responses. Outputs JSON, JUnit, SARIF, HTML for CI. Seeking engineers to test on real stacks, especially parallel calls, custom MCP, streaming. MIT licensed, runs locally.

Related event: InfernoSIM: Open Source AI Agent Failure Simulator(2 posts)→

Original post →

More from coding & agent

coding & agent channel →