Dev seeks blueprint for agent that triages incidents across PagerDuty, Datadog, GitLab and Slack
cruelcaricature · reddit · 2026-09-14
A Reddit developer is asking for architecture advice on building an agentic system for production incident triage: when a Jenkins failure or PagerDuty alert surfaces in Slack, the agent should inspect the relevant Jenkins job and Datadog monitors, infer which repo caused the issue, fix it, and open an MR.
He already uses RAG over Confluence docs and Slack conversations, but is stuck on how to structure multi-repo code and dependency information for agent querying, and asks for open-source tools that map interrelated repo dependencies. Useful thread to watch for tooling recommendations.
More from coding & agent
- Running a 100% hands-off B2B business on Muse: 3 ideas daily, $10k MRR goal in 90 days — armand_ruiz · 2026-09-14
- Scry AI search gets fast and ergonomic, making agent-wide web analysis trivial — AaronBergman18 · 2026-09-14
- CompBio Researcher Says a Qwen3.8-27B Fine-tune Beats Other Builds on Tool Calling, Cancels Claude — Usual-Carrot6352 · 2026-09-14
- Survey of 60+ image generators: retrieval, not capture, is the real pain point — shivam_dewan · 2026-09-14
- LeanDB: Theoric Labs builds a strongly typed Lean frontend for SQL databases — hargup13 · 2026-09-14
- iOS AI dev workflow: AppKit + Figma import + Claude Opus 5 gives best UI fidelity — dotey · 2026-09-14