Dev seeks blueprint for agent that triages incidents across PagerDuty, Datadog, GitLab and Slack

cruelcaricature · reddit · 2026-09-14

A Reddit developer is asking for architecture advice on building an agentic system for production incident triage: when a Jenkins failure or PagerDuty alert surfaces in Slack, the agent should inspect the relevant Jenkins job and Datadog monitors, infer which repo caused the issue, fix it, and open an MR.

He already uses RAG over Confluence docs and Slack conversations, but is stuck on how to structure multi-repo code and dependency information for agent querying, and asks for open-source tools that map interrelated repo dependencies. Useful thread to watch for tooling recommendations.

Related event: Developer Seeks Ideas for AI Agent That Auto-Diagnoses Incidents and Files Fix MRs(2 posts)→

Original post →

More from coding & agent

coding & agent channel →