Slack ran 200+ agentic E2E tests at $15-30 per run: scripts guard journeys, agents verify goals

bibryam · x · 2026-09-06

Slack Engineering published findings from 200+ agentic E2E workflow runs using Playwright MCP, Playwright CLI, and agent-generated Playwright tests.

Key takeaways:

Across runs the workflow stayed consistent (login → search → result) while paths varied in input methods, navigation patterns, and extra/skipped steps; the post details the reliability, cost, and execution-time tradeoffs.

Original post →

More from coding & agent

coding & agent channel →