Stanford CS329Z Assignments: Build a QA Agent, Then Design Its Eval Suite

jyangballin · x · 2026-09-05

Developer @codehiyouga recommends Stanford's new CS329Z: Engineering AI Agents course as a checklist for anyone who already calls models, connects tools, and builds agents — especially if your system keeps gaining features but you can't tell whether each change actually helps.

The course covers RAG, tool use, MCP, agent frameworks, memory, multi-agent systems, optimization, agent data, evaluation, plus safety and reliability of long-running agents.

The assignments are concrete: build a research-paper QA agent from scratch, then design an evaluation suite with benchmark tasks, code-based graders, an LLM-as-judge, and error analysis.

Related event: Stanford launches CS329Z on engineering AI agents(2 posts)→

Original post →

More from coding & agent

coding & agent channel →