AA-AgentPerf-Local hits GitHub: benchmark your machine's agent inference speed
ArtificialAnlys · x · 2026-09-30
Artificial Analysis published the GitHub repo and initial results for AA-AgentPerf-Local. The open-source tool replays recorded agent conversations to measure local LLM serving throughput and latency, works with any OpenAI-compatible server or auto-downloads a pinned model, and supports macOS, Linux, and Windows (managed runs need NVIDIA GPU).
More from Infra
- DeepSeek Ships TileLang Ascend Support, Signaling Ascend Now Viable for Training — teortaxesTex · 2026-09-30
- EasyPPO: just fix the critic — stable PPO for LLM post-training with zero training collapse — teortaxesTex · 2026-09-30
- Why data centers can't just pump seawater: corrosion, salt fouling, and biofouling explained — AryHHAry · 2026-09-30
- Graphsub pitches in-memory graph DB for agent data: don't trust one AI corp with it all — arthurcolle · 2026-09-30
- Even if OpenAI and Anthropic blow up at IPO, going long GPU hardware still pays — GabGarrett · 2026-09-30
- GPU dominance in deep learning is largely accidental: the case for custom chiplet inference hardware — ai · 2026-09-30