Open-source 10 Agent Skills for LLM app building lift Claude Code scores from 0.40 to 0.76
Quirky-Location7712 · reddit · 2026-09-16
A developer packaged production-oriented LLM app engineering guidance into 10 open-source Agent Skills (plain SKILL.md files) covering architecture, tool calling, structured outputs, prompting, context/memory, RAG, subagents, model selection and evals — targeting common agent mistakes like overusing multi-agent setups or skipping measurement. A first benchmark of 96 Claude Code generations across 16 tasks scored 12 wins / 4 ties / 0 losses, mean 0.40 → 0.76 (+0.36, 95% CI [+0.22, +0.51]); only 4 skills were tested and removing style-sensitive checks cuts the delta to +0.20. Free on GitHub, works with Claude Code, Codex, Cursor, OpenCode and Copilot.
More from coding & agent
- Removed from org, 5 years of commits gone: dev can't train agent on own history — DanielLockyer · 2026-09-16
- Workshop Sep 19: explainable AI apps with Neo4j, GraphRAG, Cypher and LLM agents — camerongreen95 · 2026-09-16
- LLM bug hunt finds full-stack attack chain to brick a hardware device — matthew_d_green · 2026-09-16
- Claude Code spotted testing Sessions Hub: unified local and cloud session management — testingcatalog · 2026-09-16
- Rowboat Launches as Open-Source Multiplayer AI Assistant That Ships Code via Claude Code — ycombinator · 2026-09-16
- Dev builds FailEcho, a scanner that finds agent failures repeating across runs — EvenAd1183 · 2026-09-16