mirror of
https://github.com/wshobson/agents
synced 2026-06-21 14:13:58 +00:00
2e04276283
- Expanded SKILL.md to 480 lines covering all three evaluation layers, composite scoring formula with blend weights, dimension grade interpretation, all five anti-pattern flags with fix guidance, Elo ranking mechanics, CLI reference with code examples, and a troubleshooting section - Added references/rubrics.md (510 lines) with full anchored rubrics for all four judge dimensions (triggering accuracy, orchestration fitness, output quality, scope calibration) including per-level examples and calibration norms - Updated description to include specific trigger contexts for autonomous invocation
plugin-eval
Three-layer quality evaluation framework for Claude Code plugins.