Official @ClaudeDevs: Claude Code adds `claude plugin eval`—author test cases, score a plugin/skill against a no-plugin baseline (terminal + HTML), then `claude update`. Docs at code.claude.com/docs/en/plugin-evals. Evals burn tokens; MCP/hooks run as configured.
Key Takeaways
- ✓Official claude plugin eval command to quantify plugin/skill value.
- ✓Test cases with plugin vs no-plugin baseline; terminal + HTML reports.
- ✓Requires Claude Code update; evals use tokens and run MCP/hooks.
Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.