plugin-eval
Three-layer quality evaluation framework for Claude Code plugins with Elo ranking.
🛡️ AgentReady threat assessment
MAESTRO 7-layer threat model + OWASP AIVSS risk score for plugin-eval, derived from its capabilities.
These scores are auto-generated from public information (the agent's own listing, docs, and repository) using the canonical OWASP AIVSS formula and the MAESTRO framework — an estimate for guidance, not a penetration test, audit, or certification. See the scoring methodology — every score is re-derived by the same automated method as an agent's public evidence changes.
Overview
A Claude Code plugin that evaluates other Claude Code plugins across a three-layer quality framework and ranks them with an Elo system. It inspects plugin structure and behavior, making it a meta-tool for vetting the very extensions that carry security surface.
Key features and capabilities
- Three-layer plugin evaluation
- Elo-based ranking
- Plugin quality vetting
Use cases
- Vet plugins before installing
- Rank plugin quality objectively