.claude/skills/evaluate-skillRuntime, accounts, dependencies, permissions, network behavior and task quality remain untested.
Measure a skill's reliability — run it k times for a pass@k score, design or interpret its eval, or compare it against the base agent. Use when the user wants to run, design, or interpret a skill's eval, or write an .eval.yaml spec.
These states come from the source or distribution context. None of the entries below are SkillVetAI compatibility test results.
These checks parse the fixed package against dated platform rules. They do not execute the Skill or verify task behavior.
.claude/skills/evaluate-skillRuntime, accounts, dependencies, permissions, network behavior and task quality remain untested.
.agents/skills/evaluate-skillRuntime, accounts, dependencies, permissions, network behavior and task quality remain untested.
skills/evaluate-skillRuntime, accounts, dependencies, permissions, network behavior and task quality remain untested.
This command is recorded from the source ecosystem and resolves the registry's latest release. The fixed release shown on this page should be inspected before adoption.
clawhub install @edonadei/evaluate-skillclawhub inspect @edonadei/evaluate-skill --version 1.0.12This automated, non-executing scan is bound to this release hash. It is not a safety certification and may contain false positives or false negatives.
This is registry-supplied evidence for the recorded release, not an independent SkillVetAI scan. Check the canonical source for the full report, scanner versions, scope, and current moderation state.
The catalog stores hashes and an inventory summary for change detection. It does not republish the package contents.
sha256:522f41fc0084022791622e6a0fe46a4f7882665194dac578df8c07500f53b104evaluate-skill.eval.yamlREFERENCE.mdreferences/evals/claude-code-smoke/claude-code-smoke.eval.yamlreferences/evals/claude-code-smoke/SKILL.mdreferences/evals/commit-simple/commit-simple.eval.yamlreferences/evals/commit-simple/SKILL.mdreferences/evals/screenshot/screenshot.eval.yamlreferences/evals/screenshot/SKILL.mdreferences/evals/summarize/SKILL.mdreferences/evals/summarize/summarize.eval.yamlreferences/evals/tdd/SKILL.mdreferences/evals/tdd/tdd.eval.yamlreferences/examples/simple.eval.yamlskill-card.mdSKILL.mdImprovements on the Caliper underlying CLI focused on reliability and usability (performance and retries)