Real Skill packageSource verifiedClawHub registry

evaluate-skill

Measure a skill's reliability — run it k times for a pass@k score, design or interpret its eval, or compare it against the base agent. Use when the user wants to run, design, or interpret a skill's eval, or write an .eval.yaml spec.

Identity and source

Publisher attributionEmrick Donadeiregistry owner unverified by skillvetai
Functional categoryAgent Engineering, Security & Governanceautomatically inferred · 64% rule confidence
Package forminstruction bundle15 recorded files
Canonical sourceClawHub registryclawhub:edonadei:evaluate-skill
Open canonical source ↗

Platform declarations

These states come from the source or distribution context. None of the entries below are SkillVetAI compatibility test results.

OpenClawnative officialProvenance: registry distribution

Independent structural checks

These checks parse the fixed package against dated platform rules. They do not execute the Skill or verify task behavior.

Claude Codepasses structure
Checker 0.1.0 · agent-skills-2026-08-13+claude-code-docs-2026-08-13 · 8/29/2026.claude/skills/evaluate-skill

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

OpenAI Codexpasses structure
Checker 0.1.0 · agent-skills-2026-08-13+codex-docs-2026-08-13 · 8/29/2026.agents/skills/evaluate-skill

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

OpenClawpasses structure
Checker 0.1.0 · agent-skills-2026-08-13+openclaw-docs-2026-08-13 · 8/29/2026skills/evaluate-skill

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

Installation and inspection

This command is recorded from the source ecosystem and resolves the registry's latest release. The fixed release shown on this page should be inspected before adoption.

clawhub install @edonadei/evaluate-skill
clawhub inspect @edonadei/evaluate-skill --version 1.0.12

Security evidence

SkillVetAI static result: no findings detected

This automated, non-executing scan is bound to this release hash. It is not a safety certification and may contain false positives or false negatives.

Status
completed
Coverage
full text content
Files
15 / 15 inspected as text
Checked
8/29/2026, 7:04:56 AM
Scanner
0.1.3
Policy
1.0.3
5 inferred permission indicators
  • shell execution — automatically inferred
  • network access — automatically inferred
  • filesystem read — automatically inferred
  • filesystem write — automatically inferred
  • browser control — automatically inferred
3 dependency and API indicators
  • api: clawhub.ai
  • api: mcp.example.com
  • api: youtu.be
External clawhub result: clean

This is registry-supplied evidence for the recorded release, not an independent SkillVetAI scan. Check the canonical source for the full report, scanner versions, scope, and current moderation state.

Evidence checked
8/29/2026, 6:43:43 AM
Release binding
Matches this record
  • vt: clean
  • skillspector: suspicious
  • llm: clean

Recorded files

The catalog stores hashes and an inventory summary for change detection. It does not republish the package contents.

Package content hashsha256:522f41fc0084022791622e6a0fe46a4f7882665194dac578df8c07500f53b104
Show up to 15 recorded paths
  • evaluate-skill.eval.yaml
  • REFERENCE.md
  • references/evals/claude-code-smoke/claude-code-smoke.eval.yaml
  • references/evals/claude-code-smoke/SKILL.md
  • references/evals/commit-simple/commit-simple.eval.yaml
  • references/evals/commit-simple/SKILL.md
  • references/evals/screenshot/screenshot.eval.yaml
  • references/evals/screenshot/SKILL.md
  • references/evals/summarize/SKILL.md
  • references/evals/summarize/summarize.eval.yaml
  • references/evals/tdd/SKILL.md
  • references/evals/tdd/tdd.eval.yaml
  • references/examples/simple.eval.yaml
  • skill-card.md
  • SKILL.md

Source changelog

Improvements on the Caliper underlying CLI focused on reliability and usability (performance and retries)