Real Skill packageSource verifiedClawHub registry

Patello Bench

Turn a production assistant failure into a permanent, discriminating LLM benchmark case — genericized prompt with frozen evidence, judge, model-matrix run, discrimination check.

Identity and source

Publisher attributionPatrik Ekenbergregistry owner unverified by skillvetai
Functional categoryWriting, Content & Translationautomatically inferred · 56% rule confidence
Package forminstruction with code44 recorded files
Canonical sourceClawHub registryclawhub:patello:bench-it
Open canonical source ↗

Platform declarations

These states come from the source or distribution context. None of the entries below are SkillVetAI compatibility test results.

OpenClawnative officialProvenance: registry distribution

Independent structural checks

These checks parse the fixed package against dated platform rules. They do not execute the Skill or verify task behavior.

Claude Codepasses structure
Checker 0.1.0 · agent-skills-2026-08-13+claude-code-docs-2026-08-13 · 8/27/2026.claude/skills/bench-it

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

OpenAI Codexpasses structure
Checker 0.1.0 · agent-skills-2026-08-13+codex-docs-2026-08-13 · 8/27/2026.agents/skills/bench-it

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

OpenClawpasses structure
Checker 0.1.0 · agent-skills-2026-08-13+openclaw-docs-2026-08-13 · 8/27/2026skills/bench-it

Runtime, accounts, dependencies, permissions, network behavior and task quality remain untested.

Installation and inspection

This command is recorded from the source ecosystem and resolves the registry's latest release. The fixed release shown on this page should be inspected before adoption.

clawhub install @patello/bench-it
clawhub inspect @patello/bench-it --version 0.2.1

Security evidence

SkillVetAI static result: no findings detected

This automated, non-executing scan is bound to this release hash. It is not a safety certification and may contain false positives or false negatives.

Status
completed
Coverage
partial text content
Files
40 / 44 inspected as text
Checked
8/27/2026, 10:18:59 PM
Scanner
0.1.3
Policy
1.0.3
5 inferred permission indicators
  • network access — automatically inferred
  • filesystem read — automatically inferred
  • filesystem write — automatically inferred
  • credential access — automatically inferred
  • browser control — automatically inferred
12 dependency and API indicators
  • api: api.semanticscholar.org
  • api: arxiv.org
  • api: cloudgamesdb.io
  • api: doi.org
  • api: github.com
  • api: meadowcupracing.com
  • api: microsoft.github.io
  • api: openrouter.ai
  • api: playgalaxy.gg
  • api: python.useinstructor.com
  • api: storefrontb.example
  • api: streamcheck.io
External clawhub result: suspicious

This is registry-supplied evidence for the recorded release, not an independent SkillVetAI scan. Check the canonical source for the full report, scanner versions, scope, and current moderation state.

Evidence checked
8/27/2026, 9:01:14 PM
Release binding
Matches this record
  • skillspector: suspicious
  • llm: suspicious

Recorded files

The catalog stores hashes and an inventory summary for change detection. It does not republish the package contents.

Package content hashsha256:aabc037329fe845f06c36307f8ea6fdff21c7e216391321d1264ce87eb740593
Show up to 30 recorded paths
  • _meta.json
  • bench/gen_providers.py
  • bench/intake.py
  • bench/README.md
  • bench/run_bench.py
  • bench/run_case.py
  • changelog.txt
  • data/cases/claim-follows-from-evidence-001/case.json
  • data/cases/claim-follows-from-evidence-001/notes.md
  • data/cases/claim-follows-from-evidence-001/prompt.md
  • data/cases/claim-follows-from-evidence-001/rubric.md
  • data/public/synthetic-tasks.json
  • judges/__init__.py
  • judges/base.py
  • judges/classifiers/__init__.py
  • judges/classifiers/auto/__init__.py
  • judges/classifiers/auto/_metrics.py
  • judges/classifiers/auto/_prompts.py
  • judges/classifiers/auto/core.py
  • judges/classifiers/correctness.py
  • judges/classifiers/hallucination.py
  • judges/classifiers/harmfulness.py
  • judges/classifiers/patello.py
  • judges/classifiers/query_quality.py
  • judges/classifiers/refusal.py
  • judges/cli/__init__.py
  • judges/cli/entrypoint.py
  • judges/graders/__init__.py
  • judges/graders/correctness.py
  • judges/graders/empathy.py

Source changelog

- fix CI: contents:write permission for release creation; bump 0.2.1