skillrank_skill BagelHole/agent-evalsMIT · open registry

skill detail

← registry

Agent Evals

BagelHole/agent-evals

Build automated evaluation suites for AI agents using golden datasets, rubrics, and regression gates. Use when shipping agent features, validating prompt changes, or gating deployments on quality.

communityprovisionalaibuildautomatedevaluation

skillrank score

████░░░░░░░░░░░░22

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

46

BagelHole/DevOps-Security-Agent-Skills

devops/ai/agent-evals

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install BagelHole/agent-evals
$ curl -fsSL skillrank.dev | sh