skillrank_skill oimiragieo/agent-evaluationMIT · open registry

skill detail

← registry

Agent Evaluation

oimiragieo/agent-evaluation

LLM-as-judge evaluation framework with 5-dimension rubric (accuracy, groundedness, coherence, completeness, helpfulness) for scoring AI-generated content quality with weighted composite scores and evidence citations

communityprovisionaldocumentllmjudgeevaluation

skillrank score

███░░░░░░░░░░░░░21

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

36

oimiragieo/agent-studio

.claude/skills/agent-evaluation

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install oimiragieo/agent-evaluation
$ curl -fsSL skillrank.dev | sh