skillrank_skill muratcankoylan/advanced-evaluationMIT · open registry

skill detail

← registry

advanced-evaluation

muratcankoylan/advanced-evaluation

This skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation, confidence scoring,

communityprovisionalai

skillrank score

█████████░░░░░░░58

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

17.0k

muratcankoylan/Agent-Skills-for-Context-Engineering

skills/advanced-evaluation

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install muratcankoylan/advanced-evaluation
$ curl -fsSL skillrank.dev | sh