skillrank_skill mlflow/agent-evaluationMIT · open registry

skill detail

← registry

Agent Evaluation

mlflow/agent-evaluation

Use this when you need to EVALUATE OR IMPROVE or OPTIMIZE an existing LLM agent's output quality - including improving tool selection accuracy, answer quality, reducing costs, or fixing issues where the agent gives wrong/incomplete responses. Evaluates agents systematically...

communityprovisionalaiusethiswhen

skillrank score

████░░░░░░░░░░░░24

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

69

mlflow/skills

agent-evaluation

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install mlflow/agent-evaluation
$ curl -fsSL skillrank.dev | sh