skill detail
← registryDatabricks Mlflow Evaluation
databricks/databricks-mlflow-evaluation
MLflow 3 GenAI agent evaluation. Use when writing mlflow.genai.evaluate() code, creating @scorer functions, using built-in scorers (Guidelines, Correctness, Safety, RetrievalGroundedness), building eval datasets from traces, setting up trace ingestion and production...
skillrank score
SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.
source
★ 256
databricks/databricks-agent-skills
plugins/databricks/claude/skills/databricks-mlflow-evaluation
open on GitHub ▸eval status
eval pending
Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.
install
$ skillrank install databricks/databricks-mlflow-evaluation$ curl -fsSL skillrank.dev | sh