skillrank_skill databricks/databricks-mlflow-evaluationMIT · open registry

skill detail

← registry

Databricks Mlflow Evaluation

databricks/databricks-mlflow-evaluation

MLflow 3 GenAI agent evaluation. Use when writing mlflow.genai.evaluate() code, creating @scorer functions, using built-in scorers (Guidelines, Correctness, Safety, RetrievalGroundedness), building eval datasets from traces, setting up trace ingestion and production...

communityprovisionaldocumentmlflowgenaiagent

skillrank score

█████░░░░░░░░░░░32

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

256

databricks/databricks-agent-skills

plugins/databricks/claude/skills/databricks-mlflow-evaluation

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install databricks/databricks-mlflow-evaluation
$ curl -fsSL skillrank.dev | sh