skillrank_skill davidondrej/run-deep-sweMIT · open registry

skill detail

← registry

Run Deep Swe

davidondrej/run-deep-swe

Score any AI model on the DeepSWE coding-agent benchmark via the OpenRouter API. Use when the user wants an independent, reproducible coding-agent eval — "run DeepSWE", "benchmark this model on DeepSWE", "score model X on the coding benchmark", "test a model via OpenRouter on...

communityprovisionaltestingscoreanymodel

skillrank score

████████░░░░░░░░47

SkillRank score blends community stars, real usage, and our eval lift; provisional until a skill is evaluated -- so popularity alone can't reach the top tier.

source

3.6k

davidondrej/skills

skills/agent-orchestration/run-deep-swe

open on GitHub ▸

eval status

eval pending

Success delta, token delta, and trial count are not available yet. No eval number is shown until this skill has measured results.

install

$ skillrank install davidondrej/run-deep-swe
$ curl -fsSL skillrank.dev | sh