The leaderboard is public. So are the evals.
Skills ranked by measured lift — not download counts. Every score is backed by evals you can inspect, case by case.
← Why rank by lift? See how SkillScore worksBuild your own — author, eval, govern ↓| # | Skill | SkillScore ▾ | Lift | Turns Δ | Tokens Δ | Rating | Teams |
|---|---|---|---|---|---|---|---|
| 1st | sql-result-contract ✓ benchmarked data & sql | 91 | +66% | −5% | +2% | ★ 4.9 | — |
| 2nd | secret-scanner ✓ benchmarked security | 89 | +52% | −4% | −2% | ★ 4.8 | — |
| 3rd | commit-conventions ✓ benchmarked coding | 87 | +35% | −12% | −8% | ★ 4.8 | — |
| 04 | error-triage-protocol ✓ benchmarked support | 86 | +28% | −14% | −5% | ★ 4.7 | — |
| 05 | postgres-migration-guard ✓ benchmarked ops & infra | 85 | +40% | −9% | −4% | ★ 4.7 | — |
| 06 | source-citation-guard ✓ benchmarked research & web | 84 | +44% | −10% | +6% | ★ 4.7 | — |
| 07 | query-plan-reviewer ✓ benchmarked data & sql | 83 | +38% | −8% | −2% | ★ 4.7 | — |
| 08 | pr-review-checklist ✓ benchmarked coding | 82 | +30% | −9% | −4% | ★ 4.6 | — |
| 09 | incident-postmortem ✓ benchmarked ops & infra | 82 | +34% | −13% | +7% | ★ 4.6 | — |
| 10 | refund-policy-writer ✓ benchmarked support | 81 | +30% | −11% | −4% | ★ 4.6 | — |
| 11 | prompt-injection-shield ✓ benchmarked security | 81 | +33% | −2% | −1% | ★ 4.6 | — |
| 12 | json-strict-output ✓ benchmarked coding | 80 | +22% | −6% | −3% | ★ 4.5 | — |
| 13 | search-query-planner ✓ benchmarked research & web | 80 | +29% | −8% | +4% | ★ 4.5 | — |
| 14 | pii-redactor ✓ benchmarked security | 79 | +31% | 0% | −3% | ★ 4.5 | — |
| 15 | terraform-plan-review ✓ benchmarked ops & infra | 79 | +32% | −7% | −3% | ★ 4.5 | — |
| 16 | csv-schema-inference ✓ benchmarked data & sql | 78 | +27% | −6% | +4% | ★ 4.4 | — |
| 17 | ticket-summarizer ✓ benchmarked support | 78 | +25% | −16% | −7% | ★ 4.5 | — |
| 18 | test-case-generator ✓ benchmarked coding | 77 | +24% | −7% | +3% | ★ 4.4 | — |
| 19 | schema-linter ✓ benchmarked data & sql | 77 | +26% | −7% | −2% | ★ 4.4 | — |
| 20 | page-summary-extractor ✓ benchmarked research & web | 77 | +23% | −9% | −5% | ★ 4.4 | — |
| 21 | dependency-audit ✓ benchmarked security | 76 | +25% | −6% | +5% | ★ 4.3 | — |
| 22 | k8s-manifest-lint ✓ benchmarked ops & infra | 76 | +24% | −5% | −2% | ★ 4.3 | — |
| 23 | support-triage ✓ benchmarked support | 75 | +21% | −18% | −6% | ★ 4.6 | — |
| 24 | rag-chunk-strategy ✓ benchmarked research & web | 75 | +20% | −4% | −8% | ★ 4.3 | — |
| 25 | dbt-model-conventions ✓ benchmarked data & sql | 74 | +18% | −5% | −3% | ★ 4.2 | — |
| 26 | authz-boundary-check ✓ benchmarked security | 74 | +23% | −5% | −1% | ★ 4.3 | — |
| 27 | tone-consistency-guard ✓ benchmarked support | 73 | +17% | −3% | +2% | ★ 4.2 | — |
| 28 | web-scraper-toolkit caution research & web | 73 | +22% | −3% | +5% | ★ 4.2 | — |
| 29 | changelog-writer ✓ benchmarked coding | 72 | +19% | −4% | −9% | ★ 4.1 | — |
| 30 | runbook-writer ✓ benchmarked ops & infra | 71 | +16% | −6% | −11% | ★ 4 | — |
One number, four signals — and one of them can't be bought.
Benchmark lift at 32% (pass rate with the skill minus without), live pass rate from real runs at 32%, an AI-judged quality review at 16%, and adoption at 20% — which grows logarithmically and is capped by org diversity, so volume from a single workspace can’t buy a rank. A missing signal isn’t scored as zero; the remaining weights rescale.
Nothing lists without clearing the gate.
Every submission runs a three-stage pipeline — static scan, AI security review, content check — and carries its band. Blocked skills never appear.
Know which skills your agent actually uses.
smart_route() ranks a shortlist, then compares what you offered against what the agent activated. Under 40% flags menu bloat — and shows which skills to cut.
Your team’s skills, not just everyone else’s.
Fork a public skill or write your own, prove it works on your cases, and ship it to your team — on the same rails that rank the public registry. Authoring, evals, versioning, access: skills governance without building the system.
Everything you’d build internally — already here.
Keeping a team’s skills authored, evaluated, safe, versioned, and shared is real work — skills governance. In-house it’s a stack you build and maintain. Here, it’s what the registry already does.
Adopt the skills that are proven to help.
Browse free — no sign-up. Two lines to instrument your agent.
Browse the registry