Skills registry

The leaderboard is public. So are the evals.

Skills ranked by measured lift — not download counts. Every score is backed by evals you can inspect, case by case.

← Why rank by lift? See how SkillScore worksBuild your own — author, eval, govern ↓
AllCodingData & SQLSupportSecurityResearch & webOps & infra✓ Benchmarked only
Top skillsby SkillScore · all categories
lift measured on gemini-3.5-flash · 21+ cases per skill · re-run any skill on your own model with the open runner
#SkillSkillScoreLiftTurns ΔTokens ΔRatingTeams
1st
sql-result-contract ✓ benchmarked
data & sql
91+66%−5%+2% 4.9
2nd
commit-conventions ✓ benchmarked
coding
87+35%−12%−8% 4.8
3rd
postgres-migration-guard ✓ benchmarked
ops & infra
85+40%−9%−4% 4.7
04
error-triage-protocol ✓ benchmarked
support
83+28%−14%−5% 4.7
05
support-triage ✓ benchmarked
support
82+28%−18%−6% 4.6
06
prompt-injection-shield ✓ benchmarked
security
81+33%−2%−1% 4.6
07
json-strict-output ✓ benchmarked
coding
80+22%−6%−3% 4.5
08
pii-redactor ✓ benchmarked
security
79+31%0%−3% 4.5
09
schema-linter ✓ benchmarked
data & sql
77+26%−7%−2% 4.4
10
web-scraper-toolkit caution
research & web
74+22%−3%+5% 4.2
11
changelog-writer early
coding
73 · early+19%−4%−9% 4.1
1st
sql-result-contract ✓ benchmarked
data & sql · ★ 4.9
91
+66%
2nd
commit-conventions ✓ benchmarked
coding · ★ 4.8
87
+35%
3rd
postgres-migration-guard ✓ benchmarked
ops & infra · ★ 4.7
85
+40%
04
error-triage-protocol ✓ benchmarked
support · ★ 4.7
83
+28%
05
support-triage ✓ benchmarked
support · ★ 4.6
82
+28%
06
prompt-injection-shield ✓ benchmarked
security · ★ 4.6
81
+33%
07
json-strict-output ✓ benchmarked
coding · ★ 4.5
80
+22%
08
pii-redactor ✓ benchmarked
security · ★ 4.5
79
+31%
09
schema-linter ✓ benchmarked
data & sql · ★ 4.4
77
+26%
10
web-scraper-toolkit caution
research & web · ★ 4.2
74
+22%
11
changelog-writer early
coding · ★ 4.1
73
+19%
illustrative sampleSee the live ranking →
Browse the catalog8,000+ skills · card view
sql-result-contract ✓ benchmarked
data & sql
91SkillScore
Validate query results against a typed contract before they return.
+66% lift4.9view evals →
commit-conventions ✓ benchmarked
coding
87SkillScore
Enforce your repo's commit format — consistent, parseable messages.
+35% lift4.8view evals →
postgres-migration-guard ✓ benchmarked
ops & infra
85SkillScore
Catch unsafe schema migrations before they hit production.
+40% lift4.7view evals →
error-triage-protocol ✓ benchmarked
support
83SkillScore
Triage a failing run to the right owner with a consistent, reviewable protocol.
+28% lift4.7view evals →
support-triage ✓ benchmarked
support
82SkillScore
Route support conversations to the right resolution path.
+28% lift4.6view evals →
prompt-injection-shield ✓ benchmarked
security
81SkillScore
Detect and block prompt-injection attempts in tool inputs.
+33% lift4.6view evals →
json-strict-output ✓ benchmarked
coding
80SkillScore
Lint model output against a strict JSON schema before anything downstream reads it.
+22% lift4.5view evals →
pii-redactor ✓ benchmarked
security
79SkillScore
Strip PII from traces before they leave your workspace.
+31% lift4.5view evals →
schema-linter ✓ benchmarked
data & sql
77SkillScore
Lint output schemas for missing or drifted required fields.
+26% lift4.4view evals →
web-scraper-toolkit caution
research & web
74SkillScore
Structured extraction helpers — listed with a caution note.
+22% lift4.2view evals →
changelog-writer early
coding
73SkillScore
Draft release changelogs from your merged PRs.
+19% lift4.1view evals →
SkillScore

One number, four signals — and one of them can't be bought.

Benchmark lift at 32% (pass rate with the skill minus without), live pass rate from real runs at 32%, an AI-judged quality review at 16%, and adoption at 20% — which grows logarithmically and is capped by org diversity, so volume from a single workspace can’t buy a rank. A missing signal isn’t scored as zero; the remaining weights rescale.

SkillScore · commit-conventions
Benchmark lift · 32%+35%
Live pass rate · 32%86%
AI rating · 16%A
Adoption · 20%
SkillScore87
no adoption signal yet — the other three weights rescale to 100%
SkillSafety

Nothing lists without clearing the gate.

Every submission runs a three-stage pipeline — static scan, AI security review, content check — and carries its band. Blocked skills never appear.

SkillSafety bands
passedcleared all three stages
cautionlisted with the reason shown
blockednever lists
Skill router

Know which skills your agent actually uses.

smart_route() ranks a shortlist, then compares what you offered against what the agent activated. Under 40% flags menu bloat — and shows which skills to cut.

Offered vs activated · last 7d
Offered12 skills
Activated5 skills
menu bloat42% activation — trim 7
How a skill earns its badge
01Scanned02Security review03Content check04Benchmarked05Band + score published
Build on it

Your team’s skills, not just everyone else’s.

Fork a public skill or write your own, prove it works on your cases, and ship it to your team — on the same rails that rank the public registry. Authoring, evals, versioning, access: skills governance without building the system.

Your eval · refund-policy-writer
Pass rate without54%
Pass rate with84%
Lift+30 pts
workspace-only · illustrative — run it with the open eval runner
From your draft to your team’s router
01Author or forkStart from a benchmarked skill or a blank one. The skill format, linter, and local safety scan are built in.
02Eval before you shipRun the same harness that scores public skills against your own cases — pass rate before and after, and the exact lift.
03Share with your teamPublish to your workspace. Your agents pick it up via smart_route() — and real usage feeds back into the score.
04Governed by defaultThe same safety gate and pinned versions as the public registry. Governance runs on the rails — not as a process you police.
Skills governance

Everything you’d build internally — already here.

Keeping a team’s skills authored, evaluated, safe, versioned, and shared is real work — skills governance. In-house it’s a stack you build and maintain. Here, it’s what the registry already does.

Skills governanceBuild it in-houseOn DecimalAI
Eval harness — with / without runsyou build it✓ the same harness that ranks the registry
Safety review on every versionyou build it✓ SkillSafety gate included
Quality ranking nobody gamesyou build it✓ SkillScore — measured lift
Versioning & pinningyou build it✓ pinned versions built in
Sharing & access controlyou build it✓ workspace publishing
the upkeepweeks of engineering · yours to maintainpip install decimalai · today
Publish your first skill →
FORK & EVAL ON FREE · TEAM PUBLISHING FROM CORE $49
One engine

The engine that ranks skills also versions your whole agent.

Explore Agent versioning →

Adopt the skills that are proven to help.

Browse free — no sign-up. Two lines to instrument your agent.

Browse the registry