The leaderboard is public. So are the evals.
Skills ranked by measured lift — not download counts. Every score is backed by evals you can inspect, case by case.
← Why rank by lift? See how SkillScore worksBuild your own — author, eval, govern ↓| # | Skill | SkillScore ▾ | Lift | Turns Δ | Tokens Δ | Rating | Teams |
|---|---|---|---|---|---|---|---|
| 1st | sql-result-contract ✓ benchmarked data & sql | 91 | +66% | −5% | +2% | ★ 4.9 | — |
| 2nd | commit-conventions ✓ benchmarked coding | 87 | +35% | −12% | −8% | ★ 4.8 | — |
| 3rd | postgres-migration-guard ✓ benchmarked ops & infra | 85 | +40% | −9% | −4% | ★ 4.7 | — |
| 04 | error-triage-protocol ✓ benchmarked support | 83 | +28% | −14% | −5% | ★ 4.7 | — |
| 05 | support-triage ✓ benchmarked support | 82 | +28% | −18% | −6% | ★ 4.6 | — |
| 06 | prompt-injection-shield ✓ benchmarked security | 81 | +33% | −2% | −1% | ★ 4.6 | — |
| 07 | json-strict-output ✓ benchmarked coding | 80 | +22% | −6% | −3% | ★ 4.5 | — |
| 08 | pii-redactor ✓ benchmarked security | 79 | +31% | 0% | −3% | ★ 4.5 | — |
| 09 | schema-linter ✓ benchmarked data & sql | 77 | +26% | −7% | −2% | ★ 4.4 | — |
| 10 | web-scraper-toolkit caution research & web | 74 | +22% | −3% | +5% | ★ 4.2 | — |
| 11 | changelog-writer early coding | 73 · early | +19% | −4% | −9% | ★ 4.1 | — |
One number, four signals — and one of them can't be bought.
Benchmark lift at 32% (pass rate with the skill minus without), live pass rate from real runs at 32%, an AI-judged quality review at 16%, and adoption at 20% — which grows logarithmically and is capped by org diversity, so volume from a single workspace can’t buy a rank. A missing signal isn’t scored as zero; the remaining weights rescale.
Nothing lists without clearing the gate.
Every submission runs a three-stage pipeline — static scan, AI security review, content check — and carries its band. Blocked skills never appear.
Know which skills your agent actually uses.
smart_route() ranks a shortlist, then compares what you offered against what the agent activated. Under 40% flags menu bloat — and shows which skills to cut.
Your team’s skills, not just everyone else’s.
Fork a public skill or write your own, prove it works on your cases, and ship it to your team — on the same rails that rank the public registry. Authoring, evals, versioning, access: skills governance without building the system.
Everything you’d build internally — already here.
Keeping a team’s skills authored, evaluated, safe, versioned, and shared is real work — skills governance. In-house it’s a stack you build and maintain. Here, it’s what the registry already does.
Adopt the skills that are proven to help.
Browse free — no sign-up. Two lines to instrument your agent.
Browse the registry