Skill Grader

Every installed skill's description rides in the system prompt of every message — whether or not it ever fires — and heavy descriptions dilute the triggering of everything else you have installed. Paste a GitHub repo or a SKILL.md and get a report card on what your skill costs: resident footprint, description honesty, body size, progressive disclosure, factoring, and CLI leverage. Deterministic, free, no signup.

Methodology: “@skills: Attention Is All You Have” (Yin et al., 2026) + the open seo-skill-bench footprint scorer. We graded our own skill first — it got a C+, so we rebuilt it.

Top-graded skills right now

See all 20 ranked skills ↓
  1. #1833
    A
    gemini-skill
    WJZ-P
    3.95/4 · 48t resident
    • Images & media
    • Compact body
    • Progressive disclosure
  2. #2260
    A
    ip-diagram-creator
    haloshin
    3.90/4 · 86t resident
    • Design & frontend
    • Images & media
    • Compact body
    • Progressive disclosure
    • Rules in prose
  3. #3762
    A
    web-design
    xiaopu-ai
    3.80/4 · 70t resident
    • Design & frontend
    • Progressive disclosure
    • CLI leverage

Leaderboard

Every listed skill, best score first — ties go to the lighter footprint. Blue tags say what a skill is for; green and amber say what it does well and what is costing it rank. Fix yours, grade it again above, and your rank and README badge update right away. Listed: public GitHub repos with 25+ stars.

RankSkillScoreStars
#1Agemini-skill · WJZ-P
  • Images & media
  • Compact body
  • Progressive disclosure
3.95/4833
#2Aip-diagram-creator · haloshin
  • Design & frontend
  • Images & media
  • Compact body
  • Progressive disclosure
  • Rules in prose
3.90/4260
#3Aweb-design · xiaopu-ai
  • Design & frontend
  • Progressive disclosure
  • CLI leverage
3.80/4762
#4Arepo-task-proof-loop · DenisSergeevitch
  • Coding workflow
  • Agents & automation
  • Progressive disclosure
  • CLI leverage
3.80/4729
#5ABREAK · JDArmy
  • Security
  • Data & research
  • Compact body
  • CLI leverage
  • Loads everything up front
3.70/4383
#6Amarkit · shift-labs-ai
  • Documents & files
  • Data & research
  • Compact body
  • Loads everything up front
3.65/41.3k
#7Aalgorithms · luofengmacheng
  • Compact body
  • Loads everything up front
  • Rules in prose
3.60/469
#8Alieflat-less-ai-tone · larashero3-dotcom
  • Writing
  • Compact body
  • Loads everything up front
  • Rules in prose
3.60/4743
#9Amcp-server-typescript · dataforseo
  • SEO & marketing
  • Data & research
  • CLI leverage
  • Loads everything up front
3.50/4246
#10Bguardian-cli · zakirkun
  • Security
  • Agents & automation
  • Loads everything up front
  • CLI leverage
3.40/41.9k
#11Btotal-recall · gavdalf
  • Memory & context
  • Loads everything up front
  • CLI leverage
3.40/4272
#12Bgodot-solana-sdk · Virus-Axel
  • Coding workflow
  • Loads everything up front
  • Rules in prose
3.30/478
#13Bdebugpy · microsoft
  • Coding workflow
  • Debugging & testing
  • Rules in prose
  • Compact body
  • 7 skills
3.25/42.5k
#14Bsynapse-admin · Awesome-Technologies
  • Coding workflow
  • Design & frontend
  • Long body
  • Loads everything up front
  • 2 skills
3.05/41.2k
#15BAgentRecall-X · Goldentrii
  • Memory & context
  • Agents & automation
  • Long body
  • Loads everything up front
  • CLI leverage
3.00/4371
#16Bprecise · microprediction
  • Data & research
  • Coding workflow
  • Heavy description
  • Padded description
  • 6 skills
2.70/4335
#17Bxiaobei · TeamWiseFlow
  • Agents & automation
  • SEO & marketing
  • Heavy description
  • Progressive disclosure
  • 40 skills
2.70/48.5k
#18Creact-admin · marmelab
  • Coding workflow
  • Design & frontend
  • Padded description
  • Long body
  • 2 skills
2.35/426.9k
#19Cgsd-browser · gsd-build
  • Agents & automation
  • Debugging & testing
  • Heavy description
  • Padded description
  • 2 skills
2.30/4265
#20Copenclaw-marketing-skills · LeoYeAI
  • SEO & marketing
  • Writing
  • Heavy description
  • Padded description
  • 39 skills
2.00/41k

FAQ

What does the Skill Grader measure?

Six deterministic axes: resident footprint (the tokens your skill descriptions occupy in the system prompt of every message), description honesty (statement of purpose vs. trigger-keyword padding), body size vs. the published-skill corpus (median 921 words, p90 2,207), progressive disclosure (is detail fetched at point of use?), factoring (one procedure per skill vs. monolith), and CLI leverage (rules pushed into code cost zero attention). Tokens are approximated as chars/4.

Why does an installed skill cost tokens on every message?

A skill’s frontmatter description is resident in the agent’s system prompt so it can auto-trigger — it loads on every message whether or not the skill ever fires. Research ("@skills: Attention Is All You Have", Yin et al. 2026) measured descriptions at 50–280 tokens each and showed heavy descriptions dilute the trigger reliability of everything else installed alongside them.

What grade does SEOAgent’s own skill get?

We graded ourselves first — the original skill scored a C+ (17,237-word body). We rebuilt it against this methodology: the body dropped 70% and every protocol moved into reference files fetched at point of use. Run the grader on @seoagent-official/seoagent and check.