Skill Grader / report card

D

NeoLabHQ/context-engineering-kit

  • Agents & automation
  • Coding workflow

Graded D on context cost and craft. 1,594 tokens of this skill ride in the system prompt of every message, whether or not it fires.

Ranked #85 of 85 on the leaderboard →

Score
1.00 / 4
Resident
1,594 tokens
Footprint
heavy
Per activation
28,944 tokens
Share this grade on X ↗Re-grade now →View on GitHub ↗Grade another skillLast graded Oct 11, 2026 from HEAD

Add the badge to your README

It shows the current letter grade and links back to this page. Improve the skill, hit Re-grade now, and the badge follows — GitHub may keep the old image for up to an hour.

Skill Grader: D
Markdown
[![Skill Grader: D](https://seoagent.com/skill-grader/badge/NeoLabHQ/context-engineering-kit.svg)](https://seoagent.com/skill-grader/NeoLabHQ/context-engineering-kit)
HTML
<a href="https://seoagent.com/skill-grader/NeoLabHQ/context-engineering-kit"><img src="https://seoagent.com/skill-grader/badge/NeoLabHQ/context-engineering-kit.svg" alt="Skill Grader: D" height="20"></a>

Report card

D
D overall · 1.00/4
github.com/NeoLabHQ/context-engineering-kit@HEAD · 40 skills · 1,594t resident (heavy) · 28,944t per activation · 256,755t bundle
F
Resident footprint
1,594t of description text ride in the system prompt of every message across 40 skills (paper's observed per-skill range: 50–280t).
C
Description honesty
Defensive padding detected: 56 comma-separated trigger keywords; 788 words (a quiet description is ~15–40). Padding is individually rational and collectively ruinous — it dilutes every co-installed skill.
F
Body size
Largest body is 16,240 words (~28,944t), loaded on every trigger. Corpus median 921, p90 2207.
F
Progressive disclosure
Everything is inlined into a 16,240-word body with no reference layer to disclose progressively.
A
Factoring
Scoped to a coherent procedure set.
A
CLI leverage
Drives a CLI (`git`, 103 invocations + 175 shell blocks) — verification and state live in code, not prompt.
D overall (1.00/4). 40 skill(s), 1,594t resident on every message (heavy), 28,944t loaded per activation. Weakest axis: resident footprint (F).

Six deterministic axes, no LLM, same result every run. Tokens are approximated as chars/4. Read the methodology, or see how SEO skills score on task outcomes in the open SEO skill benchmark.