Skills, tested
Ten skills across five task classes. Each is run twice on the same task pairs: once with the skill loaded as system context, once with the identical prompt and nothing injected.
Skill authoring
Deterministic, scored against a committed answer key. Own and external skills are listed together and ranked only against each other.
| Skill | Cohort | Commit | Injected context tokens | Effect | Tested |
|---|---|---|---|---|---|
| No skill loaded | baseline | not applicable | 0 | 0 by construction | not yet |
| skills/skill-creator from anthropics/skills | external | 3b3fad96af16 | 10,587 | not yet measured | not yet |
| skills/skill-creation-walkthrough from rampstackco/claude-skills (operated by this site) | own | 047924252254 | 8,687 | not yet measured | not yet |
How this external slot was filled
Decided by: two candidates named the class, so the closer output format took it. The deciding text is at: SKILL.md, section heading, in anthropics/skills skills/skill-creator at 3b3fad96af16.
Two candidates name the class. The task class is scored on a single SKILL.md whose frontmatter must parse and whose section order must match SKILL_AUTHORING.md. writing-skills prescribes that document's section order under the heading above. skill-creator prescribes an authoring and evaluation workflow, and the structure it prescribes under its own Report structure heading is an evaluation report rather than the SKILL.md. writing-skills is the closer output format.
Addendum. RECORDED AFTER THE FACT, ON 2026-08-26, AND NOT ACTED ON. This tiebreak was decided when the skill-authoring class was scored against SKILL_AUTHORING.md alone. That document is now Key B, disclosed and unranked, and the ranked key is the open Agent Skills specification. The tiebreak reasoning above therefore turns on a basis that is no longer the ranked one. It is left exactly as it was decided, because a selection record that is rewritten to agree with a later ruling is not a record of what was decided. The selection was NOT re-run against Key A, and whether it would produce the same candidate is an open question stated here rather than assumed.
Second addendum. SECOND ADDENDUM, 2026-08-26. THE TIEBREAK WAS RE-RUN WITH KEY A AS THE OUTPUT-FORMAT BASIS, AND IT SWAPPED. Basis: the selection rule is unchanged, but the output format the class is scored on is now the open Agent Skills specification rather than SKILL_AUTHORING.md. Candidates compared: the same two that qualified, obra/superpowers skills/writing-skills and anthropics/skills skills/skill-creator. Result: skill-creator. It prescribes the specification's own model of the artifact, naming the required and the optional frontmatter fields, the bundled directory layout, the progressive disclosure levels and the line ceiling. writing-skills prescribes a body section order, which is the one part of a SKILL.md the specification explicitly leaves unconstrained, and it disagrees with the specification in three places: it puts the thousand and twenty four character limit on the whole frontmatter rather than on the description, it tells an author the description must not say what the skill does where the specification asks for both, and the example name in its own template is capitalised where the specification allows lowercase only. Run through Key A, writing-skills' own template fails a predicate and skill-creator's passes every applicable one. The original record and the first addendum above are left exactly as written. obra/superpowers skills/writing-skills moves to the rejected pins, where its bytes stay committed and its quoted sentence stays resolvable.
| Candidate | Commit | Content hash | Qualified | Why, in our words |
|---|---|---|---|---|
| skills/writing-skills from obra/superpowers | b36e0829c6d0 | d34db5c8aed6 | yes | Names the class in its when-to-use directly. |
| skills/skill-creator from anthropics/skills | 3b3fad96af16 | e51ce07ea7fd | yes | Also names the class directly, so the tiebreak applies. |
| skills/skill-scout from affaan-m/ECC | d8409a4b0813 | 71b1e8a4aede | no | Searches for an existing skill before one is written. It names the moment before authoring, not authoring. |
| skills/skill-comply from affaan-m/ECC | d8409a4b0813 | e26cbbc30a32 | no | Measures whether a written skill is obeyed. It is about a skill, not about writing one. |
Accessibility audit
Deterministic, scored against a committed answer key. Own and external skills are listed together and ranked only against each other.
| Skill | Cohort | Commit | Injected context tokens | Effect | Tested |
|---|---|---|---|---|---|
| No skill loaded | baseline | not applicable | 0 | 0 by construction | not yet |
| skills/accessibility from affaan-m/ECC | external | d8409a4b0813 | 1,546 | not yet measured | not yet |
| skills/accessibility-audit from rampstackco/claude-skills (operated by this site) | own | 047924252254 | 10,239 | not yet measured | not yet |
How this external slot was filled
Decided by: one candidate named the class in its own when-to-use. The deciding text is at: SKILL.md frontmatter, description, in affaan-m/ECC skills/accessibility at d8409a4b0813.
| Candidate | Commit | Content hash | Qualified | Why, in our words |
|---|---|---|---|---|
| skills/accessibility from affaan-m/ECC | d8409a4b0813 | ab86a50717d6 | yes | Names auditing against WCAG in the when-to-use itself. |
| skills/click-path-audit from affaan-m/ECC | d8409a4b0813 | 662fc5d02483 | no | An audit of behavioural state, not of accessibility. The word audit is shared and the class is not. |
- obra/superpowers has no skill naming accessibility in any when-to-use.
- anthropics/skills has none either; webapp-testing names browser testing rather than accessibility.
Spec writing
Deterministic, scored against a committed answer key, and paired blind comparison, win rate with ties counting half. Reported as two rows, never averaged.. Own and external skills are listed together and ranked only against each other.
| Skill | Cohort | Commit | Injected context tokens | Effect | Tested |
|---|---|---|---|---|---|
| No skill loaded | baseline | not applicable | 0 | 0 by construction | not yet |
| skills/product-capability from affaan-m/ECC | external | d8409a4b0813 | 1,046 | not yet measured | not yet |
| skills/pm-spec-writing from rampstackco/claude-skills (operated by this site) | own | 047924252254 | 6,292 | not yet measured | not yet |
How this external slot was filled
Decided by: two candidates named the class, so the closer output format took it. The deciding text is at: SKILL.md, section heading, in affaan-m/ECC skills/product-capability at d8409a4b0813.
Two candidates name the class. The task class is scored on a document with required sections and acceptance criteria, so the closer output format is the one that fixes its sections. product-capability declares a canonical artifact and a section-by-section output format under the heading above. doc-coauthoring prescribes a three-stage collaboration and fixes no sections in the document it produces.
Addendum. RECORDED AFTER THE FACT, ON 2026-08-26, AND NOT ACTED ON. This tiebreak was decided when the skill-authoring class was scored against SKILL_AUTHORING.md alone. That document is now Key B, disclosed and unranked, and the ranked key is the open Agent Skills specification. The tiebreak reasoning above therefore turns on a basis that is no longer the ranked one. It is left exactly as it was decided, because a selection record that is rewritten to agree with a later ruling is not a record of what was decided. The selection was NOT re-run against Key A, and whether it would produce the same candidate is an open question stated here rather than assumed.
Second addendum. SECOND ADDENDUM, 2026-08-26. RE-EXAMINED, NOT RE-RUN, AND THE CANDIDATE IS UNCHANGED. Basis: this tiebreak was never decided on the operator's house standard. The spec-writing class is scored on the section list and the acceptance-criterion pattern that the task prompt states to BOTH arms, which is a contract the two arms are given rather than a convention one of them was taught, and R7 did not move it. Key A is the answer key for the skill-authoring class and has no bearing here. Candidates compared: the same two that qualified, affaan-m/ECC skills/product-capability and anthropics/skills skills/doc-coauthoring. Result: unchanged, product-capability. THE FIRST ADDENDUM ABOVE IS WRONG ON THIS SLOT and is left in place rather than edited. It was applied to both output-format tiebreaks at once and says this one was decided against SKILL_AUTHORING.md, which it was not. Correcting it by rewriting would hide that the error was made; this sentence is the correction.
| Candidate | Commit | Content hash | Qualified | Why, in our words |
|---|---|---|---|---|
| skills/product-capability from affaan-m/ECC | d8409a4b0813 | 3e0052802969 | yes | Names writing a specification from product intent. |
| skills/doc-coauthoring from anthropics/skills | 3b3fad96af16 | 2e47d78846fa | yes | Names technical specs among several document kinds, so the tiebreak applies. |
| skills/writing-plans from obra/superpowers | b36e0829c6d0 | 48508f44bbfd | no | Takes a spec as its input. It is downstream of the class, not the class. |
| skills/product-lens from affaan-m/ECC | d8409a4b0813 | d082be7c3dd9 | no | Rules itself out in its own words and hands the class to product-capability. |
On-page audit
Deterministic, scored against a committed answer key. Own and external skills are listed together and ranked only against each other.
| Skill | Cohort | Commit | Injected context tokens | Effect | Tested |
|---|---|---|---|---|---|
| No skill loaded | baseline | not applicable | 0 | 0 by construction | not yet |
| skills/seo from affaan-m/ECC | external | d8409a4b0813 | 1,018 | not yet measured | not yet |
| skills/seo-onpage from rampstackco/claude-skills (operated by this site) | own | 047924252254 | 5,117 | not yet measured | not yet |
How this external slot was filled
Decided by: one candidate named the class in its own when-to-use. The deciding text is at: SKILL.md frontmatter, description, in affaan-m/ECC skills/seo at d8409a4b0813.
| Candidate | Commit | Content hash | Qualified | Why, in our words |
|---|---|---|---|---|
| skills/seo from affaan-m/ECC | d8409a4b0813 | 9a655a52cfd9 | yes | Names auditing and on-page optimization in the same sentence. |
- obra/superpowers has no skill naming search or on-page work in any when-to-use.
- anthropics/skills has none either.
Voice
Paired blind comparison, win rate with ties counting half. Own and external skills are listed together and ranked only against each other.
| Skill | Cohort | Commit | Injected context tokens | Effect | Tested |
|---|---|---|---|---|---|
| No skill loaded | baseline | not applicable | 0 | 0 by construction | not yet |
| skills/brand-voice from affaan-m/ECC | external | d8409a4b0813 | 1,122 | not yet measured | not yet |
| skills/brand-voice from rampstackco/claude-skills (operated by this site) | own | 047924252254 | 5,236 | not yet measured | not yet |
How this external slot was filled
Decided by: one candidate named the class in its own when-to-use. The deciding text is at: SKILL.md frontmatter, description, in affaan-m/ECC skills/brand-voice at d8409a4b0813.
| Candidate | Commit | Content hash | Qualified | Why, in our words |
|---|---|---|---|---|
| skills/brand-voice from affaan-m/ECC | d8409a4b0813 | eae455eed766 | yes | Names writing voice and consistency of it. |
| skills/brand-guidelines from anthropics/skills | 3b3fad96af16 | 1120b3769e29 | no | Visual identity, colors and type. Brand is shared and voice is not. |
- obra/superpowers has no skill naming writing voice in any when-to-use.
Nothing has been measured yet
The cohort is pinned, the task sets are frozen, the answer keys are committed and the cost of the run is projected. No model has been called, so no effect, no interval and no cost appears above. The table shows the instrument, not a result.