skills/brand-voice from affaan-m/ECC
Voice tasks, run with the skill loaded and without it, on the same task pairs.
Identity pin
- Repository
- affaan-m/ECC
- Path
- skills/brand-voice
- Commit
d8409a4b0813771235555e32e3d8046a73988bfa- Content hash
eae455eed766cf8ab9d2859713386c56fdce39a8e6a80b60a006a0a076e1d35a- Date tested
- not yet tested
The pin is the whole of this skill's identity here. It resolves at https://github.com/affaan-m/ECC/tree/d8409a4b0813771235555e32e3d8046a73988bfa/skills/brand-voice, and the content hash is a sha256 over exactly the text the with arm was given. Nothing else about the skill appears on this site.
What was claimed
- Loading skills/brand-voice from affaan-m/ECC improves outputs on voice tasks.
Pass criterion: With-arm win rate at or above 0.60 across the same 10 briefs, with the interval excluding a rate of 0.50. Scale: unit. Same instrument and scale as S09.
Shares the voice task set with S09. Paired blind comparison only, and labelled as such on every surface that renders it.
Injected context tokens
| Arm | Context characters | Injected context tokens |
|---|---|---|
| without | 0 | 0 |
| with | 4,780 | 1,122 |
The character count is exact: it is the length of the text the with arm is given, and the content hash above is a sha256 over that same text. The token figure is an estimate at 4.26 characters per token, the ratio the phase 1 run measured over 3,120 calls, and it is labelled an estimate until a run reports its own token counts. The without arm is given the identical prompt and nothing else, so its zero is a measurement rather than a missing value.
Not yet measured
No run records exist for this skill. The instrument is built and committed, the task sets are frozen, and the projected cost of the run is published, but no model has been called. There is therefore no verdict, no effect, no interval and no cost per task on this page, and none is shown.
- S10-ecc-brand-voice: 10 task pairs per model. Paired blind comparison, win rate with ties counting half.
The committed instrument is at harness/claims/claims-skills.json, harness/claims/tasksets-skills.json and harness/claims/skill-cohort.json. The projection for the run that has not happened is at harness/results/dry-run-projection-skills.md. How a claim gets tested.