Copywriting tone of voice creator: set up for testing, not yet run
Copywriting tone of voice creator, from samber/cc-skills. Its voice tasks were run with the skill loaded and without it. Each time it is the same task, run twice.
What we tested
Whether loading this skill helps a model write in a brand's voice.
What counts as helping
A blind reader has to prefer the version written with the skill at least 60 times in 100, across the same 10 briefs.
How we scored it
A blind reader compared the two versions and picked one, without being told which was which. The score is how often the version written with the skill was the one picked, from 0 to 1.
Nothing has been measured on this page yet.
Show per-task detail
What exactly was tested, and how it was scored
- Repository
- samber/cc-skills
- Path
- skills/copywriting-tone-of-voice-creator
- Commit
62bbac4f2f0c0eceab665c6ebdc98c992009c957- Content hash
f4753aa281d4ab7529bd8ae0ae9c67dea5d9300da947b081576dceb1a6ed749e- Date tested
- not yet tested
The pin is the whole of this skill's identity here. It resolves at https://github.com/samber/cc-skills/tree/62bbac4f2f0c0eceab665c6ebdc98c992009c957/skills/copywriting-tone-of-voice-creator, and the content hash is a sha256 over exactly the text the model was given with the skill loaded. Nothing else about the skill appears on this site.
S21-samber-tone-of-voice-creator
Loading skills/copywriting-tone-of-voice-creator from samber/cc-skills improves outputs on voice tasks.
- Pass criterion
- With-arm win rate at or above 0.60 across the same 10 briefs, with the interval excluding a rate of 0.50.
- Scale
- unit. Same scorer, same instrument and same scale as S10-ecc-brand-voice. Added 2026-08-31 by the source expansion under R23; adding a row to an instrument does not touch the instrument.
- Task pairs planned per model
- 10
- Notes
- Shares the voice task set with every other claim in this class, so all of its rows are measured on identical items against one answer key. The slot this claim scores was pinned by the 2026-08-31 source expansion under R23 and NO MODEL HAS BEEN RUN AGAINST IT. The claim exists so the class holds one claim per pinned slot, which is what makes the row comparable the day it is run.
Injected context tokens
| Arm | Context characters | Injected context tokens |
|---|---|---|
| without | 0 | 0 |
| with | 88,335 | 20,736 |
The character count is exact: it is the length of the text the with arm is given, and the content hash above is a sha256 over that same text. The token figure is an estimate at 4.26 characters per token, the ratio the phase 1 run measured over 3,120 calls, and it is labelled an estimate until a run reports its own token counts. The without arm is given the identical prompt and nothing else, so its zero is a measurement rather than a missing value.
Not yet measured
No run records exist for this skill. The instrument is built and committed, the task sets are frozen, and the projected cost of the run is published, but no model has been called. There is therefore no verdict, no effect, no interval and no cost per task on this page, and none is shown.
- S21-samber-tone-of-voice-creator: 10 task pairs per model, as planned. Blind comparison of the same task, run twice, with ties counting half.
The committed instrument is at harness/claims/claims-skills.json, harness/claims/tasksets-skills.json and harness/claims/skill-cohort.json. The projection for the run that has not happened is at harness/results/dry-run-projection-skills.md. How a claim gets tested.