Tips

AI prompting tips, tested on GPT, Claude, and Gemini

I got tired of hunting for prompting tips that actually work, so I built a place to collect and test them.

18 claims tested3 models4,938 records

17 tips, from 18 tested claims: the examples family reports two claims on one page.

How to read this

Each tip rests on a claim tested as a paired comparison on a fixed task set. Across 54 tested pairs the verdicts are 23 In free drift, 16 Stable, 14 Unobservable, 1 Past the horizon.

In free drift: could not be told from no effect; Stable: held; Unobservable: the test could not measure it; Past the horizon: measured worse.

Filter tips

Showing 17 of 17 tips.

  • What the marks mean

    • Holds
    • No measured effect
    • Could not measure
    • Helped a little, under our bar
    • Scored worse

Showing 17 of 17 tips.

Ordering is derived. The lead is the one claim Stable on every model. The co-lead is the most-sampled null, computed by the rule in the methodology. Neither is chosen by hand.