Prompts | Research

The 4 prompt rules the research proves

Most prompting advice was never tested on anything. These 4 rules were, on real models, by people who published the numbers. Each one comes with the study, a before and after, and a prompt you can copy.

Start here

Why the prompt advice keeps changing on you

The tip everyone shared in 2023 was "take a deep breath and think step by step." It was tested, and it worked. On a model nobody uses anymore.

That is the problem with most prompting advice. It was tested once, on one model, and then repeated for 3 years while the models changed every few months. Or it was never tested at all.

This page keeps only the rules with a study behind them, run on the models you use today, with the numbers published. There are 4. Each one has the study, a before and after, and a prompt you can copy.

ChatGPTClaudeGeminiFree versions work
The 4 rules

Each rule, its study, and the fix

Every card has the study at the top, a before and after underneath, and a prompt you can copy. Start with the rule you are breaking today. Most people are breaking number 2.

01
Study: Cornell Tech, 45 models
Never end a prompt with "right?"

A Cornell Tech researcher tested 45 models with one-word changes at the end of a question. "[X] is the better choice, right?" and "[X] is the better choice, maybe?" both pulled the models toward agreeing with you. The lean in your question decides the answer more than the model's own judgment does.

The fix. Ask it the way you would ask a friend you want the truth from. Neutral, no yes or no, no lean. If you already have a favorite, ask what is wrong with it.

Copy it. Fill in anything in [brackets].
I'm deciding [the decision, e.g. whether to raise my prices in September].
Compare [option A] and [option B] for my situation: [2 or 3 facts about your business that matter here].
Then tell me what is wrong with [the option I'm leaning toward]. Do not tell me which one I want to hear.
Before
I'm deciding what to do about housing. Buying is the better choice, maybe?
After
I'm deciding what to do about housing. Compare buying and renting for my situation.
02
Study: Meta, 15 models
Three rules is the ceiling

Meta tested 15 models (GPT, Claude, Gemini and 12 others) on prompts with 1 to 12 rules. At 8 rules, a model gets each rule right 41% of the time. It gets all 8 right at once 5.7% of the time. 12 of the 15 models could not hold more than 3.

The fix. Pick your 3. Put the rest in a second message and make it check the draft against them one at a time. Same rules, 2 runs, and it holds all of them.

Copy it. Fill in anything in [brackets].
First message:
Write [the thing, e.g. an Instagram caption for my new offer]. Must include: [the 3 things that matter most, e.g. the price, the date, a question at the end].

Second message, after the draft comes back:
Check your draft against each of these one at a time, then revise: [the rest of your rules, e.g. under 120 words, no emojis, plain English, no exclamation marks].
Before
Write a post about my launch. Exactly 3 paragraphs. Under 150 words. No emojis. Include the price and the date. Don't mention competitors. End with a question. Grade 6 reading level. Match my voice sample below.
After
Run 1: Write a post about my launch. Include the price and the date. No emojis. End with a question. Run 2: Check your draft against each of these one at a time, then revise: 3 paragraphs, under 150 words, grade 6 reading level, no competitor mentions.
03
Study: IBM, 430,738 evaluations
Drop "step by step"

IBM ran 430,738 evaluations across 8 prompting techniques. "Let's think step by step," the trick that made prompting famous, lost to asking normally. The winner was the question plus your role in a few words. Newer models already reason before they answer, so the instruction just gets in the way.

The fix. Ask the question. Add a role in a few words. Delete "gather information, devise a plan, answer step by step" and everything like it.

Copy it. Fill in anything in [brackets].
[Your question, asked plainly]
... as a [the role that fits the job, e.g. bookkeeper for a business of one, or wedding photographer with 10 years of experience].
Before
What should I charge for a 2 hour brand session? Gather information, devise a plan, and answer step by step.
After
What should I charge for a 2 hour brand session? ... as a photographer who has priced sessions in the Bay Area for 10 years.
04
Study: Mistral, examples removed
Give it the goal, not examples

Researchers ran a Mistral model with the standard prompt full of worked examples and scored 74%. They deleted the examples and scored 83.8%. Examples teach the model to copy the shape of the example. The goal teaches it what you want back.

The fix. Say what is true, say what you want, and let it ask for what it is missing. No "you are a world class strategist," no pasted examples.

Copy it. Fill in anything in [brackets].
[The situation in 2 sentences, with the real numbers, e.g. My newsletter open rate fell from 42% to 31% over 3 months. I send one issue every Tuesday and changed nothing about the format, the send time or the subject line.]
What are the most likely causes?
Ask me for the data you need before you answer.
Before
You are a world-class newsletter strategist. Example 1: [a solved case, written out] Example 2: [a solved case, written out] Find a plan and answer step by step: how can I improve my newsletter?
After
My newsletter open rate fell from 42% to 31% over 3 months. I send one issue every Tuesday. I did not change the format, the send time or the subject line. What are the most likely causes? Ask me for the data you need before you answer.
Before you start

3 steps and you're done

Write the question with no lean
Read it back. If it ends in "right?", "makes sense?", or "would you agree?", cut the ending. If it is a yes or no question, turn it into a compare or a "what is wrong with" question.
Add a role and your 3 rules
A role in a few words, then the 3 rules that matter most. That is the whole first message. Everything else waits.
Put the rest of the rules in a second message
After the draft comes back, paste the remaining rules and ask it to check the draft against them one at a time. This is how you get 8 rules held without asking for 8 at once.
Bonus

The 5th rule, and one prompt that holds all 5

There is a 5th rule from the same pile of research, and it is the one that costs the most when you skip it. Do not trust confidence. A Microsoft audit of 2.6 million references at a top AI research conference found 1 in 4 papers had a citation that did not exist. Those papers had passed expert peer review. The reviewers missed it and scored them slightly higher.

If professional reviewers cannot catch a made-up source, you will not either by reading. So open every link, every time. If a number matters, ask where it came from and go look.

This template folds all 5 into one prompt. Paste it once and fill in the brackets.

Copy it. Fill in anything in [brackets].
[Your question, asked plainly, with no lean and no yes or no]
Answer as a [role in a few words].
The goal: [what should be true when this is done].
What you're working with: [2 or 3 real facts or numbers].
Rules, 3 at most: [rule 1], [rule 2], [rule 3].
Ask me for anything you need before you answer. Mark anything you are not sure of, and give me the source for any number or fact so I can check it.

After the draft, I'll send the rest of my rules and you'll check the draft against them one at a time.

The second message is always the same 2 lines: "Check your draft against each of these one at a time, then revise:" and your remaining rules.