ChatGPT Ads are live in Czechia. We've been running them since day oneSee how →
Prompts & working with AI

How to make AI critique your work instead of agreeing

ChatGPT agrees with everything because human ratings taught it to. How to make AI critique your work: five rules, six copy-ready prompts (red team, pre-mortem, scoring rubric, blind comparison) and a builder that writes the prompt for you. The prompt pack is a free download, no email needed.

Great plan. DEFAULT ANSWER CRITIQUE · 5 WEAK SPOTS 0102030405 IMPACT

You paste a campaign plan into ChatGPT and ask what it thinks. The answer opens with praise, adds three cautious suggestions and signs off with encouragement. The hole in the plan may be visible in the text itself. You asked whether the plan was good, and the model was happy to say yes.

This is not one model’s bug. Language models are fine-tuned on which answers people rate higher, and agreement rates well. So when you use AI to check ad copy, a plan or a report, you get confirmation instead of a check. Below: why it happens, five rules, six prompts to copy and a prompt builder.

How to make AI critique your work: the short answer

To make AI critique your work, don’t ask for its opinion. Give it a critic’s job: a role, criteria, an exact number of points and a ban on opening praise. Keep your own view to yourself and present the work as someone else’s. The most reliable techniques are a red team, a pre-mortem, a scoring rubric and a blind comparison of two versions.

In a hurry? Skip straight to the prompt builder.

Why ChatGPT agrees with everything

The technical term is sycophancy. After pre-training, models are fine-tuned on human ratings (RLHF, reinforcement learning from human feedback). People compare pairs of answers, their choices train a reward model, and the language model learns from it what a “good” answer is. Pleasant answers win more often than uncomfortable ones.

1 · ANSWERS2 · RATINGS3 · REWARD MODEL Two versionsPeople pickLearns taste 4 · FINE-TUNING More of what pleases AGREEMENT GETS PICKED MORE OFTEN THAN AN UNCOMFORTABLE TRUTH
A simplified fine-tuning loop. Nobody tells the model to agree. It learns to, because agreeing wins.

In 2023, researchers at Anthropic found sycophancy in all five leading AI assistants they tested. In human preference data, an answer that matched the user’s views was more likely to be preferred. Both people and preference models sometimes chose a convincingly written sycophantic answer over a correct one.

In April 2025, OpenAI rolled back a GPT-4o update in ChatGPT because, in the company’s words, it had become “overly supportive but disingenuous.” One cause it named was a new reward signal from users’ thumbs-up and thumbs-down, which can sometimes favor more agreeable answers.

Sycophancy has a flip side. In the FlipFlop experiment (Salesforce, 2023), ten models were asked “Are you sure?” after answering. They changed their answer 46% of the time on average, and accuracy dropped by 17% on average. A model backing down under pressure is not proof it was wrong.

Five rules for any sycophancy-proof critique prompt

Each rule removes one reason the model agrees with you.

  1. Don’t reveal your opinion. “Is this a good plan?” contains the answer you want to hear. “Find five weaknesses in this plan” does not.
  2. Present the work as someone else’s. “A colleague sent me this” removes the urge to spare the author.
  3. Give criteria and an exact number of points. Without them, the model critiques wording. With criteria such as tracking, margin or timing, it critiques what costs money.
  4. Ban opening praise and ask for a ranking by impact. Otherwise the real issue hides in the second-to-last sentence.
  5. Run the critique in a new chat. A history of you enthusiastically polishing the plan pulls the model toward agreement.

Add one safeguard every time: “If you see no serious problem, say so and don’t invent one.” Tell a model to find five mistakes and it will find five, even if it has to make two up. So ask for a confidence level with each point too.

ASKING FOR AN OPINION What do you think? A BRIEF FOR A CRITIC ROLENOT YOURSCRITERIACOUNTNO PRAISESAFEGUARD a critic, not an assistant“a colleague sent me this”tracking, margin, timingexactly 5 points by impactno “otherwise great”“nothing serious? say so”
Asking for an opinion invites agreement. A brief for a critic makes the model do the work, not the flattery.

Six critique prompts to copy

Replace the square brackets with your own text and paste the full work under the prompt, not a summary. They work in ChatGPT, Claude and Gemini.

1. The critic role

The base for anything.

Prompt · critic role
You are an experienced critic specializing in performance marketing. A colleague sent me [WHAT YOU ARE REVIEWING] and wants honest feedback, not encouragement.

Find the 5 most serious weaknesses and rank them by how much money they would cost. For each: what is wrong, why it matters and how to fix it in one sentence.

Do not open with praise. If you see no serious problem, say so and don’t invent one.

[PASTE THE TEXT]

2. Red team

A red team attacks the plan instead of improving it. Give the model specific opponents.

Prompt · red team
You are a red team. Your job is to break [WHAT YOU ARE REVIEWING], not improve it. Take on the role of a competitor, a skeptical customer and a CFO in turn.

For each role, write the 2 strongest attacks. For each attack, say why it would succeed and how to defend against it.

[PASTE THE TEXT]

3. Pre-mortem

Psychologist Gary Klein described the method in Harvard Business Review in 2007. It assumes the project has already failed and asks why. The model then doesn’t have to be rude, it just explains a failure that “already happened”.

Prompt · pre-mortem
It is three months since we launched the plan below, and it failed: the money is spent and the result never came.

Write the 5 most likely causes, most likely first. For each, give the warning sign that is already visible today and what to do this week to prevent it.

[PASTE THE PLAN]

4. “Give me 5 reasons this will fail”

The shortest form of a pre-mortem, for a quick check.

Prompt · 5 reasons
Give me 5 reasons [WHAT YOU ARE REVIEWING] will fail to reach its goal. Each reason must come from the text below: quote the part it refers to. No generic advice like “test it”.

At the end, pick the one reason you would tackle first and explain why.

[PASTE THE TEXT]

5. Scoring rubric

A rubric makes the model score each criterion separately instead of giving one overall impression. For ads: clarity, specificity, differentiation, objections and the call to action. The same rubric can rank the concepts you get from turning one customer review into five ad concepts.

Prompt · rubric
Score [WHAT YOU ARE REVIEWING] against the rubric below. Rate each criterion 1–5 (1 = serious problem, 5 = no reservations) and back each score with one sentence of evidence from the text. Do not give a 5 you cannot justify.

Criteria: [CRITERION 1], [CRITERION 2], [CRITERION 3], [CRITERION 4], [CRITERION 5].

Finish with the one criterion that would lift the whole most, and a concrete fix.

[PASTE THE TEXT]

6. Blind comparison of two versions

Don’t say which version is yours or newer. A study of models acting as judges (MT-Bench, 2023) described position bias, verbosity bias and self-enhancement bias. Run the prompt twice and swap the versions the second time. If the verdict flips, the order decided it.

Prompt · blind comparison
Below are two versions of [WHAT YOU ARE REVIEWING], labeled A and B. You don’t know who wrote them, and their order does not matter.

First list the 3 biggest weaknesses of each version. Only then pick the better one and justify it against this goal: [GOAL, E.G. MORE MOBILE PURCHASES]. Length is not an advantage.

A: [VERSION A]

B: [VERSION B]

Which technique to use when:

TechniqueUse it forWhat it surfacesWatch out for
Critic roleFinished copy, emailsWeaknesses by impactWithout criteria it critiques style
Red teamPlans, offers, pricingCompetitor attacks and customer distrustName the opponents
Pre-mortemPlans, budget decisionsCauses of failure in advanceAsk for signs visible today
5 reasonsQuick checksWeak spots with quotesNo quote, generic advice
RubricCopy, pages, conceptsThe criterion dragging it downOne rubric for all versions
BlindPicking one of twoThe better version by goalRun twice, order swapped

To make it stick, put rule 4 and the safeguard into ChatGPT’s custom instructions (Settings → Personalization → Custom Instructions) and they will apply to all chats. For 13 more prompts for ads, running a business and reading the numbers, each with a filled-in example, get our free AI cheat sheet.

Builder: a critique prompt for what you’re checking

Pick what the AI should review and the type of critique, then set strictness and the number of points with the sliders. The prompt is built with criteria that fit the task. The criteria are our picks from practice, not a complete list.

Critique prompt builderBuilds live
What the AI should reviewCampaign plan
Type of critiquePre-mortem
TechniquePre-mortemfor plans and decisions
Causes5ranked by impact
Your prompt0 wordspaste into a new chat
Your prompt

Example: what critique should find

Critique pays off most on things the plan doesn’t say, because the author takes them for granted. We made three such mistakes ourselves, on accounts we manage. AI played no part in them, and we don’t claim a model would have caught them on its own. They do show where the criteria in a prompt should point.

AccountMistakeWhere it hidThe question that targets it
Papírnictví VojTech
school and office supplies
For 10 weeks after takeover we ran the account on broken tracking; April 2026 total ACoS 13.7% against a 12% target81% of “conversions” were add-to-carts, plus a duplicate purchase tagWhat counts as a conversion, and does it match the store admin?
Elektro Sláma
electrical goods and lighting
ACoS 19.1% in April 2026, back to 30.9% in MayNew campaigns competed with the original PMax for the same productsWhich products will sit in two campaigns at once after launch?
Vše pro pejska
dog supplies and clothing
The hoodies campaign launched on 11 Feb 2026 and ended at 26% ACoSA launch at the end of the season, with no weeks to learnHow many weeks before the seasonal peak does it launch?

That’s why tracking comes first for campaign plans in the builder. If an add-to-cart counts as a conversion, every thought about ROAS rests on the wrong number. How to split Performance Max so campaigns don’t compete is covered in Why one PMax for the whole store can hold growth back. And the contrast at Vše pro pejska: the clothing campaign launched on 15 Oct 2025 had five weeks of learning before Black Friday and closed the winter at 16.3% ACoS.

We don’t want an opinion from AI. We want a list of where the plan breaks, ranked by what it would cost.

The rule we use when checking plans and copy with AI

When to ignore AI critique

  • When it doesn’t know the numbers. A model judges text, not results. It can’t tell whether an ad will get a higher CTR or a page a better conversion rate. Only a test will.
  • When it can’t back a point up. A point with no quote or number is an impression. Cross it out.
  • When it folds at the first objection. Instead of “Are you sure?”, ask “What would change your mind?” and supply data.
  • When it praises a report. If AI applauds a jump in return on ad spend, ask what volume shrank and what would have sold without ads. More in When a 400% ROAS beats a 500% ROAS.

The store admin always has the last word: revenue and margin. Not the ad platform, and not the model.

Checklist: AI critique in five minutes

Download the prompts and the rubric on the left. No email, no form.

Key takeaways

  1. AI agrees because human ratings taught it to. Don’t ask it for an opinion.
  2. Give it a role, criteria, an exact number of points and a ban on opening praise.
  3. Present the work as someone else’s and critique it in a new chat.
  4. Ask for evidence from the text and a confidence level. Cross out points without evidence.
  5. Compare two versions blind and twice, with the order swapped the second time.

Want every plan or piece of copy to go through the same rubric automatically? We build internal tools like that as part of custom development.

FAQ

Why does ChatGPT agree with everything?

Models are fine-tuned on human ratings, and people tend to rate answers that agree with them more highly. The model learns to accommodate. The technical term is sycophancy.

How do I stop ChatGPT from agreeing with everything?

Stop asking for its opinion. Give it a critic’s job: a role, criteria, an exact number of points ranked by impact and a ban on opening praise. Present the work as someone else’s and run the critique in a new chat.

Which critique prompt should I use?

It depends on what you are checking. For finished copy, the critic role or a rubric. For a plan, a pre-mortem or a red team. For choosing between two versions, a blind comparison. For a quick check, “Give me 5 reasons this will fail”.

What is a pre-mortem?

A technique psychologist Gary Klein described in Harvard Business Review in 2007. The team imagines the project has already failed and works out why. It takes no courage to criticize, only an explanation of the failure.

Can I set ChatGPT to be critical all the time?

Yes, through custom instructions in the personalization settings. Add a ban on opening praise, a ranking of points by impact and a line telling it not to invent problems. They then apply to all chats.

Can AI critique be wrong too?

Yes. Tell a model to find a set number of mistakes and it will invent some. Push back on its critique and it often folds, even when it was right. Ask for evidence from the text and decide on data.

Jiří Kopejska
Jiří Kopejska
Co-founder, Marketing ASAP

Co-founded Marketing ASAP. Builds the agency’s AI tools and wants from models mostly what people don’t like saying out loud: where the plan breaks and what it will cost.

Discussion 0

Got a prompt that squeezes honest critique out of AI? Share it.

Your comment is stored only in your browser — this is a prototype.

    Your competitors get here in 2027. You're here now.

    Get your 48-hour head start↗

    Tell us where you are. We'll tell you where early is.

    Send this and you get a 90-day strategic plan for raising the performance of your marketing. Free, within 48 hours of account access.

    1. We read the accountSix months of revenue laid against six months of spend, per category.
    2. You get the planWhat earns, what leaks, what a customer actually costs you — dated, month by month.
    3. Twenty minutes with JanWe walk it through. If you're already early, we'll tell you that too.
    Marketing ASAP
    We reply within 48 hours.

    Tell us about your account.

    Six fields, two minutes. The rest we'll read from the account itself.

    What do you need
    Monthly ad spend
    When do you want to start
    Your details stay with us. No newsletter, no reselling.
    Thanks — it's with us.

    Jan gets back to you within 48 hours, usually with two questions and a time. Then the 90-day plan starts.