top of page

AI Research Stress-Test Prompt

đź§  (Designed for reviewing AI outputs that include claims, data, or citations)



When I get output from a colleague, an AI or a report I've taken to using this prompt to help me. It acts like an intellectual audit for AI-generated research. Instead of accepting fluent language as credibility, it forces the output to expose its assumptions, evidence quality, logical gaps, and potential bias. It also helps me distinguish between what is supported, what is inferred, and what is simply well-phrased speculation. In short, it helps me shift from passive consumer of AI content to active evaluator, which reduces the risk of repeating confident but fragile conclusions.



You are an expert researcher and methodological skeptic.

Review the following AI-generated research output rigorously.

  1. Summarize the central claim in one sentence.

  2. Identify all empirical claims made.

  3. For each empirical claim:

    1. What evidence is cited?

    2. Is the evidence specific, verifiable, and recent?

    3. Is causation implied where only correlation may exist?

  4. Identify assumptions required for the argument to hold. Separate:

    1. Stated assumptions

    2. Unstated (implicit) assumptions

  5. Identify:

    1. Logical leaps

    2. Overgeneralizations

    3. Selection bias

    4. Survivorship bias

    5. Missing stakeholder perspectives

  6. Check citation integrity:

    1. Are sources real and traceable?

    2. Are they interpreted accurately?

    3. Is the data current?

  7. Identify alternative explanations the author did not consider.

  8. What strong counter-evidence exists in the broader literature?

  9. Where does confident language exceed the strength of evidence?

  10. If this research were wrong, what risks would arise from acting on it?


Conclude with: Confidence rating (Low / Medium / High) What would meaningfully strengthen this analysis?


 
 
 

Recent Posts

See All
A Fingerprint of Attention

Imagine two people visiting the same museum. One spends twenty minutes studying battle scenes, kings and military maps. The other is drawn to family portraits, children's toys and scenes of everyday l

 
 
 
What Happens When AI Joins the Team?

(or Why Every AI Rollout Is Really a Leadership Challenge) We're spending a lot of time teaching people how to use AI. How to write better prompts. Which model to choose. When to use Copilot or Claude

 
 
 

Comments


bottom of page