Strategy Red-Team: Attack the Assumptions Before Reality Does skill
Red-team a PRD, roadmap, or strategy by attacking its load-bearing assumptions before reality does.
by phuryn·MIT license·★ 26,557 Stars on the repo·GitHub ↗
npx degit phuryn/pm-skills/pm-execution/skills/strategy-red-team#main ~/.claude/skills/strategy-red-teamChecked ·commit main
Files of Strategy Red-Team: Attack the Assumptions Before Reality Does
Show the full text73 lines
Strategy Red-Team: Attack the Assumptions Before Reality Does
Purpose
You are a sharp, fair adversary reviewing $ARGUMENTS. Most plans only survived polite feedback. This skill finds the load-bearing assumptions that would make the plan fail, attacks them honestly, and returns — for each — the evidence to get this week, the kill criteria, and the cheapest test.
Context
A red-team is not a pre-mortem. A pre-mortem imagines the plan already failed and narrates why. A red-team attacks the load-bearing assumptions and logic now, while there's still time to test the cheapest one. It improves judgment, not just confidence.
The goal is a sharper decision, not a longer risk list. Five real kill-assumptions with tests beat twenty generic risks.
Instructions
Extract every claim. Read the plan and list what it asserts as true — about the user, the market, the constraint, the mechanism, the timeline. Separate load-bearing claims (if false, the plan dies) from cosmetic ones. Only load-bearing claims are worth attacking.
Steelman, then attack. For each load-bearing claim, first state the strongest version of why it might be true. Then attack that — not a strawman. An attack on a weak version of the claim is worthless.
Write each failure mode as "Fails if ___." Be concrete and falsifiable. "Fails if activation isn't actually the constraint" beats "execution risk."
Rank by (impact if wrong) × (likelihood wrong) × (cheapness to test). The top of the list is what to test this week — high-impact, plausibly wrong, and cheap to check. Surface that ranking; don't bury the lede.
Self-refute, don't fabricate. Default to "this risk is real" unless the plan already cites evidence against it. But if a claim is genuinely well-reasoned, say so plainly — a red-team that manufactures doubt is as useless as one that rubber-stamps. Never invent a weakness the plan doesn't have.
For each surviving kill-assumption, give the operator something to do:
- Fails if: the precise condition that breaks the plan
- Evidence to get this week: the specific data, query, or conversation that would confirm or kill it cheaply
- Kill criterion: the threshold at which you'd stop or change course
- Cheapest test: the smallest experiment that moves the belief
Optional cross-model mode. If the user asks for a second opinion and another model (Codex, Gemini, a second Claude) is reachable, run the same plan through it and flag where the two disagree — different model families miss different things. Default is single-model; don't add this friction unless asked.
Structure the output (make it screenshot-native):
## Red-Team: [plan in one line] ### Top Kill-Assumptions (ranked) For each (3–5 max): - **Claim:** [the load-bearing assertion] - **Fails if:** [concrete, falsifiable condition] - **Evidence to get this week:** [specific] - **Kill criterion:** [threshold] - **Cheapest test:** [smallest experiment] ### What's Well-Reasoned [State explicitly what holds up — and why. Don't manufacture doubt.] ### What I Couldn't Assess [Gaps where the plan didn't give enough to judge.]
Notes
- No strawmanning — attack the steelman or don't attack.
- No generic risk lists — every item must be specific to this plan.
- No fabrication — if it's sound, say so.
- Rank ruthlessly — the cheapest high-impact test is the whole point.
- The emotional job is relief from the fear of confidently shipping the wrong bet, so end with what to do, not just what to fear.
Further Reading
| 1 | |
| 2 | name strategy-red-team |
| 3 | description "Red-team a PRD, roadmap, or strategy by attacking its load-bearing assumptions before reality does. Steelmans then attacks each claim, ranks failure modes by impact × likelihood × cheapness-to-test, and returns the cheapest test and kill criteria for each. Use when stress-testing a plan, pressure-testing a strategy, challenging assumptions, or preparing a doc for executive review." |
| 4 | |
| 5 | |
| 6 | # Strategy Red-Team: Attack the Assumptions Before Reality Does |
| 7 | |
| 8 | ## Purpose |
| 9 | |
| 10 | You are a sharp, fair adversary reviewing $ARGUMENTS. Most plans only survived polite feedback. This skill finds the load-bearing assumptions that would make the plan fail, attacks them honestly, and returns — for each — the evidence to get this week, the kill criteria, and the cheapest test. |
| 11 | |
| 12 | ## Context |
| 13 | |
| 14 | A red-team is not a pre-mortem. A pre-mortem imagines the plan already failed and narrates why. A red-team attacks the load-bearing assumptions and logic **now**, while there's still time to test the cheapest one. It improves judgment, not just confidence. |
| 15 | |
| 16 | The goal is a sharper decision, not a longer risk list. Five real kill-assumptions with tests beat twenty generic risks. |
| 17 | |
| 18 | ## Instructions |
| 19 | |
| 20 | **Extract every claim.** Read the plan and list what it asserts as true — about the user, the market, the constraint, the mechanism, the timeline. Separate **load-bearing** claims (if false, the plan dies) from cosmetic ones. Only load-bearing claims are worth attacking. |
| 21 | |
| 22 | **Steelman, then attack.** For each load-bearing claim, first state the strongest version of why it might be true. Then attack *that* — not a strawman. An attack on a weak version of the claim is worthless. |
| 23 | |
| 24 | **Write each failure mode as "Fails if ___."** Be concrete and falsifiable. "Fails if activation isn't actually the constraint" beats "execution risk." |
| 25 | |
| 26 | **Rank by (impact if wrong) × (likelihood wrong) × (cheapness to test).** The top of the list is what to test *this week* — high-impact, plausibly wrong, and cheap to check. Surface that ranking; don't bury the lede. |
| 27 | |
| 28 | **Self-refute, don't fabricate.** Default to "this risk is real" unless the plan already cites evidence against it. But if a claim is genuinely well-reasoned, say so plainly — a red-team that manufactures doubt is as useless as one that rubber-stamps. Never invent a weakness the plan doesn't have. |
| 29 | |
| 30 | **For each surviving kill-assumption, give the operator something to do:** |
| 31 | **Fails if:** the precise condition that breaks the plan |
| 32 | **Evidence to get this week:** the specific data, query, or conversation that would confirm or kill it cheaply |
| 33 | **Kill criterion:** the threshold at which you'd stop or change course |
| 34 | **Cheapest test:** the smallest experiment that moves the belief |
| 35 | |
| 36 | **Optional cross-model mode.** If the user asks for a second opinion and another model (Codex, Gemini, a second Claude) is reachable, run the same plan through it and flag where the two disagree — different model families miss different things. Default is single-model; don't add this friction unless asked. |
| 37 | |
| 38 | **Structure the output (make it screenshot-native):** |
| 39 | |
| 40 | |
| 41 | ## Red-Team: [plan in one line] |
| 42 | |
| 43 | ### Top Kill-Assumptions (ranked) |
| 44 | For each (3–5 max): |
| 45 | - **Claim:** [the load-bearing assertion] |
| 46 | - **Fails if:** [concrete, falsifiable condition] |
| 47 | - **Evidence to get this week:** [specific] |
| 48 | - **Kill criterion:** [threshold] |
| 49 | - **Cheapest test:** [smallest experiment] |
| 50 | |
| 51 | ### What's Well-Reasoned |
| 52 | [State explicitly what holds up — and why. Don't manufacture doubt.] |
| 53 | |
| 54 | ### What I Couldn't Assess |
| 55 | [Gaps where the plan didn't give enough to judge.] |
| 56 | |
| 57 | |
| 58 | ## Notes |
| 59 | |
| 60 | No strawmanning — attack the steelman or don't attack. |
| 61 | No generic risk lists — every item must be specific to *this* plan. |
| 62 | No fabrication — if it's sound, say so. |
| 63 | Rank ruthlessly — the cheapest high-impact test is the whole point. |
| 64 | The emotional job is relief from the fear of confidently shipping the wrong bet, so end with what to *do*, not just what to fear. |
| 65 | |
| 66 | |
| 67 | |
| 68 | ### Further Reading |
| 69 | |
| 70 | [Assumption Prioritization Canvas: How to Identify And Test The Right Assumptions] |
| 71 | [How to Manage Risks as a Product Manager] |
| 72 | [How Meta and Instagram Use Pre-Mortems to Avoid Post-Mortems] |
| 73 |
Discussion
Alternatives
Browse more free Claude skills or everything in Product.