Kill-first validation
Critical Assumption First — Name It, Falsify It, Then Decide
Most founders collect a list of assumptions and then test the easiest one. That feels like progress. It is usually a delay. The useful move is the opposite: name the single claim that would make the rest of the plan irrelevant if it were false, write the decision lines before any evidence arrives, and spend the week trying to break that claim — not to decorate it.
Last updated: September 26, 2026
Direct answer
What is a critical assumption?
A critical assumption is the one claim the rest of the plan depends on. If it is false, the other risks do not matter yet. Eric Ries, in The Lean Startup (Crown Business, 2011), calls the handful of beliefs a new venture cannot survive without “leap-of-faith assumptions,” and treats validated learning as progress measured by tests of those beliefs — not by features shipped. Steve Blank’s Customer Development, published in The Four Steps to the Epiphany (2005), puts the same idea in operating form: get out of the building and treat customer, problem, and solution statements as hypotheses. This page is the field method: one sentence, three pre-written verdicts, one cheap falsifier, one decision.
Key takeaways
What to remember
- Name one assumption, not twelve. If this claim is false, the rest of the list is homework for a different idea.
- Write Continue, Refine, and Stop lines before anyone talks to you. A bar invented on Sunday is a story, not a result.
- Falsify cheaply. Interviews about a recent Tuesday, a landing page with one action, or a week of manual delivery beat a six-month build that “will prove it.”
- Compliments are not evidence. Rob Fitzpatrick’s The Mom Test (2013) is blunt on this: asking “would you buy this?” produces politeness, not a decision.
- A score is a flashlight. It can point at the weakest dimension. It cannot name the sentence, run the week, or decide for you.
Why it matters
Why testing the easy assumption wastes the month
I have watched founders spend six weeks “validating” a claim they already believed — usually “people have this problem in the abstract” — because conversations about that claim are pleasant. Meanwhile the plan still hangs on something they have not touched: that a named buyer will pay the planned price, switch a weekly workflow, or be reachable in a channel the founder can actually work. Ries’s leap-of-faith framing is useful here because it is rude. If the business dies when one belief is false, that belief is the work. Everything else is a side quest until you have a result on that one.
Vocabulary
Four terms this page uses the same way as the rest of Yibud
These match the glossary. If a sentence on this page uses one of these words, it means this — not “a risk we should keep an eye on.”
Assumption
A claim about the world the plan depends on that has not been tested. “Busy professionals want healthier dinners” is an opinion until someone who recently skipped dinner can walk through what they actually did last Tuesday.
Critical assumption
The single assumption that, if false, would invalidate the rest of the plan. Finding and testing it is the highest-leverage work in early-stage validation. A list of “also important” claims is not a substitute.
Falsify
Design a test that could prove the claim wrong. A test that can only produce a yes is a pitch. Validated learning, in Ries’s sense, requires an experiment whose failure you would accept.
Continue / Refine / Stop
Three pre-written verdicts. Continue: the claim survived this week’s bar; move to the next expensive claim. Refine: the action happened for a different person, job, or price; rewrite the sentence and re-run the same bar. Stop: the qualified people did not do the action; do not “build it anyway and see.”
The loop
The Kill-First Loop: name → cheap falsify → decide
Three moves. Do them in this order. Do not skip the writing. The loop is original to this page as an operating pack; the underlying discipline is Ries (leap of faith + validated learning), Blank (hypotheses you leave the building to test), and Fitzpatrick (questions that can produce a no).
- 1
Name the kill claim in one sentence
Finish the template below before you book a call or publish a page. If you cannot finish it, you do not have a critical assumption yet — you have a slogan. The sentence must name a person, a recent pain, an observable action, a timebox, and a reason they already spend time or money.
- 2
Pick the cheapest test that could prove it false
If you still do not know whether the pain is real in their week, start with conversations about the last incident — not a pitch. If strangers already describe the pain and the expensive unknown is reach, a one-page smoke test is enough. If they reach and the expensive unknown is whether they will use or pay for a delivered outcome, do the job by hand for a handful of people. Do not stand up a factory to answer a sentence.
- 3
Decide against the lines you wrote on Monday
Compare the week to Continue / Refine / Stop — not to a vibe, not to a compliment, not to “we learned a lot.” If the result is Refine, rewrite the sentence. If it is Stop, stop this version. If it is Continue, the next week belongs to the next kill claim, not to a feature list.
Copy this
The assumption sentence and the three decision lines
Write all four on the same page, the same day, before recruitment. Fill the brackets. Do not leave a bracket as “users.”
Assumption sentence
People who [specific role] and recently [specific pain] will [observable action] within [timebox] because they already [spend time / spend money / run a workaround] on this job.
If you cannot name the recent pain or the existing spend, the sentence is still a wish. Observable actions: used the output, requested a second cycle, paid, booked a follow-up, introduced a colleague. “Said it was interesting” is not an action.
Continue if
At least [N] of [M] qualified people do [action] by [date], and at least [one] of them [repeats / pays / introduces].
Refine if
The action happens, but for a different [role / job / price / channel] than the sentence named. Rewrite the sentence. Re-run the same bar.
Stop if
Fewer than [N] of [M] qualified people do [action] by [date]. Do not start a build to “save” the week.
Which one
How to pick THE assumption (not a list of twelve)
Founders often produce a wall of sticky notes and then test the note they already like. Use this filter instead. The first claim that fails a later question is usually the one.
- Would the rest of the plan still matter if this were false? If yes, it is not critical yet — it is a later risk.
- Is the claim about a person you can find this week? “The market is huge” cannot be falsified in seven days. “Freelance designers who chased an unpaid invoice last month will pay $29 before I automate the chase” can.
- Can a no show up? If every possible outcome still lets you keep building, you wrote a slogan.
- Are you avoiding it because it is uncomfortable? That discomfort is often the tell. The easy assumption — “people have this problem” — is the one founders already believe.
This week
Seven-day falsification checklist
Copy the seven lines. The week is a timebox, not a launch. You do not need a statistically valid sample. You need a handful of qualified people and a bar you wrote before they arrived.
- 1
Monday — write the sentence and the three verdicts
One page. Role, recent pain, action, timebox, existing spend. Continue / Refine / Stop filled in with numbers and a calendar date. If a bracket is still “users,” you are not done.
- 2
Tuesday — choose the cheapest falsifier
Conversation about the last incident, one-page smoke test, or manual delivery. Pick one. Write why the other two are the wrong question this week. If you pick all three, you are stalling.
- 3
Wednesday — recruit only qualified people
The role in the sentence, plus a recent incident. Friends who want to be supportive do not count. A stranger who lived the pain last month does. Cap the count at what you can actually serve or interview.
- 4
Thursday — run the first half without pitching
Ask about the last Tuesday, or publish the page, or deliver the first job. Do not explain the vision. Log the action you named — not the compliment.
- 5
Friday — finish the sample you wrote on Monday
Do not expand the ICP because the first two people were “meh.” That is Refine or Stop material, not a reason to hunt a friendlier audience mid-week.
- 6
Saturday — log, do not narrate
For each person: qualified or not, action or not, what you improvised, whether they asked for another cycle or a price. Leave the story for Sunday.
- 7
Sunday — decide against Monday’s lines
One sentence: Continue, Refine, or Stop. If Refine, rewrite the assumption sentence the same day. If Stop, write what you will not build. If Continue, name the next kill claim — do not open a feature backlog.
Worked example
One week you can copy (illustrative)
The week below is illustrative — names, prices, and counts are invented so you can see the setup. It is not a conversion benchmark and not a Yibud report. The method is what you copy, not the numbers.
The sentence, written on Monday
A founder wants a $29/month tool that chases unpaid invoices for freelance designers. The tempting assumption is “I can build the Stripe reminders.” That is a build task. The kill claim is: “Freelance designers who chased an unpaid invoice in the last 30 days will pay $29 this week for a chase I run by hand, because they already spend hours on follow-up emails.” Continue: 3 of 5 qualified designers pay $29 or book a second chase by Sunday. Refine: they want the chase, but only if it is a one-time $15 job, or only if they are studios, not solo designers. Stop: fewer than 2 pay or book a second chase.
- Tuesday: the cheapest falsifier is not an app. It is five conversations about the last unpaid invoice, then an offer to run this week’s chase by hand for $29. No Figma. No “AI collections engine.”
- Wednesday–Friday: five strangers from a freelance Slack, not the founder’s design-school friends. After each call: did they describe a specific invoice, what they already tried, and whether they paid or booked the manual chase.
- Saturday log: two paid $29 and sent a real invoice PDF. One wanted a $15 one-off. Two said “interesting” and did not send the PDF. The compliment column is discarded.
- Sunday: that is Refine, not Continue — the $29 subscription sentence did not survive; a paid one-off chase might. Rewrite the sentence for studios or for a one-time fee. Do not start coding reminders “so the next five have a product.”
Where the score fits
The score is a flashlight, not the verdict
If you already have a Yibud / Startup MRI report, start with the named critical assumption and the weakest dimension — not the overall number. The engine is deterministic: the same inputs produce the same scores. Optional language-model text, when it is used, only polishes prose. It does not invent the number and it does not decide Continue or Stop. A mid-band overall can still hide a kill claim. Read the score-meaning page if you are stuck on “is 60 good?” Then come back here and write the sentence.
How to read a mid-band validation score →Common mistakes
What usually ruins the week
Testing the assumption you already believe
“People have this problem” is comfortable because strangers will often agree in the abstract. The kill claim is usually payment, switching a weekly workflow, or being reachable in a channel you can actually work.
Writing the bar after the evidence arrives
Sunday’s “we got good signal” is not a result. If Continue / Refine / Stop were blank on Monday, you do not have a decision. You have a recap.
Asking “would you buy this?”
Fitzpatrick’s rule exists because politeness is cheap. Ask about the last incident, the last workaround, and the last time they spent money on this job. Then offer a real next step — a second call, a paid manual job, a page with one action.
Treating a list of twelve as progress
A canvas full of assumptions feels rigorous. It is a stall if none of them is the kill claim. Rank by “if false, does the rest still matter?” Keep one.
Using the score as a permission slip
A high overall does not mean build. A low overall does not mean the sentence is false. The score ranks relative risk in the inputs you typed. The week tests the world.
Refusing to Stop because you already told people the idea
Identity is not evidence. If the qualified people did not do the action, the honest next page is a rewritten sentence or a different idea — not a quieter version of the same build.
Sources
Where these ideas come from
- Eric Ries, The Lean Startup (Crown Business, 2011) — Leap-of-faith assumptions and validated learning: progress is a test of the beliefs the venture cannot survive without, not a feature list. Used here as the reason to pick one kill claim first — not as a conversion-rate table.
- Steve Blank, The Four Steps to the Epiphany (2005) / Customer Development — Customer, problem, and solution statements are hypotheses. The work is to leave the building and test them before you scale a product organization. Used here as the operating posture, not as a sample-size rule.
- Rob Fitzpatrick, The Mom Test (2013) — Talk about the customer’s life, not your idea; ask about specifics in the past; talk less than you think. Used here so the week cannot be saved by compliments. Fair-use attribution only — no long quotations.
With Yibud
Use the report to name the claim — then run the week yourself
Yibud is a free, no-signup startup idea validator. Scores come from a deterministic rule engine, not from a language model guessing success. A Startup MRI report names a single critical assumption and a 7-day action plan derived from the highest-risk dimension. That is a starting sentence, not a verdict. This page is the field method after you have that sentence: write the three lines, pick the cheapest falsifier, decide on Sunday.
Analyze my idea →FAQ
Questions founders actually ask
Is the critical assumption the same as the lowest score?
Often, not always. On Yibud the engine usually points at the weakest dimension, but a moderately low dimension can still hold the kill claim — for example distribution looks “fine” in the abstract while the buyer is unreachable in the channel you named. Write the sentence. If the rest of the plan dies when that sentence is false, that is the one.
Can I test two assumptions in the same week?
Only if one test produces a clean no on both, which is rare. Two claims in one week usually means you will explain away whichever one failed. Pick the kill claim. Park the other on a list for the week after Continue.
What if I cannot find five qualified people?
That can be the result. If you cannot reach the person named in the sentence in seven days, the distribution claim may be the kill claim — rewrite the sentence around reach, not around the feature. Do not substitute friends so the week “has data.”
Do I have to charge money?
Not on day one, but an observable action has to be costlier than a compliment. A booked second session, a sent file, a deposit, or a paid manual job are all stronger than “sounds great.” Only take money you can deliver or refund this week.
How is this different from a SWOT list or a risk register?
Those lists enumerate. This loop ranks by lethality and forces a pre-written decision. A register that never produces a Stop line is a diary.
Does a good Yibud score mean I can skip this week?
No. The score is a structured read of the inputs you typed. It does not interview anyone, collect a payment, or watch a workflow. Skip the week only if you already have behavioral evidence against the same sentence — not because the number looked friendly.
What if the week says Refine every time?
Then the sentence is still too wide, or you are changing the bar instead of the claim. Two Refines in a row should narrow the role or the action, not invent a new product. A third Refine with no narrower person is a Stop.
How does this connect to Yibud?
Run the analyzer if you do not yet know which dimension is expensive. The report will not operate the week for you. It will name a critical assumption and a 7-day plan. Bring this page to that plan: write the three verdicts, then decide against them.
Name the claim before you build the product
Yibud scores the weak dimensions with a deterministic engine. You write the sentence, run the cheapest falsifier, and decide Continue / Refine / Stop against a bar you wrote on Monday.
Generate a Yibud report →