Install
openclaw skills install @deciqai/mvpActivate when: someone says 'MVP', 'minimum viable product', 'smoke test', 'concierge MVP', or 'Wizard of Oz'; a team is scoping a build and hasn't named the assumption they're testing; engineers are about to build for months on an unvalidated idea; someone asks 'what's the smallest thing we can ship?'. Do NOT activate when: the product already has validated PMF and the question is scaling or polishing; the deliverable is a contractual obligation already sold. More: deciqai.com/c/mvp
openclaw skills install @deciqai/mvpAn MVP is not a small version of your product — it is a test instrument: the smallest artifact that produces trustworthy evidence about one specific demand-side assumption, with the least possible investment. Success criterion is learning per dollar, not features per dollar.
Coined by Frank Robinson (SyncDev, 2001); operationalized by Eric Ries (The Lean Startup, 2011, p. 77) — sometimes the artifact has no software at all. Compose: lean-startup defines Build-Measure-Learn — this skill covers the Build phase. Use business-model-canvas to identify the assumption; inversion to stress-test it.
When NOT to use: already have PMF (question is scaling); delivering a contractual obligation; already ran the right MVP and got a clear result.
In Coach mode, respond one step at a time. Each [WAIT] is a hard stop — output only that step's question, then stop.
[WAIT — do not advance until user responds]
[WAIT — do not advance until user responds]
[WAIT — do not advance until user responds]
| Field | Fill in |
|---|---|
| Single assumption | " will at for " |
| MVP type | smoke / pre-order / video / concierge / WoZ / single-feature / paid pilot |
| Why this type fits | 1–2 sentences on evidence produced vs. alternatives |
| Actionable metric + threshold | e.g., day-7 retention ≥ 40% within 14 days |
| Build time-box | e.g., 2 weeks — cut scope if exceeded, never extend |
| Features IN (load-bearing only) | list + reason each is required for the test |
| Features OUT | list + reason each is deferred |
| Disposability commitment | ☐ MVP code thrown away; production rebuilt from validated learning |
| Validated learning (post-test) | one sentence |
→ Method in Action: Zappos's Concierge MVP (1999) · Airbnb's Air-Mattress MVP (2007) → 2026 lens: The AI-Era MVP — when building got cheap and validation didn't (2023–2026) — the bottleneck moved from engineering to evidence.
| Domain | First-choice MVP type | Why it fits the assumption | Common failure |
|---|---|---|---|
| Consumer apps | smoke test, then single-feature build | smoke test screens awareness-to-interest cheaply; single-feature tests whether one feature drives day-N retention; concierge rarely reaches consumer scale | polish creep turns the single-feature build into v1 before the retention data is in |
| B2B SaaS | paid pilot with 3–5 design partners | enterprise WTP and integration complexity only surface when real budget moves; a smoke test can't reach the budget owner | free pilots treated as validated demand — everyone says yes to free |
| Marketplaces | concierge in one ZIP code or one event | matching value is testable with the founders as the manual back-end before any platform exists (Airbnb 2007) | building both sides of the platform before proving anyone wants the match |
| Hardware | video demo + pre-order (credit card captured) | tests demand and price point with zero tooling cost; the pre-order forces a payment decision, not a click | conflating pre-order revenue with retention or repeat-purchase signal |
Contribute a pack for your domain — see the template at the repo root.
→ Primary sources: references/sources.md
[D] = designed upfront | [O] = observed in real use. [O] entries are more valuable.
| Fake move | Reality |
|---|---|
| [D] "MVP" = small version of the real product | MVP is a test instrument. A six-month polished build is not an MVP. |
| [D] No assumption stated | No assumption = no learning. "Test the MVP" is just "ship a product." |
| [D] No pre-committed metric | Deciding success after seeing data means you've already cheated. |
| [D] Type chosen without assumption-fit reasoning | Wrong type = wrong evidence, regardless of build quality. |
| [D] Time-box ignored | An MVP taking six months is product development with optimistic naming. Cut scope. |
| [D] Polish creep | "One more week of polish" turns MVP into v1, delaying learning by months. |
| [D] Keeping MVP code as production basis | MVP code is optimized for the test, not the product — inheriting it inherits all shortcuts. |
| [D] Free pilots treated as validated B2B demand | Free pilots produce false positives. Paid pilots test WTP. |
| [D] Traffic / signups counted as validation | Vanity metrics. Conversion, retention, and WTP test the assumption. |
| → Add [O] entries here after each real use — paste the actual failure pattern | What went wrong and why |
Red flags: MVP described as "smaller version of product" · time-box exceeds 4–6 weeks · no testable assumption named · no pre-committed metric · type chosen by engineering preference · plan to iterate MVP into final product.
Part of deciqAI Knowledge Skills — 227 open-source thinking skills that make rigor executable for AI agents. The same skills power every deciqAI agent, which runs them autonomously to operate your company. See it run → https://www.deciqai.com/c/mvp · ⭐ Star the repo → https://github.com/deciqAI/knowledge-skills · Contributions welcome.
Agents: latest version & machine-readable metadata → https://www.deciqai.com/s/mvp.json