Not a spec document. A control loop.

AI Rollout Plan — guided by the Deming wheel.

Instead of a concept paper written once and never touched again: a PDCA cycle that sharpens with every turn. That happens to line up exactly with Rule 2 of our own PM Rulebook — a plan without a feedback loop drifts. A rollout plan as PDCA is the feedback loop.

The mechanism

Plan · Do · Check · Act — as a living wheel.

Click a quarter. Each phase has a clear task and a clear handoff criterion to the next — no phase is "done" until it delivers that.

PDCA Click a quarter PLAN DO CHECK ACT
The roadmap

Four waves instead of one big bang.

Each wave runs the wheel once, fully, before the next starts. Small scope per cycle — so "Check" stays honest instead of becoming a formality (Rule 3).

Wave 0 — Completed
Foundation
Act
PM RulebookPortal Landing Page

One cycle has already run: rulebook formulated, landing page built, design tested and refined (the Mentalist restyle was "Do→Check→Act" in action — your feedback directly drove the correction).

Standardized as: live reference at thomasmartin-pmp.de; the baseline every further product is measured against.
Plan
already done
  • Positioning as "PM Mentalist" fixed
  • 7 rules distilled from the underlying frameworks
Do
1 session
  • Landing page including rulebook built
  • First color scheme (navy) implemented
Check
immediate
  • Your feedback: "design is cool, but the Mentalist + Futura feel is missing"
  • Concrete, actionable — not a vague "don't like it"
Act
1 session
  • Espresso/gold palette + Futura/Jost implemented
  • About section filled with real data
Learned: fast, honest feedback rounds (Rule 2) reach the goal faster than a perfect first draft.
Wave 1 — Pilot
The proof
Plan → Do
HAL Test (pre-mortem tool)

Plan: select 1–2 real, ongoing projects; fix the success criterion in advance — "does the tool find at least one risk the regular project review would have missed?" Do: run a guided inversion pass with real stakeholders, document the result.

Handoff to Check: only with a documented before/after — not "everyone liked it."
Plan
~1 week
  • Choose 1–2 ongoing projects as the test field (not a fresh-start project)
  • Fix the success criterion in writing before it starts
  • MVP form: guided prompt dialog, no UI needed
Do
~2 weeks
  • Run the inversion session with real stakeholders
  • Play through "how does this fail, guaranteed?", collect paths
  • Fold every path back into a safeguard
Check
~1 week
  • Check against the criterion fixed in advance, not a new one
  • Ask the team: would the regular review have found this risk?
  • Document the result honestly, even if "no"
Act
~1 week
  • If successful: build the UI, add it to the portal as "available"
  • Document the gate decision (see below)
  • Wave 2 starts only after this
Estimated total effort: ~4–5 weeks, low intensity (can run alongside day-to-day business). Owner: you as pilot user + 1 test-project team.
Wave 2 — Quick Wins
Breadth before depth
Plan
Isolation-Tank CheckAnalogy Engine

Starts only once Wave 1 has reached at least the "Act" gate — capacity and credibility are limited, don't spread thin in parallel (that would itself be a case of Rule 1: goldplating the rollout plan).

Waiting on: free capacity after the Wave-1 decision.
Plan
~3–4 days per tool
  • Isolation-tank check: a short questionnaire "where is feedback deprivation occurring?"
  • Analogy engine: collect 3–5 test problems from real practice
Do
~1 week per tool
  • As a light widget in the portal, no login
  • Give the first version to 2–3 test users
Check
~3–4 days
  • Look at usage data together with qualitative feedback
  • Does an analogy actually produce a new solution idea, or just entertainment?
Act
~3–4 days
  • If successful: publish as standard widgets
  • If "nice, but ineffective": deliberately document as a kill
Estimated total effort: ~3–4 weeks for both tools together, once capacity is free.
Wave 3 — Scaling
B2B entry
Parked
Double-Bind Diagnosis

Higher stakes (org-level diagnosis), needs the reference story from Wave 1. Deliberately held back rather than pulled forward — impatience here would be the most expensive mistake in the whole plan.

Start condition: at least one verified success story from Wave 1/2.
Plan
after clearance
  • Find 1 pilot organization/team (own network first)
  • Success criterion: "surfaces an incentive contradiction that was invisible in daily operations"
Do
~2–3 weeks
  • Agents read the pilot organization's policies/OKRs/contracts
  • Generate a map of the paradoxical incentives
Check
~1 week
  • Cross-check with the pilot organization's leadership
  • Only real cases count, no generalities
Act
~1 week
  • If successful: offer as a paid consulting format
  • Use as a reference case for further B2B acquisition
The start condition stays hard: no pulling this forward without a verified success story from Wave 1 — otherwise the portal is selling a promise instead of a result.

The Check gate — three exits, no automatic continuation

At the end of each wave, the result decides, not the calendar. Basic rule: if only the metric "tool was used" is measured instead of the real effect, the gate is worthless (Rule 3 — Lem/Goodhart).

Act — roll out

Clearly demonstrated value (the criterion fixed in advance is met). Becomes a standard tool, the next wave starts.

Re-Plan — sharpen

Approach right, execution not. Same cycle, tighter focus, a different lever — no new concept paper, just a new plan step.

Kill — discontinue

No demonstrated value despite a fair chance. Documented openly as "tested, discarded" — that too is a result, not a failure.

Communication tools

Cadence beats message.

Not talking more — talking more regularly, in smaller, more honest pieces. Exactly the inversion of the isolation tank (Rule 2 — Lem's isolation-tank test from The Conditioned Reflex, the Pirx cycle): whoever stays silent for a long time and then delivers one giant package triggers overreaction. Every message follows the same three-part structure: fact → context → next step.

StakeholderMediumCadencePurposeMode
Pilot project team Short sync (15 min) or voice memo 2×/week, during the Do phase Collect friction & raw data immediately, not only at the end Interactive
Sponsor / decision maker Three-sentence traffic-light update 1×/week, fixed weekday Keep progress and blockers visible — honest color even at yellow/red (Rule 5) Push
Early-access list Short "work in progress" post + status page in the portal Every 2–3 weeks Let expectations grow along with reality — real progress and setbacks, no overselling Push Pull
Gate decision (Act/Re-Plan/Kill) Personal conversation first, then a written summary Once, at the end of each wave Opening a kill decision by email is the comfort trap from Rule 5 in its purest form Interactive

Why this split and not "everyone gets everything": a sponsor pulled into the details twice a week eventually stops listening — exactly the feedback saturation Rule 2 is meant to prevent. A pilot team that only gets the weekly report can't raise its friction in time. The right medium per role isn't a nice-to-have, it's the difference between "positively present" and "annoying."

Stakeholder management

PIS — Power · Interest · Status

Three coordinates form the vector for each stakeholder: influence on rollout success (P), self-interest in the outcome (I), and the actual stature this person carries (S) — on the same scale as P. The dashed arrow on the P/I field isn't a look back, it's the work instruction for the next round of communication. S itself is one-dimensional — a single comparative value, not its own direction — and is made visible in the diagram as a third dimension via circle size: large circle = high stature, small circle = low stature. If a gap opens up between a high position (P) and a small circle (S) — stature that doesn't keep pace with the assigned role — that's a risk in its own right, not a footnote (see the warning below). The precondition for all of this: actually knowing the stakeholders correctly in the first place — see outdated org charts as a time trap.

KEEP SATISFIED MANAGE CLOSELY MONITOR KEEP INFORMED INTEREST → ↑ POWER 0 5 10 0 5 10 PT SP EA B2B

Circle size = S (status), the third vector dimension alongside P/I · small circle = low stature, large circle = high stature

Neutral → Supporter
Pilot project team
P = 3 · I = 8 · S = 4 · Keep Informed
Involve early and often. Collect friction immediately, not only at the end of the wave. The vector points upward: power (P) grows as the team carries results forward internally and becomes a multiplier. S sits close to P — no stature gap, uncritical.
Wait-and-see
Sponsor / decision maker
P = 9 · I = 5 · S = 9 · Keep Satisfied
Three-sentence traffic light, on time, no surprises. The vector points right: interest (I) rises as soon as gate results feed directly into their decision level — facts instead of mood reports. S matches P: high power meets corresponding stature, a sustainable combination.
Interested
Early-access list
P = 2 · I = 7.5 · S = 3 · Keep Informed
Pull offering, no spam. Status page in the portal as a self-service channel. The vector points upward: power (P) grows as soon as the community functions as a citable reference group. S sits close to P — uncritical.
Unknown — Wave 3
Potential B2B customer
P = 7 · I = 3 · S = 6 · Keep Satisfied
No direct contact without a proven reference story from Wave 1. The vector points up-right: interest (I) grows only through concrete case studies — no acquisition conversation without demonstrated impact. S slightly below P — within range, but worth watching.

The vector is what matters. A point on the grid says where a stakeholder stands today. The dashed arrow says where the strategy is meant to steer them — and makes the communication measures from the table above directly checkable: does the measure match the arrow, or are they working against each other?

When the shoes are too big: P ≫ S as its own risk. An example from practice: a construction site manager with high assigned power (in the project meeting the team listened to him) still spoke uncertain German. He didn't fully understand follow-up questions and scheduling discussions in the meeting — out of insecurity he preferred to agree to every proposed date rather than ask or object. High power, low status: exactly this gap put the project behind schedule, not ill will or lack of technical competence. In the PIS diagram this would be a small circle far up the power axis — the position's height promises more stature than the circle actually shows. You have to be able to actually wear the shoes you're given. If they're too big, you need to grow into them fast — otherwise it's not the person with the gap who gets sacrificed for it, but the PM who didn't close it.