Artifact critiqueAdvancedAI Opportunity & Model Strategy / Roadmapping under model uncertainty / #22
Write the roadmap narrative you would present to a board.
SPARK the board slide that showed one number and hid which fabric it actually applied to
A board narrative isn't a features-and-dates slide with AI written on it. Loomweave Textiles runs a defect-detection system called LoomEye on its production lines. Fenn Adisa is the AI PM writing the narrative for the board's next quarterly review, after the first version of that narrative already went wrong once.
The direct answer
Anchor the whole narrative on one number: the dollar cost of defects that reached customers last year, and how much of it is actually gone. Never a blended accuracy percent standing in for the whole system. Break the roadmap into proven, piloting, and exploring, name which slice of that dollar figure each tier is responsible for, and don't promise a plant-wide date until at least two lines have cleared their bar on their own.
Do this, in order
Open on the cost of the status quo, in dollars, not the technology.Why: a board that starts with the problem's real size judges progress against it, not against a vague sense of "is it working."
Never publish one blended accuracy number across different lines or fabric types.Why: a single percent hides exactly which case is still weak, and that's the case that eventually reaches a customer.
Name three tiers, proven, piloting, exploring, and put every roadmap item in one.Why: a board that knows which tier something's in doesn't confuse a hope with a commitment.
Hold the plant-wide date until two lines clear the bar independently.Why: one line clearing it could be luck. Two is a pattern worth a promise.
Design the narrative so a single miss reads as a stalled sub-item, not a broken headline.Why: if the whole story depended on one line, a miss there sinks the whole narrative instead of one slice of it.
How to answer this, stage by stage
Nobody is scoring whether you can format a slide. They're scoring whether the narrative survives contact with the one item that doesn't go as planned.
Stage 1
Scope it to one real board
Say it like this
"I'll write the actual narrative Fenn Adisa would present to Loomweave Textiles' board about LoomEye, its defect-detection system, not a generic template for any AI roadmap slide."
Why this works
Grounds the answer in a real audience with real stakes instead of a hypothetical deck.
Stage 2
Say your structure out loud
Say it like this
"I'll build this with SPARK. Situation, what the board understands today. Payoff, the habit I want them to build. Anchor, the one structural decision. Risk, what happens the day something misses. Keep out, what I won't promise yet."
Why this works
Shows you're designing the narrative, not just describing a topic.
Stage 3
Name the situation, plainly
Say it like this
"Last quarter's narrative led with '94 percent detection accuracy achieved.' The board stopped asking questions after that slide. Nobody realized the number only applied to two of six production lines."
Why this works
Names the real starting point, including the mistake, instead of pretending this is a clean first draft.
Stage 4
Give the anchor, the one decision
Say it like this
"This narrative opens on one number: four hundred and ten thousand dollars in defect cost that reached customers last year. Every roadmap item answers one question: how much of that number does it remove, and how sure are we."
Why this works
This is the direct answer, made specific enough that a board member could actually hold it in their head.
Stage 5
Prove the anchor survives its own risk
Say it like this
"If dark synthetic fabric misses its bar again next quarter, the narrative still reads: a hundred and twenty-six thousand dollars avoided so far, one line stalled, two lines not yet started. Nobody's headline broke. One sub-item didn't move."
Why this works
Shows the anchor was designed to survive exactly the failure it will eventually face.
Stage 6
Name what you're deliberately leaving out
Say it like this
"This narrative doesn't put a percent on the slide at all, and it doesn't put a plant-wide date on it either, not until two lines clear the bar on their own."
Why this works
Shows judgment about what a board doesn't need yet, not just a wish list of everything that could be shown.
Stage 7
Close on the one line
Say it like this
"Build the narrative around one real dollar figure and how much of it is gone, not a features list with a percent bolted on, because the percent is what breaks trust the moment someone finally asks what it actually covers."
Why this works
Restates the direct answer in one breath, exactly what a live follow-up rewards.
Let's learn
Here is what happens when a board narrative leads with a single reassuring number instead of the actual size of the problem it's solving.
Before LoomEye, an inspector checked every bolt of fabric by eye at the end of each production line, roughly forty seconds a bolt, catching most real defects but missing the faint ones, the tension lines and dye bleeds that only show under the right light. Loomweave's defect returns cost the company about four hundred and ten thousand dollars a year. With LoomEye flagging likely defects for a final human check on two lines, that cost started to fall.
This is the whole job LoomEye touches. None of it needs a board slide to explain.
Here's the turn: Fenn's first board narrative led with "94 percent detection accuracy achieved," a real number, measured honestly, on the two lines running well-lit standard cotton. The board heard "94 percent" and stopped asking about the AI system in every meeting after that. Nobody in the room realized the number said nothing about the other four lines.
Defect cost exposure by production line, and how much LoomEye has actually removed
The 94 percent figure only ever described the top bar. The other two-thirds of the problem never appeared on the slide at all.
At its worst, a single confident number on a board slide doesn't just overstate progress. It teaches the board to stop asking the one question that would have caught the gap: which part of the problem does this number actually cover.
The choice I would take back
The first narrative reported one blended accuracy figure across the two lines running the AI system, without ever naming that it excluded the other four lines and their harder fabric types. That made sense when only one line existed and there was nothing to blend. It stopped making sense the moment a second, harder line joined the same slide under the same number.
What I would leave alone: the three-tier structure isn't needed for line-level scheduling decisions Fenn makes with the plant floor. That's an operational detail the board doesn't need broken out; it only needs to see the dollar-cost view.
The lesson: a board narrative isn't dishonest because it hides bad news. It's dishonest when one true number gets asked to stand in for a claim it was never measured against.
Now here is the same thing as a story
The short version above is what you'd say defending this quarter's narrative to Fenn's own manager. Read this one for how a new face in the room asked the question everyone else had stopped asking.
Elin Marchetti has inspected fabric at Loomweave for six years, on a shared tablet three shifts pass between them at the end of each line. She can spot a weave flaw from four feet away, in a light that would make anyone else miss it entirely.
When LoomEye launched on her line, standard cotton, well-lit, easy to photograph clearly, it caught things she sometimes missed at the end of a long shift. She trusted its flags within a month. The board, one floor up, trusted the 94 percent even faster.
Only the line in the upper right ever earned the number that got quoted for all of them.
For two quarters, every board meeting opened the same way: "94 percent, on track." Nobody asked which lines that covered. Nobody asked because the number sounded finished, and a finished number doesn't invite questions.
Knowledge spark: what's a blended accuracy number?
One percentage built by averaging results across very different conditions. It can be completely true and still hide that the easy cases are carrying the whole score while the hard cases are barely working at all.
Then a newly appointed board member, in her first meeting, looked at the slide and asked, simply: "Does that 94 percent apply to every line?" Nobody in the room had a clean answer. Fenn didn't either, not right away, and had to admit afterward that the number had only ever been measured on the two easiest lines.
The number was never a lie. It just got asked to carry a claim it had never actually been tested against.
The next narrative Fenn wrote opened differently: four hundred and ten thousand dollars in exposure, a hundred and twenty-six thousand of it gone, and two lines, dark synthetic and novelty print, named plainly as not there yet. Elin's line kept doing exactly what it had been doing. The narrative just finally said, out loud, what it couldn't say yet about the other four.
The first narrative only had one branch. Every line got treated as if it had already reached the leftmost one.
SPARK, the narrative built to survive its own bad newsNot a features list dressed up for a board. SPARK is what keeps one missed line from sinking the whole story.
S
Situation. What the board understands today, honestly.
The board heard "94 percent" last quarter and stopped probing, without realizing it only described two of six lines.
Naming the real starting point, mistake included, is what makes the rest of the narrative credible.
P
Payoff. The habit you want the board to build.
Judge progress by dollars of defect cost actually removed, not by a percentage that could be measured on anything.
The habit, not the slide design, is the actual thing being shipped here.
A
Anchor. The one structural decision.
Open on the four hundred ten thousand dollar exposure number, and tie every roadmap item to how much of it that item removes.
This is the hardest step, and the one the whole narrative actually turns on.
R
Risk. What happens the day something misses.
If dark synthetic misses its bar again, the narrative still reads as progress on a large number, with one stalled sub-item, not a broken headline.
Proves the anchor was built to survive exactly the failure it will eventually meet.
K
Keep out. What this narrative won't promise yet.
No blended percent on the slide, no plant-wide date before two lines clear the bar on their own.
Naming what's deliberately absent is what separates judgment from a wish list.
The recap, one line per letter: situation is naming honestly that the board stopped asking questions after one confident number, payoff is teaching the board to track dollars removed instead of a percent, anchor is opening every narrative on the same real exposure figure, risk is designing the story so one missed line reads as a stalled sub-item, and keep out is refusing a blended percent or a plant-wide date before the evidence exists.
And if you want to be sure it really works, try it somewhere elseSame five letters, a regional pharmacy chain's prescription-review tool instead of a textile mill. A different old decision breaks the second narrative.
Harrow Vale Pharmacy Group runs a tool that flags likely prescription interactions for a pharmacist's final check. Its first board narrative reported "97 percent flag accuracy," a real figure measured against the chain's busiest, best-documented stores. Mapped onto SPARK: situation is the board treating that number as chain-wide fact. Payoff is getting the board to track dollars of avoided interaction incidents, not a flag-accuracy percent. Anchor is opening the narrative on the actual cost of interaction incidents last year, with each store cluster's contribution named separately. Risk is a rural cluster with thinner records missing its bar, which reads as one named cluster still exploring, not a broken chain-wide promise. Keep out is refusing to publish any store-level rollout date until two clusters independently clear the bar. The old decision here isn't a hidden number, it's a removed affordance: the first narrative dropped the per-cluster breakdown entirely to keep the slide to one page, on the reasoning that a single summary number was cleaner for a busy board to absorb.
The same four parts anchor a textile board slide and a pharmacy board slide alike.
Swap the trigger and it still runs.
Speed: an interviewer caps you at sixty seconds. Say "open on the real dollar cost, tie every item to how much of it it removes, and never let one blended percent stand in for the whole system," and stop.
Cost: there's no time to build a full per-line cost model before the next board meeting. Say so honestly, and start with the exposure number for the two hardest lines alone, since that's where the risk of a false claim is highest.
The model gets better, for real: if dark synthetic genuinely clears its bar next quarter, that's real news to report as its own line moving from piloting to proven, not folded quietly back into one blended number.
Where people run it wrong.
They let one confident percent stand in for a system with very different performance across cases.
They promise a plant-wide date the moment one line looks good, instead of waiting for a second line to confirm the pattern.
They cut the per-line breakdown to keep the slide short, and lose the one thing that would have caught the gap early.
How to use it live. The moment someone asks you to write a board narrative, ask yourself: if I only had one number to defend under hard questions, which one would actually hold up. Build the whole page around that one.
Every point on this line is honest about what it doesn't yet know, which is exactly why none of them can break trust the way the first slide did.
Flashcards (tap any card to flip it)
1 · THE FRAMEWORK
What framework fits "design the artifact you'd present" questions?
Tap to flip
ANSWER
SPARK: situation, payoff, anchor, risk, keep out. It runs forward, building a design decision that has to survive its own risk.
2 · THE PERSON
Who is this answer about?
Tap to flip
ANSWER
Elin Marchetti, a six-year fabric inspector at Loomweave Textiles who trusted LoomEye's flags on her own well-lit line within a month.
3 · THE HABIT
What did the board stop doing after the first "94 percent" slide?
Tap to flip
ANSWER
They stopped asking which lines or fabric types the number actually covered, and stopped probing the AI system in meetings at all.
4 · THE ANCHOR
What's the one structural decision the whole narrative hangs on?
Tap to flip
ANSWER
Opening on the real dollar cost of defects reaching customers, and tying every roadmap item to how much of that figure it removes.
5 · THE OLD DECISION
What decision would you take back?
Tap to flip
ANSWER
Reporting one blended accuracy number across all lines without saying it only covered two of six, a call that made sense with one line and stopped making sense with a second, harder one.
6 · THE NUMBER
Fill in the blank: of the 410 thousand dollars in yearly defect cost, ___ thousand has actually been avoided so far.
Tap to flip
ANSWER
126 thousand, all of it from the two standard-cotton lines that had actually cleared their bar.
7 · THE REPLAY
Same new board member, same sharp question, but the dollar-anchored narrative is already in place. What changes?
Tap to flip
ANSWER
The slide already names which lines are proven, piloting, or exploring, so the question has a clean answer on the spot, and nobody discovers a hidden gap two quarters late.
8 · CROSS PRODUCT TRANSFER
Section 4 answers this same question again for a different product. Which product, and what old decision gets taken back?
Tap to flip
ANSWER
Harrow Vale Pharmacy Group's prescription-interaction tool. The reversal is a removed affordance: the per-cluster breakdown got cut to keep the slide to one page.
Check yourself Score: 0 / 0
Short answer, name the reversal
1. What old decision does this answer take back, and why did it make sense when it was first made?
Show hint
Look at "the choice I would take back."
Show answer
Model answer: Reporting one blended accuracy number across all lines. It made sense when only one line existed and there was nothing yet to blend it with.
Multiple choice
2. According to the anchor step, what should every item in the board narrative be tied to?
A. The engineering team's own confidence level.
B. How much of the real dollar-cost exposure that item actually removes.
C. How many lines the feature has been deployed to.
D. How the feature compares to a competitor's tool.
Show hint
Look at the anchor step in the SPARK recap.
Show answer
B. The whole narrative is built around one exposure number, and every roadmap item answers how much of it that item removes.
True or false
3. True or false: this answer argues the board should never see a percentage figure again, under any circumstances.
True
False
Show hint
Look at "keep out" and what it specifically refuses.
Show answer
False. It refuses one blended percent standing in for the whole system, not every number; a per-line figure tied to its own tier is fine.
Fill in the blank
4. Fill in the blank: standard cotton lines account for ___ thousand dollars of the 410 thousand total defect-cost exposure.
Show hint
Look at the stacked bar chart.
Show answer
180 thousand. Of that, 126 thousand has already been avoided.
Short answer, apply it yourself
5. Think of a status update, dashboard, or report you've seen that led with one confident summary number. What question would have exposed what that number didn't cover?
Show hint
Think about an average, satisfaction score, or uptime percentage that might hide a much worse result for one segment.
Show answer
Model answer: "Does this average include every region, or just the ones already performing well?" is usually the question that exposes a blended number's blind spot.
Short answer, where it wouldn't matter
6. Name something in Fenn's planning that genuinely doesn't need the proven-piloting-exploring breakdown on the board slide.
Show hint
Look at "what I would leave alone."
Show answer
Model answer: Day-to-day line scheduling decisions on the plant floor. Those are operational details the board doesn't need broken out; it only needs the dollar-cost view.
Before you close the answer
Why this works
Tests whether you can design a communication artifact that survives its own bad news, instead of one built only to look good in the meeting where it's first presented.
Follow-up traps
"Won't a dollar figure without any accuracy percent look less rigorous to a technical board member?" Response: the per-line detection bar still exists and gets shared on request; it's just not the headline, because the headline needs to survive a board member who only reads one line of the slide.
"What if the board specifically asks for one overall number?" Response: give them the portfolio total, dollars avoided against total exposure, which is one number, but keep the per-tier breakdown visible right underneath it, not folded away.
If pressed
The bar that followed set dark synthetic's detection threshold at 75 percent, evaluated on a held-out batch of prior returns from that line specifically, not on any data the standard-cotton lines had ever touched.
From U2xAI Academy
From answering questions to owning outcomes.
A live workshop where you ship a working AI agent, defend a launch decision, and walk away with a portfolio recruiters can't wave off, not just more questions to study.