CaseIntermediateAI Opportunity & Model Strategy / Opportunity identification for AI / #12

Describe the jobs-to-be-done framing for an AI writing assistant.

SPARKthe job screen built from a laminated style sheet Verbena never stopped needing

Marchvale builds Tonecraft, an AI writing assistant marketing teams use for ad copy, email campaigns, and landing page text. Thessalyn Colfax is the AI PM who owns Tonecraft's roadmap. Silvio Winthrop, Marchvale's CEO, keeps asking for more buttons on the toolbar to match rival copy tools. Verbena Faircastle runs marketing at Amberley Home, a candle and home fragrance brand, and has quietly gone back to drafting Amberley's biggest campaigns by hand.

The direct answer
Stop describing a marketing writing assistant by its buttons. Before a draft starts, make the writer name the one real job it's for, in the brand's own words, not a generic tone setting. Ground that specific job in the brand's actual style guide and its own best past copy, so "sound like us" is something the model can check itself against, not a slider any writing tool could offer.
Do this, in order
  1. Make the writer name the job before any draft starts.Why: skip this and the tool answers "what can it do" instead of "what was it hired for."
  2. Ground the brand-voice job in real material: the style guide and the brand's own top past copy.Why: without grounding, "sound like us" is just a friendlier name for the same generic tone dial.
  3. Keep the job list short and specific, never a catch-all "help me write better."Why: a vague job produces the exact grab bag a feature list produces, just with fewer buttons.
  4. Run a fact check pass before any grounded draft is marked ready.Why: a model that nails the brand's voice can still invent a product claim that isn't true.
  5. Leave the "fast draft, I'll rewrite it" job on simple settings, no grounding step required.Why: forcing every job through the same heavy step slows down work that never needed it.
  6. Refuse to bolt unrelated writing jobs, like blog outlines or internal Slack tone, onto the same workspace.Why: each one added dilutes what the assistant is actually good at for the job that matters most.

How to answer this, stage by stage

Nobody is grading whether you can define "jobs to be done." They are grading whether you can name the one real job, in the customer's own words, and the one design decision that makes it checkable instead of a nicer word for a feature list.

1
Scope it to one brand, one campaign, one person
Say it like this
"Let's ground this. Marchvale builds Tonecraft, a writing assistant marketing teams use for ad copy, emails, and landing pages. I'm the AI PM, Thessalyn Colfax. Verbena Faircastle runs marketing at Amberley Home, a candle brand, and our CEO just asked me why her team barely opens Tonecraft anymore."
Why this works
Naming the actual brand and person stops the answer from staying a general statement about "writing assistants."
2
Say your structure out loud
Say it like this
"I'll run this as SPARK. Situation, how Verbena writes copy today, without any of this. Payoff, the one habit I want the product to build in her. Anchor, the actual screen that makes that habit real. Risk, what breaks the first time the job I've named is wrong. Keep out, what I will not let this job absorb."
Why this works
Two seconds of structure tells the interviewer there's a plan before any single detail can bury it.
3
Reframe what's actually being tested
Say it like this
"This isn't really 'design a marketing writing tool.' It's whether I describe that tool by what it does, autocomplete, grammar, tone, or by the one job someone actually hired it for. Those two descriptions build completely different products."
Why this works
Compresses the whole answer into one breath before it can get buried under detail.
4
Give the one decision
Say it like this
"Here's what I'd do. Before Verbena drafts anything, Tonecraft makes her name the job: match our exact brand voice for this launch, get a fast draft I'll rewrite myself, or turn one great email into new versions. Pick the brand-voice job, and the model gets fed our actual style guide and our best past campaigns, not a generic tone knob."
Why this works
This is the direct answer, said plainly, before the story arrives to earn it.
5
Prove it with a compressed failure, and name the guardrail
Say it like this
"Here's why it matters. For months, Tonecraft's toolbar just grew: autocomplete, grammar check, a tone slider, a hashtag button, nine tools with no job attached to any of them. Verbena used it for a while, then quietly stopped opening it for anything that had to sound like Amberley, and went back to a laminated style sheet pinned above her monitor. Nobody filed a ticket. And even once we grounded the model in her brand's own voice, it once claimed a candle was hand poured when it wasn't, so every grounded draft now has to clear a separate fact check before it's marked ready."
Why this works
Four sentences carry the whole drift, and the fact check line proves the fix survives its own weak point.
6
Close on the one line, and say what you'd leave alone
Say it like this
"So: name the job before the draft starts, ground the brand-voice job in real material, and never make the fast-draft job sit through the same grounding step it doesn't need. That's the whole answer."
Why this works
Leaves the interviewer with the transferable method, not just a story about one candle brand.

Let's learn

Picture two ways to describe the exact same button. One is a list of everything it can do. The other is the one job somebody actually hired it for. Only one of those tells you what to build.

Tonecraft is Marchvale's AI writing assistant. Marketing teams open it to draft ad copy, email campaigns, and landing page text. Before Verbena's team leaned on it, a single big campaign, like Amberley Home's yearly Ember Sale, took her about six hours to draft by hand: pulling from a laminated one page brand voice sheet pinned above her monitor, and a folder of the brand's best performing past emails.

Hand sketched diagram titled Verbena, before Tonecraft ever asked what job it was for. A person icon at the center labeled Verbena drafts alone, with four labeled callouts around it: a laminated brand voice sheet pinned above the monitor, a folder of Amberley's best past emails, about six hours for one campaign, and no screen ever asking what job this is for.
This is the workaround the job screen was built to replace. Not a lack of judgment from Verbena. A lack of anywhere to put it.

Now, with Tonecraft's original nine button toolbar, autocomplete, grammar check, a tone slider, subject line generator, hashtag suggester, hook generator, CTA generator, readability score, and a plagiarism checker, Verbena could get a full first draft in about forty minutes. That felt like a win, for a while.

Hand sketched comparison diagram titled One job, not a grab bag of buttons. Left panel, a funnel icon labeled Nine toolbar buttons, caption reads autocomplete, tone slider, hashtags, hooks, and more, all at once. Right panel, a document icon labeled One named job, caption reads match our brand voice for the Ember Sale launch.
Two ways to describe the same product. Only one of them tells Tonecraft what it's actually for.

Here's the turn. Forty minutes was never the real problem, and neither were a few rough sentences the tone slider produced. The real problem was what came out the other end: copy that read like every other candle brand's AI written email, warm, a little breathless, nothing like the specific voice Amberley had spent five years building. Verbena started rewriting almost every draft from top to bottom, then started skipping the tool for anything that had to carry the brand's name.

We didn't save Verbena forty minutes. We spent it handing her copy that sounded like everyone else's brand instead of hers.
Knowledge spark: what does "grounded" mean here? It means the model isn't guessing at tone from a short instruction. It's reading real material, Amberley's actual style guide and its own best past campaigns, before it writes a word, so a draft can be checked against something specific instead of a vibe.

What it costs at its worst: brand sensitive accounts like Amberley are exactly the customers a generic AI writing tool can't easily win, the ones who care most about sounding like themselves and not like a template. If Tonecraft only ever offers a nicer tone slider, those accounts quietly churn to whichever competitor actually asks what job they're hiring for, and Tonecraft is left competing on price with tools that write generic copy for a fraction of the cost.

What Thessalyn built instead: before any draft starts, Tonecraft makes the writer name the job. Match our exact brand voice for this launch. Get a fast draft I'll rewrite myself. Turn one great email into new variants. Picking the brand voice job pulls in Amberley's real style guide and its twelve best performing past campaigns, and every draft has to clear a fact check pass against Amberley's actual product sheet before it can be marked ready, since a model that nails the brand's voice can still invent a claim that isn't true.

Hand sketched labeled parts diagram titled The anchor, close up: name the job before you draft. A document icon at the center labeled What job is this for, with four labeled callouts: match our exact brand voice for this launch, fast rough draft I will rewrite it, grounded in style guide plus twelve past campaigns, turn one great email into new variants.
The whole fix in one screen. Pick the brand voice job, and Tonecraft pulls in the material that makes that job checkable.
Tonecraft drafts kept without a full rewrite, old toolbar era vs named job era
100% 50% 0 22% Old toolbar, no named job 71% Named job, grounded
Feature grab bagOne named, grounded job
Nine buttons moved almost nothing. Naming one job and grounding it in real brand material moved this number 49 points.
The choice I would take back When Tonecraft first launched, its roadmap answered one question: what should a good writing assistant do. The list came from what other writing tools already offered, autocomplete, grammar check, a tone slider. That made sense when Marchvale had one generic customer profile to design for. It stopped making sense the moment brand sensitive marketing teams like Amberley's became a real, distinct kind of customer with a genuinely different job to get done.

What I would leave alone: for low stakes internal writing, meeting notes, a Slack update, the plain "help me write faster" job is fine exactly as it was, a simple tone slider, no style guide, no fact check. Forcing that work through brand grounding would just slow down something nobody was struggling with.

The lesson: a writing assistant can always do more. That was never the real question. The real question was whether Verbena's actual job, sounding like Amberley without sounding like every other brand's AI written copy, ever had a place to live inside the tool at all.

Now here is the same thing as a story

The short version above is what you'd say out loud in the room. Read this one for what it actually felt like to watch someone quietly stop trusting a tool built to help her.

Verbena Faircastle has run marketing at Amberley Home for five years. She wrote every headline, every subject line, every hero paragraph for the brand's five annual launches herself, and she was good at it. No exclamation points, ever. Never the word "cozy," even for a candle brand, because Amberley's founder banned it in year one for being what every competitor already said. Benefit before feeling, always. She kept all of it on a laminated sheet, and in her head.

Tonecraft arrived two years ago, and for a while it genuinely helped. Late on a Tuesday before a launch, she'd open it, get a rough draft in minutes instead of an hour, and spend the saved time on the parts that mattered most: the one headline the whole email hung on.

Then the habit thinned, in three beats nobody would have called a crisis on its own. First she used Tonecraft for full first drafts. Then only for autocomplete on individual sentences, because the full drafts kept needing a rewrite anyway. Then she stopped opening it at all for anything that had to carry Amberley's name, and kept it only for internal notes, where the voice never mattered.

Hand sketched horizontal timeline titled How Verbena's habit thinned, week by week. Four milestones: Full drafts, caption the good months. Autocomplete only, caption the tone slider drifts generic. Stops opening it, caption for anything brand facing, this milestone emphasized. Notes only, caption the brand job never had a home.
Nobody decided to stop trusting Tonecraft. It just never had a place to put the one job that mattered most.

There was no single moment. Thessalyn found it during an ordinary quarterly account review, when seat level usage among brand sensitive accounts like Amberley had been quietly falling for weeks, while usage among less picky teams held flat. No ticket. No complaint. Verbena had simply gone back to the laminated sheet, six hours a campaign again, and never once told anyone.

We didn't lose Verbena to a slow tool. We lost her the moment "sound like us" stopped having anywhere to go inside it.

It was never really about the nine buttons on the toolbar. It was about whether any single one of them knew what job it was doing for the person pressing it. A tone slider set to "warm" means the same thing to a brand that wants a fast rough draft and a brand that wants a launch ready send. It can't tell the two jobs apart, because it was never asked to.

Thessalyn remembers the meeting where the original toolbar list got approved, eighteen months earlier. The question on the whiteboard read: what should a good writing assistant do. Someone had already listed what two rival tools shipped, and the room mostly agreed to match it, plus one thing extra. Nobody in that room asked what job any real customer was trying to get done. It made sense at the time. Marchvale had one type of customer then, and getting the basics working was the whole point.

Hand sketched comparison diagram titled The day the model gets it wrong. Left panel, a question mark box icon labeled First draft, caption reads says the candle is hand poured, it is not. Right panel, a scale icon labeled Caught before it ships, caption reads fact check step blocks it, the job stays matched to the brand.
The anchor has to survive its own first mistake. This is what that survival looks like.

What Thessalyn put in its place: before any draft starts, name the job. Match our exact brand voice for this launch, or get a fast draft to rewrite, or turn one great email into new versions. Choosing the brand voice job feeds the model Amberley's real style guide and its twelve strongest past campaigns, and every grounded draft has to clear a fact check pass against the actual product sheet first, the same guardrail that caught a draft claiming a candle was hand poured when it wasn't.

Three weeks after the job screen shipped, the share of Tonecraft drafts Verbena kept without a full rewrite went from 22 percent to 71 percent. The Ember Sale email that used to take her six hours took ninety minutes: one grounded draft, one fact check pass, done.

One nine button toolbar chased "what can a writing tool do." One job screen asked "what job is this for, and can I prove it." The second one is the only version Verbena still opens.

What I'd tell myself, the day that first toolbar list got approved on a whiteboard: the mistake wasn't adding a tone slider. It was never asking, out loud, whose job any of those nine buttons was actually doing.

SPARK, in the words of one candle brand's inboxNot a script for sounding thorough. SPARK is what forces you to name the one job in the customer's own words, and prove the fix survives its own first false claim.

SSituation. How does Verbena get this done today, without any job screen?
She writes every campaign by hand, from a laminated brand voice sheet and a folder of Amberley's own best past emails, about six hours per launch. Without a designed job screen, Tonecraft's toolbar answers "what can this tool do," never "what job is this draft actually for."
One brand, one campaign, one real task. Never a segment called "marketing teams."
PPayoff. What habit do I want this to build?
Before drafting anything, a writer names the one real job this draft is for, in the brand's own words, so Tonecraft knows what "good" means for that specific job. Faster drafts and fewer rewrites are downstream of that habit, not the goal itself.
Name the sentence Verbena says to herself before she opens the tool. That's the payoff.
AAnchor. The one decision everything else hangs on.
A required job screen before any draft starts: match our exact brand voice for this launch, get a fast draft to rewrite, or turn one email into new versions. Picking the brand voice job grounds the model in Amberley's real style guide and its own past top copy, not a generic tone setting. Thessalyn also considered just expanding the tone slider into a more detailed per brand preset dial, and rejected it: a dial only ever describes the surface of the writing. Two teams can want the identical warm, plainspoken setting for two completely different jobs, one wanting a rough draft to rewrite and one wanting a launch ready send, and a shared dial can't tell them apart or ground the model differently for either.
Concrete enough to argue with. This is the answer to the question.
RRisk. What breaks the first time it's wrong?
Name the job too vaguely, "help me write better marketing copy," and Tonecraft falls back to the same generic grab bag a feature list produces, just with a nicer label, and brand sensitive accounts churn because nothing feels tailored to their actual moment of need. Grounding also has a real cost: naming the job adds about ninety seconds of setup, one extra screen, and pulling in the style guide and swipe file adds a few seconds to generation, a trade Thessalyn accepts because it buys drafts that don't need a full rewrite.
Not "the job screen helps." What breaks, and what it costs, in either direction.
KKeep out. What I deliberately will not build into this job.
The brand voice job never absorbs writing that serves a different job, even a reasonable one. Long blog post outlines. Rewriting the tone of an internal Slack update, where no brand voice is at stake at all. Bolting either onto the same workspace would dilute what Tonecraft is actually good at for the job that matters most.
Ties straight back to Risk: a job that tries to do everything ends up excelling at nothing.
Hand sketched numbered icon list titled What the marketing copy job deliberately leaves out. Item one, a box icon, long blog post outlines, a different job entirely. Item two, a box icon, rewriting internal Slack updates, no brand voice at stake. Item three, a document icon, ad copy, email, landing pages stay the whole job.
Keep out doesn't mean these jobs are bad. It means bolting them onto this one dilutes what it's actually good at.

The recap, one line per letter: situation is Verbena drafting alone from a laminated sheet, payoff is one habit, name the real job before opening the tool, anchor is the required job screen grounded in real brand material, risk is a vague job producing the same grab bag with a nicer label, and keep out draws the line at any writing job that isn't ad copy, email, or landing pages, no matter how reasonable it sounds on its own.

And if you want to be sure it really works, try it somewhere elseSame five letters, a county permits office instead of a candle brand, and this time the missing job isn't "sound like us." It's "explain exactly why, in words the applicant can actually use."

Almsgate County Permits Office runs Plainscript, an AI writing assistant that drafts applicant facing letters, approvals, denials, requests for more information. Sarafina Warrender owns its roadmap. Almsgate's leadership had built Plainscript the same ungrounded way Marchvale first built Tonecraft: one generic "write a clear, polite letter" setting, tuned for friendliness, with no distinction between a routine approval and a denial that has to cite an exact section of county code.

Mapped onto SPARK: the situation is permit clerks writing every denial letter by hand, pulling the right code citation from a binder, because no tool had ever separated that job from a routine approval. The payoff is the same shared habit, name the real job before drafting. The anchor is the same job screen, run again: is this a routine approval, or a denial that must name the specific code section and the specific fix. Choosing the denial job grounds Plainscript in the actual code text and a set of past letters that survived appeal, not a generic polite tone. Risk runs the same shape: a vague "write a nice letter" job produces a friendly letter with no real citation, which is exactly the kind applicants call the office to argue about. Keep out draws the same line: Plainscript never drafts internal staff memos or press statements, a different job with a different owner entirely.

Escalation calls per 100 permit letters, before and after Plainscript's job screen shipped
50% 25% 0 36 39 34 38 22 13 9 Wk 1 Wk 2 Wk 3 Wk 4 Job screen ships Wk 6 Wk 7
Generic polite letterJob screen ships mid weekDenial job, grounded in code
Four flat weeks near 37 calls per 100 letters told nobody anything was wrong. Naming the denial job and grounding it in real code text is what actually moved the number.

Swap the trigger and it still runs.
Speed: an interviewer caps you at ninety seconds. Skip straight to the anchor, name the job before the draft, ground the specific job in real material.
Cost: no budget to rebuild the whole toolbar. Ship the job screen alone first, grounding can follow once the highest value job is proven.
The model got better, for real: say a future model writes noticeably better prose on its own. The job screen still matters, because "which job is this draft for" is a question about the customer, not about how good any one model happens to be that quarter.

Where people run it wrong.
They ship one default job so vague it might as well not exist, "write good copy," and call the naming step done.
They let the job screen exist but never actually ground the top job in real material, so picking "brand voice" changes a label, not the output.
They bolt every writing task the company can think of into one job's settings, so the one job that mattered most quietly gets worse to make room for the others.

How to use it live. Before answering a "design an AI writing tool" question cold, ask yourself one thing: what's the one sentence a real person would say about why they opened it today, in their own words, not a feature name. Naming that sentence, not a phrase like "personalized writing assistance," is usually exactly what the question is listening for.

Flashcards (tap any card to flip it)

1 · THE FRAMEWORK
What framework fits a question asking you to design the job a writing assistant is actually for?
Tap to flip
ANSWER
SPARK: situation, payoff, anchor, risk, keep out. It runs forward from how the job gets done today, not backward from one failure.
2 · THE PEOPLE
Who is this answer about?
Tap to flip
ANSWER
Thessalyn Colfax, the AI PM who owns Tonecraft's roadmap at Marchvale. Silvio Winthrop is Marchvale's CEO. Verbena Faircastle runs marketing at Amberley Home.
3 · THE HABIT
What habit does naming the job exist to build?
Tap to flip
ANSWER
Before drafting anything, a writer names the one real job this draft is for, in the brand's own words, so Tonecraft knows what "good" means for that specific job.
4 · THE ANCHOR
What's the one concrete design decision in this answer?
Tap to flip
ANSWER
A required job screen before drafting starts. Picking the brand voice job grounds Tonecraft in the brand's real style guide and its own best past campaigns, not a generic tone slider.
5 · THE OLD DECISION
What old decision would Thessalyn take back?
Tap to flip
ANSWER
Tonecraft's original roadmap answered "what should a good writing assistant do" by copying features other writing tools already had, instead of asking which job any one customer had hired it for.
6 · THE NUMBER
Fill in the blank: drafts Verbena kept without a full rewrite went from ___ percent under the old toolbar to ___ percent once the job screen shipped.
Tap to flip
ANSWER
22 percent, then 71 percent. A whole toolbar of features barely moved that number. Naming one job, grounded in real brand material, moved it 49 points.
7 · THE REPLAY
Same campaign, new design, what changes?
Tap to flip
ANSWER
The Ember Sale email that used to take Verbena six hours by hand took ninety minutes instead: one grounded draft, one fact check pass, done.
8 · CROSS-PRODUCT TRANSFER
Section 4 runs SPARK again on a different product. Which one, and what's the equivalent anchor?
Tap to flip
ANSWER
Plainscript, used by the Almsgate County Permits Office. Its anchor is the same job screen, naming whether a letter is a denial that must cite an exact code section, grounded in the real code text instead of a generic polite tone.

Check yourself Score: 0 / 0

Fill in the blank
1. Fill in the blank: after the job screen shipped, the share of Tonecraft drafts Verbena kept without a full rewrite rose from 22 percent to ___ percent.
Show hint
Look at the bar chart in Let's learn.
Show answer
71 percent. A nine button toolbar had barely moved that number for months. Naming one job and grounding it in real brand material moved it 49 points.
Multiple choice
2. Why did grounding the brand voice job in Amberley's real style guide and past copy matter more than a better tone slider would have?
  • A. Tone sliders take longer to load than a job screen.
  • B. A dial only ever describes the surface of the writing. It can't tell two different jobs apart, or check a draft against real brand material the way a grounded job can.
  • C. Amberley's brand voice never changes, so grounding was never really necessary.
  • D. The tone slider caused Tonecraft to crash during long campaigns.
Show hint
Look at the Anchor step, where Thessalyn considers and rejects a bigger tone dial.
Show answer
B. Two teams can want the identical warm, plainspoken dial setting for two completely different jobs. Only naming and grounding the job tells the model which one it's actually doing.
True or false
3. True or false: naming a vague job, like "help people write better marketing copy," is just as safe as naming a specific one, since Tonecraft grounds itself either way.
  • True
  • False
Show hint
Look at the Risk step.
Show answer
False. A vague job gives grounding nothing specific to check a draft against, so the model falls back to generic output, the exact grab bag a feature list produces, and brand sensitive customers churn.
Short answer, name the reversal
4. What old decision would Thessalyn take back, and why did it make sense when it was first made?
Show hint
Look at the key point box titled "The choice I would take back," in Let's learn.
Show answer
Model answer: Tonecraft's original roadmap was built by copying features other writing tools already offered, autocomplete, grammar check, a tone slider, instead of asking which job any one customer had hired it for. It made sense when Marchvale had one generic customer profile to design for. It stopped making sense once brand sensitive marketing teams became a real, distinct kind of customer.
Short answer, apply it yourself
5. Think of a writing tool you use yourself. What's one specific job you'd want it to name before it drafts anything, versus the generic "write faster" job it probably defaults to?
Show hint
Look for a moment where the tool's generic output needed heavy rewriting because it didn't know the actual situation you were writing for.
Show answer
Model answer: A note taking app's "clean this up" button defaults to generic tidying. The real job, for a meeting recap that a manager will actually read, is "pull out the three decisions and who owns each one," a genuinely different, more specific job than "make this read better."
Short answer, work the number
6. If Tonecraft's fact check pass had NOT caught the false "hand poured" claim before it shipped, would that be evidence the job naming screen doesn't work? Why or why not?
Show hint
Look at what the job screen is actually responsible for, versus what the fact check pass is responsible for.
Show answer
Model answer: No. The job screen's job is matching voice and intent to the right job. Catching a false factual claim is a separate guardrail, the fact check pass. A gap there means that specific guardrail needs strengthening, not that naming the job was the wrong anchor.
Before you close the answer
Why this works
Tests whether you'll design around what a customer actually hired the product for, or default to a feature list, and whether the judgment underneath it is genuinely about grounding a model in real brand material, not generic writing quality any tool could offer.
Follow-up traps
"Isn't naming a job before drafting just extra friction users will skip?" Response: it costs about ninety seconds once per campaign, and it's the reason drafts kept without a full rewrite went from 22 percent to 71 percent, so the friction pays for itself inside the same session.

"What if a customer's brand voice changes over time? Doesn't the grounding go stale?" Response: yes, which is exactly why the style guide and swipe file are pulled fresh each time rather than baked into the model once. Amberley refreshes its swipe file every quarter, and grounding reads from whatever is current.
If pressed
The fact check pass runs claim by claim against Amberley's actual product spec sheet, not just a general tone check. Any sentence naming a material, a process, or a certification has to match an approved field before the draft can be marked ready. That's a separate, narrower gate than the brand voice grounding, and it exists specifically because a model confident in a brand's voice can still say something false in that voice.
From U2xAI Academy

From answering questions to owning outcomes.

A live workshop where you ship a working AI agent, defend a launch decision, and walk away with a portfolio recruiters can't wave off, not just more questions to study.

  • A live AI agent you actually shipped
  • A launch decision you can defend under pressure
  • An interview-ready portfolio, not more flashcards
Know more