Artifact critiqueIntermediateResponsible AI & Advanced Practice / Building an AI PM portfolio / #11

What should the write-up accompanying a portfolio build contain?

ORDER the portfolio piece is Keepline, an AI copilot that drafts retention offers for a telecom call center

Alderway Mobile is a fictional regional telecom carrier. Keepline is the AI copilot Desmond Okafor built as a portfolio project: it drafts a retention offer for a call-center agent the moment a customer's cancellation risk crosses a set line. Corinne Vance is a hiring manager who reads about forty of these write-ups every hiring season.

The direct answer
If you only have an hour, write the tradeoffs-you-rejected section first, with one real screenshot of a failed attempt attached. That section is the hardest one to fake after the fact, and it is the one part of a write-up that actually proves judgment instead of just execution.
Do this, in order
  1. Write the tradeoffs-you-rejected section first, with a real failure screenshot attached.Why: it's the hardest section to backfill honestly once the memory of the build fades.
  2. Decide, before you start building, which failed attempts you'll keep proof of.Why: you can't write about a rejected path you never saved a trace of.
  3. Attach one concrete piece of evidence per claim, not a paragraph of reassurance.Why: a claim with nothing attached reads as marketing, not a record.
  4. Say plainly what you left out of the build, and why.Why: naming your own scope limit is more of the same judgment the tradeoffs section is proving.
  5. Rank any extras, a demo video, a slide deck, behind the reasoning section, never ahead of it.Why: extras are nice to have; the reasoning is the one thing a reviewer actually reads for.
  6. Save the wording, layout, and font for last.Why: polish is the one part of the page you can redo anytime without losing anything.

How to answer this, stage by stage

Nobody is grading whether your write-up looks nice. They're grading whether you know which section actually proves you can think.

Stage 1
Scope it to one real write-up
Say it like this
"I'll answer this for a write-up next to Keepline, a copilot that drafts retention offers for a telecom call center."
Why this works
Grounds an abstract question about "portfolio content" in one concrete page a real reviewer would read.
Stage 2
Say your structure out loud
Say it like this
"I'll use ORDER. Outcome, what the page is proving. Reversibility, what's hardest to fake later. Dependency, what has to exist during the build. Evidence, what's cheap and concrete. Rank, what I'd write first."
Why this works
Signals a real method for choosing what goes in, not a wish list of everything a write-up could hold.
Stage 3
Name the outcome the page is proving
Say it like this
"The write-up isn't proving I can run a stack. It's proving I can reason under a real tradeoff, the same thing an AI PM does at work every week."
Why this works
Without a named outcome, every section looks equally worth including, and none of them are.
Stage 4
Find what's hardest to fake
Say it like this
"Wording and layout, I can redo any night I want. The section on what I tried and rejected, I can't honestly reconstruct once I've forgotten why I actually made that call."
Why this works
A reviewer can tell reasoning written in the moment from reasoning invented after the fact.
Stage 5
Name the real dependency
Say it like this
"For that section to be honest, I have to actually keep a trace of the failed attempt while I'm building, a bad screenshot, a low eval score, not go hunting for one afterward."
Why this works
Shows the write-up isn't a document you write at the end. It's a record you start keeping on day one.
Stage 6
Pick the cheap, concrete evidence
Say it like this
"One screenshot of Keepline drafting a discount that didn't exist, next to the rule that caught it, does more work than three paragraphs saying 'I iterated a lot.'"
Why this works
A concrete artifact is checkable. A claim of effort is not.
Stage 7
Rank it, and close on one line
Say it like this
"So, in order: tradeoffs rejected with proof, then the build itself, then the metric, then whatever's left of my hour goes to wording."
Why this works
Ends on an ordered plan, not a list, which is exactly what the question actually asked for.

Let's learn

A portfolio write-up is the page next to a build that tells a hiring manager which parts were a real decision, and which parts just happened.

Corinne Vance reads about forty of these write-ups every hiring season. Six months ago, most of what she read were two paragraphs: a headline metric and a screenshot. She'd spend maybe ninety seconds on one before moving to the next file.

Knowledge spark: what's a golden set? A small pile of real or realistic examples you test a model against before you write down a number about it. Without one, a "94% accurate" claim has nothing behind it to check.

Lately, more candidates write a full page: a process paragraph, three screenshots, a metric. Corinne now spends six or seven minutes on some of these.

Minutes Corinne spends reading, by what the write-up contains
7 min 3.5 min 0 Headline only 1.5 min Full page, no tradeoffs 4 min Tradeoffs + proof 7 min
More words alone barely moved her reading time. A rejected tradeoff with a failure screenshot attached nearly doubled it, and it's the only version that gets a real follow-up question.

A longer page is not the fix. The write-ups that actually change Corinne's mind are not the longest ones. They're the ones that admit to a real dead end, with something to prove it. Most candidates fill the extra length with more of the happy path instead.

The extra paragraphs were never the proof. The one screenshot of the thing that didn't work was.

At its worst: a candidate spends a whole evening polishing an intro paragraph and adding a fourth screenshot of the same success, and never once mentions the version he tried and threw away. Corinne finishes the page no more sure of his judgment than when she started.

The decision that mattered Write the tradeoffs-you-rejected section first, with one real screenshot attached, before touching a single word of polish.

What I would leave alone: the exact font, color, or layout of the page truly does not matter, and redoing it at the very end costs nothing. Spend none of your one hour there.

The lesson: a write-up is not a diary of everything you did. It is proof of the moments you chose one thing over another. Write the proof first, and let the polish come last, if it comes at all.

Now here is the same thing as a story

The short version above is what you'd say defending your own write-up's structure out loud. Read this one for how Desmond actually rebuilt his.

Desmond Okafor can tell a save-able customer from a lost one before Alderway Mobile's own retention dashboard flags it, four years of doing it by ear on the phones before he ever touched a model.

His first draft of the Keepline write-up went well for a while. He built the copilot on weeknights, wrote up the architecture, and felt good about a clean one-page summary he could hand anyone.

Then the habit thinned in three beats. His first draft had a short paragraph naming one alternative approach he'd considered. His second draft trimmed that paragraph to fit the page on one screen. His third and final draft cut it entirely, since it felt like the least essential part next to a clean screenshot of the copilot working.

A friend reviewing a printed copy circled the tools-and-stack list sitting at the very top of the page and wrote one line in the margin: "This proves you can use tools. It doesn't prove you can think."

Hand sketched flow diagram titled What unblocks what. Five boxes: build it, keep the miss highlighted, name tradeoff, write it up, polish last.
The second box is the one most drafts skip. You can't write about a miss you never kept a trace of.

Desmond went back through his old commits hunting for an early, worse version of Keepline, one that had drafted a retention offer using a discount tier that didn't actually exist in Alderway's pricing sheet. He'd almost deleted that screenshot months earlier, since it embarrassed him at the time.

Hand sketched comparison diagram titled Reversible or not. Left panel, a document icon labeled Reword intro, caption redo anytime, free. Right panel, a box icon labeled Keep the miss, caption decide once, or lost.
Wording, he could always redo. The decision to keep or delete that one bad screenshot, he only got to make once.

Rebuilding the write-up took about three hours, plus another ninety minutes just hunting through old commits for the screenshot he'd nearly thrown away. The real cost was never the three hours of rewriting. It was that his best piece of evidence, the one that actually proved he'd caught something real, almost didn't survive to make it into the page at all.

Back at his kitchen table on the Sunday night he wrote the first draft, putting the tools-and-stack list at the very top had felt obvious. It was the easiest section to write, and it felt like a natural place to start a technical document. That made complete sense for a page meant to organize his own thinking. It stopped making sense the moment the page's real job changed to persuading someone else he could reason under uncertainty, not just operate a stack.

Hand sketched decision tree titled Reading one metric three ways. Root, one clean metric, branching to own curated eval leads to looks great, one lucky case leads to unproven, eval keeps failures leads to actually trustworthy.
A clean number alone doesn't say which branch it came from. Only the failure case behind it does.

The next time Corinne opened one of his drafts, she reached the tradeoffs section in under ninety seconds, since it now sat first, and read the whole page in six minutes flat, ending on a real follow-up question about the pricing guardrail instead of a polite pass to the next file.

The old page asked a reviewer to take his judgment on faith. The new page put the proof of it in the very first section she'd read.

I put the tools list first because it was the easiest thing to write, not because it was the thing that mattered. It took one circled sentence in a margin to see I'd built the page for myself, not for the person actually reading it.

ORDER, the write-up broken openNot a checklist of everything a write-up could hold. ORDER is what tells you which one section to write when the clock is the real constraint.

O
Outcome. What all the sections are competing to prove.
Not that the candidate can run a stack. That they can reason under a real tradeoff.
Without a named outcome, every section looks equally worth including.
R
Reversibility. What's hardest to fake or backfill later.
The tradeoffs-rejected section. Wording and layout, redo any night. Real-time reasoning, you only get once.
This is why Rank sends you to that section first.
D
Dependency. What has to exist before the page can be honest.
Keeping the failed screenshot while building, not hunting for one after the fact.
Ties directly to the discount-tier screenshot Desmond almost deleted.
E
Evidence. What's cheap and concrete.
One real screenshot of a failed attempt beats a paragraph asserting "I iterated a lot."
A guardrail catch becomes the strongest evidence on the whole page.
R
Rank. Which section you'd write first, and why.
Tradeoffs rejected, with proof attached, before touching the tools list or a single word of wording.
The hardest step, and the one that turns a wish list into an actual hour of writing.
Hand sketched icon list titled Five things Corinne checks first. Items: a document icon labeled a real failure shown, a gauge icon labeled a named eval set, a scale icon labeled one tradeoff rejected, a funnel icon labeled cost or speed given up, a box icon labeled what got left out.
None of these five are about the build itself working. All five are about proving you knew where it didn't.
Real follow-up question rate, by where the tradeoffs section sits
70% 35% 0 Not included Appendix Mid-page First section 10% 25% 45% 70%
Moving the same section from an appendix to first place nearly tripled the rate of a real follow-up question, with no new evidence added, only a reordering.
Hand sketched quadrant titled Sorting sections by proof value. Axes how easy to fake and how much it convinces. Tools list and intro paragraph sit easy to fake and convince little. Single metric sits mid. Rejected tradeoff sits hard to fake and convinces a lot.
The tools list and the intro paragraph sit in the same corner: cheap to write, and cheap to fake. The rejected tradeoff sits alone in the corner that actually matters.

The recap, one line per letter: outcome is proving judgment, not just execution. Reversibility is the tradeoffs section being the hardest to backfill honestly. Dependency is keeping the failed screenshot while building, not after. Evidence is one concrete artifact beating a paragraph of assurance. Rank is writing that section first, before anything else, when the hour runs out.

And if you want to be sure it really works, try it somewhere elseSame five letters, a livestock health app instead of a phone call. A field with real animals in it, where a miss costs more than a customer.

Bramwell Agritech is a fictional agricultural co-op. Anwar Deeb built a livestock health-triage assistant there, an AI tool that flags animals worth a vet's attention from daily photos and weight logs. Ffion Pryce reviews his write-up.

Mapped onto ORDER: the outcome the write-up proves is that Anwar understands the real cost of a missed case, not that the model runs on a farm's data. The reversibility test lands on the one section describing an animal the model rated low-risk that later got sick, since a farmer can't un-know that, and a reader can tell a real logged miss from one reconstructed after the fact. The dependency is that Anwar had to actually log that real case, the exact threshold score, the date, the vet's later note, while it was happening, not rebuild it from memory for the write-up. The evidence is one real annotated vet chart showing the miss next to the threshold that let it through, stronger than any paragraph saying the model "isn't perfect." And the rank: write that missed-case section first, since in a field where being wrong costs an animal's health, it's the one section that proves he understands the stakes, not just the code.

Hand sketched labeled parts diagram titled The livestock write-up, ranked. Center document icon labeled Triage write-up, with four callouts: missed case, vet chart proof, threshold used, written first.
Three of these four parts exist because a real animal's case went wrong once. The fourth just says which one gets written down first.

Swap the trigger and it still runs.
Speed: an interviewer caps you at sixty seconds. Say "lead with the section that's hardest to fake: what you got wrong and caught, not what you got right," and stop.
Cost: there's no time to build the write-up twice, only once, in an evening. Even one honest paragraph naming a single dead end beats zero, done in five minutes.
The model gets better, for real: if Keepline's or Anwar's accuracy improves later, that's still no reason to delete the rejected-tradeoffs section. A better model retroactively makes the old failure more interesting, since it shows real movement, not less.

Where people run it wrong.
They write the tools list first because it's the easiest section, the exact mistake Desmond made.
They spend the whole hour polishing wording and never touch the tradeoffs section at all.
They invent a rejected alternative after the fact, and it reads like an afterthought instead of a real decision made in the moment.

How to use it live. When someone asks what a write-up should contain, ask yourself one question first: which section would be hardest to fake if you only had an hour. Write that one first, out loud, before listing anything else.

Flashcards (tap any card to flip it)

1 · THE FRAMEWORK
What framework fits "what should a portfolio write-up contain"?
Tap to flip
ANSWER
ORDER: outcome, reversibility, dependency, evidence, rank. Rank names the one section you'd write first under a real time limit.
2 · THE PERSON
Who is this answer about?
Tap to flip
ANSWER
Desmond Okafor, four years reading retention calls at Alderway Mobile, who built Keepline as his portfolio project.
3 · THE HABIT
What did Desmond stop doing across three drafts?
Tap to flip
ANSWER
Mentioning any alternative he'd considered and rejected. Draft one had a short paragraph on it; draft three cut it entirely.
4 · THE MECHANISM
Which section is hardest to fake after the fact, and why?
Tap to flip
ANSWER
The tradeoffs-rejected section. A reader can tell reasoning recorded in the moment from reasoning invented after the build is done.
5 · THE OLD DECISION
What decision would you take back?
Tap to flip
ANSWER
Putting the tools-and-stack list at the top of the page because it was the easiest section to write, back when the page was only for himself.
6 · THE NUMBER
Fill in the blank: a write-up with a rejected tradeoff and proof holds Corinne's attention for about ___ minutes, versus ___ for a full page with none.
Tap to flip
ANSWER
7 minutes versus 4 minutes. Only the first one reliably ends in a real follow-up question.
7 · THE REPLAY
Same draft, reordered. What changes for Corinne?
Tap to flip
ANSWER
She reaches the tradeoffs section in under 90 seconds and reads the whole page in 6 minutes, ending on a real question about the pricing guardrail.
8 · CROSS PRODUCT TRANSFER
Section 4 answers this again for a different project. Which one, and what's the rank there?
Tap to flip
ANSWER
Anwar Deeb's livestock health-triage write-up at Bramwell Agritech. The rank is the same: write the missed-case section first.

Check yourself Score: 0 / 0

Multiple choice
1. Why does this answer say to write the tradeoffs-rejected section first, instead of the tools-and-stack list?
  • A. Because tools lists are boring to read.
  • B. Because the tradeoffs section is hardest to fake honestly after the fact, and it proves judgment instead of just execution.
  • C. Because interviewers never read past the first section anyway.
  • D. Because listing tools reveals confidential company information.
Show hint
Look at the quadrant sorting sections by proof value.
Show answer
B. A tools list is easy to fake and convinces little. A named tradeoff with proof attached is hard to fake and convinces a lot.
True or false
2. True or false: this answer recommends spending your first hour polishing the intro paragraph and layout.
  • True
  • False
Show hint
Look at "what I would leave alone."
Show answer
False. Wording and layout come last, since they can be redone anytime without losing anything. The tradeoffs section comes first.
Fill in the blank
3. Fill in the blank: Corinne reviews about ___ portfolio write-ups every hiring season.
Show hint
Look near the start of the lede and Section 1.
Show answer
40. That volume is why she can only spend a few minutes on most of them.
Short answer, where it wouldn't matter
4. Name a part of the write-up where this exact rigor doesn't matter, even under a strict time limit.
Show hint
Look at "what I would leave alone."
Show answer
Model answer: The font, color scheme, and layout. They can be redone at any point without losing anything real, so they're safe to leave for last, or to skip.
Short answer, apply it yourself
5. Pick a project you've built or worked on. If you had one hour to write it up, which single decision would you write about first, and why is it the one you'd struggle to fake later?
Show hint
Think about a choice you made that you could still defend today, versus one you'd have to guess at reconstructing.
Show answer
Model answer: Most people land on a moment where they chose to cut scope or drop a feature, since that's the decision that's hardest to invent convincingly after the fact.
Short answer, the number question
6. If Corinne only had 90 seconds per write-up instead of 7 minutes, would leading with the tradeoffs section still make sense? Why or why not?
Show hint
Think about what a reader with almost no time actually gets to see.
Show answer
Model answer: Yes, even more so. A 90-second reader only ever reaches the first thing on the page, so whatever sits there has to be the section that proves judgment, not the easiest one to write.
Before you close the answer
Why this works
Tests whether you understand that a write-up's job is to prove judgment, not summarize effort, and whether you can prioritize what to write under real time pressure instead of listing everything a write-up "could" include.
Follow-up traps
"Isn't listing your tech stack still useful information?" Response: sure, but it proves you can operate tools, not that you can reason under a real tradeoff, so it belongs after the sections that do that.

"What if you genuinely never hit a real failure worth writing about?" Response: then the build wasn't tested hard enough yet. Every real model has at least one case it gets wrong, and the write-up should say which eval or threshold turned that up.
If pressed
Alderway Mobile's real guardrail rule was that any discount figure Keepline drafted had to be checked against a live pricing lookup table before an agent could send it to a customer. That exact rule is what caught Desmond's early hallucinated offer, and it became the strongest evidence on his whole page.
From U2xAI Academy

From answering questions to owning outcomes.

A live workshop where you ship a working AI agent, defend a launch decision, and walk away with a portfolio recruiters can't wave off, not just more questions to study.

  • A live AI agent you actually shipped
  • A launch decision you can defend under pressure
  • An interview-ready portfolio, not more flashcards
Know more