learn-ai-project-management-with-phoebe / Leader session 5 of 6
Learn AI + Project Management with Phoebe · Leader track · Session 5 of 6

The reporting cadence: a rhythm a machine can maintain and a human can defend

Reporting is where your delivery function becomes visible to the rest of the organisation. Get the cadence right and AI quietly removes a day of assembly work every week. Get it wrong and you industrialise a beautiful, confident, unreadable report that nobody challenges because nobody can. Tonight you design the week, set the quality bar, and write the one rule that makes a colour mean something.

🟠 Leader track Heads of delivery PMO leads Cadence design 45 min
0-3 · Welcome 3-20 · The rhythm and the six checks 20-42 · Your cadence + the RAG rule 42-45 · Q&A
Part 0

Where we are

In a1 you split the delivery job into what AI drafts and what only a person decides. In a2 you set the governed boundary: what goes into a workspace, what never does, who is allowed to connect what. In a3 you built a review rubric for AI-drafted plans, and in a4 you did the same for risk, dependencies and escalation. Those four sessions were about artifacts your team produces. Tonight is different. Reporting is the artifact the organisation consumes, weekly, forever. It is where your function is judged, and it is the single most common first place AI lands in a delivery team - which makes it the most common place to go fluently, confidently wrong at scale.

Live - presented in session Self-study - read after class ★ Try it now prompt Official sources covered
★ What you walk out with today A five-day reporting rhythm with the two human checkpoints marked, the six checks translated into what each one buys you and what breaks without it, a written RAG rule your sponsor can agree in one meeting, and a way to spot the new failure mode - the report that has become too good to carry a warning.
Part 1 · covers the measurement domain + where AI lands first

The week, as a loop 6 min live

Most reporting is not a process, it is a Friday afternoon. Somebody assembles four threads, a ticket export and a spreadsheet into a document, ships it at six, and starts again on Monday. That works until the person is on leave, or the programme doubles, or the sponsor asks for a second one. A cadence fixes it: five small stops instead of one large heroic one, each with an owner, and only two of them needing a human in the room.

FIVE STOPS · ONE LOOP · TWO OF THEM NEED A PERSON Mon · Refresh plan, risks and last report in PMO analyst Tue · Sweep risk and gate forecasts moved PM + risk owner Wed · Draft report drafted, gaps block read AI drafts, PM cuts Thu · Decide steerco takes the ask, logs it sponsor decides Fri · Sign decisions logged, actions sent out PM signs it and the loop starts again on Monday Three stops are assembly work a machine can hold all week without getting tired. Two are judgment. Thursday nobody can automate, Friday nobody should. If a stop needs a meeting to happen, it is not a drafting stop. Protect those two, move the rest.
🔍 Click to zoom - the reporting week as five small stops, with the two that require a person marked in warm
LiveThe five stops, and who owns each one4 min

The point of naming stops is that each one becomes small enough to survive a holiday. Nobody has to be heroic on a Friday, because Friday is now fifteen minutes of reading and signing.

  • Monday - refresh the context pack. Swap in the current plan, the current risk register, and last week's report. Five minutes, usually a PMO analyst. A stale pack is the most expensive failure in this whole cadence, because it produces confidently wrong answers about last month's programme, which is far harder to catch than obvious nonsense.
  • Tuesday - risk and dependency sweep. What moved: gate forecasts, new risks, risks that closed, dependencies that slipped underneath you. On Northwind this is the stop where R-03, the data-engineer hiring gate, gets its honest forecast before it becomes a surprise.
  • Wednesday - draft plus the gaps block. The draft is cheap. The gaps block is the valuable half: the list of things the material could not support. Read that first, then go and fill it from people.
  • Thursday - steerco and decisions. The ask gets answered, or it does not, and either way it gets logged with an id and a date. No model participates in this stop.
  • Friday - log and actions out. Decisions written into the log, actions distributed with one owner and one date each, report signed by a named human and sent.
Real world

The programme that survived a holiday. A head of delivery ran a portfolio where one PM assembled every weekly pack alone. She took two weeks off in August and the pack simply did not go out - twice. Nobody noticed until a dependent team built a sprint on a date that had moved. After the cadence went in, the Monday refresh and Wednesday draft ran without her, and the cover person's whole job was Thursday and Friday: attend, decide, sign. The pack went out both weeks she was away, and the delta section is what caught the moved date.

Self-studyWhy the two human stops are Thursday and Friday2 min read

You could put a person at all five stops. Plenty of organisations do, call it oversight, and get nothing for it - because review that happens everywhere happens nowhere. PMI's 2026 AI standard is unusually direct about this: oversight needs clear structures, roles, authority and guardrails, with real intervention triggers, rather than a perfunctory review layered over everything.

So concentrate the human where intervention is actually possible. Thursday is where a decision gets taken, and a model cannot take one. Friday is where accountability attaches: one name, and that person read it, especially the numbers.

Monday, Tuesday and Wednesday are assembly. If something goes wrong there, Thursday and Friday catch it, which is what an intervention trigger means in practice. Two real checkpoints beat five ceremonial ones every time.

Part 2 · the quality bar, from the leader's chair

The six checks, and what each one buys you 5 min live

Your PMs build these in the PM track. You need something different: the ability to pick up any report in your portfolio and know in thirty seconds whether it can be acted on. Six checks, and each missing one produces a specific, recognisable failure - not a vague drop in quality. Learn the failures and you can review by symptom. Send your PMs to session b6 in the PM track to feel it: the same six checks are scored live there, a Northwind week 9 report climbing from 24 out of 100 with every lever off to 100 with all six on. Watching the score jump when the RAG rule goes in does more for adoption than any memo you will write.

The checkWhat it buys youThe failure when it is missing
Status by a written RAG ruleA colour that two people read the same way, stated in the report itselfThe colour is a mood. Amber runs for eleven weeks and nobody acts, because nobody agreed what it meant
The delta since last weekMovement, which is the only thing a sponsor can actually steerA 14-day slip reads exactly like a stable plan. The photograph is fine, the film is missing
One named owner and one date per actionA commitment a specific person can be held to next weekOrphan actions. "Follow up on data quality" reappears for a month, reworded each time
Decisions taken, loggedAn audit trail, and an end to relitigating settled thingsThe same decision gets taken three times, slightly differently, by three different groups
Evidence behind every claimA report that can be challenged, which is the only kind that can be trustedUnfalsifiable prose. "Quality is looking reasonable" cannot be argued with or relied on
The askA steerco that decides something instead of nodding at somethingIt is a broadcast. The real escalation then happens in a corridor, off the record
LiveReading a report by symptom, in thirty seconds3 min

You are not grading prose. You are looking for six structural features, and their absence is visible before you read a sentence.

  • Scan for a colour and a rule on the same page. If the rule is not restated, it lives in someone's head and you have a mood ring.
  • Scan for two numbers of the same kind. This week and last week. One number is a photograph.
  • Scan the actions column for names. Any instance of "the team", "engineering" or "TBC" that has been TBC for three weeks is a finding.
  • Scan for ids and for the last paragraph. D-14, R-07, NWD-412, a metric with a denominator - no ids means no evidence, whatever the adjectives say. And if the closing paragraph does not name a decision, options with costs and an expiry date, the report was written to inform, not to move anything.
Real world

Amber for eleven weeks. A programme reported amber every week for a quarter with no written definition of amber. Every steerco nodded. In week twelve it went red, and the sponsor asked why nobody had flagged it. The honest answer was that everybody had - in a colour nobody had defined. The report was never lying. It simply could not carry information, and no amount of AI drafting would have changed that, because the missing thing was a rule, not a paragraph.

Self-studyWhat the checks look like on Northwind week 92 min read

Northwind is nine weeks in. M3, raw ingest signed off, was gated for 10 August and now forecasts 24 August. Six of nine sources are landing, 812k rows a night, and data quality is 96.4% against a 99% gate. Here is the same week with the checks on and off.

Check offCheck on
"Amber - some challenges with ingest""Amber: M3 forecasts 24 August against a 10 August gate, with a recovery plan attached. Amber means a gate is 2 or more weeks late and recoverable"
"Data quality is improving""DQ 96.4% this week against 96.1% last week, gate is 99%. Ticket NWD-412"
"The team will chase the vendor""Priya N. to hold the Veridian escalation call by 6 August, VN-2291"
"We discussed contract cover""D-14, 29 July, Elena V.: hold contract cover pending the 6 August vendor call"
"We may need more resource""Decision needed by 8 August: 2 contract engineers for 4 weeks, about 38k, or M4 moves. After 8 August the option lapses"

Nothing in the right column is longer prose. It is the same length with facts in it. That is the whole shift.

Part 3 · the failure nobody warns you about

When the reports get too good 5 min live

Here is the failure mode that only appears after AI drafting works. Every report in your portfolio becomes well-structured, correctly formatted, evenly toned and pleasant to read. And you quietly lose a signal you had been using for years without ever writing it down: the sound of a PM who was worried. PMBOK 7's stakeholder and measurement domains both treat the report as a two-way instrument - it informs, and the reaction to it informs you back. Uniform tone attacks the second half, so you have to replace that channel on purpose.

ONE REPORT · THREE READERS · THREE DIFFERENT JOBS The sponsor reads it once, for thirty seconds A dependent team reads it to plan their own week The delivery team reads it to see themselves in it Actually needs the ask, the date, the cost of not deciding this week Actually needs the dates that move theirs, and one name to call Actually needs the gaps block, and what was escalated for them Write it only for the sponsor and the other two stop reading and those two are the readers who used to catch your errors for free. A perfect report nobody reads is worse than a rough one three teams argue with.
🔍 Click to zoom - the same weekly report, and the three different jobs it has to do
LiveThree habits that keep the signal3 min

You cannot un-improve the writing, and you should not want to. What you can do is stop relying on tone as your early-warning system and replace it with three deliberate habits.

  • Read the gaps block, not just the report. The report is what could be supported. The gaps block is what could not - the action with no real owner, the number nobody could source, the date that came from an assumption. Three weeks of gaps blocks tells you more about a programme than three months of reports, and it is the part that has not been smoothed.
  • Ask one question no report can answer. Every week, one PM, one question: "what are you worried about that is not in here?" or "what would you tell me if this were not written down?" It takes ninety seconds and it restores the channel the tone used to carry.
  • Watch for the report nobody reads. Drafting is cheap now, so reports grow. A five-page pack that is beautiful and unread is a worse instrument than the scruffy one page people argued with. If nobody has challenged a line in a month, that is not quality. That is silence.
Real world

The PM who stopped sounding worried. A delivery lead had one PM whose reports she always read twice, because when he was anxious his sentences got shorter. After the team standardised on a drafted format, his reports read exactly like everyone else's - clean, complete, calm. Two months later a vendor issue surfaced late, and when she went back through the packs the warning was there every week, in the gaps block, in a line about an owner nobody could confirm. She had been reading the report and skipping the gaps. Now the gaps block goes first in her portfolio review, and the reports go second.

Demo 1 of 2

Design your own cadence ★ 12 min · everyone builds

Take your own reporting week and fill a five-day grid: what is drafted, what is reviewed, and who signs. Do it for one real programme, not the portfolio - the portfolio version is just this repeated. Most people find they have three stops crammed into one afternoon and no owner named for any of them.

Write down what actually happens today, honestly. Who assembles the pack, on which day, from how many sources, and how long it takes them.

Split it into the five stops. Refresh, sweep, draft, decide, sign. Some of yours will be empty, which is the finding.

Put a named person against each stop - a role is not enough, because a role cannot be on leave. Then name the cover person for each.

Mark the two stops that require a human in the room. If you have marked more than two, ask what intervention each one is actually capable of - review with no power to change anything is theatre. Then write the one thing that would break the loop, which is usually the Monday refresh, because it is small, boring and nobody's actual job.

★ Try it now - the cadence grid, fill it inPROGRAMME: .......... REPORT GOES OUT: .......... at .......... MON refresh the pack done by: .......... cover: .......... current plan, current risk register, last week's report go in TUE risk + dependency sweep done by: .......... cover: .......... gate forecasts moved, risks opened and closed, deps slipped WED draft + gaps block AI drafts gaps read first by: .......... the gaps block is read BEFORE the report, and filled from people THU steerco + decisions HUMAN REQUIRED chaired by: .......... the ask is answered or refused, and logged with an id and a date FRI log + actions out HUMAN REQUIRED signed by: .......... one owner and one date per action, a named human signs, then send HARD LENGTH: one page weekly. HARD RULE: no colour without the rule beside it.
The cadence almost everyone actually hasMON-THU nothing FRI one person assembles four threads, a ticket export and a spreadsheet into a document between 2pm and 6pm, then sends it reviewed by: nobody signed by: whoever sent it cover: none
Why the bad one is worse than it looks It is not the four hours. It is that the pack has one point of failure, no review that could change anything, and no stop where a slip gets noticed before it is written down. Adding AI to that Friday afternoon makes the document faster and the programme no safer. The cadence, not the drafting, is what buys you the safety.
Demo 2 of 2

Write the RAG rule with your sponsor ★ 10 min · build your own

This is the highest-value fifteen minutes in the leader track, and it is one meeting. An agreed, written RAG rule turns a colour from an opinion into information, and it is the only one of the six checks that a PM cannot install alone - it needs your sponsor's agreement, which means it needs you.

ColourA workable default to start from
GreenEvery live gate forecasts on or before its date, and no critical risk is open above the agreed threshold
AmberA gate forecasts 2 or more weeks late, and a credible recovery plan is attached to this report
RedA gate is unrecoverable without a steerco decision on scope, money or the date

Take the default above into the meeting already written. Never open with "how should we define amber" - you will get a philosophy session. Open with a draft and invite edits.

Test it against three real past weeks from your own portfolio. If the rule would have coloured a week differently from how you actually coloured it, that is the conversation worth having.

Agree the threshold numbers out loud: how many weeks late is amber, what counts as a critical risk, and who can overrule the rule. Somebody usually can, and it is better to name them than to discover them.

Agree that the rule gets restated in one line inside every report, then put it in the house rules file the same day, before anyone's memory of the meeting drifts. That one line removes the "what does amber mean here" conversation that eats the first ten minutes of every steerco.

★ The script for the one meeting - fifteen minutes"I want five minutes on one thing: what amber means. Right now our reports use a colour we have never written down, so you and I read it differently. That is why amber can run for a quarter without anyone acting. Here is a draft rule. Green: every live gate forecasts on or before its date and no critical risk is open above our threshold. Amber: a gate is 2 or more weeks late and there is a credible recovery plan attached. Red: the gate is unrecoverable without a decision from you on scope, money or the date. I have run our last three weeks against it. Two would have been coloured the same. Week 9 would have gone amber a fortnight earlier than it did. Two questions. Is 2 weeks the right amber threshold for you? And who can overrule the rule when the rule and the room disagree? If you are happy, it goes in every report as one line, starting Friday."
Real world

Fifteen minutes, then nine months of arguments that did not happen. A PMO lead had spent two years watching steercos open with a debate about what the colour meant. She took a drafted rule to her sponsor, tested it against three past weeks, and got it agreed before the coffee arrived. The rule itself was unremarkable. What changed was that the first ten minutes of every steerco stopped being about interpretation and started being about the ask - and the programme that had been amber for a quarter was red within two weeks, on purpose, with a decision attached.

Homework

Try it yourself - this week ◐ 30-45 min total

Source material

Official sources covered

Taught from the delivery reporting canon, PMI's public AI guidance, and the Northwind case carried through both tracks. Certification and the full normative text of PMI's standards stay with PMI. This session covers:

PMBOK Guide 7 - measurement and stakeholder performance domainsParts 1-3 · what a report is for, and who it has to serve
PMI - AI in project management, adoption reportingPart 0 · status reporting as the most common first use
PMI - Standard for AI in Portfolio, Program and Project Management (2026)Part 1 · oversight with real intervention triggers, not perfunctory review
PM track b6 - the status report live simulatorPart 2 · the same six checks, scored lever by lever
Check yourself

Three questions before you go 🎯 ◐ 90 seconds

1 · You are designing the reporting week. How many stops should require a human in the room, and which?

Oversight that happens everywhere happens nowhere. PMI's standard asks for real intervention triggers, not a review layer over everything. Concentrate the human where intervention is possible: the decision, and the signature.

2 · Every report in your portfolio is now well-structured, evenly toned and pleasant to read. What have you lost?

You used to hear when a PM was worried. Uniform drafting removes that channel, so replace it deliberately: read the gaps block before the report, and ask one PM a week what is worrying them that is not written down.

3 · A programme has reported amber for eleven consecutive weeks. What is the first thing to fix?

Eleven weeks of amber is the classic undefined-colour failure. Without a written rule the colour is a mood, so nobody can act on it and nobody can argue with it. One fifteen-minute meeting fixes it permanently.

Leader session 5 cheat sheet · pin this

The five stopsMon refresh · Tue sweep · Wed draft + gaps · Thu decide · Fri log and sign.
Two human pointsThursday nobody can automate, Friday nobody should. Everything else is assembly.
The six checksrule-based status · delta · owner and date · decisions logged · evidence · the ask.
Review by symptomno rule = a mood · no delta = a photograph · no name = an orphan · no ask = a broadcast.
The RAG rulegreen on or before the date · amber 2+ weeks late with a recovery plan · red needs a steerco decision.
Take a draft, not a questionthe rule gets agreed in one meeting if you walk in with it already written.
The too-good reportuniform tone removes your early warning. Read the gaps block first, ask one live question.
Guard the lengthdrafting is cheap, so reports grow. One page weekly, defended in the house rules.