learn-ai-project-management-with-phoebe / Leader session 1 of 6
Learn AI + Project Management with Phoebe · Leader track · Session 1 of 6

What AI changes about delivery: the labour moves, the accountability does not

Somebody above you wants an AI number in the PMO plan. Somebody below you has tried it twice and gone quiet. This session gives you the one distinction that survives both conversations: artifact labour versus judgment. By the end you have split your own function's week into those two piles and drafted the one-pager that tells your PMs what AI drafts here, what it never touches, and who signs.

🟢 Leader track Heads of delivery · PMO leads · sponsors No tooling required Start here
0-3 · Welcome 3-20 · What changes, what does not 20-42 · Two exercises you run 42-45 · Q&A
Part 0

How the leader track works

Six sessions, forty-five minutes each, and you never have to open a tool. This is thinking mode: the decisions a head of delivery has to make before a single PM types a prompt. Every session ends the same way, with the questions to ask your PMs on Monday - because your job here is not to build the artifacts, it is to know a good one from a fluent one and to say out loud where the line sits. Running underneath is Northwind, a fictional retailer building a governed data platform: six milestones, a slipping ingest gate, a vendor going quiet, and a sponsor who wants a date. The PM track (ten sessions, b1-b10) builds Northwind's charter, plan, risk register and status pack with AI sitting next to them. You are learning to govern that work, not to do it.

Live - presented in session Self-study - read after class ★ Try it now prompt Official sources covered
★ What you walk out with today An honest split of your function's week into artifact labour and judgment, the top three artifacts worth moving first, a drafted "we will / we will not" one-pager for your delivery function, and the three rails you will repeat until your PMs are sick of hearing them: context beats prompting, the dates come from the team, a named human signs it.
Part 1 · covers PMI's 2026 AI standard

The split: artifact labour and judgment 7 min live

Take your function's week and cut it in two. One pile is artifact labour: drafting the status pack, collating four threads into one update, chasing owners for a date, reformatting the plan for a different audience, turning ninety minutes of steerco into actions. The other pile is judgment: which trade-off to take, what to escalate and to whom, whether a date is real, which quiet person to go and talk to. AI collapses the first pile. It cannot touch the second, and the whole of this track lives on that line.

AI DRAFTS THIS YOU OWN THIS Draft the artifacts status packs, plans, minutes, reformatting Chase and collate threads, tickets and notes into one view Judgment calls trade-offs, priority, what to escalate A person signs is that date real? your name, your call Left pile: most of the calendar, almost none of the value your function is judged on. Right pile is the job. Your PMs cannot move the left pile until you name the artifacts. PMI's 2026 standard: human-in-the-loop is a value source, not a safety net.
🔍 Click to zoom - the line the whole course runs along: drafting moves, judgment does not
LiveWhy leaders get this backwards3 min

The instinct in most PMOs is to point AI at the visible pain: the status report is late, so automate the status report. That is the right pile but usually the wrong first step, because the status report is the artifact with the most political weight in your function. Get it wrong once in front of a sponsor and the whole idea is dead for a year.

The better read is to sort the labour pile by two things only: how many hours it eats across the whole function, and how much damage a bad first draft can do. Start where the hours are high and the blast radius is low. Meeting notes into actions. Reformatting the same plan for three audiences. Collating four tools into one view. Those are unglamorous, they are also where a delivery function quietly loses a day a week per PM.

Real world

The Friday that disappeared. A Northwind delivery lead spent every Friday afternoon assembling one status pack: 812k rows a night from the ingest dashboard, ticket NWD-412 sitting failed in Jira, a vendor thread with Priya N., and a spreadsheet Sofia K. kept separately. Same shape, every week, about four hours. Once the collation moved, the draft landed in under two minutes and the four hours went somewhere far more useful: a proper conversation with Priya about why VN-2291 had been open eleven days, and a quiet check on Marcus L., the only data engineer who understood the CDC job. Neither of those was ever going to come out of a template.

LiveThe four questions that sort any task2 min

When your PMs argue about whether something belongs in the labour pile, these four questions settle it faster than a framework:

  • Does it produce a document, or a decision? Documents move. Decisions do not.
  • Is the input already written down somewhere? If the raw material lives in tickets, threads and past reports, it is collation. If it lives in someone's head or in a corridor conversation, it is judgment.
  • Would a competent stranger produce roughly the same thing? If yes, it is labour. If the answer depends on knowing that Elena V. hates adjectives and Dan R. always under-promises, it is not.
  • What happens if the first draft is wrong? A wrong draft of meeting notes costs five minutes. A wrong committed date costs a quarter.

Then give them the boundary in one sentence: "AI drafts the artifact, the team supplies the facts, a named person signs it." If your PMs can repeat that back to you, most of the adoption arguments in the next six months resolve themselves without you in the room.

Three misreadings to head off early This is not a headcount argument - sell it upward as a cut and you will get the cut and lose the judgment with it. It is not a tool rollout: a licence with no context, no policy and no named owner produces generic project-management advice nobody needs. And it is not optional to be specific, because "use AI more" is not a mandate, it is a mood.
Part 2 · covers PMI adoption reporting

Where AI lands first in delivery, and the gap underneath 6 min live

Across delivery teams the same three uses show up before anything else, and in roughly the same order. They are not fashion. They are the three jobs that are pure collation, done weekly, and quietly hated by everyone who does them.

LiveThe three that arrive first3 min
UseWhat it actually replacesWhere your PMs build it
Automated status reportingA week of tickets, threads and dashboards compressed into something Elena V. can read in four minutesPM session b6, with a live scoring simulator
Task and owner suggestionsTurning ninety minutes of steerco into "who does what by when", before everyone forgetsPM session b7
Early-warning signalsSpotting that M3 has drifted from a 10 Aug gate to a 24 Aug forecast before it becomes an escalationPM sessions b4 and b5

Notice what is missing from that list: nothing decides anything. No date is committed, no message is sent, no risk is closed. That is not caution for its own sake - it is the same boundary PMI's standard draws, and it is the boundary that keeps the first three uses safe enough to actually adopt.

LiveThe adoption gap, and what it costs a PMO3 min

Public reporting on AI in project management points at a squeeze that most heads of delivery will recognise from their own week. Roughly half of respondents say AI already has an impact on how projects get managed. But only about a fifth of project managers claim good practical experience with it, and about half claim little or none.

Read the shape of that, not the decimals: high pressure from above, thin practical skill below, and nothing written down in the middle. That middle is you.

  • Shadow adoption. With no policy, the keen PMs use it anyway, on whatever account they have, with whatever data is on their screen. You find out during an audit.
  • Uneven quality. Two PMs, same programme, wildly different artifacts. One pastes raw output, one rewrites everything. Neither can tell you which is which, because there is no rubric.
  • The credibility tax. One fluent, empty status report at a steerco and your sponsor stops trusting the whole reporting line, including the reports that were fine.
  • Wasted licences. The most common outcome of a tool-first rollout is a per-seat spend with a usage chart that dies after week three.
Real world

Two Northwind PMs, one Monday. Both drafted the same weekly update. The first pasted it as it came: "data quality is improving and the team is making good progress toward M3." The second forced every claim to carry a source and got something usable: DQ pass rate 96.4% against a 99% gate, 6 of 9 source systems landing, NWD-412 still failed, so the M3 gate of 10 August now forecasts 24 August. Same tool, same week, same programme. The difference was not skill with prompts. It was that one of them had been told what a good artifact looks like, and the other had not.

It is not a training problem The reflex is to book a course. Training helps, but what your PMs are missing is not technique - it is permission and a standard. They do not know whether they may put the risk register in, they do not know who carries it if a number is wrong, and they have no agreed definition of a good draft. All three are yours to issue, all three fit on one page, and that page is what Demo 2 produces.
Part 3 · covers PMI's 2026 standard in depth

What does not change: the accountability chain 7 min live

PMI published the first edition of The Standard for Artificial Intelligence in Portfolio, Program, and Project Management in 2026. It is deliberately technology-agnostic - it names no vendor - and it is human-centered in a way that is genuinely useful to a head of delivery, because it tells you where to put the human rather than just insisting there is one.

THE CHAIN THAT DOES NOT MOVE The draft a first pass, made in minutes A named reviewer one person, not "the team" A signed artifact their name on it, numbers sourced Steerco asks who decided, and on what basis "The AI wrote it" Not a defence. It has never worked at a steerco, and using it costs more credibility than the original error did. Oversight means real intervention triggers and an escalation path, not a glance.
🔍 Click to zoom - the chain PMI's standard protects: draft, named reviewer, signed artifact, steerco
LiveThe principles that actually bite in delivery3 min

The standard sets out eight equally weighted principles. Reciting all eight in a leadership session is a good way to lose the room, so take the four that change what you do on Monday.

  • Governance and compliance. AI governance belongs inside your existing governance, not beside it. The standard asks for "clear structures, roles, authority, and guardrails" - which is a description of a RACI and a policy, not a new committee.
  • Data quality. A principle in its own right, and the one your PMs will feel first. A status pack drawn from a stale context pack is confidently wrong about last month's programme, which is far harder to catch than obvious nonsense.
  • People and culture, and ethics. If your PMs think this is a headcount exercise, the artifacts get quietly worse and nobody tells you why. And the accountable name does not move to a vendor. Ever.

The other four - strategic value, risk, stakeholder engagement, optimization and innovation - matter, and they are where sessions a4 to a6 land. Equal weighting is the point: there is no principle you get to trade away because the quarter is tight.

LiveHuman-in-the-loop as a value source, not a safety net2 min

This is the single most useful line in the standard for a head of delivery, and it is worth reading twice. The human is not in the loop to catch the machine's mistakes. The human is in the loop because that is where the judgment lives. Reviewing is not the point. Deciding is.

The practical test is whether your reviewer has anything to intervene with. A reviewer who can only say yes or no is a rubber stamp. A reviewer who knows Marcus L. is the single point of failure on the CDC job, who was in the room when decision D-14 descoped real-time POS to nightly batch, and who has the authority to move the M3 gate - that person is adding something the draft could never contain.

Real world

The review that was really a signature. A PMO put a review step in front of every AI-drafted report and called it governance. In practice one person approved eleven reports in nine minutes on a Thursday afternoon. When a wrong figure reached the sponsor, the process was technically followed and nobody could say who had actually read it. The fix was not more process: it was naming one reviewer per programme, giving them the authority to hold the report back, and writing down the two triggers that mean stop - a number with no source, and a date nobody on the team had confirmed.

Self-studyThe three rails, in leader language2 min read
  • Context beats prompting. When a draft comes back generic, the fix is almost never a cleverer instruction. It is a missing document. If your PMs are trading prompt tricks instead of loading the charter, plan and risk register, they are optimising the wrong thing - and session a2 is about what they are allowed to load.
  • The dates come from the team. Ask for a plan and you will get dates. They will be evenly spaced, plausible, and entirely fictional, because nobody was asked and nothing was capacity-checked. AI drafts the structure of a plan: milestones, gate criteria, dependencies, the questions to ask. The commitments come from the people who will be held to them.
  • A named human signs it. One accountable name per artifact, and that person read it. This is not a formality; it is exactly what PMI means by oversight with real triggers.
Repeat these until they are boring You will say these three lines in every session of this track, and you should say them in every PMO meeting for the next two quarters. Rails only work if everybody can recite them without looking.
Demo 1 of 2

Split your own function's week ★ 12 min · everyone fills this in

Not Northwind. Your function, your PMs, this quarter. The output is a shortlist of three artifacts worth moving first, and you will use it in every remaining session of this track.

List every recurring artifact or activity your delivery function produces in a normal week. Aim for ten to fifteen lines, including the ones nobody admits to, like reformatting the same update for three audiences. Estimate hours a week across the function, not per person - six PMs at ninety minutes each is a day and a half, and that is the number that gets attention.

Mark each line labour or judgment using the four questions from Part 1, then score blast radius - what a bad first draft costs. Where you cannot decide, split the line: most real activities are eighty percent collation with a judgment call at the end.

Pick your three: highest hours, lowest blast radius. Write the names down. Those are your pilot artifacts for session a6.

Artifact or activityHours a week, whole functionLabour or judgmentBlast radiusMove it?
Weekly status pack per programme9Labour, with a judgment call on the RAGHigh - the sponsor reads itYes, but not first
Steerco minutes into actions5LabourLowYes - start here
Deciding what to escalate and to whom3JudgmentHighNo, and say so out loud
Chasing owners for dates6Labour to chase, judgment to acceptMediumChase yes, accept no
Your lines, from here down----
The row that teaches the most "Chasing owners for dates" is the row worth arguing about in the room. Chasing is pure labour and should move tomorrow. Accepting the date that comes back is judgment and never moves. If your table has that distinction on it, your PMs will understand rail two without you explaining it again.
Demo 2 of 2

Draft the "we will / we will not" one-pager ★ 10 min · build your own

One page, three headings, your name at the bottom. This is the instruction your function has been missing, and it does more for quality than any amount of tool training. Session a2 hardens it into a governed setup with permissions and a red list; today you get the shape down while the split from Demo 1 is fresh.

Write the three artifacts from Demo 1 under "we will". Name them specifically - "the weekly status pack" beats "reporting".

Write "we will not" from the judgment pile, and be blunt. Anything that commits a date, closes a risk, or sends a message to a human outside the team belongs here.

Name the signer: one accountable person per artifact type, by role and by name. If you cannot name them, the artifact is not ready to move.

Add the two intervention triggers that mean stop and escalate, then circulate it as a draft and ask your PMs what is missing. The ones already using AI quietly will tell you more in ten minutes than a survey will.

★ The one-pager skeleton - edit hard, then sign itAI IN DELIVERY - [your function], v0.1, [date] Owner: [your name, head of delivery] WE WILL use AI to draft - Meeting notes into actions with an owner and a date on every line - The first pass of the weekly status pack, from the programme's own documents - Reformatting an approved plan for a different audience - Early-warning flags on slip and risk, raised to a human, never actioned alone WE WILL NOT use AI to - Commit or change a date. Dates come from the team who will be held to them. - Open, close or downgrade a risk - Send anything to a client, a vendor or a sponsor without a named human sending it - Decide scope, priority or what gets escalated WHO SIGNS - Status pack: the programme PM, by name, every week - Escalation or anything to a sponsor: me - Anything a client or auditor reads: me, and only after I have read it STOP AND ESCALATE IF - A draft contains a number with no traceable source - A draft contains a date nobody on the delivery team has confirmed WHAT WE LOG Which artifacts were AI-drafted, who reviewed, who signed. Reviewed at the monthly PMO.
The version that changes nothingOur PMO is committed to embracing AI to drive efficiency and innovation across delivery. Teams are encouraged to use AI responsibly and in line with company policy, with appropriate human oversight at all times. We will review as the technology matures.
Why the second one fails Read it as a PM on Monday morning. Can you put the risk register in? Unclear. Can you send an AI-drafted note to the vendor? Unclear. Who carries it if a number is wrong? Unclear. "Appropriate human oversight" is exactly the perfunctory review PMI's standard warns against, dressed up as a policy. Every line of the first one either permits something, forbids something, or names somebody - which is what makes it usable.
Real world

The line that saved a steerco. One head of delivery added a single rule: any number in a draft must carry its source. The next week a Northwind draft claimed data quality had "improved significantly" and could not produce a ticket or a metric behind it - it had smoothed a wobble into a trend. The PM caught it in about ten seconds, went back to the actual dashboard, and reported the truth: 96.4% against a 99% gate, flat for three weeks, so the M3 gate was at risk and R-11 needed raising. Elena V. got a worse number and a better relationship. That is the difference between a rule and a hope.

Homework

Before session a2 ◐ 30-45 min total

Source material

Official sources covered

Taught from PMI's public standards and AI guidance plus the delivery canon. Certification (PMP, PMI-ACP) and the full normative text of the AI standard stay with PMI. This session covers:

PMI - Standard for AI in Portfolio, Program and Project Management (2026)Part 3 · principles, governance inside GRC, human-in-the-loop as a value source
PMI - AI in project management, adoption reportingPart 2 · where AI lands first, the practitioner skill gap and what it costs
PMBOK Guide 7th edition - principles and performance domainsPart 1 · stewardship and measurement under the labour / judgment split
PM session b1 - the programme workspace and context packPart 0 · what your PMs build behind this track
Check yourself

Three questions before you go 🎯 ◐ 90 seconds

1 · You have hours-per-week and blast-radius scores for every artifact your function produces. Which three do you move first?

High hours prove the value, low blast radius keeps the first mistake survivable. The sponsor-facing status pack is the right pile but a risky first step: get it wrong once in front of Elena V. and the idea is dead for a year.

2 · A PMO puts a review step in front of every AI-drafted report. One person approved eleven reports in nine minutes. What did PMI's standard actually ask for?

Human-in-the-loop is framed as a value source, not a safety net. A reviewer who can only say yes is a rubber stamp. Name one reviewer, give them authority to hold the artifact back, and write down the triggers that mean stop.

3 · Your PMs ask for a plan and get back a clean six-milestone schedule with dates. What is the leader instruction?

Rail two. Nobody was asked and nothing was capacity-checked, so the dates are fiction with good typography. AI drafts the structure - milestones, gate criteria, dependencies, the questions to ask. Commitments come from the team who will be held to them.

Leader session 1 cheat sheet · pin this

The splitArtifact labour moves. Judgment does not. Everything else in this track sits on that line.
Sorting testDocument or decision? Is the input written down? Would a stranger produce the same? What if the draft is wrong?
Move firstHigh hours across the function, low blast radius. Not the sponsor-facing pack on day one.
Lands first everywherestatus reporting, task and owner suggestions, early-warning signals. None of them decide anything.
The gapRoughly half say AI already affects PM. About a fifth of PMs have real experience. The middle is you.
PMI 2026Eight equal principles, five performance domains. Governance inside existing GRC, not beside it.
OversightReal intervention triggers and authority to stop. Human-in-the-loop is a value source, not a safety net.
The three railsContext beats prompting. The dates come from the team. A named human signs it.