How the leader track works
Six sessions, forty-five minutes each, and you never have to open a tool. This is thinking mode: the decisions a head of delivery has to make before a single PM types a prompt. Every session ends the same way, with the questions to ask your PMs on Monday - because your job here is not to build the artifacts, it is to know a good one from a fluent one and to say out loud where the line sits. Running underneath is Northwind, a fictional retailer building a governed data platform: six milestones, a slipping ingest gate, a vendor going quiet, and a sponsor who wants a date. The PM track (ten sessions, b1-b10) builds Northwind's charter, plan, risk register and status pack with AI sitting next to them. You are learning to govern that work, not to do it.
The split: artifact labour and judgment 7 min live
Take your function's week and cut it in two. One pile is artifact labour: drafting the status pack, collating four threads into one update, chasing owners for a date, reformatting the plan for a different audience, turning ninety minutes of steerco into actions. The other pile is judgment: which trade-off to take, what to escalate and to whom, whether a date is real, which quiet person to go and talk to. AI collapses the first pile. It cannot touch the second, and the whole of this track lives on that line.
LiveWhy leaders get this backwards3 min▶
The instinct in most PMOs is to point AI at the visible pain: the status report is late, so automate the status report. That is the right pile but usually the wrong first step, because the status report is the artifact with the most political weight in your function. Get it wrong once in front of a sponsor and the whole idea is dead for a year.
The better read is to sort the labour pile by two things only: how many hours it eats across the whole function, and how much damage a bad first draft can do. Start where the hours are high and the blast radius is low. Meeting notes into actions. Reformatting the same plan for three audiences. Collating four tools into one view. Those are unglamorous, they are also where a delivery function quietly loses a day a week per PM.
The Friday that disappeared. A Northwind delivery lead spent every Friday afternoon assembling one status pack: 812k rows a night from the ingest dashboard, ticket NWD-412 sitting failed in Jira, a vendor thread with Priya N., and a spreadsheet Sofia K. kept separately. Same shape, every week, about four hours. Once the collation moved, the draft landed in under two minutes and the four hours went somewhere far more useful: a proper conversation with Priya about why VN-2291 had been open eleven days, and a quiet check on Marcus L., the only data engineer who understood the CDC job. Neither of those was ever going to come out of a template.
LiveThe four questions that sort any task2 min▶
When your PMs argue about whether something belongs in the labour pile, these four questions settle it faster than a framework:
- Does it produce a document, or a decision? Documents move. Decisions do not.
- Is the input already written down somewhere? If the raw material lives in tickets, threads and past reports, it is collation. If it lives in someone's head or in a corridor conversation, it is judgment.
- Would a competent stranger produce roughly the same thing? If yes, it is labour. If the answer depends on knowing that Elena V. hates adjectives and Dan R. always under-promises, it is not.
- What happens if the first draft is wrong? A wrong draft of meeting notes costs five minutes. A wrong committed date costs a quarter.
Then give them the boundary in one sentence: "AI drafts the artifact, the team supplies the facts, a named person signs it." If your PMs can repeat that back to you, most of the adoption arguments in the next six months resolve themselves without you in the room.
Where AI lands first in delivery, and the gap underneath 6 min live
Across delivery teams the same three uses show up before anything else, and in roughly the same order. They are not fashion. They are the three jobs that are pure collation, done weekly, and quietly hated by everyone who does them.
LiveThe three that arrive first3 min▶
| Use | What it actually replaces | Where your PMs build it |
|---|---|---|
| Automated status reporting | A week of tickets, threads and dashboards compressed into something Elena V. can read in four minutes | PM session b6, with a live scoring simulator |
| Task and owner suggestions | Turning ninety minutes of steerco into "who does what by when", before everyone forgets | PM session b7 |
| Early-warning signals | Spotting that M3 has drifted from a 10 Aug gate to a 24 Aug forecast before it becomes an escalation | PM sessions b4 and b5 |
Notice what is missing from that list: nothing decides anything. No date is committed, no message is sent, no risk is closed. That is not caution for its own sake - it is the same boundary PMI's standard draws, and it is the boundary that keeps the first three uses safe enough to actually adopt.
LiveThe adoption gap, and what it costs a PMO3 min▶
Public reporting on AI in project management points at a squeeze that most heads of delivery will recognise from their own week. Roughly half of respondents say AI already has an impact on how projects get managed. But only about a fifth of project managers claim good practical experience with it, and about half claim little or none.
Read the shape of that, not the decimals: high pressure from above, thin practical skill below, and nothing written down in the middle. That middle is you.
- Shadow adoption. With no policy, the keen PMs use it anyway, on whatever account they have, with whatever data is on their screen. You find out during an audit.
- Uneven quality. Two PMs, same programme, wildly different artifacts. One pastes raw output, one rewrites everything. Neither can tell you which is which, because there is no rubric.
- The credibility tax. One fluent, empty status report at a steerco and your sponsor stops trusting the whole reporting line, including the reports that were fine.
- Wasted licences. The most common outcome of a tool-first rollout is a per-seat spend with a usage chart that dies after week three.
Two Northwind PMs, one Monday. Both drafted the same weekly update. The first pasted it as it came: "data quality is improving and the team is making good progress toward M3." The second forced every claim to carry a source and got something usable: DQ pass rate 96.4% against a 99% gate, 6 of 9 source systems landing, NWD-412 still failed, so the M3 gate of 10 August now forecasts 24 August. Same tool, same week, same programme. The difference was not skill with prompts. It was that one of them had been told what a good artifact looks like, and the other had not.
What does not change: the accountability chain 7 min live
PMI published the first edition of The Standard for Artificial Intelligence in Portfolio, Program, and Project Management in 2026. It is deliberately technology-agnostic - it names no vendor - and it is human-centered in a way that is genuinely useful to a head of delivery, because it tells you where to put the human rather than just insisting there is one.
LiveThe principles that actually bite in delivery3 min▶
The standard sets out eight equally weighted principles. Reciting all eight in a leadership session is a good way to lose the room, so take the four that change what you do on Monday.
- Governance and compliance. AI governance belongs inside your existing governance, not beside it. The standard asks for "clear structures, roles, authority, and guardrails" - which is a description of a RACI and a policy, not a new committee.
- Data quality. A principle in its own right, and the one your PMs will feel first. A status pack drawn from a stale context pack is confidently wrong about last month's programme, which is far harder to catch than obvious nonsense.
- People and culture, and ethics. If your PMs think this is a headcount exercise, the artifacts get quietly worse and nobody tells you why. And the accountable name does not move to a vendor. Ever.
The other four - strategic value, risk, stakeholder engagement, optimization and innovation - matter, and they are where sessions a4 to a6 land. Equal weighting is the point: there is no principle you get to trade away because the quarter is tight.
LiveHuman-in-the-loop as a value source, not a safety net2 min▶
This is the single most useful line in the standard for a head of delivery, and it is worth reading twice. The human is not in the loop to catch the machine's mistakes. The human is in the loop because that is where the judgment lives. Reviewing is not the point. Deciding is.
The practical test is whether your reviewer has anything to intervene with. A reviewer who can only say yes or no is a rubber stamp. A reviewer who knows Marcus L. is the single point of failure on the CDC job, who was in the room when decision D-14 descoped real-time POS to nightly batch, and who has the authority to move the M3 gate - that person is adding something the draft could never contain.
The review that was really a signature. A PMO put a review step in front of every AI-drafted report and called it governance. In practice one person approved eleven reports in nine minutes on a Thursday afternoon. When a wrong figure reached the sponsor, the process was technically followed and nobody could say who had actually read it. The fix was not more process: it was naming one reviewer per programme, giving them the authority to hold the report back, and writing down the two triggers that mean stop - a number with no source, and a date nobody on the team had confirmed.
Self-studyThe three rails, in leader language2 min read▶
- Context beats prompting. When a draft comes back generic, the fix is almost never a cleverer instruction. It is a missing document. If your PMs are trading prompt tricks instead of loading the charter, plan and risk register, they are optimising the wrong thing - and session a2 is about what they are allowed to load.
- The dates come from the team. Ask for a plan and you will get dates. They will be evenly spaced, plausible, and entirely fictional, because nobody was asked and nothing was capacity-checked. AI drafts the structure of a plan: milestones, gate criteria, dependencies, the questions to ask. The commitments come from the people who will be held to them.
- A named human signs it. One accountable name per artifact, and that person read it. This is not a formality; it is exactly what PMI means by oversight with real triggers.
Split your own function's week ★ 12 min · everyone fills this in
Not Northwind. Your function, your PMs, this quarter. The output is a shortlist of three artifacts worth moving first, and you will use it in every remaining session of this track.
List every recurring artifact or activity your delivery function produces in a normal week. Aim for ten to fifteen lines, including the ones nobody admits to, like reformatting the same update for three audiences. Estimate hours a week across the function, not per person - six PMs at ninety minutes each is a day and a half, and that is the number that gets attention.
Mark each line labour or judgment using the four questions from Part 1, then score blast radius - what a bad first draft costs. Where you cannot decide, split the line: most real activities are eighty percent collation with a judgment call at the end.
Pick your three: highest hours, lowest blast radius. Write the names down. Those are your pilot artifacts for session a6.
| Artifact or activity | Hours a week, whole function | Labour or judgment | Blast radius | Move it? |
|---|---|---|---|---|
| Weekly status pack per programme | 9 | Labour, with a judgment call on the RAG | High - the sponsor reads it | Yes, but not first |
| Steerco minutes into actions | 5 | Labour | Low | Yes - start here |
| Deciding what to escalate and to whom | 3 | Judgment | High | No, and say so out loud |
| Chasing owners for dates | 6 | Labour to chase, judgment to accept | Medium | Chase yes, accept no |
| Your lines, from here down | - | - | - | - |
Draft the "we will / we will not" one-pager ★ 10 min · build your own
One page, three headings, your name at the bottom. This is the instruction your function has been missing, and it does more for quality than any amount of tool training. Session a2 hardens it into a governed setup with permissions and a red list; today you get the shape down while the split from Demo 1 is fresh.
Write the three artifacts from Demo 1 under "we will". Name them specifically - "the weekly status pack" beats "reporting".
Write "we will not" from the judgment pile, and be blunt. Anything that commits a date, closes a risk, or sends a message to a human outside the team belongs here.
Name the signer: one accountable person per artifact type, by role and by name. If you cannot name them, the artifact is not ready to move.
Add the two intervention triggers that mean stop and escalate, then circulate it as a draft and ask your PMs what is missing. The ones already using AI quietly will tell you more in ten minutes than a survey will.
The line that saved a steerco. One head of delivery added a single rule: any number in a draft must carry its source. The next week a Northwind draft claimed data quality had "improved significantly" and could not produce a ticket or a metric behind it - it had smoothed a wobble into a trend. The PM caught it in about ten seconds, went back to the actual dashboard, and reported the truth: 96.4% against a 99% gate, flat for three weeks, so the M3 gate was at risk and R-11 needed raising. Elena V. got a worse number and a better relationship. That is the difference between a rule and a hope.
Before session a2 ◐ 30-45 min total
- Finish the split table for your real function, and get at least two of your PMs to challenge the hours - your estimate is probably low.
- Circulate the one-pager as an explicit draft. Ask one question with it: "what is on this list that you have already been doing anyway?" Expect an uncomfortable answer, and treat it as data rather than a disciplinary matter.
- Find out who in your organisation actually owns the answer to "what may we put into an AI tool". If nobody knows, that is a genuine finding, and it is exactly where a2 starts. Then ask your PMs these three on Monday: which artifact eats your week? What would you need before you trusted an AI-drafted version of it? Who do you think signs it today?
Official sources covered
Taught from PMI's public standards and AI guidance plus the delivery canon. Certification (PMP, PMI-ACP) and the full normative text of the AI standard stay with PMI. This session covers:
Three questions before you go 🎯 ◐ 90 seconds
1 · You have hours-per-week and blast-radius scores for every artifact your function produces. Which three do you move first?
High hours prove the value, low blast radius keeps the first mistake survivable. The sponsor-facing status pack is the right pile but a risky first step: get it wrong once in front of Elena V. and the idea is dead for a year.
2 · A PMO puts a review step in front of every AI-drafted report. One person approved eleven reports in nine minutes. What did PMI's standard actually ask for?
Human-in-the-loop is framed as a value source, not a safety net. A reviewer who can only say yes is a rubber stamp. Name one reviewer, give them authority to hold the artifact back, and write down the triggers that mean stop.
3 · Your PMs ask for a plan and get back a clean six-milestone schedule with dates. What is the leader instruction?
Rail two. Nobody was asked and nothing was capacity-checked, so the dates are fiction with good typography. AI drafts the structure - milestones, gate criteria, dependencies, the questions to ask. Commitments come from the team who will be held to them.