Six working labs. One specimen AI system. You make every call — and every decision seals into a SHA-256 evidence chain. Your certificate is an audit trail, not an attendance letter.
Write the definition of success for TS-7, then watch weak criteria fail under attack.
Before judging the machine, measure the human. Your own performance becomes the baseline.
Score the agent's outputs against a rubric, then see how far your scores drift from the panel's.
Five candidate systems, five dossiers, one question each: does it enter production?
The cleared system drifts live. Catch the breach, then execute the rollback in the right order.
Examiners question everything you did. Your only defense is the evidence chain you built.
Your sealed completion evidence: every decision, hashed and chained in order.
Institutions have policies, frameworks, and committees. What they do not have is people who have ever actually written a defensible success criterion, chaired a clearance decision, or executed a rollback while the metric was falling. The failure statistics are the bill for that gap.
Participants perform the six disciplines of AI governance on a live specimen: writing success criteria that survive attack, measuring the human baseline, anchoring an evaluation, holding the production gate, executing a rollback under time pressure, and answering an examiner from the record. Skills practiced, not described.
A cohort's choices reveal the institution's real risk posture. The room that clears the candidate with the self-authored evaluation set, or fumbles the rollback order, has just shown leadership exactly where the next finding will come from, before an examiner finds it first.
Every decision each participant makes is sealed in real time into a SHA-256 hash chain. The credential is not an attendance certificate; it is a tamper-evident record of what that person actually did, verifiable by anyone and falsifiable by no one. Training that can be demonstrated, not merely asserted.
The primary audience. They are expected to challenge and approve AI systems they have never governed. They know the regulations; they have never sat in the seat. The Proving Ground puts them in it, with consequences.
The builders who will one day face the gate. Running the labs before that day means evidence gets designed in from the first sprint rather than reconstructed for the audit.
A compressed half-day session, Labs IV through VI. The people who sign the attestations experience what "we have governance" must actually mean before they put their names to it.
Licensed cohort delivery for firms who open governance engagements with their own clients. The labs create the shared vocabulary; the engagement installs the standard.
It is not a technical machine learning course, and it will not teach anyone to build or tune a model. It is not a compliance module to click through at a desk. It is designed for cohorts, because half the value is the argument that breaks out when two risk officers vote differently on the same dossier.
Six working labs, facilitated liveOn-site or remote, 90 to 120 minutes end to end
12 to 20 participantsMixed first and second line cohorts produce the strongest sessions
Labs IV to VI, half dayFor risk committees and accountable executives
Runs entirely in the browserNo installation, no data leaves the participant's machine
Success criteria survive scrutiny only when they name a number, a dataset, and a consequence
The build team wants TS-7 in production next quarter. Before anyone writes a line of governance, someone has to say what success means, in numbers, with a name attached. That someone is you.
First, learn to spot the criteria that fail under attack. Then sign your own.
No AI performance claim is testable until the human process it replaces has been measured
Your charter promises 96% recall "against the baseline." Most institutions skip this step and compare their AI to a number nobody ever measured. Not here. You will work six live alerts from the triage queue exactly as the surveillance desk does. Your accuracy becomes the sealed baseline.
Unanchored expert scoring produces a spread no audit can accept
TS-7 has processed its first alerts. Rate each disposition and rationale from 1 to 5. When you finish, your scores are laid against the review panel's, and the spread between honest experts is the whole lesson.
A single unmet gate is a hold, regardless of how strong the rest of the dossier reads
You now chair the clearance board. Each candidate arrives with a dossier summarizing its posture against the five production gates. Some look ready and are not. One looks ordinary and is ready. The record will show how you voted.
Rollback capability is only real once it has been executed under time pressure
Your charter set the retirement threshold at 90% recall. The telemetry below is live. Flag the breach the moment recall crosses the line, then execute the rollback protocol in the canonical order: freeze the queue, reroute to the desk, preserve the evidence snapshot, notify the accountable owner. In that order. Wrong order in a real incident destroys the evidence you need for Lab VI.
Examinations are withstood from the record, not from the room
Internal audit and a supervisory examiner sit across the table. They are not hostile. They are worse: they are thorough. Each question has one answer grounded in your evidence chain and two that sound fine in a meeting. Meetings are not the standard here.
Not a certificate of attendance. A verifiable chain of what you actually did.