How every lab runs
Demo, build, debrief
A demonstration of one workflow on realistic practice data, an extended hands-on build block, and shared results. Everyone works in the Claude desktop app from a single working folder; a browser fallback covers anyone whose IT blocks the install.
The data spine
Export from the protected environment with identifying details removed, do the substantive work in frontier AI tools, then return results to your own systems where the details reassemble. Safe data handling gets practiced every week, not taught once. The specific anonymization technique is a decision we make together; an options sheet accompanies this map.
Buddy pairs, real backlogs
Two participants per organization work between sessions to finish real deliverables from their own ship list: the 3 to 5 things they actually need to produce this quarter, collected before Lab 1.
The seven labs
First Safe Win
The full data round-trip on a real writing task. Everyone leaves with one usable output and the safety habit started.
AI Mindset & Foundations · Governance & RiskWhat happens in this lab
Agency framing, not a tool pitch. The fears in the room (accuracy, voice, data, jobs) get named directly, each one mapped to the specific practice that addresses it.
The round-trip end to end on a mock email-thread export: swap names for tokens, drop the file into the working folder, draft a board update with a narrated prompt, catch a planted contradiction in verification, return and re-identify. That motion is the whole program.
Pairs run the identical exercise on the mock pack first (completion, not quality), then aim the same loop at the smallest writing item on their own ship list.
Two or three pairs show what they made, struggles included. Each pair states its homework commitment out loud.
All three levels run the round-trip on real work in their own environment; everyone also gathers 3 to 4 public org documents for next week's build.
Writing + the Org Brain
Each pair builds a permanent project loaded with their organization's identity and voice, then produces real writing through it: grant applications, donor communications, memos.
Quality & AccuracyWhat happens in this lab
Homework debrief: two pairs show the writing they shipped, including what broke.
An org brain built live: a Project loaded with mission, writing samples, and outcomes, then voice instructions written by describing the samples and editing by hand. Payoff: a mock RFP becomes a compliance matrix and a first draft. The same prompt runs once without the Project; the side-by-side makes the argument.
Pairs build their own org brain from material they brought (a fictional starter pack covers anyone who arrived empty-handed), then aim it at a writing item from their ship list. Floor: Project exists, one draft out. Ceiling: voice tuned through two revisions.
Before/after read alouds. Where did it still not sound like you, and what instruction fixes it?
One more writing rep, a full shipped deliverable, or a second voice profile (funder-formal vs. community newsletter).
Board-Ready Reporting + the Second Chair
Messy inputs turned into polished narrative summaries, plus a critique pass that becomes a required step in everything that follows.
Reporting & Narrative BuildingWhat happens in this lab
Homework debrief, plus one question: did anyone's Project surprise them?
Real reporting inputs (meeting notes, a partial spreadsheet, three emails) become a plausible board summary. Then the move this lab exists for: a second prompt played as a hostile board member attacks the draft and catches planted weaknesses on screen, including an unsourced statistic and the sentence that would hurt if quoted out of context. One review, two protections: accuracy and communication risk. Plausible is the danger zone.
Pairs run the same arc on the mock pack, then on a real reporting item. The adversarial prompt and a source-accountability checklist get saved into every Project before the lab ends.
Each pair shares the single best catch its critique made. Best catches are the proof the step earns its time.
Critique something already in your outbox, ship a full report section through the arc, or tune the critique into your org's own failure-mode profile (the seed of the Lab 6 agent).
Data: Clean & Extract
Spreadsheet cleanup and insight extraction on program data through the round-trip, with an auditable change log.
Metrics & Data ReviewWhat happens in this lab
The week's best adversarial catches, fast.
A 58-row mock grants export carrying every real disease: one org spelled three ways, blanks, amounts as text, three date formats, a duplicate. First a green/yellow/red classification of what can travel and how, then two separate passes in Cowork: clean with a row-by-row change log, then insights (totals, outliers, what looks wrong). Verification is specific: spot-check five rows, reconcile totals, calculator-test the findings. The change log is the star, not the clean file.
Pairs run the mock dataset through the full arc, then write their own cleanup recipe: a saved, named prompt for the recurring mess they actually get every month.
Recipes read aloud; the room steals freely. The same five messes live in every org.
Run the recipe on real data, deliver a full clean-plus-insights memo, or chain recipe into a standing monthly review.
Numbers to Narrative to Deck
Analysis becomes an impact story with verified figures, then a presentation. Includes visual judgment: what to chart, what to say in a sentence.
Metrics + Reporting, combinedWhat happens in this lab
One real cleaned-data win from the homework, told by the pair that had it.
Three moves on the cleaned data. Numbers to narrative: an impact story for a named audience where every figure traces to a source row, then an adversarial pass played as a funder who suspects cherry-picking. Visual judgment: chart or sentence, one message per chart, honest axes, and Claude proposes the chart while the human edits the judgment. Narrative to deck: slide titles as claims, numbers in speaker notes, generated in minutes. The deck is fast because the thinking was already done.
Pairs work the ladder on their own Lab 4 output (or the reference data): narrative first, deck outline second, one chart specified per the visual guide.
Two pairs present one slide each, 60 seconds, as if to their real audience. What would the hostile funder ask?
Finish the deck (the second rep that makes it stick), build the full pack for a real upcoming meeting, or template a reusable board-pack generator.
Agents & the Wider Toolbox
Each pair turns their critique pass into a standing reviewer agent, runs one supervised research-agent task, and learns which tools complement which.
Workflow Design · Governance appliedWhat happens in this lab
One deck from the homework, 60 seconds.
Two moves. The critique prompt becomes a standing agent: named, saved, triggered on demand, with org-specific failure modes. A prompt is a thing you remember to do; an agent is a thing that happens. Then the wider toolbox, demonstrated live by the facilitator: an agentic research run (a funder-landscape scan) launches while the model-selection map gets drawn (Copilot for inside the container, Claude for the work, research agents for the outside world, NotebookLM for your own documents and notes). The scan returns and gets verified live, with the bias question asked out loud: whose perspective did these sources systematically miss?
Pairs build and test their critique agent on a Lab 5 artifact, writing failure modes from real experience (what does YOUR board push on?). The research motion runs on the facilitator's screen with pairs directing it, unless Wells Fargo opts into participant accounts; three claims get verified before anything is repeated.
Each pair states one sentence: the next thing we'd automate and the tool we'd use. Those sentences come back in Lab 7.
Use the agent on a real deliverable, write up a verified research answer, or build a second agent or prospect pipeline (the champion track that shapes what comes after the pilot).
Build Session & 90-Day Roadmap
Pairs assemble, document, and present the operational workflow they built across the program, then set 90-day priorities.
Workflow Design · Build Session & RoadmapWhat happens in this lab
The frame shifts from learning to owning. The automation sentences from Lab 6 get read back, verbatim.
One complete workflow document, filled in honestly: trigger, steps, the Project artifacts that make it run, verification gates, and real before/after time costs. A workflow nobody wrote down is a demo, not a workflow.
Assembly, not construction: pairs document the workflow from pieces already built across the homework arc (org brain, critique agent, cleanup recipe, narrative-and-deck motion). The floor version, round-trip plus one motion, is a legitimate, complete outcome. One test: could a colleague run this from the document alone?
Two minutes per pair, hard-timed, four beats: what we built, what it replaced, what it costs us now, one number. Everyone else writes one steal per presentation. No critique today; two minutes of being proud of work is part of the design.
90-day commitments written and spoken: the workflow runs at least monthly, one optional Build Ahead item, and where help lives when something breaks. No homework; the 90-day plan is the homework.
Where the other scope topics went
- Quality & Accuracy became a verification ritual inside every lab rather than a single session, deepening each week: source-tracing, the critique pass, spot-checks and reconciliation on data.
- Governance & Risk Awareness became the data round-trip itself, practiced weekly from Lab 1, plus a simple classification habit for deciding what data can travel and how.
- Bias awareness and communication risk management ride the critique pass: every adversarial review checks for missing perspectives, slanted sources, and the sentence that could be quoted out of context, from Lab 3 onward and explicitly in the agents lab.
- Workflow Design became the between-session arc: each week, buddy pairs complete one stage of their own workflow, so the final session is assembly rather than a cold start.
Homework at three levels, chosen weekly
Each week's between-session work comes in three levels with a stated output. Pairs choose based on their time and experience, and can choose differently each week.
Keep Pace
One rep of the week's motion on a small real item. The habit survives the week.
Ship It
A full ship-list deliverable through the motion plus verification. Real work out the door.
Build Ahead
Stretch builds toward a tool or agent set: showcase material for the final session and the bridge to Phase 2.
Before Lab 1
- A short intake gathers current tools, comfort levels, and the 3 to 5 real items each pair needs to produce this quarter.
- A setup guide walks every participant through installing the Claude desktop app and proving the file workflow end to end, with a browser fallback documented for anyone whose IT blocks the install. Wells Fargo participants may need IT approval, so this goes out the day the cohort is confirmed.
- A technical pre-flight confirms working access for every participant so no session is lost to login problems.
What stays the same
Seven sessions over seven weeks at 90 minutes each, live and recorded, two cohorts of 25, and every participant finishing with at least one operational workflow. The structure above is a starting point; the pieces are built to move once we hear what the cohort actually needs.