Mike's Checks/gpt-5.6-sol/05 showrunner
05 showrunner
gpt-5.6-solCodex CLIhigh effortrun 22 Aug 2026
▸Instructions — the case's current instructions; none were saved with this result
Hi, I need to create a modified Claude code for beginners session, like run sheet basically, run of show, that has the following modifications. So I'm going to upload a few different examples of things I've done in the past, just so you have them to work from. But basically, there's a few modifications for this one which make it unique. One is that this is only using Claude Cowork, so we should not be using Claude Code, either in the desktop app or in the terminal, because this team is non-technical. The team are all working in financial services slash accounting for a education company. So what kind of workplace education company called Brightmoor Education. So the sorry that spelled b the uh team is very busy and um doesn have a lot of like time to explore speculative use cases of ai they want this to be incredibly practical and based on their existing workflows. I'm uploading some transcripts from a call we had with them and also just some notes on what I think could work. Specifically, they're interested in... creating the dashboards and specifically they mean just making the information look visibly better like like creating a html dashboard in a nice kind of style from from existing say data that contained in Excel They also interested in doing financial analysis So more of a, you know, like, I guess, a cloud coworker would use Python to do the analysis and then maybe output the analysis in Excel or the input would be in Excel. And then the third thing was like making a PowerPoint where they update the PowerPoint template automatically using Claude Cowork so that they can do kind of weekly or monthly presentations on the data. So that's the goal, is to kind of make it very focused on that. I still want to keep the way that we introduce Claude Cowork, where we explain to it that the difference between Claude Cowork and ChatGPT is that it has access to the file system. It can run code. I think that important But maybe the ASCII image idea that I normally do where I ask it to create an ASCII image based on my location isn like that useful I think we need to make it more practical for them So maybe we use an Excel implementation instead. The assumption is... oh, so they use Snowflake, by the way, for all their data, and that is powered by NetSuite, which is our accounting software. I guess that's by Oracle. So when we create synthetic data, we'll be, for these tasks that we're going to demo, we'll be focused specifically on solving that, solving those problems. So like the shape of the tasks, the shape of the data should be like NetSuite shaped, if that makes sense. So make it very tight run down, make it very condensed and practical. No real like wider theory or like, hey, we're excited about AI things. I think these guys just want to get stuff done.
So yeah, let me know what the run should be.
some very specific asks in here
- definitely non-technical (claude cowork only + claude in excel)
- core focus on accounting/finance function
- use cases to demo: creating a dashboard, doing financial analysis, making a powerpoint
heavy focus on practical stuff rather than theory - they're just super busy and want to get stuff done, have heard from friends at other companies you can do it with claude
For accounting: put financials together, doing reconciliations, identifying trends in data sets, putting together repeatable compliance reports For planning & forecast: build financial models (make assumptions and pull & clean data from snowflake)create dashboards, and publish to executivesAutomate charts on slides
Answer
Claude Cowork for Finance: Brightmoor Education
Recommended run of show
Format: 3.5-hour, hands-on beginner workshop
Audience: 20–30 people across accounting, payroll, bursar, and strategic finance
Tools: Claude Cowork and Claude in Excel only
Working scenario: A monthly operating review built from a synthetic, NetSuite-shaped export from Snowflake
End products: A checked analysis workbook, an executive HTML dashboard, and an updated PowerPoint deck
The session should follow one continuous workflow rather than three unrelated demos:
NetSuite-shaped Snowflake export → Excel checks and analysis → executive dashboard → monthly PowerPoint
This mirrors Brightmoor's actual work and keeps the session focused on getting a recurring finance deliverable done. It also makes the central lesson easy to remember: Claude can do the file work, but the finance owner still verifies the numbers.
What participants should be able to do by the end
- Give Claude Cowork a folder of finance files and ask it to inspect, analyze, and create new deliverables without using a terminal or writing code.
- Use Claude in Excel to understand a workbook, add calculations, investigate variances, and build visible checks.
- Turn validated finance data into a clean, filterable HTML dashboard for leadership.
- Use an existing PowerPoint template and chart as a style reference, then ask Cowork to update the deck with the new month's numbers and commentary.
- Re-run the workflow next month while preserving source files and checking every output against control totals.
Detailed facilitator run sheet
Times below are elapsed so the schedule works regardless of the event's start time.
| Time | Segment | What happens | Participant output |
|---|---|---|---|
| 0:00–0:10 | Start with the job | Set the goal: reduce the manual work between a NetSuite/Snowflake export and a leadership-ready monthly report. Show the three finished outputs. No AI trends, icebreaker, or speculative use cases. | Clear view of the finish line |
| 0:10–0:25 | First useful result in Excel | Open the synthetic monthly workbook in Excel. Use Claude in Excel to explain the workbook, calculate an actual-versus-budget variance, and flag the five largest exceptions. Replace the usual ASCII warm-up with a real finance task. | A useful result in the first 15 minutes |
| 0:25–0:35 | Cowork in plain English | Explain only what is needed: unlike a typical ChatGPT-style conversation that mainly returns an answer in the chat, Cowork can work with files in an approved folder, run code behind the scenes, and create Excel, HTML, and PowerPoint files. Demonstrate selecting the workshop folder and requiring new versions rather than overwriting source files. | A safe working folder and a mental model |
| 0:35–1:20 | Build 1: financial analysis and reconciliation | In Cowork, inspect the export, map accounts, calculate revenue and EBITDA, compare actuals with budget, split payroll and non-payroll costs, identify campus trends, and reconcile totals. Save the result as a new Excel workbook with a Checks sheet. | Brightmoor_Monthly_Analysis.xlsx |
| 1:20–1:30 | Verify and debrief | Tie source totals to the analysis, inspect unmapped accounts and duplicates, and review the largest variances. Ask participants to fix one issue through a plain-language follow-up prompt. | A checked workbook and exception log |
| 1:30–1:40 | Break | Ten-minute break. | — |
| 1:40–2:25 | Build 2: executive dashboard | Give Cowork the checked workbook and ask for a standalone HTML dashboard. Add leadership-friendly KPI cards, campus and period filters, actual-versus-budget charts, an EBITDA bridge, and short callouts. Participants make two visible changes, such as changing a chart to horizontal bars and adding a filter. | Brightmoor_Executive_Dashboard.html |
| 2:25–3:05 | Build 3: refresh the monthly PowerPoint | Give Cowork the checked workbook, a six-slide monthly-review template, and one completed chart as a style reference. Update charts, tables, titles, and draft commentary. Demonstrate the high-value shortcut Brightmoor requested: mirror one approved chart style across the next five slides. | Brightmoor_Monthly_Review_Updated.pptx |
| 3:05–3:20 | Make it repeatable | Ask Cowork to write a one-page monthly refresh checklist: which files to replace, what prompt to reuse, what must tie out, and what requires human review. Demonstrate a new-month rerun using a second synthetic export. | Monthly_Refresh_Checklist.docx or .md |
| 3:20–3:30 | Close on the outputs | Show the three artifacts side by side. Recap the working pattern below and take only workflow-specific questions. Give participants the prompt card and checklist. | A reusable process, not just a demo |
The agenda provides about two hours of hands-on work. The facilitator should type the first prompt with the group, then pause after each major action so beginners do not fall behind.
The working pattern to repeat in every exercise
Use a finance-specific version of the existing build loop:
- Brief: State the business outcome, audience, files, and required output.
- Inspect: Ask Claude to describe the data, assumptions, and proposed checks before changing anything.
- Make: Let Claude create a new version of the deliverable.
- Verify: Tie totals to the source, review exceptions, and make targeted corrections.
- Tune: Change presentation, commentary, filters, or chart types only after the numbers pass checks.
The order matters: verify before beautifying.
Demo prompts
These are deliberately written in normal workplace language. Participants should edit the month, audience, and thresholds rather than learn prompt syntax.
1. Claude in Excel: first practical prompt
Review this workbook and explain what each sheet contains in plain English. Do not change anything yet. Identify the fields we can use to compare actuals with budget by month and campus, and tell me about any missing values, duplicate rows, unmapped accounts, or inconsistent signs that could affect the analysis.
Follow with:
Add a new sheet called Variance Review. Show actual, budget, dollar variance, and percentage variance by campus and reporting line. Clearly flag the five largest unfavorable variances. Add a total row and visible checks that tie actuals and budget back to the source sheets. Preserve the original sheets.
2. Cowork: analysis and reconciliation
Work only in the Brightmoor workshop folder. Use
Brightmoor_Monthly_Finance.xlsxas the source and do not overwrite it. First inspect the workbook and give me a short plan, including the control totals you will use.Then create
Brightmoor_Monthly_Analysis.xlsx. Include revenue, operating expenses, and EBITDA for the current month and year to date; actual versus budget in dollars and percentages; payroll versus non-payroll costs; and trends by campus. Create separate sheets called Executive Summary, Variance Analysis, Campus Trends, Reconciliation, Checks, and Assumptions.Flag duplicate records, unmapped accounts, missing campus values, and reconciliation differences above $1,000. Put every exception in the Checks sheet. Do not invent missing values. If a definition is unclear, record the assumption instead of silently choosing one. At the end, report whether source totals tie to the output and list anything a finance owner must review.
Useful follow-ups:
Explain the three largest unfavorable variances in plain English. Separate facts supported by the workbook from possible explanations that need an owner to confirm.
Reconcile the GL payroll accounts to the payroll detail. Show the difference by month and campus, and flag anything over $1,000.
Create a repeatable compliance report showing the source, transformation, control total, exception, reviewer, and review status for each check.
3. Cowork: executive HTML dashboard
Use the checked numbers in
Brightmoor_Monthly_Analysis.xlsxto create a polished standalone HTML dashboard for Brightmoor's senior leadership. Do not recalculate or alter the financial data.Include KPI cards for revenue, EBITDA, EBITDA margin, and total operating expense; actual-versus-budget comparisons; a monthly trend; an EBITDA bridge; campus performance; and a short Key Takeaways section. Add filters for month and campus. Use a restrained executive style with accessible colors, clear labels, dollar formatting, and unfavorable variances shown consistently. Include the reporting period, data source, and last refresh date.
Save it as
Brightmoor_Executive_Dashboard.html. Before finishing, compare every displayed total with the Checks sheet and list any differences.
Tuning prompts:
Change the campus comparison to horizontal bars, sort from largest unfavorable variance to largest favorable variance, and keep the values visible on the chart.
Add a campus filter and a Reset Filters button. Do not change any totals or other parts of the dashboard.
Make the page easier to scan in a leadership meeting: reduce visual clutter, increase label size, and put the three most decision-relevant insights above the fold.
For the workshop, “publish” should mean producing a self-contained HTML file and showing how it opens in a browser. Sharing it beyond finance should happen only through a Brightmoor-approved internal location. Public hosting, authentication, and live Snowflake connections are separate IT and security decisions, not beginner workshop tasks.
4. Cowork: PowerPoint refresh
Update
Brightmoor_Monthly_Review_Template.pptxusing only the checked figures inBrightmoor_Monthly_Analysis.xlsx. Preserve the slide master, fonts, colors, footers, and existing layout. Do not overwrite the template; save the result asBrightmoor_Monthly_Review_Updated.pptx.Update the reporting period and create six slides: executive summary, revenue versus budget, EBITDA bridge, campus performance, payroll versus non-payroll, and reconciliation/open items. Use the approved chart on slide 2 as the style reference for charts on the next five slides. Keep chart scales and labels honest and consistent.
Draft no more than three commentary bullets per slide. Label statements based directly on the data as findings. Put possible business explanations in speaker notes and prefix them with “To confirm:” rather than presenting them as facts. Add a small source note to each slide. Then compare the figures on every slide with the workbook's Checks sheet and provide a slide-by-slide QA summary.
Tuning prompts:
Slide 4 is too crowded. Keep the same content, but simplify it for a two-minute executive discussion without changing any numbers.
Mirror the approved chart treatment from slide 2 across slides 3–6, including font sizes, colors, labels, and spacing. Leave the underlying values unchanged.
Check the deck for clipped text, overlapping objects, inconsistent number formats, stale month labels, and unsupported commentary. Fix presentation issues and list every change you made.
Synthetic workshop pack
Use synthetic data so Snowflake access is optional and no attendee is blocked by permissions. The data should resemble Brightmoor's workflow without pretending to reproduce its exact production schema.
Finance workbook
Brightmoor_Monthly_Finance.xlsx should contain 12 months of data and these sheets:
- GL_Actuals: period, transaction date, subsidiary, campus, department, class/program, account number, account name, account type, document number, vendor, memo, debit, credit, net amount, currency, and source system.
- Budget: period, campus, department, account number, and budget amount.
- Account_Mapping: account number, management-reporting line, payroll/non-payroll classification, EBITDA inclusion, and display sign.
- Payroll_Detail: synthetic employee ID, period, campus, department, gross pay, employer tax, benefits, and total payroll cost.
- Reconciliation: period, campus, GL balance, subledger balance, difference, owner, and status.
- Instructions: fictional-data notice, business definitions, sign conventions, and expected control totals.
Include a small number of intentional issues so the checking exercise is real: one duplicate journal, two unmapped accounts, a missing campus, an inconsistent expense sign, and two payroll reconciliation differences. Provide a clean answer key to the facilitator only.
Dashboard and slide inputs
- A clean Brightmoor-style guide with approved colors, typography, logo placeholder, and number formats.
Brightmoor_Monthly_Review_Template.pptx, with six slide layouts and one fully styled example chart.- A prior-month completed deck for visual reference.
- A second month's synthetic workbook for the repeatability demo.
- A one-page participant prompt card containing the four core prompts above.
Verification standard used throughout
Every output should make these checks visible:
- Source row count and source total before transformation.
- Output row count and output total after transformation.
- Actual, budget, and payroll totals tied back to their source sheets.
- Duplicate, missing-value, unmapped-account, and sign-convention exceptions.
- Reconciliation threshold and unresolved differences.
- Assumptions separated from facts.
- Reporting period, source file, refresh time, and output version.
- Named human reviewer and review status before distribution.
The facilitator should say this plainly: Claude accelerates preparation and presentation; it does not approve a close, certify a reconciliation, or own the forecast.
Pre-work and logistics
Send this checklist at least two business days before the session:
- Claude access is provisioned for every attendee; do not spend workshop time creating accounts.
- Claude Cowork opens successfully in the desktop app.
- Claude in Excel is installed, signed in, and visible in Excel.
- Excel, PowerPoint, and a modern browser are available.
- The complete synthetic workshop folder has been downloaded locally and unzipped.
- Each attendee can open the workbook, PowerPoint template, and HTML sample.
- Attendees know not to use production data during the workshop.
- Snowflake access is explicitly optional. The session must work without it.
Facilitator backups should include completed versions of every artifact, screenshots of each major step, a spare copy of the source folder, and a version of the dashboard and deck ready to show if file creation takes longer than expected.
What to leave out
- Claude Code, the terminal, GitHub, repositories, or coding concepts.
- ASCII art or location-based warm-ups.
- General AI history, model comparisons, or “future of work” framing.
- Open-ended build time and speculative use-case brainstorming.
- Public deployment or live system integration during the session.
- Claims that the full 15-slide monthly process will be perfect after one prompt.
The realistic promise is more useful: participants will leave knowing how to produce and iterate on a checked workbook, a leadership dashboard, and a six-slide monthly review from familiar finance-shaped data. Turning that into Brightmoor's production 15-slide workflow will require the approved template, agreed definitions, repeatable control checks, and several monthly iterations.
Assumptions
- The working session is 3.5 hours inside the four-hour block mentioned in the preparation call.
- The attendee count is 20–30 and most participants are beginners.
- Brightmoor will approve the security and sharing rules for Cowork before anyone uses production data.
- The workshop uses synthetic data modeled on a generic NetSuite-to-Snowflake finance export; the exact field names and reporting definitions will be adjusted if Brightmoor supplies sanitized examples.
- Internal dashboard publishing and direct Snowflake connectivity are follow-on implementation topics because they depend on Brightmoor's access controls and IT environment.
Checks
16/16Run of show existsexists
a run-of-show document was produced
Coherent scheduleformat
it parses as a run of show: one coherent schedule, no segment that ends before it starts, >=5 time-ranged rows carrying activity text (md table rows, list items, bold headings "(9:00-9:15)", single-clock table rows, or plain lines all count)
Five-minute time markstime-grain
segment boundaries sit on 5-minute marks, at most one off-grid -- a schedule planned at :02/:08/:17 is arithmetic, not a plan
Timings add uptimings-tie-out
per schedule, segments are contiguous: no overlaps, no gap over 30 min, >=85% of the span scheduled, and a span between 45 and 600 minutes
Q1Uses Cowork throughout
Judge's reasoning
All attendee steps are Claude Cowork or Claude in Excel; a 'What to leave out' list explicitly excludes 'Claude Code, the terminal, GitHub, repositories, or coding concepts.'
▸Rubric
Cowork-only. Do all attendee-facing steps stay inside Claude Cowork (and Claude in Excel)? FAIL if any step the attendees are asked to do involves Claude Code, the terminal/CLI, git or GitHub, or installing developer tooling. The prompt: "this is only using Claude Cowork, so we should not be using Claude Code, either in the desktop app or in the terminal, because this team is non-technical."
Q2Dashboard exercise
Judge's reasoning
1:40–2:25 'Build 2: executive dashboard' has attendees produce `Brightmoor_Executive_Dashboard.html` from the checked workbook and make two visible changes, with paste-ready prompts.
▸Rubric
Dashboard use case. Is there a hands-on segment where attendees build a dashboard from existing spreadsheet data? FAIL if dashboards are only mentioned, described or promised for later rather than built in the session.
Q3Financial analysis exercise
Judge's reasoning
1:20 block 'Build 1: financial analysis and reconciliation' (45 min) plus a verify segment produces `Brightmoor_Monthly_Analysis.xlsx` with EBITDA, budget variance, payroll splits and a Checks sheet.
▸Rubric
Financial analysis use case. Is there a hands-on segment where attendees do financial analysis (Claude writing/running code over their numbers, e.g. Excel in, analysis out)? FAIL if absent or folded into the dashboard segment as a passing remark.
Q4PowerPoint exercise
Judge's reasoning
2:25–3:05 'Build 3: refresh the monthly PowerPoint' lists `Brightmoor_Monthly_Review_Updated.pptx` in the participant-output column, including the client's 'mirror one approved chart style across the next five slides' ask.
▸Rubric
PowerPoint use case. Is there a hands-on segment where attendees update a PowerPoint template/deck from the data — the client's monthly reporting pack? FAIL if absent, or if it is only a demo the facilitator drives while attendees watch.
Q5Relevant business data
Judge's reasoning
Synthetic pack is NetSuite-shaped — GL_Actuals with subsidiary, campus, department, class/program, account number, document number, vendor, memo, debit/credit — plus Budget, Account_Mapping, Payroll_Detail and Reconciliation sheets with month and YTD reporting.
▸Rubric
NetSuite/Snowflake-shaped data. Is the synthetic data used in the exercises shaped like this client's actual data — NetSuite/Snowflake-style finance records (GL export, revenue and budget vs actual, EBITDA build-up, payroll vs non-payroll, campus/entity breakdown, monthly and year-to-date columns)? FAIL if the exercises use generic sample data (a demo CSV, sales widgets, made-up SaaS metrics) or leave the data unspecified. The prompt: "the shape of the tasks, the shape of the data should be like NetSuite shaped."
Q6No ASCII icebreaker
Judge's reasoning
Opening exercise is 'First useful result in Excel' (variance vs budget, five largest exceptions), explicitly 'Replace the usual ASCII warm-up with a real finance task', and ASCII art is listed under what to leave out.
▸Rubric
No ASCII icebreaker. Is the opening hands-on exercise a practical finance/Excel task? FAIL if the run of show keeps the ASCII-image-of-your-location icebreaker or substitutes another whimsical non-work exercise. The prompt: "maybe the ASCII image idea that I normally do where I ask it to create an ASCII image based on my location isn['t] like that useful... maybe we use an Excel implementation instead."
Q7Explains Cowork versus ChatGPT
Judge's reasoning
0:25–0:35 segment keeps the framing: 'unlike a typical ChatGPT-style conversation that mainly returns an answer in the chat, Cowork can work with files in an approved folder, run code behind the scenes.'
▸Rubric
Cowork vs ChatGPT explained. Does the run of show keep the explanation of how Cowork differs from ChatGPT — that it has access to the file system and can run code? FAIL if that framing is dropped. The prompt: "I still want to keep the way that we introduce Claude Cowork, where we explain to it that the difference between Claude Cowork and ChatGPT is that it has access to the file system. It can run code. I think that important."
Q8No theory or hype
Judge's reasoning
Every segment is a finance deliverable step; the run sheet states 'No AI trends, icebreaker, or speculative use cases' and excludes 'General AI history, model comparisons, or "future of work" framing.'
▸Rubric
No theory, no AI cheerleading. Is every segment tied to a task these people do at work? FAIL if any segment is devoted to AI industry context, the future of work, model capabilities, prompt-engineering theory, or excitement-building. The prompt: "make it very tight run down, make it very condensed and practical. No real like wider theory or like, hey, we're excited about AI things. I think these guys just want to get stuff done."
Q9Mostly hands-on work
Judge's reasoning
Hands-on segments (15+45+10+45+40+15 = ~170 min) dominate a 210-minute session, with only ~20 min of framing/close and a 10-min break.
▸Rubric
Mostly hands-on. Is at least half the scheduled time attendees working with Claude themselves? FAIL if presentation, discussion and Q&A segments outweigh build segments. From the call: "we try and make at least 50% of it them actually you know working with us to do some of these tasks."
Q10Realistic segment timings
Judge's reasoning
Build segments get 45, 45 and 40 minutes with a verify/debrief block, and setup is handled by a two-day-ahead pre-work checklist plus 'type the first prompt with the group, then pause after each major action so beginners do not fall behind.'
▸Rubric
Per-segment timings realistic. Could a room of 20–30 beginners actually finish each segment in the time allotted? FAIL if any build segment is implausibly short (a dashboard, an analysis or a deck refresh built end to end in ~15 minutes), or if setup/handholding time for first-time users is ignored.
Q11Fits the booked time
Judge's reasoning
Schedule runs 0:00–3:30 and explicitly assumes 'The working session is 3.5 hours inside the four-hour block mentioned in the preparation call.'
▸Rubric
Total length fits the booked slot. Does the session run about three and a half hours, inside the four-hour block? FAIL if the total is materially shorter or longer (under ~3 hours or over ~4 hours) without the run of show explaining the change. From the call notes: "we have a four hour block for next week and the session itself is three and a half hours."
Q12Ready to run
Judge's reasoning
Each segment names its input files, participant output artifact and full copy-pasteable prompts (plus tuning follow-ups), a data-pack spec and a facilitator backup plan.
▸Rubric
Usable as a run of show. Could a facilitator run the session from this document alone? FAIL if segments are titles without content — no prompts to paste, no data files named, no statement of what attendees produce — so the facilitator would still have to design the session.