Week 1 file
diary/week-1-2026-10-08.md is the audience cut (board, investors, customers, employees) of the pause. Day 4 in this log is the same decision. Neither file is a measured result.
Simulation: FrontierOrg is a fictional infrastructure holding group; figures are fixtures or marked not measured, and no vendor sponsors this work.
Day 0 plus Days 1–22. Days 13–22 deepen consultation outcomes, draft an empty cost envelope, chase licence unknowns without inventing a region, design a stock-check remediation, bound a draft-only agent that never sends, red-team it on paper, write handoffs, rank the portfolio, and recommend revise — not go, not stop-as-erasure. Scorecard actuals stay not measured. A workflow may open an issue. It does not publish this page.
Scenario date: 3 October 2026 Gate: Simulation: FrontierOrg is fictional. This entry authorises a method, not a company.
Can one service promise be joined across CRM, worker, business partner, and invoice without inventing a benefit?
The mandate is a hypothesis, not a target percentage. A customer service promise should be traceable from a fictional CRM case to a held invoice, with a named human at each handoff. Success is a join a reader can inspect, a decision log, and empty scorecard cells left empty.
Decision: proceed to mapping in Week 1. No pilot. No agent in the path. The board triad — customer outcome, employee transition, risk and cost — is equal. A gap in any one is enough to pause.
Sponsor on the page is a fictional chief executive. Nothing here binds a real board.
Nothing is for sale. There is no valuation, no revenue, and no saving. The only asset in view is a public method. Time spent so far is uncosted. COST-001 is opened so that absence stays visible.
No customer exists. The first story will use Harbour Grid Municipal, a fictional account at Northwater Distribution, a fictional energy-distribution operating company. That choice is a working Week 1 setting, not a locked brand. Any date in the fixture is synthetic. Partners are not being told a thing, because they are not real.
Fifty thousand is a planning total for human workers. It is not a directory we provisioned. Agent and service identities will be modelled apart from that total and will not inflate it. No role is declared redundant. PEOPLE-001 asks what the service coordinator, billing analyst, and data steward actually do before any task is redesigned. There is no workforce to consult; the diary will not pretend there was.
| Item | State |
|---|---|
| Scorecards | Day 0 shells. Actuals: not measured. |
| Fixture | Not yet generated into the story (lands with Week 1 as seed 20261003). |
| Issues | DATA-001 named, not yet evidenced. PEOPLE-001 open. COST-001 open. |
Scenario date: 2026-10-05 (Monday) Gate: Gate 0 — sponsor and mandate Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional CEO (sponsor on the page only). Amina Okonkwo is named as the future service coordinator and is not yet in a case. No workforce exists to consult.
Wrote the engagement hypothesis: one service promise at Northwater Distribution must be traceable to a held invoice with a named human at each handoff. Ethics constraints recorded: simulated directory only, no vendor endorsement, no real CKI claim, pause beats a fake win. Operating company for the first story remains Northwater Distribution.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. |
| Still open | PEOPLE-001 open. COST-001 open. DATA-001 named on Day 0, not yet evidenced in a row. |
No actual moves. All three scorecards stay not measured. Decision text updated from 'commission' to 'mapping authorised for one operating company, no pilot'.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Authorise current-state mapping of a single path. Do not deploy an agent.
Can the path be opened in the fixture without inventing a customer outcome?
Scenario date: 2026-10-06 (Tuesday) Gate: Gate 1 — current-state work Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo, service coordinator (O*NET 43-5032.00, partial match). Rowan Ellis, fictional liaison at Harbour Grid Municipal. Both names are fabricated.
In the fixture, CASE-61003 opens at 09:14 Europe/London with a feeder-restoration promise of 20 October 2026. Amina acknowledges it at 11:02. Parts are not confirmed by 15:40. The defect is now a row, not a slogan.
| Movement | Detail |
|---|---|
| Opened | DATA-001 (promise before stock). |
| Closed | None. |
| Still open | PEOPLE-001, COST-001. |
Operational card records fixture cardinality: 1 case in data/sample. That count is a file fact, not a rate. Customer, people, and cost actuals remain not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep mapping. Do not put an assistant on the acknowledgement step.
What does billing do with a promise that parts cannot support?
Scenario date: 2026-10-07 (Wednesday) Gate: Gate 1 — exceptions and controls Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Jonas Berg, billing analyst (O*NET 43-3021.00, partial). Priya Raman, data steward (O*NET 15-1243.00, weak match). Amina remains the case owner.
Jonas raises INV-61003 against BP-61003 and holds it because alias BP-61003-A is a duplicate suspect. Priya reviews the join at 16:45 and does not silently merge the alias. The hold is the control.
| Movement | Detail |
|---|---|
| Opened | None. DATA-001 gains the duplicate-partner evidence. |
| Closed | None. |
| Still open | DATA-001, PEOPLE-001, COST-001. |
Operational exception text now cites the hold. Actuals remain not measured. No satisfaction score.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Do not clear the hold to make the path look finished.
Does the board pause, or does someone argue the path is 'agent-ready' because the steps are listed?
Scenario date: 2026-10-08 (Thursday) Gate: Gate 0 decision inside Gate 1 Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
The three named personas plus fictional board. Agent identity AGT-61003-DRAFT is modelled and not a person.
Wednesday diary window (8 October, 06:30 Europe/London) publishes the Week 1 cut: the promise outruns the parts. This day records the decision in the activity log. The public Wednesday file remains diary/week-1-2026-10-08.md.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. |
| Still open | DATA-001, PEOPLE-001, COST-001. |
C-suite decision on risk/cost becomes 'do not start an assistant on this path'. People decision: no task removed. Customer actual: not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
AGT-61003-DRAFT stays not deployed. Pause.
What lineage is actually proven by the six events, and what is still a label?
Scenario date: 2026-10-09 (Friday) Gate: Gate 2 — data reality, started Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Priya Raman leads. Jonas and Amina confirm they can point at their own event rows. ESCO crosswalk owner: none.
Traced SCN-61003 across crm_cases, workers, sap_business_partners, sap_invoices, and process_events. Keys hold for this one scenario. ESCO URIs are still empty. Directory rows say provisioned_in_entra = no.
| Movement | Detail |
|---|---|
| Opened | DATA-003 — ESCO v1.2.1 crosswalk absent. Do not invent URIs. |
| Closed | None. |
| Still open | DATA-001, PEOPLE-001, COST-001, DATA-003. |
Operational card notes the join keys that resolve. Still not a quality percentage.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Do not start a second operating company until this join's defects are either accepted or fixed in the fixture on purpose.
Which licences would a real pilot need, and which of those do we refuse to invent?
Scenario date: 2026-10-12 (Monday) Gate: Gate 2 — licence reality Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional CFO and a fictional technology owner. No vendor salesperson is in the story.
Listed the classes of licence a later pilot would have to name: identity, CRM analogue, ERP analogue, model hosting, and logging. Every class is 'not inventoried'. Processing location is unknown. No API quota is stored.
| Movement | Detail |
|---|---|
| Opened | RISK-001. |
| Closed | None. |
| Still open | DATA-001, DATA-003, PEOPLE-001, COST-001, RISK-001. |
Risk/cost scorecard cites RISK-001. Actual and cost_to_date remain not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Licence inventory is a precondition for Gate 5. It is not started. No sandbox.
Can O*NET tell us which tasks exist in the holding group without pretending the codes are our job descriptions?
Scenario date: 2026-10-13 (Tuesday) Gate: Gate 1 extended — work inventory from an external taxonomy Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Priya owns the extract. A fictional workforce-intelligence lead is named and has not validated a single title. Amina's code remains a partial match.
Loaded O*NET 31.0 (August 2026 text release) from the O*NET Resource Center. Generator generators/onet_frontier_map.py keeps 302 occupations relevant to an infrastructure holding group and 5,932 task statements. Department tags are a FrontierOrg guess. Essential Skills (importance) stand in for the retired Skills file.
| Movement | Detail |
|---|---|
| Opened | DATA-002. |
| Closed | None. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, COST-001, RISK-001. |
Fixture metric, not a performance actual: 302 occupations and 5,932 tasks in data/generated. Scorecard actuals for customer, people, and cost stay not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Publish the map as a draft inventory. Do not treat SOC codes as redundancy evidence.
Who would have to sit in a room before a department tag is allowed to drive a pilot?
Scenario date: 2026-10-14 (Wednesday) Gate: Gate 1 — how the map will be read Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Same three, plus the pattern borrowed from organisationalunderstanding.com: an occupation has tasks, skills, and a separate AI layer that is labelled as a modification. No worker has logged in, because the directory is simulated.
Published apps/site/org-map.html from org_map_summary.json. Departments list occupations and task counts. Documented the bridge to OrgUnderstanding and the WorkDNA work-chart idea: reporting lines do not locate the work. Wednesday draft slot may open; this entry is the human activity log, not an auto-published win.
| Movement | Detail |
|---|---|
| Opened | PEOPLE-002. |
| Closed | None. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001. |
People scorecard decision: org map is a draft. Consultation count is a fixture fact of zero sessions.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep the page public and labelled heuristic. Do not rank a build list yet.
Which tasks must stay human even when the wording looks clerical?
Scenario date: 2026-10-15 (Thursday) Gate: Gate 4 preview — no-go stays explicit. Gate 3 not opened. Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Priya reviews the suitability column. A fictional safety lead is the reason RISK-002 exists. Jonas is not asked to automate billing. Amina is not asked to let an agent promise a date.
ai_workload_candidates.csv classifies each of the 5,932 tasks as assist, augment, automate-candidate, or human-only. The heuristic defaults to augment when the words are weak, so most rows are not automation. Human-only catches a thin set of safety, emergency, statutory, and physical-control phrases. Automate-candidate is rare and still carries a human gate.
| Movement | Detail |
|---|---|
| Opened | RISK-002, BENEFIT-001. |
| Closed | None. |
| Still open | All prior issues. |
No ROI cell. Suitability counts may be cited as fixture counts of labels, not as value. Customer outcome not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Do not open Gate 5. The portfolio item 'agent on the service promise' stays an explicit no-go.
Can a larger synthetic population keep the same joins without being mistaken for the 50,000?
Scenario date: 2026-10-16 (Friday) Gate: Gate 2 — population fixture, not scale Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
The original three remain the only named characters in the diary. The 2,000 workers are numbered synthetic identities (example.com). Ten agent and service identities sit outside the human count.
generators/sample_large.py wrote data/sample_large at seed 20261003: 2,000 humans, matching directory rows, CRM accounts, cases, SAP-like partners, invoices, and events. Referential checks passed in the generator. About one case in five repeats promise-before-stock. Every eighth account has a duplicate-suspect partner. Amounts are blank. The 50,000 planning total is documented as a local command into gitignored data/sample_full and was not committed.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. Pause stands. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
Fixture cardinality on the operational card: 2,000 workers in data/sample_large, plus the original 3 in data/sample. Not a headcount result. Not an FTE saving. Actuals not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Stop the ten-day burst here. Next human review is the department tags and the human-only list, not a pilot.
Which twenty occupations would a real validation session start with, and who is allowed to overturn the heuristic?
Scenario date: 2026-10-19 (Monday) Gate: Gate 1 — the map in front of the named roles Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo (service coordinator), Jonas Berg (billing analyst), and Priya Raman (data steward). The people page now shows their O*NET occupations, a handful of tasks, the suitability label, and a link to the department guess on the org map. Consultation sessions held: 0. No label was changed, because nobody who does the work reviewed one. That zero is still a fixture fact.
The official ESCO–O*NET crosswalk (mapping project version 1, September 2022) is in the repo. The file itself says ESCO v1.1.0 and **O*NET-SOC 2019**. It is not ESCO v1.2.1. Joined to the 302 O*NET 31.0 occupations in this slice: 290 have at least one row (96.0%). Twelve codes are unmapped. Match type and a human-review flag sit on each row. A URI is not a job evaluation.
data/generated/northwater_ai_workloads.csv holds 40 task rows on the path service-promise → dispatch → parts → invoice, plus the data join and a safety-critical block. Fifteen rows are human-only. Eight of those sit on the safety-critical step (switching, grounds, climbing, line work, breaker repair, shutdown notice, site safety). The keyword pass had called most of them augment. That miss is written down. RISK-002 stays open: there is still no safety reviewer outside the fiction.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. DATA-003 has a sourced file and is not closed: version is v1.1.0, twelve codes are unmapped, and review flags remain. PEOPLE-002 is not closed: no session happened. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Crosswalk coverage is a join count, not a benefit. Consultation count remains zero. Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Leave the pages up, labelled as a draft. Do not start a pilot. Do not treat an ESCO URI as permission to redesign a role.
Who may overturn a broad match, or a human-only override, and what gets written down when they do?
Scenario date: 2026-10-20 (Tuesday) Gate: Gate 1 — simulated consultation Simulation: FrontierOrg is fictional. The sessions below are a scenario script. They are not quotes, not a survey, and not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo (service coordinator, O*NET 43-5032.00), Jonas Berg (billing analyst, O*NET 43-3021.00), and Priya Raman (data steward, O*NET 15-1243.00). The people page carries one session note each. Occupations reviewed in the script: those three codes. Labels changed: none. Tasks removed: none.
These lines are authored for the thought experiment so the method has a session shape. They are not words from a workforce.
Workforce consultations held: 0. Simulated session notes: 3. PEOPLE-001 stays open because its done-when is a session with people outside this simulation, or an honest zero. A script is not that session. PEOPLE-002 stays open: the note names occupations reviewed and states that no label changed, but the reviewers are fictional.
controls/licence_region_retention_register.md and data/generated/licence_register.json list the fictional stack. Microsoft 365 / Copilot-style wording is an assumption, labelled fictional, with no edition and no price. Entra-like identity is documented only as not provisioned. Halo-inspired CRM, SAP-like ERP, and Workday-like HCM are class assumptions. Model hosting, region, retention, and logging are unknown. No retention period was invented. RISK-001 stays open. Gate 5 stays blocked. COST-001 stays open.
A local run of generators/sample_large.py --seed 20261003 --workers 50000 --out data/sample_full passed the generator's referential checks. Counts are in data/sample_full/MANIFEST.md. The CSVs are gitignored. The count is not headcount in a directory and not a benefit.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. RISK-001, PEOPLE-001, and PEOPLE-002 are partial only. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Customer outcome, employee transition, and risk/cost stay not measured. Fixture cardinality may cite the 50,000-row regenerate. That is a file count.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep the script visible and labelled. Do not remove a task. Do not start a pilot. Do not close PEOPLE-001 on invented speech.
Who outside the fiction is allowed to confirm or overturn a label, and which system row must leave "unknown" before any training course is named?
Scenario date: 2026-10-21 (Wednesday) Gate: Gate 1 — deepen consultation outcomes Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo, Jonas Berg, and Priya Raman. The Day 12 scripts are deepened into controls/consultation_outcomes_register.md. Occupations reviewed: 43-5032.00, 43-3021.00, 15-1243.00. Labels changed: 0. Workforce consultations: 0.
Deepen Day 12 scripts into an outcomes register: fears → control shapes, training asks → course stubs that are not scheduled. Labels still unchanged. Workforce consultations: 0.
Outcomes register columns: fear → control shape → training ask → course status. Course status for every row is not scheduled. The register is a method artefact, not evidence that anyone sat down.
stock_confirmed is false. Training ask remains how to reject a draft date.| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Simulated session notes remain 3. Workforce consultations remain 0. Labels changed: 0.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep the outcomes register labelled as a scenario script. Do not schedule a course. Do not change a task label.
What cost lines must appear before anyone asks for a spend ceiling on a draft-only design?
Scenario date: 2026-10-22 (Thursday) Gate: Gate 2 — cost envelope draft Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional CFO and risk lead on the page. Amina, Jonas, and Priya are named as the roles the envelope would have to fund if a later design were ever approved. No workforce is costed.
Draft COST-001 envelope with labelled assumption lines: licence class, inference, integration, training, oversight, remediation. Every amount stays not measured. No vendor price list copied in.
File: controls/cost_envelope_draft.md and data/generated/cost_envelope.json.
Lines in the envelope (every amount not measured, every source assumption — not a contract):
No forecast range is modelled as a wish. No token-spend model is presented as actuals. Done-when for COST-001 is still unmet: a sourced figure for at least one term, or the cell stays not measured. We choose the latter for every term.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. COST-001 partial: structure without figures. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Cost to date stays not measured. Forecast range: not modelled.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Publish the empty envelope. Do not invent a price to fill it. Do not open a pilot because a spreadsheet exists.
Which licence unknown can be chased without inventing a tenant region or a retention period?
Scenario date: 2026-10-23 (Friday) Gate: Gate 2 — licence unknowns chased Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional technology owner and risk lead. Craig remains the only human who can approve public text about a real product.
Chase RISK-001 unknowns: write preferred-but-unselected notes (EMEA preference labelled preference, not selection). Model hosting and logging stay unknown. No retention period invented. Gate 5 stays blocked.
Updates to controls/licence_region_retention_register.md:
RISK-001 done-when remains unmet. Gate 5 stays blocked. No SKU named. No price.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. RISK-001 still open after the chase. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Licence register rows still 8. Unknowns remain unknowns.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Leave unknowns as unknown. Do not invent a region, a retention period, or a logging rule to unblock Gate 5.
Where does a stock-confirmed check sit on the SCN-61003 path so Amina cannot acknowledge an unsupported date?
Scenario date: 2026-10-26 (Monday) Gate: Gate 1/2 — data join remediation design Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo (would own the stock-check before promise acknowledgement), Priya Raman (duplicate BP remains unmerged), Jonas Berg (hold stays until the join is clean).
Design remediation for DATA-001: a stock-confirmed gate before Amina may acknowledge a promise date on SCN-61003. Duplicate BP stay unmerged; Priya still refuses silent merge. Design only — fixture not rewritten as fixed.
File: docs/engagement/data_join_remediation_scn61003.md.
Proposed control (design only — the published fixture still shows the defect):
stock_confirmed_before_promise must be true before coordinator_acknowledged_promise.duplicate_suspect. No silent merge. Priya's refuse-merge rule stays.Fixture counts unchanged: 1 case, promise_before_stock still true in the published sample. Remediation progress is design, not a claim that the defect is gone.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. DATA-001 partial: remediation designed, defect still in the fixture. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Exception baseline still 'promise before stock; invoice held'. Not measured as a rate.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep the defect visible in the fixture. Treat the stock-check as a required gate for any later draft-only design. Do not rewrite history to pretend the promise was safe.
What is the narrowest draft-only agent design that never sends and never acts on a case?
Scenario date: 2026-10-27 (Tuesday) Gate: Gate 3 — bounded pilot design Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
AGT-61003-DRAFT (agent identity, not a person). Bound reviewer in the design: Amina Okonkwo. Jonas and Priya are out of scope for this draft path.
Design AGT-61003-DRAFT as draft-only: may propose a schedule note for Amina; may_act_on_case stays no; never sends to customer or crew. No sandbox. No spend ceiling because nothing is spent.
File: docs/engagement/pilot_design_agt61003_draft_only.md.
Bounds written into the design:
may_act_on_case = no (unchanged in data/sample/agent_identities.csv)deployment_status = not_deployedThis is Gate 3 design work. Gate 5 (controlled pilots) is not started.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. OP-ROLLBACK baseline remains 'nothing deployed'. OP-TOKENS stays not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Accept the draft-only design as paper. Do not deploy. Do not bind a model host. Do not train anyone on a product that is not named.
What failure modes must a red team prove before the draft-only design is even rehearsed?
Scenario date: 2026-10-28 (Wednesday) Gate: Gate 3/5 — red-team and safety Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional risk lead as red-team chair. Amina, Jonas, and Priya appear as the humans the attacks would bypass. No agent runs.
Red-team the draft design against SCN-61003: promise before stock, crew relay, hold release, silent BP merge. Failure modes require pause. RISK-002 stays open. No agent runs.
File: docs/engagement/redteam_agt61003_scn61003.md.
Attack stories (tabletop only):
stock_confirmed=false. Required response: refuse; surface unconfirmed status. Pass condition: draft never becomes the customer date.RISK-002 stays open: safety-critical and payment work must not quietly automate. Red team recorded on paper. Agent executions: 0.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. RISK-002 partial: failure modes listed, automation still forbidden. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Safety remains design, not a measured QA rate.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep every red-team pass condition as a hard gate on the draft-only design. Do not rehearse with a live model while Gate 5 is blocked.
Can the human–agent handoffs be written so each persona knows what they alone may release?
Scenario date: 2026-10-29 (Thursday) Gate: Gate 3 — human–agent handoffs Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Amina Okonkwo (draft schedule reviewer), Jonas Berg (hold release authority), Priya Raman (join and access authority).
Write human–agent handoff table for Amina, Jonas, Priya: who reviews, who rejects, what never leaves a human. Transition design only. No task removed. No course ran.
File: docs/engagement/human_agent_handoffs.md. Also mirrored on the people and process pages.
| Human | May receive from draft | May release | Never delegated |
|---|---|---|---|
| Amina | Draft schedule note only | Customer-visible date after stock-check | Crew relay; unsupported promise |
| Jonas | (none in this path) | Invoice hold release | Payment; reading blank amount as money |
| Priya | (none in this path) | Partner join decision | Silent merge; access writes |
Labels changed: 0. Tasks removed: 0. Courses: 0. Workforce consultations: 0.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. PEOPLE-001 / PEOPLE-002 still open (script and design, not workforce). |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Employee transition stays not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Publish the handoff table as design. Do not treat it as consultation evidence. Do not remove a task.
How should Northwater's promise-to-invoice path be ranked against a send/act pilot?
Scenario date: 2026-10-30 (Friday) Gate: Gate 4 — portfolio ranking Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional COO and CFO for portfolio ranking. Northwater path owners remain Amina, Jonas, Priya.
Rank Northwater promise-to-invoice: explicit no-go for send/act pilots; conditional revise for draft-only behind stock-check and human release. BENEFIT-001 stays a refusal. No ROI.
File: docs/engagement/portfolio_ranking_northwater.md.
Ranking for SCN-61003 / Northwater Distribution:
| Use case | Rank | Decision |
|---|---|---|
| Agent sends restoration date to customer | Explicit no-go | DATA-001 and red-team fail conditions |
| Agent relays to crew | Explicit no-go | RISK-002 / human-only |
| Agent clears invoice hold or pays | Explicit no-go | Payment human-only |
| Draft-only schedule note for Amina, never sends | Conditional revise | Only behind stock-check, red-team gates, may_act_on_case=no |
| Silent BP merge assist that writes | Explicit no-go | Priya refuse-merge |
BENEFIT-001: no benefit claimed. No hours-to-cash conversion. No ROI.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. BENEFIT-001 remains a standing refusal. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. BA-BENEFIT stays not claimed / not measured.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Keep send/act as explicit no-go. Keep draft-only as revise-candidate only. Do not open Gate 5.
What three board options — go, revise, stop — can be written without inventing a win?
Scenario date: 2026-11-02 (Monday) Gate: Gate 4/6 prep — board options Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional board (CEO, CHRO, CFO, risk). Personas Amina, Jonas, Priya appear as the named roles in the options paper.
Draft board pack with three options: go (pilot with spend — refused), revise (continue mapping + draft-only design), stop (freeze path). Equal triad. No decision claimed as taken.
File: docs/engagement/board_options_go_revise_stop.md.
Go. Start a controlled pilot with a named model host, spend ceiling, and sandbox. Blocked: RISK-001 unknowns, COST-001 empty, Gate 5 not opened, red-team not executed on a running system.
Revise. Continue current-state mapping; keep draft-only design on paper; implement stock-check in the *next* fixture revision without erasing Day 2 evidence; keep may_act_on_case=no; keep scorecards not measured.
Stop. Freeze the Northwater path; leave diary as a public no-go case; do not design further agents.
Equal triad applies to each option. No option invents a customer outcome, a transition rate, or a net value. Decision not taken on this day — pack only.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves. Board triad remains equal. Forecast: not modelled.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Circulate the options pack as fiction for the method site. Do not claim a board vote. Do not go.
Which option should the engagement recommend, and what stays open if that option wins?
Scenario date: 2026-11-03 (Tuesday) Gate: Gate 4/6 prep — recommendation Simulation: FrontierOrg is fictional. This entry is an activity log, not a measured result. Figures are fixture counts or marked not measured.
Fictional CEO as sponsor on the page. Recommendation authored for the method; Craig still approves anything public. Amina, Jonas, Priya remain the named roles.
Editor recommendation inside the fiction: revise. Keep SCORECARD actuals not measured. Keep may_act_on_case = no. Standing issues stay open. Pause still beats a fake win.
Recommendation: revise.
Reasons under the equal triad:
Standing issues stay open. AGT-61003-DRAFT stays not deployed. Fixture still shows promise before stock. Scorecard actuals stay not measured.
Days 13–22 close this activity block with a recommendation, not a declared board minute.
| Movement | Detail |
|---|---|
| Opened | None. |
| Closed | None. All nine standing issues remain open by design. |
| Still open | DATA-001, DATA-002, DATA-003, PEOPLE-001, PEOPLE-002, COST-001, RISK-001, RISK-002, BENEFIT-001. |
No actual moves across Days 13–22. Fixture cardinality unchanged as KPIs. Recommendation text updated on CS decision fields only.
Customer outcome, employee transition, and risk/cost stay equal. Pause beats a fake win.
Recommend revise. Refuse go. Refuse stop-as-erasure. Keep pause available on any later day that invents a win.
If revise continues, which single artefact — stock-check in a new fixture revision, or a named processing region — should move first without inventing the other?
diary/week-1-2026-10-08.md is the audience cut (board, investors, customers, employees) of the pause. Day 4 in this log is the same decision. Neither file is a measured result.
Options and the Day 22 recommendation live in docs/engagement/board_options_go_revise_stop.md. Not a board minute.