Conversation Design for High-Stakes Escalations

B100's Ops team rebuilds every payment case by hand across three disconnected systems, often while handling daily calls with clients. I designed a three-level escalation ladder, ambient, copilot, and agentic, with the transitions and trust patterns that hold it together while money is moving.

Timeline

2026, in design

Role

Lead Designer, Conversational & Agentic AI

Method

Escalation-ladder framework + interaction-pattern specs

Team

1 Designer (Design & AI architecture) + 1 Engineer

Context & Opportunity

To explain why a payment is stuck, an Ops analyst rebuilds the case by hand across three disconnected systems, the card processor, the provider that moves the money, and a legacy database linking them, with no single screen that shows all three.

This is agentic work: multi-source retrieval and reasoning. But because money is moving, confidence signals, citations, approval gates, and an audit trail need to be built in from the start.

SOURCE 01
Gen 3
Card transaction system
manual lookup
SOURCE 02
Transfermate Portal
External money-transfer agent
manual lookup
SOURCE 03
Gen 2
Legacy mapping database

The flow

One case, end to end

A single scenario, a stuck Brightwater Logistics settlement, walked through all three levels and all four transitions. This is the flow the case study is built to demonstrate.

9:41
Transactions
This month
Brightwater Logistics
Tamsett Co. · due today$10,000
$10,000 vs ~$3,100 typical · flagged by Gen 3
Kestrel Mfg.Pending
Orbit Industrial · Sep 3$25,000
Halden FreightFunded
Ridgeline Fleet · settled$42,000
Verde Foods Co.Funded
Greenfield Produce · Aug 29$8,200
Larkspur SystemsError
Datapoint · failed 14:02$5,500
9:41
Brightwater · GW…826541
Why didn't this settle?
Anomaly · loaded as context
Checking sources
Gen 3 · card system
Transfermate · transfer
Gen 2 · mapping
Ask a follow-up…
9:41
Brightwater · GW…826541
G3
Gen 3
Complete
Amount
$10,000.00 · full capture
Card
••4471 · auth 14:29:58
Queried
14:31:04
TM
Transfermate
Partial
Leg 1
settled 14:12:30
Leg 2
not initiated, no trigger event
Queried
14:31:04
G2
Gen 2
No match
Expected
GW-826541
Found
no match
Mapping
table last synced Aug 28
Leg 2 was never initiated for this transaction or for 11 others, totaling $284,900.
Gen 3 and Transfermate agree, but Gen 2's mapping is stale, so the affected count may be off by one or two. I can retry the 11 confirmed transactions or re-check the count first.
Ask a follow-up…
9:41
Brightwater · GW…826541
Here's the retry, ready for your approval.
123
Proposed
Retry Leg 2 settlement
Destructive action
$284,900 · 11 of 12 transactions
1 excluded: under dispute, can't verify
Review transactions11
GW…826541$10,000
GW…826544$42,000
GW…826547under dispute
GW…826550$8,200
GW…826551$15,400
GW…826553$31,000
Approve AllReview each
Ask a follow-up…
9:41
Brightwater · GW…826541
Retrying now. You can stop the remaining at any point.
23
Executing
Retry Leg 2 settlement
9 of 11 settled · 2 in progress
Transactions11
GW…826541Settled
GW…826544Settled
GW…826547skipped
GW…826550Retrying
GW…826551Retrying
Cancel
Ask a follow-up…
9:41
Brightwater · GW…826541
Done. 11 settled, 1 skipped.
3
Done
Retry Leg 2 settlement
$269,500 cleared · 11 settled, 1 skipped
Transactions11
GW…826541Settled
GW…826544Settled
GW…826547skipped
GW…826550Settled
GW…826551Settled
GW…826553Settled
Audit trail
IntentOps · 14:29
PlanAgent · 14:30
ApprovedOps · 14:31
ExecutedAgent · 14:32
ResultSettled
Reversible until 3:32pm ·
Ask a follow-up…

tap the phone to pause · tap a number to jump

Where AI doesn't go

Direct Ops data entry

New buyer mapping stays manual until volume justifies the calibration cost.

Compliance & legal approvals

Human, non-negotiable. The agent can assemble the packet; it cannot sign off.

Terminal states of money already moved

Audit yes, action no. Once the money has settled there is nothing left to edit, a correction is a new transaction.

The B100 Operations team reviewing a payment case
Before
bAgent actionProposed
Retry Leg 2 settlement
Scope11 of 12 transactions
Total$284,900 to 11 suppliers
OrderOldest first, stop after 2 fails
Reversible for 24h
GW…826541$10,000
GW…826547Under dispute
Show all 11 ⏷
Approve 11
Review eachEdit rules
After
123
Proposed
Retry Leg 2 settlement
Destructive action
$284,900 · 11 of 12 transactions
1 excluded: under dispute, can't verify
Review transactions11
Approve AllReview each
  1. 1Positioned as a checkpoint, not a commit
  2. 2The agent classified it destructive and gated itself
  3. 3The agent set the scope: 11 of 12
  4. 4Confident enough to batch; review is the hedge

The framework

The Escalation Ladder

Three levels of AI sit on one continuum. As you move up, four things shift together: who takes initiative, whether AI can change state, how much trust the interaction requires, and how reversible the output is. What makes it a ladder, not three separate products, is the transition between levels, allowing analysts to move up or down without losing context.

L1 · AMBIENT system speaks L2 · COPILOT dialogue L3 · AGENTIC propose → approve → act expand audit trail promote explain L3 · AGENTIC propose → approve → act L2 · COPILOT dialogue L1 · AMBIENT system speaks
L1 · Ambient Help built into the screens people already use. No chat, no agent.
Initiative
System
Turn-taking
None
Mutates state
Never
Trust
Low

Explain-this-status on a badge, proactive anomaly flags on odd rows, smart empty states, inline summaries inside the drawer that already exists.

L2 · Copilot A side panel that answers questions. It reads and reasons, nothing more.
Initiative
User
Turn-taking
Multi-turn
Mutates state
Never
Trust
Medium

Every answer shows its plan, the systems it checked and when, and how sure it is.

L3 · Agentic The agent proposes an action, a person approves it, then it runs.
Initiative
Shared
Turn-taking
Propose, approve, run
Mutates state
Yes
Trust
High

An action card with a plan you can preview, approve in bulk or one by one, stop mid-run, and undo. Every step is logged.

B100 payments operations

Five patterns

01Confidence signal

Confidence is stated in words, never a bare number. At L1, the Anomaly flag opens its evidence: $10,000 vs ~$3,100 typical · flagged by Gen 3. At L2, Confidence: medium with the reasoning behind it. At L3 it limits the action: unable to verify one of the 12 transactions, the agent retries 11 and says which one it left out.

02Citation & sourcing

Every agent claim is tied to its source, Gen 3 / Transfermate / Gen 2, with the time it was queried. Hover at L1, inline at L2, in the audit trail at L3.

03Permission model

Actions are typed. Safe (retrieval, drafts) needs no approval. Reversible (label a transaction, save a filter) needs a simple approval. Destructive (retry settlement, send external comms) needs approval with details.

04Graceful degradation

If confidence drops at L3, the agent proposes L2, "I'm not sure, want me to investigate?" If L2 can't answer, it proposes a human handoff rather than guessing.

05Escape hatch

When the agent makes a claim, the reason is always one tap away. Flags show why, answers show their sources, and actions keep Review each, Cancel, and Undo. The agent is never the only way to access the information or take action.

Trade-offs & open questions

Batch approval for destructive actions

Far faster for the 12-transaction case, but one "approve all" click carries more weight. Mitigated with per-item opt-out and undo, still open whether high-value transactions should force per-item approval.

Open, whose name is on the action?

Accountability blurs between the human who approved and the agent that executed. The audit trail captures the full chain, but the org still has to decide where responsibility sits.

[status] work in progress

This is an ongoing project, with new features coming soon.

Want to know more about this project?

The framework and the prototype behind it are written up on my Medium.