Shared & manual
Begin with requirements, test design and evidence drafts. Establish stable fixtures and repeatable runs before expanding generation.
Our Banking Client / Illustrative banking scenario / Executive
A payment-release story: 75 offshore QA staff, heavy manual testing and constrained test environments.
One payment release.
A platform the next team can reuse.
Executive briefing • Strategic vision • Client scenario
Our Banking Client / Illustrative banking scenario / Executive
75 offshore QA staff across five squads; 45 focus on manual and domain testing. Environments and vendor test slots limit parallel work.
The pilot draws eight QA staff from this population. More engineers alone do not create more vendor test capacity.
Our Banking Client / Illustrative banking scenario / Executive
PAY-142 simulates a CAD 100 transfer accepted before a timeout. The customer retries; an injected duplicate-posting defect tests whether QE catches the second debit.
Payer CAD 800
Recipient CAD 200
Payer CAD 900
Recipient CAD 100
Expected: one transfer, one balanced journal and one customer-visible result. This is a QE test scenario, not a reported client incident.
Our Banking Client / Illustrative banking scenario / Executive
Modelled hands-on effort for one bounded release test pack. Eight stages add to 300 hours.
The pack spans 10 business days. Waiting time and human effort use different units and must be measured separately.
Our Banking Client / Illustrative banking scenario / Executive
Common context, fixtures and evidence serve every squad. Application teams retain business ownership.
Overview: proposed work handoffs.
Gold indicates the current handoff. All other arrows remain visible.
Existing test frameworks and pipelines execute approved work. AI supports preparation and interpretation.
Our Banking Client / Illustrative banking scenario / Executive
Overview · follow the numbered steps.
Gold = current focus. Arrows show a proposed handoff, not live activity.
Same payment scope. The two target states are proposed; no productivity or quality improvement has been measured.
Our Banking Client / Illustrative banking scenario / Executive
The same tool creates different work in a shared manual environment and a repeatable delivery system.
Begin with requirements, test design and evidence drafts. Establish stable fixtures and repeatable runs before expanding generation.
Pilot payment APIs and one web journey. Add synthetic fixtures, provider stubs and evidence links alongside AI assistance.
Extend generation and triage across more application teams. Evaluate test selection in shadow mode before changing required coverage.
These profiles illustrate effort, not adoption approval. Review twelve dependency areas before committing to scope, dates or benefits.
Our Banking Client / Illustrative banking scenario / Executive
Mixed-maturity target: 300 hours becomes 222, including 40 hours of review.
Purple segment = review. Total assisted effort = 222h.
Conditional capacity arithmetic; adoption prerequisites are unverified. Tool, cloud and vendor charges are excluded. This is not cash ROI or measured AI impact.
Our Banking Client / Illustrative banking scenario / Executive
| Owner | Accountability |
|---|---|
| Shared enablement | Own context adapters, templates, fixtures, runner integration and coaching. |
| Application QA teams | Own business scenarios, expected results, execution and defect disposition. |
| Product and release owners | Resolve requirement questions and accept the release evidence. |
| Delivery management | Provide repository access, overlap hours and an explicit backlog for freed capacity. |
Reskill domain testers through paired reviews. Agree supplier incentives and reuse ownership before scaling.
Our Banking Client / Illustrative banking scenario / Executive
Record actual pilot time from baseline through repeated API runs. Example windows remain subject to readiness.
Observation status: not yet recorded
Record API test preparation, active effort, waits and rework for one comparable pack.
Evidence gate: Reviewed API scope + usable baselineRecord setup separately; measure review, execution, triage and retest over repeated packs.
Evidence gate: Required assertions pass + comparable run evidenceMeasure second-team setup, support effort and repeat runs before choosing the next scope.
Evidence gate: Reusable pattern + evidence-based next stepPlanning windows are illustrative. Environment, data, virtualization and reviewer readiness determine when each phase can progress.
Our Banking Client / Illustrative banking scenario / Executive
| Measure | Expansion evidence |
|---|---|
| Quality | All mandatory payment assertions pass; no unexplained skips or weakened checks. |
| Effort | Include preparation, review, correction, retesting and platform operation. |
| Delivery | Report elapsed time, waiting reasons and failed-run recovery separately. |
| Reuse | A second squad runs the pattern and maintains it without the pilot authors. |
A faster draft is insufficient. Compare tasks of similar scope and retain conventional work as a reference.
Our Banking Client / Illustrative banking scenario / Executive
The next decision: assess one payment workflow with QA, development and platform leads.
Agree the bounded pilot and its evidence before promising wider rollout. See the discovery brief for owners, commercial scope and timing.
Our Banking Client / Illustrative banking scenario / Executive
Drafting
Approved rules and observable outcomes.
Diagnosis only
Retained failures, known dispositions and usable diagnostics.
Execution
Add repeatable environments, CI, fixtures, provider control and reliable runners.
Our Banking Client has not been assessed. The dependency backlog can change scope, cost and timing; the 480-hour setup allowance is not a client quote.
Our Banking Client / Illustrative banking scenario / Executive
Reviewed scenarios, expected outcomes and ownership
AI drafts with domain reviewA reviewer can identify a wrong answerReliable suites, isolated data and controlled dependencies
AI drafts tests that run in CIAnother engineer reproduces the resultReusable environments, contracts, artifacts and support
Assistance works across teamsA second team onboards and operates itEvaluated tools, bounded actions and recovery
Agents execute within approved scopeFailures stop, evidence persists, people can interveneFor Our Banking Client, fund repeatable payment tests and shared QE services alongside reviewed AI assistance. Validate this capability roadmap through client discovery.
Our Banking Client / Illustrative banking scenario / Executive
Libra Internet Bank
Test-creation timeIndex: previous time = 100
UiPath customer case. Normalized from the reported reduction; review effort is unspecified. F1
Goldman Sachs
Unit-test coverageShare of a selected module (%)
Diffblue customer case. Coverage measures exercised code; it does not establish financial correctness. F2
Fiserv
Major incidentsIndex: previous year = 100
Tricentis modernization case. AI testing was a later pilot; this is not an AI-attributed reduction. F4
Peer results support a trial. Outcomes for Our Banking Client remain unmeasured.
Our Banking Client / Illustrative banking scenario / Executive
Requirements ready
Draft scenarios in the current test-management workflow.
ProofAccepted coverage and measured review effort.
Execution repeatable
Generate candidates in the existing framework and CI pipeline.
ProofCorrect behavior passes; a known fault fails.
Diagnostics available
Draft diagnoses alongside the existing disposition process.
ProofFaster correct decisions with traceable evidence.
Choose one entry point, validate its prerequisites and agree how reviewers will judge the result. Wider rollout depends on local evidence.
Our Banking Client / Illustrative banking scenario / Executive
Authored roadmap. Expand with validated prerequisites and client evidence.
Our Banking Client / Illustrative banking scenario / Executive
Before / scenario baseline
Backend services are not virtualized
Vendor test slots limit parallel runsAfter / pilot target
Context + fixtures + runners + evidence
AI drafts and explains · people reviewQE modernization: isolated runs and controlled dependencies. AI assistance: preparation and interpretation.
Starting point: heavy manual QE and scarce vendor environments. Pilot target: controlled service tests, with real integration still required.
Our Banking Client / Illustrative banking scenario / Executive
Catch every critical seeded fault
Reviewed fault scenarios and independent balance assertions
Exercise every critical scenario
AI drafts edge cases; domain QA approves the scenario matrix
Complete every decision record
CI captures evidence; AI drafts a source-linked summary
Baseline: Not recorded · Observed after: Not recorded · Targets proposed for pilot agreement.
Business aims: fewer escaped payment defects and emergency fixes. These proposed pilot measures do not establish those outcomes.
Our Banking Client / Illustrative banking scenario / Executive
Overview · follow the numbered steps.
Gold = current focus. Arrows show a proposed handoff, not live activity.
PAY-142 is an authored scenario. Each AI output needs review; test runners execute and people decide.
Our Banking Client / Illustrative banking scenario / Executive
Overview · follow the numbered steps.
Gold = current focus. Arrows show a proposed handoff, not live activity.
Illustrative artifacts and outcomes. Preserve failed evidence; a passing retest alone does not approve release.
Prepared by Tom Wu · Contact / feedback. AI-generated illustrations and synthetic English narration. Diagrams and scenario models are authored explanations; sources retain their own attribution. Narration provenance.