Robust, not fragile

Cross-Run Results

Every simulation run side by side — which outcome patterns hold across seeds and which are per-run noise.

01

Executive summary

12 of 14 outcome metrics hold across every run — 10 for Track A, 0 for Track B, 2 tied.

Track A owns Translation debt, Exception rate, Handoff failure rate, Supplement requests, Cost per claim, Cycle time, Trust polarization, First-pass accuracy, Customer retention, and Policy bind rate. Missed recovery opportunities and Recovery dollars are a dead heat. Underwriting cycle time and Risk selection accuracy flip between runs — real noise, not signal.

02

Metric winners by runseed 42 · 43 · 44

Three runs compared. A pattern is marked stable when the same track wins it in every run. If it holds across seeds, the result is structural, not a fluke.

MetricSeed 42Seed 43Seed 44Stable
Translation debtMeaning lost when work passes between steps.Track A loses less information at handoffs25 vs 28A25 vs 31A29 vs 32A
Exception rateHow often work hits an exception needing a human decision.Track A hits fewer exceptions15 vs 25A19 vs 30A23 vs 29A
Handoff failure rateHow often a handoff between steps breaks.Track A has fewer broken handoffs40 vs 50A46 vs 58A49 vs 53A
Supplement requestsHow often agents ask for missing information.Track A requests fewer supplements64 vs 98A52 vs 112A81 vs 132A
Cost per claimDollars to process a claim, including rework.Track A spends less per claim400 vs 434A385 vs 434A501 vs 542A
Cycle timeDays from claim start to finish.Track A finishes claims faster5 vs 5A4 vs 5A7 vs 7A
Trust polarizationHow divided employees are about the AI.Track A has less internal division about AI3 vs 3A2 vs 3A3 vs 3A
Underwriting cycle timeTime to issue a policy.Track A issues policies faster4 vs 4B4 vs 4B4 vs 4A
Missed recovery opportunitiesRecovery chances subrogation left on the table.Track A misses fewer recovery chances8 vs 8=5 vs 5=3 vs 3=
First-pass accuracyShare of claims handled right the first time, no rework.Track A gets more claims right the first time57 vs 45A52 vs 37A28 vs 21A
Customer retentionShare of customers who stay after a claim.Track A keeps more customers80 vs 76A79 vs 74A73 vs 71A
Risk selection accuracyHow accurately risks are priced.Track A prices risks more accurately82 vs 92B89 vs 87A87 vs 88B
Policy bind rateHow often a quoted policy is bound.Track A closes more policies90 vs 85A88 vs 87A85 vs 84A
Recovery dollarsDollars recovered through subrogation.Track A recovers more dollars4199 vs 4199=4921 vs 4921=5394 vs 5394=
Employee AI trustHow much employees trust the AI (0–10).Trust stays roughly flat — the divergence is in outcomes, not sentiment4.9 vs 5.3B5.0 vs 5.0A4.4 vs 4.5B

A = Track A wins  ·  B = Track B wins  ·  = tied. Stable marks a pattern that holds in every run.