No model-risk evidence
The POC produces no CP24/2-aligned model card and no traceability matrix for second-line risk. Without that evidence, the system never clears review and never reaches production.
An agentic pipeline that classifies, scores, and routes inbound claims, with the per-decision audit trail FCA CP24/2 expects.
Every workload ships the same citation-bound audit trail and inline fraud signal. Start with the line of business costing you most.
Most insurers have run claims-AI POCs; most stalled. The pattern is consistent: the POC clears engineering but fails at second-line risk review.
Compliance is not a model feature. It's the end-to-end claims system.
The POC cleared engineering and died at second-line risk review.
The POC produces no CP24/2-aligned model card and no traceability matrix for second-line risk. Without that evidence, the system never clears review and never reaches production.
Even when the model auto-routes, adjusters re-review every decision 'just in case', eroding the entire business case. The AI becomes overhead instead of the leverage it promised.
Triage classifies the claim while fraud-scoring sits in a separate vendor tool. Suspicious claims slip through the gap because neither system ever sees the other's signal.
When the FCA asks why a claim was routed the way it was, the team can't reconstruct the agent's full reasoning trajectory, and confidence in the whole pipeline collapses.
Not features: outcomes a head of claims, an SIU lead, or a second-line risk function can defend in the room.
Each classification, score, and routing decision is bound to its inputs and reasoning, and replayable on demand. “Why was this claim routed this way?” has a sub-minute answer, not a forensic project.
Adjusters stop re-keying and triaging the obvious. 60–70% of claims auto-route; only complex, fraud-flagged, and material-injury claims reach a specialist, and every override sharpens the next quarter's model.
Motor, home, and SME commercial claims as they actually arrive (paper FNOLs, broker submissions, messy free text), not a cleaned demo subset. The agent matches loss-adjuster decisions on 96–98% of the standard tier against hold-out sets.
Regulator evidence is engineering output, not a final-stage scramble. Every build ships CP24/2-aligned model cards, a SAR-bearing decision pathway, and an on-prem sovereign fallback for data that can't leave the estate.
Not features: four reasons a head of claims or second-line risk lead can defend in the room.
CP24/2 model-risk evidence ships as engineering output, not a late scramble.
Every routing decision is audit-logged and replayable in under a minute.
A phased path from demo to 2LoD-approved production.
60–70% of claims auto-route; specialists handle only the complex and suspicious.
Five stages, each with a bounded output you can hold us to. The CP24/2 evidence runs in parallel with engineering, never as a final-stage scramble.
We audit the current triage flow, design the CP24/2 evidence harness, lock the agent specification, and get the model-card template signed off by second-line risk, before any code ships.
Classify, complexity-score, fraud-score, route, plus the retrieval layer and evaluation harness, with daily hold-out evaluation and prompt-regression detection from day one.
CP24/2 model cards, the FCA-format traceability matrix, per-decision trajectory replay, and pen-test, produced as engineering output alongside the build, not after it.
Run in shadow against the live FNOL stream for two weeks, clear final 2LoD review, then submit the FCA notification with the parity report attached.
Phased rollout (10% → 50% → 100%), then extend to commercial claims and SIU fraud triage under quarterly retrain governance.
Success stories
BritonOne Technology is a full-cycle engineering company that builds and operates production software for regulated estates. Since 2017, we have shipped programmes that clear audit on the first pass across banking, insurance, wealth, healthcare, and biotech. Our teams pair deep domain knowledge with disciplined engineering, treating compliance, security, and resilience as first-class deliverables. From architecture through to live operations, we stay accountable for the systems we build, measuring success by uptime, audit outcomes, and defensible business results.
Anonymised under MNDA, verifiable on reference call. Each quote is from a senior owner who carried the engagement through second-line review.
“Three previous internal AI projects had failed at second-line risk review. BritonOne Technology shipped the FCA-cleared agentic pipeline in eleven weeks. Twelve months later, 68% of motor claims auto-route end-to-end and we have taken £4.2M of loss-adjuster cost out of the run-rate. The most surprising thing: the FCA supervisory meeting was almost boring.”
