Financial services and credit
Scorecards and fraud models decide who is lent to. We test them on thin-file applicants and on households paid in cash.
We help operators, regulators, and the public understand the performance and risk of deployed AI systems.
Six domains where AI already carries weight in decisions at scale.
Scorecards and fraud models decide who is lent to. We test them on thin-file applicants and on households paid in cash.
Triage and screening decide who is seen first. We measure them against the presentations of the clinic where they will run.
Identity and eligibility systems establish who the state recognises. We measure how often recognition holds, and for whom.
Earth observation sits behind early warnings and insurance payouts. We establish how much of the population at risk it covers.
Detection and targeting systems mark who is treated as a threat. We test how often that mark falls on the wrong person.
Models are consulted as advisors before anyone certifies them. We measure how often they are confident and wrong, language by language.
Describe the system, the decision it takes part in, and the access you can give us. We reply within three working days.