Removing protected characteristics from an underwriting model does not guarantee fair outcomes.
For years, removing protected characteristics from insurance data has been viewed as a simple solution to algorithmic bias. The assumption is that if a model does not access race, gender, or ethnicity, it cannot
Machine learning makes that assumption difficult to defend.
The
More data can make proxy bias harder to see
Auto insurance provides a familiar example. Geographic location can be closely associated with demographic composition.
A
That correlation does not establish that race itself caused the difference. Geography can capture legitimate differences in insurance risk. The concern arises when geographic and other variables carry demographic information into a model, contributing to outcomes that disproportionately affect protected groups.
The same problem has appeared in healthcare. A widely used algorithm used healthcare spending as a proxy for patients' health needs.
That matters as insurers move beyond traditional actuarial models and use increasingly complex machine learning systems. A model can identify relationships that were not obvious during feature selection or model design. More variables create more opportunities for those relationships to emerge.
Reviewing the feature list is only the starting point. Insurers also need to examine the outcomes those features produce across protected groups.
Regulators are looking beyond the model itself
U.S. insurance regulators are moving in that direction.
The
Colorado's Regulation 10-1-1 establishes governance and risk-management requirements for insurers that use external consumer data and information sources, including algorithms and predictive models that use such data.
Virginia has taken a similar approach. Its Administrative Letter 2024-01 sets expectations for insurers to govern and manage risks associated with the development, acquisition, and use of AI systems.
The
In the U.K., the Financial Conduct Authority's review of insurance firms' outcomes monitoring under the Consumer Duty says firms should monitor the outcomes customers receive and identify whether different groups are experiencing different outcomes.
For insurers, reviewing the variables going into an AI system is only part of the job. The resulting decisions need scrutiny too.
Explainability can reveal bias without removing it
This is where explainability tools such as SHAP (Shapley Additive exPlanations) and LIME (Local Interpretable Model-agnostic Explanations) come into play.
They can show which features influenced an individual prediction or help teams understand how a model behaves.
They do not, by themselves, establish that an outcome is fair.
A model can be transparent about the factors influencing its decision and still produce disparate outcomes. An explanation shows how features contributed to a prediction. It does not determine whether the resulting decision is fair.
Fairness needs to be addressed during model development and after deployment.
Taking three technical measures can help eliminate bias:
Debias the data before training. Sample weighting, rebalancing, and other preprocessing techniques can reduce the influence of historical disparities before a model learns from them.
Test for proxy relationships. An adversarial model can test whether protected characteristics can be inferred from the information available to the primary model. High inference accuracy can indicate that ostensibly neutral inputs carry demographic information.
Monitor outcomes after deployment. MLOps (machine learning operations) processes can track relevant fairness measures alongside model performance and business metrics. Changes in those measures can trigger further investigation.
The goal is not to reduce predictive accuracy, but to ensure that risk signals are not accompanied by hidden indicators of protected characteristics.
An insurer may remove race from datasets, restrict access to sensitive attributes, and document model features to explain individual decisions. But none of these steps tells the insurer whether the model is producing different outcomes for different groups.
This is where AI fairness in insurance must be evaluated.










