How Proxy Race Distorts Regression-Based Fairness Audits
作者
Authors
Xi Xin|Giles Hooker|Fei Huang
期刊
Journal
暂无期刊信息
年份
Year
2026
分类
Category
国家
Country
日本Japan
📝 摘要
Abstract
Proxy-based race inference is increasingly used to conduct fairness assessments when protected-class data are unavailable or legally restricted -- most prominently in U.S. fair-lending enforcement, and now explicitly contemplated in emerging insurance regulation, including Colorado's draft SB21-169 testing framework and New York's Insurance Circular Letter No. 7. Despite this growing regulatory relevance, little is known about how standard regression-based discrimination analyses behave when race is measured with error through proxies such as Bayesian Improved Surname Geocoding (BISG) or Bayesian Improved First Name and Surname Geocoding (BIFSG). This paper studies the consequences of using proxy-imputed race as a categorical regressor in regression-based fairness assessments. Treating proxy race as a categorical covariate subject to misclassification, we show that proxy-based coefficients become weighted mixtures of true group effects, systematically shrinking estimated disparities toward the majority group -- even when overall classification accuracy is high. Empirically, using a linked North Carolina voter-insurance dataset with self-reported race and ZIP-level auto insurance premiums, we demonstrate two mechanisms through which it distorts inference: (i) the intrinsic mixing of group effects implied by misclassification, and (ii) structured errors that vary with ZIP-level racial composition and socioeconomic conditions and remain correlated with pricing residuals after controls. As a result, regression-based disparity estimates can be attenuated or amplified relative to analogous analyses based on self-reported race. Our findings caution against treating proxy race as a plug-in substitute in regulatory testing and highlight design implications for proxy-based audit frameworks in insurance and other high-stakes domains.
📊 文章统计
Article Statistics
基础数据
Basic Stats
372
浏览
Views
0
下载
Downloads
44
引用
Citations
引用趋势
Citation Trend
阅读国家分布
Country Distribution
阅读机构分布
Institution Distribution
月度浏览趋势
Monthly Views