Usage-Based Insurance, Algorithmic Fairness, and Risk-Based Pricing: A Socio-Technical Perspective on Telematics Regulation

DOI:

https://doi.org/10.63646/SWHP6298

Keywords:

Usage-based insurance; algorithmic fairness; telematics regulation; generalized additive models; risk-based pricing; territory clustering; proxy discrimination

Abstract

The rapid proliferation of telematics technology has fundamentally transformed auto insurance pricing through Usage-Based Insurance (UBI) programmes, which link premium rates to directly measured driving behaviour. While these developments offer significant potential for actuarial precision and risk differentiation, they simultaneously introduce profound challenges related to algorithmic fairness, proxy discrimination, and regulatory transparency. This paper adopts a socio-technical perspective to examine how telematics-derived variables interact with socio-economic and demographic factors in ways that may systematically disadvantage certain policyholder groups. Drawing on a synthetic UBI dataset, we apply Generalized Additive Models (GAMs) and Generalized Linear Models (GLMs) to model both claim frequency and severity, and we employ interpretable K-means clustering to develop a territory risk classification framework. Our analysis demonstrates that while annual miles driven, credit score, and years without claims are the most statistically significant predictors of insurance risk, their joint effects can create indirect disparate impacts across socio-economic strata. We further conduct a fairness decomposition analysis that distinguishes direct risk-based differentiation from indirect proxy discrimination. On the regulatory side, we propose a three-pillar governance framework—transparency, auditability, and adaptive oversight—designed to reconcile the competing imperatives of actuarial fairness and social equity in UBI pricing. Our findings contribute to an emerging interdisciplinary literature at the intersection of actuarial science, algorithmic fairness, and insurance regulation.

How to Cite

Xu, X., Wang, H., & Wu, J. (2026). Usage-Based Insurance, Algorithmic Fairness, and Risk-Based Pricing: A Socio-Technical Perspective on Telematics Regulation. Journal of Technology Innovation and Society, 2(2), 93-110. https://doi.org/10.63646/SWHP6298

References

Antonio, K., & Valdez, E. A. (2012). Statistical concepts of a priori and a posteriori risk classification in insurance. AStA Advances in Statistical Analysis, 96(2), 187–224. https://doi.org/10.1007/s10182-011-0152-7

Arthur, D., & Vassilvitskii, S. (2007). K-means++: The advantages of careful seeding. In Proceedings of the Eighteenth Annual ACM-SIAM Symposium on Discrete Algorithms (pp. 1027–1035).

Arumugam, S., & Bhargavi, R. (2019). A survey on driving behavior analysis in usage-based insurance using big data. Journal of Big Data, 6, Article 86. https://doi.org/10.1186/s40537-019-0249-5

Aseervatham, V., Lex, C., & Spindler, M. (2016). How do unisex rating regulations affect gender differences in insurance premiums? The Geneva Papers on Risk and Insurance—Issues and Practice, 41(1), 128–160. https://doi.org/10.1057/gpp.2015.22

Ayuso, M., Guillen, M., & Pérez-Marín, A. M. (2016). Telematics and gender discrimination: Some usage-based evidence on whether men's risk of accidents differs from women's. Risks, 4(2), Article 10. https://doi.org/10.3390/risks4020010

Baecke, P., & Bocca, L. (2017). The value of vehicle telematics data in insurance risk selection processes. Decision Support Systems, 98, 69–79. https://doi.org/10.1016/j.dss.2017.04.009

Barocas, S., Hardt, M., & Narayanan, A. (2023). Fairness and machine learning: Limitations and opportunities. MIT Press.

Baudry, M., & Robert, C. Y. (2019). A machine learning approach for individual claims reserving in insurance. Applied Stochastic Models in Business and Industry, 35(5), 1127–1155. https://doi.org/10.1002/asmb.2455

Boucher, J.-P., Denuit, M., & Guillen, M. (2007). Risk classification for claim counts: A comparative analysis of various zero-inflated mixed Poisson and hurdle models. North American Actuarial Journal, 11(4), 110–131. https://doi.org/10.1080/10920277.2007.10597487

Boucher, J.-P., Côté, S., & Guillen, M. (2017). Exposure as duration and distance in telematics motor insurance using generalized additive models. Risks, 5(4), Article 54. https://doi.org/10.3390/risks5040054

Calders, T., & Verwer, S. (2010). Three naive Bayes approaches for discrimination-free classification. Data Mining and Knowledge Discovery, 21(2), 277–292. https://doi.org/10.1007/s10618-010-0190-x

Charpentier, A. (Ed.). (2014). Computational actuarial science with R. CRC Press. https://doi.org/10.1201/b17230

Chouldechova, A. (2017). Fair prediction with disparate impact: A study of bias in recidivism prediction instruments. Big Data, 5(2), 153–163. https://doi.org/10.1089/big.2016.0047

Corbett-Davies, S., & Goel, S. (2018). The measure and mismeasure of fairness: A critical review of fair machine learning. arXiv. https://doi.org/10.48550/arXiv.1808.00023

Denuit, M., Guillen, M., & Trufin, J. (2019). Multivariate credibility modelling for usage-based motor insurance pricing with behavioural data. Annals of Actuarial Science, 13(2), 378–399. https://doi.org/10.1017/S1748499518000349

Dwork, C., Hardt, M., Pitassi, T., Reingold, O., & Zemel, R. (2012). Fairness through awareness. In Proceedings of the 3rd Innovations in Theoretical Computer Science Conference (pp. 214–226). https://doi.org/10.1145/2090236.2090255

Eling, M., & Lehmann, M. (2018). The impact of digitalization on the insurance value chain and the insurability of risks. The Geneva Papers on Risk and Insurance—Issues and Practice, 43(3), 359–396. https://doi.org/10.1057/s41288-017-0073-0

Ferrario, A., Noll, A., & Wüthrich, M. V. (2020). Insights from inside neural networks. SSRN Electronic Journal. https://doi.org/10.2139/ssrn.3226852

Frees, E. W., & Valdez, E. A. (2008). Hierarchical insurance claims modeling. Journal of the American Statistical Association, 103(484), 1457–1469. https://doi.org/10.1198/016214508000000823

Frees, E. W., Meyers, G., & Cummings, A. D. (2014). Insurance ratemaking and a Gini index. Journal of Risk and Insurance, 81(2), 335–366. https://doi.org/10.1111/j.1539-6975.2012.01507.x

Gao, G., & Wüthrich, M. V. (2019). Convolutional neural network classification of telematics car driving data. Risks, 7(1), Article 6. https://doi.org/10.3390/risks7010006

Gao, G., Wang, H., & Wüthrich, M. V. (2022). Boosting Poisson regression models with telematics car driving data. Machine Learning, 111(1), 243–272. https://doi.org/10.1007/s10994-021-05957-0

Guelman, L. (2012). Gradient boosting trees for auto insurance loss cost modeling and prediction. Expert Systems with Applications, 39(3), 3659–3667. https://doi.org/10.1016/j.eswa.2011.09.058

Guillen, M., Nielsen, J. P., Ayuso, M., & Pérez-Marín, A. M. (2019). The use of telematics devices to improve automobile insurance rates. Risk Analysis, 39(3), 662–672. https://doi.org/10.1111/risa.13172

Hardt, M., Price, E., & Srebro, N. (2016). Equality of opportunity in supervised learning. Advances in Neural Information Processing Systems, 29, 3315–3323.

Hastie, T. J., & Tibshirani, R. J. (1990). Generalized additive models. Chapman & Hall.

Henckaerts, R., Antonio, K., Clijsters, M., & Verbelen, R. (2018). A data-driven binning strategy for the construction of insurance tariff classes. Scandinavian Actuarial Journal, 2018(8), 681–705. https://doi.org/10.1080/03461238.2018.1429300

Huang, Y., & Meng, S. (2019). Automobile insurance classification ratemaking based on telematics driving data. Decision Support Systems, 127, Article 113156. https://doi.org/10.1016/j.dss.2019.113156

Kaufman, L., & Rousseeuw, P. J. (1990). Finding groups in data: An introduction to cluster analysis. John Wiley & Sons. https://doi.org/10.1002/9780470316801

Kieschnick, R., & McCullough, B. D. (2003). Regression analysis of variates observed on (0, 1): Percentages, proportions and fractions. Statistical Modelling, 3(3), 193–213. https://doi.org/10.1191/1471082X03st053oa

Kou, G., & Lu, Y. (2025). FinTech: A literature review of emerging financial technologies and applications. Financial Innovation, 11(1), 1–34. https://doi.org/10.1186/s40854-024-00668-6

Lindholm, M., Richman, R., Tsanakas, A., & Wüthrich, M. V. (2022). Discrimination-free insurance pricing. ASTIN Bulletin, 52(1), 55–89. https://doi.org/10.1017/asb.2021.23

Lou, Y., Caruana, R., & Gehrke, J. (2012). Intelligible models for classification and regression. In Proceedings of the 18th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (pp. 150–158). https://doi.org/10.1145/2339530.2339556

Lu, Y. (2019). Artificial intelligence: A survey on evolution, models, applications and future trends. Journal of Management Analytics, 6(1), 1–29. https://doi.org/10.1080/23270012.2019.1570365

Lu, Y., & Xu, L. D. (2019). Internet of Things (IoT) cybersecurity research: A review of current research topics. IEEE Internet of Things Journal, 6(2), 2103–2115. https://doi.org/10.1109/JIOT.2018.2869847

MacQueen, J. (1967). Some methods for classification and analysis of multivariate observations. In Proceedings of the Fifth Berkeley Symposium on Mathematical Statistics and Probability (Vol. 1, pp. 281–297).

Mayer, M., Meier, D., & Wüthrich, M. V. (2023). SHAP for actuaries: Explain any model. SSRN Electronic Journal. https://doi.org/10.2139/ssrn.4389797

Mehrabi, N., Morstatter, F., Saxena, N., Lerman, K., & Galstyan, A. (2021). A survey on bias and fairness in machine learning. ACM Computing Surveys, 54(6), Article 115, 1–35. https://doi.org/10.1145/3457607

Molnar, C. (2022). Interpretable machine learning: A guide for making black box models explainable (2nd ed.). Independently published.

Noll, A., Salzmann, R., & Wüthrich, M. V. (2020). Case study: French motor third-party liability claims. SSRN Electronic Journal. https://doi.org/10.2139/ssrn.3164764

Nori, H., Jenkins, S., Koch, P., & Caruana, R. (2019). InterpretML: A unified framework for machine learning interpretability. arXiv. https://doi.org/10.48550/arXiv.1909.09223

Ohlsson, E., & Johansson, B. (2010). Non-life insurance pricing with generalized linear models. Springer. https://doi.org/10.1007/978-3-642-10791-7

Paefgen, J., Staake, T., & Thiesse, F. (2013). Evaluation and aggregation of pay-as-you-drive insurance rate factors: A classification analysis approach. Decision Support Systems, 56, 192–201. https://doi.org/10.1016/j.dss.2013.06.001

Pesantez-Narvaez, J., Guillen, M., & Alcañiz, M. (2019). Predicting motor insurance claims using telematics data—XGBoost versus logistic regression. Risks, 7(2), Article 70. https://doi.org/10.3390/risks7020070

Richman, R. (2021). AI in actuarial science—A review of recent advances—Part 1. Annals of Actuarial Science, 15(2), 207–229. https://doi.org/10.1017/S1748499520000238

Shi, P., Feng, X., & Boucher, J.-P. (2016). Multilevel modeling of insurance claims using copulas. The Annals of Applied Statistics, 10(2), 834–863. https://doi.org/10.1214/16-AOAS914

So, B., Boucher, J.-P., & Valdez, E. A. (2021). Synthetic dataset generation of driver telematics. Risks, 9(4), Article 58. https://doi.org/10.3390/risks9040058

Taylor, G., & McGuire, G. (2004). Loss reserving with GLMs: A case study. Casualty Actuarial Society Discussion Paper Program, 327–392.

Tselentis, D. I., Yannis, G., & Vlahogianni, E. I. (2016). Innovative insurance schemes: Pay as/how you drive. Transportation Research Procedia, 14, 362–371. https://doi.org/10.1016/j.trpro.2016.05.088

Verbelen, T., Antonio, K., & Claeskens, G. (2018). Unravelling the predictive power of telematics data in car insurance pricing. Journal of the Royal Statistical Society: Series C (Applied Statistics), 67(5), 1275–1304. https://doi.org/10.1111/rssc.12283

Verma, S., & Rubin, J. (2018). Fairness definitions explained. In Proceedings of the International Workshop on Software Fairness (pp. 1–7). https://doi.org/10.1145/3194770.3194776

Wood, S. N. (2017). Generalized additive models: An introduction with R (2nd ed.). CRC Press. https://doi.org/10.1201/9781315370279

Wüthrich, M. V., & Buser, C. (2016). Data analytics for non-life insurance pricing. Swiss Finance Institute Research Paper No. 16-68. https://doi.org/10.2139/ssrn.2870308

Xin, X., & Huang, F. (2024). Antidiscrimination insurance pricing: Regulations, fairness criteria, and models. North American Actuarial Journal, 28(2), 285–319. https://doi.org/10.1080/10920277.2023.2190528

Xu, L. D., Lu, Y., & Li, L. (2021). Embedding blockchain technology into IoT for security: A survey. IEEE Internet of Things Journal, 8(13), 10452–10473. https://doi.org/10.1109/JIOT.2021.3060508

Zhang, C., & Lu, Y. (2021). Study on artificial intelligence: The state of the art and future prospects. Journal of Industrial Information Integration, 23, Article 100224. https://doi.org/10.1016/j.jii.2021.100224