AI and the new frontiers in the insurance sector

Column

Artificial Intelligence

Time Series

Technology

April 30, 2025

AI and the new frontiers in the insurance sector

Data science techniques improve prediction accuracy, optimize processes, and create fairer and more sustainable products

The current era is based on data analysis, which imposes unprecedented challenges and opportunities for the insurance sector. Traditional factors continue to influence claims and policy pricing, while new technologies, such as artificial intelligence, remote sensing and telematics data analysis, transform risk assessment and management. This article explores recent advances in claims and pricing insurance modeling, focusing on how data science techniques improve prediction accuracy, optimize processes, and create fairer and more sustainable products.

The insurance industry has always been linked to statistical analysis. For decades, policy pricing and claims forecasting were based on traditional models such as linear regressions and generalized linear models (GLM), which offer high interpretability in scenarios with linear relationships between variables.

However, the complexity of modern risks, driven by climate change, new mobility patterns, and technological transformations, has highlighted the limitations of conventional methods. Many phenomena began to require models capable of capturing complex relationships and hidden patterns in the data.

In this context, techniques such as Random Forest, XGBoost, and deep neural networks began to be incorporated into the claims prediction process, offering relevant gains in accuracy. Another important movement was the expansion of data sources, including telemetry, satellite imagery, and climate data to enrich risk models.

The widespread adoption of these methodologies still faces challenges such as the need to ensure model interpretability, protect sensitive data, and adapt practices to regulatory requirements.

Car insurance

The automotive insurance sector has been one of the most dynamic in adopting new technologies. With detailed telematics data on speed, acceleration, and driving patterns, insurers have a richer basis for understanding driver behavior.

Recent studies show that the integration of telematics data surpasses traditional variables in accident prediction. Gao, Meng, and Wüthrich (2019) demonstrated that variables derived from speed and acceleration heatmaps were more effective than demographic factors in modeling claim frequency. Subsequently, Gao, Wang, and Wüthrich (2022) showed that combining this data with machine learning significantly improves predictive accuracy.

Yu et al. (2021) used neural networks optimized by genetic algorithms to predict automotive claims, achieving accuracy above 95%. Decision tree-based models have also consolidated as robust alternatives, with Hanafy and Ming (2021) demonstrating that Random Forest obtained 86.77% accuracy in predicting claims.

The search for models that combine accuracy with interpretability has led to architectures such as TabNet. McDonnell et al. (2023) observed that TabNet outperformed traditional models in recall and F1-score, offering a balance between performance and transparency.

The sector also incorporates alternative sources of information, such as social networks, which can increase the accuracy of automotive warranty request prediction by up to 21.9% (Shokouhyar et al., 2021).

Agricultural insurance

In rural insurance, information asymmetry represents a significant challenge. Climate unpredictability and the difficulty of large-scale monitoring limit the efficiency of traditional models. In this scenario, the combination of artificial intelligence with remote sensing emerges as an innovative solution.

Barros and Freitas (2023) integrated satellite imagery with machine learning algorithms to predict agricultural claims. Analyzing over 9,500 contracts in Paraná, the study identified Random Forest as the most effective model, with an accuracy of 71.35%.

The differential of this approach lies in the reduction of informational asymmetries. With satellite images and adequate predictive models, it is possible to continuously monitor crops and anticipate risks with greater precision than traditional methods.

This methodology also promotes more inclusive practices, allowing small and medium-sized producers to be evaluated more fairly, based on objective evidence.

Life insurance and pension

The life insurance and pension segment also incorporates advances in risk modeling. These products require long-term forecasts, sensitive to demographic and economic changes.

Gonçalves and Pandolfi (2024) demonstrated the effectiveness of models of time series by comparing linear regression and ARIMA models. Using ten years of data from major Brazilian insurance companies, they concluded that the regression model showed a lower mean squared error, proving to be slightly superior.

Já Neves, Fernandes e Melo (2014) proposed a more elaborate statistical model for predicting investment redemption rates, combining different techniques: they used generalized linear models (GLM) to understand how explanatory variables (those that have the potential to influence or predict the response of an experiment) influence redemptions, ARMA-GARCH processes to simultaneously model the trend and volatility of these rates over time, and elliptical copulas to analyze the relationship between redemptions and financial market performance. This approach allowed them to observe that redemption rates increase when the stock market performs worse, indicating an inverse correlation. The study reinforced the idea that sophisticated statistical modeling is essential for adequately assessing the financial risks associated with these products.

Marine and climate insurance

In the marine insurance sector, data integration in real-time has become essential. Operations in dynamic environments are impacted by weather conditions and rapidly changing factors, requiring models capable of capturing this volatility.

Adland et al. (2021) demonstrated the potential of meteorological data combined with records from the Automatic Identification System of Ships to predict naval accidents. Analyzing over 42,000 voyages in the North Pacific, they employed models such as LASSO regression and XGBoost, showing that the inclusion of this data significantly improved predictive capability.

This approach represents a qualitative leap compared to traditional methodologies. With high-frequency data, it is possible to adjust prices more precisely and implement risk mitigation measures in real-time. These practices have potential application for other types of insurance affected by environmental variables, pointing to a future with more adaptive pricing and management.

Trends and challenges

The technological transformation in the insurance sector points to a dynamic and challenging future. The adoption of advanced models offers gains in efficiency and precision, but issues arise related to ethics, governance of data and regulation.

A promising trend is the development of discrimination free models. Lindholm et al. (2024) proposed multitask neural networks to calculate prices without using sensitive variables or their proxies, ensuring that the algorithms do not perpetuate historical biases .

Another challenge is the balance between personalization and predictive robustness. Hosein (2024) highlighted that, although personalization brings competitive advantages, it increases the risk of overfitting (when the model adjusts too much to the training data, losing its ability to generalize to new data) and can reduce the robustness of the models in out-of-sample scenarios.

The interpretability of models also becomes increasingly relevant. Solutions like TabNet demonstrate that it is possible to develop accurate models without sacrificing the ability to explain predictive decisions.

The use of alternative data sources requires careful management of privacy and security of information. Compliance with regulations such as LGPD and GDPR will be increasingly central to the sustainability of data-driven business models. In summary, the next frontier for the sector will not only be technological, but also ethical and regulatory.

The evidence presented shows that the combination between data science and technological innovation generates substantial gains in precision, efficiency, and sustainability. The sector is advancing towards more personalized, fair, and resilient practices.

However, technical advancements bring new challenges. The need to ensure algorithmic equity, data protection, transparency of models, and regulatory compliance imposes a strategic agenda beyond mere technology adoption. Innovating means not only improving predictive capability but building ethical and responsible business models.

In an increasingly competitive and data-driven market, insurers that can align technological innovation with solid governance practices will have a significant strategic advantage.

To access the references of this text click here.

Who wrote this column

José Erasmo Silva

José Erasmo Silva é professor, formado em Matemática e Administração, com mais de 25 anos de experiência em gestão empresarial e de pessoas. É mestre e doutor em Administração, com foco em Finanças, e especialista em Data Science e Analytics e em Finanças e Controladoria. Realizou pós-doutorado na Universidade Federal da Bahia (UFBA). Atualmente, atua como professor orientador no MBA em Data Science, Inteligência Artificial e Analytics da USP/Esalq e leciona na EEP/FUMEP e na rede estadual de ensino de São Paulo.

You may also like

October 05, 2026

Prediction of default on tax debt installments

The installment payment of tax debts constitutes a relevant fiscal recovery instrument, allowing taxpayers to regularize their obligations in installments, while exposing the tax administration to the risk of cancellation due to non-compliance. The objective was to develop and evaluate machine learning models for predicting the cancellation of ICMS installments, aiming to support proactive collection strategies. A database of historical installment plans from the Secretariat of Economy of the State of Goiás was used. Decision Tree, Random Forest, and XGBoost algorithms were trained and compared, with hyperparameter optimization via GridSearchCV and evaluation by confusion matrices and ROC curves. XGBoost showed superior performance, with an AUC of 0.869 for the model that included the variable Number of Installments and 0.778 without it, highlighting the centrality of this feature. Applied to the active portfolio of R$ 2.752 billion, the model identified that 92.9% of the total value presented a risk of cancellation equal to or greater than 50%, with R$ 1.488 billion concentrated in extreme risk installments, a result consistent with the historical behavior of cancellations in the initial phases of agreements. The results demonstrated that the adoption of predictive models in the management of tax installments enables the transition from a reactive stance to a preventive approach, with the potential to substantially increase fiscal recovery.

Keywords: Tax administration; Machine learning; Binary classification; Fiscal recovery; XGBoost.

October 05, 2026

Interactive system for generation and comparison of predictive models of monthly rural credit concessions

The ability to predict the future volume of rural credit concessions is essential for efficient resource allocation and setting disbursement targets. The work aimed to develop an interactive web platform to generate, analyze, and compare predictive models of time series of rural credit concessions, covering the period from March 2011 to January 2026. Econometric techniques (ARIMA, SARIMA, SARIMAX) and linear regression were confronted with machine learning methods (Random Forest, XGBoost). The methodology included collecting monthly data on rural credit concessions, macroeconomic variables, and agricultural commodity indicators, followed by exploratory analysis and platform development in a client-server architecture with Python and web technologies. The platform’s application to rural credit forecasting in three scenarios (total, individual, and corporate) evaluated by temporal cross-validation revealed that no technique proved universally superior. Linear regression models with seasonal lags showed the most consistent results across all folds, maintaining stable performance even in periods of level shift, where decision tree algorithms registered significant degradation. The corporate scenario showed low predictability in all models. The results demonstrated that the platform fulfilled its objective, revealing that algorithmic complexity does not guarantee predictive superiority and that the choice of model should be guided by the series’ characteristics and the application context.

Keywords: Agribusiness; Forecasting; Machine learning; Predictive modeling; Time series.

October 05, 2026

Music festivals as a platform for brand engagement with Gen Z consumers

Music festivals have consolidated themselves as complex experience ecosystems, where the convergence between physical entertainment and digital narratives redefines brand positioning strategies. In the post-pandemic scenario, the events sector showed a significant recovery, establishing itself as a strategic platform for engaging with Generation Z, an audience that prioritizes authenticity and shared experiences over traditional advertising formats. The study aimed to analyze how brand activations in these environments impacted the engagement of consumers born between 1995 and 2010. The methodology was characterized by a quantitative and descriptive research, conducted through the application of a structured questionnaire that obtained the participation of 163 respondents. The results showed that experience marketing strategies, especially those that integrated aesthetic attributes and digital sharing potential, presented the highest averages of positive perception. It was found that brand trust was strengthened after successful physical interactions, revealing a symbiosis between the emotional environment of the event and the young person’s digital journey. In contrast, the perception of authenticity mediated by digital influencers obtained the lowest agreement index. It was concluded that Generation Z’s engagement was enhanced by hybrid strategies that allowed consumers to take on the role of protagonist in building the brand narrative, prioritizing the authenticity of direct experience.

Keywords: Consumer behavior; Cultural consumption; Transmedia strategies; Digital influencers; Experience marketing.

Compliance And Esg

October 05, 2026

Compliance to mitigate and prevent theft in construction sites: Integrity program applied to civil construction

Thefts at construction sites represent a significant challenge for the civil construction industry, generating financial, operational, and reputational impacts. In this context, integrity and compliance programs have emerged as management tools to strengthen governance and mitigate risks. The study aimed to propose a compliance program applied to civil construction, focused on preventing and mitigating thefts at construction sites. A qualitative approach, with quantitative support, was adopted, developed in two stages. Firstly, a field survey was conducted between February and March 2026, applying an electronic questionnaire to industry professionals. Subsequently, a fictitious case study was developed, based on the author’s professional experiences, integrating the research results with risk management and governance practices. The results highlighted the importance of implementing compliance programs in the sector and indicated that the combination of control mechanisms, structured processes, team training, reporting channels, continuous monitoring, and strengthening of an ethical culture can reduce vulnerabilities. It was concluded that the adoption of integrity practices enhances organizations’ preventive capacity and improves governance and risk management in the sector.

Keywords: Compliance; Civil construction; Theft; Risk management; Corporate governance.

October 05, 2026

Multi-signal panel for optimizing vulnerability prioritization in cybersecurity

The growing proliferation of cyber vulnerabilities has rendered prioritization based solely on static severity metrics inadequate. A multi-signal analytical panel was developed and evaluated to optimize cyber vulnerability prioritization, integrating Common Vulnerability Scoring System (CVSS), Exploit Prediction Scoring System (EPSS), criteria derived from Stakeholder-Specific Vulnerability Categorization (SSVC), and the Known Exploited Vulnerabilities (KEV) catalog as the supervised target variable. 137,854 vulnerability records from NVD (2022–2025) were consolidated, with 607 confirmed KEVs (0.44%), characterizing a classification problem with a highly imbalanced class. A supervised Random Forest model was trained with ten non-circular variables and evaluated using metrics suitable for imbalance. The model achieved an AUC-ROC of 0.9869 and AUC-PR of 0.4959, outperforming isolated EPSS on both metrics (AUC-ROC=0.9383 and AUC-PR=0.3231). Discrepancy analysis with the deterministic SSVC approach revealed that 83.4% of vulnerabilities were over-prioritized. Of the 4,584 CVEs classified as P0, 4,029 represented high-imminent-risk vulnerabilities not confirmed as exploited, flagged by the model. It was concluded that the multi-signal approach goes beyond querying the KEV catalog, identifying vulnerabilities with high future exploitation potential and guiding proactive remediation prioritization.

Keywords: Exploit prediction; Risk management; Vulnerability management; Known Exploited Vulnerabilities (KEV); Random Forest.

Neuroscience And Learning In Education

October 05, 2026

Meditation: a tool that collaborates with the teacher in the classroom

The study aimed to identify meditative practices, their benefits, limitations, and application possibilities in the school context, with a scientific focus and based on existing publications, seeking to expand strategies for promoting mental health and well-being of students and teachers. An exploratory and descriptive research was conducted, with a qualitative approach, based on a literature review of scientific articles, books, and documents published in the last ten years, and on the researcher’s experience reports. The methodology included the scientific definition of meditation and the analysis of studies on meditation and mindfulness in school settings, incorporating contributions from neuroscience, but avoiding biological reductionism of learning. As a main result, a Booklet of Meditative Practices for Teachers was developed, conceived as a complementary pedagogical support material. The findings indicated that meditation can contribute to the reduction of symptoms of stress, anxiety, and depression, in addition to fostering attention, self-reflection, empathy, and coexistence. It was concluded that the booklet offers simple and adaptable guidelines for the classroom, serving as a complementary tool that does not replace pedagogical interventions or public policies. The sustainability of these practices requires articulation with the pedagogical project, institutional support, and teacher training, integrating neuroscience, pedagogy, and practical experience for an education more attentive to the cognitive, emotional, social, and existential dimensions of students.

Keywords: School environment; Mindfulness; Meditation; Neuroscience; Mental health.