This study compares the largest claims from two insurance portfolios using stochastic orderings.
problem Comparing the largest claims from two heterogeneous insurance portfolios.
method Used various stochastic orderings and established sufficient conditions associated with model parameters.
result Established sufficient conditions for comparing the largest claims from two insurance portfolios.
The paper analyzes systemic risk in an insurance model with multiple business lines and heterogeneous claims.
problem Analyzing systemic risk in a multi-dimensional insurance model with heterogeneous claims.
method A multi-dimensional Lévy process-based renewal risk model with pairwise asymptotic independence (PAI).
result Asymptotic formulas for tail probabilities and systemic risk measures are derived.
The paper examines stochastic inequalities involving minimum and maximum claim amounts.
problem Investigating stochastic inequalities for claim amounts with random number of claims.
method Analyzing stochastic order and reversed hazard rate order for minimum and maximum claim amounts.
result Strengthening and generalizing existing results in the literature.
Let Xλ1,…,Xλn be a set of dependent and non-negative random variables share a survival copula and let Yi=IpiXλi, i=1,…,n, where Ip1,…,Ipn be independent Bernoulli random variables independent of Xλi's, with E[Ipi]=pi, i=1,…,n. In actuarial scie…
A federated minimax framework for heterogeneous clients.
problem Training with edge devices having different datasets and capabilities.
method Proposes a federated minimax optimization framework with normalized updates.
result Improves convergence and communication complexity for nonconvex functions.
Method proposed for pricing insurance products covering both foreseeable and unforeseeable risks.
problem Pricing insurance products that include unforeseeable risks.
method Mixed Poisson process with Bayesian setup and linear exponential family distributions.
result Bayesian premiums are more reactive to claim trends than traditional ones.
Let Xλ1,…,Xλn be dependent non-negative random variables and Yi=IpiXλi, i=1,…,n, where Ip1,…,Ipn are independent Bernoulli random variables independent of Xλi's, with E[Ipi]=pi, i=1,…,n. In actuarial sciences, Yi corresponds to the claim amo…
SUOD accelerates OD for large, diverse models.
problem Training and scoring new samples with many unsupervised, heterogeneous OD models.
method Data reduction, model approximation, and taskload optimization.
result SUOD accelerates OD for over 20 benchmark datasets and a real-world case.
Method completes mixed matrix from complex surveys with heterogeneous missingness.
problem Recovering a mixed dataframe matrix from complex survey sampling with different missingness patterns.
method Two-stage procedure: logistic regression for missingness modeling, and weighted log-likelihood maximization with low-rank constraint.
result The proposed method achieves sublinear convergence and shows superior performance compared to existing methods.
We propose a novel approach for loss reserving based on deep neural networks. The approach allows for joint modeling of paid losses and claims outstanding, and incorporation of heterogeneous inputs. We validate the models on loss reserving data across lines of business, and show that they improve on the predictive accu…
This paper considers the problems of modeling and predicting a long-term and ``blurry'' relapse that occurs after a medical act, such as a surgery. The relapse is observed only indirectly, in a ``blurry'' fashion, through longitudinal prescriptions of drugs over a long period of time after the medical act. We introduce…
New method optimizes costly functions with unknown costs and budget constraints.
problem Optimizing functions with unknown and heterogeneous evaluation costs under a budget constraint.
method Budgeted multi-step expected improvement acquisition function.
result Our method outperforms existing approaches in various synthetic and real problems.
Deviance Voronoi residuals improve earthquake insurance risk assessment.
problem Assessing earthquake insurance risk using spatio-temporal point process models.
method Extended Voronoi residuals and created simulation-based approach.
result Proposed formula for country-wide minimum capital test.
The opioid epidemic in the United States claims over 40,000 lives per year, and it is estimated that well over two million Americans have an opioid use disorder. Over-prescription and misuse of prescription opioids play an important role in the epidemic. Individuals who are prescribed opioids, and who are diagnosed wit…
New distances for comparing heterogeneous probability measures efficiently.
problem Comparing probability measures across different spaces.
method Introducing Anchor Energy (AE) and Anchor Wasserstein (AW) distances, and a sweep line algorithm for exact computation.
result Exact computation of AE and AW distances in log-quadratic time, significantly faster than GW.
The primary goal of this study is doing a meta-analysis research on two groups of published studies. First, the ones that focus on the evaluation of the United States Department of Agriculture (USDA) forecasts and second, the ones that evaluate the market reactions to the USDA forecasts. We investigate four questions. …
Reinforcement learning improves insurance claims reserving by learning from all claim trajectories.
problem Traditional reserving models learn only from settled claims, missing valuable data from ongoing claims.
method Formulated as a Markov decision process, uses reinforcement learning to update OCL estimates sequentially.
result Soft Actor-Critic implementation achieves competitive claim-level accuracy and strong aggregate performance.
The study analyzes how bonus-malus systems and delayed claims settlement affect insurance companies' financial stability.
problem Analyzing the impact of bonus-malus systems and delayed claims settlement on insurance companies' financial stability.
method Examined a discrete-time risk model with time-varying premiums, evaluating two types of claims and settlement delays.
result Delayed settlement of by-claims leads to lower ruin probabilities under specific assumptions.
Deep Claim predicts payer responses from claims data using deep learning.
problem Predicting payer responses from claims data to improve healthcare performance.
method Learning complex dependencies in claim inputs to create a compact representation, then using deep learning to predict responses.
result Deep Claim improves claim denial prediction by 22.21%.
We consider an ad hoc network where multiple users access the same set of channels. The channel characteristics are unknown and could be different for each user (heterogeneous). No controller is available to coordinate channel selections by the users, and if multiple users select the same channel, they collide and none…
New method for individual claims reserving using machine learning.
problem Traditional claims reserving methods are limited in individual claim prediction.
method Restructured data utilization for CL prediction, using multi-period factors.
result Neural networks applied for individual claims reserving.
The tail of the distribution of a sum of a random number of independent and identically distributed nonnegative random variables depends on the tails of the number of terms and of the terms themselves. This situation is of interest in the collective risk model, where the total claim size in a portfolio is the sum of a …
Two machine learning models detect anomalies in ER claims, saving up to 40% in improper payments.
problem Improper health insurance payments from fraud and upcoding.
method Two machine learning models: an upcoding model based on severity code distributions and a random forest model for claim sorting.
result Random forest model saved 12% to 40% in improper payments compared to a baseline approach.
Optimizes insurance processing capacity to minimize costs.
problem Processing delays and backlogs in insurance claims.
method Optimal capacity selection to minimize delay-adjusted and fixed costs.
result Minimizes claims costs by balancing processing capacity and delays.
New model bridges pricing and reserving for insurance claims.
problem Incomplete claim data due to reporting and settlement delays.
method Develops an occurrence and development model to estimate both claims and premiums.
result Effective resolution of pricing and reserving inconsistencies.
Federated Learning (FL) refers to learning a high quality global model based on decentralized data storage, without ever copying the raw data. A natural scenario arises with data created on mobile phones by the activity of their users. Given the typical data heterogeneity in such situations, it is natural to ask how ca…
Model detects insurance fraud using social network analysis.
problem Fraudulent insurance claims by exaggeration or intentional damage.
method Network construction linking claims and parties, BiRank algorithm for fraud score computation, feature extraction from network and claims, supervised model building.
result Network features improve fraud detection performance.
We consider trading in a financial market with proportional transaction costs. In the frictionless case, claims are maximal if and only if they are priced by a consistent price process--the equivalent of an equivalent martingale measure. This result fails in the presence of transaction costs. A properly maximal claim i…
Investor maximizes utility from an unknown claim using robust optimization.
problem Maximizing utility from an unknown contingent claim.
method Robust optimization with quantile formulation and variational inequalities.
result Optimal trading strategy and utility indifference price determined.
Model predicts individual insurance claim reserves using activation patterns.
problem Accurately predicting individual claim reserves in insurance contracts.
method Multinomial logistic regression to model claim activation and development.
result The model generates accurate predictions of total and per coverage reserves.
BERT learns claim descriptions to identify patent novelty.
problem Identifying novel patent claims among existing documents.
method Training BERT on concatenated claims and descriptions, scoring BERT's output.
result BERT identifies relevant X documents for patent novelty.
Insurance companies must manage millions of claims per year. While most of these claims are non-fraudulent, fraud detection is core for insurance companies. The ultimate goal is a predictive model to single out the fraudulent claims and pay out the non-fraudulent ones immediately. Modern machine learning methods are we…
Paper introduces EEMs for pricing contingent claim returns.
problem Computing expected future prices of contingent claims.
method Dynamic change of measure approach to construct EEMs.
result EEMs provide physical and pricing expectations of contingent claim prices.
Traditional non-life reserving models largely neglect the vast amount of information collected over the lifetime of a claim. This information includes covariates describing the policy, claim cause as well as the detailed history collected during a claim's development over time. We present the hierarchical reserving mod…
A new method for modeling insurance claim frequencies using random proportions.
problem Inaccurate fitting of classical distributions to insurance claim frequency data.
method Modeling claim frequencies using random proportions of insurance contracts and applying goodness-of-fit tests.
result A new statistical approach for better modeling insurance claim frequencies.
The paper introduces BCART models for aggregate claim amount, improving frequency-severity and joint modeling.
problem Modeling aggregate claim amount with frequency-severity and joint dependencies.
method Developed three types of BCART models: frequency-severity, sequential, and joint models. Used various distributions for claim severity data.
result Weibull distribution outperforms gamma and lognormal for right-skewed, heavy-tailed claim severity data.
In this work, we focus on fine-tuning an OpenAI GPT-2 pre-trained model for generating patent claims. GPT-2 has demonstrated impressive efficacy of pre-trained language models on various tasks, particularly coherent text generation. Patent claim language itself has rarely been explored in the past and poses a unique ch…
FiNCAT tool automatically identifies financial numerals in documents.
problem Differentiating between in-claim and out-of-claim numerals in financial documents.
method Extracts context embeddings of numerals using BERT, then uses Logistic Regression to classify.
result Achieved a Macro F1 score of 0.8223 on validation set.
LLMs help automate extraction of actuarial variables from unstructured claims data.
problem Manual processing of unstructured claims data is time-consuming and inconsistent.
method Two-stage processing architecture using LLMs, modular Python pipeline.
result LLM-based extraction achieved high accuracy and practical actuarial value.
New method simplifies individual claims reserving.
problem Insufficient flexibility and robustness in existing methods.
method Building on classical chain-ladder method, introduces new perspective.
result Advances toward a new standard for micro-level reserving.
Meta learning works well with overparameterized models, a phenomenon called 'benign overfitting'.
problem Understanding why overparameterized models perform well in few-shot learning.
method Analyzed the generalization performance of gradient-based meta learning with an overparameterized meta linear regression model.
result Demonstrated that overparameterized meta learning can still generalize well, a phenomenon called 'benign overfitting'.
Using a suitable change of probability measure, we obtain a novel Poisson series representation for the arbitrage- free price process of vulnerable contingent claims in a regime-switching market driven by an underlying continuous- time Markov process. As a result of this representation, along with a short-time asymptot…
Study tackles imbalanced data in car insurance claims prediction.
problem Predicting rare events (claims) in car insurance with imbalanced data.
method Various machine learning techniques (logistic-regression, decision tree, random forest, xgBoost, feed-forward network) applied to imbalanced dataset.
result Comparison of machine learning algorithms' performance in claim occurrence prediction.
We derive asymptotic expansions for the prices of a variety of European and barrier-style claims in a general local-stochastic volatility setting. Our method combines Taylor series expansions of the diffusion coefficients with an expansion in the correlation parameter between the underlying asset and volatility process…
Time-aware fact-checking improves veracity predictions for time-sensitive claims.
problem Fact-checking decisions should consider temporal information of claims and evidence.
method Investigated four temporal ranking methods to optimize evidence ranking for fact-checking models.
result Time-aware evidence ranking surpasses relevance assumptions and improves veracity predictions for time-sensitive claims.
This paper considers the problem of predicting the number of events that have occurred in the past, but which are not yet observed due to a delay. Such delayed events are relevant in predicting the future cost of warranties, pricing maintenance contracts, determining the number of unreported claims in insurance and in …
We consider a classical risk process with arrival of claims following a non-stationary Hawkes process. We study the asymptotic regime when the premium rate and the baseline intensity of the claims arrival process are large, and claim size is small. The main goal of the article is to establish a diffusion approximation …
We investigate, focusing on the ruin probability, an adaptation of the Cramer-Lundberg model for the surplus process of an insurance company, in which, conditionally on their intensities, the two mixed Poisson processes governing the arrival times of the premiums and of the claims respectively, are independent. Such a …