Gradient boosting enhances existing Mendelian models for genetic disease risk prediction.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Genome-wide association studies (GWAS) have achieved great success in the genetic study of Alzheimer's disease (AD). Collaborative imaging genetics studies across different research institutions show the effectiveness of detecting genetic risk factors. However, the high dimensionality of GWAS data poses significant cha…
New methods improve genetic studies of complex diseases.
Develops a SAS approach for high-dimensional risk prediction using unlabeled data.
We study the performance of various agent strategies in an artificial investment scenario. Agents are equipped with a budget, , and at each time step invest a particular fraction, , of their budget. The return on investment (RoI), , is characterized by a periodic function with different types and leve…
With the emergence of the Hospital Readmission Reduction Program of the Center for Medicare and Medicaid Services on October 1, 2012, forecasting unplanned patient readmission risk became crucial to the healthcare domain. There are tangible works in the literature emphasizing on developing readmission risk prediction m…
Optimal capital allocation between different assets is an important financial problem, which is generally framed as the portfolio optimization problem. General models include the single-period and multi-period cases. The traditional Mean-Variance model introduced by Harry Markowitz has been the basis of many models use…
Deep models improve GWAS by identifying genetic interactions.
In this work, a novel approach is proposed for joint analysis of high dimensional time-resolved cardiac motion features obtained from segmented cardiac MRI and low dimensional clinical risk factors to improve survival prediction in heart failure. Different methods are evaluated to find the optimal way to insert convent…
Semi-supervised GAN creates synthetic genetic data for disease prediction.
Study uses machine learning to predict future health from various health data types.
ENN method uses expectile regression for genetic data analysis of complex diseases.
Paper uses AI to predict stock market volatility with neural networks and genetic algorithms.
One component of precision medicine is to construct prediction models with their predictive ability as high as possible, e.g. to enable individual risk prediction. In genetic epidemiology, complex diseases have a polygenic basis and a common assumption is that biological and genetic features affect the outcome under co…
U-aggregation combines multiple models without labels for better risk prediction.
Lapse-supported life insurance exacerbates adverse selection risks.
In plant and animal breeding studies a distinction is made between the genetic value (additive + epistatic genetic effects) and the breeding value (additive genetic effects) of an individual since it is expected that some of the epistatic genetic effects will be lost due to recombination. In this paper, we argue that t…
Advances of modern sensing and sequencing technologies generate a deluge of high dimensional space-temporal physiological and next-generation sequencing (NGS) data. Physiological traits are observed either as continuous random functions, or on a dense grid and referred to as function-valued traits. Both physiological a…
We introduce the binacox, a prognostic method to deal with the problem of detecting multiple cut-points per features in a multivariate setting where a large number of continuous features are available. The method is based on the Cox model and combines one-hot encoding with the binarsity penalty, which uses total-variat…
A new model selects low-carbon mutual funds considering ESG criteria, risk, and investor preferences.
Neural networks improve cancer risk prediction from family history data.
Diagnosing an inherited disease often requires identifying the pattern of inheritance in a patient's family. We represent family trees with genetic patterns of inheritance using hypergraphs and latent state space models to provide explainable inheritance pattern predictions. Our approach allows for exact causal inferen…
Bayesian transfer learning improves predictive performance with limited source data.
Genome-wide association studies (GWAS) offer new opportunities to identify genetic risk factors for Alzheimer's disease (AD). Recently, collaborative efforts across different institutions emerged that enhance the power of many existing techniques on individual institution data. However, a major barrier to collaborative…
This study introduces a new GAS blending ensemble model for Bitcoin price prediction.
Metaheuristics optimize portfolios with pre-assignment and margin trading for better risk-adjusted returns.
VEGN uses graph neural networks to predict disease-causing mutations from genetic variants.
New algorithm predicts lung cancer progression and mortality.
As the amount and complexity of genetic information increases it is necessary that we explore some efficient ways of handling these data. This study takes the "divide and conquer" approach for analyzing high dimensional genomic data. Our aims include reducing the dimensionality of the problem that has to be dealt one a…
In statistical genetics an important task involves building predictive models for the genotype-phenotype relationships and thus attribute a proportion of the total phenotypic variance to the variation in genotypes. Numerous models have been proposed to incorporate additive genetic effects into models for prediction or …
Given genetic variations and various phenotypical traits, such as Magnetic Resonance Imaging (MRI) features, we consider two important and related tasks in biomedical research: i)to select genetic and phenotypical markers for disease diagnosis and ii) to identify associations between genetic and phenotypical data. Thes…
We analyze large, multi-dimensional, sparse counting data sets, finding unsupervised groups to provide unique insights into genetic data. We create gene and biological pathway groups based on patients' variants to find common risk factors for four common types of cancer (breast, lung, prostate, and colorectal) and auti…
Optimizes stock portfolios with profit, risk, and sustainability.
New benchmark predicts cardiometabolic risk from accelerometer data, with varying accuracy.
Paper uses AI to predict tail risks in US financial markets.
Improved genetic algorithm optimizes SVR for robust long-term stock index forecasting.
New method uses DNN for genetic variant identification, controlling randomness and improving interpretability.
EB-VAE combines tumor growth and dropout data for personalized treatment response modeling.
Machine learning predicts obesity causes using genetic and imaging data.
We utilize a recently developed genetic algorithm, in conjunction with discrete wavelets, for carrying out successful forecasts of the trend in financial time series, that includes the NASDAQ composite index. Discrete wavelets isolate the local, small scale variations in these non-stationary time series, after which th…
Paper addresses class imbalance in disk SMART dataset using GANs and genetic algorithms.
Technical analysis is used to discover investment opportunities. To test this hypothesis we propose an hybrid system using machine learning techniques together with genetic algorithms. Using technical analysis there are more ways to represent a currency exchange time series than the ones it is possible to test computat…
New causal models perform poorly when evaluated on biased training sets.
Enhances genetic programming for stock alpha discovery with warm start and structural constraints.
This paper demonstrates how to apply machine learning algorithms to distinguish good stocks from the bad stocks. To this end, we construct 244 technical and fundamental features to characterize each stock, and label stocks according to their ranking with respect to the return-to-volatility ratio. Algorithms ranging fro…
The central aim in this paper is to address variable selection questions in nonlinear and nonparametric regression. Motivated by statistical genetics, where nonlinear interactions are of particular interest, we introduce a novel and interpretable way to summarize the relative importance of predictor variables. Methodol…
For precision medicine and personalized treatment, we need to identify predictive markers of disease. We focus on Alzheimer's disease (AD), where magnetic resonance imaging scans provide information about the disease status. By combining imaging with genome sequencing, we aim at identifying rare genetic markers associa…
BayesMR estimates causal effects and directionality from genetic data.