Cross-GCN models cross features in GCN for better performance.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Graph cross network improves graph classification accuracy.
DNN2LR bridges DNN power and LR interpretability.
Various factorization-based methods have been proposed to leverage second-order, or higher-order cross features for boosting the performance of predictive models. They generally enumerate all the cross features under a predefined maximum order, and then identify useful feature interactions through model training, which…
Feature engineering has been the key to the success of many prediction models. However, the process is non-trivial and often requires manual feature engineering or exhaustive searching. DNNs are able to automatically learn feature interactions; however, they generate all the interactions implicitly, and are not necessa…
DCN-V2 improves deep & cross network for web-scale learning to rank systems.
We describe how cross-kernel matrices, that is, kernel matrices between the data and a custom chosen set of `feature spanning points' can be used for learning. The main potential of cross-kernels lies in the fact that (a) only one side of the matrix scales with the number of data points, and (b) cross-kernels, as oppos…
The paper tackles feature cross search for linear models, providing approximation algorithms and structural results.
Set-Sequence model learns cross-sectional dynamics directly from time series data.
This paper improves image super-resolution by integrating cross-scale non-local attention.
Cross-balancing improves causal inference by balancing features with outcome data.
Paper proposes BSP to find stable bimodules of cross-correlated features.
Unified model learns from both time-series and cross-sectional momentum features.
Feature crossing captures interactions among categorical features and is useful to enhance learning from tabular data in real-world businesses. In this paper, we present AutoCross, an automatic feature crossing tool provided by 4Paradigm to its customers, ranging from banks, hospitals, to Internet corporations. By perf…
Cross-modal hashing has been receiving increasing interests for its low storage cost and fast query speed in multi-modal data retrievals. However, most existing hashing methods are based on hand-crafted or raw level features of objects, which may not be optimally compatible with the coding process. Besides, these hashi…
With the proliferation of social media platforms and e-commerce sites, several cross-domain collaborative filtering strategies have been recently introduced to transfer the knowledge of user preferences across domains. The main challenge of cross-domain recommendation is to weigh and learn users' different behaviors in…
This paper extends neural collapse to imbalanced data under cross-entropy loss.
Structural correspondence learning (SCL) is an effective method for cross-lingual sentiment classification. This approach uses unlabeled documents along with a word translation oracle to automatically induce task specific, cross-lingual correspondences. It transfers knowledge through identifying important features, i.e…
Dual-structured method improves cross-domain imitation learning.
MetFA aligns source and target domains for cross-device image classification.
Multivariate functional data from a complex system are naturally high-dimensional and have complex cross-correlation structure. The complexity of data structure can be observed as that (1) some functions are strongly correlated with similar features, while some others may have almost no cross-correlations with quite di…
Improved 3D ECG feature attributions for clinical interpretation.
It is ubiquitous in natural and social sciences that two variables, recorded temporally or spatially in a complex system, are cross-correlated and possess multifractal features. We propose a new method called multifractal detrended cross-correlation analysis (MF-DXA) to investigate the multifractal behaviors in the pow…
DHEN improves CVR prediction for ads with multitask learning and auxiliary loss.
New method corrects bias in feature importance measures of GBM.
Study improves stock movement prediction using multimodal data.
Pipeline integrates cross-sectional and longitudinal multi-omics data for IBD research.
Cross-validation is the de facto standard for predictive model evaluation and selection. In proper use, it provides an unbiased estimate of a model's predictive performance. However, data sets often undergo various forms of data-dependent preprocessing, such as mean-centering, rescaling, dimensionality reduction, and o…
In general, a self-attention mechanism has been applied for speaker embedding encoding. Previous studies focused on training the self-attention in a high-level layer, such as the last pooling layer. However, the effect of low-level features was reduced in the speaker embedding encoding. Therefore, we propose masked cro…
Deep learning predicts cross-sectional stock prices for practical investment.
UBMF tackles fault diagnosis in imbalanced industrial data with enhanced accuracy and adaptability.
The study identifies features making cross-impact relevant in explaining price variance of US assets.
Used to estimate the risk of an estimator or to perform model selection, cross-validation is a widespread strategy because of its simplicity and its apparent universality. Many results exist on the model selection performances of cross-validation procedures. This survey intends to relate these results to the most recen…
The paper evaluates company investment value using machine learning models.
Improved survival analysis using square root Cox's models and neural networks.
In deep neural network, the cross-entropy loss function is commonly used for classification. Minimizing cross-entropy is equivalent to maximizing likelihood under assumptions of uniform feature and class distributions. It belongs to generative training criteria which does not directly discriminate correct class from co…
In recent years, there have been numerous developments towards solving multimodal tasks, aiming to learn a stronger representation than through a single modality. Certain aspects of the data can be particularly useful in this case - for example, correlations in the space or time domain across modalities - but should be…
Develops a prediction method based on sampling design.
Automatic classification of epileptic seizure types in electroencephalograms (EEGs) data can enable more precise diagnosis and efficient management of the disease. This task is challenging due to factors such as low signal-to-noise ratios, signal artefacts, high variance in seizure semiology among epileptic patients, a…
In this paper, we focus on the separability of classes with the cross-entropy loss function for classification problems by theoretically analyzing the intra-class distance and inter-class distance (i.e. the distance between any two points belonging to the same class and different classes, respectively) in the feature s…
New method improves unsupervised feature learning for natural data.
Regularization improves stability and consistency of sparse autoencoders.
SBCA optimizes portfolios by fusing price data and text sentiment.
Least-squares models such as linear regression and Linear Discriminant Analysis (LDA) are amongst the most popular statistical learning techniques. However, since their computation time increases cubically with the number of features, they are inefficient in high-dimensional neuroimaging datasets. Fortunately, for k-fo…
Transportation modes prediction is a fundamental task for decision making in smart cities and traffic management systems. Traffic policies designed based on trajectory mining can save money and time for authorities and the public. It may reduce the fuel consumption and commute time and moreover, may provide more pleasa…
This paper extends neural collapse to class-imbalanced datasets using an unconstrained ReLU feature model.
Identifying epileptic seizures through analysis of the electroencephalography (EEG) signal becomes a standard method for the diagnosis of epilepsy. Manual seizure identification on EEG by trained neurologists is time-consuming, labor-intensive and error-prone, and a reliable automatic seizure/non-seizure classification…
Real analytic functions can be extended on manifolds with normal crossings.