New method measures common information in high-dimensional data.
problem Measuring common information among many variables.
method Information sieve decomposition to formulate a scalable common information problem.
result Scalable approach demonstrates common information's usefulness in high-dimensional learning.
A new model generates samples with a succinct common representation using Wyner's common information.
problem Generating samples with a succinct common representation.
method Proposes a variational Wyner model trained to minimize symmetric Kullback-Leibler divergence with regularization terms.
result Demonstrates utility through joint and conditional generation experiments.
We simplify information measure computation using learned features.
problem Computing information measures from raw data is computationally expensive.
method Developed a separable design for computing information measures from learned feature representations.
result A variety of information measures can be computed efficiently through learned feature representations.
Optimizes portfolios using neural network approximations of asset sensitivities to common drivers.
problem Optimizing portfolios with complex asset dynamics and common drivers.
method Model asset dynamics with PDEs, approximate sensitivities with neural networks, and use hierarchical clustering on sensitivity matrix for optimization.
result Achieves over-performance in portfolio optimization across various markets and datasets.
New method detects latent common causes from observational data.
problem Detecting latent common causes in observational data.
method Modified causal discovery algorithms to detect latent common causes.
result Successfully detects latent common causes in various noise regimes and real data.
We extend common entropy concept and propose algorithms to distinguish causation from correlation.
problem Discovering the simplest latent variable for conditional independence of observed variables.
method Renyi common entropy, iterative algorithm, constraint-based methods modification.
result Improved constraint-based methods for causal inference in small samples.
MCCA extracts shared structure from multiple tensor datasets.
problem Extracting shared structure from multiple tensor datasets.
method Multilinear common component analysis (MCCA) using Kronecker products of mode-wise covariance matrices.
result MCCA constructs a common basis that retains information from multiple tensor datasets.
Study finds stocks with common firm fears earn lower returns.
problem Identifying and quantifying firm-level investor fears.
method Analysis of equity options to identify common firm-level fears and their impact on stock returns.
result Stocks with exposure to common bad fears earn lower returns and require higher compensation.
Hopformer combines common trends with series-specific details for better time series forecasting.
problem Forecasting multiple time-series with high-dimensional covariates while retaining series-specific information.
method Hopformer uses a two-stage framework: SPA for common trends and LoRA-fine-tuned Transformer for residual dependencies.
result Improves MASE by an average of 6.56% across synthetic and real-world benchmarks.
Proposes MV-Co-VH for multi-view clustering using visible and hidden views.
problem Lack of efficient algorithms for fully utilizing multi-view data.
method Projects multiple views to a common hidden space using NMF, then applies collaborative learning.
result Competitive clustering performance on UCI and real-world datasets.
NestedVAE isolates common factors from paired images without additional supervision.
problem Reduction of data-driven biases in machine learning models.
method Combines deep latent variable models with information bottleneck theory.
result NestedVAE significantly outperforms alternative methods in various tasks.
Solves a game between brokers and informed traders using stochastic differential equations.
problem Optimizing wealth in a game between brokers and informed traders with private signals.
method Closed-form solutions to a mean-field game using forward-backward SDEs.
result Optimal trading strategies for both brokers and informed traders are found.
MAGMA uses a common mean process to improve multi-step-ahead time series forecasting.
problem Improving multiple-step-ahead predictions for time series data.
method Proposes a novel multi-task Gaussian process framework with a common mean process for sharing information across tasks.
result Significantly improves predictive performances, even far from observations, and reduces computational complexity.
Study shows publicly available news impacts financial markets.
problem Impact of publicly available news on financial markets.
method Extracted news from Common Crawl, identified relevant companies, used sentiment analysis and information theory.
result Publicly available news has significant impact on financial markets.
Proposes adversarial normalization for multi-domain image segmentation.
problem Current image normalization is per-dataset, limiting multi-domain segmentation.
method Adversarial training to learn common normalizing functions across multiple datasets.
result Optimal normalizer improves segmentation accuracy and realism.
Spatial information is not always necessary for spatio-temporal models.
problem The necessity of including spatial information in spatio-temporal models.
method Comparison of spatial agnostic neural networks with state-of-the-art models on ten datasets.
result Spatial information is not always needed in most spatio-temporal models.
Improved ranking method for scarce data with feature info.
problem Ranking items with limited comparisons and feature data.
method Modified RankCentrality using diffusion methods for feature info.
result Meaningful rankings even with scarce comparisons.
WideDTA predicts drug-target binding affinity using text-based information.
problem Predicting drug-target binding affinity is a major challenge in drug discovery.
method WideDTA uses chemical and biological textual sequence information, including protein sequence, ligand SMILES, protein domains and motifs, and maximum common substructure words.
result WideDTA outperformed DeepDTA on the KIBA dataset, indicating the word-based sequence representation is a promising alternative.
This paper relaxes the common prior assumption in the public and private information game of Morris and Shin (2000, 2004). For the generalized game, where the agent's prior expectations are heterogenous, it derives a sharp condition for the emergence of unique/multiple equilibria. This condition indicates that unique e…
We consider the statistical problem of learning common source of variability in data which are synchronously captured by multiple sensors, and demonstrate that Siamese neural networks can be naturally applied to this problem. This approach is useful in particular in exploratory, data-driven applications, where neither …
Relationships between entities in datasets are often of multiple nature, like geographical distance, social relationships, or common interests among people in a social network, for example. This information can naturally be modeled by a set of weighted and undirected graphs that form a global multilayer graph, where th…
Study finds adding more information to robust option pricing does not improve bounds.
problem Exploring robust pricing of financial claims using minimal assumptions.
method Empirical study of variance options, incorporating intermediate market data.
result Incorporating more information does not improve robust pricing bounds.
Proposes CMP method for reducing tensor object dimensions in binary classification.
problem Reduction of tensor object dimensions while maintaining class separability.
method Proposes Common Mode Patterns (CMP) method considering class labels.
result CMP method increases inter-class separability compared to MPCA.
HPCA improves PCA for portfolio management by interpreting sector-specific factors.
problem Difficult interpretation of PCA's higher eigenportfolios in practical portfolio management.
method Partitioning the market into sectors and applying Hierarchical PCA.
result HPCA leads to no loss of information and interpretable factors.
A new graph kernel uses LCS and Wasserstein distance for better graph comparisons.
problem Graph learning methods can be limited by information from distant vertices and path length constraints.
method Proposes a Graph Kernel based on LCS similarity and Wasserstein distance in a novel metric space.
result The new kernel emphasizes comparisons between similar paths and reduces information loss.
For common people, in contrast to brokers, bankers, and those who play on rising and falling prices of stocks, the stock market law is based on the simple fact that the depositors aim for financial profit at any given concrete stage. The common depositor cannot cause any significant variations in prices. This concept s…
Paper presents a technique using Spearman's Rank Correlation Coefficient for KE in TDs.
problem Extracting common characteristics and grouping similar TDs.
method Spearman's Rank Correlation Coefficient (SRCC) for KE.
result SRCC proves a comprehensive measure for high-quality KE.
Privacy-preserving distributed deep learning method for multiple classification.
problem Privacy issues in training deep learning models for various fields.
method Split learning into common extractor, cloud model, and local classifier.
result Average performance improvement of 2.63% over existing local training models.
A new approach simplifies Sliced-Wasserstein distances to improve learning performance.
problem The concentration of measure phenomenon makes random projections uninformative in high dimensions.
method Propose rescaling the 1D Wasserstein distance to make all slices equally informative.
result The classical Sliced-Wasserstein, properly configured, can match or surpass complex variants.
Proposes ESCA model to analyze mixed data types in multiple sets of measurements.
problem Separating common and distinct information in mixed data types from multiple sources.
method Exponential Family Simultaneous Component Analysis (ESCA) model with structured sparse loading matrix.
result The proposed method effectively disentangles global, local common and distinct information.
Proposes an MTL method with clustering to improve regression accuracy.
problem Improving regression accuracy by sharing information among related tasks.
method Centroid parameter for clustering tasks, separating regression and clustering parameters.
result Improves estimation and prediction accuracy for regression coefficient vectors.
This paper presents an original approach for jointly fitting survival times and classifying samples into subgroups. The Coxlogit model is a generalized linear model with a common set of selected features for both tasks. Survival times and class labels are here assumed to be conditioned by a common risk score which depe…
MALI aligns distinct domains using labeled data.
problem Aligning multi-domain data for machine learning.
method MALI learns manifold structure via diffusion and uses labeled data to guide alignment.
result MALI outperforms state-of-the-art methods across multiple datasets.
Methods for analysis of principal components in discrete data have existed for some time under various names such as grade of membership modelling, probabilistic latent semantic analysis, and genotype inference with admixture. In this paper we explore a number of extensions to the common theory, and present some applic…
Paper designs a penalty for model order selection using information criteria.
problem Selecting the correct model order from a set of candidate models.
method Designs a penalty for the generalized information criterion (GIC) to minimize underestimation.
result Optimal penalty minimizes underestimation while keeping overestimation below a specified level.
Eluder dimension and information gain are equivalent for reproducing kernel Hilbert spaces.
problem Complexity measures in bandit and reinforcement learning.
method Equivalence of eluder dimension and information gain for reproducing kernel Hilbert spaces.
result Eluder dimension and information gain are equivalent for reproducing kernel Hilbert spaces.
Boosted tree method improves MTL in heterogeneous domains.
problem Improving MTL in diverse, domain-specific tasks.
method Two-stage approach: common model for shared features, specific models for task-specific instances.
result Enhanced multi-task learning performance with interpretability.
Ensembles of classification and regression trees remain popular machine learning methods because they define flexible non-parametric models that predict well and are computationally efficient both during training and testing. During induction of decision trees one aims to find predicates that are maximally informative …
Deep learning improves skin disease diagnosis for primary care.
problem Limited dermatological expertise in primary care settings.
method Supervised deep learning for nine dermatological conditions.
result 80% accuracy compared to 57% by human doctors.
The study challenges the notion that partial data annotation is inferior, suggesting it can sometimes outperform complete annotation.
problem The inefficiency and high cost of completely annotating structured data.
method Information theoretic formulation applied to three diverse structured learning tasks.
result Learning from partial structures can sometimes outperform learning from complete ones.
Better investment strategies identified through a network metric of asset commonality.
problem Identifying investment strategies based on fund portfolio asset popularity.
method Bipartite network analysis of mutual funds and their holdings, calculating the Average Commonality Coefficient (ACC).
result Funds investing in less popular assets outperform those in more popular ones, even after adjusting for standard factors.
New research shows existing information-theoretic methods can't establish minimax rates for gradient descent in stochastic convex optimization.
problem Establishing minimax rates for gradient descent in stochastic convex optimization using information-theoretic methods.
method Examined several information-theoretic frameworks including input-output mutual information bounds, conditional mutual information bounds, PAC-Bayes bounds, and their variants.
result Proved that none of the examined information-theoretic frameworks can establish minimax rates for gradient descent in stochastic convex optimization.
Brain signals predict user interest in digital content.
problem Finding relevant information from large document collections.
method A brain-information interface using EEG to infer user interest from reading Wikipedia.
result Users' interests can be modeled from brain signals, enabling information recommendation.
A novel graph-regularized CCA approach for datasets with a common source graph.
problem Discovering hidden sources in datasets with common geometry.
method Graph regularizer to encode common sources' geometry in CCA.
result Improved classification performance over competing methods.
Improved method for encoding contingency tables reduces mutual information bias.
problem Mutual information bias in measuring label similarity.
method Improved method for encoding contingency tables to reduce information cost.
result Better bound on reduced mutual information in typical use cases.
New approach combines multi-view learning for improved convergence in knowledge transfer.
problem Improving convergence in knowledge transfer settings like learning with privileged information and distillation.
method Adopting a multi-view approach under reasonable assumptions about hypothesis spaces, encouraging agreement between teacher and student.
result Improved convergence rate achieved with regularized empirical risk minimization.
The study explores special submanifolds in the probability simplex with unique geometric properties.
problem Exploring special submanifolds in the probability simplex.
method Algebraic characterization and classification of doubly autoparallel submanifolds.
result Characterization and classification of doubly autoparallel submanifolds on the probability simplex.
Extends information bottleneck to multi-view unsupervised learning.
problem Identifying superfluous information in unlabeled multi-view data.
method Multi-view information bottleneck model, leveraging data augmentation.
result State-of-the-art results on Sketchy and MIR-Flickr datasets.