New method disentangles shared and private latent factors in multimodal data.
problem Challenges in disentangling shared and private latent factors in multimodal data.
method Proposes a modification to existing multimodal Variational Autoencoders (MMVAE) to better handle modality-specific variation.
result Demonstrates improved robustness of modified MMVAE to modality-specific variation.
We propose a novel classification model for weak signal data, building upon a recent model for Bayesian multi-view learning, Group Factor Analysis (GFA). Instead of assuming all data to come from a single GFA model, we allow latent clusters, each having a different GFA model and producing a different class distribution…
Method estimates shared and study-specific factors for multi-study data.
problem Covariance estimation for multi-study data with shared and study-specific components.
method Spectral decomposition for latent factors, surrogate Bayesian regressions for loadings and variances.
result Strong frequentist guarantees and superior performance in simulations and real data.
N-VAE separates class-specific and shared factors in data.
problem Multimodal distribution of generative factors due to class distinction.
method N-VAE model with class-conditioned and shared latent spaces.
result Model effectively disentangles class-dependent and shared factors.
There is a growing interest in joint multi-subject fMRI analysis. The challenge of such analysis comes from inherent anatomical and functional variability across subjects. One approach to resolving this is a shared response factor model. This assumes a shared and time synchronized stimulus across subjects. Such a model…
SharedMF uses secret sharing to protect privacy in distributed recommendation systems.
problem Privacy issues in multi-source data for recommendation systems.
method Federated learning and secret sharing technology.
result SharedMF achieves faster execution speed and better data adaptability compared to homomorphic encryption methods.
This paper presents a novel approach to speaker subspace modelling based on Gaussian-Binary Restricted Boltzmann Machines (GRBM). The proposed model is based on the idea of shared factors as in the Probabilistic Linear Discriminant Analysis (PLDA). GRBM hidden layer is divided into speaker and channel factors, herein t…
MCPCA analyzes shared factors across multiple data contexts.
problem No tools to recover shared factors across multiple contexts.
method Developed a theoretical and algorithmic framework (MCPCA).
result Reveals shared axes of variation across subsets of contexts.
The MAXFLAT low-pass filter improves factor adjustment for better portfolio performance in China's stock market.
problem Improving factor adjustment for better portfolio performance in China's stock market.
method Using MAXFLAT low-pass volatility model to adjust factors and construct portfolios.
result Adjusted factors by MAXFLAT volatility model show better performance in both large and small cap universes.
Margin trading in which investors purchase shares with money borrowed from brokers is blamed to be a major cause of the 2015 Chinese stock market crash. We propose a cascading failure model and examine how an increase in margin trading increases share price vulnerability. The model is based on a bipartite graph of inve…
New model extracts shared brain activity patterns from fMRI data.
problem Challenges in aggregating multi-subject fMRI data due to variability.
method Shared Gaussian Process Factor Analysis (S-GPFA) incorporating temporal information.
result Model reveals ground truth latent structures and replicates experimental performance.
Method learns shared and specific factors in multi-study gene expression data.
problem Understanding shared and specific factors in high-dimensional multi-study data.
method Nonlinear multi-study factor model with sparse variational autoencoder.
result Method recovers meaningful shared and specific factors in platelet gene expression data.
HCL learns shared and modality-specific latent representations for multimodal data.
problem Binary shared-private decomposition inadequately represents shared information across subsets of modalities.
method Hierarchical Contrastive Learning framework combining latent-variable formulation, structural sparsity, and contrastive objective.
result HCL accurately recovers hierarchical structure and improves predictive performance on multimodal data.
New results show contrastive learning can recover shared factors in multimodal data.
problem Understanding when contrastive learning can recover shared latent factors in multimodal data.
method New identifiability results for multimodal contrastive learning, distinguishing between multi-view and multimodal settings.
result Contrastive learning can block-identify shared latent factors in multimodal data, even with dependencies.
Introduces factor risk measures to assess risk relative to multiple factors.
problem Measuring risk relative to multiple factors.
method Introduces a double-argument mapping as a risk measure to assess risk relative to a vector of factors.
result Characterizes various types of factor risk measures including distortion, quantile, linear, and coherent measures.
dCMF learns shared latent representations from multiple matrices, improving predictive modeling.
problem Learning from multiple heterogeneous data sources, especially non-linear interactions.
method Develops a deep-learning based method (dCMF) for unsupervised learning of multiple shared representations.
result dCMF significantly outperforms previous CMF algorithms in integrating heterogeneous data.
A new method for analyzing multi-source, multi-way data reduces dimensionality and reveals shared and individual structures.
problem Analyzing multi-source, multi-way data from different high-throughput technologies.
method Multiple Linked Tensor Factorization (MULTIFAC) extending CP decomposition with L2 penalties and EM algorithm for incomplete data.
result MULTIFAC approximates underlying signal, identifies shared and unshared structures, and imputes missing data.
BIDIFAC integrates multi-platform, multi-cohort data for shared and unique patterns.
problem Integration of multi-platform, multi-cohort data for shared and unique patterns.
method BIDIFAC integrates bidimensionally linked matrices into four components: globally shared, row-shared, column-shared, and single-matrix structural components.
result BIDIFAC reveals shared and unique patterns of variability in multi-platform, multi-cohort data.
K-FAC speeds up training of modern neural networks with linear weight-sharing.
problem Efficiently training modern neural networks with linear weight-sharing layers.
method Kronecker-Factored Approximate Curvature (K-FAC) applied to linear weight-sharing layers.
result K-FAC-reduce is generally faster than K-FAC-expand for deep linear networks.
Proposes MD-LiNA for multi-domain latent factor causal discovery.
problem Discovering causal structures among latent factors from multi-domain data.
method Multi-Domain Linear Non-Gaussian Acyclic Models (MD-LiNA) with an integrated two-phase algorithm.
result Locally consistent estimators of causal structure among shared latent factors.
Sector specific multifactor CES elasticity of substitution and the corresponding productivity growths are jointly measured by regressing the growths of factor-wise cost shares against the growths of factor prices. We use linked input-output tables for Japan and the Republic of Korea as the data source for factor price …
The purpose of this study is to measure the Total Factor Productivity (TFP) growth and determine the share of each of the economic growth sources in the mining sector of Iran. The time period of this study is 1355-1385 of the Solar Hijri calendar (roughly overlaying with the time period of 1976-2006 of the Gregorian ca…
Enhances LLM quantization with MDBF, improving perplexity and accuracy.
problem Limited performance of Double Binary Factorization in extreme quantization.
method Introduces Multi-envelope DBF, retaining sign matrices and replacing single envelope with rank-l envelope. result Improves perplexity and zero-shot accuracy over previous binary formats.
CARE improves LLM aggregation by accounting for shared confounders.
problem LLM judges' correlated errors due to shared latent confounders.
method CARE explicitly models judges' scores as true quality and confounders.
result CARE reduces aggregation error by up to 26.8% across various benchmarks.
BIDIFAC+ factorizes linked matrices for cancer studies.
problem Integrating multiple omics platforms across various cancer types.
method Flexible approach to simultaneous factorization and decomposition of linked matrices using BIDIFAC+.
result Identifies shared and specific modes of variability across multiple omics platforms and cancer types.
Paper tackles skill transfer in RL for morphologically different agents.
problem Transfer skills between morphologically different reinforcement learning agents.
method Proposes a paired variational encoder-decoder model (PVED) for subspace learning.
result Demonstrates improved skill transfer efficiency compared to state-of-the-art methods.
Empirical study of CAPM and Fama-French model in Chinese A-share market.
problem Testing and validating CAPM and Fama-French model in Chinese A-share market.
method Used Fama-MacBeth regression and Fama-French three-factor model to analyze Chinese A-share trading data from 2000 to 2019, adjusting for IPO shell value contamination.
result Fama-French model captures most of A-share market returns, with adjusted R-squared > 0.88.
We introduce Bayesian multi-tensor factorization, a model that is the first Bayesian formulation for joint factorization of multiple matrices and tensors. The research problem generalizes the joint matrix-tensor factorization problem to arbitrary sets of tensors of any depth, including matrices, can be interpreted as u…
Unified MTL framework for heterogeneous data integrates shared and task-specific encoders.
problem Efficiently sharing information across multiple tasks with heterogeneous data.
method Dual-encoder framework with task-shared and task-specific encoders.
result Unified algorithm alternates learning task-specific and shared encoders and coefficients.
This paper analyzes stock market data to predict share prices using regression models.
problem Predicting stock prices in the share market of Bangladesh.
method Thorough linear regression analysis on Dhaka Stock Exchange data, compared with random forest.
result Random forest model performs better than linear regression for predicting stock prices.
Paper proposes VAE-BPTF for better tensor factorization of sparse, imbalanced count data.
problem Inference of Bayesian Poisson-Gamma models for sparse and imbalanced count data is challenging.
method Variational auto-encoder framework with multi-layer perceptron networks for complex update information sharing and reweighting.
result VAE-BPTF outperforms current models in reconstruction errors and latent factor coherence across real-world datasets.
DISCoVeR learns disentangled representations by separating shared and condition-specific factors.
problem Learning disentangled representations for multi-condition data.
method Dual-latent architecture, parallel reconstructions, max-min objective.
result DISCoVeR achieves improved disentanglement on various datasets.
FACTM combines FA with correlated topic modeling for structured data integration.
problem Integrating structured data modalities like text and single cell sequencing.
method Bayesian FACTM model combining FA and correlated topic modeling with variational inference.
result FACTM outperforms other methods in identifying clusters in structured data and integrating them with simple modalities.
A method learns matrix factorization from diverse matrices and applies the knowledge to unseen matrices.
problem Matrix factorization without shared rows or columns.
method Neural network meta-learned to minimize expected imputation error using MAP estimation.
result The method can impute missing values from unseen matrices efficiently.
The paper defines fair profit sharing ratios in Islamic PL contracts.
problem Determining fair profit sharing ratios in Islamic PL contracts.
method Introduces c-fair profit sharing ratios and uses econometrics models to compute or approximate them. result Elucidates the relation between profit sharing ratios and economic factors.
We propose a privacy-enhanced matrix factorization recommender that exploits the fact that users can often be grouped together by interest. This allows a form of "hiding in the crowd" privacy. We introduce a novel matrix factorization approach suited to making recommendations in a shared group (or nym) setting and the …
CMF is a technique for simultaneously learning low-rank representations based on a collection of matrices with shared entities. A typical example is the joint modeling of user-item, item-property, and user-feature matrices in a recommender system. The key idea in CMF is that the embeddings are shared across the matrice…
We develop necessary and sufficient conditions and a novel provably consistent and efficient algorithm for discovering topics (latent factors) from observations (documents) that are realized from a probabilistic mixture of shared latent factors that have certain properties. Our focus is on the class of topic models in …
This paper investigates weight-sharing in NAS, revealing its impact and providing solutions.
problem Reducing the time and computational cost of training neural networks.
method Comprehensive experiments on weight-sharing in NAS, analyzing variance and interference.
result Properly reducing weight sharing can reduce variance and improve model performance.
Bayesian hypergraph inference models disease pathways from EHR data.
problem Modeling rare diseases influenced by shared risk factors.
method Bayesian hypergraph inference framework reframing multi-disease modeling.
result Interpretable disease pathways and well-calibrated uncertainty quantification.
New tensor approach models global fixed income risks across maturities and economies.
problem Lack of models capturing multi-dimensional data in global fixed income markets.
method Introduces tensor-valued approach to model shared risks among multiple interest rate curves.
result Estimates risk factors decomposable into maturity and country domains, enabling tailored portfolio management.
DPFact preserves privacy while collaboratively factorizing EHR tensors.
problem Privacy-preserving tensor factorization for EHRs.
method Differential privacy and collaborative learning.
result DPFact achieves higher accuracy and efficiency under privacy constraints.
Tensor factorization models offer an effective approach to convert massive electronic health records into meaningful clinical concepts (phenotypes) for data analysis. These models need a large amount of diverse samples to avoid population bias. An open challenge is how to derive phenotypes jointly across multiple hospi…
A new model optimizes portfolios by learning stock return distributions conditioned on factors.
problem Optimizing portfolios with high-dimensional asset-specific factors.
method Conditional Diffusion Transformer architecture linking each asset's return to its factor vector.
result The model outperforms benchmarks in mean-variance and mean-CVaR optimization.
Study examines pricing strategies in competitive supply chains with discrete prices.
problem Inaccurate assumptions in traditional SC models for pricing decisions.
method Examines a SC model with one supplier and two manufacturers, considering customer demand segmentation and discrete price setting.
result Nash equilibria among manufacturers are not unique, and low denomination factors can lead to instability.
We present a novel factor analysis method that can be applied to the discovery of common factors shared among trajectories in multivariate time series data. These factors satisfy a precedence-ordering property: certain factors are recruited only after some other factors are activated. Precedence-ordering arise in appli…
This paper compares two stock factor models in China's A-share market.
problem Contradicting results in existing research on stock factor models.
method Empirical analysis using China's A-share data from 2005-2020, orthogonalizing redundant factors, and 25-group portfolio returns calculation.
result The five-factor model outperforms the three-factor model in explaining excess return rates.
Proposes MVMC for multi-view clustering, enhancing diversity and quality.
problem Leveraging multi-view data for diverse clustering.
method Adapts multi-view self-representation learning, HSIC for redundancy reduction, matrix factorization.
result Generates multiple high-quality and diverse clusterings from multi-view data.