This paper addresses the problem of scalable optimization for L1-regularized conditional Gaussian graphical models. Conditional Gaussian graphical models generalize the well-known Gaussian graphical models to conditional distributions to model the output network influenced by conditioning input variables. While highly …
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
SGM combines deep learning and planning for robust long-horizon tasks.
New framework models complex spatial data with basis functions and graphical vectors.
Matrix Factorization (MF) on large scale matrices is computationally as well as memory intensive task. Alternative convergence techniques are needed when the size of the input matrix is higher than the available memory on a Central Processing Unit (CPU) and Graphical Processing Unit (GPU). While alternating least squar…
Memory-efficient learning for large-scale imaging systems.
An associative memory is a framework of content-addressable memory that stores a collection of message vectors (or a dataset) over a neural network while enabling a neurally feasible mechanism to recover any message in the dataset from its noisy version. Designing an associative memory requires addressing two main task…
Efficiently discovers Bayesian network structure with reduced memory usage.
GPU (graphics processing unit) has been used for many data-intensive applications. Among them, deep learning systems are one of the most important consumer systems for GPU nowadays. As deep learning applications impose deeper and larger models in order to achieve higher accuracy, memory management becomes an important …
Estimating multiple sparse Gaussian Graphical Models (sGGMs) jointly for many related tasks (large ) under a high-dimensional (large ) situation is an important task. Most previous studies for the joint estimation of multiple sGGMs rely on penalized log-likelihood estimators that involve expensive and difficult n…
Service-induced congestion in memory-constrained LLM serving
Novel SVAE learns interpretable discrete data representations from deep learning.
Develops a fast method to learn graph structures from large datasets.
New GPU algorithm speeds up Gaussian Process analysis.
With the abundance of data in recent years, interesting challenges are posed in the area of recommender systems. Producing high quality recommendations with scalability and performance is the need of the hour. Singular Value Decomposition(SVD) based recommendation algorithms have been leveraged to produce better result…
Surface parameterizations and registrations are important in computer graphics and imaging, where 1-1 correspondences between meshes are computed. In practice, surface maps are usually represented and stored as 3D coordinates each vertex is mapped to, which often requires lots of storage memory. This causes inconvenien…
The sparse inverse covariance estimation problem is commonly solved using an -regularized Gaussian maximum likelihood estimator known as "graphical lasso", but its computational cost becomes prohibitive for large data sets. A recent line of results showed--under mild assumptions--that the graphical lasso esti…
GPU optimization speeds up large-scale classification tasks.
Deep-neural-network-based image reconstruction has demonstrated promising performance in medical imaging for under-sampled and low-dose scenarios. However, it requires large amount of memory and extensive time for the training. It is especially challenging to train the reconstruction networks for three-dimensional comp…
DREAM model improves computational efficiency for non-linear effects in relational event models.
For large matrix factorisation problems, we develop a distributed Markov Chain Monte Carlo (MCMC) method based on stochastic gradient Langevin dynamics (SGLD) that we call Parallel SGLD (PSGLD). PSGLD has very favourable scaling properties with increasing data size and is comparable in terms of computational requiremen…
Deep neural network models used for medical image segmentation are large because they are trained with high-resolution three-dimensional (3D) images. Graphics processing units (GPUs) are widely used to accelerate the trainings. However, the memory on a GPU is not large enough to train the models. A popular approach to …
SELD-TCN improves sound event localization and detection efficiency.
Efficiently infers time-varying sparse MRFs with strong statistical guarantees.
Graphical lasso may fail to fit models when data points are insufficient.
Study on rigidity of translating hypersurfaces not in graphical direction.
This is a short description of graphic lambda calculus, with special emphasis on a duality suggested by the two different appearances of knot diagrams, in lambda calculus and emergent algebra sectors of the graphic lambda calculus respectively. This duality leads to the introduction of the dual of the graphic beta move…
LSTM-MDNs improve risk forecasting during turbulent periods.
We consider the problem of learning high-dimensional Gaussian graphical models. The graphical lasso is one of the most popular methods for estimating Gaussian graphical models. However, it does not achieve the oracle rate of convergence. In this paper, we propose the graphical nonconvex optimization for optimal estimat…
New method for tuning Graphical Lasso hyperparameters.
Consider a mean curvature flow of hypersurfaces in Euclidean space, that is initially graphical inside a cylinder. There exists a period of time during which the flow is graphical inside the cylinder of half the radius. Here we prove a lower bound on this period depending on the Lipschitz-constant of the initial graphi…
Undirected graphical models, or Markov networks, are a popular class of statistical models, used in a wide variety of applications. Popular instances of this class include Gaussian graphical models and Ising models. In many settings, however, it might not be clear which subclass of graphical models to use, particularly…
Bayesian method for estimating functional graphical models from neuroimaging data.
Paper introduces a nonparametric functional graphical model for random functions.
Graphical models improve portfolio optimization for financial time series.
Paper estimates non-causal graphical models using covariance extension and transportation distance.
In this paper, we prove a generalization of Rado's Theorem, a fundamental result of minimal surface theory, which says that minimal surfaces over a convex domain with graphical boundaries must be disks which are themselves graphical. We will show that, for a minimal surface of any genus, whose boundary is "almost graph…
Probabilistic graphical models combine the graph theory and probability theory to give a multivariate statistical modeling. They provide a unified description of uncertainty using probability and complexity using the graphical model. Especially, graphical models provide the following several useful properties: - Graphi…
rags2ridges simplifies graphical modeling of high-dimensional data.
We consider the task of estimating a Gaussian graphical model in the high-dimensional setting. The graphical lasso, which involves maximizing the Gaussian log likelihood subject to an l1 penalty, is a well-studied approach for this task. We begin by introducing a surprising connection between the graphical lasso and hi…
Deep RNNs excel at capturing long-term dependencies in sequential data.
Efficient algorithms solve joint graphical lasso problems.
The study proves stability of various graphical translators in mean curvature flow.
Nonparametric undirected graphical model selection using diffusion models
Graphical notation simplifies tensor operations and decompositions.
Method solves Gaussian graphical models on ladder graphs efficiently.
New graphical criteria for efficient covariate adjustment in non-parametric causal models.
Graphically discrete groups have strong rigidity properties.
Continual Learning in artificial neural networks suffers from interference and forgetting when different tasks are learned sequentially. This paper introduces the Active Long Term Memory Networks (A-LTM), a model of sequential multi-task deep learning that is able to maintain previously learned association between sens…