Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,878 papers · 148 categories

Trend · papers per month

0.3%0.5%0.8%0.4% · Dec 201019922001200920172026
42 results for two-block

The study examines MCMC methods for arbitrary objectives and finds likelihood sharpness impacts performance and regularization.

problem Limitations of MCMC methods for arbitrary objective functions.
method Two-block MCMC framework with Metropolis-Hastings and Gibbs sampling, exploring likelihood curvature and sharpness.
result Likelihood sharpness governs in-sample performance and regularization inferred by training data.

We consider the community detection problem in sparse random hypergraphs. Angelini et al. (2015) conjectured the existence of a sharp threshold on model parameters for community detection in sparse hypergraphs generated by a hypergraph stochastic block model. We solve the positive part of the conjecture for the case of…

2019-04-11abs ↗pdf ↗

New method speeds up solving machine learning problems by splitting variables randomly.

problem Solving large-scale machine learning and signal processing problems efficiently.
method Randomized Multi-Block ADMM (RAC-MBADMM) for convex and nonconvex quadratic optimization.
result RAC-MBADMM converges linearly and outperforms other algorithms in solution time and quality.

The study examines algebraic structures of specific tensor forms in four-dimensional spacetimes.

problem Investigating algebraic features of certain tensor forms in spacetimes.
method General treatment followed by specialization to four-dimensional spacetimes, focusing on invariant subspaces and generalizing relations.
result Generalized relations such as the Ruse-Lanczos identity, Bel-Matte decomposition, and Lovelock-like quadratic identities.

Construction of (colored) knot polynomials for double-fat graphs is further generalized to the case when "fingers" and "propagators" are substituting R-matrices in arbitrary closed braids with m-strands. Original version of arXiv:1504.00371 corresponds to the case m=2, and our generalizations sheds additional light on …

2015-06-01abs ↗pdf ↗

Co-Clustering, the problem of simultaneously identifying clusters across multiple aspects of a data set, is a natural generalization of clustering to higher-order structured data. Recent convex formulations of bi-clustering and tensor co-clustering, which shrink estimated centroids together using a convex fusion penalt…

2019-01-18abs ↗pdf ↗

DAG models with hidden variables present many difficulties that are not present when all nodes are observed. In particular, fully observed DAG models are identified and correspond to well-defined sets ofdistributions, whereas this is not true if nodes are unobserved. Inthis paper we characterize exactly the set of dist…

2013-01-10abs ↗pdf ↗

Unified sampling approach for Bayesian imaging problems.

problem Sampling from complex prior and posterior distributions in Bayesian imaging.
method Gaussian latent machine model for efficient prior and posterior sampling.
result Unified and generalized sampling algorithms for various imaging problems.

Improved VGG networks enhance image classification accuracy.

problem Enhancing image classification accuracy using modified VGG architectures.
method Two improved VGG architectures were created by freezing the first two blocks and applying different dilation rates in the last three blocks.
result Significant out-performance on image classification tasks on CIFAR-10 and CIFAR-100 datasets.

MTCNet uses MTL to estimate crowd density and count.

problem Crowd count estimation challenges due to scale variations and perspective.
method MTL deep neural network architecture with two tasks: density estimation and count classification.
result Achieves lower MAE than state-of-the-art methods on multiple datasets.

The performance of spectral clustering can be considerably improved via regularization, as demonstrated empirically in Amini et. al (2012). Here, we provide an attempt at quantifying this improvement through theoretical analysis. Under the stochastic block model (SBM), and its extensions, previous results on spectral c…

2013-12-05abs ↗pdf ↗

A common approach to analyze a covariate-sample count matrix, an element of which represents how many times a covariate appears in a sample, is to factorize it under the Poisson likelihood. We show its limitation in capturing the tendency for a covariate present in a sample to both repeat itself and excite related ones…

2016-04-25abs ↗pdf ↗

Many modern data mining applications are concerned with the analysis of datasets in which the observations are described by paired high-dimensional vectorial representations or "views". Some typical examples can be found in web mining and genomics applications. In this article we present an algorithm for data clusterin…

2012-02-02abs ↗pdf ↗

Develops a new model to better predict corporate bond yields.

problem Persistent shifts in interest rates undermine single-regime models.
method Regime-switching generalized CIR model with two-state short-rate process and credit factors.
result The model improves joint curve fit and delivers interpretable probabilities.

A deep learning model corrects precipitation bias without expert knowledge.

problem Precipitation bias in numerical predictions due to limited observation and models.
method Data-driven deep learning model with Denoising Autoencoder and Ordinal Regression blocks.
result The model achieves the best correcting performance and TS compared to classical methods.

Proposes a new neural network architecture combining MLP and basis functions.

problem Function approximation and operator learning in scientific machine learning.
method Combines robust MLP inner functions with flexible basis functions outer functions.
result KKAN outperforms MLPs and KANs in function approximation and operator learning tasks.

Random matrix theory explains transient signal detectability in early-stopped gradient flow.

problem Transient signal detectability in early-stopped gradient flow.
method Random matrix theory applied to gradient flow in a linear teacher-student setting.
result Transient Baik-Ben Arous-Péché (BBP) transition in learning dynamics due to anisotropy and noise.

Model clusters networks and their communities simultaneously.

problem Clustering networks and their communities in unlabeled, heterogeneous networks.
method Nested Stochastic Block Model (NSBM) with Bayesian approach and NDP prior.
result Model accurately estimates both within and across network clustering structures.

Two new FedMF algorithms improve data clustering in federated learning.

problem Challenges in federated matrix factorization with non-convex and non-smooth problems.
method Proposed FedMAvg and FedMGS algorithms based on model averaging and gradient sharing principles.
result Convergence analyses and experiment results show improved performance over existing distributed clustering algorithms.

The paper proposes AIS for Bayesian inversion of multioutput signals with covariance estimation.

problem Performing uncertainty analysis of covariance matrices in Bayesian inversion problems for multioutput signals.
method Adaptive Importance Sampling (AIS) scheme, split variables, frequentist approach for noise covariance, prior density over covariance matrix.
result Estimation of model parameters and covariance matrix of noise.

Study extends neural network approximation to time-varying PDEs using Fourier-Lebesgue spaces.

problem Limitation to static PDEs and different time-domain regularity.
method Extend spectral Barron spaces to anisotropic weighted Fourier-Lebesgue spaces, measure approximation error in Bochner-Sobolev norm.
result Established bound on approximation rate for functions in anisotropic weighted Fourier-Lebesgue spaces.

Profile graphical models represent multivariate dependence under varying risk factors.

problem Capturing varying conditional independence structures across different levels of a risk factor.
method Introducing a novel class of graphical models (profile graphical models) that represent multivariate dependence under varying risk factors, and developing a Bayesian approach for learning shared sparsity structures.
result Demonstrated enhanced ability to capture subject-specific differences in protein network data from acute myeloid leukemia.

Alternative approach to generative modeling using convex conjugates and optimal transport.

problem Traditional generative modeling splits sampling and mapping; this work explores an alternative.
method Inspired by moment measures, proposes a new factorization and uses optimal transport for recovery.
result Intuitive results on factorized distributions, showing potential for practical tasks.

New gradient coding schemes reduce decoding error in both random and adversarial straggler settings.

problem Creating efficient approximate gradient coding schemes for distributed optimization.
method Introduced novel approximate gradient codes based on expander graphs, achieving optimal decoding coefficients.
result Achieved nearly optimal error in random setting and nearly half the error in adversarial setting compared to existing codes.

New method recovers sparse signals from nonlinear observations with robust error bounds.

problem Recovering two sparse vectors from nonlinearly mixed observations with limited data.
method Regularization-based framework combining Huberized data fidelity and generalized folded-concave penalties with a proximal alternating algorithm.
result Estimation error bounds of order σslog(n)/mσ\sqrt{s\log(n)/m} at every localized stationary point, with oracle rate σs/mσ\sqrt{s/m} under beta-min condition.

We analyze mixing times of three DA algorithms for regression models.

problem Analyzing mixing times of data augmentation algorithms for regression models.
method Modified conductance-based method to study mixing times of Gibbs samplers.
result Prove non-asymptotic polynomial upper bounds on mixing times for DA algorithms.