Single global merging boosts decentralized learning performance.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Federated learning for Bayesian clustering of large datasets.
Discrete knot theory models use lattice-filtered graphs to detect merging knot components.
FedFMC improves federated learning on non-iid data without sharing data or increasing communication costs.
There are many statistical tests that verify the null hypothesis: the variable of interest has the same distribution among k-groups. But once the null hypothesis is rejected, how to present the structure of dissimilarity between groups? In this article, we introduce The Merging Path Plot - a methodology, and factorMerg…
LEWIS merges LLMs without training, improving performance on specific tasks.
A method merges two pretrained diffusion experts to improve image quality and likelihood.
In big data image/video analytics, we encounter the problem of learning an overcomplete dictionary for sparse representation from a large training dataset, which can not be processed at once because of storage and computational constraints. To tackle the problem of dictionary learning in such scenarios, we propose an a…
Study shows training duration impacts model merging quality, suggesting joint selection of duration and method.
Study shows training duration affects model merging quality, suggesting joint selection of duration and method.
EpiMer merges models by solving Fréchet mean on a Riemannian manifold.
New Bayesian method improves Pareto front estimation in multitask finetuning.
Deep reinforcement learning (DRL) on Markov decision processes (MDPs) with continuous action spaces is often approached by directly training parametric policies along the direction of estimated policy gradients (PGs). Previous research revealed that the performance of these PG algorithms depends heavily on the bias-var…
Study on merging predictors in causal and anticausal directions using CMAXENT.
In this paper, a similarity-driven cluster merging method is proposed for unsuper-vised fuzzy clustering. The cluster merging method is used to resolve the problem of cluster validation. Starting with an overspecified number of clusters in the data, pairs of similar clusters are merged based on the proposed similarity-…
New method merges MCMC samples without distributional assumptions.
Securely evaluates the benefits of merging datasets for causal estimation.
A new method merges neural networks using CCA to improve model performance.
This paper reverses a construction by merging boundary critical points into an interior one.
Budgeted Stochastic Gradient Descent (BSGD) is a state-of-the-art technique for training large-scale kernelized support vector machines. The budget constraint is maintained incrementally by merging two points whenever the pre-defined budget is exceeded. The process of finding suitable merge partners is costly; it can a…
NAMEx merges experts using Nash bargaining for improved performance.
Efficiently estimates longitudinal networks by merging sparse networks.
A new method reduces task interference in model merging.
Boosting strategies for merging vs. ensembling studies analyzed.
We propose a novel neural network architecture, named the Global Workspace Network (GWN), which addresses the challenge of dynamic and unspecified uncertainties in multimodal data fusion. Our GWN is a model of attention across modalities and evolving through time, and is inspired by the well-established Global Workspac…
Method constructs finance LLMs without instruction data using pretraining and model merging.
Deep learning algorithms for connectomics rely upon localized classification, rather than overall morphology. This leads to a high incidence of erroneously merged objects. Humans, by contrast, can easily detect such errors by acquiring intuition for the correct morphology of objects. Biological neurons have complicated…
Gradient boosted decision trees (GBDT) is the leading algorithm for many commercial and academic data applications. We give a deep analysis of this algorithm, especially the histogram technique, which is a basis for the regulized distribution with compact support. We present three new modifications. 1) Share memory tec…
Although a key driver of Earth's climate system, global land-atmosphere energy fluxes are poorly constrained. Here we use machine learning to merge energy flux measurements from FLUXNET eddy covariance towers with remote sensing and meteorological data to estimate net radiation, latent and sensible heat and their uncer…
DPSM clusters nodes in data and graph spaces via density propagation and subcluster merging.
Bayesian Federated Inference improves survival model analysis without merging data.
In order to drive safely and efficiently under merging scenarios, autonomous vehicles should be aware of their surroundings and make decisions by interacting with other road participants. Moreover, different strategies should be made when the autonomous vehicle is interacting with drivers having different level of coop…
In this paper, we propose a nonlinear distance metric learning scheme based on the fusion of component linear metrics. Instead of merging displacements at each data point, our model calculates the velocities induced by the component transformations, via a geodesic interpolation on a Lie transfor- mation group. Such vel…
New methods merge discrete gradient fields from patches to correct errors.
The hierarchical Dirichlet process (HDP) has become an important Bayesian nonparametric model for grouped data, such as document collections. The HDP is used to construct a flexible mixed-membership model where the number of components is determined by the data. As for most Bayesian nonparametric models, exact posterio…
The article improves the display of acceptable exchange ratios for merging companies.
NetFuse merges different DNN models with varying weights for faster inference.
The MAXENT principle helps merge datasets to infer causal effects.
Signed-permutation coordinate transport improves model alignment across checkpoints.
A critical decision point when training predictors using multiple studies is whether studies should be combined or treated separately. We compare two multi-study prediction approaches in the presence of potential heterogeneity in predictor-outcome relationships across datasets: 1) merging all of the datasets and traini…
Split-Merge MCMC (Monte Carlo Markov Chain) is one of the essential and popular variants of MCMC for problems when an MCMC state consists of an unknown number of components. It is well known that state-of-the-art methods for split-merge MCMC do not scale well. Strategies for rapid mixing requires smart and informative …
Paper proposes a forecasting model combining autoregressive models with spectral attention.
Effects connected with the world globalization affect also the financial markets. On a way towards quantifying the related characteristics we study the financial empirical correlation matrix of the 60 companies which both the Deutsche Aktienindex (DAX) and the Dow Jones (DJ) industrial average comprised during the year…
Finding the optimal -means clustering is NP-hard in general and many heuristics have been designed for minimizing monotonically the -means objective. We first show how to extend Lloyd's batched relocation heuristic and Hartigan's single-point relocation heuristic to take into account empty-cluster and single-poin…
Merging datasets is a key operation for data analytics. A frequent requirement for merging is joining across columns that have different surface forms for the same entity (e.g., the name of a person might be represented as "Douglas Adams" or "Adams, Douglas"). Similarly, ontology alignment can require recognizing disti…
HPSGD speeds up DNN training by paralleling data sync with local training.
Improved sampling for network community detection.
Recommender systems leverage product and community information to target products to consumers. Researchers have developed collaborative recommenders, content-based recommenders, and (largely ad-hoc) hybrid systems. We propose a unified probabilistic framework for merging collaborative and content-based recommendations…