Typical large-scale recommender systems use deep learning models that are stored on a large amount of DRAM. These models often rely on embeddings, which consume most of the required memory. We present Bandana, a storage system that reduces the DRAM footprint of embeddings, by using Non-volatile Memory (NVM) as the prim…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
GrateTile optimizes CNN feature map storage for efficient data access.
Convolutional neural networks (CNNs) demand huge DRAM bandwidth for computational imaging tasks, and block-based processing has recently been applied to greatly reduce the bandwidth. However, the induced additional computation for feature recomputing or the large SRAM for feature reusing will degrade the performance or…
Deep learning is attracting significant interest in the neuroimaging community as a means to diagnose psychiatric and neurological disorders from structural magnetic resonance images. However, there is a tendency amongst researchers to adopt architectures optimized for traditional computer vision tasks, rather than des…
CKA with Gaussian RBF kernels converges linearly as bandwidth increases.
Changing kernel bandwidth during training improves kernel regression performance.
A new method for faster bandwidth selection in Gaussian kernel ridge regression.
We provide a way to infer about existence of topological circularity in high-dimensional data sets in from its projection in obtained through a fast manifold learning map as a function of the high-dimensional dataset and a particular choice of a positive real known as band…
The paper bounds bandwidth and focal radius for manifolds with positive isotropic curvature.
Support vector data description (SVDD) is a popular technique for detecting anomalies. The SVDD classifier partitions the whole space into an inlier region, which consists of the region near the training data, and an outlier region, which consists of points away from the training data. The computation of the SVDD class…
Proposes GRAB-MDM for robust multiview data fusion.
Consistency of the kernel density estimator requires that the kernel bandwidth tends to zero as the sample size grows. In this paper we investigate the question of whether consistency is possible when the bandwidth is fixed, if we consider a more general class of weighted KDEs. To answer this question in the affirmativ…
Study bandwidth-limited training and inference of language models.
Algorithm selects variables and bandwidths for geographically weighted regression.
New method selects optimal bandwidth for price return density estimation, impacting efficient market hypothesis evaluation.
An accurate and fast estimation of the available bandwidth in a network with varying cross-traffic is a challenging task. The accepted probing tools, based on the fluid-flow model of a bottleneck link with first-in, first-out multiplexing, estimate the available bandwidth by measuring packet dispersions. The estimation…
Support vector data description (SVDD) is a popular anomaly detection technique. The SVDD classifier partitions the whole data space into an inlier region, which consists of the region near the training data, and an outlier region, which consists of points away from the training data. The computation of the SVDD classi…
This paper improves bandwidth selectors for SPBNs to enhance their performance.
New technique reduces memory usage and boosts neural network training speed.
We consider recommendation systems that need to operate under wireless bandwidth constraints, measured as number of broadcast transmissions, and demonstrate a (tight for some instances) tradeoff between regret and bandwidth for two scenarios: the case of multi-armed bandit with context, and the case where there is a la…
This paper proposes a new method for automatically selecting the optimal kernel bandwidth in density estimation.
Study shows that ridgeless Gaussian kernel regression overfits even with varying bandwidth or dimensionality.
New approach to adaptively select bandwidths in nonparametric regression.
Study provides bounds for estimating intrinsic dimension using Gaussian kernels.
New method tightens federated probe-logit distillation rates under varying bandwidths.
We propose a flexible nonparametric regression method for ultrahigh-dimensional data. As a first step, we propose a fast screening method based on the favored smoothing bandwidth of the marginal local constant regression. Then, an iterative procedure is developed to recover both the important covariates and the regress…
Local Gaussian correlation struggles in tails but a new method improves it.
We explore the performance of several automatic bandwidth selectors, originally designed for density gradient estimation, as data-based procedures for nonparametric, modal clustering. The key tool to obtain a clustering from density gradient estimators is the mean shift algorithm, which allows to obtain a partition not…
Support Vector Data Description (SVDD) provides a useful approach to construct a description of multivariate data for single-class classification and outlier detection with various practical applications. Gaussian kernel used in SVDD formulation allows flexible data description defined by observations designated as sup…
Study neural communication systems with bandwidth-limited channels.
Extends FC-RAG to anytime-valid sequential coverage for language model swarms.
A streaming algorithm estimates quadratic covariation from financial data efficiently.
PCA-Triage optimizes sensor data sampling for IoT networks.
The study optimizes Gaussian process approximations for finite-rank models.
SmartDeal reduces energy and storage costs for deep neural networks.
Paper proves MS convergence for radially symmetric kernels with large bandwidths.
A-FADMM improves FL scalability and privacy via wireless channel perturbations and interference.
Estimates bandwidth for CMC initial data sets.
Estimators of information theoretic measures such as entropy and mutual information are a basic workhorse for many downstream applications in modern data science. State of the art approaches have been either geometric (nearest neighbor (NN) based) or kernel based (with a globally chosen bandwidth). In this paper, we co…
Support Vector Data Description (SVDD) is a machine-learning technique used for single class classification and outlier detection. SVDD formulation with kernel function provides a flexible boundary around data. The value of kernel function parameters affects the nature of the data boundary. For example, it is observed …
We address the problem of setting the kernel bandwidth used by Manifold Learning algorithms to construct the graph Laplacian. Exploiting the connection between manifold geometry, represented by the Riemannian metric, and the Laplace-Beltrami operator, we set the bandwidth by optimizing the Laplacian's ability to preser…
Large-scale distributed training requires significant communication bandwidth for gradient exchange that limits the scalability of multi-node training, and requires expensive high-bandwidth network infrastructure. The situation gets even worse with distributed training on mobile devices (federated learning), which suff…
In kernel methods, the median heuristic has been widely used as a way of setting the bandwidth of RBF kernels. While its empirical performances make it a safe choice under many circumstances, there is little theoretical understanding of why this is the case. Our aim in this paper is to advance our understanding of the …
Kernel Density Estimation is a very popular technique of approximating a density function from samples. The accuracy is generally well-understood and depends, roughly speaking, on the kernel decay and local smoothness of the true density. However concrete statements in the literature are often invoked in very specific …
The problem of adaptive noisy clustering is investigated. Given a set of noisy observations , , the goal is to design clusters associated with the law of 's, with unknown density with respect to the Lebesgue measure. Since we observe a corrupted sample, a direct approach as the popular …
Nanowire field-effect sensors have recently been developed for label-free detection of biomolecules. In this work, we introduce a computational technique based on Bayesian estimation to determine the physical parameters of the sensor and, more importantly, the properties of the analyte molecules. To that end, we first …
In this paper, we study how to solve resource allocation problems in ultra-reliable and low-latency communications by unsupervised deep learning, which often yield functional optimization problems with quality-of-service (QoS) constraints. We take a joint power and bandwidth allocation problem as an example, which mini…
Paper provides an upper bound for bias of Nadaraya-Watson kernel regression.