We present a novel method for neural network quantization that emulates a non-uniform k-quantile quantizer, which adapts to the distribution of the quantized parameters. Our approach provides a novel alternative to the existing uniform quantization techniques for neural networks. We suggest to compare the results as …
Convolutional Neural Networks (CNN) has become more popular choice for various tasks such as computer vision, speech recognition and natural language processing. Thanks to their large computational capability and throughput, GPUs ,which are not power efficient and therefore does not suit low power systems such as mobil…
We propose Additive Powers-of-Two~(APoT) quantization, an efficient non-uniform quantization scheme for the bell-shaped and long-tailed distribution of weights and activations in neural networks. By constraining all quantization levels as the sum of Powers-of-Two terms, APoT quantization enjoys high computational effic…
A method for reducing neural network size using look-up tables.
problem Reducing memory and computational footprint of deep neural networks.
method Iteratively learns value dictionaries and assignment matrices for network weights.
result General framework for network reduction that can handle various reduction problems.
DBQ quantizes lightweight networks efficiently for resource-constrained devices.
problem High computational and storage complexity of deep neural networks on resource-constrained devices.
method A differentiable non-uniform quantizer that can be mapped onto efficient ternary-based dot product engines.
result Achieves state-of-the-art results with minimal training overhead and best accuracy-complexity trade-off.
This study analyzes quantization in deep learning models using statistical physics methods.
problem The computational resource requirements for large-scale data analysis models.
method Typical case analysis from statistical physics, specifically the replica method.
result Optimal quantization width minimizes error and delays overfitting.
Degree-Quant improves GNN efficiency by quantizing them without losing accuracy.
problem Efficiency of graph neural networks at inference time.
method Architecturally-agnostic method, Degree-Quant, for quantizing GNNs.
result Degree-Quant trained models perform as well as full-precision models and achieve up to 26% gains.
SQWA improves low-precision DNNs with model averaging and quantization.
problem Designing good generalization DNNs with quantized weights.
method Floating-point model training, direct quantization, multiple low-precision models, weight averaging, re-quantization, fine-tuning, loss visualization.
result State-of-the-art results for 2-bit QDNNs on CIFAR-100 and ImageNet datasets.
Improved deep learning model deployment on tiny MCUs with mixed-precision quantization.
problem Memory limitations prevent accurate deployment of DNN models on tiny MCUs.
method Automated mixed-precision quantization using Reinforcement Learning for MCU constraints.
result Mixed-precision models achieve high accuracy with uniform quantization policies.
Paper proposes a learning-based sparse Bayesian method for accurate off-grid DOA estimation.
problem One-bit off-grid direction of arrival (DOA) estimation in a single snapshot scenario.
method Formulated off-grid DOA estimation model, used Sparse Bayesian framework, proposed Learning-based Sparse Bayesian approach.
result Improved computational efficiency and accuracy in off-grid DOA estimation.
QuantEase optimizes LLMs with CD-based quantization, achieving state-of-the-art performance.
problem Efficiently quantize large language models for deployment.
method Layer-wise quantization using CD-based algorithms with matrix and vector operations.
result State-of-the-art performance in perplexity and zero-shot accuracy.
Unified framework for uniform signal recovery in nonlinear GCS with 1-bit/quantized measurements.
problem Uniform recovery guarantees for nonlinear generative compressed sensing.
method Unified framework using generalized Lasso and Lipschitz approximation.
result Uniform recovery of all signals in the ball up to an error of ε using approximately O(k/ε^2) samples.
The task of estimating a matrix given a sample of observed entries is known as the \emph{matrix completion problem}. Most works on matrix completion have focused on recovering an unknown real-valued low-rank matrix from a random sample of its entries. Here, we investigate the case of highly quantized observations when …
Paper proposes LC-Checkpoint for efficient deep learning model checkpoints.
problem Efficient construction of checkpoints for deep learning models.
method Lossy compression scheme using quantization and priority promotion with Huffman coding.
result LC-Checkpoint achieves up to 28x compression and 5.77x speedup over SCAR.
The study finds non-uniform lattices with thin Hitchin representations in specific Lie groups.
problem Finding thin Hitchin representations in non-uniform lattices of Lie groups.
method Arithmetic methods to construct thin Hitchin representations.
result Infinitely many orbits of thin Hitchin representations in non-uniform lattices.
New approach to adversarial robustness with non-uniform perturbations.
problem Real-world adversaries craft adversarial examples with non-uniform perturbations.
method Proposes non-uniform perturbations based on feature dependencies and data distribution.
result Shows improved robustness to real-world attacks compared to uniform perturbations.
This work studies the robustness certification problem of neural network models, which aims to find certified adversary-free regions as large as possible around data points. In contrast to the existing approaches that seek regions bounded uniformly along all input features, we consider non-uniform bounds and use it to …
New method upsamples sparse, non-uniform point clouds more accurately.
problem Suboptimal results from existing point cloud upsampling methods.
method Imposes manifold distribution constraints using Gaussian functions.
result Generates higher-quality, more uniformly distributed dense point clouds.
New approach finds minima of geodesic lengths for non-uniform fillings.
problem Finding minima of geodesic length functions for non-uniform fillings.
method Elementary optimization for 4-regular topological fillings, analysis of fat graphs and optimization techniques.
result Minima of geodesic length functions are found to be at triangle surfaces in both analyzed classes of non-uniform fillings.
New method uses graphene transistors for efficient non-uniform random number generation.
problem Generating non-uniform random variates efficiently.
method GFET-based hardware non-uniform random number generator.
result Demonstrated speedup of Monte Carlo integration by up to 2x.
Study shows how non-uniform scaling affects persistence diagrams.
problem Stability of persistence diagrams under non-uniform scaling.
method Explicit bounds on bottleneck distance derived for Euclidean scaling.
result Explicit bounds on the stability of persistence diagrams under non-uniform scaling.
Unified framework for non-uniform materials evolving over time.
problem Dealing with non-uniform materials evolving over time.
method Constructing a material groupoid and material distribution.
result Unified framework for general non-uniform evolution materials.
New rigidity theorem for product of lattices.
problem Understanding quasi-isometry of product lattices.
method Demonstrated rigidity for product of non-uniform rank one lattice and nilpotent lattice.
result Any quasi-isometric group is an extension of a non-uniform rank one lattice by a nilpotent lattice.
Non-uniform lattices in PU(n,1) cannot geometrically act on CAT(0) cube complexes.
problem Non-uniform lattices in PU(n,1) cannot geometrically act on CAT(0) cube complexes.
method Proving non-uniform lattices in PU(n,1) cannot geometrically act on CAT(0) cube complexes.
result Non-uniform lattices in PU(n,1) cannot geometrically act on CAT(0) cube complexes.
We apply stochastic average gradient (SAG) algorithms for training conditional random fields (CRFs). We describe a practical implementation that uses structure in the CRF gradient to reduce the memory requirement of this linearly-convergent stochastic gradient method, propose a non-uniform sampling scheme that substant…
We study primal-dual type stochastic optimization algorithms with non-uniform sampling. Our main theoretical contribution in this paper is to present a convergence analysis of Stochastic Primal Dual Coordinate (SPDC) Method with arbitrary sampling. Based on this theoretical framework, we propose Optimality Violation-ba…
New method detects communities in complex hypergraphs, matching theoretical limits.
problem Detecting communities in non-uniform hypergraphs with varying hyperedge sizes.
method Developed a spectral theory for weighted non-backtracking operators on non-uniform hypergraphs.
result Achieved the Kesten-Stigum bound for weak recovery in a general class of non-uniform HSBMs.
Sharp threshold for exact recovery in non-uniform hypergraph stochastic block model.
problem Community detection in random hypergraphs with non-uniform hyperedge probabilities.
method Sharp threshold established; two efficient algorithms for exact recovery.
result Sharp threshold for exact recovery; information-theoretic lower bound on misclassification.
We prove that if G is a non-uniform lattice in a rank-one semi-simple Lie group $\ne Isom(\H^2_\R)$ then G is quasi-isometrically co-Hopf. This means that every quasi-isometric embedding G→G is coarsely onto and thus is a quasi-isometry.
Improved sampling accuracy in SG-MCMC methods via non-uniform gradient subsampling.
problem Computational inefficiency and sampling error in stochastic gradient MCMC methods.
method Proposes a non-uniform subsampling scheme to reduce sampling error in EWSG, a variant of SG-MCMC.
result EWSG reduces sampling error compared to uniform subsampling, improving accuracy without sacrificing convergence speed.
A groupoid Ω(B) called material groupoid is naturally associated to any simple body B. The material distribution is introduced due to the (possible) lack of differentiability of the material groupoid. Thus, the inclusion of these new objects in the theory of material bodies opens th…
In this note, we study deformations of a non-uniform real hyperbolic lattice in quaternionic hyperbolic spaces. Specially we show that the representations of the fundamental group of the figure eight knot complement into PU(2,1) cannot be deformed in PSp(2,1) out of PU(2,1) up to conjugacy.
We construct a three-point compact finite difference scheme on a non-uniform mesh for the time-fractional Black-Scholes equation. We show that for special graded meshes used in finance, the Tavella-Randall and the quadratic meshes the numerical solution has a fourth-order accuracy in space. Numerical experiments are di…
We improve GANs by enforcing reproducibility and using non-uniform sampling.
problem Overrepresentation of certain samples in GANs' marginal log-likelihood.
method Enforce reproducibility through matching empirical distribution to prior, use non-uniform sampling for mini-batch selection.
result Improved quality and variety in generated samples, validated on CIFAR10, Fashion MNIST, and CelebA.
We study the effectiveness of non-uniform randomized feature selection in decision tree classification. We experimentally evaluate two feature selection methodologies, based on information extracted from the provided dataset: (i) \emph{leverage scores-based} and (ii) \emph{norm-based} feature selection. Experimenta…
Improved matrix completion for non-uniformly sampled data.
problem Estimating unobserved entries in a matrix with varying sampling probabilities.
method Developed entry-specific bounds for low-rank matrix completion under structured non-uniform sampling.
result Error bounds for each entry match minimax lower bounds under certain conditions.
New loss function equivalence reveals PER's uniform sampling can be improved.
problem Improving Prioritized Experience Replay (PER) for better learning efficiency.
method Transforming non-uniformly sampled data loss functions into uniformly sampled ones.
result Some environments can replace PER with a new loss function without performance loss.
Two algorithms converge to dictionary learning with geometric rate for non-uniform data.
problem Dictionary learning for non-uniform data models.
method Derivation of convergence conditions for MOD and ODL.
result Both algorithms converge to the generating dictionary with geometric rate under certain conditions.
Proof shows volumes of certain geometric representations are always integers.
problem Integrality of volumes of specific geometric representations.
method Elementary, combinatorial-geometrical proof.
result Volumes of representations are integers when n≥2. Spectral algorithm recovers community structure in sparse hypergraphs.
problem Community detection in sparse random hypergraphs with community structure and higher-order interactions.
method Spectral algorithm with three steps: hyperedge selection, spectral partition, and correction/merging.
result Weak consistency achieved for weak signal-to-noise ratio.
The paper proves Zimmer's conjecture for non-uniform lattices by controlling mass escape and Lyapunov exponents.
problem Proving Zimmer's conjecture for non-uniform lattices in higher-rank semisimple Lie groups.
method Establishes finiteness of low-dimensional actions, introduces novel techniques to control mass escape and Lyapunov exponents.
result Proves Zimmer's conjecture for many non-uniform lattices, improving previous results.
Survey on quantization methods on Kähler manifolds.
problem None explicitly stated; focuses on methods.
method Deformation quantization, geometric quantization, Berezin-Toeplitz quantization, BV quantization.
result New relationships among quantization methods on Kähler manifolds.
The study finds that certain hyperbolic manifolds contain subgroups isomorphic to surface groups.
problem The existence of thin surface subgroups in non-uniform arithmetic lattices.
method Analyzes arithmetic hyperbolic manifolds and their fundamental groups.
result Fundamental groups of non-compact arithmetic hyperbolic manifolds contain thin surface subgroups.
Validates conformal prediction for network data under non-uniform sampling.
problem Validity of conformal prediction for network data under non-representative sampling.
method Interprets sampling mechanisms as selection rules, studies validity conditional on selection events, uses permutation invariance and joint exchangeability.
result Finite-sample validity of conformal prediction for certain selection events and asymptotic validity for random walk sampling.
The paper classifies quantizable functions and explores symmetry in quantization methods.
problem Classifying quantizable functions and understanding symmetry in quantization methods.
method Deformation quantization and geometric quantization methods are compared and classified.
result Formal quantizable functions are of a specific form and relate to Hamiltonian Killing vector fields.
A new method for matrix completion with model-free weights.
problem Matrix completion under non-uniform missing structures.
method Constructs weights via convex optimization to adjust for non-uniformity without modeling observation probabilities.
result Recover matrix with stronger theoretical guarantees, especially in heterogeneous missing settings.
We derive high-order compact finite difference schemes for option pricing in stochastic volatility models on non-uniform grids. The schemes are fourth-order accurate in space and second-order accurate in time for vanishing correlation. In our numerical study we obtain high-order numerical convergence also for non-zero …
Let G and G′ be simple Lie groups of equal real rank and real rank at least 2. Let Γ<G and Λ<G′ be non-uniform lattices. We prove a theorem that often implies that any quasi-isometric embedding of Γ into Λ is at bounded distance from a homomorphism. For example, any quasi-isometric embedding of $SL(n,\ma…