Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

79158237316 · Jun 202019922001200920172026
48 results for Weight preservation

Holonomy-preserving transformations help recover Alexander polynomials from graph zeta functions.

problem Recovering Alexander polynomials from graph zeta functions.
method Introducing holonomy to preserve zeta functions of matrix-weighted graphs and extending to group elements and quandles.
result Holonomy-preserving transformations correspond to transformations of group presentations and preserve the twisted Alexander polynomial.

Pion optimizes LLMs by preserving weight matrix singular values.

problem Training large language models (LLMs) with standard optimizers leads to unstable weight matrices.
method Pion uses orthogonal transformations to update weight matrices, preserving their singular values.
result Pion offers a stable alternative to standard optimizers for LLM pretraining and finetuning.

The use of inverse probability weighting (IPW) methods to estimate the causal effect of treatments from observational studies is widespread in econometrics, medicine and social sciences. Although these studies often involve sensitive information, thus far there has been no work on privacy-preserving IPW methods. We add…

2019-05-29abs ↗pdf ↗

Econometric framework integrates heavy-tailed distributions with behavioral probability weighting for better asset pricing.

problem Underestimation of Value-at-Risk by traditional models in asset pricing.
method Developed an econometric framework combining heavy-tailed Student's tt distributions with behavioral probability weighting.
result Student's tt specifications outperform Gaussian models in 88.4% of cases, reducing underestimation of Value-at-Risk by 16.5 percentage points.

Smoothly conjugate Anosov flows on 3D manifolds are actually smoothly conjugate.

problem Smoothly conjugate 3D Anosov flows are not always smoothly conjugate.
method Proved smooth rigidity for volume preserving Anosov flows on 3-manifolds.
result Smooth conjugacy implies smooth conjugacy for volume preserving Anosov flows.

New method improves tensor completion by selectively preserving important elements.

problem Recovering corrupted high-dimensional tensor data with missing entries and noise.
method Tensor weighted correlated total variation (TWCTV) regularizer with ADMM algorithm.
result Superior performance in image completion, denoising, and background subtraction tasks.

New federated method preserves privacy and estimates treatment effects.

problem Privacy-preserving causal inference for multi-site studies.
method Multiply robust nuisance function estimation, transfer learning.
result Efficient and optimal treatment effect estimation under different scenarios.

In this work we study the properties of deep neural networks (DNN) with random weights. We formally prove that these networks perform a distance-preserving embedding of the data. Based on this we then draw conclusions on the size of the training data and the networks' structure. A longer version of this paper with more…

2014-12-18abs ↗pdf ↗

BiTAT improves neural network quantization for edge devices by focusing on weight dependencies and disentangling them.

problem Performance degradation of compact neural networks under extreme quantization.
method Task-dependent Aggregated Transformation (BiTAT) method that orthonormalizes weights and progressively quantizes them.
result BiTAT effectively preserves model performance on ImageNet and CIFAR-100 with compact backbones.

Before training a neural net, a classic rule of thumb is to randomly initialize the weights so the variance of activations is preserved across layers. This is traditionally interpreted using the total variance due to randomness in both weights \emph{and} samples. Alternatively, one can interpret the rule of thumb as pr…

2019-02-13abs ↗pdf ↗

We study the interplay between sequential decision making and avoiding discrimination against protected groups, when examples arrive online and do not follow distributional assumptions. We consider the most basic extension of classical online learning: "Given a class of predictors that are individually non-discriminato…

2018-10-28abs ↗pdf ↗

It was shown by Kaup that every origin-preserving automorphism of quasi-circular domains is a polynomial mapping. In this paper, we study how the weight of quasi-circular domains and the degree of such automorphisms are related. By using the Bergman mapping, we prove that every origin-preserving automorphism of normal …

2014-03-17abs ↗pdf ↗

The study shows finite measure-preserving isometry groups for certain metric measure spaces.

problem Understanding the structure of isometry groups in metric measure spaces.
method Analyzing synthetic negative Ricci curvature and Bakry-Émery Ricci curvature.
result The measure-preserving isometry group is finite for compact metric measure spaces with specific curvature conditions.

Study mass transport in low-diffusivity using Lagrangian coordinates.

problem Mass preserving transport of passive tracers in low-diffusivity limit.
method Lagrangian coordinates, time-averaged diffusion equation, weighted manifold structure.
result Leading order asymptotics extend to dominant nontrivial singular value in low-diffusivity limit.

Improved diffusion models for image synthesis with better training dynamics.

problem Uneven and ineffective training in diffusion models.
method Redesigned network layers to preserve activation, weight, and update magnitudes.
result Significantly better networks at equal computational complexity, improving FID to 1.81.

TransNet improves community detection on target networks using privacy-preserved source networks.

problem Community detection on sensitive network data with privacy constraints.
method Spectral clustering framework leveraging locally distributed privacy-preserved auxiliary networks via randomized response and adaptive weighting.
result TransNet delivers strong gains in community detection across various privacy levels and heterogeneity patterns.

The paper generalizes K-stability results to singular and weighted settings.

problem Generalizing K-stability to singular and weighted settings.
method Generalization of results in \cite{Li22a} to singular and weighted settings.
result The \(\mathbb{G}\)-uniform weighted K-stability for models implies \(\mathbb{G}\)-coercivity of the weighted Mabuchi functional.

New method preserves privacy by aggregating feature-vectors with weighted sums, ensuring label differential privacy.

problem Ensuring privacy in training data aggregation for sensitive labels.
method Learning from bag aggregates (LBA) with weighted Gaussian sums, preserving label differential privacy (label-DP).
result Weighted LBA using iid Gaussian weights with mm randomly sampled disjoint kk-sized bags provides (ε,δ)(\varepsilon, δ)-label-DP.

We propose a novel approach to addressing the vanishing (or exploding) gradient problem in deep neural networks. We construct a new architecture for deep neural networks where all layers (except the output layer) of the network are a combination of rotation, permutation, diagonal, and activation sublayers which are all…

2019-11-21abs ↗pdf ↗

A new algorithm improves federated learning by combining knowledge distillation and weighted combination loss.

problem Non-IID client data in federated learning leads to model drift and poor generalization.
method pFedKD-WCL integrates knowledge distillation with bi-level optimization to address non-IID challenges.
result pFedKD-WCL outperforms state-of-the-art algorithms in accuracy and convergence speed.

This paper solves aggregation of Pareto optimal models by using Bayesian priors and weighted averaging.

problem How to rationally aggregate Pareto optimal models while preserving Pareto efficiency.
method Four logical steps: 1) Bayesian models, 2) Prior as preference ranking, 3) Consistent aggregation, 4) Weighted average of priors.
result All rational/consistent aggregation rules follow a generalized hierarchical Bayesian model.

Given a compact Riemannian manifold (M, g) and two positive functions ρρ and σσ, we are interested in the eigenvalues of the Dirichlet energy functional weighted by σσ, with respect to the L 2 inner product weighted by ρρ. Under some regularity conditions on ρρ and σσ, these eigenvalues are those of the operator …

2016-06-12abs ↗pdf ↗

We show that from an even degree symplectic NQ-manifold, whose homological vector field Q preserves the symplectic form, one can construct a weight system for tri-valent graphs with values in the Q-cohomology ring, satisfying the IHX relation. Likewise, given a representation of the homological vector field, one can co…

2011-10-24abs ↗pdf ↗

The study examines stable regions in weighted manifolds with boundary properties.

problem Studying stable regions in weighted manifolds with boundary properties.
method Using deformations constructed from parallel vector fields tangent to the boundary, the study deduces rigidity properties for stable sets.
result The classification of stable sets in some Riemannian cylinders and uniqueness results for minimizers.

R2D2-Net improves Bayesian neural networks by preventing over-shrinkage of important weights.

problem Bayesian neural networks struggle with choosing appropriate priors, leading to over-shrinkage or poor predictive performance.
method Proposes R2D2-Net with an R^2-induced Dirichlet Decomposition prior and variational Gibbs inference algorithm.
result R2D2-Net effectively shrinks irrelevant coefficients while preventing key features from over-shrinkage.

Incorporation of a new knowledge into neural networks with simultaneous preservation of the previous one is known to be a nontrivial problem. This problem becomes even more complex when new knowledge is contained not in new training examples, but inside the parameters (connection weights) of another neural network. Her…

2018-09-25abs ↗pdf ↗

We introduce a quotient of the affine Temperley-Lieb category that encodes all weight-preserving linear maps between finite-dimensional sl(2)-representations. We study the diagrammatic idempotents that correspond to projections onto extremal weight spaces and find that they satisfy similar properties as Jones-Wenzl pro…

2017-01-09abs ↗pdf ↗

Constructs a support-preserving homotopy for differential forms with boundary decay estimates.

problem Non-uniqueness of chain homotopies in de Rham complexes with boundary decay properties.
method Constructs a specific chain homotopy with desirable support propagation and boundary decay estimates.
result Obtains a support-preserving right inverse of the divergence operator with optimal decay estimates.

New mechanism protects neural network weights from privacy attacks during self-supervised learning.

problem Privacy risks during fine-tuning stage of self-supervised learning.
method Proposes a novel differential privacy mechanism using additive logistic noise.
result Reduces membership inference attack accuracy to 50% while maintaining below 5% performance loss.

Binary autoencoder with sparse hidden layer preserves information and zero reconstruction error.

problem Preserving information and zero reconstruction error in binary neural networks.
method Binary autoencoder with random binary weights, sparse hidden layer, and varying neuron thresholds.
result Zero reconstruction error for any input with a large hidden layer and varying neuron thresholds.

Since nn-dimensional λλ-hypersurfaces in the Euclidean space Rn+1\mathbb {R}^{n+1} are critical points of the weighted area functional for the weighted volume-preserving variations, in this paper, we study the rigidity properties of complete λλ-hypersurfaces. We give a gap theorem of complete λλ-hypersurfaces with po…

2014-03-17abs ↗pdf ↗

One of the difficulties of training deep neural networks is caused by improper scaling between layers. Scaling issues introduce exploding / gradient problems, and have typically been addressed by careful scale-preserving initialization. We investigate the value of preserving scale, or isometry, beyond the initial weigh…

2016-04-26abs ↗pdf ↗

This paper considers the scenario that multiple data owners wish to apply a machine learning method over the combined dataset of all owners to obtain the best possible learning output but do not want to share the local datasets owing to privacy concerns. We design systems for the scenario that the stochastic gradient d…

2018-09-10abs ↗pdf ↗