Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

4.2%8.3%12.5%16.7% · Apr 199519922001200920182026
48 results for Non-Differentiable Constraints

New method solves non-convex constrained optimization problems with non-differentiable constraints.

problem Training non-convex models with non-differentiable constraints.
method Proxy-Lagrangian formulation and semi-coarse correlated equilibrium.
result Solves non-convex constrained optimization problems with theoretical guarantees.

VaR-CPO optimizes VaR-constrained RL problems with conservative policy updates.

problem Optimizing VaR-constrained reinforcement learning problems.
method Combines Cantelli's inequality and trust-region framework for efficient and conservative optimization.
result Achieves zero constraint violations during training in feasible environments.

In recent years, constrained optimization has become increasingly relevant to the machine learning community, with applications including Neyman-Pearson classification, robust optimization, and fair machine learning. A natural approach to constrained optimization is to optimize the Lagrangian, but this is not guarantee…

2018-04-17abs ↗pdf ↗

Tutorial on combining latent variable models with deep learning.

problem Combining latent variable models with deep learning to model natural language.
method Exploring variational inference to address intractable posterior inference and non-differentiability issues.
result Exploration of variational inference techniques to handle deep latent variable models.

Paper proves autodiff systems are correct for non-differentiable functions.

problem Correctness of autodiff systems for non-differentiable functions in deep learning.
method Investigation of PAP functions and introduction of intensional derivatives.
result Intensional derivatives always exist and coincide with standard derivatives for almost all inputs.

New method relaxes PCA orthogonality constraints using explained variance of correlated components.

problem Difficulty in using PCA for sparse design due to orthogonality constraints and non-differentiable penalty.
method Introduce expvar(Y) to measure variance explained by correlated components, relax orthogonality constraints.
result Two expvar(Y) definitions suitable for block PCA formulations without orthogonality constraints.

We generalize stochastic smoothing for gradient estimation of non-differentiable functions.

problem Gradient estimation for non-differentiable functions.
method Developed a general framework for relaxation and gradient estimation of non-differentiable black-box functions using stochastic smoothing with reduced assumptions.
result Empirically validated the effectiveness of variance reduction strategies for various non-differentiable tasks.

The study connects Hilbert entropy to non-differentiability points of limit sets in flag spaces.

problem Understanding non-differentiability points in limit sets of convex projective structures.
method Introduces hyperplane conicality for θθ-Anosov representations and uses it to prove properties of boundary maps.
result Hilbert entropy is linked to the Hausdorff dimension of non-differentiability points in flag spaces.

ES for non-differentiable parameters scales to large models.

problem Learning non-differentiable parameters in large models.
method Hybrid approach combining ES for non-differentiable and gradient-based methods for differentiable parameters.
result Hybrid approach is competitive and allows training sparse models from the start.

Algorithm finds optimal investment strategies for non-differentiable preferences.

problem Optimal investment strategies under non-differentiable preferences.
method Reduces problem to a discrete grid, uses efficient method to find strategies.
result Optimal strategies lie on a discrete grid, allowing efficient computation.

We present a new algorithm for stochastic variational inference that targets at models with non-differentiable densities. One of the key challenges in stochastic variational inference is to come up with a low-variance estimator of the gradient of a variational objective. We tackle the challenge by generalizing the repa…

2018-06-01abs ↗pdf ↗

Study shows AD for neural nets with machine-representable numbers can be incorrect.

problem Correctness of AD for neural nets with machine-representable numbers.
method Analyzed two sets of parameters: incorrect and non-differentiable. Proved bounds and conditions for AD correctness.
result AD can be incorrect for machine-representable numbers, but provides a Clarke subderivative on non-differentiable set.

Unified approach for sampling non-differentiable and heavy-tailed targets.

problem Sampling non-differentiable and heavy-tailed distributions using Langevin algorithms.
method Anchored Langevin dynamics, which modifies the Langevin diffusion with a smooth reference potential and multiplicative scaling.
result Non-asymptotic guarantees in the 2-Wasserstein distance to the target distribution.

Study improves understanding of non-differentiable penalties in high-dimensional settings.

problem Theoretical understanding of non-differentiable penalties like generalized LASSO and nuclear norm in high-dimensional settings.
method Proportional high-dimensional regime analysis with finite sample upper bounds on expected squared error.
result LO provides accurate estimation of out-of-sample risk in high-dimensional settings.

Study identifies conditions for proxy adjustment in confounded binary treatment outcomes.

problem Average causal effect estimation with a non-differentially mismeasured binary confounder.
method Identifies conditions for proxy adjustment in the presence of a non-differentially mismeasured binary confounder.
result Adjusting for a non-differentially mismeasured binary proxy can improve estimation of the average causal effect.

Develops LF-PPL for non-differentiable models with automatic boundary checks.

problem Handling non-differentiable models in probabilistic programming.
method Introduces LF-PPL with automatic boundary checks and a formalism ensuring measure zero discontinuities.
result Demonstrates efficient inference for non-differentiable models using DHMC.

Differentiable pipeline replaces non-differentiable CAE components for shape optimization.

problem Gradient-based optimization is limited by non-differentiable components in CAE workflows.
method Surrogate models replace non-differentiable pipeline components, enabling gradient-based optimization.
result Gradient-based shape optimization possible without differentiable solvers.

This paper is an attempt at understanding the quantum-like dynamics of financial markets in terms of non-differentiable price-time continuum having fractal properties. The main steps of this development are the statistical scaling, the non-differentiability hypothesis, and the equations of motion entailed by this hypot…

2013-12-11abs ↗pdf ↗

Smooth Contextual Bandits bridge two previously studied extremes of non-differentiable and parametric-response bandits.

problem Nonparametric contextual bandits with Hölder smoothness.
method Developed a novel algorithm that optimally balances between non-differentiable and parametric-response bandits.
result Proved the algorithm achieves rate-optimal regret for all smoothness settings.

SoDeep learns approximations of ranking metrics for deep learning tasks.

problem Non-differentiable metrics in machine learning tasks.
method Sorting deep (SoDeep) net trained to approximate sorting of scores.
result Competitive results on Cross-modal text-image retrieval, multi-label image classification, and visual memorability ranking tasks.

A new fuzzy clustering method using hyperbolic smoothing for large datasets.

problem Building fuzzy clusters for large data sets efficiently.
method A novel smoothing numerical approach to relax the sum-of-squares criterion, converting the problem into a differentiable optimization problem.
result The method produces better fuzzy partitions compared to traditional fuzzy CC-means.

Paper generalizes Hardy-Rogers maps for market equilibrium analysis in duopoly markets.

problem Existence and uniqueness of market equilibrium in duopoly markets with non-differentiable, nonlinear response functions.
method Coupled fixed points approach for generalized Hardy-Rogers maps.
result Enriched understanding of market equilibrium in duopoly markets with non-differentiable response functions.

Proposes a method for inference in high-dimensional classification with non-differentiable surrogate losses.

problem Lack of inference procedures for identifying driving factors in high-dimensional classification with non-differentiable surrogate losses.
method Kernel-smoothed decorrelated score and cross-fitted version for hypothesis tests and interval estimators.
result Valid and superior inference methods for high-dimensional classification with non-differentiable surrogate losses.

Complex computer simulators are increasingly used across fields of science as generative models tying parameters of an underlying theory to experimental observations. Inference in this setup is often difficult, as simulators rarely admit a tractable density or likelihood function. We introduce Adversarial Variational O…

2017-07-22abs ↗pdf ↗

This paper extends geometric study of neural networks to non-differentiable layers and random walks.

problem Understanding the geometric properties of neural networks, especially those with non-differentiable activation functions.
method Singular Riemannian geometry approach to convolutional, residual, and recursive neural networks.
result Illustrated geometric findings with numerical experiments on image classification and thermodynamic problems.

New methods approximate LOOCV for high-dimensional, non-differentiable learning problems.

problem Finding optimal regularization parameters in high-dimensional learning problems.
method Three frameworks based on primal, dual, and proximal formulations of a convex optimization problem.
result Equivalence of three methods under smoothness conditions, validated by empirical results.

New stochastic algorithms solve DC functions and non-convex problems efficiently.

problem Solving non-convex, non-smooth, and non-differentiable functions efficiently.
method Proposed new stochastic optimization algorithms for DC functions and non-convex problems.
result First non-asymptotic convergence for non-convex optimization with general non-convex non-differentiable regularizers.

Study calculates slope gaps on polygon surfaces, finding non-unimodal distributions.

problem Understanding the distribution of slope gaps on polygon surfaces.
method Explicit computation of slope gap distributions for 2n-gons, providing bounds on non-differentiability points.
result Slope gap distributions are not always unimodal, answering a question by Athreya.

Bayesian optimization uses triangulation candidates for better performance.

problem Non-convex and multi-modal optimization challenges in Bayesian optimization.
method Proposes using Delaunay triangulation candidates for discrete search over continuous optimization.
result Triangulation candidates outperform numerically optimized and random alternatives.

This paper addresses the scalability challenge of architecture search by formulating the task in a differentiable manner. Unlike conventional approaches of applying evolution or reinforcement learning over a discrete and non-differentiable search space, our method is based on the continuous relaxation of the architectu…

2018-06-24abs ↗pdf ↗

Extends batch active learning to non-differentiable models.

problem Efficiently training machine learning models on large, initially unlabelled datasets.
method Black-box batch active learning for regression tasks that relies solely on model predictions.
result Achieves strong performance on regression datasets compared to white-box approaches for deep learning models.

New machine learning method uses algorithmic complexity for non-differentiable spaces.

problem Machine learning on non-differentiable spaces.
method Introduces complexity theory in machine learning, using algorithmic complexity for regression and classification.
result More generalizable and resilient to random attacks compared to traditional methods.

Proposes a differentiable hypergeometric distribution for learning group importance.

problem Learning the sizes of subsets in applications like clustering and weakly-supervised learning.
method Introduces a reparameterizable hypergeometric distribution to model group sizes and learn their relative importance.
result Outperforms previous methods in weakly-supervised learning and clustering.

The paper improves ALO for 1\ell_1-regularized models.

problem Estimating out-of-sample error for 1\ell_1-regularized models.
method Developed a novel theory for 1\ell_1-regularized problems, bounding ALO error.
result For 1\ell_1-regularized problems, ALO error goes to zero as p goes to infinity.

Neural painters learn to generate brushstrokes from a non-deterministic painting program.

problem Training an agent to generate realistic brushstrokes from a non-differentiable painting program.
method A differentiable neural painter model trained on brushstrokes, optimizing for human-like strokes and intrinsic style transfer.
result Direct optimization of brushstrokes can visualize ImageNet categories and generate ideal paintings.

Improves discrete latent representations using differentiable approximation bridges.

problem Improving discrete latent representations in neural networks.
method Training with a differentiable approximation bridge (DAB) neural network.
result Improves state-of-the-art performance in various domains.

Paper proposes FONE for efficient distributed estimation and inference.

problem Efficient distributed estimation and inference for non-differentiable convex losses.
method Proposes a multi-round distributed estimation procedure using a First-Order Newton-type Estimator (FONE).
result FONE efficiently estimates Σ1wΣ^{-1} w for non-differentiable losses, facilitating inference.

New algorithm trains communication systems without a differentiable channel model.

problem Training communication systems with unknown or non-differentiable channel models.
method Iterative training between receiver and transmitter using true and approximated gradients.
result Works as well as model-based training and achieves state-of-the-art performance.

Local LMO optimizes constrained problems using local linear minimization.

problem Constrained optimization problems with complex feasible sets.
method Designs a new projection-free gradient method using local linear minimization.
result Transfers convergence rates of Projected Gradient Descent to the projection-free world.