Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

103206309412 · Jun 202019922001200920172026
48 results for Absolute Value activations

Neural networks learn distance-based representations, not just intensity.

problem Understanding how neural networks interpret and learn from internal activations.
method Manipulated ReLU and Absolute Value activations to observe sensitivity to distance and intensity perturbations.
result Neural networks are highly sensitive to small distance-based perturbations, challenging the intensity-based interpretation.

The paper introduces Absolute Shapley Value to handle negative contributions in machine learning model training.

problem Negative marginal contributions in machine learning model training.
method Investigates three philosophies: Original Shapley Value, Zero Shapley Value, and Absolute Shapley Value.
result Absolute Shapley Value significantly outperforms other definitions in evaluating data importance.

Neural decoders were shown to outperform classical message passing techniques for short BCH codes. In this work, we extend these results to much larger families of algebraic block codes, by performing message passing with graph neural networks. The parameters of the sub-network at each variable-node in the Tanner graph…

2019-09-05abs ↗pdf ↗

A complex-valued convolutional network (convnet) implements the repeated application of the following composition of three operations, recursively applying the composition to an input vector of nonnegative real numbers: (1) convolution with complex-valued vectors followed by (2) taking the absolute value of every entry…

2015-03-11abs ↗pdf ↗

In this manuscript we analyse the leading statistical properties of fluctuations of (log) 3-month US Treasury bill quotation in the secondary market, namely: probability density function, autocorrelation, absolute values autocorrelation, and absolute values persistency. We verify that this financial instrument, in spit…

2007-06-08abs ↗pdf ↗

Optimal trading patterns adjust based on market efficiency and slippage costs.

problem Balancing active alphas and trading costs in active portfolios.
method Maximization of utility including projected alpha-based profits, slippage costs, and risk aversion.
result Optimal trading involves a no-trade zone width that scales as Δc1/2Δ\sim c^{1/2}, differing from stochastic settings.

Deep neural nets on 1-D data are convex Lasso models with reflection features.

problem Training neural networks on 1-D data.
method Proving equivalence to convex Lasso problems with discrete, explicitly defined dictionary matrices.
result Reflection features in neural networks with certain activations.

Identification of regions of interest (ROI) associated with certain disease has a great impact on public health. Imposing sparsity of pixel values and extracting active regions simultaneously greatly complicate the image analysis. We address these challenges by introducing a novel region-selection penalty in the framew…

2016-05-27abs ↗pdf ↗

We classify isotopy classes of irreducible Heegaard splittings of solvmanifolds. If the monodromy of the solvmanifold can be expressed as a 2 x 2 matrix with 0 in the lower right hand corner (as always is true when the absolute value of the trace is 3), then any irreducible splitting is strongly irreducible and of genu…

1998-03-31abs ↗pdf ↗

The folk result in Kyle-Back models states that the value function of the insider remains unchanged when her admissible strategies are restricted to absolutely continuous ones. In this paper we show that, for a large class of pricing rules used in current literature, the value function of the insider can be finite when…

2018-12-18abs ↗pdf ↗

We present a nonlinear stochastic differential equation (SDE) which mimics the probability density function (PDF) of the return and the power spectrum of the absolute return in financial markets. Absolute return as a measure of market volatility is considered in the proposed model as a long-range memory stochastic vari…

2009-01-07abs ↗pdf ↗

Monotone aggregation of dependent random vectors has an absolutely continuous distribution under certain conditions.

problem Monotone aggregation of dependent random vectors
method Coordinatewise monotonicity and uniform lower-increment conditions
result One-dimensional push-forwards of dependent random vectors have an absolutely continuous distribution

We consider the problem of portfolio selection within the classical Markowitz mean-variance framework, reformulated as a constrained least-squares regression problem. We propose to add to the objective function a penalty proportional to the sum of the absolute values of the portfolio weights. This penalty regularizes (…

2007-07-31abs ↗pdf ↗

We study principal curvatures of fibers and Heegaard surfaces smoothly embedded in hyperbolic 3-manifolds. It is well known that a fiber or a Heegaard surface in a hyperbolic 3-manifold cannot have principal curvatures everywhere less than one in absolute value. We show that given an upper bound on the genus of a minim…

2010-02-04abs ↗pdf ↗

In this work, we propose an order book model with herd behavior. The proposed model is built upon two distinct approaches: a recent empirical study of the detailed order book records by Kanazawa et al. [Phys. Rev. Lett. 120, 138301] and financial herd behavior model. Combining these approaches allows us to propose a mo…

2018-09-08abs ↗pdf ↗

Previous literature on unsupervised learning focused on designing structural priors with the aim of learning meaningful features. However, this was done without considering the description length of the learned representations which is a direct and unbiased measure of the model complexity. In this paper, first we intro…

2019-07-12abs ↗pdf ↗

Paper studies geometric properties of nonlinear Lebesgue spaces.

problem Geometric and analytic properties of nonlinear Lebesgue spaces.
method Formalizes pointwise description of geometric properties using a nonlinear Fubini-Lebesgue theorem.
result Definition of length structure, Alexandrov curvature bounds, and speed for absolutely continuous curves in nonlinear Lebesgue spaces.

We prove uniqueness results for a Calderon type inverse problem for the Hodge Laplacian acting on graded forms on certain manifolds in three dimensions. In particular, we show that partial measurements of the relative-to-absolute or absolute-to-relative boundary value maps uniquely determine a zeroth order potential. T…

2013-10-17abs ↗pdf ↗

Deep neural networks favor symmetric structures, enabling multilevel symmetries.

problem Understanding and optimizing deep neural networks.
method Formulating DNN training as convex Lasso problems with geometric algebra.
result Deep networks inherently favor symmetric structures, enabling multilevel symmetries.

Optimal stock price prediction model using recurrent neural networks with RMSprop optimizer.

problem Stock price prediction using neural networks.
method Comparison of fully connected, convolutional, and recurrent architectures; inclusion of three optimization techniques.
result Single layer recurrent neural network with RMSprop optimizer produces optimal results with validation and test MAE of 0.0150 and 0.0148 respectively.

Active learning improves ordering of items with contextual attributes.

problem Learning accurate item orderings from pairwise comparisons, especially when exhaustive comparisons are impractical.
method Proposes an active learning strategy that samples items to minimize expected ordering error, accounting for uncertainty in comparisons.
result Superior sample efficiency and generalization compared to non-contextual ranking approaches and active preference learning baselines.

In this paper, we consider the problem of recovering a sparse signal based on penalized least squares formulations. We develop a novel algorithm of primal-dual active set type for a class of nonconvex sparsity-promoting penalties, including 0\ell^0, bridge, smoothly clipped absolute deviation, capped 1\ell^1 and mini…

2013-10-04abs ↗pdf ↗

The visibility transformation embeds data position into signature features for efficient pattern recognition.

problem Embedding absolute position into signature features for efficient pattern recognition.
method The visibility transformation is put on a theoretical footing and used to embed absolute position into signature features efficiently.
result The generated feature set simplifies pattern recognition by accommodating nonlinear functions of absolute and relative values.

Sharp bounds on neural network approximation rates and widths.

problem Estimating approximation rates, metric entropy, and n-widths of shallow neural networks.
method Introducing smoothly parameterized dictionaries and providing upper and lower bounds.
result Sharp bounds on approximation rates, metric entropy, and n-widths for neural networks with various activation functions.

The paper proposes methods to find a shared active subspace for multivariate vector-valued functions.

problem Minimizing the deviation between function evaluations in the original and reconstructed spaces.
method Manipulating gradients or SPD matrices to identify a shared structure.
result Summing SPD matrices often identifies the best shared active subspace.

We present an efficient algorithm for calculating the number of components of an integral lamination on an nn-punctured disk, given its Dynnikov coordinates. The algorithm requires O(n2M)O(n^2M) arithmetic operations, where MM is the sum of the absolute values of the Dynnikov coordinates.

2015-12-28abs ↗pdf ↗

Model predicts cannabis use disorder risk for adolescents and young adults.

problem Predicting cannabis use disorder progression in adolescents and young adults.
method Bayesian machine learning model trained on longitudinal data.
result Model provides personalized risk assessment with AUC of 0.68-0.75 and E/O ratio of 0.95-1.

An interesting approach to analyzing neural networks that has received renewed attention is to examine the equivalent kernel of the neural network. This is based on the fact that a fully connected feedforward network with one hidden layer, a certain weight distribution, an activation function, and an infinite number of…

2017-11-24abs ↗pdf ↗

Recurrent neural networks (RNNs) are notoriously difficult to train. When the eigenvalues of the hidden to hidden weight matrix deviate from absolute value 1, optimization becomes difficult due to the well studied issue of vanishing and exploding gradients, especially when trying to learn long-term dependencies. To cir…

2015-11-20abs ↗pdf ↗

Establishes a link between risk measures and uniform integrability in finance.

problem Understanding uniform integrability in the context of financial risk measures.
method Introduces the folding score of distortion risk measures to study uniform integrability directly with gains and losses.
result Obtains three sets of equivalent conditions for uniform integrability involving coherent risk measures.

Complex-valued neural networks can approximate any continuous function with bounded widths and depths.

problem Approximating continuous functions with complex-valued neural networks of bounded widths and depths.
method Analyzing activation functions and proving universality for complex-valued networks.
result Deep narrow complex-valued networks are universal if and only if their activation function is neither holomorphic, nor antiholomorphic, nor R\mathbb{R}-affine.

S2D selectively decays large singular values to improve quantization of neural activations.

problem Large activation outliers in transformer models cause accuracy drops during quantization.
method Selective Spectral Decay (S2DS^2D) that surgically regularizes only the largest singular values.
result Significantly reduces activation outliers and produces well-conditioned representations.