We introduce a new weight-decay scaling rule to maintain sublayer gains across different widths in modern scale-invariant architectures.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Gated Linear Units (arXiv:1612.08083) consist of the component-wise product of two linear projections, one of which is first passed through a sigmoid function. Variations on GLU are possible, using different nonlinear (or even linear) functions in place of sigmoid. We test these variants in the feed-forward sublayers o…
We propose a novel approach to addressing the vanishing (or exploding) gradient problem in deep neural networks. We construct a new architecture for deep neural networks where all layers (except the output layer) of the network are a combination of rotation, permutation, diagonal, and activation sublayers which are all…
Transformers cluster meaningless words around leaders for sentiment analysis.
Randomly initialized transformers show extreme token preferences.
Introduces relative information gain for improving Gaussian process regression rates.
Decision trees algorithms use a gain function to select the best split during the tree's induction. This function is crucial to obtain trees with high predictive accuracy. Some gain functions can suffer from a bias when it compares splits of different arities. Quinlan proposed a gain ratio in C4.5's information gain fu…
The gain-loss ratio is known to enjoy very good properties from a normative point of view. As a confirmation, we show that the best market gain-loss ratio in the presence of a random endowment is an acceptability index and we provide its dual representation for general probability spaces. However, the gain-loss ratio w…
Eluder dimension and information gain are equivalent for reproducing kernel Hilbert spaces.
PC-GAIN improves GAIN's imputation by incorporating category information.
Investment horizon approach has been used to analyze indexes of Polish stock market.Optimal time horizon for each return value is evaluated by fitting appropriate function form of the distribution. Strong asymmetry of gain-loss curves is observed for WIG index, whereas gain and loss curves look similar for WIG20 and fo…
We study the problem of online path learning with non-additive gains, which is a central problem appearing in several applications, including ensemble structured prediction. We present new online algorithms for path learning with non-additive count-based gains for the three settings of full information, semi-bandit and…
Deep FPF approximates gain function for high-dimensional particle filtering.
Previous research has shown that for stock indices, the most likely time until a return of a particular size has been observed is longer for gains than for losses. We establish that this so-called gain/loss asymmetry is present also for individual stocks and show that the phenomenon is closely linked to the well-known …
New measures for causal entropy and information gain studied.
Study adds investment gains and losses to recursive utility model, proving existence and uniqueness of utility process.
New method improves feature selection in tree-based models.
Isobenefit Lines can offer a certain range of applicability in Location Theory and Gravitational Models for Urban and Geography Economics, in positional decision processes made by citizens, and, last but not least, in land value and property market theories and analysis. The value of a land, or a property, in a generic…
A new protocol evaluates small machine learning improvements conservatively.
REGAIN learns optimal auxiliary directions for forecast reconciliation.
We pursue an early stopping technique that helps Gaussian Restricted Boltzmann Machines (GRBMs) to gain good natural image representations in terms of overcompleteness and data fitting. GRBMs are widely considered as an unsuitable model for natural images because they gain non-overcomplete representations which include…
We demonstrate that the gain/loss asymmetry observed for stock indices vanishes if the temporal dependence structure is destroyed by scrambling the time series. We also show that an artificial index constructed by a simple average of a number of individual stocks display gain/loss asymmetry - this allows us to explicit…
Paper calculates greeks for DeFi LPs and introduces Impermanent Gain.
Proposes a new method to enhance neural learning by maximizing information gain.
We experimentally achieve a 19% capacity gain per Watt of electrical supply power in a 12-span link by eliminating gain flattening filters and optimizing launch powers using machine learning by deep neural networks in a massively parallel fiber context.
Generalizes optimal portfolio theory to include capital gains taxes.
ICYM2I corrects missingness bias in multimodal learning.
We develop a tractable model of realization utility that studies the role of reference-dependent S-shaped preferences in a dynamic investment setting with reinvestment. Our model generates both voluntarily realized gains and losses. It makes specific predictions about the volume of gains and losses, the holding periods…
Active inference selects actions to maximize information gain, aiding structure learning.
We consider a financial contract that delivers a single cash flow given by the terminal value of a cumulative gains process. The problem of modelling and pricing such an asset and associated derivatives is important, for example, in the determination of optimal insurance claims reserve policies, and in the pricing of r…
An econometric or statistical model may undergo a marginal gain if we admit a new variable to the model, and a marginal loss if we remove an existing variable from the model. Assuming equality of opportunity among all candidate variables, we derive a valuation framework by the expected marginal gain and marginal loss i…
Study shows gain-loss asymmetry in stock indices using a q-spin Potts model.
AdaS adapts SGD learning rate based on knowledge gain metrics.
Study optimal stopping problems with finite-time horizon and proves continuity and strict monotonicity of the boundary.
Perfect tracking control for real-world Euler-Lagrange systems is challenging due to uncertainties in the system model and external disturbances. The magnitude of the tracking error can be reduced either by increasing the feedback gains or improving the model of the system. The latter is clearly preferable as it allows…
New federated learning framework reduces model complexity and improves performance.
Study calculates arbitrage gains between two markets with limited liquidity.
We note a simple mechanism that may at least partially resolve several outstanding economic puzzles, including why the cyclically adjusted price to earnings ratio of the S&P 500 index has been oddly high for the past two decades, why gains to capital have outpaced gains to wages, and the persistence of the equity premi…
Autoencoder-based geometric shaping is proposed that includes optimizing bit mappings. Up to 0.2 bits/QAM symbol gain in GMI is achieved for a variety of data rates and in the presence of transceiver impairments. The gains can be harvested with standard binary FEC at no cost w.r.t. conventional BICM.
Active sampling algorithm improves accuracy of inferred scores from pairwise comparisons.
Upper bound derived for informed traders' gains in a model, akin to thermodynamics.
DO-IQS recovers optimal stopping region from expert trajectories, addressing specific challenges.
Unified framework for portfolio optimization using gain PDF.
New method improves robustness of Bayesian experimental design.
Theory of symmetric rigidity in hyperbolic geometry.
Bayesian analysis reveals asymmetry in financial data.
Generalized statistical arbitrage concepts are introduced corresponding to trading strategies which yield positive gains on average in a class of scenarios rather than almost surely. The relevant scenarios or market states are specified via an information system given by a -algebra and so this notion contains classi…
Proposes EPIG for active learning to improve predictive performance.