A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Analyzes neural networks using linear models to understand their behavior.
problem Understanding multi-layer neural networks through linear models.
method Recalls and reviews four models: linear regression with concentrated features, kernel ridge regression, random feature model, and neural tangent model.
result Highlights limitations of linear theory and discusses approaches to overcome them.
Yu. I. Merzljakov developed a method of splittable coordinates which helps to verify the linearity of some groups, he established some fundamental results using this method. In this paper we use the method of splittable coordinates and find some sufficient condition under which the semi--direct product of two linear gr…
In this paper we give an example of a linear group such that its tensor square is not linear. Also, we formulate some sufficient conditions for the linearity of non-abelian tensor products G⊗H and tensor squares G⊗G. Using these results we prove that tensor squares of some groups with one relation a…
We prove that the semistability growth of hyperbolic groups is linear, which implies that hyperbolic groups which are sci (simply connected at infinity) have linear sci growth. Based on the linearity of the end-depth of finitely presented groups we show that the linear sci is preserved under amalgamated products over f…
In this paper we derive a differential identity for linearized gravity on the Kerr spacetime and more generally on vacuum spacetimes of Petrov type D. We show that a linear combination of second derivatives of the linearized Weyl tensor can be formed into a complex symmetric 2-tensor Mab which solves the…
Linear Transformer Block combines MLP and linear attention for near-optimal ICL in linear regression.
problem Achieving near-optimal in-context learning (ICL) risk for linear regression with a Gaussian prior.
method Combines linear attention and MLP components in a Linear Transformer Block (LTB). Establishes correspondence with one-step gradient descent estimators (GDext−β).
result LTB achieves nearly Bayes optimal ICL risk for linear regression with a Gaussian prior.
In this paper, we give complete classifications of linear ∞-harmonic maps between Euclidean and Heisenberg spaces, between Nil and Sol spaces. We also classify all ∞-harmonic linear endomorphisms of Sol space and show that there is a subgroup of ∞-harmonic linear automorphisms in the group of linea…
Region-specific linear models are widely used in practical applications because of their non-linear but highly interpretable model representations. One of the key challenges in their use is non-convexity in simultaneous optimization of regions and region-specific models. This paper proposes novel convex region-specific…
In a paper with Jean-Paul Dufour in 1999 \cite{DufourZung-Nambu1999}, we gave a classification of linear Nambu structures, and obtained linearization results for Nambu structures with a nondegenerate linear part. There was a case left open in \cite{DufourZung-Nambu1999}, namely the case of smooth linearization of Nambu…
In many European countries the growth of the real GDP per capita has been linear since 1950. An explanation for this linearity is still missing. We propose that in artificial intelligence we may find models for a linear growth of performance. We also discuss possible consequences of the fact that in systems with linear…
The parallel linear transports defined by flat linear connection are axiomatically described. On this basis a number of properties, some of which are new, of these transports and connections are derived.
Gaussian processes retain the linear model either as a special case, or in the limit. We show how this relationship can be exploited when the data are at least partially linear. However from the perspective of the Bayesian posterior, the Gaussian processes which encode the linear model either have probability of nearly…
We consider linear slices of the space of Kleinian once-punctured torus groups; a linear slice is obtained by fixing the value of the trace of one of the generators. The linear slice for trace 2 is called the Maskit slice. We will show that if traces converge `horocyclically' to 2 then associated linear slices converge…
Let F_n denote the free group generated by n letters. The purpose of this article is to show that Hol(F_2), the holomorph of the free group on two generators, is linear. Consequently, any split group extension of F_2 by a linear group H is linear. This result gives a large linear subgroup of Aut(F_3). A second applicat…
Study non-linear combinatorial bandits with polynomial rewards, finding significant differences from linear cases.
problem Adversarial combinatorial bandits with general non-linear reward functions.
method Extending existing work on adversarial linear combinatorial bandits, analyzing minimax optimal regret for polynomial and non-polynomial reward functions.
result Minimax optimal regret bounds for adversarial combinatorial bandits with general non-linear reward functions.
S. Bigelow proved that the braid groups are linear. That is, there is a faithful representation of the braid group into the general linear group of some field. Using this, we deduce from previously known results that the mapping class group of a sphere with punctures and hyperelliptic mapping class groups are linear. I…
We describe the ringed-space structure of moduli spaces of jets of linear connections (at a point) as orbit spaces of certain linear representations of the general linear group. Then, we use this fact to prove that the only (scalar) differential invariants associated to linear connections are constant functions, as wel…