Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

61121182242 · Jun 202019922001200920172026
48 results for algebraic priors

We algorithmically construct multi-output Gaussian process priors which satisfy linear differential equations. Our approach attempts to parametrize all solutions of the equations using Gröbner bases. If successful, a push forward Gaussian process along the paramerization is the desired prior. We consider several exampl…

2018-01-28abs ↗pdf ↗

Open problem: Establishing bounds for Cayley-table completion to discover discrete algorithmic axioms.

problem Discovering discrete algorithmic axioms missing in deep learning.
method Cayley-table completion as a testbed for algorithmic complexity minimization.
result Formal exact recovery bounds for Cayley-table completion.

When solving data analysis problems it is important to integrate prior knowledge and/or structural invariances. This paper contributes by a novel framework for incorporating algebraic invariance structure into kernels. In particular, we show that algebraic properties such as sign symmetries in data, phase independence,…

2014-11-28abs ↗pdf ↗

A new method uses algebraic insights to create approximately equivariant networks without complex architectures.

problem Designing equivariant neural networks with complex architectures and high computational cost.
method Imposes the group's regular representation as an inductive bias via an auxiliary loss, adding no learnable parameters.
result Matches or outperforms specialized models in several cases, even for infinite groups.

GS-B3^3SE improves label shift estimation by smoothing priors on a graph.

problem Label shift adaptation when source and target distributions share conditional but not marginal probabilities.
method Graph-Smoothed Bayesian Black-Box Shift Estimator (GS-B3^3SE) places Laplacian-Gaussian priors on log-priors and confusion-matrix columns tied by a label-similarity graph.
result GS-B3^3SE produces a tractable posterior with HMC or Newton-CG schemes, proving identifiability, contraction, and robustness.

Delta-unlinking number measures how to unlink algebraically split links.

problem Measuring unlinking complexity of algebraically split links.
method Defining delta-unlinking number as minimum delta-moves to unlink, proving bounds and calculating specific values.
result Precise delta-unlinking numbers for algebraically split prime links up to 9 crossings, and 4-genus values for most.

Geometrically classifies maps from R^0|2 to any manifold, unifying theories.

problem Classifying maps from R^0|2 to any manifold without auxiliary structures.
method Relates maps to pullback of decomposable bivector bundle over S via algebraic constraints.
result Reduced manifold has fiber dimension dim(S) + 1, unifying topological and algebraic views.

This work introduces a method for almost equivariance in neural networks using Lie algebra convolutions.

problem Real-world data often does not conform to strict group equivariances, leading to underperformance in models.
method Definition and practical implementation of almost equivariance through Lie algebra convolutions.
result Demonstrated the validity of the approach through benchmarking against fully equivariant settings.

We characterize value functions in partially observable MDPs as semi-algebraic sets.

problem Understanding feasible value functions in partially observable Markov decision processes.
method Characterization of feasible value functions as semi-algebraic sets defined by polynomial inequalities.
result The feasible set of value functions in POMDPs is a semi-algebraic set, not a polytope as in MDPs.

Dynamic paired comparison models, such as Elo and Glicko, are frequently used for sports prediction and ranking players or teams. We present an alternative dynamic paired comparison model which uses a Gaussian Process (GP) as a prior for the time dynamics rather than the Markovian dynamics usually assumed. In addition,…

2019-02-20abs ↗pdf ↗

Proves constant scalar curvature Kähler metrics are very general.

problem Existence of constant scalar curvature Kähler metrics on smooth polarized varieties.
method Combining uniform arc K-stability and algebraic properties in families.
result The constant scalar curvature Kähler locus is very general.

This paper concerns cluster algebras with principal coefficients A(S,M) associated to bordered surfaces (S,M), and is a companion to a concurrent work of the authors with Schiffler [MSW2]. Given any (generalized) arc or loop in the surface -- with or without self-intersections -- we associate an element of (the fractio…

2011-08-17abs ↗pdf ↗

Paper improves AIRL by enhancing policy imitation and addressing reward recovery issues.

problem Inadequate policy imitation and limited transferable reward recovery in AIRL.
method Substituted built-in algorithm with SAC for policy updating and proposed PPO-AIRL + SAC hybrid framework.
result SAC improves policy imitation but hinders reward recovery; PPO-AIRL + SAC achieves satisfactory transfer effect.

A new method discovers equations from data using Bayesian and kernel techniques.

problem Discovering equations from data is hard due to sparsity and noise.
method Kernel regression for function estimation and Bayesian spike-and-slab prior for uncertainty quantification.
result KBASS method outperforms state-of-the-art methods on benchmark tasks.

AlgebraNets use alternative algebras for neural networks, improving performance on image and language tasks.

problem Improving neural network performance on large-scale image and language tasks.
method Considered alternative algebras (C, H, M2(R), M2(C), M3(R), M4(R)) for activations and weights, and studied their performance on ImageNet and enwiki8 datasets.
result Alternative algebras deliver better parameter and computational efficiency compared with real numbers, especially in sparse and auto-regressive inference scenarios.

We provide a general construction of integral TQFTs over a general commutative ring, k\mathbf{k}, starting from a finite Hopf algebra over k\mathbf{k} which is Frobenius and double balanced. These TQFTs specialize to the Hennings invariants of the respective doubles on closed 3-manifolds. We show the construction app…

2013-05-31abs ↗pdf ↗

We introduce a Bayesian Gaussian process latent variable model that explicitly captures spatial correlations in data using a parameterized spatial kernel and leveraging structure-exploiting algebra on the model covariance matrices for computational tractability. Inference is made tractable through a collapsed variation…

2018-05-22abs ↗pdf ↗

Graded Transformers embed algebraic structure in neural networks through graded transformations.

problem Efficiently modeling hierarchical and structured data in neural networks.
method Introduces Linearly Graded Transformer (LGT) and Exponentially Graded Transformer (EGT) with graded scaling operators.
result Establishes rigorous guarantees and improved efficiency for structured data.

Gaussian processes are rich distributions over functions, with generalization properties determined by a kernel function. When used for long-range extrapolation, predictions are particularly sensitive to the choice of kernel parameters. It is therefore critical to account for kernel uncertainty in our predictive distri…

2018-02-02abs ↗pdf ↗

Our main result is that for all sufficiently large x0>0x_0>0, the set of commensurability classes of arithmetic hyperbolic 2- or 3-orbifolds with fixed invariant trace field kk and systole bounded below by x0x_0 has density one within the set of all commensurability classes of arithmetic hyperbolic 2- or 3-orbifolds wit…

2015-04-20abs ↗pdf ↗

We introduce new obstructions to topological knot concordance. These are obtained from amenable groups in Strebel's class, possibly with torsion, using a recently suggested L2L^2-theoretic method due to Orr and the author. Concerning (h)(h)-solvable knots which are defined in terms of certain Whitney towers of height $h…

2010-10-06abs ↗pdf ↗

This paper continues our exploration of homology cobordism of 3-manifolds using our recent results on Cheeger-Gromov rho-invariants associated to amenable representations. We introduce a new type of torsion in 3-manifold groups we call hidden torsion, and an algebraic approximation we call local hidden torsion. We cons…

2011-01-21abs ↗pdf ↗

Continues work on derived manifolds and symplectic schemes, constructing virtual classes.

problem Constructing virtual fundamental classes for derived manifolds and schemes.
method Cosection localization, reduced virtual fundamental classes, and applications to Donaldson-Thomas theory.
result Virtual fundamental classes for (2)(-2)-shifted symplectic derived schemes are consistent with algebraic and differential geometric constructions.

New framework for probabilistic linear solvers reduces manual effort.

problem Manual implementation of probabilistic iterative methods is laborious.
method Affine Tracing: Automatically constructs PIMs from standard implementations.
result Any realistic affine PIM is calibrated, motivating their adoption.

A new Weyl prior is proposed for Bayesian statistics, offering a more canonical choice for parameter α.

problem Choosing a prior distribution for Bayesian inference.
method Proposed a new Weyl prior based on the Weyl structure on a statistical manifold.
result The Weyl prior is a special case of the α-parallel prior with α = -n, where n is the dimension of the statistical manifold.

New kernel interprets 3D anisotropic data with rotations and improved predictions.

problem Capturing rotated anisotropy in 3D spatial fields.
method Introduces a Lie-algebraic kernel with three principal length-scales and an explicit rotation.
result Posterior recovers rotated anisotropy and improves prediction over axis-aligned kernels.

Informative Bayesian priors are often difficult to elicit, and when this is the case, modelers usually turn to noninformative or objective priors. However, objective priors such as the Jeffreys and reference priors are not tractable to derive for many models of interest. We address this issue by proposing techniques fo…

2017-04-04abs ↗pdf ↗

While Bayesian methods are praised for their ability to incorporate useful prior knowledge, in practice, convenient priors that allow for computationally cheap or tractable inference are commonly used. In this paper, we investigate the following question: for a given model, is it possible to compute an inference result…

2016-06-02abs ↗pdf ↗

We introduce an approach for imposing physically motivated inductive biases on graph networks to learn interpretable representations and improved zero-shot generalization. Our experiments show that our graph network models, which implement this inductive bias, can learn message representations equivalent to the true fo…

2019-09-12abs ↗pdf ↗

Semantic word embeddings represent the meaning of a word via a vector, and are created by diverse methods. Many use nonlinear operations on co-occurrence statistics, and have hand-tuned hyperparameters and reweighting methods. This paper proposes a new generative model, a dynamic version of the log-linear topic model o…

2015-02-12abs ↗pdf ↗

Researchers derive exact priors for finite Bayesian neural networks.

problem Understanding non-Gaussian priors in finite Bayesian neural networks.
method Analytical derivation of function space priors for finite fully-connected feedforward networks.
result Exact solutions for priors of finite networks, including Meijer G-function for linear networks and mixtures for ReLU networks.

PRCD-MAP learns to trust imperfect priors in causal discovery, improving accuracy and robustness.

problem Tackles the brittle trade-off between blind trust and rejection of external priors in causal discovery.
method Proposes PRCD-MAP, a soft prior-consumption layer that assigns per-edge trust to imperfect priors and modulates regularization in a MAP objective.
result Enjoys a population-level safety guarantee and outperforms existing methods on real-world causal discovery tasks.

Bayesian metalearning improves performance in linear bandits with misspecified priors.

problem Improper priors lead to suboptimal performance in sequential decision-making.
method Proves performance bounds for metalearning priors in stochastic linear bandits and develops a metalearning algorithm.
result Metalearning can improve performance by learning the prior from multiple tasks.

The paper extends and applies a new shrinkage prior in Bayesian factor analysis.

problem Estimating the number of factors in sparse Bayesian factor analysis.
method Introduces and extends a generalized cumulative shrinkage process (CUSP) prior.
result Exchangeable spike-and-slab shrinkage priors imply increasing shrinkage as the column index increases.

Bayesian method corrects for model selection multiplicity in regression.

problem Model selection multiplicity in regression analysis.
method Developed a Bayesian prior distribution based on Holm procedure analogy.
result Adequate multiplicity correction requires sparsity not provided by recommended priors.