Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

3517021,0531,404 · Jun 202019922001200920172026
48 results for Predecessor Models

The paper addresses Dyna-style RL's value hallucination issue by proposing a new algorithm.

problem Value hallucination in Dyna-style RL due to bootstrapping simulated states.
method Introduces a new Dyna algorithm using predecessor models with multi-step updates.
result Evidence supports the Hallucinated Value Hypothesis (HVH), suggesting predecessor models with multi-step updates are promising.

This work explores a social learning problem with agents having nonidentical noise variances and mismatched beliefs. We consider an NN-agent binary hypothesis test in which each agent sequentially makes a decision based not only on a private observation, but also on preceding agents' decisions. In addition, the agents…

2018-11-23abs ↗pdf ↗

We propose Generative Predecessor Models for Imitation Learning (GPRIL), a novel imitation learning algorithm that matches the state-action distribution to the distribution observed in expert demonstrations, using generative models to reason probabilistically about alternative histories of demonstrated states. We show …

2019-04-01abs ↗pdf ↗

This paper shows that only finitely many knots can be ribbon concordant to any given knot.

problem Understanding the relationship between ribbon concordance and fibered knots.
method Using knot Floer homology and bordered Heegaard Floer homology, the paper proves an inequality relating the knot Floer homology of a satellite knot and its companion.
result Every knot in S3S^3 has only finitely many fibered predecessors under ribbon concordance.

These are lecture notes on the rigidity of submanifolds of projective space "resembling" compact Hermitian symmetric spaces in their homogeneous embeddings. Recent results are surveyed, along with their classical predecessors. The notes include an introduction to moving frames in projective geometry, an exposition of t…

2006-09-18abs ↗pdf ↗

This paper provides a new version of the condition of Di Nunno et al. (2003), Ankirchner and Imkeller (2005) and Biagini and \{O}ksendal (2005) ensuring the semimartingale property for a large class of continuous stochastic processes. Unlike our predecessors, we base our modeling framework on the concept of portfolio p…

2007-06-04abs ↗pdf ↗

This work concludes a series of four papers on the foundational theory of orbifolds and stacks. We apply the abstract theory, developed in its predecessors, to orbifolds derived from manifolds. Specifically, we show how the very concrete topological base spaces associated to such orbifolds can be described and manipula…

1995-03-14abs ↗pdf ↗

Bidirectional attention is shown to be equivalent to a continuous bag of words model with mixture-of-experts.

problem Understanding the statistical underpinnings of bidirectional attention.
method Exploring bidirectional attention as a mixture-of-experts model and reparameterizing it.
result Bidirectional attention can be viewed as a continuous bag of words model with mixture-of-experts weights.

We present new, unified proofs for the cell-like, Z/p\mathbb{Z}/p-, and Q\mathbb{Q}-resolution theorems. Our arguments employ extensions that are much simpler then those used by our predecessors. The techniques allow us to solve problems involving cohomology groups by converting them into problems about homology groups…

2017-06-05abs ↗pdf ↗

We consider the problem of optimising functions in the reproducing kernel Hilbert space (RKHS) of a Matérn kernel with smoothness parameter νν over the domain [0,1]d[0,1]^d under noisy bandit feedback. Our contribution, the ππ-GP-UCB algorithm, is the first practical approach with guaranteed sublinear regret for all $ν>1…

2020-01-28abs ↗pdf ↗

Numerical computations suggest that each point on a certain optimized shape called the ideal trefoil is in contact with two other points. We consider sequences of such contact points, such that each point is in contact with its predecessor and call it a billiard. Our numerics suggest that a particular billiard on the i…

2011-03-18abs ↗pdf ↗

This paper extends geometric structure theory to infinite type structures.

problem Calculating characteristic class relations in complex Cartan geometries.
method Improves representation theory for infinite type structures.
result Direct calculation of characteristic class relations from structure group representation.

Transformers can perform well with less long-range memory.

problem The need for deep long-range memory in Transformers for language modeling.
method Performed interventions to show performance can be achieved with fewer long-range memories and by limiting attention range.
result Comparable performance can be achieved with 6X fewer long-range memories and better performance with limited attention range.

Four decades after their invention, quasi-Newton methods are still state of the art in unconstrained numerical optimization. Although not usually interpreted thus, these are learning algorithms that fit a local quadratic approximation to the objective function. We show that many, including the most popular, quasi-Newto…

2012-06-18abs ↗pdf ↗

The natural partial ordering of the orbit types of the action of the group of local gauge transformations on the space of connections in space-time dimension d<=4 is investigated. For that purpose, a description of orbit types in terms of cohomology elements of space-time, derived earlier, is used. It is shown that, on…

2000-09-12abs ↗pdf ↗

We survey the main ideas in the early history of the subjects on which Riemann worked and that led to some of his most important discoveries. The subjects discussed include the theory of functions of a complex variable, elliptic and Abelian integrals, the hypergeometric series, the zeta function, topology, differential…

2017-10-11abs ↗pdf ↗

TOLD++ improves convergence of diffusion models by critically damping the forward transition matrix.

problem Improving the convergence of Denoising Diffusion Probabilistic Models.
method Critically damping the Third-Order Langevin Dynamics (TOLD) forward transition matrix using eigen-analysis.
result TOLD++ converges faster than TOLD, verified on toy and real datasets.

Study ribbon concordance and minimal compressions, proving new results about fibered knots.

problem Understanding ribbon concordance and minimal compressions of surface homeomorphisms.
method Proving monotonicity of simplicial volume and dilatation under ribbon concordance, algorithmic enumeration of minimal compressions.
result Every fibered knot has only finitely many predecessors in the ribbon-concordance partial order.

Many real-world applications require robust algorithms to learn point processes based on a type of incomplete data --- the so-called short doubly-censored (SDC) event sequences. We study this critical problem of quantitative asynchronous event sequence analysis under the framework of Hawkes processes by leveraging the …

2017-02-22abs ↗pdf ↗

VICE embeds concepts in a vector space using human data.

problem Developing numerical models for mental representations of object concepts.
method Variational Interpretable Concept Embeddings (VICE) using variational inference and triplet odd-one-out task data.
result VICE outperforms SPoSE at predicting human behavior and provides more reproducible object representations.

Riemann's mathematical papers contain many ideas that arise from physics, and some of them are motivated by problems from physics. In fact, it is not easy to separate Riemann's ideas in mathematics from those in physics. Furthermore, Riemann's philosophical ideas are often in the background of his work on science. The …

2017-11-06abs ↗pdf ↗

We use the contact invariant defined in [2] to construct a new invariant of Legendrian knots in Kronheimer and Mrowka's monopole knot homology theory (KHM), following a prescription of Stipsicz and Vértesi. Our Legendrian invariant improves upon an earlier Legendrian invariant in KHM defined by the second author in sev…

2014-05-13abs ↗pdf ↗

Coherent uncertainty quantification is a key strength of Bayesian methods. But modern algorithms for approximate Bayesian posterior inference often sacrifice accurate posterior uncertainty estimation in the pursuit of scalability. This work shows that previous Bayesian coreset construction algorithms---which build a sm…

2018-02-05abs ↗pdf ↗

New EP variants improve inference stability and efficiency.

problem Inference stability and efficiency issues in EP.
method Motivated by natural-gradient optimization, new EP variants are introduced that are robust to Monte Carlo noise and efficient with single samples.
result Improved stability and efficiency in inference tasks.

One-Shot Neural Architecture Search (NAS) is a promising method to significantly reduce search time without any separate training. It can be treated as a Network Compression problem on the architecture parameters from an over-parameterized network. However, there are two issues associated with most one-shot NAS methods…

2019-05-13abs ↗pdf ↗

This paper reviews deep time-series forecasting focusing on autocorrelation modeling.

problem Modeling autocorrelation in history and label sequences for time-series forecasting.
method Proposes a novel taxonomy for model architectures and learning objectives.
result Provides a comprehensive review and analysis of deep time-series forecasting.

New study shows MLE can avoid model collapse with gradual synthetic data addition.

problem Model collapse in generative models trained on synthetic data.
method Theoretical study of maximum likelihood estimation (MLE) under iterative training with accumulating synthetic data.
result Non-asymptotic bounds show MLE can avoid model collapse even as real data fraction vanishes.

RE enhances DL by learning model behavior, enabling iterative self-improvement.

problem Static data representations limit DL's potential for evolving models.
method RE uses multiple mappings of data through identical deep architectures, analyzing internal representations and performance signals.
result Models can gain insight from predecessors, leading to iterative self-improvement.

Improved neural image compression with refined latent representations.

problem Sub-optimal results from variational autoencoders due to imperfect optimization and capacity limitations.
method Stochastic Gumbel Annealing (SGA) and its extensions (SGA+), including three different methods.
result Significant improvement in compression performance, especially on the R-D trade-off.

Kolmogorov-Arnold Networks improve deep learning adaptivity and can approximate Besov functions optimally.

problem Improving deep learning adaptivity and understanding approximation rates.
method Analyzing Besov norms and using Res-KANs for approximation.
result KANs can optimally approximate Besov functions at the optimal rate.

The generalized volume conjecture relates asymptotic behavior of the colored Jones polynomials to objects naturally defined on an algebraic curve, the zero locus of the A-polynomial A(x,y)A(x,y). Another "family version" of the volume conjecture depends on a quantization parameter, usually denoted qq or \hbar; this quan…

2012-03-09abs ↗pdf ↗

Proposes ΠΠ-Nets, polynomial neural networks, for improved representation power.

problem Improving representation power in deep learning models.
method Introduces ΠΠ-Nets, a new class of deep polynomial neural networks.
result Demonstrates ΠΠ-Nets outperform standard DCNNs and achieve state-of-the-art results.