Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,932 papers · 148 categories

Trend · papers per month

134269403537 · Jun 202019922001200920172026
48 results for single class

Single neurons can perform as well as dense networks in binary and multi-class recognition tasks.

problem Designing efficient neural networks for recognition tasks.
method Investigated the use of single or multiple neurons in neural networks for binary and multi-class recognition tasks.
result Sparse networks can be as efficient as dense networks in both binary and multi-class tasks.

We formulate a new class of conditional generative models based on probability flows. Trained with maximum likelihood, it provides efficient inference and sampling from class-conditionals or the joint distribution, and does not require a priori knowledge of the number of classes or the relationships between classes. Th…

2019-02-05abs ↗pdf ↗

Develops a new method for efficient stochastic bilevel optimization.

problem Stochastic bilevel optimization problems in machine learning applications.
method Single-Timescale stochAstic BiLevEl optimization (STABLE) method.
result Achieves the same order of sample complexity as stochastic gradient descent for single-level optimization.

A new classifier encodes local neighborhoods for each class using Fly Bloom Filters.

problem Efficiently classify data with single-pass learning.
method Proposes a new classifier that encodes local neighborhoods for each class with per-class Fly Bloom Filters.
result The proposed classifier's performance is competitive with nearest-neighbor classifiers and other single-pass classifiers.

Efficiently learns Single-Index Models with constant factor approximation.

problem Learning Single-Index Models under L22L_2^2 loss with unknown link functions.
method An efficient algorithm using alignment sharpness for optimization.
result Achieves constant factor approximation to optimal loss for various distributions and link functions.

RL in MFGs is as hard as solving many single-agent RL problems.

problem Learning Nash Equilibrium in Mean-Field Games (MFGs).
method Introduce P-MBED to measure model complexity, develop a novel exploration strategy, and establish polynomial sample complexity results.
result Learning Nash Equilibrium in MFGs is no more statistically challenging than solving a logarithmic number of single-agent RL problems.

Despite their ability to memorize large datasets, deep neural networks often achieve good generalization performance. However, the differences between the learned solutions of networks which generalize and those which do not remain unclear. Additionally, the tuning properties of single directions (defined as the activa…

2018-03-19abs ↗pdf ↗

Optimizes natural frequencies of cellular composites with various microstructures.

problem Designing cellular composites with diverse microstructures for maximizing natural frequencies.
method Data-driven topology optimization with a latent-variable Gaussian process model.
result Cellular designs with multiclass microstructures achieve higher natural frequencies.

The classical Godbillon-Vey invariant is an odd degree cohomology class that is a cobordism invariant of a single foliation. Here we investigate cohomology classes of even degree that are cobordism invariants of (germs of) 1-parameter families of foliations.

2001-11-12abs ↗pdf ↗

Kernel alignment measures the degree of similarity between two kernels. In this paper, inspired from kernel alignment, we propose a new Linear Discriminant Analysis (LDA) formulation, kernel alignment LDA (kaLDA). We first define two kernels, data kernel and class indicator kernel. The problem is to find a subspace to …

2016-10-14abs ↗pdf ↗

Study SGD dynamics in sequence models, revealing training phases and influence of sequence length.

problem Understanding SGD in sequence models like attention networks.
method Derived closed-form population loss and analyzed SGD dynamics for SSI models.
result Two distinct training phases: escape from uninformative initialization and alignment with target subspace.

The paper provides guarantees for learning switching non-linear systems from a single trajectory.

problem Learning non-linear dynamical systems with switching dynamics.
method Non-asymptotic bounds derived under stability assumptions for i.i.d. switching modes.
result Explicit convergence rates for Hölder and linear function classes based on effective sample size.

SGD shows distinct phases in learning single-index models, achieving optimal sample complexity and regret.

problem Learning single-index models with SGD in adaptive data settings.
method Stochastic gradient descent (SGD) with an optimal learning rate schedule.
result SGD achieves near-optimal sample complexity and regret guarantees across both burn-in and learning phases.

We prove that the deRham cohomology classes of Lee forms of locally conformally symplectic structures taming the complex structure of a compact complex surface SS with first Betti number equal to 11 is either a non-empty open subset of HdR1(S,R)H^1_{dR}(S, \mathbb R), or a single point. In the latter case, we show that SS

2016-10-31abs ↗pdf ↗

We carry out a Painlevé analysis to find the cases where the cohomogeneity one steady Ricci soliton equation can be integrable. We concentrate on two classes of solitons: warped products and complex line bundles over a Fano Kähler Einstein base. For warped products, the analysis singles out the case with one factor whe…

2018-02-28abs ↗pdf ↗

A new teacher-class network method compresses DNNs by distributing knowledge to multiple student networks.

problem Overwhelming size of Deep Neural Networks (DNNs).
method Single teacher with multiple student networks, transferring knowledge to each student.
result The combined knowledge of the class of students achieves better performance and reduces parameters.

Methods for automated discovery of causal relationships from non-interventional data have received much attention recently. A widely used and well understood model family is given by linear acyclic causal models (recursive structural equation models). For Gaussian data both constraint-based methods (Spirtes et al., 199…

2012-05-09abs ↗pdf ↗

Study shows computational and statistical gaps in Gaussian Single-Index Models.

problem Statistical and computational trade-offs in high-dimensional regression problems.
method Analysis of SQ and LDP frameworks, partial-trace algorithm.
result Computational algorithms require significantly more samples than information-theoretic limits.

Estimates joint causal effects using single-variable interventions on nonlinear models.

problem Estimating joint causal effects from single-variable interventions.
method Identifiability result and practical estimator for decomposing causal effects.
result Joint effects can be inferred without joint interventional data for nonlinear additive models.

Multi-expert L2D underfits more severely, requiring new methods.

problem Underfitting in multi-expert L2D settings.
method PiCCE (Pick the Confident and Correct Expert), a surrogate-based method.
result PiCCE effectively reduces multi-expert L2D to a single-expert-like problem, resolving underfitting.

Modifying the method of [21], we compute the perturbed HF+HF^+ for some special classes of fibered three manifolds in the second highest spinc^c-structures Sg2S_{g-2}. The special classes considered in this paper include the mapping tori of Dehn twists along a single non-separating curve and along a transverse pair of c…

2009-03-02abs ↗pdf ↗

Improves conformal prediction by combining multiple score functions and optimizing weights.

problem Limitations of single-score conformal predictors in multi-class classification.
method Combines multiple score functions and optimizes weights to minimize prediction set size.
result Consistently outperforms single-score conformal predictors while maintaining valid coverage.

The pullback approach to global Finsler geometry is adopted. Three classes of recurrence in Finsler geometry are introduced and investigated: simple recurrence, Ricci recurrence and concircular recurrence. Each of these classes consists of four types of recurrence. The interrelationships between the different types of …

2016-07-25abs ↗pdf ↗

A new model explains asset returns with a single factor, improving cross-sectional performance.

problem Understanding the cross-section of asset returns with complex models.
method Proposes a non-linear single-factor asset pricing model with a nonparametric link function estimated jointly with sieve-based estimators.
result The model delivers superior cross-sectional performance with a low-dimensional approximation of the link function.

We show that on a nonorientable surface of genus at least 7 any power of a Dehn twist is equal to a single commutator in the mapping class group and the same is true, under additional assumptions, for the twist subgroup, and also for the extended mapping class group of an orientable surface of genus at least 3.

2010-07-01abs ↗pdf ↗

This paper initiates a systematic study of the relation of commensurability of surface automorphisms, or equivalently, fibered commensurability of 3-manifolds fibering over the circle. We show that every hyperbolic fibered commensurability class contains a unique minimal element, whereas the class of Seifert manifolds …

2010-03-01abs ↗pdf ↗

Single proxy variable helps estimate causal effects from confounders.

problem Estimating causal effects from treatment to outcome when unobserved confounders are present.
method Assumes a single, potentially multi-dimensional proxy variable of the unobserved confounder and a known mechanism generating the proxy from the confounder. Proves causal effects are identifiable under completeness assumption.
result Causal effects are identifiable under SPICE assumption.

Paper tackles offline RL with weak assumptions on both function classes and data coverage.

problem Achieve sample-efficient offline RL with weak assumptions on both factors.
method Simple algorithm based on primal-dual formulation of MDPs, with density-ratio function modeling dual variables.
result Polynomial sample complexity achieved under realizability and single-policy concentrability.

This paper shows that there are symplectic four-manifolds M with the following property: a single isotopy class of smooth embedded two-spheres in M contains infinitely many Lagrangian submanifolds, no two of which are isotopic as Lagrangian submanifolds. The examples are constructed using a special class of symplectic …

1998-03-19abs ↗pdf ↗

Paper proposes methods to learn with multiple incorrect labels per example.

problem Learning with a single incorrect label per example limits potential.
method Proposes a novel problem setting allowing multiple incorrect labels per example and two learning methods.
result Demonstrates improved learning with multiple incorrect labels compared to single incorrect labels.