Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

85170254339 · Jun 202019922001200920172026
48 results for fewer examples

Method trains deep models to explain predictions with fewer examples.

problem Difficulty in humans understanding deep model predictions.
method Simultaneously trains prediction and explanation models with sparse regularization.
result Improves faithfulness of explanations with fewer examples while maintaining predictive performance.

One of the biggest bottlenecks in a machine learning workflow is waiting for models to train. Depending on the available computing resources, it can take days to weeks to train a neural network on a large dataset with many classes such as ImageNet. For researchers experimenting with new algorithmic approaches, this is …

2019-06-12abs ↗pdf ↗

Information planning enables faster learning with fewer training examples. It is particularly applicable when training examples are costly to obtain. This work examines the advantages of information planning for text data by focusing on three supervised models: Naive Bayes, supervised LDA and deep neural networks. We s…

2018-02-09abs ↗pdf ↗

Several large classes of homogeneous spaces are known to be formal---in the sense of Rational Homotopy Theory. However, it seems that far fewer examples of non-formal homogeneous spaces are known. In this article we provide several construction principles and characterisations for non-formal homogeneous spaces, which w…

2012-06-04abs ↗pdf ↗

GGAN improves audio representation learning with fewer labels.

problem Learning representations for specific tasks from unlabelled data.
method Guided Generative Adversarial Neural Network (GGAN).
result GGAN learns better representations with fewer labelled data.

Error bounds based on worst likely assignments use permutation tests to validate classifiers. Worst likely assignments can produce effective bounds even for data sets with 100 or fewer training examples. This paper introduces a statistic for use in the permutation tests of worst likely assignments that improves error b…

2015-03-31abs ↗pdf ↗

The ability to learn from a small number of examples has been a difficult problem in machine learning since its inception. While methods have succeeded with large amounts of training data, research has been underway in how to accomplish similar performance with fewer examples, known as one-shot or more generally few-sh…

2017-08-22abs ↗pdf ↗

A core aspect of human intelligence is the ability to learn new tasks quickly and switch between them flexibly. Here, we describe a modular continual reinforcement learning paradigm inspired by these abilities. We first introduce a visual interaction environment that allows many types of tasks to be unified in a single…

2017-11-20abs ↗pdf ↗

We establish a correspondence between trisections of smooth, compact, oriented 44--manifolds with connected boundary and diagrams describing these trisected 44--manifolds. Such a diagram comes in the form of a compact, oriented surface with boundary together with three tuples of simple closed curves, with possibly fe…

2016-10-20abs ↗pdf ↗

We show how the success of deep learning could depend not only on mathematics but also on physics: although well-known mathematical theorems guarantee that neural networks can approximate arbitrary functions well, the class of functions of practical interest can frequently be approximated through "cheap learning" with …

2016-08-29abs ↗pdf ↗

We consider the "intrinsic" symmetry group of a two-component link LL, defined to be the image Σ(L)Σ(L) of the natural homomorphism from the standard symmetry group $\MCG(S^3,L)$ to the product $\MCG(S^3) \cross \MCG(L)$. This group, first defined by Whitten in 1969, records directly whether LL is isotopic to a link $L…

2012-01-13abs ↗pdf ↗

Efficient kernel method learns differential equations with fewer data.

problem Learning differential equations with limited data and computational resources.
method Kernel-based framework for differential equations with theoretical error bounds.
result Significant improvements in accuracy and computational efficiency.

For a knot K, let b_n(K) be the minimum length of an n-stranded braid representative of K. Examples of knots exist for which b_n(K) is a non-increasing function. We investigate the behavior of b_n(K). We develop bounds on the function in terms of the genus of K, with stronger results for homogeneous knots and braid pos…

2006-05-17abs ↗pdf ↗

The Turaev genus and dealternating number of a link are two invariants that measure how far away a link is from alternating. We determine the Turaev genus of a torus knot with five or fewer strands either exactly or up to an error of at most one. We also determine the dealternating number of a torus knot with five or f…

2017-03-07abs ↗pdf ↗

Machine learning has made major advances in categorizing objects in images, yet the best algorithms miss important aspects of how people learn and think about categories. People can learn richer concepts from fewer examples, including causal models that explain how members of a category are formed. Here, we explore the…

2019-04-17abs ↗pdf ↗

In this paper, we define the primitive/Seifert-fibered property for a knot in S^3. If satisfied, the property ensures that the knot has a Dehn surgery that yields a small Seifert-fibered space (i.e. base S^2 and three or fewer critical fibers). Next we describe the twisted torus knots, which provide an abundance of exa…

2003-06-15abs ↗pdf ↗

We show that there exist infinitely many examples of pairs of knots, K_1 and K_2, that have no epimorphism π1(S3K1)π1(S3K2)π_1(S^3\setminus K_1) \to π_1(S^3\setminus K_2) preserving peripheral structure although their A-polynomials have the factorization AK2(L,M)AK1(L,M)A_{K_2}(L,M) \mid A_{K_1}(L,M). Our construction accounts for most of the kno…

2011-07-13abs ↗pdf ↗

We show that the proportion of hyperbolic knots among all of the prime knots of nn or fewer crossings does not converge to 11 as nn approaches infinity. Moreover, we show that if KK is a nontrivial knot then the proportion of satellites of KK among all of the prime knots of nn or fewer crossings does not converge…

2019-08-16abs ↗pdf ↗

Active feature selection uses mutual information to choose fewer labels for better feature selection.

problem Selecting features with limited labeled data.
method Uses active feature selection with mutual information criterion, optimizing label selection for better feature quality.
result Algorithm selects features with higher mutual information using fewer labels than the data set size.

The concordance orders of many algebraic order two knots of ten or fewer crossings have been heretofore unknown. We use Casson-Gordon invariants and twisted Alexander polynomials to find that, in all but one case, these knots do not have concordance order two. We also find that a certain family of algebraic order two t…

2000-08-08abs ↗pdf ↗

We investigate the bi-orderability of two-bridge knot groups and the groups of knots with 12 or fewer crossings by applying recent theorems of Chiswell, Glass and Wilson. Amongst all knots with 12 or fewer crossings (of which there are 2977), previous theorems were only able to determine bi-orderability of 599 of the c…

2014-10-21abs ↗pdf ↗

Note that this paper is superceded by "Black-Box Adversarial Attacks with Limited Queries and Information." Current neural network-based image classifiers are susceptible to adversarial examples, even in the black-box setting, where the attacker is limited to query access without access to gradients. Previous methods -…

2017-12-19abs ↗pdf ↗

Most of Markov Chain Monte Carlo (MCMC) and sequential Monte Carlo (SMC) algorithms in existing probabilistic programming systems suboptimally use only model priors as proposal distributions. In this work, we describe an approach for training a discriminative model, namely a neural network, in order to approximate the …

2015-12-14abs ↗pdf ↗

We show that if KK is a nontrivial knot then the proportion of satellites of KK among all of the prime non-split links of nn or fewer crossings does not converge to 00 as nn approaches infinity. This implies in particular that the proportion of hyperbolic links among all of the prime non-split links of nn or fewe…

2019-07-09abs ↗pdf ↗

The sparse group lasso optimization problem is solved using a coordinate gradient descent algorithm. The algorithm is applicable to a broad class of convex loss functions. Convergence of the algorithm is established, and the algorithm is used to investigate the performance of the multinomial sparse group lasso classifi…

2012-05-06abs ↗pdf ↗

A graph is 2-apex if it is planar after the deletion of at most two vertices. Such graphs are not intrinsically knotted, IK. We investigate the converse, does not IK imply 2-apex? We determine the simplest possible counterexample, a graph on nine vertices and 21 edges that is neither IK nor 2-apex. In the process, we s…

2009-10-08abs ↗pdf ↗

Fewer data weight updates lead to faster convergence in machine learning models.

problem Improving robustness of machine learning models through data mixing.
method Analyzing convergence behavior of data mixing with a finite number of inner steps.
result The optimal number of inner steps scales with the budget and type of gradients used.

Deep neural networks excel at learning the training data, but often provide incorrect and confident predictions when evaluated on slightly different test examples. This includes distribution shifts, outliers, and adversarial examples. To address these issues, we propose Manifold Mixup, a simple regularizer that encoura…

2018-06-13abs ↗pdf ↗

The goal of a decision-based adversarial attack on a trained model is to generate adversarial examples based solely on observing output labels returned by the targeted model. We develop HopSkipJumpAttack, a family of algorithms based on a novel estimate of the gradient direction using binary information at the decision…

2019-04-03abs ↗pdf ↗

Examples are given of prime Legendrian knots in the standard contact 3-space that have arbitrarily many distinct Chekanov polynomials, refuting a conjecture of Lenny Ng. These are constructed using a new `Legendrian tangle replacement' technique. This technique is then used to show that the phenomenon of multiple Cheka…

2004-11-09abs ↗pdf ↗

Study shows semi-supervised learning can be more robust with fewer labeled examples.

problem Learning robust predictors in semi-supervised PAC model with minimal labeled data.
method Characterizes the minimal labeled and unlabeled data required for robust learning.
result Proves nearly matching upper and lower bounds on labeled sample complexity.

Recent work on deep neural network pruning has shown there exist sparse subnetworks that achieve equal or improved accuracy, training time, and loss using fewer network parameters when compared to their dense counterparts. Orthogonal to pruning literature, deep neural networks are known to be susceptible to adversarial…

2019-12-05abs ↗pdf ↗