Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

35810 · May 201819922001200920172026
48 results for audio-transcript entailment

End-to-end ASR error detection using audio-transcript entailment.

problem Detecting transcription errors in ASR systems to prevent error propagation.
method Proposes a novel end-to-end approach using audio-transcript entailment, with acoustic and linguistic encoders.
result Achieves CER of 26.2% on all transcription errors and 23% on medical errors specifically, improving by 12% and 15.4% respectively over a strong baseline.

In training a deep learning system to perform audio transcription, two practical problems may arise. Firstly, most datasets are weakly labelled, having only a list of events present in each recording without any temporal information for training. Secondly, deep neural networks need a very large amount of labelled train…

2018-07-10abs ↗pdf ↗

The application of deep recurrent networks to audio transcription has led to impressive gains in automatic speech recognition (ASR) systems. Many have demonstrated that small adversarial perturbations can fool deep neural networks into incorrectly predicting a specified target with high confidence. Current work on fool…

2018-05-20abs ↗pdf ↗

Neural language models are a critical component of state-of-the-art systems for machine translation, summarization, audio transcription, and other tasks. These language models are almost universally autoregressive in nature, generating sentences one token at a time from left to right. This paper studies the influence o…

2018-08-23abs ↗pdf ↗

The paper proposes methods to control errors in language generation models using textual entailment.

problem The lack of a correctness metric hinders applying principled methods to language generation tasks.
method The paper leverages textual entailment to evaluate correctness and proposes two selective generation algorithms: SGen^Sup and SGen^Semi.
result The proposed algorithms control the false discovery rate with respect to textual entailment and achieve comparable selection efficiency to baselines.

Word embeddings provide point representations of words containing useful semantic information. We introduce multimodal word distributions formed from Gaussian mixtures, for multiple word meanings, entailment, and rich uncertainty information. To learn these distributions, we propose an energy-based max-margin objective…

2017-04-27abs ↗pdf ↗

Training modern deep learning models requires large amounts of computation, often provided by GPUs. Scaling computation from one GPU to many can enable much faster training and research progress but entails two complications. First, the training library must support inter-GPU communication. Depending on the particular …

2018-02-15abs ↗pdf ↗

Learning graph representations via low-dimensional embeddings that preserve relevant network properties is an important class of problems in machine learning. We here present a novel method to embed directed acyclic graphs. Following prior work, we first advocate for using hyperbolic spaces which provably model tree-li…

2018-04-03abs ↗pdf ↗

In this article we extend cutting and blowing up to the nonrational symplectic toric setting. This entails the possibility of cutting and blowing up for symplectic toric manifolds and orbifolds in nonrational directions.

2016-06-02abs ↗pdf ↗

By representing words with probability densities rather than point vectors, probabilistic word embeddings can capture rich and interpretable semantic information and uncertainty. The uncertainty information can be particularly meaningful in capturing entailment relationships -- whereby general words such as "entity" co…

2018-04-26abs ↗pdf ↗

ICP improves text infilling and POS tagging with valid confidence sets.

problem Statistical reliability of machine learning predictions.
method Inductive conformal prediction algorithms for text infilling and POS tagging.
result Valid set-valued predictions with small size for real-world applications.

This paper is an attempt at understanding the quantum-like dynamics of financial markets in terms of non-differentiable price-time continuum having fractal properties. The main steps of this development are the statistical scaling, the non-differentiability hypothesis, and the equations of motion entailed by this hypot…

2013-12-11abs ↗pdf ↗

We prove that a wide class of correlated stochastic volatility models exactly measure an empirical fact in which past returns are anticorrelated with future volatilities: the so-called ``leverage effect''. This quantitative measure allows us to fully estimate all parameters involved and it will entail a deeper study on…

2002-02-12abs ↗pdf ↗

Modern neural networks are often augmented with an attention mechanism, which tells the network where to focus within the input. We propose in this paper a new framework for sparse and structured attention, building upon a smoothed max operator. We show that the gradient of this operator defines a mapping from real val…

2017-05-22abs ↗pdf ↗

We reformulate Lehmer's question from 1933 and a question due to Schinzel and Zassenhaus from 1965 in terms of a comparison of the Mahler measures and the houses, respectively, of monic integer reciprocal and skew-reciprocal polynomials of the same degree. This entails that understanding the difference between orientat…

2018-12-12abs ↗pdf ↗

The Killing operator on a Riemannian manifold is a linear differential operator on vector fields whose kernel provides the infinitesimal Riemannian symmetries. The Killing operator is best understood in terms of its prolongation, which entails some simple tensor identities. These simple identities can be viewed as aris…

2010-06-08abs ↗pdf ↗

Embedding methods which enforce a partial order or lattice structure over the concept space, such as Order Embeddings (OE) (Vendrov et al., 2016), are a natural way to model transitive relational data (e.g. entailment graphs). However, OE learns a deterministic knowledge base, limiting expressiveness of queries and the…

2018-05-17abs ↗pdf ↗

We discuss the construction of Sp(2)Sp(1)-structures whose fundamental form is closed. In particular, we find 10 new examples of 8-dimensional nilmanifolds that admit an invariant closed 4-form with stabiliser Sp(2)Sp(1). Our constructions entail the notion of SO(4)-structures on 7-manifolds. We present a thorough inve…

2013-08-19abs ↗pdf ↗

Design of reliable systems must guarantee stability against input perturbations. In machine learning, such guarantee entails preventing overfitting and ensuring robustness of models against corruption of input data. In order to maximize stability, we analyze and develop a computationally efficient implementation of Jac…

2019-08-07abs ↗pdf ↗

Paper characterizes causal graphs from hard interventions and proposes a learning algorithm.

problem Discovering causal structure from hard interventions and observational data.
method Proposes graphical constraints and a learning algorithm based on do-calculus.
result Characterizes interventional equivalence classes of causal graphs with latent variables.

We study the first uniformly finite homology group of Block and Weinberger for uniformly locally finite graphs, with coefficients in Z\mathbb{Z} and Z2\mathbb{Z}_2. When the graph is a tree, or coefficients are in Z2\mathbb{Z}_2, a characterisation of the group is obtained. In the general case, we describe three pheno…

2020-01-14abs ↗pdf ↗

Multi-hop inference is necessary for machine learning systems to successfully solve tasks such as Recognising Textual Entailment and Machine Reading. In this work, we demonstrate the effectiveness of adaptive computation for learning the number of inference steps required for examples of different complexity and that l…

2016-10-24abs ↗pdf ↗

VAEs struggle with surjective multimodal data, especially class labels describing images.

problem VAEs struggle to capture variability in surjective multimodal data.
method Theoretical and empirical demonstration of VAEs with a mixture of experts posterior.
result VAEs with a mixture of experts posterior can disregard variation in surjective multimodal data.

We study vector fields generating a local flow by automorphisms of a parabolic geometry with higher order fixed points. We develop general tools extending the techniques of [1], [2], and [3]. We apply these tools to almost Grassmannian, almost quaternionic, and contact parabolic geometries, including CR structures, to …

2012-08-27abs ↗pdf ↗

The pseudo-Finsleroid relativistic metric was constructed upon assuming that the involved vector field bib_i is time-like. In the present paper it is shown that the metric admits just the alternative counterpart in which the field is space-like. The entailed pseudo-Finsleroid-spatial framework is systematically describ…

2008-06-16abs ↗pdf ↗

SNN architecture shows gradient descent converges to regularized solution in matrix sensing problems.

problem Understanding implicit regularization in neural networks for matrix sensing.
method Developed Spectral Neural Networks (SNN) for matrix learning problems, rigorously demonstrating implicit regularization.
result Gradient descent converges to the solution of a regularized learning problem in matrix sensing problems.

Paper summarizes unsupervised learning challenges for disentangled representations.

problem Unsupervised learning of disentangled representations without inductive biases.
method Theoretical and practical analysis of existing approaches.
result Unsupervised disentanglement is fundamentally impossible without inductive biases.

In this note we consider homogeneous Willmore surfaces in Sn+2S^{n+2}. The main result is that a homogeneous Willmore two-sphere is conformally equivalent to a homogeneous minimal two-sphere in Sn+2S^{n+2}, i.e., either a round two-sphere or one of the Borůvka-Veronese 2-spheres in S2mS^{2m}. This entails a classification o…

2018-05-09abs ↗pdf ↗

Obtaining continuous representations of structural data such as directed acyclic graphs (DAGs) has gained attention in machine learning and artificial intelligence. However, embedding complex DAGs in which both ancestors and descendants of nodes are exponentially increasing is difficult. Tackling in this problem, we de…

2019-02-12abs ↗pdf ↗

Machine learning tasks entail the use of complex computational pipelines to reach quantitative and qualitative conclusions. If some of the activities in a pipeline produce erroneous or uninformative outputs, the pipeline may fail or produce incorrect results. Inferring the root cause of failures and unexpected behavior…

2020-02-11abs ↗pdf ↗