This paper reports the performances of shallow word-level convolutional neural networks (CNN), our earlier work (2015), on the eight datasets with relatively large training data that were used for testing the very deep character-level CNN in Conneau et al. (2016). Our findings are as follows. The shallow word-level CNN…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
End-to-end automatic speech recognition (ASR) commonly transcribes audio signals into sequences of characters while its performance is evaluated by measuring the word-error rate (WER). This suggests that predicting sequences of words directly may be helpful instead. However, training with word-level supervision can be …
We present Listen, Attend and Spell (LAS), a neural network that learns to transcribe speech utterances to characters. Unlike traditional DNN-HMM models, this model learns all the components of a speech recognizer jointly. Our system has two components: a listener and a speller. The listener is a pyramidal recurrent ne…
Paper proposes a method to maintain ASR performance on new tasks without forgetting old ones.
A tailored HTR system improves CER to 0.015 for medieval Latin.
Bayesian method improves neural net convergence for character recognition.
In this paper, we address the task of Optical Character Recognition(OCR) for the Telugu script. We present an end-to-end framework that segments the text image, classifies the characters and extracts lines using a language model. The segmentation is based on mathematical morphology. The classification module, which is …
In this paper, we propose a refined multi-stage multi-task training strategy to improve the performance of online attention-based encoder-decoder (AED) models. A three-stage training based on three levels of architectural granularity namely, character encoder, byte pair encoding (BPE) based encoder, and attention decod…
Predicting the click-through rate of an advertisement is a critical component of online advertising platforms. In sponsored search, the click-through rate estimates the probability that a displayed advertisement is clicked by a user after she submits a query to the search engine. Commercial search engines typically rel…
Many important forms of data are stored digitally in XML format. Errors can occur in the textual content of the data in the fields of the XML. Fixing these errors manually is time-consuming and expensive, especially for large amounts of data. There is increasing interest in the research, development, and use of automat…
Paper introduces new bounds linking data compressibility to generalization error.
Decoding language representations directly from the brain can enable new Brain-Computer Interfaces (BCI) for high bandwidth human-human and human-machine communication. Clinically, such technologies can restore communication in people with neurological conditions affecting their ability to speak. In this study, we prop…
Charmer improves character-level adversarial attacks for language models.
A transformer model improves spell correction with hierarchical attention.
The recently proposed Sequence-to-Sequence (seq2seq) framework advocates replacing complex data processing pipelines, such as an entire automatic speech recognition system, with a single neural network trained in an end-to-end fashion. In this contribution, we analyse an attention-based seq2seq speech recognition syste…
Bayesian SHMM models speech units from unannotated speech.
Direct acoustics-to-word (A2W) models in the end-to-end paradigm have received increasing attention compared to conventional sub-word based automatic speech recognition models using phones, characters, or context-dependent hidden Markov model states. This is because A2W models recognize words from speech without any de…
This paper tackles worst-class error rate in classification tasks.
AV-ASR system improves speech recognition with visual context.
The theory of differential characters is developed completely from a de Rham - Federer viewpoint. Characters are defined as equivalence classes of special currents, called sparks, which appear naturally in the theory of singular connections. There are many different spaces of currents which yield the character groups. …
Study controls error rates of binary classifiers using hypothesis testing.
We show how to compute the Bayes error-rate for speaker verifiers.
This research examines how the error rate of nearest neighbor classifiers varies with dataset size.
Character scheme of small Seifert 3-manifolds is reduced if and only if no exceptional abelian character exists.
Quantum character varieties unify four construction methods.
We show that for any knot there exist only finitely many irreducible metabelian characters in the -character variety of the knot group, and the number is given explicitly by using the determinant of the knot. Then it turns out that for any 2-bridge knot a section of the -character va…
This paper reviews methods for constructing confidence intervals for error rates in 1:1 matching tasks.
Character varieties get a natural Poisson structure.
Study of orbifold Chern character using superconnections.
In this paper, we give a quantum interpretation of the Bismut-Chern character form (the loop space lifting of the Chern character form) as well as the Chern character form associated to a complex vector bundle with connection over a smooth manifold in the framework of supersymmetric quantum field theories developed by …
Study shows how certain knots and tori are detected by ideal points in character varieties.
Study SL(2,C) character schemes for finitely generated groups.
Background elimination for noisy character images or character images from real scene is still a challenging problem, due to the bewildering backgrounds, uneven illumination, low resolution and different distortions. We propose a stroke-based character reconstruction(SCR) method that use a weighted quadratic Bezier cur…
Character variety of Borromean link solved, Alexander polynomial formula found.
RoyalFlush system improves multi-speaker ASR in M2MeT challenge.
Ex ante forecast outcomes should be interpreted as counterfactuals (potential histories), with errors as the spread between outcomes. Reapplying measurements of uncertainty about the estimation errors of the estimation errors of an estimation leads to branching counterfactuals. Such recursions of epistemic uncertainty …
Grapheme-based acoustic modeling has recently been shown to outperform phoneme-based approaches in both hybrid and end-to-end automatic speech recognition (ASR), even on non-phonemic languages like English. However, graphemic ASR still has problems with rare long-tail words that do not follow the standard spelling conv…
This research analyzes the error convergence rate of GAN models.
Computes dimensions of representation and character varieties for 2 and 3-dimensional orbifolds.
Paper generates personalized fonts from a few characters.
We study some basic properties of the variety of characters in PSL(2,C) of a finitely generated group. In particular we give an interpretation of its points as characters of representations. We construct 3-manifolds whose variety of characters has arbitrarily many components that do not lift to SL(2,C). We also study t…
Study of quantum decorated character stacks and their quantizations.
Let K be a knot in an integral homology 3-sphere and let B denote the 2-fold branched cover of the integral homology sphere branched along K. We construct a map from the slice of characters with trace free along meridians in the SL(2, C)-character variety of the knot exterior to the SL(2, C)-character variety of 2-fold…
Overrides of credit ratings are important correctives of ratings that are determined by statistical rating models. Financial institutions and banking regulators agree on this because on the one hand errors with ratings of corporates or banks can have fatal consequences for the lending institutions and on the other hand…
We consider least squares estimation in a general nonparametric regression model. The rate of convergence of the least squares estimator (LSE) for the unknown regression function is well studied when the errors are sub-Gaussian. We find upper bounds on the rates of convergence of the LSE when the errors have uniformly …
Study character varieties of hyperbolic 3-manifolds using bundle methods.
Researchers describe character varieties for Hopf links, proving geometric properties.
Study of SU(2,1) character varieties on one-holed torus.