System predicts vehicle interactions and trajectories with uncertainty.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Normal and almost normal surfaces are essential tools for algorithmic 3-manifold topology, but to use them requires exponentially slow enumeration algorithms in a high-dimensional vector space. The quadrilateral coordinates of Tollefson alleviate this problem considerably for normal surfaces, by reducing the dimension …
GDMaps reduces high-dimensional data to lower dimensions for better classification.
Acoustic Neighbor Embeddings map speech and text to fixed dimensions for phonetic confusability.
Spectral dimensionality reduction algorithms are widely used in numerous domains, including for recognition, segmentation, tracking and visualization. However, despite their popularity, these algorithms suffer from a major limitation known as the "repeated Eigen-directions" phenomenon. That is, many of the embedding co…
This paper introduces a new method for semi-supervised learning on high dimensional nonlinear manifolds, which includes a phase of unsupervised basis learning and a phase of supervised function learning. The learned bases provide a set of anchor points to form a local coordinate system, such that each data point on…
CTM uses conjunctive clauses for image recognition, achieving high accuracy.
In this paper, a novel architecture of Recurrent Neural Network (RNN) is designed and experimented. The proposed RNN adopts a computational memory based on the concept of stigmergy. The basic principle of a Stigmergic Memory (SM) is that the activity of deposit/removal of a quantity in the SM stimulates the next activi…
Charts are an excellent way to convey patterns and trends in data, but they do not facilitate further modeling of the data or close inspection of individual data points. We present a fully automated system for extracting the numerical values of data points from images of scatter plots. We use deep learning techniques t…
Improved speech recognition using EEG and video.
We describe the Fast Greedy Sparse Subspace Clustering (FGSSC) algorithm providing an efficient method for clustering data belonging to a few low-dimensional linear or affine subspaces. The main difference of our algorithm from predecessors is its ability to work with noisy data having a high rate of erasures (missed e…
Study shows emotion affects speaker recognition and vice versa.
With the recent renaissance of deep convolution neural networks, encouraging breakthroughs have been achieved on the supervised recognition tasks, where each class has sufficient training data and fully annotated training data. However, to scale the recognition to a large number of classes with few or now training samp…
End-to-end speech recognition using EEG without speech input.
New proof for sphere recognition algorithm.
Continuous speech recognition from brain activity without vocalization.
VoxCeleb 2019 challenge assesses speaker recognition in uncontrolled settings.
Learning representation on graph plays a crucial role in numerous tasks of pattern recognition. Different from grid-shaped images/videos, on which local convolution kernels can be lattices, however, graphs are fully coordinate-free on vertices and edges. In this work, we propose a Gaussian-induced convolution (GIC) fra…
Paper proposes Roweisposes for 3D action recognition using generalized eigenvalue problem.
This is a survey article on recognition problem of frontal singularities. We specify geometrically several frontal singularities and then we solve the recognition problem of such singularities, giving explicit normal forms. We combine the recognition results by K. Saji and several arguments on openings, which was perfo…
Paper tackles zero-shot activity recognition using video features and text embeddings.
A new deep neural network improves short-speech recognition.
Paper explores EEG-based speech recognition using transformers, showing faster training and better performance for smaller vocabularies.
EmbraceNet fusion model for multi-sensor activity recognition.
Paper improves EEG-based speech recognition using CTC and beam search.
TransFall uses transfer learning to improve activity recognition from mobile sensors.
AV-CPL uses continuous pseudo-labels for AVSR combining labeled and unlabeled data.
Neural network framework for language recognition considers sequence information and improves accuracy.
The paper presents a recognition system for Pashto letters using KNN and ANN.
Fawkes protects images from unauthorized facial recognition models.
In science and engineering, intelligent processing of complex signals such as images, sound or language is often performed by a parameterized hierarchy of nonlinear processing layers, sometimes biologically inspired. Hierarchical systems (or, more generally, nested systems) offer a way to generate complex mappings usin…
Paper shows continuous speech recognition with EEG features, no speech input.
01 loss is robust to outliers and adversarial attacks, outperforming convex losses.
Paper proposes a framework to protect user anonymity in emotion recognition.
Learned feature representations and sub-phoneme posteriors from Deep Neural Networks (DNNs) have been used separately to produce significant performance gains for speaker and language recognition tasks. In this work we show how these gains are possible using a single DNN for both speaker and language recognition. The u…
Study develops sign recognition system for DHH users.
Hybrid and end-to-end models compare in syllable recognition.
Random forest can be adapted for open-set recognition with improved performance.
This paper is a sequel of arxiv:1709.09045 and deals with privileged coordinates and nilpotent approximation of Carnot manifolds. By a Carnot manifold it is meant a manifold equipped with a filtration by subbundles of the tangent bundle which is compatible with the Lie bracket of vector fields. In this paper, we single…
This paper presents a novel method for structural data recognition using a large number of graph models. In general, prevalent methods for structural data recognition have two shortcomings: 1) Only a single model is used to capture structural variation. 2) Naive recognition methods are used, such as the nearest neighbo…
Enhances 2D face recognition with 3D features using active illumination.
We build CSI-Net, a unified Deep Neural Network~(DNN), to learn the representation of WiFi signals. Using CSI-Net, we jointly solved two body characterization problems: biometrics estimation (including body fat, muscle, water, and bone rates) and person recognition. We also demonstrated the application of CSI-Net on tw…
Improved speech emotion recognition using pre-trained language models.
Coordinate descent methods usually minimize a cost function by updating a random decision variable (corresponding to one coordinate) at a time. Ideally, we would update the decision variable that yields the largest decrease in the cost function. However, finding this coordinate would require checking all of them, which…
Proves NP and co-NP status for knot core recognition in solid torus.
The performance of automatic speech recognition systems(ASR) degrades in the presence of noisy speech. This paper demonstrates that using electroencephalography (EEG) can help automatic speech recognition systems overcome performance loss in the presence of noise. The paper also shows that distillation training of auto…
Deep learning improves speaker recognition verification and identification.
NIST CTS Superset offers a large dataset for telephony speaker recognition.