Parallel sentence extraction is a task addressing the data sparsity problem found in multilingual natural language processing applications. We propose a bidirectional recurrent neural network based approach to extract parallel sentences from collections of multilingual texts. Our experiments with noisy parallel corpora…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Parallel sentences are a relatively scarce but extremely useful resource for many applications including cross-lingual retrieval and statistical machine translation. This research explores our methodology for mining such data from previously obtained comparable corpora. The task is highly practical since non-parallel m…
Machine learning is being increasingly used by individuals, research institutions, and corporations. This has resulted in the surge of Machine Learning-as-a-Service (MLaaS) - cloud services that provide (a) tools and resources to learn the model, and (b) a user-friendly query interface to access the model. However, suc…
This study presents a new lossy image compression method that utilizes the multi-scale features of natural images. Our model consists of two networks: multi-scale lossy autoencoder and parallel multi-scale lossless coder. The multi-scale lossy autoencoder extracts the multi-scale image features to quantized variables a…
Cardiac motion modeling using LDDMM and shape splines.
CURE extracts relations without supervision by clustering similar entity pairs.
Parallel texts are a relatively rare language resource, however, they constitute a very useful research material with a wide range of applications. This study presents and analyses new methodologies we developed for obtaining such data from previously built comparable corpora. The methodologies are automatic and unsupe…
This paper automates mining of COVID-19 scholarly articles using machine learning.
We prove that Fefferman spaces, associated to non--degenerate CR structures of hypersurface type, are characterised, up to local conformal isometry, by the existence of a parallel orthogonal complex structure on the standard tractor bundle. This condition can be equivalently expressed in terms of conformal holonomy. Ex…
Grammatical Error Correction (GEC) has been recently modeled using the sequence-to-sequence framework. However, unlike sequence transduction problems such as machine translation, GEC suffers from the lack of plentiful parallel data. We describe two approaches for generating large parallel datasets for GEC using publicl…
TSML tackles anomaly detection and pattern discovery in industrial time series data.
This paper proposes a framework to predict long-term trends and short-term fluctuations in multivariate time series.
We propose a method to predict the subject-specific longitudinal progression of brain structures extracted from baseline MRI, and evaluate its performance on Alzheimer's disease data. The disease progression is modeled as a trajectory on a group of diffeomorphisms in the context of large deformation diffeomorphic metri…
Topic modeling, a method for extracting the underlying themes from a collection of documents, is an increasingly important component of the design of intelligent systems enabling the sense-making of highly dynamic and diverse streams of text data. Traditional methods such as Dynamic Topic Modeling (DTM) do not lend the…
Probabilistic embeddings improve speaker diarization accuracy.
SDD improves DD for estimating treatment effects by adjusting for confounding.
Word2vec is a widely used algorithm for extracting low-dimensional vector representations of words. State-of-the-art algorithms including those by Mikolov et al. have been parallelized for multi-core CPU architectures, but are based on vector-vector operations with "Hogwild" updates that are memory-bandwidth intensive …
New framework extracts useful information from tensor data with structural properties.
Given a multivariate data set, sparse principal component analysis (SPCA) aims to extract several linear combinations of the variables that together explain the variance in the data as much as possible, while controlling the number of nonzero loadings in these combinations. In this paper we consider 8 different optimiz…
Word2Vec is a widely used algorithm for extracting low-dimensional vector representations of words. It generated considerable excitement in the machine learning and natural language processing (NLP) communities recently due to its exceptional performance in many NLP applications such as named entity recognition, sentim…
We introduce BilBOWA (Bilingual Bag-of-Words without Alignments), a simple and computationally-efficient model for learning bilingual distributed representations of words which can scale to large monolingual datasets and does not require word-aligned parallel training data. Instead it trains directly on monolingual dat…
In the era of big data, practical applications in various domains continually generate large-scale time-series data. Among them, some data show significant or potential periodicity characteristics, such as meteorological and financial data. It is critical to efficiently identify the potential periodic patterns from mas…
Generalizes holographic method to higher codimension submanifolds.
We propose a mixed deep neural network strategy, incorporating parallel combination of Convolutional (CNN) and Recurrent Neural Networks (RNN), cascaded with deep autoencoders and fully connected layers towards automatic identification of imagined speech from EEG. Instead of utilizing raw EEG channel data, we compute t…
Harer-Zagier formulas generalized to knot matrix models.
The counting grid is a grid of microtopics, sparse word/feature distributions. The generative model associated with the grid does not use these microtopics individually. Rather, it groups them in overlapping rectangular windows and uses these grouped microtopics as either mixture or admixture components. This paper bui…
We present a structural clustering algorithm for large-scale datasets of small labeled graphs, utilizing a frequent subgraph sampling strategy. A set of representatives provides an intuitive description of each cluster, supports the clustering process, and helps to interpret the clustering results. The projection-based…
We introduce topological parallelisms of oriented lines (briefly called oriented parallelisms). Every topological parallelism (of lines) on PG(3,R) gives rise to a parallelism of oriented lines, but we show that even the most homogeneous parallelisms of oriented lines other than the Clifford parallelism do not necessar…
Parallelizes MCTS for continuous domains using leaf and root parallelization.
A new data-level recombination strategy improves RGB-D salient object detection.
The paper parallelizes HMM inference for efficient long-term computations.
The scale of functional magnetic resonance image data is rapidly increasing as large multi-subject datasets are becoming widely available and high-resolution scanners are adopted. The inherent low-dimensionality of the information in this data has led neuroscientists to consider factor analysis methods to extract and a…
UniPhyNet improves cognitive load classification accuracy using EEG, ECG, and EDA signals.
New rational parallelisms found on complex manifolds that are not flat.
Introduces a natural parallel translation for navigation data.
We prove a conjecture formulated by Pablo M. Chacon and Guillermo A. Lobos in [Pseudo-parallel Lagrangian submanifolds in complex space forms, Differential Geom. Appl.] stating that every Lagrangian pseudo-parallel submanifold of a complex space form of dimension at least 3 is semi-parallel.
This study compares parallel SMC and MCMC for Bayesian deep learning, showing SMC parallel is faster.
The paper explores parallel 1-forms on special Finsler manifolds and their properties.
This paper surveys parallel submanifolds in Riemannian and pseudo-Riemannian manifolds.
We propose a new integrated method of exploiting model, batch and domain parallelism for the training of deep neural networks (DNNs) on large distributed-memory computers using minibatch stochastic gradient descent (SGD). Our goal is to find an efficient parallelization strategy for a fixed batch size using process…
Characterizes regular parallelisms in 3D space with 2-torus action.
We propose a nonparallel data-driven emotional speech conversion method. It enables the transfer of emotion-related characteristics of a speech signal while preserving the speaker's identity and linguistic content. Most existing approaches require parallel data and time alignment, which is not available in most real ap…
Paper studies second order symmetric parallel tensors in generalized f.pk-space forms.
Classifies simply-connected pluriclosed manifolds with parallel Bismut torsion.
Study on generalized ξ-parallel maps in Riemannian geometry.
This paper improves parallel belief propagation for scalable machine learning.
Cyclic Data Parallelism reduces memory usage and balances gradient communications.
Betten and Riesinger constructed Parallelisms of with automorphism group by applying the reducible -action to a rotational Betten spread. This was generalized by the present author so as to include oriented parallelisms (i.e., p…