New method identifies valid IVs for bi-directional MR with invalid instruments.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Improved Q&A model with LSTM and bi-directional attention.
BICompFL tackles bi-directional compression challenges in stochastic FL, reducing communication costs by an order of magnitude.
Graphical models with bi-directed edges (<->) represent marginal independence: the absence of an edge between two vertices indicates that the corresponding variables are marginally independent. In this paper, we consider maximum likelihood estimation in the case of continuous variables with a Gaussian joint distributio…
Novel technique detects adversarial samples in face recognition models.
New risk class penalizes loss deviations from mean on both sides.
The covariance graph (aka bi-directed graph) of a probability distribution is the undirected graph where two nodes are adjacent iff their corresponding random variables are marginally dependent in . In this paper, we present a graphical criterion for reading dependencies from , under the assumption that $…
DoCoFL compresses model updates for cross-device federated learning.
We discuss two parameterizations of models for marginal independencies for discrete distributions which are representable by bi-directed graph models, under the global Markov property. Such models are useful data analytic tools especially if used in combination with other graphical models. The first parameterization, i…
Bi-directional Curriculum Learning improves graph anomaly detection by considering both homogeneity and heterogeneity.
By establishing a connection between bi-directional Helmholtz machines and information theory, we propose a generalized Helmholtz machine. Theoretical and experimental results show that given \textit{shallow} architectures, the generalized model outperforms the previous ones substantially.
PML-GAN tackles noisy multi-label annotations using adversarial learning.
This work compares NN architectures for spectrum sensing.
BiHRNN predicts inflation by leveraging hierarchical structure and bidirectional RNNs.
We investigate deep generative models that can exchange multiple modalities bi-directionally, e.g., generating images from corresponding texts and vice versa. A major approach to achieve this objective is to train a model that integrates all the information of different modalities into a joint representation and then t…
FedSGM tackles constrained federated learning with unified framework.
I-BERT extends Transformer's self-attention to arbitrary input lengths.
Improved NER performance on imbalanced data.
Bi-directional LSTMs are a powerful tool for text representation. On the other hand, they have been shown to suffer various limitations due to their sequential nature. We investigate an alternative LSTM structure for encoding text, which consists of a parallel state for each word. Recurrent steps are used to perform lo…
Confidentiality of patient information is an essential part of Electronic Health Record System. Patient information, if exposed, can cause a serious damage to the privacy of individuals receiving healthcare. Hence it is important to remove such details from physician notes. A system is proposed which consists of a deep…
Study neural networks learning from noisy examples via reverberation.
We replace the Hidden Markov Model (HMM) which is traditionally used in in continuous speech recognition with a bi-directional recurrent neural network encoder coupled to a recurrent neural network decoder that directly emits a stream of phonemes. The alignment between the input and output sequences is established usin…
Recently, a technique called Layer-wise Relevance Propagation (LRP) was shown to deliver insightful explanations in the form of input space relevances for understanding feed-forward neural network classification decisions. In the present work, we extend the usage of LRP to recurrent neural networks. We propose a specif…
The work investigates deep generative models, which allow us to use training data from one domain to build a model for another domain. We propose the Variational Bi-domain Triplet Autoencoder (VBTA) that learns a joint distribution of objects from different domains. We extend the VBTAs objective function by the relativ…
AER combines auto-encoder and LSTM for better time series anomaly detection.
A new method predicts protein functions using variable-length sequences.
Non-autoregressive transformer improves speech recognition speed and accuracy.
Project classifies Hinglish social content on platforms like Twitter, Reddit.
Proposes a stochastic model for South African actuarial use.
Bi-LSTM with attention generates jazz music with rich nuances.
Bi-GAN model for imputing and predicting irregular time-series data.
Motivated by the need to automate medical information extraction from free-text radiological reports, we present a bi-directional long short-term memory (BiLSTM) neural network architecture for modelling radiological language. The model has been used to address two NLP tasks: medical named-entity recognition (NER) and …
In this paper, we propose a probabilistic parsing model, which defines a proper conditional probability distribution over non-projective dependency trees for a given sentence, using neural representations as inputs. The neural network architecture is based on bi-directional LSTM-CNNs which benefits from both word- and …
State-of-the-art sequence labeling systems traditionally require large amounts of task-specific knowledge in the form of hand-crafted features and data pre-processing. In this paper, we introduce a novel neutral network architecture that benefits from both word- and character-level representations automatically, by usi…
We present our first efforts in building an automatic speech recognition system for Somali, an under-resourced language, using 1.57 hrs of annotated speech for acoustic model training. The system is part of an ongoing effort by the United Nations (UN) to implement keyword spotting systems supporting humanitarian relief…
Study tackles ranking fraud in online platforms by learning robust rankings.
Knots can be constructed and decomposed using Murasugi sums of Seifert surfaces.
New framework using Jensen-Shannon divergence improves domain adaptation theory.
ALAD uses GANs to detect anomalies in complex data.
The Gaussian process state space model (GPSSM) is a non-linear dynamical system, where unknown transition and/or measurement mappings are described by GPs. Most research in GPSSMs has focussed on the state estimation problem, i.e., computing a posterior of the latent state given the model. However, the key challenge in…
Grammatical error correction, like other machine learning tasks, greatly benefits from large quantities of high quality training data, which is typically expensive to produce. While writing a program to automatically generate realistic grammatical errors would be difficult, one could learn the distribution of naturally…
CCHM algorithm learns BN structure with latent variables, improving causal effect measurement.
A comparison of SLDS and LSTM for pedestrian behavior prediction shows SLDS works better with shorter sequences.
Adversarial Reprogramming has demonstrated success in utilizing pre-trained neural network classifiers for alternative classification tasks without modification to the original network. An adversary in such an attack scenario trains an additive contribution to the inputs to repurpose the neural network for the new clas…
BiSHop tackles tabular data challenges with sparse Hopfield layers.
Navigated 2D multi-slice dynamic Magnetic Resonance (MR) imaging enables high contrast 4D MR imaging during free breathing and provides in-vivo observations for treatment planning and guidance. Navigator slices are vital for retrospective stacking of 2D data slices in this method. However, they also prolong the acquisi…
Several methods exist to infer causal networks from massive volumes of observational data. However, almost all existing methods require a considerable length of time series data to capture cause and effect relationships. In contrast, memory-less transition networks or Markov Chain data, which refers to one-step transit…
Framework generates pop song melodies and piano accompaniment.