Paraphrasing exemplifies the ability to abstract semantic content from surface forms. Recent work on automatic paraphrasing is dominated by methods leveraging Machine Translation (MT) as an intermediate step. This contrasts with humans, who can paraphrase without being bilingual. This work proposes to learn paraphrasin…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Generating paraphrases that are lexically similar but semantically different is a challenging task. Paraphrases of this form can be used to augment data sets for various NLP tasks such as machine reading comprehension and question answering with non-trivial negative examples. In this article, we propose a deep variatio…
MARGE learns to reconstruct text by paraphrasing, achieving strong performance across multiple tasks.
Semantic paraphrases can fool financial sentiment classifiers due to geometric shifts in model representations.
TGLS generates text by optimizing search results and learning from them.
We give a explicit computation of the pointed harmonic volumes of hyperelliptic curves with Weierstrass base points, which are paraphrased into a combinatorial formula.
The following paper presents a method of comparing two sets of vectors. The method can be applied in all tasks, where it is necessary to measure the closeness of two objects presented as sets of vectors. It may be applicable when we compare the meanings of two sentences as part of the problem of paraphrasing. This is t…
Recently, generating adversarial examples has become an important means of measuring robustness of a deep learning model. Adversarial examples help us identify the susceptibilities of the model and further counter those vulnerabilities by applying adversarial training techniques. In natural language domain, small pertu…
Word embeddings generated by neural network methods such as word2vec (W2V) are well known to exhibit seemingly linear behaviour, e.g. the embeddings of analogy "woman is to queen as man is to king" approximately describe a parallelogram. This property is particularly intriguing since the embeddings are not trained to a…
New approach detects sensitive info in text, outperforming previous methods.
We present a syntax-infused variational autoencoder (SIVAE), that integrates sentences with their syntactic trees to improve the grammar of generated sentences. Distinct from existing VAE-based text generative models, SIVAE contains two separate latent spaces, for sentences and syntactic trees. The evidence lower bound…
Word2vec (Mikolov et al., 2013) has proven to be successful in natural language processing by capturing the semantic relationships between different words. Built on top of single-word embeddings, paragraph vectors (Le and Mikolov, 2014) find fixed-length representations for pieces of text with arbitrary lengths, such a…
In real-world applications of natural language generation, there are often constraints on the target sentences in addition to fluency and naturalness requirements. Existing language generation techniques are usually based on recurrent neural networks (RNNs). However, it is non-trivial to impose constraints on RNNs whil…
Modeling human language learning with multi-checkpoint machine translation.
Recently, generative adversarial networks have gained a lot of popularity for image generation tasks. However, such models are associated with complex learning mechanisms and demand very large relevant datasets. This work borrows concepts from image and video captioning models to form an image generative framework. The…
Convolutional neural networks (CNNs) have recently emerged as a popular building block for natural language processing (NLP). Despite their success, most existing CNN models employed in NLP share the same learned (and static) set of filters for all input sentences. In this paper, we consider an approach of using a smal…
Word2Vec (W2V) and GloVe are popular, fast and efficient word embedding algorithms. Their embeddings are widely used and perform well on a variety of natural language processing tasks. Moreover, W2V has recently been adopted in the field of graph embedding, where it underpins several leading algorithms. However, despit…
Recurrent neural networks have become ubiquitous in computing representations of sequential data, especially textual data in natural language processing. In particular, Bidirectional LSTMs are at the heart of several neural models achieving state-of-the-art performance in a wide variety of tasks in NLP. However, BiLSTM…
Adversarial examples are carefully constructed modifications to an input that completely change the output of a classifier but are imperceptible to humans. Despite these successful attacks for continuous data (such as image and audio samples), generating adversarial examples for discrete structures such as text has pro…
DeepSubQE estimates translation quality for subtitles, improving on existing methods.
We aim to design strategies for sequential decision making that adjust to the difficulty of the learning problem. We study this question both in the setting of prediction with expert advice, and for more general combinatorial decision tasks. We are not satisfied with just guaranteeing minimax regret rates, but we want …
This paper presents an unsupervised method to learn a neural network, namely an explainer, to interpret a pre-trained convolutional neural network (CNN), i.e., the explainer uses interpretable visual concepts to explain features in middle conv-layers of a CNN. Given feature maps of a conv-layer of the CNN, the explaine…
NMIXX fine-tunes embeddings for finance, outperforming general models in Korean.
A new framework explains GNN predictions by simulating graph structure and feature changes.
Study on estimating Gumbel--Max watermark proportions in edited documents.
The semantic map calibrates uncertainty from language model probabilities.
Ultra-fast search algorithm for trillion-scale corpora with semantic flexibility.
Many tasks in natural language understanding require learning relationships between two sequences for various tasks such as natural language inference, paraphrasing and entailment. These aforementioned tasks are similar in nature, yet they are often modeled individually. Knowledge transfer can be effective for closely …
Paper introduces SDM for detecting LLM hallucinations, improving on entropy tests.
WISER detects watermarked segments in text via epidemic change-point analysis.
Online monitoring system for safety classifiers with shift detection and conformal adaptation
LFD method improves text classification by making features clearer and less label-leaking.
In this on-going work, I explore certain theoretical and empirical implications of data transformations under the PCA. In particular, I state and prove three theorems about PCA, which I paraphrase as follows: 1). PCA without discarding eigenvector rows is injective, but looses this injectivity when eigenvector rows are…
Paper defines generalized braids and proves their subgroup status.
Defines a new Poisson structure for generalized Sasakian spaces.
Improved image generation through iterative flow matching to reduce hallucinations.
Framework generates personalized insulin treatment strategies using deep models.
OptiGAN uses GAN and RL to optimize sequence generation for specific goals.
Survey on deep models for graph generation.
Improves deep generative models to generate images of any size.
Develops a unified theory of Yang-Mills and GR using generalized principal bundles.
Meta-CoTGAN improves adversarial text generation by preventing mode collapse.
Generative models can still learn from contaminated data, but with limitations.
Generative AI tasks analyzed for text, images, audio, video, code, and molecules.
Defines Kahler angle for a broader context.
Established a generalized Boothby-Wang theorem in contact geometry.
The twistor construction for Riemannian manifolds is extended to the case of manifolds endowed with generalized metrics (in the sense of generalized geometry à la Hitchin). The generalized twistor space associated to such a manifold is defined as the bundle of generalized complex structures on the tangent spaces of the…
We present a characterization, in terms of torsion-free generalized connections, for the integrability of various generalized structures (generalized almost complex structures, generalized almost hypercomplex structures, generalized almost Hermitian structures and generalized almost hyper-Hermitian structures) defined …