A graph model improves short text classification by integrating sentence relationships.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Recent approaches based on artificial neural networks (ANNs) have shown promising results for short-text classification. However, many short texts occur in sequences (e.g., sentences in a document or utterances in a dialog), and most existing ANN-based systems do not leverage the preceding short texts when classifying …
BBM models short texts using biterms to improve coherence.
Paper introduces new indicators for forecasting crude oil prices using short news headlines.
POTA improves short text clustering by generating reliable pseudo-labels.
Linking authors of short-text contents has important usages in many applications, including Named Entity Recognition (NER) and human community detection. However, certain challenges lie ahead. Firstly, the input short-text contents are noisy, ambiguous, and do not follow the grammatical rules. Secondly, traditional tex…
AOBTM adapts online topic modeling for short app reviews, revealing coherent topics over time.
Paper improves short text clustering by integrating semantic relationships into Optimal Transport.
We give a short proof of a theorem of Handel and Mosher stating that any finitely generated subgroup of either contains a fully irreducible automorphism, or virtually fixes the conjugacy class of a proper free factor of , and we extend their result to non finitely generated subgroups of $\text{Ou…
New Gamma-Poisson model improves topic selection for short text.
Paper introduces a new text clustering model using Beta-Liouville priors.
Study examines how cluster number affects short-text clustering, introducing a stability metric.
Model clusters authors and topics in short texts like social media posts.
As the emergence and the thriving development of social networks, a huge number of short texts are accumulated and need to be processed. Inferring latent topics of collected short texts is useful for understanding its hidden structure and predicting new contents. Unlike conventional topic models such as latent Dirichle…
System identifies language of transliterated text.
In this paper we consider the problem of clustering collections of very short texts using subspace clustering. This problem arises in many applications such as product categorisation, fraud detection, and sentiment analysis. The main challenge lies in the fact that the vectorial representation of short texts is both hi…
Improves video search by balancing text and visual modalities.
SCROLLS benchmarks long text NLP tasks, improving existing models.
The increasing volume of short texts generated on social media sites, such as Twitter or Facebook, creates a great demand for effective and efficient topic modeling approaches. While latent Dirichlet allocation (LDA) can be applied, it is not optimal due to its weakness in handling short texts with fast-changing topics…
We apply text analysis approaches for a specialized search engine for 3D CAD models and associated products. The main goals are to distinguish between actual product descriptions and other text on a website, as well as to decide whether a given text is or contains a product name. For this we use paragraph vectors for t…
A latent-variable model is introduced for text matching, inferring sentence representations by jointly optimizing generative and discriminative objectives. To alleviate typical optimization challenges in latent-variable models for text, we employ deconvolutional networks as the sequence decoder (generator), providing l…
Interventional cancer clinical trials are generally too restrictive, and some patients are often excluded on the basis of comorbidity, past or concomitant treatments, or the fact that they are over a certain age. The efficacy and safety of new treatments for patients with these characteristics are, therefore, not defin…
RNNs are crucial for text and speech tasks, explained in this overview.
Social media is increasingly used by humans to express their feelings and opinions in the form of short text messages. Detecting sentiments in the text has a wide range of applications including identifying anxiety or depression of individuals and measuring well-being or mood of a community. Sentiments can be expressed…
Let us denote by the hyperspace of all convex bodies of equipped with the Hausdorff distance topology. An affine invariant point is a continuous and Aff(n)-equivariant map , where Aff(n) denotes the group of all nonsingular affine maps of . Fo…
Proposes a flexible neural recommendation framework for better prediction performance.
We consider various notions of strains; quantitative measures for the deviation of a linear transformation from an isometry. The main approach, which is motivated by physical applications and follows the work of Patrizio Neff and co-workers , is to select a Riemannian metric on , and use its induced geodes…
We give a short proof of Masbaum and Reid's result that mapping class groups involve any finite group, appealing to free quotients of surface groups and a result of Gilman, following Dunfield-Thurston.
Many methods have been used to recognize author personality traits from text, typically combining linguistic feature engineering with shallow learning models, e.g. linear regression or Support Vector Machines. This work uses deep-learning-based models and atomic features of text, the characters, to build hierarchical, …
This is a short, elementary survey article about taut submanifolds. In order to simplify the exposition, we restrict to the case of compact smooth submanifolds of Euclidean or spherical spaces. Some new, partial results concerning taut 4-manifolds are discussed at the end of the text.
The Generative Adversarial Network (GAN) has achieved great success in generating realistic (real-valued) synthetic data. However, convergence issues and difficulties dealing with discrete data hinder the applicability of GAN to text. We propose a framework for generating realistic text via adversarial training. We emp…
Study shows text-based news veracity models don't generalize across U.S. and U.K.
We give a short proof of a theorem of Guth relating volume of balls and Uryson width. The same approach applies to Hausdorff content implying a recent result of Liokumovich-Lishak-Nabutovsky-Rotman. We show also that for any there is a Riemannian metric on a 3-sphere such that and for an…
Study proves short-term existence for harmonic maps under evolving metrics.
Shorter adversarial prompts help protect LLMs from jailbreak attacks.
We present PubMed 200k RCT, a new dataset based on PubMed for sequential sentence classification. The dataset consists of approximately 200,000 abstracts of randomized controlled trials, totaling 2.3 million sentences. Each sentence of each abstract is labeled with their role in the abstract using one of the following …
One-hot CNN (convolutional neural network) has been shown to be effective for text categorization (Johnson & Zhang, 2015). We view it as a special case of a general framework which jointly trains a linear model with a non-linear feature generator consisting of `text region embedding + pooling'. Under this framework, we…
In this paper, we prove that there exists a dimensional constant such that given any background Kähler metric , the Calabi flow with initial data satisfying \begin{equation*} \partial \bar \partial u_0 \in L^\infty (M) \text{ and } (1- δ)ω< ω_{u_0} < (1+δ)ω, \end{equation*} admits a unique short time so…
A new neural topic model using optimal transport improves document representation and topic coherence.
Dual-attention GCN improves text classification by adapting to textual complexity.
Paper proposes deep learning for fake claim detection on social media.
Study proposes a multimodal model for cardiovascular risk prediction using EHRs.
Entity linking is the task of mapping potentially ambiguous terms in text to their constituent entities in a knowledge base like Wikipedia. This is useful for organizing content, extracting structured data from textual documents, and in machine learning relevance applications like semantic search, knowledge graph const…
Introduces basics of supergeometry for PhD students.
Study enhances cryptocurrency sentiment analysis using TikTok and Twitter data.
AudioPaLM combines text and speech models to improve speech processing and translation.
Framework incorporates prior knowledge into Bayesian models for data streams.
Ontology learning is a critical task in industry, dealing with identifying and extracting concepts captured in text data such that these concepts can be used in different tasks, e.g. information retrieval. Ontology learning is non-trivial due to several reasons with limited amount of prior research work that automatica…