Improves medical note processing by training model on related concepts and global context.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
In supervised machine learning for author name disambiguation, negative training data are often dominantly larger than positive training data. This paper examines how the ratios of negative to positive training data can affect the performance of machine learning algorithms to disambiguate author names in bibliographic …
EviTrack improves sequential prediction in delayed disambiguation scenarios.
Paper proposes an algorithm to recover full supervision from weakly labeled data.
This paper describes a conditional neural network architecture for Mandarin Chinese polyphone disambiguation. The system is composed of a bidirectional recurrent neural network component acting as a sentence encoder to accumulate the context correlations, followed by a prediction network that maps the polyphonic charac…
A new method for name disambiguation in academic networks using multi-view attention and recurrent neural networks.
A coloring scheme improves graph neural networks for node disambiguation.
Author name disambiguation in bibliographic databases is the problem of grouping together scientific publications written by the same person, accounting for potential homonyms and/or synonyms. Among solutions to this problem, digital libraries are increasingly offering tools for authors to manually curate their publica…
We consider a scenario where an artificial agent is reading a stream of text composed of a set of narrations, and it is informed about the identity of some of the individuals that are mentioned in the text portion that is currently being read. The agent is expected to learn to follow the narrations, thus disambiguating…
We address the problem of disambiguating large scale catalogs through the definition of an unknown artist clustering task. We explore the use of metric learning techniques to learn artist embeddings directly from audio, and using a dedicated homonym artists dataset, we compare our method with a recent approach that lea…
We present a new local entity disambiguation system. The key to our system is a novel approach for learning entity representations. In our approach we learn an entity aware extension of Embedding for Language Model (ELMo) which we call Entity-ELMo (E-ELMo). Given a paragraph containing one or more named entity mentions…
DivDis learns diverse hypotheses from underspecified data to improve robustness.
In this paper, we first develope the concept of Lyapunov graph to weighted Lyapunov graph (abbreviated as WLG) for nonsingular Morse-Smale flows (abbreviated as NMS flows) on . WLG is quite sensitive to NMS flows on . For instance, WLG detect the indexed links of NMS flows. Then we use WLG and some other tool…
The translation is not verbatim, many parts have been abbreviated and in some case alternative proofs were devised emphasizing intuition.
Vector representations of words have heralded a transformational approach to classical problems in NLP; the most popular example is word2vec. However, a single vector does not suffice to model the polysemous nature of many (frequent) words, i.e., words with multiple meanings. In this paper, we propose a three-fold appr…
Sparse-mode DMD disambiguates local and global modes in spatiotemporal data.
Following a particular news story online is an important but difficult task, as the relevant information is often scattered across different domains/sources (e.g., news articles, blogs, comments, tweets), presented in various formats and language styles, and may overlap with thousands of other stories. In this work we …
This work addresses the problem of author name homonymy in the Web of Science. Aiming for an efficient, simple and straightforward solution, we introduce a novel probabilistic similarity measure for author name disambiguation based on feature overlap. Using the researcher-ID available for a subset of the Web of Science…
Entity linking is the task of mapping potentially ambiguous terms in text to their constituent entities in a knowledge base like Wikipedia. This is useful for organizing content, extracting structured data from textual documents, and in machine learning relevance applications like semantic search, knowledge graph const…
Efficient autoregressive entity linking with correction for faster, more accurate results.
This paper is devoted to discussing affine Hirsch foliations on -manifolds. First, we prove that up to isotopic leaf-conjugacy, every closed orientable -manifold admits , or affine Hirsch foliations. Furthermore, every case is possible. Then, we analyze the -manifolds admitting two affine Hirsch…
FONDUE identifies ambiguous nodes in networks for better analysis.
Improves biomedical entity linking with latent type modeling.
In this paper, we present our method of using fixed-size ordinally forgetting encoding (FOFE) to solve the word sense disambiguation (WSD) problem. FOFE enables us to encode variable-length sequence of words into a theoretically unique fixed-size representation that can be fed into a feed forward neural network (FFNN),…
Partial multi-label learning (PML), which tackles the problem of learning multi-label prediction models from instances with overcomplete noisy annotations, has recently started gaining attention from the research community. In this paper, we propose a novel adversarial learning model, PML-GAN, under a generalized encod…
Due to recent technical and scientific advances, we have a wealth of information hidden in unstructured text data such as offline/online narratives, research articles, and clinical reports. To mine these data properly, attributable to their innate ambiguity, a Word Sense Disambiguation (WSD) algorithm can avoid numbers…
One of the major problems in natural language processing (NLP) is the word sense disambiguation (WSD) problem. It is the task of computationally identifying the right sense of a polysemous word based on its context. Resolving the WSD problem boosts the accuracy of many NLP focused algorithms such as text classification…
New solver for MKL-SVM with 0/1 loss function.
Active inference selects actions to maximize information gain, aiding structure learning.
AI beats 95% of humans in Rock-Paper-Scissors.
Researchers tackle the globalization problem of locally cosymplectic Hamiltonian dynamics.
AXE evaluates explanations to avoid misleading Rashomon set model selection.
We present an LDA approach to entity disambiguation. Each topic is associated with a Wikipedia article and topics generate either content words or entity mentions. Training such models is challenging because of the topic and vocabulary size, both in the millions. We tackle these problems using a novel distributed infer…
Abstracts discuss a common framework for constructing homology theories.
Paper introduces MKL--SVM for SVM with loss.
Paper introduces a modified Allen-Cahn equation for better energy equipartition.
This article is the second part of the article we promised to write at the end of Section 1 of [FOOO15] (arXiv:1209.4410). (Part I appeared in [Part I] (arXiv:1503.07631).) We discuss the foundation of the virtual fundamental chain and cycle technique, especially its version that appeared in [FOn] and also in Section A…
Study of ACM(3)S structures on manifolds with G2 structure.
The present paper develops a novel aggregated gradient approach for distributed machine learning that adaptively compresses the gradient communication. The key idea is to first quantize the computed gradients, and then skip less informative quantized gradient communications by reusing outdated gradients. Quantizing and…
We describe our language-independent unsupervised word sense induction system. This system only uses topic features to cluster different word senses in their global context topic space. Using unlabeled data, this system trains a latent Dirichlet allocation (LDA) topic model then uses it to infer the topics distribution…
The paper computes KV cochain differentials and their geometric implications.
Paper provides a rigorous proof of the index theorem for economists.
We define an algebraic/combinatorial object on the front projection of a Legendrian knot called a Morse complex sequence, abbreviated MCS. This object is motivated by the theory of generating families and provides new connections between generating families, normal rulings, and augmentations of the Chekanov-Eliashb…
In this work we introduce a mixture of GPs to address the data association problem, i.e. to label a group of observations according to the sources that generated them. Unlike several previously proposed GP mixtures, the novel mixture has the distinct characteristic of using no gating function to determine the associati…
Transparency, user trust, and human comprehension are popular ethical motivations for interpretable machine learning. In support of these goals, researchers evaluate model explanation performance using humans and real world applications. This alone presents a challenge in many areas of artificial intelligence. In this …
A new knot move preserves pass-move equivalence and differs in count.
We study the projected gradient descent method on low-rank matrix problems with a strongly convex objective. We use the Burer-Monteiro factorization approach to implicitly enforce low-rankness; such factorization introduces non-convexity in the objective. We focus on constraint sets that include both positive semi-defi…
Clinical notes in electronic health records contain highly heterogeneous writing styles, including non-standard terminology or abbreviations. Using these notes in predictive modeling has traditionally required preprocessing (e.g. taking frequent terms or topic modeling) that removes much of the richness of the source d…