Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

12.5%25.0%37.5%50.0% · Sep 199319922001200920182026
48 results for word identities

This paper assesses biases in contextualized word representations.

problem Analyzing biases in contextualized word representations.
method Proposes assessing bias at the contextual word level, capturing contextual effects of bias.
result Demonstrates evidence of bias in contextual word models, including racial bias and exacerbated effects for intersectional minorities.

We bound the value of the Casson invariant of any integral homology 3-sphere MM by a constant times the distance-squared to the identity, measured in any word metric on the Torelli group $\T$, of the element of $\T$ associated to any Heegaard splitting of MM. We construct examples which show this bound is asymptotica…

2007-07-16abs ↗pdf ↗

Study polynomial trace identities in $SL(2,\IC)$ using quaternion algebras.

problem Understanding polynomial trace identities in $SL(2,\IC)$ and their applications.
method Use quaternion algebras over indefinites and their units to study discrete subgroups of $SL(2,\IC)$.
result Obtained structure theorems for quaternion algebras and new polynomial trace identities.

New post-processing methods improve word embedding performance.

problem Boosting the performance of word embeddings for similarity and analogy tasks.
method Optimizing a semi-Riemannian manifold with Centralised Kernel Alignment (CKA) to shrink the covariance matrix towards a scaled identity matrix.
result Improved performance on downstream tasks after smoothing the spectrum of word vectors.

We show that for any given n, there exists a sequence of words a_k in the generators sigma_1, ... sigma_{n-1} of the braid group B_n, representing the identity element of B_n, such that the number of braid relations of the form sigma_i sigma_{i+1} sigma_i = sigma_{i+1} sigma_i sigma_{i+1} needed to pass from a_k to the…

2009-05-31abs ↗pdf ↗

We show that every co--orientable taut foliation F of an orientable, atoroidal 3-manifold admits a transverse essential lamination. If this transverse lamination is a foliation G, the pair F,G are the unstable and stable foliation respectively of an Anosov flow. Otherwise, F admits a pair of transverse very full genuin…

2002-10-10abs ↗pdf ↗

In this paper we prove that the space of flat metrics (nonpositively curved Euclidean cone metrics) on a closed, oriented surface is marked length spectrally rigid. In other words, two flat metrics assigning the same lengths to all closed curves differ by an isometry isotopic to the identity. The novel proof suggests a…

2015-04-05abs ↗pdf ↗

We show that one can skip the skew-symmetry assumption in the definition of Nambu-Poisson brackets. In other words, a n-ary bracket on the algebra of smooth functions which satisfies the Leibniz rule and a n-ary version of the Jacobi identity must be skew-symmetric. A similar result holds for a non-antisymmetric versio…

2001-04-11abs ↗pdf ↗

Unsupervised MT struggles with morphologically rich languages.

problem Limitations of unsupervised machine translation on morphologically rich languages.
method Adversarial unsupervised alignment of word embedding spaces for bilingual dictionary induction.
result A simple trick exploiting weak supervision from identical words improves unsupervised bilingual dictionary induction performance.

Improved neural keyphrase generation by beam search with reward functions.

problem Sequence length bias and beam diversity issues in neural keyphrase generation.
method Beam search decoding strategy with word-level and ngram-level reward functions.
result Significant improvement in generating diverse and accurate keyphrases.

Seq-CVAE learns a latent space for each word position to capture sentence intention.

problem Capturing diversity in image captioning models.
method Seq-CVAE learns a sequential latent space for each word position, mimicking future sentence summaries.
result Significantly improves diversity metrics on MSCOCO dataset compared to baselines.

Mitigates bias in text classification by weighting instances.

problem Unintended biases in text classification datasets based on demographic terms.
method Instance weighting to recover non-discrimination distribution.
result Effective mitigation of unintended biases without sacrificing generalization.

Given a principal GG-bundle PMP \to M and two C1C^1 curves in MM with coinciding endpoints, we say that the two curves are holonomically equivalent if the parallel transport along them is identical for any smooth connection on PP. The main result in this paper is that if GG is semi-simple, then the two curves are h…

2013-11-26abs ↗pdf ↗

A new method selects anchor words for better topic discovery in text corpora.

problem Selecting anchor words for improved topic modeling in text corpora.
method Proposes a new greedy method to find a minimum edge-weight anchor clique in a word similarity graph.
result The proposed method outperforms existing methods on topic quality and is faster.

There are certain families of words and word sequences (words in the generators of a two-generator group) that arise frequently in the Teichm{ü}ller theory of hyperbolic three-manifolds and Kleinian and Fuchsian groups and in the discreteness problem for two generator matrix groups. We survey some of the families of su…

2007-01-20abs ↗pdf ↗

Probabilistic FastText captures multiple word senses and sub-word structures.

problem Capturing multiple word senses and sub-word structures in word embeddings.
method Probabilistic FastText uses Gaussian mixture densities to represent words, sharing statistical strength across sub-word structures and capturing different word senses.
result Probabilistic FastText outperforms existing models on word-similarity benchmarks and discerning different meanings.

We prove that two countable locally finite-by-abelian groups G,H endowed with proper left-invariant metrics are coarsely equivalent if and only if their asymptotic dimensions coincide and the groups are either both finitely-generated or both are infinitely generated. On the other hand, we show that each countable group…

2008-07-07abs ↗pdf ↗

We discuss a topological approach to words introduced by the author. Words on an arbitrary alphabet are approximated by Gauss words and then studied up to natural modifications inspired by the Reidemeister moves on knot diagrams. This leads us to a notion of homotopy for words. We introduce several homotopy invariants …

2006-09-19abs ↗pdf ↗

Random groups prove length constraints on product of conjugates.

problem Quantify products of conjugates in random groups.
method Sharp van Kampen diagram argument and boundary block-counting.
result Prove a sharp inequality for products of conjugates in random groups.

Discriminative model identifies readers and assesses comprehension from eye movements.

problem Inferring readers' identities and estimating their text comprehension from eye movements.
method Generative model of gaze patterns, Fisher-score representation, Fisher-SVM with Fisher kernel.
result SVM with Fisher kernel excels at identifying readers, but not comprehending text.

The abstract explains how word and relation representations capture semantic meaning.

problem Understanding how word and relation representations capture semantic meaning.
method Theoretical justification and extension of geometric relationships between word embeddings and knowledge graph representations.
result The geometric relationships between word embeddings correspond to semantic relations between words and entities in knowledge graphs.

Approaches KL divergence for learning multi-sense word distributions.

problem Capturing the polysemy and uncertainty of words in word embeddings.
method Modeling words as multi-sense Gaussian mixtures and using KL divergence for learning.
result The proposed approach effectively captures word entailment and distribution similarity.

Word2vec improved but lacks multi-meaning words; ConEc creates new embeddings.

problem Lack of meaningful embeddings for words with multiple meanings and OOV words.
method Context encoders (ConEc) extend word2vec by multiplying embeddings with context vectors.
result ConEc creates embeddings for OOV words and words with multiple meanings based on local contexts.

Proposes MorphMine for unsupervised morpheme segmentation to improve word embeddings.

problem Lack of semantic information in word-level analysis for infrequent and out-of-vocabulary words.
method MorphMine applies a parsimony criterion to hierarchically segment words into the fewest number of morphemes.
result MorphMine segments words into human-verified morphemes and improves word embedding quality.

End-to-end ASR model combines word and character representation for improved performance.

problem Difficulty in training with word-level supervision due to sparsity of examples.
method Multi-task learning framework combining word and character representations.
result Improved word-error rate (WER) by interpolating between word-level and character-level models.

Paper analyzes word embedding composition using tensor decomposition.

problem Given vector representations of two words, compute a vector for the entire phrase.
method Generative model with low rank Tucker decomposition of word embedding correlations.
result Word embeddings and a core tensor can be derived from the Tucker decomposition.