Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

4896143191 · Jun 202019922001200920182026
48 results for text normalization

Neural model improves text normalization for non-English languages.

problem Improving text normalization in non-English languages with limited data.
method Sequence-to-sequence model with character and word embeddings, using pre-trained word embeddings with subword information.
result Achieved state-of-the-art F1 score on Arabic language correction dataset.

Improves text clustering by incorporating sequential features and word embeddings.

problem Lack of sequential information and synonym handling in current text clustering methods.
method SiDPMM model that models documents as joint of bags of words, sequential features, and word embeddings.
result Significant improvement in performance and accurate inference of cluster numbers.

FlowGMM uses normalizing flows for semi-supervised learning, showing promising results across various data types.

problem Semi-supervised learning with limited labeled data.
method Normalizing flows combined with latent Gaussian mixture models for generative modeling.
result FlowGMM achieves promising results on multiple data types, including text and tabular data.

A new model uses normalizing flows for discrete sequences, improving generation speed.

problem Modeling discrete sequences like text using normalizing flows poses challenges.
method Proposes a VAE-based model with autoregressive and non-autoregressive flow architectures.
result Flow-based models can match or improve on autoregressive baselines for discrete sequence tasks.

We propose a theoretical framework for thinking about score normalization, which confirms that normalization is not needed under (admittedly fragile) ideal conditions. If, however, these conditions are not met, e.g. under data-set shift between training and runtime, our theory reveals dependencies between scores that c…

2017-09-28abs ↗pdf ↗

Guided Flows enhance sample quality in conditional image generation and text-to-speech.

problem Improving sample quality in conditional generative models.
method Integrating classifier-free guidance into Flow Matching (FM) models for Continuous Normalizing Flows (CNFs).
result Guided Flows significantly improve sample quality in conditional image generation and text-to-speech synthesis.

This is an expanded version of the lecture notes for a minicourse that I gave at a summer school called "Advanced Course on Geometry and Dynamics of Integrable Systems" at CRM Barcelona, 9--14/September/2013. In this text we study the following aspects of integrable non-Hamiltonian systems: local and semi-local normal …

2014-07-16abs ↗pdf ↗

In this paper, we prove that the two well-known natural normalizations of Hamiltonian functions on the symplectic manifold (M,ω)(M,ω) canonically relates the action spectra of different normalized Hamiltonians on {\it arbitrary} symplectic manifolds (M,ω)(M,ω). The natural class of normalized Hamiltonians consists of those w…

2002-06-10abs ↗pdf ↗

Fine-tuning normalization layers can reconstruct smaller networks.

problem Understanding the expressive power of fine-tuning normalization layers.
method Random ReLU networks and sparsified networks were fine-tuned to reconstruct target networks.
result Fine-tuning normalization layers can reconstruct networks that are O(extwidth)O(\sqrt{ ext{width}}) times smaller.

A new test assesses text similarity between two groups of documents.

problem Comparing similarity between two groups of documents.
method Neural network-based language models estimate entropy, and a test statistic derived from an estimation-and-inference framework is used.
result The proposed test maintains the nominal Type one error rate while offering greater power compared to existing methods.

This new research explores the effects of various training methods on a Polish to English Statistical Machine Translation system for medical texts. Various elements of the EMEA parallel text corpora from the OPUS project were used as the basis for training of phrase tables and language models and for development, tunin…

2015-09-29abs ↗pdf ↗

Motivated by manifold learning techniques, we give an explicit lower bound for how far a smoothly embedded compact submanifold in RN{\mathbb R}^N can move in a normal direction and remain an embedding. In addition, given a penalty function P:Emb(M,RN)RP : \text{Emb}(M,\mathbb{R}^N) \rightarrow \mathbb{R} on the space of embeddi…

2015-04-08abs ↗pdf ↗

APo-VAE generates text in hyperbolic space for better hierarchical representation.

problem Lack of hierarchical structure in Euclidean embeddings for natural language.
method Adversarial Poincare Variational Autoencoder (APo-VAE) in hyperbolic latent space.
result APo-VAE outperforms Euclidean VAEs in capturing latent language hierarchies.

EBMs improve text discrimination by generating negatives from auto-regressive models.

problem Discriminating machine-generated text from human-generated text.
method Use energy-based models to discriminate text, generating negatives using pre-trained auto-regressive language models.
result EBMs can generalize well to changes in generator architectures but are sensitive to training set.

The study examines the structure of certain subgroups of quasi-isometry groups of Euclidean spaces, proving their nontriviality and properties.

problem Investigating the algebraic and dynamical structure of certain subgroups of quasi-isometry groups of Euclidean spaces.
method Analyzing the normal subgroups \(H\) and \(H_\alpha\) of \(QI(\mathbb{R}^n)\) and proving their properties.
result The centers of \(QI(\mathbb{R}^n)/H\) and \(QI(\mathbb{R}^n)/H_\alpha\) are trivial, while these quotients admit nontrivial torsion elements.

We consider maximum solution g(t)g(t), t[0,+)t\in [0, +\infty), to the normalized Ricci flow. Among other things, we prove that, if (M,ω)(M, ω) is a smooth compact symplectic 4-manifold such that b2+(M)>1b_2^+(M)>1 and let g(t),t[0,)g(t),t\in[0,\infty), be a solution to (1.3) on MM whose Ricci curvature satisfies that $|\text{Ric}(g(t))|\l…

2007-04-05abs ↗pdf ↗

Proposes a novel framework for multi-label text classification.

problem Lack of coherent consideration of non-consecutive and long-distance semantics and hierarchical relations among labels.
method Hierarchical taxonomy-aware and attentional graph capsule recurrent CNNs framework.
result Significantly improves multi-label text classification performance.

Non-normal subgroups of certain groups grow homologically exponentially.

problem Homological torsion growth in non-normal subgroups of specific groups.
method Proving exponential growth of homological torsion in a sequence of non-normal subgroups.
result Exponential homological torsion growth in a sequence of non-normal subgroups.

SHMM models human mobility from GPS and text data, overcoming text sparsity.

problem Modeling human mobility from semantic trace data, especially addressing text sparsity.
method SHMM is a multi-modal spherical hidden Markov model that jointly models location, time, and text embeddings on a unit sphere using vMF distribution.
result SHMM outperforms state-of-the-art models in next location prediction and has lower training cost.

We let φ\varphi be an ageometric fully irreducible outer automorphism so that its Handel-Mosher axis bundle consists of a single unique axis. We show that the centralizer Cen(φ)Cen(\langle\varphi\rangle) of the cyclic subgroup generated by φ\varphi equals the stabilizer Stab(Λφ+)\text{Stab}(Λ^+_\varphi) of the attracting lamina…

2016-03-23abs ↗pdf ↗

Let (M,g)(M,g) be a nn-dimensional compact Riemannian manifold with boundary. We consider the Yamabe type problem \begin{equation} \left\{ \begin{array}{ll} -Δ_{g}u+au=0 & \text{ on }M \\ \partial_νu+\frac{n-2}{2}bu= u^{{n\over n-2}\pm\varepsilon} & \text{ on }\partial M \end{array}\right. \end{equation} where $a\in C^1(…

2015-06-30abs ↗pdf ↗

Normal solutions found for a specific curvature equation in 4D space.

problem Blow-up behavior in the Nirenberg problem and prescribed Q-curvature equation in R^4.
method Analyzing the integral form of solutions and proving existence and non-existence results.
result Normal solutions exist if and only if p ∈ (0, 4) and a specific range for Λ.

Critical volatility triggers log-normal to power-law transitions in interconnected systems.

problem Understanding the transition from log-normal to power-law distributions in interconnected systems.
method Analyzing an infinite option-on-option chain model, deriving a critical volatility threshold.
result A critical volatility threshold of approximately 250.66% for unconditional cases, dropping to 125.3% with selective survival.

Deep Learning predicts e-commerce activity from Italian enterprise websites.

problem Predicting e-commerce activity from Italian enterprise websites.
method Developed a sophisticated processing pipeline using Convolutional Neural Networks and Word Embeddings.
result Deep Learning outperforms traditional Machine Learning methods for text classification.

The study examines the normal growth exponent of submanifolds in negatively curved manifolds.

problem Understanding the normal growth exponent of submanifolds in negatively curved manifolds.
method Analyzing the geodesic flow and operator norms on submanifolds bi-Lipschitz to hyperbolic spaces.
result If a submanifold's normal growth exponent is at most 1, the ambient manifold is bi-Lipschitz to hyperbolic space.

We propose a family of near-metrics based on local graph diffusion to capture similarity for a wide class of data sets. These quasi-metametrics, as their names suggest, dispense with one or two standard axioms of metric spaces, specifically distinguishability and symmetry, so that similarity between data points of arbi…

2017-07-21abs ↗pdf ↗

New method predicts quasar continuum near Lyman-α with high precision and accuracy.

problem Precise measurement of quasar red damping wing for epoch of reionization.
method Fully probabilistic approach using conditional neural spline flows.
result Achieved state-of-the-art precision and accuracy in predicting quasar continua.

Discrete flows extend normalizing flows to discrete data, improving various applications.

problem Applying normalizing flows to discrete data distributions.
method Developed discrete autoregressive and bipartite flows, showing their effectiveness on various discrete data tasks.
result Discrete autoregressive flows outperform autoregressive baselines on synthetic discrete distributions and Potts models.

New RL algorithm for linear MDPs with nearly optimal regret.

problem Optimizing reinforcement learning for linear mixture Markov decision processes.
method Proposed a new Bernstein-type concentration inequality for self-normalized martingales and a computationally efficient algorithm UCRL-VTR+.
result UCRL-VTR+ achieves nearly minimax optimal regret of ildeO(dHT) ilde O(dH\sqrt{T}).

TFiLM expands convolutional models' receptive field with minimal overhead.

problem Capturing long-range dependencies in sequential data.
method A novel architectural component using a recurrent neural network to modulate convolutional model activations.
result TFiLM significantly improves learning speed and accuracy on various tasks.

We point out important problems with the common practice of using the best single model performance for comparing deep learning architectures, and we propose a method that corrects these flaws. Each time a model is trained, one gets a different result due to random factors in the training process, which include random …

2018-07-05abs ↗pdf ↗

New method proves asymptotic normality for matrix sensing problems.

problem Proving asymptotic normality for matrix sensing under general convex losses.
method Riemannian geometry to handle degeneracy of the Hessian due to rotational symmetry.
result Proves n(φ0φ)DN(0,(H)1)\sqrt{n}(φ^0-φ^*)\xrightarrow{D}N(0,(H^*)^{-1}) as non o\infty.