Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

6.3%12.5%18.8%25.0% · Jul 199319922001200920172026
48 results for forward translation

Forward translation improves neural machine translation for sentences originally in source language.

problem Improving neural machine translation quality using synthetic data.
method Case study with French-English news translation, separating test sets by original language, analyzing domains, translationese, and noise.
result Forward translation delivers superior gains on sentences originally in source language, complementing back-translation on target language sentences.

Let SO+(p,q)\mathrm{SO}^+(p,q) denote the identity connected component of the real orthogonal group with signature (p,q)(p,q). We give a complete description of the spaces of continuous and generalized translation- and SO+(p,q)\mathrm{SO}^+(p,q)-invariant valuations, generalizing Hadwiger's classification of Euclidean isometry-invari…

2016-02-28abs ↗pdf ↗

This paper uses LLMs and cycle consistency for better machine translation evaluation.

problem Evaluating translation quality and LLM capabilities without ground truth.
method Generate translation candidates, back-translate, and evaluate cycle consistency.
result Larger LLMs or more inference passes improve cycle consistency.

Demographic projections of future mortality rates involve a high level of uncertainty and require stochastic mortality models. The current paper investigates forward mortality models driven by a (possibly infinite dimensional) Wiener process and a compensated Poisson random measure. A major innovation of the paper is t…

2019-07-11abs ↗pdf ↗

Transformer networks have lead to important progress in language modeling and machine translation. These models include two consecutive modules, a feed-forward layer and a self-attention layer. The latter allows the network to capture long term dependencies and are often regarded as the key ingredient in the success of…

2019-07-02abs ↗pdf ↗

End-to-end optimization has achieved state-of-the-art performance on many specific problems, but there is no straight-forward way to combine pretrained models for new problems. Here, we explore improving modularity by learning a post-hoc interface between two existing models to solve a new task. Specifically, we take i…

2019-02-21abs ↗pdf ↗

Recurrent neural networks (RNNs) sequentially process data by updating their state with each new data point, and have long been the de facto choice for sequence modeling tasks. However, their inherently sequential computation makes them slow to train. Feed-forward and convolutional architectures have recently been show…

2018-07-10abs ↗pdf ↗

Ancient grain boundaries resemble atoms in their formation and properties.

problem Understanding the formation and properties of ancient grain boundaries.
method Analyzing ancient grain boundaries as analogous to atoms and using geometric flow techniques.
result New examples of convex ancient and translating solutions to mean curvature flow.

A new method uses normalizing flows to approximate optimal transport between empirical distributions.

problem Learning an optimal transport map between two empirical distributions.
method Relaxing the Monge formulation of optimal transport, using normalizing flows to approximate the solution.
result The method provides a good approximation of the true optimal transport.

The theory of convex risk functions has now been well established as the basis for identifying the families of risk functions that should be used in risk averse optimization problems. Despite its theoretical appeal, the implementation of a convex risk function remains difficult, as there is little guidance regarding ho…

2016-07-24abs ↗pdf ↗

We introduce dropout compaction, a novel method for training feed-forward neural networks which realizes the performance gains of training a large model with dropout regularization, yet extracts a compact neural network for run-time efficiency. In the proposed method, we introduce a sparsity-inducing prior on the per u…

2016-11-18abs ↗pdf ↗

Paper forecasts stock correlations using a hybrid model combining graph neural networks and transformers.

problem Improving stock correlation forecasts for better portfolio management.
method Hybrid model combining Transformer and graph attention networks for forecasting residual deviations from historical data.
result The hybrid model reduces correlation forecasting error compared to rolling-window estimates.

We introduce two models of taxation, the latent and natural tax processes, which have both been used to represent loss-carry-forward taxation on the capital of an insurance company. In the natural tax process, the tax rate is a function of the current level of capital, whereas in the latent tax process, the tax rate is…

2018-11-05abs ↗pdf ↗

Deep learning's success requires vast computing power, making future progress unsustainable.

problem Deep learning's success is heavily dependent on computing power, making future progress unsustainable.
method Cataloging and extrapolating the dependency on computing power for various deep learning applications.
result Continued progress in deep learning applications will require more computationally-efficient methods.

The vast majority of successful deep neural networks are trained using variants of stochastic gradient descent (SGD) algorithms. Recent attempts to improve SGD can be broadly categorized into two approaches: (1) adaptive learning rate schemes, such as AdaGrad and Adam, and (2) accelerated schemes, such as heavy-ball an…

2019-07-19abs ↗pdf ↗

Study on rigidity of translating hypersurfaces not in graphical direction.

problem Rigidity of translating hypersurfaces not in graphical direction.
method Proved rigidity results for complete graphical translating hypersurfaces under specific conditions.
result Entire graphical translating surfaces are flat under certain conditions.

Bayesian deep learning improves seismic imaging uncertainty.

problem Uncertainty in seismic imaging due to data noise and linearization errors.
method Combines Bayesian inference and deep neural networks to quantify uncertainty in horizon tracking.
result Uncertainty in automatically tracked horizons can be quantified and visualized.

The natural automorphism group of a translation surface is its group of translations. For finite translation surfaces of genus g > 1 the order of this group is naturally bounded in terms of g due to a Riemann-Hurwitz formula argument. In analogy with classical Hurwitz surfaces, we call surfaces which achieve the maxima…

2013-11-28abs ↗pdf ↗

Study on stable translation lengths of surface homeomorphisms and their approximations.

problem Understanding stable translation lengths of homeomorphisms and their finite approximations.
method Comparing stable translation lengths of homeomorphisms and their finite approximations on curve graphs.
result Stable translation length of homeomorphisms with dense periodic points equals the supremum of their approximations.

Study measures gender bias in machine translation using multiple reference points.

problem Measuring and identifying gender bias in machine translation.
method Used an optimal non-biased translator, reference points from occupational statistics and survey.
result Found bias against both genders, but more against women, and found occupations have a greater effect than adjectives.

Paper finds a non-existence theorem for certain translators in high dimensions.

problem Non-existence of certain translators in high-dimensional spaces.
method Developed a non-existence theorem and found an example of a translator.
result Non-existence of entire Qn1Q_{n-1}-translators in Rn+1\mathbb{R}^{n+1}.

Constructing translating solitons from Lagrangian Grim Reapers.

problem Creating Lagrangian translating solitons from intersections of Grim Reapers.
method Desingularizing intersections with special Lagrangian Lawlor necks.
result Constructing Lagrangian translating solitons with multiple ends and loops.

Study on singular points of translation surfaces under linearly dependent conditions.

problem Investigate singular points of translation surfaces under linearly dependent conditions.
method Use theories of generalised framed surfaces and framed surfaces.
result Introduce translation generalised framed surfaces and investigate their singular points.

In this article we prove two non-existence results for translating solitons of the mean curvature flow (translators for short) in Rm+1\mathbb{R}^{m+1}. We also obtain an upper bound to the maximum height that a compact embedded translator in R3\mathbb{R}^{3} can achieve. On the other hand, we study graphical perturbation…

2016-01-27abs ↗pdf ↗

Neural machine translation is a recently proposed approach to machine translation. Unlike the traditional statistical machine translation, the neural machine translation aims at building a single neural network that can be jointly tuned to maximize the translation performance. The models proposed recently for neural ma…

2014-09-01abs ↗pdf ↗

Study invariant λλ-translators in Lorentz-Minkowski space.

problem Characterize λλ-translators invariant under translations and rotations.
method Analyze 1-parameter group of translations and rotations, find explicit parametrizations, and solve non-linear autonomous systems.
result Explicit parametrizations and qualitative properties of invariant λλ-translators.