Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

265278104 · Jun 202019922001200920182026
48 results for weak absolute-valued feedback

New algorithm for duelling bandits with weak regret in adversarial settings.

problem Improving performance in duelling bandits with weak regret.
method Developed an algorithm for duelling bandits in adversarial environments, considering the Borda winner.
result Algorithm provides theoretical guarantees in both utility-based and unrestricted settings.

Simple uniqueness proof for McKean-Vlasov systems with weak feedback.

problem Global and short-time uniqueness for McKean-Vlasov systems with small feedback.
method Simple probabilistic comparison argument robust to solution regularity.
result Global uniqueness for a broad class of McKean-Vlasov problems in the weak feedback regime.

Online boosting for multiclass classification with limited feedback.

problem Online multiclass classification with bandit feedback.
method Proposed unbiased loss estimate and extended full information boosting algorithms to bandit setting.
result Asymptotic error bounds match full information counterparts, with larger sample complexity due to limited feedback.

Interactive weak supervision learns useful heuristics from user feedback.

problem Creating useful heuristics for large labeled datasets is tedious and subjective.
method Develops an interactive framework for learning heuristics from user feedback.
result Only a few feedback iterations are needed to train models without ground truth labels.

This paper studies the Glosten Milgrom model whose risky asset value admits an arbitrary discrete distribution. Contrast to existing results on insider's models, the insider's optimal strategy in this model, if exists, is not of feedback type. Therefore a weak formulation of equilibrium is proposed. In this weak formul…

2013-10-18abs ↗pdf ↗

Framework identifies population quantities from MNAR feedback using weak shadow variables from pretrained models.

problem Estimating mean outcomes from MNAR user feedback with bias and lack of identification.
method Develops a partial identification framework using linear programs and weak shadow variables from pretrained models.
result Bounds on estimand are obtained by solving linear programs incorporating pretrained model predictions.

Efficient algorithms for online multiclass linear classification with bandit feedback under linear separability conditions.

problem Efficient online multiclass linear classification with bandit feedback for separable data.
method Design of efficient algorithms based on kernel Perceptron for strong and weak linear separability conditions.
result Near-optimal mistake bounds of $O\left( K/γ^2 ight)$ for strong separability and min(2O~(Klog2(1/γ)),2O~(1/γlogK))\min (2^{\widetilde{O}(K \log^2 (1/γ))}, 2^{\widetilde{O}(\sqrt{1/γ} \log K)}) for weak separability.

Novel ramp loss method improves weakly supervised machine translation and parsing.

problem Training neural models without gold labels in weak supervision scenarios.
method Adapted ramp loss objectives to promote positive outputs and discourage negative ones.
result Bipolar ramp loss objectives outperform other methods on weakly supervised tasks.

Formalizes weak and strong verification for LLMs, controlling errors without assumptions.

problem Balancing cost and reliability in reasoning with LLMs.
method Formalizes weak-strong verification policies, introduces metrics, develops online algorithm.
result Optimal policies admit a two-threshold structure, and calibration and sharpness govern value of weak verifiers.

New graph feedback model for bandits with improved regret bounds.

problem Understanding how graph structure affects regret in bandit problems.
method Introduced fractional weak domination number and kk-packing independence number to capture upper and lower bounds on regret. Used strong duality theorem to derive upper and lower bounds.
result Proved general upper and lower bounds on regret for various graph structures, showing tightness up to a logarithmic factor.

The paper constrains families of smooth 4-manifolds using Seiberg-Witten invariants.

problem Understanding the topology of families of smooth 4-manifolds.
method Finite dimensional approximation of the Seiberg-Witten monopole map.
result Constructs examples of continuous Zp\mathbb{Z}_p-actions and shows non-smoothability.

In this manuscript we analyse the leading statistical properties of fluctuations of (log) 3-month US Treasury bill quotation in the secondary market, namely: probability density function, autocorrelation, absolute values autocorrelation, and absolute values persistency. We verify that this financial instrument, in spit…

2007-06-08abs ↗pdf ↗

This paper extends a Kyle model to include price-responsive traders, revealing new dynamics and equilibria.

problem Real-world market dynamics involve price-responsive traders, affecting market equilibrium and insider profits.
method Developed a continuous-time Kyle model with two types of price-responsive traders (momentum and contrarian), leading to a forward-backward Riccati system for equilibrium.
result The model shows that feedback effects can lead to multiple equilibria and amplify price informativeness.

Proposes a model to classify nodes in networks using weighted feedback relations.

problem Challenges in predicting node labels in sparse networks with implicit feedback.
method Weighted personalized two-stage matrix factorization model with Bayesian ranking loss.
result Significantly outperforms state-of-the-art models on various datasets.

We classify isotopy classes of irreducible Heegaard splittings of solvmanifolds. If the monodromy of the solvmanifold can be expressed as a 2 x 2 matrix with 0 in the lower right hand corner (as always is true when the absolute value of the trace is 3), then any irreducible splitting is strongly irreducible and of genu…

1998-03-31abs ↗pdf ↗

Proposes estimators for complex dose-response curves using kernel methods.

problem Estimating complex dose-response curves with continuous treatments, mediators, and covariates.
method Kernel ridge regression with sequential kernel embedding technique.
result Simple estimators for mediated and time-varying dose response curves with nonasymptotic uniform rates.

Neural networks learn distance-based representations, not just intensity.

problem Understanding how neural networks interpret and learn from internal activations.
method Manipulated ReLU and Absolute Value activations to observe sensitivity to distance and intensity perturbations.
result Neural networks are highly sensitive to small distance-based perturbations, challenging the intensity-based interpretation.

We study principal curvatures of fibers and Heegaard surfaces smoothly embedded in hyperbolic 3-manifolds. It is well known that a fiber or a Heegaard surface in a hyperbolic 3-manifold cannot have principal curvatures everywhere less than one in absolute value. We show that given an upper bound on the genus of a minim…

2010-02-04abs ↗pdf ↗

We explore whether useful temporal neural generative models can be learned from sequential data without back-propagation through time. We investigate the viability of a more neurocognitively-grounded approach in the context of unsupervised generative modeling of sequences. Specifically, we build on the concept of predi…

2017-11-30abs ↗pdf ↗

Algorithm learns to bid optimally in repeated first-price auctions with censored feedback.

problem Learning to bid optimally in repeated first-price auctions with incomplete feedback.
method Developed an algorithm exploiting the specific feedback structure and payoff function of first-price auctions.
result Achieved a near-optimal O~(T)\widetilde{O}(\sqrt{T}) regret bound for first-price auctions.

Learning the right graph representation from noisy, multisource data has garnered significant interest in recent years. A central tenet of this problem is relational learning. Here the objective is to incorporate the partial information each data source gives us in a way that captures the true underlying relationships.…

2014-01-14abs ↗pdf ↗

Model user preferences for conversational LLMs using weak rewards.

problem Lack of persistent user models in conversational LLMs leading to repeated user restatements.
method Vector-Adapted Retrieval Scoring (VARS) framework that updates user vectors online from weak scalar rewards.
result Full VARS agent achieves strongest overall performance, matches strong Reflection baseline in task success, and reduces user effort.

There is increasing interest in learning algorithms that involve interaction between human and machine. Comparison-based queries are among the most natural ways to get feedback from humans. A challenge in designing comparison-based interactive learning algorithms is coping with noisy answers. The most common fix is to …

2018-02-20abs ↗pdf ↗

Algorithm minimizes regret in dueling bandits with contextualized utilities.

problem Minimizing regret in dueling bandits with context-dependent utilities.
method Proposes CoLSTIM algorithm based on perturbed utility estimates.
result Achieves regret of order ildeO(dT) ilde O(\sqrt{dT}).

Improved mistake bound for group linear separable cases in online multiclass linear classification.

problem Improving mistake bounds for online multiclass linear classification under group linear separable conditions.
method Refined group weak linear separability condition and rational kernel approach.
result Achieved a mistake bound of K2ildeO(1/γlogL))K\cdot 2^{ ilde{O}(\sqrt{1/γ}\log L)}) under group weak linear separable condition.

xAI-GAN improves GANs by providing richer feedback, enhancing image quality.

problem Challenges in GAN training, especially resource-intensive and data-intensive.
method Integrates xAI systems to provide richer corrective feedback from discriminators to generators.
result Improves image quality by up to 23.18% on MNIST and FMNIST datasets.

We present an efficient algorithm for calculating the number of components of an integral lamination on an nn-punctured disk, given its Dynnikov coordinates. The algorithm requires O(n2M)O(n^2M) arithmetic operations, where MM is the sum of the absolute values of the Dynnikov coordinates.

2015-12-28abs ↗pdf ↗

We propose a method to build quantum memristors in quantum photonic platforms. We firstly design an effective beam splitter, which is tunable in real-time, by means of a Mach-Zehnder-type array with two equal 50:50 beam splitters and a tunable retarder, which allows us to control its reflectivity. Then, we show that th…

2017-09-22abs ↗pdf ↗

Recurrent neural networks (RNNs) are notoriously difficult to train. When the eigenvalues of the hidden to hidden weight matrix deviate from absolute value 1, optimization becomes difficult due to the well studied issue of vanishing and exploding gradients, especially when trying to learn long-term dependencies. To cir…

2015-11-20abs ↗pdf ↗

Establishes a link between risk measures and uniform integrability in finance.

problem Understanding uniform integrability in the context of financial risk measures.
method Introduces the folding score of distortion risk measures to study uniform integrability directly with gains and losses.
result Obtains three sets of equivalent conditions for uniform integrability involving coherent risk measures.

Active learning improves inspection systems by using weakly labeled data.

problem Rapidly updating machine vision inspection systems in evolving manufacturing processes.
method Developed a methodology for active learning from weakly labeled data, addressing covariate shift with domain-adversarial training.
result Demonstrated that active learning can accelerate the annotation process and reduce false positives.

In this paper, we study the power of Gaussian curvature flow of a compact convex hypersurface and establish its Harnack inequality when the power is negative. In the Harnack inequality, we require that the absolute value of the power is strictly positive and strictly less than the inverse of the dimension of the hypers…

2011-02-22abs ↗pdf ↗

A finitely generated module over the ring L=Z[t, t^{-1}] of integer Laurent polynomials that has no Z-torsion is determined by a pair of sub-lattices of L^d. Their indices are the absolute values of the leading and trailing coefficients of the order of the module. This description has applications in knot theory.

2010-06-21abs ↗pdf ↗

A complex-valued convolutional network (convnet) implements the repeated application of the following composition of three operations, recursively applying the composition to an input vector of nonnegative real numbers: (1) convolution with complex-valued vectors followed by (2) taking the absolute value of every entry…

2015-03-11abs ↗pdf ↗