Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

16314762 · Jun 202019922001200920172026
48 results for feedback unification

Paper fine-tunes LLMs using user edits, unifying preference, supervision, and reward feedback.

problem Adapting LLMs to user preferences and feedback types.
method Derives bounds for learning algorithms from user edits, proposes an ensembling procedure.
result Ensembling procedure outperforms individual feedback methods and robustly adapts to different user-edit distributions.

A new formulation of the Anomaly flow in the case of vanishing slope parameter is given, where the dependence on the global section of the canonical bundle appears only in the initial data. This allows a natural unification of the Anomaly flow with the Kähler-Ricci flow.

2019-05-06abs ↗pdf ↗

We investigate unification of two systems of identical elements having different dimensions which may be of interest for both physics and economics. Characteristic parameters as well as explicit formulae for the temperature (in economics - capital turnover) and dimension of the united system are obtained as functions o…

2015-10-03abs ↗pdf ↗

In the paper [4] is presented a theory which unifies the gravitation theory and the mechanical effects, which is different from the Riemannian theories like GTR. Moreover it is built in the style of the electomagnetic field theory. This paper is a continuation of [4] such that the complex variant of that theory yields …

2001-10-11abs ↗pdf ↗

Human reasoning involves recognising common underlying principles across many examples. The by-products of such reasoning are invariants that capture patterns such as "if someone went somewhere then they are there", expressed using variables "someone" and "somewhere" instead of mentioning specific people or places. Hum…

2019-09-16abs ↗pdf ↗

Unified study of Riemannian and sub-Riemannian geometries with synthetic Ricci curvature bounds.

problem Unified framework for Riemannian and sub-Riemannian geometries.
method Study of gauge metric measure spaces.
result Unified synthetic Ricci curvature lower bounds for both Riemannian and sub-Riemannian structures.

Unified field theory from higher-order Riemannian geometry.

problem Field-theoretical unification of fundamental forces.
method Exploiting higher-order Riemannian geometry and Einstein-Hilbert action, deriving gauge theories and predicting physical constants.
result Theoretical predictions for Weinberg angle and Coulomb's constant match experimental values.

The paper glosses different forms of an introducing of higher order tangent-like functors, especially functors derived from higher order nonholonomic tangent functors. A special attention is devoted to higher order osculating bundles: their identification with higher order tangent bundles is demonstrated as the main re…

2012-02-13abs ↗pdf ↗

The paper improves on existing algorithms for minimizing different types of regret in online learning.

problem Minimizing external, internal, and swap regret in online learning with multiple experts.
method Develops a single algorithm using φ-regret minimization and Haar-wavelet-inspired matrix features to achieve optimal bounds in various scenarios.
result Achieves optimal bounds for external, internal, and swap regrets in different expert scenarios.

Unified framework for spectral methods, kernel learning, and manifold unfolding.

problem Tackles the unification and optimization of spectral dimensionality reduction methods.
method Unified spectral methods as kernel PCA, kernel learning by SDP, and detailed explanation of MVU variants.
result Unified understanding and optimization of manifold learning techniques.

We use a new approach that we call unification to prove that standard weighted double bubbles in nn-dimensional Euclidean space minimize immiscible fluid surface energy, that is, surface area weighted by constants. The result is new for weighted area, and also gives the simplest known proof to date of the (unit weight…

2012-12-19abs ↗pdf ↗

Non-associtive algebras is a research direction gaining much attention these days. New developments show that associative algebras and some not-associative structures can be unified at the level of Yang-Baxter structures. In this paper, we present a unification for associative algebras, Jordan algebras and Lie algebras…

2014-08-16abs ↗pdf ↗

New insights into cascade feedback linearization of control systems.

problem Obtaining a cascade feedback linearization for invariant control systems.
method Introducing truncated versions of operators from the calculus of variations to prove new theorems.
result Established new geometry and foundational theorems for future work.

New method learns from either positive or negative feedback alone.

problem Limited applicability of existing preference optimization methods in scenarios with only unpaired feedback.
method Decouples learning from positive and negative feedback, using expectation-maximization (EM) to optimize probability of positive outcomes and explicitly incorporate negative examples.
result Stable learning from negative feedback alone demonstrated.

User preferences for items can be inferred from either explicit feedback, such as item ratings, or implicit feedback, such as rental histories. Research in collaborative filtering has concentrated on explicit feedback, resulting in the development of accurate and scalable models. However, since explicit feedback is oft…

2011-09-27abs ↗pdf ↗

This paper improves image retrieval accuracy through novel relevance feedback methods.

problem Improving image retrieval accuracy in Content-Based Image Retrieval (CBIR).
method Novel addition to feature re-weighting and classification techniques, focusing on 0-th iteration improvement.
result Significantly improved retrieval accuracy from relevance feedback.

Paper connects two portfolio methods, HRP and Minimum Variance, revealing their underlying similarity.

problem Inability to universally adopt optimization-based portfolio construction methods.
method Unifies Hierarchical Risk Parity and Minimum Variance approaches.
result Schur complementary allocation reveals the connection between HRP and Minimum Variance.

We introduce a new class of quantum enhancements we call biquandle brackets, which are customized skein invariants for biquandle colored links.Quantum enhancements of biquandle counting invariants form a class of knot and link invariants that includes biquandle cocycle invariants and skein invariants such as the HOMFLY…

2015-08-26abs ↗pdf ↗

In 2006 Habiro initiated a construction of generating functions for Witten-Reshetikhin-Turaev (WRT) invariants known as unified WRT invariants. In a series of papers together with Irmgard Buehler and Christian Blanchet we extended his construction to a larger class of 3-manifolds. The unified invariants provide a stron…

2011-06-30abs ↗pdf ↗

Develops a new model for RLHF accounting for partially observed states and intermediate feedback.

problem Lack of models for partially observed states and intermediate feedback in RLHF.
method PORRL model with cardinal and dueling feedback methods.
result Demonstrates improved learning and alignment with new model-based and model-free methods.

Classifier learns to ignore unreliable feedback from end users.

problem Improving classifier performance by filtering unreliable feedback.
method Modeling end users as autonomous agents, periodically retraining classifier with filtered feedback.
result Classifier can identify and filter out unreliable feedback, improving performance.

Advocates a local feedback approach for RL in unknown systems.

problem Finding optimal feedback laws in unknown nonlinear dynamical systems.
method Searches over a local feedback representation consisting of an open-loop sequence and an optimal linear feedback law.
result Results in highly efficient training and superior performance compared to global methods.

Study how communication and feedback graphs affect learning outcomes.

problem Understanding the impact of feedback graphs on cooperative online learning.
method Analyzed network regret in terms of the independence number of the strong product of communication and feedback graphs.
result Proved bounds for network regret and demonstrated the non-improvable nature of positive results in pathological cases.

New method uses correlated auxiliary feedback to reduce regret in parameterized bandits.

problem Reducing regret in parameterized bandits with correlated auxiliary feedback.
method Develops a reward estimator using auxiliary feedback with tight confidence bounds.
result Shows significant reduction in regret compared to standard methods.

In this paper we describe market in projective geometry language and give definition of a matrix of market rate, which is related to the matrix rate of return and the matrix of judgements in the Analytic Hierarchy Process (AHP). We use these observations to extend the AHP model to projective geometry formalism and gene…

2007-09-27abs ↗pdf ↗

The problem of feedback equivalence for control systems is considered. An algebra of differential invariants and criteria for the feedback equivalence for regular control systems are found.

2008-12-07abs ↗pdf ↗