Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

51102152203 · Jun 202019922001200920172026
48 results for behavior modification

Behavior modification improves prediction accuracy by nudging user behavior.

problem Improving prediction accuracy using behavior modification techniques.
method Combining prediction and behavior modification with reinforcement learning algorithms.
result Behavior modification can make predictions more certain but may not generalize.

In this paper we investigate the asymptotic behavior of the colored HOMFLY polynomial of the figure eight knot associated with the symmetric representation. We establish an analogous asymptotic expansion for the colored HOMFLY polynomial. From the asymptotic behavior we show that the Chern-Simons invariants and twisted…

2017-11-13abs ↗pdf ↗

New bounds ensure reliable deep learning performance without model changes.

problem Certifying deep neural networks' reliability without altering them.
method Data-dependent generalization bounds that apply directly to trained models.
result Achieves meaningful generalization guarantees for large, unaltered deep networks.

A new framework for offline RL improves policy flexibility and regularity.

problem Lack of environmental interactions in offline RL leads to poor policy performance.
method Proposes a behavior-regularized implicit policy framework with modified policy-matching methods.
result The framework improves policy effectiveness and robustness beyond static datasets.

Study on new hyperbolicity notions for non-Kähler manifolds and their deformations.

problem Analyzing new hyperbolicity notions for non-Kähler complex manifolds.
method Introducing and analyzing two new notions of hyperbolicity for compact complex non-Kähler manifolds, and studying their behavior under smooth modifications.
result Established openness results for pp-HS hyperbolicity and pp-Kähler hyperbolicity under holomorphic deformations.

ModHiFi identifies critical components for model modification without gradients or loss function.

problem Modifying open weight models without access to training data or loss function.
method Theoretical analysis of Lipschitz-continuous networks, Subset Fidelity metric, and ModHiFi algorithm.
result ModHiFi-P and ModHiFi-U achieve significant performance improvements in model pruning and unlearning.

This work improves testing of machine learning model modifications using novel statistical methods.

problem Overfitting and conservative Bonferroni correction when testing multiple model modifications.
method Introduces alpha-recycling and SRGPs to control error rate and approve more beneficial modifications.
result Novel statistical methods approve a higher number of beneficial modifications than previous approaches.

Unintended effects from scaling neural network outputs with adaptive learning rates.

problem Adaptive learning rate optimization's behavior is altered by output scaling, leading to misinterpretation.
method Presented a modified optimization algorithm to mitigate unintended effects.
result Adaptive learning rate's effectiveness is significantly impacted by output scaling, especially for small scaling factors.

To widen their accessibility and increase their utility, intelligent agents must be able to learn complex behaviors as specified by (non-expert) human users. Moreover, they will need to learn these behaviors within a reasonable amount of time while efficiently leveraging the sparse feedback a human trainer is capable o…

2019-02-12abs ↗pdf ↗

Many studies assume stock prices follow a random process known as geometric Brownian motion. Although approximately correct, this model fails to explain the frequent occurrence of extreme price movements, such as stock market crashes. Using a large collection of data from three different stock markets, we present evide…

2009-12-30abs ↗pdf ↗

New method detects RNA modifications without prior training, revealing novel sites.

problem Detecting RNA modifications with high accuracy and sensitivity.
method Anomaly detection using nanopore raw ionic current signals and nearest neighbor comparison.
result Detects diverse RNA modifications without prior training, including a novel 2'-O-methylated site in DENV.

Federated Learning is a distributed learning paradigm with two key challenges that differentiate it from traditional distributed optimization: (1) significant variability in terms of the systems characteristics on each device in the network (systems heterogeneity), and (2) non-identically distributed data across the ne…

2018-12-14abs ↗pdf ↗

We define an operation on homology B4{B}^4 which we call an nn-twist annulus modification. We give a new construction of smoothly slice knots and exotically slice knots via nn-twist annulus modifications. As an application, we present a new example of a smoothly slice knot with non-slice derivatives. Such examples we…

2015-12-01abs ↗pdf ↗

InfoSFT improves LLMs by focusing on informative, medium-confidence tokens.

problem Overfitting to unlikely samples and degradation of prior capabilities in SFT.
method InfoSFT uses a principled weighting scheme to concentrate learning signals on medium-confidence tokens.
result InfoSFT improves generalization and preserves pre-existing capabilities over vanilla SFT and likelihood-weighted baselines.

We present a modification of the so-called Parrondo's paradox where one is allowed to choose in each turn the game that a large number of individuals play. It turns out that, by choosing the game which gives the highest average earnings at each step, one ends up with systematic loses, whereas a periodic or random seque…

2002-12-16abs ↗pdf ↗

Multi-agent systems exhibit complex behaviors that emanate from the interactions of multiple agents in a shared environment. In this work, we are interested in controlling one agent in a multi-agent system and successfully learn to interact with the other agents that have fixed policies. Modeling the behavior of other …

2020-01-29abs ↗pdf ↗

BCO* improves BCO by concurrently training inverse dynamics and expert policy.

problem Efficiently learn from unlabeled demonstrations without requiring many initial interactions.
method Introduce BCO* that concurrently trains an inverse dynamics model and expert policy.
result BCO* eliminates the need for initial interactions and improves sample complexity.

A modification of the confidence screening mechanism based on adaptive weighing of every training instance at each cascade level of the Deep Forest is proposed. The idea underlying the modification is very simple and stems from the confidence screening mechanism idea proposed by Pang et al. to simplify the Deep Forest …

2019-01-04abs ↗pdf ↗

For a manifold with nonpositive curvature, the Martin boundary is described by the behavior of normalized Green's functions at infinity. A classical result by Anderson and Schoen states that if the manifold has pinched negative curvature, the geometric boundary is the same as the Martin boundary. In this paper, we stud…

2017-06-14abs ↗pdf ↗

Study of unimodular Sasaki and Vaisman Lie groups, determining all modifications explicitly.

problem Classifying unimodular Sasaki and Vaisman Lie groups.
method Applying the technique of modification to determine all homogeneous Sasaki and Vaisman manifolds of unimodular Lie groups explicitly.
result Complete classification of unimodular Sasaki and Vaisman Lie groups.

We consider a two-valued function uu that is either Dirichlet energy minimizing, C1,μC^{1,μ} harmonic, or in C1,μC^{1,μ} with an area-stationary graph such that Almgren's frequency (restricted to the singular set) is continuous at a singular point Y0Y_0. As a corollary of recent work of Wickramasekera and the author, if t…

2014-10-27abs ↗pdf ↗

Linear Q-learning converges to a bounded set without divergence.

problem Proving linear Q-learning does not diverge and converges to a bounded set.
method No modifications to the original linear Q-learning algorithm, no Bellman completeness or near-optimality assumptions, only an ε-softmax behavior policy with adaptive temperature.
result First L2L^2 convergence rate of linear Q-learning iterates to a bounded set.

A simple approach to offline RL without additional complexity.

problem Learning from a fixed dataset of actions with value estimation errors.
method Adding a behavior cloning term to the policy update of an online RL algorithm and normalizing the data.
result Matches the performance of state-of-the-art offline RL algorithms with minimal changes.

Contextual bandit algorithms are applied in a wide range of domains, from advertising to recommender systems, from clinical trials to education. In many of these domains, malicious agents may have incentives to attack the bandit algorithm to induce it to perform a desired behavior. For instance, an unscrupulous ad publ…

2020-02-10abs ↗pdf ↗

Paper proposes using pairwise feature comparisons to infer modification costs for user recourse.

problem Learning and inferring user preferences for modifying features in black-box models.
method Bradley-Terry model for inferring feature-wise costs from non-exhaustive human comparison surveys.
result Non-exhaustive human surveys can efficiently learn feature costs, enabling recourse finding.

We briefly review the approach to optimization of portfolios according to the theory of Markowitz and propose a further modification that can improve the outcome of the optimization process. The modification takes account of the entropic contribution from the time series used to compute the parameters in the Markowitz …

2014-08-01abs ↗pdf ↗

Max-Cut decision tree improves classification accuracy and reduces computation time.

problem Improving decision tree accuracy and efficiency for complex classification tasks.
method Alternative splitting metric (max cut) and PCA-based feature selection at each node.
result 49% improvement in accuracy with 94% reduction in CPU time on CIFAR-100 data.

This paper considers the invariance of knot Floer homology in a purely algebraic setting, without reference to Heegaard diagrams, holomorphic disks, or grid diagrams. We show that (a small modification of) Ozsváth and Szabó's cube of resolutions for knot Floer homology, which is assigned to a braid presentation with a …

2010-07-15abs ↗pdf ↗

In this work we present a modification in the conventional flow of information through a LSTM network, which we consider well suited for RNNs in general. The modification leads to a iterative scheme where the computations performed by the LSTM cell are repeated over a constant input and cell state values, while updatin…

2018-07-11abs ↗pdf ↗

Kapustin and Witten associate a Hecke modification of a holomorphic bundle over a Riemann surface to a singular monopole on a Riemannian surface times an interval satisfying prescribed boundary conditions. We prove existence and uniqueness of singular monopoles satisfying prescribed boundary conditions for any given He…

2008-04-23abs ↗pdf ↗