A new reinforcement learning method for robots thinking and moving simultaneously.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Training neural network often uses a machine learning framework such as TensorFlow and Caffe2. These frameworks employ a dataflow model where the NN training is modeled as a directed graph composed of a set of nodes. Operations in neural network training are typically implemented by the frameworks as primitives and rep…
Paper proposes a method to learn and exceed expert demonstrations in unknown reward environments.
New TTP framework fuses control arms while controlling Type-I error.
This paper examines a generalized Kropina metric and its geometric properties.
The paper studies Finsler spaces with semi-concurrent vector fields and their equivalence to Riemannian spaces.
A Ricci soliton on a Riemannian manifold is said to have concurrent potential field if its potential field is a concurrent vector field. In the first part of this paper we completely classify Ricci solitons with concurrent potential fields. In the second part we derive a necessary and suffic…
A Ricci soliton on a Riemannian manifold is said to have concurrent potential field if its potential field is a concurrent vector field. Ricci solitons arisen from concurrent vector fields on Riemannian manifolds were studied recently in \cite{CD2}. The most important concurrent vector field is …
In the present paper, we introduce and investigate the notion of a semi concurrent vector field on a Finsler manifold. We show that some special Finsler manifolds admitting such vector fields turn out to be Riemannian. We prove that Tachibana's characterization of Finsler manifolds admitting a concurrent vector field l…
We consider the problem of downlink power control in wireless networks, consisting of multiple transmitter-receiver pairs communicating with each other over a single shared wireless medium. To mitigate the interference among concurrent transmissions, we leverage the network topology to create a graph neural network arc…
MeanFlow training is unstable due to misusing conditional velocity, leading to variance issues.
A decentralized deep RL controller improves hexapod locomotion learning.
Speeds up deep neural networks training by 10x using GPU concurrency.
DeepFDR uses deep learning for better FDR control in neuroimaging data.
BCO* improves BCO by concurrently training inverse dynamics and expert policy.
The present paper deals with an \emph{intrinsic} investigation of the notion of a concurrent -vector field on the pullback bundle of a Finsler manifold . The effect of the existence of a concurrent -vector field on some important special Finsler spaces is studied. An intrinsic investigation of a particular…
For large-scale industrial processes under closed-loop control, process dynamics directly resulting from control action are typical characteristics and may show different behaviors between real faults and normal changes of operating conditions. However, conventional distributed monitoring approaches do not consider the…
We consider the problem of concurrent portfolio losses in two non-overlapping credit portfolios. In order to explore the full statistical dependence structure of such portfolio losses, we estimate their empirical pairwise copulas. Instead of a Gaussian dependence, we typically find a strong asymmetry in the copulas. Co…
New algorithm provably converges to second-order stationary points in NMF.
Novel framework controls FDR in high-dimensional, dependent data.
Novel framework for data sharing and coordinated exploration in concurrent RL with non-identical environments.
In this paper, we completely classify almost Yamabe solitons on hypersurfaces in Euclidean spaces arisen from the position vector field. Some results of almost Yamabe solitons with a concurrent vector field and almost Yamabe solitons on submanifolds in Riemannian manifolds equipped with a concurrent vector field are al…
MEC-Cox: A Machine-Learning-Assisted Generalized Entropy Calibration Method for Estimating ATT Marginal Hazard-Ratio
We generalize Matsumoto metrics with a special π-form and explore their geometric properties.
Deep learning model classifies concurrent human interactions from WiFi data with high accuracy.
New algorithm trains neural nets on simple skills to learn complex tasks faster.
A novel method optimizes variable-stiffness structures for better strength and weight.
We study the equilibrium positions of three points on a convex curve under influence of the Coulomb potential. We identify these positions as orthotripods, three points on the curve having concurrent normals. This relates the equilibrium positions to the caustic (evolute) of the curve. The concurrent normals can only m…
We consider a team of reinforcement learning agents that concurrently operate in a common environment, and we develop an approach to efficient coordinated exploration that is suitable for problems of practical scale. Our approach builds on seed sampling (Dimakopoulou and Van Roy, 2018) and randomized value function lea…
Optimizes query routing to LLMs under cost and resource constraints.
The study confirms conjectures about normals to convex polytopes in 3D space.
This work addresses unstable MeanFlow training by optimizing a coefficient in the loss function.
Catastrophic forgetting occurs when a neural network loses the information learned in a previous task after training on subsequent tasks. This problem remains a hurdle for artificial intelligence systems with sequential learning capabilities. In this paper, we propose a task-based hard attention mechanism that preserve…
Deep reinforcement learning enables algorithms to learn complex behavior, deal with continuous action spaces and find good strategies in environments with high dimensional state spaces. With deep reinforcement learning being an active area of research and many concurrent inventions, we decided to focus on a relatively …
Max-rank improves multiple testing in conformal prediction.
The paper discusses the impossibility of eliminating surplus intersections in Lagrangian submanifolds.
Given a similarity graph between items, correlation clustering (CC) groups similar items together and dissimilar ones apart. One of the most popular CC algorithms is KwikCluster: an algorithm that serially clusters neighborhoods of vertices, and obtains a 3-approximation ratio. Unfortunately, KwikCluster in practice re…
PolySwarm uses a swarm of LLMs to predict and arbitrage prediction markets.
Study on vector fields on Lie groups reveals surprising algebraic coincidences.
The aim of this paper is to train an RBF neural network and select centers under concurrent faults. It is well known that fault tolerance is a very attractive property for neural networks. And center selection is an important procedure during the training process of an RBF neural network. In this paper, we devise two n…
In dialogues, an utterance is a chain of consecutive sentences produced by one speaker which ranges from a short sentence to a thousand-word post. When studying dialogues at the utterance level, it is not uncommon that an utterance would serve multiple functions. For instance, "Thank you. It works great." expresses bot…
A new method creates simpler, more interpretable decision trees from complex ensembles.
Study presents a low-cost local motion planner for vineyard navigation.
The presence of data corruption in user-generated streaming data, such as social media, motivates a new fundamental problem that learns reliable regression coefficient when features are not accessible entirely at one time. Until now, several important challenges still cannot be handled concurrently: 1) corrupted data e…
New algorithms improve tensor CP decomposition under mild conditions.
We introduce economic models based on Boolean Delay Equations: this formalism makes easier to take into account the complexity of the interactions between firms and is particularly appropriate for studying the propagation of an initial damage due to a catastrophe. Here we concentrate on simple cases, which allow to und…
This paper describes Plumbing for Optimization with Asynchronous Parallelism (POAP) and the Python Surrogate Optimization Toolbox (pySOT). POAP is an event-driven framework for building and combining asynchronous optimization strategies, designed for global optimization of expensive functions where concurrent function …
In artificial multi-agent systems, the ability to learn collaborative policies is predicated upon the agents' communication skills: they must be able to encode the information received from the environment and learn how to share it with other agents as required by the task at hand. We present a deep reinforcement learn…