Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,341 papers · 148 categories

Trend · papers per month

0.3%0.5%0.8%0.2% · Nov 201619922001200920182026
9 results for Goal-Driven

GoBOED optimizes experiments for specific decision-making objectives, improving downstream outcomes.

problem Reducing parameter uncertainty does not always improve decision-making in critical settings.
method Combines variational posterior surrogate and differentiable convex decision layer for gradient-based design optimization.
result GoBOED identifies designs that better align with specific decision objectives and reveals wider optimal design windows.

This paper explores how hierarchical agent policies affect exploration in goal-driven navigation environments.

problem Understanding how hierarchical agent policies influence exploration in goal-driven navigation.
method Design of EscapeRoom environments, measuring complexity with hitting times of dependency graphs, evaluating PPO and hierarchical PPO.
result Analytically estimated hitting time in goal dependency graphs is a metric of environment complexity and hierarchical approaches are necessary for complex environments.

RNNs trained on head direction task mimic brain's compass and shifter neurons.

problem Modeling brain's head direction system using neural networks.
method Optimized recurrent neural networks trained on angular velocity integration.
result RNNs naturally emerge with compass and shifter neuron-like properties.

BetaDataWeighter learns weights for unlabelled data to improve self-supervised learning accuracy.

problem Improving unsupervised representations with domain shift between unlabelled and target data.
method Learning Bayesian instance weights for unlabelled data to prioritize useful instances.
result BetaDataWeighter achieves highest average accuracy and prunes up to 78% of images without significant loss in accuracy.

AI learns to perform experiments on objects to discover their properties.

problem Teaching AI to perform scientific experiments and infer physical properties.
method Deep reinforcement learning in a simulated environment.
result AI can learn to perform experiments to discover hidden physical properties.

This paper collects a large dataset for conversational recommendation research.

problem Creating dialogue systems for conversational recommendation.
method Collects a large dataset (ReDial) and explores neural architectures, mechanisms, and methods for conversational recommendation systems.
result Demonstrates the utility of the collected dataset for systematic probing of model sub-components in conversational recommendation.

Unified perspective unites Bayesian optimization and active learning for efficient goal-oriented optimization.

problem Efficiently optimize expensive engineering and scientific problems with limited data.
method Unified framework linking Bayesian infill criteria and active learning criteria.
result Unified approach formalizes Bayesian infill criteria and active learning criteria.