New methodology controls synthetic data bias for neural program synthesis.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
This article introduces the degenerate special Lagrangian equation (DSL) and develops the basic analytic tools to construct and study its solutions. The DSL governs geodesics in the space of positive graph Lagrangians in Existence of geodesics in the space of positive Lagrangians is an important step in…
DSL uses supervised learning to optimize portfolios, improving stability and performance.
New findings on convexity of special Lagrangian geodesics.
The process of designing neural architectures requires expert knowledge and extensive trial and error. While automated architecture search may simplify these requirements, the recurrent neural network (RNN) architectures generated by existing methods are limited in both flexibility and components. We propose a domain-s…
Proposes PA-DSL for correcting noisy human labels in automated data labeling.
DSL estimates heterogeneous treatment effects over time in survival settings.
We show that the degenerate special Lagrangian equation, recently introduced by Rubinstein-Solomon, induces a global equation on every Riemannian manifold, and that for certain associated geometries this equation governs, as it does in the Euclidean setting, geodesics in the space of positive Lagrangians. For example, …
DynamicPPL speeds up probabilistic modeling in Julia.
The goal in network state prediction (NSP) is to classify the global state (label) associated with features embedded in a graph. This graph structure encoding feature relationships is the key distinctive aspect of NSP compared to classical supervised learning. NSP arises in various applications: gene expression samples…
The results of data mining endeavors are majorly driven by data quality. Throughout these deployments, serious show-stopper problems are still unresolved, such as: data collection ambiguities, data imbalance, hidden biases in data, the lack of domain information, and data incompleteness. This paper is based on the prem…
New method uses imperfect LLM annotations for valid statistical inference in social science.
We present a training system, which can provably defend significantly larger neural networks than previously possible, including ResNet-34 and DenseNet-100. Our approach is based on differentiable abstract interpretation and introduces two novel concepts: (i) abstract layers for fine-tuning the precision and scalabilit…
This paper characterizes and designs loss functions for robust classification with abstention.
Recent work has shown how to embed differentiable optimization problems (that is, problems whose solutions can be backpropagated through) as layers within deep learning architectures. This method provides a useful inductive bias for certain problems, but existing software for differentiable optimization layers is rigid…
Study optimal rates for multiclass classification, resolving open questions.
A new approach for specifying and synthesizing subroutines for optimizing metrics.
LLM agents discover cryptocurrency factors under reproducible constraints.
We introduce SPFlow, an open-source Python library providing a simple interface to inference, learning and manipulation routines for deep and tractable probabilistic models called Sum-Product Networks (SPNs). The library allows one to quickly create SPNs both from data and through a domain specific language (DSL). It e…
Improved genetic programming by optimizing mutation operators for continuous program search.