Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

4285127169 · Jun 202019922001200920172026
48 results for testable independence

Polynomial-time tester-learner for general halfspaces with Gaussian adversarial noise.

problem Learning general halfspaces with adversarial label noise.
method Reduction to testable learning of nearly homogeneous halfspaces.
result First polynomial time tester-learner for general halfspaces with dimension-independent misclassification error.

New approach for testable learning using moment matching and Rademacher complexity.

problem Replacing hard-to-verify distributional assumptions with testable ones.
method Moment matching and metric distances in probability.
result Improved sample complexity bounds for various concept classes and distributions.

Local learning method selects covariates for causal effect estimation in the presence of latent variables.

problem Estimating causal effects from nonexperimental data with latent variables.
method Local learning approach that identifies valid adjustment sets for causal relationships.
result Ensures soundness and completeness of causal effect estimation under standard assumptions.

Efficient algorithm for learning halfspaces in a new model with polynomial time complexity.

problem Learning halfspaces in the testable learning model with distributional constraints.
method Developed new tests using labels and combined with moment-matching approach.
result Achieved near optimal error rates for Gaussian and strongly log-concave distributions.

Polynomial-time algorithm for learning halfspaces with Gaussian-distributed data and adversarial noise.

problem Learning halfspaces in the presence of adversarial label noise.
method Iterative soft localization technique enhanced with appropriate testers.
result Output a halfspace with misclassification error $O(\opt)+\eps$.

We introduce the hemicubic codes, a family of quantum codes obtained by associating qubits with the pp-faces of the nn-cube (for n>pn>p) and stabilizer constraints with faces of dimension (p±1)(p\pm1). The quantum code obtained by identifying antipodal faces of the resulting complex encodes one logical qubit into $N = 2^…

2019-11-08abs ↗pdf ↗

Detect hidden confounding in observational data using multiple environments.

problem Detect hidden confounding in observational data.
method Theoretical framework and simulation studies to test for hidden confounding.
result The proposed procedure correctly predicts hidden confounding, especially when bias is large.

This paper studies structure detection problems in high temperature ferromagnetic (positive interaction only) Ising models. The goal is to distinguish whether the underlying graph is empty, i.e., the model consists of independent Rademacher variables, versus the alternative that the underlying graph contains a subgraph…

2018-09-21abs ↗pdf ↗

Identifying components and estimating mixing weights in unlabeled finite mixtures under marginal independence.

problem Identifying components and estimating mixing weights in unlabeled finite mixtures.
method Proving structural results and extending them to observable mixtures.
result Identifying components and estimating mixing weights under marginal independence.

Significant pattern mining, the problem of finding itemsets that are significantly enriched in one class of objects, is statistically challenging, as the large space of candidate patterns leads to an enormous multiple testing problem. Recently, the concept of testability was proposed as one approach to correct for mult…

2015-08-24abs ↗pdf ↗

We propose a method to learn causal response representations through direct effect analysis.

problem Uncovering direct causal effects in complex, multivariate settings.
method Our method bridges conditional independence testing with causal representation learning, formulating an optimisation problem to maximise evidence against conditional independence.
result The largest eigenvalue distribution can be bounded by an FF-distribution, providing testable conditional independence.

The paper introduces tests for missing data models based on graph assumptions.

problem Verification of assumptions in missing data models is insufficiently addressed.
method The paper explores three classes of missing data models and designs goodness-of-fit tests.
result The paper provides new insights and tests for missing data graphical models.

New theory for local parameterization of deep ReLU networks.

problem Determining local parameters of deep ReLU neural networks.
method Introducing local lifting operators and charts of a manifold, deriving necessary and sufficient conditions for local identifiability.
result Sharp and testable conditions for local identifiability of deep ReLU networks.

A new approach to learning in brain-like networks using adversarial algorithms.

problem Complex inter-dependencies in brain-like networks not compatible with conditional independence assumptions.
method Adversarial algorithm for learning models of perceptual processing.
result The approach can mimic known neural phenomena and yields testable hypotheses.

Universal tester-learner for halfspaces over structured distributions.

problem Learning halfspaces over a wide class of structured distributions.
method Uses a fully polynomial tester-learner based on hypercontractivity and sum-of-squares (SOS) programs.
result Achieves error O(opt)+εO(\mathrm{opt}) + ε on any labeled distribution that the tester accepts.

Probabilistic graphical models (PGMs) have become a popular tool for computational analysis of biological data in a variety of domains. But, what exactly are they and how do they work? How can we use PGMs to discover patterns that are biologically relevant? And to what extent can PGMs help us formulate new hypotheses t…

2007-06-14abs ↗pdf ↗

Reinforcement learning (RL) tasks are challenging to implement, execute and test due to algorithmic instability, hyper-parameter sensitivity, and heterogeneous distributed communication patterns. We argue for the separation of logical component composition, backend graph definition, and distributed execution. To this e…

2018-10-21abs ↗pdf ↗

We study random 2-dimensional complexes in the Linial - Meshulam model and find torsion in their fundamental groups at various regimes. We find a simple algorithmically testable criterion for a subcomplex of a random 2-complex to be aspherical; this implies that any aspherical subcomplex of a random 2-complex satisfies…

2013-07-13abs ↗pdf ↗

Tests whether a treatment's effect is fully mediated by observed outcomes and identifies causal mechanisms.

problem Understanding how a treatment affects an outcome through intermediate variables.
method Proposes a test to evaluate full mediation and causal mechanism identification, extending to non-randomly assigned treatments.
result A conditionally random treatment is conditionally independent of the outcome given mediators and covariates if full mediation and causal mechanism identification hold.

The semantic map calibrates uncertainty from language model probabilities.

problem Uncertainty in language model probabilities for professional decisions.
method Prespecified semantic map linking probabilities of verbal responses to probabilities of declared states.
result Language-derived probabilities outperform printed numerical probabilities and recover valid uncertainty coverage.

Finding statistically significant interactions between binary variables is computationally and statistically challenging in high-dimensional settings, due to the combinatorial explosion in the number of hypotheses. Terada et al. recently showed how to elegantly address this multiple testing problem by excluding non-tes…

2014-07-04abs ↗pdf ↗

A new framework for adaptive behavior using reusable value profiles.

problem Adaptive behavior in changing environments requires switching among value-control regimes, but maintaining separate parameters for each situation is impractical.
method Introduces value profiles: reusable bundles of parameters assigned to hidden states, allowing for state-conditional strategy recruitment without independent parameters for each context.
result Profile-based models outperform simpler alternatives in probabilistic reversal learning, suggesting belief-dependent control of adaptive behavior.

We use standard physics techniques to model trading and price formation in a market under the assumption that order arrival and cancellations are Poisson random processes. This model makes testable predictions for the most basic properties of a market, such as the diffusion rate of prices, which is the standard measure…

2001-12-23abs ↗pdf ↗

Representation and learning of long-range dependencies is a central challenge confronted in modern applications of machine learning to sequence data. Yet despite the prominence of this issue, the basic problem of measuring long-range dependence, either in a given data source or as represented in a trained deep model, r…

2019-04-08abs ↗pdf ↗

New framework explains adversarial examples in neural nets.

problem Adversarial examples in deep neural networks are counterintuitive and hard to explain.
method Introduces Dimpled Manifold Model to explain training phases and adversarial examples.
result Shows how adversarial examples arise from training dynamics and dimpling phase.

Criteria for extending degree-2 Azumaya algebras with C2-actions over curves.

problem Determining when degree-2 Azumaya algebras with C2-actions extend to entire curves.
method Criteria for extension of algebra and new condition for extension with action, testable by computer algebra systems.
result New conditions for extending degree-2 Azumaya algebras with C2-actions over curves.

The literature on structured prediction for NLP describes a rich collection of distributions and algorithms over sequences, segmentations, alignments, and trees; however, these algorithms are difficult to utilize in deep learning frameworks. We introduce Torch-Struct, a library for structured prediction designed to tak…

2020-02-03abs ↗pdf ↗

As machine learning algorithms are increasingly applied to high impact yet high risk tasks, such as medical diagnosis or autonomous driving, it is critical that researchers can explain how such algorithms arrived at their predictions. In recent years, a number of image saliency methods have been developed to summarize …

2017-04-11abs ↗pdf ↗

AI systems need reliable testing to ensure safety and trustworthiness.

problem Current AI Act lacks functional trustworthiness for AI systems.
method Define technical application distribution, set risk-based performance, and conduct statistically valid testing.
result Reliable functional trustworthiness is essential for AI systems.

ZDP detects drift in large language models without labels, proving key theorems and metrics.

problem Detecting drift in large language models without task labels or output evaluations.
method Zero-Direction Probing (ZDP) framework based on null directions of transformer activations, proving theoretical guarantees.
result Proves the Variance--Leak Theorem, Fisher Null-Conservation, Rank--Leak bound, and logarithmic-regret guarantee.

We provide a microfoundation for linear price impact models in a stationary market.

problem Deriving linear price impact models in a stationary market with asymmetric information.
method Deriving linear price impact models as the equilibrium of an agent-based system.
result The model shows compatibility with universal price diffusion at small times and non-universal mean-reversion at larger times.

New method identifies whether equity return predictability is due to magnitude shrinkage or directional reversal.

problem Determining the nature of equity return predictability (directional reversal vs magnitude shrinkage).
method Developed the Fourier-Residue Identity (FRI) to decompose return autocorrelation into sign and magnitude channels.
result The lag-1 autocorrelation in SPY is driven entirely by magnitude shrinkage, not directional reversal.

Traditional anatomical analyses captured only a fraction of real phenomic information. Here, we apply deep learning to quantify total phenotypic similarity across 2468 butterfly photographs, covering 38 subspecies from the polymorphic mimicry complex of Heliconius erato\textit{Heliconius erato} and Heliconius melpomene\textit{Heliconius melpomene}. E…

2019-08-15abs ↗pdf ↗

Generative model solves financial market equilibria with stable reinforcement learning.

problem Financial market equilibria under realistic frictions and multiple agents.
method Generative adversarial reinforcement learning with decoupling feedback.
result Algorithm learns and predicts asset returns and volatilities.

NFT royalties boost creator earnings by sharing risk, reducing info asymmetry, and enabling price discrimination.

problem NFTs' royalties are criticized for being neutralized by speculators.
method Analyzes NFTs' royalties in various market conditions and their effects on creators.
result Royalties enable creators to capitalize on speculators' presence through risk sharing, info reduction, and price discrimination.

The two-dimensional renormalization group acting as the Ricci flow ΛΛgμν=RμνΛ\frac{\partial}{\partialΛ} g_{μν} = R_{μν} produces a specific 1+3 dimensional space-time metric which describes an expanding universe that starts with a big bang at1/3a \sim t^{\scriptscriptstyle 1/\sqrt3} then decelerates until z=0.2z=0.2 then accelerate…

2019-09-03abs ↗pdf ↗

The paper clarifies the approximation of SGD with Ito SDEs for finite learning rates.

problem Theoretical justification and experimental verification of the Ito SDE approximation for finite learning rates in SGD.
method An efficient simulation algorithm SVAG and a necessary condition test for the SDE approximation.
result The Ito SDE approximation can meaningfully capture training and generalization properties of deep nets with finite learning rates.

We develop a theory for the market impact of large trading orders, which we call metaorders because they are typically split into small pieces and executed incrementally. Market impact is empirically observed to be a concave function of metaorder size, i.e., the impact per share of large metaorders is smaller than that…

2011-02-26abs ↗pdf ↗