Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

68137205273 · May 202619922001200920172026
48 results for synthetic equivalence

Timelike curvature and Brunn-Minkowski inequality linked in non-smooth spacetimes.

problem Equivalence between timelike Ricci curvature and Brunn-Minkowski inequality in synthetic Lorentzian spaces.
method Introducing strong qq-timelike Brunn-Minkowski condition and proving equivalence to curvature conditions.
result Timelike curvature dimension condition equivalent to timelike Brunn-Minkowski inequality in specific settings.

Equivalence found between smooth and synthetic timelike curvature bounds.

problem Understanding timelike sectional curvature bounds in spacetime geometry.
method Established equivalence between sectional curvature bounds on timelike planes and synthetic timelike bounds.
result Equivalence proved for sectional curvature bounds on timelike planes and synthetic timelike bounds on strongly causal spacetimes.

Paper proves Brunn-Minkowski inequality and curvature dimension condition are equivalent in weighted Riemannian manifolds.

problem Proving equivalence between Brunn-Minkowski inequality and curvature dimension condition.
method Analyzes weighted Riemannian manifolds, proving equivalence without optimal transport or differential structure.
result Brunn-Minkowski inequality and curvature dimension condition are equivalent in weighted Riemannian manifolds.

We study the sample complexity of private synthetic data generation over an unbounded sized class of statistical queries, and show that any class that is privately proper PAC learnable admits a private synthetic data generator (perhaps non-efficient). Previous work on synthetic data generators focused on the case that …

2019-02-09abs ↗pdf ↗

Alexandrov spaces with non-negative curvature are characterized by the matrix displacement convexity of an entropy tensor.

problem Characterizing non-negative curvature in Alexandrov spaces
method Constructing a parallel trivialization of the entropy tensor
result The entropy tensor is matrix displacement convex on Alexandrov spaces

A new algorithm for robust causal discovery in small sample sizes.

problem Limited data leads to weak conditional independence tests in causal discovery.
method Proposes a kk-PC algorithm that bounds conditioning set size for robust causal discovery.
result The kk-PC algorithm enables more robust causal discovery in small sample sizes.

The paper proves metrizability and dynamics of Weil bundles.

problem Metrizability and dynamics of Weil bundles in differential geometry.
method Investigation of metrizability and dynamics of Weil bundles for smooth compact manifolds and Weil algebras.
result A canonical, complete, weighted metric \(\mathfrak{d}_w\) on \(M^\mathbf{A}\) that encodes geometry and deformations.

Principal component regression (PCR) is a simple, but powerful and ubiquitously utilized method. Its effectiveness is well established when the covariates exhibit low-rank structure. However, its ability to handle settings with noisy, missing, and mixed-valued, i.e., discrete and continuous, covariates is not understoo…

2019-02-28abs ↗pdf ↗

Deep networks learn hierarchical data by invariant representations.

problem How many examples are needed for deep networks to learn hierarchical data?
method Random Hierarchy Model: synthetic tasks inspired by language and images hierarchy.
result Deep networks learn by invariant representations and require a detectable number of correlations between low-level features and classes.

Proposes a new method for generating synthetic data using copula flows.

problem Challenges of current synthetic data generation methods, especially with mixed real and categorical variables.
method Uses normalizing flows to learn copula density and univariate marginals based on copula theory.
result Demonstrates improved synthetic data generation and density estimation.

Synthetic framework for null hypersurfaces in non-smooth spacetimes.

problem Analyzing null hypersurfaces in non-smooth spacetimes.
method Develops synthetic null hypersurfaces using optimal transport and Lorentzian geometry.
result Synthetic null energy condition stabilizes under convergence and applies to low-regularity spacetimes.

We find a deterministic equivalent for random feature regression's test error, independent of feature map dimension.

problem Understanding the generalization performance of random feature ridge regression.
method We derive a deterministic equivalent for the test error of RFRR under a concentration property, showing it can be approximated by a closed-form expression dependent on feature map eigenvalues.
result Our approximation guarantee is non-asymptotic, multiplicative, and independent of the feature map dimension, providing a tight result for the smallest number of features achieving optimal minimax error rate.

This paper proposes and evaluates the k-greedy equivalence search algorithm (KES) for learning Bayesian networks (BNs) from complete data. The main characteristic of KES is that it allows a trade-off between greediness and randomness, thus exploring different good local optima. When greediness is set at maximum, KES co…

2012-10-19abs ↗pdf ↗

New findings show different cost functions yield equivalent curvature bounds.

problem Establishing equivalence of curvature bounds under various transport costs.
method Needle decomposition and localization technique for optimal transport.
result All CDp(K,N)\mathrm{CD}_{p}(K,N) conditions are equivalent for p>1p>1.

The paper develops methods to bound causal effects using Partial Ancestral Graphs.

problem Bounding causal effects from observational data when true causal diagrams are unknown.
method Proposes a method using Partial Ancestral Graphs to derive bounds on causal effects from observational data.
result Demonstrates the effectiveness of the method with synthetic and real data examples.

This paper provides a method for noise-calibrated inference from DP synthetic data.

problem Inference from DP synthetic data is often miscalibrated and lacks principled uncertainty quantification.
method Release DP sufficient statistics, perform noise-calibrated likelihood-based inference, and optional synthetic data generation.
result Asymptotic normality and valid confidence intervals for the plug-in DP MLE.

FWC creates fair synthetic samples for machine learning tasks.

problem Addressing biases in machine learning models for fair decision-making.
method FWC uses an efficient majority minimization algorithm to minimize Wasserstein distance while enforcing demographic parity.
result FWC achieves a competitive fairness-utility tradeoff and reduces biases in predictions from large language models.

Paper proposes scalable algorithm to estimate intervention targets in linear models.

problem Estimating intervention targets in linear models from observational and interventional data.
method The paper proposes a scalable algorithm that estimates intervention sites from the difference between precision matrices of observational and interventional datasets.
result The algorithm consistently identifies all intervention targets and updates observational Markov equivalence classes to interventional ones.

Latent feature models (LFM)s are widely employed for extracting latent structures of data. While offering high, parameter estimation is difficult with LFMs because of the combinational nature of latent features, and non-identifiability is a particularly difficult problem when parameter estimation is not unique and ther…

2018-09-11abs ↗pdf ↗

Synthetic speech data improves keyword spotting models with fewer real examples.

problem Training models for recognizing spoken keywords with limited real data.
method Used a pre-trained speech embedding model to extract features for training a small keyword spotting model.
result A model trained on synthetic speech data can detect 10 keywords with the same accuracy as a model trained on over 500 real examples.

Proposes ENVAR for causal discovery in structural VAR models with equal noise variance.

problem Challenges in causal discovery from multivariate time series with contemporaneous effects.
method Introduces observational equivalence and the observational alignment discrepancy for structural VAR models with equal noise variance.
result Shows that multiple structural VAR parameterizations can induce the same stationary observed process law.

K-Means and RBF networks are shown to be equivalent under certain conditions.

problem Discrete clustering vs. continuous optimization in machine learning.
method Established variational and gradient-based equivalence between K-Means and RBF networks.
result Gradient-based updates of RBF centers recover K-Means centroid update rule.

We will study metric measure spaces (X,d,m)(X,d,m) beyond the scope of spaces with synthetic lower Ricci bounds. In particular, we introduce distribution-valued lower Ricci bounds BE1(κ,)_1(κ,\infty) \bullet for which we prove the equivalence with sharp gradient estimates, \bullet the class of which will be preserved under…

2019-10-30abs ↗pdf ↗

Refines d'Alembertian for signed Lorentz distance functions in metric measure spacetimes.

problem Exact representation and bounds of d'Alembertian for signed Lorentz distance functions.
method Metric geometry techniques, localization, Sobolev calculus.
result Distributional d'Alembertian is a signed measure with integration by parts formula.

We discuss various characterizations of synthetic upper Ricci bounds for metric measure spaces in terms of heat flow, entropy and optimal transport. In particular, we present a characterization in terms of semiconcavity of the entropy along certain Wasserstein geodesics which is stable under convergence of mm-spaces. A…

2017-11-06abs ↗pdf ↗

We develop a method to summarize causal models with cycles in cubic time.

problem Cycles in high-dimensional causal models limit applicability of existing methods.
method We relax the acyclicity assumption in LiNG models and develop a low-dimensional DAG summary.
result Our method allows recovery of a low-dimensional DAG from high-dimensional data with cycles.

Model identifies causal structure from paired observational and interventional data with unknown soft interventions.

problem Identifying causal structure from observational and interventional data with unknown soft interventions.
method Proposes a scalable causal discovery model that aggregates subset-level PDAGs and applies contrastive cross-regime orientation rules.
result The model asymptotically recovers the identifiable PDAG and can orient additional edges compared to non-contrastive subset-restricted methods.

Extends field theory foundations to infinitesimal spaces, simplifying complex concepts.

problem Develop rigorous foundations for field theory, especially for infinitesimal spaces.
method Formulates local Lagrangian field theory in a new category of thickened smooth sets.
result Establishes a firm foundation for field theory, including tangent bundles and perturbative considerations.

We discretize a cost functional for image registration problems by deriving Taylor expansions for the matching term. Minima of the discretized cost functionals can be computed with no spatial discretization error, and the optimal solutions are equivalent to minimal energy curves in the space of kk-jets. We show that t…

2014-12-23abs ↗pdf ↗

We propose a framework for constructing and analyzing multiclass and multioutput classification metrics, i.e., involving multiple, possibly correlated multiclass labels. Our analysis reveals novel insights on the geometry of feasible confusion tensors -- including necessary and sufficient conditions for the equivalence…

2019-08-24abs ↗pdf ↗

Algorithm generates private continuous-time data for sensitive domains.

problem Private generation of continuous-time data for sensitive domains.
method Mean-field Langevin dynamics and noisy particle gradient descent.
result Strong privacy guarantees for one-time data contributions.

We present the Causal Gaussian Process Convolution Model (CGPCM), a doubly nonparametric model for causal, spectrally complex dynamical phenomena. The CGPCM is a generative model in which white noise is passed through a causal, nonparametric-window moving-average filter, a construction that we show to be equivalent to …

2018-02-22abs ↗pdf ↗