MWGAN tackles multi-marginal matching problem with Wasserstein GAN.
problem Learning mappings to match a source domain to multiple target domains with cross-domain correlations.
method Develops a novel Multi-marginal Wasserstein GAN (MWGAN) with inner- and inter-domain constraints to minimize Wasserstein distance.
result Theoretical and empirical evaluations show MWGAN's effectiveness on balanced and imbalanced translation tasks.
3MSBM learns smooth trajectories from multiple snapshots.
problem Capturing long-range temporal dependencies in complex systems.
method Lifts dynamics to phase space, generalizes stochastic bridges to multi-marginal conditional problems, learns transport maps preserving intermediate marginals.
result Significantly improves convergence and scalability in capturing complex dynamics.
GMC benchmark isolates retrieval in Transformers, revealing max-margin alignment.
problem Understanding how Transformers develop match-and-copy behavior on natural data.
method Introducing Gaussian Match-and-Copy (GMC) as a minimalist benchmark.
result Gradient descent drives parameters to diverge while aligning with max-margin separator.
New margin bound improves generalization for voting classifiers.
problem Improving generalization bounds for voting classifiers.
method Established a new margin-based generalization bound.
result Derives an optimal weak-to-strong learner with matching theoretical lower bound.
Generative models often fail to preserve joint structure despite matching marginals.
problem Generative models fail to capture complex dependencies beyond univariate marginals.
method Introduced D_Sigma(P,Q) = ||Sigma_P - Sigma_Q||_F to measure covariance-level dependence fidelity.
result Covariance-level divergence can lead to structural instability in downstream inference.
New lower bounds nearly match existing upper bounds for boosted classifiers.
problem Understanding the generalization performance of boosted classifiers.
method Margin-based lower bounds on boosted classifiers.
result Lower bounds nearly match the kth margin bound, settling the generalization performance of boosted classifiers. Matching only the marginal distribution of latent style variables in factorized models fails to prevent class leakage.
problem Class leakage in factorized generative models despite matching marginal distributions.
method Derive an exact decomposition showing four conditions required for factorized sampling, and demonstrate that matching only the marginal distribution is insufficient.
result Class labels can be recovered with high accuracy (74%--100%) from factorized generative models, indicating leakage.
MSBM extends SB for multi-marginal trajectory inference.
problem Trajectory inference from multiple discrete snapshots.
method Multi-Marginal Schrödinger Bridge Matching (MSBM) using iterative Markovian fitting (IMF).
result MSBM effectively captures complex trajectories and respects intermediate distributions.
A new method for efficient exploration in reinforcement learning.
problem Improving exploration in reinforcement learning agents.
method State Marginal Matching (SMM) to learn policies matching a target state distribution.
result Agents that optimize SMM explore faster and adapt quicker to new tasks.
ScoreMatchingRiesz improves debiased machine learning and policy effects estimation.
problem Improving debiased machine learning and policy effects estimation.
method Score matching and Riesz representer estimation.
result Estimates policy path for continuous treatments, improving interpretability.
We give polynomial-time algorithms for the exact computation of lowest-energy (ground) states, worst margin violators, log partition functions, and marginal edge probabilities in certain binary undirected graphical models. Our approach provides an interesting alternative to the well-known graph cut paradigm in that it …
The paper improves SVM margin-based generalization bounds.
problem Improving generalization bounds for SVMs.
method Revisiting and improving classic generalization bounds in terms of margins, complementing with a nearly matching lower bound.
result Almost settles the generalization performance of SVMs in terms of margins.
Improves EM algorithm for better local optima in mixture models.
problem EM algorithm's sensitivity to initialization and bad local optima.
method Big Learning principle applied to upgrade EM algorithm.
result BigLearn-EM delivers optimal solution with high probability.
Score matching errors are not sufficient for measuring diffusion model quality.
problem The L2 score matching error is not a reliable measure of diffusion model performance. method Decomposed score errors into gradient and solenoidal components and analyzed their geometric properties.
result Only the gradient component of the score error affects the marginal distributional quality.
Improved flow matching using Gaussian processes for better sample quality.
problem Training continuous normalizing flows with reduced variance and flexibility.
method Extending conditional flow matching to streams modeled with Gaussian processes.
result Improved quality of generated samples with moderate computational cost.
A new imputation method estimates missing values by matching observed marginals from masked data.
problem Missing values in data undermine statistical and machine learning analysis.
method Estimates a distribution from masked observations using positive semi-definite kernel density estimation.
result The method yields both single and multiple imputations from the same fitted density, with statistical consistency and fast adaptive excess risk.
New findings show score matching's accuracy doesn't ensure numerical stability in diffusion sampling.
problem Numerical stability issues in diffusion sampling despite small forward-marginal error.
method Constructing a smooth score field with arbitrarily small forward-marginal L2 error, showing nonexplosive behavior and moments of every order. result Euler--Maruyama discretizations can converge in probability even when moments diverge, demonstrating failure of weak convergence.
New algorithm preserves transport maps for better diffusion model training.
problem Training diffusion models with task-specific optimality structures.
method Generalized Schrödinger Bridge Matching (GSBM), inspired by conditional stochastic optimal control.
result GSBM better preserves transport maps, enabling stable convergence and improved scalability.
New geometric analysis shows L2 score error is flawed for diffusion models.
problem Score matching errors in diffusion models do not fully capture distributional quality.
method Decomposed score errors into gradient and solenoidal components, focusing on gradient's role in Fokker-Planck dynamics.
result Only gradient component affects marginal distributional quality; solenoidal component is structurally invisible.
New framework using Jensen-Shannon divergence improves domain adaptation theory.
problem Incoherence between empirical domain adversarial training and theoretical H-divergence. method Established new theoretical framework based on Jensen-Shannon divergence, derived bi-directional upper bounds.
result Framework exhibits flexibilities for various transfer learning problems.
Study examines time-varying betas and their volatility in bank interest income and expense margins.
problem Understanding the variability of bank betas and their impact on net interest margins.
method Used state-space methods to estimate time-varying betas and conditional volatility.
result Substantial variation in interest income and expense betas, leading to varying net interest margin coefficients.
Paper proves tight lower bounds for online multicalibration, separating it from marginal calibration.
problem Proving lower bounds for online multicalibration in relation to marginal calibration.
method Information-theoretic approach, constructing group families from orthonormal bases.
result Establishes tight lower bounds for online multicalibration, matching upper bounds up to logarithmic factors.
The paper models asset prices with random volatility to match option prices.
problem Matching asset price dynamics with observed option prices.
method Uses a mixture of diffusion processes with random volatility.
result Derives explicit pricing formulas for derivatives.
The paper exposes VAEs' limitations in learning marginal distributions and proposes VAE-GAN hybrids as a solution.
problem VAEs fail to learn marginal distributions in latent and visible spaces.
method Analyzed VAEs and proposed VAE-GAN hybrids as a solution.
result VAE-GAN hybrids are harder to scale, evaluate, and use for inference compared to VAEs.
We address the problem of learning the parameters in graphical models when inference is intractable. A common strategy in this case is to replace the partition function with its Bethe approximation. We show that there exists a regime of empirical marginals where such Bethe learning will fail. By failure we mean that th…
New research shows the maximum ℓ1-margin classifier doesn't adapt to sparse ground truths.
problem Understanding the limitations of the maximum ℓ1-margin classifier in high-dimensional settings.
method Analyzing convergence and prediction error rates of the maximum ℓ1-margin classifier.
result Proves tight upper and lower bounds for prediction error, showing benign overfitting.
New method learns flows between multiple distributions efficiently.
problem Learning dynamic transport maps between multiple empirical distributions.
method Combining flow matching and dynamic optimal transport with potential terms.
result OTP-FM achieves state-of-the-art performance on various datasets.
Analyzes large-margin classifiers under high-dimensional data.
problem Selecting the best classifier among various margin-based methods.
method Investigates asymptotic performance of large-margin classifiers under two component mixture models.
result Analytical results closely match with Monte Carlo simulations.
New method improves calibration of neural networks by targeting robust margins and local smoothness.
problem Poor calibration of neural networks, leading to unreliable confidence estimates.
method Intervene on training procedure by targeting robust margins and local smoothness.
result Improved out-of-sample calibration without sacrificing accuracy.
Novel proof shows continuity of optimal transport feasible set mapping.
problem Continuity of feasible set mapping in optimal transport problems.
method Presented a novel and shorter proof of continuity.
result Established continuity of the feasible set mapping.
Unified framework learns matching from noisy data.
problem Learning adaptive interaction costs from incomplete data.
method Inverse optimal transport with marginal relaxation.
result Efficiently predicts new matching in various contexts.
Framework for private, noise-tolerant, and efficient learning algorithms.
problem Private and efficient learning of large-margin halfspaces in noisy environments.
method Simple framework using differential privacy and noise tolerance conditions.
result Noise-tolerant and private PAC learners for large-margin halfspaces with sample complexity independent of dimension.
Framework learns continuous dynamics from sparse trajectories.
problem Learning dynamics from sparsely sampled and high-dimensional trajectories.
method Interpolative Multi-Marginal Flow Matching (IMMFM) framework.
result IMMFM outperforms existing methods in forecasting and downstream tasks.
Causal invariance can improve finite-sample domain adaptation, but only when the target risk margins are large.
problem Finite-sample domain adaptation
method Linear regression with causal knowledge
result Adaptive aggregation can match best candidate predictor while avoiding negative transfer
New method controls error in low-dimensional marginals of spatial models.
problem Inaccurate approximation of low-dimensional marginals in spatial models.
method Stein's method with δ-locality condition for spatial models.
result Uniform error bound for marginals of approximate distributions.
We develop a framework for post model selection inference, via marginal screening, in linear regression. At the core of this framework is a result that characterizes the exact distribution of linear functions of the response y, conditional on the model being selected (``condition on selection" framework). This allows…
Projects Markovian processes from Itô semimartingales with jumps.
problem Modeling Itô semimartingales with jumps using Markovian projections.
method Construct Markovian projections for Itô semimartingales with jumps using non-local FPKEs.
result Markovian projections match the marginal laws of the original process.
Generative models learn latent process to match target distributions.
problem Training flow-matching models with auxiliary stochastic dynamics.
method Introduces latent process generator matching, treating generative state as a deterministic image of a Markov process.
result Learn generator of a stochastic process with same marginal distributions.
MARGINATTACK improves zero-confidence adversarial attacks' accuracy and efficiency.
problem Improving zero-confidence adversarial attacks' accuracy and efficiency.
method Proposes MARGINATTACK, a zero-confidence attack framework that computes margin with improved accuracy and efficiency.
result MARGINATTACK computes a smaller margin than state-of-the-art zero-confidence attacks and matches state-of-the-art fix-perturbation attacks.
Augmented bridge matching preserves coupling information between distributions.
problem Preserving the original empirical pairing in flow and bridge matching processes.
method Augmenting the velocity field with initial sample point information.
result Simple modification recovers coupling information without losing Markovian property.
Paper proposes a new method for training diffusion models using Markov operators.
problem Training efficiency and accuracy in diffusion models.
method Operator-informed score matching using spectral decomposition of Markov operators.
result Improved score matching for both low and high-dimensional distributions.
New method solves tree-structured Schrödinger Bridge problems.
problem Computing Schrödinger Bridge between tree-structured distributions.
method Iterative Markovian Fitting (IMF) procedure for tree-structured costs.
result Extends IMF to tree-structured Schrödinger Bridge problems.
This paper introduces Gumbel-Sinkhorn networks for learning latent matchings.
problem Learning in latent variable models with permutations is difficult due to combinatorial intractability.
method Approximates maximum-weight matching using the Sinkhorn operator, extending Gumbel-Softmax.
result Demonstrates effectiveness on sorting, jigsaw puzzles, and neural signal identification tasks.
New method uncovers zero entropy in dependent observations after finite samples.
problem Understanding uncertainty reduction in dependent observations.
method Minimum list entropy coupling, greedy algorithm.
result Zero entropy achieved with O(log(1/P_min)) samples for dependent observations.
This study examines biases in flow matching samplers using finite-sample estimation.
problem Biases in flow matching samplers when using finite-sample surrogates.
method Finite-sample plug-in estimation and hierarchy of empirical FM models.
result Exact empirical minimizer and smoothed plug-in regime identified for affine conditional flows.
We improve GANs by enforcing reproducibility and using non-uniform sampling.
problem Overrepresentation of certain samples in GANs' marginal log-likelihood.
method Enforce reproducibility through matching empirical distribution to prior, use non-uniform sampling for mini-batch selection.
result Improved quality and variety in generated samples, validated on CIFAR10, Fashion MNIST, and CelebA.
A novel method for Bayesian predictive distribution modeling with neural nets.
problem Modeling and quantifying prediction uncertainty in neural networks.
method Evidential Deep Learning, Bayesian Neural Net, progressive moment matching, PAC bound.
result Improves model fit and uncertainty quantification on various benchmarks.
A new method for unsupervised domain adaptation using Gaussian processes.
problem Reducing target domain error by aligning input and output distributions.
method Max-margin Gaussian process approach to achieve hypothesis consistency.
result Our method effectively minimizes maximum discrepancy and maximizes margins.