Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

61123184245 · May 202619922001200920172026
48 results for product matching

A new machine learning model uses score matching to estimate probability densities efficiently.

problem Estimating probability density functions is challenging.
method Introduced a product Jacobi-Theta Boltzmann machine (pJTBM) and used score matching for efficient fitting.
result The pJTBM can fit probability densities more efficiently than the RTBM using score matching.

Given a matched pair of Lie groups, we show that the tangent bundle of the matched pair group is isomorphic to the matched pair of the tangent groups. We thus obtain the Euler-Lagrange equations on the trivialized matched pair of tangent groups, as well as the Euler-Poincaré equations on the matched pair of Lie algebra…

2015-12-21abs ↗pdf ↗

OmniMatch algorithm perfectly matches graphs without edge correlation.

problem Graph matching in the absence of edge correlation.
method OmniMatch algorithm for seeded multiple graph matching.
result OmniMatch aligns O(sα)O(s^α) unseeded vertices across multiple networks efficiently and perfectly.

New projection operators for multipatch spaces with stable properties.

problem Problems with non-matching interfaces in multipatch spaces.
method Construction of commuting projection operators on de Rham sequences of multipatch spaces with local tensor-product parametrization.
result Local and stable projection operators in any LpL^p norm for shape-regular spline patches with different mappings and local refinements.

Study dynamic assortment and positioning of products with varying display effects.

problem Dynamic assortment and positioning of products with varying display effects.
method Design round-based learning algorithms for both multiplicative and general position effects models, and develop efficient subroutines for optimization.
result First regret-optimal characterization for both models, with matching upper and lower bounds.

A new method for estimating complex models and high-dimensional data.

problem Difficulty in computing Hessian of log-density functions for complex models and high-dimensional data.
method Sliced score matching, which projects scores onto random vectors before comparison.
result Sliced score matching can learn deep energy-based models and produce accurate score estimates.

A new method uses a product of experts with Dirichlet variables to approximate complex distributions.

problem Approximating complex distributions with tractable models.
method A product of experts with auxiliary Dirichlet variables, using a Feynman identity to sample and optimize.
result The method efficiently approximates complex distributions using a product of experts and Dirichlet variables.

Research decouples Lie algebroids using bicocycle double cross product theory.

problem Understanding decoupling and coupling phenomena in Lie algebroids.
method Bicocycle double cross product realization method.
result Unified product, double cross product, semi-direct product, and cocycle extension frameworks are instances of the general method.

Normal distributions ensure asymptotic variance reduction in moment matching Monte Carlo.

problem Asymptotic variance reduction in general integration problems.
method Characterization of conditions for asymptotic variance reduction using normal distributions.
result Asymptotic variance reduction is guaranteed for normal distributions in moment matching Monte Carlo.

DeepCF combines representation learning and matching function learning for better recommendation.

problem Matching users and items with semantic gap in initial spaces.
method Unified framework combining representation learning and matching function learning.
result Demonstrates effectiveness on four datasets.

We propose graph kernels based on subgraph matchings, i.e. structure-preserving bijections between subgraphs. While recently proposed kernels based on common subgraphs (Wale et al., 2008; Shervashidze et al., 2009) in general can not be applied to attributed graphs, our approach allows to rate mappings of subgraphs by …

2012-06-27abs ↗pdf ↗

DIAL learns embeddings to maximize recall and accuracy for entity resolution.

problem Low resource settings for entity resolution with large Cartesian product search space.
method DIAL uses an Index-By-Committee framework with pre-trained transformer language models to jointly learn embeddings for recall and accuracy.
result DIAL achieves high precision, recall, and efficiency on benchmark datasets.

Diversifies reply suggestions for IM systems using M-CVAE.

problem Improving diversity of automated reply suggestions in instant messaging systems.
method Formulated a generative latent variable model with Conditional Variational Auto-Encoder (M-CVAE) to diversify responses.
result Increased diversity by ~30-40% without significant impact on relevance.

Adaptive algorithm for online evaluation of targeted audiences in advertising.

problem Determining the right match between advertising creatives and target audiences.
method Contextual bandit approach to address audience overlap and learn optimal display policies.
result The proposed method is more efficient than traditional split-testing methods.

New private identity testers for high-dimensional distributions with improved sample complexity.

problem Testing goodness-of-fit for high-dimensional product distributions under differential privacy.
method Developed novel differentially private testers for multivariate product distributions, including Gaussians and binary product distributions.
result Achieved sample complexity matching the minimax sample complexity of O(d1/2/α2)O(d^{1/2}/α^2) in many parameter regimes.

We propose a robust elastic net (REN) model for high-dimensional sparse regression and give its performance guarantees (both the statistical error bound and the optimization bound). A simple idea of trimming the inner product is applied to the elastic net model. Specifically, we robustify the covariance matrix by trimm…

2015-11-15abs ↗pdf ↗

AR-CSM models use derivatives of univariate log-conditionals to estimate joint distributions efficiently.

problem Scalability and stability issues in training autoregressive models.
method Parameterize joint distribution using derivatives of univariate log-conditionals and introduce Composite Score Matching (CSM) for efficient training.
result AR-CSM models are more scalable and stable compared to previous score matching algorithms.

Improves retrieval accuracy for hierarchical documents, especially for distant matches.

problem Limited expressive power of dual encoder models in hierarchical retrieval.
method Proves feasibility of DEs for HR, introduces pretrain-finetune recipe to improve long-distance retrieval.
result Pretrain-finetune boosts recall on long-distance pairs from 19% to 76%.

Researchers prove inner product recovery is impossible in latent space models.

problem Recovering inner products in latent space models with random geometric graphs.
method Rate-distortion theory applied to Gaussian or spherical latent locations.
result Impossible to recover inner products if dimensionality exceeds nh(p)n h(p), matching positive results' conditions.

A new method for flow matching reduces computational costs and improves performance.

problem Efficiently matching flow models to target data distributions.
method Semidiscrete formulation of optimal transport (SD-OT) using SGD and maximum inner product search (MIPS).
result Semidiscrete FM (SD-FM) outperforms batch-OT and traditional flow matching methods.

Prime Match protects client stock trades from market price manipulation.

problem Protecting client stock trades from market price manipulation.
method Prime Match uses a two-round secure linear comparison protocol to match orders without revealing information.
result Prime Match reduces market impact and maintains client privacy.

Framework for monitoring ML model performance in production without labels.

problem Monitoring real-time prediction quality of ML models in production without labels.
method ML Health framework using diagnostic methods to generate alerts for further investigation.
result The method outperforms standard distance metrics at detecting issues with mismatched data sets.

An online labor platform faces an online learning problem in matching workers with jobs and using the performance on these jobs to create better future matches. This learning problem is complicated by the rise of complex tasks on these platforms, such as web development and product design, that require a team of worker…

2018-09-18abs ↗pdf ↗

Proposes UICR to improve novelty in recommendation systems without sacrificing relevance.

problem Balancing relevance and novelty in recommendation systems is challenging, especially for long-tail items.
method Introduces uncertainty modeling in the matching stage and multi-task modeling of model and index uncertainty.
result Improves novelty without sacrificing relevance, as shown by experimental results and online A/B tests.

We consider the problem faced by a service platform that needs to match limited supply with demand but also to learn the attributes of new users in order to match them better in the future. We introduce a benchmark model with heterogeneous "workers" (demand) and a limited supply of "jobs" that arrive over time. Job typ…

2016-03-15abs ↗pdf ↗

GRAMPA spectral method solves graph matching problem with high probability.

problem Finding vertex correspondence between unlabeled graphs.
method GRAMPA constructs a similarity matrix from weighted eigenvector comparisons, rounding to produce a matching.
result GRAMPA exactly recovers correct vertex correspondence with high probability for Gaussian models.

It is argued that arguments for strict prohibition of interests must be based on the use of arguments from authority. This is carried out by first making a survey of so-called dialectical roots for interest prohibition and then demonstrating that for at least one important positive interest bearing financial product, t…

2011-05-14abs ↗pdf ↗

Combining deep learning and ensemble smoothers for better history matching.

problem Dealing with complex facies distributions in history matching.
method Using autoencoders and generative adversarial networks to parameterize facies models, applying distance-based localization.
result Improved history matching performance with deep learning parameterizations.

To compare entities of differing types and structural components, the artificial neural network paradigm was used to cross-compare structural components between heterogeneous documents. Trainable weighted structural components were input into machine-learned activation functions of the neurons. The model was used for m…

2018-01-09abs ↗pdf ↗

Let g\mathfrak{g} be a Leibniz algebra and EE a vector space containing g\mathfrak{g} as a subspace. All Leibniz algebra structures on EE containing g\mathfrak{g} as a subalgebra are explicitly described and classified by two non-abelian cohomological type objects: ${\mathcal H}{\mathcal L}^{2}_{\mathfrak{g}} \, (…

2013-07-09abs ↗pdf ↗