Study finds many Lagrangian fillings for certain Legendrian links.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New method uses cluster shapes to improve track finding in particle collisions.
Paper tackles graph matching with partially correct seeds, improving performance guarantees.
The paper shows how shared random seeds can reduce variance in machine learning evaluations.
New method stabilizes machine learning predictions across random seeds.
Fairness audits fail under missing protected labels, especially at zero access.
New method fills cluster seeds with exact Lagrangian structures.
Many methods for automated software test generation, including some that explicitly use machine learning (and some that use ML more broadly conceived) derive new tests from existing tests (often referred to as seeds). Often, the seed tests from which new tests are derived are manually constructed, or at least simpler t…
Relation extraction aims to extract relational facts from sentences. Previous models mainly rely on manually labeled datasets, seed instances or human-crafted patterns, and distant supervision. However, the human annotation is expensive, while human-crafted patterns suffer from semantic drift and distant supervision sa…
Bayesian optimization is a powerful tool for expensive stochastic black-box optimization problems such as simulation-based optimization or machine learning hyperparameter tuning. Many stochastic objective functions implicitly require a random number seed as input. By explicitly reusing a seed a user can exploit common …
In this paper, we focus on quantifying model stability as a function of random seed by investigating the effects of the induced randomness on model performance and the robustness of the model in general. We specifically perform a controlled study on the effect of random seeds on the behaviour of attention, gradient-bas…
New method finds 198,846 toric-colorable seeds of Picard number 5.
Optimized biopharmaceutical seed train design reduces variability and saves time.
Study on how intraclass variability affects Temporal Ensembling accuracy.
PNN-smoothing improves -means clustering by merging subsets' clusterings.
Given two graphs, the graph matching problem is to align the two vertex sets so as to minimize the number of adjacency disagreements between the two graphs. The seeded graph matching problem is the graph matching problem when we are first given a partial alignment that we are tasked with completing. In this paper, we m…
Data-aware methods for dimensionality reduction and matrix decomposition aim to find low-dimensional structure in a collection of data. Classical approaches discover such structure by learning a basis that can efficiently express the collection. Recently, "self expression", the idea of using a small subset of data vect…
Community detection is, at its core, an attempt to attach an interpretable function to an otherwise indecipherable form. The importance of labeling communities has obvious implications for identifying clusters in social networks, but it has a number of equally relevant applications in product recommendations, biologica…
We present a novel approximate graph matching algorithm that incorporates seeded data into the graph matching paradigm. Our Joint Optimization of Fidelity and Commensurability (JOFC) algorithm embeds two graphs into a common Euclidean space where the matching inference task can be performed. Through real and simulated …
Efficiently selects seed nodes to maximize content influence in unknown social networks.
Eradicating hunger and malnutrition is a key development goal of the 21st century. We address the problem of optimally identifying seed varieties to reliably increase crop yield within a risk-sensitive decision-making framework. Specifically, we introduce a novel hierarchical machine learning mechanism for predicting c…
We prove that for a generic -dimensional integrable rolling distribution of contact elements (excluding developable seed and isotropic developable leaves) isometric correspondence of leaves of a general nature (independent of the shape of the seed) requires the Bäcklund transformation.
Paper proves uniqueness of a complex construction.
PPM improves graph matching for correlated Gaussian Wigner models with high probability.
The well-known Ising model used in statistical physics was adapted to a social dynamics context to simulate the adoption of a technological innovation. The model explicitly combines (a) an individual's perception of the advantages of an innovation and (b) social influence from members of the decision-maker's social net…
We present a parallelized bijective graph matching algorithm that leverages seeds and is designed to match very large graphs. Our algorithm combines spectral graph embedding with existing state-of-the-art seeded graph matching procedures. We justify our approach by proving that modestly correlated, large stochastic blo…
Discrete knot theory models use lattice-filtered graphs to detect merging knot components.
TOO optimizes stochastic epidemiological models by finding both parameter settings and random seeds.
OmniMatch algorithm perfectly matches graphs without edge correlation.
A cluster variety of Fock and Goncharov is a scheme constructed from the data related to the cluster algebras of Fomin and Zelevinsky. A seed is a combinatorial data which can be encoded as an matrix with integer entries, or as a quiver in special cases, together with formal variables. A mutation is a c…
We study a well known noisy model of the graph isomorphism problem. In this model, the goal is to perfectly recover the vertex correspondence between two edge-correlated Erdős-Rényi random graphs, with an initial seed set of correctly matched vertex pairs revealed as side information. For seeded problems, our result pr…
Regularization improves stability and consistency of sparse autoencoders.
We provide initial seedings to the Quick Shift clustering algorithm, which approximate the locally high-density regions of the data. Such seedings act as more stable and expressive cluster-cores than the singleton modes found by Quick Shift. We establish statistical consistency guarantees for this modification. We then…
The paper proves there are many Lagrangian fillings for Legendrian links of affine type.
The paper proposes a method to stabilize predictions by identifying causal variables using a seed variable.
Consistently checking the statistical significance of experimental results is one of the mandatory methodological steps to address the so-called "reproducibility crisis" in deep reinforcement learning. In this tutorial paper, we explain how the number of random seeds relates to the probabilities of statistical errors. …
The study finds many Lagrangian fillings for Legendrian links of specific types.
Study uses deep learning to predict mycotoxin levels in Irish oats.
A fast regime-split Black-Scholes implied volatility solver
The paper defines and studies discrete p-density and compression-radius profiles of lattice knots.
Recently, deep learning models play more and more important roles in contents recommender systems. However, although the performance of recommendations is greatly improved, the "Matthew effect" becomes increasingly evident. While the head contents get more and more popular, many competitive long-tail contents are diffi…
Clustering is an extensive research area in data science. The aim of clustering is to discover groups and to identify interesting patterns in datasets. Crisp (hard) clustering considers that each data point belongs to one and only one cluster. However, it is inadequate as some data points may belong to several clusters…
A new protocol evaluates small machine learning improvements conservatively.
New method expands seed genes to functionally related clusters.
Consider two networks on overlapping, non-identical vertex sets. Given vertices of interest in the first network, we seek to identify the corresponding vertices, if any exist, in the second network. While in moderately sized networks graph matching methods can be applied directly to recover the missing correspondences,…
Our analysis of financial data, in terms of super-exponential growth, suggests that the seed of the 2002/03 crisis of the Dutch supermarket giant AHOLD was planted in 1996. It became quite visible in 1999 when the post-bubble destabilization regime was well-developed and acted as the precursor of an inevitable collapse…
This article contains the first published example of a real economic balance sheet where the Solvency II ratio substantially depends on the seed selected for the random number generator (RNG) used. The theoretical background and the main quality criteria for RNGs are explained in detail. To serve as a gauge for RNGs, a…
We present a modern scalable reinforcement learning agent called SEED (Scalable, Efficient Deep-RL). By effectively utilizing modern accelerators, we show that it is not only possible to train on millions of frames per second but also to lower the cost of experiments compared to current methods. We achieve this with a …