Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

150300450600 · Jun 202019922001200920182026
48 results for Sampling patterns

Flexics samples patterns with guarantees, addressing flexibility and accuracy issues.

problem Pattern explosion and limited sampling accuracy with existing methods.
method Leverages SAT sampling and pattern mining algorithms to support flexible quality measures and constraints.
result Flexics provides strong guarantees on sampling accuracy while being flexible and efficient.

LOUPE optimizes MRI sub-sampling patterns using machine learning.

problem Optimizing sub-sampling patterns for MRI scans to improve reconstruction accuracy.
method End-to-end learning strategy combining sub-sampling pattern optimization and reconstruction model training.
result LOUPE yields more accurate reconstructions compared to standard under-sampling schemes.

LetSIP learns relevant patterns for user interests in data mining.

problem Redundancy in pattern mining makes it hard for analysts to identify relevant patterns.
method Combines pattern sampling with interactive data mining, using user feedback to learn sampling distribution.
result Favourable trade-offs in quality-diversity and exploitation-exploration compared to existing methods.

MCRapper efficiently computes patterns in data using Monte-Carlo Rademacher Averages.

problem Finding statistically significant patterns in data with limited samples.
method Monte-Carlo Empirical Rademacher Averages (MCERA) for poset families.
result MCRapper provides upper bounds to the discrepancy of functions, enabling efficient pattern mining.

R-GPM enables efficient graph pattern mining through user-defined relations.

problem Efficient graph pattern mining through user-defined relations.
method Parallel computing framework with MCMC sampling algorithm and optimizations.
result Efficient estimators for graph pattern statistics with up to 3-orders-of-magnitude computational cost reduction.

The paper analyzes conditions for low-rank tensor completion using TT decomposition.

problem Conditions for finite completability of low-rank tensors.
method Algebraic geometric analysis on the TT manifold, focusing on the independence of polynomials defined by sampling patterns and TT decompositions.
result Deterministic and probabilistic conditions for finite completability of tensors with high probability.

Two supervised methods classify single-molecule patterns from X-ray imaging.

problem Classifying high-quality patterns from noisy, stochastic XFEL data.
method Supervised template-based learning methods: Eigen-Image and Log-Likelihood classifiers.
result Classifiers can find best-matched templates within milliseconds and parallelize for XFEL repetition rate.

PbP strategy improves logistic model prediction with missing values.

problem Predicting with missing inputs in logistic models.
method Pattern-by-Pattern (PbP) strategy for logistic models with missing values.
result PbP accurately approximates Bayes probabilities under GPMM across various missing data scenarios.

The paper analyzes how sampling design affects machine learning model generalization.

problem The impact of sampling properties on machine learning model generalization.
method Spectral analysis of the generalization error in Euclidean space using Fourier analysis.
result Estimation of expected error bounds and convergence rates for various sampling patterns.

Moon phases added to stock market analysis for better pattern recognition.

problem Finding meaningful patterns in stock market data using irregular time sampling.
method Incorporating Moon phases into the Gregorian calendar time sampling methods for stock market analysis.
result Moon phases provide unique, irregular sampling features for stock market pattern recognition.

MPPN network improves long-term time series forecasting accuracy.

problem Inaccurate long-term time series forecasting due to noise and lack of interpretability.
method MPPN network constructs context-aware multi-resolution semantic units and employs multi-periodic pattern mining and channel adaptive module.
result MPPN significantly outperforms state-of-the-art methods on nine real-world benchmarks.

New method for tensor completion with reduced sampling requirements.

problem Tensors with low CP rank completion under sampling.
method Analyzing manifold structure and defining polynomials based on sampling pattern and CP decomposition.
result Deterministic and probabilistic conditions for finite and unique completability with reduced sampling requirements.

CAG method predicts nonlinear solid mechanics responses in real-time with high accuracy and efficiency.

problem Real-time prediction of nonlinear solid mechanics responses.
method Clustering adaptive Gaussian process regression (CAG) method.
result Offers predictions within a second with high precision using only 20 samples.

Generative Adversarial Networks create realistic geology from sparse measurements.

problem Building models of subsurface geology from sparse physical measurements.
method Semantic inpainting with Generative Adversarial Networks.
result Generated samples mimic a distribution of geological patterns, not a single image.

New guarantees for matrix completion from any deterministic sampling patterns.

problem Proving guarantees for low-rank matrix completion from non-random sampling schemes.
method Introduced a graph with observed entries as edges to analyze the performance of constrained nuclear norm minimization algorithm.
result The algorithm can successfully complete the matrix if the observation graph is well-connected and has similar node degrees.

CDPA identifies common and distinctive patterns in high-dimensional datasets.

problem Existing methods fail to capture the common pattern between coefficient matrices of shared latent factors.
method Proposes CDPA, an unsupervised learning method that incorporates both common and distinctive patterns of coefficient matrices.
result CDPA provides better characterization of common and distinctive patterns in high-dimensional datasets.

New framework for unbiased sampling of temporal networks.

problem Challenges in analyzing and modeling large, continuous temporal networks.
method General framework for unbiased temporal network sampling with online, single-pass algorithms and unbiased estimators.
result Effective algorithms for fast, accurate, and memory-efficient statistical estimation of temporal network patterns and properties.

Paper proposes efficient algorithm for recovering sparsity pattern from deterministic missing data.

problem Recovering sparsity pattern from datasets with deterministic missing structure.
method Proposes an efficient algorithm for missing value imputation using topological property of censorship filter.
result Consistently recovers the sparsity pattern with high probability in polynomial time and logarithmic sample complexity.

The standard approach to compressive sampling considers recovering an unknown deterministic signal with certain known structure, and designing the sub-sampling pattern and recovery algorithm based on the known structure. This approach requires looking for a good representation that reveals the signal structure, and sol…

2016-02-01abs ↗pdf ↗

Randomization helps verify if data mining results are due to inherent patterns.

problem Verify if data mining results are due to inherent patterns or coincidental findings.
method Metropolis sampling based on local swaps to randomize data while preserving discovered patterns.
result Randomized data often reveals that clustering results imply frequent pattern discovery.

Paper proposes active learning for hotspot detection in VLSI design.

problem Hotspot detection in VLSI design is computationally expensive and relies on costly reference libraries.
method Active learning-based layout pattern sampling and hotspot detection flow.
result Significantly reduces lithography simulation overhead with satisfactory detection accuracy.

FSR efficiently discovers significant patterns with few resampled datasets.

problem Mining significant patterns in transactional data, especially subgroups.
method FSR uses resampling to bound the supremum deviation of quality statistics, providing rigorous guarantees on false discoveries.
result FSR effectively discovers significant subgroups with a small number of resampled datasets.

New statistical measures assess group separability in low-dimensional geometrical spaces.

problem Lack of statistical measures to evaluate group separability in low-dimensional geometrical spaces.
method Proposed three statistical measures (PSI-ROC, PSI-PR, PSI-P) based on Projection Separability rationale.
result Statistical-based measures outperform traditional cluster validity indices in evaluating group separability.

Paper proposes TRA to learn multiple stock trading patterns.

problem Inconsistent i.i.d. assumption limits stock prediction performance.
method TRA architecture with Optimal Transport for pattern assignment.
result Improves information coefficient (IC) by 0.04-0.06 compared to baselines.

A distributed algorithm learns patterns in large images and signals.

problem High-dimensional optimization in large images and signals.
method Distributed asynchronous algorithm with locally greedy coordinate descent.
result Patterns can be learned on large scales images from the Hubble Space Telescope.

Guided warping augments time series data by aligning features with a teacher.

problem Small time series datasets limit neural network performance.
method Guided warping with a discriminative teacher to augment data deterministically.
result Significant improvement in performance on various time series datasets.

RotatE embeds knowledge graphs using rotations in complex space for better link prediction.

problem Predicting missing links in knowledge graphs.
method RotatE models relations as rotations in complex vector space, using self-adversarial negative sampling.
result RotatE models and infers various relation patterns, outperforming existing models.