Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

52104155207 · Jun 202019922001200920172026
48 results for long gaps

GACELA fills long gaps in musical audio with a GAN and context conditioning.

problem Restoring long gaps in musical audio with varying complexity and duration.
method Generative adversarial network (GAN) with five parallel discriminators and context conditioning.
result Reduced artifacts in inpaintings from unacceptable to mildly disturbing.

Sparse attention model reduces long-context inference time with exponential accuracy guarantees.

problem Efficiently processing long-context queries in large language models.
method Formalizes attention as a projection onto key vectors, analyzes entropic relaxation, and introduces Vashista Sparse Attention.
result Sparse attention concentrates on a constant-size active face, leading to exponential decay of inactive tokens' mass and linear scaling of active face error.

Generative adversarial network improves audio inpainting for long gaps.

problem Generating missing audio content in long-range gaps using WGAN.
method Proposed WGAN architecture with short-range and long-range neighboring borders.
result The proposed model outperforms classical WGAN in reconstructing high-frequency content.

LLapDiff models irregular multivariate time series without step-by-step integration.

problem Trade-off between discrete and continuous methods for long-horizon forecasting.
method Generative framework that models target as a low-dimensional latent trajectory, guided by modal parameterization and Laplace domain poles.
result Improves long-horizon forecasting over baselines and supports missing-value imputation.

New theorem shows curvature concentration depends linearly on volume ratio.

problem Gap theorem for nonnegative Ricci curvature manifolds with small curvature concentration.
method Exhibited Ricci flow solution with faster than 1/t curvature decay.
result Curvature concentration depends linearly on asymptotic volume ratio.

As a testament to their success, the theory of random forests has long been outpaced by their application in practice. In this paper, we take a step towards narrowing this gap by providing a consistency result for online random forests.

2013-02-20abs ↗pdf ↗

Study reveals 1/f1/f noise in signals made from nonoverlapping rectangular pulses.

problem Analyzing 1/f1/f noise in signals composed of nonoverlapping pulses.
method Derived a general formula for power spectral density, analyzed rectangular pulse case.
result Observed pure 1/f1/f noise until very low frequencies with long pulse durations.

Study shows exponential gap in sample complexity between noisy and non-noisy recurrent neural networks.

problem Understanding the impact of noise on the sample complexity of recurrent neural networks.
method Analyzing noisy multi-layered sigmoid recurrent neural networks with independent noise and proving lower bounds.
result Exponential gap in sample complexity between noisy and non-noisy networks, even for small noise values.

New method fuses optical and SAR data to fill LAI gaps during cloudy periods.

problem Cloudy periods mask key crop growth stages, leading to unreliable yield predictions.
method Multi-Output Gaussian Process (MOGP) regression for fusing Sentinel-1 RVI and Sentinel-2 LAI time series.
result MOGP provides improved LAI estimations even during cloudy periods, especially for long gaps.

TimeBridge addresses non-stationarity in long-term time series forecasting.

problem Non-stationarity in multivariate time series leads to spurious regressions and obscures long-term relationships.
method TimeBridge segments series into patches, applying Integrated Attention for short-term non-stationarity and Cointegrated Attention for long-term cointegration.
result TimeBridge achieves state-of-the-art performance in both short-term and long-term forecasting.

Neural ARFIMA model improves exchange rate forecasting for BRIC economies.

problem Forecasting exchange rates for emerging markets with long-term memory and nonlinear dynamics.
method Integrates ARFIMA for long-memory with neural networks for nonlinear approximation.
result NARFIMA model outperforms benchmarks in BRIC exchange rate forecasting.

The paper addresses estimating long-term treatment effects with monotone missing data.

problem Estimating long-term treatment effects with missing data, especially monotone missing.
method The paper introduces the sequential missingness assumption for identification and proposes three novel estimation methods: inverse probability weighting, sequential regression imputation, and SeqMSM. It also introduces a balancing-enhanced approach, BalanceNet, to improve estimation accuracy.
result The proposed methods, including BalanceNet, effectively estimate long-term treatment effects with monotone missing data.

TAET tackles long-tailed distributions in adversarial robustness.

problem Long-tailed distributions complicate adversarial robustness in real-world applications.
method TAET integrates an initial stabilization phase followed by a stratified equalization adversarial training phase.
result TAET achieves significant improvements in robustness and efficiency.

Chaos in cerebellar cells enhances complexity of neural patterns.

problem Understanding how cerebellar granular layer represents complex information.
method Constructed a model of cerebellar granular layer with gap junctions, evaluated using reservoir computing.
result Chaotic dynamics in the cerebellar granular layer produce complex and diverse output patterns.

Persistent homology enhances graph classification by capturing long-range graph properties.

problem Lack of formal assessment of persistent homology in graph learning.
method Brief introduction and theoretical discussion of persistent homology in graph context, followed by empirical analysis.
result Persistent homology improves graph classification, especially for data with prominent topological structures.

The study confirms a conjecture about Kähler manifolds with quasi-negative curvature.

problem Confirming a long-standing conjecture about Kähler manifolds with quasi-negative curvature.
method Introducing (ε,δ)(\varepsilon,δ)--quasi-negativity and applying gap-type theorems.
result Obtained gap-type theorems for Xc1(KX)n>0\int_X c_1(K_X)^n>0 in terms of real bisectional curvature and weighted orthogonal Ricci curvature.

This paper examines multifractal dynamics in cryptocurrencies using two methodologies.

problem Understanding the multifractal nature of cryptocurrencies and their stochastic processes.
method Two alternative multi-scaling methodologies applied to 84 cryptocurrencies.
result Cryptocurrencies exhibit different degrees of long-range dependence and stochastic processes.

Novel approach predicts long-term seasonal component of electricity prices for improved forecasting.

problem Improving day-ahead electricity price forecasting accuracy.
method Extracts trend-seasonal pattern from extrapolated price series using autoregressive and LASSO models.
result Improves predictive accuracy by 3-15% in root mean squared error and 1% in profits.

New convergence rates for shuffling gradient methods without strong convexity.

problem Theoretical gap between shuffling gradient methods' empirical success and established convergence rates.
method Proved last-iterate convergence rates for shuffling gradient methods using function value gap.
result First last-iterate convergence rates for shuffling gradient methods without strong convexity.

RNNs struggle with in-context retrieval, while Transformers excel.

problem In-context retrieval capability of RNNs.
method Theoretical analysis and experimental techniques (CoT, RAG, Transformer layer).
result Enhancing RNNs with techniques improves their in-context retrieval capability, closing the representation gap with Transformers.

Improves spatio-temporal forecasting by reducing errors between training and inference.

problem Accumulation of small errors in Seq2Seq models during inference due to different distributions of training and inference phases.
method Curriculum learning based on Temporal Progressive Growing Sampling to replace some ground-truth context with generated predictions.
result Better models long-term dependencies and outperforms baseline approaches on two datasets.

New algorithm reduces regret in combinatorial causal bandits without graph structure.

problem Minimizing regret in combinatorial causal bandits without graph structure.
method Design of algorithms for binary general causal models and BGLMs without graph skeleton.
result Achieves O(TlnT)O(\sqrt{T}\ln T) expected regret for causal models and O(T23lnT)O(T^{\frac{2}{3}}\ln T) for BGLMs.

The liberalization of electricity markets and the development of renewable energy sources has led to new challenges for decision makers. These challenges are accompanied by an increasing uncertainty about future electricity price movements. The increasing amount of papers, which aim to model and predict electricity pri…

2017-03-31abs ↗pdf ↗

Time-related features improve time series forecasting models.

problem Lack of explicit time-related encoding in current forecasting models limits their ability to capture cyclical and seasonal trends.
method Introducing Time Stamp Forecaster (TimeSter) to encode time-related features and integrating it with a linear backbone.
result TimeLinear model reduces MSE by 23% on benchmark datasets, improving performance with exceptional efficiency.

New framework uses dynamics to justify Gaussian process for turbulent flows.

problem Lack of rigorous justification for Gaussian process priors in turbulent flows.
method Introduces a dynamics-informed Gaussian process framework based on quasi-Gaussianity.
result Provides a principled, long-time dynamical justified GP prior for turbulent flows.

Long-term lane change prediction model predicts maneuvers with 75% accuracy.

problem Predicting long-term lane changes for safer autonomous driving.
method Introduced three models: logistic regression, MLP, and RNN. Used NGSIM dataset with new labeling scheme.
result Developed model predicts 75% of lane changes with an average advanced time of 8.05 seconds.

Self-supervised learning performs better than supervised learning on imbalanced datasets.

problem The performance gap between balanced and imbalanced pre-training with self-supervised learning is smaller than with supervised learning.
method Systematic investigation of self-supervised learning under dataset imbalance, including experiments and theoretical analyses.
result Self-supervised representations are more robust to class imbalance than supervised representations.