Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

25.0%50.0%75.0%100.0% · Jun 199319922001200920172026
48 results for pseudo data

Doubly robust self-training improves semi-supervised learning by balancing labeled and pseudo-labeled data.

problem Improving semi-supervised learning performance with limited labeled data.
method Introduces doubly robust self-training, a method that combines labeled and pseudo-labeled data to balance between labeled-only and pseudo-labeled-only training.
result Demonstrates superior performance of doubly robust self-training on ImageNet and nuScenes datasets.

Paper explores embedding methods for detecting pseudo-cliques in random graphs, showing limitations and potential.

problem Detecting planted pseudo-cliques in random dot product graphs.
method Adjacency Spectral Embedding (ASE) and Graph Encoder Embedding (GEE).
result These methods can localize pseudo-cliques with additional clean network data, but not without it.

Researchers solve a Plateau problem for maximal surfaces in pseudo-hyperbolic spaces.

problem Finding maximal surfaces with given boundary curves in pseudo-hyperbolic spaces.
method Defined and proved the existence of unique solutions using asymptotic Plateau problem and analysis of pseudo-holomorphic curves.
result Existence and uniqueness of maximal surfaces with specified boundary conditions.

Solves biased pseudo-labels in imbalanced SSL by refining them.

problem Imbalanced class distributions in semi-supervised learning lead to biased pseudo-labels.
method Formulates a convex optimization problem to refine pseudo-labels and develops an efficient algorithm, DARP.
result Demonstrates the effectiveness of DARP in various imbalanced semi-supervised scenarios.

Pseudo-label selection affects semi-supervised learning performance.

problem Selection of pseudo-labeled data impacts semi-supervised learning's generalization performance.
method Embedding pseudo-label selection into decision theory, deriving a novel selection criterion based on posterior predictive.
result BPLS (Bayesian pseudo-label selection) outperforms traditional methods in overfitting-prone data.

The paper improves self-training in semi-supervised learning by selecting more robust pseudo-labeled data.

problem Improving the reliability of pseudo-labeled data selection in self-training for semi-supervised learning.
method Proposes a multi-objective utility function to select pseudo-labeled data that maximizes reliability, considering model selection, accumulation of errors, and covariate shift uncertainties.
result Robustness towards model choice can lead to substantial accuracy gains in self-training.

A new framework for semi-supervised learning using pseudo-representation labeling.

problem Improving deep learning models with limited labeled data.
method Pseudo-representation labeling framework integrating pseudo-labeling and self-supervised representation learning.
result Outperforms state-of-the-art semi-supervised learning methods in industrial classification problems.

Pseudo rehearsal uses non-photo-realistic images to save resources without sacrificing performance.

problem Catastrophic forgetting in neural networks when learning new tasks.
method Synthetically generate non-photo-realistic images to rehearse previous tasks.
result Non-photo-realistic images can be used for rehearsal without sacrificing performance and significantly reduce resource consumption.

Framework for domain adaptation using pseudo-labels from unlabeled data.

problem Improving prediction accuracy in target domain with covariate shift.
method Kernel GLMs with labeled and pseudo-labeled data, using imputation model for target data.
result Non-asymptotic excess-risk bounds for effective labeled sample size.

A new pseudo-metric uses data depth to compare probability distributions.

problem Designing a metric between probability distributions for machine learning applications.
method Extension of univariate quantiles to multivariate spaces, using data depth and Hausdorff distance.
result The pseudo-metric is robust, factorizes translations, and has good behavior under transformations.

There has been increasing interest in modelling survival data using deep learning methods in medical research. Current approaches have focused on designing special cost functions to handle censored survival data. We propose a very different method with two steps. In the first step, we transform each subject's survival …

2019-08-06abs ↗pdf ↗

A method for selecting pseudo-labeled data in semi-supervised learning using generalized Bayes and soft revision.

problem Selecting pseudo-labeled data for semi-supervised learning with robustness to uncertainty.
method Using credal sets and the Gamma-Maximin method with soft revision to update priors and select pseudo-labeled data.
result The Gamma-Maximin method with soft revision can achieve promising results, especially in scenarios with low labeled data proportions.

MTL method uses unlabeled data with pseudo labels to improve classification with disjoint datasets.

problem Improving classification performance with disjoint labeled datasets using unlabeled data.
method Proposes MTL-SA method to select and augment unlabeled data with confident pseudo labels and close distribution to labeled data.
result Extensive experiments show the effectiveness of MTL-SA method in improving classification performance.

Method uses pseudo-samples to improve RCT data in ride-hailing pricing studies.

problem Small and biased RCT data leads to significant bias when generalizing to broader user base.
method Pseudo-sample matching to expand and match RCT data with observational data.
result 0.41% improvement in profit through pseudo-sample matching.

New insights into pseudo-Anosov flows with special periodic orbits.

problem Understanding pseudo-Anosov flows with periodic orbits in 3-manifolds.
method Analyzing the topological features corresponding to trees of scalloped regions and classifying flows with the same free homotopy data.
result Explicit examples of flows with the same free homotopy data but not orbit equivalent.

SurvFM-RMST converts survival outcomes into pseudo-observation targets for tabular models.

problem Right-censored follow-up prevents direct use of survival labels in tabular patient data.
method SurvFM-RMST framework that converts survival outcomes into jackknife pseudo-observation targets for restricted mean survival time.
result SurvFM-RMST accurately recovered restricted event-free time in simulations and outperformed naive targets in static datasets.

Contrastive regularization improves semi-supervised learning by better propagating confident pseudo-labels.

problem Consistency regularization's limitation in high performance and efficiency.
method Proposes contrastive regularization to update model features, pushing confident labels into unlabeled samples.
result Improves semi-supervised learning tasks with fewer training iterations and robust performance.

We first define Pseudo-Calabi flow, as {equation*} {{aligned}{{\partial \varphi}\over {\partial t}}&= -f(\varphi), \triangle_varphi f(\varphi) &= S(\varphi) - \ul S.{aligned}. \end{equation*} Then we prove the well-posedness of this flow including the short time existence, the regularity of the solution and the continu…

2010-04-15abs ↗pdf ↗

Paper establishes convergence rates for learning elliptic pseudo-differential operators.

problem Learning elliptic pseudo-differential operators in partial differential equations.
method Wavelet-Galerkin framework, structured infinite-dimensional regression problem, sparse estimator, matrix compression, nested-support strategy.
result Obtained convergence rates for the estimator and efficient Galerkin solver.

Combines pseudo-point and state space approximations for scalable GPs.

problem Handling large numbers of off-the-grid spatial data-points and long time-series.
method Combines pseudo-point approximations for spatial data with state space GP approximations for temporal data.
result Combined approach is more scalable and applicable to a greater range of spatio-temporal problems.

In this work, we study the pseudo-Riemannian submanifolds of a pseudo-sphere with 1-type pseudo-spherical Gauss map. First, we classify the Lorentzian surfaces in a 4-dimensional pseudo-sphere Ss4(1)\mathbb{S}^4_s(1) with index s, s=1,2s=1, 2, and having harmonic pseudo-spherical Gauss map. Then we give a characterization the…

2015-10-28abs ↗pdf ↗

Extends tangle theory to include undetermined crossings in periodic structures.

problem Classical tangle theory's limitations in handling undetermined crossings.
method Introduces pseudo DP tangles, defined as liftings of pseudo motifs in the thickened torus, and analyzes them through diagrammatic methods.
result Defines equivalence for pseudo DP tangles and proves an analogue of Reidemeister theorem.

Mitigates confirmation bias in SSL by adjusting pseudo labels dynamically.

problem Confirmation bias in semi-supervised learning leads to errors in pseudo labels.
method TaMatch framework adjusts scaling ratio to debias pseudo labels and dynamically adjusts target distribution.
result TaMatch significantly outperforms existing methods in SSL tasks.

In this paper, we derived biharmonic equations for pseudo-Riemannian submanifolds of pseudo-Riemannian manifolds which includes the biharmonic equations for submanifolds of Riemannian manifolds as a special case. As applications, we proved that a pseudo-umbilical biharmonic pseudo-Riemannian submanifold of a pseudo-Rie…

2015-12-08abs ↗pdf ↗

Proposes ConstraintMatch for semi-supervised clustering with unconstrained data.

problem Leveraging unconstrained data alongside constraints for clustering models.
method Semi-supervised context with pseudo-constraining and pseudo-labeling mechanisms.
result Demonstrates effectiveness of ConstraintMatch over baselines.

The paper studies pseudo links in genus g handlebodies, generalizing knot theory.

problem Modeling DNA knots with missing crossing information.
method Introducing pseudo links as mixed pseudo links in S^3, generalizing Kauffman bracket polynomial and Alexander theorem.
result The theory of pseudo links is closely related to singular links and can be applied to study singular links in genus g handlebodies.

Paper constructs a HOMFLYPT-type invariant for pseudo links.

problem Inability to construct polynomial invariants for pseudo links using Hecke algebra techniques.
method Using a resolution homomorphism and pseudo Hecke algebra of type \(A\), the paper constructs a HOMFLYPT-type invariant for oriented pseudo links.
result The constructed invariant satisfies a natural pseudo skein relation and admits a state-sum formulation.