Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,341 papers · 148 categories

Trend · papers per month

23456890 · May 202619922001200920182026
48 results for unique items

This report clarifies the distribution of unique items in bootstrap samples.

problem Understanding the role of duplicate items in bootstrap samples.
method Analyzes the distribution of unique items in bootstrap samples and derives a heuristic for normal approximation.
result Derives a heuristic for when a normal approximation is permissible for the distribution of unique items in bootstrap samples.

We consider the problem of learning soft assignments of NN items to KK categories given two sources of information: an item-category similarity matrix, which encourages items to be assigned to categories they are similar to (and to not be assigned to categories they are dissimilar to), and an item-item similarity mat…

2014-05-23abs ↗pdf ↗

Adaptive cascade submodular maximization tackles sequential selection under uncertainty.

problem Maximizing expected utility from a set of items with unknown states and continuation probabilities.
method Proposed adaptive cascade submodular functions and a 0.12 approximation algorithm.
result Identified a class of functions (adaptive cascade submodular) that many practical applications satisfy.

Expressive recommender models deliver accurate top-N recommendations.

problem Improving recommendation accuracy while maintaining interpretability.
method Normalized nonnegative models that assign probability distributions to users and items.
result Performance matches PureSVD, providing interpretable user and item representations.

HybridSVD combines user and item info for efficient, flexible recommendations.

problem Lack of effective methods for incorporating both user and item side information in collaborative filtering.
method Hybrid algorithm using PureSVD with generalized singular value decomposition and cold start solution.
result Superior performance compared to similar hybrid models on various datasets.

Study finds existence and non-uniqueness of cone spherical metrics on compact Riemann surfaces.

problem Existence and non-uniqueness of cone spherical metrics with prescribed singularities.
method Utilizing polystable extensions of line bundles, the study establishes three primary results concerning these metrics.
result Existence of multiple irreducible and reducible cone spherical metrics for certain effective divisors.

CSEAL uses cognitive structure to personalize learning paths.

problem Personalized learning paths based on learners' evolving knowledge levels and item structures.
method CSEAL integrates knowledge levels and item structures using a Markov Decision Process and actor-critic algorithm.
result CSEAL effectively personalizes learning paths, improving learning outcomes.

Framework learns item representations from text data for complementary and similar items.

problem Generating accurate complementary item recommendations from textual data.
method Quadruplet network learning framework for latent space representation of items.
result Items are placed closer together in latent space for similar and complementary items compared to non-complementary items.

WCF uses Wasserstein distance to recommend cold-start items based on content similarity.

problem Recommendation performance drops for new items with little interaction history.
method Applies Wasserstein distance to map interaction history to contents, inferring user preferences.
result WCF outperforms state-of-the-art methods in cold-start recommendation.

Bayesian method improves adaptive testing item selection, ensuring full item exposure.

problem Adaptive testing selects items to estimate ability, but must also ensure diverse item exposure.
method Formulated as Bayesian model averaging, deriving optimal item sampling probabilities.
result Stochastic method achieves full item bank exposure without sacrificing accuracy.

FBSM improves item recommendation for cold-start users by modeling feature interactions.

problem Cold-start item recommendation for new users.
method Factorized bilinear similarity model learning interactions among item features.
result Improves TOP-n recommendation performance compared to traditional methods.

The paper analyzes and optimizes recommendation systems using user-user and item-item collaborative filtering.

problem Optimizing recommendation systems to minimize disliked recommendations.
method Proposes algorithms inspired by user-user and item-item collaborative filtering, proving performance guarantees in terms of expected regret.
result Information-theoretic lower bounds on regret match upper bounds up to logarithmic factors in two model parameter regimes.

Paper presents a fast framework for root cause analysis in large-scale systems.

problem Challenges in reviewing logs for identifying issues in large-scale production environments.
method Automates root cause analysis on structured logs with improved scalability using frequent item-set mining and association rule learning.
result Proposes a framework that selects unique item-sets for target failures, improving interpretability and scalability.

NNMs improve item recommendation by providing interpretable user and item representations.

problem Creating recommender systems that are both accurate and understandable.
method Normalized nonnegative models (NNMs) for item recommendation.
result NNM-based recommender systems provide high predictive power, computational tractability, and expressive user and item representations.

Active learning improves ordering of items with contextual attributes.

problem Learning accurate item orderings from pairwise comparisons, especially when exhaustive comparisons are impractical.
method Proposes an active learning strategy that samples items to minimize expected ordering error, accounting for uncertainty in comparisons.
result Superior sample efficiency and generalization compared to non-contextual ranking approaches and active preference learning baselines.

The study measures similarity in introductory programming items, offering a method and evaluation.

problem Measuring similarity in a diverse pool of programming items for personalized learning.
method General approach to measuring similarity, specific measures for introductory programming, three levels of abstraction evaluation.
result Evaluation of similarity measures using diverse programming environments.

This paper optimizes the number of comparisons needed to find the best k items from pairwise comparisons.

problem Finding the best k items from pairwise comparisons with limited comparisons.
method Developed algorithms for finding probably approximately correct and exact best k items under stochastic conditions.
result Upper and lower bounds on the number of comparisons for finding the best k items, with matching upper bounds for PAC best k items.

Two methods improve 10-K item segmentation using large language models.

problem Challenges in extracting specific items from 10-K reports due to variations in document formats and item presentation.
method Two advanced item segmentation methods: GPT4ItemSeg and BERT4ItemSeg.
result BERT4ItemSeg achieves a macro-F1 of 0.9825, surpassing other methods.

Much of the data being created on the web contains interactions between users and items. Stochastic blockmodels, and other methods for community detection and clustering of bipartite graphs, can infer latent user communities and latent item clusters from this interaction data. These methods, however, typically ignore t…

2015-05-25abs ↗pdf ↗

New model improves recommendation systems by analyzing user-item interactions.

problem Improving recommendation systems for better user-item interactions.
method Sliced Anti-symmetric Decomposition (SAD) model using tensor decomposition.
result SAD produces the most consistent personalized preferences compared to SOTA models.

A new Bayesian model improves forecasting for intermittent demand.

problem Sparse observations, cold-start items, and obsolescence in intermittent demand forecasting.
method Hierarchical Bayesian TSB model with partial pooling and calibrated probabilistic configuration.
result TSB-HB achieves the lowest RMSE and RMSSE on the UCI Online Retail dataset.

Eigenvalue analogy explains item-based recommender system accuracy.

problem Lack of theoretical explanation for item-based recommender system success.
method Formalized as an eigenvalue problem, estimating ratings as true ratings multiplied by user-specific eigenvalues.
result Eigenvalue magnitude correlates with user's recommendation accuracy and can measure confidence.

Algorithm identifies best item from subsets with random utility model feedback.

problem PAC learning the best item from subsets with random utility model feedback.
method Pairwise relative counts and hierarchical elimination for learning algorithm.
result Near-optimal PAC sample complexity guarantee for identifying ε-optimal item.

In this paper, we consider decentralized sequential decision making in distributed online recommender systems, where items are recommended to users based on their search query as well as their specific background including history of bought items, gender and age, all of which comprise the context information of the use…

2013-09-26abs ↗pdf ↗

Etsy uses novel embeddings to improve user recommendations based on item interactions.

problem Improving personalized recommendations for users based on diverse item interactions.
method Learning interaction-based item embeddings to encode co-occurrence patterns of item and interaction types.
result Taking interaction type into account improves user shopping behavior modeling accuracy.

Collaborative filtering is used to recommend items to a user without requiring a knowledge of the item itself and tends to outperform other techniques. However, collaborative filtering suffers from the cold-start problem, which occurs when an item has not yet been rated or a user has not rated any items. Incorporating …

2014-06-09abs ↗pdf ↗

DPPNets approximate DPP sampling with deep learning for efficient subset selection.

problem Efficiently sampling from Determinantal Point Processes (DPPs) with high diversity and quality.
method Developed DPPNets using transformer networks with an inhibitive attention mechanism.
result Samples from DPPNets receive high likelihood under the more expensive DPP alternative, demonstrating efficiency.

Matrix factorization simplifies user-item co-occurrence analysis.

problem Understanding the meaning of low-dimensional matrices in matrix factorization.
method Showed matrix factorization equals calculating eigenvectors of co-occurrence matrices, using RMT insights.
result Low-dimension matrices represent a reduced noise user and item co-occurrence space.

New estimators improve Rasch model item parameter estimation for sparse data.

problem Estimating item parameters in sparse Rasch model data.
method Random pairing maximum likelihood estimator (RP-MLE) and its bootstrapped variant (MRP-MLE).
result RP-MLE and MRP-MLE are minimax optimal and provide precise item parameter estimates.

TransCF improves recommendation by modeling user-item relationships with translation vectors.

problem Triangle inequality violation in matrix factorization-based recommendation methods.
method TransCF uses translation vectors to model latent user-item relationships in implicit feedback.
result TransCF outperforms state-of-the-art methods by up to 17% in hit ratio.

Gradient optimization improves preference elicitation for large item spaces.

problem Computational infeasibility of EVOI for large item spaces in recommender systems.
method Continuous formulation of EVOI as a differentiable network, optimized using gradient methods.
result Gradient-based EVOI optimization achieves state-of-the-art performance and scalability.