Complementary products recommendation is an important problem in e-commerce. Such recommendations increase the average order price and the number of products in baskets. Complementary products are typically inferred from basket data. In this study, we propose the BB2vec model. The BB2vec model learns vector representat…
Online shopping caters to the needs of millions of users daily. Search, recommendations, personalization have become essential building blocks for serving customer needs. Efficacy of such systems is dependent on a thorough understanding of products and their representation. Multiple information sources and data types p…
We consider a context-based dynamic pricing problem of online products, which have low sales. Sales data from Alibaba, a major global online retailer, illustrate the prevalence of low-sale products. For these products, existing single-product dynamic pricing algorithms do not work well due to insufficient data samples.…
Linear classifiers in product space forms improve scRNA-seq data classification.
problem Linear classification in products of Euclidean, spherical, and hyperbolic spaces.
method Novel formulations of linear classifiers on Riemannian manifolds, proving expressive power, and formalizing perceptron and SVM classifiers.
result Linear classifiers in product space forms have the same expressive power as in Euclidean space of the same dimension.
The easy access to large data sets has allowed for leveraging methodology in network physics and complexity science to disentangle patterns and processes directly from the data, leading to key insights in the behavior of systems. Here we use to country specific food production data to study binary and weighted topologi…
Improved 3D LiDAR data classification using product coefficients.
problem Enhancing accuracy in 3D LiDAR data classification.
method Introducing product coefficients derived from measure theory as additional features in the classification process, alongside PCA.
result Significant improvement in classification accuracy with product coefficients.
The growing conflicts in and about oil exporting regions and speculations about volatile oil prices during the last decade have renewed the public interest in predictions for the near future oil production and consumption. Unfortunately, studies from only 10 years ago, which tried to forecast the oil production during …
Labor productivity was studied at the microscopic level in terms of distributions based on individual firm financial data from Japan and the US. A power-law distribution in terms of firms and sector productivity was found in both countries' data. The labor productivities were not equal for nation and sectors, in contra…
The paper tackles imbalance in production data by proposing sampling methods to improve model performance on underrepresented observations.
problem Imbalance in production data negatively impacts model predictive performance on underrepresented observations.
method Three sampling approaches are investigated to adjust for imbalance in training data and improve model performance.
result Fitting a model using sampled data yields a small reduction in overall predictive performance but a better performance on underrepresented observations.
Improved sales forecasting for new products using transfer learning.
problem Insufficient training data for new products leads to inaccurate sales forecasts.
method Network-based Transfer Learning approach for deep neural networks.
result Deep neural networks' prediction accuracy for food sales forecasting can be effectively increased.
Study spherical T-duality and Massey products in iterated sphere bundles.
problem Understanding spherical T-duality and Massey products in iterated sphere bundles.
method Analyzing Gysin sequences and Massey products to find T-dual iterated sphere bundles.
result For certain iterated sphere bundles, spherical T-duality can be represented by Massey products.
Heterogeneity of economic agents is emphasized in a new trend of macroeconomics. Accordingly the new emerging discipline requires one to replace the production function, one of key ideas in the conventional economics, by an alternative which can take an explicit account of distribution of firms' production activities. …
Framework for pricing data products in data-poor markets.
problem Challenges in pricing advanced data products due to lack of transaction data.
method Prior-predictive Monte Carlo framework for generating probabilistic price bands.
result Stable probabilistic price bands for data products in data-poor markets.
Proposes using entity embedding vectors to improve Gaussian Process models for knowledge transfer across cell lines.
problem Lack of reuse of experimental data for predicting novel processes.
method Hybrid Gaussian Process models with entity embedding vectors to represent product identity.
result Improved performance in predicting novel processes compared to traditional methods.
New method maps global value chains at product level from trade data.
problem Lack of detailed product-level value chain information in existing datasets.
method Machine learning and trade theory applied to international trade data.
result Approximate product-level value chain information inferred from trade patterns.
Product categorization using text data for eCommerce is a very challenging extreme classification problem with several thousands of classes and several millions of products to classify. Even though multi-class text classification is a well studied problem both in academia and industry, most approaches either deal with …
Adaptive time decay functions improve financial product recommendation accuracy.
problem Inaccurate recommendations due to static historical data in finance.
method Time-dependent collaborative filtering with personalized decay functions.
result Significant improvements over state-of-the-art benchmarks in financial product recommendation.
New method uses product embeddings to predict bundle success.
problem Designing effective product bundles in large retail settings.
method Leverage historical purchases and clickstream data to generate product embeddings, then use heuristics for complementarity and substitutability.
result Embeddings-based heuristics predict bundle success, robust across categories and retailers.
Spherical T-duality for iterated sphere bundles
problem T-duality for iterated sphere bundles
method Repackaging cohomological data into Massey products
result Found T-dual iterated sphere bundles associated to Massey products
This paper proposes a method for estimating consumer preferences among discrete choices, where the consumer chooses at most one product in a category, but selects from multiple categories in parallel. The consumer's utility is additive in the different categories. Her preferences about product attributes as well as her…
Production forecasting is a key step to design the future development of a reservoir. A classical way to generate such forecasts consists in simulating future production for numerical models representative of the reservoir. However, identifying such models can be very challenging as they need to be constrained to all a…
Paper proves translating solutions for a specific flow in a product manifold.
problem Existence of translating solutions for nonparametric mean curvature flow with Neumann boundary data.
method Proves existence using product manifold MnimesR with specific conditions. result Existence of translating solutions for the flow in the product manifold.
This study analyses, through cross-section estimation methods, the influence of spatial effects in the conditional product convergence in the parishes' economies of mainland Portugal between 1991 and 2001 (the last year with data available for this spatial disaggregation level). To analyse the data, Moran's I statistic…
Harmonic maps from hyperbolic planes to hyperbolic space exist with given boundary data.
problem Existence of harmonic maps from product of hyperbolic planes to hyperbolic space.
method Existence result for asymptotic Dirichlet problem.
result Existence of harmonic maps with given boundary data.
The constant growth of the e-commerce industry has rendered the problem of product retrieval particularly important. As more enterprises move their activities on the Web, the volume and the diversity of the product-related information increase quickly. These factors make it difficult for the users to identify and compa…
Technological improvement is the most important cause of long-term economic growth. We study the effects of technology improvement in the setting of a production network, in which each producer buys input goods and converts them to other goods, selling the product to households or other producers. We show how this netw…
ProductNet is a collection of high-quality product datasets for better product understanding. Motivated by ImageNet, ProductNet aims at supporting product representation learning by curating product datasets of high quality with properly chosen taxonomy. In this paper, the two goals of building high-quality product dat…
New algorithm reduces cold-start costs in multi-armed bandits for many products.
problem High burn-in costs in multi-armed bandits for new products.
method Two-phase bandit algorithm using subsampling and low-rank matrix estimation.
result Reduces burn-in costs and expedites experiment in large product sets.
Paper proposes a new method to optimize feature coordinates for better image classification.
problem Improving feature extraction for better machine learning classification.
method Mutual-energy inner product optimization method.
result The method enhances low-frequency features and suppresses high-frequency noise, leading to better classification results.
New product structures encode superintegrable Hamiltonian systems in Euclidean spaces.
problem Encoding superintegrable Hamiltonian systems using product structures.
method Introducing commutative and associative product structures on Euclidean spaces of dimension at least three, satisfying specific conditions.
result All abundant superintegrable Hamiltonian systems on Euclidean space of dimension at least three arise from these product structures.
Method identifies potential customers from limited data.
problem Efficiently market products to interested but non-loyal customers.
method Double Positive and Unlabeled (PU) learning approach.
result Proposed algorithm achieves efficient marketing.
The paper proposes a new method for product recommendation that considers revenue contributions and user similarity.
problem High dimensionality and sparsity in user-item data, especially in terms of revenue contributions.
method The approach encodes revenue contributions in the user-item matrix and computes customer similarity using suitable distance measures.
result The method segments users based on revenue-based similarity and supports recommendations aligned with profitability objectives.
In this paper, we describe a solution to tackle a common set of challenges in e-commerce, which arise from the fact that new products are continually being added to the catalogue. The challenges involve properly personalising the customer experience, forecasting demand and planning the product range. We argue that the …
Unified theory for neural scaling laws in hierarchically compositional data.
problem Understanding neural scaling laws in hierarchically compositional data.
method Probabilistic context-free grammars and power-law distributed production rules.
result Unified learning curve behavior for classification and next-token prediction tasks.
This work characterizes topological descriptors of graph products and their expressive power.
problem Capturing multiscale structural information in graph products using topological descriptors.
method Analysis of various filtrations on graph products, including Euler characteristic and persistent homology.
result Persistent homology of graph products contains more information than individual graphs.
Bayesian networks improve product risk assessment by handling uncertainty and causality.
problem Limited handling of uncertainty and inability to incorporate causal explanations in existing methods.
method Bayesian Networks (BNs) for improved systematic product risk assessment.
result BN approach provides more powerful and flexible risk assessments.
The paper optimizes exceptions in a statistical production system using machine learning.
problem Lack of curated and labeled training data for machine learning in data quality assurance.
method Explainable supervised machine learning to identify and prioritize exceptions.
result Improvement in the quality and efficiency of exceptions generated and authenticated by users.
Proves existence and uniqueness of CMC solutions in product manifolds.
problem Existence and uniqueness of solutions to CMC equation with Neumann boundary data.
method Analyzes product manifold MnimesR with specific curvature conditions. result Proves existence and uniqueness of solutions.
Study uses TDA to improve OEE forecasting in manufacturing.
problem Volatility and nonlinearity in OEE time series data.
method Topological Data Analysis (TDA) to transform raw data into structured knowledge.
result Improves OEE forecasting accuracy by at least 17%.
Focuses on monitoring and explaining models in real-world applications.
problem Ensuring high quality machine learning services in production environments.
method Statistical techniques for model performance and data monitoring, explanations of predictions.
result Challenges and solutions for implementing monitoring and explanation in production models.
Estimates modes and ridges in mixed Euclidean and directional spaces.
problem Estimating local modes and density ridges in product spaces combining Euclidean and directional metrics.
method Extends mean shift algorithm to product spaces, addressing challenges in generalization.
result Established convergence of the proposed methods and demonstrated effectiveness on real-world datasets.
New method calibrates Gaussian product experts for better predictions.
problem Erratic predictions and uncalibrated uncertainty in Gaussian product experts.
method Calibration via tempered softmax and Wasserstein barycenter for predictions.
result Improved predictions with better mean and uncertainty quantification.
Derives equations for capital deepening in a competitive economy without assuming a production function.
problem Understanding capital deepening and firm survival in a competitive economy.
method Derives equations of motion from accounting identities, without assuming a production function. Uses four coupled relaxation equations to govern capital productivity, labor share, and new investment productivity.
result A 1% improvement in new-capital productivity nearly doubles the aggregate growth rate within one capital lifetime.
Product diversity of large US firms has declined steadily since 1997.
problem Lack of data on global product diversity makes investigation difficult.
method Text mining of US firms' product descriptions from 1997-2017.
result Product diversity of large US firms has been declining since 1997.
AI agent predicts industry and product/service codes for companies.
problem Manual curation of company data is expensive and prone to errors.
method Hierarchical multi-class industry code classifier with multi-label product/service code classifier.
result High accuracy (92-96%) achieved with limited labeled data.
Recently developed machine learning techniques, in association with the Internet of Things (IoT) allow for the implementation of a method of increasing oil production from heavy-oil wells. Steam flood injection, a widely used enhanced oil recovery technique, uses thermal and gravitational potential to mobilize and dilu…
Improved product recommendations using deep learning.
problem Sparse customer purchasing data for personalized recommendations.
method Deep Collaborative Filtering (NCF) with latent variables and Bayesian Optimization.
result NCF achieved highest NDCG performance on proprietary dataset.
Homotopy on nanophrases is an equivalence relation defined using some data called a homotopy data triple. We define a product on homotopy data triples. We show that any homotopy data triple can be factorized into a product of prime homotopy data triples and this factorization is unique up to isomorphism and order. If a…