Enhances reinforcement learning with partial state information.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Study on Privileged ERM showing limitations and providing capacity analysis.
Privileged Information Dropout improves RL performance without distillation.
Introduces privilege scores to measure and interpret protected attribute-related privilege in machine learning models.
CMRM improves robustness in noisy label settings without requiring privileged knowledge.
Kernel method-based one-class classifier is mainly used for outlier or novelty detection. In this letter, kernel ridge regression (KRR) based one-class classifier (KOC) has been extended for learning using privileged information (LUPI). LUPI-based KOC method is referred to as KOC+. This privileged information is availa…
In this paper we attempt to give a systematic account on privileged coordinates and the nilpotent approximation of Carnot manifolds. By a Carnot manifold it is meant a manifold with a distinguished filtration of subbundles of the tangent bundle which is compatible with the Lie bracket of vector fields. This paper lies …
Study shows using time-series privileged information improves model efficiency.
Proposed SMO algorithm for OC-SVM+ significantly outperforms non-sequential algorithms.
Learning using privileged information is an attractive problem setting that helps many learning scenarios in the real world. A state-of-the-art method of Gaussian process classification (GPC) with privileged information is GPC+, which incorporates privileged information into a noise term of the likelihood. A drawback o…
The learning with privileged information setting has recently attracted a lot of attention within the machine learning community, as it allows the integration of additional knowledge into the training process of a classifier, even when this comes in the form of a data modality that is not available at test time. Here, …
Prior knowledge can be used to improve predictive performance of learning algorithms or reduce the amount of data required for training. The same goal is pursued within the learning using privileged information paradigm which was recently introduced by Vapnik et al. and is aimed at utilizing additional information avai…
We introduce a new unsupervised anomaly detection ensemble called SPI which can harness privileged information - data available only for training examples but not for (future) test examples. Our ideas build on the Learning Using Privileged Information (LUPI) paradigm pioneered by Vapnik et al. [19,17], which we extend …
This paper presents privileged multi-label learning (PrML) to explore and exploit the relationship between labels in multi-label learning problems. We suggest that for each individual label, it cannot only be implicitly connected with other labels via the low-rank constraint over label predictors, but also its performa…
We introduce a learning framework called learning using privileged information (LUPI) to the computer vision field. We focus on the prototypical computer vision problem of teaching computers to recognize objects in images. We want the computers to be able to learn faster at the expense of providing extra information du…
Unsupervised domain adaptation improves with privileged information.
New insights and algorithms improve prediction models with time-series privileged information.
Proposes a new method to selectively access privileged information in reinforcement learning.
Learning using privileged information (LUPI) is a powerful heterogenous feature space machine learning framework that allows a machine learning model to learn from highly informative or privileged features which are available during training only to generate test predictions using input space features which are availab…
This paper explores conformal prediction in the learning under privileged information (LUPI) paradigm. We use the SVM+ realization of LUPI in an inductive conformal predictor, and apply it to the MNIST benchmark dataset and three datasets in drug discovery. The results show that using privileged information produces va…
New research suggests privileged information doesn't improve model performance.
For well over a quarter century, detection systems have been driven by models learned from input features collected from real or simulated environments. An artifact (e.g., network event, potential malware sample, suspicious email) is deemed malicious or non-malicious based on its similarity to the learned model at runt…
Unlike machines, humans learn through rapid, abstract model-building. The role of a teacher is not simply to hammer home right or wrong answers, but rather to provide intuitive comments, comparisons, and explanations to a pupil. This is what the Learning Under Privileged Information (LUPI) paradigm endeavors to model b…
This paper is a sequel of arxiv:1709.09045 and deals with privileged coordinates and nilpotent approximation of Carnot manifolds. By a Carnot manifold it is meant a manifold equipped with a filtration by subbundles of the tangent bundle which is compatible with the Lie bracket of vector fields. In this paper, we single…
Improved meta-learning for dynamics using additional structured knowledge.
Distillation (Hinton et al., 2015) and privileged information (Vapnik & Izmailov, 2015) are two techniques that enable machines to learn from other machines. This paper unifies these two techniques into generalized distillation, a framework to learn from multiple machines and data representations. We provide theoretica…
Many machine learning algorithms assume that all input samples are independently and identically distributed from some common distribution on either the input space X, in the case of unsupervised learning, or the input and output space X x Y in the case of supervised and semi-supervised learning. In the last number of …
We adopt a multi-view approach for analyzing two knowledge transfer settings---learning using privileged information (LUPI) and distillation---in a common framework. Under reasonable assumptions about the complexities of hypothesis spaces, and being optimistic about the expected loss achievable by the student (in disti…
SAGE enhances reinforcement learning by injecting hints to prevent model stagnation.
MPWTSVM improves multi-view learning by reducing redundancy and enhancing accuracy.
Joint training improves model accuracy by selectively using privileged information.
Study null hypersurfaces with privileged vector fields, extending surface gravity and defining new horizon types.
A number of important applied problems in engineering, finance and medicine can be formulated as a problem of anomaly detection. A classical approach to the problem is to describe a normal state using a one-class support vector machine. Then to detect anomalies we quantify a distance from a new observation to the const…
Proposes methods to correct bias and missing data in regression models.
Most existing learning to hash methods assume that there are sufficient data, either labeled or unlabeled, on the domain of interest (i.e., the target domain) for training. However, this assumption cannot be satisfied in some real-world applications. To address this data sparsity issue in hashing, inspired by transfer …
Improves deep neural networks using soft labels through alternating minimization.
DARTS- improves robustness by factoring out skip connections' advantage.
The concept of an objective spatial direction in special relativity is investigated and theories assuming light-speed isotropy while accepting the existence of a privileged spatial direction are classified. A natural generalization of the proper time principle is introduced which makes it possible to devise experimenta…
We construct a privileged system of coordinates with respect to the controlling distribution of a trident snake robot and, furthermore, we construct a nilpotent approximation with respect to the given filtration. Note that all constructions are local in the neighbourhood of a particular point. We compare the motions co…
This paper is about the geometry of flip-graphs associated to triangulations of surfaces. More precisely, we consider a topological surface with a privileged boundary curve and study the spaces of its triangulations with n vertices on the boundary curve. The surfaces we consider topologically fill this boundary curve s…
ADVISOR dynamically balances imitation and reinforcement learning to overcome the imitation gap.
This survey reviews Heterogeneous Representation Learning (HRL) for diverse data types.
IReEn reveals functionality of black-box agents via iterative neural synthesis.
Study optimal portfolios for traders with asymmetric information and delay.
Knowledge transfer speeds up neural classifier training.
Short selling is key to exploiting arbitrage opportunities in financial markets.
The paper characterizes the efficiency of transferring knowledge from a teacher to a student classifier over finite domains.
Completed classification of Einstein spaces with specific metric properties.