Unified framework for learning with indirect supervision signals.
problem Learning from indirect supervision signals when gold labels are missing or costly.
method Developed a unified theoretical framework for multi-class classification with variable supervision.
result Introduced the concept of separation to characterize learnability and generalization bounds.
Paper tackles RUL prediction with scarce data using indirect supervision.
problem Predicting RUL with indirect supervision and scarce time series data.
method Unified framework called parameterized static regression, handling data scarcity without interpolation.
result Competitive performance in prediction accuracy with simulated data scarcity.
PLRM synthesizes labels from mismatched sources for better training sets.
problem Creating labeled training sets is a major challenge in machine learning.
method PLRM uses probabilistic modeling to synthesize labels from indirect supervision sources with different output spaces.
result PLRM outperforms baselines by 2%-9% on various tasks.
Weakly-supervised learning is a paradigm for alleviating the scarcity of labeled data by leveraging lower-quality but larger-scale supervision signals. While existing work mainly focuses on utilizing a certain type of weak supervision, we present a probabilistic framework, learning from indirect observations, for learn…
In structured prediction problems where we have indirect supervision of the output, maximum marginal likelihood faces two computational obstacles: non-convexity of the objective and intractability of even a single gradient computation. In this paper, we bypass both obstacles for a class of what we call linear indirectl…
Study stability of trading strategy under market perturbations.
problem Dynamic stability of trading strategy under market changes.
method Established reverse conjugacy characterizations, proved continuity and convergence of indirect utility process.
result Continuity and first-order convergence of indirect utility process under market perturbations.
Paper develops models to forecast private equity fund cash flows.
problem Limited literature on illiquid alternative asset cash flow forecasting.
method Develops benchmark model and two novel approaches (direct vs. indirect) using LSTM/GRU models and macroeconomic indicators.
result Direct model performs better and aligns with actual cash flows, but indirect model's performance is less clear.
In this paper, we propose a simple model referred as Contradistinguisher (CTDR) for unsupervised domain adaptation whose objective is to jointly learn to contradistinguish on unlabeled target domain in a fully unsupervised manner along with prior knowledge acquired by supervised learning on an entirely different domain…
Maximizes image representation dependence for self-supervised learning.
problem Learning meaningful image representations from unlabeled data.
method Maximizes Hilbert-Schmidt Independence Criterion (HSIC) between image transformations and identity.
result Matches state-of-the-art performance on ImageNet and other vision tasks.
Machine learning models trained on indirect data labels can fail on real-world examples.
problem Validity issues in machine learning when target labels are indirectly defined.
method Identification of problematic datasets and models using a general procedure.
result Machine learning models trained on indirect data labels will fail on real-world examples.
Direct and indirect RL methods classified and compared.
problem Classifying RL methods for sequential decision making.
method Direct RL solves optimal policy directly, indirect RL solves Bellman equation.
result Direct and indirect RL methods are equivalent and can be unified.
We address the problem of gauging the influence exerted by a given country on the global trade market from the viewpoint of complex networks. In particular, we apply the PWP method for computing indirect influences on the world trade network.
Study shows cooperation can improve everyone's market efficiency.
problem Understanding when collective cooperation improves market efficiency.
method General semimartingale framework, deriving necessary and sufficient conditions.
result Strict improvement in each agent's indirect utility when cooperation is beneficial.
Unified framework for estimating indirect effects in observational studies with unmeasured confounding.
problem Challenges in evaluating indirect effects due to unmeasured confounding and unethical exposures.
method Developed a unified identification and estimation framework using proximal causal inference.
result Unified identification and estimation of PIIE and causal effect of an intervening variable in settings with pervasive unmeasured confounding.
New methods resolve conflicting treatment effect estimates in health tech assessments.
problem Conflicting conclusions from different sponsors analyzing the same data.
method Arbitrated indirect treatment comparisons (ArMAIC) targeting a common target population.
result Estimates treatment effects in a common target population, resolving the MAIC paradox.
This study compares direct and indirect methods for estimating own funds in life insurance, finding indirect methods more effective under realistic asset-liability coupling.
problem Computing own funds for life insurers using direct and indirect methods in a risk-neutral pricing framework.
method Introduced a novel family of mixed estimators including both direct and indirect methods, integrated into a control variate framework for variance reduction.
result The indirect method is more effective under realistic asset-liability coupling, but neither method is universally superior.
Estimates causal effects using machine learning for binary treatment and mediator.
problem Estimating direct and indirect quantile treatment effects under selection-on-observables.
method Double/debiased machine learning estimators based on efficient score functions.
result Uniform consistency and asymptotic normality of effect estimators.
The relationship between international trade and foreign direct investment (FDI) is one of the main features of globalization. In this paper we investigate the effects of FDI on trade from a network perspective, since FDI takes not only direct but also indirect channels from origin to destination countries because of f…
New method for learning indirectly through control variables.
problem Learning relationships when direct manipulation of variables is impossible.
method Study of indirect active learning under nonparametric models with fixed budget.
result Minimax rates for estimating relationships between variables.
Indirect attacks can fool graph classifiers even with poisoned neighbors.
problem How to evaluate and defend graph convolutional neural networks against indirect adversarial attacks.
method Proposed a method to generate adversarial perturbations on a single node far from the target.
result 99% attack success rate within two-hops from the target in two datasets.
Existing methods for CWS usually rely on a large number of labeled sentences to train word segmentation models, which are expensive and time-consuming to annotate. Luckily, the unlabeled data is usually easy to collect and many high-quality Chinese lexicons are off-the-shelf, both of which can provide useful informatio…
Develops exact and invariant study-based decompositions for network meta-analysis.
problem Lack of exact contribution decompositions in network meta-analysis.
method Contrast-space projection formulation of NMA, study-based definition of direct and indirect evidence.
result Exact covariance-aware decompositions of NMA estimator into direct and indirect contributions.
Financial markets are exposed to systemic risk, the risk that a substantial fraction of the system ceases to function and collapses. Systemic risk can propagate through different mechanisms and channels of contagion. One important form of financial contagion arises from indirect interconnections between financial insti…
Deep learning based task systems normally rely on a large amount of manually labeled training data, which is expensive to obtain and subject to operator variations. Moreover, it does not always hold that the manually labeled data and the unlabeled data are sitting in the same distribution. In this paper, we alleviate t…
This paper learns prior models from indirect data efficiently.
problem Learning prior models from indirect data in Bayesian inversion.
method Generative model of prior as pushforward of Gaussian in latent space, learned by minimizing loss function.
result Efficient residual-based neural operator approximation for forward model learning.
AI detects 38% NFT trades likely manipulated, improving on indirect methods.
problem Detecting crypto wash trading using indirect methods and leaked data.
method Public NFT data analysis, direct estimation, AI-based estimator.
result AI reduces estimation errors in NFT markets, improving on indirect methods.
The study tackles indirect discrimination in insurance pricing models.
problem Indirect discrimination in insurance pricing models.
method Presented a statistical model free of proxy discrimination.
result The canonical price in the model does not satisfy group fairness axioms.
Noise2Inverse removes artifacts in noisy CT images without needing clean data.
problem Removing artifacts in noisy CT images.
method Noise2Inverse uses a deep CNN trained on multiple statistically independent reconstructions of the same noisy data.
result Noise2Inverse improves peak signal-to-noise ratio and structural similarity index compared to existing methods.
We study the ever more integrated and ever more unbalanced trade relationships between European countries. To better capture the complexity of economic networks, we propose two global measures that assess the trade integration and the trade imbalances of the European countries. These measures are the network (or indire…
Driven by the goal to enable sleep apnea monitoring and machine learning-based detection at home with small mobile devices, we investigate whether interpretation-based indirect knowledge transfer can be used to create classifiers with acceptable performance. Interpretation-based indirect knowledge transfer means that a…
This paper tackles structure learning in indirect observations of Gaussian and non-Gaussian random vectors.
problem Learning the graphical structure of random vectors indirectly observed through a sensing matrix and corrupted noise.
method Parametric and non-parametric approaches for Gaussian and non-Gaussian distributions, respectively.
result Correct graphical structure can be recovered under indefinite sensing systems with insufficient samples.
Paper proposes an algorithm to learn DAGs with indirect dependencies.
problem Learning DAGs misses indirect dependencies in local variables.
method Two-phase algorithm using high-order HSIC for local optimization.
result OT algorithm outperforms existing methods in structure estimation.
New benchmark PVR tests neural network reasoning about indirection.
problem Understanding neural network generalization limits.
method Introducing Pointer Value Retrieval (PVR) benchmark.
result Large variations in performance across different conditions.
This paper studies communication efficiency in federated learning by optimizing the sum-rate-distortion function for indirect multiterminal source coding.
problem Indirect multiterminal source coding in federated learning where edge devices send noisy gradients to the server.
method Analyzes the rate region for the quadratic vector Gaussian CEO problem under unbiased estimator and derives an explicit formula for the sum-rate-distortion function.
result Derives an explicit formula for the sum-rate-distortion function in the special case of identical gradients over edge devices.
New method estimates corporate default probabilities using indirect data.
problem Lack of direct default rate data for corporate companies.
method Modeling default probability dynamics using Bank of Russia overdue debt data.
result Validated method produces trustworthy default probability series.
Study relaxes identification assumptions for natural direct effects in non-randomized settings.
problem Identifying causal direct effects under unmeasured confounding.
method Developed relaxed conditions for identifying natural direct effects in non-randomized settings.
result Identified natural direct effect under unmeasured confounding conditions.
Our goal is to learn a semantic parser that maps natural language utterances into executable programs when only indirect supervision is available: examples are labeled with the correct execution result, but not the program itself. Consequently, we must search the space of programs for those that output the correct resu…
Support vector machine (SVM) is a particularly powerful and flexible supervised learning model that analyzes data for both classification and regression, whose usual algorithm complexity scales polynomially with the dimension of data space and the number of data points. To tackle the big data challenge, a quantum SVM a…
Researchers show how to secretly train models with hidden data, detect usage with high confidence.
problem Protecting training data from traceability in large language models.
method Gradient-based optimization to learn secret sequences absent from training data.
result Secret sequences can be learned by models without performance degradation, detectable with high confidence.
Optimal reinsurance contracts designed for a continuum of risk types.
problem Designing optimal reinsurance contracts with a continuum of risk types.
method Principal-agent model, VaR at risk tolerance level, change of variables, univariate approach.
result Optimal reinsurance contracts are in stop-loss form, classifying agents into high and low risk groups.
Machine learning has become pervasive in multiple domains, impacting a wide variety of applications, such as knowledge discovery and data mining, natural language processing, information retrieval, computer vision, social and health informatics, ubiquitous computing, etc. Two essential problems of machine learning are …
Contradistinguisher learns to distinguish target domain without aligning source and target domains.
problem Difficulty in aligning source and target domains for domain adaptation.
method Direct approach to unsupervised domain adaptation that learns contrastive features and improves classification performance.
result Achieves state-of-the-art performance on Office-31 and VisDA-2017 datasets.
End-to-end algorithm for controlling bilinear systems with probabilistic noise.
problem Controlling bilinear systems with noisy data.
method Proposes an end-to-end algorithm using statistical learning theory and robust controller design.
result Derived finite sample identification error bounds and structurally suitable for control.
This study measures price risk aversion using indirect utility functions in a lab experiment.
problem Measuring risk aversion with uncertain prices in experimental economics.
method Using indirect utility functions and a multiple price list method in a lab experiment.
result Price risk aversion is statistically greater than payoff risk aversion.
It is important to learn various types of classifiers given training data with noisy labels. Noisy labels, in the most popular noise model hitherto, are corrupted from ground-truth labels by an unknown noise transition matrix. Thus, by estimating this matrix, classifiers can escape from overfitting those noisy labels. …
ProAGAN stabilizes GANs for learning SOMs from noisy medical imaging data.
problem Learning stochastic object models from noisy and indirect medical imaging measurements.
method Developed Progressive Growing of AmbientGANs (ProAGAN) to stabilize GANs training.
result Signal detection performance improved using ProAGAN-generated images.
A fundamental problem in geostatistical modeling is to infer the heterogeneous geological field based on limited measurements and some prior spatial statistics. Semantic inpainting, a technique for image processing using deep generative models, has been recently applied for this purpose, demonstrating its effectiveness…
A simple banking network model is proposed which features multiple waves of bank defaults and is analytically solvable in the limiting case of an infinitely large homogeneous network. The model is a collection of nodes representing individual banks; associated with each node is a balance sheet consisting of assets and …