s-RBFN integrates multiple hypotheses for efficient and diverse prediction.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We propose a method to generate multiple diverse and valid human pose hypotheses in 3D all consistent with the 2D detection of joints in a monocular RGB image. We use a novel generative model uniform (unbiased) in the space of anatomically plausible 3D poses. Our model is compositional (produces a pose by combining par…
There is a significant literature on methods for incorporating knowledge into multiple testing procedures so as to improve their power and precision. Some common forms of prior knowledge include (a) beliefs about which hypotheses are null, modeled by non-uniform prior weights; (b) differing importances of hypotheses, m…
In many practical applications of multiple hypothesis testing using the False Discovery Rate (FDR), the given hypotheses can be naturally partitioned into groups, and one may not only want to control the number of false discoveries (wrongly rejected null hypotheses), but also the number of falsely discovered groups of …
Attention-based encoder decoder network uses a left-to-right beam search algorithm in the inference step. The current beam search expands hypotheses and traverses the expanded hypotheses at the next time step. This traversal is implemented using a for-loop program in general, and it leads to speed down of the recogniti…
Current statistical inference problems in areas like astronomy, genomics, and marketing routinely involve the simultaneous testing of thousands -- even millions -- of null hypotheses. For high-dimensional multivariate distributions, these hypotheses may concern a wide range of parameters, with complex and unknown depen…
DivDis learns diverse hypotheses from underspecified data to improve robustness.
The problem of finding itemsets that are statistically significantly enriched in a class of transactions is complicated by the need to correct for multiple hypothesis testing. Pruning untestable hypotheses was recently proposed as a strategy for this task of significant itemset mining. It was shown to lead to greater s…
rMCL improves on MCL by preserving diversity in predictions for regression problems.
aMCL uses annealing to improve hypothesis diversity in ambiguous tasks.
Private online FDR control for adaptive testing under differential privacy.
The paper confirms two groups of gamma-ray bursts using a new nonparametric metric.
Finding statistically significant interactions between binary variables is computationally and statistically challenging in high-dimensional settings, due to the combinatorial explosion in the number of hypotheses. Terada et al. recently showed how to elegantly address this multiple testing problem by excluding non-tes…
Do two data samples come from different distributions? Recent studies of this fundamental problem focused on embedding probability distributions into sufficiently rich characteristic Reproducing Kernel Hilbert Spaces (RKHSs), to compare distributions by the distance between their embeddings. We show that Regularized Ma…
One important partition of algorithms for controlling the false discovery rate (FDR) in multiple testing is into offline and online algorithms. The first generally achieve significantly higher power of discovery, while the latter allow making decisions sequentially as well as adaptively formulating hypotheses based on …
We prove under suitable hypotheses that convergence of integral varifolds implies convergence of associated mod 2 flat chains and subsequential convergence of associated integer-multiplicity rectifiable currents. The convergence results imply restrictions on the kinds of singularities that can occur in mean curvature f…
In this paper a lower bound for the ADM mass is given in terms of the angular momenta and charges of black holes present in axisymmetric initial data sets for the Einstein-Maxwell equations. This generalizes the mass-angular momentum-charge inequality obtained by Chrusciel and Costa to the case of multiple black holes.…
In this paper, we explore various approaches for semi supervised learning in an end to end automatic speech recognition (ASR) framework. The first step in our approach involves training a seed model on the limited amount of labelled data. Additional unlabelled speech data is employed through a data selection mechanism …
The paper predicts and explains the decay of stock anomaly performance over time.
We learn multiple hypotheses for related tasks under a latent hierarchical relationship between tasks. We exploit the intuition that for domain adaptation, we wish to share classifier structure, but for multitask learning, we wish to share covariance structure. Our hierarchical model is seen to subsume several previous…
Coordination recognition and subtle pattern prediction of future trajectories play a significant role when modeling interactive behaviors of multiple agents. Due to the essential property of uncertainty in the future evolution, deterministic predictors are not sufficiently safe and robust. In order to tackle the task o…
Multiple hypothesis testing is a core problem in statistical inference and arises in almost every scientific field. Given a set of null hypotheses , Benjamini and Hochberg introduced the false discovery rate (FDR), which is the expected proportion of false positives among rejected nu…
We show that, in the Teichmüller metric, "thin-framed triangles are thin"---that is, under suitable hypotheses, the variation of geodesics obeys a hyperbolic-like inequality. This theorem has applications to the study of random walks on Teichmüller space. In particular, an application is worked out for the action of th…
In the online multiple testing problem, p-values corresponding to different null hypotheses are observed one by one, and the decision of whether or not to reject the current hypothesis must be made immediately, after which the next p-value is observed. Alpha-investing algorithms to control the false discovery rate (FDR…
Almost-isometries are quasi-isometries with multiplicative constant one. Lifting a pair of metrics on a compact space gives quasi-isometric metrics on the universal cover. Under some additional hypotheses on the metrics, we show that there is no almost-isometry between the universal covers. We show that Riemannian mani…
We consider the problem of undirected graphical model inference. In many applications, instead of perfectly recovering the unknown graph structure, a more realistic goal is to infer some graph invariants (e.g., the maximum degree, the number of connected subgraphs, the number of isolated nodes). In this paper, we propo…
Max-rank improves multiple testing in conformal prediction.
We consider the problem of asynchronous online testing, aimed at providing control of the false discovery rate (FDR) during a continual stream of data collection and testing, where each test may be a sequential test that can start and stop at arbitrary times. This setting increasingly characterizes real-world applicati…
Proposes handling ambiguity in sequential data predictions.
Biological research often involves testing a growing number of null hypotheses as new data is accumulated over time. We study the problem of online control of the familywise error rate (FWER), that is testing an apriori unbounded sequence of hypotheses (p-values) one by one over time without knowing the future, such th…
Calibration without labels in multiple testing
Paper tackles efficient learning of non-convex hypotheses in metric spaces.
In part \textit{I} we proposed a structure for a general Hypotheses Space , the Learning Space , which can be employed to avoid \textit{overfitting} when estimating in a complex space with relative shortage of examples. Also, we presented the U-curve property, which can be taken ad…
Many real-world vision problems suffer from inherent ambiguities. In clinical applications for example, it might not be clear from a CT scan alone which particular region is cancer tissue. Therefore a group of graders typically produces a set of diverse but plausible segmentations. We consider the task of learning a di…
New method controls false discoveries in financial asset pricing.
e-LOND algorithm controls FDR in online testing with arbitrary dependencies.
A new approach to learning in brain-like networks using adversarial algorithms.
Machine Learning benefits from prior information and computational power for better performance and understanding.
Boosting is a celebrated machine learning approach which is based on the idea of combining weak and moderately inaccurate hypotheses to a strong and accurate one. We study boosting under the assumption that the weak hypotheses belong to a class of bounded capacity. This assumption is inspired by the common convention t…
New method uses LLMs to generate detailed scientific hypotheses.
Hybrid approach combines transformer and Bayesian filtering for robust multiple particle tracking.
Sequential tests for nonparametric hypotheses using supermartingales.
In this paper, we diagnose deep neural networks for 3D point cloud processing to explore utilities of different intermediate-layer network architectures. We propose a number of hypotheses on the effects of specific intermediate-layer network architectures on the representation capacity of DNNs. In order to prove the hy…
Recently, there has been growing interest in multi-speaker speech recognition, where the utterances of multiple speakers are recognized from their mixture. Promising techniques have been proposed for this task, but earlier works have required additional training data such as isolated source signals or senone alignments…
In this paper we discuss a novel framework for multiclass learning, defined by a suitable coding/decoding strategy, namely the simplex coding, that allows to generalize to multiple classes a relaxation approach commonly used in binary classification. In this framework, a relaxation error analysis can be developed avoid…
GANF uses normalizing flows to detect anomalies in multiple time series.
Develops a new theory of localization in algebraic geometry.
LASLA improves multiple testing accuracy with network-structured data.