Markov Chain Monte Carlo (MCMC) algorithms are a workhorse of probabilistic modeling and inference, but are difficult to debug, and are prone to silent failure if implemented naively. We outline several strategies for testing the correctness of MCMC algorithms. Specifically, we advocate writing code in a modular way, w…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Synthetic control method improves policy evaluation in high-dimensional settings.
New term ADS describes how machine learning can change user behavior.
The paper addresses selection bias in conformal prediction for focal units.
This study applies old and new generations of panel unit root tests to test the validity of long-run real interest rate parity (RIP) hypothesis for ten Central and Eastern European Countries (CEECs) with respect to the Euro area and an average of the CEECs' real interest rates, respectively. When the panel unit root te…
Deep-MIL models fail to respect key MIL assumption, leading to incorrect learning.
The paper examines A/B tests in recommendation systems to detect biased algorithm comparisons due to shared data.
In this paper we investigate the performance of different types of rectified activation functions in convolutional neural network: standard rectified linear unit (ReLU), leaky rectified linear unit (Leaky ReLU), parametric rectified linear unit (PReLU) and a new randomized leaky rectified linear units (RReLU). We evalu…
Study shows IRM framework can be unstable with small changes, leading to worse generalization.
Sharp inequalities in unit ball with constraints on moments.
In this work, we propose an infinite restricted Boltzmann machine~(RBM), whose maximum likelihood estimation~(MLE) corresponds to a constrained convex optimization. We consider the Frank-Wolfe algorithm to solve the program, which provides a sparse solution that can be interpreted as inserting a hidden unit at each ite…
Regularizing for or against class selectivity in DNNs improves test accuracy.
Binary testing for softmax models requires many samples, similar to leverage score models.
(ABRIDGED) In previous work, two platforms have been developed for testing computer-vision algorithms for robotic planetary exploration (McGuire et al. 2004b,2005; Bartolo et al. 2007). The wearable-computer platform has been tested at geological and astrobiological field sites in Spain (Rivas Vaciamadrid and Riba de S…
We investigate the behavior of the Shanghai Stock Exchange Composite (SSEC) index for the period from 1990:12 to 2007:06 using an unconstrained two-regime threshold autoregressive (TAR) model with an unit root developed by Caner and Hansen. The method allows us to simultaneously consider non-stationarity and nonlineari…
In this paper we presented a novel constructive approach for training deep neural networks using geometric approaches. We show that a topological covering can be used to define a class of distributed linear matrix inequalities, which in turn directly specify the shape and depth of a neural network architecture. The key…
Optimizes balanced treatment assignment for experiments.
Estimates network causal effects considering contagion and latent confounding.
We prove an explicit characterization of the points in Thurston's Master Teapot. This description can be implemented algorithmically to test whether a point in belongs to the complement of the Master Teapot. As an application, we show that the intersection of the Master Teapot with the un…
Study shows text-based news veracity models don't generalize across U.S. and U.K.
A hybrid K-NN and SVM technique improves classification accuracy.
The hypothesis that high dimensional data tend to lie in the vicinity of a low dimensional manifold is the basis of manifold learning. The goal of this paper is to develop an algorithm (with accompanying complexity guarantees) for fitting a manifold to an unknown probability distribution supported in a separable Hilber…
Dropout training improves neural networks' performance.
We present a deep learning system for testing graphics units by detecting novel visual corruptions in videos. Unlike previous work in which manual tagging was required to collect labeled training data, our weak supervision method is fully automatic and needs no human labelling. This is achieved by reproducing driver bu…
New framework improves text watermark detection under imperfect pseudorandomness.
FROCC uses random projections for fast one-class classification.
Over 150,000 new people in the United States are diagnosed with colorectal cancer each year. Nearly a third die from it (American Cancer Society). The only approved noninvasive diagnosis tools currently involve fecal blood count tests (FOBTs) or stool DNA tests. Fecal blood count tests take only five minutes and are av…
Recurrent neural networks with various types of hidden units have been used to solve a diverse range of problems involving sequence data. Two of the most recent proposals, gated recurrent units (GRU) and minimal gated units (MGU), have shown comparable promising results on example public datasets. In this paper, we int…
A deep generative model is developed for representation and analysis of images, based on a hierarchical convolutional dictionary-learning framework. Stochastic {\em unpooling} is employed to link consecutive layers in the model, yielding top-down image generation. A Bayesian support vector machine is linked to the top-…
Bayesian units improve speech recognition with minimal parameters.
This paper proposes an improved design of the perceptron unit to mitigate the vanishing gradient problem. This nuisance appears when training deep multilayer perceptron networks with bounded activation functions. The new neuron design, named auto-rotating perceptron (ARP), has a mechanism to ensure that the node always…
New method tests independence using ROC analysis and bipartite ranking.
Automates phased release strategy to balance risk and speed.
We propose a new family of specification tests called kernel conditional moment (KCM) tests. Our tests are built on a novel representation of conditional moment restrictions in a reproducing kernel Hilbert space (RKHS) called conditional moment embedding (CMME). After transforming the conditional moment restrictions in…
Examines how central bank policies affect stock markets and asset prices.
The concept of SCN offers a fast framework with universal approximation guarantee for lifelong learning of non-stationary data streams. Its adaptive scope selection property enables for proper random generation of hidden unit parameters advancing conventional randomized approaches constrained with a fixed scope of rand…
Our work focuses on the problem of predicting the transfer of pediatric patients from the general ward of a hospital to the pediatric intensive care unit. Using data collected over 5.5 years from the electronic health records of two medical facilities, we develop classifiers based on adaptive boosting and gradient tree…
SliceOut speeds up deep learning training without sacrificing accuracy.
Training data-driven approaches for complex industrial system health monitoring is challenging. When data on faulty conditions are rare or not available, the training has to be performed in a unsupervised manner. In addition, when the observation period, used for training, is kept short, to be able to monitor the syste…
FF algorithm uses goodness as a likelihood-ratio test for scalar normalization.
In this paper we study the problems of estimating heterogeneity in causal effects in experimental or observational studies and conducting inference about the magnitude of the differences in treatment effects across subsets of the population. In applications, our method provides a data-driven approach to determine which…
This technical report describes a practical field test on word-image classification in a very large collection of more than 300 diverse handwritten historical manuscripts, with 1.6 million unique labeled images and more than 11 million images used in testing. Results indicate that several deep-learning tests completely…
FF algorithm uses goodness as a measure of input quality, derived from likelihood-ratio tests.
I studied the convergence of regional house prices to national prices in USA by analyzing time-series of house price indices of 9 Census Divisions. I found the evidence of the convergence in some parts of the country using asymmetric unit root tests. The fact that the evidence of the convergence is not present in large…
Many researchers implicitly assume that neural networks learn relations and generalise them to new unseen data. It has been shown recently, however, that the generalisation of feed-forward networks fails for identity relations.The proposed solution for this problem is to create an inductive bias with Differential Recti…
DNPUs improve neural network performance with high-capacity nanoelectronic nodes.
This paper certifies cluster assignments from sum-of-norms clustering algorithms.
Hierarchical causal models help understand cause and effect in nested data.