New method extracts hidden phases in binary mixtures using tubular tilings.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We describe a model for capturing the statistical structure of local amplitude and local spatial phase in natural images. The model is based on a recently developed, factorized third-order Boltzmann machine that was shown to be effective at capturing higher-order structure in images by modeling dependencies among squar…
Proposes a method to estimate sparse Gaussian graphical models with hidden clustering structure.
In the last few years there has been a growing interest in Human Activity Recognition~(HAR) topic. Sensor-based HAR approaches, in particular, has been gaining more popularity owing to their privacy preserving nature. Furthermore, due to the widespread accessibility of the internet, a broad range of streaming-based app…
We consider a random sparse graph with bounded average degree, in which a subset of vertices has higher connectivity than the background. In particular, the average degree inside this subset of vertices is larger than outside (but still bounded). Given a realization of such graph, we aim at identifying the hidden subse…
We propose a new algorithm to learn a dictionary for reconstructing and sparsely encoding signals from measurements without phase. Specifically, we consider the task of estimating a two-dimensional image from squared-magnitude measurements of a complex-valued linear transformation of the original image. Several recent …
We propose a novel online learning algorithm for Restricted Boltzmann Machines (RBM), namely, the Online Generative Discriminative Restricted Boltzmann Machine (OGD-RBM), that provides the ability to build and adapt the network architecture of RBM according to the statistics of streaming data. The OGD-RBM is trained in…
One of the longstanding open problems in spectral graph clustering (SGC) is the so-called model order selection problem: automated selection of the correct number of clusters. This is equivalent to the problem of finding the number of connected components or communities in an undirected graph. We propose automated mode…
We present a theoretical analysis of Maximum a Posteriori (MAP) sequence estimation for binary symmetric hidden Markov processes. We reduce the MAP estimation to the energy minimization of an appropriately defined Ising spin model, and focus on the performance of MAP as characterized by its accuracy and the number of s…
New bounds on neural network capacity for treelike sign perceptrons using RDT.
The concept of SCN offers a fast framework with universal approximation guarantee for lifelong learning of non-stationary data streams. Its adaptive scope selection property enables for proper random generation of hidden unit parameters advancing conventional randomized approaches constrained with a fixed scope of rand…
The Denoising Autoencoder (DAE) enhances the flexibility of the data stream method in exploiting unlabeled samples. Nonetheless, the feasibility of DAE for data stream analytic deserves an in-depth study because it characterizes a fixed network capacity that cannot adapt to rapidly changing environments. Deep evolving …
Gradient descent trains both layers of a ReLU network to fit a linear model.
We consider a binary sequence generated by thresholding a hidden continuous sequence. The hidden variables are assumed to have a compound symmetry covariance structure with a single parameter characterizing the common correlation. We study the parameter estimation problem under such one-parameter models. We demonstrate…
Study reconstructs hidden perfect matchings in random graphs with specific edge weights.
This study analyzes VAEs using ID and II, revealing a transition in behaviour and distinct training phases.
Restricted Boltzmann Machines are described by the Gibbs measure of a bipartite spin glass, which in turn corresponds to the one of a generalised Hopfield network. This equivalence allows us to characterise the state of these systems in terms of retrieval capabilities, both at low and high load. We study the paramagnet…
Modeling financial markets as gas molecules, the paper predicts phase transitions similar to water and steam.
We propose convex relaxations for convolutional neural nets with one hidden layer where the output weights are fixed. For convex activation functions such as rectified linear units, the relaxations are convex second order cone programs which can be solved very efficiently. We prove that the relaxation recovers the glob…
SGD learns sparse parities near computational limits with discontinuous phase transitions.
Reinforcement learning aims at searching the best policy model for decision making, and has been shown powerful for sequential recommendations. The training of the policy by reinforcement learning, however, is placed in an environment. In many real-world applications, however, the policy training in the real environmen…
We combine Riemannian geometry with the mean field theory of high dimensional chaos to study the nature of signal propagation in generic, deep neural networks with random weights. Our results reveal an order-to-chaos expressivity phase transition, with networks in the chaotic phase computing nonlinear functions whose g…
Paper proposes a tensor model for clustering noisy multi-view data.
We study the flow of information and the evolution of internal representations during deep neural network (DNN) training, aiming to demystify the compression aspect of the information bottleneck theory. The theory suggests that DNN training comprises a rapid fitting phase followed by a slower compression phase, in whic…
Previous studies have found that an adversary attacker can often infer unintended input information from intermediate-layer features. We study the possibility of preventing such adversarial inference, yet without too much accuracy degradation. We propose a generic method to revise the neural network to boost the challe…
New method identifies common cause in causal insufficiency, revealing complex phase transitions.
Graphs avoid oversmoothing with properly initialized weights.
RENNs protect input privacy by rotating d-ary features.
The generative learning phase of Autoencoder (AE) and its successor Denosing Autoencoder (DAE) enhances the flexibility of data stream method in exploiting unlabelled samples. Nonetheless, the feasibility of DAE for data stream analytic deserves in-depth study because it characterizes a fixed network capacity which can…
New insights into neural networks reveal unexpected phase transitions.
This paper analyzes the training dynamics of binary neural networks using information bottleneck.
Extracting automatically the complex set of features composing real high-dimensional data is crucial for achieving high performance in machine--learning tasks. Restricted Boltzmann Machines (RBM) are empirically known to be efficient for this purpose, and to be able to generate distributed and graded representations of…
This paper describes a novel energy-based probabilistic distribution that represents complex-valued data and explains how to apply it to direct feature extraction from complex-valued spectra. The proposed model, the complex-valued restricted Boltzmann machine (CRBM), is designed to deal with complex-valued visible unit…
Society's drive toward ever faster socio-technical systems, means that there is an urgent need to understand the threat from 'black swan' extreme events that might emerge. On 6 May 2010, it took just five minutes for a spontaneous mix of human and machine interactions in the global trading cyberspace to generate an unp…
In this paper we study the valuation problem of an insurance company by maximizing the expected discounted future dividend payments in a model with partial information that allows for a changing economic environment. The surplus process is modeled as a Brownian motion with drift. This drift depends on an underlying Mar…
New algorithm IAC recovers hidden communities in labeled SBM with optimal performance.
The dynamics of an ideal fluid or plasma is constrained by topological invariants such as the circulation of (canonical) momentum or, equivalently, the flux of the vorticity or magnetic fields. In the Hamiltonian formalism, topological invariants restrict the orbits to submanifolds of the phase space. While the coadjoi…
We consider the problem of learning a one-hidden-layer neural network with non-overlapping convolutional layer and ReLU activation, i.e., , in which both the convolutional weights and the output weights are paramete…
Optimal data-driven formulations are found for learning and decision-making with historical data.
In the jet-bundle description of first-order classical field theories there are some elements, such as the lagrangian energy and the construction of the hamiltonian formalism, which require the prior choice of a connection. Bearing these facts in mind, we analyze the situation in the jet-bundle description of time-depe…
This paper resolves the all-or-nothing phase transition in graph matching.
Neural networks can approximate positive homogeneous functions, especially with multiple hidden layers.
Graph alignment problem solved with convex relaxations for correlated matrices.
A method uses non-autonomous equations to classify time signals efficiently.
A simple and elegant arrangement of stock components of a portfolio (market index-DJIA) in a recent paper [1], has led to the construction of crossing of stocks diagram. The crossing stocks method revealed hidden remarkable algebraic and geometrical aspects of stock market. The present paper continues to uncover new ma…
Solves Dirichlet problem for Lagrangian phase equation with critical and supercritical phase.
In this paper, we propose a general framework for tensor singular value decomposition (tensor SVD), which focuses on the methodology and theory for extracting the hidden low-rank structure from high-dimensional tensor data. Comprehensive results are developed on both the statistical and computational limits for tensor …
Analyzing deep neural networks (DNNs) via information plane (IP) theory has gained tremendous attention recently as a tool to gain insight into, among others, their generalization ability. However, it is by no means obvious how to estimate mutual information (MI) between each hidden layer and the input/desired output, …