We establish linear regret bounds for convex smooth losses using Fenchel-Young losses.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We propose a projection pursuit (PP) algorithm based on Gaussian mixture models (GMMs). The negentropy obtained from a multivariate density estimated by GMMs is adopted as the PP index to be maximised. For a fixed dimension of the projection subspace, the GMM-based density estimation is projected onto that subspace, wh…
OT-ICA uses optimal transport to find independent components, outperforming traditional methods.
MGD combines maximum entropy and diffusion methods for efficient sampling.
New loss functions based on f-divergences improve language model performance.
This paper develops sparse alternatives to continuous distributions, including new types of Gaussians and attention mechanisms.
A new method improves ICA performance by approximating MDI.
Validation is one of the most important aspects of clustering, but most approaches have been batch methods. Recently, interest has grown in providing incremental alternatives. This paper extends the incremental cluster validity index (iCVI) family to include incremental versions of Calinski-Harabasz (iCH), I index and …
iCVI-ARTMAP accelerates clustering with adaptive resonance theory and validity indices.
Introduces Finslerian convolution metrics and their properties.
In this paper, we introduce a Deep Convolutional Analysis Dictionary Model (DeepCAM) by learning convolutional dictionaries instead of unstructured dictionaries as in the case of deep analysis dictionary model introduced in the companion paper. Convolutional dictionaries are more suitable for processing high-dimensiona…
Spectral graph convolutional neural networks (CNNs) require approximation to the convolution to alleviate the computational complexity, resulting in performance loss. This paper proposes the topology adaptive graph convolutional network (TAGCN), a novel graph convolutional network defined in the vertex domain. We provi…
Generative flows are attractive because they admit exact likelihood optimization and efficient image synthesis. Recently, Kingma & Dhariwal (2018) demonstrated with Glow that generative flows are capable of generating high quality images. We generalize the 1 x 1 convolutions proposed in Glow to invertible d x d convolu…
VC dimensions of group CNNs are infinite for certain kernels and groups.
We introduce a guide to help deep learning practitioners understand and manipulate convolutional neural network architectures. The guide clarifies the relationship between various properties (input shape, kernel shape, zero padding, strides and output shape) of convolutional, pooling and transposed convolutional layers…
Convolution Neural Network (CNN) has gained tremendous success in computer vision tasks with its outstanding ability to capture the local latent features. Recently, there has been an increasing interest in extending convolution operations to the non-Euclidean geometry. Although various types of convolution operations h…
We introduce Group equivariant Convolutional Neural Networks (G-CNNs), a natural generalization of convolutional neural networks that reduces sample complexity by exploiting symmetries. G-CNNs use G-convolutions, a new type of layer that enjoys a substantially higher degree of weight sharing than regular convolution la…
In recent times, the use of separable convolutions in deep convolutional neural network architectures has been explored. Several researchers, most notably (Chollet, 2016) and (Ghosh, 2017) have used separable convolutions in their deep architectures and have demonstrated state of the art or close to state of the art pe…
Proves DCNNs with expansive convolution are strongly universally consistent.
New framework for manifold convolutions using toric embeddings.
Functor connects Lie groupoid algebras to bornological structures.
Convolution has been playing a prominent role in various applications in science and engineering for many years. It is the most important operation in convolutional neural networks. There has been a recent growth of interests of research in generalizing convolutions on curved domains such as manifolds and graphs. Howev…
Convolutional Neural Networks, as most artificial neural networks, are commonly viewed as methods different in essence from kernel-based methods. We provide a systematic translation of Convolutional Neural Networks (ConvNets) into their kernel-based counterparts, Convolutional Kernel Networks (CKNs), and demonstrate th…
New method enforces orthogonality in convolutional layers for improved robustness.
New method improves grouped convolutions on edge devices.
GCNs improve regression tasks by aggregating neighbor signals.
Convolution and pooling improve kernel methods in image classification.
Convolutional networks outperform fully-connected ones in certain tasks.
Although group convolutional networks are able to learn powerful representations based on symmetry patterns, they lack explicit means to learn meaningful relationships among them (e.g., relative positions and poses). In this paper, we present attentive group equivariant convolutions, a generalization of the group convo…
This work proposes hyperbolic deep convolutional neural networks for better pattern recognition.
New linear flows using exponential of linear transformations improve generative models.
Convolutional neural networks (CNNs) have achieved breakthrough performances in a wide range of applications including image classification, semantic segmentation, and object detection. Previous research on characterizing the generalization ability of neural networks mostly focuses on fully connected neural networks (F…
We describe convolutional networks using harmonic functions.
Proposes a new convolutional neural network for non-grid data.
New mechanism discovered for feature learning in CNNs.
Coordinate-independent convolutions on manifolds avoid reference frame ambiguity.
Recently, many researchers have been focusing on the definition of neural networks for graphs. The basic component for many of these approaches remains the graph convolution idea proposed almost a decade ago. In this paper, we extend this basic component, following an intuition derived from the well-known convolutional…
Yes, they do. This paper provides the first empirical demonstration that deep convolutional models really need to be both deep and convolutional, even when trained with methods such as distillation that allow small or shallow models of high accuracy to be trained. Although previous research showed that shallow feed-for…
Unified theory for adaptive image convolutions using metric perspectives.
Ensemble learning is a method of combining multiple trained models to improve model accuracy. We propose the usage of such methods, specifically ensemble average, inside Convolutional Neural Network (CNN) architectures by replacing the single convolutional layers with Inner Average Ensembles (IEA) of multiple convoluti…
Simplifies convolutions using tensor networks and einsum for efficient second-order methods.
Introduces new algebraic structures for relational groupoids and proves a reduction theorem.
The paper proposes an ensemble of convolution-based methods for fault detection in gearboxes.
Study convolution of invariant valuations on Lie groups.
Corrected graph convolutions improve node classification on graphs.
Convolutional neural network is an important model in deep learning. To avoid exploding/vanishing gradient problems and to improve the generalizability of a neural network, it is desirable to have a convolution operation that nearly preserves the norm, or to have the singular values of the transformation matrix corresp…
This study reveals a Min-Max property in LeNet's convolutional layers, enhancing adversarial robustness.
Convolutional networks struggle with repeating patterns in ECGs.