Active learning improves GP regression on complex, high-dimensional data.
problem Improving Gaussian Process regression in high-dimensional spaces with discontinuous functions.
method Combines manifold learning with active learning to optimize data selection and reduce dimensionality.
result Superior performance over random learning in synthetic data experiments.
A new method identifies key dimensions for function analysis.
problem Limitations in Active Subspaces for high-dimensional functions.
method Active Manifolds (AM) method for C1(Rm) functions. result AM reduces approximation error by an order of magnitude compared to AS.
In this paper we study a symmetry group of vector space. Basis manifold is a homogeneous space of a symmetry group. This concept leads us to the definition of active and passive transformations on basis manifold. Active transformation can be expressed as a transformation of vector space. Passive transformation gives ab…
New framework tackles high-dimensional reliability analysis using surrogate models and active subspaces.
problem High computational cost and curse of dimensionality in reliability analysis of high-dimensional systems.
method Sparse Active Subspace (SAS) algorithm for identifying low-dimensional manifolds and constructing efficient surrogate models.
result Proposed framework significantly improves accuracy and efficiency of reliability analysis compared to existing methods.
Active subspaces on Riemannian manifolds generalize Euclidean principles.
problem Understanding how scalar-valued quantities change over Riemannian manifolds.
method Generalization of active subspaces from Euclidean to Riemannian spaces using parallel transport.
result The method provides a new way to study scalar-valued quantities on manifolds, differing from extrinsic approaches.
Paper tackles zero-shot activity recognition using video features and text embeddings.
problem Zero-shot activity recognition with videos.
method Auto-encoder model for multimodal joint embedding, 3D convolutional action recognition for visual features, GloVe word embeddings for textual features.
result Improved zero-shot recognition results with top-n accuracy and mean Nearest Neighbor Overlap.
LASER compresses recursive model activations by exploiting their low-dimensional structure.
problem Understanding and optimizing the geometric structure of recursive reasoning trajectories.
method Dynamic low-rank basis tracking via matrix-free subspace tracking with a fidelity-triggered reset mechanism.
result Recursive activations occupy a linear, low-dimensional subspace that can be compressed efficiently.
Deep neural networks can interpolate any dataset in the overparametrized regime.
problem Interpolating any dataset with deep neural networks in the overparametrized regime.
method Proving universal approximations and interpolating any dataset with deep neural networks, considering specific conditions on activation functions.
result Interpolation of any dataset is possible in the overparametrized regime with deep neural networks.
Study shows how correlations between neural activity affect classification capacity.
problem Understanding how correlations between neural activity impact classification performance.
method Calculated the capacity of neural activity on spherical manifolds with and without correlations between centroids and axes.
result Introducing correlations between neural activity centroids pushes spheres closer together, while correlations between axes shrink their radii, revealing a duality between correlations and geometry in classification.
New activation improves deep learning accuracy and robustness.
problem Improving accuracy and robustness of deep neural nets with limited data.
method Replaces softmax with graph Laplacian-based interpolating function.
result Significantly improves natural and robust accuracy.
Method identifies cardiac ectopic activity sites from 12-lead ECG.
problem Locate dangerous ectopic activation sites in the heart.
method Bayesian optimization on cardiac model and ECG data.
result Method converges to minimum after 11.7-3.5 iterations.
New insights into data geometry reveal manifold structure in grid-cell activity.
problem Understanding the roles of different dimensions in data geometry.
method Generalised Hanson-Wright inequality and random function model analysis.
result Persistence diagrams reveal latent homology and manifold structure.
SAEs struggle with curved activation manifolds, revealing layer-dependent scaling laws.
problem Sparse autoencoders' reconstruction error varies across layers, not fitting existing scaling laws.
method Cross-layer study of 844 SAE checkpoints, fitting and regressing on manifold geometry.
result Manifold geometry predicts layer-dependent width exponents in SAEs, with transferable coefficients.
A new method for reducing model complexity using neural active manifolds.
problem Uncertainty quantification in computationally expensive models.
method Autoencoders and surrogate models to discover a neural active manifold.
result Neural active manifolds reduce model variance in multifidelity sampling.
GPMI method interpolates uncertain atrial conduction velocity on non-Euclidean manifolds.
problem Uncertainty in atrial conduction velocity calculations.
method Gaussian Process Manifold Interpolation (GPMI) on human atrial manifolds.
result GPMI accounts for atrial topology and calculates CV uncertainty.
REMAL: Residual Equilibrium Manifold Active Learning for Surrogate-Based Multidisciplinary Design Analysis
problem Multidisciplinary design analysis of coupled engineering systems requires solving equilibrium states where all disciplinary coupling variables are consistent.
method Residual manifold surrogate modeling framework for coupled systems.
result REMAL learns a surrogate model of the joint residual manifold via multitask Gaussian process models.
Active learning selects high-quality examples for text-to-SQL systems.
problem Efficiently annotate large language models for text-to-SQL systems.
method Formalizes example selection as a constrained experimental design problem over semantic query embeddings, proposing a stratified greedy algorithm that maximizes heteroscedastic mutual information.
result Proposed method significantly reduces labeling effort while maintaining high text-to-SQL retrieval accuracy.
High frequency oscillations (HFOs) are a promising biomarker of epileptic brain tissue and activity. HFOs additionally serve as a prototypical example of challenges in the analysis of discrete events in high-temporal resolution, intracranial EEG data. Two primary challenges are 1) dimensionality reduction, and 2) asses…
Enhances Gaussian process regression with multi-fidelity models and active subspaces for high-dimensional problems.
problem Data scarcity and high-dimensional input spaces with low intrinsic dimensionality.
method Employ Gaussian processes in a Bayesian setting, augmenting with low-fidelity models, and exploiting active subspaces.
result Improves model accuracy through multi-fidelity Gaussian process regression with active subspaces.
Generative models can approximate any data manifold well under certain conditions.
problem The universality of generative models and their ability to approximate any data manifold.
method Theoretical analysis of neural networks and activation functions, proving the ability to map latent spaces to data manifolds within specified distances.
result Neural networks can map latent spaces onto data manifolds within specified Hausdorff distances under mild assumptions on the activation function.
Improved neural population modeling using shared features and ensemble detection.
problem Missing shared coding properties in neural latent variable models.
method Feature sharing across tuning curves and soft clustering of neurons.
result More interpretable and better-performing neural population models.
We replace the output layer of deep neural nets, typically the softmax function, by a novel interpolating function. And we propose end-to-end training and testing algorithms for this new architecture. Compared to classical neural nets with softmax function as output activation, the surrogate with interpolating function…
Proposes a stratified sampling method for high-dimensional models using neural active manifolds.
problem Uncertainty propagation in computationally expensive models with many inputs.
method Neural active manifolds for nonlinear dimensionality reduction, followed by stratification in the reduced space.
result Effective variance reduction in high-dimensional models using stratified sampling.
Stochastic subgradient descent avoids critical points in definable functions.
problem Finding local minima in definable functions.
method Stochastic subgradient descent with density-like perturbation.
result SGD converges to a local minimum in definable functions.
Scientists and engineers rely on accurate mathematical models to quantify the objects of their studies, which are often high-dimensional. Unfortunately, high-dimensional models are inherently difficult, i.e. when observations are sparse or expensive to determine. One way to address this problem is to approximate the or…
Latent FxLMS accelerates ANC by adapting along low-dimensional filter weights.
problem Improving active noise control with neural adaptive filters.
method Training an auto-encoder on filter coefficients, constraining weights to latent variables, and updating in latent space.
result Latent FxLMS converges in fewer steps with comparable error to standard FxLMS.
This paper proposes a geometry-aware active learning framework for spatiotemporal dynamic systems.
problem Challenges in modeling complex dynamic systems with 3D geometries and time evolution.
method Geometry-aware spatiotemporal Gaussian Process (G-ST-GP) and adaptive active learning strategy.
result The proposed framework outperforms traditional methods in predicting high-dimensional dynamic behaviors.
New neural network with RePU activation approximates smooth functions and their derivatives.
problem Approximating smooth functions and their derivatives with neural networks.
method Differentiable neural networks with RePU activation functions.
result Improved approximation error bounds for RePU-activated neural networks.
Theory of MoE Transformers' generalization and scaling.
problem Understanding the generalization and scaling of Mixture-of-Experts (MoE) Transformers.
method Developed a theory that separates active capacity from routing combinatorics, derived a sup-norm covering-number bound, and proved a constructive approximation theorem.
result Generalization and scaling laws for MoE Transformers, showing how active capacity and routing structure affect performance.
New flatness measure for deep networks invariant to scaling.
problem Lack of invariant flatness measures for deep networks under parameter rescaling.
method Introduced a quotient manifold structure and Hessian-based invariant flatness measure.
result Confirms that Large-Batch SGD minima are sharper than Small-Batch SGD minima.
In this paper, we consider the problem of fast and efficient indexing techniques for sequences evolving in non-Euclidean spaces. This problem has several applications in the areas of human activity analysis, where there is a need to perform fast search, and recognition in very high dimensional spaces. The problem is ma…
Many activation functions have been proposed in the past, but selecting an adequate one requires trial and error. We propose a new methodology of designing activation functions within a neural network at each layer. We call this technique an "activation ensemble" because it allows the use of multiple activation functio…
This paper studies activation sparsity in large language models, finding key trends and implications.
problem Activation sparsity in large language models (LLMs) can be improved for efficiency and interpretability.
method Proposes PPL-p% sparsity, analyzes trends with training data, width-depth ratio, and parameter scale. result ReLU is more efficient for sparsity than SiLU, and deeper architectures can improve sparsity.
Study uses HMM for real-time activity recognition from sensor data.
problem Real-time activity recognition from streaming sensor data.
method Online hierarchical hidden Markov model.
result Improved activity recognition accuracy compared to existing methods.
First the title could be also understood as ``3-manifolds related by non-zero degree maps" or "Degrees of maps between 3-manifolds" for some aspects in this survey talk. The topology of surfaces was completely understood at the end of 19th century, but maps between surfaces kept to be an active topic in the 20th centur…
Deep networks can adapt to intrinsic dimensionality beyond domain constraints.
problem Approximating functions on low-dimensional manifolds with high-dimensional data.
method Two-layer compositions with ReLU activation, using dimensionality reducing feature maps.
result Near optimal approximation rates depend on the complexity of the dimensionality reducing map, not the ambient dimension.
Drop-Activation reduces overfitting by randomly setting activations to identity.
problem Overfitting in deep learning models.
method Randomly sets activations to identity during training and uses a deterministic network during testing.
result Improves generalization and performance of neural networks.
Adapts neural network neurons' activation functions for better predictions.
problem Training neural networks with fixed activation functions limits their performance.
method Proposes training over a shape parameter, allowing neurons to adapt their own activation functions.
result Improves prediction accuracy by allowing neurons to tune their activation functions.
In the wake of recent advances in experimental methods in neuroscience, the ability to record in-vivo neuronal activity from awake animals has become feasible. The availability of such rich and detailed physiological measurements calls for the development of advanced data analysis tools, as commonly used techniques do …
BinaryDuo improves BNNs by coupling binary activations, outperforming state-of-the-art models.
problem Gradient mismatch in BNNs due to binarizing activations.
method Using gradient of smoothed loss function to estimate gradient mismatch, proposing BinaryDuo scheme with coupled ternary activations.
result BinaryDuo outperforms state-of-the-art BNNs on various benchmarks.
Active learning reduces spin network inference complexity by 10^6-fold.
problem Difficulty in inferring direct interactions in complex networks.
method Information geometry framework to quantify inference difficulty and information gain from perturbations.
result Designed perturbations reduce sampling complexity by 10^6-fold across various network architectures.
Evolutionary algorithms improve neural network performance by discovering better activation functions.
problem The choice of activation function affects neural network performance, but ReLU remains dominant.
method Defined a tree-based search space of candidate activation functions and used evolutionary algorithms (mutation, crossover, exhaustive search) to explore and discover better functions.
result Replacing ReLU with evolved activation functions statistically significantly increases network accuracy.
Active learning method balances bias and variance under class imbalance.
problem Active learning under label shift when class proportions differ.
method Mediated Active Learning under Label Shift (MALLS) using a 'medial distribution'.
result MALLS reduces asymptotic sample complexity under arbitrary label shift.
New method learns neural network activation functions from data.
problem Learning activation functions for neural networks.
method Model each neuron's activation function as a small neural network.
result Learned activation functions improve network performance.
Study active learning of PTFs with derivative access.
problem Active learning of polynomial threshold functions (PTFs).
method Algorithm for active learning degree-d univariate PTFs with derivative access. result Computational efficient algorithm for active learning degree-d univariate PTFs. Collaborative filtering is a useful technique for exploiting the preference patterns of a group of users to predict the utility of items for the active user. In general, the performance of collaborative filtering depends on the number of rated examples given by the active user. The more the number of rated examples giv…
Paper proposes algorithms for active learning of reject option classifiers.
problem Active learning of reject option classifiers is unaddressed in machine learning.
method Developed novel algorithms using double ramp and double sigmoid loss functions.
result Proposed algorithms efficiently reduce the number of labeled examples required.
Bayesian adaptive designs can be biased by active learning, especially with misspecified models.
problem Active learning bias in Bayesian adaptive experimental designs.
method Analysis of linear and preference learning models, empirical testing.
result Model misspecification and noise influence active learning bias in Bayesian designs.