Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

12.5%25.0%37.5%50.0% · May 199319922001200920182026
48 results for tensor hypernetworks

Task-conditioned hypernetworks help neural networks learn multiple tasks without forgetting.

problem Catastrophic forgetting in neural networks when sequentially trained on multiple tasks.
method Task-conditioned hypernetworks that generate target model weights based on task identity.
result Task-conditioned hypernetworks achieve state-of-the-art performance on CL benchmarks and retain long memories.

Graph hypernetworks improve molecule property prediction and classification.

problem Improving molecule property prediction and classification using graph neural networks.
method Replacing underlying networks with hypernetworks and addressing training instability.
result Demonstrated state-of-the-art performance in various benchmarks.

Two methods improve tensor recovery in Ising models, revealing gene interactions.

problem Improving tensor recovery in Ising models for complex data structures.
method Pseudolikelihood and interaction screening approaches for tensor learning.
result Both methods achieve tensor recovery with sample size logarithmic in nodes, exponential in strength and degree.

Improved hypernetwork for efficient neural network hyperparameter tuning.

problem Efficiently optimizing hyperparameters in neural networks.
method Proposed ΔΔ-STN architecture focusing on accurate best-response Jacobian approximation.
result Significantly improved hyperparameter tuning accuracy and stability.

Graph HyperNetworks (GHN) speed up neural architecture search.

problem Expensive neural architecture search (NAS) requiring training thousands of networks.
method GHN models architecture topology and generates weights via graph neural network.
result GHNs can search nearly 10 times faster than other methods on CIFAR-10 and ImageNet.

A new image representation method using hypernetworks.

problem Representing images in a way that allows for continuous manipulation and analysis.
method Constructing a hypernetwork that maps pixel positions to colors, allowing for continuous image manipulation.
result Comparable image super-resolution results to existing methods using a single model.

This work introduces an integrative approach based on Q-analysis with machine learning. The new approach, called Neural Hypernetwork, has been applied to a case study of pulmonary embolism diagnosis. The objective of the application of neural hyper-network to pulmonary embolism (PE) is to improve diagnose for reducing …

2014-09-19abs ↗pdf ↗

Bayesian hypernetworks learn to transform noise to parameter distributions for neural networks.

problem Approximate Bayesian inference in neural networks with complex parameter correlations.
method Train a Bayesian hypernetwork to transform a simple noise distribution to a complex posterior distribution over neural network parameters using variational inference.
result Bayesian hypernetworks can represent multimodal approximate posteriors with correlations between parameters and enable efficient sampling.

Generative model creates diverse neural network weights efficiently.

problem Creating high-performance and diverse weights for neural networks.
method Trains a hypernetwork mapping latent vectors to high-performance weights, balancing accuracy and diversity.
result Generated weights form a diverse manifold, improving classification accuracy.

HyperST-Net uses hypernetworks to improve spatio-temporal forecasting.

problem Forecasting spatio-temporal data is challenging due to complex spatial and temporal factors.
method Proposes a framework based on hypernetworks with three modules: spatial, temporal, and deduction.
result Models achieve significant improvements over state-of-the-art baselines.

End-to-end framework classifies cognitive workload in real-time driving scenarios.

problem Challenging task of classifying human cognitive states from behavioral and physiological signals.
method End-to-end framework using mixture Hyper Long Short Term Memory Networks (HyperNetworks).
result Framework outperforms previous methods with 83.9% precision and 87.8% recall.

Single training run learns optimal VAE parameters for various β values.

problem Training VAEs with varying β values for optimal trade-off between distortion and rate.
method Introduced Multi-Rate VAE (MR-VAE) using hypernetworks to map β to optimal parameters.
result MR-VAEs can construct the full rate-distortion curve without additional training.

We compress large neural networks for quick adaptation to specific contexts.

problem How to quickly adapt a pretrained large neural network to specific contexts.
method Propose a Bayesian hypernetwork framework to compress the network and encourage sparsity.
result Generated compressed networks are significantly smaller than baseline methods.

Proposes a framework for semi-supervised continual learning from sequentially arriving data.

problem Learning from data with changing task distribution over time, especially in domains with a mix of labeled and unlabeled data.
method Meta-Consolidation for Continual Semi-Supervised Learning (MCSSL) framework with a hypernetwork and semi-supervised auxiliary classifier.
result Significant improvements in continual semi-supervised learning setting.

This paper introduces a new formulation of the Conic Gromov-Wasserstein distance for comparing complex network structures.

problem Comparing measures of unequal mass and complex network structures.
method Novel semi-coupling formulation and extension to hypernetworks.
result Establishes fundamental properties and robustness of CGW metric.

SVH-PSL uses Stein Variational Gradient Descent and Hypernetworks to improve Pareto set learning for expensive MOO.

problem Fragmented surrogate models and pseudo-local optima in expensive multi-objective optimization problems.
method SVH-PSL integrates Stein Variational Gradient Descent (SVGD) with Hypernetworks to address fragmentation and pseudo-local optima.
result SVH-PSL significantly improves the quality of the learned Pareto set, offering a promising solution for expensive MOO.

A new framework enables real-time task trade-off control.

problem Conflict between multiple related tasks in a fixed model capacity.
method Formulates MTL as a preference-conditioned multiobjective optimization problem; uses a hypernetwork-based neural network.
result A single model can handle different trade-off preferences among multiple tasks.

This paper won 1st place in forecasting and investment challenges, improving on meta-learning and parametric models.

problem Forecasting and investment challenges in time-series data.
method Hypernetworks and adversarial portfolios to design time-series models.
result Outperformed state-of-the-art meta-learning methods and conventional parametric models.

Bayes by Hypernet improves neural network uncertainty measures.

problem Overconfidence and lack of meaningful uncertainty measures in neural networks.
method Bayes by Hypernet (BbH) uses implicit distributions and neural networks to model complex distributions.
result Bayes by Hypernet achieves competitive accuracies and predictive uncertainties on MNIST and CIFAR5 tasks.

A scalable framework uses Langevin sampling to approximate neural network models of evolving processes.

problem Uncertainty quantification in neural network models of dynamic systems.
method Flexible data model based on NODE, joint learning of data model and posterior parameters, Langevin sampling.
result Demonstrated performance on chemical reaction and material physics data, compared favorably to variational inference.

CoDA adapts dynamics models to new physical systems by conditioning on context.

problem Generalizing to new physical systems with shared dynamics but different contexts.
method Context-informed dynamics adaptation (CoDA) using multiple environments and a hypernetwork.
result State-of-the-art generalization results on nonlinear dynamics.

DISCO predicts system states from short trajectories using an evolved operator.

problem Predicting next states of dynamical systems governed by unknown PDEs.
method DISCO uses a hypernetwork to generate parameters of a smaller operator network for state prediction.
result DISCO achieves state-of-the-art performance with fewer training epochs and generalizes well.

New MTPP model offers interpretable predictions with state-of-the-art performance.

problem Inexpressive models lack interpretability, while neural models sacrifice interpretability for performance.
method Extends Hawkes process to a hypernetwork with a latent space, making it flexible and interpretable.
result Achieves state-of-the-art performance across various tasks and metrics.

We simplify Volterra process predictions by reducing dimensionality and using a tailored deep learning model.

problem Predicting the conditional law of Volterra processes with stochastic volatility is challenging due to high dimensionality and non-smoothness.
method We developed a stable dimension reduction technique onto a low-dimensional statistical manifold of non-positive curvature and introduced a sequentially deep learning model tailored to this geometry.
result Our model can approximate the conditional law of Volterra processes with approximation rates achievable only with very large networks.

TerraNova models Earth and societies as a unified system.

problem Modeling the physical Earth and human societies as a coupled system.
method TerraNova integrates physical Earth fields and societal indicators in their native geometries using encoders, cross-modal transformers, and a hypernetwork.
result TerraNova represents the physical Earth and societies without lossy averaging over borders, achieving competitive performance and spanning axes not represented by purpose-built encoders.

Novel NAS method balances performance and hardware metrics efficiently.

problem Challenging multi-objective optimization in neural architecture search.
method Parameterizes joint architectural distribution via hypernetwork conditioned on hardware features and preferences.
result Zero-shot transferability to new devices with representative and diverse architectures.

Study evaluates CL methods in RNNs, highlighting differences from feedforward networks.

problem Preventing catastrophic forgetting in RNNs processing sequential data.
method Comprehensive evaluation of CL methods, including elastic weight consolidation and hypernetworks.
result Weight-importance methods perform similarly regardless of sequence length but require more stability for high working memory demands.