How can we build agents that keep learning from experience, quickly and efficiently, after their initial training? Here we take inspiration from the main mechanism of learning in biological brains: synaptic plasticity, carefully tuned by evolution to produce efficient lifelong learning. We show that plasticity, just li…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Oja's rule improves neural network training without engineered tricks.
Unified approach for neural networks with multi-compartmental neurons and non-Hebbian plasticity.
A model retains learned knowledge for longer by adding a plastic component to neural networks.
Selective reinitialization improves adaptability of neural bandits in dynamic environments.
Biological neurons learn tensor decompositions of higher-order correlations using nonlinear Hebbian plasticity.
Physics-informed deep learning approximates strain gradient plasticity solutions.
Learning and memory in the brain are implemented by complex, time-varying changes in neural circuitry. The computational rules according to which synaptic weights change over time are the subject of much research, and are not precisely understood. Until recently, limitations in experimental methods have made it challen…
Elite ONNs learn better with synaptic plasticity, improving performance over CNNs.
Introduces generalized almost plastic structures on manifolds.
Dropout helps a stable network learn new tasks without forgetting old ones.
Sleep-based regularization stabilizes STDP in recurrent neural networks.
Model for material elasticity and plasticity using networks.
In the domain of machine learning, Neural Memory Networks (NMNs) have recently achieved impressive results in a variety of application areas including visual question answering, trajectory prediction, object tracking, and language modelling. However, we observe that the attention based knowledge retrieval mechanisms us…
We identify a phenomenon, which we refer to as multi-model forgetting, that occurs when sequentially training multiple deep networks with partially-shared parameters; the performance of previously-trained models degrades as one optimizes a subsequent one, due to the overwriting of shared parameters. To overcome this, w…
Proposes a new neural network approach to credit assignment.
Natural gradient learning improves synaptic plasticity in spiking neurons.
Catastrophic forgetting/interference is a critical problem for lifelong learning machines, which impedes the agents from maintaining their previously learned knowledge while learning new tasks. Neural networks, in particular, suffer plenty from the catastrophic forgetting phenomenon. Recently there has been several eff…
Improved neural ODEs learn adaptable flows.
AANets balance stability and plasticity in CIL.
Flashback Learning balances model stability and plasticity in continual learning.
Neural plasticity is an important functionality of human brain, in which number of neurons and synapses can shrink or expand in response to stimuli throughout the span of life. We model this dynamic learning process as an -norm regularized binary optimization problem, in which each unit of a neural network (e.g., …
Vision transformers benefit from non-smooth components in adaptation.
Metric anomalies arising from a distribution of point defects (intrinsic interstitials, vacancies, point stacking faults), thermal deformation, biological growth, etc. are well known sources of material inhomogeneity and internal stress. By emphasizing the geometric nature of such anomalies we seek their representation…
Synaptic strength can be seen as probability to propagate impulse, and according to synaptic plasticity, function could exist from propagation activity to synaptic strength. If the function satisfies constraints such as continuity and monotonicity, neural network under external stimulus will always go to fixed point, a…
Unified framework for strain-gradient plasticity from dislocations.
Study examines how training regime affects neural networks' forgetting.
We propose a particularly structured Boltzmann machine, which we refer to as a dynamic Boltzmann machine (DyBM), as a stochastic model of a multi-dimensional time-series. The DyBM can have infinitely many layers of units but allows exact and efficient inference and learning when its parameters have a proposed structure…
This paper presents a method to improve continual learning stability and plasticity.
This paper suggests a learning-theoretic perspective on how synaptic plasticity benefits global brain functioning. We introduce a model, the selectron, that (i) arises as the fast time constant limit of leaky integrate-and-fire neurons equipped with spiking timing dependent plasticity (STDP) and (ii) is amenable to the…
Study evaluates CL methods in RNNs, highlighting differences from feedforward networks.
A dynamic Boltzmann machine (DyBM) has been proposed as a model of a spiking neural network, and its learning rule of maximizing the log-likelihood of given time-series has been shown to exhibit key properties of spike-timing dependent plasticity (STDP), which had been postulated and experimentally confirmed in the fie…
Identifies learning rules from neural network observables.
Neuro-inspired recurrent neural network algorithms, such as echo state networks, are computationally lightweight and thereby map well onto untethered devices. The baseline echo state network algorithms are shown to be efficient in solving small-scale spatio-temporal problems. However, they underperform for complex task…
Spiking Neural Networks (SNNs) are brain-inspired, event-driven machine learning algorithms that have been widely recognized in producing ultra-high-energy-efficient hardware. Among existing SNNs, unsupervised SNNs based on synaptic plasticity, especially Spike-Timing-Dependent Plasticity (STDP), are considered to have…
Spiking neural networks (SNNs) could play a key role in unsupervised machine learning applications, by virtue of strengths related to learning from the fine temporal structure of event-based signals. However, some spike-timing-related strengths of SNNs are hindered by the sensitivity of spike-timing-dependent plasticit…
We introduce SIM-CE, an advanced, user-friendly modeling and simulation environment in Simulink for performing multi-scale behavioral analysis of the nervous system of Caenorhabditis elegans (C. elegans). SIM-CE contains an implementation of the mathematical models of C. elegans's neurons and synapses, in Simulink, whi…
New insights into continual learning with task similarity.
Analyze and predict complex 3D shape deformations using LSTM autoencoders and oriented bounding boxes.
VSML unifies meta learning concepts and enables simple backpropagation.
Biological neural network mimics CCA for multi-channel data.
A brain-inspired spiking Transformer reduces energy consumption and enhances interpretability.
Although Deep Neural Networks have seen great success in recent years through various changes in overall architectures and optimization strategies, their fundamental underlying design remains largely unchanged. Computational neuroscience on the other hand provides more biologically realistic models of neural processing…
Neural machine learning methods, such as deep neural networks (DNN), have achieved remarkable success in a number of complex data processing tasks. These methods have arguably had their strongest impact on tasks such as image and audio processing - data processing domains in which humans have long held clear advantages…
This work tackles catastrophic forgetting in neural networks by mimicking brain's metaplasticity.
AI learns to learn sequentially without forgetting.
Olshausen and Field (OF) proposed that neural computations in the primary visual cortex (V1) can be partially modeled by sparse dictionary learning. By minimizing the regularized representation error they derived an online algorithm, which learns Gabor-filter receptive fields from a natural image ensemble in agreement …
Symmetry in loss functions constrains model parameters, leading to specific learning outcomes.