A model of operant learning based on chaotically varying synaptic strength.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 30176514.
- Also identified by DOI 10.1016/j.neunet.2018.08.006.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Operant learning is learning based on reinforcement of behaviours. We propose a new hypothesis for operant learning at the single neuron level based on spontaneous fluctuations of synaptic strength caused by receptor dynamics. These fluctuations allow the neural system to explore a space of outputs. If the receptor dynamics are altered by a reinforcement signal the neural system settles to better states, i.e., to match the environmental dynamics that determine reward. Simulations show that this mechanism can support operant learning in a feed-forward neural circuit, a recurrent neural circuit, and a spiking neural circuit controlling an agent learning in a dynamic reward and punishment situation. We discuss how the new principle relates to existing learning rules and observed phenomena of short and long-term potentiation.
Medical subject headings
- Conditioning, Operant
- Models, Neurological
- Nerve Net
- Nonlinear Dynamics
- Synapses