A biologically plausible embodied model of action discovery. 2013

Rufino Bolado-Gomez, and Kevin Gurney
Department of Psychology, Adaptive Behaviour Research Group, University of Sheffield Sheffield, UK.

During development, animals can spontaneously discover action-outcome pairings enabling subsequent achievement of their goals. We present a biologically plausible embodied model addressing key aspects of this process. The biomimetic model core comprises the basal ganglia and its loops through cortex and thalamus. We incorporate reinforcement learning (RL) with phasic dopamine supplying a sensory prediction error, signalling "surprising" outcomes. Phasic dopamine is used in a cortico-striatal learning rule which is consistent with recent data. We also hypothesized that objects associated with surprising outcomes acquire "novelty salience" contingent on the predicability of the outcome. To test this idea we used a simple model of prediction governing the dynamics of novelty salience and phasic dopamine. The task of the virtual robotic agent mimicked an in vivo counterpart (Gancarz et al., 2011) and involved interaction with a target object which caused a light flash, or a control object which did not. Learning took place according to two schedules. In one, the phasic outcome was delivered after interaction with the target in an unpredictable way which emulated the in vivo protocol. Without novelty salience, the model was unable to account for the experimental data. In the other schedule, the phasic outcome was reliably delivered and the agent showed a rapid increase in the number of interactions with the target which then decreased over subsequent sessions. We argue this is precisely the kind of change in behavior required to repeatedly present representations of context, action and outcome, to neural networks responsible for learning action-outcome contingency. The model also showed cortico-striatal plasticity consistent with learning a new action in basal ganglia. We conclude that action learning is underpinned by a complex interplay of plasticity and stimulus salience, and that our model contains many of the elements for biological action discovery to take place.

UI MeSH Term Description Entries

Related Publications

Rufino Bolado-Gomez, and Kevin Gurney
January 2010, Annual International Conference of the IEEE Engineering in Medicine and Biology Society. IEEE Engineering in Medicine and Biology Society. Annual International Conference,
Rufino Bolado-Gomez, and Kevin Gurney
July 2006, Vision research,
Rufino Bolado-Gomez, and Kevin Gurney
January 2010, Journal of vision,
Rufino Bolado-Gomez, and Kevin Gurney
January 2002, Physical therapy,
Rufino Bolado-Gomez, and Kevin Gurney
January 2009, Journal of neurophysiology,
Rufino Bolado-Gomez, and Kevin Gurney
August 2011, Physical review. E, Statistical, nonlinear, and soft matter physics,
Rufino Bolado-Gomez, and Kevin Gurney
February 2010, Journal of computational neuroscience,
Rufino Bolado-Gomez, and Kevin Gurney
October 2013, Cognitive neurodynamics,
Rufino Bolado-Gomez, and Kevin Gurney
January 1989, Arteriosclerosis (Dallas, Tex.),
Rufino Bolado-Gomez, and Kevin Gurney
January 1997, Bio Systems,
Copied contents to your clipboard!