Search bioRxivSearch

bioRxiv · 10.1101/045559

Neural mechanisms for reward-modulated vector learning and navigation: from social insects to embodied agents

Abstract

Despite their small size, insect brains are able to produce robust and efficient navigation in complex environments. Specifically in social insects, such as ants and bees, these navigational capabilities are guided by orientation directing vectors generated by a process called path integration. During this process, they integrate compass and odometric cues to estimate their current location as a vector, called home vector for guiding them back home on a straight path. They further acquire and retrieve path integration-based vector memories anchored globally to the nest or visual landmarks. Although existing computational models reproduced similar behaviors, they largely neglected evidence for possible neural substrates underlying the generated behavior. Therefore, we present here a model of neural mechanisms in a modular closed-loop control - enabling vector navigation in embodied agents. The model consists of a path integration mechanism, reward-modulated global and local vector learning, random search, and action selection. The path integration mechanism integrates compass and odometric cues to compute a vectorial representation of the agents current location as neural activity patterns in circular arrays. A reward-modulated learning rule enables the acquisition of vector memories by associating the local food reward with the path integration state. A motor output is computed based on the combination of vector memories and random exploration. In sim-ulation, we show that the neural mechanisms enable robust homing and localization, even in the presence of external sensory noise. The proposed learning rules lead to goal-directed navigation and route formation performed under realistic conditions. This provides an explanation for, how view-based navigational strategies are guided by path integration. Consequently, we provide a novel approach for vector learning and navigation in a simulated embodied agent linking behavioral observations to their possible underlying neural substrates.\n\nAuthor SummaryDesert ants survive under harsh conditions by foraging for food in temperatures over 60{degrees} C. In this extreme environment, they cannot, like other ants, use pheromones to track their long-distance journeys back to their nests. Instead they apply a computation called path integration, which involves integrating skylight compass and odometric stimuli to estimate its current position. Path integration is not only used to return safely to their nests, but also helps in learning so-called vector memories. Such memories are sufficient to produce goal-directed and landmark-guided navigation in social insects. How can small insect brains generate such complex behaviors? Computational models are often useful for studying behavior and their underlying control mechanisms. Here we present a novel computational framework for the acquisition and expression of vector memories based on path integration. It consists of multiple neural networks and a reward-based learning rule, where vectors are represented by the activity patterns of circular arrays. Our model not only reproduces goal-directed navigation and route formation in a simulated agent, but also offers predictions about neural implementations. Taken together, we believe that it demonstrates the first complete model of vector-guided navigation linking observed behaviors of navigating social insects to their possible underlying neural mechanisms.

Explore related subjects

Keep this discovery

BibTeXRIS

Dennis Goldschmidt, Poramate Manoonpong, Sakyasingha Dasgupta. 2016-03-24. Neural mechanisms for reward-modulated vector learning and navigation: from social insects to embodied agents. https://doi.org/10.1101/045559

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

FUNCTIONAL MRI IN AWAKE DOGS PREDICTS SUITABILITY FOR ASSISTANCE WORK

The overall goal of this work was to measure the efficacy of fMRI for predicting whether a dog would be a successful service dog. The training and imaging were performed in 50 dogs entering advanced training at 17-21 months of age. FMRI responses were measured while each dog observed hand signals indicating either reward or no reward and given by both a familiar handler and a stranger. 49 dogs successfully completed fMRI training and scanning. Of these, 33 eventually completed service training and were matched with a person, while 10 were released for behavioral reasons. Using anatomically defined regions-of-interest in the ventral caudate, amygdala, and visual cortex, we developed a classifier based on the dogs' outcomes. We found that responses in the stranger condition were sufficient to develop an accurate brain-based classifier. On all data, the classifier had a positive predictive value of 96% with 10% false positives. The area under the receiver operating characteristic curve was 0.90 (0.79 with 4-fold cross-validation, P=0.02), indicating a significant diagnostic capability. Within the stranger condition, the differential response to [reward - no reward] in ventral caudate was positively correlated with a successful outcome, while the differential response in the amygdala was negatively correlated to outcome. These results show that successful service dogs transfer knowledge to strangers as indexed by ventral caudate activity without excessive arousal as measured in the amygdala.

Animal Behavior and Cognition

Automatic Head Tracking of The Common Marmoset.

New technologies for manipulating and recording the nervous system allow us to perform unprecedented experiments. However, the influence of our experimental manipulations on psychological processes must be inferred from their effects on behavior. Today, quantifying behavior has become the bottleneck for large-scale, high-throughput, experiments. The method presented here addresses this issue by using deep learning algorithms for video-based animal tracking. Here we describe a reliable automatic method for tracking head position and orientation from simple video recordings of the common marmoset (Callithrix jacchus). This method for measuring marmoset behavior allows for the estimation of gaze within foveal error, and can easily be adapted to a wide variety of similar tasks in biomedical research. In particular, the method has great potential for the simultaneous tracking of multiple marmosets to quantify social behaviors.

Animal Behavior and Cognition

Behavior of Caenorhabditis elegans in a nicotine gradient modulated by food

Nicotine decreases food intake, and smokers often report that they smoke to control their weight. To see whether similar phenomena could be observed in the model organism Caenorhabditis elegans, we challenged drug-naive nematodes with a chronic low (0.01 mM) and high (1 mM) nicotine concentration for 55 h (from hatching to adulthood). After that, we recorded changes in their behavior in a nicotine gradient, where they could choose a desired nicotine concentration. By using a combination of behavioral and morphometric methods, we found that both nicotine and food modulate worm behavior. In the presence of food the nematodes adapted to the low nicotine concentration, when placed in the gradient, chose a similar nicotine concentration like C. elegans adapted to the high nicotine concentration. However, in the absence of food, the nematodes adapted to the low nicotine concentration, when placed in the gradient of this alkaloid, chose a similar nicotine concentration like naive worms. The nematodes growing up in the presence of high concentrations of nicotine had a statistically smaller body size, compared to the control condition, and the presence of food did not cause any enhanced slowing movement. These results provide a platform for more detailed molecular and cellular studies of nicotine addiction and food intake in this model organism.

Animal Behavior and Cognition