bioRxiv · 10.1101/2021.09.30.462676
Experience resetting in reinforcement learning facilitates exploration-exploitation transitions during a behavioral task for primates
Abstract
The exploration-exploitation trade-off is a fundamental problem in re-inforcement learning. To study the neural mechanisms involved in this problem, a target search task in which exploration and exploitation phases appear alternately is useful. Monkeys well trained in this task clearly understand that they have entered the exploratory phase and quickly acquire new experiences by resetting their previous experiences. In this study, we used a simple model to show that experience resetting in the exploratory phase improves performance rather than decreasing the greediness of action selection, and we then present a neural network-type model enabling experience resetting.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sakamoto, K., Okuzaki, H., Sato, A., Mushiake, H.. 2021-10-02. Experience resetting in reinforcement learning facilitates exploration-exploitation transitions during a behavioral task for primates. https://doi.org/10.1101/2021.09.30.462676
Cite the original work for its findings. Save a collection to share your selection of sources.