bioRxiv · 10.1101/2023.12.03.569774
Adaptive algorithms for shaping behavior
Abstract
Dogs and laboratory mice are commonly trained to perform complex tasks by guiding them through a curriculum of simpler tasks ( shaping). What are the principles behind effective shaping strategies? Here, we propose a machine learning framework for shaping animal behavior, where an autonomous teacher agent decides its students task based on the students transcript of successes and failures on previously assigned tasks. Using autonomous teachers that plan a curriculum in a common sequence learning task, we show that near-optimal shaping algorithms adaptively alternate between simpler and harder tasks to carefully balance reinforcement and extinction. Based on this intuition, we derive an adaptive shaping heuristic with minimal parameters, which we show is near-optimal on the sequence learning task and robustly trains deep reinforcement learning agents on navigation tasks that involve sparse, delayed rewards. Extensions to continuous curricula are explored. Our work provides a starting point towards a general computational framework for shaping animal behavior.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Tong, W. L., Iyer, A., Murthy, V. N., Reddy, G.. 2023-12-05. Adaptive algorithms for shaping behavior. https://doi.org/10.1101/2023.12.03.569774
Cite the original work for its findings. Save a collection to share your selection of sources.