Search bioRxivSearch

Biology subjects

Zenon, A.

Publications and source records attributed to Zenon, A..

4 recordsLinked to original sources

Learning and Forgetting Using Reinforced Bayesian Change Detection

Agents living in volatile environments must be able to detect changes in contingencies while refraining to adapt to unexpected events that are caused by noise. In Reinforcement Learning (RL) frameworks, this requires learning rates that adapt to past reliability of the model. The observation that behavioural flexibility in animals tends to decrease following prolonged training in stable environment provides experimental evidence for such adaptive learning rates. However, in classical RL models, learning rate is either fixed or scheduled and can thus not adapt dynamically to environmental changes. Here, we propose a new Bayesian learning model, using variational inference, that achieves adaptive change detection by the use of Stabilized Forgetting, updating its current belief based on a mixture of fixed, initial priors and previous posterior beliefs. The weight given to these two sources is optimized alongside the other parameters, allowing the model to adapt dynamically to changes in environmental volatility and to unexpected observations. This approach is used to implement the \"critic\" of an actor-critic RL model, while the actor samples the resulting value distributions to choose which action to undertake. We show that our model can emulate different adaptation strategies to contingency changes, depending on its prior assumptions of environmental stability, and that model parameters can be fit to real data with high accuracy. The model also exhibits trade-offs between flexibility and computational costs that mirror those observed in real data. Overall, the proposed method provides a general framework to study learning flexibility and decision making in RL contexts.\n\nAuthor summaryIn stable contexts, animals and humans exhibit automatic behaviour that allows them to make fast decisions. However, these automatic processes exhibit a lack of flexibility when environmental contingencies change. In the present paper, we propose a model of behavioural automatization that is based on adaptive forgetting and that emulates these properties. The model builds an estimate of the stability of the environment and uses this estimate to adjust its learning rate and the balance between exploration and exploitation policies. The model performs Bayesian inference on latent variables that represent relevant environmental properties, such as reward functions, optimal policies or environment stability. From there, the model makes decisions in order to maximize long-term rewards, with a noise proportional to environmental uncertainty. This rich model encompasses many aspects of Reinforcement Learning (RL), such as Temporal Difference RL and counterfactual learning, and accounts for the reduced computational cost of automatic behaviour. Using simulations, we show that this model leads to interesting predictions about the efficiency with which subjects adapt to sudden change of contingencies after prolonged training.

neuroscience

Variational Treatment of Trial-by-Trial Drift-Diffusion Models of Behaviour

The Full Drift Diffusion Model (DDM) is challenging to fit to behavioural data. Precision of the fits are usually poor for some if not all parameters, and computationally expensive to obtain. Moreover, inference at the trial level for each and every parameters, threshold included, has so far been considered as impossible. Most approaches rely on strong assumptions about the model structure, such as selection of the parameters that are subject to trial-to-trial variability or prior distribution of these parameters, that usually lack precise mathematical or empirical justifications. The fact that, in most versions of the DDM, the sequence of participants choices are considered as independent and identically distributed (i.i.d.), has been mainly overlooked so far. Our contribution to the field is threefold: first, we introduce Variational Bayes as a method to fit the full DDM. Second, we relax the i.i.d. assumption, and propose a data-driven algorithm based on a Recurrent Auto-Encoder, that estimates the local posterior probability of the DDM parameters at each trial based on the sequence of parameters and data preceding the data. Finally, we show that inference at the trial level can be achieved efficiently for each and every parameter of the DDM, threshold included. This data-driven approach is highly generic and self-contained, in the sense that no external input (e.g. regressors or physiological measure) is necessary to fit the data. Using simulations and real-world examples, we show that this method outperforms by several order of magnitude the i.i.d.-based ones, either Markov Chain Monte Carlo or i.i.d.-VB.

animal behavior and cognition

An information-theoretic perspective on the costs of cognition

In statistics and machine learning, model accuracy is traded off with complexity, which can be viewed as the amount of information extracted from the data. Here, we discuss how cognitive costs can be expressed in terms of similar information costs, i.e. as a function of the amount of information required to update a persons prior knowledge (or internal model) to effectively solve a task. We then examine the theoretical consequences that ensue from this assumption. This framework naturally explains why some tasks - for example, unfamiliar or dual tasks - are costly and permits to quantify these costs using information-theoretic measures. Finally, we discuss brain implementation of this principle and show that subjective cognitive costs can originate either from local or global capacity limitations on information processing or from increased rate of metabolic alterations. These views shed light on the potential adaptive value of cost-avoidance mechanisms.

neuroscience

Objective but not subjective fatigue increases cognitive task avoidance

Mentally demanding tasks feel effortful and are usually avoided. Furthermore, prolonged cognitive engagement leads to mental fatigue, consisting of subjective feeling of exhaustion and decline in performance. Despite the intuitive characterization of fatigue as an increase in subjective effort perception, the effect of fatigue on effort cost has never been tested experimentally. To this end, sixty participants in 2 separate experiments underwent a forced-choice working memory task following either a fatigue-inducing (i.e. Stroop task) or a control manipulation. We measured subjective fatigue and effort as well as their objective behavioral signatures: performance decline and task avoidance, respectively. We found that fatigue-induced performance decline was correlated with task avoidance, while the feelings of fatigue and effort were unrelated to each other. Our findings highlight the discrepancy between subjective and objective manifestations of fatigue and effort, and provide valuable evidence feeding the ongoing theoretical debate on the nature of these constructs.

neuroscience