Dynamic refinement of behavioural restructure mediates dopamine-dependent credit assignment

Jonathan C.Y. Tang,Vitor Paixao, Filipe Carvalho,Artur Silva,Andreas Klaus, Joaquim Alves da Silva,Rui M. Costa

biorxiv(2023)

引用 0|浏览4
暂无评分
摘要
Animals exhibit a diverse behavioral repertoire when exploring new environments and can learn which actions or action sequences produce positive outcomes. Dopamine release upon encountering reward is critical for reinforcing reward-producing actions[1][1]–[3][2]. However, it has been challenging to understand how credit is assigned to the exact action that produced dopamine release during continuous behavior. We investigated this problem with a novel self-stimulation paradigm in which specific spontaneous movements triggered optogenetic stimulation of dopaminergic neurons. Dopamine self-stimulation rapidly and dynamically changes the structure of the entire behavioral repertoire. Initial stimulations reinforced not only the stimulation-producing target action, but also actions similar to target and actions that occurred a few seconds before stimulation. Repeated pairings led to gradual refinement of the behavioral repertoire to home in on the target. Reinforcement of action sequences revealed further temporal dependencies of refinement. Action pairs spontaneously separated by long time intervals promoted a stepwise credit assignment, with early refinement of actions most proximal to stimulation and subsequent refinement of more distal actions. Thus, a retrospective reinforcement mechanism promotes not only reinforcement, but gradual refinement of the entire behavioral repertoire to assign credit to specific actions and action sequences that lead to dopamine release. ### Competing Interest Statement F.C. is the Director of Open Ephys Production Site. [1]: #ref-1 [2]: #ref-3
更多
查看译文
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要