Model-based reinforcement learning under concurrent schedules of reinforcement in rodents

被引：22

作者：

Huh, Namjung

Jo, Suhyun

Kim, Hoseok

Sul, Jung Hoon

Jung, Min Whan ^{[1
]}

机构：

[1] Ajou Univ, Sch Med, Neurobiol Lab, Inst Med Sci, Suwon 443721, South Korea

来源：

LEARNING & MEMORY | 2009年 / 16卷 / 05期

关键词：

ANTERIOR CINGULATE CORTEX; MIXED-STRATEGY GAME; DECISION-MAKING; PREFRONTAL CORTEX; DOPAMINE NEURONS; MATCHING LAW; HUMANS; CHOICE; REPRESENTATION; STRIATUM;

D O I：

10.1101/lm.1295509

中图分类号：

Q189 [神经科学];

学科分类号：

071006 ;

摘要：

Reinforcement learning theories postulate that actions are chosen to maximize a long-term sum of positive outcomes based on value functions, which are subjective estimates of future rewards. In simple reinforcement learning algorithms, value functions are updated only by trial-and-error, whereas they are updated according to the decision-maker's knowledge or model of the environment in model-based reinforcement learning algorithms. To investigate how animals update value functions, we trained rats under two different free-choice tasks. The reward probability of the unchosen target remained unchanged in one task, whereas it increased over time since the target was last chosen in the other task. The results show that goal choice probability increased as a function of the number of consecutive alternative choices in the latter, but not the former task, indicating that the animals were aware of time-dependent increases in arming probability and used this information in choosing goals. In addition, the choice behavior in the latter task was better accounted for by a model-based reinforcement learning algorithm. Our results show that rats adopt a decision-making process that cannot be accounted for by simple reinforcement learning models even in a relatively simple binary choice task, suggesting that rats can readily improve their decision-making strategy through the knowledge of their environments.

引用

页码：315 / 323

页数：9

共 50 条

[41] Learning to Reweight Imaginary Transitions for Model-Based Reinforcement Learning
Huang, Wenzhen
Yin, Qiyue
Zhang, Junge
Huang, Kaiqi
THIRTY-FIFTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, THIRTY-THIRD CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE AND THE ELEVENTH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2021, 35 : 7848 - 7856
[42] Reward Shaping for Model-Based Bayesian Reinforcement Learning
Kim, Hyeoneun
Lim, Woosang
Lee, Kanghoon
Noh, Yung-Kyun
Kim, Kee-Eung
PROCEEDINGS OF THE TWENTY-NINTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, 2015, : 3548 - 3555
[43] Model-based Adversarial Meta-Reinforcement Learning
Lin, Zichuan
Thomas, Garrett
Yang, Guangwen
Ma, Tengyu
ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 33, NEURIPS 2020, 2020, 33
[44] On the Importance of Hyperparameter Optimization for Model-based Reinforcement Learning
Zhang, Baohe
Rajan, Raghu
Pineda, Luis
Lambert, Nathan
Biedenkapp, Andre
Chua, Kurtland
Hutter, Frank
Calandra, Roberto
24TH INTERNATIONAL CONFERENCE ON ARTIFICIAL INTELLIGENCE AND STATISTICS (AISTATS), 2021, 130
[45] Safe Model-based Reinforcement Learning with Stability Guarantees
Berkenkamp, Felix
Turchetta, Matteo
Schoellig, Angela P.
Krause, Andreas
ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 30 (NIPS 2017), 2017, 30
[46] A Brief Survey of Model-Based Reinforcement Learning Techniques
Pal, Constantin-Valentin
Leon, Florin
2020 24TH INTERNATIONAL CONFERENCE ON SYSTEM THEORY, CONTROL AND COMPUTING (ICSTCC), 2020, : 92 - 97
[47] Model-based reinforcement learning for approximate optimal regulation
Kamalapurkar, Rushikesh
Walters, Patrick
Dixon, Warren E.
AUTOMATICA, 2016, 64 : 94 - 104
[48] The Value Equivalence Principle for Model-Based Reinforcement Learning
Grimm, Christopher
Barreto, Andre
Singh, Satinder
Silver, David
ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 33, NEURIPS 2020, 2020, 33
[49] Continuous-Time Model-Based Reinforcement Learning
Yildiz, Cagatay
Heinonen, Markus
Lahdesmaki, Harri
INTERNATIONAL CONFERENCE ON MACHINE LEARNING, VOL 139, 2021, 139
[50] Model-based Bayesian Reinforcement Learning for Dialogue Management
Lison, Pierre
14TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2013), VOLS 1-5, 2013, : 475 - 479

← 1 2 3 4 5 →