首页 | 本学科首页   官方微博 | 高级检索  
   检索      


Exploration and exploitation during sequential search
Authors:Dam Gregory  Körding Konrad
Institution:Department of Physiology, Feinberg School of Medicine, Northwestern University;
Rehabilitation Institute of Chicago
Abstract:When we learn how to throw darts we adjust how we throw based on where the darts stick. Much of skill learning is computationally similar in that we learn using feedback obtained after the completion of individual actions. We can formalize such tasks as a search problem; among the set of all possible actions, find the action that leads to the highest reward. In such cases our actions have two objectives: we want to best utilize what we already know (exploitation), but we also want to learn to be more successful in the future (exploration). Here we tested how participants learn movement trajectories where feedback is provided as a monetary reward that depends on the chosen trajectory. We mathematically derived the optimal search policy for our experiment using decision theory. The search behavior of participants is well predicted by an ideal searcher model that optimally combines exploration and exploitation.
Keywords:Human search behavior  Neuroeconomics  Skill acquisition  Decision making  Motor control  Mathematical modeling
本文献已被 PubMed 等数据库收录!
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号