聖塔非研究所

摘要 Understanding how individuals learn in an unknown

2018-02-14 · 已發表論文 · 更新 2026/08/30 下午12:48

摘要 Understanding how individuals learn in an unknown environment is an important problem in 經濟s. We model and examine experimentally behavior in a very simple multi armed bandit framework in…

本頁只刊出中文翻譯與中文說明;英文原文請見下方原文連結。

原文連結

論文資訊

  • 類型:已發表論文
  • 日期:2018-02-14

摘要

Understanding how individuals learn in an unknown environment is an important problem in 經濟s. We model and examine experimentally behavior in a very simple multi armed bandit framework in which participants do not know the inter-temporal payoff structure. We propose a baseline reinforcement learning model that allows for pattern recognition and change in the strategy space. We also analyse three augmented versions that accommodate observational learning from the actions and/or payoffs of another player. The models successfully reproduce the distributional properties of observed discovery times and total payoffs. Our study further shows that when one of the pair discovers the hidden pattern, observing another's actions and/or payoffs improves discovery time compared to the baseline case.

※ 此為已發表論文,全文需透過期刊付費取得