Article
Finding structure in multi-armed bandits
2018-10-03
Abstract excerpt
How do humans search for rewards? This question is commonly studied using multi-armed bandit tasks, which require participants to trade off exploration and exploitation. Standard multi-armed bandits assume that each option has an independent reward distribution. However, learning about options independently is unrealistic, since in the real world options often share an underlying structure. We study a class of st...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 58424937-bf29-5d4c-b952-35788e58ed8e
- DOI
- 10.1101/432534
