Back to search

Article

Finding structure in multi-armed bandits

2018-10-03

Abstract excerpt

How do humans search for rewards? This question is commonly studied using multi-armed bandit tasks, which require participants to trade off exploration and exploitation. Standard multi-armed bandits assume that each option has an independent reward distribution. However, learning about options independently is unrealistic, since in the real world options often share an underlying structure. We study a class of st...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
58424937-bf29-5d4c-b952-35788e58ed8e
DOI
10.1101/432534
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Select a neighboring publication to make it the new centre.