Published version of an article in the journal: Applied Intelligence. Also available from the publisher at: http://dx.doi.org/10.1007/s10489-012-0346-zThe two-armed bandit problem is a classical optimization problem where a decision maker sequentially pulls one of two arms attached to a gambling machine, with each pull resulting in a random reward. The reward distributions are unknown, and thus, one must balance between exploiting existing knowledge about the arms, and obtaining new information. Bandit problems are particularly fascinating because a large class of real world problems, including routing, Quality of Service (QoS) control, game playing, and resource allocation, can be solved in a decentralized manner when modeled as a system o...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2010 – Universitetet i Agder, GrimstadMulti...
Published version of an article from Lecture Notes in Computer Science. Also available at SpringerLi...
Abstract—We present a formal model of human decision-making in explore-exploit tasks using the conte...
The two-armed bandit problem is a classical optimization problem where a decision maker sequentially...
Published version of an article in the journal: Applied Intelligence. Also available from the publis...
Published version of a chapter from the book: Modern Approaches in Applied Intelligence. Also availa...
The two-armed bandit problem is a classical optimization problem where a decision maker sequentially...
Published version of a chapter from the book: Modern Approaches in Applied Intelligence. Also availa...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2009 – Universitetet i Agder, GrimstadThe t...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2009 – Universitetet i Agder, GrimstadThe t...
The two-armed bandit problem is a classical optimization problem where a player sequentially selects...
Multi-Armed bandit problem is a classic example of the exploration vs. exploitation dilemma in which...
Published version of a chapter in the book: IFIP Advances in Information and Communication Technolog...
Published version of an article from Lecture Notes in Computer Science. Also available at SpringerLi...
The multi-armed bandit problem is a classical optimization problem where an agent sequentially pulls...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2010 – Universitetet i Agder, GrimstadMulti...
Published version of an article from Lecture Notes in Computer Science. Also available at SpringerLi...
Abstract—We present a formal model of human decision-making in explore-exploit tasks using the conte...
The two-armed bandit problem is a classical optimization problem where a decision maker sequentially...
Published version of an article in the journal: Applied Intelligence. Also available from the publis...
Published version of a chapter from the book: Modern Approaches in Applied Intelligence. Also availa...
The two-armed bandit problem is a classical optimization problem where a decision maker sequentially...
Published version of a chapter from the book: Modern Approaches in Applied Intelligence. Also availa...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2009 – Universitetet i Agder, GrimstadThe t...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2009 – Universitetet i Agder, GrimstadThe t...
The two-armed bandit problem is a classical optimization problem where a player sequentially selects...
Multi-Armed bandit problem is a classic example of the exploration vs. exploitation dilemma in which...
Published version of a chapter in the book: IFIP Advances in Information and Communication Technolog...
Published version of an article from Lecture Notes in Computer Science. Also available at SpringerLi...
The multi-armed bandit problem is a classical optimization problem where an agent sequentially pulls...
Masteroppgave i informasjons- og kommunikasjonsteknologi 2010 – Universitetet i Agder, GrimstadMulti...
Published version of an article from Lecture Notes in Computer Science. Also available at SpringerLi...
Abstract—We present a formal model of human decision-making in explore-exploit tasks using the conte...