Approximate optimality and the risk/reward tradeoff in a class of bandit problems

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Chen, Zengjing, Epstein, Larry G., Zhang, Guodong
Natura: Preprint
Pubblicazione: 2022
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910283892523008
author Chen, Zengjing
Epstein, Larry G.
Zhang, Guodong
author_facet Chen, Zengjing
Epstein, Larry G.
Zhang, Guodong
contents This paper studies a sequential decision problem where payoff distributions are known and where the riskiness of payoffs matters. Equivalently, it studies sequential choice from a repeated set of independent lotteries. The decision-maker is assumed to pursue strategies that are approximately optimal for large horizons. By exploiting the tractability afforded by asymptotics, conditions are derived characterizing when specialization in one action or lottery throughout is asymptotically optimal and when optimality requires intertemporal diversification. The key is the constancy or variability of risk attitude. The main technical tool is a new central limit theorem.
format Preprint
id arxiv_https___arxiv_org_abs_2210_08077
institution arXiv
publishDate 2022
record_format arxiv
spellingShingle Approximate optimality and the risk/reward tradeoff in a class of bandit problems
Chen, Zengjing
Epstein, Larry G.
Zhang, Guodong
Theoretical Economics
Probability
This paper studies a sequential decision problem where payoff distributions are known and where the riskiness of payoffs matters. Equivalently, it studies sequential choice from a repeated set of independent lotteries. The decision-maker is assumed to pursue strategies that are approximately optimal for large horizons. By exploiting the tractability afforded by asymptotics, conditions are derived characterizing when specialization in one action or lottery throughout is asymptotically optimal and when optimality requires intertemporal diversification. The key is the constancy or variability of risk attitude. The main technical tool is a new central limit theorem.
title Approximate optimality and the risk/reward tradeoff in a class of bandit problems
topic Theoretical Economics
Probability
url https://arxiv.org/abs/2210.08077