Closing the Gap between Bandit and Full-Information Online Optimization: High-Probability Regret Bound

Bartlett, Peter; Rakhlin, Alexander; Tewari, Ambuj; EECS Department, University of California

PDF

Description

We demonstrate a modification of the algorithm of Dani et al. for the online linear optimization problem in the bandit setting, which allows us to achieve an O(sqrt(T ln T)) regret bound in high probability against an adaptive adversary, as opposed to the in expectation result against an oblivious adversary of Dani et al. We obtain the same dependence on the dimension as that exhibited by Dani et al. The results of this paper rest firmly on those of Dani et al. and the remarkable technique of Auer et al. for obtaining high-probability bounds via optimistic estimates. This paper answers an open question: it eliminates the gap between the high-probability bounds obtained in the full-information vs bandit settings.

Details

Title

Closing the Gap between Bandit and Full-Information Online Optimization: High-Probability Regret Bound

Creator

Bartlett, Peter, Author
Rakhlin, Alexander, Author
Tewari, Ambuj, Author
EECS Department, University of California, Publisher

Published

2007-08-26

Full Collection Name

Electrical Engineering & Computer Sciences Technical Reports

Other Identifiers

EECS-2007-109

Type

Text

Format

technical reports

Extent

14 p

Archive

The Engineering Library

Usage Statement

Researchers may make free and open use of the UC Berkeley Library’s digitized public domain materials. However, some materials in our online collections may be protected by U.S. copyright law (Title 17, U.S.C.). Use or reproduction of materials protected by copyright beyond that allowed by fair use (Title 17, U.S.C. § 107) requires permission from the copyright owners. The use or reproduction of some materials may also be restricted by terms of University of California gift or purchase agreements, privacy and publicity rights, or trademark law. Responsibility for determining rights status and permissibility of any use or reproduction rests exclusively with the researcher. To learn more or make inquiries, please see our permissions policies (https://www.lib.berkeley.edu/about/permissions-policies).

Collection

EECS Technical Reports

Files

Statistics

Download Full History

Download

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DublinCore
EndNote
NLM
RefWorks
RIS

Add to Basket