Search Machine Learning Repository:
Tracking Adversarial Targets
Authors: Yasin Abbasi-yadkori, Peter Bartlett and Varun Kanade
Conference: Proceedings of the 31st International Conference on Machine Learning (ICML-14)
Abstract: We study linear control problems with quadratic losses and adversarially chosen tracking targets. We present an efficient algorithm for this problem and show that, under standard conditions on the linear system, its regret with respect to an optimal linear policy grows as $O(\log^2 T)$, where $T$ is the number of rounds of the game. We also study a problem with adversarially chosen transition dynamics; we present an exponentially-weighted average algorithm for this problem, and we give regret bounds that grow as $O(\sqrt T)$.
authors venues years
Suggest Changes to this paper.
Brought to you by the WUSTL Machine Learning Group. We have open faculty positions (tenured and tenure-track).