Search CORE

1,020 research outputs found

Online learning with graph-structured feedback against adaptive adversaries

Author: Feng Zhili
Loh Po-Ling
Publication venue
Publication date: 01/04/2018
Field of study

We derive upper and lower bounds for the policy regret of

T

-round online learning problems with graph-structured feedback, where the adversary is nonoblivious but assumed to have a bounded memory. We obtain upper bounds of

\widetilde O(T^{2/3})

and

\widetilde O(T^{3/4})

for strongly-observable and weakly-observable graphs, respectively, based on analyzing a variant of the Exp3 algorithm. When the adversary is allowed a bounded memory of size 1, we show that a matching lower bound of

\widetilde\Omega(T^{2/3})

is achieved in the case of full-information feedback. We also study the particular loss structure of an oblivious adversary with switching costs, and show that in such a setting, non-revealing strongly-observable feedback graphs achieve a lower bound of

\widetilde\Omega(T^{2/3})

, as well.Comment: This paper has been accepted to ISIT 201

arXiv.org e-Print Archive

Optimal Allocation Strategies for the Dark Pool Problem

Author: Agarwal Alekh
Bartlett Peter
Dama Max
Publication venue
Publication date: 01/01/2010
Field of study

We study the problem of allocating stocks to dark pools. We propose and analyze an optimal approach for allocations, if continuous-valued allocations are allowed. We also propose a modification for the case when only integer-valued allocations are possible. We extend the previous work on this problem to adversarial scenarios, while also improving on their results in the iid setup. The resulting algorithms are efficient, and perform well in simulations under stochastic and adversarial inputs

arXiv.org e-Print Archive

CiteSeerX