ERC Advanced Grant · 2019
Computational learning theory has been highly successful over the last three decades, both in providing deep mathematical theories and in influencing the practice of machine learning. One of the great recent successes of computational learning theory has been the study of online learning and multi-arm bandits. This line of research has been highly successful, both theoretically and practically, addressing many important applications. Unfortunately, the recent theoretical progress in Markov Decision Process and reinforcement learning has been slower. Based on my fundamental contributions to reinforcement learning (e.g. policy gradient, sparse sampling and trajectory trees), to online…
From the public funding record at EU CORDIS. Describes the funded project, not the reviews below.