Reinforcement Learning, Bit by BitXiuyuan Lu, Benjamin Van Roy, Vikranth Dwaracherla, Morteza Ibrahimi, Ian Osband · First published 2023Open the Tome
Tutorial on Thompson SamplingDaniel J. Russo, Benjamin Van Roy, Abbas Kazerouni, Ian Osband, Zheng Wen · First published 2018Open the Tome