papersSEP 10 04:00 UTC
Meta-LinEXP3: Meta-Learning for Adversarial Linear Contextual Bandits
A new arXiv paper presents Meta-LinEXP3, an algorithm that brings meta-learning to adversarial linear contextual bandits, a setting where earlier results largely covered only stochastic or non-contextual variants. The method uses an online-within-online framework to transfer knowledge across a sequence of bandit tasks.