Online learning with noisy side observations

Tom\'a\v{s} Koc\'ak; Gergely Neu; Michal Valko

arXiv:2604.13740·cs.LG·April 16, 2026·26 cites

Online learning with noisy side observations

Tom\'a\v{s} Koc\'ak, Gergely Neu, Michal Valko

PDF

Abstract

We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback about the other actions, depending on the underlying structure of the problem. We represent this structure by a weighted directed graph, where the edge weights are related to the quality of the feedback shared by the connected nodes. Our main contribution is an efficient algorithm that guarantees a regret of $O (α^{*} T)$ after $T$ rounds, where $α^{*}$ is a novel graph property that we call the effective independence number. Our algorithm is completely parameter-free and does not require knowledge (or even estimation) of $α^{*}$ . For the special case of binary edge weights, our setting reduces to the partial-observability models of Mannor and Shamir (2011) and Alon et al. (2013) and our algorithm recovers the…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.