Cascading Bandits Robust to Adversarial Corruptions

Jize Xie; Cheng Chen; Zhiyong Wang; Shuai Li

arXiv:2502.08077·cs.LG·February 13, 2025

Cascading Bandits Robust to Adversarial Corruptions

Jize Xie, Cheng Chen, Zhiyong Wang, Shuai Li

PDF

Open Access

TL;DR

This paper introduces robust algorithms for cascading bandits that can withstand adversarial feedback corruptions, maintaining low regret in the presence of manipulation, which is crucial for reliable online ranking systems.

Contribution

The paper formulates the CBAC problem and proposes two algorithms that are robust to adversarial corruptions, with proven regret bounds and empirical validation.

Findings

01

Algorithms achieve logarithmic regret without attack.

02

Regret increases linearly with corruption level.

03

Experimental results confirm robustness.

Abstract

Online learning to rank sequentially recommends a small list of items to users from a large candidate set and receives the users' click feedback. In many real-world scenarios, users browse the recommended list in order and click the first attractive item without checking the rest. Such behaviors are usually formulated as the cascade model. Many recent works study algorithms for cascading bandits, an online learning to rank framework in the cascade model. However, the performance of existing methods may drop significantly if part of the user feedback is adversarially corrupted (e.g., click fraud). In this work, we study how to resist adversarial corruptions in cascading bandits. We first formulate the ``\textit{Cascading Bandits with Adversarial Corruptions}" (CBAC) problem, which assumes that there is an adaptive adversary that may manipulate the user feedback. Then we propose two…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Adversarial Robustness in Machine Learning · Imbalanced Data Classification Techniques

MethodsSparse Evolutionary Training