PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 25, 20240 citationsOpen Access

Multi-Player Approaches for Dueling Bandits

View Full Paper
OROr RavehJHJunya HondaMSMasashi Sugiyama

Key Points

Key points are not available for this paper at this time.

Abstract

Various approaches have emerged for multi-armed bandits in distributed systems. The multiplayer dueling bandit problem, common in scenarios with only preference-based information like human feedback, introduces challenges related to controlling collaborative exploration of non-informative arm pairs, but has received little attention. To fill this gap, we demonstrate that the direct use of a Follow Your Leader black-box approach matches the lower bound for this setting when utilizing known dueling bandit algorithms as a foundation. Additionally, we analyze a message-passing fully distributed approach with a novel Condorcet-winner recommendation protocol, resulting in expedited exploration in many cases. Our experimental comparisons reveal that our multiplayer algorithms surpass single-player benchmark algorithms, underscoring their efficacy in addressing the nuanced challenges of the multiplayer dueling bandit setting.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Raveh et al. (2024) studied this question.

synapsesocial.com/papers/68e686d2b6db64358760fe17https://doi.org/10.48550/arxiv.2405.16168
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Multi-agent Multi-armed Bandits with Stochastic Sharable Arm Capacities2024
  2. 2Biased Dueling Bandits with Stochastic Delayed Feedback2024
  3. 3Conversational Dueling Bandits in Generalized Linear Models2024 · 6 citations
  4. 4Multi-player Multi-armed Bandits with Delayed Feedback2025
  5. 5A Study of Exploration-Exploitation Strategies in Unconventional Situations2025