Authors
Loading...
Research proposes algorithms improving regret bounds in lexicographic multi-armed bandits, highlighting adversarial corruptions and varying objectives.
Xue et al. (2025) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: