PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 5, 2026Machine Learning and Knowledge Extraction0 citationsOpen Access

When to Explore and When to Exploit: Adaptive Decisions in Bayesian Optimization

View Full Paper
ACAntonio CandelieriFAFrancesco ArchettiISIman Seyedi

Key Points

  • The aim is to enhance the exploration-exploitation trade-off in Bayesian optimization through a novel adaptive acquisition function.
  • Proposed a surrogate-agnostic acquisition function that adjusts exploration-exploitation dynamics based on optimization evolution.
  • Implemented within Gaussian processes for Bayesian optimization, using inverse distance weighting for uncertainty quantification.
  • Tested on benchmark functions to assess Pareto optimality and balance between convergence and exploration.
  • The adaptive acquisition function achieved near-perfect Pareto optimality compared to state-of-the-art methods.
  • Demonstrated a superior balance between convergence to the global optimum and exploration capabilities in tests.

Abstract

Gaussian process-based Bayesian optimization (BO) is a sample-efficient sequential strategy for optimizing expensive black-box functions. The Gaussian process provides a probabilistic approximation of the unknown function, while an acquisition function balances exploration and exploitation to select the next evaluation point. Despite significant research efforts, no master acquisition function has been identified. This paper proposes a novel adaptive acquisition function that dynamically adjusts the exploration–exploitation trade-off based on the evolution of the optimization process, rather than using fixed or random scheduling. While implemented here within a GP-based BO framework, the core switching mechanism is surrogate-agnostic: the exploitative component requires only a surrogate point prediction, and the explorative component is entirely model-free. Unlike traditional approaches, where mechanisms like UCB/LCB lean toward exploration over iterations, or fixed strategies that switch from exploratory (EI) to exploitative (PI) behavior at predetermined points, the proposed method makes purely exploitative decisions using only the GP’s prediction. However, it discards these decisions when they have low potential for significant improvement, instead focusing on uncertainty reduction. Notably, this approach uses inverse distance weighting for uncertainty quantification rather than the GP’s predictive uncertainty, avoiding bias from the GP’s predictions. Testing on benchmark functions demonstrates that the proposed acquisition function is almost always Pareto optimal, offering the most balanced trade-off between convergence to the global optimum and exploration capability compared to state-of-the-art alternatives.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Candelieri et al. (2026) studied this question.

synapsesocial.com/papers/6a49f464f5d1d45b287ffe62https://doi.org/10.3390/make8070193
Ask AI
Helpful
Bookmark
Share
View Full Paper