PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 22, 2026Mathematical Methods of Operations Research0 citationsOpen Access

Learning the follower’s objective function in sequential bilevel games

IMIoana MolanMSMartin SchmidtJTJohannes Thürauf

Key Points

  • This research investigates how a leader can learn the follower’s objective function in bilevel optimization problems through repeated interactions.
  • Utilized a multiplicative weight update (MWU) method and an inverse optimization approach using KKT conditions.
  • Analyzed two specific applications: continuous knapsack interdiction problem and sequential bilevel pricing game.
  • Emphasized learning through the leader's observations of the follower's choices over time.
  • The MWU method shows convergence towards follower's objective function values over time, while the inverse KKT method aligns learned functions with previous interactions.
  • Both methods help the leader make more informed decisions by approximating outcomes similar to those achievable with full information.

Abstract

Abstract We consider bilevel optimization problems in which the leader has no or only partial knowledge about the objective function of the follower. The studied setting is a sequential one in which the bilevel game is played repeatedly. This allows the leader to learn the objective function (values) of the follower over time. We focus on two methods: a multiplicative weight update (MWU) method and one based on the lower-level’s KKT conditions that are used in the sense of inverse optimization. The MWU method requires less assumptions but the convergence guarantee is also only on the follower’s objective function values, whereas the inverse KKT method requires stronger assumptions but actually allows to learn objective functions that are consistent with the already observed interactions between the two players. Although the theory we present is only related to the lower-level and not to the upper-level problem, we show that the gained information are practically useful for the leader by illustrating that, over time, the leader’s objective function values tend to those that would be obtained under full information. The applicability of the proposed methods is shown using two case studies. First, we study a repeatedly played continuous knapsack interdiction problem and, second, a sequential bilevel pricing game in which the leader needs to learn the utility function of the follower. For both problems, we further illustrate the impact of this learning on the leader’s decisions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Molan et al. (2026) studied this question.

synapsesocial.com/papers/699a9ded482488d673cd429ahttps://doi.org/10.1007/s00186-025-00908-0
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Neur2BiLO: Neural Bilevel Optimization2024 · 2 citations
  2. 2Bilevel Programming Problems2015 · 274 citations
  3. 3Untitled2012 · 928 citations
  4. 4The Cost of Subsistence1945 · 494 citations
  5. 5Sequential Shortest Path Interdiction with Incomplete Information2015 · 54 citations