PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 24, 2026Transactions on Emerging Telecommunications Technologies0 citations

A Reinforcement Learning–Driven Payoff‐Adaptive Game‐Theoretic Framework for Secure and Reliable Operation of Mobile Ad Hoc Networks

View Full Paper
SZShaik Khader ZelaniKRK. V. S. S. RamakrishnaFAF. Asiri

Key Points

  • The research aims to develop a framework for secure operation in mobile ad hoc networks (MANETs) facing various adversarial threats.
  • Integrated a reinforcement learning-driven, payoff-adaptive game-theoretic framework.
  • Employed online payoff learning and Boltzmann-guided Q-learning for dynamic updates.
  • Utilized feature-aware utility estimation to optimize security, throughput, energy, and latency.
  • Achieved near-optimal reliability with high packet delivery ratios and low latency.
  • Demonstrated strong resilience against seven types of adversarial threats.
  • Improved utility metrics and reduced impacts of misreporting and collusion.

Abstract

ABSTRACT Mobile ad hoc networks (MANETs) are inherently vulnerable to selfish forwarding, collusion, stealthy oscillations, and resource‐draining attacks. Due to their decentralized, dynamic, and adversarial nature, existing anomaly‐based and deep‐learning intrusion detection systems rely on static training distributions and fixed payoff structures. This limits their adaptability under dynamic network conditions. This paper presents a secure and reliable operation of MANETs under adversarial environments. The model integrates a reinforcement learning–driven (RL‐driven), payoff‐adaptive, three‐action game‐theoretic (GT) framework. This combination helps learning both the utility structure and the optimal defense strategies for MANET nodes. The model employs online payoff learning, Boltzmann‐guided Q‐learning, and feature‐aware utility estimation to dynamically update security, throughput, energy, and latency trade‐offs. Experiments on a large‐scale NS‐3 mobility dataset ( trace snapshots) over 80‐node MANETs, 500 rounds, and 10 random seeds demonstrate a strong stability and resilience. The system achieves near‐optimal reliability (packet delivery ratio, , low latency , efficient energy usage controlled energy consumption and high average utility . Under seven adversarial threats, the framework consistently restores optimal operation. misreporting impact is reduced, with utility improving from , collusion inflation suppression remains , and regret normalization goes from to . Overall, the proposed GT–RL model provides adaptive equilibrium formation, multi‐metric flexibility, and theoretical convergence guarantees for secure MANET operation.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zelani et al. (2026) studied this question.

synapsesocial.com/papers/69c2299aaeb5a845df0d4493https://doi.org/10.1002/ett.70394
Ask AI
Helpful
Bookmark
Share
View Full Paper