English description :Self-improving LLM agents are advancing rapidly (Reflexion, Voyager, DSPy, AI Scientist, STaR, Self-Rewarding Language Models), but four structural problems limit their scaling · instability under self-modification, lack of a rigorous modification acceptance criterion, intrinsic adversarial vulnerability, and evaluation cost incompatible with real-time execution. The current ML literature addresses these problems case by case, with ad hoc heuristics. This note proposes another reading · all four share a common formal structure with that of structural control of distributed networks under profiled KPIs, a problem long studied in applied graph theory and resilient networks. The defended hypothesis is that mechanisms proven in this field · efficient structural proxies, multi-signal acceptance logic, adversarial defenses through topological invariants, bounded execution without a global solver · can be transposed to the control of self-improving agents, providing a rigor that the current literature does not offer. The note states the common grammar, proposes four concrete contributions, identifies areas of uncertainty, and traces a falsifiable research program in three stages. French description : Les agents auto-améliorants progressent rapidement (Reflexion, Voyager, DSPy, AI Scientist, STaR, Self-Rewarding Language Models), mais quatre problèmes structurels limitent leur passage à l'échelle · instabilité sous auto-modification, absence de critère rigoureux d'acceptation des modifications, vulnérabilité adversariale intrinsèque, coût d'évaluation incompatible avec une exécution temps réel. La littérature ML actuelle traite ces quatre problèmes au cas par cas, par heuristiques ad hoc. Cette note propose une autre lecture · ils partagent une structure formelle commune avec celui du contrôle structurel de réseaux distribués sous KPIs profilés, problème étudié de longue date en théorie des graphes appliquée et en réseaux résilients. L'hypothèse défendue est que des mécanismes éprouvés dans ce champ · proxys structurels efficaces, logiques d'acceptation multi-signal, défenses adversariales par invariants topologiques, exécution bornée sans solveur global · peuvent être transposés au contrôle d'agents auto-améliorants, et y apporter une rigueur que la littérature actuelle ne fournit pas. La note pose la grammaire commune, propose quatre apports concrets, identifie les zones d'incertitude, et trace un programme de recherche falsifiable en trois étapes (formalisation, prototype minimal, études comparées). Ce n'est pas un résultat. C'est une hypothèse de recherche, qui se veut falsifiable.
Mohammed ZERROUK (Fri,) studied this question.