Dynamic allocation and optimization strategy of communication network resources driven by reinforcement learning | Synapse