Optimal Consensus Tracking Control for Nonlinear Multi-Agent Systems via Actor–Critic Reinforcement Learning | Synapse