Reinforcement Learning for Synchronization of Heterogeneous Multiagent Systems by Improved Q-Functions | Synapse