Deep reinforcement learning-based resource allocation for D2D communications in heterogeneous cellular networks | Synapse