Hidden-layer configurations in reinforcement learning models for stock portfolio optimization | Synapse