Intrinsic gradient oxygen-driven second-order memristors for continual reinforcement learning | Synapse