Reinforcement learning framework for computerized adaptive testing using multi armed bandit approach | Synapse