結果 : reinforcement learning basic code