Q-learning is a typical Reinforcement Learning (RL) method with a slow convergence speed especially as the scales of the state space and action space increase.
学习是一种典型的强化学习,其学习效率较低,尤其是当状态空间和决策空间较大时。
youdao
应用推荐
模块上移
模块下移
不移动