Taking the decision situation with the lowest information constraints, only reinforcement learning model is chosen.
考虑场景特征和各种学习理论要求的最低信息条件,只有强化模型可以在本文中使用。
Honeybees are frequently used as a neurobiological model for learning-they can be trained, using positive or negative reinforcement, to retain information.
蜜蜂经常被作为学习神经生物学的模型——它们可以被训练,用正负强化法保留信息。
The model adopts a new reinforcement function. All the methods can improve agent's learning velocity.
新的奖励函数和表示方法减少了学习空间、增加了学习速度。
应用推荐