go top

reinforcement learning

  • n. 强化学习,增强学习:机器学习的一种范式,智能体通过与环境的交互进行试错学习,依据奖励或惩罚反馈调整策略,以最大化长期累积回报。

专业释义英英释义

  • 强化学习 - 引用次数:375

    Q learning is of great importance in reinforcement learning.

    Q 学习是一种重要的强化学习算法。

    参考来源 - 一种多步Q强化学习方法 in C
    增强学习 - 引用次数:46

    Thus,reinforcement learning methods have wide application areas in solving complex optimization and decision problems,where teacher signals are hard to be obtained.

    因此增强学习在求解无法获得教师信号的复杂优化与决策问题中具有广泛的应用前景。

    参考来源 - 增强学习及其在移动机器人导航与控制中的应用研究
    激励学习 - 引用次数:23

    That presented in this paper is a new utility clustering based reinforcement learning algorithm called U-Clustering. Unlike the U-Tree,it does not use fringe and related statistical test at all.

    提出了一个新的效用聚类激励学习算法U—Clustering。

    参考来源 - 期刊学术社区
    再励学习 - 引用次数:19

    3. Since traditional sparse coding does not consider the intra-class variations of the features, two new sparse coding algorithms based on reinforcement learning are proposed to increase the classification property of the features.

    3.针对传统稀疏编码未对特征的类内距离进行有效约束这个缺点,提出了两种 基于再励学习的稀疏编码算法。

    参考来源 - 人脸图像分析和识别方法研究
    加强学习
  • 强化学习 - 引用次数:24

    参考来源 - 生产调度问题的智能优化方法研究及应用
    增强学习 - 引用次数:2

    参考来源 - 无线移动网络中智能位置管理策略研究
    自加强学习算法
  • 再励学习 - 引用次数:4

    The concepts of Markov decision process and reinforcement learning are introduced firstly.

    论文首先介绍了马尔可夫决策过程的基本概念和再励学习的框架。

    参考来源 - 再励学习在通信网络最优控制中的应用研究
  • 激励学习 - 引用次数:3

    Computer game playing is an interesting subject in artificial intelligence research. Reinforcement learning is a machine learning method to study Superior strategy through trial-and-error search and delayed reward from Environment-feedback.

    人机博弈是人工智能领域中的一个重要主题,激励学习是一种智能体通过不断地试错,从环境反馈中得到延迟奖惩信息,积累经验,最终学习到最优策略的机器学习方法。

    参考来源 - 基于激励学习的中国象棋研究

·2,447,543篇论文数据,部分数据来源于NoteExpress

Reinforcement learning

  • abstract: Inspired by behaviorist psychology, reinforcement learning is an area of machine learning in computer science, concerned with how software agents ought to take actions in an environment so as to maximize some notion of cumulative reward. The problem, due to its generality, is studied in many other disciplines, such as game theory, control theory, operations research, information theory, simulation-based optimization, statistics, and genetic algorithms.

以上来源于: WordNet

双语例句权威例句

  • Learning is of great importance in reinforcement learning.

    学习一种重要强化学习算法。

    youdao

  • Without reinforcement learning is only short term and easily lost.

    没有巩固学习只能短期的,很快遗忘的。

    youdao

  • This sample graph is from a simple reinforcement learning application that USES Q learning.

    这个示例是从使用Q学习一个简单增强式学习应用程序中得到的。

    youdao

更多双语例句
$firstVoiceSent
- 来自原声例句
小调查
请问您想要如何调整此模块?

感谢您的反馈,我们会尽快进行适当修改!
进来说说原因吧 确定
小调查
请问您想要如何调整此模块?

感谢您的反馈,我们会尽快进行适当修改!
进来说说原因吧 确定