• Learning is of great importance in reinforcement learning.

    学习是一种重要的强化学习算法。

    youdao

  • Without reinforcement learning is only short term and easily lost.

    没有巩固的学习只能是短期的,很快遗忘的。

    youdao

  • Can you explain the A. I. technique called reinforcement learning?

    你能解释一下什么是“强化学习”技术吗?

    youdao

  • What makes a task more appropriate for incorporating reinforcement learning?

    什么样的任务更适合应用强化学习技术?

    youdao

  • What are the differences between supervised learning and reinforcement learning?

    监督学习与强化学习的区别是什么?

    youdao

  • This sample graph is from a simple reinforcement learning application that USES Q learning.

    这个示例图是从使用Q学习的一个简单增强式学习应用程序中得到的。

    youdao

  • Contemporary theories of reinforcement learning are rooted in the dopaminergic reward system.

    当代的强化学习理论是基于多巴胺奖赏系统。

    youdao

  • The former one is a new approach combining reinforcement learning with feedback control.

    基于强化学习的多指手控制方法,该方法将反馈控制与强化学习相结合。

    youdao

  • Reinforcement learning (RL) to motion planning of dynamic manipulation tasks was applied.

    提出增强学习(RL)解决机器人动态操作任务运动规划的方法。

    youdao

  • An average reward reinforcement learning algorithm for control Markov chains is presented.

    讨论平均准则控制马氏链的强化学习算法。

    youdao

  • This paper adopts reinforcement learning method to accomplish robot soccer cooperation strategy.

    利用强化学习方法实现足球机器人协作策略。

    youdao

  • Multilog correlation is a core log-data analysis use case for unsupervised and reinforcement learning.

    多样相关性是用与无监视和强化学习的焦点日志数据剖析使用案例。

    youdao

  • MAXQ, a hierarchical reinforcement learning method for multi-agent system, is proposed in recent years.

    MAXQ分层多智能体学习方法是近年来被提出的一种新方法。

    youdao

  • Simulation machine car through reinforcement learning algorithm, learning optimal navigation strategies.

    说明:模拟智能机器小车,通过强化学习算法,学习最优导航策略。

    youdao

  • Research on local path planning of mobile robot based on Q reinforcement learning and CMAC neural networks.

    基于Q强化学习与CMAC神经网络的移动机器人局部路径规划研究。

    youdao

  • Taking the decision situation with the lowest information constraints, only reinforcement learning model is chosen.

    考虑场景特征和各种学习理论要求的最低信息条件,只有强化模型可以在本文中使用。

    youdao

  • Several approaches applying reinforcement learning techniques to game playing have been described in the literature.

    将强化学习技术运用于游戏的集中方法在文献里都有记载。

    youdao

  • This paper discusses reinforcement learning(RL)algorithm and its application to technical action learning of soccer robot.

    主要研究了强化学习算法及其在机器人足球比赛技术动作学习问题中的应用。

    youdao

  • The thesis mainly focuses on the dynamic scheduling method based on the averaged rewards reinforcement learning algorithms.

    论文主要研究了基于平均型强化学习算法的动态调度方法。

    youdao

  • Reinforcement learning has the ability to learn from experience as opposed to supervised learning which learns from examples.

    与监督学习从范例中学习的方式不同,强化学习不需要先验知识,而是具有从经验中学习的能力。

    youdao

  • I think where reinforcement learning has some challenges is when the action-state you may take is incredibly broad and large.

    我觉得强化学习面临的一些挑战主要集中在当你可以采取的行为状态极为宽泛的时候。

    youdao

  • For vector control AC drive system, the thesis presented a fuzzy neural network speed controller based on reinforcement learning.

    针对矢量控制交流调速系统,该文提出并设计了一种基于再励学习的模糊神经网络速度控制器。

    youdao

  • It is rational to adopt the average reward reinforcement learning algorithms for solving the absorbing goal states cyclical tasks.

    对于有吸收目标状态的循环任务,比较合理的方法是采用基于平均报酬模型的强化学习。

    youdao

  • A reinforcement learning algorithm based on process reward and prioritized sweeping is presented as interference solving strategy.

    本文提出了基于过程奖赏和优先扫除的强化学习算法作为多机器人系统的冲突消解策略。

    youdao

  • On the basis of theoretical analysis, the cooperative game reinforcement learning method is proposed and its convergence is proved.

    在理论分析的基础上,提出了协同博弈的强化学习算法,并证明了算法的收敛性。

    youdao

  • Reinforcement learning is an important machine learning method. However, slow convergence has been one of main problem in practice.

    强化学习是一种重要的机器学习方法,然而在实际应用中,收敛速度缓慢是其主要不足之一。

    youdao

  • Reinforcement learning based on Markov decision process is a way of on-line learning, which can be applied to single agent environment.

    基于马尔科夫过程的强化学习作为一种在线学习方式,能够很好地应用于单智能体环境中。

    youdao

  • In this paper, the approximate theorem of average reward reinforcement learning is proven by means of the theory of performance potentials.

    文中基于性能势理论,证明了平均奖赏强化学习的逼近定理。

    youdao

  • Reinforcement learning is a common technique for this scenario as well as the more traditional scenario of actually learning the utility function.

    强化学习是这种情况下的常用技术,而更多的传统情形下需要使用效用函数。

    youdao

  • Here the computational principle is reinforcement learning and active exploration, which may also be behind learning motor movements in an infant.

    在这里计算原理是加强学习过程和主动探索过程,这些也许也是婴儿学习动机背后的原因。

    youdao

$firstVoiceSent
- 来自原声例句
小调查
请问您想要如何调整此模块?

感谢您的反馈,我们会尽快进行适当修改!
进来说说原因吧 确定
小调查
请问您想要如何调整此模块?

感谢您的反馈,我们会尽快进行适当修改!
进来说说原因吧 确定