When combined with the Markov decision process, it provides a new formalization suitable for multi-agent system. That is stochastic game concerning the interactive learning system of multi-agent.
对策论与马尔可夫决策过程相结合便构建了一个用于研究交互式多agent学习的理论框架——随机对策。
参考来源 - 结合围捕问题的合作多智能体强化学习研究·2,447,543篇论文数据,部分数据来源于NoteExpress
以上来源于: WordNet
The dynamic CRM model is presented by using stochastic game theory and estimable structural dynamic programming technologies.
运用随机博弈及可评估的结构动态规划技术建立动态客户关系管理的数学模型。
It is proved that under certain condition the stochastic game has a value and both players have optimal strategies for discounted rewards.
在一定条件下,我们证明随机对策有值函数,两个局中人相对于折扣报酬都有最优策略。
We study dynamic measure of risk problem in incomplete market when stock appreciation rates are uncertainty. We also study a related stochastic game problem.
本文讨论不完全市场中股票收益率不确定时的动态风险度量问题和一个相关的随机对策问题。
应用推荐