Sample PATH modification to include the JITI library.
PATH修改为包含jiti库的样例。
In the optimization of service policy for closed queuing networks, the method of optimizing policy that is based on single sample path of the systems is of practical interest.
在闭排队网络服务策略的优化中,基于对系统一条样本轨道的仿真进行策略优化是一种很有实用意义的方法。
By the simulation of a sample path, and the approximation to performance potentials via neural networks based on reinforcement learning (RL), the systems optimization methods are provided.
在样本轨道仿真的基础上,利用神经网络进行强化学习仿真逼近系统的性能势,进而对系统进行优化。
应用推荐