期刊文献+
共找到1篇文章
< 1 >
每页显示 20 50 100
Deep Reinforcement Learning with Fuse Adaptive Weighted Demonstration Data
1
作者 Baofu Fang taifeng guo 《国际计算机前沿大会会议论文集》 2022年第1期163-177,共15页
Traditional multi-agent deep reinforcement learning has difficulty obtaining rewards,slow convergence,and effective cooperation among agents in the pretraining period due to the large joint state space and sparse rewa... Traditional multi-agent deep reinforcement learning has difficulty obtaining rewards,slow convergence,and effective cooperation among agents in the pretraining period due to the large joint state space and sparse rewards for action.Therefore,this paper discusses the role of demonstration data in multiagent systems and proposes a multi-agent deep reinforcement learning algorithm from fuse adaptive weight fusion demonstration data.The algorithm sets the weights according to the performance and uses the importance sampling method to bridge the deviation in the mixed sampled data to combine the expert data obtained in the simulation environment with the distributed multi-agent reinforcement learning algorithm to solve the difficult problem.The problem of global exploration improves the convergence speed of the algorithm.The results in the RoboCup2D soccer simulation environment show that the algorithm improves the ability of the agent to hold and shoot the ball,enabling the agent to achieve a higher goal scoring rate and convergence speed relative to demonstration policies and mainstream multi-agent reinforcement learning algorithms. 展开更多
关键词 Multiagent deep reinforcement learning Exploration Offline reinforcement learning Importance sampling
原文传递
上一页 1 下一页 到第
使用帮助 返回顶部