期刊文献+
共找到1篇文章
< 1 >
每页显示 20 50 100
A policy iteration method for improving robot assembly trajectory efficiency
1
作者 Qi ZHANG Zongwu XIE +1 位作者 Baoshi CAO Yang LIU 《Chinese Journal of Aeronautics》 SCIE EI CAS CSCD 2023年第3期436-448,共13页
Bolt assembly by robots is a vital and difficult task for replacing astronauts in extravehicular activities(EVA),but the trajectory efficiency still needs to be improved during the wrench insertion into hex hole of bo... Bolt assembly by robots is a vital and difficult task for replacing astronauts in extravehicular activities(EVA),but the trajectory efficiency still needs to be improved during the wrench insertion into hex hole of bolt.In this paper,a policy iteration method based on reinforcement learning(RL)is proposed,by which the problem of trajectory efficiency improvement is constructed as an issue of RL-based objective optimization.Firstly,the projection relation between raw data and state-action space is established,and then a policy iteration initialization method is designed based on the projection to provide the initialization policy for iteration.Policy iteration based on the protective policy is applied to continuously evaluating and optimizing the action-value function of all state-action pairs till the convergence is obtained.To verify the feasibility and effectiveness of the proposed method,a noncontact demonstration experiment with human supervision is performed.Experimental results show that the initialization policy and the generated policy can be obtained by the policy iteration method in a limited number of demonstrations.A comparison between the experiments with two different assembly tolerances shows that the convergent generated policy possesses higher trajectory efficiency than the conservative one.In addition,this method can ensure safety during the training process and improve utilization efficiency of demonstration data. 展开更多
关键词 Bolt assembly policy initialization policy iteration Reinforcement learning(RL) Robotic assembly Trajectory efficiency
原文传递
上一页 1 下一页 到第
使用帮助 返回顶部