摘要
Obstacle avoidance is an important issue in the motion planning of autonomous unmanned systems. Therefore, designing an effective avoidance control method is crucial. For further improving the decision-making process, this paper presents a novel autonomous obstacle avoidance control method based on reinforcement learning that generates a safe motion trajectory in an adaptive manner. First, the barrier function is utilized to design a smooth penalty function in the cost function, thereby transforming the avoidance problem into an unconstrained optimal control problem. Then, adaptive reinforcement learning is implemented by using an actor-critic neural network architecture and policy iteration, in which the critic network uses the state-following kernel function to approximate the cost function while the actor network provides an approximate optimal control policy. During this learning process, the simulated experience is obtained through state extrapolation such that the critic network can use experience replay for reliable local exploration. Finally, simulation experiments on simplified drone systems and a nonlinear numerical system are provided. The proposed method can generate a safe motion trajectory in real time with comparable performance.
| 投稿的翻译标题 | Autonomous obstacle avoidance control method based on safe adaptive reinforcement learning |
|---|---|
| 源语言 | 繁体中文 |
| 页(从-至) | 1672-1686 |
| 页数 | 15 |
| 期刊 | Scientia Sinica Informationis |
| 卷 | 52 |
| 期 | 9 |
| DOI | |
| 出版状态 | 已出版 - 2022 |
| 已对外发布 | 是 |
关键词
- autonomous unmanned systems
- experience replay
- neural networks
- obstacle avoidance control
- reinforcement learning
学术指纹
探究 '基于安全自适应强化学习的自主避障控制方法' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver