Abstract
In this article, we propose a cascade control framework to attenuate the residual vibration of the underactuated manipulator. The control framework is divided into two phases. In the first phase, a path generator trained by the reinforcement learning produces the leading signal for the tracking controller. In the second phase, the leading signal stabilizes the underactuated manipulator, and the adaptive proportional derivative controller is implemented to reduce the vibration. In the process, a novel path planning method is proposed to improve exploration efficiency, and a negative reward is introduced to avoid unsafe strategies and simulation instability. The effectiveness of the proposed control scheme is verified in the simulations of the double pendulum crane and the two-link flexible manipulator.
Get full access to this article
View all access options for this article.
References
Supplementary Material
Please find the following supplemental material available below.
For Open Access articles published under a Creative Commons License, all supplemental material carries the same license as the article it is associated with.
For non-Open Access articles published, all supplemental material carries a non-exclusive license, and permission requests for re-use of supplemental material or any part of supplemental material shall be sent directly to the copyright owner as specified in the copyright notice associated with the article.
