Cable SCARA Robot Controlled by a Neural Network Using Reinforcement LearningSource: Journal of Computational and Nonlinear Dynamics:;2023:;volume( 018 ):;issue: 010::page 104501-1DOI: 10.1115/1.4063222Publisher: The American Society of Mechanical Engineers (ASME)
Abstract: In this work, three reinforcement learning algorithms (Proximal Policy Optimization, Soft Actor-Critic, and Twin Delayed Deep Deterministic Policy Gradient) are employed to control a two link selective compliance articulated robot arm (SCARA) robot. This robot has three cables attached to its end-effector, which creates a triangular shaped workspace. Positioning the end-effector in the workspace is a relatively simple kinematic problem, but moving outside this region, although possible, requires a nonlinear dynamic model and a state-of-the-art controller. To solve this problem in a simple manner, reinforcement learning algorithms are used to find possible trajectories for three targets out of the workspace. Additionally, the SCARA mechanism offers two possible configurations for each end-effector position. The algorithm results are compared in terms of displacement error, velocity, and standard deviation among ten trajectories provided by the trained network. The results indicate the Proximal Policy Algorithm as the most consistent in the analyzed situations. Still, the Soft Actor-Critic presented better solutions, and Twin Delayed Deep Deterministic Policy Gradient provided interesting and more unusual trajectories.
|
Collections
Show full item record
| contributor author | Okabe, Eduardo | |
| contributor author | Paiva, Victor | |
| contributor author | Silva-Teixeira, Luis H. | |
| contributor author | Izuka, Jaime | |
| date accessioned | 2023-11-29T19:46:35Z | |
| date available | 2023-11-29T19:46:35Z | |
| date copyright | 9/1/2023 12:00:00 AM | |
| date issued | 9/1/2023 12:00:00 AM | |
| date issued | 2023-09-01 | |
| identifier issn | 1555-1415 | |
| identifier other | cnd_018_10_104501.pdf | |
| identifier uri | http://yetl.yabesh.ir/yetl1/handle/yetl/4295022 | |
| description abstract | In this work, three reinforcement learning algorithms (Proximal Policy Optimization, Soft Actor-Critic, and Twin Delayed Deep Deterministic Policy Gradient) are employed to control a two link selective compliance articulated robot arm (SCARA) robot. This robot has three cables attached to its end-effector, which creates a triangular shaped workspace. Positioning the end-effector in the workspace is a relatively simple kinematic problem, but moving outside this region, although possible, requires a nonlinear dynamic model and a state-of-the-art controller. To solve this problem in a simple manner, reinforcement learning algorithms are used to find possible trajectories for three targets out of the workspace. Additionally, the SCARA mechanism offers two possible configurations for each end-effector position. The algorithm results are compared in terms of displacement error, velocity, and standard deviation among ten trajectories provided by the trained network. The results indicate the Proximal Policy Algorithm as the most consistent in the analyzed situations. Still, the Soft Actor-Critic presented better solutions, and Twin Delayed Deep Deterministic Policy Gradient provided interesting and more unusual trajectories. | |
| publisher | The American Society of Mechanical Engineers (ASME) | |
| title | Cable SCARA Robot Controlled by a Neural Network Using Reinforcement Learning | |
| type | Journal Paper | |
| journal volume | 18 | |
| journal issue | 10 | |
| journal title | Journal of Computational and Nonlinear Dynamics | |
| identifier doi | 10.1115/1.4063222 | |
| journal fristpage | 104501-1 | |
| journal lastpage | 104501-7 | |
| page | 7 | |
| tree | Journal of Computational and Nonlinear Dynamics:;2023:;volume( 018 ):;issue: 010 | |
| contenttype | Fulltext |