| Title: | Reinforcement learning for robot manipulation tasks in human-robot collaboration using the CQL/SAC algorithms |
|---|
| Authors: | ID Husaković, A. (Author) ID Banjanović-Mehmedović, Lejla (Author) ID Gurdić-Ribić, A. (Author) ID Prljača, Naser (Author) ID Karabegović, Isak (Author) |
| Files: | APEM20-1_005-017.pdf (1,75 MB) MD5: 352CB7FED41017CED5AF79BA67B02A7E
https://apem-journal.org/Archives/2025/VOL20-ISSUE01.html
|
|---|
| Language: | English |
|---|
| Work type: | Article |
|---|
| Typology: | 1.01 - Original Scientific Article |
|---|
| Organization: | FS - Faculty of Mechanical Engineering
|
|---|
| Abstract: | The integration of human-robot collaboration (HRC) into industrial and service environments demands efficient and adaptive robotic systems capable of executing diverse tasks, including pick-and-place operations. This paper investigates the application of Soft Actor-Critic (SAC) and Conservative Q-Learning (CQL)—two deep reinforcement learning (DRL) algorithms—for the learning and optimization of pick-and-place actions within HRC scenarios. By leveraging SAC’s capability to balance exploration and exploitation, the robot autonomously learns to perform pick-and-place tasks while adapting to dynamic environments and human interactions. Moreover, the integration of CQL ensures more stable learning by mitigating Q-value overestimation, which proves particularly advantageous in offline and suboptimal data scenarios. The combined use of CQL and SAC enhances policy robustness, facilitating safer and more efficient decision-making in continually evolving environments. The proposed framework combines simulation-based training with transfer learning techniques, enabling seamless deployment in real-world environments. The critical challenge of trajectory completion is addressed through a meticulously designed reward function that promotes efficiency, precision, and safety. Experimental validation demonstrates a 100 % success rate in simulation and an 80 % success rate on real hardware, confirming the practical viability of the proposed model. This work underscores the pivotal role of DRL in enhancing the functionality of collaborative robotic systems, illustrating its applicability across a range of industrial environments. |
|---|
| Keywords: | human-robot collaboration, robot learning, deep reinforcement learning, soft actor-critic algorithm, Conservative Q-learning, robot manipulation tasks |
|---|
| Publication status: | Published |
|---|
| Publication version: | Version of Record |
|---|
| Submitted for review: | 10.02.2025 |
|---|
| Article acceptance date: | 07.03.2025 |
|---|
| Publication date: | 29.04.2025 |
|---|
| Publisher: | Fakulteta za strojništvo |
|---|
| Year of publishing: | 2025 |
|---|
| Number of pages: | str. 5-17 |
|---|
| Numbering: | Vol. 20, no. 1 |
|---|
| PID: | 20.500.12556/DKUM-96524  |
|---|
| UDC: | 007.52 |
|---|
| ISSN on article: | 1854-6250 |
|---|
| COBISS.SI-ID: | 264960259  |
|---|
| DOI: | 10.14743/apem2025.1.523  |
|---|
| Publication date in DKUM: | 16.01.2026 |
|---|
| Views: | 164 |
|---|
| Downloads: | 3 |
|---|
| Metadata: |  |
|---|
| Categories: | Misc.
|
|---|
|
:
|
Copy citation |
|---|
| | | | Average score: | (0 votes) |
|---|
| Your score: | Voting is allowed only for logged in users. |
|---|
| Share: |  |
|---|
Hover the mouse pointer over a document title to show the abstract or click
on the title to get all document metadata. |