Garuda - Garba Rujukan Digital

Article Per Year (5 Year)

p-Index From 2021 - 2026

0.23

P-Index

This Author published in this journals

All Journal Journal of Robotics and Control (JRC)

Neshat Elhami Fard

Concordia University

Author-ID : 2771375

Aerospace Engineering Computer Science & IT Control & Systems Engineering Electrical & Electronics Engineering Mechanical Engineering

Published : 1 Documents Claim Missing Document

Claim Missing Document

Articles

Title

Consensus of Multi-agent Reinforcement Learning Systems: The Effect of Immediate Rewards Neshat Elhami Fard; Rastko Selmic
Journal of Robotics and Control (JRC) Vol 3, No 2 (2022): March
Publisher : Universitas Muhammadiyah Yogyakarta

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.18196/jrc.v3i2.13082

This paper studies the consensus problem of a leaderless, homogeneous, multi-agent reinforcement learning (MARL) system using actor-critic algorithms with and without malicious agents. The goal of each agent is to reach the consensus position with the maximum cumulative reward. Although the reward function converges in both scenarios, in the absence of the malicious agent, the cumulative reward is higher than with the malicious agent present. We consider here various immediate reward functions. First, we study the immediate reward function based on Manhattan distance. In addition to proposing three different immediate reward functions based on Euclidean, $n$-norm, and Chebyshev distances, we have rigorously shown which method has a better performance based on a cumulative reward for each agent and the entire team of agents. Finally, we present a combination of various immediate reward functions that yields a higher cumulative reward for each agent and the team of agents. By increasing the agents’ cumulative reward using the combined immediate reward function, we have demonstrated that the cumulative team reward in the presence of a malicious agent is comparable with the cumulative team reward in the absence of the malicious agent. The claims have been proven theoretically, and the simulation confirms theoretical findings.

Co-Authors Rastko Selmic

Title Search

Found 1 Documents Search Journal : Journal of Robotics and Control (JRC)

Abstract

Title

Found 1 Documents
Search
Journal : Journal of Robotics and Control (JRC)