Multi-agent learning via gradient ascent activity-based credit assignment

Oussama Sabri; Luc Lehéricy; Alexandre Muzy

doi:10.1038/s41598-023-42448-9

Multi-agent learning via gradient ascent activity-based credit assignment

Sci Rep. 2023 Sep 14;13(1):15256. doi: 10.1038/s41598-023-42448-9.

Authors

Oussama Sabri^#^{1

2}, Luc Lehéricy^#^{3

4}, Alexandre Muzy^#^{5

4}

Affiliations

¹ CNRS, I3S, Sophia Antipolis, France. ou.sabri@outlook.com.
² Université Côte d'Azur, Nice, France. ou.sabri@outlook.com.
³ CNRS, JAD, Nice, France.
⁴ Université Côte d'Azur, Nice, France.
⁵ CNRS, I3S, Sophia Antipolis, France.

^# Contributed equally.

Abstract

We consider the situation in which cooperating agents learn to achieve a common goal based solely on a global return that results from all agents' behavior. The method proposed is based on taking into account the agents' activity, which can be any additional information to help solving multi-agent decentralized learning problems. We propose a gradient ascent algorithm and assess its performance on synthetic data.

Grants and funding

ANR-15-IDEX-01/Agence Nationale de la Recherche