Ali H. Sayed, Mert Kayaalp
Most works on multi-agent reinforcement learning focus on scenarios where the state of the environment is fully observable. In this work, we consider a cooperative policy evaluation task in which agents are not assumed to observe the environment state dire ...
IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC2023