Fast and Scalable Global Convergence in Single-Optimum Decentralized Coordination Problems
Over the past few years, the scientific community has been studying the usefulness of evolutionary game theory to solve distributed control problems. In this paper we analyze a simple version of the Best Experienced Payoff (BEP) algorithm, a revision protocol recently proposed in the evolutionary ga...
| Autores: | , , |
|---|---|
| Tipo de documento: | artigo |
| Estado: | Versión aceptada para publicación |
| Data de publicação: | 2022 |
| País: | España |
| Recursos: | Universidad de Burgos (UBU) |
| Repositório: | Repositorio Institucional de la Universidad de Burgos (RIUBU) |
| OAI Identifier: | oai:riubu.ubu.es:10259/6776 |
| Acesso em linha: | http://hdl.handle.net/10259/6776 |
| Access Level: | Acceso aberto |
| Palavra-chave: | Best experienced payoff Decentralized algorithms Distributed control Evolutionary dynamics Evolutionary game theory Large population double limit Small noise limit Ingeniería Engineering |
| Resumo: | Over the past few years, the scientific community has been studying the usefulness of evolutionary game theory to solve distributed control problems. In this paper we analyze a simple version of the Best Experienced Payoff (BEP) algorithm, a revision protocol recently proposed in the evolutionary game theory literature. This revision protocol is simple, completely decentralized and has minimum information requirements. Here we prove that adding some noise to this protocol can lead to efficient results in single-optimum coordination problems in little time, even in large populations of agents. We also test the algorithm under a wide range of different conditions using computer simulation. In particular, we consider different numbers of agents and of strategies, and we analyze the robustness of the algorithm to different updating schemes (e.g. synchronous vs asynchronous) and to different types of interaction networks (e.g. ring, preferential attachment, small world and complete). In all cases, using the noisy version of BEP, the agents quickly approach a small neighborhood of the optimal state from every initial condition, and spend most of the time in that neighborhood. |
|---|