Özet
Sequential Social Dilemmas are gaining attention in recent years. The current trends either focus on engineering incentive functions for modifying rewards to reach general welfare, or develop learning based approaches to modify the reward function by accounting for the impact of the incentive on policy updates. One of the most significant works in the learning based approach is LIO, which enables independent self-interested agents to incentivize each other by an additive incentive reward and demonstrates the method’s success in several sequential social dilemma environments. We investigate LIO’s performance under a variety of different setups in public goods game Cleanup in order to analyse its robustness against necessity of including inductive bias in incentive function, randomness in initial agent position with an option of asymmetric incentive potential, and assess its stability under frozen incentive functions after agents’ explorations are reset. We observe and demonstrate empirically that LIO is indeed sensitive to these settings and it is not reliable for obtaining good incentives that would let the system stay stable when it is static. We conclude with some research directions that would improve the robustness of the method and incentive learning research.
| Orijinal dil | İngilizce |
|---|---|
| Ana bilgisayar yayını başlığı | Computational Science and Computational Intelligence - 11th International Conference, CSCI 2024, Proceedings |
| Editörler | Hamid R. Arabnia, Leonidas Deligiannidis, Farzan Shenavarmasouleh, Soheyla Amirian, Farid Ghareh Mohammadi |
| Yayınlayan | Springer Science and Business Media Deutschland GmbH |
| Sayfalar | 116-125 |
| Sayfa sayısı | 10 |
| ISBN (Basılı) | 9783031995880 |
| DOI'lar | |
| Yayın durumu | Yayınlandı - 2025 |
| Etkinlik | 11th International Conference on Computational Science and Computational Intelligence, CSCI 2024 - Las Vegas, United States Süre: 11 Ara 2024 → 13 Ara 2024 |
Yayın serisi
| Adı | Communications in Computer and Information Science |
|---|---|
| Hacim | 2512 CCIS |
| ISSN (Basılı) | 1865-0929 |
| ISSN (Elektronik) | 1865-0937 |
???event.eventtypes.event.conference???
| ???event.eventtypes.event.conference??? | 11th International Conference on Computational Science and Computational Intelligence, CSCI 2024 |
|---|---|
| Ülke/Bölge | United States |
| Şehir | Las Vegas |
| Periyot | 11/12/24 → 13/12/24 |
Bibliyografik not
Publisher Copyright:© The Author(s), under exclusive license to Springer Nature Switzerland AG 2025.
Parmak izi
Empirical Robustness Analysis of Learning to Incentivize Other Self-interested Agents' araştırma başlıklarına git. Birlikte benzersiz bir parmak izi oluştururlar.Alıntı Yap
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver