MADECM: A Curiosity‐Augmented Evolutionary Algorithm for Multi‐Agent Policy Diversity Optimization.
Saved in:
| Title: | MADECM: A Curiosity‐Augmented Evolutionary Algorithm for Multi‐Agent Policy Diversity Optimization. |
|---|---|
| Authors: | Wu, Jianyang1 (AUTHOR), Fu, Yv2 (AUTHOR), Wang, Xinning2 (AUTHOR), Yang, Xin2 (AUTHOR) xinyang@dlut.edu.cn |
| Source: | Computer Animation & Virtual Worlds. May/Jun2026, Vol. 37 Issue 3, p1-12. 12p. |
| Subjects: | Evolutionary algorithms, Reinforcement learning, Intrinsic motivation |
| Abstract: | Multi‐agent reinforcement learning (MARL) often suffers from low sample efficiency and limited behavioral diversity, leading to policy homogenization, insufficient exploration, and reduced robustness. To address these challenges, we propose MADECM, a curiosity‐augmented evolutionary framework built upon MADDPG that integrates curiosity‐driven updates with evolutionary quality‐diversity optimization. MADECM employs random network distillation (RND) to estimate the novelty of each agent's local observations and uses the resulting novelty signal to dynamically allocate additional update frequencies, thereby emphasizing exploration‐relevant experience during training. In addition, MADECM combines population‐based diversification with a quality‐diversity (QD) archive through a staged optimization procedure, enabling the joint improvement of task return and policy diversity. We evaluate MADECM on the multi‐agent particle environment (MPE), including Spread and Reference, which capture cooperative and partially observable dynamics, and on google research football (GRF), which emphasizes long‐horizon sequential decision‐making. Results show that MADECM consistently outperforms strong MADDPG‐based baselines. The modular design of MADECM, consisting of RND‐based novelty estimation and staged QD optimization, further supports consistent generalization across these structurally distinct environments without task‐specific hyperparameter tuning. [ABSTRACT FROM AUTHOR] |
| Copyright of Computer Animation & Virtual Worlds is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) | |
| Database: | Engineering Source |
| FullText | Text: Availability: 0 |
|---|---|
| Header | DbId: egs DbLabel: Engineering Source An: 194920585 AccessLevel: 6 PubType: Academic Journal PubTypeId: academicJournal PreciseRelevancyScore: 0 |
| IllustrationInfo | |
| Items | – Name: Title Label: Title Group: Ti Data: MADECM: A Curiosity‐Augmented Evolutionary Algorithm for Multi‐Agent Policy Diversity Optimization. – Name: Author Label: Authors Group: Au Data: <searchLink fieldCode="AR" term="%22Wu%2C+Jianyang%22">Wu, Jianyang</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Fu%2C+Yv%22">Fu, Yv</searchLink><relatesTo>2</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Wang%2C+Xinning%22">Wang, Xinning</searchLink><relatesTo>2</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Yang%2C+Xin%22">Yang, Xin</searchLink><relatesTo>2</relatesTo> (AUTHOR)<i> xinyang@dlut.edu.cn</i> – Name: TitleSource Label: Source Group: Src Data: <searchLink fieldCode="JN" term="%22Computer+Animation+%26+Virtual+Worlds%22">Computer Animation & Virtual Worlds</searchLink>. May/Jun2026, Vol. 37 Issue 3, p1-12. 12p. – Name: Subject Label: Subjects Group: Su Data: <searchLink fieldCode="DE" term="%22Evolutionary+algorithms%22">Evolutionary algorithms</searchLink><br /><searchLink fieldCode="DE" term="%22Reinforcement+learning%22">Reinforcement learning</searchLink><br /><searchLink fieldCode="DE" term="%22Intrinsic+motivation%22">Intrinsic motivation</searchLink> – Name: Abstract Label: Abstract Group: Ab Data: Multi‐agent reinforcement learning (MARL) often suffers from low sample efficiency and limited behavioral diversity, leading to policy homogenization, insufficient exploration, and reduced robustness. To address these challenges, we propose MADECM, a curiosity‐augmented evolutionary framework built upon MADDPG that integrates curiosity‐driven updates with evolutionary quality‐diversity optimization. MADECM employs random network distillation (RND) to estimate the novelty of each agent's local observations and uses the resulting novelty signal to dynamically allocate additional update frequencies, thereby emphasizing exploration‐relevant experience during training. In addition, MADECM combines population‐based diversification with a quality‐diversity (QD) archive through a staged optimization procedure, enabling the joint improvement of task return and policy diversity. We evaluate MADECM on the multi‐agent particle environment (MPE), including Spread and Reference, which capture cooperative and partially observable dynamics, and on google research football (GRF), which emphasizes long‐horizon sequential decision‐making. Results show that MADECM consistently outperforms strong MADDPG‐based baselines. The modular design of MADECM, consisting of RND‐based novelty estimation and staged QD optimization, further supports consistent generalization across these structurally distinct environments without task‐specific hyperparameter tuning. [ABSTRACT FROM AUTHOR] – Name: AbstractSuppliedCopyright Label: Group: Ab Data: <i>Copyright of Computer Animation & Virtual Worlds is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.) |
| PLink | https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=194920585 |
| RecordInfo | BibRecord: BibEntity: Identifiers: – Type: doi Value: 10.1002/cav.70121 Languages: – Code: eng Text: English PhysicalDescription: Pagination: PageCount: 12 StartPage: 1 Subjects: – SubjectFull: Evolutionary algorithms Type: general – SubjectFull: Reinforcement learning Type: general – SubjectFull: Intrinsic motivation Type: general Titles: – TitleFull: MADECM: A Curiosity‐Augmented Evolutionary Algorithm for Multi‐Agent Policy Diversity Optimization. Type: main BibRelationships: HasContributorRelationships: – PersonEntity: Name: NameFull: Wu, Jianyang – PersonEntity: Name: NameFull: Fu, Yv – PersonEntity: Name: NameFull: Wang, Xinning – PersonEntity: Name: NameFull: Yang, Xin IsPartOfRelationships: – BibEntity: Dates: – D: 01 M: 05 Text: May/Jun2026 Type: published Y: 2026 Identifiers: – Type: issn-print Value: 15464261 Numbering: – Type: volume Value: 37 – Type: issue Value: 3 Titles: – TitleFull: Computer Animation & Virtual Worlds Type: main |
| ResultId | 1 |