<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="research-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Robot. AI</journal-id>
<journal-title>Frontiers in Robotics and AI</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Robot. AI</abbrev-journal-title>
<issn pub-type="epub">2296-9144</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1470886</article-id>
<article-id pub-id-type="doi">10.3389/frobt.2024.1470886</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Robotics and AI</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Global progress in competitive co-evolution: a systematic comparison of alternative methods</article-title>
<alt-title alt-title-type="left-running-head">Nolfi and Pagliuca</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/frobt.2024.1470886">10.3389/frobt.2024.1470886</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Nolfi</surname>
<given-names>Stefano</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/666/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/conceptualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/data-curation/"/>
<role content-type="https://credit.niso.org/contributor-roles/formal-analysis/"/>
<role content-type="https://credit.niso.org/contributor-roles/funding-acquisition/"/>
<role content-type="https://credit.niso.org/contributor-roles/investigation/"/>
<role content-type="https://credit.niso.org/contributor-roles/methodology/"/>
<role content-type="https://credit.niso.org/contributor-roles/software/"/>
<role content-type="https://credit.niso.org/contributor-roles/supervision/"/>
<role content-type="https://credit.niso.org/contributor-roles/validation/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-original-draft/"/>
<role content-type="https://credit.niso.org/contributor-roles/Writing - review &#x26; editing/"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Pagliuca</surname>
<given-names>Paolo</given-names>
</name>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/1028490/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/data-curation/"/>
<role content-type="https://credit.niso.org/contributor-roles/formal-analysis/"/>
<role content-type="https://credit.niso.org/contributor-roles/investigation/"/>
<role content-type="https://credit.niso.org/contributor-roles/methodology/"/>
<role content-type="https://credit.niso.org/contributor-roles/software/"/>
<role content-type="https://credit.niso.org/contributor-roles/validation/"/>
<role content-type="https://credit.niso.org/contributor-roles/Writing - review &#x26; editing/"/>
</contrib>
</contrib-group>
<aff>
<institution>Laboratory of Autonomous Robotics and Artificial Life (LARAL)</institution>, <institution>Institute of Cognitive Sciences and Technologies (ISTC)</institution>, <institution>National Research Council (CNR)</institution>, <addr-line>Rome</addr-line>, <country>Italy</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/147408/overview">Giovanni Iacca</ext-link>, University of Trento, Italy</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/134035/overview">Jos&#xe9; Antonio Becerra Permuy</ext-link>, University of A Coru&#xf1;a, Spain</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1166692/overview">Eric Medvet</ext-link>, University of Trieste, Italy</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Paolo Pagliuca, <email>paolo.pagliuca@istc.cnr.it</email>
</corresp>
</author-notes>
<pub-date pub-type="epub">
<day>21</day>
<month>01</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2024</year>
</pub-date>
<volume>11</volume>
<elocation-id>1470886</elocation-id>
<history>
<date date-type="received">
<day>26</day>
<month>07</month>
<year>2024</year>
</date>
<date date-type="accepted">
<day>23</day>
<month>12</month>
<year>2024</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2025 Nolfi and Pagliuca.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Nolfi and Pagliuca</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>The usage of broad sets of training data is paramount to evolve adaptive agents. In this respect, competitive co-evolution is a widespread technique in which the coexistence of different learning agents fosters adaptation, which in turn makes agents experience continuously varying environmental conditions. However, a major pitfall is related to the emergence of endless limit cycles where agents discover, forget and rediscover similar strategies during evolution. In this work, we investigate the use of competitive co-evolution for synthesizing progressively better solutions. Specifically, we introduce a set of methods to measure historical and global progress. We discuss the factors that facilitate genuine progress. Finally, we compare the efficacy of four qualitatively different algorithms, including two newly introduced methods. The selected algorithms promote genuine progress by creating an archive of opponents used to evaluate evolving individuals, generating archives that include high-performing and well-differentiated opponents, identifying and discarding variations that lead to local progress only (i.e., progress against the opponents experienced and retrogressing against others). The results obtained in a predator-prey scenario, commonly used to study competitive evolution, demonstrate that all the considered methods lead to global progress in the long term. However, the rate of progress and the ratio of progress versus retrogressions vary significantly among algorithms. In particular, our outcomes indicate that the Generalist method introduced in this work outperforms the other three considered methods and represents the only algorithm capable of producing global progress during evolution.</p>
</abstract>
<kwd-group>
<kwd>competitive co-evolution</kwd>
<kwd>evolutionary robotics</kwd>
<kwd>local historical and global progress</kwd>
<kwd>open-ended evolution</kwd>
<kwd>predator-prey robots</kwd>
</kwd-group>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Robot Learning and Evolution</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="s1">
<title>1 Introduction</title>
<p>Recent advances in machine learning have demonstrated the importance of using large corpora of training data. In the context of embodied and situated agents, this implies placing the agents in complex and diversified environments. However, manually designing environments of this kind is both challenging and costly. A convenient alternative is constituted by multi-agent scenarios where adaptive agents are situated in environments with other adaptive agents with conflicting goals&#x2014;a method known as competitive co-evolution (<xref ref-type="bibr" rid="B43">Rosin and Belew, 1995</xref>) or self-play (<xref ref-type="bibr" rid="B1">Bansal et al., 2017</xref>). In these settings, learning agents are exposed to continuously varying environmental conditions due to the behavioral changes of other adaptive agents. In other words, these settings allow for the automatic generation of a large corpus of training data.</p>
<p>Competitive settings also offer other important advantages. They can spontaneously produce convenient learning curricula (<xref ref-type="bibr" rid="B44">Rosin and Belew, 1997</xref>; <xref ref-type="bibr" rid="B52">Wang et al., 2021</xref>) in which the complexity of the training conditions progressively increases while the skills of the agents&#x2014;and consequently their ability to master complex conditions&#x2014;also improve. Finally, such settings may naturally generate a form of adversarial learning (<xref ref-type="bibr" rid="B34">Lowd and Meek, 2005</xref>), where the training data are shaped to challenge the weaknesses of the adaptive agents.</p>
<p>Unfortunately, competitive co-evolution does not necessarily result in a progressive complexification of agents&#x2019; skills or environmental conditions. As highlighted by Dawkins and Krebs in the context of natural co-evolution (<xref ref-type="bibr" rid="B11">Dawkins and Krebs, 1979</xref>), the evolutionary process can lead to four distinct long-term dynamics: (1) extinction: one side may drive the other to extinction, (2) definable optimum: one side might reach a definable optimum, preventing the other side from reaching its own optimum, (3) mutual local optimum: both sides may reach a mutual local optimum, and (4) endless limit cycle: the race may persist in a theoretically endless limit cycle, in which similar strategies are abandoned and rediscovered over and over again.</p>
<p>Regrettably, pioneering attempts to evolve competing robots consistently yield the last undesirable outcome described by Dawkins and Krebs. Initially, there is true progress, but subsequently agents modify their strategies, resulting in apparent progress. In other words, they improve against the strategies exhibited by their current opponents while retrogressing against other strategies that are later adopted by opponents (<xref ref-type="bibr" rid="B35">Miconi, 2008</xref>). Consequently, limit cycle dynamics emerge, where the same strategies are abandoned and rediscovered repeatedly (<xref ref-type="bibr" rid="B49">Sinervo and Lively, 1996</xref>; <xref ref-type="bibr" rid="B36">Miconi, 2009</xref>; <xref ref-type="bibr" rid="B38">Nolfi, 2012</xref>).</p>
<p>The generation of genuine progress requires the usage of special algorithms that: (i) expose the adapting agents to both current and ancient opponents (<xref ref-type="bibr" rid="B44">Rosin and Belew, 1997</xref>), (ii) expose the adaptive agents to a diverse range of opponents (<xref ref-type="bibr" rid="B13">De Jong, 2005</xref>; <xref ref-type="bibr" rid="B48">Simione and Nolfi, 2021</xref>), and (iii) identify and retain variations that lead to true progress only (<xref ref-type="bibr" rid="B48">Simione and Nolfi, 2021</xref>). Furthermore, analyzing competitive settings necessitates the formulation of suitable measures to differentiate between true and apparent progress and to evaluate the efficacy of the obtained solutions.</p>
<p>In this article, we present a systematic comparison of alternative competitive co-evolutionary algorithms, including novel variations of state-of-the-art algorithms. We describe the measures that can be used to analyze the co-evolutionary process, discriminate between apparent and true progress, and compare the efficacy of alternative algorithms. Finally, we analyze whether the proposed methods manage to avoid limit-cycle dynamics.</p>
</sec>
<sec sec-type="materials|methods" id="s2">
<title>2 Materials and methods</title>
<sec id="s2-1">
<title>2.1 Measuring progress</title>
<p>In evolutionary experiments in which the evolving individuals are situated in solitary environments, the fitness measured during individuals&#x2019; evaluation provides a direct and absolute measure of performance. Fitness evaluation typically includes a stochastic factor due to random variations of the initial state of the robot and of the environment. However, these variations are not adversarial in nature. Consequently, their impact can be usually be mitigated by evolving solutions that are robust to these variations and by averaging the fitness obtained during multiple evaluation episodes (<xref ref-type="bibr" rid="B26">Jakobi et al., 1995</xref>; <xref ref-type="bibr" rid="B6">Carvalho and Nolfi, 2023</xref>; <xref ref-type="bibr" rid="B41">Pagliuca and Nolfi, 2019</xref>).</p>
<p>In competitive social settings, instead, the fitness obtained during individuals&#x2019; evaluation crucially depends on the opponent(s) situated in the same environment. Consequently, the method used to select opponents has a pivotal effect on the course of the co-evolutionary process.</p>
<p>The dependency of the fitness measure on the characteristics of the opponents also affects: (1) the identification of the best solution generated during an evolutionary process, (2) the estimation of the overall effectiveness of a solution, and (3) the comparison of the efficacy of alternative experimental conditions.</p>
<p>The former two problems can be tackled by identifying a specific representative set of opponents known as &#x201c;champions,&#x201d; which typically include the best opponents produced during independent evolutionary experiments. The third issue can be resolved through a technique called &#x201c;cross-test.&#x201d; In a cross-test, the best solutions obtained from N independent evolutionary experiments conducted under one experimental condition are evaluated against the best opponents obtained from N different independent evolutionary experiments.</p>
<p>The dependence of the fitness measure on the opponent characteristics impacts the method used to measure evolutionary progress. In non-competitive settings, progress and retrogression can be straightforwardly measured by computing the variation of fitness across generations. Instead, in competitive settings, measuring progress becomes more challenging.</p>
<p>As highlighted by <xref ref-type="bibr" rid="B36">Miconi (2009)</xref>, we need to differentiate between three types of progress: (i) local progress, i.e., progress against current opponents, (ii) historical progress, i.e., progress against opponents of previous generations, and (iii) global progress, i.e., progress against all possible opponents. Local progress can be measured by evaluating agents against opponents from recent preceding generations. Historical progress can be assessed by evaluating agents against opponents from previous generations. These data can be effectively visualized using the &#x201c;Current Individual against Ancestral Opponents&#x201d; (CIAO) plots introduced by (<xref ref-type="bibr" rid="B9">Cliff and Miller 1995</xref>; <xref ref-type="bibr" rid="B10">Cliff and Miller, 2006</xref>). Finally, global progress can be estimated by evaluating agents against opponents generated in independent evolutionary experiments&#x2014;opponents that differ from those encountered during the evolutionary process. Additionally, an indication of global progress can be obtained by evaluating agents against opponents from future generations. The data obtained by post-evaluating agents against opponents of previous and future generations can be conveniently visualized using the master tournament plots introduced by <xref ref-type="bibr" rid="B39">Nolfi and Floreano (1998)</xref>.</p>
</sec>
<sec id="s2-2">
<title>2.2 Competitive evolutionary algorithms</title>
<p>In this section we will review the most interesting co-evolutionary algorithms described in the literature and the methods that we will compare in our experiments. These algorithms can be applied to the evolution of two species that reciprocally affect each other.</p>
<p>We focus our analysis on evolutionary algorithms attempting to maximize the expected utility, i.e., the expected fitness against a randomly selected opponent or the average fitness against all possible opponents. Other researchers have explored the use of competitive evolution for synthesizing Nash equilibrium solutions (<xref ref-type="bibr" rid="B14">Ficici and Pollack, 2003</xref>; <xref ref-type="bibr" rid="B53">Wiegand et al., 2002</xref>) and Pareto-optimal solutions (<xref ref-type="bibr" rid="B12">De Jong, 2004</xref>).</p>
<p>As previously mentioned, achieving true progress necessitates the utilization of specialized algorithms that: (i) expose adapting agents to both current and ancient opponents, (ii) evaluate adaptive agents against a diverse set of opponents, or (iii) identify and retain variations that lead to genuine progress.</p>
<p>The first method we consider is the Archive algorithm, as introduced by <xref ref-type="bibr" rid="B44">Rosin and Belew (1997)</xref>, see also (<xref ref-type="bibr" rid="B51">Stolfi et al., 2021</xref>). The pseudocode of the Archive algorithm is provided in <xref ref-type="statement" rid="Algorithm_1">Algorithm 1</xref>. In this algorithm, a copy of the best individual from each generation is stored in a &#x201c;hall-of-fame&#x201d; archive. Opponents are then randomly selected from this archive using a uniform distribution. Evaluating evolving agents against opponents from previous generations clearly promotes historical progress. While the production of global progress is not guaranteed, it can be expected as a form of generalization. Indeed, the need to defeat an increasing number of ancient opponents should encourage the development of strategies that generalize to opponents not yet encountered.</p>
<p>
<statement content-type="algorithm" id="Algorithm_1">
<label>Algorithm 1</label>
<p>Archive algorithm.<list list-type="simple">
<list-item>
<p>1.&#x2003;n_parents &#x3d; 1, n_offspring &#x3d; 40, mutation_range &#x3d; 0.02, learning_rate &#x3d; 0.01, max_total_steps &#x3d; 75 &#x2022; 10<sup>9</sup>
</p>
</list-item>
<list-item>
<p>2.&#x2003;<bold>initialize</bold>the genotype of the parents (predator and prey) randomly</p>
</list-item>
<list-item>
<p>3.&#x2003;<bold>initialize</bold>the archives (predator and prey) with 10 genotypes generated randomly</p>
</list-item>
<list-item>
<p>4.&#x2003;<bold>while</bold>tot_steps &#x3c; max_total_steps</p>
</list-item>
<list-item>
<p>5.&#x2003;&#x2003;&#x2003;selected_opponents &#x3d; <bold>select</bold>10 opponents&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;randomly from the archive</p>
</list-item>
<list-item>
<p>6.&#x2003;&#x2003;&#x2003;<bold>generate</bold>offspring</p>
</list-item>
<list-item>
<p>7.&#x2003;&#x2003;&#x2003;fitness <bold>&#x3d; evaluate</bold>offspring against&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;selected_opponents</p>
</list-item>
<list-item>
<p>8.&#x2003;&#x2003;&#x2003;<bold>compute</bold>gradient</p>
</list-item>
<list-item>
<p>9.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>update</bold>parent</p>
</list-item>
<list-item>
<p>10.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>append</bold>the fittest offspring to the&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;archive</p>
</list-item>
<list-item>
<p>11.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;current_generation &#x2b; &#x3d; 1</p>
</list-item>
</list>
</p>
</statement>
</p>
<p>The second method that we will consider is the Maxsolve&#x2a; algorithm (see <xref ref-type="statement" rid="Algorithm_2">Algorithm 2</xref>), a variation of the original method introduced by <xref ref-type="bibr" rid="B13">De Jong (2005)</xref>, see also (<xref ref-type="bibr" rid="B47">Samothrakis et al., 2013</xref>; <xref ref-type="bibr" rid="B33">Liskowski and Krawiec, 2016</xref>). The Maxsolve algorithm operates by using an archive containing a predetermined maximum number of opponents. The size of the archive is kept bounded by removing dominated and redundant opponents from it. The domination criterion works for transitive problems, i.e., in problems where agents A outperforming agents B, which in turn outperform agents C, necessarily outperform agents C. To allow the algorithm to operate with non-transitive problems, like the predator and prey task considered in this paper, we implemented the Maxsolve&#x2a; algorithm, a variation of the original Maxsolve algorithm, which retains in the archive the individuals achieving the highest performance, on average, against the opponents of the 10 preceding phases, where each phase corresponds to <inline-formula id="inf1">
<mml:math id="m1">
<mml:mrow>
<mml:mfrac>
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mn>10</mml:mn>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</inline-formula> of the total number of generations. The method thus attempts to automatically select high-quality champions that can promote the discovery of high-quality solutions and minimize the time spent evaluating agents against poor opponents (for a related approach, see <xref ref-type="bibr" rid="B2">Bari et al., 2018</xref>). Clearly, the size of the archive plays an important role in this method. Indeed, the smaller the size of the archive is, the smaller the training data is. On the other hand, the smaller the size of the archive is, the higher the minimization of the time spent against poor opponents is. By systematically varying the size of the archive (data not shown), we observed that the best results are obtained by using relatively large archives, i.e., archives containing up to 25,000 opponents.</p>
<p>
<statement content-type="algorithm" id="Algorithm_2">
<label>Algorithm 2</label>
<p>Maxsolve&#x2a; algorithm.<list list-type="simple">
<list-item>
<p>1.&#x2003;n_parents &#x3d; 1, n_offspring &#x3d; 40, max_archive_size &#x3d; 25000, mutation_range &#x3d; 0.02, learning_rate &#x3d; 0.01, max_total_steps &#x3d; 75 &#x2022; 10<sup>9</sup>
</p>
</list-item>
<list-item>
<p>2.&#x2003;<bold>initialize</bold>the genotype of the parents (predator and prey) randomly</p>
</list-item>
<list-item>
<p>3.&#x2003;<bold>initialize</bold>the archives (predator and prey) with 10 genotypes generated randomly</p>
</list-item>
<list-item>
<p>4.&#x2003;<bold>while</bold>tot_steps &#x3c; max_total_steps</p>
</list-item>
<list-item>
<p>5.&#x2003;&#x2003;&#x2003;selected_opponents &#x3d; <bold>select</bold>10 opponents&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;randomly from the archive</p>
</list-item>
<list-item>
<p>6.&#x2003;&#x2003;&#x2003;<bold>generate</bold>offspring</p>
</list-item>
<list-item>
<p>7.&#x2003;&#x2003;&#x2003;fitness <bold>&#x3d; evaluate</bold>offspring against&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;selected_opponents</p>
</list-item>
<list-item>
<p>8.&#x2003;&#x2003;&#x2003;store fitness_data[agent][opponent]</p>
</list-item>
<list-item>
<p>9.&#x2003;&#x2003;&#x2003;<bold>compute</bold>gradient</p>
</list-item>
<list-item>
<p>10.&#x2003;&#x2003;&#x2003;<bold>update</bold>parent</p>
</list-item>
<list-item>
<p>11.&#x2003;&#x2003;&#x2003;<bold>append</bold>the fittest offspring to the archive</p>
</list-item>
<list-item>
<p>12.&#x2003;&#x2003;&#x2003;current_generation &#x2b; &#x3d; 1</p>
</list-item>
<list-item>
<p>13.&#x2003;&#x2003;&#x2003;<bold>if</bold>size(archive) &#x3e; 25000</p>
</list-item>
<list-item>
<p>14.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>remove</bold>a dominated agent from the archive&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;(see text for details)</p>
</list-item>
</list>
</p>
</statement>
</p>
<p>The third method is the Archive&#x2a; algorithm, a novel approach introduced in this paper as a variation of the original Archive algorithm described earlier. The Archive&#x2a; algorithm operates by evolving N independent populations, each contributing to the generation of a single archive. These evolving populations are then evaluated against opponents selected from the shared archive. The pseudocode of the method is reported in <xref ref-type="statement" rid="Algorithm_3">Algorithm 3</xref>.</p>
<p>
<statement content-type="algorithm" id="Algorithm_3">
<label>Algorithm 3</label>
<p>Archive&#x2a; algorithm.<list list-type="simple">
<list-item>
<p>1.&#x2003;n_parents &#x3d; 10, n_offspring &#x3d; 40, mutation_range &#x3d; 0.02, learning_rate &#x3d; 0.01, max_total_steps &#x3d; 75 &#x2022; 10<sup>9</sup>
</p>
</list-item>
<list-item>
<p>2.&#x2003;<bold>initialize</bold>the genotype of the parents (predator and prey) randomly</p>
</list-item>
<list-item>
<p>3.&#x2003;<bold>initialize</bold>the archives (predator and prey) with 10 genotypes generated randomly</p>
</list-item>
<list-item>
<p>4.&#x2003;<bold>while</bold>tot_steps &#x3c; max_total_steps</p>
</list-item>
<list-item>
<p>5.&#x2003;&#x2003;&#x2003;<bold>for</bold>p in range (n_parents)</p>
</list-item>
<list-item>
<p>6.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;selected_opponents &#x3d; <bold>select</bold>10 opponents&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;randomly from the archive</p>
</list-item>
<list-item>
<p>7.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>generate</bold>offspring [p]</p>
</list-item>
<list-item>
<p>8.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;fitness [p] <bold>&#x3d; evaluate</bold>offspring [p]&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;against selected_opponents</p>
</list-item>
<list-item>
<p>9.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>compute</bold>gradient [p]</p>
</list-item>
<list-item>
<p>10.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>update</bold>parent [p]</p>
</list-item>
<list-item>
<p>11.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;<bold>append</bold>the fittest offspring [p] to&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;the archive</p>
</list-item>
<list-item>
<p>12.&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;current_generation &#x2b; &#x3d; 1</p>
</list-item>
</list>
</p>
</statement>
</p>
<p>The rationale behind this method lies in the use of multiple populations, which enhances the diversification of opponents included in the archive. Specifically, the Archive&#x2a; method automatically generates multiple families of diversified champions, representing alternative challenges for the evolving agent. The Archive&#x2a; algorithm shares similarities with the population-based reinforcement learning method proposed by <xref ref-type="bibr" rid="B24">Jaderberg et al. (2018)</xref>.</p>
<p>Finally, the fourth method we propose is the Generalist algorithm, introduced by <xref ref-type="bibr" rid="B48">Simione and Nolfi (2021)</xref> (see <xref ref-type="statement" rid="Algorithm_4">Algorithm 4</xref>). In this approach, agents are evaluated against a subset of opponents, while the remaining opponents serve to discriminate between agents who retained variations leading to global progress and those who retained variations leading to local progress. This information guides the preservation or discarding of individuals. The subset of opponents used for agent evaluation is randomly selected at regular intervals (every N generations) to maximize the functional diversity of the subset. Unlike the previous algorithms, the Generalist method does not rely on archives.</p>
<p>
<statement content-type="algorithm" id="Algorithm_4">
<label>Algorithm 4</label>
<p>Generalist algorithm.<list list-type="simple">
<list-item>
<p>1.&#x2003;n&#x5f;parents &#x3d; 80, n&#x5f;offspring &#x3d; 40, mutation&#x5f;range &#x3d; 0.02, learning&#x5f;rate &#x3d; 0.01, max&#x5f;total&#x5f;steps &#x3d; 75 &#x2022; 10<sup>9</sup>
</p>
</list-item>
<list-item>
<p>2.&#x2003;<bold>initialize</bold>the genotype of the parents (predator and prey) randomly</p>
</list-item>
<list-item>
<p>3.&#x2003;<bold>while</bold>tot&#x5f;steps &#x3c; max&#x5f;total&#x5f;steps</p>
</list-item>
<list-item>
<p>4.&#x2003;&#x2003;<bold>every</bold>20 generations</p>
</list-item>
<list-item>
<p>5.&#x2003;&#x2003;&#x2003;performance [n&#x5f;parents] [n&#x5f;parents] &#x3d; <bold>evaluate</bold>(all&#x5f;parents, all&#x5f;opponents)</p>
</list-item>
<list-item>
<p>6.&#x2003;&#x2003;&#x2003;selected&#x5f;parents &#x3d; <bold>select</bold>10 parents randomly</p>
</list-item>
<list-item>
<p>7.&#x2003;&#x2003;&#x2003;selected&#x5f;opponents &#x3d; <bold>select</bold>the 10 opponents with the highest performance against the 10 selected parents</p>
</list-item>
<list-item>
<p>8.&#x2003;&#x2003;&#x2003;candidate&#x5f;parents [] <bold>create-a-copy-of</bold>selected&#x5f;parents []</p>
</list-item>
<list-item>
<p>9.&#x2003;&#x2003;<bold>for</bold>p in range (candidate&#x5f;parents)</p>
</list-item>
<list-item>
<p>10.&#x2003;&#x2003;&#x2003;<bold>generate</bold>offspring [p]</p>
</list-item>
<list-item>
<p>11.&#x2003;&#x2003;&#x2003;fitness [p] <bold>&#x3d; evaluate</bold>offspring [p] against selected selected&#x5f;opponents</p>
</list-item>
<list-item>
<p>12.&#x2003;&#x2003;&#x2003;<bold>compute</bold>gradient [p]</p>
</list-item>
<list-item>
<p>13.&#x2003;&#x2003;&#x2003;<bold>update</bold>candidate&#x5f;parent [p]</p>
</list-item>
<list-item>
<p>14.&#x2003;&#x2003;&#x2003;<bold>every</bold>20 generations</p>
</list-item>
<list-item>
<p>15.&#x2003;&#x2003;&#x2003;performance [n&#x5f;selected&#x5f;parents] [n&#x5f;parents] &#x3d; <bold>evaluate</bold>(selected&#x5f;parents, all&#x5f;opponents)</p>
</list-item>
<list-item>
<p>16.&#x2003;&#x2003;&#x2003;<bold>replace</bold>the worst parents with the best candidate&#x5f;parents which outperform them.</p>
</list-item>
<list-item>
<p>17.&#x2003;&#x2003;&#x2003;current&#x5f;generation &#x2b; &#x3d; 1</p>
</list-item>
</list>
</p>
</statement>
</p>
<p>Another potential approach involves using randomly generated opponents (<xref ref-type="bibr" rid="B8">Chong et al., 2009</xref>; <xref ref-type="bibr" rid="B7">2012</xref>; <xref ref-type="bibr" rid="B28">Ja&#x15b;kowski et al., 2013</xref>). The primary advantage of this technique lies in the direct promotion of global progress because agents are consistently evaluated against new opponents. However, there is a significant drawback: the efficacy of these opponents does not improve over generations. Consequently, this method does not allow for the development of agents capable of defeating strong opponents. In fact, the performance obtained using this approach by <xref ref-type="bibr" rid="B47">Samothrakis et al. (2013)</xref> were considerably lower than the performance achieved with a variation of the Maxsolve algorithm described earlier.</p>
<p>The methods described above are meta-algorithms that should be combined with an evolutionary algorithm to determine how populations of individuals vary across generations. In previous studies, standard evolutionary algorithms or evolutionary strategies were employed. Instead, in this work, we utilize the OpenAI-ES algorithm (<xref ref-type="bibr" rid="B46">Salimans et al., 2017</xref>), which represents a &#x201c;modern&#x201d; evolutionary strategy (<xref ref-type="bibr" rid="B40">Pagliuca et al., 2020</xref>). The OpenAI-ES algorithm leverages the matrix of variations introduced within the population and the fitness values obtained by corresponding individuals to estimate the gradient of fitness. It then guides the population&#x2019;s movement in the direction of this gradient using a stochastic optimizer (<xref ref-type="bibr" rid="B30">Kingma and Ba, 2014</xref>). Importantly, this algorithm is well-suited for non-stationary environments. This is because the momentum vectors also guide the population in the direction of previously estimated gradients, thereby enhancing the possibility of generating solutions effective against opponents encountered in previous generations.</p>
<p>Below we include the pseudo-code of the algorithms. In all cases the connection weights of the controllers of the robots were evolved by using the OpenAI-ES algorithm (<xref ref-type="bibr" rid="B46">Salimans et al., 2017</xref>) and by using the hyperparameters indicated in the reference. More specifically, observation vectors were normalized by using virtual batch normalization (<xref ref-type="bibr" rid="B45">Salimans et al., 2016</xref>; <xref ref-type="bibr" rid="B46">Salimans et al., 2017</xref>), the connections weights were normalized by using weight decay, the distribution of the perturbations of parameters was set to 0.02, and the step size of the Adam optimizer was set to 0.01. The fitness gradient was estimated by generating 40 offspring, i.e., 40 perturbed versions of the parent. The fitness of the evolving individuals corresponds to the average fitness obtained during 10 episodes in which they are evaluated against 10 different opponents in 10 corresponding evaluation episodes. The fitness and the perturbation vectors are used to compute the gradient of the expected fitness, which is used to update the parameters of the parent through the Adam (<xref ref-type="bibr" rid="B30">Kingma, and Ba, 2014</xref>) stochastic optimizer. The predator and prey robots evolved in parallel. The evolutionary process is continued until the total number of evaluation steps performed exceeds 75 &#x2022; 10<sup>9</sup>.</p>
<p>In the case of the Archive and Maxsolve&#x2a; algorithms, the parameters of the evolving robots were generated from a single parent (a parent for the predator and a parent for the prey robots). In the case of the Archive&#x2a; algorithm, they are generated from 10 parents. The generalist algorithm, instead, uses a population of 80 parents, evolves 10 candidate parents generated by creating a copy of 10 parents selected randomly every 20 generations, and replace the worst parents with the candidate parents outperforming them. The performance of parents and of candidate parents are computed by evaluating the parents against the full set of 80 opponents, i.e., by evaluating the candidate parents also against opponents not encountered in the preceding generations. The steps performed to evaluate performance contribute to increase the total number of steps performed.</p>
<p>The Archive and Archive&#x2a; algorithms preserve the best robots of each generation in archives that keep increasing in size during the evolutionary process. The Maxsolve&#x2a; algorithm uses archives that can grow up to a maximum size only. This is realized by: (i) storing in a fitness_data [X][Y] matrix the average fitness obtained by agents of generation X against opponents of generation Y, (ii) computing the average fitness obtained by agents contained in the archive against opponents of 10 subsequent evolutionary phases, and (iii) eliminating one dominated agent selected randomly from the archive in each generation. An agent is dominated when the average fitness obtained against opponents of different phases is consistently equal or lower than the average fitness obtained by another agent included in the archive.</p>
</sec>
<sec id="s2-3">
<title>2.3 The predator and prey problem</title>
<p>We chose to compare alternative algorithms using a predator and prey problem because it represents a challenging scenario (<xref ref-type="bibr" rid="B37">Miller and Cliff, 1994</xref>). Additionally, this problem is widely used for studying competitive evolutionary algorithms (<xref ref-type="bibr" rid="B37">Miller and Cliff, 1994</xref>; <xref ref-type="bibr" rid="B15">Floreano and Nolfi, 1997a</xref>; <xref ref-type="bibr" rid="B16">Floreano and Nolfi, 1997b</xref>; <xref ref-type="bibr" rid="B17">Floreano et al., 1998</xref>; <xref ref-type="bibr" rid="B39">Nolfi and Floreano, 1998</xref>; <xref ref-type="bibr" rid="B50">Stanley and Miikkulainen, 2002</xref>; <xref ref-type="bibr" rid="B5">Buason and Ziemke, 2003</xref>; <xref ref-type="bibr" rid="B4">Buason et al., 2005</xref>; <xref ref-type="bibr" rid="B25">Jain et al., 2012</xref>; <xref ref-type="bibr" rid="B42">Palmer and Chou, 2012</xref>; <xref ref-type="bibr" rid="B23">Ito et al., 2013</xref>; <xref ref-type="bibr" rid="B18">Georgiev et al., 2019</xref>; <xref ref-type="bibr" rid="B31">Lan et al., 2019</xref>; <xref ref-type="bibr" rid="B32">Lee et al., 2021</xref>; <xref ref-type="bibr" rid="B48">Simione and Nolfi, 2021</xref>; <xref ref-type="bibr" rid="B51">Stolfi et al., 2021</xref>).</p>
<p>The predator and prey problem presents extremely dynamic, highly unpredictable, and hostile environmental conditions. Consequently, it necessitates the development of solutions that are rapid, resilient and adaptable. Moreover, agents must exhibit a range of integrated behavioral and cognitive capabilities, including avoiding stationary and moving obstacles, optimizing motion trajectories under multiple constraints, integrating sensory information over time, anticipating opponent behavior, disorienting opponents, and adapting behavior in real time based on the opponent&#x2019;s actions (<xref ref-type="bibr" rid="B22">Humphries and Driver, 1970</xref>).</p>
<p>The robots in our study are simulated MarXbots (<xref ref-type="bibr" rid="B3">Bonani et al., 2010</xref>) equipped with neural network controllers. The connection strengths within the robots&#x2019; neural networks, which determine their behavior, are encoded in artificial genotypes and evolved. Specifically, predators are evolved to enhance their ability to capture prey (i.e., reach and physically touch the prey) as quickly as possible, while prey are evolved to maximize their ability to avoid being captured for as long as possible. The fitness of predators corresponds to the fraction of time required to capture the prey (see <xref ref-type="disp-formula" rid="e1">Equation 1</xref>). The fitness of preys is the inverse of the fraction of time required by the predator to capture them (see <xref ref-type="disp-formula" rid="e2">Equation 2</xref>).<disp-formula id="e1">
<mml:math id="m2">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold-italic">F</mml:mi>
<mml:mrow>
<mml:mi mathvariant="bold-italic">p</mml:mi>
<mml:mi mathvariant="bold-italic">r</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">d</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi mathvariant="bold-italic">N</mml:mi>
<mml:mi mathvariant="bold-italic">S</mml:mi>
<mml:mi mathvariant="bold-italic">t</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">p</mml:mi>
<mml:mi mathvariant="bold-italic">s</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="bold-italic">C</mml:mi>
<mml:mi mathvariant="bold-italic">S</mml:mi>
<mml:mi mathvariant="bold-italic">t</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">p</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi mathvariant="bold-italic">N</mml:mi>
<mml:mi mathvariant="bold-italic">S</mml:mi>
<mml:mi mathvariant="bold-italic">t</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">p</mml:mi>
<mml:mi mathvariant="bold-italic">s</mml:mi>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
<label>(1)</label>
</disp-formula>
<disp-formula id="e2">
<mml:math id="m3">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold-italic">F</mml:mi>
<mml:mrow>
<mml:mi mathvariant="bold-italic">p</mml:mi>
<mml:mi mathvariant="bold-italic">r</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">y</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mn mathvariant="bold">1.0</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold-italic">F</mml:mi>
<mml:mrow>
<mml:mi mathvariant="bold-italic">p</mml:mi>
<mml:mi mathvariant="bold-italic">r</mml:mi>
<mml:mi mathvariant="bold-italic">e</mml:mi>
<mml:mi mathvariant="bold-italic">d</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
<label>(2)</label>
</disp-formula>
</p>
<p>In the equations above, <italic>NSteps</italic> denotes the maximum length of the evaluation episode (1000 in our experimental scenarios)<italic>, CStep</italic> represents the step in which the predator succeeds in capturing the prey, <italic>F</italic>
<sub>
<italic>pred</italic>
</sub> is the fitness of the predator, while <italic>F</italic>
<sub>
<italic>prey</italic>
</sub> indicates the performance of the prey.</p>
<p>The simulated MarXbots are circular robots with a 17 cm diameter. They are equipped with a differential drive motion system, a ring of 24 color LEDs, 24 infrared sensors, 4 ground sensors, an omnidirectional camera, and a traction sensor. In the experiments, the LEDs of predator robots are set in red, while the LEDs of prey robots are set in green. The robots were placed in a 3 &#xd7; 3 m square arena surrounded by black walls. The arena floor was grayscale, varying from white to black from the center to the periphery (see <xref ref-type="fig" rid="F1">Figure 1</xref>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>The robots and the environment in simulation. The red and green robots correspond to the predator and prey robots, respectively.</p>
</caption>
<graphic xlink:href="frobt-11-1470886-g001.tif"/>
</fig>
<p>The maximum wheel speed that the wheels of the differential drive motion system could assume was 10 rad/s for the prey and 8.5 rad/s for the predators respectively. The relative speed of the two robot types was adjusted to balance the overall complexity of the problem faced by predators and preys, i.e., preventing one species from reaching maximum or minimum fitness. The environment state, robot sensors, motors, and neural network were updated at a frequency of 10 Hz.</p>
<p>The neural network controller of the robot consists of a LSTM (Long Short-Term Memory, see <xref ref-type="bibr" rid="B21">Hochreiter and Schmidhuber, 1997</xref>; <xref ref-type="bibr" rid="B19">Gers and Schmidhuber, 2001</xref>) recurrent neural network with 23 sensory neurons, 25 internal units, and 2 motor neurons. The sensory layer includes 8 sensory neurons encoding the average activation state of eight groups of three adjacent infrared sensors, 8 neurons that encode the fraction of green or red light perceived in the eight 45&#xb0; sectors of the visual field of the camera, 1 neuron that encodes the average amount of green or red light detected in the entire visual field of the camera, 4 neurons encoding the state of the four ground sensors, 1 neuron that encodes the average activation of the four ground sensors, 1 neuron that encodes whether the robot collides with an obstacle (i.e., whether the traction force detected by the traction sensor exceeds a threshold). The sensory neuron states are normalized in the range [0.0, 1.0]. The motor layer includes two neurons encoding the desired translational and rotational motion of the robot in the range [0.0, 1.0].</p>
<p>To compare the relative effectiveness of the considered methods, we continued the evolutionary process until a total of 75 billion evaluation steps were performed. This ensures that the comparison of the different techniques is fair.</p>
<p>The experiments included in this article can be replicated using the Evorobotpy2 tool which is available from the Github repository <ext-link ext-link-type="uri" xlink:href="https://github.com/snolfi/evorobotpy2">https://github.com/snolfi/evorobotpy2</ext-link>. The source code of the co-evolutionary algorithms and of the associated experiments is available from the Github repository <ext-link ext-link-type="uri" xlink:href="https://github.com/snolfi/competitive-evolution">https://github.com/snolfi/competitive-evolution</ext-link>.</p>
</sec>
</sec>
<sec sec-type="results" id="s3">
<title>3 Results</title>
<p>In this section, we present the results obtained using the Archive, Maxsolve&#x2a;, Archive&#x2a;, and Generalist algorithms. The results include data collected from 10 replication experiments conducted with each algorithm, resulting in a total of 40 experiments.</p>
<p>
<xref ref-type="fig" rid="F2">Figure 2</xref> displays master tournament data, i.e., the performance of predator and prey robots of different generations evaluated against opponents of previous, current, and future generations. The results demonstrate that all considered methods exhibit historical progress overall. Notably, the robots perform better against opponents from previous generations than against those from successive generations in most cases. Additionally, these data provide an indication of global progress, as the robots of generation X &#x2b; N perform better against opponents from future generations than robots of generation X in most of the cases. However, it is worth noting that only the Generalist algorithm consistently produces progress across all phases, resulting in monotonically better robots. The other algorithms also exhibit retrogressions, albeit less frequently than progress.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>Performance of predators of different evolutionary phases evaluated against opponents of all phases. Results obtained with the Archive, Maxsolve&#x2a;, Archive&#x2a; and Generalist algorithms. Data was collected by evaluating the predators of phase 0 (generation 0) and the subsequent 10 phases against prey from phase 0 (generation 0) and the following 10 phases. The vertical and horizontal axes represent the phases of the predator and prey, respectively. The color of each cell indicates the average performance of the predator against prey from the corresponding phases. The performance of the prey is the inverse of the predator&#x2019;s performance, i.e., it corresponds to 1.0 minus the predator&#x2019;s performance (see <xref ref-type="disp-formula" rid="e2">Equation 2</xref>). The phases are separated by <inline-formula id="inf2">
<mml:math id="m4">
<mml:mrow>
<mml:mfrac>
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mn>10</mml:mn>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</inline-formula> of the total generations. The results are averaged over 10 evolutionary experiments. Generally, the performance of predators from any given phase (shown in any given row) improves (becomes redder) against opponents from previous generations (displayed in the first columns) and declines (becomes bluer) against opponents from successive generations (displayed in the last columns). Similarly, the performance of prey from any given phase (displayed in any given column) improves (becomes bluer) against opponents from previous phases (shown in the first rows) and worsens (becomes less blue) against opponents from successive phases (displayed in the last rows).</p>
</caption>
<graphic xlink:href="frobt-11-1470886-g002.tif"/>
</fig>
<p>This qualitative difference is confirmed by <xref ref-type="fig" rid="F3">Figure 3</xref>, which illustrates the performance of the robots from last generation against opponents from both the current and previous generations. The Archive and Maxsolve&#x2a; robots of the last generations consistently exhibit better and better performance against opponents of the last four preceding phases only. The Archive&#x2a; robots of the last generations consistently display better and better performance against opponents of the last 3 phases (for predator robots) and the last 5 phases (for prey robots) only. Only robots evolved with the Generalist algorithm consistently exhibit improved performance against older opponents across all phases.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>Performance of predators (red lines) and prey (green lines) from the last generation against opponents from the same generation (0) and opponents from the 10 preceding phases (1&#x2013;10), where 10 corresponds to opponents from the initial generation. The 10 phases are separated by <inline-formula id="inf3">
<mml:math id="m5">
<mml:mrow>
<mml:mfrac>
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mn>10</mml:mn>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</inline-formula> of the total generations. The vertical axes represent the average performance while the horizontal axes represent the preceding phase of the opponents, where 0 correspond to opponents from the last generation and 10 corresponds to opponent from generation 0. Results were obtained using the Archive, Maxsolve&#x2a;, Archive&#x2a;, and Generalist algorithms. Each plot represents the average results of 10 evolutionary experiments. These plots display the same data as shown in the last row and last column of the matrices presented in <xref ref-type="fig" rid="F2">Figure 2</xref>.</p>
</caption>
<graphic xlink:href="frobt-11-1470886-g003.tif"/>
</fig>
<p>To identify the best performing method, we conducted cross-tests comparing the champions of each algorithm against those of each other algorithms. Each champion was selected from the best predators and prey of the last 40 generations of the corresponding replication. Consequently, we have 10 champion predators and 10 champion prey for each experimental condition. The cross-tests were conducted by comparing the performance of a set of 10 champion agents evolved under one experimental condition against champion opponents evolved under the same or a different experimental condition. More specifically, cross-tests values were computed according to <xref ref-type="disp-formula" rid="e3">Equation 3</xref>:<disp-formula id="e3">
<mml:math id="m6">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold-italic">c</mml:mi>
<mml:mn mathvariant="bold">12</mml:mn>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mrow>
<mml:mstyle displaystyle="true">
<mml:munderover>
<mml:mo>&#x2211;</mml:mo>
<mml:mrow>
<mml:mi mathvariant="bold-italic">i</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn mathvariant="bold">1</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi mathvariant="bold-italic">j</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:mrow>
<mml:mn mathvariant="bold">10</mml:mn>
</mml:munderover>
</mml:mstyle>
<mml:mrow>
<mml:mi mathvariant="bold-italic">F</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msubsup>
<mml:mi mathvariant="bold-italic">A</mml:mi>
<mml:mi mathvariant="bold-italic">i</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msubsup>
<mml:mtext>&#x2009;</mml:mtext>
<mml:mo>&#x7c;</mml:mo>
<mml:mtext>&#x2009;</mml:mtext>
<mml:msubsup>
<mml:mi mathvariant="bold-italic">O</mml:mi>
<mml:mi mathvariant="bold-italic">j</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msubsup>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="bold-italic">F</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msubsup>
<mml:mi mathvariant="bold-italic">A</mml:mi>
<mml:mi mathvariant="bold-italic">i</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msubsup>
<mml:mtext>&#x2009;</mml:mtext>
<mml:mo>&#x7c;</mml:mo>
<mml:mtext>&#x2009;</mml:mtext>
<mml:msubsup>
<mml:mi mathvariant="bold-italic">O</mml:mi>
<mml:mi mathvariant="bold-italic">j</mml:mi>
<mml:mn mathvariant="bold">2</mml:mn>
</mml:msubsup>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(3)</label>
</disp-formula>where <italic>c</italic> is the cross-test value, <italic>1</italic> and <italic>2</italic> denote the experimental conditions being compared, <italic>F()</italic> indicates the performance (fitness), <italic>A</italic> represents the agents, <italic>O</italic> represents the opponents, <italic>i</italic> and <italic>j</italic> are indices for the 10 champion agents and opponents, respectively. The notation (x &#x7c;y) indicates the evaluation of an individual x against the opponent y. The symbol <inline-formula id="inf4">
<mml:math id="m7">
<mml:mrow>
<mml:msubsup>
<mml:mi>x</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>k</mml:mi>
</mml:msubsup>
</mml:mrow>
</mml:math>
</inline-formula> denotes the <italic>i-th</italic> individual (x) evolved under the experimental condition <italic>k</italic>. Positive and negative cross-test values indicate the superiority and inferiority, respectively, of the first experimental condition over the second.</p>
<p>
<xref ref-type="table" rid="T1">Table 1</xref> presents the results of cross-tests conducted among the four experimental conditions. Notably, agents evolved using the Archive&#x2a; method significantly outperform those evolved with the Archive method, for both predators and preys. Furthermore, agents evolved using the Generalist method significantly outperform agents from the other three methods, again for both predators and preys.</p>
<table-wrap id="T1" position="float">
<label>TABLE 1</label>
<caption>
<p>The cross-test of champion agents conducted using the Archive, Maxsolve&#x2a;, Archive&#x2a;, and Generalist algorithms.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th colspan="5" align="center">Predators</th>
</tr>
<tr>
<th align="left"/>
<th align="left">Archive</th>
<th align="left">Maxsolve&#x2a;</th>
<th align="left">Archive&#x2a;</th>
<th align="left">Generalist</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td colspan="5" align="left">Archive</td>
</tr>
<tr>
<td align="left">MaxSolve&#x2a;</td>
<td align="right">0.05, p &#x3d; .270</td>
<td align="left"/>
<td align="left"/>
<td align="left"/>
</tr>
<tr>
<td align="left">Archive&#x2a;</td>
<td align="right">
<bold>0.15, p &#x3d; .001</bold>
</td>
<td align="right">0.04, p &#x3d; .112</td>
<td align="left"/>
<td align="left"/>
</tr>
<tr>
<td align="left">Generalist</td>
<td align="right">
<bold>0.12, p &#x3d; .007</bold>
</td>
<td align="right">
<bold>0.13, p &#x3d; .001</bold>
</td>
<td align="right">
<bold>0.14, p &#x3d; .001</bold>
</td>
<td align="left"/>
</tr>
</tbody>
</table>
<table>
<thead valign="top">
<tr>
<th colspan="5" align="center">Preys</th>
</tr>
<tr>
<th align="left"/>
<th align="left">Archive</th>
<th align="left">Maxsolve&#x2a;</th>
<th align="left">Archive&#x2a;</th>
<th align="left">Generalist</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td colspan="5" align="left">Archive</td>
</tr>
<tr>
<td align="left">MaxSolve&#x2a;</td>
<td align="right">0.04, p &#x3d; .933</td>
<td align="left"/>
<td align="left"/>
<td align="left"/>
</tr>
<tr>
<td align="left">Archive&#x2a;</td>
<td align="right">
<bold>0.17, p &#x3d; .003</bold>
</td>
<td align="right">0.12, p &#x3d; .044</td>
<td align="left"/>
<td align="left"/>
</tr>
<tr>
<td align="left">Generalist</td>
<td align="right">
<bold>0.27, p &#x3d; .001</bold>
</td>
<td align="right">
<bold>0.23, p &#x3d; .001</bold>
</td>
<td align="left">
<bold>0.24 p &#x3d; 0.001</bold>
</td>
<td align="left"/>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn>
<p>In each row, we evaluate the performance of agents evolved under the condition indicated in that row against opponents evolved under the condition indicated in the corresponding column. We then subtract the performance obtained by evaluating the agents against the opponents indicated in the same row. Positive values indicate that the condition indicated in the row outperforms the condition indicated in the column. The numbers denoted by &#x201c;p &#x3d; &#x201d; represent the probability that the performance obtained against the two sets of opponents belongs to the same distribution. Values in bold indicate cases where the difference in performance is statistically significant (Mann&#x2013;Whitney U-test with Bonferroni correction, p-value &#x3c;0.0167). The table is divided into two parts: the top section displays cross-tests using predators as agents and preys as opponents, while the bottom section reverses the roles.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>To further validate the effectiveness of the alternative methods and assess the generality of the solutions, we conducted an additional analysis. Specifically, we evaluated the 40 champion predators (obtained from each method in corresponding 10 replications) against the 40 champion preys. The results, displayed in <xref ref-type="fig" rid="F4">Figures 4</xref>, <xref ref-type="fig" rid="F5">5</xref> and <xref ref-type="table" rid="T2">Table 2</xref>, reveal significant differences. Notably, Generalist predator and prey champions outperform the champions obtained with all other methods (Mann&#x2013;Whitney U-test with Bonferroni correction, p-value &#x3c;0.0167). Moreover, Archive&#x2a; predator and prey champions outperform Archive predators and preys (Mann&#x2013;Whitney U-test with Bonferroni correction, p-value &#x3c;0.0167). Instead, the performances of Maxsolve&#x2a; and Archive predator and prey champions do not significantly differ (Mann&#x2013;Whitney U-test with Bonferroni correction, p-value &#x3e;0.0167).</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption>
<p>Post-evaluation of the 40 predator champions obtained using four different methods against the 40 prey champions obtained using the same methods. The table consists of four sets of 10 rows and columns, representing the performance of predator and prey champions from the Archive (rows 0&#x2013;9 and columns 0&#x2013;9), Maxsolve&#x2a; (rows 10&#x2013;19 and columns 10&#x2013;19), Archive&#x2a; (rows 20&#x2013;29 and columns 20&#x2013;29), and Generalist (rows 30&#x2013;39 and columns 30&#x2013;39) methods. Each champion is selected from a corresponding replication. Each pixel&#x2019;s color indicates the predator&#x2019;s performance, while the prey&#x2019;s performance corresponds to the inverse of the predator&#x2019;s one.</p>
</caption>
<graphic xlink:href="frobt-11-1470886-g004.tif"/>
</fig>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption>
<p>The performance achieved by the 40 predator champions (top figure) and the 40 prey champions (bottom figure), evolved using the four different methods across 10 replications of the experiments. The data refer to the post-evaluation phase (see also <xref ref-type="fig" rid="F4">Figure 4</xref>). Boxes denote the interquartile range of the data. The horizontal line inside the boxes represents the median value, while the green triangle in each box indicates the average fitness value (see also <xref ref-type="table" rid="T2">Table 2</xref>). The whiskers extend to the most extreme data points within 1.5 times the interquartile range from the.</p>
</caption>
<graphic xlink:href="frobt-11-1470886-g005.tif"/>
</fig>
<table-wrap id="T2" position="float">
<label>TABLE 2</label>
<caption>
<p>Average performance achieved by the champion agents evolved with the four different methods.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th colspan="4" align="center">Predators</th>
</tr>
<tr>
<th align="left">Archive</th>
<th align="left">Maxsolve&#x2a;</th>
<th align="left">Archive&#x2a;</th>
<th align="left">Generalist</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">0.282 [0.382]</td>
<td align="left">0.335 [0.401]</td>
<td align="left">0.415 [0.414]</td>
<td align="left">
<bold>0.572 [0.381]</bold>
</td>
</tr>
</tbody>
</table>
<table>
<thead valign="top">
<tr>
<th colspan="4" align="center">Preys</th>
</tr>
<tr>
<th align="left">Archive</th>
<th align="left">Maxsolve&#x2a;</th>
<th align="left">Archive&#x2a;</th>
<th align="left">Generalist</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">0.511 [0.415]</td>
<td align="left">0.552 [0.419]</td>
<td align="left">0.613 [0.403]</td>
<td align="left">
<bold>0.720 [0.369]</bold>
</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn>
<p>Data refer to the cross-test (green triangles in <xref ref-type="fig" rid="F4">Figure 4</xref>). Data within squared brackets denote the standard deviations. Bold values indicate the best results. The table is split in two parts: the top section shows the performance obtained by the predator champions, while the bottom one displays the fitness achieved by the prey champions.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>As shown in <xref ref-type="fig" rid="F4">Figure 4</xref>, the predators and preys obtained using the Generalist method (displayed in rows and columns 30&#x2013;39, respectively) achieve the best performance. Notably, the fifth and sixth predator champions obtained with the Generalist algorithm (<xref ref-type="fig" rid="F4">Figure 4</xref>, rows 34 and 35) achieve a performance of at least 0.5 against 29 out of the 30 prey champions obtained using the other three methods. Additionally, the seventh and eighth prey champions obtained with the generalist algorithm (<xref ref-type="fig" rid="F4">Figure 4</xref>, columns 36 and 37) also achieve a performance of at least 0.5 against 29 out of 30 predator champions obtained using the other three methods.</p>
<p>The analysis of the behavior displayed by the champions reveals their acquisition of sophisticated behavioral skills. Specifically, some of the champions demonstrate the ability to move both toward the front and rear directions, skillfully alternating their direction of motion based on the circumstances (as shown in Video 1, <xref ref-type="app" rid="app1">Appendix</xref>). They exhibit the capability to capture and evade a wide range of opponents. Moreover, they remain robust against adversarial behaviors exhibited by opponents in most cases; in other words, they are rarely fooled by opponent strategies, despite those strategies have been fine-tuned against them. The ability to alternate their direction of motion depending on the circumstances is more commonly observed in agents evolved using the Generalist algorithm.</p>
<p>An illustrative example of an agent vulnerable to specific opponent behavior is the 8th champion predator obtained using the Generalist method. While this predator is generally effective, it proves fragile when confronted with the adversarial strategy employed by the 7th champion prey. This prey displays an oscillatory behavior that triggers a harmless spinning-in-place response from the predator (see Video 2, <xref ref-type="app" rid="app1">Appendix</xref>). Another instance of an agent vulnerable to specific opponent behavior is the predator shown in Video 3 (<xref ref-type="app" rid="app1">Appendix</xref>). The 4th prey champion, obtained through the Archive&#x2a; algorithm, effectively neutralizes this specific predator by moving counterclockwise around it, consistently eliciting the same avoidable attacking behavior in the opponent.</p>
</sec>
<sec sec-type="conclusion" id="s4">
<title>4 Conclusion</title>
<p>In this article, we delve into the conditions that drive competitive evolution toward genuine progress, i.e., toward solutions that become better and better against all possible opponents. Specifically, we introduced a set of methods for measuring historical and global progress, we discussed factors that facilitate genuine progress, and we compared the efficacy of four algorithms.</p>
<p>The methods considered were the follows: (1) the Archive algorithm (<xref ref-type="bibr" rid="B44">Rosin and Belew, 1997</xref>) that promotes global progress by maintaining an archive of the best individuals from previous generations. This permits to evaluate evolving individuals against opponents of current and previous generations. (2) The Maxsolve&#x2a; algorithm, i.e., a variation of the original De Jong&#x2019;s algorithm (<xref ref-type="bibr" rid="B13">De Jong, 2005</xref>) adapted for both transitive and non-transitive problems. This method also relies on an archive that, however, is used to preserve only diversified individuals. (3) The Archive&#x2a; algorithm, introduced in this paper, which extends the vanilla Archive method by leveraging multiple evolving populations. This extension allows for the inclusion of more diverse individuals in the archive. (4) The Generalist algorithm (<xref ref-type="bibr" rid="B48">Simione and Nolfi, 2021</xref>) that does not use an archive but incorporates a mechanism for identifying and discarding variations leading to local progress only. To analyze the long-term dynamics of the co-evolutionary process, the methods were compared by performing long-lasting experiments.</p>
<p>The results obtained in a predator-prey scenario, commonly used to study competitive co-evolution, demonstrate that all the considered methods lead to global progress in the long term. However, the rate of progress and the ratio of progress versus retrogressions vary significantly.</p>
<p>The Generalist method outperforms the other three methods and is the only one capable of producing solutions that consistently score better and better across generations against previous and future opponents in successive evolutionary phases. The other three methods also exhibit retrogression phases, although less frequently than progression phases. Additionally, the Archive&#x2a; algorithm, introduced in this paper, outperforms both the vanilla Archive Algorithm and the MaxSolve&#x2a; algorithm. Overall, our results demonstrate the utilization of proper methods can prevent the convergence of the evolutionary process in limit-cycle dynamics, which jeopardizes the appealing properties of competitive co-evolution.</p>
<p>The superiority of the Generalist algorithm is also demonstrated through visual comparisons of the behavior exhibited by the evolving robots. Indeed, the ability to move bi-directionally and appropriately alternate the direction of motion depending on the circumstances, providing a significant advantage, is more commonly observed among the robots evolved with the Generalist algorithm than among the robots evolved with other algorithms.</p>
<p>Future research should verify whether our results generalize to other competitive settings and whether the continuation of evolutionary progress can lead to an open-ended dynamic in which the efficacy of the evolved solutions keeps increasing in an unbounded manner. The remarkable results recently achieved with Large Language Models also open the possibility to leverage the knowledge acquired by these systems and their online learning capabilities to select useful opponents. Pioneering attempts in this direction are reported in <xref ref-type="bibr" rid="B29">Jorgensen et al. (2024)</xref> and <xref ref-type="bibr" rid="B54">Zala et al. (2024)</xref>.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s5">
<title>Data availability statement</title>
<p>The raw data supporting the conclusions of this article will be made available by the authors, without undue reservation.</p>
</sec>
<sec sec-type="author-contributions" id="s6">
<title>Author contributions</title>
<p>SN: Conceptualization, Data curation, Formal Analysis, Funding acquisition, Investigation, Methodology, Software, Supervision, Validation, Writing&#x2013;original draft, Writing&#x2013;review and editing. PP: Data curation, Formal Analysis, Investigation, Methodology, Software, Validation, Writing&#x2013;review and editing.</p>
</sec>
<sec sec-type="funding-information" id="s7">
<title>Funding</title>
<p>The author(s) declare that financial support was received for the research, authorship, and/or publication of this article. We acknowledge financial support from PNRR MUR project PE0000013-FAIR.</p>
</sec>
<sec sec-type="COI-statement" id="s8">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s9">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bansal</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Pachocki</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sidor</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Sutskever</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Mordatch</surname>
<given-names>I.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Emergent complexity via multi-agent competition</article-title>. <source>arXiv Prepr. arXiv:1710.03748</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1710.03748</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Bari</surname>
<given-names>A. G.</given-names>
</name>
<name>
<surname>Gaspar</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Wiegand</surname>
<given-names>R. P.</given-names>
</name>
<name>
<surname>Bucci</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2018</year>). &#x201c;<article-title>Selection methods to relax strict acceptance condition in test-based coevolution</article-title>,&#x201d; in <source>2018 IEEE congress on evolutionary computation (CEC)</source> (<publisher-loc>IEEE</publisher-loc>), <fpage>1</fpage>&#x2013;<lpage>8</lpage>.</citation>
</ref>
<ref id="B3">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Bonani</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Longchamp</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Magnenat</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>R&#xe9;tornaz</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Burnier</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Roulet</surname>
<given-names>G.</given-names>
</name>
<etal/>
</person-group> (<year>2010</year>). &#x201c;<article-title>The marXbot, a miniature mobile robot opening new perspectives for the collective-robotic research</article-title>,&#x201d; in <source>IEEE/RSJ 2010 international conference on intelligent robots and systems, IROS 2010 - conference proceedings</source>, <fpage>4187</fpage>&#x2013;<lpage>4193</lpage>.</citation>
</ref>
<ref id="B4">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Buason</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Bergfeldt</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Ziemke</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2005</year>). &#x201c;<article-title>Brains, bodies, and beyond: competitive co-evolution of robot controllers, morphologies and environments</article-title>,&#x201d; in <source>Genetic programming and evolvable machines</source> (<publisher-name>Springer Science &#x2b; Business Media, Inc. Manufactured in The Netherlands</publisher-name>), <fpage>25</fpage>&#x2013;<lpage>51</lpage>.</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Buason</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Ziemke</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2003</year>). <article-title>Co-evolving task-dependent visual morphologies in predator-prey experiments</article-title>. <source>Lect. Notes Comput. Sci.</source> <volume>2723</volume>, <fpage>458</fpage>&#x2013;<lpage>469</lpage>. <pub-id pub-id-type="doi">10.1007/3-540-45105-6_58</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Carvalho</surname>
<given-names>J. T.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2023</year>). <article-title>The role of morphological variation in evolutionary robotics: maximizing performance and robustness</article-title>. <source>Evol. Comput.</source>, <fpage>1</fpage>&#x2013;<lpage>18</lpage>. <pub-id pub-id-type="doi">10.1162/evco_a_00336</pub-id>
</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chong</surname>
<given-names>S. Y.</given-names>
</name>
<name>
<surname>Ti&#x148;o</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Ku</surname>
<given-names>D. C.</given-names>
</name>
<name>
<surname>Yao</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>Improving generalization performance in co-evolutionary learning</article-title>. <source>IEEE Trans. Evol. Comput.</source> <volume>16</volume>, <fpage>70</fpage>&#x2013;<lpage>85</lpage>. <pub-id pub-id-type="doi">10.1109/tevc.2010.2051673</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chong</surname>
<given-names>S. Y.</given-names>
</name>
<name>
<surname>Ti&#x148;o</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Yao</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Relationship between generalization and diversity in coevolutionary learning</article-title>. <source>IEEE Transaction Comput. Intell AI Games</source> <volume>1</volume>, <fpage>214</fpage>&#x2013;<lpage>232</lpage>. <pub-id pub-id-type="doi">10.1109/TCIAIG.2009.2034269</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Cliff</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Miller</surname>
<given-names>G. F.</given-names>
</name>
</person-group> (<year>1995</year>). &#x201c;<article-title>Tracking the red queen: measurements of adaptive progress in co-evolutionary simulations</article-title>,&#x201d; in <source>Lecture notes in computer science (including subseries lecture notes in artificial intelligence and lecture notes in bioinformatics)</source> (<publisher-name>Springer Verlag</publisher-name>), <fpage>200</fpage>&#x2013;<lpage>218</lpage>.</citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cliff</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Miller</surname>
<given-names>G. F.</given-names>
</name>
</person-group> (<year>2006</year>). <article-title>Visualizing coevolution with CIAO plots</article-title>. <source>Artif. Life</source> <volume>12</volume>, <fpage>199</fpage>&#x2013;<lpage>202</lpage>. <pub-id pub-id-type="doi">10.1162/artl.2006.12.2.199</pub-id>
</citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dawkins</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Krebs</surname>
<given-names>J. R.</given-names>
</name>
</person-group> (<year>1979</year>). <article-title>Arms races between and within species</article-title>. <source>Proc. R. Soc. B Biol. Sci.</source> <volume>205</volume>, <fpage>489</fpage>&#x2013;<lpage>511</lpage>. <pub-id pub-id-type="doi">10.1098/rspb.1979.0081</pub-id>
</citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>De Jong</surname>
<given-names>E.</given-names>
</name>
</person-group> (<year>2004</year>). <article-title>The incremental pareto-coevolution archive</article-title>. <source>Lect. Notes Comput. Sci.</source> <volume>3102</volume>, <fpage>525</fpage>&#x2013;<lpage>536</lpage>. <pub-id pub-id-type="doi">10.1007/978-3-540-24854-5_55</pub-id>
</citation>
</ref>
<ref id="B13">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>De Jong</surname>
<given-names>E.</given-names>
</name>
</person-group> (<year>2005</year>). &#x201c;<article-title>The MaxSolve algorithm for coevolution</article-title>,&#x201d; in <source>Genetic and evolutionary computation (GECCO 2005), lecture notes in computer science</source>, <fpage>483</fpage>&#x2013;<lpage>489</lpage>.</citation>
</ref>
<ref id="B14">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Ficici</surname>
<given-names>S. G.</given-names>
</name>
<name>
<surname>Pollack</surname>
<given-names>J. B.</given-names>
</name>
</person-group> (<year>2003</year>). &#x201c;<article-title>A game-theoretic memory mechanism for coevolution</article-title>,&#x201d; in <source>Genetic and evolutionary computation (GECCO-2003), lectures notes on computer sciences</source> (<publisher-loc>Chicago, USA</publisher-loc>: <publisher-name>Springer Verlag</publisher-name>), <fpage>286</fpage>&#x2013;<lpage>297</lpage>.</citation>
</ref>
<ref id="B15">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Floreano</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>1997a</year>). &#x201c;<article-title>Adaptive behavior in competing co-evolving species</article-title>,&#x201d; in <source>Proceeding of the fourth European conference on artificial life</source>, <fpage>378</fpage>&#x2013;<lpage>387</lpage>.</citation>
</ref>
<ref id="B16">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Floreano</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>1997b</year>). &#x201c;<article-title>God save the red queen! Competition in co-evolutionary robotics</article-title>,&#x201d; in <source>Genetic programming 1997: proceedings of the second annual conference</source>, <fpage>398</fpage>&#x2013;<lpage>406</lpage>.</citation>
</ref>
<ref id="B17">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Floreano</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Mondada</surname>
<given-names>F.</given-names>
</name>
</person-group> (<year>1998</year>). &#x201c;<article-title>Competitive Co-evolutionary robotics: from theory to practice</article-title>,&#x201d; in <source>Proc. Of the fifth international conference on simulation of adaptive behavior (SAB), from animals to animats</source> (<publisher-loc>Z&#xfc;rich</publisher-loc>: <publisher-name>ETH</publisher-name>).</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Georgiev</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Tanev</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Shimohara</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Ray</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Evolution, robustness and generality of a team of simple agents with asymmetric morphology in predator-prey pursuit problem</article-title>. <source>Information</source> <volume>10</volume> (<issue>2</issue>), <fpage>72</fpage>. <pub-id pub-id-type="doi">10.3390/info10020072</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gers</surname>
<given-names>F. A.</given-names>
</name>
<name>
<surname>Schmidhuber</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2001</year>). <article-title>LSTM recurrent networks learn simple context free and context sensitive languages</article-title>. <source>IEEE Trans. Neural Netw.</source> <volume>12</volume> (<issue>6</issue>), <fpage>1333</fpage>&#x2013;<lpage>1340</lpage>. <pub-id pub-id-type="doi">10.1109/72.963769</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hochreiter</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Schmidhuber</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>1997</year>). <article-title>Long short-term memory</article-title>. <source>Neural Comput.</source> <volume>9</volume> (<issue>8</issue>), <fpage>1735</fpage>&#x2013;<lpage>1780</lpage>. <pub-id pub-id-type="doi">10.1162/neco.1997.9.8.1735</pub-id>
</citation>
</ref>
<ref id="B22">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Humphries</surname>
<given-names>D. A.</given-names>
</name>
<name>
<surname>Driver</surname>
<given-names>P. M.</given-names>
</name>
</person-group> (<year>1970</year>). <article-title>Protean defence by prey animals</article-title>. <source>Oecologia</source> <volume>5</volume>, <fpage>285</fpage>&#x2013;<lpage>302</lpage>. <pub-id pub-id-type="doi">10.1007/BF00815496</pub-id>
</citation>
</ref>
<ref id="B23">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ito</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Pilat</surname>
<given-names>M. L.</given-names>
</name>
<name>
<surname>Suzuki</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Arita</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>ALife approach for body-behavior predator&#x2013;prey coevolution: body first or behavior first?</article-title> <source>Artif. Life Robotics</source> <volume>18</volume> (<issue>1</issue>), <fpage>36</fpage>&#x2013;<lpage>40</lpage>. <pub-id pub-id-type="doi">10.1007/s10015-013-0096-y</pub-id>
</citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Jaderberg</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Czarnecki</surname>
<given-names>W. M.</given-names>
</name>
<name>
<surname>Dunning</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Marris</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Lever</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Castaneda</surname>
<given-names>A. G.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Human-level performance in first-person multiplayer games with population-based deep reinforcement learning</article-title>. <source>Arxiv. arXiv Prepr. arXiv:1807.01281</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1807.01281</pub-id>
</citation>
</ref>
<ref id="B25">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Jain</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Subramoney</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Miikulainen</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2012</year>). &#x201c;<article-title>Task decomposition with neuroevolution in extended predator-prey domain</article-title>,&#x201d; in <source>Alife 2012: the thirteenth international conference on the synthesis and simulation of living systems</source> (<publisher-name>MIT Press</publisher-name>), <fpage>341</fpage>&#x2013;<lpage>348</lpage>.</citation>
</ref>
<ref id="B26">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Jakobi</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Husbands</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Harvey</surname>
<given-names>I.</given-names>
</name>
</person-group> (<year>1995</year>). &#x201c;<article-title>Noise and the reality gap: the use of simulation in evolutionary robotics</article-title>,&#x201d; in <source>Advances in artificial life: third European conference on artificial life</source> (<publisher-name>Springer Berlin Heidelberg</publisher-name>), <fpage>704</fpage>&#x2013;<lpage>720</lpage>.</citation>
</ref>
<ref id="B28">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Ja&#x15b;kowski</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Liskowski</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Szubert</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Krawiec</surname>
<given-names>K.</given-names>
</name>
</person-group> (<year>2013</year>). &#x201c;<article-title>Improving coevolution by random sampling</article-title>,&#x201d; in <source>Proceedings of the 15th annual conference on Genetic and evolutionary computation</source>, <fpage>1141</fpage>&#x2013;<lpage>1148</lpage>.</citation>
</ref>
<ref id="B29">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Jorgensen</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Nadizar</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Pietropolli</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Manzoni</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Medvet</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>O&#x27;Reilly</surname>
<given-names>U. M.</given-names>
</name>
<etal/>
</person-group> (<year>2024</year>). &#x201c;<article-title>Large Language model-based test case generation for GP agents</article-title>,&#x201d; in <source>Proceedings of the genetic and evolutionary computation conference</source>, <fpage>914</fpage>&#x2013;<lpage>923</lpage>.</citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kingma</surname>
<given-names>D. P.</given-names>
</name>
<name>
<surname>Ba</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>Adam: a method for stochastic optimization</article-title>. <source>arXiv Prepr. arXiv:1412.6980</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1412.6980</pub-id>
</citation>
</ref>
<ref id="B31">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Lan</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Eiben</surname>
<given-names>A. E.</given-names>
</name>
</person-group> (<year>2019</year>). &#x201c;<article-title>Evolutionary predator-prey robot systems: from simulation to real world</article-title>,&#x201d; in <source>Proceedings of the genetic and evolutionary computation conference companion</source>, <fpage>123</fpage>&#x2013;<lpage>124</lpage>.</citation>
</ref>
<ref id="B32">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lee</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Kim</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Shim</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Simulation of sustainable co-evolving predator-prey system controlled by neural network</article-title>. <source>J. Korea Soc. Comput. Inf.</source> <volume>26</volume> (<issue>9</issue>), <fpage>27</fpage>&#x2013;<lpage>35</lpage>. <pub-id pub-id-type="doi">10.9708/jksci.2021.26.09.027</pub-id>
</citation>
</ref>
<ref id="B33">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Liskowski</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Krawiec</surname>
<given-names>K.</given-names>
</name>
</person-group> (<year>2016</year>). &#x201c;<article-title>Online discovery of search objectives for test-based problems</article-title>,&#x201d; in <source>Proceedings of the 2016 on genetic and evolutionary computation conference companion</source>, <fpage>163</fpage>&#x2013;<lpage>164</lpage>.</citation>
</ref>
<ref id="B34">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Lowd</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Meek</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2005</year>). &#x201c;<article-title>Adversarial learning</article-title>,&#x201d; in <source>Proceedings of the eleventh ACM SIGKDD international conference on Knowledge discovery in data mining</source> (<publisher-loc>New York, USA</publisher-loc>: <publisher-name>ACM Press</publisher-name>), <fpage>641</fpage>&#x2013;<lpage>647</lpage>.</citation>
</ref>
<ref id="B35">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Miconi</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2008</year>). <article-title>Evolution and complexity: the double-edged sword</article-title>. <source>Artif. Life</source> <volume>14</volume>, <fpage>325</fpage>&#x2013;<lpage>344</lpage>. <pub-id pub-id-type="doi">10.1162/artl.2008.14.3.14307</pub-id>
</citation>
</ref>
<ref id="B36">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Miconi</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2009</year>). &#x201c;<article-title>Why coevolution doesn&#x2019;t &#x201c;Work&#x201d;: superiority and progress in coevolution</article-title>,&#x201d; in <source>Proceedings of the 12th European conference on genetic programming, lecture notes in computer science</source> (<publisher-loc>Berlin</publisher-loc>: <publisher-name>Springer Verlag</publisher-name>), <fpage>49</fpage>&#x2013;<lpage>60</lpage>.</citation>
</ref>
<ref id="B37">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Miller</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Cliff</surname>
<given-names>D.</given-names>
</name>
</person-group> (<year>1994</year>). &#x201c;<article-title>Protean behavior in dynamic games: arguments for the co-evolution of pursuit-evasion tactics</article-title>,&#x201d; in <source>From animals to animats III: proceedings of the third international conference on simulation of adaptive behavior</source>. Editors <person-group person-group-type="editor">
<name>
<surname>Cliff</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Husbands</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Meyer</surname>
<given-names>J. R.</given-names>
</name>
<name>
<surname>Wilson</surname>
<given-names>S. W.</given-names>
</name>
</person-group> (<publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press-Bradford Books</publisher-name>).</citation>
</ref>
<ref id="B38">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>Co-evolving predator and prey robots</article-title>. <source>Adapt. Behav.</source> <volume>20</volume>, <fpage>10</fpage>&#x2013;<lpage>15</lpage>. <pub-id pub-id-type="doi">10.1177/1059712311426912</pub-id>
</citation>
</ref>
<ref id="B39">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Floreano</surname>
<given-names>D.</given-names>
</name>
</person-group> (<year>1998</year>). <article-title>Co-evolving predator and prey robots: do `arms-races&#x2019; arise in artificial evolution?</article-title> <source>Artif. Life</source> <volume>4</volume>, <fpage>1</fpage>&#x2013;<lpage>26</lpage>. <pub-id pub-id-type="doi">10.1162/106454698568620</pub-id>
</citation>
</ref>
<ref id="B40">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Pagliuca</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Milano</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Efficacy of modern neuro-evolutionary strategies for continuous control optimization</article-title>. <source>Front. Robotics AI</source> <volume>7</volume>, <fpage>98</fpage>. <pub-id pub-id-type="doi">10.3389/frobt.2020.00098</pub-id>
</citation>
</ref>
<ref id="B41">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Pagliuca</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Robust optimization through neuroevolution</article-title>. <source>PloS one</source> <volume>14</volume> (<issue>3</issue>), <fpage>e0213193</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0213193</pub-id>
</citation>
</ref>
<ref id="B42">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Palmer</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Chou</surname>
<given-names>A. K.</given-names>
</name>
</person-group> (<year>2012</year>). &#x201c;<article-title>An artificial visual cortex drives behavioral evolution in co-evolved predator and prey robots</article-title>,&#x201d; in <source>Proceedings of the 14th annual conference companion on Genetic and evolutionary computation</source>, <fpage>361</fpage>&#x2013;<lpage>364</lpage>.</citation>
</ref>
<ref id="B43">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Rosin</surname>
<given-names>C. D.</given-names>
</name>
<name>
<surname>Belew</surname>
<given-names>R. K.</given-names>
</name>
</person-group> (<year>1995</year>). &#x201c;<article-title>Methods for competitive co-evolution: finding opponents worth beating</article-title>,&#x201d; in <source>Proceedings of the 6th international conference on genetic algorithms</source> (<publisher-loc>Pittsburgh, PA, USA</publisher-loc>: <publisher-name>Morgan Kaufmann</publisher-name>), <volume>1995</volume>, <fpage>373</fpage>&#x2013;<lpage>381</lpage>.</citation>
</ref>
<ref id="B44">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rosin</surname>
<given-names>C. D.</given-names>
</name>
<name>
<surname>Belew</surname>
<given-names>R. K.</given-names>
</name>
</person-group> (<year>1997</year>). <article-title>New methods for competitive coevolution</article-title>. <source>Evol. Comput.</source> <volume>5</volume>, <fpage>1</fpage>&#x2013;<lpage>29</lpage>. <pub-id pub-id-type="doi">10.1162/evco.1997.5.1.1</pub-id>
</citation>
</ref>
<ref id="B45">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Salimans</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Goodfellow</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Zaremba</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Cheung</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Radford</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>Improved techniques for training gans</article-title>. <source>Adv. neural Inf. Process. Syst.</source> <volume>29</volume>. <pub-id pub-id-type="doi">10.5555/3157096.3157346</pub-id>
</citation>
</ref>
<ref id="B46">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Salimans</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Ho</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Sidor</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Sutskever</surname>
<given-names>I.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Evolution strategies as a scalable alternative to reinforcement learning</article-title>. <source>arXiv:1703</source>, <fpage>03864v0382</fpage>. <pub-id pub-id-type="doi">10.48550/arXiv.1703.03864</pub-id>
</citation>
</ref>
<ref id="B47">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Samothrakis</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Lucas</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Runarsson</surname>
<given-names>T. P.</given-names>
</name>
<name>
<surname>Robles</surname>
<given-names>D.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Coevolving game-playing agents: measuring performance and intransitivities</article-title>. <source>IEEE Trans. Evol. Comput.</source> <volume>17</volume>, <fpage>213</fpage>&#x2013;<lpage>226</lpage>. <pub-id pub-id-type="doi">10.1109/TEVC.2012.2208755</pub-id>
</citation>
</ref>
<ref id="B48">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Simione</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Nolfi</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Long-term progress and behavior complexification in competitive coevolution</article-title>. <source>Artif. Life</source> <volume>26</volume>, <fpage>409</fpage>&#x2013;<lpage>430</lpage>. <pub-id pub-id-type="doi">10.1162/artl_a_00329</pub-id>
</citation>
</ref>
<ref id="B49">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sinervo</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Lively</surname>
<given-names>C. M.</given-names>
</name>
</person-group> (<year>1996</year>). <article-title>The rock-paper-scissors game and the evolution of alternative male strategies</article-title>. <source>Nature</source> <volume>380</volume>, <fpage>240</fpage>&#x2013;<lpage>243</lpage>. <pub-id pub-id-type="doi">10.1038/380240a0</pub-id>
</citation>
</ref>
<ref id="B50">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stanley</surname>
<given-names>K. O.</given-names>
</name>
<name>
<surname>Miikkulainen</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>The dominance tournament method of monitoring progress in coevolution</article-title>. <source>GECCO 2002 Proc. Bird. a Feather Work Genet. Evol. Comput. Conf.</source>, <fpage>242</fpage>&#x2013;<lpage>248</lpage>.</citation>
</ref>
<ref id="B51">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stolfi</surname>
<given-names>D. H.</given-names>
</name>
<name>
<surname>Brust</surname>
<given-names>M. R.</given-names>
</name>
<name>
<surname>Danoy</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Bouvry</surname>
<given-names>P.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>UAV-UGV-UMV multi-swarms for cooperative surveillance</article-title>. <source>Front. Robotics AI</source> <volume>8</volume>, <fpage>616950</fpage>. <pub-id pub-id-type="doi">10.3389/frobt.2021.616950</pub-id>
</citation>
</ref>
<ref id="B52">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Zhu</surname>
<given-names>W.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>A survey on curriculum learning</article-title>. <source>IEEE Trans. Pattern Analysis Mach. Intell.</source> <volume>44</volume>, <fpage>4555</fpage>&#x2013;<lpage>4576</lpage>. <pub-id pub-id-type="doi">10.1109/tpami.2021.3069908</pub-id>
</citation>
</ref>
<ref id="B53">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Wiegand</surname>
<given-names>R. P.</given-names>
</name>
<name>
<surname>Liles</surname>
<given-names>W. C.</given-names>
</name>
<name>
<surname>De Jong</surname>
<given-names>K. A.</given-names>
</name>
</person-group> (<year>2002</year>). &#x201c;<article-title>Analyzing cooperative coevolution with evolutionary game theory</article-title>,&#x201d; in <source>Proceedings of the 2002 congress on evolutionary computation, CEC 2002</source>. (<publisher-name>IEEE Press</publisher-name>), <fpage>1600</fpage>&#x2013;<lpage>1605</lpage>.</citation>
</ref>
<ref id="B54">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zala</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Cho</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Yoon</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Bansal</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2024</year>). <article-title>EnvGen: generating and adapting environments via LLMs for training embodied agents</article-title>. <source>arXiv Prepr. arXiv:2403.12014</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2403.12014</pub-id>
</citation>
</ref>
</ref-list>
<app-group>
<app id="app1">
<title>Appendix</title>
<p>
<bold>Video 1</bold>. Available from <ext-link ext-link-type="uri" xlink:href="https://youtube.com/shorts/prpFwtN-v3Y?feature=share">https://youtube.com/shorts/prpFwtN-v3Y?feature&#x3d;share</ext-link>
</p>
<p>
<bold>Video 2</bold>. Available from <ext-link ext-link-type="uri" xlink:href="https://youtube.com/shorts/77mYMT6CKnI?feature=share">https://youtube.com/shorts/77mYMT6CKnI?feature&#x3d;share</ext-link>
</p>
<p>
<bold>Video 3</bold>. Available from <ext-link ext-link-type="uri" xlink:href="https://youtube.com/shorts/hvrs2rcJVdU?feature=share">https://youtube.com/shorts/hvrs2rcJVdU?feature&#x3d;share</ext-link>
</p>
</app>
</app-group>
</back>
</article>