<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">ESD</journal-id><journal-title-group>
    <journal-title>Earth System Dynamics</journal-title>
    <abbrev-journal-title abbrev-type="publisher">ESD</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Earth Syst. Dynam.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">2190-4987</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/esd-10-789-2019</article-id><title-group><article-title>Improving weather and climate predictions <?xmltex \hack{\break}?> by training of supermodels</article-title><alt-title>Training of supermodels</alt-title>
      </title-group><?xmltex \runningtitle{Training of supermodels}?><?xmltex \runningauthor{F.~Schevenhoven et al}?>
      <contrib-group>
        <contrib contrib-type="author" corresp="yes" rid="aff1 aff2">
          <name><surname>Schevenhoven</surname><given-names>Francine</given-names></name>
          <email>francine.schevenhoven@uib.no</email>
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff3">
          <name><surname>Selten</surname><given-names>Frank</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff4 aff1 aff2 aff5">
          <name><surname>Carrassi</surname><given-names>Alberto</given-names></name>
          
        <ext-link>https://orcid.org/0000-0003-0722-5600</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff2">
          <name><surname>Keenlyside</surname><given-names>Noel</given-names></name>
          
        <ext-link>https://orcid.org/0000-0002-8708-6868</ext-link></contrib>
        <aff id="aff1"><label>1</label><institution>Geophysical Institute, University of Bergen, Bergen, Norway</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>Bjerknes Centre for Climate Research, Bergen, Norway</institution>
        </aff>
        <aff id="aff3"><label>3</label><institution>Royal Netherlands Meteorological Institute, De Bilt, the Netherlands</institution>
        </aff>
        <aff id="aff4"><label>4</label><institution>Nansen Environmental and Remote Sensing Center, Bergen, Norway</institution>
        </aff>
        <aff id="aff5"><label>5</label><institution>Mathematical Institute, Utrecht University, Utrecht, the Netherlands</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Francine Schevenhoven (francine.schevenhoven@uib.no)</corresp></author-notes><pub-date><day>28</day><month>November</month><year>2019</year></pub-date>
      
      <volume>10</volume>
      <issue>4</issue>
      <fpage>789</fpage><lpage>807</lpage>
      <history>
        <date date-type="received"><day>3</day><month>June</month><year>2019</year></date>
           <date date-type="rev-request"><day>19</day><month>June</month><year>2019</year></date>
           <date date-type="rev-recd"><day>11</day><month>October</month><year>2019</year></date>
           <date date-type="accepted"><day>18</day><month>October</month><year>2019</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2019 Francine Schevenhoven et al.</copyright-statement>
        <copyright-year>2019</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019.html">This article is available from https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019.html</self-uri><self-uri xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019.pdf">The full text article is available as a PDF file from https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019.pdf</self-uri>
      <abstract><title>Abstract</title>
    <p id="d1e137">Recent studies demonstrate that weather and climate predictions potentially improve by dynamically combining different models into a so-called “supermodel”. Here, we focus on the weighted supermodel – the supermodel's time derivative is a weighted superposition of the time derivatives of the imperfect models, referred to as weighted supermodeling. A crucial step is to train the weights of the supermodel on the basis of historical observations. Here, we apply two different training methods to a supermodel of up to four different versions of the global atmosphere–ocean–land model SPEEDO. The standard version is regarded as truth. The first training method is based on an idea called cross pollination in time (CPT), where models exchange states during the training. The second method is a synchronization-based learning rule, originally developed for parameter estimation. We demonstrate that both training methods yield climate simulations and weather predictions of superior quality as compared to the individual model versions. Supermodel predictions also outperform predictions based on the commonly used multi-model ensemble (MME) mean. Furthermore, we find evidence that negative weights can improve predictions in cases where model errors do not cancel (for instance, all models are warm with respect to the truth). In principle, the proposed training schemes are applicable to state-of-the-art models and historical observations. A prime advantage of the proposed training schemes is that in the present context relatively short training periods suffice to find good solutions. Additional work needs to be done to assess the limitations due to incomplete and noisy data, to combine models that are structurally different (different resolution and state representation, for instance) and to evaluate cases for which the truth falls outside of the model class.</p>
  </abstract>
    </article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
<sec id="Ch1.S1.SS1">
  <label>1.1</label><title>Premises and the multi-model ensemble</title>
      <?pagebreak page790?><p id="d1e156">Although weather and climate models continue to improve, they will inevitably remain imperfect <xref ref-type="bibr" rid="bib1.bibx2" id="paren.1"/>. Nature is so complex that it is impossible to model all relevant physical processes solely based on the fundamental laws of physics (think, for instance, about the microphysical properties of clouds that determine the cloud radiational properties). Progress in predictive power crucially depends on further improving our knowledge and the numerical representation of the physical processes the model is intended to describe. Nevertheless, with the best possible models in hand, more accurate predictions can be obtained by making good use of all of them, thus exploiting multi-model information. In order to reduce the impact of model errors on predictions, it is common practice to combine the predictions of a collection of different models in a statistical fashion. This is referred to as the multi-model ensemble (MME) approach: the MME mean prediction is often more skillful as model errors tend to average out <xref ref-type="bibr" rid="bib1.bibx25" id="paren.2"/>, whereas the spread between the model predictions is naturally interpreted as a measure of the uncertainty about the mean <xref ref-type="bibr" rid="bib1.bibx10" id="paren.3"/>. Although MME tends to improve predictions of climate statistics (i.e., mean and variance), a major drawback is that it is not designed to produce an improved trajectory that can be seen as a specific climate forecast, given that averaging uncorrelated climate trajectories from different models leads to variance reduction and smoothing.</p>
      <p id="d1e168">The foundation of modern weather and climate prediction rests on the assumption that when an estimate of the climate state is at disposal at a particular instance in time, its time evolution can be calculated by a proper application of a numerical discretization of the fundamental laws of physics, supplemented by empirical relationships describing unresolved scales and a complete specification of the external forcing and boundary conditions. Integration in time subsequently yields a predicted climate trajectory into the future and formally frames the climate prediction endeavor as a mixed initial and boundary conditions problem <xref ref-type="bibr" rid="bib1.bibx5 bib1.bibx9" id="paren.4"><named-content content-type="pre">see, e.g.,</named-content></xref>. Initial conditions, but also boundary conditions and external forcing, are usually estimated by combining data with models via data assimilation techniques <xref ref-type="bibr" rid="bib1.bibx3" id="paren.5"><named-content content-type="pre">see, e.g.,</named-content><named-content content-type="post">for a review</named-content></xref>. Errors in the time derivative (i.e., the model error) propagate into errors in the predicted trajectory but model error also affects the model statistics, so that the model and observed mean and variance differ, giving rise to model biases.</p>
      <p id="d1e183">An illustrative example of this propagation of model errors is presented in <xref ref-type="bibr" rid="bib1.bibx15" id="text.6"/> in relation to a change in the model's prescribed aerosol concentrations in the region of the Sahara. Already, within the first few hours of prediction, the different aerosol concentration leads to changing the stability and convection in the region. This in turn changes the upper air divergence and promotes the generation of large-scale Rossby waves that travel horizontally eastward and northward into the Northern Hemisphere during the subsequent week and finally impact the surface air temperatures in Siberia. This example demonstrates that a specific model error can impact model prediction skills on far regions and diverse variables. Furthermore, it suggests that, in order to mitigate or in the best case to prevent model error from growing and affecting the whole model phase space, it is better to intervene at each model computational time step rather than a posteriori by combining outputs after a prediction is completed as in the MME approach.</p>
</sec>
<sec id="Ch1.S1.SS2">
  <label>1.2</label><title>Supermodeling</title>
      <p id="d1e197">Reducing model errors early in the prediction is precisely what supermodeling attempts to achieve <xref ref-type="bibr" rid="bib1.bibx23" id="paren.7"/>. In a supermodel, different models exchange information during the simulation at every time step and form a consensus on a single best prediction. An advantage over the standard MME approach is that the supermodel produces a trajectory with improved long-term statistics. Improved trajectories are extremely valuable for calculations of the impact of climate on society. For instance, crop yields, spread of diseases and river discharge all depend on the specific sequences of weather events, not just on statistics <xref ref-type="bibr" rid="bib1.bibx4 bib1.bibx22 bib1.bibx24" id="paren.8"/>.</p>
      <p id="d1e206">The supermodeling approach was originally developed using low-order dynamical systems <xref ref-type="bibr" rid="bib1.bibx23 bib1.bibx13" id="paren.9"/> and subsequently applied to a global atmosphere model <xref ref-type="bibr" rid="bib1.bibx17 bib1.bibx26" id="paren.10"/> and to a coupled atmosphere–ocean–land model <xref ref-type="bibr" rid="bib1.bibx18" id="paren.11"/>. A partial implementation of the supermodeling concept using real-world observations was presented in <xref ref-type="bibr" rid="bib1.bibx20" id="text.12"/>. In the original supermodeling concept, model equations are connected by nudging terms such that each model in the ensemble is nudged to the state of every other model at every time step. For appropriate connections, the ensemble of models eventually synchronizes on a common solution that depends on the strength of the connections. For instance, if all models are nudged to a particular model that is not nudged to any other model, the ensemble will follow that particular solution. By training connections on observed data, an optimal solution is found that is produced by the connected ensemble of models. This type of supermodel is referred to as connected supermodeling. <xref ref-type="bibr" rid="bib1.bibx26" id="text.13"/> showed that in the limit of strong connections the connected supermodel solution converges to the solution of a weighted superposition of the individual model equations, referred to as a weighted supermodel.</p>
      <p id="d1e224">A crucial step in supermodeling is the training of the connection coefficients (for connected supermodels) or weights (for weighted supermodels) based on data, the observations. The first training schemes of supermodels were based on the minimization of a cost function dependent on long simulations with the supermodel <xref ref-type="bibr" rid="bib1.bibx23 bib1.bibx13 bib1.bibx20" id="paren.14"/>. Given that iterations, and thus many evaluations of the cost function, were necessary in the minimization procedure, this approach turned out to be computationally very expensive. <xref ref-type="bibr" rid="bib1.bibx17" id="text.15"/> developed a computationally very efficient training scheme based on cross pollination in time (CPT), a concept originally introduced by <xref ref-type="bibr" rid="bib1.bibx21" id="text.16"/> in the context of ensemble weather forecasting. In CPT, the models in a multi-model ensemble exchange states during the simulation, generating mixed trajectories that exponentially increase in number in the course of time. As a consequence, a larger area of phase space is explored, thus increasing the chance that the observed trajectory is shadowed within the span of all of the mixed model trajectories. Given the above, CPT training is then based on the selection of the trajectory that remains closest to an observed trajectory. Another alternative efficient approach or training was introduced in <xref ref-type="bibr" rid="bib1.bibx18" id="text.17"/> to learn the connections coefficients in a supermodel. Their method, hereafter referred to as the “synch rule”, is<?pagebreak page791?> based on synchronization and it is inspired by an idea originally proposed in <xref ref-type="bibr" rid="bib1.bibx7" id="text.18"/> for general parameter learning.</p>
      <p id="d1e242">Before supermodeling becomes suitable for the class of large-dimensional state-of-the-art weather and climate models, we need to have training schemes that are computationally suitable for that context. In this paper, we develop, apply and compare CPT and the synch rule to train a weighted supermodel based on the intermediate complexity global coupled atmosphere–ocean–land model SPEEDO <xref ref-type="bibr" rid="bib1.bibx19" id="paren.19"/>. Short-term supermodel prediction skill as well as long-term climate statistics show that both training methods result in supermodels that outperform the individual models. Furthermore, novel experiments with negative weights, as opposed to the standard case of weights larger than or equal to zero, suggest that even when the individual model biases do not compensate for each other an improved supermodel solution can be achieved.</p>
      <p id="d1e249">In Sect. <xref ref-type="sec" rid="Ch1.S2"/>, the two types of supermodels, connected and weighted, are introduced in detail. Section <xref ref-type="sec" rid="Ch1.S3"/> describes the global coupled atmosphere–ocean–land model SPEEDO and the construction of a SPEEDO supermodel. The two training strategies are described in Sect. <xref ref-type="sec" rid="Ch1.S4"/> with specific details when applied to the SPEEDO model in Sect. <xref ref-type="sec" rid="Ch1.S5"/>. The results of the training are shown in Sect. <xref ref-type="sec" rid="Ch1.S6"/>. The final section discusses the results and lists further steps to be taken towards training a supermodel based on state-of-the-art weather and climate models using real-world observations.</p>
</sec>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Weighted and connected supermodeling</title>
      <p id="d1e271">To make the supermodeling approach more explicit, we formally write the model equations of a weather or climate model <inline-formula><mml:math id="M1" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> as
          <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M2" display="block"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
        where <inline-formula><mml:math id="M3" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is a high-dimensional state vector, and <inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> a non-linear evolution function depending on the state <inline-formula><mml:math id="M5" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and on a number of adjustable parameters <inline-formula><mml:math id="M6" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. In practice, weather and climate models generally differ in the representation of the climate state, i.e., the phase where <inline-formula><mml:math id="M7" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is defined, the evolution function and parameter values. In this stage of developing the supermodeling approach and training schemes, we simplify the context and focus on a situation where the models share the same evolution function, <inline-formula><mml:math id="M8" display="inline"><mml:mi mathvariant="bold-italic">f</mml:mi></mml:math></inline-formula>, and the same phase space, so that <inline-formula><mml:math id="M9" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>∈</mml:mo><mml:msup><mml:mi mathvariant="double-struck">R</mml:mi><mml:mi>n</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> for all <inline-formula><mml:math id="M10" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>. However, the models differ in the parameters, <inline-formula><mml:math id="M11" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>≠</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> if <inline-formula><mml:math id="M12" display="inline"><mml:mrow><mml:mi>i</mml:mi><mml:mo>≠</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:math></inline-formula>. The approach can be generalized using data assimilation approaches <xref ref-type="bibr" rid="bib1.bibx6" id="paren.20"/>. We will furthermore denote the “truth” as given by the model <inline-formula><mml:math id="M13" display="inline"><mml:mi mathvariant="bold-italic">f</mml:mi></mml:math></inline-formula> with a specific set of parameters. An ensemble of imperfect models can be dynamically combined in a weighted or connected supermodel.</p><?xmltex \hack{\newpage}?>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Weighted supermodeling</title>
      <p id="d1e455">A weighted supermodel based on two imperfect models is given by

                <disp-formula id="Ch1.E2" specific-use="align" content-type="subnumberedsingle"><mml:math id="M14" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E2.3"><mml:mtd><mml:mtext>2a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E2.4"><mml:mtd><mml:mtext>2b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E2.5"><mml:mtd><mml:mtext>2c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            where <inline-formula><mml:math id="M15" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>∈</mml:mo><mml:msup><mml:mi mathvariant="double-struck">R</mml:mi><mml:mi>n</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> represents the supermodel state vector and diagonal matrices <inline-formula><mml:math id="M16" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="normal">diag</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> with <inline-formula><mml:math id="M17" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>∈</mml:mo><mml:msup><mml:mi mathvariant="double-struck">R</mml:mi><mml:mi>n</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> denote the weights. In the weighted supermodel, the states are imposed to be perfectly synchronized. Training a weighted supermodel implies training the weights <inline-formula><mml:math id="M18" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">w</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.</p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>Connected supermodeling</title>
      <p id="d1e662">For completeness and for comparison of the weighted supermodels with the connected supermodel from <xref ref-type="bibr" rid="bib1.bibx18" id="text.21"/>, we introduce the equations for the connected supermodel. A connected supermodel based on two imperfect models is given by

                <disp-formula id="Ch1.E6" specific-use="align" content-type="subnumberedsingle"><mml:math id="M19" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E6.7"><mml:mtd><mml:mtext>3a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold">C</mml:mi><mml:mn mathvariant="normal">12</mml:mn></mml:msub><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E6.8"><mml:mtd><mml:mtext>3b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold">C</mml:mi><mml:mn mathvariant="normal">21</mml:mn></mml:msub><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E6.9"><mml:mtd><mml:mtext>3c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mn mathvariant="normal">2</mml:mn></mml:mfrac></mml:mstyle><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            Note the nudging terms (the rightmost terms in Eq. <xref ref-type="disp-formula" rid="Ch1.E6.7"/> and <xref ref-type="disp-formula" rid="Ch1.E6.8"/>) that push the state of each model to the state of the other at every time step. The size of the nudging terms <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold">C</mml:mi><mml:mn mathvariant="normal">12</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M21" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold">C</mml:mi><mml:mn mathvariant="normal">21</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reflects the strength of the coupling between the two models. They have the form of diagonal matrices
<inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:mo>∈</mml:mo><mml:msup><mml:mi mathvariant="double-struck">R</mml:mi><mml:mrow><mml:mi>n</mml:mi><mml:mo>×</mml:mo><mml:mi>n</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> and can thus be written as <inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold">C</mml:mi><mml:mn mathvariant="normal">12</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="normal">diag</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">c</mml:mi><mml:mn mathvariant="normal">12</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> with <inline-formula><mml:math id="M24" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">c</mml:mi><mml:mn mathvariant="normal">12</mml:mn></mml:msub><mml:mo>∈</mml:mo><mml:msup><mml:mi mathvariant="double-struck">R</mml:mi><mml:mi>n</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>. The diagonal form reflects the fact that each model state vector component is nudged towards the same component of the other model. The approach can be extended to be multivariate allowing for cross nudging, but this will require careful scaling of the variables. For appropriate connections, the models fall into a synchronized motion <xref ref-type="bibr" rid="bib1.bibx14" id="paren.22"/>. Because in general the synchronization will not be perfect due to the different parameter values, the supermodel solution <inline-formula><mml:math id="M25" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is defined as the average of the different model states. Note that the states will be close for strong connections so that smoothing and loss of variance due to the averaging will be limited. The supermodel solution depends on the relative strengths of connection coefficients. Training a connected supermodel implies training the value of the connection coefficients.</p>
      <p id="d1e934">A connected supermodel allows for more flexibility in the event that the ensemble is not perfectly synchronized <xref ref-type="bibr" rid="bib1.bibx27" id="paren.23"/>. In regions of phase space of strong divergence, for instance, one model can pull the ensemble along if it takes a very different trajectory. However, in <xref ref-type="bibr" rid="bib1.bibx27" id="text.24"/>, it is noted that the size of the connection coefficients after training is typically quite large.<?pagebreak page792?> The larger the coefficients, the stronger the models converge on a synchronized trajectory, which can be described by a weighted superposition of the models <xref ref-type="bibr" rid="bib1.bibx27" id="paren.25"/>. Since for some training applications, perfect synchronization is required as we shall see in Sect. <xref ref-type="sec" rid="Ch1.S4"/>, only weighted supermodels are considered in this paper. We do not limit ourselves to combining only two imperfect models into a supermodel; also, combining four imperfect models will be discussed.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>SPEEDO climate model</title>
      <p id="d1e957">The SPEEDO global climate model consists of an atmospheric component (SPEEDY) that exchanges information with a land (LBM) and an ocean–sea-ice component (CLIO) using coupling routines (Fig. <xref ref-type="fig" rid="Ch1.F1"/>). The coupling routines perform re-gridding operations between the computational grids of the different modules. A detailed description of SPEEDO can be found in <xref ref-type="bibr" rid="bib1.bibx19 bib1.bibx18" id="text.26"/>.</p>
      <p id="d1e965">The atmospheric model SPEEDY   describes the evolution of the two horizontal wind components <inline-formula><mml:math id="M26" display="inline"><mml:mi>U</mml:mi></mml:math></inline-formula> (east–west) and <inline-formula><mml:math id="M27" display="inline"><mml:mi>V</mml:mi></mml:math></inline-formula> (north–south), temperature <inline-formula><mml:math id="M28" display="inline"><mml:mi>T</mml:mi></mml:math></inline-formula> and specific humidity <inline-formula><mml:math id="M29" display="inline"><mml:mi>q</mml:mi></mml:math></inline-formula> at eight vertical levels and the surface pressure <inline-formula><mml:math id="M30" display="inline"><mml:mrow><mml:msub><mml:mi>p</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. Relatively simple calculations of heating and cooling rates due to radiation, convective transports, cloud amounts, precipitation and turbulent heat, water and momentum exchange at the surface are performed at a computational grid of approximately 3.75<inline-formula><mml:math id="M31" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> horizontal spacing (<inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:mn mathvariant="normal">48</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">96</mml:mn></mml:mrow></mml:math></inline-formula> grid cells).</p>
      <p id="d1e1029">SPEEDY exchanges water and heat with the land model LBM that uses three soil layers and up to two snow layers to close the hydrological cycle over land and a heat budget equation that controls the land temperatures. The horizontal discretization is the same as for the atmosphere model. The land surface reflection coefficient for solar radiation is prescribed using a monthly climatology. Each land bucket has a maximum soil water capacity. The runoff is collected in river basins and drained into the ocean at specific locations of the major river outflows.</p>
      <p id="d1e1032">SPEEDY exchanges heat, water and momentum with the ocean model CLIO <xref ref-type="bibr" rid="bib1.bibx8" id="paren.27"/>. CLIO describes the evolution of ocean currents, temperature and salinity on a computational grid of 3<inline-formula><mml:math id="M33" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> horizontal resolution and 20 unevenly spaced layers in the vertical. A three-layer thermodynamic–dynamic sea-ice model describes the evolution of sea ice in the event that ocean temperatures drop below freezing levels. Heat storage in the snow–ice system is accounted for and snow amounts and ice thickness evolve in response to surface and bottom heat fluxes. Sea ice is considered to behave as a viscous–plastic continuum as it moves under the action of winds and ocean currents.</p>
      <p id="d1e1048">Formally, the SPEEDO equations can be written as
<?xmltex \hack{\newpage}?><?xmltex \hack{\vspace*{-6mm}}?>

              <disp-formula id="Ch1.E10" specific-use="align" content-type="subnumberedsingle"><mml:math id="M34" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E10.11"><mml:mtd><mml:mtext>4a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msup></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E10.12"><mml:mtd><mml:mtext>4b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E10.13"><mml:mtd><mml:mtext>4c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi>l</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi>l</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where <inline-formula><mml:math id="M35" display="inline"><mml:mi mathvariant="bold-italic">a</mml:mi></mml:math></inline-formula> is the atmospheric state vector, <inline-formula><mml:math id="M36" display="inline"><mml:mi mathvariant="bold-italic">o</mml:mi></mml:math></inline-formula> the ocean/sea-ice state vector, <inline-formula><mml:math id="M37" display="inline"><mml:mi mathvariant="bold-italic">l</mml:mi></mml:math></inline-formula> the land state vector, <inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> the heat exchange vector between atmosphere and surface, <inline-formula><mml:math id="M39" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> the water exchange vector, <inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> the momentum exchange vector and <inline-formula><mml:math id="M41" display="inline"><mml:mi mathvariant="bold-italic">r</mml:mi></mml:math></inline-formula> the river outflow vector describing the flow of water from land to ocean. The exchange vectors depend on the state of the atmosphere and the surface but this dependency is not made explicit in Eq. (4) to simplify the notation. The projection operators <inline-formula><mml:math id="M42" display="inline"><mml:mi mathvariant="script">P</mml:mi></mml:math></inline-formula> represent the regridding operations between the computational grids. These operations are conservative so that the globally integrated heat and water loss of the atmosphere at any time at the surface equals the integrated heat and water gain of the land and ocean. The non-linear functions <inline-formula><mml:math id="M43" display="inline"><mml:mi mathvariant="bold-italic">f</mml:mi></mml:math></inline-formula> represent the cumulative contribution of the modeled physical processes to the change in the climate state vector and depend on the values of the parameter vectors <inline-formula><mml:math id="M44" display="inline"><mml:mi mathvariant="bold-italic">p</mml:mi></mml:math></inline-formula>. Some of these parameters go through a daily and/or seasonal cycle and/or have a spatial dependence like the reflectivity of the surface. The non-linear functions <inline-formula><mml:math id="M45" display="inline"><mml:mi mathvariant="bold-italic">g</mml:mi></mml:math></inline-formula> describe how the exchange of heat, water and momentum between the subsystems affects the change of the climate state vector.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F1"><?xmltex \currentcnt{1}?><label>Figure 1</label><caption><p id="d1e1366">Schematic representation of the SPEEDO climate model. The atmosphere needs surface characteristics (temperature, roughness, reflectivity, soil moisture) in order to calculate the exchange of heat, water and momentum. Coupler software communicates this information between the components and interpolates between the computational grids.</p></caption>
        <?xmltex \igopts{width=113.811024pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f01.png"/>

      </fig>

<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>SPEEDO supermodel</title>
      <p id="d1e1382">The training experiments of this study are evaluated in a noise-free observation framework, with perfect observations generated by sampling a reference model trajectory. This “perfect model” provides a set of time-ordered observations,<?pagebreak page793?> called the “truth”. We consider the SPEEDO climate model with standard parameter values as truth and create imperfect models by perturbing parameter values in the atmospheric component. A supermodel is formed by combining the imperfect atmosphere models through a weighted superposition of the time derivatives of the imperfect models (Eq. 2) which are each coupled to the same ocean and land model (Fig. <xref ref-type="fig" rid="Ch1.F2"/>). All atmosphere models receive the same state information from the ocean and land model but each calculates their own water, heat and momentum exchange. On the other hand, the ocean and land model receive the multi-model weighted average of these atmospheric components; this follows the interactive ensemble approach <xref ref-type="bibr" rid="bib1.bibx11" id="paren.28"/>. Following Eq. (2), the SPEEDO weighted supermodel equations are given by

                <disp-formula id="Ch1.E14" specific-use="align" content-type="subnumberedsingle"><mml:math id="M46" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E14.15"><mml:mtd><mml:mtext>5a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>;</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msubsup></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E14.16"><mml:mtd><mml:mtext>5b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mover accent="true"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup></mml:mrow><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mover accent="true"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup></mml:mrow><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mover accent="true"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msup></mml:mrow><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E14.17"><mml:mtd><mml:mtext>5c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi>l</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi>l</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:mover accent="true"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">h</mml:mi></mml:msup></mml:mrow><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:mover accent="true"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:msup></mml:mrow><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E14.18"><mml:mtd><mml:mtext>5d</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:munder><mml:mo movablelimits="false">∑</mml:mo><mml:mi>i</mml:mi></mml:munder><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            where <inline-formula><mml:math id="M47" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> denotes the time derivative of the supermodel, <inline-formula><mml:math id="M48" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> values denote diagonal matrices with weights on the diagonal, <inline-formula><mml:math id="M49" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> refers to imperfect model <inline-formula><mml:math id="M50" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>, and the overbar denotes a weighted average over the models.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F2"><?xmltex \currentcnt{2}?><label>Figure 2</label><caption><p id="d1e1722">Schematic representation of the SPEEDO climate supermodel based on two imperfect atmosphere models. The two atmosphere models exchange water, heat and momentum with the perfect ocean and land model. The ocean and land models send their state information to both atmosphere models. The atmosphere models exchange state information in order to combine their time derivatives.</p></caption>
          <?xmltex \igopts{width=142.26378pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f02.png"/>

        </fig>

</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Learning methods</title>
      <p id="d1e1740">Two different learning strategies are evaluated in this study in order to train the SPEEDO weighted supermodel: learning based on CPT as developed and applied to low-order dynamical systems in <xref ref-type="bibr" rid="bib1.bibx17" id="text.29"/>, and learning based on synchronization as applied to a connected SPEEDO supermodel in <xref ref-type="bibr" rid="bib1.bibx18" id="text.30"/>.</p><?xmltex \hack{\newpage}?>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>Cross pollination in time</title>
      <p id="d1e1757">The CPT learning approach is based on an idea proposed by <xref ref-type="bibr" rid="bib1.bibx21" id="text.31"/>. CPT “crosses” trajectories of different models in order to create a larger solution space. The aim is to generate trajectories that follow the truth more closely. The training phase of CPT starts from an observed initial condition in state space. For simplicity, assume the model is one-dimensional. From the same initial state, the imperfect models compute one time step each ending in a different state. Next, all models compute another time step from each of these new states. Continuing this process leads to a rapid increase in the number of trajectories with time (Fig. <xref ref-type="fig" rid="Ch1.F3"/>a) that will ultimately cover a larger area of the state space. Among the full set of mixed trajectories, the one which is closest to the truth (i.e., to the data) is continued; the others are discarded, resulting in a pruned ensemble, as is depicted in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F3" specific-use="star"><?xmltex \currentcnt{3}?><label>Figure 3</label><caption><p id="d1e1769">Adapted from <xref ref-type="bibr" rid="bib1.bibx17" id="text.32"/>. A one-dimensional schematic of CPT for three models, a full ensemble <bold>(a)</bold> and a pruned ensemble <bold>(b)</bold>. Note that the “truth” has been drawn here as a continuous line for illustrative purpose. In practice, the truth is only known at discrete times (the observation times) and the distance with respect to model trajectories is computed at those times only.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f03.png"/>

        </fig>

      <p id="d1e1787">In the case of a multi-dimensional model, such as SPEEDO, it is possible that at each time step different models are closest to the truth for different state variables and at different grid locations. In this case, we continue per state variable with the model that is closest. This means that the initial state for the next time step can consist of a combination of models. As the values for the different state variables might not be in agreement with each other, this creates imbalances that can lead to numerical instabilities. A (partial) solution is to decrease the time step, as we shall see in Sect. <xref ref-type="sec" rid="Ch1.S5"/>.</p>
      <p id="d1e1793">The training period is terminated when the CPT trajectory starts to deviate from the truth beyond a given pre-specified threshold. After training, an optimal trajectory is obtained that is produced by a combination of different imperfect models (Fig. <xref ref-type="fig" rid="Ch1.F4"/>). Next, we count how often during training each model has produced the best prediction of a particular component of the state vector. This frequency of occurrences is used to compute weights <inline-formula><mml:math id="M51" display="inline"><mml:mi mathvariant="bold">W</mml:mi></mml:math></inline-formula> for the corresponding time derivative of the state vector. This superposition of weighted imperfect models forms a weighted supermodel, as expressed in the example of Eq. (2).</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F4" specific-use="star"><?xmltex \currentcnt{4}?><label>Figure 4</label><caption><p id="d1e1807">CPT trajectory after a training period of 20 time steps. Model 1 is used for 6 out of 20 time steps; hence, model 1 will get a weight of 0.3.</p></caption>
          <?xmltex \igopts{width=341.433071pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f04.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS2">
  <label>4.2</label><title>Synchronization-based learning</title>
      <p id="d1e1824">For the training of a supermodel based on synchronization, a learning rule (the synch rule) is used that updates the weights such that synchronization errors between truth and supermodel are minimized. In contrast to CPT learning, initial values for the weights need to be chosen and the weights are updated during training. Under certain conditions, the supermodel will fall into synchronized motion with the truth as the weights are updated and the supermodel is nudged to the truth (black arrows in Fig. <xref ref-type="fig" rid="Ch1.F5"/>).</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F5" specific-use="star"><?xmltex \currentcnt{5}?><label>Figure 5</label><caption><p id="d1e1831">At each observation (dots) of the truth (continuous black line), the weights of the imperfect models (red, blue) are updated which gives a new supermodel solution (green dotted line). The black arrows indicate the nudging to the truth.</p></caption>
          <?xmltex \igopts{width=341.433071pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f05.png"/>

        </fig>

      <?pagebreak page794?><p id="d1e1840">The synch rule for the weights is an application of the general synchronization-based parameter estimation approach suggested in <xref ref-type="bibr" rid="bib1.bibx7" id="paren.33"/>. Recently, the synch rule was applied to train the connections in a connected SPEEDO supermodel <xref ref-type="bibr" rid="bib1.bibx18" id="paren.34"/>. We follow a similar strategy and implement the synch rule to train the weights of a weighted SPEEDO supermodel.</p>
      <p id="d1e1850">In the context of two dynamical systems that differ in parameter values only, the general synch rule for parameter estimation is given by

                <disp-formula id="Ch1.E19" specific-use="align" content-type="subnumberedsingle"><mml:math id="M52" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E19.20"><mml:mtd><mml:mtext>6a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>;</mml:mo><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E19.21"><mml:mtd><mml:mtext>6b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo>;</mml:mo><mml:mi mathvariant="bold-italic">q</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi mathvariant="bold">K</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E19.22"><mml:mtd><mml:mtext>6c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi>q</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi>j</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:munder><mml:mo movablelimits="false">∑</mml:mo><mml:mi>i</mml:mi></mml:munder><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">q</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>q</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            where <inline-formula><mml:math id="M53" display="inline"><mml:mi mathvariant="bold-italic">p</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M54" display="inline"><mml:mi mathvariant="bold-italic">q</mml:mi></mml:math></inline-formula> are vectors of parameters.  <inline-formula><mml:math id="M55" display="inline"><mml:mrow><mml:mi mathvariant="bold">K</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is a connecting term between the two systems that nudges <inline-formula><mml:math id="M56" display="inline"><mml:mi mathvariant="bold-italic">y</mml:mi></mml:math></inline-formula> towards <inline-formula><mml:math id="M57" display="inline"><mml:mi mathvariant="bold-italic">x</mml:mi></mml:math></inline-formula>. <inline-formula><mml:math id="M58" display="inline"><mml:mi mathvariant="bold">K</mml:mi></mml:math></inline-formula> is a diagonal matrix of nudging coefficients, <inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:mi mathvariant="bold">K</mml:mi><mml:mo>=</mml:mo><mml:mi mathvariant="normal">diag</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">k</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. Suppose the two systems (Eq. <xref ref-type="disp-formula" rid="Ch1.E19.20"/> and <xref ref-type="disp-formula" rid="Ch1.E19.21"/>) synchronize if <inline-formula><mml:math id="M60" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mo>=</mml:mo><mml:mi mathvariant="bold-italic">q</mml:mi></mml:mrow></mml:math></inline-formula>;  that is, as <inline-formula><mml:math id="M61" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>→</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M62" display="inline"><mml:mrow><mml:mtext mathvariant="bold">y</mml:mtext><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>→</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. We further assume that the parameters appear only linearly in the model equations. Then it can be proven that, using the learning rule (Eq. <xref ref-type="disp-formula" rid="Ch1.E19.22"/>), even if the two systems are not identical, <inline-formula><mml:math id="M63" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mo>≠</mml:mo><mml:mi mathvariant="bold-italic">q</mml:mi></mml:mrow></mml:math></inline-formula>, the systems will still synchronize and the parameters will become equal, <inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">q</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>→</mml:mo><mml:mi mathvariant="bold-italic">p</mml:mi></mml:mrow></mml:math></inline-formula> as <inline-formula><mml:math id="M65" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>→</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow></mml:math></inline-formula>. Here, <inline-formula><mml:math id="M66" display="inline"><mml:mrow><mml:msub><mml:mi>q</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> denotes the parameter values, with <inline-formula><mml:math id="M67" display="inline"><mml:mi>j</mml:mi></mml:math></inline-formula> indexing the elements of the parameter vector. Furthermore, <inline-formula><mml:math id="M68" display="inline"><mml:mrow><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> denotes the synchronization error at the current time step with <inline-formula><mml:math id="M69" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> indexing the elements of the state vector and <inline-formula><mml:math id="M70" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> an adjustable rate of learning scaling factor. At every time step, the update <inline-formula><mml:math id="M71" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>q</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> for the weight <inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:msub><mml:mi>q</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is calculated.</p>
      <p id="d1e2253">In training a supermodel, we assume that the truth can be described by a weighted dynamical combination of imperfect models with the weights as adjustable parameters. In this case, the function <inline-formula><mml:math id="M73" display="inline"><mml:mi mathvariant="bold-italic">f</mml:mi></mml:math></inline-formula> corresponds to the supermodel time derivative, <inline-formula><mml:math id="M74" display="inline"><mml:mi mathvariant="bold-italic">q</mml:mi></mml:math></inline-formula> corresponds to the weights of the supermodel, <inline-formula><mml:math id="M75" display="inline"><mml:mi mathvariant="bold-italic">x</mml:mi></mml:math></inline-formula> denotes the truth and <inline-formula><mml:math id="M76" display="inline"><mml:mi mathvariant="bold-italic">y</mml:mi></mml:math></inline-formula> the supermodel solution. The derivative of <bold>f</bold> with respect to a certain weight is the tendency of the imperfect model belonging to that weight (see Eq. <xref ref-type="disp-formula" rid="Ch1.E2.5"/>). In our SPEEDO case, the truth cannot exactly be described as a weighted superposition of imperfect<?pagebreak page795?> models since the perturbed parameters do not appear linearly in the equations, yet the approximation is close enough for the learning rule to work well.</p>
      <p id="d1e2290">Integration of the synch rule implies that as long as the time series of the synchronization error <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and the effect of the parameter on the imperfect model evolution <inline-formula><mml:math id="M78" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">y</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">q</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>q</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula> are correlated, the parameter will be updated. For instance, when a parameter update systematically enhances warming in the model when the model is colder than the truth and the same holds when the parameter update systematically cools the model when it is too warm, then the updated parameter will decrease the synchronization error between the model and truth over time. When this correlation vanishes, there is no systematic relation anymore between updating the parameter and the state of the model; then systematic updates cease. When perfect synchronization is reached (hence, <inline-formula><mml:math id="M79" display="inline"><mml:mrow><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>), naturally updates also stop.</p>
</sec>
</sec>
<sec id="Ch1.S5">
  <label>5</label><title>Training in SPEEDO</title>
      <p id="d1e2362">In training the SPEEDO supermodel, we regard the atmospheric model with standard parameter values as truth, whereas imperfect atmospheric models are created by perturbing those parameter values. Figure <xref ref-type="fig" rid="Ch1.F6"/> depicts the configuration during training. All atmosphere models are independently coupled to the same ocean and land model. They each calculate their own water, heat and momentum fluxes and receive the information from the ocean and the land model from the truth only.</p>
      <p id="d1e2367">During training, the truth and imperfect models all share their states. In the case of CPT, this state information is used by each imperfect model to check which model is closest to the truth and continue the integration from that state. In the case of the synch rule, this state information is used to calculate the synchronization error between the supermodel and the truth.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F6"><?xmltex \currentcnt{6}?><label>Figure 6</label><caption><p id="d1e2372">Schematic representation of the SPEEDO system during training <xref ref-type="bibr" rid="bib1.bibx18" id="paren.35"/>.</p></caption>
        <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f06.png"/>

      </fig>

      <p id="d1e2385">Application of the synch rule to a weighted SPEEDO supermodel of two imperfect models implies integration of the following set of equations:
<?xmltex \hack{\newpage}?><?xmltex \hack{\vspace*{-6mm}}?>

              <disp-formula id="Ch1.E23" specific-use="align" content-type="subnumberedsingle"><mml:math id="M80" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E23.24"><mml:mtd><mml:mtext>7a</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>;</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">a</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">m</mml:mi></mml:msubsup></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.25"><mml:mtd><mml:mtext>7b</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>;</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="normal">a</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="normal">m</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:mi mathvariant="bold">K</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.26"><mml:mtd><mml:mtext>7c</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>;</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="normal">a</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="normal">m</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:mi mathvariant="bold">K</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.27"><mml:mtd><mml:mtext>7d</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi mathvariant="bold-italic">o</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">m</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msup><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.28"><mml:mtd><mml:mtext>7e</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">f</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:mi mathvariant="bold-italic">l</mml:mi><mml:mo>;</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">p</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">g</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">h</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="script">P</mml:mi><mml:mi mathvariant="normal">l</mml:mi></mml:msup><mml:msubsup><mml:mi mathvariant="bold-italic">e</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">w</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">r</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.29"><mml:mtd><mml:mtext>7f</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="bold">W</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.30"><mml:mtd><mml:mtext>7g</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mover accent="true"><mml:mi>W</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mrow><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi>a</mml:mi><mml:mrow><mml:mi mathvariant="normal">s</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mrow><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfenced><mml:msub><mml:mover accent="true"><mml:mi>a</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mrow><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where index 0 refers to the truth and <inline-formula><mml:math id="M81" display="inline"><mml:mrow><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> refers to the weight of model <inline-formula><mml:math id="M82" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> and state vector element <inline-formula><mml:math id="M83" display="inline"><mml:mi>j</mml:mi></mml:math></inline-formula>. During training, we choose a uniform nudging strength corresponding to a 24 h timescale, as motivated by <xref ref-type="bibr" rid="bib1.bibx18" id="text.36"/>. They showed that, for this value of <inline-formula><mml:math id="M84" display="inline"><mml:mi mathvariant="bold">K</mml:mi></mml:math></inline-formula>, two connected identical SPEEDO models (perfect model scenario) almost perfectly synchronize with very small synchronization errors in temperatures of the order of 0.01 <inline-formula><mml:math id="M85" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C. Imperfect models, on the other hand, have synchronization errors with respect to the truth that are usually 10 times larger.</p><?xmltex \hack{\newpage}?>
<?pagebreak page796?><sec id="Ch1.S5.SS1">
  <label>5.1</label><title>Construction of imperfect models</title>
      <p id="d1e2991">In order to be able to compare results of the weighted supermodels of this study to the connected supermodels in <xref ref-type="bibr" rid="bib1.bibx18" id="text.37"/>, we choose the same parameter values for the imperfect models. These parameters are the convection relaxation timescale, the relative humidity threshold and the momentum diffusion timescale. The reason to perturb these parameters is because the uncertainty in climate models mostly lies in the parameterization of clouds and convection, and perturbing these parameters in the SPEEDO model results in a spread in the simulated climate that characterizes this uncertainty. The parameters are listed in Table <xref ref-type="table" rid="Ch1.T1"/>, where model 1 and model 2 correspond to the imperfect models of <xref ref-type="bibr" rid="bib1.bibx18" id="text.38"/>. The impact of the parameter perturbations on the climate (i.e., long-term behavior) of the models is assessed on the basis of 40-year simulations initiated on 1 January  of model year 2001 of a long control simulation as in <xref ref-type="bibr" rid="bib1.bibx18" id="text.39"/>. Table <xref ref-type="table" rid="Ch1.T2"/> shows the global mean average difference between the truth and the imperfect models of Table <xref ref-type="table" rid="Ch1.T1"/> for different variables. From the table, it appears evident how the imperfect models all drift away from the truth giving rise to biases. For example, the global mean temperature of imperfect model 1 rises about 1.4 <inline-formula><mml:math id="M86" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C  within a couple of decades, whereas model 2 cools around 0.4 <inline-formula><mml:math id="M87" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C. These global mean temperature biases are comparable to the biases of state-of-the-art global climate models compared to real-world observations <xref ref-type="bibr" rid="bib1.bibx10" id="paren.40"/>.</p>
      <p id="d1e3031">The first supermodel that we will train will consist of a weighted superposition of models 1 and 2. The second supermodel will consist of a weighted superposition of models 1, 3, 4 and 5. The parameter values of these models are chosen such that they form a so-called convex hull around the true parameter values (see <xref ref-type="bibr" rid="bib1.bibx17" id="altparen.41"/> for a discussion on the convex hull principle). Note that we use only two perturbed values for each parameter; the imperfect models differ only in the combination of these values, such that in the four-model supermodel, a convex hull is formed. This implies that, provided the model functional dependence on the parameters is linear, the true parameter values can be obtained as a linear combination with positive coefficients/weights of the four parameter values of the imperfect models. While these conditions do not perfectly hold in this case, we expect that we can create a weighted supermodel based on these four models that will be close to the truth.  All of these four models overestimate the global mean temperature and precipitation (Table <xref ref-type="table" rid="Ch1.T2"/>). Therefore, simply taking the MME mean with positive weights will not produce a climatology closer to the truth. However, we expect that, based on the convex hull principle, the weighted supermodel will nevertheless be able to produce a climatology that is closer to the truth.</p>
      <p id="d1e3039">The third supermodel consists of a weighted superposition of models 1 and 6. In this case, both imperfect models have parameter values that are smaller than the corresponding true values. A weighted superposition with positive weights does not correspond to a model with parameter values that are closer to the truth. Note that both models overestimate the average temperature and precipitation (Table <xref ref-type="table" rid="Ch1.T2"/>); hence, taking the MME mean with positive weights also does not produce a climatology closer to the truth. In this case, we will explore whether a weighted supermodel with negative weights can be trained in order to improve the climatology and short-term forecasts.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T1"><?xmltex \currentcnt{1}?><label>Table 1</label><caption><p id="d1e3048">Parameter values of perfect and imperfect models.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="4">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="center"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2">Convection</oasis:entry>
         <oasis:entry colname="col3">Relative</oasis:entry>
         <oasis:entry colname="col4">Momentum</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2">relaxation</oasis:entry>
         <oasis:entry colname="col3">humidity</oasis:entry>
         <oasis:entry colname="col4">diffusion</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2">timescale</oasis:entry>
         <oasis:entry colname="col3">threshold</oasis:entry>
         <oasis:entry colname="col4">timescale</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Perfect</oasis:entry>
         <oasis:entry colname="col2">6 h</oasis:entry>
         <oasis:entry colname="col3">0.9</oasis:entry>
         <oasis:entry colname="col4">24 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">4 h</oasis:entry>
         <oasis:entry colname="col3">0.85</oasis:entry>
         <oasis:entry colname="col4">18 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 2</oasis:entry>
         <oasis:entry colname="col2">8 h</oasis:entry>
         <oasis:entry colname="col3">0.95</oasis:entry>
         <oasis:entry colname="col4">30 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 3</oasis:entry>
         <oasis:entry colname="col2">4 h</oasis:entry>
         <oasis:entry colname="col3">0.95</oasis:entry>
         <oasis:entry colname="col4">30 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 4</oasis:entry>
         <oasis:entry colname="col2">8 h</oasis:entry>
         <oasis:entry colname="col3">0.95</oasis:entry>
         <oasis:entry colname="col4">18 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 5</oasis:entry>
         <oasis:entry colname="col2">8 h</oasis:entry>
         <oasis:entry colname="col3">0.85</oasis:entry>
         <oasis:entry colname="col4">30 h</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 6</oasis:entry>
         <oasis:entry colname="col2">3 h</oasis:entry>
         <oasis:entry colname="col3">0.75</oasis:entry>
         <oasis:entry colname="col4">14 h</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T2" specific-use="star"><?xmltex \currentcnt{2}?><label>Table 2</label><caption><p id="d1e3223">Global mean average difference between the imperfect models and the perfect model, calculated over the last 30 years of the simulation.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="7">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:colspec colnum="7" colname="col7" align="right"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2">Temperature</oasis:entry>
         <oasis:entry colname="col3">Precipitation</oasis:entry>
         <oasis:entry colname="col4">Wind at</oasis:entry>
         <oasis:entry colname="col5">Wind at</oasis:entry>
         <oasis:entry colname="col6">Solar</oasis:entry>
         <oasis:entry colname="col7">Cloud cover</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M88" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula><inline-formula><mml:math id="M89" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C<inline-formula><mml:math id="M90" display="inline"><mml:mo>)</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M91" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>mm d<inline-formula><mml:math id="M92" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">200 hPa</oasis:entry>
         <oasis:entry colname="col5">850 hPa</oasis:entry>
         <oasis:entry colname="col6">surface</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M93" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>%<inline-formula><mml:math id="M94" display="inline"><mml:mo>)</mml:mo></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M95" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>m s<inline-formula><mml:math id="M96" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M97" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>m s<inline-formula><mml:math id="M98" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6">radiation</oasis:entry>
         <oasis:entry colname="col7"/>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4"/>
         <oasis:entry colname="col5"/>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M99" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>W m<inline-formula><mml:math id="M100" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col7"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">1.37</oasis:entry>
         <oasis:entry colname="col3">0.11</oasis:entry>
         <oasis:entry colname="col4">1.04</oasis:entry>
         <oasis:entry colname="col5">0.07</oasis:entry>
         <oasis:entry colname="col6">2.06</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M101" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.59</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 2</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M102" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.38</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M103" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.04</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M104" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.31</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M105" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.03</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.13</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col7">0.87</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 3</oasis:entry>
         <oasis:entry colname="col2">0.99</oasis:entry>
         <oasis:entry colname="col3">0.10</oasis:entry>
         <oasis:entry colname="col4">1.14</oasis:entry>
         <oasis:entry colname="col5">0.06</oasis:entry>
         <oasis:entry colname="col6">1.21</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M107" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.03</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 4</oasis:entry>
         <oasis:entry colname="col2">0.45</oasis:entry>
         <oasis:entry colname="col3">0.04</oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M108" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.04</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M109" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M110" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.20</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col7">0.10</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 5</oasis:entry>
         <oasis:entry colname="col2">0.86</oasis:entry>
         <oasis:entry colname="col3">0.08</oasis:entry>
         <oasis:entry colname="col4">0.72</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M111" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M112" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.19</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M113" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.12</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 6</oasis:entry>
         <oasis:entry colname="col2">3.20</oasis:entry>
         <oasis:entry colname="col3">0.26</oasis:entry>
         <oasis:entry colname="col4">2.25</oasis:entry>
         <oasis:entry colname="col5">0.03</oasis:entry>
         <oasis:entry colname="col6">3.95</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">3.37</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S5.SS2">
  <label>5.2</label><title>Global weights</title>
      <p id="d1e3724">For both CPT and the synch rule, we choose to work with global weights, which means that for each meteorological variable we use the same weight at every grid point. In principle, one could allow different weights per each grid point but it could induce dynamic imbalances that pull the model away from its attractor. The model's reaction is then to restore the dynamical balances and return to its own attractor <xref ref-type="bibr" rid="bib1.bibx14" id="paren.42"/>. In SPEEDO, this leads to the generation of fast gravity waves and fast convective adjustments. An adequately small time step is required in order to prevent numerical instabilities. We choose instead to use global weights in order to limit the computational time.</p>
</sec>
<sec id="Ch1.S5.SS3">
  <label>5.3</label><title>Exchange of state information</title>
      <p id="d1e3738">The SPEEDO model has five prognostic variables: temperature, vorticity, divergence, specific humidity and surface pressure (<inline-formula><mml:math id="M115" display="inline"><mml:mi>T</mml:mi></mml:math></inline-formula>, VOR, DIV, TR, PS). Best results were obtained by limiting the weighted averaging of state information to temperature, vorticity and divergence only. We suspect that exchanging specific humidity and surface pressure leads to imbalances and fast spurious adjustments that deteriorate the supermodel solution. We found that a perfect  SPEEDY atmosphere only fully synchronizes with the truth when at least temperature, vorticity and divergence are nudged to the truth (not shown). Therefore, in a weighted supermodel, at least these variables need to be exchanged.</p><?xmltex \hack{\newpage}?>
</sec>
<?pagebreak page797?><sec id="Ch1.S5.SS4">
  <label>5.4</label><title>Required time step</title>
      <p id="d1e3757">We found that smaller time steps were required during CPT training as compared to standard integrations. Gravity waves induced by the state replacement during training require a smaller time step in order to prevent numerical instabilities. We found that a 15 min time step was sufficient with our choice of imperfect models, which is half the time step of the standard integration.</p>
</sec>
<sec id="Ch1.S5.SS5">
  <label>5.5</label><title>Initialization of the weights for the synch rule</title>
      <p id="d1e3769">In CPT training, the sum of the weights is normalized to 1. In the application of the synch rule, on the other hand, the sum of the weights is not explicitly constrained. One can start from zero weights and let the synch rule find the optimal set of weights. Initializing weights with a sum larger than 1 easily leads to numerical instabilities because the weighted mean state becomes more energetic. Imposing the constraint of the sum of weights being 1 during the training also led to numerical instabilities. We chose to initialize with equal weights that sum to 1.</p>
</sec>
<sec id="Ch1.S5.SS6">
  <label>5.6</label><title>Rate of learning in the synch rule</title>
      <p id="d1e3780">The synch rule contains an adjustable rate of learning scaling factor <inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, with <inline-formula><mml:math id="M117" display="inline"><mml:mi>j</mml:mi></mml:math></inline-formula> the index of the state vector. A large rate of learning is desirable since it leads to faster convergence and shorter training periods. However, the parameters should vary on a slower timescale than the dynamical variables and this provides an upper bound for the value of <inline-formula><mml:math id="M118" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. Furthermore, it turns out that if <inline-formula><mml:math id="M119" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is too large, the sum of the weights can become greater than 1, which easily leads to numerical instabilities. The size of <inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> in the synch rule depends on the variable that is being exchanged and was determined by trial and error during the training experiments. The largest values for <inline-formula><mml:math id="M121" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> that resulted in converged weights were on the order of 10<inline-formula><mml:math id="M122" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">7</mml:mn></mml:msup></mml:math></inline-formula> for divergence and vorticity and 10<inline-formula><mml:math id="M123" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> for temperature. With these scaling factors, approximately similar rates of learning were achieved for the different variables. This makes sense since the state values for divergence and vorticity are much smaller than for temperature, so the product of <inline-formula><mml:math id="M124" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and the state values in the synch rule is of the same order of magnitude.</p>
</sec>
<sec id="Ch1.S5.SS7">
  <label>5.7</label><title>Weights for the heat, water and momentum fluxes</title>
      <p id="d1e3886">In the experimental setup during training, we assume a perfect ocean and land models which receive fluxes from the perfect atmosphere. However, in the supermodel setup, perfect fluxes are not available and we use a weighted combination of the fluxes from both imperfect models instead. In the connected supermodel of <xref ref-type="bibr" rid="bib1.bibx18" id="text.43"/>, the fluxes are averaged using equal weights. In this paper, we further optimize the weights for the fluxes, because we found they have a big influence on the supermodel's performance. In particular, we selected weights given by the average of the three weights for the prognostic variables. To check whether this choice was optimal, we used a least squares minimization method in order to optimize the weights for the fluxes. During 1 year of training, the fluxes from the perfect and imperfect models were saved at every time step. The weights were determined by a least squares fit of a weighted sum of the imperfect fluxes to the perfect fluxes. The flux weights obtained from the minimization method did differ slightly per flux (heat, water or momentum flux), but the average weights were close to the average of the weights for the prognostic variables.</p>
</sec>
</sec>
<sec id="Ch1.S6">
  <label>6</label><title>Results</title>
      <p id="d1e3901">We describe the learning results and the forecast short- and long-term capabilities of the three supermodel configurations separately.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F7" specific-use="star"><?xmltex \currentcnt{7}?><label>Figure 7</label><caption><p id="d1e3906">Calculation of weights for a supermodel constructed from two imperfect models using two different training schemes. <bold>(a)</bold> CPT weights calculated during a training period of 1 week estimated for each week of a year. <bold>(b)</bold> Weights for the synch rule during a training period of 1 year.</p></caption>
        <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f07.png"/>

      </fig>

<sec id="Ch1.S6.SS1">
  <label>6.1</label><title>Supermodels based on two imperfect models</title>
      <?pagebreak page798?><p id="d1e3928">We first trained a weighted supermodel based on imperfect models 1 and 2 (see Table <xref ref-type="table" rid="Ch1.T1"/>), applying both CPT and the synch rule. As a benchmark, we compare the quality of the weighted supermodel after training with the connected SPEEDO supermodel of <xref ref-type="bibr" rid="bib1.bibx18" id="text.44"/>. This supermodel is based on the same imperfect models and was trained by the synch rule.</p>
      <p id="d1e3936">Ideally, both CPT and the synch rule should produce converged weights, i.e., weights that remain stable if the training period is extended. The required length of the training period for the convergence of the two methods turns out to be very different. For CPT, a training period as short as a couple of days produces converged weights, whereas for the synch rule it takes about a year. Note that we limit the CPT training period to a week, as the CPT trajectory starts to deviate significantly from the truth after approximately 10 days. The reason that CPT diverges from the truth is because we have a limited ensemble size. With non-linear processes causing rapid error growth, the truth soon falls outside the limited ensemble. The problem is exacerbated by replacing a model state with state variables mixed from different models which introduces imbalances that cause additional error growth.</p>
      <p id="d1e3939">In order to check the difference between the CPT weights during a year, the CPT method is applied for each week during 1 year. After each week, the values for all prognostic variables are reset to the truth, and the procedure is repeated. Figure <xref ref-type="fig" rid="Ch1.F7"/>a shows the values of the weights during training. The weights for both temperature and vorticity remain fairly constant. The weights for divergence vary within 0.04 of a mean value. For the final supermodel weights, we just take the average over the whole year (Table <xref ref-type="table" rid="Ch1.T3"/>).</p>
      <p id="d1e3946">Using the synch rule, weights for temperature and vorticity converge within the first couple of weeks, whereas for divergence the weights cannot be learned faster than within a year in order to avoid numerical instabilities (see Fig. <xref ref-type="fig" rid="Ch1.F7"/>b). When using the synch rule, the weights converge to similar values as compared to the CPT training (Table <xref ref-type="table" rid="Ch1.T3"/>). Converged values of both methods are within 0.05. Whether these small differences matter for climate and weather forecasts will be assessed in the next two sections. Although not imposed, the training yields sum of weights equal to 1 as an optimal solution.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F8" specific-use="star"><?xmltex \currentcnt{8}?><label>Figure 8</label><caption><p id="d1e3956">Global mean time series for the perfect model, the imperfect models and the two supermodels trained by CPT and the synch rule. The normalized root mean squared error (RMSE) in the climatology of model years 2011–2040 with respect to the climatology of the truth is given in each panel. The normalization is such that the expected value of the perfect model error is 1.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f08.png"/>

        </fig>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T3"><?xmltex \currentcnt{3}?><label>Table 3</label><caption><p id="d1e3968">Weights for the supermodel trained by CPT and the synch rule. Between brackets, the standard deviation over the year (CPT) or the standard deviation over the last 10 weeks of training (synch rule) is given.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.87}[.87]?><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2">Method</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M125" display="inline"><mml:mi>T</mml:mi></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">VOR</oasis:entry>
         <oasis:entry colname="col5">DIV</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">CPT</oasis:entry>
         <oasis:entry colname="col3">0.30 (0.016)</oasis:entry>
         <oasis:entry colname="col4">0.39 (0.007)</oasis:entry>
         <oasis:entry colname="col5">0.35 (0.031)</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Model 2</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.70 (0.016)</oasis:entry>
         <oasis:entry colname="col4">0.61 (0.007)</oasis:entry>
         <oasis:entry colname="col5">0.65 (0.031)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">Synch rule</oasis:entry>
         <oasis:entry colname="col3">0.35 (0.0043)</oasis:entry>
         <oasis:entry colname="col4">0.38 (0.0018)</oasis:entry>
         <oasis:entry colname="col5">0.34 (0.0052)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 2</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.65 (0.0043)</oasis:entry>
         <oasis:entry colname="col4">0.62 (0.0018)</oasis:entry>
         <oasis:entry colname="col5">0.66 (0.0053)</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

<sec id="Ch1.S6.SS1.SSS1">
  <label>6.1.1</label><title>Climate measures</title>
      <?pagebreak page799?><p id="d1e4093">The imperfect models and the supermodel are integrated for 40 years in time, starting from 1 January of model year 2001. The climatology is defined as the average over years 11–40. The error in the climatology is defined as the root of the global mean squared error (RMSE) between the model and the truth. In addition, the perfect model is integrated for 40 years from a slightly perturbed initial condition, in order to obtain an estimate of the sampling error, i.e., to estimate the representativeness of the errors of the different models. Global mean time series for surface air temperature, precipitation, surface solar radiation and cloud cover for the different models show that both weighted supermodels behave very similar and remain close to the perfect model (Fig. <xref ref-type="fig" rid="Ch1.F8"/>). The errors in the climatologies of the various fields of both supermodels are much reduced as compared to both imperfect models and are indistinguishable from the statistical sampling error of the perfect model. Both training methods succeed in greatly improving the simulation of the climate. Compared to the trained connected supermodels of <xref ref-type="bibr" rid="bib1.bibx18" id="text.45"/>, the weighted supermodels have reduced climatological errors (see Fig. <xref ref-type="fig" rid="Ch1.F15"/>). Training of a connected supermodel by the synch rule on the other hand is more efficient as faster learning rates could be used, leading to convergence within 2 weeks of training.</p>
      <p id="d1e4103">A spatial characterization of the performance of the supermodel in simulating the climatology of the zonal wind at 200 hPa is given in Fig. <xref ref-type="fig" rid="Ch1.F9"/>. Clearly, both supermodels outperform the imperfect models and their local errors are of similar magnitude as the sampling error of the perfect model. We computed an optimal weighted average of the climatology of both imperfect models (optimal in the sense that the RMSE in the climatology is minimized) as in <xref ref-type="bibr" rid="bib1.bibx18" id="text.46"/>. This MME mean climatology (Fig. <xref ref-type="fig" rid="Ch1.F9"/>f) has errors of the same order of magnitude as both trained weighted supermodels due to fact that the imperfect model errors are near-mirror images of each other.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F9" specific-use="star"><?xmltex \currentcnt{9}?><label>Figure 9</label><caption><p id="d1e4115">Difference in the zonal wind at 200 hPa averaged over model years 2011–2040 for the various models with respect to the truth. Contours denote areas where the difference is larger than the sampling error at 95 % confidence (solid for positive difference; dotted for negative). Positive values imply stronger mean winds blowing eastward.
Units: m s<inline-formula><mml:math id="M126" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>.</p></caption>
            <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f09.png"/>

          </fig>

      <p id="d1e4137">In the context of simpler models, <xref ref-type="bibr" rid="bib1.bibx17" id="text.47"/> noted that CPT training of a couple of days duration was sufficient to reduce climatological errors substantially, and this result carries over to the complex SPEEDO model used here. This notion that errors in fast processes contribute substantially to errors in the long-term mean state is also supported by other studies, for example, by <xref ref-type="bibr" rid="bib1.bibx16" id="text.48"/>. Since the climatological errors are reduced, we expect the trained supermodels to produce better short-term forecasts as compared to the imperfect models.</p>
</sec>
<sec id="Ch1.S6.SS1.SSS2">
  <label>6.1.2</label><title>Forecast quality</title>
      <p id="d1e4154">In order to assess the quality of short-term forecasts, we initialized the various models from slightly perturbed states of the truth and integrated the models for 2 weeks. We selected 25 initial states, 2 weeks apart, starting 1 January, so the forecasts cover almost 1 year. The quality of the forecast is measured by the RMSE in the global surface air temperature forecast, averaged over the 25 forecasts, and is shown in Fig. <xref ref-type="fig" rid="Ch1.F10"/>. In these forecasts, the atmosphere models are forced by the ocean and land conditions of the truth; this is to exclude error growth related to the coupled interactions. As expected, the RMSE in surface air temperature of the perfect model is the one growing the slowest, and it is still as small as about 0.3 <inline-formula><mml:math id="M127" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C  at day 14. On the other hand, the forecast errors of both imperfect models is 0.3 <inline-formula><mml:math id="M128" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C around day 3 and grow to over 3 <inline-formula><mml:math id="M129" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C at day 14. Both trained weighted supermodels reach 0.3 <inline-formula><mml:math id="M130" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C around day 8 and over 1 <inline-formula><mml:math id="M131" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C at day 14. For comparison, we computed the forecast error of the weighted mean forecast of both imperfect models using the same weights as those used in Fig. <xref ref-type="fig" rid="Ch1.F9"/> in the calculation of the optimal climatology. This MME mean forecast has<?pagebreak page800?> smaller forecast errors than the imperfect models, yet both supermodels are clearly superior.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F10"><?xmltex \currentcnt{10}?><label>Figure 10</label><caption><p id="d1e4209">Forecast quality as measured by the RMSE of the truth and a model with a perturbed initial condition. The control is the difference between the perfect model and the perfect model with a perturbed initial condition.</p></caption>
            <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f10.png"/>

          </fig>

</sec>
</sec>
<sec id="Ch1.S6.SS2">
  <label>6.2</label><title>Supermodels based on four imperfect models forming a convex hull</title>
      <p id="d1e4228">As explained in Sect. <xref ref-type="sec" rid="Ch1.S5.SS1"/>, the parameter perturbations of models 1, 3, 4 and 5 form a convex hull around the true parameter values (Table <xref ref-type="table" rid="Ch1.T1"/>). We therefore expect to be able to create a weighted supermodel based on these four models that will be close to the truth, despite the fact that all four have a warmer climatology than the truth (see Table <xref ref-type="table" rid="Ch1.T2"/>). The weights are trained using both CPT and the synch rule in the same way as in the previous case with two imperfect models and are shown in Fig. <xref ref-type="fig" rid="Ch1.F11"/>; nevertheless, given that the supermodels are now based on four imperfect models, the number of weights is doubled. Again, the weights during CPT training vary from week to week within 0.05 and converge within a year using the synch rule. Weights for vorticity turn out to be a special case, since the change in vorticity as calculated by imperfect models 1 and 3 is equal to, respectively, models 4 and 5. The reason is that only the perturbation in the momentum diffusion timescale affects the vorticity change, and models 1 and 3 have the same diffusion timescale as in models 4 and 5, respectively. Therefore, their weights are equal. Table <xref ref-type="table" rid="Ch1.T4"/> denotes the final supermodel weights, where for vorticity the weight is equally distributed over models 1 and 4 and models 3 and 5.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F11" specific-use="star"><?xmltex \currentcnt{11}?><label>Figure 11</label><caption><p id="d1e4243">CPT weights calculated during a training period of 1 week for 1 year <bold>(a)</bold> and the weights for the synch rule for a training period of 1 year <bold>(b)</bold> with four imperfect models.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f11.png"/>

        </fig>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T4"><?xmltex \currentcnt{4}?><label>Table 4</label><caption><p id="d1e4261">Weights for the supermodel trained by CPT and the synch rule. Between brackets, the standard deviation over the year (CPT) or the standard deviation over the last 10 weeks of training (synch rule) is given.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.87}[.87]?><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2">Method</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M132" display="inline"><mml:mi>T</mml:mi></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">VOR</oasis:entry>
         <oasis:entry colname="col5">DIV</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">CPT</oasis:entry>
         <oasis:entry colname="col3">0.01 (0.005)</oasis:entry>
         <oasis:entry colname="col4">0.19 (0.022)</oasis:entry>
         <oasis:entry colname="col5">0.12 (0.017)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 3</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.33 (0.030)</oasis:entry>
         <oasis:entry colname="col4">0.31 (0.037)</oasis:entry>
         <oasis:entry colname="col5">0.28 (0.024)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 4</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.40 (0.009)</oasis:entry>
         <oasis:entry colname="col4">0.19 (0.022)</oasis:entry>
         <oasis:entry colname="col5">0.41 (0.028)</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Model 5</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.26 (0.026)</oasis:entry>
         <oasis:entry colname="col4">0.31 (0.040)</oasis:entry>
         <oasis:entry colname="col5">0.19 (0.023)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">Synch rule</oasis:entry>
         <oasis:entry colname="col3">0.01 (0.0070)</oasis:entry>
         <oasis:entry colname="col4">0.19 (0.0007)</oasis:entry>
         <oasis:entry colname="col5">0.00 (0.0064)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 3</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.42 (0.0072)</oasis:entry>
         <oasis:entry colname="col4">0.31 (0.0007)</oasis:entry>
         <oasis:entry colname="col5">0.27 (0.0047)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 4</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.37 (0.0065)</oasis:entry>
         <oasis:entry colname="col4">0.19 (0.0007)</oasis:entry>
         <oasis:entry colname="col5">0.44 (0.0044)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 5</oasis:entry>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">0.20 (0.0062)</oasis:entry>
         <oasis:entry colname="col4">0.31 (0.0007)</oasis:entry>
         <oasis:entry colname="col5">0.29 (0.0029)</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

      <?pagebreak page801?><p id="d1e4449"><?xmltex \hack{\newpage}?>The values of the weights for vorticity trained by the synch rule are very close to the values obtained by CPT training. This is not the case for temperature and divergence. For temperature, CPT puts 10 % less weight on imperfect model 3 compared to the synch rule and a 10 % stronger weight on model 1 for divergence. The synch rule puts (almost) zero weight on imperfect model 1. This is because imperfect model 1 calculates exactly the same vorticity change as imperfect model 4; hence, the synch rule suggests that imperfect model 1 has no added value in the weighted supermodel. Again, the synch rule training yields sum of weights equal to 1 as an optimal solution. Using these weights, we will compare the climatology and forecast skill of both supermodels.</p>
<sec id="Ch1.S6.SS2.SSS1">
  <label>6.2.1</label><title>Climate measures and forecast quality</title>
      <p id="d1e4460">We repeated similar climate integrations as in the case of the supermodels based on two imperfect models and assessed the climatological errors. By comparing the 40-year time series of global mean values in Fig. <xref ref-type="fig" rid="Ch1.F12"/>, both supermodels remain close to the perfect model and are clearly superior to all the imperfect models. Despite all imperfect models becoming too warm and precipitating too much on the global scale, the supermodels balance model deficiencies and produce climate simulations that are close to the truth. Inspection of the RMSE of the 30-year mean fields in the different figure panels indicates that for temperature the supermodel with weights from the CPT training is substantially better than the supermodel with weights from the synch rule. Recall that while imperfect model 1 almost does not contribute to the supermodel with the weights from the synch rule, it does so for the supermodel with the weights from the CPT training. Although imperfect model 1 has larger climatological errors than the other imperfect models (Fig. <xref ref-type="fig" rid="Ch1.F12"/>), it nevertheless improves the quality of the CPT supermodel.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F12" specific-use="star"><?xmltex \currentcnt{12}?><label>Figure 12</label><caption><p id="d1e4469">Global mean time series for the truth, the perfect model, the imperfect models and the two supermodels trained by CPT and the synch rule. Included is the RMSE of the model years 2011–2040 with respect to the truth. The normalized RMSE in the climatology of model years 2011–2040 with respect to the climatology of the truth is given in each panel. The normalization is such that the expected value of the perfect model error is 1.</p></caption>
            <?xmltex \igopts{width=455.244094pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f12.png"/>

          </fig>

      <p id="d1e4478">This experiment demonstrates the potential of supermodels to mitigate common errors and thereby clearly outperform the standard MME approach. Since all imperfect models overestimate the global average temperature and simulate too much precipitation, a standard weighted MME approach results in a climatological forecast worse than the best imperfect model. In the case that the imperfect parameters form a convex hull around the true parameter values, we may expect that a supermodel can be constructed with a climatology much closer to the truth as compared to the best imperfect model. In the case that the imperfect models do not form a convex hull around the true parameter values, allowing negative weights in the weighted supermodel might still improve the climatology and forecast skill. This will be explored in the next section.</p>
      <p id="d1e4482">We repeated the same forecast experiment as in the case of the supermodel based on two imperfect models. Also, in this case, the supermodels have forecast errors that are substantially reduced as compared to the imperfect models, up to a factor of 3 smaller (not shown). Both supermodels have comparable forecast skill in this measure.</p>
</sec>
</sec>
<sec id="Ch1.S6.SS3">
  <label>6.3</label><title>Negative weights</title>
      <p id="d1e4494">The CPT training method only produces positive weights, since the weights are defined as being equal to the frequency that the solution of a particular model is closest to the truth during the training period. The synch rule training, on the other hand, does not impose any constraint on the weights. The weights came out positive due to the convex hull principle: the imperfect models considered so far surrounded the truth and with positive weights the effect of the true parameter values can be approximated. But in the event that the imperfect models have parameter values that are all smaller or larger than the truth, only by allowing negative weights one can construct a linear superposition of imperfect models that is closer to the truth. To test if such a supermodel with negative weights indeed shows the desired physical behavior and to test if we can obtain such a model with the synch rule, we construct a weighted supermodel based on two imperfect models (models 1 and 6) with parameter values on the same side of the true parameter values (Table <xref ref-type="table" rid="Ch1.T1"/>).</p>
      <?pagebreak page802?><p id="d1e4499"><?xmltex \hack{\newpage}?>After a training period of 1 year using the synch rule, stable weights are obtained, which indicates that at least a local minimum is reached. And as expected, the training produces negative weights (Table <xref ref-type="table" rid="Ch1.T5"/>). In contrast to the previous experiments, however, the weights for temperature, divergence and vorticity are quite different. The weights for divergence are positive and do not substantially differ from the weights of the previous experiments. The weights for temperature and vorticity are negative for one of the imperfect models and larger than 1 for the other such that the sum is again close to 1.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T5"><?xmltex \currentcnt{5}?><label>Table 5</label><caption><p id="d1e4508">Weights for the supermodel trained by the synch rule. Between brackets the standard deviation over the last 10 weeks of training is given.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="4">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M133" display="inline"><mml:mi>T</mml:mi></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3">VOR</oasis:entry>
         <oasis:entry colname="col4">DIV</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">1.30 (0.016)</oasis:entry>
         <oasis:entry colname="col3">2.00 (0.011)</oasis:entry>
         <oasis:entry colname="col4">0.40 (0.010)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 2</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M134" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.30</mml:mn></mml:mrow></mml:math></inline-formula> (0.016)</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M135" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.00</mml:mn></mml:mrow></mml:math></inline-formula> (0.010)</oasis:entry>
         <oasis:entry colname="col4">0.60 (0.009)</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T6"><?xmltex \currentcnt{6}?><label>Table 6</label><caption><p id="d1e4603">Global mean average difference with the perfect model, calculated over the last 30 years of the simulation.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.88}[.88]?><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="center"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:colspec colnum="5" colname="col5" align="center"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Model</oasis:entry>
         <oasis:entry colname="col2">Temperature</oasis:entry>
         <oasis:entry colname="col3">Precipitation</oasis:entry>
         <oasis:entry colname="col4">Solar</oasis:entry>
         <oasis:entry colname="col5">Cloud cover</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M136" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula><inline-formula><mml:math id="M137" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C<inline-formula><mml:math id="M138" display="inline"><mml:mo>)</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M139" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>mm d<inline-formula><mml:math id="M140" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">surface</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M141" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>%<inline-formula><mml:math id="M142" display="inline"><mml:mo>)</mml:mo></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4">radiation</oasis:entry>
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M143" display="inline"><mml:mo>(</mml:mo></mml:math></inline-formula>W m<inline-formula><mml:math id="M144" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Model 1</oasis:entry>
         <oasis:entry colname="col2">1.37</oasis:entry>
         <oasis:entry colname="col3">0.11</oasis:entry>
         <oasis:entry colname="col4">2.06</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M145" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.59</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Model 6</oasis:entry>
         <oasis:entry colname="col2">3.20</oasis:entry>
         <oasis:entry colname="col3">0.26</oasis:entry>
         <oasis:entry colname="col4">3.95</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M146" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">3.37</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Supermodel</oasis:entry>
         <oasis:entry colname="col2">0.64</oasis:entry>
         <oasis:entry colname="col3">0.06</oasis:entry>
         <oasis:entry colname="col4">1.68</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M147" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1.16</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

      <p id="d1e4841">Stable climate simulations turn out to be possible with a weighted supermodel using negative weights. The climatology of the supermodel has improved significantly compared to both imperfect models, as displayed in Table <xref ref-type="table" rid="Ch1.T6"/>. Global mean values of the various fields are closer to the truth, despite the fact that the global mean climatological errors of both imperfect models have the same sign. Also, local model errors largely have the same sign but are smallest for the supermodel as shown in Fig. <xref ref-type="fig" rid="Ch1.F13"/> for the zonal wind at 200 hPa. Nevertheless, despite the improvement, substantial errors still remain in the supermodel solution.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F13" specific-use="star"><?xmltex \currentcnt{13}?><label>Figure 13</label><caption><p id="d1e4850">Difference in the east–west component of the wind at the 200 hPa pressure level averaged over model years 2011–2040 for the various models with respect to the truth. Contours denote areas where the difference is larger than the sampling error at 95 % confidence (solid for positive difference; dotted for negative). Positive values imply stronger mean winds blowing eastward. Units: m s<inline-formula><mml:math id="M148" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f13.png"/>

        </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F14"><?xmltex \currentcnt{14}?><label>Figure 14</label><caption><p id="d1e4873">Forecast quality as measured by the RMSE of the truth and a model with a perturbed initial condition. The control is the difference between the perfect model and the perfect model with a perturbed initial condition.</p></caption>
          <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f14.png"/>

        </fig>

      <p id="d1e4882">The forecast errors are evaluated in a similar fashion as in the previous cases and shown in Fig. <xref ref-type="fig" rid="Ch1.F14"/>. Although there is a significant improvement in quality for the supermodel as compared to the imperfect models, the forecast error is still quite large. Closer correspondence to the truth can only be expected if all prognostic variables are exchanged, hence also specific humidity and surface pressure, and if the perturbed parameters appear linearly in the equations. Both conditions are not fulfilled in this case.</p>
</sec>
<sec id="Ch1.S6.SS4">
  <label>6.4</label><title>Summary of supermodel climate errors</title>
      <p id="d1e4895">We conclude this section with a summary of the climatological errors of the weighted supermodels of this study and the connected supermodel of <xref ref-type="bibr" rid="bib1.bibx18" id="text.49"/> in Fig. <xref ref-type="fig" rid="Ch1.F15"/>. The climatological errors of the weighted supermodels of this study based on two imperfect models are of the order of the sampling error of the perfect model, whereas the connected supermodel based on the same two imperfect models has substantially larger errors. Also, the weighted supermodel based on the four imperfect models trained by CPT is indistinguishable from the truth with respect to its climatological errors, whereas the synch-rule-trained weighted supermodel has substantially larger errors. These results suggest that CPT training might yield more robust results. The largest climatological errors remain for the supermodel with negative weights.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F15"><?xmltex \currentcnt{15}?><label>Figure 15</label><caption><p id="d1e4905">Overview the RMSE of the different supermodels (the connected supermodel of <xref ref-type="bibr" rid="bib1.bibx18" id="altparen.50"/>, and the weighted supermodels from the experiments of this paper) over the model years 2011–2040 with respect to the truth.</p></caption>
          <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://esd.copernicus.org/articles/10/789/2019/esd-10-789-2019-f15.png"/>

          <?xmltex \hack{\vspace*{10mm}}?>
        </fig>

<?xmltex \hack{\newpage}?>
</sec>
</sec>
<?pagebreak page803?><sec id="Ch1.S7" sec-type="conclusions">
  <label>7</label><title>Discussion and conclusions</title>
      <p id="d1e4930">We have demonstrated the potential of weighted supermodeling to improve weather and climate predictions using the global coupled atmosphere–ocean–land model SPEEDO in the presence of parametric error. Weighted supermodels are constructed based on SPEEDO with perturbed parameters. The perturbations are chosen such that the spread in imperfect models reflects the uncertainty in climate models realistically. The weights are trained using data from the perfect model (i.e., our reference simulated truth) using two different training schemes having low computational cost. The first method is based on CPT, where different model trajectories are “crossed” in order to create a larger ensemble of possible trajectories. The second method is a synchronization-based learning rule (synch rule), which adapts the weights of the<?pagebreak page804?> different imperfect models during training such that the supermodel synchronizes with the perfect model.</p>
      <?pagebreak page805?><p id="d1e4933">Both training methods yield supermodels that outperform the individual imperfect models in short-term forecasts as well as in long-term climate simulations. CPT training required shorter training periods (1 week as opposed to a year for the synch rule), but both are much more efficient than cost-function-based approaches that are known to require many climate simulations in an iterative process to reach convergence on optimal weights <xref ref-type="bibr" rid="bib1.bibx23 bib1.bibx20" id="paren.51"/>. An advantage of the synch rule is that it allows for negative weights that can potentially improve the weighted supermodel in the event that model errors do not compensate for positive weights. In addition, CPT requires fairly good models such that mixed trajectories are able to track an observed trajectory for some time. During the synch rule training, on the other hand, the nudging terms keep the supermodel in the neighborhood of the observed trajectory and is therefore more robust (i.e., less sensitive) with respect to the quality of the imperfect models.</p>
      <p id="d1e4939">In the application of CPT in this study, we encountered numerical issues due to the partial state replacement. A possible solution is the use of data assimilation techniques to combine state information from different models in a dynamical consistent manner <xref ref-type="bibr" rid="bib1.bibx1 bib1.bibx3" id="paren.52"/>. One straightforward solution along this line could be based on the idea of <xref ref-type="bibr" rid="bib1.bibx6" id="text.53"/>, in which pseudo-orbit data assimilation is used instead of replacement of the entire state. <xref ref-type="bibr" rid="bib1.bibx6" id="text.54"/> have already used this approach successfully for low-order dynamical systems. These data assimilation techniques would also allow application of CPT in the event that the different models differ in state representation, for instance, different numerical grids.</p>
      <p id="d1e4951">The weighted supermodels of this study have smaller climatological errors as compared to the connected supermodel based on the same two imperfect models in <xref ref-type="bibr" rid="bib1.bibx18" id="text.55"/>. Also, in the four-model experiment, the CPT supermodel has substantially better climatology than the supermodel trained by the synch rule. This suggests that synchronization with the truth can be difficult to obtain, especially when the imperfect models that form the supermodel are not fully synchronized in the case of a connected supermodel or when the weighted supermodel consists of several imperfect models. Although it is a common result in synchronization theory that identical systems will synchronize if the nudging strength is strong enough and if there are enough observations from the truth, in practice, this can be a challenge. The issues with synchronized-based learning can be easily demonstrated using a low-order dimensional system (not shown).</p>
      <p id="d1e4958">In the second supermodel experiment of this paper, the parameter perturbations of four imperfect models were chosen such that they formed a so-called convex hull around the true parameter values. This implies that a linear combination with positive weights of these four models is able to reproduce the model equations with the true parameter values, provided that the parameters appear only linear in the equations. This is not exactly true in this case, but the trained weighted supermodel based on these four models turned out to have a climatology close to the truth. As all four imperfect models have a warmer and wetter climatology than the truth, simply taking the MME mean with positive weights thus does not improve the climatology. This experiment is a clear example of the potential benefit of the supermodeling approach to ameliorate common model errors. This benefit arises due to the fact that model errors are compensated at an early stage, in the time derivative, and not a posteriori, as in the MME approach where model errors have propagated spatially across the globe, across scales and across the different meteorological fields and other components of the climate system.</p>
      <p id="d1e4961">In the final supermodel experiment, we have explored the use of negative weights in order to improve predictions in the case that model errors do not compensate; i.e., both imperfect models have parameter perturbations and climatological errors of the same sign. A supermodel trained using the synch rule yielded negative weights. With these weights, stable and credible simulations turn out to be possible and forecast errors as well as climatological errors are reduced with respect to the imperfect models. Substantial errors remain as not all prognostic equations are combined (only temperature, vorticity and divergence, not humidity and surface pressure) and the parameters do not appear linearly in the equations.</p>
      <p id="d1e4964">Although the synch rule training does not impose that the weights sum to 1, the training inevitably yielded sum of weights equal to 1. An example based on the Lorenz 1963 equations <xref ref-type="bibr" rid="bib1.bibx12" id="paren.56"/> serves to illustrate why this might be the case. The Lorenz 1963 equations are

              <disp-formula specific-use="align"><mml:math id="M149" display="block"><mml:mtable displaystyle="true"><mml:mtr><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>-</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>y</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mover accent="true"><mml:mi>z</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mi>z</mml:mi><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>

          where the standard parameter values are <inline-formula><mml:math id="M150" display="inline"><mml:mrow><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">10</mml:mn></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M151" display="inline"><mml:mrow><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">28</mml:mn></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M152" display="inline"><mml:mrow><mml:mi mathvariant="italic">β</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">8</mml:mn><mml:mn mathvariant="normal">3</mml:mn></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula>. Assume we have two imperfect models with imperfect parameters <inline-formula><mml:math id="M153" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M154" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>. Then, <inline-formula><mml:math id="M155" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, with “s” denoting the supermodel solution. We can rewrite this as <inline-formula><mml:math id="M156" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mi mathvariant="normal">s</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>)</mml:mo><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mi>z</mml:mi><mml:mo>+</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. To reproduce the standard parameter model solution, two conditions must be satisfied: <inline-formula><mml:math id="M157" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M158" display="inline"><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>. Not only should the linear combination of imperfect parameter values match the true parameter value, but also the weights have to sum to 1.</p>
      <p id="d1e5327">The ultimate goal of our research is to apply supermodeling to realistic climate models. But will it work? Based on the current results, we believe that this is possible, although the application is not as straightforward as for SPEEDO. First, state-of-the-art models are far bigger and more complex, making their numerical computation a substantial burden. This makes numerical efficiency a key aspect to consider. Second, the real world is not simply a perturbed parameter version of these complex models. In this paper, we have worked under the hypothesis that model error only originates by error in the model parameters in the atmosphere. The imperfect atmosphere models were coupled to the same ocean and land model, which constrains the variability on longer timescales. So far we have demonstrated that the long-term behavior of the supermodel improves while training only<?pagebreak page806?> short-term prediction errors. It remains to be seen how much the long-term evolution will improve in the presence of imperfections in the slow components of the climate system. Furthermore, it is essential to extend the approach to other sources of model error towards the application with real climate models. In that case, on top of parametric error, model error can arise from the presence of unresolved scale, numerical discretization or incorrect physics.</p>
      <p id="d1e5330">Together with the realisms of the models (and of the related model error), those of the observations are also of central importance. In all previous studies with supermodeling, including the current, observations were assumed to be perfect, i.e., to be complete and noise free. To use real data, it will thus be necessary to study the robustness of the supermodeling approach to noisy and unevenly distributed observations and to extend the methods to account for the observational noise. This latter problem is the subject of ongoing research of scientists which are making use of ideas and techniques from data assimilation. Data-assimilation-based supermodeling is also envisioned to account for generic source of model error in the construction of the supermodel, and it will be the subject of future research.</p>
</sec>

      
      </body>
    <back><notes notes-type="dataavailability"><title>Data availability</title>

      <p id="d1e5337">No data sets were used in this article.</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d1e5343">FSc conceived the study, carried out the research and led the writing of the manuscript. FSe provided codes and technical advice and provided with AC and NK input for the interpretation of the results and the writing.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d1e5349">The authors declare that they have no conflict of interest.</p>
  </notes><ack><title>Acknowledgements</title><p id="d1e5355">Alberto Carrassi has been funded by the Trond Mohn Foundation under the project no. BFS2018TMT0.</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d1e5360">This research has been supported by the H2020 European Research Council (grant no. STERCP (648982)).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d1e5367">This paper was edited by Andrey Gritsun and reviewed by two anonymous referees.</p>
  </notes><?xmltex \hack{\newpage}?><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>Asch et al.(2016)Asch, Bocquet, and Nodet</label><?label asch2016data?><mixed-citation>Asch, M., Bocquet, M., and Nodet, M.: Data Assimilation: Methods, Algorithms,
and Applications, Fundamentals of Algorithms, Society for Industrial and
Applied Mathematics, Philadelphia, PA, <ext-link xlink:href="https://doi.org/10.1137/1.9781611974546" ext-link-type="DOI">10.1137/1.9781611974546</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Bauer et al.(2015)Bauer, Thorpe, and Brunet</label><?label cite-key?><mixed-citation>Bauer, P., Thorpe, A., and Brunet, G.: The quiet revolution of numerical
weather prediction, Nature, 525, 47–55, <ext-link xlink:href="https://doi.org/10.1038/nature14956" ext-link-type="DOI">10.1038/nature14956</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Carrassi et al.(2018)Carrassi, Bocquet, Bertino, and
Evensen</label><?label Carrassi2018?><mixed-citation>Carrassi, A., Bocquet, M., Bertino, L., and Evensen, G.: Data assimilation in
the geosciences: An overview of methods, issues, and perspectives, Wiley
Interdisciplin. Rev.: Clim. Change, 9, e535, <ext-link xlink:href="https://doi.org/10.1002/wcc.535" ext-link-type="DOI">10.1002/wcc.535</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>Challinor and Wheeler(2008)</label><?label wrro78034?><mixed-citation>
Challinor, A. and Wheeler, T.: Crop yield reduction in the tropics under
climate change: processes and uncertainties, Agr. Forest Meteorol., 148, 343–356, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Collins and Allen(2002)</label><?label Collins2002?><mixed-citation>Collins, M. and Allen, M. R.: Assessing the Relative Roles of Initial and
Boundary Conditions in Interannual to Decadal Climate Predictability, J.
Climate, 15, 3104–3109,
<ext-link xlink:href="https://doi.org/10.1175/1520-0442(2002)015&lt;3104:ATRROI&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0442(2002)015&lt;3104:ATRROI&gt;2.0.CO;2</ext-link>, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Du and Smith(2017)</label><?label DU201731?><mixed-citation>Du, H. and Smith, L. A.: Multi-model cross-pollination in time, Physica D, 353–354, 31–38, <ext-link xlink:href="https://doi.org/10.1016/j.physd.2017.06.001" ext-link-type="DOI">10.1016/j.physd.2017.06.001</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Duane et al.(2007)Duane, Yu, and Kocarev</label><?label Duane2007?><mixed-citation>Duane, G. S., Yu, D., and Kocarev, L.: Identical synchronization, with
translation invariance, implies parameter estimation, Phys. Lett. A, 371,
416–420, <ext-link xlink:href="https://doi.org/10.1016/j.physleta.2007.06.059" ext-link-type="DOI">10.1016/j.physleta.2007.06.059</ext-link>, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Goosse and Fichefet(1999)</label><?label Goosse1999?><mixed-citation>Goosse, H. and Fichefet, T.: Importance of ice-ocean interactions for the
global ocean circulation: A model study, J. Geophys. Res.-Oceans, 104, 23337–23355, <ext-link xlink:href="https://doi.org/10.1029/1999JC900215" ext-link-type="DOI">10.1029/1999JC900215</ext-link>, 1999.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Hawkins and Sutton(2009)</label><?label Hawkins2009?><mixed-citation>Hawkins, E. and Sutton, R.: The Potential to Narrow Uncertainty in Regional
Climate Predictions, B. Am. Meteorol. Soc., 90, 1095–1108, <ext-link xlink:href="https://doi.org/10.1175/2009BAMS2607.1" ext-link-type="DOI">10.1175/2009BAMS2607.1</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>IPCC(2013)</label><?label IPCCreport?><mixed-citation>
IPCC: Climate Change 2013: The Physical Science Basis, in: Contribution of Working Group I to the Fifth Assessment Report of the Intergovernmental Panel on Climate Change, edited by: Stocker, T. F., Qin, D., Plattner, G.-K., Tignor, M., Allen, S. K., Boschung, J., Nauels, A., Xia, Y., Bex, V., and  Midgley, P. M., Cambridge University Press, Cambridge, UK and New York, NY, USA, 1535 pp., 2013.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Kirtman and Shukla(2002)</label><?label Kirtman?><mixed-citation>Kirtman, B. P. and Shukla, J.: Interactive coupled ensemble: A new coupling
strategy for CGCMs, Geophys. Res. Lett., 29, 5-1–5-4, <ext-link xlink:href="https://doi.org/10.1029/2002GL014834" ext-link-type="DOI">10.1029/2002GL014834</ext-link>, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Lorenz(1963)</label><?label Lorenz63?><mixed-citation>
Lorenz, E.: Deterministic nonperiodic flow, J. Atmos. Sci., 20, 130–140, 1963.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Mirchev et al.(2012)Mirchev, Duane, Tang, and Kocarev</label><?label Mirchev2012?><mixed-citation>Mirchev, M., Duane, G. S., Tang, W. K., and Kocarev, L.: Improved modeling by
coupling imperfect models, Commun. Nonlin. Sci. Numer. Simul., 17, 2741–2751, <ext-link xlink:href="https://doi.org/10.1016/j.cnsns.2011.11.003" ext-link-type="DOI">10.1016/j.cnsns.2011.11.003</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Pecora and Carroll(1990)</label><?label PhysRevLett.64.821?><mixed-citation>Pecora, L. M. and Carroll, T. L.: Synchronization in chaotic systems, Phys.
Rev. Lett., 64, 821–824, <ext-link xlink:href="https://doi.org/10.1103/PhysRevLett.64.821" ext-link-type="DOI">10.1103/PhysRevLett.64.821</ext-link>, 1990.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Rodwell and Jung(2008)</label><?label RodwellJung?><mixed-citation>Rodwell, M. J. and Jung, T.: Understanding the local and global impacts of
model physics changes: an aerosol example, Q. J. Roy. Meteorol. Soc., 134, 1479–1497, <ext-link xlink:href="https://doi.org/10.1002/qj.298" ext-link-type="DOI">10.1002/qj.298</ext-link>, 2008.</mixed-citation></ref>
      <?pagebreak page807?><ref id="bib1.bibx16"><label>Rodwell and Palmer(2007)</label><?label Rodwell?><mixed-citation>Rodwell, M. J. and Palmer, T. N.: Using numerical weather prediction to assess climate models, Q. J. Roy. Meteorol. Soc., 133, 129–146, <ext-link xlink:href="https://doi.org/10.1002/qj.23" ext-link-type="DOI">10.1002/qj.23</ext-link>, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Schevenhoven and Selten(2017)</label><?label esd-8-429-2017?><mixed-citation>Schevenhoven, F. J. and Selten, F. M.: An efficient training scheme for
supermodels, Earth Syst. Dynam., 8, 429–438, <ext-link xlink:href="https://doi.org/10.5194/esd-8-429-2017" ext-link-type="DOI">10.5194/esd-8-429-2017</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>Selten et al.(2017)Selten, Schevenhoven, and
Duane</label><?label Selten?><mixed-citation>Selten, F. M., Schevenhoven, F. J., and Duane, G. S.: Simulating climate with a synchronization-based supermodel, Chaos, 27, 126903, <ext-link xlink:href="https://doi.org/10.1063/1.4990721" ext-link-type="DOI">10.1063/1.4990721</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Severijns and Hazeleger(2010)</label><?label Severijns2010?><mixed-citation>Severijns, C. A. and Hazeleger, W.: The efficient global primitive equation
climate model SPEEDO V2.0, Geosci. Model Dev., 3, 105–122,
<ext-link xlink:href="https://doi.org/10.5194/gmd-3-105-2010" ext-link-type="DOI">10.5194/gmd-3-105-2010</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Shen et al.(2016)Shen, Keenlyside, Selten, Wiegerinck, and
Duane</label><?label GRL:GRL53840?><mixed-citation>Shen, M.-L., Keenlyside, N., Selten, F., Wiegerinck, W., and Duane, G. S.:
Dynamically combining climate models to “supermodel” the tropical
Pacific, Geophys. Res. Lett., 43, 359–366, <ext-link xlink:href="https://doi.org/10.1002/2015GL066562" ext-link-type="DOI">10.1002/2015GL066562</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Smith(2001)</label><?label Smith2001?><mixed-citation>Smith, L. A.: Nonlinear Dynamics and Statistics, in: chap. Disentangling
Uncertainty and Error: On the Predictability of Nonlinear Systems, edited by:
Mees, A. I., Birkhäuser Boston, Boston, MA, 31–64,
<ext-link xlink:href="https://doi.org/10.1007/978-1-4612-0177-9_2" ext-link-type="DOI">10.1007/978-1-4612-0177-9_2</ext-link>, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Sterl et al.(2009)Sterl, van den Brink, de Vries, Haarsma, and van
Meijgaard</label><?label os-5-369-2009?><mixed-citation>Sterl, A., van den Brink, H., de Vries, H., Haarsma, R., and van Meijgaard, E.: An ensemble study of extreme storm surge related water levels in the North Sea in a changing climate, Ocean Sci., 5, 369–378,
<ext-link xlink:href="https://doi.org/10.5194/os-5-369-2009" ext-link-type="DOI">10.5194/os-5-369-2009</ext-link>, 2009.
</mixed-citation></ref><?xmltex \hack{\newpage}?>
      <ref id="bib1.bibx23"><label>van den Berge et al.(2011)van den Berge, Selten, Wiegerinck, and
Duane</label><?label esd-2-161-2011?><mixed-citation>van den Berge, L. A., Selten, F. M., Wiegerinck, W., and Duane, G. S.: A
multi-model ensemble method that combines imperfect models through learning,
Earth Syst. Dynam., 2, 161–177, <ext-link xlink:href="https://doi.org/10.5194/esd-2-161-2011" ext-link-type="DOI">10.5194/esd-2-161-2011</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Van der Wiel et al.(2019)van der Wiel, Wanders, Selten, and
Bierkens</label><?label vanderWiel?><mixed-citation>Van der Wiel, K., Wanders, N., Selten, F. M., and Bierkens, M. F. P.: Added
value of large ensemble simulations for assessing extreme river discharge in
a 2 <inline-formula><mml:math id="M159" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C warmer world, Geophys. Res. Lett., 46, 2093–2102, <ext-link xlink:href="https://doi.org/10.1029/2019GL081967" ext-link-type="DOI">10.1029/2019GL081967</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Weigel et al.(2008)Weigel, Liniger, and
Appenzeller</label><?label Weigel?><mixed-citation>Weigel, A. P., Liniger, M. A., and Appenzeller, C.: Can multi-model
combination really enhance the prediction skill of probabilistic ensemble
forecasts?, Q. J. Roy. Meteorol. Soc., 134, 241–260, <ext-link xlink:href="https://doi.org/10.1002/qj.210" ext-link-type="DOI">10.1002/qj.210</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Wiegerinck and Selten(2017)</label><?label Wiegerinck2017?><mixed-citation>Wiegerinck, W. and Selten, F. M.: Attractor learning in synchronized chaotic
systems in the presence of unresolved scales, Chaos, 27, 126901, <ext-link xlink:href="https://doi.org/10.1063/1.4990660" ext-link-type="DOI">10.1063/1.4990660</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>Wiegerinck et al.(2013)Wiegerinck, Mirchev, Burgers, and
Selten</label><?label Wiegerinck2013?><mixed-citation>Wiegerinck, W., Mirchev, M., Burgers, W., and Selten, F.: Consensus and
Synchronization in Complex Networks, in: chap. Supermodeling Dynamics and
Learning Mechanisms, Springer, Berlin, Heidelberg, 227–255, <ext-link xlink:href="https://doi.org/10.1007/978-3-642-33359-0_9" ext-link-type="DOI">10.1007/978-3-642-33359-0_9</ext-link>, 2013.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Improving weather and climate predictions  by training of supermodels</article-title-html>
<abstract-html><p>Recent studies demonstrate that weather and climate predictions potentially improve by dynamically combining different models into a so-called <q>supermodel</q>. Here, we focus on the weighted supermodel – the supermodel's time derivative is a weighted superposition of the time derivatives of the imperfect models, referred to as weighted supermodeling. A crucial step is to train the weights of the supermodel on the basis of historical observations. Here, we apply two different training methods to a supermodel of up to four different versions of the global atmosphere–ocean–land model SPEEDO. The standard version is regarded as truth. The first training method is based on an idea called cross pollination in time (CPT), where models exchange states during the training. The second method is a synchronization-based learning rule, originally developed for parameter estimation. We demonstrate that both training methods yield climate simulations and weather predictions of superior quality as compared to the individual model versions. Supermodel predictions also outperform predictions based on the commonly used multi-model ensemble (MME) mean. Furthermore, we find evidence that negative weights can improve predictions in cases where model errors do not cancel (for instance, all models are warm with respect to the truth). In principle, the proposed training schemes are applicable to state-of-the-art models and historical observations. A prime advantage of the proposed training schemes is that in the present context relatively short training periods suffice to find good solutions. Additional work needs to be done to assess the limitations due to incomplete and noisy data, to combine models that are structurally different (different resolution and state representation, for instance) and to evaluate cases for which the truth falls outside of the model class.</p></abstract-html>
<ref-html id="bib1.bib1"><label>Asch et al.(2016)Asch, Bocquet, and Nodet</label><mixed-citation>
Asch, M., Bocquet, M., and Nodet, M.: Data Assimilation: Methods, Algorithms,
and Applications, Fundamentals of Algorithms, Society for Industrial and
Applied Mathematics, Philadelphia, PA, <a href="https://doi.org/10.1137/1.9781611974546" target="_blank">https://doi.org/10.1137/1.9781611974546</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Bauer et al.(2015)Bauer, Thorpe, and Brunet</label><mixed-citation>
Bauer, P., Thorpe, A., and Brunet, G.: The quiet revolution of numerical
weather prediction, Nature, 525, 47–55, <a href="https://doi.org/10.1038/nature14956" target="_blank">https://doi.org/10.1038/nature14956</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Carrassi et al.(2018)Carrassi, Bocquet, Bertino, and
Evensen</label><mixed-citation>
Carrassi, A., Bocquet, M., Bertino, L., and Evensen, G.: Data assimilation in
the geosciences: An overview of methods, issues, and perspectives, Wiley
Interdisciplin. Rev.: Clim. Change, 9, e535, <a href="https://doi.org/10.1002/wcc.535" target="_blank">https://doi.org/10.1002/wcc.535</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Challinor and Wheeler(2008)</label><mixed-citation>
Challinor, A. and Wheeler, T.: Crop yield reduction in the tropics under
climate change: processes and uncertainties, Agr. Forest Meteorol., 148, 343–356, 2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Collins and Allen(2002)</label><mixed-citation>
Collins, M. and Allen, M. R.: Assessing the Relative Roles of Initial and
Boundary Conditions in Interannual to Decadal Climate Predictability, J.
Climate, 15, 3104–3109,
<a href="https://doi.org/10.1175/1520-0442(2002)015&lt;3104:ATRROI&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0442(2002)015&lt;3104:ATRROI&gt;2.0.CO;2</a>, 2002.
</mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Du and Smith(2017)</label><mixed-citation>
Du, H. and Smith, L. A.: Multi-model cross-pollination in time, Physica D, 353–354, 31–38, <a href="https://doi.org/10.1016/j.physd.2017.06.001" target="_blank">https://doi.org/10.1016/j.physd.2017.06.001</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Duane et al.(2007)Duane, Yu, and Kocarev</label><mixed-citation>
Duane, G. S., Yu, D., and Kocarev, L.: Identical synchronization, with
translation invariance, implies parameter estimation, Phys. Lett. A, 371,
416–420, <a href="https://doi.org/10.1016/j.physleta.2007.06.059" target="_blank">https://doi.org/10.1016/j.physleta.2007.06.059</a>, 2007.
</mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Goosse and Fichefet(1999)</label><mixed-citation>
Goosse, H. and Fichefet, T.: Importance of ice-ocean interactions for the
global ocean circulation: A model study, J. Geophys. Res.-Oceans, 104, 23337–23355, <a href="https://doi.org/10.1029/1999JC900215" target="_blank">https://doi.org/10.1029/1999JC900215</a>, 1999.
</mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Hawkins and Sutton(2009)</label><mixed-citation>
Hawkins, E. and Sutton, R.: The Potential to Narrow Uncertainty in Regional
Climate Predictions, B. Am. Meteorol. Soc., 90, 1095–1108, <a href="https://doi.org/10.1175/2009BAMS2607.1" target="_blank">https://doi.org/10.1175/2009BAMS2607.1</a>, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>IPCC(2013)</label><mixed-citation>
IPCC: Climate Change 2013: The Physical Science Basis, in: Contribution of Working Group I to the Fifth Assessment Report of the Intergovernmental Panel on Climate Change, edited by: Stocker, T. F., Qin, D., Plattner, G.-K., Tignor, M., Allen, S. K., Boschung, J., Nauels, A., Xia, Y., Bex, V., and  Midgley, P. M., Cambridge University Press, Cambridge, UK and New York, NY, USA, 1535&thinsp;pp., 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Kirtman and Shukla(2002)</label><mixed-citation>
Kirtman, B. P. and Shukla, J.: Interactive coupled ensemble: A new coupling
strategy for CGCMs, Geophys. Res. Lett., 29, 5-1–5-4, <a href="https://doi.org/10.1029/2002GL014834" target="_blank">https://doi.org/10.1029/2002GL014834</a>, 2002.
</mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Lorenz(1963)</label><mixed-citation>
Lorenz, E.: Deterministic nonperiodic flow, J. Atmos. Sci., 20, 130–140, 1963.
</mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Mirchev et al.(2012)Mirchev, Duane, Tang, and Kocarev</label><mixed-citation>
Mirchev, M., Duane, G. S., Tang, W. K., and Kocarev, L.: Improved modeling by
coupling imperfect models, Commun. Nonlin. Sci. Numer. Simul., 17, 2741–2751, <a href="https://doi.org/10.1016/j.cnsns.2011.11.003" target="_blank">https://doi.org/10.1016/j.cnsns.2011.11.003</a>, 2012.
</mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Pecora and Carroll(1990)</label><mixed-citation>
Pecora, L. M. and Carroll, T. L.: Synchronization in chaotic systems, Phys.
Rev. Lett., 64, 821–824, <a href="https://doi.org/10.1103/PhysRevLett.64.821" target="_blank">https://doi.org/10.1103/PhysRevLett.64.821</a>, 1990.
</mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Rodwell and Jung(2008)</label><mixed-citation>
Rodwell, M. J. and Jung, T.: Understanding the local and global impacts of
model physics changes: an aerosol example, Q. J. Roy. Meteorol. Soc., 134, 1479–1497, <a href="https://doi.org/10.1002/qj.298" target="_blank">https://doi.org/10.1002/qj.298</a>, 2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Rodwell and Palmer(2007)</label><mixed-citation>
Rodwell, M. J. and Palmer, T. N.: Using numerical weather prediction to assess climate models, Q. J. Roy. Meteorol. Soc., 133, 129–146, <a href="https://doi.org/10.1002/qj.23" target="_blank">https://doi.org/10.1002/qj.23</a>, 2007.
</mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Schevenhoven and Selten(2017)</label><mixed-citation>
Schevenhoven, F. J. and Selten, F. M.: An efficient training scheme for
supermodels, Earth Syst. Dynam., 8, 429–438, <a href="https://doi.org/10.5194/esd-8-429-2017" target="_blank">https://doi.org/10.5194/esd-8-429-2017</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Selten et al.(2017)Selten, Schevenhoven, and
Duane</label><mixed-citation>
Selten, F. M., Schevenhoven, F. J., and Duane, G. S.: Simulating climate with a synchronization-based supermodel, Chaos, 27, 126903, <a href="https://doi.org/10.1063/1.4990721" target="_blank">https://doi.org/10.1063/1.4990721</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Severijns and Hazeleger(2010)</label><mixed-citation>
Severijns, C. A. and Hazeleger, W.: The efficient global primitive equation
climate model SPEEDO V2.0, Geosci. Model Dev., 3, 105–122,
<a href="https://doi.org/10.5194/gmd-3-105-2010" target="_blank">https://doi.org/10.5194/gmd-3-105-2010</a>, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Shen et al.(2016)Shen, Keenlyside, Selten, Wiegerinck, and
Duane</label><mixed-citation>
Shen, M.-L., Keenlyside, N., Selten, F., Wiegerinck, W., and Duane, G. S.:
Dynamically combining climate models to “supermodel” the tropical
Pacific, Geophys. Res. Lett., 43, 359–366, <a href="https://doi.org/10.1002/2015GL066562" target="_blank">https://doi.org/10.1002/2015GL066562</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Smith(2001)</label><mixed-citation>
Smith, L. A.: Nonlinear Dynamics and Statistics, in: chap. Disentangling
Uncertainty and Error: On the Predictability of Nonlinear Systems, edited by:
Mees, A. I., Birkhäuser Boston, Boston, MA, 31–64,
<a href="https://doi.org/10.1007/978-1-4612-0177-9_2" target="_blank">https://doi.org/10.1007/978-1-4612-0177-9_2</a>, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Sterl et al.(2009)Sterl, van den Brink, de Vries, Haarsma, and van
Meijgaard</label><mixed-citation>
Sterl, A., van den Brink, H., de Vries, H., Haarsma, R., and van Meijgaard, E.: An ensemble study of extreme storm surge related water levels in the North Sea in a changing climate, Ocean Sci., 5, 369–378,
<a href="https://doi.org/10.5194/os-5-369-2009" target="_blank">https://doi.org/10.5194/os-5-369-2009</a>, 2009.

</mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>van den Berge et al.(2011)van den Berge, Selten, Wiegerinck, and
Duane</label><mixed-citation>
van den Berge, L. A., Selten, F. M., Wiegerinck, W., and Duane, G. S.: A
multi-model ensemble method that combines imperfect models through learning,
Earth Syst. Dynam., 2, 161–177, <a href="https://doi.org/10.5194/esd-2-161-2011" target="_blank">https://doi.org/10.5194/esd-2-161-2011</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Van der Wiel et al.(2019)van der Wiel, Wanders, Selten, and
Bierkens</label><mixed-citation>
Van der Wiel, K., Wanders, N., Selten, F. M., and Bierkens, M. F. P.: Added
value of large ensemble simulations for assessing extreme river discharge in
a 2&thinsp;°C warmer world, Geophys. Res. Lett., 46, 2093–2102, <a href="https://doi.org/10.1029/2019GL081967" target="_blank">https://doi.org/10.1029/2019GL081967</a>, 2019.
</mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Weigel et al.(2008)Weigel, Liniger, and
Appenzeller</label><mixed-citation>
Weigel, A. P., Liniger, M. A., and Appenzeller, C.: Can multi-model
combination really enhance the prediction skill of probabilistic ensemble
forecasts?, Q. J. Roy. Meteorol. Soc., 134, 241–260, <a href="https://doi.org/10.1002/qj.210" target="_blank">https://doi.org/10.1002/qj.210</a>, 2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Wiegerinck and Selten(2017)</label><mixed-citation>
Wiegerinck, W. and Selten, F. M.: Attractor learning in synchronized chaotic
systems in the presence of unresolved scales, Chaos, 27, 126901, <a href="https://doi.org/10.1063/1.4990660" target="_blank">https://doi.org/10.1063/1.4990660</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Wiegerinck et al.(2013)Wiegerinck, Mirchev, Burgers, and
Selten</label><mixed-citation>
Wiegerinck, W., Mirchev, M., Burgers, W., and Selten, F.: Consensus and
Synchronization in Complex Networks, in: chap. Supermodeling Dynamics and
Learning Mechanisms, Springer, Berlin, Heidelberg, 227–255, <a href="https://doi.org/10.1007/978-3-642-33359-0_9" target="_blank">https://doi.org/10.1007/978-3-642-33359-0_9</a>, 2013.
</mixed-citation></ref-html>--></article>
