<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "https://jats.nlm.nih.gov/nlm-dtd/publishing/3.0/journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0" article-type="research-article">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">NPG</journal-id><journal-title-group>
    <journal-title>Nonlinear Processes in Geophysics</journal-title>
    <abbrev-journal-title abbrev-type="publisher">NPG</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Nonlin. Processes Geophys.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1607-7946</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/npg-32-353-2025</article-id><title-group><article-title>Long-window tandem variational data assimilation methods for chaotic climate models tested with the Lorenz 63 system</article-title><alt-title>Long-window tandem variational data assimilation methods</alt-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author" corresp="yes" rid="aff1">
          <name><surname>Kennedy</surname><given-names>Philip David</given-names></name>
          <email>philipdkennedy@physics.org</email>
        <ext-link>https://orcid.org/0000-0002-8491-2570</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Banerjee</surname><given-names>Abhirup</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Köhl</surname><given-names>Armin</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Stammer</surname><given-names>Detlef</given-names></name>
          
        </contrib>
        <aff id="aff1"><label>1</label><institution>Fakultät für Mathematik, Informatik und Naturwissenschaften, Fernerkundung &amp; Assimilation, Universität Hamburg, Bundesstr. 53, 20146 Hamburg, Germany</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Philip David Kennedy (philipdkennedy@physics.org)</corresp></author-notes><pub-date><day>24</day><month>September</month><year>2025</year></pub-date>
      
      <volume>32</volume>
      <issue>3</issue>
      <fpage>353</fpage><lpage>365</lpage>
      <history>
        <date date-type="received"><day>19</day><month>November</month><year>2024</year></date>
           <date date-type="rev-request"><day>9</day><month>December</month><year>2024</year></date>
           <date date-type="rev-recd"><day>18</day><month>April</month><year>2025</year></date>
           <date date-type="accepted"><day>9</day><month>June</month><year>2025</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2025 Philip David Kennedy et al.</copyright-statement>
        <copyright-year>2025</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025.html">This article is available from https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025.html</self-uri><self-uri xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025.pdf">The full text article is available as a PDF file from https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025.pdf</self-uri>
      <abstract><title>Abstract</title>

      <p id="d2e110">We apply 4D variational data assimilation to the Lorenz 63 model to introduce a new method for parameter estimation in chaotic climate models. The approach aims to optimise an Earth system model (ESM) for which no adjoint exists by utilising the adjoint of a different, potentially simpler ESM. This relies on the synchronisation of the model to observed data. Dynamical state and parameter estimation (DSPE) is used to stabilise the tangent linear system by reducing all positive Lyapunov exponents to negative values, thereby improving parameter estimation by enabling long assimilation windows. The method introduces a second layer of synchronisation between the two models, with and without an adjoint, to facilitate linearisation around the trajectory of the model for which no adjoint exists. This is achieved by synchronising two Lorenz 63 systems, one with and the other without an adjoint model. Results are presented for an idealised case of identical, perfect models and for a more realistic case in which they differ from one another. If employed in a high-resolution ESM for which a coarse-resolution adjoint exists, the method will save computational resources as only one forward run with the full high-resolution ESM per iteration is needed. It is demonstrated that there is negligible error and uncertainty change compared to the traditional optimisation of a full ESM with an adjoint. Stemming from this approach, it is shown that the synchronisation between two identical models can be used to filter noisy data in a dynamical way which reduces the parametric uncertainty of the optimised model by approximately one-third. Such a precision gain could prove to be valuable for seasonal, annual, and decadal predictions.</p>
  </abstract>
    
<funding-group>
<award-group id="gs1">
<funding-source>Deutsche Forschungsgemeinschaft</funding-source>
<award-id>66492934</award-id>
</award-group>
</funding-group>
</article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d2e122">The time evolution of the Earth system can be simulated using numerical Earth system models (ESMs). Provided these models skilfully describe the system's time evolution and observed processes, they can be used to forecast future states of the system as long as accurate initial conditions exist. Data assimilation is a powerful tool to bring ESMs into agreement with the observed climatic state by combining data with the numerical model while preserving dynamic principles governing the system <xref ref-type="bibr" rid="bib1.bibx67 bib1.bibx45" id="paren.1"/> while also attempting to further improve the ESM's predictive skills.</p>
      <p id="d2e128">There are two common assimilation approaches typically used to incorporate observations into a model: sequential and variational data assimilation schemes <xref ref-type="bibr" rid="bib1.bibx66" id="paren.2"/>. Sequential data assimilation <xref ref-type="bibr" rid="bib1.bibx5" id="paren.3"/> involves the application of a filter, most commonly Kalman filters <xref ref-type="bibr" rid="bib1.bibx31 bib1.bibx14 bib1.bibx15 bib1.bibx61 bib1.bibx29" id="paren.4"/>. This technique merges a predicted state with observations at each analysis time step by estimating a joint probability distribution between the two by taking into account their respective modelling and observational uncertainties. Variants of the Kalman filter technique include extended Kalman filters, ensemble Kalman filters, and square-root filters <xref ref-type="bibr" rid="bib1.bibx4 bib1.bibx52 bib1.bibx15 bib1.bibx64 bib1.bibx62" id="paren.5"/>. They all share a similar basic procedure while differing in terms of specific variations in the methodology. The strength of all filtering techniques is that the sequential procedure allows for real-time assimilation of observations, for example, in initialised numerical weather forecasting.</p>
      <p id="d2e143">In contrast, variational data assimilation <xref ref-type="bibr" rid="bib1.bibx37" id="paren.6"/> estimates a joint probability distribution over an extended period of time by minimising a scalar cost function, defined as the quadratic misfit between the model trajectory and all available observations within a defined time window. The most common approaches include four-dimensional variational assimilation (4D-var) <xref ref-type="bibr" rid="bib1.bibx50" id="paren.7"/>; three-dimensional variational data assimilation (3D-var) <xref ref-type="bibr" rid="bib1.bibx23" id="paren.8"/>; weak and strong constraint 4D-var <xref ref-type="bibr" rid="bib1.bibx63 bib1.bibx17" id="paren.9"/>; and ensemble variational filters, including 4DEnVar <xref ref-type="bibr" rid="bib1.bibx11" id="paren.10"/>. Variational data assimilation is a useful technique for solving both initial-value and parameter estimation problems <xref ref-type="bibr" rid="bib1.bibx16 bib1.bibx20 bib1.bibx51 bib1.bibx70" id="paren.11"/>. It will be exclusively used in this study.</p>
      <p id="d2e165">The 4D-var approach utilises an adjoint of the model to iteratively minimise the model–data misfit by adjusting control variables <xref ref-type="bibr" rid="bib1.bibx59 bib1.bibx39 bib1.bibx33 bib1.bibx3 bib1.bibx44" id="paren.12"/>. The adjoint equations of a fully non-linear model are derived from the forward equations using integration by parts. In the data assimilation context, this can be used to numerically calculate the gradient of the cost function which is subsequently used to find the cost function minimum in an iterative procedure. Adjoint models have also been widely used for sensitivity analysis in meteorology and oceanography <xref ref-type="bibr" rid="bib1.bibx26 bib1.bibx25 bib1.bibx24 bib1.bibx40 bib1.bibx53" id="paren.13"/>; this includes calculating sensitivity with respect to lateral boundary conditions <xref ref-type="bibr" rid="bib1.bibx22" id="paren.14"/> and estimating the sensitivity of the 2 m surface temperature with respect to the sea surface temperature, sea ice, and sea surface salinity <xref ref-type="bibr" rid="bib1.bibx54" id="paren.15"/>. In practice, the primary limitation in finding the minimum of the cost function is the large amount of computational resources required due to non-linear or chaotic elements of the system.</p>
      <p id="d2e181">In the context of a full non-linear ESM, the use of adjoint models faces several challenges. Applying an adjoint model to a state-of-the-art Earth system problem is primarily limited by the very large number of state variables <inline-formula><mml:math id="M1" display="inline"><mml:mi mathvariant="script">O</mml:mi></mml:math></inline-formula>(10<sup>7</sup>–10<sup>8</sup>), requiring significant computational resources and observational constraints. However, more fundamental is the fact that non-linear dynamics of the system limit the applicability of adjoint methods to the Earth system predictability timescale. This can lead to exponentially growing adjoint sensitivities as a result of multiple local minima in the cost function. Under such circumstances, spikes occur in the estimated gradients, and the cost function becomes very rough by showing an increasing number of local minima <xref ref-type="bibr" rid="bib1.bibx33 bib1.bibx36" id="paren.16"/>. Fortunately, the problem can be mitigated through synchronisation, which removes the non-linear or chaotic dynamics, leading to a smooth cost function <xref ref-type="bibr" rid="bib1.bibx2" id="paren.17"/>. This method allows for the extension of the assimilation window beyond the predictability timescale, provided that sufficient observations are available. However, this solution comes at the expense of a violation of the original model equations.</p>
      <p id="d2e215">The creation of an adjoint model code from the forward code usually requires considerable effort. Automatic differentiation tools, such as those of <xref ref-type="bibr" rid="bib1.bibx19" id="text.18"/> and <xref ref-type="bibr" rid="bib1.bibx27" id="text.19"/>, were developed to aid in this step. But substantial changes to the forward model code are required unless it was already developed with the adjoint modelling in mind. <xref ref-type="bibr" rid="bib1.bibx54" id="text.20"/> created the first adjoint of an intermediate-complexity fully coupled Earth system model that is automatically created from the forward model by automatic differentiation using the Transformation of Algorithms in FORTRAN (TAF) compiler, called the Centrum für Erdsystemforschung und Nachhaltigkeit (CEN) Earth System Assimilation Model (CESAM). The adjoint of this intermediate-complexity model is intended to be utilised for tuning more complex CMIP-type models through parameter estimation since the basic underlying physics are very similar. Otherwise, this is a manual process with considerable ambiguity in the choice of parameters <xref ref-type="bibr" rid="bib1.bibx42" id="paren.21"/>.</p>
      <p id="d2e230">Therefore, we propose a novel framework in which we use two climate models that are both coupled through synchronisation, one with a high complexity and the other of intermediate complexity, for which an adjoint exists to address the second problem. The technique also has a much wider range of additional applications since resolutions using the adjoint method lag behind those applications featuring simpler assimilation methods as variational methods are typically a factor of 100 more costly than running the associated forward model. For example, the global GECCO3 ocean reanalysis based on the adjoint method <xref ref-type="bibr" rid="bib1.bibx32" id="paren.22"/> features only a nominal resolution of 0.4°, while the GOFS 3.1 <xref ref-type="bibr" rid="bib1.bibx35" id="paren.23"/> based on 3D-var <xref ref-type="bibr" rid="bib1.bibx8" id="paren.24"/> features a <inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">12</mml:mn></mml:mrow></mml:math></inline-formula>° resolution. Employing coarser versions of the adjoint while still running the forward model with full resolution could significantly reduce the cost of the assimilation effort. Therefore, the objective of this paper is to investigate the accuracy and precision of such a synchronised data assimilation approach. We perform this test using a Lorenz 63 model.</p>
      <p id="d2e254">The Lorenz 63 system <xref ref-type="bibr" rid="bib1.bibx38" id="paren.25"/> is a well-established proxy model of chaotic fluid systems, such as the atmosphere <xref ref-type="bibr" rid="bib1.bibx18 bib1.bibx43 bib1.bibx48 bib1.bibx55 bib1.bibx34 bib1.bibx30 bib1.bibx68 bib1.bibx9 bib1.bibx13" id="paren.26"/>. The advantage is that it can be used to rapidly evaluate parameter estimation techniques in data assimilation schemes prior to their application in a full ESM with low computational resource requirements. New modelling techniques can thus be trialled in fast experiments <xref ref-type="bibr" rid="bib1.bibx46 bib1.bibx58 bib1.bibx21 bib1.bibx41 bib1.bibx69" id="paren.27"/>. It can also be used in a wide range of other applications including, but not limited to, data assimilation, stochastic modelling terms, and predictions <xref ref-type="bibr" rid="bib1.bibx12 bib1.bibx7 bib1.bibx47" id="paren.28"/>. The system generates a three-dimensional, time-varying trajectory which, with variation of both model parameters and initial conditions, will produce very different trajectories. Thus, it is an ideal test bed for non-linear modelling in a number of fields <xref ref-type="bibr" rid="bib1.bibx28" id="paren.29"/>. The Lyapunov exponent of the Lorenz 63 model is directly dependent upon its parameters, making it ideal for climatological parameter estimation experiments. For our specific case, these properties make it ideal to evaluate our technique's merits. In a previous study, <xref ref-type="bibr" rid="bib1.bibx39" id="text.30"/> used the Lorenz 63 model and its adjoint to fit a single parameter <inline-formula><mml:math id="M5" display="inline"><mml:mi mathvariant="italic">ρ</mml:mi></mml:math></inline-formula> and the initial conditions <inline-formula><mml:math id="M6" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> to observations. This present study builds on <xref ref-type="bibr" rid="bib1.bibx39" id="text.31"/> to simultaneously fit all three model parameters and uses a model with an adjoint to optimise the parameters of another model without one.</p>
      <p id="d2e306">The structure of the paper is as follows: in Sect. <xref ref-type="sec" rid="Ch1.S2"/>, we outline the methodology of synchronisation, the cost function, and the adjoint method. Section <xref ref-type="sec" rid="Ch1.S3"/> introduces the Lorenz 63 model and describes our reference setup before introducing our two novel multi-model methods, describing our minimisation algorithm, and detailing our statistical metrics for evaluating results. Section <xref ref-type="sec" rid="Ch1.S4"/> shows and discusses the results of our multi-model setups using a single-model setup as a baseline for comparison. The results of introducing a mismodelling term into the adjoint model are also included. A summary and concluding remarks are given in Sect. <xref ref-type="sec" rid="Ch1.S5"/>.</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Methodology</title>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Synchronisation</title>
      <p id="d2e332">In chaotic systems, integrating over periods longer than the predictability timescale creates problems for accurate parameter estimation. This is due to exponentially growing gradients and a maximum likelihood estimate with an increasing number of local maxima <xref ref-type="bibr" rid="bib1.bibx33 bib1.bibx36" id="paren.32"/>. The non-linear or chaotic dynamics, which detrimentally effect the maximum likelihood estimate, can be removed by synchronisation <xref ref-type="bibr" rid="bib1.bibx2 bib1.bibx56" id="paren.33"/>, which transforms the chaotic model into one with linear dynamics and without positive Lyapunov exponents, leading to maximum likelihood functions with one unique maxima. This can be implemented into a generic model of ordinary differential equations:

            <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M7" display="block"><mml:mrow><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M8" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the state vector, <inline-formula><mml:math id="M9" display="inline"><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:math></inline-formula> is the parameter vector, and <inline-formula><mml:math id="M10" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula> is the time. Synchronisation can be incorporated by adding a term which penalises the difference between the model and observations. This term is simply added to the equations:

            <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M11" display="block"><mml:mrow><mml:mover accent="true"><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M12" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> is the synchronisation coefficient, and <inline-formula><mml:math id="M13" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the observation state vector.</p>
      <p id="d2e512">According to the law of large numbers, both with perfect models and in the presence of noise, the precision of the recovered parameters will improve with increasing window length since more data are integrated into the estimation. Similar benefits could be achieved by averaging estimates obtained over short windows, for which no synchronisation is necessary. However, underlying restrictions differ. For synchronisation, noise affects the state over the entire window, whereas, for short windows, noise effects are transported. Short window assimilation can be of benefit in perfect model settings from the error growth, as suggested by the quasi-static variational assimilation (QSVA) framework <xref ref-type="bibr" rid="bib1.bibx48" id="paren.34"/>, due to fact that sensitivities increase exponentially with time in chaotic models. The analogue of this QSVA effect in the dynamical state and parameter estimation (DSPE) method <xref ref-type="bibr" rid="bib1.bibx1" id="paren.35"/> is the attempt to reduce the synchronisation parameter as the optimisation progresses and parameters move closer to their true values. Since errors and sensitivities grow exponentially, feasible window lengths in QSVA have a maximum value due to limited numerical precision. Similarly, synchronisation parameters cannot approach zero for assimilation windows much larger than the predictability limit because synchronisation will eventually fail if positive Lyapunov exponents exist <xref ref-type="bibr" rid="bib1.bibx49" id="paren.36"/>. We note that the reasoning regarding the need for long assimilation windows is somewhat different in the context of a full ESM, for which it is essential to resolve long-timescale physical mechanisms impacted by the specific choice of parameters, such as air–sea interactions of advection timescales in the ocean.</p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>The cost function</title>
      <p id="d2e532">As previously mentioned, in the context of variational data assimilation, a cost function <inline-formula><mml:math id="M14" display="inline"><mml:mi>J</mml:mi></mml:math></inline-formula> must be introduced, measuring the quadratic misfit between the model trajectory and observations. For the case of perfectly known initial conditions but uncertain parameters <inline-formula><mml:math id="M15" display="inline"><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M16" display="inline"><mml:mi>J</mml:mi></mml:math></inline-formula> takes the generic form

            <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M17" display="block"><mml:mtable class="split" rowspacing="0.2ex" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:mi>J</mml:mi></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub></mml:mrow></mml:mfenced><mml:mi>T</mml:mi></mml:msup><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close=")" open="("><mml:mrow><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mo>+</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mi>N</mml:mi></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi mathvariant="bold-italic">h</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mtext mathvariant="italic">T</mml:mtext></mml:msup><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi mathvariant="bold-italic">h</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>

          where <inline-formula><mml:math id="M18" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> is the total integration time, <inline-formula><mml:math id="M19" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is the known uncertainty associated with the observational noise, <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">h</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the measurement function of the model's predicted state vector <inline-formula><mml:math id="M21" display="inline"><mml:mi mathvariant="bold-italic">x</mml:mi></mml:math></inline-formula>, and <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the observation state vector. The prior parameter information with the associated uncertainty is denoted by <inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M24" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, respectively. The global minimum of this function is the maximum likelihood estimate of the model parameter values relative to the observations and prior information.</p>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>The adjoint method and the cost function gradient</title>
      <p id="d2e823">To aid in the minimisation of the cost function, it is standard practice to calculate its gradient and to use this to iteratively adjust control parameters. The adjoint model is introduced to provide these cost function gradients and requires the generation of the adjoint of the forward model equations. The resulting adjoint model can then be integrated in the reverse direction to give the gradient of the cost function. The background term in Eq. (<xref ref-type="disp-formula" rid="Ch1.E3"/>) can be omitted assuming a well-posed problem without prior information on the parameter. Therefore, the gradient of the cost function with respect to the parameters is

            <disp-formula id="Ch1.E4" content-type="numbered"><label>4</label><mml:math id="M25" display="block"><mml:mrow><mml:msub><mml:mi mathvariant="bold">∇</mml:mi><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:msub><mml:mi>J</mml:mi><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mi>N</mml:mi></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:msub><mml:mo>∂</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:msub><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M26" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> is again the total integration time period, <inline-formula><mml:math id="M27" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the adjoint vector at time <inline-formula><mml:math id="M28" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>, and <inline-formula><mml:math id="M29" display="inline"><mml:mrow><mml:msub><mml:mo>∂</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:msub><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the partial differential of the model with respect to the model parameters at time <inline-formula><mml:math id="M30" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Experimental setup</title>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Lorenz 63 model</title>
      <p id="d2e987">In this study, we use the Lorenz 63 system for all of our experiments <xref ref-type="bibr" rid="bib1.bibx38" id="paren.37"/>. The  model is defined by the following equations:
          

                <disp-formula id="Ch1.E5" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M31" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E5.6"><mml:mtd><mml:mtext>5a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E5.7"><mml:mtd><mml:mtext>5b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi>x</mml:mi><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mi>z</mml:mi><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E5.8"><mml:mtd><mml:mtext>5c</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mi>z</mml:mi><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where <inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> denotes the state variables at each given time step, and <inline-formula><mml:math id="M33" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> denotes the model parameters. Throughout this article, we integrate all of our models using the fourth-order Runge–Kutta method with a step size of <inline-formula><mml:math id="M34" display="inline"><mml:mrow><mml:mi mathvariant="normal">Δ</mml:mi><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula> and a total time period of <inline-formula><mml:math id="M35" display="inline"><mml:mn mathvariant="normal">100</mml:mn></mml:math></inline-formula> time units (TUs). This system of equations will be subsequently referred to as the true model with the parameters <inline-formula><mml:math id="M36" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:mn mathvariant="normal">10</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">28</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">8</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">3</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. This true model is used to generate pseudo-observations which will be used for synchronisation, data assimilation, and parameter estimation. Noise is included in these pseudo-observations by adding random values from a Gaussian distribution centred at zero relative to the true trajectory. The random noise magnitudes are bounded to 25 % of the Lorenz 63 system's standard deviation. These pseudo-observations will be labelled as <inline-formula><mml:math id="M37" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>.</p>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Reference setup (single model)</title>
      <p id="d2e1248">To quantify the efficacy of our novel method, we outline a reference setup for a synchronised Lorenz 63 framework similar to those described in <xref ref-type="bibr" rid="bib1.bibx68" id="text.38"/> and <xref ref-type="bibr" rid="bib1.bibx39" id="text.39"/>. We expand the Lorenz 63 model (Eq. <xref ref-type="disp-formula" rid="Ch1.E5"/>) by adding synchronisation terms which then read as follows:
          

                <disp-formula id="Ch1.E9" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M38" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E9.10"><mml:mtd><mml:mtext>6a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E9.11"><mml:mtd><mml:mtext>6b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi>x</mml:mi><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi>x</mml:mi><mml:mi>z</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E9.12"><mml:mtd><mml:mtext>6c</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mi>z</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi>z</mml:mi><mml:mo>)</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          Here, <inline-formula><mml:math id="M39" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> is the synchronisation coefficient, and <inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> denotes the pseudo-observations generated from the true model. We set the initial parameter values of this model at the start of the optimisation as the true system values plus a <inline-formula><mml:math id="M41" display="inline"><mml:mn mathvariant="normal">10</mml:mn></mml:math></inline-formula> % error. This gives <inline-formula><mml:math id="M42" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">11</mml:mn></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M43" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">30.8</mml:mn></mml:mrow></mml:math></inline-formula>, and <inline-formula><mml:math id="M44" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">44</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">15</mml:mn></mml:mrow></mml:math></inline-formula>, which will act as our initial values for the parametric fit. The initial conditions will remain unchanged compared to the true model as our interests are exclusively in climatic parameter estimation. Synchronisation will occur at every time step in all of our setups, and its coefficient will also be present in the adjoint equations. The significance of this will be discussed in Sect. <xref ref-type="sec" rid="Ch1.S4"/> as <inline-formula><mml:math id="M45" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> has a critical role in the precision and accuracy to which parameters can be estimated due to its influence on both the cost function and its gradient.</p>
      <p id="d2e1533">For each state variable in Eq. (<xref ref-type="disp-formula" rid="Ch1.E9"/>), a synchronisation term is included. There are seven possible combinations of these state variables which can be synchronised. The effect of each of the possible choices on the root mean squared error (RMSE) between the true and adjoint systems by varying the synchronisation constant <inline-formula><mml:math id="M46" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> from 0 to 30 is shown in Fig. <xref ref-type="fig" rid="F1"/>. Noise was added (with zero mean and <inline-formula><mml:math id="M47" display="inline"><mml:msqrt><mml:mn mathvariant="normal">2</mml:mn></mml:msqrt></mml:math></inline-formula> standard deviation) to the true system when constructing the pseudo-observations. The figure demonstrates that synchronising the <inline-formula><mml:math id="M48" display="inline"><mml:mi>z</mml:mi></mml:math></inline-formula> component is ineffective at reducing the RMSE <xref ref-type="bibr" rid="bib1.bibx68" id="paren.40"/>. In contrast, synchronising both <inline-formula><mml:math id="M49" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M50" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> proves to be effective, with <inline-formula><mml:math id="M51" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> leading with the lowest RMSE values of the single variable for all values of <inline-formula><mml:math id="M52" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. Synchronising <inline-formula><mml:math id="M53" display="inline"><mml:mrow><mml:mi>x</mml:mi><mml:mi>y</mml:mi><mml:mi>z</mml:mi></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M54" display="inline"><mml:mrow><mml:mi>x</mml:mi><mml:mi>y</mml:mi></mml:mrow></mml:math></inline-formula> achieves the most effective reduction in RMSE for the lowest value of <inline-formula><mml:math id="M55" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. It can be seen in the figure that synchronising <inline-formula><mml:math id="M56" display="inline"><mml:mi>z</mml:mi></mml:math></inline-formula> can lead to model instability. Thus, we choose to only synchronise <inline-formula><mml:math id="M57" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M58" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> in the following research to achieve more stable and accurate results with negligible precision loss.</p>

      <fig id="F1"><label>Figure 1</label><caption><p id="d2e1647">The RMSE between the true and reference model trajectories in seven different synchronisation scenarios. The synchronisation constant <inline-formula><mml:math id="M59" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> is varied from 0 to 30 in steps of 0.5. The solid line is the median, and the shaded area is the 68 % percentile interval for the ensemble of 100 experiments carried out.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f01.png"/>

        </fig>

      <p id="d2e1664">The Lorenz 63 attractors for the trajectories of the true model and that with an adjoint are shown in Fig. <xref ref-type="fig" rid="F2"/>a without synchronisation. A large divergence is visible between the trajectories. However, if synchronisation is introduced, the trajectories become very similar, as shown in Fig. <xref ref-type="fig" rid="F2"/>b. There is now significant overlap between their kernel density estimations (KDEs). KDEs represent a smoothed estimate of the PDF for the model trajectory over a given time period. This allows for convenient visual comparison of trajectories. A more numerically rigorous method to check for effective synchronisation will be discussed in Sect. <xref ref-type="sec" rid="Ch1.S4"/>.</p>

      <fig id="F2"><label>Figure 2</label><caption><p id="d2e1675">The bottom-left quadrants show the Lorenz 63 true and model attractors from the main three variable orientations. The diagonal plots show kernel density estimations (KDEs). Panel <bold>(a)</bold> shows the trajectories without synchronisation. Panel <bold>(b)</bold> shows the trajectories with synchronisation.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f02.png"/>

        </fig>

      <p id="d2e1690">For all subsequent experiments with our setups, a parallel experiment will be performed with this reference setup. The differences in the results can then be compared to evaluate the advantages and disadvantages of the novel techniques.</p>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title>Multi-model data assimilation</title>
      <p id="d2e1701">A multi-model tandem technique is now considered, which consecutively synchronises two forward models before running the adjoint of the second model backward in time. For this purpose, Eq. (<xref ref-type="disp-formula" rid="Ch1.E9"/>) must be modified to incorporate a consecutive synchronisation. A schematic of this setup is provided in Fig. <xref ref-type="fig" rid="F3"/>, and the implications of the two possible ways to calculate the cost function are discussed in the subsequent subsections.</p>
      <p id="d2e1708">The first model has no adjoint equations and is the target model for which we wish to optimise the parameters. The equations of model 1, which is run only in forward mode, are
          

                <disp-formula id="Ch1.E13" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M60" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E13.14"><mml:mtd><mml:mtext>7a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E13.15"><mml:mtd><mml:mtext>7b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>f</mml:mi></mml:msub><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E13.16"><mml:mtd><mml:mtext>7c</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where the subscript f denotes the forward run of model 1, and the subscript o denotes observations generated from the true model. The system of equations for model 2, which has an adjoint, will now be modified to synchronise with the forward-only model and not the observations:
          

                <disp-formula id="Ch1.E17" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M61" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E17.18"><mml:mtd><mml:mtext>8a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E17.19"><mml:mtd><mml:mtext>8b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E17.20"><mml:mtd><mml:mtext>8c</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where the subscript a denotes model 2, which has an adjoint. This model synchronises with model 1 but never directly with the observations.</p>

      <fig id="F3"><label>Figure 3</label><caption><p id="d2e2089">Illustration of the multi-model setup where each pseudo-observation generated from the true model includes random additive Gaussian noise. The cost function can measure the difference between the observations and either model 1 or model 2 depending on the assumptions made. Both options are discussed in the text.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f03.png"/>

        </fig>

<sec id="Ch1.S3.SS3.SSS1">
  <label>3.3.1</label><title>Setup 1 – state-filtered data assimilation (SFDA)</title>
      <p id="d2e2106">Assuming that both model 1 and model 2 can be thought of as representing two identical climate models, the cost function can be placed on model 2. This allows model 1 to filter out some of the background noise in the observations before they are given to the cost function attached to model 2. Such a filtering setup would theoretically reduce parametric uncertainty below that of traditional single-model data assimilation because model 1 should act to reduce the amount of noise synchronised into model 2. We will subsequently refer to this setup as state-filtered data assimilation (SFDA.)</p>
      <p id="d2e2109">In SFDA the cost function acts to constrain model 2. The cost function is

              <disp-formula id="Ch1.E21" content-type="numbered"><label>9</label><mml:math id="M62" display="block"><mml:mrow><mml:msub><mml:mi>J</mml:mi><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mi>T</mml:mi></mml:msup><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>

            <inline-formula><mml:math id="M63" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> is again the total number of time steps of the assimilation window, and <inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is the uncertainty associated with the observation noise. The adjoint matrix includes terms arising from the second tandem layer of synchronisation for model 2. This is given by

              <disp-formula id="Ch1.E22" content-type="numbered"><label>10</label><mml:math id="M65" display="block"><mml:mrow><mml:mtable rowspacing="0.2ex" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:msubsup><mml:mi mathvariant="bold">M</mml:mi><mml:mtext>SFDA</mml:mtext><mml:mo>∗</mml:mo></mml:msubsup><mml:mo>=</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mfenced open="(" close=")"><mml:mtable class="matrix" columnalign="center center center center center center" framespacing="0em"><mml:mtr><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi mathvariant="italic">σ</mml:mi></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi></mml:mrow></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi mathvariant="italic">α</mml:mi></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mi mathvariant="italic">α</mml:mi></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mi mathvariant="italic">σ</mml:mi></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced></mml:mrow></mml:mtd></mml:mtr></mml:mtable><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

            which in practice is numerically evaluated using the automatic differentiation (AD) package JAX to calculate the vector Jacobian product.</p>
      <p id="d2e2483">The adjoint equation for SFDA is given by
            

                  <disp-formula id="Ch1.E23" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M66" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E23.24"><mml:mtd><mml:mtext>11a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mtable rowspacing="0.2ex" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close=")" open="("><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msubsup><mml:mi mathvariant="bold">M</mml:mi><mml:mtext>SFDA</mml:mtext><mml:mo>∗</mml:mo></mml:msubsup><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mtext>for</mml:mtext><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mi>N</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E23.25"><mml:mtd><mml:mtext>11b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mtext>with</mml:mtext><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mspace width="0.25em" linebreak="nobreak"/><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>N</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mn mathvariant="bold">0</mml:mn><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            These equations were derived using the method detailed in <xref ref-type="bibr" rid="bib1.bibx57" id="text.41"/>. The gradient can then be calculated with respect to the parameters <inline-formula><mml:math id="M67" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> notated by the subscript <inline-formula><mml:math id="M68" display="inline"><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:math></inline-formula>. This yields

              <disp-formula id="Ch1.E26" content-type="numbered"><label>12</label><mml:math id="M69" display="block"><mml:mrow><mml:msub><mml:mi mathvariant="bold">∇</mml:mi><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:msub><mml:msub><mml:mi>J</mml:mi><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mi>N</mml:mi></mml:mrow><mml:mn mathvariant="normal">0</mml:mn></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>SFDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mfenced close=")" open="("><mml:mtable class="matrix" columnalign="center" framespacing="0em"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

            which is a component-wise multiplication at each time step.</p>
</sec>
<sec id="Ch1.S3.SS3.SSS2">
  <label>3.3.2</label><title>Setup 2 – tandem data assimilation (TDA)</title>
      <p id="d2e2887">In this section, we want to explore if using an existing adjoint from one model could be utilised to optimise a different target model without an adjoint. This will be referred to as TDA. In TDA, we assume that both models may differ in terms of resolution or numerical formulation but are governed by the same continuum dynamics. Instead of interpolating or transforming the original model variables onto the adjoint model grid, formulation of the adjoint model through synchronisation would provide a simpler means to do this as only essential parameters need to be interpolated. Auxiliary variables and parameters will be generated by the synchronised model, including those that may not exist in the target model.</p>
      <p id="d2e2890">The cost function of TDA is

              <disp-formula id="Ch1.E27" content-type="numbered"><label>13</label><mml:math id="M70" display="block"><mml:mrow><mml:msub><mml:mi>J</mml:mi><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mi>T</mml:mi></mml:msup><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>

            <inline-formula><mml:math id="M71" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> is the total number of time steps of the assimilation window, and <inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is the uncertainty associated with the observation noise. This measures the quadratic misfit between the forward-only model 1 and the observations. Model 1, <inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, will be constrained by this cost function, and its gradient will be calculated using the adjoint of model 2, <inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. In this formulation, the two systems are no longer considered to be a single synchronised one but rather two separate models, one for the calculation of the trajectory and the other for calculating the gradient from its adjoint. The algorithm is also no longer exact as we only assume that model 2's adjoint will provide a good approximation as long as its trajectory follows model 1 closely and is driven by the model–data differences of model 1.  The adjoint matrix is

              <disp-formula id="Ch1.E28" content-type="numbered"><label>14</label><mml:math id="M75" display="block"><mml:mrow><mml:msubsup><mml:mi mathvariant="bold">M</mml:mi><mml:mtext>TDA</mml:mtext><mml:mo>∗</mml:mo></mml:msubsup><mml:mo>=</mml:mo><mml:mfenced close=")" open="("><mml:mtable class="matrix" columnalign="center center center" framespacing="0em"><mml:mtr><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi mathvariant="italic">σ</mml:mi></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

            which is numerically evaluated using AD. The synchronisation with model 1 ensures that the trajectory of model 2 <inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> closely follows that of model 1.</p>
      <p id="d2e3156">The adjoint equation for TDA is given by
            

                  <disp-formula id="Ch1.E29" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M77" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E29.30"><mml:mtd><mml:mtext>15a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mtable rowspacing="0.2ex" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mo mathvariant="normal">˙</mml:mo></mml:mover><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">o</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msubsup><mml:mi mathvariant="bold">M</mml:mi><mml:mtext>TDA</mml:mtext><mml:mo>∗</mml:mo></mml:msubsup><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mtext>for</mml:mtext><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mi>N</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E29.31"><mml:mtd><mml:mtext>15b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mtext>with</mml:mtext><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mspace linebreak="nobreak" width="0.25em"/><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>N</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mn mathvariant="bold">0</mml:mn><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            The gradient with respect to the parameters <inline-formula><mml:math id="M78" display="inline"><mml:mrow><mml:mi mathvariant="bold-italic">θ</mml:mi><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>  is calculated to be

              <disp-formula id="Ch1.E32" content-type="numbered"><label>16</label><mml:math id="M79" display="block"><mml:mrow><mml:msub><mml:mi mathvariant="bold">∇</mml:mi><mml:mi mathvariant="bold-italic">θ</mml:mi></mml:msub><mml:msub><mml:mi>J</mml:mi><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mi>N</mml:mi></mml:mrow><mml:mn mathvariant="normal">0</mml:mn></mml:munderover><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi><mml:msub><mml:mi mathvariant="bold-italic">λ</mml:mi><mml:mtext>TDA</mml:mtext></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mfenced open="(" close=")"><mml:mtable class="matrix" columnalign="center" framespacing="0em"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

            which is again a component-wise multiplication at each time step. For TDA and SFDA, the trajectories of both models and adjoint vectors are stored for evaluation of the gradient.</p>
</sec>
</sec>
<sec id="Ch1.S3.SS4">
  <label>3.4</label><title>Minimisation algorithm</title>
      <p id="d2e3461">To assimilate the data, we fit one of the synchronised models to the observations by optimising the model parameters. A cost function is constructed to calculate the misfit between observations and the model of interest. The gradient of the cost function with respect to the model parameters is always calculated using the adjoint method. However, the form of the adjoint will vary between the two methods we presented in Eqs. (<xref ref-type="disp-formula" rid="Ch1.E23"/>) and (<xref ref-type="disp-formula" rid="Ch1.E29"/>). The adjoint model is numerically evaluated by AD of model 2. This is done in the Python package JAX, which numerically evaluates the vector Jacobian product of the model with respect to its state variable vector <xref ref-type="bibr" rid="bib1.bibx6" id="paren.42"/>. This is then integrated using an inverse Runge–Kutta scheme. Our code stores the state variables and adjoint vectors at each time step. It is also possible to carry out the entire integration using JAX. The process of synchronising all models, calculating the cost function and its gradient, and then adjusting model parameters is carried out iteratively by our chosen minimisation algorithm. Throughout all steps the parameter values of forward-only and adjoint models are identical and optimised simultaneously.</p>
</sec>
<sec id="Ch1.S3.SS5">
  <label>3.5</label><title>Statistical metrics</title>
      <p id="d2e3480">To get a more robust quantification of our setup's behaviour, it is necessary to repeat our study over a number of data sets to calculate medians and percentile intervals (PIs). This allows us to examine general traits of our model without an individual noise event obscuring trends and features of significance. Here this is done by generating 100 pseudo-data sets and assimilating each set independently. The plotting package is then directly applied to these 100 outputs to plot the median and <inline-formula><mml:math id="M80" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula> % PIs. The PIs are included to illustrate the statistical spread of the results and reproducibility and not to explicitly indicate uncertainty. Hence, we choose <inline-formula><mml:math id="M81" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula> % for our PIs to give a concise visualisation of the central <inline-formula><mml:math id="M82" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="italic">σ</mml:mi></mml:mrow></mml:math></inline-formula> of results. The mean percentage error and uncertainty are plotted separately to allow for quantification of both the accuracy and precision of our results. These are calculated by

            <disp-formula id="Ch1.E33" content-type="numbered"><label>17</label><mml:math id="M83" display="block"><mml:mrow><mml:mtable class="split" rowspacing="0.2ex" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mtext>mean</mml:mtext><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="italic">%</mml:mi><mml:mtext>error</mml:mtext><mml:mo>=</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mn mathvariant="normal">100</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="italic">%</mml:mi><mml:mo>⋅</mml:mo><mml:msqrt><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mn mathvariant="normal">3</mml:mn></mml:mfrac></mml:mstyle><mml:mo>⋅</mml:mo><mml:mfenced close="]" open="["><mml:mrow><mml:msup><mml:mfenced open="(" close=")"><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mfenced close=")" open="("><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mfenced open="(" close=")"><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="italic">β</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfenced></mml:mrow></mml:msqrt></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>

          and

            <disp-formula id="Ch1.E34" content-type="numbered"><label>18</label><mml:math id="M84" display="block"><mml:mrow><mml:mtable rowspacing="0.2ex" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mtext>mean</mml:mtext><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="italic">%</mml:mi><mml:mtext>uncertainty</mml:mtext><mml:mo>=</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mn mathvariant="normal">100</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="italic">%</mml:mi><mml:mo>⋅</mml:mo><mml:msqrt><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mn mathvariant="normal">3</mml:mn></mml:mfrac></mml:mstyle><mml:mo>⋅</mml:mo><mml:mfenced close="]" open="["><mml:mrow><mml:msup><mml:mfenced open="(" close=")"><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">Δ</mml:mi><mml:mi mathvariant="italic">σ</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mfenced close=")" open="("><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">Δ</mml:mi><mml:mi mathvariant="italic">ρ</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">ρ</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mfenced close=")" open="("><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">Δ</mml:mi><mml:mi mathvariant="italic">β</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfenced></mml:mrow></mml:msqrt><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>

          The error is calculated by percentage difference between the fitted and true parameter values. The parametric uncertainty is calculated by the minimisation algorithm using a Hessian estimate.</p>
      <p id="d2e3728">After parameter estimation, the optimised parameters are used to initialise a free unsynchronised run of the model. The attractors are plotted against the attractor of the true model. In all cases with synchronisation greater than or equal to the optimum value, the attractors' KDE shows precise and consistent agreement with that of the true model. Results are not displayed as there are no differences which merit discussion. Thus, our focus in the subsequent results is on comparing how the examined setups differ in terms of the accuracy and precision of optimised parameters recovered.</p>
</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Results</title>
      <p id="d2e3741">Throughout the following section, we will use the single model described in <xref ref-type="bibr" rid="bib1.bibx39" id="text.43"/> as our benchmark to compare the new setups against. To understand the behaviour of the setups at different operating extremes, assimilations are carried out for variations of observational noise and <inline-formula><mml:math id="M85" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. This will help establish the optimal synchronisation strength dependent on the noise amplitude. We will also be able to compare the errors and uncertainties of the single model with our multi-model setups.</p>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>SFDA (setup 1)</title>
      <p id="d2e3761">The results from a scan of <inline-formula><mml:math id="M86" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> are shown in Fig. <xref ref-type="fig" rid="F4"/>. The single-model scan has two main regions. The first, for <inline-formula><mml:math id="M87" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>≤</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>, is where the system is poorly synchronised, leading to an inaccurate fit of the parameters and an unstable median value. The second, where <inline-formula><mml:math id="M88" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>, is where the system is fully synchronised and recovers the true model parameters very effectively. SFDA has a higher onset of effective synchronisation than the single-model setup, beginning at <inline-formula><mml:math id="M89" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">11</mml:mn></mml:mrow></mml:math></inline-formula>. Above <inline-formula><mml:math id="M90" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">12.5</mml:mn></mml:mrow></mml:math></inline-formula>, SFDA has consistently more accurate parameter recovery than the single-model setup, while the opposite holds below <inline-formula><mml:math id="M91" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">12.5</mml:mn></mml:mrow></mml:math></inline-formula>. The minimum errors at the respective optimal <inline-formula><mml:math id="M92" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> values are nearly identical.</p>

      <fig id="F4"><label>Figure 4</label><caption><p id="d2e3843">The percentage error between the true values of <inline-formula><mml:math id="M93" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and the fitted value from SFDA. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (lines) and <inline-formula><mml:math id="M94" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted. The noise level is <inline-formula><mml:math id="M95" display="inline"><mml:mn mathvariant="normal">25</mml:mn></mml:math></inline-formula> %.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f04.png"/>

        </fig>

      <p id="d2e3886">Figure <xref ref-type="fig" rid="F5"/> shows the results of two fits carried out for data with an applied noise of <inline-formula><mml:math id="M96" display="inline"><mml:mn mathvariant="normal">25</mml:mn></mml:math></inline-formula> % relative to the systems' standard deviations. The mean percentage uncertainty over the three parameters is plotted for both setups. Noticing in particular the spread of the percentile intervals, the single-model setup is found to be synchronised and has a high precision from <inline-formula><mml:math id="M97" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula> and an SFDA from <inline-formula><mml:math id="M98" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">11.5</mml:mn></mml:mrow></mml:math></inline-formula>. Once the SFDA setup is synchronised, it is found to have a reduced uncertainty compared with the single model. SFDA is found to be approximately one-third more precise than the single-model setup for all values of <inline-formula><mml:math id="M99" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> investigated. However, since SFDA requires larger <inline-formula><mml:math id="M100" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>, uncertainty values for the same <inline-formula><mml:math id="M101" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> cannot be directly compared.</p>
      <p id="d2e3945">Figure <xref ref-type="fig" rid="F4"/> suggests that the SFDA technique is more accurate than a standard single-model setup at higher values of <inline-formula><mml:math id="M102" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. Figure <xref ref-type="fig" rid="F5"/> additionally shows that SFDA is more precise than a single model for all values of <inline-formula><mml:math id="M103" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> after the onset of effective synchronisation. In cases where accuracy and precision are desirable and where computational resources and time are available, this would support the use of SFDA. However,  Fig. <xref ref-type="fig" rid="F4"/> suggests that the error estimates do not represent the actual achieved accuracy of the parameter estimation. Both metrics suggest that the accuracy is less sensitive to the choice of <inline-formula><mml:math id="M104" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> for SFDA. It is also important to note that, in the context where precision is the priority, the lowest value of <inline-formula><mml:math id="M105" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> after effective synchronisation should be chosen. Increasing values of <inline-formula><mml:math id="M106" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> increase the parametric uncertainty because of the associated declining sensitivity of the trajectory to parameter changes. This is also the reason why, for the same <inline-formula><mml:math id="M107" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>, SFDA shows consistently better performance: the less efficient indirect constraint to the observations makes it more sensitive to the parameters.</p>
      <p id="d2e3997">Figure <xref ref-type="fig" rid="F6"/> shows the results of varying the noise levels on the fitted parameter values. For all applied noise levels, the quality of the fit can be considered to be good as the median of the mean percentage uncertainty of the parameters remains below 0.5 % even with noise levels of up to 50 %. SFDA is found to have a mean error performance similar to the single-model system across the range of noise levels tested. However, the spread of the error is slightly improved in the double-model setup at low noise due to the forward-only model smoothing outlying observation better than a single-model setup. The parametric uncertainty is found to be consistently reduced in the double-model system for all noise levels. This demonstrates the precision improvement achieved by running the forward model twice to smooth the observations before carrying out data assimilation. The consequences of this are that, for smaller models, where computational resources are available and where improved precision or accuracy are desirable, SFDA can reduce error and, in particular, decrease uncertainty.</p>

      <fig id="F5"><label>Figure 5</label><caption><p id="d2e4004">The average percentage uncertainty of the three parameters <inline-formula><mml:math id="M108" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> after SFDA from the minimisation algorithm for different values of <inline-formula><mml:math id="M109" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (line) and <inline-formula><mml:math id="M110" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f05.png"/>

        </fig>

      <fig id="F6"><label>Figure 6</label><caption><p id="d2e4049">The percentage error between the true values of <inline-formula><mml:math id="M111" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and those from SFDA, as well as average percentage uncertainty of the SFDA parameters. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (line) and <inline-formula><mml:math id="M112" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted. The noise level varies between 5 % and 50 % in steps of 5 %.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f06.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS2">
  <label>4.2</label><title>TDA (setup 2)</title>
      <p id="d2e4093">Similarly to SFDA, TDA results are evaluated in terms of percentage error and uncertainty estimates against the single model with a data noise level of <inline-formula><mml:math id="M113" display="inline"><mml:mn mathvariant="normal">25</mml:mn></mml:math></inline-formula> %. Error estimates are shown in Fig. <xref ref-type="fig" rid="F7"/> as a function of <inline-formula><mml:math id="M114" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>. For <inline-formula><mml:math id="M115" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:math></inline-formula>, synchronisation starts to set in, and parameter estimation begins to improve. The system only synchronises effectively for <inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula> to consistently recover the true model parameters, as is visible from the small spread of the error. The TDA scan follows the behaviour of the primary model very closely, with no visible disadvantage.</p>

      <fig id="F7"><label>Figure 7</label><caption><p id="d2e4138">The percentage error between the true values of <inline-formula><mml:math id="M117" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and those from TDA. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different pseudo-data sets. The median (lines) and <inline-formula><mml:math id="M118" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted. The noise level is <inline-formula><mml:math id="M119" display="inline"><mml:mn mathvariant="normal">25</mml:mn></mml:math></inline-formula> %.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f07.png"/>

        </fig>

      <p id="d2e4181">The mean percentage uncertainty over the three parameters is plotted in Fig. <xref ref-type="fig" rid="F8"/> for both setups. TDA is found to have almost identical uncertainty compared to the single model. The plot consists of two regions. The first, for <inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>&lt;</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>, is the region where the model is not yet consistently synchronised producing high variability depending on the specific noise. The second, for <inline-formula><mml:math id="M121" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>, is where the system is consistently synchronised. The minimum median of the mean parametric uncertainty, after consistent synchronisation begins, is <inline-formula><mml:math id="M122" display="inline"><mml:mrow><mml:mo>≈</mml:mo><mml:mn mathvariant="normal">0.35</mml:mn></mml:mrow></mml:math></inline-formula> % and achieved at <inline-formula><mml:math id="M123" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>. The subsequent increase in uncertainty is due to the reduced the parametric sensitivity associated with the increased <inline-formula><mml:math id="M124" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula>, thereby reducing the curvature of the cost function at the minimum.</p>

      <fig id="F8"><label>Figure 8</label><caption><p id="d2e4243">The percentage uncertainty, from the minimisation algorithm, averaged over all three parameters <inline-formula><mml:math id="M125" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> after TDA. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (lines) and <inline-formula><mml:math id="M126" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f08.png"/>

        </fig>

      <p id="d2e4279">Figure <xref ref-type="fig" rid="F9"/> shows the results of varying the noise levels on the fitted parameter values. For all noise levels studied, the fit can be considered to be accurate as the mean percentage error of the parameters remains below 1 % even with noise levels of 50 %. The increased spread of the error at low noise is thought to be due to the fixed value of <inline-formula><mml:math id="M127" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> used for all noise levels impacting the synchronisation of the system. The TDA setup is found to have extremely consistent uncertainty compared to the single-model system.</p>

      <fig id="F9"><label>Figure 9</label><caption><p id="d2e4293">The percentage error between the true values of <inline-formula><mml:math id="M128" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and those from TDA. A single-model assimilation is included for comparison. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (lines) and <inline-formula><mml:math id="M129" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted. The noise level varies between 5 % and 50 % in steps of 5 %.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f09.png"/>

        </fig>

      <p id="d2e4329">The consistent results of TDA in Figs. <xref ref-type="fig" rid="F7"/>, <xref ref-type="fig" rid="F8"/>, and <xref ref-type="fig" rid="F9"/> relative to the standard single-model setup show that transferring information via synchronisation does not compromise precision. Figures <xref ref-type="fig" rid="F7"/> and <xref ref-type="fig" rid="F8"/> also concur with those of SFDA in suggesting that the optimal value of <inline-formula><mml:math id="M130" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> is the smallest value after the onset of effective synchronisation. An increase in <inline-formula><mml:math id="M131" display="inline"><mml:mi mathvariant="italic">α</mml:mi></mml:math></inline-formula> beyond this point can lead to a significant reduction in the precision of the parameters. In cases where only a simpler but similar model with an adjoint is available, results are likely to degrade. In the following section, we will study the potential impact of model inconsistencies on the precision of the parameter estimation.</p>
</sec>
<sec id="Ch1.S4.SS3">
  <label>4.3</label><title>Mismodelling in TDA</title>
      <p id="d2e4365">In this section, the tandem data assimilation (setup 2) is be used with different forward and adjoint models that share common physics to examine the impact of introducing model discrepancies. We construct a test case where the equations of model 2, which has an adjoint (see Eq. <xref ref-type="disp-formula" rid="Ch1.E17"/>), are modified to give an oscillatory difference to both the true model and model 1. This can be done in a number of ways. We choose to introduce a multiplicative sine function into the equations in such a way that it is also included in the adjoint matrix and thus modifies the gradient values returned to the fitting algorithm. Model 2 with an adjoint is now
          

                <disp-formula id="Ch1.E35" specific-use="gather" content-type="subnumberedsingle"><mml:math id="M132" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E35.36"><mml:mtd><mml:mtext>19a</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E35.37"><mml:mtd><mml:mtext>19b</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mi mathvariant="italic">α</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">f</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E35.38"><mml:mtd><mml:mtext>19c</mml:mtext></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>a</mml:mi></mml:msub><mml:msub><mml:mi>y</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:msub><mml:mi>z</mml:mi><mml:mi mathvariant="normal">a</mml:mi></mml:msub><mml:mo>⋅</mml:mo><mml:mfenced close=")" open="("><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mi mathvariant="italic">ϵ</mml:mi><mml:mi>sin⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="italic">π</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

          where <inline-formula><mml:math id="M133" display="inline"><mml:mi mathvariant="italic">ϵ</mml:mi></mml:math></inline-formula> is the term which determines the strength of the oscillation term. The effect of this term on the attractor is shown in Fig. <xref ref-type="fig" rid="F10"/> without synchronisation. When compared to Fig. <xref ref-type="fig" rid="F2"/>a, it can be seen that this term is successful in distorting the shape and probability density of the attractor.</p>

      <fig id="F10"><label>Figure 10</label><caption><p id="d2e4592">The Lorenz true-model and model-2 attractors in the case of <inline-formula><mml:math id="M134" display="inline"><mml:mrow><mml:mi mathvariant="italic">ϵ</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1.0</mml:mn></mml:mrow></mml:math></inline-formula>. The bottom-left and top-right quadrants show the attractors from all possible coordinate orientations. The diagonal plots show kernel density estimations (KDEs). No noise is added.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f10.png"/>

        </fig>

      <p id="d2e4613">The consequences of varying this term based on the accuracy of parameter optimisation after assimilation are shown in Fig. <xref ref-type="fig" rid="F11"/>. With increasing <inline-formula><mml:math id="M135" display="inline"><mml:mi mathvariant="italic">ϵ</mml:mi></mml:math></inline-formula>, the percentage error and uncertainty between fitted and true systems remain stable. In spite of the large impact this term has on the attractor shape, the figure demonstrates a resilience of TDA to modelling differences between the forward-only model 1 and model 2 with an adjoint.</p>

      <fig id="F11"><label>Figure 11</label><caption><p id="d2e4628">The percentage error <bold>(a)</bold> and uncertainty <bold>(b)</bold> between the true values of <inline-formula><mml:math id="M136" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">ρ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">β</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and those from TDA. An ensemble of 100 assimilations is carried out over 100 different data sets. The median (line) and <inline-formula><mml:math id="M137" display="inline"><mml:mn mathvariant="normal">68</mml:mn></mml:math></inline-formula>th percentile intervals (shaded areas) are plotted. The noise level is <inline-formula><mml:math id="M138" display="inline"><mml:mn mathvariant="normal">25</mml:mn></mml:math></inline-formula> %, and <inline-formula><mml:math id="M139" display="inline"><mml:mrow><mml:mi mathvariant="italic">α</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">7.5</mml:mn></mml:mrow></mml:math></inline-formula>.</p></caption>
          <graphic xlink:href="https://npg.copernicus.org/articles/32/353/2025/npg-32-353-2025-f11.png"/>

        </fig>

</sec>
</sec>
<sec id="Ch1.S5" sec-type="conclusions">
  <label>5</label><title>Conclusions</title>
      <p id="d2e4699">In this paper, we have demonstrated the ability to constrain a Lorenz 63 model using a second model with similar physics and an adjoint by 4D-var data assimilation. Such an approach removes the need to generate an adjoint for a forward model if such an adjoint already exists for a separate yet dynamically similar system. An important application of this technique in Earth system modelling would be a situation where a low-resolution ESM with an adjoint shares a parameterisation with a high-resolution ESM for which no adjoint exists.  Moreover, using a lower-resolution version of the same model could computationally make data assimilation much faster. We have shown that, in both cases, the low-resolution ESM with an adjoint could, through synchronisation, follow the trajectory of the more complex and high-resolution model while, at the same time, providing all necessary variables to run its tangent linear adjoint model. This can then be utilised to estimate parameters in the complex high-resolution ESM. We have also shown that running a forward model twice before beginning data assimilation can act to smooth the data and reduce the parametric uncertainty. Our focus is on optimising the parameters of a full ESM, which will be tested as a next step. It would also be possible to optimise the initial condition of the state variables, which is more applicable to weather forecasting techniques. Future work will examine the resilience of such setups to spatially and temporally sparse data.</p>
</sec>

      
      </body>
    <back><notes notes-type="dataavailability"><title>Data availability</title>

      <p id="d2e4706">The pseudo-data samples used in this study are available on request from the authors. The figures in this research were plotted by the seaborn plotting package <xref ref-type="bibr" rid="bib1.bibx65" id="paren.44"/> (<ext-link xlink:href="https://doi.org/10.21105/joss.03021" ext-link-type="DOI">10.21105/joss.03021</ext-link>, last access: 17 April 2025) utilising our output, which was managed using pandas <xref ref-type="bibr" rid="bib1.bibx60" id="paren.45"/> (<ext-link xlink:href="https://doi.org/10.5281/zenodo.3509134" ext-link-type="DOI">10.5281/zenodo.3509134</ext-link>, last access: 17 April 2025). Automatic differentiation of our model was carried out using JAX <xref ref-type="bibr" rid="bib1.bibx6" id="paren.46"/> (<uri>http://github.com/google/jax</uri>, last access: 17 April 2025). The minimisation of our cost function and the uncertainty evaluation was done by iminuit <xref ref-type="bibr" rid="bib1.bibx10" id="paren.47"/> (<ext-link xlink:href="https://doi.org/10.5281/zenodo.3949207" ext-link-type="DOI">10.5281/zenodo.3949207</ext-link>, last access: 17 April 2025).</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d2e4737">PDK carried out the research and wrote and edited the initial paper draft. AB assisted with the research and edited the paper. AK and DS designed the research concept, supervised the work, and participated in the writing of the paper.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d2e4743">The contact author has declared that none of the authors has any competing interests.</p>
  </notes><notes notes-type="disclaimer"><title>Disclaimer</title>

      <p id="d2e4749">Publisher’s note: Copernicus Publications remains neutral with regard to jurisdictional claims made in the text, published maps, institutional affiliations, or any other geographical representation in this paper. While Copernicus Publications makes every effort to include appropriate place names, the final responsibility lies with the authors.</p>
  </notes><ack><title>Acknowledgements</title><p id="d2e4755">The authors would like to thank the two anonymous referees for providing very detailed and constructive reviews which significantly contributed to the final form of this paper. This work used resources of the Deutsches Klimarechenzentrum (DKRZ), granted by its Scientific Steering Committee (WLA). This research is a contribution to the Centrum für Erdsystemforschung und Nachhaltigkeit (CEN) of Universität Hamburg.</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d2e4760">This research was supported, in part, through the Koselleck grant <inline-formula><mml:math id="M140" display="inline"><mml:mrow><mml:mi>E</mml:mi><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>t</mml:mi><mml:msup><mml:mi>h</mml:mi><mml:mrow><mml:mi>R</mml:mi><mml:mi>A</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> funded by the Deutsche Forschungsgemeinschaft (grant no. 66492934).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d2e4788">This paper was edited by Wansuo Duan and reviewed by two anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>Abarbanel et al.(2009)Abarbanel, Creveling, Farsian, and Kostuk</label><mixed-citation> Abarbanel, H. D., Creveling, D. R., Farsian, R., and Kostuk, M.: Dynamical state and parameter estimation, SIAM J. Appl. Dynam. Syst., 8, 1341–1381, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Abarbanel et al.(2010)Abarbanel, Kostuk, and Whartenby</label><mixed-citation> Abarbanel, H. D., Kostuk, M., and Whartenby, W.: Data assimilation with regularized nonlinear instabilities, Q. J. Roy. Meteor. Soc. A, 136, 769–783, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Allaire(2015)</label><mixed-citation> Allaire, G.: A review of adjoint methods for sensitivity analysis, uncertainty quantification and optimization in numerical codes, Ingénieurs de l'Automobile, 836, 33–36, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>Bar-Shalom et al.(2004)Bar-Shalom, Li, and Kirubarajan</label><mixed-citation>Bar-Shalom, Y., Li, X. R., and Kirubarajan, T.: Estimation with applications to tracking and navigation: theory algorithms and software, John Wiley &amp; Sons, <ext-link xlink:href="https://doi.org/10.1002/0471221279" ext-link-type="DOI">10.1002/0471221279</ext-link>, 2004.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Bertino et al.(2003)Bertino, Evensen, and Wackernagel</label><mixed-citation> Bertino, L., Evensen, G., and Wackernagel, H.: Sequential data assimilation techniques in oceanography, Int. Stat. Rev., 71, 223–241, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Bradbury et al.(2018)Bradbury, Frostig, Hawkins, Johnson, Leary, Maclaurin, Necula, Paszke, VanderPlas, Wanderman-Milne, and Zhang</label><mixed-citation>Bradbury, J., Frostig, R., Hawkins, P., Johnson, M. J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., and Zhang, Q.: JAX: composable transformations of Python+NumPy programs, Github [code], <uri>http://github.com/google/jax</uri>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Cameron and Yang(2019)</label><mixed-citation> Cameron, M. and Yang, S.: Computing the quasipotential for highly dissipative and chaotic SDEs an application to stochastic Lorenz’63, Commun. Appl. Math. Comput. Sci., 14, 207–246, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Cummings and Smedstad(2013)</label><mixed-citation>Cummings, J. A. and Smedstad, O. M.: Variational data assimilation for the global ocean, in: Data assimilation for atmospheric, oceanic and hydrologic applications (Vol. II),  Springer, 303–343, <ext-link xlink:href="https://doi.org/10.1007/978-3-642-35088-7_13" ext-link-type="DOI">10.1007/978-3-642-35088-7_13</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Daron and Stainforth(2015)</label><mixed-citation>Daron, J. and Stainforth, D. A.: On quantifying the climate of the nonautonomous Lorenz-63 model, Chaos, 25, 043103, <ext-link xlink:href="https://doi.org/10.1063/1.4916789" ext-link-type="DOI">10.1063/1.4916789</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>Dembinski and et al.(2020)</label><mixed-citation>Dembinski, H., Ongmongkolkul, P., Deil, C., Schreiner, H., Feickert, M., Burr, C., Watson, J., Rost, F., Pearce, A., Geiger, L., Abdelmotteleb, A., Desai, A., Wiedemann, B. M., Gohlke, C., Sanders, J., Drotleff, J., Eschle, J., Neste, L., Gorelli, M. E., and Zapata, O.: scikit-hep/iminuit, Zenodo [code], <ext-link xlink:href="https://doi.org/10.5281/zenodo.3949207" ext-link-type="DOI">10.5281/zenodo.3949207</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Desroziers et al.(2014)Desroziers, Camino, and Berre</label><mixed-citation> Desroziers, G., Camino, J.-T., and Berre, L.: 4DEnVar: link with 4D state formulation of variational assimilation and different possible implementations, Q. J. Roy. Meteor. Soc., 140, 2097–2110, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Du and Shiue(2021)</label><mixed-citation>Du, Y. J. and Shiue, M.-C.: Analysis and computation of continuous data assimilation algorithms for Lorenz 63 system based on nonlinear nudging techniques, J. Comput. Appl. Math., 386, 113246, <ext-link xlink:href="https://doi.org/10.1016/j.cam.2020.113246" ext-link-type="DOI">10.1016/j.cam.2020.113246</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Errico(1997)</label><mixed-citation> Errico, R. M.: What is an adjoint model?, B. Am. Meteorol. Soc., 78, 2577–2592, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Evensen(1994)</label><mixed-citation> Evensen, G.: Sequential data assimilation with a nonlinear quasi-geostrophic model using Monte Carlo methods to forecast error statistics, J. Geophys. Res.-Oceans, 99, 10143–10162, 1994.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Evensen(2003)</label><mixed-citation> Evensen, G.: The ensemble Kalman filter: Theoretical formulation and practical implementation, Ocean Dynam., 53, 343–367, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Evensen et al.(2022)Evensen, Vossepoel, and Van Leeuwen</label><mixed-citation>Evensen, G., Vossepoel, F. C., and Van Leeuwen, P. J.: Data assimilation fundamentals: A unified formulation of the state and parameter estimation problem, Springer Nature, <ext-link xlink:href="https://doi.org/10.1007/978-3-030-96709-3" ext-link-type="DOI">10.1007/978-3-030-96709-3</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Fisher et al.(2011)Fisher, Tremolet, Auvinen, Tan, and Poli</label><mixed-citation>Fisher, M., Tremolet, Y., Auvinen, H., Tan, D., and Poli, P.: Weak-constraint and long-window 4D-Var, ECMWF Technical Memoranda, 655, 47, <ext-link xlink:href="https://doi.org/10.21957/9ii4d4dsq" ext-link-type="DOI">10.21957/9ii4d4dsq</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>Gauthier(1992)</label><mixed-citation> Gauthier, P.: Chaos and quadri-dimensional data assimilation: A study based on the Lorenz model, Tellus A, 44, 2–17, 1992.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Giering and Kaminski(1998)</label><mixed-citation> Giering, R. and Kaminski, T.: Recipes for adjoint code construction, ACM T. Math. Softw., 24, 437–474, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Goodliff et al.(2015)Goodliff, Amezcua, and Van Leeuwen</label><mixed-citation>Goodliff, M., Amezcua, J., and Van Leeuwen, P. J.: Comparing hybrid data assimilation methods on the Lorenz 1963 model with increasing non-linearity, Tellus A, <ext-link xlink:href="https://doi.org/10.3402/tellusa.v67.26928" ext-link-type="DOI">10.3402/tellusa.v67.26928</ext-link>,  67, 1–12, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Goodliff et al.(2020)Goodliff, Fletcher, Kliewer, Forsythe, and Jones</label><mixed-citation>Goodliff, M., Fletcher, S., Kliewer, A., Forsythe, J., and Jones, A.: Detection of Non-Gaussian Behavior Using Machine Learning Techniques: A Case Study on the Lorenz 63 Model, J. Geophys. Res.-Atmos., 125, e2019JD031551, <ext-link xlink:href="https://doi.org/10.1029/2019JD031551" ext-link-type="DOI">10.1029/2019JD031551</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Gustafsson et al.(1998)Gustafsson, Källén, and Thorsteinsson</label><mixed-citation> Gustafsson, N., Källén, E., and Thorsteinsson, S.: Sensitivity of forecast errors to initial and lateral boundary conditions, Tellus A, 50, 167–185, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>Gustafsson et al.(2001)Gustafsson, Berre, Hörnquist, Huang, Lindskog, Navascues, Mogensen, and Thorsteinsson</label><mixed-citation> Gustafsson, N., Berre, L., Hörnquist, S., Huang, X.-Y., Lindskog, M., Navascues, B., Mogensen, K., and Thorsteinsson, S.: Three-dimensional variational data assimilation for a limited area model: Part I: General formulation and the background error constraint, Tellus A, 53, 425–446, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Hall(1986)</label><mixed-citation> Hall, M. C.: Application of adjoint sensitivity theory to an atmospheric general circulation model, J. Atmos. Sci., 43, 2644–2652, 1986.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Hall and Cacuci(1983)</label><mixed-citation> Hall, M. C. and Cacuci, D. G.: Physical interpretation of the adjoint functions for sensitivity analysis of atmospheric models, J. Atmos. Sci., 40, 2537–2546, 1983.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Hall et al.(1982)Hall, Cacuci, and Schlesinger</label><mixed-citation> Hall, M. C., Cacuci, D. G., and Schlesinger, M. E.: Sensitivity analysis of a radiative-convective model by the adjoint method, J. Atmos. Sci., 39, 2038–2050, 1982.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>Hascoet and Pascual(2013)</label><mixed-citation> Hascoet, L. and Pascual, V.: The Tapenade automatic differentiation tool: Principles, model, and specification, ACM T. Math. Softw., 39, 1–43, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx28"><label>Hirsch et al.(2013)Hirsch, Smale, and Devaney</label><mixed-citation>Hirsch, M. W., Smale, S., and Devaney, R. L.: 14 – The Lorenz System, in: Differential Equations, Dynamical Systems, and an Introduction to Chaos (Third Edition), edited by: Hirsch, M. W., Smale, S., and Devaney, R. L.,     3rd Edn., Academic Press, Boston,  305–328, ISBN 978-0-12-382010-5, <ext-link xlink:href="https://doi.org/10.1016/B978-0-12-382010-5.00014-2" ext-link-type="DOI">10.1016/B978-0-12-382010-5.00014-2</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx29"><label>Houtekamer and Mitchell(2001)</label><mixed-citation> Houtekamer, P. L. and Mitchell, H. L.: A sequential ensemble Kalman filter for atmospheric data assimilation, Mon. Weather Rev., 129, 123–137, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Huai et al.(2017)Huai, Li, Ding, Feng, and Liu</label><mixed-citation> Huai, X.-W., Li, J.-P., Ding, R.-Q., Feng, J., and Liu, D.-Q.: Quantifying local predictability of the Lorenz system using the nonlinear local Lyapunov exponent, Atmos. Ocean. Sci. Lett., 10, 372–378, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Kalman(1960)</label><mixed-citation>Kalman, R. E.: A New Approach to Linear Filtering and Prediction Problems, J. Basic Eng., 82, 35–45, <ext-link xlink:href="https://doi.org/10.1115/1.3662552" ext-link-type="DOI">10.1115/1.3662552</ext-link>, 1960.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Köhl(2020)</label><mixed-citation> Köhl, A.: Evaluating the GECCO3 1948–2018 ocean synthesis – a configuration for initializing the MPI-ESM climate model, Q. J. Roy. Meteor. Soc., 146, 2250–2273, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Köhl and Willebrand(2002)</label><mixed-citation> Köhl, A. and Willebrand, J.: An adjoint method for the assimilation of statistical characteristics into eddy-resolving ocean models, Tellus A, 54, 406–425, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Kravtsov and Tsonis(2021)</label><mixed-citation>Kravtsov, S. and Tsonis, A. A.: Lorenz-63 Model as a Metaphor for Transient Complexity in Climate, Entropy, 23, 951, <ext-link xlink:href="https://doi.org/10.3390/e23080951" ext-link-type="DOI">10.3390/e23080951</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>Laboratory(2016)</label><mixed-citation>Naval Research Laboratory.: Global HYCOM+CICE 1/12 degree page, <uri>https://www7320.nrlssc.navy.mil/GLBhycomcice1-12/</uri> (last access: 17 April 2025), 2016.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>Lea et al.(2000)Lea, Allen, and Haine</label><mixed-citation> Lea, D. J., Allen, M. R., and Haine, T. W.: Sensitivity analysis of the climate of a chaotic system, Tellus A, 52, 523–532, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Le Dimet and Talagrand(1986)</label><mixed-citation> Le Dimet, F.-X. and Talagrand, O.: Variational algorithms for analysis and assimilation of meteorological observations: theoretical aspects, Tellus A, 38, 97–110, 1986.</mixed-citation></ref>
      <ref id="bib1.bibx38"><label>Lorenz(1963)</label><mixed-citation> Lorenz, E. N.: Deterministic nonperiodic flow, J. Atmos. Sci., 20, 130–141, 1963.</mixed-citation></ref>
      <ref id="bib1.bibx39"><label>Lyu et al.(2018)Lyu, Köhl, Matei, and Stammer</label><mixed-citation> Lyu, G., Köhl, A., Matei, I., and Stammer, D.: Adjoint-based climate model tuning: Application to the planet simulator, J. Adv. Model. Earth Sy., 10, 207–222, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx40"><label>Marotzke et al.(1999)Marotzke, Giering, Zhang, Stammer, Hill, and Lee</label><mixed-citation> Marotzke, J., Giering, R., Zhang, K. Q., Stammer, D., Hill, C., and Lee, T.: Construction of the adjoint MIT ocean general circulation model and application to Atlantic heat transport sensitivity, J. Geophys. Res.-Oceans, 104, 29529–29547, 1999.</mixed-citation></ref>
      <ref id="bib1.bibx41"><label>Marzban(2013)</label><mixed-citation> Marzban, C.: Variance-based sensitivity analysis: An illustration on the Lorenz'63 model, Mon. Weather Rev., 141, 4069–4079, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx42"><label>Mauritsen et al.(2012)Mauritsen, Stevens, Roeckner, Crueger, Esch, Giorgetta, Haak, Jungclaus, Klocke, Matei et al.</label><mixed-citation>Mauritsen, T., Stevens, B., Roeckner, E., Crueger, T., Esch, M., Giorgetta, M., Haak, H., Jungclaus, J., Klocke, D., Matei, D., Mikolajewicz, U., Notz, D., Pincus, R., Schmidt, H., and Tomassini, L.: Tuning the climate of a global model, J. Adv. Model. Earth Sy., 4, M00A01,  <ext-link xlink:href="https://doi.org/10.1029/2012MS000154" ext-link-type="DOI">10.1029/2012MS000154</ext-link>,2012.</mixed-citation></ref>
      <ref id="bib1.bibx43"><label>Miller et al.(1994)Miller, Ghil, and Gauthiez</label><mixed-citation> Miller, R. N., Ghil, M., and Gauthiez, F.: Advanced data assimilation in strongly nonlinear dynamical systems, J. Atmos. Sci., 51, 1037–1056, 1994.</mixed-citation></ref>
      <ref id="bib1.bibx44"><label>Navon(2009)</label><mixed-citation>Navon, I. M.: Data assimilation for numerical weather prediction: a review, Data assimilation for atmospheric, oceanic and hydrologic applications, 21–65, <ext-link xlink:href="https://doi.org/10.1007/978-3-540-71056-1_2" ext-link-type="DOI">10.1007/978-3-540-71056-1_2</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx45"><label>Nichols(2010)</label><mixed-citation>Nichols, N. K.: Mathematical Concepts of Data Assimilation, Springer Berlin Heidelberg, Berlin, Heidelberg,  13–39, ISBN 978-3-540-74703-1, <ext-link xlink:href="https://doi.org/10.1007/978-3-540-74703-1_2" ext-link-type="DOI">10.1007/978-3-540-74703-1_2</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx46"><label>Pasini and Pelino(2005)</label><mixed-citation>Pasini, A. and Pelino, V.: Can we estimate atmospheric predictability by performance of neural network forecasting? The toy case studies of unforced and forced Lorenz models, in: CIMSA, 2005 IEEE International Conference on Computational Intelligence for Measurement Systems and Applications, 2005, IEEE, 69–74,  <ext-link xlink:href="https://doi.org/10.1109/CIMSA.2005.1522829" ext-link-type="DOI">10.1109/CIMSA.2005.1522829</ext-link>,2005.</mixed-citation></ref>
      <ref id="bib1.bibx47"><label>Pelino and Maimone(2007)</label><mixed-citation>Pelino, V. and Maimone, F.: Energetics, skeletal dynamics, and long-term predictions on Kolmogorov-Lorenz systems, Phys. Rev. E, 76, 046214, <ext-link xlink:href="https://doi.org/10.1103/PhysRevE.76.046214" ext-link-type="DOI">10.1103/PhysRevE.76.046214</ext-link>, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx48"><label>Pires et al.(1996)Pires, Vautard, and Talagrand</label><mixed-citation> Pires, C., Vautard, R., and Talagrand, O.: On extending the limits of variational assimilation in nonlinear chaotic systems, Tellus A, 48, 96–121, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx49"><label>Quinn et al.(2009)Quinn, Bryant, Creveling, Klein, and Abarbanel</label><mixed-citation>Quinn, J. C., Bryant, P. H., Creveling, D. R., Klein, S. R., and Abarbanel, H. D.: Parameter and state estimation of experimental chaotic systems using synchronization, Phys. Rev. E, 80, 016201, <ext-link xlink:href="https://doi.org/10.1103/PhysRevE.80.016201" ext-link-type="DOI">10.1103/PhysRevE.80.016201</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx50"><label>Rabier and Liu(2003)</label><mixed-citation>Rabier, F. and Liu, Z.: Variational data assimilation: theory and overview, in: Proc. ECMWF Seminar on Recent Developments in Data Assimilation for Atmosphere and Ocean, Reading, UK,  8–12 September,  29–43, <uri>https://www.ecmwf.int/en/elibrary/76079-variational-data-assimiltion-theory-and-overview</uri>, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx51"><label>Ruiz et al.(2013)Ruiz, Pulido, and Miyoshi</label><mixed-citation> Ruiz, J. J., Pulido, M., and Miyoshi, T.: Estimating model parameters with ensemble-based data assimilation: A review, J. Meteor. Soc. Jpn. Ser. II, 91, 79–99, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx52"><label>Simon(2006)</label><mixed-citation>Simon, D.: Optimal state estimation: Kalman, H infinity, and nonlinear approaches, John Wiley &amp; Sons, <ext-link xlink:href="https://doi.org/10.1002/0470045345" ext-link-type="DOI">10.1002/0470045345</ext-link>, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx53"><label>Stammer et al.(2016)Stammer, Balmaseda, Heimbach, Köhl, and Weaver</label><mixed-citation> Stammer, D., Balmaseda, M., Heimbach, P., Köhl, A., and Weaver, A.: Ocean data assimilation in support of climate applications: status and perspectives, Annu. Rev. Mar. Sci., 8, 491–518, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx54"><label>Stammer et al.(2018)Stammer, Köhl, Vlasenko, Matei, Lunkeit, and Schubert</label><mixed-citation> Stammer, D., Köhl, A., Vlasenko, A., Matei, I., Lunkeit, F., and Schubert, S.: A pilot climate sensitivity study using the CEN coupled adjoint model (CESAM), J. Climate, 31, 2031–2056, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx55"><label>Stensrud and Bao(1992)</label><mixed-citation> Stensrud, D. J. and Bao, J.-W.: Behaviors of variational and nudging assimilation techniques with a chaotic low-order model, Mon. Weather Rev., 120, 3016–3028, 1992.</mixed-citation></ref>
      <ref id="bib1.bibx56"><label>Sugiura et al.(2014)Sugiura, Masuda, Fujii, Kamachi, Ishikawa, and Awaji</label><mixed-citation> Sugiura, N., Masuda, S., Fujii, Y., Kamachi, M., Ishikawa, Y., and Awaji, T.: A framework for interpreting regularized state estimation, Mon. Weather Rev., 142, 386–400, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx57"><label>Talagrand(2010)</label><mixed-citation>Talagrand, O.: Variational Assimilation. In: Lahoz, W., Khattatov, B., Menard, R. (eds) Data Assimilation. Springer, Berlin, Heidelberg. <ext-link xlink:href="https://doi.org/10.1007/978-3-540-74703-1_3" ext-link-type="DOI">10.1007/978-3-540-74703-1_3</ext-link>,  2010.</mixed-citation></ref>
      <ref id="bib1.bibx58"><label>Tandeo et al.(2015)Tandeo, Ailliot, Ruiz, Hannart, Chapron, Cuzol, Monbet, Easton, and Fablet</label><mixed-citation>Tandeo, P., Ailliot, P., Ruiz, J., Hannart, A., Chapron, B., Cuzol, A., Monbet, V., Easton, R., and Fablet, R.: Combining analog method and ensemble data assimilation: application to the Lorenz-63 chaotic system, in: Machine Learning and Data Mining Approaches to Climate Science: proceedings of the 4th International Workshop on Climate Informatics,  Springer,  3–12, <ext-link xlink:href="https://doi.org/10.1007/978-3-319-17220-0_1" ext-link-type="DOI">10.1007/978-3-319-17220-0_1</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx59"><label>Tett et al.(2017)Tett, Yamazaki, Mineter, Cartis, and Eizenberg</label><mixed-citation>Tett, S. F. B., Yamazaki, K., Mineter, M. J., Cartis, C., and Eizenberg, N.: Calibrating climate models using inverse methods: case studies with HadAM3, HadAM3P and HadCM3, Geosci. Model Dev., 10, 3567–3589, <ext-link xlink:href="https://doi.org/10.5194/gmd-10-3567-2017" ext-link-type="DOI">10.5194/gmd-10-3567-2017</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx60"><label>The pandas development team(2020)</label><mixed-citation>The pandas development team: pandas-dev/pandas: Pandas, Zenodo [code], <ext-link xlink:href="https://doi.org/10.5281/zenodo.3509134" ext-link-type="DOI">10.5281/zenodo.3509134</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx61"><label>Tippett and Chang(2003)</label><mixed-citation> Tippett, M. K. and Chang, P.: Some theoretical considerations on predictability of linear stochastic dynamics, Tellus A, 55, 148–157, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx62"><label>Tippett et al.(2003)Tippett, Anderson, Bishop, Hamill, and Whitaker</label><mixed-citation>Tippett, M. K., Anderson, J. L., Bishop, C. H., Hamill, T. M., and Whitaker, J. S.: Ensemble square root filters, Mon. Weather Rev., 131, 1485–1490, 2003.  </mixed-citation></ref>
      <ref id="bib1.bibx63"><label>Tremolet(2006)</label><mixed-citation> Tremolet, Y.: Accounting for an imperfect model in 4D-Var, Q. J. Roy. Meteor. Soc., 132, 2483–2504, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx64"><label>Van Der Merwe and Wan(2001)</label><mixed-citation>Van Der Merwe, R. and Wan, E. A.: The square-root unscented Kalman filter for state and parameter-estimation, in: 2001 IEEE international conference on acoustics, speech, and signal processing, Proceedings (Cat. No. 01CH37221), IEEE,  vol. 6,  3461–3464, <ext-link xlink:href="https://doi.org/10.1109/ICASSP.2001.940586" ext-link-type="DOI">10.1109/ICASSP.2001.940586</ext-link>, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx65"><label>Waskom(2021)</label><mixed-citation>Waskom, M. L.: seaborn: statistical data visualization, J. Open Source Softw. [code], 6, 3021, <ext-link xlink:href="https://doi.org/10.21105/joss.03021" ext-link-type="DOI">10.21105/joss.03021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx66"><label>Wunsch(1996)</label><mixed-citation> Wunsch, C.: The ocean circulation inverse problem, Cambridge University Press, ISBN 9780521480901, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx67"><label>Wunsch and Heimbach(2006)</label><mixed-citation>Wunsch, C. and Heimbach, P.: Estimated Decadal Changes in the North Atlantic Meridional Overturning Circulation and Heat Flux 1993–2004, J. Phys. Oceanogr., 36, 2012–2024, <ext-link xlink:href="https://doi.org/10.1175/JPO2957.1" ext-link-type="DOI">10.1175/JPO2957.1</ext-link>, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx68"><label>Yang et al.(2006)Yang, Baker, Li, Cordes, Huff, Nagpal, Okereke, Villafane, Kalnay, and Duane</label><mixed-citation> Yang, S.-C., Baker, D., Li, H., Cordes, K., Huff, M., Nagpal, G., Okereke, E., Villafane, J., Kalnay, E., and Duane, G. S.: Data assimilation as synchronization of truth and model: Experiments with the three-variable Lorenz system, J. Atmos. Sci., 63, 2340–2354, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx69"><label>Yin et al.(2014)Yin, Wang, Liu, and Tan</label><mixed-citation>Yin, X., Wang, B., Liu, J., and Tan, X.: Evaluation of conditional non-linear optimal perturbation obtained by an ensemble-based approach using the Lorenz-63 model, Tellus A, 66, 22773, <ext-link xlink:href="https://doi.org/10.3402/tellusa.v66.22773" ext-link-type="DOI">10.3402/tellusa.v66.22773</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx70"><label>Zou et al.(1992)Zou, Navon, and Ledimet</label><mixed-citation> Zou, X., Navon, I., and Ledimet, F.: An optimal nudging data assimilation scheme using parameter estimation, Q. J. Roy. Meteor. Soc., 118, 1163–1186, 1992.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Long-window tandem variational data assimilation methods for chaotic climate models tested with the Lorenz 63 system</article-title-html>
<abstract-html/>
<ref-html id="bib1.bib1"><label>Abarbanel et al.(2009)Abarbanel, Creveling, Farsian, and
Kostuk</label><mixed-citation>
      
Abarbanel, H. D., Creveling, D. R., Farsian, R., and Kostuk, M.: Dynamical
state and parameter estimation, SIAM J. Appl. Dynam. Syst., 8,
1341–1381, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Abarbanel et al.(2010)Abarbanel, Kostuk, and
Whartenby</label><mixed-citation>
      
Abarbanel, H. D., Kostuk, M., and Whartenby, W.: Data assimilation with
regularized nonlinear instabilities, Q. J. Roy.
Meteor. Soc. A, 136, 769–783, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Allaire(2015)</label><mixed-citation>
      
Allaire, G.: A review of adjoint methods for sensitivity analysis, uncertainty
quantification and optimization in numerical codes, Ingénieurs de
l'Automobile, 836, 33–36, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Bar-Shalom et al.(2004)Bar-Shalom, Li, and
Kirubarajan</label><mixed-citation>
      
Bar-Shalom, Y., Li, X. R., and Kirubarajan, T.: Estimation with applications to
tracking and navigation: theory algorithms and software, John Wiley &amp; Sons, <a href="https://doi.org/10.1002/0471221279" target="_blank">https://doi.org/10.1002/0471221279</a>,
2004.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Bertino et al.(2003)Bertino, Evensen, and
Wackernagel</label><mixed-citation>
      
Bertino, L., Evensen, G., and Wackernagel, H.: Sequential data assimilation
techniques in oceanography, Int. Stat. Rev., 71, 223–241,
2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Bradbury et al.(2018)Bradbury, Frostig, Hawkins, Johnson, Leary,
Maclaurin, Necula, Paszke, VanderPlas, Wanderman-Milne, and
Zhang</label><mixed-citation>
      
Bradbury, J., Frostig, R., Hawkins, P., Johnson, M. J., Leary, C., Maclaurin,
D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., and
Zhang, Q.: JAX: composable transformations of Python+NumPy programs, Github [code],
<a href="http://github.com/google/jax" target="_blank"/>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Cameron and Yang(2019)</label><mixed-citation>
      
Cameron, M. and Yang, S.: Computing the quasipotential for highly dissipative
and chaotic SDEs an application to stochastic Lorenz’63, Commun.
Appl. Math. Comput. Sci., 14, 207–246, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Cummings and Smedstad(2013)</label><mixed-citation>
      
Cummings, J. A. and Smedstad, O. M.: Variational data assimilation for the
global ocean, in: Data assimilation for atmospheric, oceanic and hydrologic
applications (Vol. II),  Springer, 303–343, <a href="https://doi.org/10.1007/978-3-642-35088-7_13" target="_blank">https://doi.org/10.1007/978-3-642-35088-7_13</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Daron and Stainforth(2015)</label><mixed-citation>
      
Daron, J. and Stainforth, D. A.: On quantifying the climate of the
nonautonomous Lorenz-63 model, Chaos, 25, 043103, <a href="https://doi.org/10.1063/1.4916789" target="_blank">https://doi.org/10.1063/1.4916789</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Dembinski and et al.(2020)</label><mixed-citation>
      
Dembinski, H., Ongmongkolkul, P., Deil, C., Schreiner, H., Feickert, M., Burr, C., Watson, J., Rost, F., Pearce, A., Geiger, L., Abdelmotteleb, A., Desai, A., Wiedemann, B. M., Gohlke, C., Sanders, J., Drotleff, J., Eschle, J., Neste, L., Gorelli, M. E., and Zapata, O.: scikit-hep/iminuit, Zenodo [code],
<a href="https://doi.org/10.5281/zenodo.3949207" target="_blank">https://doi.org/10.5281/zenodo.3949207</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Desroziers et al.(2014)Desroziers, Camino, and
Berre</label><mixed-citation>
      
Desroziers, G., Camino, J.-T., and Berre, L.: 4DEnVar: link with 4D state
formulation of variational assimilation and different possible
implementations, Q. J. Roy. Meteor. Soc., 140,
2097–2110, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Du and Shiue(2021)</label><mixed-citation>
      
Du, Y. J. and Shiue, M.-C.: Analysis and computation of continuous data
assimilation algorithms for Lorenz 63 system based on nonlinear nudging
techniques, J. Comput. Appl. Math., 386, 113246, <a href="https://doi.org/10.1016/j.cam.2020.113246" target="_blank">https://doi.org/10.1016/j.cam.2020.113246</a>,
2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Errico(1997)</label><mixed-citation>
      
Errico, R. M.: What is an adjoint model?, B. Am.
Meteorol. Soc., 78, 2577–2592, 1997.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Evensen(1994)</label><mixed-citation>
      
Evensen, G.: Sequential data assimilation with a nonlinear quasi-geostrophic
model using Monte Carlo methods to forecast error statistics, J.
Geophys. Res.-Oceans, 99, 10143–10162, 1994.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Evensen(2003)</label><mixed-citation>
      
Evensen, G.: The ensemble Kalman filter: Theoretical formulation and practical
implementation, Ocean Dynam., 53, 343–367, 2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Evensen et al.(2022)Evensen, Vossepoel, and
Van Leeuwen</label><mixed-citation>
      
Evensen, G., Vossepoel, F. C., and Van Leeuwen, P. J.: Data assimilation
fundamentals: A unified formulation of the state and parameter estimation
problem, Springer Nature, <a href="https://doi.org/10.1007/978-3-030-96709-3" target="_blank">https://doi.org/10.1007/978-3-030-96709-3</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Fisher et al.(2011)Fisher, Tremolet, Auvinen, Tan, and
Poli</label><mixed-citation>
      
Fisher, M., Tremolet, Y., Auvinen, H., Tan, D., and Poli, P.: Weak-constraint
and long-window 4D-Var, ECMWF Technical Memoranda, 655, 47, <a href="https://doi.org/10.21957/9ii4d4dsq" target="_blank">https://doi.org/10.21957/9ii4d4dsq</a>, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Gauthier(1992)</label><mixed-citation>
      
Gauthier, P.: Chaos and quadri-dimensional data assimilation: A study based on
the Lorenz model, Tellus A, 44, 2–17,
1992.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Giering and Kaminski(1998)</label><mixed-citation>
      
Giering, R. and Kaminski, T.: Recipes for adjoint code construction, ACM
T. Math. Softw., 24, 437–474, 1998.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Goodliff et al.(2015)Goodliff, Amezcua, and
Van Leeuwen</label><mixed-citation>
      
Goodliff, M., Amezcua, J., and Van Leeuwen, P. J.: Comparing hybrid data
assimilation methods on the Lorenz 1963 model with increasing non-linearity,
Tellus A,
<a href="https://doi.org/10.3402/tellusa.v67.26928" target="_blank">https://doi.org/10.3402/tellusa.v67.26928</a>,  67, 1–12, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Goodliff et al.(2020)Goodliff, Fletcher, Kliewer, Forsythe, and
Jones</label><mixed-citation>
      
Goodliff, M., Fletcher, S., Kliewer, A., Forsythe, J., and Jones, A.: Detection
of Non-Gaussian Behavior Using Machine Learning Techniques: A Case Study on
the Lorenz 63 Model, J. Geophys. Res.-Atmos., 125,
e2019JD031551, <a href="https://doi.org/10.1029/2019JD031551" target="_blank">https://doi.org/10.1029/2019JD031551</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Gustafsson et al.(1998)Gustafsson, Källén, and
Thorsteinsson</label><mixed-citation>
      
Gustafsson, N., Källén, E., and Thorsteinsson, S.: Sensitivity of
forecast errors to initial and lateral boundary conditions, Tellus A, 50, 167–185, 1998.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Gustafsson et al.(2001)Gustafsson, Berre, Hörnquist, Huang,
Lindskog, Navascues, Mogensen, and Thorsteinsson</label><mixed-citation>
      
Gustafsson, N., Berre, L., Hörnquist, S., Huang, X.-Y., Lindskog, M.,
Navascues, B., Mogensen, K., and Thorsteinsson, S.: Three-dimensional
variational data assimilation for a limited area model: Part I: General
formulation and the background error constraint, Tellus A, 53, 425–446,
2001.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Hall(1986)</label><mixed-citation>
      
Hall, M. C.: Application of adjoint sensitivity theory to an atmospheric
general circulation model, J. Atmos. Sci., 43, 2644–2652,
1986.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Hall and Cacuci(1983)</label><mixed-citation>
      
Hall, M. C. and Cacuci, D. G.: Physical interpretation of the adjoint functions
for sensitivity analysis of atmospheric models, J. Atmos.
Sci., 40, 2537–2546, 1983.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Hall et al.(1982)Hall, Cacuci, and Schlesinger</label><mixed-citation>
      
Hall, M. C., Cacuci, D. G., and Schlesinger, M. E.: Sensitivity analysis of a
radiative-convective model by the adjoint method, J. Atmos.
Sci., 39, 2038–2050, 1982.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Hascoet and Pascual(2013)</label><mixed-citation>
      
Hascoet, L. and Pascual, V.: The Tapenade automatic differentiation tool:
Principles, model, and specification, ACM T. Math.
Softw., 39, 1–43, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Hirsch et al.(2013)Hirsch, Smale, and Devaney</label><mixed-citation>
      
Hirsch, M. W., Smale, S., and Devaney, R. L.: 14 – The Lorenz System, in:
Differential Equations, Dynamical Systems, and an Introduction to Chaos
(Third Edition), edited by: Hirsch, M. W., Smale, S., and Devaney, R. L.,     3rd Edn., Academic Press, Boston,  305–328, ISBN 978-0-12-382010-5,
<a href="https://doi.org/10.1016/B978-0-12-382010-5.00014-2" target="_blank">https://doi.org/10.1016/B978-0-12-382010-5.00014-2</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Houtekamer and Mitchell(2001)</label><mixed-citation>
      
Houtekamer, P. L. and Mitchell, H. L.: A sequential ensemble Kalman filter for
atmospheric data assimilation, Mon. Weather Rev., 129, 123–137, 2001.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Huai et al.(2017)Huai, Li, Ding, Feng, and Liu</label><mixed-citation>
      
Huai, X.-W., Li, J.-P., Ding, R.-Q., Feng, J., and Liu, D.-Q.: Quantifying
local predictability of the Lorenz system using the nonlinear local Lyapunov
exponent, Atmos. Ocean. Sci. Lett., 10, 372–378, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Kalman(1960)</label><mixed-citation>
      
Kalman, R. E.: A New Approach to Linear Filtering and Prediction Problems,
J. Basic Eng., 82, 35–45, <a href="https://doi.org/10.1115/1.3662552" target="_blank">https://doi.org/10.1115/1.3662552</a>, 1960.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Köhl(2020)</label><mixed-citation>
      
Köhl, A.: Evaluating the GECCO3 1948–2018 ocean synthesis – a configuration
for initializing the MPI-ESM climate model, Q. J. Roy.
Meteor. Soc., 146, 2250–2273, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Köhl and Willebrand(2002)</label><mixed-citation>
      
Köhl, A. and Willebrand, J.: An adjoint method for the assimilation of
statistical characteristics into eddy-resolving ocean models, Tellus A, 54, 406–425, 2002.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Kravtsov and Tsonis(2021)</label><mixed-citation>
      
Kravtsov, S. and Tsonis, A. A.: Lorenz-63 Model as a Metaphor for Transient
Complexity in Climate, Entropy, 23, 951, <a href="https://doi.org/10.3390/e23080951" target="_blank">https://doi.org/10.3390/e23080951</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Laboratory(2016)</label><mixed-citation>
      
Naval Research Laboratory.: Global HYCOM+CICE 1/12 degree page,
<a href="https://www7320.nrlssc.navy.mil/GLBhycomcice1-12/" target="_blank"/> (last access: 17 April 2025), 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Lea et al.(2000)Lea, Allen, and Haine</label><mixed-citation>
      
Lea, D. J., Allen, M. R., and Haine, T. W.: Sensitivity analysis of the climate
of a chaotic system, Tellus A, 52,
523–532, 2000.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Le Dimet and Talagrand(1986)</label><mixed-citation>
      
Le Dimet, F.-X. and Talagrand, O.: Variational algorithms for analysis and
assimilation of meteorological observations: theoretical aspects, Tellus A, 38, 97–110, 1986.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Lorenz(1963)</label><mixed-citation>
      
Lorenz, E. N.: Deterministic nonperiodic flow, J. Atmos. Sci.,
20, 130–141, 1963.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Lyu et al.(2018)Lyu, Köhl, Matei, and Stammer</label><mixed-citation>
      
Lyu, G., Köhl, A., Matei, I., and Stammer, D.: Adjoint-based climate model
tuning: Application to the planet simulator, J. Adv. Model.
Earth Sy., 10, 207–222, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Marotzke et al.(1999)Marotzke, Giering, Zhang, Stammer, Hill, and
Lee</label><mixed-citation>
      
Marotzke, J., Giering, R., Zhang, K. Q., Stammer, D., Hill, C., and Lee, T.:
Construction of the adjoint MIT ocean general circulation model and
application to Atlantic heat transport sensitivity, J. Geophys.
Res.-Oceans, 104, 29529–29547, 1999.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>Marzban(2013)</label><mixed-citation>
      
Marzban, C.: Variance-based sensitivity analysis: An illustration on the
Lorenz'63 model, Mon. Weather Rev., 141, 4069–4079, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>Mauritsen et al.(2012)Mauritsen, Stevens, Roeckner, Crueger, Esch,
Giorgetta, Haak, Jungclaus, Klocke, Matei et al.</label><mixed-citation>
      
Mauritsen, T., Stevens, B., Roeckner, E., Crueger, T., Esch, M., Giorgetta, M.,
Haak, H., Jungclaus, J., Klocke, D., Matei, D., Mikolajewicz, U., Notz, D., Pincus, R., Schmidt, H., and Tomassini, L.: Tuning the climate of
a global model, J. Adv. Model. Earth Sy., 4, M00A01,  <a href="https://doi.org/10.1029/2012MS000154" target="_blank">https://doi.org/10.1029/2012MS000154</a>,2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Miller et al.(1994)Miller, Ghil, and Gauthiez</label><mixed-citation>
      
Miller, R. N., Ghil, M., and Gauthiez, F.: Advanced data assimilation in
strongly nonlinear dynamical systems, J. Atmos. Sci., 51,
1037–1056, 1994.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>Navon(2009)</label><mixed-citation>
      
Navon, I. M.: Data assimilation for numerical weather prediction: a review,
Data assimilation for atmospheric, oceanic and hydrologic applications,
21–65, <a href="https://doi.org/10.1007/978-3-540-71056-1_2" target="_blank">https://doi.org/10.1007/978-3-540-71056-1_2</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Nichols(2010)</label><mixed-citation>
      
Nichols, N. K.: Mathematical Concepts of Data Assimilation,
Springer Berlin Heidelberg, Berlin, Heidelberg,  13–39, ISBN 978-3-540-74703-1,
<a href="https://doi.org/10.1007/978-3-540-74703-1_2" target="_blank">https://doi.org/10.1007/978-3-540-74703-1_2</a>, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Pasini and Pelino(2005)</label><mixed-citation>
      
Pasini, A. and Pelino, V.: Can we estimate atmospheric predictability by
performance of neural network forecasting? The toy case studies of unforced
and forced Lorenz models, in: CIMSA, 2005 IEEE International Conference on
Computational Intelligence for Measurement Systems and Applications, 2005,
IEEE, 69–74,  <a href="https://doi.org/10.1109/CIMSA.2005.1522829" target="_blank">https://doi.org/10.1109/CIMSA.2005.1522829</a>,2005.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Pelino and Maimone(2007)</label><mixed-citation>
      
Pelino, V. and Maimone, F.: Energetics, skeletal dynamics, and long-term
predictions on Kolmogorov-Lorenz systems, Phys. Rev. E, 76, 046214, <a href="https://doi.org/10.1103/PhysRevE.76.046214" target="_blank">https://doi.org/10.1103/PhysRevE.76.046214</a>,
2007.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>Pires et al.(1996)Pires, Vautard, and Talagrand</label><mixed-citation>
      
Pires, C., Vautard, R., and Talagrand, O.: On extending the limits of
variational assimilation in nonlinear chaotic systems, Tellus A, 48, 96–121,
1996.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>Quinn et al.(2009)Quinn, Bryant, Creveling, Klein, and
Abarbanel</label><mixed-citation>
      
Quinn, J. C., Bryant, P. H., Creveling, D. R., Klein, S. R., and Abarbanel,
H. D.: Parameter and state estimation of experimental chaotic systems using
synchronization, Phys. Rev. E, 80, 016201, <a href="https://doi.org/10.1103/PhysRevE.80.016201" target="_blank">https://doi.org/10.1103/PhysRevE.80.016201</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Rabier and Liu(2003)</label><mixed-citation>
      
Rabier, F. and Liu, Z.: Variational data assimilation: theory and overview, in:
Proc. ECMWF Seminar on Recent Developments in Data Assimilation for
Atmosphere and Ocean, Reading, UK,  8–12 September,  29–43, <a href="https://www.ecmwf.int/en/elibrary/76079-variational-data-assimiltion-theory-and-overview&#xA;" target="_blank"/>, 2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Ruiz et al.(2013)Ruiz, Pulido, and Miyoshi</label><mixed-citation>
      
Ruiz, J. J., Pulido, M., and Miyoshi, T.: Estimating model parameters with
ensemble-based data assimilation: A review, J. Meteor.
Soc. Jpn. Ser. II, 91, 79–99, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Simon(2006)</label><mixed-citation>
      
Simon, D.: Optimal state estimation: Kalman, H infinity, and nonlinear
approaches, John Wiley &amp; Sons, <a href="https://doi.org/10.1002/0470045345" target="_blank">https://doi.org/10.1002/0470045345</a>, 2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Stammer et al.(2016)Stammer, Balmaseda, Heimbach, Köhl, and
Weaver</label><mixed-citation>
      
Stammer, D., Balmaseda, M., Heimbach, P., Köhl, A., and Weaver, A.: Ocean
data assimilation in support of climate applications: status and
perspectives, Annu. Rev. Mar. Sci., 8, 491–518, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Stammer et al.(2018)Stammer, Köhl, Vlasenko, Matei, Lunkeit, and
Schubert</label><mixed-citation>
      
Stammer, D., Köhl, A., Vlasenko, A., Matei, I., Lunkeit, F., and Schubert,
S.: A pilot climate sensitivity study using the CEN coupled adjoint model
(CESAM), J. Climate, 31, 2031–2056, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>Stensrud and Bao(1992)</label><mixed-citation>
      
Stensrud, D. J. and Bao, J.-W.: Behaviors of variational and nudging
assimilation techniques with a chaotic low-order model, Mon. Weather
Rev., 120, 3016–3028, 1992.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Sugiura et al.(2014)Sugiura, Masuda, Fujii, Kamachi, Ishikawa, and
Awaji</label><mixed-citation>
      
Sugiura, N., Masuda, S., Fujii, Y., Kamachi, M., Ishikawa, Y., and Awaji, T.: A
framework for interpreting regularized state estimation, Mon. Weather
Rev., 142, 386–400, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Talagrand(2010)</label><mixed-citation>
      
Talagrand, O.: Variational Assimilation. In: Lahoz, W., Khattatov, B., Menard, R. (eds) Data Assimilation. Springer, Berlin, Heidelberg. <a href="https://doi.org/10.1007/978-3-540-74703-1_3" target="_blank">https://doi.org/10.1007/978-3-540-74703-1_3</a>,  2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Tandeo et al.(2015)Tandeo, Ailliot, Ruiz, Hannart, Chapron, Cuzol,
Monbet, Easton, and Fablet</label><mixed-citation>
      
Tandeo, P., Ailliot, P., Ruiz, J., Hannart, A., Chapron, B., Cuzol, A., Monbet,
V., Easton, R., and Fablet, R.: Combining analog method and ensemble data
assimilation: application to the Lorenz-63 chaotic system, in: Machine
Learning and Data Mining Approaches to Climate Science: proceedings of the
4th International Workshop on Climate Informatics,  Springer,  3–12, <a href="https://doi.org/10.1007/978-3-319-17220-0_1" target="_blank">https://doi.org/10.1007/978-3-319-17220-0_1</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib59"><label>Tett et al.(2017)Tett, Yamazaki, Mineter, Cartis, and
Eizenberg</label><mixed-citation>
      
Tett, S. F. B., Yamazaki, K., Mineter, M. J., Cartis, C., and Eizenberg, N.: Calibrating climate models using inverse methods: case studies with HadAM3, HadAM3P and HadCM3, Geosci. Model Dev., 10, 3567–3589, <a href="https://doi.org/10.5194/gmd-10-3567-2017" target="_blank">https://doi.org/10.5194/gmd-10-3567-2017</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib60"><label>The pandas development team(2020)</label><mixed-citation>
      
The pandas development team: pandas-dev/pandas: Pandas, Zenodo [code],
<a href="https://doi.org/10.5281/zenodo.3509134" target="_blank">https://doi.org/10.5281/zenodo.3509134</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib61"><label>Tippett and Chang(2003)</label><mixed-citation>
      
Tippett, M. K. and Chang, P.: Some theoretical considerations on predictability
of linear stochastic dynamics, Tellus A, 55, 148–157, 2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib62"><label>Tippett et al.(2003)Tippett, Anderson, Bishop, Hamill, and
Whitaker</label><mixed-citation>
      
Tippett, M. K., Anderson, J. L., Bishop, C. H., Hamill, T. M., and Whitaker,
J. S.: Ensemble square root filters, Mon. Weather Rev., 131, 1485–1490,
2003.


    </mixed-citation></ref-html>
<ref-html id="bib1.bib63"><label>Tremolet(2006)</label><mixed-citation>
      
Tremolet, Y.: Accounting for an imperfect model in 4D-Var, Q. J.
Roy. Meteor. Soc., 132, 2483–2504, 2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib64"><label>Van Der Merwe and Wan(2001)</label><mixed-citation>
      
Van Der Merwe, R. and Wan, E. A.: The square-root unscented Kalman filter for
state and parameter-estimation, in: 2001 IEEE international conference on
acoustics, speech, and signal processing, Proceedings (Cat. No. 01CH37221),
IEEE,  vol. 6,  3461–3464, <a href="https://doi.org/10.1109/ICASSP.2001.940586" target="_blank">https://doi.org/10.1109/ICASSP.2001.940586</a>, 2001.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib65"><label>Waskom(2021)</label><mixed-citation>
      
Waskom, M. L.: seaborn: statistical data visualization, J. Open Source
Softw. [code], 6, 3021, <a href="https://doi.org/10.21105/joss.03021" target="_blank">https://doi.org/10.21105/joss.03021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib66"><label>Wunsch(1996)</label><mixed-citation>
      
Wunsch, C.: The ocean circulation inverse problem, Cambridge University Press, ISBN 9780521480901, 1996.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib67"><label>Wunsch and Heimbach(2006)</label><mixed-citation>
      
Wunsch, C. and Heimbach, P.: Estimated Decadal Changes in the North Atlantic
Meridional Overturning Circulation and Heat Flux 1993–2004, J.
Phys. Oceanogr., 36, 2012–2024,
<a href="https://doi.org/10.1175/JPO2957.1" target="_blank">https://doi.org/10.1175/JPO2957.1</a>, 2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib68"><label>Yang et al.(2006)Yang, Baker, Li, Cordes, Huff, Nagpal, Okereke,
Villafane, Kalnay, and Duane</label><mixed-citation>
      
Yang, S.-C., Baker, D., Li, H., Cordes, K., Huff, M., Nagpal, G., Okereke, E.,
Villafane, J., Kalnay, E., and Duane, G. S.: Data assimilation as
synchronization of truth and model: Experiments with the three-variable
Lorenz system, J. Atmos. Sci., 63, 2340–2354, 2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib69"><label>Yin et al.(2014)Yin, Wang, Liu, and Tan</label><mixed-citation>
      
Yin, X., Wang, B., Liu, J., and Tan, X.: Evaluation of conditional non-linear
optimal perturbation obtained by an ensemble-based approach using the
Lorenz-63 model, Tellus A, 66, 22773, <a href="https://doi.org/10.3402/tellusa.v66.22773" target="_blank">https://doi.org/10.3402/tellusa.v66.22773</a>,
2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib70"><label>Zou et al.(1992)Zou, Navon, and Ledimet</label><mixed-citation>
      
Zou, X., Navon, I., and Ledimet, F.: An optimal nudging data assimilation
scheme using parameter estimation, Q. J. Roy.
Meteor. Soc., 118, 1163–1186, 1992.

    </mixed-citation></ref-html>--></article>
