<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0"><?xmltex \makeatother\@nolinetrue\makeatletter?>
  <front>
    <journal-meta><journal-id journal-id-type="publisher">NPG</journal-id><journal-title-group>
    <journal-title>Nonlinear Processes in Geophysics</journal-title>
    <abbrev-journal-title abbrev-type="publisher">NPG</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Nonlin. Processes Geophys.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1607-7946</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/npg-27-473-2020</article-id><title-group><article-title>Statistical postprocessing of ensemble forecasts for<?xmltex \hack{\break}?> severe weather at Deutscher Wetterdienst</article-title><alt-title>Statistical Postprocessing for Severe Weather</alt-title>
      </title-group><?xmltex \runningtitle{Statistical Postprocessing for Severe Weather}?><?xmltex \runningauthor{R. Hess}?>
      <contrib-group>
        <contrib contrib-type="author" corresp="yes">
          <name><surname>Hess</surname><given-names>Reinhold</given-names></name>
          <email>reinhold.hess@dwd.de</email>
        <ext-link>https://orcid.org/0000-0001-8559-2516</ext-link></contrib>
        <aff id="aff1"><institution>Deutscher Wetterdienst, Offenbach, Germany</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Reinhold Hess (reinhold.hess@dwd.de)</corresp></author-notes><pub-date><day>6</day><month>October</month><year>2020</year></pub-date>
      
      <volume>27</volume>
      <issue>4</issue>
      <fpage>473</fpage><lpage>487</lpage>
      <history>
        <date date-type="received"><day>7</day><month>January</month><year>2020</year></date>
           <date date-type="rev-request"><day>20</day><month>January</month><year>2020</year></date>
           <date date-type="rev-recd"><day>13</day><month>August</month><year>2020</year></date>
           <date date-type="accepted"><day>19</day><month>August</month><year>2020</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2020 Reinhold Hess</copyright-statement>
        <copyright-year>2020</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020.html">This article is available from https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020.html</self-uri><self-uri xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020.pdf">The full text article is available as a PDF file from https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020.pdf</self-uri>
      <abstract><title>Abstract</title>
    <p id="d1e79">This paper gives an overview of Deutscher Wetterdienst's (DWD's) postprocessing system called
Ensemble-MOS together with its motivation and the design consequences
for probabilistic forecasts of extreme events based on ensemble data.
Forecasts of the ensemble systems COSMO-D2-EPS and ECMWF-ENS
are statistically optimised and calibrated by Ensemble-MOS
with a focus on severe weather in order to support the warning decision management at DWD.</p>
    <p id="d1e82">Ensemble mean and spread are used as predictors for linear and logistic multiple regressions
to correct for conditional biases.
The predictands are derived from synoptic observations and include temperature, precipitation amounts, wind gusts and many more and are statistically estimated in a comprehensive model output statistics (MOS) approach.
Long time series and collections of stations are used as
training data that capture a sufficient number of observed events, as required for robust statistical modelling.</p>
    <p id="d1e85">Logistic regressions are applied to probabilities that predefined meteorological events occur.
Details of the implementation including the selection of predictors with testing for significance are presented.
For probabilities of severe wind gusts global logistic parameterisations are developed that
depend on local estimations of wind speed. In this way, robust probability forecasts for extreme events are obtained
while local characteristics are preserved.</p>
    <p id="d1e88">The problems of Ensemble-MOS, such as model changes and consistency requirements,
which occur with the operative MOS systems of the DWD are addressed.</p>
  </abstract>
    </article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d1e100">Ensemble forecasting rose with the understanding of the limited predictability of weather.
This limitation is caused by sparse and imperfect observations, approximating numerical data assimilation and modelling, and by the chaotic physical nature of the atmosphere.
The basic idea of ensemble forecasting is to vary observations, initial and boundary conditions,
and physical parameterisations within their assumed scale of uncertainty and rerun the forecast model with these changes.</p>
      <p id="d1e103">The obtained ensemble of forecasts expresses the distribution of possible weather scenarios to be expected.
Probabilistic forecasts can be derived from the ensemble, like
forecast errors, probabilities for special weather events,
quantiles of the distribution or even estimations of the full distribution.
The ensemble spread is often used as estimation for forecast errors.
In a perfect ensemble system the spread
is statistically consistent with the forecast error of the ensemble mean against observations
<xref ref-type="bibr" rid="bib1.bibx36" id="paren.1"><named-content content-type="pre">e.g.</named-content></xref>; however, it is experienced as being often too small, especially for near-surface weather elements and short lead times. Typically, an optimal spread–skill relationship close to 1 and its involved forecast reliability
are obtained much more easily for atmospheric variables in higher vertical layers, e.g. 500 hPa geopotential height, than for screen-level variables like 2 <inline-formula><mml:math id="M1" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> temperature, 10 <inline-formula><mml:math id="M2" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speed or precipitation <xref ref-type="bibr" rid="bib1.bibx6 bib1.bibx8 bib1.bibx5" id="paren.2"><named-content content-type="pre">e.g.</named-content></xref>; see also Sect. <xref ref-type="sec" rid="Ch1.S4"/>.</p>
      <p id="d1e134">In order to make best use of the probabilistic information contained in the ensembles,
e.g. by relating probabilities for harmful weather events to economical value in cost–loss evaluations <xref ref-type="bibr" rid="bib1.bibx35 bib1.bibx2" id="paren.3"><named-content content-type="pre">e.g.</named-content></xref>,
the ensemble forecasts should be calibrated to observed relative frequencies as motivated by <xref ref-type="bibr" rid="bib1.bibx5" id="text.4"/>.
Warning<?pagebreak page474?> thresholds are the levels of probabilities at which meteorological warnings are to be issued.
These thresholds may be tailored to the public depending on categorical scores such as probability of detection (POD) and false alarm ratio (FAR).
Statistical reliability of forecast probabilities is considered essential for qualified threshold definitions
and for automated warning guidances.</p>
      <p id="d1e145">For deterministic forecasts statistical postprocessing is used for optimisation and interpretation.
This is likewise true for ensemble forecasts, where statistical calibration is an additional application of postprocessing.
<xref ref-type="bibr" rid="bib1.bibx11" id="text.5"/> describe probabilistic postprocessing as a method to maximise the sharpness of a predictive distribution under the condition of calibration
(the climatologic average is calibrated too; however, it has no sharpness and is useless as a forecast). Nevertheless, optimisation is still an issue for ensemble forecasts.
In general, the systematic errors of the underlying numerical model turn up
in each forecast member and thus are retained in the ensemble mean.
Averaging only reduces the random errors of the ensemble members.</p>
      <p id="d1e152">Due to its ability to improve skill and reliability of probabilistic forecasts,
many different postprocessing methods exist for both single- and multi-model ensembles.
There are comprehensive multivariate systems and univariate systems that are specific to a certain forecast element. Length of training data generally depends on the statistical method and application;
however, the availability of data is also often a serious limitation.
Some systems perform individual training for different locations in order to account for
local characteristics, whilst others apply the same statistical model to collections of stations or grid points.
Global modelling improves statistical sampling at the cost of orographic and climatologic disparities.</p>
      <p id="d1e155">Classical MOS systems tend to underestimate forecast errors if corrections are applied to each ensemble member individually.
In order to maintain forecast variability, <xref ref-type="bibr" rid="bib1.bibx33" id="text.6"/> suggests considering observation errors.
<xref ref-type="bibr" rid="bib1.bibx10" id="text.7"/> propose non-homogeneous Gaussian regression (NGR) that
relies on Gaussian distributions.
The location and scale parameters of the Gaussian distributions correspond to a linear function of the ensemble mean
and ensemble spread, respectively.
The NGR coefficients
are trained by minimising the continuous ranked probability score (CRPS).
In Bayesian model averaging (BMA) <xref ref-type="bibr" rid="bib1.bibx24 bib1.bibx20" id="paren.8"><named-content content-type="pre">e.g.</named-content></xref>
distributions of already bias-corrected forecasts are combined as weighted averages using kernel functions.</p>
      <p id="d1e169">Many different postprocessing methods tailored to different variables exist; only some are mentioned here. For 24-hourly precipitation <xref ref-type="bibr" rid="bib1.bibx12" id="text.9"/> present a multimodel ensemble postprocessing based on extended logistic regression and 8 years of training data. <xref ref-type="bibr" rid="bib1.bibx13" id="text.10"/> describe a method to blend high-resolution multimodel ensembles by quantile mapping
with short training periods of about 2 months for 6- and 12-hourly precipitations. Postprocessing methods specialising in wind speed have been developed as well;
e.g. <xref ref-type="bibr" rid="bib1.bibx31" id="text.11"/> use BMA in combination with Gamma distributions. An overview of conventional univariate postprocessing approaches is given in <xref ref-type="bibr" rid="bib1.bibx37" id="text.12"/>.</p>
      <p id="d1e184">In addition to the univariate postprocessing methods mentioned above,
there exist also approaches to model spatio-temporal dependence structures and hence to produce ensembles of forecast scenarios. This enables, for instance, to estimate area-related probabilities. <xref ref-type="bibr" rid="bib1.bibx30" id="text.13"/> and <xref ref-type="bibr" rid="bib1.bibx29" id="text.14"/>
use ensemble copula coupling (ECC) and
Schaake shuffle-based approaches
in order to generate postprocessed forecast scenarios for temperature, precipitation and wind.
Ensembles of ECC forecast scenarios provide
high flexibility in product generation to the constraint
that all ensemble data are accessible.</p>
      <p id="d1e193">Fewer methods focus on extreme events of precipitation and wind gusts
that are essential for automated warning support.
<xref ref-type="bibr" rid="bib1.bibx7" id="text.15"/> use the tails of generalised extreme-value distributions
in order to estimate conditional probabilities of extreme events.
As extreme meteorological events are (fortunately) rare,
long time series are required to capture a sufficiently large number of occurred events
in order to derive statistically significant estimations.
For example, strong precipitation events with rain amounts of more than 15 <inline-formula><mml:math id="M3" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">mm</mml:mi></mml:mrow></mml:math></inline-formula> h<inline-formula><mml:math id="M4" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> are captured only about once a year
at each rain gauge within Germany.
Extreme events with more than 40 and 50 <inline-formula><mml:math id="M5" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">mm</mml:mi></mml:mrow></mml:math></inline-formula> rarely appear; nevertheless, warnings are essential when they do.</p>
      <p id="d1e227">With long time series,
a significant portion of the data consists of
calm weather without relevance for warnings.
It is problematic, however, to restrict or focus training data on severe events.
In doing so, predictors might be selected that are highly correlated with the selected series of severe events
but accidentally also to calm scenarios that are not contained in the training data.
In order to exclude these spurious predictors and to derive skilful statistical models, more general
training data need to be used, since otherwise overforecasting presumably results and frequency bias (FB) and FAR increase.
This basically corresponds to the idea of the <italic>forecaster's dilemma</italic> (see <xref ref-type="bibr" rid="bib1.bibx19" id="altparen.16"/>) that states that overforecasting is a promising strategy when forecasts are evaluated
mainly for extreme events.</p>
      <p id="d1e237">The usage of probabilistic forecasts for warnings of severe weather also influences the way the forecasts need to be evaluated.
Also for verification, long time periods are required to capture enough extreme and rare events to derive
statistically significant results.
Verification scores like root mean square error (RMSE) or CRPS <xref ref-type="bibr" rid="bib1.bibx14 bib1.bibx10" id="paren.17"><named-content content-type="pre">e.g.</named-content></xref>
are highly dominated by the overwhelming majority of cases when no event occurred.
Excellent but irrelevant forecasts of calm weather can pretend good verification results,
although the few relevant extreme cases might not be forecasted well.
Categorical scores like POD and FAR are considered more relevant for rare and extreme cases,
along with other more complex scores
like Heidke skill score (HSS) or<?pagebreak page475?> equitable threat score (ETS). Also, scatter diagrams reveal outliers and are sensitive to extreme values.</p>
      <p id="d1e245">Here we present a MOS approach that has been tailored to postprocessing ensemble forecasts
for extreme and rare events.
It is named Ensemble-MOS and has been set up at DWD in order to support warning management
with probabilistic forecasts of potentially harmful weather events
within AutoWARN; see <xref ref-type="bibr" rid="bib1.bibx28" id="text.18"/> and <xref ref-type="bibr" rid="bib1.bibx26 bib1.bibx27" id="text.19"/>. Altogether 37 different warning elements exist at DWD, including heavy rain and strong wind gusts,
both at several levels of intensity, thunderstorms, snowfall, fog, limited visibility, frost and others.
Currently the ensemble systems COSMO-D2-EPS and ECMWF-ENS are statistically optimised and calibrated
using several years of training data, but Ensemble-MOS is applicable to other ensembles in general.</p>
      <p id="d1e254">At DWD, statistically postprocessed forecasts of the ensemble systems COSMO-D2-EPS and ECMWF-ENS
and also of the deterministic models ICON and ECMWF-IFS are combined in a second step in order to provide a
consistent data set and a seamless transition from very short-term to medium-range forecasts. This combined product provides a single voice basis for the generation of warning proposals; see <xref ref-type="bibr" rid="bib1.bibx28" id="text.20"/>. The combination is based on a second MOS approach similar to the system described here; it uses the individual statistical forecasts of the numerical models as predictors.
As a linear combination of calibrated forecasts does not necessarily preserve calibration <xref ref-type="bibr" rid="bib1.bibx25" id="paren.21"><named-content content-type="pre">e.g.</named-content></xref>, additional constant predictors are added to the MOS equations as a remedy.
<xref ref-type="bibr" rid="bib1.bibx23" id="text.22"/> and <xref ref-type="bibr" rid="bib1.bibx27" id="text.23"/> state that automated warnings of wind gusts based on the combined product
achieve a performance that is comparable to that of human forecasters.</p>
      <p id="d1e271">The further outline of the paper is as follows:
after the introduction, the used observations and ensemble systems are introduced in Sect. <xref ref-type="sec" rid="Ch1.S2"/>. Thereafter, Sect. <xref ref-type="sec" rid="Ch1.S3"/> describes the conceptual design of Ensemble-MOS
with the definition of predictands and predictors and
provides technical details of the stepwise linear and logistic regressions.
Especially for extreme wind gusts, a global logistic regression is presented
that uses statistical forecasts of the speed of wind gusts as predictors for probabilities of strong events.
General caveats of MOS like model changes and forecast consistency are addressed at the end of that section.
The results shown here focus on wind gusts and are provided in Sect. <xref ref-type="sec" rid="Ch1.S4"/>.
Finally, Sect. <xref ref-type="sec" rid="Ch1.S5"/> provides a summary and conclusions.</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Observations and ensemble data</title>
      <p id="d1e290">Synoptic observations and model data from the ensemble systems COSMO-D2-EPS and ECMWF-ENS are used as training data and
for current statistical forecasts. Time series of 8 years of observations and model data have been gathered for training at the time of writing.
The used data are introduced in the following.</p>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Synoptic observations</title>
      <p id="d1e300">Observations of more than 320 synoptic stations within Germany and its surroundings are used
as part of the training data.
For short time forecasts the latest available observations at run time
are used also as predictors for the statistical modelling, which is described in more detail in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>.</p>
      <p id="d1e305">The synoptic observations include measurements of temperature, dew point, precipitation amounts,
wind speed and direction, speed of wind gusts, surface pressure, global radiation,
visibility, cloud coverage at several height levels, past and present weather and many more.
Past weather and present weather also contain observations of thunderstorm, kind of precipitation and fog amongst others. Ensemble-MOS derives all predictands that are relevant for weather warnings based on synoptic measurements
in a comprehensive approach
in order to provide the corresponding statistical forecasts.
In this paper, we focus on the speeds of wind gusts and on probabilities for severe storms.</p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>COSMO-D2-EPS and upscaled precipitation probabilities</title>
      <p id="d1e316">The ensemble system COSMO-D2-EPS of DWD consists of 20 members of the numerical model COSMO-D2.
It provides short-term weather forecasts for Germany, with runs every 3 h (i.e. 00:00, <inline-formula><mml:math id="M6" display="inline"><mml:mi mathvariant="normal">…</mml:mi></mml:math></inline-formula>, 21:00 UTC) with forecast steps of 1 up to 27 <inline-formula><mml:math id="M7" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> ahead (up to 45 <inline-formula><mml:math id="M8" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> for 03:00 UTC).
COSMO-D2 was upgraded from its predecessor model COSMO-DE on 15 May 2018, together with its ensemble system COSMO-D2-EPS;
the upgrade included an increase in horizontal resolution from 2.8 to 2.2 <inline-formula><mml:math id="M9" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> and an adapted orography.
Detailed descriptions of COSMO-DE and its ensemble system COSMO-DE-EPS are
provided in <xref ref-type="bibr" rid="bib1.bibx1" id="text.24"/>, <xref ref-type="bibr" rid="bib1.bibx8" id="text.25"/> and <xref ref-type="bibr" rid="bib1.bibx22" id="text.26"/>, respectively.
For the ensemble system initial and boundary conditions as well as physical parameterisations are
varied according to their assumed levels of uncertainty.</p>
      <p id="d1e360">For the postprocessing of COSMO-D2-EPS, 8 years of data have been gathered, including data from the predecessor system COSMO-DE-EPS, which has been available since 8 December 2010.
Thus, a number of model changes and updates are included in the data;
impacts on statistical forecasting are addressed later in Sect. <xref ref-type="sec" rid="Ch1.S3.SS4"/>.
Each run of Ensemble-MOS starts 2 h after the corresponding run of COSMO-D2-EPS to ensure that the ensemble system has finished and the data are available.</p>
      <p id="d1e365">Forecast probabilities of meteorological events can be estimated as the
relative frequency of the ensemble members that show the event of interest.
If the relative frequencies are evaluated grid point by grid point,
the probabilities imply that the event occurs within areas of the sizes of the grid cells,
which are 2.2 <inline-formula><mml:math id="M10" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 2.2 <inline-formula><mml:math id="M11" display="inline"><mml:mrow class="unit"><mml:msup><mml:mi mathvariant="normal">km</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> for COSMO-D2-EPS.
It is therefore<?pagebreak page476?> not straightforwardly possible to compare event probabilities of ensembles of numerical
models with different grid resolutions.</p>
      <p id="d1e386">For near-surface elements and short lead times, the COSMO-D2-EPS is often underdispersive and underestimates forecast errors.
Figure <xref ref-type="fig" rid="Ch1.F1"/> shows a rank histogram for 1-hourly precipitation amounts of COSMO-DE-EPS.
Too many observations have either less or more precipitation than
all members of the ensemble.
Using these relative frequencies as estimations of event probabilities statistically results
in too many probabilities with values 0 and 1.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F1"><?xmltex \currentcnt{1}?><label>Figure 1</label><caption><p id="d1e394">Rank/Talagrand histogram for 1-hourly precipitation amounts of COSMO-DE-EPS and forecast lead time 3 <inline-formula><mml:math id="M12" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>; data for 18 stations from 2011 to 2017.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f01.png"/>

        </fig>

      <p id="d1e411">Because of the high spatial variability
of precipitation, upscaled precipitation products are also derived from COSMO-D2-EPS, which are relative frequencies of precipitation events within areas of 10 <inline-formula><mml:math id="M13" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 10 grid points (i.e. 22 <inline-formula><mml:math id="M14" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 22 <inline-formula><mml:math id="M15" display="inline"><mml:mrow class="unit"><mml:msup><mml:mi mathvariant="normal">km</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula>).
A meteorological event (e.g. that the precipitation rate exceeds a certain threshold) is considered to occur within an area if the event occurs at at least one of its grid points. Area probabilities are therefore estimated straightforwardly as the relative number of ensemble members
predicting the area event, not requiring the event to take place at exactly the same grid point for all ensemble members.</p>
      <p id="d1e439">Certainly, these raw ensemble-based estimates are also affected by systematic errors of the numerical model COSMO-D2. <xref ref-type="bibr" rid="bib1.bibx16" id="text.27"/> observed a bias
of <inline-formula><mml:math id="M16" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">6.2</mml:mn></mml:mrow></mml:math></inline-formula> percentage points for the upscaled precipitation product of COSMO-DE-EPS for the probability that hourly precipitation rate exceeds 0.1 <inline-formula><mml:math id="M17" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">mm</mml:mi></mml:mrow></mml:math></inline-formula>.
Verification has been done against gauge-adjusted radar observations, which is a suitable observation system for areas.</p><?xmltex \hack{\newpage}?>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>ECMWF-ENS and TIGGE-data</title>
      <p id="d1e472">The ECMWF-ENS is a global ensemble system based on the Integrated Forecasting System (IFS)
of the European Centre for Medium-Range Weather Forecasts (ECMWF).
It consists of 50 perturbed members plus one control run and is computed twice a day for 00:00 and 12:00 UTC up to
15 d ahead (and even further with reduced resolution).
Postprocessing of Ensemble-MOS at DWD is based on the 00:00 UTC run with forecast lead times up to 10 d in steps of 3 h. Forecasts of ECMWF-ENS are interpolated from their genuine spectral resolution to a regular grid with 28 <inline-formula><mml:math id="M18" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> (0.25<inline-formula><mml:math id="M19" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>) mesh size.
Data have been gathered according to the availability of COSMO-DE/2-EPS since 8 December 2010.</p>
      <p id="d1e492">TIGGE data from 2002 to 2013 (see <xref ref-type="bibr" rid="bib1.bibx3" id="altparen.28"/> and <xref ref-type="bibr" rid="bib1.bibx32" id="altparen.29"/>) of ECMWF-ENS were used in a study to demonstrate the benefits of Ensemble-MOS prior to
unarchiving and downloading the gridded ensemble data mentioned above.
This study was restricted to the available set of model variables of TIGGE
(2 <inline-formula><mml:math id="M20" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> temperature, mean wind, cloud coverage and 24 <inline-formula><mml:math id="M21" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> precipitation);
results are given in Sect. <xref ref-type="sec" rid="Ch1.S4.SS2"/>.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Postprocessing by Ensemble-MOS</title>
      <p id="d1e528">The Ensemble-MOS of DWD is a model output statistics (MOS) system
specialised to postprocess the probabilistic information of NWP ensembles. Besides calibrating probabilistic forecasts,
Ensemble-MOS also simultaneously optimises continuous variables, e.g. precipitation amounts and the speeds of wind gusts.
Moreover, statistical interpretations also exist for meteorological elements that are not available from numerical forecasts (e.g. thunderstorm, fog or range of visibility).
In principle, all meteorological parameters and events that are regularly observed
can be forecasted statistically.
This includes temperature, dew point, wind speed and direction, wind gusts, surface pressure, global radiation,
visibility, cloud coverage at several height levels as well as
the synoptic weather with events of thunderstorm, special kinds of precipitation, fog and more.</p>
      <p id="d1e531">The basic concept of Ensemble-MOS is to use ensemble mean, spread, and other ensemble statistics as predictors
in multiple linear and logistic regressions.
The use of ensemble products as predictors instead of
processing each ensemble member individually prevents difficulties with
underdispersive statistical results and underestimated errors, especially for longer forecast horizons. Since MOS systems usually tend to converge towards
climatology due to the fading accuracy of numerical models
and the limited predictability of meteorological events <xref ref-type="bibr" rid="bib1.bibx33" id="paren.30"><named-content content-type="pre">e.g.</named-content></xref>,
individually processed members converge accordingly.
Moreover, multivariate MOS systems perform corrections
that depend on the selected set of predictors in order to reduce conditional biases.
If the postprocessing of the individual ensemble<?pagebreak page477?> members uses the same set of predictors,
the resulting statistical forecasts become correlated and underdispersive also for this reason.</p>
      <p id="d1e539">Ensemble-MOS is based on a MOS system originally set up for postprocessing
deterministic forecasts of the former global numerical model GME of DWD and of the deterministic, high-resolution IFS of ECMWF; see <xref ref-type="bibr" rid="bib1.bibx18" id="text.31"/>.
Using the ensemble mean and spread as model predictors allows application of the original MOS approach for deterministic NWP models to ensembles in a straightforward way.</p>
      <p id="d1e545">For continuous variables, such as temperature, precipitation amount or wind speed,
deterministic Ensemble-MOS forecasts and estimates of the associated forecast errors are obtained by multiple linear regression.
The MOS equation of a statistical estimate <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> with <inline-formula><mml:math id="M23" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula> predictors <inline-formula><mml:math id="M24" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M25" display="inline"><mml:mrow><mml:mi>k</mml:mi><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> coefficients <inline-formula><mml:math id="M26" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>
of a continuous predictand <inline-formula><mml:math id="M27" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> is
          <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M28" display="block"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula></p>
      <p id="d1e685">For events like thunderstorms, heavy precipitation or strong wind speed,
calibration of event occurrence or threshold exceedance probability is performed using multiple logistic regression.
For this an estimate
          <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M29" display="block"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></disp-formula>
        of the predictand <inline-formula><mml:math id="M30" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> is determined using a maximum likelihood approach.
The predictand <inline-formula><mml:math id="M31" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> now is a binary variable that is 1 in case the event was observed and 0 if not, whereas
the estimate <inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is considered a probability that takes values from 0 to 1.
Logistic regression <xref ref-type="bibr" rid="bib1.bibx17" id="paren.32"><named-content content-type="pre">e.g.</named-content></xref> is a classical approach for probabilistic postprocessing.</p>
      <p id="d1e787">Details of the implementation of linear and logistic regression are presented
in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and <xref ref-type="sec" rid="Ch1.S3.SS2"/>, respectively.
Especially for probabilities of strong and extreme wind gusts a global regression is applied that is presented in
Sect. <xref ref-type="sec" rid="Ch1.S3.SS3"/>.
For an introduction to MOS in general we refer to <xref ref-type="bibr" rid="bib1.bibx9" id="text.33"/>, <xref ref-type="bibr" rid="bib1.bibx36" id="text.34"/> and <xref ref-type="bibr" rid="bib1.bibx34" id="text.35"/>.</p>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Optimisation and interpretation by linear regression</title>
      <p id="d1e813">Ensemble-MOS derives altogether some 150 predictands
from synoptic observations (for precipitation  gauge-adjusted radar products can also be used) for statistical modelling. For the speeds of wind gusts and for precipitation amounts individual predictands for various reference periods
(e.g. 1-hourly, 3-hourly, 6-hourly and longer) are defined.
As usual, these predictands are modelled by individual linear regressions. The resulting statistical estimates
are added to the list of available predictors for subsequent regressions during postprocessing.
They are selected as predictors especially for probabilities that the speeds of wind gusts or precipitation amounts
exceed predefined thresholds within the corresponding time frames (i.e. 1, 3 <inline-formula><mml:math id="M33" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>, etc.).</p>
      <p id="d1e824">In order to estimate the error of the current forecast,
an error predictand
            <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M34" display="block"><mml:mrow><mml:msup><mml:mi>y</mml:mi><mml:mi>e</mml:mi></mml:msup><mml:mo>=</mml:mo><mml:mo>|</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi>y</mml:mi><mml:mo>|</mml:mo></mml:mrow></mml:math></disp-formula>
          is defined as the absolute value of the residuum.
The corresponding estimate <inline-formula><mml:math id="M35" display="inline"><mml:mrow><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi><mml:mi>e</mml:mi></mml:msubsup></mml:mrow></mml:math></inline-formula> is defined according to Eq. (<xref ref-type="disp-formula" rid="Ch1.E1"/>).
This error predictand can be evaluated as soon as the estimate <inline-formula><mml:math id="M36" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is available.
The absolute value is preferred over the root mean square (RMS) of the residuum, since it shows higher correlations with many predictors and a better linear fitting. For Gaussian distributions with density <inline-formula><mml:math id="M37" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">φ</mml:mi><mml:mrow><mml:mi mathvariant="italic">μ</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="italic">σ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> the absolute error <inline-formula><mml:math id="M38" display="inline"><mml:mi>e</mml:mi></mml:math></inline-formula>
(or mean absolute deviation) of the distribution can be estimated from the standard deviation <inline-formula><mml:math id="M39" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> as
            <disp-formula id="Ch1.E4" content-type="numbered"><label>4</label><mml:math id="M40" display="block"><mml:mtable class="split" columnspacing="1em" rowspacing="0.2ex" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:mi>e</mml:mi></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow><mml:mi mathvariant="normal">∞</mml:mi></mml:munderover><mml:mo>|</mml:mo><mml:mi>x</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="italic">μ</mml:mi><mml:mo>|</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msub><mml:mi mathvariant="italic">φ</mml:mi><mml:mrow><mml:mi mathvariant="italic">μ</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="italic">σ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msub><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo><mml:mi mathvariant="normal">d</mml:mi><mml:mi>x</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">2</mml:mn><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mi mathvariant="normal">∞</mml:mi></mml:munderover><mml:mi>x</mml:mi><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msqrt><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="italic">π</mml:mi></mml:mrow></mml:msqrt><mml:mi mathvariant="italic">σ</mml:mi></mml:mrow></mml:mfrac></mml:mstyle><mml:msup><mml:mi>e</mml:mi><mml:mstyle scriptlevel="+1"><mml:mfrac><mml:mrow><mml:mo>-</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:msup><mml:mi mathvariant="italic">σ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mstyle></mml:msup><mml:mi mathvariant="normal">d</mml:mi><mml:mi>x</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:msqrt><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">2</mml:mn><mml:mi mathvariant="italic">π</mml:mi></mml:mfrac></mml:mstyle></mml:msqrt><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>≈</mml:mo><mml:mn mathvariant="normal">0.8</mml:mn><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
      <p id="d1e1055">For each predictand the most relevant predictors are selected
from a predefined set of independent variables by stepwise regression.
Statistical modelling is performed for each predictand, station, season, forecast run and forecast time individually in general.
For rare events, however, nine zones with similar climatology are defined
(e.g. coastal strip, north German plain, various height zones in southern Germany, high mountain areas) and the stations are clustered together
in order to increase the number of observed events and the statistical significance of the training data.
All stations of a cluster are modelled together for those events.</p>
      <p id="d1e1058">Most potential predictors are based on
forecasts of the ensemble system,
which are interpolated to the locations of the observation sites.
Additional to the model values at the nearest grid point, the mean and standard deviation of the 6 <inline-formula><mml:math id="M41" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 6 and 11 <inline-formula><mml:math id="M42" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 11 surrounding grid points are also evaluated and provided as medium- and large-scale predictors,
respectively.
Moreover, extra variables are also derived from the NWP-model fields to be used as predictors, e.g. potential temperature, various atmospheric layer thicknesses, rotation and divergence of wind velocity,
dew-point spread and even special parameters, such as convective available potential energy (CAPE) and
severe weather threat index (SWEAT).
These variables are computed from the ensemble means of the required fields.</p>
      <p id="d1e1076">Statistical forecasts of the same variable of the last forecast step and
also of other variables of the current forecast step can be used as well.
For example, forecasts of 2 <inline-formula><mml:math id="M43" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> temperature may use statistical forecasts of precipitation amounts
of the same time step as predictors.
The order of the statistical modelling and of the forecasting is relevant in such cases to make sure the
required data are available.</p>
      <p id="d1e1087">Further predictors are derived from the latest observations that are available
at the time when the statistical forecast<?pagebreak page478?> is computed.
Generally, the latest observation is an excellent projection for short-term forecasts up to about 4 to 6 h, which is therefore added to the set of available predictors.
Special care has to be taken to process these
predictors for training, however. Only those observations
can be used that are available at run time of the forecast.
In case forecasts are computed for arbitrary locations
apart from observation and training sites,
these observations or persistency predictors have to be processed in exactly the same way in training and forecasting.
At locations other than observation sites, the required values need to be interpolated from the surrounding stations.
As interpolation generally is a weighted average based on horizontal and vertical distance,
it introduces smoothing and, with it, a systematic statistical change in the use of the observations.
If the training is performed using observations at the stations and the forecasting is
using interpolation, the statistical forecasts can be affected.
As a remedy, Ensemble-MOS uses observations as persistence predictors for training that are interpolated from
up to five surrounding stations in exactly the same way as when computing the forecast at arbitrary locations,
even if an observation at the correct location was available.</p>
      <p id="d1e1090">Special orographic predictors also exist, like height of station or height difference between station and model at a specific location. In order to address model changes, indicators or binary variables are also provided (see Sect. <xref ref-type="sec" rid="Ch1.S3.SS4"/> for details).
Altogether more than 300 independent variables are defined,
from which up to 10 predictors are selected for each predictand during multiple regression.</p>
      <p id="d1e1095">During stepwise regression,
the predictor with the highest correlation with the predictand is selected first from the set of available independent variables.
Next, the linear regression with the previously chosen set of predictors is computed
and the next predictor with the highest correlation with the residuum is selected, and so on.
Selection stops if no further predictor exists with a statistically significant correlation according to a Student's <inline-formula><mml:math id="M44" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>-test. The level of significance of the test is 0.18 divided by the number of available independent variables.
This division is used because of the high number of potential predictors.
With a type I error of e.g. 0.05 and a number of 300 available predictors, 15 predictors on average would be selected randomly
without providing significant information.
The value 0.18 is found to be a good compromise in order to select a meaningful number of predictors
and to prevent overfitting in this scenario.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T1" specific-use="star"><?xmltex \currentcnt{1}?><label>Table 1</label><caption><p id="d1e1108">Predictors for statistical forecasts of the maximal speed of wind gusts within 1 <inline-formula><mml:math id="M45" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> that have relative weights
higher than 1 %.
The relative weights are aggregated for all stations, seasons, forecast runs, and forecast lead times.
The ensemble mean of COSMO-D2-EPS is denoted by DMO (direct model output).
Parentheses within the predictor names denote time shifts.
For a time shift of <inline-formula><mml:math id="M46" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">30</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M47" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">min</mml:mi></mml:mrow></mml:math></inline-formula>, denoted as (<monospace>-0:30</monospace>), the predictors are interpolated
based on values for the previous and current forecast hours. The required statistical forecasts have to be evaluated in advance.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="3">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="left"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Predictor name</oasis:entry>
         <oasis:entry colname="col2">Rel. weight</oasis:entry>
         <oasis:entry colname="col3">Description</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2">(%)</oasis:entry>
         <oasis:entry colname="col3"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>FF(-0:30)StF</monospace></oasis:entry>
         <oasis:entry colname="col2">35.7</oasis:entry>
         <oasis:entry colname="col3">Statistical forecast of mean wind speed in 10 <inline-formula><mml:math id="M48" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> height</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>FX1(-1)StF</monospace></oasis:entry>
         <oasis:entry colname="col2">12.1</oasis:entry>
         <oasis:entry colname="col3">Statistical forecast of speed of wind gusts of the previous hour</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>VMAX_10M_LS</monospace></oasis:entry>
         <oasis:entry colname="col2">8.9</oasis:entry>
         <oasis:entry colname="col3">DMO of speed of wind gusts for a large surrounding area (mean of 11 <inline-formula><mml:math id="M49" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 11 grid points)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>VMAX_10M</monospace></oasis:entry>
         <oasis:entry colname="col2">8.6</oasis:entry>
         <oasis:entry colname="col3">DMO of speed of wind gusts for next model grid point</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>FF_850(-0:30)</monospace></oasis:entry>
         <oasis:entry colname="col2">4.8</oasis:entry>
         <oasis:entry colname="col3">DMO of wind speed in 800 hPa height</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>VMAX_10M_MS</monospace></oasis:entry>
         <oasis:entry colname="col2">4.1</oasis:entry>
         <oasis:entry colname="col3">DMO of speed of wind gusts for a medium surrounding area (mean of 6 <inline-formula><mml:math id="M50" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 6 grid points)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>Oa_D_0.5</monospace></oasis:entry>
         <oasis:entry colname="col2">3.2</oasis:entry>
         <oasis:entry colname="col3">Latest observation of the speed of wind gusts</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>FF_10m(-0:30)</monospace></oasis:entry>
         <oasis:entry colname="col2">2.8</oasis:entry>
         <oasis:entry colname="col3">DMO of mean wind speed in 10 <inline-formula><mml:math id="M51" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> height</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>StFT2m_T950</monospace></oasis:entry>
         <oasis:entry colname="col2">1.9</oasis:entry>
         <oasis:entry colname="col3">Statistical forecast of temperature difference between 2 <inline-formula><mml:math id="M52" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> and 950 hPa height (stability index)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>Location-height</monospace></oasis:entry>
         <oasis:entry colname="col2">1.6</oasis:entry>
         <oasis:entry colname="col3">Height of station</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><monospace>FF_1000(-0:30)</monospace></oasis:entry>
         <oasis:entry colname="col2">1.1</oasis:entry>
         <oasis:entry colname="col3">DMO of wind speed in 1000 hPa height</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d1e1362">Table <xref ref-type="table" rid="Ch1.T1"/> lists the most important predictors for statistical forecasts
of the maximal speed of wind gusts within 1 <inline-formula><mml:math id="M53" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>.
The relative weights are aggregated over 5472 equations,
one for each cluster, season, forecast run, and forecast lead time.
Note that predictors that are highly correlated usually exclude each other from appearing within one equation. Only the predictor with the highest correlation with the predictand is selected and supplants other correlated predictors that
do not provide enough additional information according to the <inline-formula><mml:math id="M54" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>-test.</p>
      <p id="d1e1382">The MOS equations are determined by stepwise regression for individual locations and, in case of rare events, for clusters. In order to compute statistical forecasts on a regular grid, these equations need to be evaluated for locations apart
from the training and observation sites.
In case of rare events and cluster equations, the appropriate cluster is determined for each grid point and the equation of that cluster is used. The equations for individual locations are interpolated to the required grid point by
linear interpolation of their coefficients.
In all cases, the required values of the numerical model for these equations are evaluated for the exact location.
Observations that are used as persistency predictors are interpolated from surrounding sites.
In this way, gridded forecast maps can be obtained as displayed in Fig. <xref ref-type="fig" rid="Ch1.F2"/>
for wind gust probabilities (see Sect. <xref ref-type="sec" rid="Ch1.S3.SS2"/> for probabilistic forecasts).
For computational efficiency, the forecasts are initially computed on a regular grid of 20 <inline-formula><mml:math id="M55" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> resolution
and are downscaled thereafter to 1 <inline-formula><mml:math id="M56" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> while taking into account
the various height zones in southern Germany.
The details of the downscaling are beyond the scope of the paper.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F2"><?xmltex \currentcnt{2}?><label>Figure 2</label><caption><p id="d1e1407">Probabilities of wind gusts higher than 14 <inline-formula><mml:math id="M57" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> on a regular 1 <inline-formula><mml:math id="M58" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> grid over Germany; 13 <inline-formula><mml:math id="M59" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecast lead time from Ensemble-MOS for COSMO-DE-EPS from 29 October 2018.</p></caption>
          <?xmltex \igopts{width=213.395669pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f02.png"/>

        </fig>

</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Calibration of probabilistic forecasts by logistic regression</title>
      <?pagebreak page479?><p id="d1e1457">Event probabilities are calibrated using logistic regression.
Equation (<xref ref-type="disp-formula" rid="Ch1.E2"/>) is solved using a maximum likelihood approach.
The likelihood function
            <disp-formula id="Ch1.E5" content-type="numbered"><label>5</label><mml:math id="M60" display="block"><mml:mrow><mml:mi>P</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∏</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi><mml:mi>i</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mrow><mml:msup><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msup></mml:mrow></mml:msup><mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi><mml:mi>i</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:msup><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msup></mml:mrow></mml:msup></mml:mrow></mml:math></disp-formula>
          expresses the probability that the predictand <inline-formula><mml:math id="M61" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> is realised given
the estimate <inline-formula><mml:math id="M62" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> via the coefficients <inline-formula><mml:math id="M63" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>
(and by now with fixed predictors <inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>),
with <inline-formula><mml:math id="M65" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula> being the time dimension (sample size) and <inline-formula><mml:math id="M66" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> the time index.
The predictand <inline-formula><mml:math id="M67" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> of an event probability is binomially distributed; its time-series values are defined as 1 in case the event was observed and 0 if not, whereby
conditional independence is assumed in Eq. (<xref ref-type="disp-formula" rid="Ch1.E5"/>).</p>
      <p id="d1e1642">It is mathematically equivalent and computationally more efficient to maximise the logarithm of the likelihood function
            <disp-formula id="Ch1.E6" content-type="numbered"><label>6</label><mml:math id="M68" display="block"><mml:mtable rowspacing="0.2ex" columnspacing="1em" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:mi>ln⁡</mml:mi></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mfenced close=")" open="("><mml:mrow><mml:mi>P</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mo>=</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msup><mml:mi>ln⁡</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi><mml:mi>i</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mo>+</mml:mo><mml:mfenced open="(" close=")"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:msup><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msup></mml:mrow></mml:mfenced><mml:mi>ln⁡</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi><mml:mi>i</mml:mi></mml:msubsup></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
          This maximisation is implemented by calling the routine <monospace>G02GBF</monospace> of the NAG library in FORTRAN 90; see <xref ref-type="bibr" rid="bib1.bibx21" id="text.36"/>.
The resulting fit of the estimate <inline-formula><mml:math id="M69" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> can be evaluated by the deviance
            <disp-formula id="Ch1.E7" content-type="numbered"><label>7</label><mml:math id="M70" display="block"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn><mml:mi>ln⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>P</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
          which is a measure analogous to the squared sum of residua in linear regression.</p>
      <p id="d1e1821">The selection of predictors is again performed stepwise.
Initially, the coefficient <inline-formula><mml:math id="M71" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> of the null model
<inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula>
that fits the mean of the predictand is determined
and the null deviance <inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn><mml:mi>ln⁡</mml:mi><mml:mo>(</mml:mo><mml:mi>P</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>)</mml:mo><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is computed. The coefficient <inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> is often called the intercept.
Starting from the null model the predictor
that is selected first is the one that shows the smallest deviance
<inline-formula><mml:math id="M75" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn><mml:mi>ln⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>P</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow></mml:math></inline-formula>. The difference <inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>D</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> is <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">χ</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:math></inline-formula>-distributed with 1 degree of freedom
and is used to check the statistical significance of the predictor.
This check replaces the <inline-formula><mml:math id="M78" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>-test in linear regression and uses the same statistical level.
If the predictor shows a significant contribution, it is accepted and further predictors are tested based on the
new model in the same way.
Otherwise, the predictor is rejected and the previous fitting <inline-formula><mml:math id="M79" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mrow><mml:mi>k</mml:mi><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is accepted as the final statistical model.</p>
      <p id="d1e2019">As a rule of thumb, for each selected predictor in the statistical model
at least 10 events need to be captured within the observation data (<italic>one in ten rule</italic>) to find stable coefficients. For example, with only 30 events in the training set, the number of predictors should be restricted to three.
This rule is critical especially for rare events such as extreme wind gusts or heavy precipitation.</p>
      <p id="d1e2026">Since testing all candidate predictors from the set of about 300 variables by computing their deviances is very costly,
the score test (Lagrange multiplier test) is actually applied to Ensemble-MOS.
Given a fitted logistic regression with <inline-formula><mml:math id="M80" display="inline"><mml:mrow><mml:mi>k</mml:mi><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> selected predictors,
the predictor is chosen next as <inline-formula><mml:math id="M81" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, which shows the steepest gradient of the log-likelihood function Eq. (<xref ref-type="disp-formula" rid="Ch1.E6"/>) in an absolute sense when introduced,
normalised by its standard deviation <inline-formula><mml:math id="M82" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, i.e.
            <disp-formula id="Ch1.E8" content-type="numbered"><label>8</label><mml:math id="M83" display="block"><mml:mtable rowspacing="0.2ex" columnspacing="1em" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mfenced close="|" open="|"><mml:mrow><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mo>∂</mml:mo><mml:mi>ln⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>P</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:msub><mml:mo mathsize="1.5em">|</mml:mo><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfenced><mml:mo>=</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mspace linebreak="nobreak" width="0.25em"/><mml:mfenced close="|" open="|"><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:mfenced open="(" close=")"><mml:mrow><mml:msup><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msup><mml:mo>-</mml:mo><mml:msubsup><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mrow><mml:mi>k</mml:mi><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>i</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msubsup><mml:mi>x</mml:mi><mml:mi>k</mml:mi><mml:mi>i</mml:mi></mml:msubsup></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
          This equation results from basic calculus
including the identity <inline-formula><mml:math id="M84" display="inline"><mml:mrow><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mi>k</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.
The right-hand side of Eq. (<xref ref-type="disp-formula" rid="Ch1.E8"/>) is basically the correlation of the current residuum with the new predictor. The score test thus results in the same selection<?pagebreak page480?> criterion as applied to stepwise linear regression.
Once the predictor <inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is selected, the coefficients <inline-formula><mml:math id="M86" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> are updated to maximise Eq. (<xref ref-type="disp-formula" rid="Ch1.E6"/>).</p>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title>Global logistic regression of wind gust probabilities</title>
      <p id="d1e2325">For extreme events, the number of observed occurrences can still be too small
to derive stable MOS equations, although time series of several years have been gathered and the stations
are clustered within climatologic zones in Germany.
The eight warning thresholds of DWD for wind gusts range from 12.9 <inline-formula><mml:math id="M87" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> (25.0 <inline-formula><mml:math id="M88" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">kn</mml:mi></mml:mrow></mml:math></inline-formula>, proper wind gusts)
up to 38.6 <inline-formula><mml:math id="M89" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> (75.0 <inline-formula><mml:math id="M90" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">kn</mml:mi></mml:mrow></mml:math></inline-formula>, extreme gales),
whereas the maximal observed speed of wind gusts in the training data for a cluster in the northern German plains
is only 25.4 <inline-formula><mml:math id="M91" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>.
Especially for probabilities of extreme wind gusts
global logistic regressions are developed that use events at the coastal strip or at mountains in southern Germany and allow for
meaningful statistical forecasts of extreme events also in climatologically calm areas.
The statistical forecasts of the continuous speed of wind gusts are used for these logistic regressions as the only predictors.
They are modelled by stepwise linear regression for each station individually,
as described in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>.
In this way, rare occurrences of extreme events are gathered globally while concurrently a certain degree of locality is maintained.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F3" specific-use="star"><?xmltex \currentcnt{3}?><label>Figure 3</label><caption><p id="d1e2400">Observed cumulative distributions of wind gusts exceeding threshold 13.9 <inline-formula><mml:math id="M92" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> (blue) and fit of logistic distribution (green) depending on statistically optimised forecasts of wind gusts for
forecast lead times 1 <inline-formula><mml:math id="M93" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> <bold>(a)</bold> and 7 <inline-formula><mml:math id="M94" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> <bold>(b)</bold>. The threshold is dashed.</p></caption>
          <?xmltex \igopts{width=497.923228pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f03.png"/>

        </fig>

      <p id="d1e2448">The locally optimised and unbiased forecasts of wind gust speeds are excellent predictors
for wind gust probabilities.
The logistic regressions according to Eq. (<xref ref-type="disp-formula" rid="Ch1.E2"/>) with <inline-formula><mml:math id="M95" display="inline"><mml:mrow><mml:mi>k</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>
fit the distributions of observed wind gusts quite well, as shown in Fig. <xref ref-type="fig" rid="Ch1.F3"/> for threshold <inline-formula><mml:math id="M96" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">13.9</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M97" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> (27.0 <inline-formula><mml:math id="M98" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">kn</mml:mi></mml:mrow></mml:math></inline-formula>) and forecast lead times of 1 and 7 <inline-formula><mml:math id="M99" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>, respectively.</p>
      <p id="d1e2514">The statistical modelling of wind gust probabilities is performed for each threshold <inline-formula><mml:math id="M100" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula> individually and is described in the
following.
The logistic regressions represent logistic distributions with
mean <inline-formula><mml:math id="M101" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula> and variance <inline-formula><mml:math id="M102" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:msup><mml:mi mathvariant="italic">π</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow><mml:mrow><mml:mn mathvariant="normal">3</mml:mn><mml:msubsup><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula>, which
are computed for various lead times <inline-formula><mml:math id="M103" display="inline"><mml:mi>h</mml:mi></mml:math></inline-formula> and are listed in Table <xref ref-type="table" rid="Ch1.T2"/>.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T2"><?xmltex \currentcnt{2}?><label>Table 2</label><caption><p id="d1e2598">Parameters of fitted logistic distributions as shown in Fig. <xref ref-type="fig" rid="Ch1.F3"/> for various
forecast lead times <inline-formula><mml:math id="M104" display="inline"><mml:mi>h</mml:mi></mml:math></inline-formula>,  with coefficients of logistic regressions <inline-formula><mml:math id="M105" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and resulting means <inline-formula><mml:math id="M107" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and standard deviations <inline-formula><mml:math id="M108" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> for threshold <inline-formula><mml:math id="M109" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">13.9</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M110" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. Estimated uncertainties are given in brackets.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.89}[.89]?><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="center"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:colspec colnum="5" colname="col5" align="center"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"><inline-formula><mml:math id="M111" display="inline"><mml:mi>h</mml:mi></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M112" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M113" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M115" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">1</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">17.74</mml:mn></mml:mrow></mml:math></inline-formula> (<inline-formula><mml:math id="M117" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.17</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3">1.30 (<inline-formula><mml:math id="M118" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col4">13.68 (<inline-formula><mml:math id="M119" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col5">1.40 (<inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">4</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M121" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">13.76</mml:mn></mml:mrow></mml:math></inline-formula> (<inline-formula><mml:math id="M122" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.11</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3">1.01 (<inline-formula><mml:math id="M123" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col4">13.66 (<inline-formula><mml:math id="M124" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col5">1.80 (<inline-formula><mml:math id="M125" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.02</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">7</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M126" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">13.32</mml:mn></mml:mrow></mml:math></inline-formula> (<inline-formula><mml:math id="M127" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.11</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3">0.98 (<inline-formula><mml:math id="M128" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col4">13.66 (<inline-formula><mml:math id="M129" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col5">1.86 (<inline-formula><mml:math id="M130" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.02</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">10</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M131" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">13.00</mml:mn></mml:mrow></mml:math></inline-formula> (<inline-formula><mml:math id="M132" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.10</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3">0.95 (<inline-formula><mml:math id="M133" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col4">13.66 (<inline-formula><mml:math id="M134" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col5">1.91 (<inline-formula><mml:math id="M135" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.02</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">16</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M136" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">12.51</mml:mn></mml:mrow></mml:math></inline-formula> (<inline-formula><mml:math id="M137" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.10</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3">0.92 (<inline-formula><mml:math id="M138" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col4">13.66 (<inline-formula><mml:math id="M139" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.01</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col5">1.98 (<inline-formula><mml:math id="M140" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">0.02</mml:mn></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

      <p id="d1e3100">The expectations <inline-formula><mml:math id="M141" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> are slightly smaller than the threshold <inline-formula><mml:math id="M142" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">13.9</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M143" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>,
almost independently of lead time.
The reason is that
for given statistical forecasts of wind gusts the distribution of observations is
almost Gaussian (see Fig. <xref ref-type="fig" rid="Ch1.F4"/>) albeit a little left skewed with a small number of very weak wind observations.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F4"><?xmltex \currentcnt{4}?><label>Figure 4</label><caption><p id="d1e3147">Distribution of wind gust observations from 2011 to 2016 for 178 synoptic stations for cases
where statistical forecast (fitted for that period) is between 21 and 25 <inline-formula><mml:math id="M144" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. Lead time is 1 <inline-formula><mml:math id="M145" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>.
Gaussian fit and mean of observations (green) and mean of forecasts (red).</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f04.png"/>

        </fig>

      <p id="d1e3181">The standard deviations <inline-formula><mml:math id="M146" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> increase with forecast lead time, reflecting the loss of accuracy of the statistical forecasts.
Consequently, the graph of the cumulative distribution function in Fig. <xref ref-type="fig" rid="Ch1.F3"/>
is more tilted for a forecast lead time of 7 <inline-formula><mml:math id="M147" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> than for 1 <inline-formula><mml:math id="M148" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>.</p>
      <p id="d1e3214">Figure <xref ref-type="fig" rid="Ch1.F5"/> shows fitted variances <inline-formula><mml:math id="M149" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:math></inline-formula> of the eight individual forecast runs of Ensemble-MOS for COSMO-DE-EPS
and their mean depending on lead time.
In order to reduce the number of coefficients and to increase consistency and robustness of the forecasts,
the variance <inline-formula><mml:math id="M150" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:math></inline-formula> is parameterised
depending on forecast lead time <inline-formula><mml:math id="M151" display="inline"><mml:mi>h</mml:mi></mml:math></inline-formula>
by fitting the function
            <disp-formula id="Ch1.E9" content-type="numbered"><label>9</label><mml:math id="M152" display="block"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup><mml:mo>(</mml:mo><mml:mi>h</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mi>log⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi>a</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mi>h</mml:mi><mml:mo>+</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:math></disp-formula>
          with its parameters <inline-formula><mml:math id="M153" display="inline"><mml:mrow><mml:msub><mml:mi>a</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M154" display="inline"><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, and <inline-formula><mml:math id="M155" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> for threshold <inline-formula><mml:math id="M156" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F5"><?xmltex \currentcnt{5}?><label>Figure 5</label><caption><p id="d1e3339">Variances <inline-formula><mml:math id="M157" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:math></inline-formula> of logistic distributions fitted to cumulative distributions
of observed wind gusts for threshold <inline-formula><mml:math id="M158" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">13.9</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M159" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> depending on forecast lead time.
Individual runs of Ensemble-MOS for COSMO-DE-EPS-MOS starting at 02:00, 05:00, <inline-formula><mml:math id="M160" display="inline"><mml:mi mathvariant="normal">…</mml:mi></mml:math></inline-formula>, 23:00 UTC in colours,
mean of all runs in black, fitted parameterisation of variances dashed in black.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f05.png"/>

        </fig>

      <p id="d1e3397">The fitted expectations and variances show weak dependencies on the time of the day and are neglected.
The logistic regressions of wind gust probabilities thus
can be expressed for each threshold <inline-formula><mml:math id="M161" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula> by the mean <inline-formula><mml:math id="M162" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M163" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> for
all start times of the forecasts in the same way.</p>
      <p id="d1e3429">Even for very rare gales of 38.6 <inline-formula><mml:math id="M164" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>
more than 130 events are captured using 6 years of training data when modelling all stations and forecast runs together, which is sufficient for logistic regression.
Training for these extreme events is based mainly on coastal and mountain stations, but
the statistical regressions are applied to less exposed locations in calmer regions as well.
Small threshold probabilities will be predicted for those locations in general.
However, meaningful estimations will be generated
once the statistical forecasts of local wind speed rise induced by the numerical model.</p>
</sec>
<sec id="Ch1.S3.SS4">
  <label>3.4</label><title>Specific issues and caveats of MOS</title>
      <p id="d1e3457">Ensemble-MOS optimises and calibrates
ensemble forecasts using synoptic observations.
Being a statistical method, it is vulnerable to systematic changes in input data,
since it assumes that errors and characteristics of the past persist in future.
An important part of the input are observations, whose measurement instruments sometimes change.
It is recommended to use quality-checked observations in order to avoid the use of defective values for training.
Especially observation sites that are automatised need to be screened.
Furthermore, numerical models change with new versions and updates that can affect statistical postprocessing,
as further discussed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS4.SSS1"/>.</p>
      <p id="d1e3462">Although statistical forecasts generally improve the model output when
verified against observations, the results are not always consistent in time, space
and between the forecast variables (e.g. between temperature and dew point),
if they are optimised individually.
This issue is addressed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS4.SSS2"/>.</p>
<?pagebreak page481?><sec id="Ch1.S3.SS4.SSS1">
  <label>3.4.1</label><title>Model changes</title>
      <p id="d1e3474">Statistical methods like Ensemble-MOS detect systematic errors and deficiencies of NWP models during a past training period in order to improve topical operational forecasts.
Implicitly it is assumed that the systematic characteristics of the NWP models persist. Note that multiple regressions correct not only for model bias but also for conditional biases that depend on other meteorological variables.
Multiple regressions are more vulnerable to model changes than simple regressions, therefore.
Systematic changes in NWP models can affect statistical forecasts, even if the NWP forecasts are objectively improved as confirmed by verification.
Given that the statistical modelling provides unbiased estimations, any systematic change in
NWP-model predictors will be reflected in biases in the statistical forecasts. The resulting biases depend on the magnitudes of the changes of the predictors and on their weights in the MOS equations.</p>
      <p id="d1e3477">One remedy for jumps in input data is the use of indicator (binary) predictors.
These predictors are related to the date of the change of the NWP model and are defined as 1 before<?pagebreak page482?> and 0 after. When they are selected during stepwise regression, they account for sudden jumps in the training data and can prevent the introduction of unconditional biases
in the statistical forecasts. Conditional biases depending on other forecast variables, however,
are not corrected.</p>
      <p id="d1e3480">In order to process extreme and very rare events for weather warnings,
long time series of 7 years of data for COSMO-D2-EPS have been gathered at the time of writing. Hence, the time series are subject to a number of model changes.
A significant model upgrade from COSMO-DE to COSMO-D2, including an increase in horizontal resolution from 2.8 to 2.2 <inline-formula><mml:math id="M165" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:math></inline-formula> and an update of orography, took place in May 2018.
Since reforecasting of COSMO-D2-EPS for more than 1 year was technically not possible, the existing COSMO-DE-EPS database was used further and extended with reforecasts of COSMO-D2-EPS of the year
before operational introduction.
However, statistical experiments using these reforecasts of COSMO-D2-EPS
(and the use of binary predictors; see above) revealed only insignificant improvements compared to training with data of COSMO-DE-EPS only.
For rare events, longer time series are considered more important than the use of unaltered model versions.</p>
</sec>
<sec id="Ch1.S3.SS4.SSS2">
  <label>3.4.2</label><title>Forecast consistency</title>
      <p id="d1e3499">As weather warnings are issued for a certain period of time and a specified region,
continuity of probabilistic forecasts in time and space is important.
It should be accepted, however,
that maps of probabilistic forecasts do not comply with deterministic runs of numerical models,
as probabilistic forecasts are smoothed according to forecast uncertainty.
For example, there are hardly convective cells in probabilistic forecasts, but rather
areas exist where convection might occur with a certain probability within a given time period.</p>
      <p id="d1e3502">The statistical modelling of Ensemble-MOS is carried out for each forecast variable, forecast lead time
and location independently and individual MOS equations are derived. For rare meteorological events clusters of stations are grouped together
that are similar in climatology in order to derive individual cluster equations.
This local and individual fitting results in optimal statistical forecasts for the specific time, location and variable
as measured with the RMSE compared to observations.
However, it does not guarantee that obtained forecast fields are consistent in space, time or between variables.</p>
      <p id="d1e3505">In forecast time, spurious jumps of statistical forecasts can appear, and variables with different reference periods usually do not match. For example, the sum of 12 successive 1-hourly precipitation amounts would not
equal the corresponding 12-hourly amount if the latter is modelled as an individual predictand. Statistical forecasts of temperature cannot be guaranteed to exceed those of dew point.
Maps of statistical forecasts show high variability from station to station and
unwanted anomalies in case of cluster equations.
Cluster edges turn up and it may appear that there are higher wind gusts
in a valley than on a mountain nearby, in cases where the locations are arranged
in different clusters, for example.
For consistency in time and space the situation can be improved by using the same equations for several lead times
and for larger clusters or by elaborate subsequent smoothing.
However, forecast quality for a given space and time will be degraded consequently.
For consistency between all forecast variables multivariate regressions are required that model the relevant
predictands simultaneously.</p>
      <p id="d1e3508">From the point of view of probabilistic forecasting, however, statistical forecasts are
random variables with statistical distributions,
although commonly only their expectations are considered <italic>the</italic> statistical forecast. In case forecast consistency is violated from a deterministic point of view,
this is not the case if statistical errors are taken into account.
The statistical forecasts remain valid
as long as the probability distributions of the variables overlap.
As this is a mathematical point of view, the question remains how to communicate this nature of probabilistic forecasts to the public or traditional meteorologists in terms
of useful and accepted products.</p>
</sec>
</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Results</title>
      <p id="d1e3524">Evaluation of Ensemble-MOS for COSMO-DE-EPS and ECMWF-ENS is
provided in the following.
Although Ensemble-MOS of DWD provides statistical forecasts of many
forecast variables that are relevant for warnings,
evaluation is focused on wind gusts for COSMO-DE-EPS and
on temperature for ECMWF-ENS in order to limit the scope of the paper.</p>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>Evaluation of Ensemble-MOS for COSMO-DE-EPS</title>
      <p id="d1e3534">Verifications of the continuous speed of wind gusts are presented in
Figs. <xref ref-type="fig" rid="Ch1.F6"/>–<xref ref-type="fig" rid="Ch1.F8"/> by various scatter diagrams
including forecast means (solid line) and their standard errors (dashed lines).
Figures <xref ref-type="fig" rid="Ch1.F6"/> (right) and <xref ref-type="fig" rid="Ch1.F7"/> (right) show
the statistical fit of the speed of wind gusts against synoptic observations during a training period of 6 years of Ensemble-MOS for COSMO-DE-EPS for lead times of 1 and 6 <inline-formula><mml:math id="M166" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula>, respectively.
The fit is almost unbiased for all forecast speed levels.
The raw ensemble means show overforecasting for high wind gusts (same figures, left) and the
standard errors are considerably larger.
If no overfitting occurs, out-of-sample forecasts are expected to behave accordingly,
which is verified in Fig. <xref ref-type="fig" rid="Ch1.F8"/> (right) for a test period of 3 months (at least for wind gusts up to about 20 <inline-formula><mml:math id="M167" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>).</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F6" specific-use="star"><?xmltex \currentcnt{6}?><label>Figure 6</label><caption><p id="d1e3575">Scatter plots of ensemble means of 3 <inline-formula><mml:math id="M168" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecasts of the speeds of wind gusts of COSMO-DE-EPS versus observations <bold>(a)</bold>
and corresponding statistical fits of 1 <inline-formula><mml:math id="M169" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecasts of Ensemble-MOS versus the same observations <bold>(b)</bold>. Means of observations (solid) and confidence intervals (means <inline-formula><mml:math id="M170" display="inline"><mml:mo>±</mml:mo></mml:math></inline-formula> standard deviations, dashed) are shown.
Six years of data (2011–2016) are used; number of cases are given by histograms.</p></caption>
          <?xmltex \igopts{width=426.791339pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f06.png"/>

        </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F7" specific-use="star"><?xmltex \currentcnt{7}?><label>Figure 7</label><caption><p id="d1e3615">As Fig. <xref ref-type="fig" rid="Ch1.F6"/> but for 8 <inline-formula><mml:math id="M171" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecasts of COSMO-DE-EPS <bold>(a)</bold> and 6 <inline-formula><mml:math id="M172" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecasts of Ensemble-MOS <bold>(b)</bold>.</p></caption>
          <?xmltex \igopts{width=426.791339pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f07.png"/>

        </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F8" specific-use="star"><?xmltex \currentcnt{8}?><label>Figure 8</label><caption><p id="d1e3651">As Fig. <xref ref-type="fig" rid="Ch1.F6"/> but for 3 months of data (May–July 2016) and forecasts of COSMO-DE-EPS <bold>(a)</bold> and Ensemble-MOS <bold>(b)</bold>. Data of this period were not used for training.</p></caption>
          <?xmltex \igopts{width=426.791339pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f08.png"/>

        </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F9" specific-use="star"><?xmltex \currentcnt{9}?><label>Figure 9</label><caption><p id="d1e3670">Scatter plots of 3 <inline-formula><mml:math id="M173" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> forecasts of the absolute errors of COSMO-DE-EPS forecasts of wind gust speeds (estimated as ensemble standard deviations <inline-formula><mml:math id="M174" display="inline"><mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">0.8</mml:mn></mml:mrow></mml:math></inline-formula>) versus observed absolute errors of the ensemble means <bold>(a)</bold> and corresponding 1 <inline-formula><mml:math id="M175" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> error forecasts of Ensemble-MOS versus observed absolute errors of Ensemble-MOS
(statistical fit of training period, <bold>b</bold>). Means of observed absolute errors (solid) and confidence intervals (means <inline-formula><mml:math id="M176" display="inline"><mml:mo>±</mml:mo></mml:math></inline-formula> standard deviations, dashed) are shown. Six years of data (2011–2016) are used.</p></caption>
          <?xmltex \igopts{width=426.791339pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f09.png"/>

        </fig>

      <p id="d1e3719">Ensemble-MOS can predict its own current forecast errors by using error predictands according to Eq. (<xref ref-type="disp-formula" rid="Ch1.E3"/>).
Forecasts of the absolute errors of the speeds of wind gusts
are related to observed errors in Fig. <xref ref-type="fig" rid="Ch1.F9"/> (right).
The biases are small, although individual observed errors are much larger than their predictions.
The absolute errors of the ensemble mean<?pagebreak page483?> versus ensemble spread (normalised to absolute error)
strongly underestimate the observed errors of the ensemble mean; see Fig. <xref ref-type="fig" rid="Ch1.F9"/> (left). This is another example of underestimated dispersion of COSMO-DE-EPS as shown in Fig. <xref ref-type="fig" rid="Ch1.F1"/> for precipitation.</p>
      <p id="d1e3730">The statistical forecasts of the speeds of the wind gusts are excellent predictors for
the probabilities that certain warning thresholds are exceeded.
This is demonstrated by the fits of the observed distributions by logistic regression
as shown in Fig. <xref ref-type="fig" rid="Ch1.F3"/>.
The global logistic regression presented in Sect. <xref ref-type="sec" rid="Ch1.S3.SS3"/>
is prepared for extreme and rare events; nevertheless, it is applicable to lower thresholds as well. The reliability diagram in Fig. <xref ref-type="fig" rid="Ch1.F10"/> shows well-calibrated probabilities for wind gusts exceeding 7.7 <inline-formula><mml:math id="M177" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> for a
zone in the northern German plains with calmer winds in climatology.
The COSMO-DE-EPS shows strong overforecasting in these situations.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F10"><?xmltex \currentcnt{10}?><label>Figure 10</label><caption><p id="d1e3758">Reliability diagram for probabilities of wind gusts exceeding 7.7 <inline-formula><mml:math id="M178" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> (15.0 <inline-formula><mml:math id="M179" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">kn</mml:mi></mml:mrow></mml:math></inline-formula>). Probabilistic forecasts of Ensemble-MOS with a lead time of 6 <inline-formula><mml:math id="M180" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> (green) and corresponding relative frequencies of COSMO-DE-EPS with a lead time of 8 <inline-formula><mml:math id="M181" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> (blue). Verification is done for 3 months of data (May–July 2016) and 18 stations in Germany at about N52<inline-formula><mml:math id="M182" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> latitude, including Berlin for example. Vertical lines are 5 %–95 % consistency bars according to <xref ref-type="bibr" rid="bib1.bibx4" id="text.37"/>.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f10.png"/>

        </fig>

</sec>
<?pagebreak page484?><sec id="Ch1.S4.SS2">
  <label>4.2</label><title>Evaluation of Ensemble-MOS for ECMWF-ENS</title>
      <p id="d1e3829">In order to motivate the use of Ensemble-MOS for ECMWF-ENS, a study has been carried out with a restricted set of model variables of TIGGE; see Sect. <xref ref-type="sec" rid="Ch1.S2.SS3"/>. Training is based on ensemble data and corresponding observations from 2002 to 2012, whereas
statistical forecasting and verification is performed for 2013; see <xref ref-type="bibr" rid="bib1.bibx15" id="text.38"/> for details.</p>
      <p id="d1e3837">Results for 2 <inline-formula><mml:math id="M183" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> temperature forecasts are shown in Fig. <xref ref-type="fig" rid="Ch1.F11"/>, which
illustrates essential improvements of postprocessed forecasts of Ensemble-MOS compared to
raw ensemble output. The statistical forecast (blue) not only improves the raw ensemble mean (red), but it also outperforms the high-resolution ECMWF-IFS (these data have not been used for training).
Also, the statistical estimation of Ensemble-MOS of its own errors (pink) (see Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>) is more realistic over<?pagebreak page485?> the first few days than the estimate of the ensemble mean errors by the ensemble spread (yellow).
Improvements of ECMWF-ENS with Ensemble-MOS were also obtained for 24 <inline-formula><mml:math id="M184" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> precipitation and cloud coverage.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F11"><?xmltex \currentcnt{11}?><label>Figure 11</label><caption><p id="d1e3862">Mean absolute error (MAE) of 2 <inline-formula><mml:math id="M185" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> temperature forecast and error estimations depending on forecast lead time. Spread (yellow): spread of ECMWF-ENS (normalised to MAE); MAE Ensemble Mean (red): MAE of the mean of ECMWF-ENS; MAE Ctrl (grey): MAE of the ECMWF-ENS control run; MA EMOS (pink): Ensemble-MOS forecast of its own absolute errors (see Eq. <xref ref-type="disp-formula" rid="Ch1.E3"/>: estimations of MAE EMOS, blue); MAE MOS (blue): MAE of Ensemble-MOS for ECMWF-ENS; MAE HR (green): MAE of high-resolution ECMWF-IFS.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://npg.copernicus.org/articles/27/473/2020/npg-27-473-2020-f11.png"/>

        </fig>

</sec>
</sec>
<sec id="Ch1.S5" sec-type="conclusions">
  <label>5</label><title>Conclusions</title>
      <p id="d1e3891">This paper describes the Ensemble-MOS system of DWD, which is set up to
postprocess the ensemble systems COSMO-D2-EPS and ECMWF-ENS
with respect to severe weather to support warning management.
MOS in general is a mature and sound method and, in combination with logistic regression, it can provide optimised and calibrated statistical forecasts.
Stepwise multiple regression allows reduction of conditional biases that depend on the meteorological situation, which is defined by the selected predictors.
The setup of Ensemble-MOS to use ensemble mean and spread as predictors
is computationally efficient and simplifies forecasting of calibrated event probabilities
and error estimates on longer forecast lead times.
Ensemble-MOS is operationally applicable with regard to its robustness and computational costs and runs in trial mode
in order to support warning management at DWD.</p>
      <p id="d1e3894">The ensemble spread is less often detected as an important predictor, as might be expected, however. One reason is that the spread actually carries less information about forecast accuracy than
originally intended. It is often too small and too steady to account for current forecast errors.
Another reason is that some forecast variables correlate with their own forecast errors (e.g. precipitation and wind gusts).
If the ensemble spread does not provide enough independent information,
it is not selected additionally to the ensemble mean during stepwise regression.
Currently, only ensemble mean and spread are provided as predictors for Ensemble-MOS.
The implementation of various ensemble quantiles as additional predictors is technically straightforward
and could improve the exploitation of the probabilistic information of the ensemble.</p>
      <p id="d1e3897">Statistical forecasts of the speed of the wind gusts are excellent predictors for
probabilities that given thresholds are exceeded and are used as predictors within logistic regressions.
The same approach could be advantageous for probabilities of heavy precipitation as well,
where estimated precipitation amounts would be used as predictors.</p>
      <p id="d1e3900">An important further step in probabilistic forecasting is the estimation of complete (calibrated) distributions of
forecast variables rather than forecasting only discrete threshold probabilities.
For wind gusts with Gaussian conditional errors as shown in Fig. <xref ref-type="fig" rid="Ch1.F4"/>
this seems possible but certainly requires additional research.</p>
      <p id="d1e3906">With its inherent linearity (also in the case of logistic regressions, there are linear combinations of predictors only) MOS has its restrictions in modelling but supports traceability and robustness, which are important features in operational weather forecasting.
Therefore, MOS is considered a possible baseline for future statistical approaches based
on neural networks and machine learning that allow for more<?pagebreak page486?> general statistical modelling.
Many of the statistical problems will remain however, such as finding suitable reactions to changes in the NWP models, (deterministic) consistency and the definition of useful probabilistic products (see Sect. <xref ref-type="sec" rid="Ch1.S3.SS4.SSS2"/>)
and the verification of rare events.
In all cases, training data are considered of utmost importance, including the NWP-model output, as well as quality-checked historic observations.</p>
</sec>

      
      </body>
    <back><notes notes-type="dataavailability"><title>Data availability</title>

      <p id="d1e3915">COSMO-DE-EPS data and synoptic observations are stored in DWD archives and can be made accessible under certain conditions.
Further information is available at <uri>https://opendata.dwd.de</uri> (DWD, 2020).
TIGGE data are available free of charge, see <uri>https://confluence.ecmwf.int/display/TIGGE</uri> (ECMWF, 2020).</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d1e3927">Conceptual design of Ensemble-MOS, enhancement of software for probabilistic forecasting of ensembles (including logistic regression), processing of forecasts, verification and writing was done by the author.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d1e3933">The author declares that there is no conflict of interest.</p>
  </notes><notes notes-type="sistatement"><title>Special issue statement</title>

      <p id="d1e3939">This article is part of the special issue “Advances in post-processing and blending of deterministic and ensemble forecasts”. It is not associated with a conference.</p>
  </notes><ack><title>Acknowledgements</title><p id="d1e3945">The author thanks two anonymous reviewers and the editor for their constructive comments, which helped to improve the structure and clarity of the manuscript. Thanks to James Paul for improving the use of English.</p></ack><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d1e3950">This paper was edited by Stephan Hemri and reviewed by two anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><?xmltex \def\ref@label{{Baldauf et~al.(2011)Baldauf, Seifert, F\"{o}rstner, Majewski,
Raschendorfer, and Reinhardt}}?><label>Baldauf et al.(2011)Baldauf, Seifert, Förstner, Majewski,
Raschendorfer, and Reinhardt</label><?label baldauf2011?><mixed-citation>Baldauf, M., Seifert, A., Förstner, J., Majewski, D., Raschendorfer, M., and
Reinhardt, T.: Operational convective-scale numerical weather prediction  with the COSMO model: description and sensitivities, Mon. Weather Rev.,
139, 3887–3905, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-10-05013.1" ext-link-type="DOI">10.1175/MWR-D-10-05013.1</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx2"><?xmltex \def\ref@label{{{Ben~Bouall\`{e}gue} et~al.(2015){Ben~Bouall\`{e}gue}, Pinson, and
Friederichs}}?><label>Ben Bouallègue et al.(2015)Ben Bouallègue, Pinson, and
Friederichs</label><?label benbeouallegue2015?><mixed-citation>Ben Bouallègue, Z., Pinson, P., and Friederichs, P.: Quantile forecast
discrimination ability and value, Q. J. R. Meteorol. Soc., 141, 3415–3424, <ext-link xlink:href="https://doi.org/10.1002/qj.2624" ext-link-type="DOI">10.1002/qj.2624</ext-link>,  2015.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Bougeault et al.(2010)</label><?label bougeault2010?><mixed-citation>Bougeault, P.,  Toth, Z.,  Bishop, C., et al.: The THORPEX interactive grand global ensemble, Bull. Amer. Meteor. Soc., 91, 1059–1072, <ext-link xlink:href="https://doi.org/10.1175/2010BAMS2853.1" ext-link-type="DOI">10.1175/2010BAMS2853.1</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx4"><?xmltex \def\ref@label{{Br\"{o}cker and Smith(2006)}}?><label>Bröcker and Smith(2006)</label><?label broecker2006?><mixed-citation>Bröcker, J. and Smith, L. A.: Increasing the Reliability of Reliability
Diagrams, Weather Forecast., 22, 651–661, <ext-link xlink:href="https://doi.org/10.1175/WAF993.1" ext-link-type="DOI">10.1175/WAF993.1</ext-link> 2006.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Buizza(2018)</label><?label buizza2018?><mixed-citation>
Buizza, R.: Ensemble forecasting and the need for calibration, in:
Statistical Postprocessing of Ensemble Forecasts, edited by
Vannitsem, S., Wilks, D. S., and Messner, J. W., chap. 2, pp. 15–48,
Elsevier, Amsterdam, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Buizza et al.(2005)Buizza, Houtekamer, Toth, Pellerin, Wei, and
Zhu</label><?label buizza2005?><mixed-citation>Buizza, R., Houtekamer, P. L., Toth, Z., Pellerin, G., Wei, M., and Zhu, Y.: A
comparison of the ECMWF, MSC, and NCEP global ensemble prediction systems,
Mon. Weather Rev., 133, 1076–1097, <ext-link xlink:href="https://doi.org/10.1175/MWR2905.1" ext-link-type="DOI">10.1175/MWR2905.1</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bib1"><label>1</label><?label 1?><mixed-citation>DWD: COSMO-DE-EPS data, information available at: <uri>https://opendata.dwd.de</uri>, last access: 30 September 2020.</mixed-citation></ref>
      <ref id="bib1.bib2"><label>2</label><?label 1?><mixed-citation>ECMWF: TIGGE data, available at: <uri>https://confluence.ecmwf.int/display/TIGGE</uri>, last access: 30 September 2020.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Friederichs et al.(2018)Friederichs, Wahl, and
Buschow</label><?label friederichs2018?><mixed-citation>
Friederichs, P., Wahl, S., and Buschow, S.: Postprocessing for Extreme Events,
in: Statistical Postprocessing of Ensemble Forecasts, edited by
Vannitsem, S., Wilks, D. S., and Messner, J. W., chap. 5, pp. 128–154,
Elsevier, Amsterdam, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx8"><?xmltex \def\ref@label{{Gebhardt et~al.(2011)Gebhardt, Theis, Paulat, and
{Ben~Bouall\`{e}gue}}}?><label>Gebhardt et al.(2011)Gebhardt, Theis, Paulat, and
Ben Bouallègue</label><?label gebhardt2011?><mixed-citation>Gebhardt, C., Theis, S. E., Paulat, M., and Ben Bouallègue, Z.:
Uncertainties in COSMO-DE precipitation forecasts introduced by model
perturbations and variation of lateral boundaries, Atmos. Res., 100,
168–177, <ext-link xlink:href="https://doi.org/10.1016/j.atmosres.2010.12.008" ext-link-type="DOI">10.1016/j.atmosres.2010.12.008</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Glahn and Lowry(1972)</label><?label glahn1972?><mixed-citation>Glahn, H. R. and Lowry, D. A.: The use of model output statistics (MOS) in
objective weather forecasting, J. Appl. Meteorol., 11, 1203–1211, <ext-link xlink:href="https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2</ext-link>, 1972.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>Gneiting et al.(2005)Gneiting, Raftery, Westveld, and
Goldman</label><?label gneiting2005?><mixed-citation>Gneiting, T., Raftery, A. E., Westveld, A. H., and Goldman, T.: Calibrated
probabilistic forecasting using ensemble model output statistics and minimum
CRPS estimation, Mon. Weather Rev., 133, 1098–1118, <ext-link xlink:href="https://doi.org/10.1175/MWR2904.1" ext-link-type="DOI">10.1175/MWR2904.1</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Gneiting et al.(2007)Gneiting, Balabdaoui, and
Raftery</label><?label gneiting2007?><mixed-citation>Gneiting, T., Balabdaoui, F., and Raftery, A. E.: Probabilistic forecasts,
calibration and sharpness, J. R. Statist. Soc: B, 69, 243–268, <ext-link xlink:href="https://doi.org/10.1111/j.1467-9868.2007.00587.x" ext-link-type="DOI">10.1111/j.1467-9868.2007.00587.x</ext-link>, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Hamill(2012)</label><?label hamill2012?><mixed-citation>Hamill, T.: Verification of TIGGE multimodel and ECMWF reforecast-calibrated
probabilistic precipitation forecasts over the conterminous United States,
Mon. Weather Rev., 140, 2232–2252, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-11-00220.1" ext-link-type="DOI">10.1175/MWR-D-11-00220.1</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Hamill et al.(2017)Hamill, Engle, Myrick, Peroutka, Finan, and
Scheuerer</label><?label hamill2017?><mixed-citation>Hamill, T. M., Engle, E., Myrick, D., Peroutka, M., Finan, C., and Scheuerer,
M.: The U.S. national blend of models for statistical postprocessing of
probability of precipitation and deterministic precipitation amount, Mon.
Weather Rev., 145, 3441–3463, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-16-0331.1" ext-link-type="DOI">10.1175/MWR-D-16-0331.1</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Hersbach(2000)</label><?label hersbach2000?><mixed-citation>Hersbach, H.: Decomposition of the continuous ranked probability score for
ensemble prediction systems, Wea. Forecasting, 15, 559–570, <ext-link xlink:href="https://doi.org/10.1175/1520-0434(2000)015&lt;0559:DOTCRP&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0434(2000)015&lt;0559:DOTCRP&gt;2.0.CO;2</ext-link>, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Hess et al.(2015)Hess, Glashoff, and Reichert</label><?label hess2015?><mixed-citation>
Hess, R., Glashoff, J., and Reichert, B. K.: The Ensemble-MOS of Deutscher
Wetterdienst, in: EMS Annual Meeting Abstracts, 12, Sofia, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Hess et al.(2018)Hess, Kriesche, Schaumann, Reichert, and
Schmidt</label><?label hess2018?><mixed-citation>Hess, R., Kriesche, B., Schaumann, P., Reichert, B. K., and Schmidt, V.: Area
precipitation probabilities derived from point forecasts for operational
weather and warning service applications, Q. J. R. Meteorol. Soc., 144,
2392–2403, <ext-link xlink:href="https://doi.org/10.1002/qj.3306" ext-link-type="DOI">10.1002/qj.3306</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Hosmer et al.(2013)Hosmer, Lemenshow, and Sturdivant</label><?label hosmer2013?><mixed-citation>
Hosmer, D. W., Lemenshow, S., and Sturdivant, R. X.: Applied Logistic
Regression, Wiley Series in Probability and Statistics, Wiley, New Jersey,
3rd edn., 2013.</mixed-citation></ref>
      <?pagebreak page487?><ref id="bib1.bibx18"><?xmltex \def\ref@label{{Kn{\"{u}}pffer(1996)}}?><label>Knüpffer(1996)</label><?label knuepffer1996?><mixed-citation>
Knüpffer, K.: Methodical and predictability aspects of MOS systems, in:
Proceedings of the 13th Conference on Probability and Statistics in the
Atmospheric Sciences, pp. 190–197, San Francisco, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Lerch et al.(2017)Lerch, Thorarinsdottir, Ravazzolo, and
Gneiting</label><?label lerch2017?><mixed-citation>Lerch, S., Thorarinsdottir, T. L., Ravazzolo, F., and Gneiting, T.:
Forecaster's dilemma: Extreme events and forecast evaluation, Stat.  Sci., 32, 106–127, <ext-link xlink:href="https://doi.org/10.1214/16-STS588" ext-link-type="DOI">10.1214/16-STS588</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx20"><?xmltex \def\ref@label{{M{\"{o}}ller et~al.(2013)M{\"{o}}ller, Lenkoski, and
Thorarinsdottir}}?><label>Möller et al.(2013)Möller, Lenkoski, and
Thorarinsdottir</label><?label moller2013?><mixed-citation>Möller, A., Lenkoski, A., and Thorarinsdottir, T. L.: Multivariate
probabilistic forecasting using ensemble Bayesian model averaging and
copulas, Q. J. R. Meteorol. Soc., 139, 982–991, <ext-link xlink:href="https://doi.org/10.1002/qj.2009" ext-link-type="DOI">10.1002/qj.2009</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Numerical Algorithms Group(1990)</label><?label naglib?><mixed-citation>Numerical Algorithms Group: The NAG Library, Oxford, UK, <uri>https://www.nag.com/</uri> (last access: 30 September 2020),  1990.</mixed-citation></ref>
      <ref id="bib1.bibx22"><?xmltex \def\ref@label{{Peralta et~al.(2012)Peralta, {Ben~Bouall\`{e}gue}, Theis, Gebhardt,
and Buchhold}}?><label>Peralta et al.(2012)Peralta, Ben Bouallègue, Theis, Gebhardt,
and Buchhold</label><?label peralta2012?><mixed-citation>Peralta, C., Ben Bouallègue, Z., Theis, S. E., Gebhardt, C., and
Buchhold, M.: Accounting for initial condition uncertainties in
COSMO-DE-EPS, J. Geophys. Res.-Atmos., 117, D07108, <ext-link xlink:href="https://doi.org/10.1029/2011JD016581" ext-link-type="DOI">10.1029/2011JD016581</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>Primo(2016)</label><?label primo2016?><mixed-citation>Primo, C.: Wind gust warning verification, Adv. Sci. Res., 13, 113–120, <ext-link xlink:href="https://doi.org/10.5194/asr-13-113-2016" ext-link-type="DOI">10.5194/asr-13-113-2016</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Raftery et al.(2005)Raftery, Gneiting, Balabdaoui, and
Polakowski</label><?label raftery2005?><mixed-citation>Raftery, A. E., Gneiting, T., Balabdaoui, F., and Polakowski, M.: Using
Bayesian model averaging to calibrate forecast ensembles, Mon. Weather Rev.,
133, 1155–1174, <ext-link xlink:href="https://doi.org/10.1175/MWR2906.1" ext-link-type="DOI">10.1175/MWR2906.1</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Ranjan and Gneiting(2010)</label><?label ranjan2010?><mixed-citation>Ranjan, R. and Gneiting, T.: Combining probability forecasts, J.  Roy. Stat. Soc. B, 72, 71–91, <ext-link xlink:href="https://doi.org/10.1111/j.1467-9868.2009.00726.x" ext-link-type="DOI">10.1111/j.1467-9868.2009.00726.x</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Reichert(2016)</label><?label reichert2016?><mixed-citation>
Reichert, B. K.: The operational warning decision support system AutoWARN at
DWD, in: 27th Meeting of the European Working Group on Operational
Meteorological Workstation Systems (EGOWS), Helsinki, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>Reichert(2017)</label><?label reichert2017?><mixed-citation>
Reichert, B. K.: Forecasting and Nowcasting Severe Weather Using the
Operational Warning Decision Support System AutoWARN at DWD, in: 9th
Europ. Conf. on Severe Storms ECSS, Pula, Croatia, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx28"><?xmltex \def\ref@label{{Reichert et~al.(2015)Reichert, Glashoff, Hess, Hirsch, James,
Lenhart, Paller, Primo, Raatz, Schleinzer, and
Schr{\"{o}}der}}?><label>Reichert et al.(2015)Reichert, Glashoff, Hess, Hirsch, James,
Lenhart, Paller, Primo, Raatz, Schleinzer, and
Schröder</label><?label reichertetal2015?><mixed-citation>Reichert, B. K., Glashoff, J., Hess, R., Hirsch, T., James, P., Lenhart, C.,
Paller, J., Primo, C., Raatz, W., Schleinzer, T., and Schröder, G.: The
decision support system AutoWARN for the weather warning service at DWD,
in: EMS Annual Meeting Abstracts, 12, Sofia, 2015.
 </mixed-citation></ref><?xmltex \hack{\newpage}?>
      <ref id="bib1.bibx29"><?xmltex \def\ref@label{{Schefzik and M{\"{o}}ller(2018)}}?><label>Schefzik and Möller(2018)</label><?label schefzik2018?><mixed-citation>
Schefzik, R. and Möller, A.: Ensemble postprocesing methods incorporating
dependence structures, in: Statistical Postprocessing of Ensemble
Forecasts, edited by Vannitsem, S., Wilks, D. S., and Messner, J. W.,
chap. 4, pp. 91–125, Elsevier, Amsterdam, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Schefzik et al.(2013)Schefzik, Thorarinsdottir, and
Gneiting</label><?label schefzik2013?><mixed-citation>Schefzik, R., Thorarinsdottir, T. L., and Gneiting, T.: Uncertainty
quantification in complex simulation models using ensemble copula coupling,
Statist. Sci., 28, 616–640, <ext-link xlink:href="https://doi.org/10.1214/13-STS443" ext-link-type="DOI">10.1214/13-STS443</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Sloughter et al.(2013)Sloughter, Gneiting, and
Raftery</label><?label sloughter2013?><mixed-citation>Sloughter, J. M., Gneiting, T., and Raftery, A. E.: Probabilistic wind vector
forecasting using ensembles and Bayesian model averaging, Mon. Weather Rev.,
141, 2107–2119, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-12-00002.1" ext-link-type="DOI">10.1175/MWR-D-12-00002.1</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Swinbank et al.(2016)</label><?label swinbank2016?><mixed-citation>Swinbank, R., Kyouda, M., Buchanan, P., et al.: The TIGGE project and its achievements, Bull. Amer. Meteor. Soc., 97, 49–67, <ext-link xlink:href="https://doi.org/10.1175/BAMS-D-13-00191.1" ext-link-type="DOI">10.1175/BAMS-D-13-00191.1</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Vannitsem(2009)</label><?label vannitsem2009?><mixed-citation>Vannitsem, S.: A unified linear Model Output Statistics scheme for both
deterministic and ensemble forecasts, Q. J. R. Meteorol. Soc., 135,
1801–1815, <ext-link xlink:href="https://doi.org/10.1002/qj.491" ext-link-type="DOI">10.1002/qj.491</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Vannitsem et al.(2018)Vannitsem, Wilks, and Messner</label><?label vannitsem2018?><mixed-citation>
Vannitsem, S., Wilks, D. S., and Messner, J. W., eds.: Statistical
Postprocessing of Ensemble Forecasts, Elsevier, Amsterdam, Oxford,
Cambridge, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>Wilks(2001)</label><?label wilks2001?><mixed-citation>Wilks, D. S.: A skill score based on economic value for probability forecasts,
Meteorol. Appl., 8, 209–219, <ext-link xlink:href="https://doi.org/10.1017/S1350482701002092" ext-link-type="DOI">10.1017/S1350482701002092</ext-link>, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>Wilks(2011)</label><?label wilks2011?><mixed-citation>
Wilks, D. S.: Statistical Methods in the Atmospheric Sciences, Academic
Press, San Diego, 3rd edn., 2011.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Wilks(2018)</label><?label wilks2018?><mixed-citation>
Wilks, D. S.: Univariate ensemble postprocessing, in: Statistical
Postprocessing of Ensemble Forecasts, edited by Vannitsem, S., Wilks,
D. S., and Messner, J. W., chap. 3, pp. 49–89, Elsevier, Amsterdam, 2018.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Statistical postprocessing of ensemble forecasts for severe weather at Deutscher Wetterdienst</article-title-html>
<abstract-html><p>This paper gives an overview of Deutscher Wetterdienst's (DWD's) postprocessing system called
Ensemble-MOS together with its motivation and the design consequences
for probabilistic forecasts of extreme events based on ensemble data.
Forecasts of the ensemble systems COSMO-D2-EPS and ECMWF-ENS
are statistically optimised and calibrated by Ensemble-MOS
with a focus on severe weather in order to support the warning decision management at DWD.</p><p>Ensemble mean and spread are used as predictors for linear and logistic multiple regressions
to correct for conditional biases.
The predictands are derived from synoptic observations and include temperature, precipitation amounts, wind gusts and many more and are statistically estimated in a comprehensive model output statistics (MOS) approach.
Long time series and collections of stations are used as
training data that capture a sufficient number of observed events, as required for robust statistical modelling.</p><p>Logistic regressions are applied to probabilities that predefined meteorological events occur.
Details of the implementation including the selection of predictors with testing for significance are presented.
For probabilities of severe wind gusts global logistic parameterisations are developed that
depend on local estimations of wind speed. In this way, robust probability forecasts for extreme events are obtained
while local characteristics are preserved.</p><p>The problems of Ensemble-MOS, such as model changes and consistency requirements,
which occur with the operative MOS systems of the DWD are addressed.</p></abstract-html>
<ref-html id="bib1.bib1"><label>Baldauf et al.(2011)Baldauf, Seifert, Förstner, Majewski,
Raschendorfer, and Reinhardt</label><mixed-citation>
Baldauf, M., Seifert, A., Förstner, J., Majewski, D., Raschendorfer, M., and
Reinhardt, T.: Operational convective-scale numerical weather prediction  with the COSMO model: description and sensitivities, Mon. Weather Rev.,
139, 3887–3905, <a href="https://doi.org/10.1175/MWR-D-10-05013.1" target="_blank">https://doi.org/10.1175/MWR-D-10-05013.1</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Ben Bouallègue et al.(2015)Ben Bouallègue, Pinson, and
Friederichs</label><mixed-citation>
Ben Bouallègue, Z., Pinson, P., and Friederichs, P.: Quantile forecast
discrimination ability and value, Q. J. R. Meteorol. Soc., 141, 3415–3424, <a href="https://doi.org/10.1002/qj.2624" target="_blank">https://doi.org/10.1002/qj.2624</a>,  2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Bougeault et al.(2010)</label><mixed-citation>
Bougeault, P.,  Toth, Z.,  Bishop, C., et al.: The THORPEX interactive grand global ensemble, Bull. Amer. Meteor. Soc., 91, 1059–1072, <a href="https://doi.org/10.1175/2010BAMS2853.1" target="_blank">https://doi.org/10.1175/2010BAMS2853.1</a>, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Bröcker and Smith(2006)</label><mixed-citation>
Bröcker, J. and Smith, L. A.: Increasing the Reliability of Reliability
Diagrams, Weather Forecast., 22, 651–661, <a href="https://doi.org/10.1175/WAF993.1" target="_blank">https://doi.org/10.1175/WAF993.1</a> 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Buizza(2018)</label><mixed-citation>
Buizza, R.: Ensemble forecasting and the need for calibration, in:
Statistical Postprocessing of Ensemble Forecasts, edited by
Vannitsem, S., Wilks, D. S., and Messner, J. W., chap. 2, pp. 15–48,
Elsevier, Amsterdam, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Buizza et al.(2005)Buizza, Houtekamer, Toth, Pellerin, Wei, and
Zhu</label><mixed-citation>
Buizza, R., Houtekamer, P. L., Toth, Z., Pellerin, G., Wei, M., and Zhu, Y.: A
comparison of the ECMWF, MSC, and NCEP global ensemble prediction systems,
Mon. Weather Rev., 133, 1076–1097, <a href="https://doi.org/10.1175/MWR2905.1" target="_blank">https://doi.org/10.1175/MWR2905.1</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>1</label><mixed-citation>
DWD: COSMO-DE-EPS data, information available at: <a href="https://opendata.dwd.de" target="_blank"/>, last access: 30 September 2020.
</mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>2</label><mixed-citation>
ECMWF: TIGGE data, available at: <a href="https://confluence.ecmwf.int/display/TIGGE" target="_blank"/>, last access: 30 September 2020.
</mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Friederichs et al.(2018)Friederichs, Wahl, and
Buschow</label><mixed-citation>
Friederichs, P., Wahl, S., and Buschow, S.: Postprocessing for Extreme Events,
in: Statistical Postprocessing of Ensemble Forecasts, edited by
Vannitsem, S., Wilks, D. S., and Messner, J. W., chap. 5, pp. 128–154,
Elsevier, Amsterdam, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Gebhardt et al.(2011)Gebhardt, Theis, Paulat, and
Ben Bouallègue</label><mixed-citation>
Gebhardt, C., Theis, S. E., Paulat, M., and Ben Bouallègue, Z.:
Uncertainties in COSMO-DE precipitation forecasts introduced by model
perturbations and variation of lateral boundaries, Atmos. Res., 100,
168–177, <a href="https://doi.org/10.1016/j.atmosres.2010.12.008" target="_blank">https://doi.org/10.1016/j.atmosres.2010.12.008</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Glahn and Lowry(1972)</label><mixed-citation>
Glahn, H. R. and Lowry, D. A.: The use of model output statistics (MOS) in
objective weather forecasting, J. Appl. Meteorol., 11, 1203–1211, <a href="https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2</a>, 1972.
</mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Gneiting et al.(2005)Gneiting, Raftery, Westveld, and
Goldman</label><mixed-citation>
Gneiting, T., Raftery, A. E., Westveld, A. H., and Goldman, T.: Calibrated
probabilistic forecasting using ensemble model output statistics and minimum
CRPS estimation, Mon. Weather Rev., 133, 1098–1118, <a href="https://doi.org/10.1175/MWR2904.1" target="_blank">https://doi.org/10.1175/MWR2904.1</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Gneiting et al.(2007)Gneiting, Balabdaoui, and
Raftery</label><mixed-citation>
Gneiting, T., Balabdaoui, F., and Raftery, A. E.: Probabilistic forecasts,
calibration and sharpness, J. R. Statist. Soc: B, 69, 243–268, <a href="https://doi.org/10.1111/j.1467-9868.2007.00587.x" target="_blank">https://doi.org/10.1111/j.1467-9868.2007.00587.x</a>, 2007.
</mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Hamill(2012)</label><mixed-citation>
Hamill, T.: Verification of TIGGE multimodel and ECMWF reforecast-calibrated
probabilistic precipitation forecasts over the conterminous United States,
Mon. Weather Rev., 140, 2232–2252, <a href="https://doi.org/10.1175/MWR-D-11-00220.1" target="_blank">https://doi.org/10.1175/MWR-D-11-00220.1</a>, 2012.
</mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Hamill et al.(2017)Hamill, Engle, Myrick, Peroutka, Finan, and
Scheuerer</label><mixed-citation>
Hamill, T. M., Engle, E., Myrick, D., Peroutka, M., Finan, C., and Scheuerer,
M.: The U.S. national blend of models for statistical postprocessing of
probability of precipitation and deterministic precipitation amount, Mon.
Weather Rev., 145, 3441–3463, <a href="https://doi.org/10.1175/MWR-D-16-0331.1" target="_blank">https://doi.org/10.1175/MWR-D-16-0331.1</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Hersbach(2000)</label><mixed-citation>
Hersbach, H.: Decomposition of the continuous ranked probability score for
ensemble prediction systems, Wea. Forecasting, 15, 559–570, <a href="https://doi.org/10.1175/1520-0434(2000)015&lt;0559:DOTCRP&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0434(2000)015&lt;0559:DOTCRP&gt;2.0.CO;2</a>, 2000.
</mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Hess et al.(2015)Hess, Glashoff, and Reichert</label><mixed-citation>
Hess, R., Glashoff, J., and Reichert, B. K.: The Ensemble-MOS of Deutscher
Wetterdienst, in: EMS Annual Meeting Abstracts, 12, Sofia, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Hess et al.(2018)Hess, Kriesche, Schaumann, Reichert, and
Schmidt</label><mixed-citation>
Hess, R., Kriesche, B., Schaumann, P., Reichert, B. K., and Schmidt, V.: Area
precipitation probabilities derived from point forecasts for operational
weather and warning service applications, Q. J. R. Meteorol. Soc., 144,
2392–2403, <a href="https://doi.org/10.1002/qj.3306" target="_blank">https://doi.org/10.1002/qj.3306</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Hosmer et al.(2013)Hosmer, Lemenshow, and Sturdivant</label><mixed-citation>
Hosmer, D. W., Lemenshow, S., and Sturdivant, R. X.: Applied Logistic
Regression, Wiley Series in Probability and Statistics, Wiley, New Jersey,
3rd edn., 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Knüpffer(1996)</label><mixed-citation>
Knüpffer, K.: Methodical and predictability aspects of MOS systems, in:
Proceedings of the 13th Conference on Probability and Statistics in the
Atmospheric Sciences, pp. 190–197, San Francisco, 1996.
</mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Lerch et al.(2017)Lerch, Thorarinsdottir, Ravazzolo, and
Gneiting</label><mixed-citation>
Lerch, S., Thorarinsdottir, T. L., Ravazzolo, F., and Gneiting, T.:
Forecaster's dilemma: Extreme events and forecast evaluation, Stat.  Sci., 32, 106–127, <a href="https://doi.org/10.1214/16-STS588" target="_blank">https://doi.org/10.1214/16-STS588</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Möller et al.(2013)Möller, Lenkoski, and
Thorarinsdottir</label><mixed-citation>
Möller, A., Lenkoski, A., and Thorarinsdottir, T. L.: Multivariate
probabilistic forecasting using ensemble Bayesian model averaging and
copulas, Q. J. R. Meteorol. Soc., 139, 982–991, <a href="https://doi.org/10.1002/qj.2009" target="_blank">https://doi.org/10.1002/qj.2009</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Numerical Algorithms Group(1990)</label><mixed-citation>
Numerical Algorithms Group: The NAG Library, Oxford, UK, <a href="https://www.nag.com/" target="_blank"/> (last access: 30 September 2020),  1990.
</mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Peralta et al.(2012)Peralta, Ben Bouallègue, Theis, Gebhardt,
and Buchhold</label><mixed-citation>
Peralta, C., Ben Bouallègue, Z., Theis, S. E., Gebhardt, C., and
Buchhold, M.: Accounting for initial condition uncertainties in
COSMO-DE-EPS, J. Geophys. Res.-Atmos., 117, D07108, <a href="https://doi.org/10.1029/2011JD016581" target="_blank">https://doi.org/10.1029/2011JD016581</a>, 2012.
</mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Primo(2016)</label><mixed-citation>
Primo, C.: Wind gust warning verification, Adv. Sci. Res., 13, 113–120, <a href="https://doi.org/10.5194/asr-13-113-2016" target="_blank">https://doi.org/10.5194/asr-13-113-2016</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Raftery et al.(2005)Raftery, Gneiting, Balabdaoui, and
Polakowski</label><mixed-citation>
Raftery, A. E., Gneiting, T., Balabdaoui, F., and Polakowski, M.: Using
Bayesian model averaging to calibrate forecast ensembles, Mon. Weather Rev.,
133, 1155–1174, <a href="https://doi.org/10.1175/MWR2906.1" target="_blank">https://doi.org/10.1175/MWR2906.1</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Ranjan and Gneiting(2010)</label><mixed-citation>
Ranjan, R. and Gneiting, T.: Combining probability forecasts, J.  Roy. Stat. Soc. B, 72, 71–91, <a href="https://doi.org/10.1111/j.1467-9868.2009.00726.x" target="_blank">https://doi.org/10.1111/j.1467-9868.2009.00726.x</a>, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Reichert(2016)</label><mixed-citation>
Reichert, B. K.: The operational warning decision support system AutoWARN at
DWD, in: 27th Meeting of the European Working Group on Operational
Meteorological Workstation Systems (EGOWS), Helsinki, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Reichert(2017)</label><mixed-citation>
Reichert, B. K.: Forecasting and Nowcasting Severe Weather Using the
Operational Warning Decision Support System AutoWARN at DWD, in: 9th
Europ. Conf. on Severe Storms ECSS, Pula, Croatia, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Reichert et al.(2015)Reichert, Glashoff, Hess, Hirsch, James,
Lenhart, Paller, Primo, Raatz, Schleinzer, and
Schröder</label><mixed-citation>
Reichert, B. K., Glashoff, J., Hess, R., Hirsch, T., James, P., Lenhart, C.,
Paller, J., Primo, C., Raatz, W., Schleinzer, T., and Schröder, G.: The
decision support system AutoWARN for the weather warning service at DWD,
in: EMS Annual Meeting Abstracts, 12, Sofia, 2015.

</mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Schefzik and Möller(2018)</label><mixed-citation>
Schefzik, R. and Möller, A.: Ensemble postprocesing methods incorporating
dependence structures, in: Statistical Postprocessing of Ensemble
Forecasts, edited by Vannitsem, S., Wilks, D. S., and Messner, J. W.,
chap. 4, pp. 91–125, Elsevier, Amsterdam, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Schefzik et al.(2013)Schefzik, Thorarinsdottir, and
Gneiting</label><mixed-citation>
Schefzik, R., Thorarinsdottir, T. L., and Gneiting, T.: Uncertainty
quantification in complex simulation models using ensemble copula coupling,
Statist. Sci., 28, 616–640, <a href="https://doi.org/10.1214/13-STS443" target="_blank">https://doi.org/10.1214/13-STS443</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Sloughter et al.(2013)Sloughter, Gneiting, and
Raftery</label><mixed-citation>
Sloughter, J. M., Gneiting, T., and Raftery, A. E.: Probabilistic wind vector
forecasting using ensembles and Bayesian model averaging, Mon. Weather Rev.,
141, 2107–2119, <a href="https://doi.org/10.1175/MWR-D-12-00002.1" target="_blank">https://doi.org/10.1175/MWR-D-12-00002.1</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Swinbank et al.(2016)</label><mixed-citation>
Swinbank, R., Kyouda, M., Buchanan, P., et al.: The TIGGE project and its achievements, Bull. Amer. Meteor. Soc., 97, 49–67, <a href="https://doi.org/10.1175/BAMS-D-13-00191.1" target="_blank">https://doi.org/10.1175/BAMS-D-13-00191.1</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Vannitsem(2009)</label><mixed-citation>
Vannitsem, S.: A unified linear Model Output Statistics scheme for both
deterministic and ensemble forecasts, Q. J. R. Meteorol. Soc., 135,
1801–1815, <a href="https://doi.org/10.1002/qj.491" target="_blank">https://doi.org/10.1002/qj.491</a>, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Vannitsem et al.(2018)Vannitsem, Wilks, and Messner</label><mixed-citation>
Vannitsem, S., Wilks, D. S., and Messner, J. W., eds.: Statistical
Postprocessing of Ensemble Forecasts, Elsevier, Amsterdam, Oxford,
Cambridge, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Wilks(2001)</label><mixed-citation>
Wilks, D. S.: A skill score based on economic value for probability forecasts,
Meteorol. Appl., 8, 209–219, <a href="https://doi.org/10.1017/S1350482701002092" target="_blank">https://doi.org/10.1017/S1350482701002092</a>, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Wilks(2011)</label><mixed-citation>
Wilks, D. S.: Statistical Methods in the Atmospheric Sciences, Academic
Press, San Diego, 3rd edn., 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Wilks(2018)</label><mixed-citation>
Wilks, D. S.: Univariate ensemble postprocessing, in: Statistical
Postprocessing of Ensemble Forecasts, edited by Vannitsem, S., Wilks,
D. S., and Messner, J. W., chap. 3, pp. 49–89, Elsevier, Amsterdam, 2018.
</mixed-citation></ref-html>--></article>
