<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">NPG</journal-id><journal-title-group>
    <journal-title>Nonlinear Processes in Geophysics</journal-title>
    <abbrev-journal-title abbrev-type="publisher">NPG</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Nonlin. Processes Geophys.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1607-7946</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/npg-26-339-2019</article-id><title-group><article-title>Statistical post-processing of ensemble forecasts of the height<?xmltex \hack{\break}?> of new snow</article-title><alt-title>Statistical post-processing of ensemble forecasts of the height of new snow</alt-title>
      </title-group><?xmltex \runningtitle{Statistical post-processing of ensemble forecasts of the height of new snow}?><?xmltex \runningauthor{J.-P. Nousu et al.}?>
      <contrib-group>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff2">
          <name><surname>Nousu</surname><given-names>Jari-Pekka</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="yes" rid="aff1">
          <name><surname>Lafaysse</surname><given-names>Matthieu</given-names></name>
          <email>matthieu.lafaysse@meteo.fr</email>
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Vernay</surname><given-names>Matthieu</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff3 aff4">
          <name><surname>Bellier</surname><given-names>Joseph</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff5">
          <name><surname>Evin</surname><given-names>Guillaume</given-names></name>
          
        <ext-link>https://orcid.org/0000-0003-3456-9441</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff6">
          <name><surname>Joly</surname><given-names>Bruno</given-names></name>
          
        </contrib>
        <aff id="aff1"><label>1</label><institution>Univ. Grenoble Alpes – Université de Toulouse – Météo-France – CNRS – CNRM, Centre d'Etudes de la Neige,<?xmltex \hack{\break}?> Grenoble, France</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>University of Oulu, Water, Energy and Environmental Engineering Research Unit, Oulu, Finland</institution>
        </aff>
        <aff id="aff3"><label>3</label><institution>Cooperative Institute for Research in Environmental Sciences, University of Colorado Boulder, and NOAA Earth System Research Laboratory, Physical Sciences Division, Boulder, Colorado, USA</institution>
        </aff>
        <aff id="aff4"><label>4</label><institution>Univ. Grenoble Alpes, CNRS, IRD, Grenoble INP, IGE, Grenoble, France</institution>
        </aff>
        <aff id="aff5"><label>5</label><institution>Univ. Grenoble Alpes – IRSTEA, UR ETNA, Grenoble, France</institution>
        </aff>
        <aff id="aff6"><label>6</label><institution>CNRM – Université de Toulouse – Météo-France – CNRS, GMAP, Toulouse, France</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Matthieu Lafaysse (matthieu.lafaysse@meteo.fr)</corresp></author-notes><pub-date><day>26</day><month>September</month><year>2019</year></pub-date>
      
      <volume>26</volume>
      <issue>3</issue>
      <fpage>339</fpage><lpage>357</lpage>
      <history>
        <date date-type="received"><day>16</day><month>May</month><year>2019</year></date>
           <date date-type="rev-request"><day>5</day><month>June</month><year>2019</year></date>
           <date date-type="rev-recd"><day>29</day><month>August</month><year>2019</year></date>
           <date date-type="accepted"><day>2</day><month>September</month><year>2019</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2019 Jari-Pekka Nousu et al.</copyright-statement>
        <copyright-year>2019</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019.html">This article is available from https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019.html</self-uri><self-uri xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019.pdf">The full text article is available as a PDF file from https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019.pdf</self-uri>
      <abstract><title>Abstract</title>
    <p id="d1e164">Forecasting the height of new snow (HN) is crucial for avalanche hazard forecasting, road viability, ski resort management and tourism attractiveness. Météo-France operates the PEARP-S2M probabilistic forecasting system, including 35 members of the PEARP Numerical Weather Prediction system, where the SAFRAN downscaling tool refines the elevation resolution and the Crocus snowpack model represents the main physical processes in the snowpack. It provides better HN forecasts than direct NWP diagnostics but exhibits significant biases and underdispersion. We applied a statistical post-processing to these ensemble forecasts, based on non-homogeneous regression with a censored shifted Gamma distribution. Observations come from manual measurements of 24 h HN in the French Alps and Pyrenees. The calibration is tested at the station scale and the massif scale (i.e. aggregating different stations over areas of 1000 km<inline-formula><mml:math id="M1" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula>). Compared to the raw forecasts, similar improvements are obtained for both spatial scales. Therefore, the post-processing can be applied at any point of the massifs. Two training datasets are tested: (1) a 22-year homogeneous reforecast for which the NWP model resolution and physical options are identical to the operational system but without the same initial perturbations; (2) 3-year real-time forecasts with a heterogeneous model configuration but the same perturbation methods.   The impact of the training dataset depends on lead time and on the evaluation criteria. The long-term reforecast improves the reliability of severe snowfall but leads to overdispersion due to the discrepancy in real-time perturbations. Thus, the development of reliable automatic forecasting products of HN needs long reforecasts as homogeneous as possible with the operational systems.</p>
  </abstract>
    </article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d1e185">Forecasting the height of new snow <xref ref-type="bibr" rid="bib1.bibx23" id="paren.1"><named-content content-type="pre">HN,</named-content></xref> is essential in the mountainous areas as well as in the northern regions due to various safety issues and economic activities. For instance, avalanche hazard forecasting, road viability, ski resort management and tourism attractiveness rely on the forecasts of HN. Automatic predictions are increasingly developed for that purpose, based on numerical weather prediction (NWP) model output. Nevertheless, accurate forecasting of this variable is still challenging for several reasons. First, the precipitation forecasts in NWP models have significant errors which increase with longer lead times. These forecast uncertainties have to be considered. Second, the high variability of HN as a function of elevation is difficult to describe in mountainous areas, even at the best spatial resolution available in NWP models (i.e. 1 km or a few kilometres). Finally, several processes, such as density of falling snow, mechanical compaction during the deposition and variations of the rain–snow limit elevation during some storm events, are not or are poorly represented in NWP models. Several recent scientific advances can help to face these challenges.
<list list-type="bullet"><list-item>
      <p id="d1e195">To estimate forecast uncertainty, ensemble forecasting has become an important method in NWP. Probabilistic forecasts have been in operational use for a number of years in several meteorological centres  <xref ref-type="bibr" rid="bib1.bibx45 bib1.bibx66 bib1.bibx49" id="paren.2"/>. Ensemble forecasting has increased the confidence of forecast users in predicting possible future occurrence, or non-occurrence, of unusually strong events <xref ref-type="bibr" rid="bib1.bibx12" id="paren.3"/>. In many cases, an estimation of the probability density of future weather-related variables may present more value for the forecast user than a single deterministic forecast does <xref ref-type="bibr" rid="bib1.bibx52 bib1.bibx51" id="paren.4"/>. The forecast uncertainties depend on the atmospheric flow and vary from day to day <xref ref-type="bibr" rid="bib1.bibx42" id="paren.5"/>. Therefore, ensemble forecasting aims at estimating the probability density of the future state of the atmosphere.</p></list-item><list-item>
      <p id="d1e211">In NWP, snowpack modelling is necessary since the presence of snow on the ground has a major impact on all the fluxes taking place at the interface between the Earth's atmosphere and its surface. However, NWP models often use single-layer snow schemes with homogeneous physical properties because they are relatively inexpensive, have relatively few parameters and capture first-order processes <xref ref-type="bibr" rid="bib1.bibx17" id="paren.6"/>. Models with more complexity have also been developed but are not yet implemented in most NWP systems. The most detailed ones are able to represent a detailed stratigraphy of the snowpack with an explicit description of the time evolution of the snow microstructure <xref ref-type="bibr" rid="bib1.bibx40 bib1.bibx69" id="paren.7"/>. Snow model intercomparison projects <xref ref-type="bibr" rid="bib1.bibx38" id="paren.8"/> suggest that detailed snowpack models are among the most accurate models in the reproduction of the snowpack evolution in various climates and environments. Operationally, these snow models are sometimes forced by NWP outputs to forecast the risk of avalanche <xref ref-type="bibr" rid="bib1.bibx21" id="paren.9"/>. Concerning the topic of this study, it is known that these models also provide better estimates of the height of new snow than direct NWP outputs <xref ref-type="bibr" rid="bib1.bibx14" id="paren.10"/>. This is explained by the ability of these schemes to simulate the mechanical compaction of snow on the ground occurring during the snowfall, the possible impact of changes in precipitation phase during a storm event, the possible occurrence of melting at the surface or at the bottom of the snowpack and the dependence of falling snow density on meteorological conditions.</p></list-item></list></p>
      <p id="d1e229">To benefit from both the advantages of ensemble NWP and detailed snowpack modelling, <xref ref-type="bibr" rid="bib1.bibx68" id="text.11"/> developed the PEARP-S2M modelling system (PEARP: Prévision d'Ensemble ARPEGE; ARPEGE: Action de Recherche Petite Echelle Grande Echelle; S2M: SAFRAN-SURFEX-MEPRA; SAFRAN: Système Atmosphérique Fournissant des Renseignements Atmosphériques à la Neige; SURFEX: SURFace EXternalisée; MEPRA: Modèle Expert pour la Prévision du Risque d'Avalanches). In this system, the Crocus detailed snowpack model <xref ref-type="bibr" rid="bib1.bibx69" id="paren.12"/> implemented in the SURFEX surface modelling platform is forced by the ensemble version of the ARPEGE NWP model <xref ref-type="bibr" rid="bib1.bibx16" id="paren.13"/> after an elevation adjustment of the meteorological fields by the SAFRAN downscaling tool <xref ref-type="bibr" rid="bib1.bibx20" id="paren.14"/>. However, the PEARP-S2M system still suffers from various biases and deficiencies <xref ref-type="bibr" rid="bib1.bibx68 bib1.bibx14" id="paren.15"/>. Biases in atmospheric ensemble forecasts may be caused by insufficient model resolutions <xref ref-type="bibr" rid="bib1.bibx71 bib1.bibx46 bib1.bibx61 bib1.bibx11" id="paren.16"/>, suboptimal physical parameterizations <xref ref-type="bibr" rid="bib1.bibx48 bib1.bibx72" id="paren.17"/> or suboptimal methods for generating the initial conditions <xref ref-type="bibr" rid="bib1.bibx4 bib1.bibx5 bib1.bibx31 bib1.bibx32 bib1.bibx60" id="paren.18"/>. In the case of HN forecasts, the errors also originate from the snow models <xref ref-type="bibr" rid="bib1.bibx22 bib1.bibx39" id="paren.19"/>. Due to the systematic biases in ensemble forecasts and the challenge of detecting and correcting their origins, many methods of statistical post-processing have been developed that leverage archives of past forecast errors <xref ref-type="bibr" rid="bib1.bibx67" id="paren.20"/>. In the literature, these probabilistic post-processing methods are often referred to as ensemble model output statistics (EMOS) as an extension to ensemble approaches of the traditional model output statistics (MOS) applied for several decades to deterministic forecasts <xref ref-type="bibr" rid="bib1.bibx27" id="paren.21"/>. EMOS are now routinely applied for meteorological predictands such as temperature, precipitation and wind speed.   The techniques are for instance non-homogeneous regression methods <xref ref-type="bibr" rid="bib1.bibx36 bib1.bibx28 bib1.bibx73 bib1.bibx65 bib1.bibx41 bib1.bibx54 bib1.bibx55 bib1.bibx65 bib1.bibx3 bib1.bibx25" id="paren.22"/>, logistic regression methods <xref ref-type="bibr" rid="bib1.bibx33 bib1.bibx34 bib1.bibx44" id="paren.23"/>, Bayesian model averaging <xref ref-type="bibr" rid="bib1.bibx50" id="paren.24"/>, rank histogram recalibration <xref ref-type="bibr" rid="bib1.bibx30" id="paren.25"/>, ensemble dressing approaches (i.e. kernel density) <xref ref-type="bibr" rid="bib1.bibx53 bib1.bibx70 bib1.bibx24" id="paren.26"/>, and quantile regression forests <xref ref-type="bibr" rid="bib1.bibx62 bib1.bibx63" id="paren.27"/>.</p>
      <p id="d1e285">However, statistical post-processing of ensemble HN forecasts is rarely reviewed in the literature. <xref ref-type="bibr" rid="bib1.bibx59" id="text.28"/> and <xref ref-type="bibr" rid="bib1.bibx57" id="text.29"/> are the first studies to the best of our knowledge to present post-processed ensemble forecasts of HN. However, they only considered direct ensemble NWP output as predictors (precipitation and temperature) and did not incorporate physical modelling of the snowpack. It can be expected that physical modelling could<?pagebreak page341?> capture some complex features explaining the variability of HN. This variability is difficult to reach by multivariate statistical relationships, especially the common high temporal variations of temperature and precipitation intensity during a storm event with highly non-linear impacts on the height of new snow. Furthermore, because they do not consider direct predictors of HN, these recent studies partly rely on precipitation observations in their calibration procedure, whereas solid precipitation is particularly prone to very high measurement errors <xref ref-type="bibr" rid="bib1.bibx37" id="paren.30"/>. The physical simulation of HN enables observations of this variable for the post-processing to be considered directly. This is a major advantage because HN measurement errors <xref ref-type="bibr" rid="bib1.bibx75" id="paren.31"><named-content content-type="pre">typically 0.5 cm,</named-content></xref> are considerably lower than errors in solid precipitation measurements.</p>
      <p id="d1e302">The goal of this study is to test the ability of a non-homogeneous regression method to improve the ensemble forecasts of HN from the PEARP-S2M ensemble snowpack modelling system. More precisely, the regression method of <xref ref-type="bibr" rid="bib1.bibx55" id="text.32"/> based on the censored shifted Gamma distribution was chosen in this work for the advantages identified by the authors in the case of precipitation forecasts. In particular, this method allows one to extrapolate the statistical relationship between predictors and predictands from common events to more unusual events. Considering the specificities of the available datasets in terms of predictands and predictors, two other scientific questions are considered: (1) can statistical post-processing be applied at a larger spatial scale than the observation points? (2) What are the requirements of a robust training forecast dataset for statistical post-processing?</p>
      <p id="d1e309">The structure of the paper is as follows. Section <xref ref-type="sec" rid="Ch1.S2"/> describes the model components of the PEARP-S2M system, the observation and forecast datasets used in this study, the non-homogeneous regression method chosen for post-processing and the evaluation metrics. In Sect. <xref ref-type="sec" rid="Ch1.S3"/>, the results of the post-processing method are presented for different training configurations. The discussion in Sect. <xref ref-type="sec" rid="Ch1.S4"/> focuses on the implications of our study for the possibility of implementing such post-processing in operational automatic forecast products and recommendations for improvements.</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Data and methods</title>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Models</title>
<sec id="Ch1.S2.SS1.SSS1">
  <label>2.1.1</label><title>PEARP ensemble NWP system</title>
      <p id="d1e340">PEARP is a short-range ensemble prediction system operated by Météo-France up to 4.5 d, fully described in <xref ref-type="bibr" rid="bib1.bibx16" id="text.33"/>. It includes 35 forecast members of the ARPEGE NWP model. In 2019, it is based on a 25-member ensemble assimilation combined with the singular vector perturbation methods <xref ref-type="bibr" rid="bib1.bibx10 bib1.bibx45" id="paren.34"/> to provide 35 initial states. The singular vector perturbations are designed to optimize the spread of the large-scale atmospheric fields at a 24 h lead time. Finally, the 35 members are randomly associated with 10 different sets of physical parameterizations, including different deep and shallow convection schemes, among others. The current horizontal resolution is about 10 km over France (truncature T798C2.4) with 90 atmospheric levels. All these features have been improved over time with a new operational configuration provided almost every year.</p>
</sec>
<sec id="Ch1.S2.SS1.SSS2">
  <label>2.1.2</label><title>SAFRAN downscaling tool</title>
      <p id="d1e357">SAFRAN <xref ref-type="bibr" rid="bib1.bibx18 bib1.bibx20" id="paren.35"/> is a downscaling and surface analysis tool specifically designed to provide meteorological fields in mountainous areas (i.e. with high-elevation gradients). The principle of SAFRAN is to perform a spatialization of the available weather data in mountain ranges with so-called “massifs” of about 1000 km<inline-formula><mml:math id="M2" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula> where meteorological conditions are assumed to depend only on altitude. SAFRAN variables include precipitation (rainfall and snowfall rate), air temperature, relative humidity, wind speed as well as incoming longwave and shortwave radiations. Although SAFRAN was initially designed to work as an analysis system adjusting a guess from NWP outputs with the available meteorological observations, SAFRAN also comes with a forecast mode which can be considered simply to be a downscaling tool to convert an NWP model grid (PEARP in our case) to the massif geometry. The originality of this system is the use of different vertical levels of the NWP model to obtain surface fields at different elevations.</p>
</sec>
<sec id="Ch1.S2.SS1.SSS3">
  <label>2.1.3</label><title>Crocus snowpack model</title>
      <p id="d1e380">Crocus <xref ref-type="bibr" rid="bib1.bibx69" id="paren.36"/> is a one-dimensional multilayer physical snow scheme which simulates the evolution of the snow cover affected by both the atmosphere and the ground below. It is implemented in the SURFEX surface modelling platform <xref ref-type="bibr" rid="bib1.bibx43" id="paren.37"/> as the most detailed snow scheme of the ISBA (Interactions between the Soil Biosphere and Atmosphere) land surface model. Each snow layer is described by its mass, density, enthalpy (temperature and liquid water content) and age. The evolution of snow grains is described with additional variables (optical diameter and sphericity) using metamorphism laws from <xref ref-type="bibr" rid="bib1.bibx9" id="text.38"/> and <xref ref-type="bibr" rid="bib1.bibx13" id="text.39"/>. Snow density is a particularly important property for the height of new snow. It is mainly affected by two key processes: the density of falling snow and the compaction of snow on the ground. Falling snow density was empirically parameterized as a function of air temperature and wind speed <xref ref-type="bibr" rid="bib1.bibx47" id="paren.40"/>. This parameterization is associated with significant uncertainties <xref ref-type="bibr" rid="bib1.bibx39 bib1.bibx35" id="paren.41"/>. Snow compaction is modelled with a visco-elastic scheme in which the snow viscosity of each layer is parameterized depending mainly on the<?pagebreak page342?> layer density and temperature. The parameterization of snow viscosity is also uncertain as various expressions were formulated in the literature <xref ref-type="bibr" rid="bib1.bibx64" id="paren.42"/>. Furthermore, the compaction velocity actually has a high dependence on snow microstructure <xref ref-type="bibr" rid="bib1.bibx40" id="paren.43"/>. This complex dependence cannot be described in Crocus by the visco-elastic concept and microstructure-dependent models of compaction are only available for very specific conditions <xref ref-type="bibr" rid="bib1.bibx58" id="paren.44"/>. These limitations partly explain the errors of simulated HN identified by <xref ref-type="bibr" rid="bib1.bibx14" id="text.45"/> in combination with the known errors and underdispersion of precipitation input.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F1" specific-use="star"><?xmltex \currentcnt{1}?><label>Figure 1</label><caption><p id="d1e416">Map of massifs in the French Alps <bold>(a)</bold> and Pyrenees <bold>(b)</bold>, with altitude in metres and the observation stations represented as white dots.</p></caption>
            <?xmltex \igopts{width=426.791339pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f01.png"/>

          </fig>

</sec>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>Data</title>
<sec id="Ch1.S2.SS2.SSS1">
  <label>2.2.1</label><title>Study area</title>
      <p id="d1e447">The study area covers the French Alps and Pyrenees. In all operational productions of avalanche hazard forecasting, these regions are divided respectively into 23 and 11 massifs (Fig. <xref ref-type="fig" rid="Ch1.F1"/>), identical to the ones used in SAFRAN discretization (Sect. <xref ref-type="sec" rid="Ch1.S2.SS1.SSS2"/>). The climate is contrasted, colder and wetter in the northern Alps and much drier in the southern Alps and eastern Pyrenees due to the Mediterranean influence <xref ref-type="bibr" rid="bib1.bibx19" id="paren.46"/>. White dots correspond to stations where daily meteorological and snow observations are available in winter in the so-called nivo-métérologique observation network.</p>
</sec>
<sec id="Ch1.S2.SS2.SSS2">
  <label>2.2.2</label><title>Predictors</title>
      <p id="d1e465">Two separate sets of training data for the statistical post-processing were used in this study as predictors. Their specifications detailed below are also summarized in Table <xref ref-type="table" rid="Ch1.T1"/>.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T1" specific-use="star"><?xmltex \currentcnt{1}?><label>Table 1</label><caption><p id="d1e473">Summary of the predictor dataset used for training and evaluation.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="7">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="left"/>
     <oasis:colspec colnum="5" colname="col5" align="left"/>
     <oasis:colspec colnum="6" colname="col6" align="left"/>
     <oasis:colspec colnum="7" colname="col7" align="left"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Predictor</oasis:entry>
         <oasis:entry colname="col2">Use</oasis:entry>
         <oasis:entry colname="col3">Members</oasis:entry>
         <oasis:entry colname="col4">Perturbation of</oasis:entry>
         <oasis:entry colname="col5">Model physics</oasis:entry>
         <oasis:entry colname="col6">Snow simulations geometry</oasis:entry>
         <oasis:entry colname="col7">Period</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4">initial conditions</oasis:entry>
         <oasis:entry colname="col5">and resolution</oasis:entry>
         <oasis:entry colname="col6"/>
         <oasis:entry colname="col7"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">HN reforecasts</oasis:entry>
         <oasis:entry colname="col2">Training</oasis:entry>
         <oasis:entry colname="col3">10</oasis:entry>
         <oasis:entry colname="col4">no</oasis:entry>
         <oasis:entry colname="col5">constant</oasis:entry>
         <oasis:entry colname="col6">stations (Alps and Pyrenees)</oasis:entry>
         <oasis:entry colname="col7">1994–2016</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">HN real-time forecasts</oasis:entry>
         <oasis:entry colname="col2">Training</oasis:entry>
         <oasis:entry colname="col3">35</oasis:entry>
         <oasis:entry colname="col4">yes</oasis:entry>
         <oasis:entry colname="col5">variable</oasis:entry>
         <oasis:entry colname="col6">300 m elevation bands (Alps)</oasis:entry>
         <oasis:entry colname="col7">2014–2017</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">HN real-time forecasts</oasis:entry>
         <oasis:entry colname="col2">Evaluation</oasis:entry>
         <oasis:entry colname="col3">35</oasis:entry>
         <oasis:entry colname="col4">yes</oasis:entry>
         <oasis:entry colname="col5">constant</oasis:entry>
         <oasis:entry colname="col6">stations (Alps and Pyrenees)</oasis:entry>
         <oasis:entry colname="col7">2017–2018</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S2.SS2.SSSx1" specific-use="unnumbered">
  <title>Reforecasts used for training</title>
      <p id="d1e628">First, the PEARP reforecasts consist of 10 members, including 1 control member, and were issued in 2018. The reforecasts <xref ref-type="bibr" rid="bib1.bibx7" id="paren.47"/> are based on a homogeneous model configuration identical to the operational release of 5 December 2017 (same resolution and physical parameterizations), but they only include physical perturbations and no perturbation of the initial state, contrary to operational PEARP forecasts. The initial states are built with ERA-Interim reanalysis <xref ref-type="bibr" rid="bib1.bibx15" id="paren.48"/> for the atmospheric variables and by the 24 h stand-alone coupled forecasts of the SURFEX/ARPEGE model for the Earth parameters. These reforecasts were downscaled with SAFRAN for all stations in the French Alps and Pyrenees where snow observations are available. These downscaled forecasts were used to force the Crocus snowpack model to provide simulated heights of new snow. The training period length is 22 seasons (from 1994 till 2016).</p><?xmltex \hack{\newpage}?>
</sec>
<sec id="Ch1.S2.SS2.SSSx2" specific-use="unnumbered">
  <title>Real-time forecasts used for training</title>
      <p id="d1e644">Second, the real-time forecasts of PEARP consist of 35 members, including a control member. In contrast to the reforecasts, model configuration has changed over time and the earlier versions were different from the currently operational version (lower horizontal and vertical resolution, different set of model physics). Both physical perturbations and initial state perturbations are included in the real-time forecasts. These forecasts have experimentally forced the S2M snowpack modelling chain in real time since 2014. However, these real-time snow forecasts were only issued for the French Alps massifs at specific elevations of 1200, 1500, 1800, 2100, 2400 and 2700 m. They are used for training over the 2014–2017 period.</p>
</sec>
<sec id="Ch1.S2.SS2.SSSx3" specific-use="unnumbered">
  <title>Real-time forecasts used for verification</title>
      <p id="d1e654">Statistical methods have to be evaluated on datasets independent of the ones used for the calibration. Hence, the 35-member real-time forecasts of PEARP-S2M covering the 2017–2018 winter were used as predictors for verification (last line of Table <xref ref-type="table" rid="Ch1.T1"/>). The version of PEARP is homogeneous over the verification period and identical to the reforecast configuration (resolution and physics), but it also accounts for the initial perturbations. Thus, there are more members than in the reforecasts. These verification forecasts were downscaled and forced the Crocus snowpack model for all the stations available in the snow reforecasts. These verification forecasts were used to evaluate all the training scenarios described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS5"/>.</p>
</sec>
<sec id="Ch1.S2.SS2.SSSx4" specific-use="unnumbered">
  <title>Common features</title>
      <p id="d1e667">Note that in all cases (snow reforecasts and real-time snow forecasts used for training and verification), each snow forecast is initialized by a SURFEX/ISBA-Crocus run forced by SAFRAN analysis (assimilating meteorological observations) from the beginning of the season. For each season, the snow reforecasts and real-time forecasts were issued only for months from November to April. Months outside of this time window were neglected due to insufficient observation data. In this study, we consider only HN snow reforecasts and real-time forecasts at four different lead times (<inline-formula><mml:math id="M3" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">24</mml:mn></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">48</mml:mn></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M5" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">72</mml:mn></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M6" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">96</mml:mn></mml:mrow></mml:math></inline-formula> h).</p>
</sec>
<sec id="Ch1.S2.SS2.SSSx5" specific-use="unnumbered">
  <title>Summary</title>
      <p id="d1e716">These two training datasets correspond to two different approaches to estimate the operational model statistical properties. The real-time forecasts represent the closest version to the operational system, but on a short period so that unusual events cannot be taken into account. The reforecasts are based on a simpler version of the ensemble system which does not contain all sources of error of the system but over a long <italic>climatological</italic> period. The theoretical version of a<?pagebreak page343?> reforecast would be the exact reproduction of the operational system over a long period, but it is currently not performed within the computing facilities of any national weather service.</p>
</sec>
<sec id="Ch1.S2.SS2.SSS3">
  <label>2.2.3</label><title>Observations</title>
      <p id="d1e730">The observation data used in this study have been collected from a network of stations mainly located in French ski resorts. These observation stations are illustrated as white dots in Fig. <xref ref-type="fig" rid="Ch1.F1"/>. All the observations have been manually measured<?pagebreak page344?> by the staff members of these resorts. Measurement of the variable of 24 h height of new snow (HN24) is done daily by measuring the fresh snow height on top of a measuring board. After each measurement, the board is cleaned. These observations are usually carried out every morning. They can be compared directly to the snow reforecasts which are available at each station. This yielded a total of 113 stations in the French Alps and Pyrenees. However, since the real-time snow forecasts used for training are issued only at specific elevations with a 300 m resolution, the observations were in this case associated with the closest standard elevation level in the simulations when the altitude difference was lower than <inline-formula><mml:math id="M7" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">100</mml:mn></mml:mrow></mml:math></inline-formula> m, and were ignored for higher differences. This procedure yielded a total of 47 stations only in the French Alps.</p>
</sec>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>Post-processing method</title>
      <p id="d1e754">Non-homogeneous Gaussian regression (NGR) is one of the most commonly used EMOS methods. NGR was first proposed by <xref ref-type="bibr" rid="bib1.bibx36" id="text.49"/>, <xref ref-type="bibr" rid="bib1.bibx28" id="text.50"/>, and <xref ref-type="bibr" rid="bib1.bibx73" id="text.51"/>. In these early applications of non-homogeneous regression, the predictive (i.e. post-processed) distributions are specified as Gaussian. The mean and variance of the Gaussian distribution are typically modelled with a linear regression model using the raw ensemble mean and variance as predictors. Unlike in ordinary regression-based methods, the dependence of the predictive variance on the ensemble variance in non-homogeneous regressions allows one to exhibit less uncertainty when the ensemble dispersion is small and more uncertainty when the ensemble dispersion is large <xref ref-type="bibr" rid="bib1.bibx67" id="paren.52"/>. The regression coefficients can be estimated from the training data by using optimization techniques based on the maximum likelihood or minimum continuous ranked probability score (CRPS) <xref ref-type="bibr" rid="bib1.bibx26" id="paren.53"/>. However, Gaussian predictive distributions are not adequate for certain meteorological predictands such as precipitation. This can be solved by transforming the target predictand and its predictors such that it is approximately normal <xref ref-type="bibr" rid="bib1.bibx1 bib1.bibx2" id="paren.54"/> or by using non-Gaussian predictive distributions. Many alternative predictive distributions have been proposed. <xref ref-type="bibr" rid="bib1.bibx44" id="text.55"/> applied logistic distribution for modelling square-root transformed wind speeds. Generalized extreme-value distributions were used by <xref ref-type="bibr" rid="bib1.bibx41" id="text.56"/> for forecasting the maximum daily wind speed and by <xref ref-type="bibr" rid="bib1.bibx54" id="text.57"/> for precipitation. <xref ref-type="bibr" rid="bib1.bibx55" id="text.58"/> also proposed a non-homogeneous regression with gamma distribution. Non-negative predictands such as precipitation have high probability mass at zero, and thus the use of a transformation comes with a number of problems. To address this, non-homogeneous regression methods based on truncating and censoring of predictive probabilities have been developed. For example, <xref ref-type="bibr" rid="bib1.bibx65" id="text.59"/> applied a non-homogeneous regression approach with zero-truncated Gaussian distributions for wind speed forecasting. A zero-truncated distribution is a distribution where a random variable has non-zero probability only for positive values and the negative values are excluded. Censoring instead allows a probability distribution to represent values falling below a chosen threshold. Commonly, in the case of ensemble post-processing of precipitation forecasts, the censoring threshold is set to zero and any negative probability is assigned to zero, providing a probability spike at zero. <xref ref-type="bibr" rid="bib1.bibx54" id="text.60"/> used zero-censored GEV predictive distribution in a non-homogeneous regression model for precipitation. A similar approach with a zero-censored shifted-gamma distribution (CSGD) for non-homogeneous regression was introduced by <xref ref-type="bibr" rid="bib1.bibx55 bib1.bibx56" id="text.61"/> and <xref ref-type="bibr" rid="bib1.bibx3" id="text.62"/>.</p>
      <?pagebreak page345?><p id="d1e801">Indeed, as precipitation occurrence/non-occurrence and quantity are modelled together, <xref ref-type="bibr" rid="bib1.bibx55" id="text.63"/> argued for using a continuous distribution that permits negative values and left-censors it at zero. According to their exploratory data analysis, a predictor variable which often takes small values (e.g. the ensemble-mean precipitation forecast) calls for a strongly right-skewed distribution. But as the magnitude of this predictor variable increases, the skewness becomes smaller. This sort of behaviour can be reproduced to some extent by using gamma distributions. The gamma distributions can be defined by a shape parameter <inline-formula><mml:math id="M8" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula> and a scale parameter <inline-formula><mml:math id="M9" display="inline"><mml:mi mathvariant="italic">θ</mml:mi></mml:math></inline-formula>, which are related to the mean <inline-formula><mml:math id="M10" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula> and the standard deviation <inline-formula><mml:math id="M11" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> of the distribution <xref ref-type="bibr" rid="bib1.bibx74" id="paren.64"/>:
            <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M12" display="block"><mml:mrow><mml:mi>k</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msup><mml:mi mathvariant="italic">μ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow><mml:mrow><mml:msup><mml:mi mathvariant="italic">σ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>;</mml:mo><mml:mi mathvariant="italic">θ</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msup><mml:mi mathvariant="italic">σ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow><mml:mi mathvariant="italic">μ</mml:mi></mml:mfrac></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
          <xref ref-type="bibr" rid="bib1.bibx55" id="text.65"/> introduce an additional parameter, the shift <inline-formula><mml:math id="M13" display="inline"><mml:mrow><mml:mi mathvariant="italic">δ</mml:mi><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>. The purpose of this parameter is to deal with the non-negativity of gamma distribution by shifting the CDF of gamma distribution somewhat to the left. Therefore, the CSGD model is defined by
            <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M14" display="block"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>G</mml:mi><mml:mo mathvariant="normal">̃</mml:mo></mml:mover><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">δ</mml:mi></mml:mrow></mml:msub><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mfenced open="{" close=""><mml:mtable rowspacing="0.2ex" columnspacing="1em" class="cases" columnalign="left left" framespacing="0em"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>G</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mstyle displaystyle="false"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mi>y</mml:mi><mml:mo>-</mml:mo><mml:mi mathvariant="italic">δ</mml:mi></mml:mrow><mml:mi mathvariant="italic">θ</mml:mi></mml:mfrac></mml:mstyle></mml:mstyle><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mtext>for </mml:mtext><mml:mi>y</mml:mi><mml:mi mathvariant="italic">⩾</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mtext>for </mml:mtext><mml:mi>y</mml:mi><mml:mo>&lt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced></mml:mrow></mml:math></disp-formula>
          where <inline-formula><mml:math id="M15" display="inline"><mml:mrow><mml:msub><mml:mi>G</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> denotes the cumulated distribution function of a gamma distribution with unit scale and shape parameter <inline-formula><mml:math id="M16" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>. This distribution can be parameterized with <inline-formula><mml:math id="M17" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M18" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M19" display="inline"><mml:mi mathvariant="italic">δ</mml:mi></mml:math></inline-formula> by using Eq. (<xref ref-type="disp-formula" rid="Ch1.E1"/>).</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F2" specific-use="star"><?xmltex \currentcnt{2}?><label>Figure 2</label><caption><p id="d1e1025">Distribution functions of positive HN observations over all dates and stations. <bold>(a)</bold> Frequency histogram of the raw data (grey) and probability density function of the fit with a gamma law (blue). <bold>(b)</bold> Cumulative distribution functions of observations (black) and gamma law (blue) with a focus on the distribution tail.</p></caption>
          <?xmltex \igopts{width=497.923228pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f02.png"/>

        </fig>

      <p id="d1e1041">In the non-homogeneous regression model defined by <xref ref-type="bibr" rid="bib1.bibx56" id="text.66"/>,
when the predictability becomes weak, the forecast CSG distribution converges towards a CSG distribution of mean <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, standard deviation <inline-formula><mml:math id="M21" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, and shift <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, corresponding to the best fit of the gamma law with the climatological distribution of observations. The validity of the adjustment of the climatology with a gamma law was verified over the whole observation dataset in Fig. <xref ref-type="fig" rid="Ch1.F2"/>. It only exhibits a small underestimation of extreme values (for frequencies of exceedance lower than <inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">1000</mml:mn></mml:mrow></mml:math></inline-formula>).</p>
      <p id="d1e1095">Thus, for a given day, <inline-formula><mml:math id="M24" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M25" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M26" display="inline"><mml:mi mathvariant="italic">δ</mml:mi></mml:math></inline-formula> are linked to the raw ensemble forecasts with the regression model of <xref ref-type="bibr" rid="bib1.bibx56" id="text.67"/>:

                <disp-formula specific-use="align" content-type="numbered"><mml:math id="M27" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E3"><mml:mtd><mml:mtext>3</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi mathvariant="italic">μ</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mtext>log1p</mml:mtext><mml:mfenced close="]" open="["><mml:mrow><mml:mtext>expm1</mml:mtext><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub><mml:mi mathvariant="normal">POP</mml:mi><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E4"><mml:mtd><mml:mtext>4</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mi mathvariant="italic">σ</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub><mml:msqrt><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mi mathvariant="italic">μ</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:msqrt><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mi mathvariant="normal">MD</mml:mi><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E5"><mml:mtd><mml:mtext>5</mml:mtext></mml:mtd><mml:mtd><mml:mstyle displaystyle="true" class="stylechange"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi mathvariant="italic">δ</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="italic">δ</mml:mi><mml:mi mathvariant="normal">cl</mml:mi></mml:msub><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            In Eq. (<xref ref-type="disp-formula" rid="Ch1.E3"/>), <inline-formula><mml:math id="M28" display="inline"><mml:mrow><mml:mtext>log1p</mml:mtext><mml:mo>(</mml:mo><mml:mi>u</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mtext>log</mml:mtext><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>+</mml:mo><mml:mi>u</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M29" display="inline"><mml:mrow><mml:mtext>expm1</mml:mtext><mml:mo>(</mml:mo><mml:mi>u</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mtext>exp</mml:mtext><mml:mo>(</mml:mo><mml:mi>u</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>. In this regression model, the ensemble forecasts are summarized by the ensemble mean <inline-formula><mml:math id="M30" display="inline"><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> (normalized by the climatological mean of the forecasts <inline-formula><mml:math id="M31" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi mathvariant="normal">cl</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>), the probability of precipitation <inline-formula><mml:math id="M32" display="inline"><mml:mi mathvariant="normal">POP</mml:mi></mml:math></inline-formula>, and the ensemble mean difference <inline-formula><mml:math id="M33" display="inline"><mml:mi mathvariant="normal">MD</mml:mi></mml:math></inline-formula> (a metric of ensemble spread), as defined by Eqs. (<xref ref-type="disp-formula" rid="Ch1.E6"/>), (<xref ref-type="disp-formula" rid="Ch1.E7"/>) and (<xref ref-type="disp-formula" rid="Ch1.E8"/>):

                <disp-formula specific-use="align" content-type="numbered"><mml:math id="M34" display="block"><mml:mtable displaystyle="true"><mml:mlabeledtr id="Ch1.E6"><mml:mtd><mml:mtext>6</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>M</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>M</mml:mi></mml:munderover><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E7"><mml:mtd><mml:mtext>7</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi mathvariant="normal">POP</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>M</mml:mi></mml:mfrac></mml:mstyle><mml:msub><mml:mi mathvariant="script">I</mml:mi><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr><mml:mlabeledtr id="Ch1.E8"><mml:mtd><mml:mtext>8</mml:mtext></mml:mtd><mml:mtd><mml:mstyle class="stylechange" displaystyle="true"/></mml:mtd><mml:mtd><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi mathvariant="normal">MD</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msup><mml:mi>M</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>M</mml:mi></mml:munderover><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:msup><mml:mi>m</mml:mi><mml:mo>′</mml:mo></mml:msup><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>M</mml:mi></mml:munderover><mml:mo>|</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi>m</mml:mi><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:msub><mml:mo>|</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mlabeledtr></mml:mtable></mml:math></disp-formula>

            with <inline-formula><mml:math id="M35" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> the forecast of each member <inline-formula><mml:math id="M36" display="inline"><mml:mi>m</mml:mi></mml:math></inline-formula> among the <inline-formula><mml:math id="M37" display="inline"><mml:mi>M</mml:mi></mml:math></inline-formula> members, and <inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="script">I</mml:mi><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> if <inline-formula><mml:math id="M39" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>, and 0 otherwise.</p>
      <p id="d1e1596">The regression coefficients <inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M41" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M42" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M43" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M44" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M45" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> are estimated by the optimization process used by <xref ref-type="bibr" rid="bib1.bibx55" id="text.68"/> as described in the next section.</p>
      <p id="d1e1669">In addition to the convergence of this model towards the climatological distribution for weak predictability, this model also includes several advantages compared to standard non-homogeneous regressions. First, the <inline-formula><mml:math id="M46" display="inline"><mml:mi mathvariant="normal">POP</mml:mi></mml:math></inline-formula> predictor can improve the forecast distribution compared to models based only on the ensemble mean by providing complementary information  about the expected precipitation occurrence. Then, the links between <inline-formula><mml:math id="M47" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M48" display="inline"><mml:mover accent="true"><mml:mi>x</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> and between <inline-formula><mml:math id="M49" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M50" display="inline"><mml:mi mathvariant="normal">POP</mml:mi></mml:math></inline-formula> are not supposed to be linear in Eq. (<xref ref-type="disp-formula" rid="Ch1.E3"/>) (the model tends to the linear case when <inline-formula><mml:math id="M51" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>→</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>). Finally, Eq. (<xref ref-type="disp-formula" rid="Ch1.E4"/>) introduces an explicit heteroscedasticity (<inline-formula><mml:math id="M52" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> does not only depend on <inline-formula><mml:math id="M53" display="inline"><mml:mi mathvariant="normal">MD</mml:mi></mml:math></inline-formula>, but also on <inline-formula><mml:math id="M54" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula>). This important property for precipitation (or snowfall) may not be sufficiently reproduced by the spread of the raw ensemble. Extended justifications of the form of this regression model are provided in <xref ref-type="bibr" rid="bib1.bibx55 bib1.bibx56" id="text.69"/>.</p>
</sec>
<sec id="Ch1.S2.SS4">
  <label>2.4</label><title>Evaluation metrics</title>
      <p id="d1e1763">It is commonly admitted that reliability and resolution are the two main properties to qualify the skill of a probabilistic prediction system <xref ref-type="bibr" rid="bib1.bibx12" id="paren.70"/>. Reliability is defined as a statistical consistency between the predicted probabilities and the subsequent observations. For instance, a probabilistic prediction system is reliable if a given snowfall occurs with frequency <inline-formula><mml:math id="M55" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula> when it is predicted to occur with the probability <inline-formula><mml:math id="M56" display="inline"><mml:mrow><mml:mi>p</mml:mi><mml:mo>∀</mml:mo><mml:mi>p</mml:mi><mml:mo>∈</mml:mo><mml:mo>[</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>]</mml:mo></mml:mrow></mml:math></inline-formula>. A system can be reliable if it would always predict the climatological distribution of the atmospheric variable under consideration. However, that would lack practical usefulness, and therefore the second property, resolution, implies that the individual spread of the predicted distributions must be smaller than the climatological spread.</p>
<sec id="Ch1.S2.SS4.SSS1">
  <label>2.4.1</label><title>CRPS</title>
      <p id="d1e1807">The continuous ranked probability score (CRPS) is one of the most common probabilistic tools to evaluate the ensemble skill in terms of both reliability (unbiased probabilities) and resolution (ability to separate the probability classes) <xref ref-type="bibr" rid="bib1.bibx12" id="paren.71"/>. For a given forecast, the CRPS corresponds to the integrated quadratic distance between the CDF of ensemble forecast and the CDF of observation. Commonly, the CRPS is averaged over <inline-formula><mml:math id="M57" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> available forecasts following Eq. (<xref ref-type="disp-formula" rid="Ch1.E9"/>):
              <disp-formula id="Ch1.E9" content-type="numbered"><label>9</label><mml:math id="M58" display="block"><mml:mrow><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:munder><mml:mo movablelimits="false">∫</mml:mo><mml:mi mathvariant="double-struck">R</mml:mi></mml:munder><mml:mo>(</mml:mo><mml:msub><mml:mi>F</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>H</mml:mi><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mi>o</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mi mathvariant="normal">d</mml:mi><mml:mi>x</mml:mi><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
            where <inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the cumulative distribution function of the ensemble simulation at time <inline-formula><mml:math id="M60" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M61" display="inline"><mml:mrow><mml:msub><mml:mi>o</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> the observation at time <inline-formula><mml:math id="M62" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>, and <inline-formula><mml:math id="M63" display="inline"><mml:mrow><mml:mi>H</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is the Heaviside function (<inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:mi>H</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula> if <inline-formula><mml:math id="M65" display="inline"><mml:mrow><mml:mi>y</mml:mi><mml:mo>≤</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>; <inline-formula><mml:math id="M66" display="inline"><mml:mrow><mml:mi>H</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> if <inline-formula><mml:math id="M67" display="inline"><mml:mrow><mml:mi>y</mml:mi><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>). The CRPS value has the same unit as the evaluated variable and tends towards 0 for a perfect system. Note that in the case of a CSG distribution (when <inline-formula><mml:math id="M68" display="inline"><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>G</mml:mi><mml:mo mathvariant="normal">̃</mml:mo></mml:mover><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">θ</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="italic">δ</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>), an analytic expression of CRPS allows one to directly compute the score from the parameters <inline-formula><mml:math id="M69" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M70" display="inline"><mml:mi mathvariant="italic">θ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M71" display="inline"><mml:mi mathvariant="italic">δ</mml:mi></mml:math></inline-formula> <xref ref-type="bibr" rid="bib1.bibx55" id="paren.72"/>. In this study, <inline-formula><mml:math id="M72" display="inline"><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> is used to optimize the six regression parameters (<inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M75" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">α</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M78" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) of Eqs. (<xref ref-type="disp-formula" rid="Ch1.E3"/>) and (<xref ref-type="disp-formula" rid="Ch1.E4"/>) by minimizing this score on the training<?pagebreak page346?> data. We remind the reader that the correspondence between (<inline-formula><mml:math id="M79" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M80" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula>) and (<inline-formula><mml:math id="M81" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M82" display="inline"><mml:mi mathvariant="italic">θ</mml:mi></mml:math></inline-formula>) is given by Eq. (<xref ref-type="disp-formula" rid="Ch1.E1"/>). Then, <inline-formula><mml:math id="M83" display="inline"><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> is also used to evaluate the overall skill of the HN raw forecasts of the PEARP-S2M system. Finally, to assess the improvement obtained by the post-processing compared to the raw forecasts, we compute the continuous ranked probability skill score:
              <disp-formula id="Ch1.E10" content-type="numbered"><label>10</label><mml:math id="M84" display="block"><mml:mrow><mml:mi mathvariant="normal">CRPSS</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
            where <inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the reference mean CRPS of the raw ensemble forecasts. Therefore, positive CRPSS values indicate an improvement compared to the raw forecasts. In this work, <inline-formula><mml:math id="M86" display="inline"><mml:mover accent="true"><mml:mi mathvariant="normal">CRPS</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> and CRPSS were computed separately for each station, and we present the distribution of these scores among stations.</p>
</sec>
<sec id="Ch1.S2.SS4.SSS2">
  <label>2.4.2</label><title>Rank histograms and quantile–quantile plots</title>
      <p id="d1e2260">Statistical post-processing is mainly expected to improve the reliability of ensemble forecast systems. Therefore, we chose to present complementary diagnostics to better illustrate the improvement obtained by the post-processing in terms of reliability compared to the raw forecasts. For that purpose, we used rank histograms and quantile–quantile plots.</p>
      <p id="d1e2263">Rank histograms <xref ref-type="bibr" rid="bib1.bibx29" id="paren.73"/> illustrate the occurrence frequency of the different possible ranks of the observations <inline-formula><mml:math id="M87" display="inline"><mml:mrow><mml:msub><mml:mi>o</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> among the sorted ensemble members. The flatness of this histogram is a condition of the system reliability (if the simulated probabilities are unbiased regardless of the probability level,  the different ranks should have a uniform occurrence frequency). It is also an indicator of the spread skill as underdispersion will result in a U-shaped rank histogram and overdispersion in a bell-shaped rank histogram. Rank histograms are commonly computed for the whole forecast dataset, but this can hide contrasted behaviours between the different parts of the distribution. Forecast stratification <xref ref-type="bibr" rid="bib1.bibx8" id="paren.74"/>, as the process of dividing the whole dataset into different subsets and computing verification metrics for each subset, has been introduced as a way to better diagnose where the deficiencies of the forecast system lie. <xref ref-type="bibr" rid="bib1.bibx6" id="text.75"/> compared different strategies for the stratification criteria based on either the observations or the forecasts, and justified the use of a forecast-based stratification criterion for verification rank histograms. Indeed, they showed that conditioning the rank histogram to observations is likely to draw erroneous conclusions about the real behaviour of ensemble forecasts. Therefore, in this study, a forecast-based stratification is used by considering three HN intervals <inline-formula><mml:math id="M88" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">10</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M89" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">10</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">30</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M90" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">30</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="normal">∞</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula> for the ensemble mean. To guarantee a sufficient sample size for rank histograms, they are computed for the whole evaluation dataset by considering all dates and stations to be independent.</p>
      <p id="d1e2352">To understand the connection between forecast errors and the magnitude of observed HN, quantile–quantile plots of sorted observations as a function of the forecast quantiles for the equivalent frequency levels are also presented. Contrary to the rank histograms, quantile–quantile plots do not discriminate the probability classes in the ensemble forecasts, but they allow one to verify that the post-processing removes the biases for any value of the forecast variable, with a reduced constraint on sample size compared to the stratified rank histogram. Similarly to the rank histograms, quantile–quantiles plots are computed for the whole evaluation dataset (all dates and stations).</p>
</sec>
</sec>
<sec id="Ch1.S2.SS5">
  <label>2.5</label><title>Experiments</title>
      <p id="d1e2365">The post-processing method described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS3"/> was calibrated on the data listed in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2.SSS2"/>. Several experiments were performed. First, for each station, the predictor is the simulated HN from the snow reforecast and the predictand is the observed HN at the station. This leads to a different calibration for each station. Then, for each massif, the same<?pagebreak page347?> predictors and predictands of all stations inside the massif boundaries are mixed in the same training vectors as independent events. This method will be further referred to as massif-scale calibration because it leads to a unique calibration for each massif. The results of these first two experiments are described in Sect. <xref ref-type="sec" rid="Ch1.S3.SS2.SSS1"/>. Finally, the same massif-scale calibration is applied by using the real-time forecasts as predictors. The comparison of both training datasets is analysed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS2.SSS2"/>. The skill of the raw forecast and all post-processing experiments is assessed with the independent evaluation dataset described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2.SSSx3"/> with the metrics of Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/>.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Results and discussion</title>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Evaluation of raw forecasts</title>

      <?xmltex \floatpos{t}?><fig id="Ch1.F3"><?xmltex \currentcnt{3}?><label>Figure 3</label><caption><p id="d1e2399">Evaluation of raw HN forecasts from PEARP-S2M during winter 2017–2018. <bold>(a)</bold> CRPS of HN as a function of prediction lead time. The boxplot represents the variability of scores between the 113 stations. <bold>(b)</bold> Rank histograms of HN forecasts for three classes of HN ensemble mean (indigo: <inline-formula><mml:math id="M91" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">10</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula>, cyan: <inline-formula><mml:math id="M92" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">10</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">30</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula>, red: <inline-formula><mml:math id="M93" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mn mathvariant="normal">30</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="normal">cm</mml:mi><mml:mo>,</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="normal">∞</mml:mi><mml:mo>[</mml:mo></mml:mrow></mml:math></inline-formula>). <bold>(c)</bold> Quantile–quantile plot: the black dots represent sorted observations as a function of the forecast quantiles for the equivalent frequency levels. Red line illustrates the ideal distribution.</p></caption>
          <?xmltex \igopts{width=170.716535pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f03.png"/>

        </fig>

      <p id="d1e2483">Raw HN real-time forecasts of winter 2017–2018 (Sect. <xref ref-type="sec" rid="Ch1.S2.SS2.SSSx3"/>) are evaluated. The CRPS at different lead times of the raw forecast is given in Fig. <xref ref-type="fig" rid="Ch1.F3"/>a. The boxplots represent the variability of the score between stations, which is relatively large at all lead times. The mean CRPS slightly deteriorates with longer lead times.</p>
      <p id="d1e2490">The rank histogram of the raw forecast is presented in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b and stratified according to the ensemble mean with three different categories (low subset in blue: 0–10 cm, medium subset in green: 10–30 cm, high subset in red: above 30 cm). The raw HN forecasts are biased with high underdispersion on all three subsets (U shape). Above 10 cm forecasts, about 50 % of the observed values are not included in the ensemble (rank 1 or rank 36). Note that the small sample size of the high subset (79 events) causes a high sampling variability in the different probability classes.</p>
      <p id="d1e2496">To understand the link between forecast errors and the magnitude of HN, a Q–Q plot of sorted observations as a function of the forecast quantiles for the equivalent probability levels is presented in Fig. <xref ref-type="fig" rid="Ch1.F3"/>c. The systematic bias in the forecast increases as the observed HN increases. However, since the sample size of the high observed HN is small, we can expect significant sampling variability in the upper tail.</p>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Evaluation of post-processed forecasts</title>
<sec id="Ch1.S3.SS2.SSS1">
  <label>3.2.1</label><title>Comparison of local-scale and massif-scale training</title>

      <?xmltex \floatpos{t}?><fig id="Ch1.F4" specific-use="star"><?xmltex \currentcnt{4}?><label>Figure 4</label><caption><p id="d1e2518">Comparison of post-processing skill between local-scale training (left column) and massif-scale training (right column) for post-processed HN forecasts calibrated with the reforecast dataset (1994–2016) and evaluated during winter 2017–2018. <bold>(a, b)</bold> CRPS of HN (cm) as a function of prediction lead time; the boxplot represents the variability of scores between the 113 stations. <bold>(c, d)</bold> Rank histograms; the three HN classes are the same as in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b. <bold>(e, f)</bold> Quantile–quantile plot.</p></caption>
            <?xmltex \igopts{width=341.433071pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f04.png"/>

          </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F5" specific-use="star"><?xmltex \currentcnt{5}?><label>Figure 5</label><caption><p id="d1e2540">Time series of the raw (blue) and post-processed (grey) ensemble forecasts during January 2018 from a training based on the snow reforecasts. The envelopes represent the interval between the 10th and 90th percentiles and the solid lines represent the median. These ensemble forecasts are compared to time series of HN observations (red lines). <bold>(a, b)</bold> Example of station 73 034 400 (Arêches) at the <inline-formula><mml:math id="M94" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">48</mml:mn></mml:mrow></mml:math></inline-formula> h lead time. <bold>(c, d)</bold> Example of station 73 235 400 (Saint-François-Longchamp) at the <inline-formula><mml:math id="M95" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">96</mml:mn></mml:mrow></mml:math></inline-formula> h lead time. <bold>(a, c)</bold> Local-scale training. <bold>(b, d)</bold> Massif-scale training.</p></caption>
            <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f05.png"/>

          </fig>

      <p id="d1e2582">In this section, we present only the results obtained by using the reforecast dataset as training and we compare the impact of a local-scale calibration for each station to a massif-scale calibration where all observations in the same massif are mixed in the same training vector in order to obtain only one set of parameters by massif for Eqs. (<xref ref-type="disp-formula" rid="Ch1.E3"/>) and (<xref ref-type="disp-formula" rid="Ch1.E4"/>) that can be applied at any point in the massifs. The evaluation is performed on 113 stations in 30 different massifs.</p>
      <p id="d1e2590"><?xmltex \hack{\newpage}?>The CRPSS of each station for the verification period and with the raw forecast as a reference is presented in Fig. <xref ref-type="fig" rid="Ch1.F4"/>a and b. Post-processing with both local-scale and massif-scale training significantly improves the CRPS in the majority of the stations (positive skill scores for a large majority of stations), although in both cases, the improvement decreases with longer lead times. The CRPSS of local-scale training is slightly better than the CRPSS of massif-scale training on smaller lead times (24 and 48 h), but for longer lead times (72 and 96 h), the difference between local scale and massif scale decreases. Overall, the difference between local-scale and massif-scale training according to the CRPSS is limited.</p>
      <?pagebreak page348?><p id="d1e2596">The rank histograms of the post-processed ensembles with local-scale and massif-scale snow reforecast training are presented in Fig. <xref ref-type="fig" rid="Ch1.F4"/>c and d. In both cases the shapes of the histograms are similar and show that the reliability has been greatly improved  compared to the raw forecast (Fig. <xref ref-type="fig" rid="Ch1.F3"/>b). This is the expected behaviour of the post-processing, and obtaining such a result on the validation period independent of the training period proves the robustness of the model. However, a relative overdispersion of the post-processed forecasts can be noticed (slight bell shape of the histograms). As for the rank histogram of the raw forecast, the sample size of the high subset is small and causes variability in the corresponding rank histogram (red bars).</p>
      <p id="d1e2603">Q–Q plots of both post-processing training scenarios with the snow reforecast are presented in Fig. <xref ref-type="fig" rid="Ch1.F4"/>e and f. In both cases, a significant improvement compared to the raw Q–Q plot (Fig. <xref ref-type="fig" rid="Ch1.F3"/>c) can be noted over all quantiles. Indeed, the Q–Q plot shows that the post-processed forecasts and the observations have almost the same climatological distribution. Again, the sample size of high observed HN is small and causes sampling variability in the upper tail.</p>
      <?pagebreak page349?><p id="d1e2610">Examples of raw and post-processed ensembles with the snow reforecast training at local scale and massif scale are given in Fig. <xref ref-type="fig" rid="Ch1.F5"/>. In all four cases, the CRPSS are around <inline-formula><mml:math id="M96" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">30</mml:mn></mml:mrow></mml:math></inline-formula> %, showing a clear improvement of the forecasts by the post-processing over January 2018. Note that the scores over this short period are provided for the example, but only the previous scores computed over the whole evaluation period (Fig. <xref ref-type="fig" rid="Ch1.F4"/>) should be considered for robust conclusions. The better improvement is obtained with the local-scale training in the first example (Fig. <xref ref-type="fig" rid="Ch1.F5"/>a and b) but with the massif-scale training in the second example (Fig. <xref ref-type="fig" rid="Ch1.F5"/>c and d). These examples show that some differences can be observed between the skill of the local-scale and massif-scale training but with variability between stations and no systematic improvement or deterioration. This is consistent with the similar scores presented before between both spatial scales. In both examples and regardless of the spatial scale of training, the post-processing increases the median and the spread compared to the raw ensemble, consistent with the systematic negative bias and underdispersion of the raw forecast observed in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b and c. Thus, for most days with observed snowfall, the observations fall inside the EMOS quantiles, whereas they frequently fall outside the raw ensemble. Nevertheless, the method causes overdispersion especially visible by adding spread even for the days when all the raw forecast members predict no snowfall.</p>
</sec>
<sec id="Ch1.S3.SS2.SSS2">
  <label>3.2.2</label><title>Comparison of real-time snow forecast and snow reforecast training</title>

      <?xmltex \floatpos{t}?><fig id="Ch1.F6" specific-use="star"><?xmltex \currentcnt{6}?><label>Figure 6</label><caption><p id="d1e2644">Comparison of post-processing skill between a training with the reforecasts dataset (1994–2016, left column) and a training with the real-time forecasts dataset (2014–2017, right column) for post-processed HN forecasts calibrated at the massif scale and evaluated during winter 2017–2018. <bold>(a, b)</bold> CRPS of HN (cm) as a function of prediction lead time; the boxplot represents the variability of scores between the 47 stations. <bold>(c, d)</bold> Rank histograms; the three HN classes are the same as in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b. <bold>(e, f)</bold> Quantile–quantile plot.</p></caption>
            <?xmltex \igopts{width=341.433071pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f06.png"/>

          </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F7" specific-use="star"><?xmltex \currentcnt{7}?><label>Figure 7</label><caption><p id="d1e2666">Time series of the raw (blue) and post-processed (grey) ensemble forecasts during January 2018 from a massif-scale training. The envelopes represent the interval between the 10th and 90th percentiles and the solid lines represent the median. These ensemble forecasts are compared to time series of HN observations (red lines). <bold>(a, b)</bold> Example of station 73 235 400 (Saint-François-Longchamp) at the <inline-formula><mml:math id="M97" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">24</mml:mn></mml:mrow></mml:math></inline-formula> h lead time. <bold>(c, d)</bold> Example of station 73 023 401 (Aussois) at the <inline-formula><mml:math id="M98" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mn mathvariant="normal">72</mml:mn></mml:mrow></mml:math></inline-formula> h lead time.  <bold>(a, c)</bold> Training with the snow reforecasts. <bold>(b, d)</bold> Training with the snow real-time forecasts.</p></caption>
            <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f07.png"/>

          </fig>

      <p id="d1e2708">In this section, we present only the results obtained with a massif-scale training and we compare the impact of using the reforecast dataset or the real-time forecast dataset as training, which do not have the same advantages and disadvantages. As mentioned in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2.SSS2"/>, only 47 stations in 21 different massifs were included in this comparison. Note that the results obtained in Sects. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and <xref ref-type="sec" rid="Ch1.S3.SS2.SSS1"/> are not significantly different between the 113 stations and this subset of 47 stations, as shown by the similar diagnostics obtained for massif-scale reforecast training in both sections (Fig. <xref ref-type="fig" rid="Ch1.F4"/>b, d and f just differ from Fig. <xref ref-type="fig" rid="Ch1.F6"/>a, c and e by the number of stations considered.) The CRPSS of each station for the verification period, with the raw forecast as a reference, is presented in Fig. <xref ref-type="fig" rid="Ch1.F6"/>a and b. In both cases, the CRPS is improved in the majority of the stations. Similarly to Sect. <xref ref-type="sec" rid="Ch1.S3.SS2.SSS1"/>, the CRPSS decreases with longer lead times. However, this decrease is smaller with the real-time snow forecast training, and it performs better at 96 h lead time compared to the snow reforecast training. As can be noted, the variability of mean CRPSS among the stations is generally higher with the real-time snow forecast training.</p>
      <p id="d1e2727">The rank histograms of post-processed forecasts for both training scenarios are given in Fig. <xref ref-type="fig" rid="Ch1.F6"/>c and d. The calibration of the post-processed ensembles between these two different training scenarios is different. In the case of a snow reforecast training, similar overdispersive behaviour for the low subset (blue) can be noted as was in the previous comparison with a higher number of stations, whereas such issue is not obtained with the real-time snow forecast training. However, the real-time snow forecast training causes a significant positive bias for the medium and high subsets which is not observed with the snow reforecast training.</p>
      <p id="d1e2732">Q–Q plots with the snow reforecast and the real-time snow forecast are presented in Fig. <xref ref-type="fig" rid="Ch1.F6"/>e and f. Similarly to the Q–Q plots in the previous comparison, there is a significant improvement over all quantiles compared to the raw Q–Q plot (Fig. <xref ref-type="fig" rid="Ch1.F3"/>c). The biases in the ranks from Fig. <xref ref-type="fig" rid="Ch1.F6"/>d are not translated into a bias in the simulated quantiles.</p>
      <?pagebreak page350?><p id="d1e2741">Examples of two different stations and lead times of post-processed ensembles with snow reforecast massif-scale training and the real-time snow forecast massif-scale training are given in Fig. <xref ref-type="fig" rid="Ch1.F7"/>.  In both examples with the snow reforecast training (Fig. <xref ref-type="fig" rid="Ch1.F7"/>a and c), post-processing increases the spread by stretching the distribution below and above the raw ensemble, whereas with the real-time snow forecast training (Fig. <xref ref-type="fig" rid="Ch1.F7"/>b and d), the distribution is mostly stretched above the raw ensemble. In the example of Fig. <xref ref-type="fig" rid="Ch1.F7"/>c and d, the real-time snow forecast training performs better since the raw forecast underestimates the snowfall magnitude. However, in the example of Fig. <xref ref-type="fig" rid="Ch1.F7"/>a and b, post-processing is improved with the snow reforecast training because the raw forecast overestimates the magnitude of the snowfall and the observations fall multiple times below the raw forecast. In this example, the post-processing based on the real-time snow forecast training deteriorates the raw forecast by only stretching the distribution towards higher values.
Similarly to the impact of the spatial scale on the training data, there is not any systematic positive or negative impact of the training dataset on the skill of the post-processing. The main advantages and disadvantages of the real-time snow forecast vs. the snow reforecast training identified in the rank histograms are also emphasized in these examples. First, the overdispersion on dry days obtained by the snow reforecast training can be again observed in Fig. <xref ref-type="fig" rid="Ch1.F7"/>c. This issue disappears with the real-time snow forecast consistently with the satisfactory shape of the low subset (blue bars) in the rank histograms. However, the reliability of the forecasts for severe snowfall events is better with the reforecast training in the example of Fig. <xref ref-type="fig" rid="Ch1.F7"/>a and b, consistent with the systematic bias of the medium and high subsets (green and red bars) obtained in Fig. <xref ref-type="fig" rid="Ch1.F6"/>d.</p>
</sec>
</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Discussion</title>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>Implications for operational automatic forecasts</title>
<sec id="Ch1.S4.SS1.SSS1">
  <label>4.1.1</label><title>Added value of post-processed HN forecasts</title>
      <?pagebreak page351?><p id="d1e2785">Evaluations of the HN raw forecast from the PEARP-S2M ensemble snowpack modelling system in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> exhibit a significant underdispersion over all subsets as well as an increasing systematic bias as a function of the height of new snow. It is the result of a bias and underdispersion of the PEARP precipitation forecasts <xref ref-type="bibr" rid="bib1.bibx68" id="paren.76"/>, but also of errors in recent snow density in the Crocus snowpack model (Sect. <xref ref-type="sec" rid="Ch1.S2.SS1.SSS3"/>) and a lack of accounting for uncertainty in the associated processes in the raw forecasts. Therefore, we recommend avoiding the development of automatic products of HN forecasts based on the raw simulations.</p>
      <p id="d1e2795">Statistical processing can help improve the reliability of the forecasts in such products, while the correction of these errors and underdispersion is too challenging to be quickly solved in the NWP and snowpack models. According to the results of this study, we can state that the use of statistical post-processing with the CSGD method in the case of ensemble HN forecasts is beneficial in most of the evaluated stations in all of the experiments conducted. The extent of these improvements was more or less similar to what had already been found by several authors in the case of statistical post-processing of ensemble precipitation forecasts <xref ref-type="bibr" rid="bib1.bibx25 bib1.bibx55 bib1.bibx56" id="paren.77"/>. However, since statistical post-processing of ensemble forecasts had never been applied in the literature on the outputs of a detailed snowpack model, the findings of this study are very promising in terms of automatic HN forecast developments. Thanks to many advantages of the physical modelling of the snowpack, the method represents an alternative to the more complex statistical frameworks developed by <xref ref-type="bibr" rid="bib1.bibx59" id="text.78"/> and <xref ref-type="bibr" rid="bib1.bibx57" id="text.79"/> from direct NWP diagnostics as predictors.</p>
</sec>
<sec id="Ch1.S4.SS1.SSS2">
  <label>4.1.2</label><title>Spatial scale</title>
      <p id="d1e2815">Due to the similar improvements when training data were considered at local scale or massif scale (Sect. <xref ref-type="sec" rid="Ch1.S3.SS2.SSS1"/>), the use of massif-scale training is justified. Indeed, it means that the post-processing can be applied at any point of the massifs because homogeneous sets of calibration parameters are obtained for each massif. This is especially interesting for the operational HN forecasting, which has no reason to be limited to observation stations. Note that a potential limitation for applying the post-processing anywhere is the relatively limited elevation range of observations used in the evaluations (50 % of observation stations are between 1400 and 2000 m a.s.l.). Nevertheless, local-scale post-processing is interesting as well and can be applied especially if the objective is to gain more reliable HN forecasts for specific locations (e.g. ski resorts).</p>
</sec>
<sec id="Ch1.S4.SS1.SSS3">
  <label>4.1.3</label><title>Training dataset</title>
      <p id="d1e2828">Even though local-scale and massif-scale trainings resulted in similar post-processing performances, that was not exactly the case when training forecasts with different lengths and characteristics were compared in Sect. <xref ref-type="sec" rid="Ch1.S3.SS2.SSS2"/>. This comparison resulted in large differences between the snow reforecast training and the real-time snow forecast training. The reliability of severe snowfalls was not satisfactory with the real-time snow forecast training. This may be due to the small<?pagebreak page352?> training length, making highly possible the fact that the climatology of the training period differs from the climatology of the verification period. To understand this issue, Fig. <xref ref-type="fig" rid="Ch1.F8"/> shows the differences in terms of the amount of snow coverage in the northern Alps at 1800 m between the successive seasons. As can be noted, the difference between the snow coverage during the evaluation period (2018) and the training period (2015, 2016 and 2017) is significant. Even when compared to all of the winters in Fig. <xref ref-type="fig" rid="Ch1.F8"/>, the year of 2018 is exceptional. Figure <xref ref-type="fig" rid="Ch1.F9"/> presents a rank histogram obtained by cross-validation of three different post-processing calibrations in which the training periods were 3 years of the 2014–2018 period excluding the evaluation year, repeating this process with 2015, 2016 and 2017 as evaluation years. The rank histogram is the mean of the rank frequencies of these three simulations. Such a cross-validation procedure reduces the impact of seasonal differences in the shape of a rank histogram. The shape of the highest subset (red bars) in the cross-validated rank histogram of 2014–2017 (Fig. <xref ref-type="fig" rid="Ch1.F9"/>) is completely different from the rank histogram obtained for the verification period of 2017–2018 (Fig. <xref ref-type="fig" rid="Ch1.F6"/>d). Instead of positive bias, the cross-validated verification rank histogram indicates a negative bias. Such behaviour supports the previous arguments about the impact of the seasonal differences, which is especially problematic for operational use in case only short training periods are available for the post-processing. Indeed, it is highly possible that the upcoming season is significantly different from the past few seasons. Such an issue can be avoided or minimized by using longer training periods when reforecasts are available. This conclusion is fully consistent with a significant decrease in the forecast skill obtained by <xref ref-type="bibr" rid="bib1.bibx55" id="text.80"/> for the highest precipitation amount when reducing the training data length among the same reforecast dataset.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F8" specific-use="star"><?xmltex \currentcnt{8}?><label>Figure 8</label><caption><p id="d1e2849">Qualification of snow coverage at 1800 m on northern aspects in the French northern Alps: for each winter (November–April), percentage of days with total snow depth much below average (lower than the 20th climatological percentile), below average (between the 20th and 40th percentiles), near average (between the 40th and 60th percentiles), above average (between the 60th and 80th percentiles), and much above average (higher than the 80th climatological percentile).</p></caption>
            <?xmltex \igopts{width=341.433071pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f08.png"/>

          </fig>

      <p id="d1e2858">However, the main limitation to the use of the PEARP-S2M reforecast instead of the real-time forecasts is the overdispersion generated in the post-processed forecast. It may be due to the discrepancy in the perturbations and model configurations between these two forecasts. As explained in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2.SSS2"/>, the snow reforecast and the real-time snow forecast have different perturbations and model configurations. The snow reforecast accounts only for the physical perturbations and, thus, it has 10 members, whereas the real-time snow forecast accounts for both physical and initial perturbations, making it a 35-member forecast. Due to the discrepancy in the perturbations, the real-time snow forecast has a higher spread than the snow reforecast. Hence, this can lead to an overcorrection of the spread in case the model is trained with the snow reforecast (lower spread) and verified with the real-time snow forecast (higher spread). To understand the importance of homogeneity between training and verification forecasts, Fig. <xref ref-type="fig" rid="Ch1.F10"/> shows CRPSS comparison between two cases where the post-processing was applied with the same training data (the snow reforecast local-scale) and for the same verification period (winter 2015–2016), but first the post-processing was verified with the snow reforecast as a predictor and in the second case the verification was done with the real-time snow forecast as a predictor. For this analysis, the verification period had to be included in the training period due to the limited recovery between both datasets. According to the skill scores, the difference between these two forecast evaluations was considerable. At all lead times, the skill score with the snow reforecast verification is higher than with the real-time snow forecast verification: the improvement in the median CRPSS is about 0.12. Similar skill scores were also obtained for different seasons when the snow reforecast was used for verification. Hence, the ideal training forecast for statistical post-processing should be as homogeneous as possible with the verification (operational) forecast. This is an important feedback of this work for research teams in charge of atmospheric modelling: even if the numerical costs are higher, applications of ensemble NWP need reforecasts which include all perturbations implemented in the operational system. Before such a dataset is available, it is difficult to decide whether an operational post-processing should be based on the snow reforecast or on the real-time snow forecast training as it depends on whether we prefer to optimize for severe events with high socio-economic impacts or to optimize the spread during dry days (which can be an important factor of confidence for the end-user in an automatic product).</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F9"><?xmltex \currentcnt{9}?><label>Figure 9</label><caption><p id="d1e2868">Cross-validated rank histograms of the post-processed HN forecasts from local-scale calibration with the real-time forecasts dataset. The evaluation is done separately for winters 2014, 2015, and 2016 with the training period 2014–2018 excluding the evaluation year. The three HN classes are the same as in Fig. <xref ref-type="fig" rid="Ch1.F3"/>b.</p></caption>
            <?xmltex \igopts{width=170.716535pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f09.png"/>

          </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F10"><?xmltex \currentcnt{10}?><label>Figure 10</label><caption><p id="d1e2881">CRPS of HN (cm) as a function of prediction lead time for post-processed forecasts from local-scale calibration with the reforecast dataset (1994–2016) and applied during winter 2015–2016. <bold>(a)</bold> Predictors for the verification are taken from the reforecast; <bold>(b)</bold> predictors for the verification are taken from the real-time forecasts. The boxplot represents the variability of scores between the 47 stations in the French Alps.</p></caption>
            <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://npg.copernicus.org/articles/26/339/2019/npg-26-339-2019-f10.png"/>

          </fig>

</sec>
</sec>
<sec id="Ch1.S4.SS2">
  <label>4.2</label><title>Possible refinements</title>
      <p id="d1e2905">To maximize the performance of the statistical post-processing method used in this study, some refinements or extensions could be considered. According to the personal communications with the forecasters of Météo-France, the systematic biases in NWP models may depend on circulation regimes. Hence, categorizing the training data by weather types and computing the regression parameters accordingly may be interesting. However, this would decrease the training length and could be problematic especially in the case of the real-time snow forecast training which already had a relatively short training period.</p>
      <p id="d1e2908">Another extension of the method could be the addition of new predictors. The use of a physical snowpack model reduces the need to consider both precipitation and temperature variables compared to <xref ref-type="bibr" rid="bib1.bibx59" id="text.81"/> and <xref ref-type="bibr" rid="bib1.bibx57" id="text.82"/>. Indeed, situations close to the critical threshold of 0 <inline-formula><mml:math id="M99" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula>C are likely to already result in an increased spread in the raw ensemble forecasts (some members are going to forecast rain, some others to forecast snow). Therefore the post-processed forecasts naturally exhibit more spread in this case as it is linked to ensemble spread by Eq. (<xref ref-type="disp-formula" rid="Ch1.E4"/>). Nevertheless, it is still true that the system may not have the same biases and skill scores depending on various meteorological variables such as temperature or wind speed, or even depending on the month of the year. These variables might be able to improve the statistical relationship in more<?pagebreak page353?> complex statistical models. Quantile random forests for instance could be tested <xref ref-type="bibr" rid="bib1.bibx62" id="paren.83"/> as they do not need to presume the required predictors in advance and because they could allow combination of the different available training datasets by adding a categorical variable. <xref ref-type="bibr" rid="bib1.bibx63" id="text.84"/> showed that hybrid forest-based procedures produce the largest skill improvements for forecasting heavy rainfall events over France.</p><?xmltex \hack{\newpage}?>
</sec>
</sec>
<sec id="Ch1.S5" sec-type="conclusions">
  <label>5</label><title>Conclusion</title>
      <p id="d1e2946">Various weather services are trying to increase the part of automatic forecasts in their production. This includes the challenging forecast of the height of new snow. The PEARP-S2M modelling system, designed for avalanche hazard forecasting, can also help for this application. Indeed, the PEARP ensemble numerical weather prediction (NWP) model quantifies the uncertainty of the forecast, the SAFRAN downscaling tool refines the elevation resolution, and the Crocus snowpack model represents the main physical processes responsible for the variability of the height of new snow.<?pagebreak page354?> However, the raw outputs of PEARP-S2M are biased and underdispersive. The origins of these biases in atmospheric ensembles and snow models are challenging to detect and correct, and hence a statistical post-processing of the HN output is necessary. In this study, a non-homogeneous regression method based on censored shifted gamma distributions was tested to calibrate HN forecasts.
The predictands are snow-board measurements of the height of 24 h new snow from a network of stations located in the French Alps and Pyrenees. HN outputs from the PEARP-S2M model chain were statistically post-processed by considering local-scale and massif-scale training vectors. The method was applied with two different predictor datasets for training (snow reforecast and real-time snow forecast).</p>
      <p id="d1e2949">The chosen statistical post-processing method was found to be successful as the forecast skills were improved for the majority of the stations in all the conducted experiments. Local-scale and massif-scale trainings had similar improvements and therefore the use of massif-scale training can be preferred for its ability to be applied at a larger spatial scale than the observation points. However, a potential limitation comes with the relatively limited elevation range of the observations since most of the stations are between 1400 and 2000 m a.s.l. Comparison between the snow reforecast training and the real-time snow forecast training revealed two main challenges. First, due to the higher spread in the verification (operational) forecast than in the snow reforecast training, the statistical post-processing ended up overcorrecting the spread. This was found to be especially problematic in the case of dry days when it was nearly certain that no snowfall would occur, but still the post-processed forecast indicated a small probability of snowfall. Second, because of the short training length of the real-time snow forecast, the impact of seasonal differences was found to be significant. In this case, as the training period for the statistical post-processing was significantly drier than the verification period, the statistical post-processing did not perform well with higher snowfall events.</p>
      <p id="d1e2952">An ideal training forecast was identified as being as homogeneous as possible with the operational forecast and as having a long training length. However, such a dataset was not available in our case, and before it becomes available, it is difficult to decide whether an operational application of post-processing should be based on the snow reforecast or on the real-time snow forecast since both have advantages and disadvantages. The possibility of initializing an incoming version of PEARP reforecast with an ensemble of initial states coming for instance from ERA5 reanalyses should be investigated in the future. This should reduce the discrepancy with the operational ensemble system and encourage post-processing based on the reforecast rather than on real-time forecasts. The main limitation remains the high computational time consumption of these reforecasts <xref ref-type="bibr" rid="bib1.bibx67" id="paren.85"/> and finding the balance with the frequency of operational changes in NWP and snowpack modelling systems.</p>
</sec>

      
      </body>
    <back><notes notes-type="codedataavailability"><title>Code and data availability</title>

      <p id="d1e2962">The R code used for post-processing was originally developed by Michael Scheuerer (Cooperative Institute for Research in Environmental Sciences – University of Colorado Boulder – and NOAA Earth System Research Laboratory, Physical Sciences Division, Boulder-Colorado, USA). The modified version can be provided on request, with the agreement of the original author. The Crocus snowpack model is developed inside the open-source SURFEX project (<uri>http://www.umr-cnrm.fr/surfex/</uri>, last access: 23 September 2019). The most up-to-date version of the code can be downloaded from the specific branch of the git repository maintained by the Centre d'Études de la Neige. For reproducibility of results, the version used in this work is tagged as “s2m_reanalysis_2018” on the SURFEX git repository (git.umr-cnrm.fr/git/Surfex_Git2.git, last access: 23 September 2019). The full procedure and documentation to access this git repository can be found at <uri>https://opensource.cnrm-game-meteo.fr/projects/snowtools_git/wiki</uri> (last access: 23 September 2019). The codes of PEARP and SAFRAN are not currently open source. For reproducibility of results, the PEARP version used in this study is “cy42_peace-op2.18”, and the SAFRAN version is tagged as “reforecast_2018” in the private SAFRAN git repository. The raw data of HN forecasts and reforecasts of the PEARP-S2M system can be obtained on request. The HN observations used in this work are public data available at <uri>https://donneespubliques.meteofrance.fr</uri> (last access: 23 September 2019).</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d1e2977">BJ developed and ran the PEARP reforecast. MV developed and ran the SAFRAN downscaling of the PEARP reforecast and real-time forecasts. ML developed and ran the SURFEX-Crocus snowpack simulations forced by PEARP-SAFRAN outputs and supervised the study. JPN set up the statistical framework, with scientific contributions of JB and GE. JPN and ML produced the figures and wrote the publication, with contributions of all the authors.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d1e2983">The authors declare that they have no conflict of interest.</p>
  </notes><notes notes-type="sistatement"><title>Special issue statement</title>

      <p id="d1e2989">This article is part of the special issue “Advances in post-processing and blending of deterministic and ensemble forecasts”. It is not associated with a conference.</p>
  </notes><ack><title>Acknowledgements</title><p id="d1e2995">The authors would like to thank Michael Scheuerer for providing the initial code of the EMOS-CSGD, Maxime Taillardat (Météo-France, DirOP/COMPAS) and Marie Dumont (same affiliation as ML and MV) for useful discussions and comments relevant to this work, Frédéric Sassier (Météo-France, DCSC/AVH) for providing Fig. <xref ref-type="fig" rid="Ch1.F8"/>, and César Deschamps-Berger for his help on Fig. <xref ref-type="fig" rid="Ch1.F1"/>. They also thank the two anonymous referees and the editor Stephan Hemri for their review of the manuscript.<?pagebreak page355?> CNRM/CEN, IGE and IRSTEA are part of LabEX OSUG@2020 (ANR10 LABX56).</p></ack><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d1e3004">This paper was edited by Stephan Hemri and reviewed by two anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>Baran and Lerch(2015)</label><?label baran2015?><mixed-citation>Baran, S. and Lerch, S.: Log-normal distribution based Ensemble Model Output
Statistics models for probabilistic wind-speed forecasting, Q. J. Roy.
Meteorol. Soc., 141, 2289–2299, <ext-link xlink:href="https://doi.org/10.1002/qj.2521" ext-link-type="DOI">10.1002/qj.2521</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Baran and Lerch(2016)</label><?label baran2016a?><mixed-citation>Baran, S. and Lerch, S.: Mixture EMOS model for calibrating ensemble forecasts
of wind speed, Environmetrics, 27, 116–130, <ext-link xlink:href="https://doi.org/10.1002/env.2380" ext-link-type="DOI">10.1002/env.2380</ext-link>,
2016.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Baran and Nemoda(2016)</label><?label baran2016b?><mixed-citation>Baran, S. and Nemoda, D.: Censored and shifted gamma distribution based EMOS
model for probabilistic quantitative precipitation forecasting,
Environmetrics, 27, 280–292, <ext-link xlink:href="https://doi.org/10.1002/env.2391" ext-link-type="DOI">10.1002/env.2391</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>Barkmeijer et al.(1998)</label><?label barkmeijer1998?><mixed-citation>Barkmeijer, J., Van Gijzen, M., and Bouttier, F.: Singular vectors and
estimates of the analysis-error covariance metric, Q. J. Roy. Meteorol.
Soc., 124, 1695–1713, <ext-link xlink:href="https://doi.org/10.1256/smsqj.54915" ext-link-type="DOI">10.1256/smsqj.54915</ext-link>, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Barkmeijer et al.(1999)</label><?label barkmeijer1999?><mixed-citation>Barkmeijer, J., Buizza, R., and Palmer, T.: 3D-Var Hessian singular vectors
and their potential use in the ECMWF Ensemble Prediction System, Q. J.
Roy. Meteorol. Soc., 125, 2333–2351, <ext-link xlink:href="https://doi.org/10.1256/smsqj.55817" ext-link-type="DOI">10.1256/smsqj.55817</ext-link>,
1999.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Bellier et al.(2017)</label><?label bellier2017?><mixed-citation>Bellier, J., Zin, I., and Bontron, G.: Sample Stratification in Verification
of Ensemble Forecasts of Continuous Scalar Variables: Potential Benefits and
Pitfalls, Mon. Weather Rev., 145, 3529–3544,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-16-0487.1" ext-link-type="DOI">10.1175/MWR-D-16-0487.1</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Boisserie et al.(2016)</label><?label boisserie2015?><mixed-citation>Boisserie, M., Decharme, B., Descamps, L., and Arbogast, P.: Land surface
initialization strategy for a global reforecast dataset, Q. J. Roy.
Meteorol. Soc., 142, 880–888, <ext-link xlink:href="https://doi.org/10.1002/qj.2688" ext-link-type="DOI">10.1002/qj.2688</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Broecker(2008)</label><?label brocker2008?><mixed-citation>Bröcker, J.: On reliability analysis of multi-categorical forecasts, Nonlin. Processes Geophys., 15, 661–673, <ext-link xlink:href="https://doi.org/10.5194/npg-15-661-2008" ext-link-type="DOI">10.5194/npg-15-661-2008</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Brun et al.(1992)</label><?label brun1992?><mixed-citation>Brun, E., David, P., Sudul, M., and Brunot, G.: A numerical model to simulate
snow-cover stratigraphy for operational avalanche forecasting, J. Glaciol.,
38, 13–22, <ext-link xlink:href="https://doi.org/10.3189/S0022143000009552" ext-link-type="DOI">10.3189/S0022143000009552</ext-link>, 1992.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>Buizza and Palmer(1995)</label><?label buizza1995?><mixed-citation>Buizza, R. and Palmer, T.: The singular-vector structure of the atmospheric
global circulation, J. Atmos. Sci., 52, 1434–1456,
<ext-link xlink:href="https://doi.org/10.1175/1520-0469(1995)052&lt;1434:TSVSOT&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0469(1995)052&lt;1434:TSVSOT&gt;2.0.CO;2</ext-link>, 1995.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Buizza et al.(2003)</label><?label buizza2003?><mixed-citation>Buizza, R., Richardson, D., and Palmer, T.: Benefits of increased resolution
in the ECMWF ensemble system and comparison with poor-man's ensembles,
Q. J. Roy. Meteorol. Soc., 129, 1269–1288, <ext-link xlink:href="https://doi.org/10.1256/qj.02.92" ext-link-type="DOI">10.1256/qj.02.92</ext-link>,
2003.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Candille and Talagrand(2005)</label><?label candille2005?><mixed-citation>Candille, G. and Talagrand, O.: Evaluation of probabilistic prediction systems
for a scalar variable, Q. J. Roy. Meteorol. Soc., 131, 2131–2150,
<ext-link xlink:href="https://doi.org/10.1256/qj.04.71" ext-link-type="DOI">10.1256/qj.04.71</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Carmagnola et al.(2014)</label><?label carmagnola2014?><mixed-citation>Carmagnola, C. M., Morin, S., Lafaysse, M., Domine, F., Lesaffre, B., Lejeune, Y., Picard, G., and Arnaud, L.: Implementation and evaluation of prognostic representations of the optical diameter of snow in the SURFEX/ISBA-Crocus detailed snowpack model, The Cryosphere, 8, 417–437, <ext-link xlink:href="https://doi.org/10.5194/tc-8-417-2014" ext-link-type="DOI">10.5194/tc-8-417-2014</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Champavier et al.(2018)</label><?label champavier2018?><mixed-citation>Champavier, R., Lafaysse, M., Vernay, M., and Coléou, C.: Comparison of
various forecast products of height of new snow in 24 hours on French ski
resorts at different lead times, in: Proceedings of the International Snow
Science Workshop – Innsbruck, Austria,  1150–1155,
available at: <uri>http://arc.lib.montana.edu/snow-science/objects/ISSW2018_O12.11.pdf</uri> (last access: 23 September 2019),
2018.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Dee et al.(2011)</label><?label dee2011?><mixed-citation>Dee, D. P., Uppala, S. M., Simmons, A. J., Berrisford, P., Poli, P., Kobayashi,
S., Andrae, U., Balmaseda, M. A., Balsamo, G., Bauer, P., Bechtold, P.,
Beljaars, A. C. M., van de Berg, L., Bidlot, J., Bormann, N., Delsol, C.,
Dragani, R., Fuentes, M., Geer, A. J., Haimberger, L., Healy, S. B.,
Hersbach, H., Holm, E. V., Isaksen, L., Kallberg, P., Koehler, M.,
Matricardi, M., McNally, A. P., Monge-Sanz, B. M., Morcrette, J. J., Park,
B. K., Peubey, C., de Rosnay, P., Tavolato, C., Thepaut, J. N., and Vitart,
F.: The ERA-Interim reanalysis: configuration and performance of the data
assimilation system, Q. J. Roy. Meteorol. Soc., 137, 553–597,
<ext-link xlink:href="https://doi.org/10.1002/qj.828" ext-link-type="DOI">10.1002/qj.828</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Descamps et al.(2015)</label><?label descamps2014?><mixed-citation>Descamps, L., Labadie, C., Joly, A., Bazile, E., Arbogast, P., and Cébron,
P.: PEARP, the Météo-France short-range ensemble prediction system,
Q. J. Roy. Meteorol. Soc., 141, 1671–1685, <ext-link xlink:href="https://doi.org/10.1002/qj.2469" ext-link-type="DOI">10.1002/qj.2469</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Douville et al.(1995)</label><?label douville1995?><mixed-citation>Douville, H., Royer, J.-F., and Mahfouf, J.-F.: A new snow parameterization for
the Meteo-France climate model, Part I: Validation in stand-alone
experiments, Clim. Dyn., 12, 21–35, <ext-link xlink:href="https://doi.org/10.1007/BF00208760" ext-link-type="DOI">10.1007/BF00208760</ext-link>, 1995.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>Durand et al.(1993)</label><?label durand1993?><mixed-citation>
Durand, Y., Brun, E., Mérindol, L., Guyomarc'h, G., Lesaffre, B., and
Martin, E.: A meteorological estimation of relevant parameters for snow
models, Ann. Glaciol., 18, 65–71, 1993.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Durand et al.(2009)</label><?label durand2009a?><mixed-citation>Durand, Y., Giraud, G., Laternser, M., Etchevers, P., Mérindol, L., and
Lesaffre, B.: Reanalysis of 47 Years of Climate in the French Alps
(1958–2005): Climatology and Trends for Snow Cover, J. Appl. Meteorol.
Climatol., 48, 2487–2512, <ext-link xlink:href="https://doi.org/10.1175/2009JAMC1810.1" ext-link-type="DOI">10.1175/2009JAMC1810.1</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Durand et al.(1998)</label><?label durand1998?><mixed-citation>
Durand, Y., Giraud, G., and Merindol, L.: Short-term numerical avalanche
forecast used operationally at Meteo-France over the Alps and Pyrenees, Ann.
Glaciol., 26, 357–366, International Symposium on Snow and Avalanches,
Chamonix Mont Blanc, France,  26–30 May, 1997, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Durand et al.(1999)</label><?label durand1999?><mixed-citation>Durand, Y., Giraud, G., Brun, E., Merindol, L., and Martin, E.: A
computer-based system simulating snowpack structures as a tool for regional
avalanche forecasting, J. Glaciol., 45, 469–484,
<ext-link xlink:href="https://doi.org/10.3189/S0022143000001337" ext-link-type="DOI">10.3189/S0022143000001337</ext-link>, 1999.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Essery et al.(2013)</label><?label essery2013?><mixed-citation>Essery, R., Morin, S., Lejeune, Y., and Bauduin-Ménard, C.: A comparison of
1701 snow models using observations from an alpine site, Adv. Water Res., 55,
131–148, <ext-link xlink:href="https://doi.org/10.1016/j.advwatres.2012.07.013" ext-link-type="DOI">10.1016/j.advwatres.2012.07.013</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>Fierz et al.(2009)</label><?label fierz2009?><mixed-citation>
Fierz, C., Armstrong, R. L., Durand, Y., Etchevers, P., Greene, E., McClung,
D. M., Nishimura, K., Satyawali, P. K., and Sokratov, S. A.: The
international classification for seasonal snow on the ground, IHP-VII
Technical Documents in Hydrology n 83, IACS Contribution n 1, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Fortin et al.(2006)</label><?label fortin2006?><mixed-citation>Fortin, V., Favre, A.-C., and Said, M.: Probabilistic forecasting from
ensemble prediction systems: Improving upon the best-member method by using a
different weight and dressing kernel for each member, Q. J. Roy.
Meteorol. Soc., 132, 1349–1369, <ext-link xlink:href="https://doi.org/10.1256/qj.05.167" ext-link-type="DOI">10.1256/qj.05.167</ext-link>, 2006.</mixed-citation></ref>
      <?pagebreak page356?><ref id="bib1.bibx25"><label>Gebetsberger et al.(2017)</label><?label gebetsberger2017?><mixed-citation>Gebetsberger, M., Messner, J. W., Mayr, G. J., and Zeileis, A.: Fine-Tuning
Nonhomogeneous Regression for Probabilistic Precipitation Forecasts:
Unanimous Predictions, Heavy Tails, and Link Functions, Mon. Weather Rev.,
145, 4693–4708, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-16-0388.1" ext-link-type="DOI">10.1175/MWR-D-16-0388.1</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Gebetsberger et al.(2018)</label><?label gebetsberger2018?><mixed-citation>Gebetsberger, M., Messner, J. W., Mayr, G. J., and Zeileis, A.: Estimation
Methods for Nonhomogeneous Regression Models: Minimum Continuous Ranked
Probability Score versus Maximum Likelihood, Mon. Weather Rev., 146,
4323–4338, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-17-0364.1" ext-link-type="DOI">10.1175/MWR-D-17-0364.1</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>Glahn and Lowry(1972)</label><?label glahn1972?><mixed-citation>Glahn, H. and Lowry, D.: The Use of Model Output Statistics (MOS) in Objective
Weather Forecasting, J. Appl. Meteorol., 11, 1203–1211,
<ext-link xlink:href="https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2</ext-link>, 1972.</mixed-citation></ref>
      <ref id="bib1.bibx28"><label>Gneiting et al.(2005)</label><?label gneiting2005?><mixed-citation>Gneiting, T., Raftery, A., Westveld, A., and Goldman, T.: Calibrated
probabilistic forecasting using ensemble model output statistics and minimum
CRPS estimation, Mon. Weather Rev., 133, 1098–1118,
<ext-link xlink:href="https://doi.org/10.1175/MWR2904.1" ext-link-type="DOI">10.1175/MWR2904.1</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx29"><label>Hamill(2001)</label><?label hamill2001?><mixed-citation>Hamill, T.: Interpretation of rank histograms for verifying ensemble
forecasts, Mon. Weather Rev., 129, 550–560,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(2001)129&lt;0550:IORHFV&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2001)129&lt;0550:IORHFV&gt;2.0.CO;2</ext-link>, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Hamill and Colucci(1997)</label><?label hamill1997?><mixed-citation>Hamill, T. and Colucci, S.: Verification of Eta–RSM Short-Range Ensemble
Forecasts, Mon. Weather Rev., 125, 1312–1327,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(1997)125&lt;1312:VOERSR&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(1997)125&lt;1312:VOERSR&gt;2.0.CO;2</ext-link>, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Hamill et al.(2000)</label><?label hamill2000?><mixed-citation>Hamill, T., Snyder, C., and Morss, R.: A comparison of probabilistic forecasts
from bred, singular-vector, and perturbed observation ensembles, Mon.
Weather Rev., 128, 1835–1851,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(2000)128&lt;1835:ACOPFF&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2000)128&lt;1835:ACOPFF&gt;2.0.CO;2</ext-link>, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Hamill et al.(2003)</label><?label hamill2003?><mixed-citation>Hamill, T., Snyder, C., and Whitaker, J.: Ensemble forecasts and the
properties of flow-dependent analysis-error covariance singular vectors,
Mon. Weather Rev., 131, 1741–1758, <ext-link xlink:href="https://doi.org/10.1175//2559.1" ext-link-type="DOI">10.1175//2559.1</ext-link>, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Hamill et al.(2004)</label><?label hamill2004?><mixed-citation>Hamill, T., Whitaker, J., and Wei, X.: Ensemble reforecasting: Improving
medium-range forecast skill using retrospective forecasts, Mon. Weather
Rev., 132, 1434–1447,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(2004)132&lt;1434:ERIMFS&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2004)132&lt;1434:ERIMFS&gt;2.0.CO;2</ext-link>, 2004.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Hamill and Whitaker(2006)</label><?label hamill2006?><mixed-citation>Hamill, T. M. and Whitaker, J. S.: Probabilistic quantitative precipitation
forecasts based on reforecast analogs: Theory and application, Mon. Weather
Rev., 134, 3209–3229, <ext-link xlink:href="https://doi.org/10.1175/MWR3237.1" ext-link-type="DOI">10.1175/MWR3237.1</ext-link>, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>Helfricht et al.(2018)</label><?label helfricht2018?><mixed-citation>Helfricht, K., Hartl, L., Koch, R., Marty, C., and Olefs, M.: Obtaining sub-daily new snow density from automated measurements in high mountain regions, Hydrol. Earth Syst. Sci., 22, 2655–2668, <ext-link xlink:href="https://doi.org/10.5194/hess-22-2655-2018" ext-link-type="DOI">10.5194/hess-22-2655-2018</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>Jewson et al.(2004)</label><?label jewson2004?><mixed-citation>Jewson, S., Brix, A., and Ziehmann, C.: A new parametric model for the
assessment and calibration of medium-range ensemble temperature forecasts,
Atmos. Sci. Lett., 5, 96–102, <ext-link xlink:href="https://doi.org/10.1002/asl.69" ext-link-type="DOI">10.1002/asl.69</ext-link>, 2004.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Kochendorfer et al.(2017)</label><?label kochendorfer2017?><mixed-citation>Kochendorfer, J., Rasmussen, R., Wolff, M., Baker, B., Hall, M. E., Meyers, T., Landolt, S., Jachcik, A., Isaksen, K., Brækkan, R., and Leeper, R.: The quantification and correction of wind-induced precipitation measurement errors, Hydrol. Earth Syst. Sci., 21, 1973–1989, <ext-link xlink:href="https://doi.org/10.5194/hess-21-1973-2017" ext-link-type="DOI">10.5194/hess-21-1973-2017</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx38"><label>Krinner et al.(2018)</label><?label krinner2018?><mixed-citation>Krinner, G., Derksen, C., Essery, R., Flanner, M., Hagemann, S., Clark, M., Hall, A., Rott, H., Brutel-Vuilmet, C., Kim, H., Ménard, C. B., Mudryk, L., Thackeray, C., Wang, L., Arduini, G., Balsamo, G., Bartlett, P., Boike, J., Boone, A., Chéruy, F., Colin, J., Cuntz, M., Dai, Y., Decharme, B., Derry, J., Ducharne, A., Dutra, E., Fang, X., Fierz, C., Ghattas, J., Gusev, Y., Haverd, V., Kontu, A., Lafaysse, M., Law, R., Lawrence, D., Li, W., Marke, T., Marks, D., Ménégoz, M., Nasonova, O., Nitta, T., Niwano, M., Pomeroy, J., Raleigh, M. S., Schaedler, G., Semenov, V., Smirnova, T. G., Stacke, T., Strasser, U., Svenson, S., Turkov, D., Wang, T., Wever, N., Yuan, H., Zhou, W., and Zhu, D.: ESM-SnowMIP: assessing snow models and quantifying snow-related climate feedbacks, Geosci. Model Dev., 11, 5027–5049, <ext-link xlink:href="https://doi.org/10.5194/gmd-11-5027-2018" ext-link-type="DOI">10.5194/gmd-11-5027-2018</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx39"><label>Lafaysse et al.(2017)</label><?label lafaysse2017?><mixed-citation>Lafaysse, M., Cluzet, B., Dumont, M., Lejeune, Y., Vionnet, V., and Morin, S.: A multiphysical ensemble system of numerical snow modelling, The Cryosphere, 11, 1173–1198, <ext-link xlink:href="https://doi.org/10.5194/tc-11-1173-2017" ext-link-type="DOI">10.5194/tc-11-1173-2017</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx40"><label>Lehning et al.(2002)</label><?label lehning2002?><mixed-citation>Lehning, M., Bartelt, P., Brown, B., Fierz, C., and Satyawali, P.: A physical
SNOWPACK model for the Swiss avalanche warning, Part
II: snow microstructure., Cold Reg. Sci. Technol., 35, 147–167,
<ext-link xlink:href="https://doi.org/10.1016/S0165-232X(02)00073-3" ext-link-type="DOI">10.1016/S0165-232X(02)00073-3</ext-link>, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx41"><label>Lerch and Thorarinsdottir(2013)</label><?label lerch2013?><mixed-citation>Lerch, S. and Thorarinsdottir, T. L.: Comparison of non-homogeneous regression
models for probabilistic wind speed forecasting, Tellus A, 65, 21206,
<ext-link xlink:href="https://doi.org/10.3402/tellusa.v65i0.21206" ext-link-type="DOI">10.3402/tellusa.v65i0.21206</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx42"><label>Leutbecher and Palmer(2008)</label><?label leutbecher2008?><mixed-citation>Leutbecher, M. and Palmer, T. N.: Ensemble forecasting, J. Comput. Phys.,
227, 3515–3539, <ext-link xlink:href="https://doi.org/10.1016/j.jcp.2007.02.014" ext-link-type="DOI">10.1016/j.jcp.2007.02.014</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx43"><label>Masson et al.(2013)</label><?label masson2013?><mixed-citation>Masson, V., Le Moigne, P., Martin, E., Faroux, S., Alias, A., Alkama, R., Belamari, S., Barbu, A., Boone, A., Bouyssel, F., Brousseau, P., Brun, E., Calvet, J.-C., Carrer, D., Decharme, B., Delire, C., Donier, S., Essaouini, K., Gibelin, A.-L., Giordani, H., Habets, F., Jidane, M., Kerdraon, G., Kourzeneva, E., Lafaysse, M., Lafont, S., Lebeaupin Brossier, C., Lemonsu, A., Mahfouf, J.-F., Marguinaud, P., Mokhtari, M., Morin, S., Pigeon, G., Salgado, R., Seity, Y., Taillefer, F., Tanguy, G., Tulet, P., Vincendon, B., Vionnet, V., and Voldoire, A.: The SURFEXv7.2 land and ocean surface platform for coupled or offline simulation of earth surface variables and fluxes, Geosci. Model Dev., 6, 929–960, <ext-link xlink:href="https://doi.org/10.5194/gmd-6-929-2013" ext-link-type="DOI">10.5194/gmd-6-929-2013</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx44"><label>Messner et al.(2014)</label><?label messner2014?><mixed-citation>Messner, J. W., Mayr, G. J., Wilks, D. S., and Zeileis, A.: Extending Extended
Logistic Regression: Extended versus Separate versus Ordered versus
Censored, Mon. Weather Rev., 142, 3003–3014,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-13-00355.1" ext-link-type="DOI">10.1175/MWR-D-13-00355.1</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx45"><label>Molteni et al.(1996)</label><?label molteni1996?><mixed-citation>Molteni, F., Buizza, R., Palmer, T., and Petroliagis, T.: The ECMWF ensemble
prediction system: Methodology and validation, Q. J. Roy. Meteorol.
Soc., 122, 73–119, <ext-link xlink:href="https://doi.org/10.1002/qj.49712252905" ext-link-type="DOI">10.1002/qj.49712252905</ext-link>, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx46"><label>Mullen and Buizza(2002)</label><?label mullen2002?><mixed-citation>Mullen, S. and Buizza, R.: The impact of horizontal resolution and ensemble
size on probabilistic forecasts of precipitation by the ECMWF Ensemble
Prediction System, Weather Forecast., 17, 173–191,
<ext-link xlink:href="https://doi.org/10.1175/1520-0434(2002)017&lt;0173:TIOHRA&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0434(2002)017&lt;0173:TIOHRA&gt;2.0.CO;2</ext-link>, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx47"><label>Pahaut(1975)</label><?label pahaut1975?><mixed-citation>
Pahaut, E.: La métamorphose des cristaux de neige (Snow crystal
metamorphosis), vol. 96 of Monographies de la Météorologie
Nationale, Météo France, 1975.</mixed-citation></ref>
      <ref id="bib1.bibx48"><label>Palmer(2001)</label><?label palmer2001?><mixed-citation>Palmer, T.: A nonlinear dynamical perspective on model error: A proposal for
non-local stochastic-dynamic parametr<?pagebreak page357?>ization in weather and climate
prediction models, Q. J. Roy. Meteorol. Soc., 127, 279–304,
<ext-link xlink:href="https://doi.org/10.1002/qj.49712757202" ext-link-type="DOI">10.1002/qj.49712757202</ext-link>, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx49"><label>Pellerin et al.(2003)</label><?label pellerin2003?><mixed-citation>Pellerin, G., Lefaivre, L., Houtekamer, P., and Girard, C.: Increasing the horizontal resolution of ensemble forecasts at CMC, Nonlin. Processes Geophys., 10, 463–468, <ext-link xlink:href="https://doi.org/10.5194/npg-10-463-2003" ext-link-type="DOI">10.5194/npg-10-463-2003</ext-link>, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx50"><label>Raftery et al.(2005)</label><?label raftery2005?><mixed-citation>Raftery, A., Gneiting, T., Balabdaoui, F., and Polakowski, M.: Using Bayesian
model averaging to calibrate forecast ensembles, Mon. Weather Rev., 133,
1155–1174, <ext-link xlink:href="https://doi.org/10.1175/MWR2906.1" ext-link-type="DOI">10.1175/MWR2906.1</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx51"><label>Ramos et al.(2013)</label><?label ramos2013?><mixed-citation>Ramos, M. H., van Andel, S. J., and Pappenberger, F.: Do probabilistic forecasts lead to better decisions?, Hydrol. Earth Syst. Sci., 17, 2219–2232, <ext-link xlink:href="https://doi.org/10.5194/hess-17-2219-2013" ext-link-type="DOI">10.5194/hess-17-2219-2013</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx52"><label>Richardson(2000)</label><?label richardson2000?><mixed-citation>Richardson, D.: Skill and relative economic value of the ECMWF ensemble
prediction system, Q. J. Roy. Meteorol. Soc., 126, 649–667,
<ext-link xlink:href="https://doi.org/10.1256/smsqj.56312" ext-link-type="DOI">10.1256/smsqj.56312</ext-link>, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx53"><label>Roulston and Smith(2002)</label><?label roulston2002?><mixed-citation>Roulston, M. and Smith, L.: Evaluating probabilistic forecasts using
information theory, Mon. Weather Rev., 130, 1653–1660,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(2002)130&lt;1653:EPFUIT&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2002)130&lt;1653:EPFUIT&gt;2.0.CO;2</ext-link>, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx54"><label>Scheuerer(2014)</label><?label scheuerer2014?><mixed-citation>Scheuerer, M.: Probabilistic quantitative precipitation forecasting using
Ensemble Model Output Statistics, Q. J. Roy. Meteorol. Soc., 140,
1086–1096, <ext-link xlink:href="https://doi.org/10.1002/qj.2183" ext-link-type="DOI">10.1002/qj.2183</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx55"><label>Scheuerer and Hamill(2015)</label><?label scheuerer2015?><mixed-citation>Scheuerer, M. and Hamill, T. M.: Statistical Postprocessing of Ensemble
Precipitation Forecasts by Fitting Censored, Shifted Gamma Distributions,
Mon. Weather Rev., 143, 4578–4596, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-15-0061.1" ext-link-type="DOI">10.1175/MWR-D-15-0061.1</ext-link>,
2015.</mixed-citation></ref>
      <ref id="bib1.bibx56"><label>Scheuerer and Hamill(2018)</label><?label scheuerer2018?><mixed-citation>Scheuerer, M. and Hamill, T. M.: Generating Calibrated Ensembles of Physically
Realistic, High-Resolution Precipitation Forecast Fields Based on GEFS Model
Output, J. Hydrometeorol., 19, 1651–1670,
<ext-link xlink:href="https://doi.org/10.1175/JHM-D-18-0067.1" ext-link-type="DOI">10.1175/JHM-D-18-0067.1</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx57"><label>Scheuerer and Hamill(2019)</label><?label scheuerer2019?><mixed-citation>Scheuerer, M. and Hamill, T. M.: Probabilistic Forecasting of Snowfall Amounts
Using a Hybrid between a Parametric and an Analog Approach, Mon. Weather
Rev., 147, 1047–1064, <ext-link xlink:href="https://doi.org/10.1175/MWR-D-18-0273.1" ext-link-type="DOI">10.1175/MWR-D-18-0273.1</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx58"><label>Schleef et al.(2014)</label><?label schleef2014?><mixed-citation>Schleef, S., Löwe, H., and Schneebeli, M.: Influence of stress, temperature and crystal morphology on isothermal densification and specific surface area decrease of new snow, The Cryosphere, 8, 1825–1838, <ext-link xlink:href="https://doi.org/10.5194/tc-8-1825-2014" ext-link-type="DOI">10.5194/tc-8-1825-2014</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx59"><label>Stauffer et al.(2018)</label><?label stauffer2018?><mixed-citation>Stauffer, R., Mayr, G. J., Messner, J. W., and Zeileis, A.: Hourly
probabilistic snow forecasts over complex terrain: a hybrid ensemble
postprocessing approach, Adv. Stat. Climatol. Meteorol. Oceanogr., 4, 65–86,
<ext-link xlink:href="https://doi.org/10.5194/ascmo-4-65-2018" ext-link-type="DOI">10.5194/ascmo-4-65-2018</ext-link>,
2018.</mixed-citation></ref>
      <ref id="bib1.bibx60"><label>Sutton et al.(2006)</label><?label sutton2006?><mixed-citation>Sutton, C., Hamill, T. M., and Warner, T. T.: Will perturbing soil moisture
improve warm-season ensemble forecasts? A proof of concept, Mon. Weather
Rev., 134, 3174–3189, <ext-link xlink:href="https://doi.org/10.1175/MWR3248.1" ext-link-type="DOI">10.1175/MWR3248.1</ext-link>, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx61"><label>Szunyogh and Toth(2002)</label><?label szynyough2002?><mixed-citation>Szunyogh, I. and Toth, Z.: The effect of increased horizontal resolution on
the NCEP global ensemble mean forecasts, Mon. Weather Rev., 130,
1125–1143, <ext-link xlink:href="https://doi.org/10.1175/1520-0493(2002)130&lt;1125:TEOIHR&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2002)130&lt;1125:TEOIHR&gt;2.0.CO;2</ext-link>,
2002.
</mixed-citation></ref><?xmltex \hack{\newpage}?>
      <ref id="bib1.bibx62"><label>Taillardat et al.(2016)</label><?label taillardat2016?><mixed-citation>Taillardat, M., Mestre, O., Zamo, M., and Naveau, P.: Calibrated Ensemble
Forecasts Using Quantile Regression Forests and Ensemble Model Output
Statistics, Mon. Weather Rev., 144, 2375–2393,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-15-0260.1" ext-link-type="DOI">10.1175/MWR-D-15-0260.1</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx63"><label>Taillardat et al.(2019)</label><?label taillardat2019?><mixed-citation>Taillardat, M., Fougères, A., Naveau, P., and Mestre, O.: Forest-based and
semi-parametric methods for the postprocessing of rainfall ensemble
forecasting, Weather Forecast., in press, <ext-link xlink:href="https://doi.org/10.1175/WAF-D-18-0149.1" ext-link-type="DOI">10.1175/WAF-D-18-0149.1</ext-link>,
2019.</mixed-citation></ref>
      <ref id="bib1.bibx64"><label>Teufelsbauer(2011)</label><?label teufelsbauer2011?><mixed-citation>Teufelsbauer, H.: A two-dimensional snow creep model for alpine terrain,
Nat. Hazards, 56, 481–497, <ext-link xlink:href="https://doi.org/10.1007/s11069-010-9515-8" ext-link-type="DOI">10.1007/s11069-010-9515-8</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx65"><label>Thorarinsdottir and Gneiting(2010)</label><?label thorarinsdottir2010?><mixed-citation>Thorarinsdottir, T. L. and Gneiting, T.: Probabilistic forecasts of wind
speed: ensemble model output statistics by using heteroscedastic censored
regression, J. Roy. Statist. Soc., 173, 371–388,
<ext-link xlink:href="https://doi.org/10.1111/j.1467-985X.2009.00616.x" ext-link-type="DOI">10.1111/j.1467-985X.2009.00616.x</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx66"><label>Toth and Kalnay(1997)</label><?label toth1997?><mixed-citation>Toth, Z. and Kalnay, E.: Ensemble forecasting at NCEP and the breeding
method, Mon. Weather Rev., 125, 3297–3319,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(1997)125&lt;3297:EFANAT&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(1997)125&lt;3297:EFANAT&gt;2.0.CO;2</ext-link>, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx67"><label>Vannitsem et al.(2018)</label><?label vannitsem2018?><mixed-citation>Vannitsem, S., Wilks, D. S., and Messner, J. W.: Statistical postprocessing of
ensemble forecasts, 1 edn., Elsevier, Amsterdam, <ext-link xlink:href="https://doi.org/10.1016/C2016-0-03244-8" ext-link-type="DOI">10.1016/C2016-0-03244-8</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx68"><label>Vernay et al.(2015)</label><?label vernay2015?><mixed-citation>Vernay, M., Lafaysse, M., Merindol, L., Giraud, G., and Morin, S.: Ensemble
Forecasting of snowpack conditions and avalanche hazard, Cold. Reg. Sci.
Technol., 120, 251–262, <ext-link xlink:href="https://doi.org/10.1016/j.coldregions.2015.04.010" ext-link-type="DOI">10.1016/j.coldregions.2015.04.010</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx69"><label>Vionnet et al.(2012)</label><?label vionnet2012?><mixed-citation>Vionnet, V., Brun, E., Morin, S., Boone, A., Faroux, S., Le Moigne, P., Martin, E., and Willemet, J.-M.: The detailed snowpack scheme Crocus and its implementation in SURFEX v7.2, Geosci. Model Dev., 5, 773–791, <ext-link xlink:href="https://doi.org/10.5194/gmd-5-773-2012" ext-link-type="DOI">10.5194/gmd-5-773-2012</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx70"><label>Wang and Bishop(2005)</label><?label wang2005?><mixed-citation>Wang, X. and Bishop, C.: Improvement of ensemble reliability with a new
dressing kernel, Q. J. Roy. Meteorol. Soc., 131, 965–986,
<ext-link xlink:href="https://doi.org/10.1256/qj.04.120" ext-link-type="DOI">10.1256/qj.04.120</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx71"><label>Weisman et al.(1997)</label><?label weisman1997?><mixed-citation>Weisman, M., Skamarock, W., and Klemp, J.: The resolution dependence of
explicitly modeled convective systems, Mon. Weather Rev., 125, 527–548,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(1997)125&lt;0527:TRDOEM&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(1997)125&lt;0527:TRDOEM&gt;2.0.CO;2</ext-link>, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx72"><label>Wilks(2005)</label><?label wilks2005?><mixed-citation>Wilks, D.: Effects of stochastic parametrizations in the Lorenz `96 system,
Q. J. Roy. Meteorol. Soc., 131, 389–407, <ext-link xlink:href="https://doi.org/10.1256/qj.04.03" ext-link-type="DOI">10.1256/qj.04.03</ext-link>,
2005.</mixed-citation></ref>
      <ref id="bib1.bibx73"><label>Wilks and Hamill(2007)</label><?label wilks2006?><mixed-citation>Wilks, D. S. and Hamill, T. M.: Comparison of ensemble-MOS methods using GFS
reforecasts, Mon. Weather Rev., 135, 2379–2390,
<ext-link xlink:href="https://doi.org/10.1175/MWR3402.1" ext-link-type="DOI">10.1175/MWR3402.1</ext-link>, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx74"><label>Wilks(2011)</label><?label wilks2011?><mixed-citation>
Wilks, D. S.: Statistical Methods in the Atmospheric Sciences,
3 edn., Academic Press, Amsterdam, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx75"><label>WMO(2018)</label><?label wmo2018?><mixed-citation>WMO: Preliminary 2018 Edition of the Guide to Meteorological Instruments and
Methods of Observation, Tech. Rep. 8, World Meteorological Organization,
available at: <uri>http://www.wmo.int/pages/prog/www/IMOP/publications/CIMO-Guide/Prelim_2018_ed/8_cryo_2_en_MR.pdf</uri> (last access: 23 September 2019),
cIMO Guide, Volume II, chapter 2, page 34, 2018.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Statistical post-processing of ensemble forecasts of the height of new snow</article-title-html>
<abstract-html><p>Forecasting the height of new snow (HN) is crucial for avalanche hazard forecasting, road viability, ski resort management and tourism attractiveness. Météo-France operates the PEARP-S2M probabilistic forecasting system, including 35 members of the PEARP Numerical Weather Prediction system, where the SAFRAN downscaling tool refines the elevation resolution and the Crocus snowpack model represents the main physical processes in the snowpack. It provides better HN forecasts than direct NWP diagnostics but exhibits significant biases and underdispersion. We applied a statistical post-processing to these ensemble forecasts, based on non-homogeneous regression with a censored shifted Gamma distribution. Observations come from manual measurements of 24&thinsp;h HN in the French Alps and Pyrenees. The calibration is tested at the station scale and the massif scale (i.e. aggregating different stations over areas of 1000&thinsp;km<sup>2</sup>). Compared to the raw forecasts, similar improvements are obtained for both spatial scales. Therefore, the post-processing can be applied at any point of the massifs. Two training datasets are tested: (1) a 22-year homogeneous reforecast for which the NWP model resolution and physical options are identical to the operational system but without the same initial perturbations; (2) 3-year real-time forecasts with a heterogeneous model configuration but the same perturbation methods.   The impact of the training dataset depends on lead time and on the evaluation criteria. The long-term reforecast improves the reliability of severe snowfall but leads to overdispersion due to the discrepancy in real-time perturbations. Thus, the development of reliable automatic forecasting products of HN needs long reforecasts as homogeneous as possible with the operational systems.</p></abstract-html>
<ref-html id="bib1.bib1"><label>Baran and Lerch(2015)</label><mixed-citation>
Baran, S. and Lerch, S.: Log-normal distribution based Ensemble Model Output
Statistics models for probabilistic wind-speed forecasting, Q. J. Roy.
Meteorol. Soc., 141, 2289–2299, <a href="https://doi.org/10.1002/qj.2521" target="_blank">https://doi.org/10.1002/qj.2521</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Baran and Lerch(2016)</label><mixed-citation>
Baran, S. and Lerch, S.: Mixture EMOS model for calibrating ensemble forecasts
of wind speed, Environmetrics, 27, 116–130, <a href="https://doi.org/10.1002/env.2380" target="_blank">https://doi.org/10.1002/env.2380</a>,
2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Baran and Nemoda(2016)</label><mixed-citation>
Baran, S. and Nemoda, D.: Censored and shifted gamma distribution based EMOS
model for probabilistic quantitative precipitation forecasting,
Environmetrics, 27, 280–292, <a href="https://doi.org/10.1002/env.2391" target="_blank">https://doi.org/10.1002/env.2391</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Barkmeijer et al.(1998)</label><mixed-citation>
Barkmeijer, J., Van Gijzen, M., and Bouttier, F.: Singular vectors and
estimates of the analysis-error covariance metric, Q. J. Roy. Meteorol.
Soc., 124, 1695–1713, <a href="https://doi.org/10.1256/smsqj.54915" target="_blank">https://doi.org/10.1256/smsqj.54915</a>, 1998.
</mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Barkmeijer et al.(1999)</label><mixed-citation>
Barkmeijer, J., Buizza, R., and Palmer, T.: 3D-Var Hessian singular vectors
and their potential use in the ECMWF Ensemble Prediction System, Q. J.
Roy. Meteorol. Soc., 125, 2333–2351, <a href="https://doi.org/10.1256/smsqj.55817" target="_blank">https://doi.org/10.1256/smsqj.55817</a>,
1999.
</mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Bellier et al.(2017)</label><mixed-citation>
Bellier, J., Zin, I., and Bontron, G.: Sample Stratification in Verification
of Ensemble Forecasts of Continuous Scalar Variables: Potential Benefits and
Pitfalls, Mon. Weather Rev., 145, 3529–3544,
<a href="https://doi.org/10.1175/MWR-D-16-0487.1" target="_blank">https://doi.org/10.1175/MWR-D-16-0487.1</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Boisserie et al.(2016)</label><mixed-citation>
Boisserie, M., Decharme, B., Descamps, L., and Arbogast, P.: Land surface
initialization strategy for a global reforecast dataset, Q. J. Roy.
Meteorol. Soc., 142, 880–888, <a href="https://doi.org/10.1002/qj.2688" target="_blank">https://doi.org/10.1002/qj.2688</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Broecker(2008)</label><mixed-citation>
Bröcker, J.: On reliability analysis of multi-categorical forecasts, Nonlin. Processes Geophys., 15, 661–673, <a href="https://doi.org/10.5194/npg-15-661-2008" target="_blank">https://doi.org/10.5194/npg-15-661-2008</a>, 2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Brun et al.(1992)</label><mixed-citation>
Brun, E., David, P., Sudul, M., and Brunot, G.: A numerical model to simulate
snow-cover stratigraphy for operational avalanche forecasting, J. Glaciol.,
38, 13–22, <a href="https://doi.org/10.3189/S0022143000009552" target="_blank">https://doi.org/10.3189/S0022143000009552</a>, 1992.
</mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Buizza and Palmer(1995)</label><mixed-citation>
Buizza, R. and Palmer, T.: The singular-vector structure of the atmospheric
global circulation, J. Atmos. Sci., 52, 1434–1456,
<a href="https://doi.org/10.1175/1520-0469(1995)052&lt;1434:TSVSOT&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0469(1995)052&lt;1434:TSVSOT&gt;2.0.CO;2</a>, 1995.
</mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Buizza et al.(2003)</label><mixed-citation>
Buizza, R., Richardson, D., and Palmer, T.: Benefits of increased resolution
in the ECMWF ensemble system and comparison with poor-man's ensembles,
Q. J. Roy. Meteorol. Soc., 129, 1269–1288, <a href="https://doi.org/10.1256/qj.02.92" target="_blank">https://doi.org/10.1256/qj.02.92</a>,
2003.
</mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Candille and Talagrand(2005)</label><mixed-citation>
Candille, G. and Talagrand, O.: Evaluation of probabilistic prediction systems
for a scalar variable, Q. J. Roy. Meteorol. Soc., 131, 2131–2150,
<a href="https://doi.org/10.1256/qj.04.71" target="_blank">https://doi.org/10.1256/qj.04.71</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Carmagnola et al.(2014)</label><mixed-citation>
Carmagnola, C. M., Morin, S., Lafaysse, M., Domine, F., Lesaffre, B., Lejeune, Y., Picard, G., and Arnaud, L.: Implementation and evaluation of prognostic representations of the optical diameter of snow in the SURFEX/ISBA-Crocus detailed snowpack model, The Cryosphere, 8, 417–437, <a href="https://doi.org/10.5194/tc-8-417-2014" target="_blank">https://doi.org/10.5194/tc-8-417-2014</a>, 2014.
</mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Champavier et al.(2018)</label><mixed-citation>
Champavier, R., Lafaysse, M., Vernay, M., and Coléou, C.: Comparison of
various forecast products of height of new snow in 24 hours on French ski
resorts at different lead times, in: Proceedings of the International Snow
Science Workshop – Innsbruck, Austria,  1150–1155,
available at: <a href="http://arc.lib.montana.edu/snow-science/objects/ISSW2018_O12.11.pdf" target="_blank">http://arc.lib.montana.edu/snow-science/objects/ISSW2018_O12.11.pdf</a> (last access: 23 September 2019),
2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Dee et al.(2011)</label><mixed-citation>
Dee, D. P., Uppala, S. M., Simmons, A. J., Berrisford, P., Poli, P., Kobayashi,
S., Andrae, U., Balmaseda, M. A., Balsamo, G., Bauer, P., Bechtold, P.,
Beljaars, A. C. M., van de Berg, L., Bidlot, J., Bormann, N., Delsol, C.,
Dragani, R., Fuentes, M., Geer, A. J., Haimberger, L., Healy, S. B.,
Hersbach, H., Holm, E. V., Isaksen, L., Kallberg, P., Koehler, M.,
Matricardi, M., McNally, A. P., Monge-Sanz, B. M., Morcrette, J. J., Park,
B. K., Peubey, C., de Rosnay, P., Tavolato, C., Thepaut, J. N., and Vitart,
F.: The ERA-Interim reanalysis: configuration and performance of the data
assimilation system, Q. J. Roy. Meteorol. Soc., 137, 553–597,
<a href="https://doi.org/10.1002/qj.828" target="_blank">https://doi.org/10.1002/qj.828</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Descamps et al.(2015)</label><mixed-citation>
Descamps, L., Labadie, C., Joly, A., Bazile, E., Arbogast, P., and Cébron,
P.: PEARP, the Météo-France short-range ensemble prediction system,
Q. J. Roy. Meteorol. Soc., 141, 1671–1685, <a href="https://doi.org/10.1002/qj.2469" target="_blank">https://doi.org/10.1002/qj.2469</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Douville et al.(1995)</label><mixed-citation>
Douville, H., Royer, J.-F., and Mahfouf, J.-F.: A new snow parameterization for
the Meteo-France climate model, Part I: Validation in stand-alone
experiments, Clim. Dyn., 12, 21–35, <a href="https://doi.org/10.1007/BF00208760" target="_blank">https://doi.org/10.1007/BF00208760</a>, 1995.
</mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Durand et al.(1993)</label><mixed-citation>
Durand, Y., Brun, E., Mérindol, L., Guyomarc'h, G., Lesaffre, B., and
Martin, E.: A meteorological estimation of relevant parameters for snow
models, Ann. Glaciol., 18, 65–71, 1993.
</mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Durand et al.(2009)</label><mixed-citation>
Durand, Y., Giraud, G., Laternser, M., Etchevers, P., Mérindol, L., and
Lesaffre, B.: Reanalysis of 47 Years of Climate in the French Alps
(1958–2005): Climatology and Trends for Snow Cover, J. Appl. Meteorol.
Climatol., 48, 2487–2512, <a href="https://doi.org/10.1175/2009JAMC1810.1" target="_blank">https://doi.org/10.1175/2009JAMC1810.1</a>, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Durand et al.(1998)</label><mixed-citation>
Durand, Y., Giraud, G., and Merindol, L.: Short-term numerical avalanche
forecast used operationally at Meteo-France over the Alps and Pyrenees, Ann.
Glaciol., 26, 357–366, International Symposium on Snow and Avalanches,
Chamonix Mont Blanc, France,  26–30 May, 1997, 1998.
</mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Durand et al.(1999)</label><mixed-citation>
Durand, Y., Giraud, G., Brun, E., Merindol, L., and Martin, E.: A
computer-based system simulating snowpack structures as a tool for regional
avalanche forecasting, J. Glaciol., 45, 469–484,
<a href="https://doi.org/10.3189/S0022143000001337" target="_blank">https://doi.org/10.3189/S0022143000001337</a>, 1999.
</mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Essery et al.(2013)</label><mixed-citation>
Essery, R., Morin, S., Lejeune, Y., and Bauduin-Ménard, C.: A comparison of
1701 snow models using observations from an alpine site, Adv. Water Res., 55,
131–148, <a href="https://doi.org/10.1016/j.advwatres.2012.07.013" target="_blank">https://doi.org/10.1016/j.advwatres.2012.07.013</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Fierz et al.(2009)</label><mixed-citation>
Fierz, C., Armstrong, R. L., Durand, Y., Etchevers, P., Greene, E., McClung,
D. M., Nishimura, K., Satyawali, P. K., and Sokratov, S. A.: The
international classification for seasonal snow on the ground, IHP-VII
Technical Documents in Hydrology n 83, IACS Contribution n 1, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Fortin et al.(2006)</label><mixed-citation>
Fortin, V., Favre, A.-C., and Said, M.: Probabilistic forecasting from
ensemble prediction systems: Improving upon the best-member method by using a
different weight and dressing kernel for each member, Q. J. Roy.
Meteorol. Soc., 132, 1349–1369, <a href="https://doi.org/10.1256/qj.05.167" target="_blank">https://doi.org/10.1256/qj.05.167</a>, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Gebetsberger et al.(2017)</label><mixed-citation>
Gebetsberger, M., Messner, J. W., Mayr, G. J., and Zeileis, A.: Fine-Tuning
Nonhomogeneous Regression for Probabilistic Precipitation Forecasts:
Unanimous Predictions, Heavy Tails, and Link Functions, Mon. Weather Rev.,
145, 4693–4708, <a href="https://doi.org/10.1175/MWR-D-16-0388.1" target="_blank">https://doi.org/10.1175/MWR-D-16-0388.1</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Gebetsberger et al.(2018)</label><mixed-citation>
Gebetsberger, M., Messner, J. W., Mayr, G. J., and Zeileis, A.: Estimation
Methods for Nonhomogeneous Regression Models: Minimum Continuous Ranked
Probability Score versus Maximum Likelihood, Mon. Weather Rev., 146,
4323–4338, <a href="https://doi.org/10.1175/MWR-D-17-0364.1" target="_blank">https://doi.org/10.1175/MWR-D-17-0364.1</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Glahn and Lowry(1972)</label><mixed-citation>
Glahn, H. and Lowry, D.: The Use of Model Output Statistics (MOS) in Objective
Weather Forecasting, J. Appl. Meteorol., 11, 1203–1211,
<a href="https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0450(1972)011&lt;1203:TUOMOS&gt;2.0.CO;2</a>, 1972.
</mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Gneiting et al.(2005)</label><mixed-citation>
Gneiting, T., Raftery, A., Westveld, A., and Goldman, T.: Calibrated
probabilistic forecasting using ensemble model output statistics and minimum
CRPS estimation, Mon. Weather Rev., 133, 1098–1118,
<a href="https://doi.org/10.1175/MWR2904.1" target="_blank">https://doi.org/10.1175/MWR2904.1</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Hamill(2001)</label><mixed-citation>
Hamill, T.: Interpretation of rank histograms for verifying ensemble
forecasts, Mon. Weather Rev., 129, 550–560,
<a href="https://doi.org/10.1175/1520-0493(2001)129&lt;0550:IORHFV&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2001)129&lt;0550:IORHFV&gt;2.0.CO;2</a>, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Hamill and Colucci(1997)</label><mixed-citation>
Hamill, T. and Colucci, S.: Verification of Eta–RSM Short-Range Ensemble
Forecasts, Mon. Weather Rev., 125, 1312–1327,
<a href="https://doi.org/10.1175/1520-0493(1997)125&lt;1312:VOERSR&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(1997)125&lt;1312:VOERSR&gt;2.0.CO;2</a>, 1997.
</mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Hamill et al.(2000)</label><mixed-citation>
Hamill, T., Snyder, C., and Morss, R.: A comparison of probabilistic forecasts
from bred, singular-vector, and perturbed observation ensembles, Mon.
Weather Rev., 128, 1835–1851,
<a href="https://doi.org/10.1175/1520-0493(2000)128&lt;1835:ACOPFF&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2000)128&lt;1835:ACOPFF&gt;2.0.CO;2</a>, 2000.
</mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Hamill et al.(2003)</label><mixed-citation>
Hamill, T., Snyder, C., and Whitaker, J.: Ensemble forecasts and the
properties of flow-dependent analysis-error covariance singular vectors,
Mon. Weather Rev., 131, 1741–1758, <a href="https://doi.org/10.1175//2559.1" target="_blank">https://doi.org/10.1175//2559.1</a>, 2003.
</mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Hamill et al.(2004)</label><mixed-citation>
Hamill, T., Whitaker, J., and Wei, X.: Ensemble reforecasting: Improving
medium-range forecast skill using retrospective forecasts, Mon. Weather
Rev., 132, 1434–1447,
<a href="https://doi.org/10.1175/1520-0493(2004)132&lt;1434:ERIMFS&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2004)132&lt;1434:ERIMFS&gt;2.0.CO;2</a>, 2004.
</mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Hamill and Whitaker(2006)</label><mixed-citation>
Hamill, T. M. and Whitaker, J. S.: Probabilistic quantitative precipitation
forecasts based on reforecast analogs: Theory and application, Mon. Weather
Rev., 134, 3209–3229, <a href="https://doi.org/10.1175/MWR3237.1" target="_blank">https://doi.org/10.1175/MWR3237.1</a>, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Helfricht et al.(2018)</label><mixed-citation>
Helfricht, K., Hartl, L., Koch, R., Marty, C., and Olefs, M.: Obtaining sub-daily new snow density from automated measurements in high mountain regions, Hydrol. Earth Syst. Sci., 22, 2655–2668, <a href="https://doi.org/10.5194/hess-22-2655-2018" target="_blank">https://doi.org/10.5194/hess-22-2655-2018</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Jewson et al.(2004)</label><mixed-citation>
Jewson, S., Brix, A., and Ziehmann, C.: A new parametric model for the
assessment and calibration of medium-range ensemble temperature forecasts,
Atmos. Sci. Lett., 5, 96–102, <a href="https://doi.org/10.1002/asl.69" target="_blank">https://doi.org/10.1002/asl.69</a>, 2004.
</mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Kochendorfer et al.(2017)</label><mixed-citation>
Kochendorfer, J., Rasmussen, R., Wolff, M., Baker, B., Hall, M. E., Meyers, T., Landolt, S., Jachcik, A., Isaksen, K., Brækkan, R., and Leeper, R.: The quantification and correction of wind-induced precipitation measurement errors, Hydrol. Earth Syst. Sci., 21, 1973–1989, <a href="https://doi.org/10.5194/hess-21-1973-2017" target="_blank">https://doi.org/10.5194/hess-21-1973-2017</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Krinner et al.(2018)</label><mixed-citation>
Krinner, G., Derksen, C., Essery, R., Flanner, M., Hagemann, S., Clark, M., Hall, A., Rott, H., Brutel-Vuilmet, C., Kim, H., Ménard, C. B., Mudryk, L., Thackeray, C., Wang, L., Arduini, G., Balsamo, G., Bartlett, P., Boike, J., Boone, A., Chéruy, F., Colin, J., Cuntz, M., Dai, Y., Decharme, B., Derry, J., Ducharne, A., Dutra, E., Fang, X., Fierz, C., Ghattas, J., Gusev, Y., Haverd, V., Kontu, A., Lafaysse, M., Law, R., Lawrence, D., Li, W., Marke, T., Marks, D., Ménégoz, M., Nasonova, O., Nitta, T., Niwano, M., Pomeroy, J., Raleigh, M. S., Schaedler, G., Semenov, V., Smirnova, T. G., Stacke, T., Strasser, U., Svenson, S., Turkov, D., Wang, T., Wever, N., Yuan, H., Zhou, W., and Zhu, D.: ESM-SnowMIP: assessing snow models and quantifying snow-related climate feedbacks, Geosci. Model Dev., 11, 5027–5049, <a href="https://doi.org/10.5194/gmd-11-5027-2018" target="_blank">https://doi.org/10.5194/gmd-11-5027-2018</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Lafaysse et al.(2017)</label><mixed-citation>
Lafaysse, M., Cluzet, B., Dumont, M., Lejeune, Y., Vionnet, V., and Morin, S.: A multiphysical ensemble system of numerical snow modelling, The Cryosphere, 11, 1173–1198, <a href="https://doi.org/10.5194/tc-11-1173-2017" target="_blank">https://doi.org/10.5194/tc-11-1173-2017</a>, 2017.
</mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Lehning et al.(2002)</label><mixed-citation>
Lehning, M., Bartelt, P., Brown, B., Fierz, C., and Satyawali, P.: A physical
SNOWPACK model for the Swiss avalanche warning, Part
II: snow microstructure., Cold Reg. Sci. Technol., 35, 147–167,
<a href="https://doi.org/10.1016/S0165-232X(02)00073-3" target="_blank">https://doi.org/10.1016/S0165-232X(02)00073-3</a>, 2002.
</mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>Lerch and Thorarinsdottir(2013)</label><mixed-citation>
Lerch, S. and Thorarinsdottir, T. L.: Comparison of non-homogeneous regression
models for probabilistic wind speed forecasting, Tellus A, 65, 21206,
<a href="https://doi.org/10.3402/tellusa.v65i0.21206" target="_blank">https://doi.org/10.3402/tellusa.v65i0.21206</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>Leutbecher and Palmer(2008)</label><mixed-citation>
Leutbecher, M. and Palmer, T. N.: Ensemble forecasting, J. Comput. Phys.,
227, 3515–3539, <a href="https://doi.org/10.1016/j.jcp.2007.02.014" target="_blank">https://doi.org/10.1016/j.jcp.2007.02.014</a>, 2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Masson et al.(2013)</label><mixed-citation>
Masson, V., Le Moigne, P., Martin, E., Faroux, S., Alias, A., Alkama, R., Belamari, S., Barbu, A., Boone, A., Bouyssel, F., Brousseau, P., Brun, E., Calvet, J.-C., Carrer, D., Decharme, B., Delire, C., Donier, S., Essaouini, K., Gibelin, A.-L., Giordani, H., Habets, F., Jidane, M., Kerdraon, G., Kourzeneva, E., Lafaysse, M., Lafont, S., Lebeaupin Brossier, C., Lemonsu, A., Mahfouf, J.-F., Marguinaud, P., Mokhtari, M., Morin, S., Pigeon, G., Salgado, R., Seity, Y., Taillefer, F., Tanguy, G., Tulet, P., Vincendon, B., Vionnet, V., and Voldoire, A.: The SURFEXv7.2 land and ocean surface platform for coupled or offline simulation of earth surface variables and fluxes, Geosci. Model Dev., 6, 929–960, <a href="https://doi.org/10.5194/gmd-6-929-2013" target="_blank">https://doi.org/10.5194/gmd-6-929-2013</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>Messner et al.(2014)</label><mixed-citation>
Messner, J. W., Mayr, G. J., Wilks, D. S., and Zeileis, A.: Extending Extended
Logistic Regression: Extended versus Separate versus Ordered versus
Censored, Mon. Weather Rev., 142, 3003–3014,
<a href="https://doi.org/10.1175/MWR-D-13-00355.1" target="_blank">https://doi.org/10.1175/MWR-D-13-00355.1</a>, 2014.
</mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Molteni et al.(1996)</label><mixed-citation>
Molteni, F., Buizza, R., Palmer, T., and Petroliagis, T.: The ECMWF ensemble
prediction system: Methodology and validation, Q. J. Roy. Meteorol.
Soc., 122, 73–119, <a href="https://doi.org/10.1002/qj.49712252905" target="_blank">https://doi.org/10.1002/qj.49712252905</a>, 1996.
</mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Mullen and Buizza(2002)</label><mixed-citation>
Mullen, S. and Buizza, R.: The impact of horizontal resolution and ensemble
size on probabilistic forecasts of precipitation by the ECMWF Ensemble
Prediction System, Weather Forecast., 17, 173–191,
<a href="https://doi.org/10.1175/1520-0434(2002)017&lt;0173:TIOHRA&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0434(2002)017&lt;0173:TIOHRA&gt;2.0.CO;2</a>, 2002.
</mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Pahaut(1975)</label><mixed-citation>
Pahaut, E.: La métamorphose des cristaux de neige (Snow crystal
metamorphosis), vol. 96 of Monographies de la Météorologie
Nationale, Météo France, 1975.
</mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>Palmer(2001)</label><mixed-citation>
Palmer, T.: A nonlinear dynamical perspective on model error: A proposal for
non-local stochastic-dynamic parametrization in weather and climate
prediction models, Q. J. Roy. Meteorol. Soc., 127, 279–304,
<a href="https://doi.org/10.1002/qj.49712757202" target="_blank">https://doi.org/10.1002/qj.49712757202</a>, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>Pellerin et al.(2003)</label><mixed-citation>
Pellerin, G., Lefaivre, L., Houtekamer, P., and Girard, C.: Increasing the horizontal resolution of ensemble forecasts at CMC, Nonlin. Processes Geophys., 10, 463–468, <a href="https://doi.org/10.5194/npg-10-463-2003" target="_blank">https://doi.org/10.5194/npg-10-463-2003</a>, 2003.
</mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Raftery et al.(2005)</label><mixed-citation>
Raftery, A., Gneiting, T., Balabdaoui, F., and Polakowski, M.: Using Bayesian
model averaging to calibrate forecast ensembles, Mon. Weather Rev., 133,
1155–1174, <a href="https://doi.org/10.1175/MWR2906.1" target="_blank">https://doi.org/10.1175/MWR2906.1</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Ramos et al.(2013)</label><mixed-citation>
Ramos, M. H., van Andel, S. J., and Pappenberger, F.: Do probabilistic forecasts lead to better decisions?, Hydrol. Earth Syst. Sci., 17, 2219–2232, <a href="https://doi.org/10.5194/hess-17-2219-2013" target="_blank">https://doi.org/10.5194/hess-17-2219-2013</a>, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Richardson(2000)</label><mixed-citation>
Richardson, D.: Skill and relative economic value of the ECMWF ensemble
prediction system, Q. J. Roy. Meteorol. Soc., 126, 649–667,
<a href="https://doi.org/10.1256/smsqj.56312" target="_blank">https://doi.org/10.1256/smsqj.56312</a>, 2000.
</mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Roulston and Smith(2002)</label><mixed-citation>
Roulston, M. and Smith, L.: Evaluating probabilistic forecasts using
information theory, Mon. Weather Rev., 130, 1653–1660,
<a href="https://doi.org/10.1175/1520-0493(2002)130&lt;1653:EPFUIT&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2002)130&lt;1653:EPFUIT&gt;2.0.CO;2</a>, 2002.
</mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Scheuerer(2014)</label><mixed-citation>
Scheuerer, M.: Probabilistic quantitative precipitation forecasting using
Ensemble Model Output Statistics, Q. J. Roy. Meteorol. Soc., 140,
1086–1096, <a href="https://doi.org/10.1002/qj.2183" target="_blank">https://doi.org/10.1002/qj.2183</a>, 2014.
</mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>Scheuerer and Hamill(2015)</label><mixed-citation>
Scheuerer, M. and Hamill, T. M.: Statistical Postprocessing of Ensemble
Precipitation Forecasts by Fitting Censored, Shifted Gamma Distributions,
Mon. Weather Rev., 143, 4578–4596, <a href="https://doi.org/10.1175/MWR-D-15-0061.1" target="_blank">https://doi.org/10.1175/MWR-D-15-0061.1</a>,
2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Scheuerer and Hamill(2018)</label><mixed-citation>
Scheuerer, M. and Hamill, T. M.: Generating Calibrated Ensembles of Physically
Realistic, High-Resolution Precipitation Forecast Fields Based on GEFS Model
Output, J. Hydrometeorol., 19, 1651–1670,
<a href="https://doi.org/10.1175/JHM-D-18-0067.1" target="_blank">https://doi.org/10.1175/JHM-D-18-0067.1</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Scheuerer and Hamill(2019)</label><mixed-citation>
Scheuerer, M. and Hamill, T. M.: Probabilistic Forecasting of Snowfall Amounts
Using a Hybrid between a Parametric and an Analog Approach, Mon. Weather
Rev., 147, 1047–1064, <a href="https://doi.org/10.1175/MWR-D-18-0273.1" target="_blank">https://doi.org/10.1175/MWR-D-18-0273.1</a>, 2019.
</mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Schleef et al.(2014)</label><mixed-citation>
Schleef, S., Löwe, H., and Schneebeli, M.: Influence of stress, temperature and crystal morphology on isothermal densification and specific surface area decrease of new snow, The Cryosphere, 8, 1825–1838, <a href="https://doi.org/10.5194/tc-8-1825-2014" target="_blank">https://doi.org/10.5194/tc-8-1825-2014</a>, 2014.
</mixed-citation></ref-html>
<ref-html id="bib1.bib59"><label>Stauffer et al.(2018)</label><mixed-citation>
Stauffer, R., Mayr, G. J., Messner, J. W., and Zeileis, A.: Hourly
probabilistic snow forecasts over complex terrain: a hybrid ensemble
postprocessing approach, Adv. Stat. Climatol. Meteorol. Oceanogr., 4, 65–86,
<a href="https://doi.org/10.5194/ascmo-4-65-2018" target="_blank">https://doi.org/10.5194/ascmo-4-65-2018</a>,
2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib60"><label>Sutton et al.(2006)</label><mixed-citation>
Sutton, C., Hamill, T. M., and Warner, T. T.: Will perturbing soil moisture
improve warm-season ensemble forecasts? A proof of concept, Mon. Weather
Rev., 134, 3174–3189, <a href="https://doi.org/10.1175/MWR3248.1" target="_blank">https://doi.org/10.1175/MWR3248.1</a>, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib61"><label>Szunyogh and Toth(2002)</label><mixed-citation>
Szunyogh, I. and Toth, Z.: The effect of increased horizontal resolution on
the NCEP global ensemble mean forecasts, Mon. Weather Rev., 130,
1125–1143, <a href="https://doi.org/10.1175/1520-0493(2002)130&lt;1125:TEOIHR&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2002)130&lt;1125:TEOIHR&gt;2.0.CO;2</a>,
2002.

</mixed-citation></ref-html>
<ref-html id="bib1.bib62"><label>Taillardat et al.(2016)</label><mixed-citation>
Taillardat, M., Mestre, O., Zamo, M., and Naveau, P.: Calibrated Ensemble
Forecasts Using Quantile Regression Forests and Ensemble Model Output
Statistics, Mon. Weather Rev., 144, 2375–2393,
<a href="https://doi.org/10.1175/MWR-D-15-0260.1" target="_blank">https://doi.org/10.1175/MWR-D-15-0260.1</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib63"><label>Taillardat et al.(2019)</label><mixed-citation>
Taillardat, M., Fougères, A., Naveau, P., and Mestre, O.: Forest-based and
semi-parametric methods for the postprocessing of rainfall ensemble
forecasting, Weather Forecast., in press, <a href="https://doi.org/10.1175/WAF-D-18-0149.1" target="_blank">https://doi.org/10.1175/WAF-D-18-0149.1</a>,
2019.
</mixed-citation></ref-html>
<ref-html id="bib1.bib64"><label>Teufelsbauer(2011)</label><mixed-citation>
Teufelsbauer, H.: A two-dimensional snow creep model for alpine terrain,
Nat. Hazards, 56, 481–497, <a href="https://doi.org/10.1007/s11069-010-9515-8" target="_blank">https://doi.org/10.1007/s11069-010-9515-8</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib65"><label>Thorarinsdottir and Gneiting(2010)</label><mixed-citation>
Thorarinsdottir, T. L. and Gneiting, T.: Probabilistic forecasts of wind
speed: ensemble model output statistics by using heteroscedastic censored
regression, J. Roy. Statist. Soc., 173, 371–388,
<a href="https://doi.org/10.1111/j.1467-985X.2009.00616.x" target="_blank">https://doi.org/10.1111/j.1467-985X.2009.00616.x</a>, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib66"><label>Toth and Kalnay(1997)</label><mixed-citation>
Toth, Z. and Kalnay, E.: Ensemble forecasting at NCEP and the breeding
method, Mon. Weather Rev., 125, 3297–3319,
<a href="https://doi.org/10.1175/1520-0493(1997)125&lt;3297:EFANAT&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(1997)125&lt;3297:EFANAT&gt;2.0.CO;2</a>, 1997.
</mixed-citation></ref-html>
<ref-html id="bib1.bib67"><label>Vannitsem et al.(2018)</label><mixed-citation>
Vannitsem, S., Wilks, D. S., and Messner, J. W.: Statistical postprocessing of
ensemble forecasts, 1 edn., Elsevier, Amsterdam, <a href="https://doi.org/10.1016/C2016-0-03244-8" target="_blank">https://doi.org/10.1016/C2016-0-03244-8</a>, 2018.
</mixed-citation></ref-html>
<ref-html id="bib1.bib68"><label>Vernay et al.(2015)</label><mixed-citation>
Vernay, M., Lafaysse, M., Merindol, L., Giraud, G., and Morin, S.: Ensemble
Forecasting of snowpack conditions and avalanche hazard, Cold. Reg. Sci.
Technol., 120, 251–262, <a href="https://doi.org/10.1016/j.coldregions.2015.04.010" target="_blank">https://doi.org/10.1016/j.coldregions.2015.04.010</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib69"><label>Vionnet et al.(2012)</label><mixed-citation>
Vionnet, V., Brun, E., Morin, S., Boone, A., Faroux, S., Le Moigne, P., Martin, E., and Willemet, J.-M.: The detailed snowpack scheme Crocus and its implementation in SURFEX v7.2, Geosci. Model Dev., 5, 773–791, <a href="https://doi.org/10.5194/gmd-5-773-2012" target="_blank">https://doi.org/10.5194/gmd-5-773-2012</a>, 2012.
</mixed-citation></ref-html>
<ref-html id="bib1.bib70"><label>Wang and Bishop(2005)</label><mixed-citation>
Wang, X. and Bishop, C.: Improvement of ensemble reliability with a new
dressing kernel, Q. J. Roy. Meteorol. Soc., 131, 965–986,
<a href="https://doi.org/10.1256/qj.04.120" target="_blank">https://doi.org/10.1256/qj.04.120</a>, 2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib71"><label>Weisman et al.(1997)</label><mixed-citation>
Weisman, M., Skamarock, W., and Klemp, J.: The resolution dependence of
explicitly modeled convective systems, Mon. Weather Rev., 125, 527–548,
<a href="https://doi.org/10.1175/1520-0493(1997)125&lt;0527:TRDOEM&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(1997)125&lt;0527:TRDOEM&gt;2.0.CO;2</a>, 1997.
</mixed-citation></ref-html>
<ref-html id="bib1.bib72"><label>Wilks(2005)</label><mixed-citation>
Wilks, D.: Effects of stochastic parametrizations in the Lorenz `96 system,
Q. J. Roy. Meteorol. Soc., 131, 389–407, <a href="https://doi.org/10.1256/qj.04.03" target="_blank">https://doi.org/10.1256/qj.04.03</a>,
2005.
</mixed-citation></ref-html>
<ref-html id="bib1.bib73"><label>Wilks and Hamill(2007)</label><mixed-citation>
Wilks, D. S. and Hamill, T. M.: Comparison of ensemble-MOS methods using GFS
reforecasts, Mon. Weather Rev., 135, 2379–2390,
<a href="https://doi.org/10.1175/MWR3402.1" target="_blank">https://doi.org/10.1175/MWR3402.1</a>, 2007.
</mixed-citation></ref-html>
<ref-html id="bib1.bib74"><label>Wilks(2011)</label><mixed-citation>
Wilks, D. S.: Statistical Methods in the Atmospheric Sciences,
3 edn., Academic Press, Amsterdam, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib75"><label>WMO(2018)</label><mixed-citation>
WMO: Preliminary 2018 Edition of the Guide to Meteorological Instruments and
Methods of Observation, Tech. Rep. 8, World Meteorological Organization,
available at: <a href="http://www.wmo.int/pages/prog/www/IMOP/publications/CIMO-Guide/Prelim_2018_ed/8_cryo_2_en_MR.pdf" target="_blank">http://www.wmo.int/pages/prog/www/IMOP/publications/CIMO-Guide/Prelim_2018_ed/8_cryo_2_en_MR.pdf</a> (last access: 23 September 2019),
cIMO Guide, Volume II, chapter 2, page 34, 2018.
</mixed-citation></ref-html>--></article>
