<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0" article-type="research-article"><?xmltex \makeatother\@nolinetrue\makeatletter?><?xmltex \bartext{Model experiment description paper}?>
  <front>
    <journal-meta><journal-id journal-id-type="publisher">GMD</journal-id><journal-title-group>
    <journal-title>Geoscientific Model Development</journal-title>
    <abbrev-journal-title abbrev-type="publisher">GMD</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Geosci. Model Dev.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1991-9603</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/gmd-16-2737-2023</article-id><title-group><article-title>CLGAN: a generative adversarial network (GAN)-based video prediction model for precipitation nowcasting</article-title><alt-title>CLGAN: a GAN-based video prediction model for precipitation nowcasting</alt-title>
      </title-group><?xmltex \runningtitle{CLGAN: a GAN-based video prediction model for precipitation nowcasting}?><?xmltex \runningauthor{Y.~Ji et al.}?>
      <contrib-group>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff2">
          <name><surname>Ji</surname><given-names>Yan</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="yes" rid="aff2">
          <name><surname>Gong</surname><given-names>Bing</given-names></name>
          <email>b.gong@fz-juelich.de</email>
        <ext-link>https://orcid.org/0000-0001-7770-2738</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff2">
          <name><surname>Langguth</surname><given-names>Michael</given-names></name>
          
        <ext-link>https://orcid.org/0000-0003-3354-5333</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff2">
          <name><surname>Mozaffari</surname><given-names>Amirpasha</given-names></name>
          
        <ext-link>https://orcid.org/0000-0001-6719-0425</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Zhi</surname><given-names>Xiefei</given-names></name>
          
        </contrib>
        <aff id="aff1"><label>1</label><institution>Key Laboratory of Meteorological Disaster, Ministry of Education (KLME), Nanjing University of Information Science and Technology (NUIST), Nanjing 210044, China</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>Jülich Supercomputing Centre, Forschungszentrum Jülich, 52425 Jülich, Germany</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Bing Gong (b.gong@fz-juelich.de)</corresp></author-notes><pub-date><day>23</day><month>May</month><year>2023</year></pub-date>
      
      <volume>16</volume>
      <issue>10</issue>
      <fpage>2737</fpage><lpage>2752</lpage>
      <history>
        <date date-type="received"><day>31</day><month>August</month><year>2022</year></date>
           <date date-type="rev-request"><day>14</day><month>November</month><year>2022</year></date>
           <date date-type="rev-recd"><day>26</day><month>March</month><year>2023</year></date>
           <date date-type="accepted"><day>28</day><month>March</month><year>2023</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2023 </copyright-statement>
        <copyright-year>2023</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://gmd.copernicus.org/articles/.html">This article is available from https://gmd.copernicus.org/articles/.html</self-uri><self-uri xlink:href="https://gmd.copernicus.org/articles/.pdf">The full text article is available as a PDF file from https://gmd.copernicus.org/articles/.pdf</self-uri>
      <abstract><title>Abstract</title>

      <p id="d1e125">The prediction of precipitation patterns up to 2 h ahead, also known as precipitation nowcasting, at high spatiotemporal resolutions is of great relevance in weather-dependent decision-making and early warning systems.
In this study, we are aiming to provide an efficient and easy-to-understand deep neural network – CLGAN (convolutional long short-term memory generative adversarial network) – to improve the nowcasting skills of heavy precipitation events.
The model constitutes a generative adversarial network (GAN) architecture, whose generator is built upon a u-shaped encoder–decoder network (U-Net) and is equipped with recurrent long short-term memory (LSTM) cells to capture spatiotemporal features.
The optical flow model DenseRotation and the competitive video prediction models ConvLSTM (convolutional LSTM) and PredRNN-v2 (predictive recurrent neural network version 2) are used as the competitors.
A series of evaluation metrics, including the root mean square error, the critical success index, the fractions skill score, and object-based diagnostic evaluation, are utilized for a comprehensive comparison against competing baseline models. We show that CLGAN outperforms the competitors in terms of scores for dichotomous events and object-based diagnostics.
A sensitivity analysis on the weight of the GAN component indicates that the GAN-based architecture helps to capture heavy precipitation events. The results encourage future work based on the proposed CLGAN architecture to improve the precipitation nowcasting and early warning systems.</p>
  </abstract>
    
<funding-group>
<award-group id="gs1">
<funding-source>Bundesministerium für Bildung und Forschung</funding-source>
<award-id>955513</award-id>
<award-id>01 IS18047A</award-id>
</award-group>
<award-group id="gs2">
<funding-source>H2020 European Research Council</funding-source>
<award-id>787576</award-id>
</award-group>
</funding-group>
</article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d1e137">Heavy precipitation can lead to numerous hazards, cause damage to infrastructure, and even increase risk to human life <xref ref-type="bibr" rid="bib1.bibx13 bib1.bibx49 bib1.bibx29" id="paren.1"/>.
Accurate short-term predictions of precipitation events at high spatiotemporal resolutions, also known as precipitation nowcasting, are therefore critical in establishing early warning systems. These warning systems can in turn help authorities in weather-dependent decision-making and enhance risk-governance capabilities <xref ref-type="bibr" rid="bib1.bibx8 bib1.bibx25 bib1.bibx5" id="paren.2"/>.</p>
      <p id="d1e146">Current precipitation nowcasting systems mainly rely on convective-permitting numeric weather prediction (NWP) or on extrapolation techniques of precipitation patterns with the help of composite radar observations. However, NWP models suffer from difficulties in capturing these patterns in the nowcasting time range due to the spin-up effect and the challenges of handling non-Gaussian data in assimilation <xref ref-type="bibr" rid="bib1.bibx37" id="paren.3"/>. Also, a quick model run cycle would be required. For instance, ICON-D2 only initializes every 3 h <xref ref-type="bibr" rid="bib1.bibx32" id="paren.4"/>, which makes it impossible to get quick updates in light of rapidly growing precipitation patterns. Observation-based extrapolation methods, such as optical flow, are commonly superior to NWP models for precipitation nowcasting but also fail to capture the underlying nonlinear processes of precipitation formation, e.g., secondary triggering and aggregation <xref ref-type="bibr" rid="bib1.bibx56" id="paren.5"/>.</p>
      <p id="d1e158">Deep neural networks have gained increasing attention in the meteorological community over the last few years <xref ref-type="bibr" rid="bib1.bibx38 bib1.bibx43" id="paren.6"/>. The growing<?pagebreak page2738?> interest can be attributed to the success stories in other domains where deep learning (DL) has been proven to leverage high-level information from complex and highly nonlinear data in several applications, such as autonomous driving <xref ref-type="bibr" rid="bib1.bibx20" id="paren.7"/>, anomaly detection <xref ref-type="bibr" rid="bib1.bibx30" id="paren.8"/>, and semantic segmentation <xref ref-type="bibr" rid="bib1.bibx14" id="paren.9"/>. Recently, video prediction models, developed in the computer vision community, have been explored for precipitation nowcasting. Contemporary studies mainly applied model architectures based on u-shaped convolutional networks <xref ref-type="bibr" rid="bib1.bibx3 bib1.bibx42" id="paren.10"><named-content content-type="pre">U-Net; e.g.,</named-content></xref>, convolutional long short-term memory cells <xref ref-type="bibr" rid="bib1.bibx45" id="paren.11"><named-content content-type="pre">ConvLSTM; e.g.,</named-content></xref>, and generative <xref ref-type="bibr" rid="bib1.bibx37" id="paren.12"><named-content content-type="pre">see, e.g.,</named-content></xref> and attention models <xref ref-type="bibr" rid="bib1.bibx47" id="paren.13"><named-content content-type="pre">see, e.g.,</named-content></xref>. U-Nets are thereby considered to be beneficial since they are capable of extracting multi-scale features of the atmospheric processes <xref ref-type="bibr" rid="bib1.bibx42" id="paren.14"/>.
To also explicitly capture temporal dependencies in the underlying formation process of precipitation, recurrent ConvLSTM models are an appealing choice <xref ref-type="bibr" rid="bib1.bibx45" id="paren.15"/>. Thus, combining convolutional and recurrent networks with ConvLSTM layers is advantageous in generating stable precipitation nowcasting by encoding the spatial and temporal dependencies from the historical frames.</p>
      <p id="d1e200">Nevertheless, these models have problems with handling the statistical nature of precipitation, especially when a pixel-wise loss function is applied for the optimization process during training <xref ref-type="bibr" rid="bib1.bibx46 bib1.bibx3" id="paren.16"/>. Although log transformation, importance sampling, and weighting towards heavier precipitation targets are appropriate to govern the right-skewed gamma distribution of precipitation rates <xref ref-type="bibr" rid="bib1.bibx37" id="paren.17"><named-content content-type="pre">e.g.,</named-content></xref>, the inherent uncertainty in quasi-chaotic processes at mesoscale typically leads to unrealistically smooth precipitation patterns in the forecasts. While this issue is well known in many other video prediction tasks <xref ref-type="bibr" rid="bib1.bibx31 bib1.bibx11" id="paren.18"/>, it is of particular relevance in precipitation nowcasting. The high spatiotemporal variability seen in the observational (real) data cannot be maintained, and thus, heavy precipitation events are barely captured by models applying a pixel-wise loss. Generative models, which train a generator and a discriminator adversarially (generative adversarial network – GAN – models) are considered to be a potential solution for such applications <xref ref-type="bibr" rid="bib1.bibx17" id="paren.19"/>. By forcing the generator to fool the discriminator, which aims to distinguish between real and generated data, these models succeed in maintaining the statistical properties of the underlying data <xref ref-type="bibr" rid="bib1.bibx35" id="paren.20"/>.</p>
      <p id="d1e221">Although great progress has been achieved in a series of recent studies <xref ref-type="bibr" rid="bib1.bibx37 bib1.bibx16" id="paren.21"><named-content content-type="pre">e.g.,</named-content></xref>, there is controversy regarding how different components of sophisticated model architectures contribute to the predictions.
Motivated by this, we build a simple but efficient and easy-to-understand video prediction model, CLGAN (convolutional long short-term memory generative adversarial network; see Fig. <xref ref-type="fig" rid="Ch1.F1"/>), for the nowcasting task.
CLGAN is proposed to leverage the advantages of different DL model architectures. The generator combines the U-Net with a ConvLSTM cell to abstract spatial features on multiple scales, while the temporal dependency of precipitation patterns is also preserved. The generator network is then trained adversarially to attain precipitation forecasts resembling observed data.
For our nowcasting application, we deploy a gridded dataset with a temporal resolution of 10 min aggregated from automatic weather station (AWS) gauges over Guizhou, China. The predictive performance of the proposed model architecture is then accessed in a comprehensive evaluation based on metrics designed for precipitation nowcasting. The evaluation also involves a comparison against a simplistic persistence forecast, the conventional optical flow model DenseRotation <xref ref-type="bibr" rid="bib1.bibx2" id="paren.22"/>, as well as two baseline video prediction models, a standard ConvLSTM <xref ref-type="bibr" rid="bib1.bibx45" id="paren.23"/> network and an up-to-date competing model PredRNN-v2 (predictive recurrent neural network version 2) <xref ref-type="bibr" rid="bib1.bibx51" id="paren.24"/>.</p>
      <p id="d1e240">With this, the main contributions of our study are as follows.
<list list-type="bullet"><list-item>
      <p id="d1e245">An efficient and easy-to-understand architecture, CLGAN, leveraging the merits of U-Net, ConvLSTM, and GAN models is proposed to generate perceptually realistic precipitation forecasts.</p></list-item><list-item>
      <p id="d1e249">A new 10 min level precipitation dataset based on AWS gauges (Guizhou AWS_ML precipitation dataset) is built for machine learning experiments.</p></list-item><list-item>
      <p id="d1e253">Nowcasting of heavy precipitation events is improved with comprehensive verification.</p></list-item><list-item>
      <p id="d1e257">A sensitivity analysis is performed to assess the importance of adversarial training for generating forecasts with closer statistical properties of the observed precipitation.</p></list-item></list></p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F1" specific-use="star"><?xmltex \currentcnt{1}?><?xmltex \def\figurename{Figure}?><label>Figure 1</label><caption><p id="d1e262">The details of the proposed CLGAN model.
<bold>(a)</bold> Generator: the illustration is presented for a given forecast step <inline-formula><mml:math id="M1" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>. If <inline-formula><mml:math id="M2" display="inline"><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>, the inputs are observed sequences <inline-formula><mml:math id="M3" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>. Otherwise the inputs are combined sequences of observed ones <inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mo>:</mml:mo><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and predicted ones <inline-formula><mml:math id="M5" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mi>i</mml:mi><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>.
The output is the model prediction <inline-formula><mml:math id="M6" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>,
<inline-formula><mml:math id="M7" display="inline"><mml:mi>c</mml:mi></mml:math></inline-formula> is the number of channels of inputs and here is 1, and
ngf denotes the number of filters in the first layer of U-Net.
<bold>(b)</bold> Discriminator: <inline-formula><mml:math id="M8" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula> is the length of output sequences.  <bold>(c)</bold> CLGAN: the outputs of the generator are the predicted sequences <inline-formula><mml:math id="M9" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>. <inline-formula><mml:math id="M10" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">Y</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> constitutes the corresponding ground truth. <inline-formula><mml:math id="M11" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M12" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> are the reconstruction loss and adversarial loss, respectively.</p></caption>
        <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f01.png"/>

      </fig>

</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Related work and baseline models</title>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Conventional methods</title>
      <p id="d1e498">The simplest approach to generate precipitation “forecasts” is to deploy the Eulerian persistence. For this, the most recent available observation, usually a radar composite, is used and then replicated several times for the future steps. This approach is quite accurate for very short lead times but obviously fails to provide meaningful forecasts in a quickly evolving system for timescales beyond several minutes such as the atmosphere. Thus, the related forecast quality can be considered as the minimum level for a prediction model to be useful.</p>
      <?pagebreak page2740?><p id="d1e501">Conventional precipitation nowcasting systems typically use a Lagrangian framework to predict the development of precipitation patterns. Although this framework often assumes the persistence of the precipitation features' intensity and displacement, it is still capable of outperforming mesoscale NWP models in precipitation nowcasting <xref ref-type="bibr" rid="bib1.bibx48" id="paren.25"/>. The Lagrangian method applies a two-step approach where the precipitation features are first tracked and then extrapolated to future time steps <xref ref-type="bibr" rid="bib1.bibx1" id="paren.26"/>.
Typically, the tracking step is accomplished with the help of optical flow methods that infer the motion of patterns from consecutive images. For precipitation nowcasting, radar composite images are subject to a tracking algorithm such as cross-correlation tracking <xref ref-type="bibr" rid="bib1.bibx39 bib1.bibx18 bib1.bibx57" id="paren.27"/> or centroid tracking techniques <xref ref-type="bibr" rid="bib1.bibx58" id="paren.28"/>. The tracked objects are then applied to different extrapolation schemes, e.g., image warping <xref ref-type="bibr" rid="bib1.bibx54" id="paren.29"/>, constant-vector advection <xref ref-type="bibr" rid="bib1.bibx4" id="paren.30"/>, or semi-Lagrangian schemes <xref ref-type="bibr" rid="bib1.bibx15" id="paren.31"/>. With this two-step approach, several operational precipitation nowcasting systems have been established over the globe in the last 3 decades, such as the Thunderstorm Identification Tracking Nowcasting <xref ref-type="bibr" rid="bib1.bibx8" id="paren.32"><named-content content-type="pre">TITAN;</named-content></xref>, the Storm Cell Identification and Tracking <xref ref-type="bibr" rid="bib1.bibx25" id="paren.33"><named-content content-type="pre">SCIT;</named-content></xref>, and the Short‐Term Ensemble Prediction System <xref ref-type="bibr" rid="bib1.bibx5" id="paren.34"><named-content content-type="pre">STEPS;</named-content></xref> <xref ref-type="bibr" rid="bib1.bibx53" id="paren.35"><named-content content-type="pre">see</named-content><named-content content-type="post">for a review on operational systems</named-content></xref>.</p>
      <p id="d1e548">Recently, <xref ref-type="bibr" rid="bib1.bibx2" id="text.36"/> implemented a set of advanced optical flow models into an open-source Python library called <italic>rainymotion</italic>. Two different groups of methods are part of this library from which we select the DenseRotation model that performs best in their study. The tracking algorithm of this model is based on the dense inverse search algorithm proposed in <xref ref-type="bibr" rid="bib1.bibx27" id="text.37"/> providing an estimate of the motion of each pixel based on two consecutive radar images. The extrapolation is then performed with a semi-Lagrangian advection scheme <xref ref-type="bibr" rid="bib1.bibx15" id="paren.38"/> capable of representing rotational motions.</p>
      <p id="d1e563">In our study, the Eulerian persistence model and the DenseRotation model are used to show how well the traditional methods can perform for the precipitation nowcasting task and how much benefit can be further obtained by using DL-based video prediction methods.</p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>Video prediction method</title>
      <p id="d1e574">As already mentioned, the application of deep learning techniques in the meteorological community has gained momentum over the recent years. In particular, several studies have started to explore these techniques to tackle the precipitation nowcasting problem.
By formulating precipitation nowcasting as a sequence prediction task, <xref ref-type="bibr" rid="bib1.bibx45" id="text.39"/> proposed a network of ConvLSTM cells which apply a convolution in the recurrent layers of the vanilla LSTM to capture spatiotemporal features in the underlying data. Their two-layer ConvLSTM network was able to outperform the Variational Methods for Echoes of Radar (ROVER) by <xref ref-type="bibr" rid="bib1.bibx55" id="text.40"/>, an operational precipitation nowcasting system based on optical flow methods with a semi-Lagrangian advection scheme. <xref ref-type="bibr" rid="bib1.bibx46" id="text.41"/> further extended the recurrent cells of gated recurrent units (GRUs) with non-local neural connections and proposed the Trajectory GRU (TrajGRU) model which enables the learning of location-variant structures of precipitation.</p>
      <p id="d1e586">Besides, <xref ref-type="bibr" rid="bib1.bibx50" id="text.42"/> advanced the application of ConvLSTM networks and proposed the predictive recurrent neural network (PredRNN). They deployed a stack of recurrent layers that feature a “zigzag memory flow” and involve an explicit spatiotemporal memory state. In this way, they enable an explicit communication of abstracted spatiotemporal features between different levels of the recurrent network which yields improved precipitation predictions. While this approach already provided promising results, the PredRNN model was updated to PredRNN-v2 <xref ref-type="bibr" rid="bib1.bibx51" id="paren.43"/>. The updates comprise the implementation of a “decoupling loss”, named ST-LSTM, to enhance the featuring of the spatiotemporal variations and a new, improved long-term modeling strategy. The model attains remarkable improvements when applied to multiple datasets including radar observations.</p>
      <p id="d1e595">Meanwhile, other network architectures were explored in the scope of precipitation nowcasting.
One of them is the fully convolutional U-Net architecture, which is a u-shaped hierarchical encoder–decoder network with skip connections. The architecture enables the abstraction of features on different spatial scales.
Notably, the RainNet architecture proposed by <xref ref-type="bibr" rid="bib1.bibx3" id="text.44"/> proved to significantly outperform optical-flow-based nowcasting methods for weak precipitation events. However, their network tends to provide too smooth precipitation fields and therefore fails to provide added value for more intense precipitation events with a rain rate above 10 mm h<inline-formula><mml:math id="M13" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. Recently, a deep generative model for the probabilistic precipitation nowcasting was proposed and showed state-of-the-art performance for the task <xref ref-type="bibr" rid="bib1.bibx37" id="paren.45"/>.</p>
      <p id="d1e616">All these studies demonstrate that deep neural networks have the potential to provide added value for precipitation nowcasting. In our study, we focus on further improving the predictions of strong precipitation events and therefore choose a simple ConvLSTM <xref ref-type="bibr" rid="bib1.bibx45" id="paren.46"/> and the advanced PredRNN-v2 <xref ref-type="bibr" rid="bib1.bibx51" id="paren.47"/> model for competing with our newly proposed model architecture.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Method and data</title>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Our model CLGAN</title>
      <p id="d1e641">In the following, we present our proposed CLGAN architecture in more detail. Since CLGAN aims to benefit from ConvLSTM models, the U-Net architecture, and GAN models, we first introduce its components separately to provide a deeper understanding of and reasoning for the chosen components.</p>
<sec id="Ch1.S3.SS1.SSS1">
  <label>3.1.1</label><title>ConvLSTM</title>
      <?pagebreak page2741?><p id="d1e651">The ConvLSTM network was proposed as an extension of LSTM layers which embedded the convolution operation to explicitly encode complex spatiotemporal features in a data sequence.
The basic formulas of the ConvLSTM cell which describe the gated update procedure for the hidden and cell state are provided in <xref ref-type="bibr" rid="bib1.bibx45" id="text.48"/> and are therefore not repeated here.
The objective function of a ConvLSTM model typically constitutes the classical <inline-formula><mml:math id="M14" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> reconstruction loss. This loss measures the distance between the predicted and the target (ground truth) data on the grid-point (or pixel-wise) level and can be written as
              <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M15" display="block"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mfenced open="∥" close="∥"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">Y</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
            where <inline-formula><mml:math id="M16" display="inline"><mml:mi mathvariant="bold-italic">Y</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M17" display="inline"><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover></mml:math></inline-formula> are 2D tensors for the ground truth and the predicted data, respectively. <inline-formula><mml:math id="M18" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> represents the end of the input sequence, and <inline-formula><mml:math id="M19" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula> is the forecast time step, so the model is optimized on the loss over the prediction sequence from <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> to <inline-formula><mml:math id="M21" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula>.
These tensors comprise <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:mi>w</mml:mi><mml:mo>×</mml:mo><mml:mi>h</mml:mi></mml:mrow></mml:math></inline-formula> grid points in zonal and meridional directions of the domain of interest.</p>
</sec>
<sec id="Ch1.S3.SS1.SSS2">
  <label>3.1.2</label><title>U-Net</title>
      <p id="d1e811">The U-Net model was originally applied for biomedical image segmentation <xref ref-type="bibr" rid="bib1.bibx42" id="paren.49"/> and is therefore designed as a powerful feature extractor on various spatial scales. As illustrated in Fig. <xref ref-type="fig" rid="Ch1.F1"/>a, it can be decomposed into a compressing and an expansive path that are bridged by skip connections. The contracting path can be seen as an encoder which converts the highly resolved data into coarse-grained features using convolutional and pooling layers. The expansive path, acting as a decoder, applies deconvolutional layers to convert back to the original spatial resolution, of which the number of data points are <inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:mi>w</mml:mi><mml:mo>×</mml:mo><mml:mi>h</mml:mi></mml:mrow></mml:math></inline-formula>. Usually, several pooling and deconvolutional layers are applied to allow feature extraction on different spatial scales.
To avoid the so-called vanishing gradients issue and to allow a direct information flow of specific spatial features, skip connections are implemented at every scale-specific feature extraction level <xref ref-type="bibr" rid="bib1.bibx9" id="paren.50"/>.</p>
      <p id="d1e834">In a video prediction application, the data at time step <inline-formula><mml:math id="M24" display="inline"><mml:mi>t</mml:mi></mml:math></inline-formula> enter the encoder to produce a forecast at time step <inline-formula><mml:math id="M25" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> with the decoder. By doing so, no long-term information is explicitly conveyed as with the ConvLSTM model. Since heavy precipitation events are rare but of high relevance for nowcasting, different techniques are usually applied to encourage deep neural networks in predicting events on the right tail of the underlying probability density function. Log transformation converts the right-skewed gamma distribution of precipitation data <xref ref-type="bibr" rid="bib1.bibx3" id="paren.51"><named-content content-type="pre">e.g., RainNet in</named-content></xref> into a Gaussian-like distribution which puts strong precipitation events closer to the center of mass in probability space.
Stronger weighting on higher precipitation rates and importance sampling can further support the optimization efficiency with respect to heavy precipitation events <xref ref-type="bibr" rid="bib1.bibx37" id="paren.52"/>. Nonetheless, U-Nets and ConvLSTM modes still tend to produce too smooth precipitation patterns, thereby failing to capture the relevant strong precipitation events.</p><?xmltex \hack{\newpage}?>
</sec>
<sec id="Ch1.S3.SS1.SSS3">
  <label>3.1.3</label><title>Generative adversarial networks</title>
      <p id="d1e873">To enforce a closer agreement of the generated data with the ground truth, GAN models were proposed by <xref ref-type="bibr" rid="bib1.bibx17" id="text.53"/>.
A GAN model consists of a generative network <inline-formula><mml:math id="M26" display="inline"><mml:mi>G</mml:mi></mml:math></inline-formula> (generator) and a discriminative network <inline-formula><mml:math id="M27" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula> (discriminator) which aims to assign a probability of 1 to real and a probability of 0 to generated data.
While the discriminator is optimized to distinguish between both kinds of inputted data, the generator is encouraged to fool the discriminator. Thus, the GAN applies the binary cross-entropy loss as the objective function which enters a minimax game:
              <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M28" display="block"><mml:mtable class="split" rowspacing="0.2ex" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi>G</mml:mi><mml:mo>⋆</mml:mo></mml:msup></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>=</mml:mo><mml:mi>arg⁡</mml:mi><mml:munder><mml:mo movablelimits="false">min⁡</mml:mo><mml:mi>G</mml:mi></mml:munder><mml:munder><mml:mo movablelimits="false">max⁡</mml:mo><mml:mi>D</mml:mi></mml:munder><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mtext>with </mml:mtext></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi mathvariant="double-struck">E</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mfenced open="[" close="]"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mi>D</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mo>+</mml:mo><mml:msub><mml:mi mathvariant="double-struck">E</mml:mi><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mfenced open="[" close="]"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mi>D</mml:mi><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo>)</mml:mo><mml:mo>)</mml:mo><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
            Here, the generator is conditioned on the input data sequence <inline-formula><mml:math id="M29" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>.
Since generator and discriminator are trained adversarially, the generator is encouraged to create predictions that share the same statistical properties as the ground truth data. This is considered to be useful for generating realistic precipitation forecasts which should exhibit the high spatial variability in the observed data <xref ref-type="bibr" rid="bib1.bibx37 bib1.bibx36 bib1.bibx19" id="paren.54"/>.</p>
</sec>
<sec id="Ch1.S3.SS1.SSS4">
  <label>3.1.4</label><title>Convolutional LSTM GAN (CLGAN)</title>
      <p id="d1e1095">To combine the merits of a GAN model with the strong spatiotemporal feature extraction capacities of U-Nets and ConvLSTM models, we set up the generator <inline-formula><mml:math id="M30" display="inline"><mml:mi>G</mml:mi></mml:math></inline-formula> as follows (see Fig. <xref ref-type="fig" rid="Ch1.F1"/>a). The generator constitutes a three-level U-Net following <xref ref-type="bibr" rid="bib1.bibx44" id="text.55"/>. Each level of the encoder comprises two convolutional layers followed by max pooling with a <inline-formula><mml:math id="M31" display="inline"><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:math></inline-formula> kernel to reduce the spatial dimensionality in the next layer. The number of channels is thereby increased by a factor of 2 in each level. A ConvLSTM cell with 64 filters is deployed to implement recurrency at the bridge between the encoder and decoder. The decoder then reverts the encoded data to the input resolution with the help of deconvolutional layers. Furthermore, skip connections among the encoder and decoder are added at each level of the U-Net.
The discriminator <inline-formula><mml:math id="M32" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula> consists of 3D fully convolutional layers with batch normalization which allow us to encode both the temporal and spatial dimensions of the data sequence. Again, max pooling is used to compress the data which finally get concatenated to fully connected layers (see Fig. <xref ref-type="fig" rid="Ch1.F1"/>b).
The  forecast sequence of <inline-formula><mml:math id="M33" display="inline"><mml:mi>G</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M34" display="inline"><mml:mrow><mml:msub><mml:mover accent="true"><mml:mi mathvariant="bold-italic">X</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, and the corresponding ground truth sequence, <inline-formula><mml:math id="M35" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="bold-italic">Y</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, are taken as the inputs for the discriminator <inline-formula><mml:math id="M36" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula> (see Fig. <xref ref-type="fig" rid="Ch1.F1"/>c).</p>
      <?pagebreak page2742?><p id="d1e1197">In this study, the generator is trained by combining the adversarial loss <inline-formula><mml:math id="M37" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> with the reconstruction <inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss:
              <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M39" display="block"><mml:mrow><mml:msup><mml:mi>G</mml:mi><mml:mo>⋆</mml:mo></mml:msup><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mi mathvariant="italic">λ</mml:mi><mml:mo>)</mml:mo><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo><mml:mo>+</mml:mo><mml:mi mathvariant="italic">λ</mml:mi><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>(</mml:mo><mml:mi>G</mml:mi><mml:mo>)</mml:mo><mml:mspace width="0.25em" linebreak="nobreak"/><mml:mtext>with </mml:mtext><mml:mi mathvariant="italic">λ</mml:mi><mml:mo>∈</mml:mo><mml:mo>[</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>]</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
            This ensures that the prediction remains close to the ground truth.
The relative weight of the reconstruction loss <inline-formula><mml:math id="M40" display="inline"><mml:mi mathvariant="italic">λ</mml:mi></mml:math></inline-formula> is set to <inline-formula><mml:math id="M41" display="inline"><mml:mn mathvariant="normal">0.99</mml:mn></mml:math></inline-formula>,  which proves to balance the contributions from both loss components in the following experiments. Training of the model is performed with the Adam optimizer <xref ref-type="bibr" rid="bib1.bibx26" id="paren.56"/> over eight epochs with a batch size of 32.</p>
</sec>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><?xmltex \opttitle{Guizhou AWS\_ML precipitation dataset}?><title>Guizhou AWS_ML precipitation dataset</title>
      <p id="d1e1324">In addition to the widely used remote sensing data, e.g., radar composite images, measurements from densely distributed automatic weather stations can serve as an alternative in the data-driven weather forecasting task. In this study, minute-level precipitation measurements by rain gauges of AWSs over Guizhou, China (Guizhou AWS_ML precipitation dataset), are collected for the precipitation nowcasting task.
Guizhou is a mountainous and rainy province located in southwest China (see Fig. <xref ref-type="fig" rid="Ch1.F2"/>a) where mudslides happen frequently during summertime. For instance, the region was affected by a severe rainstorm in September 2020, in which some regions experienced more than 1500 mm rainfall within 20 d. Accurate precipitation nowcasting, especially for heavy precipitation, is crucial to reduce damage from these events. Hence, the Guizhou AWS_ML precipitation dataset is established for better simulation of precipitation with data-driven approaches. The AWS locations comprise 93 basic national stations and 1740 automatic weather stations (see Fig. <xref ref-type="fig" rid="Ch1.F2"/>b). Among other meteorological quantities (2 m temperature, 10 m wind, surface pressure, and relative humidity), the AWSs measure precipitation at a high observation frequency (every minute), and the data are provided between 1 January 2015 and 31 December 2019 by Guizhou Meteorological Bureau.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F2" specific-use="star"><?xmltex \currentcnt{2}?><?xmltex \def\figurename{Figure}?><label>Figure 2</label><caption><p id="d1e1333"><bold>(a)</bold> Annual average cumulative precipitation in Guizhou from 2015 to 2019. <bold>(b)</bold> The spatial distribution of AWSs over Guizhou.</p></caption>
          <?xmltex \igopts{width=483.69685pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f02.png"/>

        </fig>

      <p id="d1e1347">Several preprocessing steps are conducted for preparing the dataset of our experiment. First, the precipitation data are accumulated over 10 min, which still constitutes a reasonably high temporal resolution.
To obtain a gridded dataset, the observations are then interpolated bilinearly onto a regular, spherical grid. The target grid comprises <inline-formula><mml:math id="M42" display="inline"><mml:mrow><mml:mi>w</mml:mi><mml:mo>×</mml:mo><mml:mi>h</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">48</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">40</mml:mn></mml:mrow></mml:math></inline-formula> data points in zonal and meridional directions, respectively, and covers a domain from 24.625 to 29.5<inline-formula><mml:math id="M43" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> N and 103.625 to 109.5<inline-formula><mml:math id="M44" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> E with 0.125<inline-formula><mml:math id="M45" display="inline"><mml:msup><mml:mi/><mml:mo>∘</mml:mo></mml:msup></mml:math></inline-formula> resolution.
To obtain the data needed for training our CLGAN and the baseline models (see Sects. <xref ref-type="sec" rid="Ch1.S2"/> and <xref ref-type="sec" rid="Ch1.S3.SS1.SSS4"/>), we generate sliding sequences of 24 consecutive gridded data samples (frames) which comprise a temporal period of 240 min, and 120 min (12 frames) of each sequence serves as input to predict the next 120 min (12 frames).
Since there are many periods with no or only weak precipitation, we furthermore only select sequences whose averaged precipitation rate exceeds the empirical 60 % quantile of the complete dataset. This results in 35 054 sequences for the subsampled dataset.
Finally, a log transformation is applied to each sequence to make the data more Gaussian-like. The log transformation reads as follows: <inline-formula><mml:math id="M46" display="inline"><mml:mrow><mml:msup><mml:mi>x</mml:mi><mml:mo>′</mml:mo></mml:msup><mml:mo>=</mml:mo><mml:mi>ln⁡</mml:mi><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="italic">ε</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>ln⁡</mml:mi><mml:mo>(</mml:mo><mml:mi mathvariant="italic">ε</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, where <inline-formula><mml:math id="M47" display="inline"><mml:mi mathvariant="italic">ε</mml:mi></mml:math></inline-formula> is a small constant (here 0.01). We use the data from 2015 to 2017 for training, the data of 2018 for validating, and the data of 2019 for testing.</p>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title>Verification methods</title>
      <p id="d1e1453">As pointed out in <xref ref-type="bibr" rid="bib1.bibx43" id="text.57"/> and more specifically for precipitation in <xref ref-type="bibr" rid="bib1.bibx28" id="text.58"/>, precipitation nowcasting should be evaluated in terms of application-specific scores. This is due to the unique statistical properties of precipitation rates, as well as the chaotic atmospheric processes which underpin the formation of precipitation. Additionally, we would like to emphasize that a single score alone can barely evaluate the model performance applied to high-dimensional data <xref ref-type="bibr" rid="bib1.bibx52" id="paren.59"/>. Therefore, we take several evaluation metrics into account to provide a comprehensive overview.</p>
      <p id="d1e1465">The first considered family of evaluation metrics is established for continuous quantities in the meteorological community. The root mean square error (RMSE) measures the distance between the predicted and the observed field on a grid-point level. The correlation coefficient (CC) measures the association or the linear relationship between the two fields. A perfect correlation would result in <inline-formula><mml:math id="M48" display="inline"><mml:mrow><mml:mtext>CC</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>, while <inline-formula><mml:math id="M49" display="inline"><mml:mrow><mml:mtext>CC</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula> indicates no linear relationship between forecast and observation on the grid-point level.</p>
      <p id="d1e1492">The second set of scores is built on dichotomous events which are obtained by thresholding the gridded precipitation fields.
A <inline-formula><mml:math id="M50" display="inline"><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:math></inline-formula> contingency table is commonly used to show the frequency of “yes” and “no” forecasts and occurrences and give a joint distribution for events with a precipitation rate exceeding a given threshold <inline-formula><mml:math id="M51" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">pr</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.
According to the elements in the contingency table, a variety of categorical statistics can be computed to evaluate the dichotomous forecasts in particular aspects.
Critical success index (CSI), also known as threat score, measures the fraction of hits with respect to the number of occurrences where the events are either forecasted or observed. The frequently applied equitable threat score (ETS) is a variant of the CSI and explicitly accounts for random forecasts which perform well just by chance <xref ref-type="bibr" rid="bib1.bibx52" id="paren.60"/>.</p>
      <p id="d1e1521">However, due to the highly nonlinear and complex processes causing precipitation formation, scores acting on grid-point level are prone to penalize predictions which recover the high spatial variability but fail to match exactly the observed precipitation field. The issue leads to the double penalty problem where the model gets penalized twice, once for missing the exact placement of a precipitation event and once for shifting it spatially <xref ref-type="bibr" rid="bib1.bibx10" id="paren.61"/>.
To relax the requirement for exact spatial matching, the fractions skill score (FSS) is computed here as a fuzzy verification metric <xref ref-type="bibr" rid="bib1.bibx40" id="paren.62"/>. Similar to the CSI and ETS, the FSS operates on dichotomous events but allows for spatial shifts by considering a local neighborhood around each grid point. Within this neighborhood, the fractional coverage of the precipitation events is calculated for both the predictions and the observations. Let <inline-formula><mml:math id="M52" display="inline"><mml:mrow><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M53" display="inline"><mml:mrow><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>o</mml:mi><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> denote the fraction of event grid boxes within the local neighborhood of size <inline-formula><mml:math id="M54" display="inline"><mml:mi>s</mml:mi></mml:math></inline-formula> around the grid point <inline-formula><mml:math id="M55" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> in the model prediction and observation.
The fractions Brier score (FBS) is given by
            <disp-formula id="Ch1.E4" content-type="numbered"><label>4</label><mml:math id="M56" display="block"><mml:mrow><mml:mtext>FBS</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mo>(</mml:mo><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:msubsup><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>o</mml:mi><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:msubsup><mml:mo>)</mml:mo><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
         <?pagebreak page2743?> which quantifies the quadratic difference between the prediction and the observation for all <inline-formula><mml:math id="M57" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> grid points over the domain. The final FSS is then obtained with
            <disp-formula id="Ch1.E5" content-type="numbered"><label>5</label><mml:math id="M58" display="block"><mml:mrow><mml:mtext>FSS</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mtext>FBS</mml:mtext><mml:mo>/</mml:mo><mml:msub><mml:mtext>FBS</mml:mtext><mml:mi mathvariant="normal">worst</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
          where <inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:msub><mml:mtext>FBS</mml:mtext><mml:mi mathvariant="normal">worst</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the sum of the squared fractions of events in the prediction and in the observation. Higher FSS values indicate better forecast, while it can be shown that a forecast becomes “useful” when <inline-formula><mml:math id="M60" display="inline"><mml:mrow><mml:mtext>FSS</mml:mtext><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">0.5</mml:mn></mml:mrow></mml:math></inline-formula> is attained for a given neighborhood scale <inline-formula><mml:math id="M61" display="inline"><mml:mi>s</mml:mi></mml:math></inline-formula> (typically expressed in terms of squares with an edge length of <inline-formula><mml:math id="M62" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> grid points).</p>
      <p id="d1e1720">Nonetheless, FSS also does not capture the spatial precipitation patterns since each grid point in the neighborhood is treated equally and no check for spatial coherence is undertaken. Thus, we additionally perform an object-based diagnostic evaluation, called MODE <xref ref-type="bibr" rid="bib1.bibx23 bib1.bibx24 bib1.bibx21" id="paren.63"/>, to focus on pattern attributes such as location, area, and shape. To obtain the desired attributes, a convolutional filter of size <inline-formula><mml:math id="M63" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula> is first applied over the precipitation field. Afterwards, objects are defined by applying a threshold on the precipitation rate <inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">Pr</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and on the object area <inline-formula><mml:math id="M65" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">A</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.
A fuzzy logic scheme is then used to merge and pair precipitation objects in the predicted and observed precipitation field. Finally, the object-based threat score <xref ref-type="bibr" rid="bib1.bibx23" id="paren.64"><named-content content-type="pre">OTS;</named-content></xref> is computed to verify how well the predicted precipitation patterns match the observed ones. Here, we choose the object area, the centroid location, and the object shape (aspect ratio and orientation angle) as target attributes for computing the OTS.</p>
      <p id="d1e1760">All the mentioned verification methods are listed in Table <xref ref-type="table" rid="Ch1.T1"/> with a brief description. The details can be found in the corresponding references.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T1" specific-use="star"><?xmltex \currentcnt{1}?><label>Table 1</label><caption><p id="d1e1768">Summary of the verification methods used in the paper.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="4">
     <oasis:colspec colnum="1" colname="col1" align="justify" colwidth="3cm"/>
     <oasis:colspec colnum="2" colname="col2" align="justify" colwidth="5cm"/>
     <oasis:colspec colnum="3" colname="col3" align="justify" colwidth="4cm"/>
     <oasis:colspec colnum="4" colname="col4" align="justify" colwidth="4cm"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Verification method</oasis:entry>
         <oasis:entry colname="col2">Description</oasis:entry>
         <oasis:entry colname="col3">Formula or reference</oasis:entry>
         <oasis:entry colname="col4">Notes</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Root mean square error</oasis:entry>
         <oasis:entry colname="col2">The average magnitude of the forecast errors</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M66" display="inline"><mml:msqrt><mml:mrow><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:msubsup><mml:mo>∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:msubsup><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mo>′</mml:mo></mml:msubsup><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msqrt></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">[0, <inline-formula><mml:math id="M67" display="inline"><mml:mrow><mml:mo>+</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow></mml:math></inline-formula>)</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Correlation coefficient</oasis:entry>
         <oasis:entry colname="col2">The correspondence between the forecast and observed values</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M68" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:msubsup><mml:mo>∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:msubsup><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:msub><mml:mi>Y</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo mathvariant="normal">¯</mml:mo></mml:mover><mml:mo>)</mml:mo><mml:mo>(</mml:mo><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mo>′</mml:mo></mml:msubsup><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mo>′</mml:mo></mml:msubsup></mml:mrow><mml:mo mathvariant="normal">¯</mml:mo></mml:mover><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:msqrt><mml:mrow><mml:msubsup><mml:mo>∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:msubsup><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:msub><mml:mi>Y</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo mathvariant="normal">¯</mml:mo></mml:mover><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msqrt><mml:msqrt><mml:mrow><mml:msubsup><mml:mo>∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:msubsup><mml:mo>(</mml:mo><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mo>′</mml:mo></mml:msubsup><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mo>′</mml:mo></mml:msubsup></mml:mrow><mml:mo mathvariant="normal">¯</mml:mo></mml:mover><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msqrt></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">[<inline-formula><mml:math id="M69" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>, 1]</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Critical success index</oasis:entry>
         <oasis:entry colname="col2">The correspondence between the forecast “yes” events and observed “yes” events</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M70" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mtext>hits</mml:mtext><mml:mrow><mml:mtext>hits</mml:mtext><mml:mo>+</mml:mo><mml:mtext>misses</mml:mtext><mml:mo>+</mml:mo><mml:mtext>false alarms</mml:mtext></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">[0, 1], 0 means no skills</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Equitable threat score</oasis:entry>
         <oasis:entry colname="col2">The correspondence between the forecast “yes” events and observed “yes” events (accounting for hits due to chance)</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M71" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mtext>hits</mml:mtext><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="normal">hits</mml:mi><mml:mi mathvariant="normal">random</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mtext>hits</mml:mtext><mml:mo>+</mml:mo><mml:mtext>false alarms</mml:mtext></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula> <inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="normal">hits</mml:mi><mml:mi mathvariant="normal">random</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mo>(</mml:mo><mml:mtext>hits</mml:mtext><mml:mo>+</mml:mo><mml:mtext>misses</mml:mtext><mml:mo>)</mml:mo><mml:mo>(</mml:mo><mml:mtext>hits</mml:mtext><mml:mo>+</mml:mo><mml:mtext>false alarms</mml:mtext><mml:mo>)</mml:mo></mml:mrow><mml:mtext>total</mml:mtext></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4">[<inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:math></inline-formula>, 1], 0 means no skills</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Fractions skill score</oasis:entry>
         <oasis:entry colname="col2">The spatial scales at which the forecast resembles the observations</oasis:entry>
         <oasis:entry colname="col3">
                      <xref ref-type="bibr" rid="bib1.bibx41" id="text.65"/>
                    </oasis:entry>
         <oasis:entry colname="col4">[0, 1], the smallest window size for which FSS <inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">0.5</mml:mn></mml:mrow></mml:math></inline-formula> can be considered as “skillful scale”</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Object-based threat score</oasis:entry>
         <oasis:entry colname="col2">The similarity between the forecast objects with the observed ones according to a series of attributes</oasis:entry>
         <oasis:entry colname="col3">
                      <xref ref-type="bibr" rid="bib1.bibx6" id="text.66"/>
                    </oasis:entry>
         <oasis:entry colname="col4">[0, 1], 0 means complete mismatch and 1 means perfect match</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table><?xmltex \gdef\@currentlabel{1}?></table-wrap>

      <p id="d1e2210">To ease the comparison between the baseline models and the simplistic persistence forecast, we furthermore calculate skill scores (except for the FSS). In general, a skill score (SS) can be constructed by considering the target score <inline-formula><mml:math id="M75" display="inline"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> of the model, the score obtained with the reference forecast <inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, and the perfect score <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">perf</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>:
            <disp-formula id="Ch1.E6" content-type="numbered"><label>6</label><mml:math id="M78" display="block"><mml:mrow><mml:mtext>SS</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">perf</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
          The higher the SS is, the better the model performs against the reference score. Perfect models thereby obtain <inline-formula><mml:math id="M79" display="inline"><mml:mrow><mml:mtext>SS</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>, while inferior models show up with <inline-formula><mml:math id="M80" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="normal">∞</mml:mi><mml:mo>&lt;</mml:mo><mml:mtext>SS</mml:mtext><mml:mo>&lt;</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula>. Note that <inline-formula><mml:math id="M81" display="inline"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">perf</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula> holds for the RMSE, whereas the other scores under consideration attain <inline-formula><mml:math id="M82" display="inline"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi mathvariant="normal">perf</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula>.
Since the size of our dataset is not unlimited, we also apply a block bootstrapping procedure to estimate sampling uncertainty <xref ref-type="bibr" rid="bib1.bibx12" id="paren.67"/>. The block bootstrapping procedure accounts for autocorrelation between the sliding sequences and thus divides the dataset into non-overlapping blocks before the resampling of the blocks with replacement is performed. Here, we set the block length to 10 h (60 frames) and perform 1000 block bootstrapping steps.</p>
</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Results</title>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>Quantitative evaluation</title>
<sec id="Ch1.S4.SS1.SSS1">
  <label>4.1.1</label><title>Point-wise evaluation metrics</title>
      <p id="d1e2373">In Fig. <xref ref-type="fig" rid="Ch1.F3"/>a–d, our model is compared to the baseline models in terms of the skill scores for the grid-point-level evaluation metrics (CC, RMSE, CSI, and ETS). The skill scores are calculated by defining the Eulerian persistence as reference<?pagebreak page2744?> forecast.
It is seen that the deep learning models (i.e., ConvLSTM, PredRNN-v2 and CLGAN), as well as the optical flow model DenseRotation, outperform the persistence forecast after 20 min lead time in terms of the continuous scores (CC and RMSE). Among the video prediction models, PredRNN-v2 is superior over the others after the first 20 min, while ConvLSTM performs best for the longer lead times. CLGAN is not so competitive for RMSE and CC as PredRNN-v2 and ConvLSTM, while it still outperforms the traditional optical flow model DenseRotation.
Note that the Eulerian persistence performs well in the first 10 min. One possible reason why these complex models can barely beat the persistence forecast in the first lead step is that the precipitation systems are relatively invariant within this very short time period. In our case, the Eulerian persistence forecast is the latest of the observations available, which is hence highly correlated to the ground truth at short lead times. With increasing lead times, its performance degrades quickly.</p>
      <p id="d1e2378">The comparisons in terms of the dichotomous scores (CSI and ETS) are given in Fig. <xref ref-type="fig" rid="Ch1.F3"/>c and d. They demonstrate that CLGAN is superior to the other competitors at  all the lead times for simulating heavy precipitation events (the threshold <inline-formula><mml:math id="M83" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">Pr</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is set to 8 mm h<inline-formula><mml:math id="M84" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> here).
The optical flow model DenseRotation performs well in the first 40 lead minutes, while its skill scores decrease rapidly afterwards. By contrast, the advanced deep learning model PredRNN-v2 shows more potential for longer lead times. Although ConvLSTM outperforms on the continuous scores, it can barely capture the heavy precipitation events.
A large performance degradation for the ConvLSTM is diagnosed at a lead time of 20 min.
One reason of the difference is that the model performance is evaluated with the skill scores, which are affected by the choice of the reference model (here the Eulerian persistence). For the first time step (lead time of 10 min), both ConvLSTM and Eulerian persistence can capture strong precipitation events, and ConvLSTM is even better. However, ConvLSTM models are prone to produce blurry predictions in an autoregressive prediction task, where the errors in the prior forecasts are inherited to the later ones. Hence, the ConvLSTM model gets less efficient in the next few lead steps, while the Eulerian persistence performs fairly well. For longer lead times, the performance of the Eulerian persistence forecasts quickly degrades, and ConvLSTM again outperforms the persistence model with positive skill scores.</p>
      <p id="d1e2406">The comparisons of the model performance show that CLGAN is superior in terms of scores for dichotomous forecasts (CSI and ETS), while it is less competitive in terms of RMSE.
This is due to the fact that our CLGAN encourages the model to generate forecasts which have a similar distribution as the ground truth data rather than just reducing the averaged point-wise loss. Hence, more heavy precipitation events are predicted by the CLGAN model, which improves the dichotomous forecast scores. However, more predicted high-value precipitation could cause larger biases, compared to the models only generating low-value forecasts. The problem is magnified with the use of the point-by-point scores, i.e., the RMSE, which suffers from the double penalty issue.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F3" specific-use="star"><?xmltex \currentcnt{3}?><?xmltex \def\figurename{Figure}?><label>Figure 3</label><caption><p id="d1e2412">Box–whisker plots for skill scores of  <bold>(a)</bold> CC, <bold>(b)</bold> RMSE, <bold>(c)</bold> CSI, and <bold>(d)</bold> ETS averaged over the testing period with the Eulerian persistence as the reference forecast. The boxes show the range of the first quartile (upper) to the third quartile (bottom) of the skill scores, and the whiskers denote the 95th percentile (upper) and 5th percentile (bottom), respectively. The threshold <inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">pr</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> of CSI and ETS is 8 mm h<inline-formula><mml:math id="M86" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>.</p></caption>
            <?xmltex \igopts{width=483.69685pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f03.png"/>

          </fig>

</sec>
<?pagebreak page2745?><sec id="Ch1.S4.SS1.SSS2">
  <label>4.1.2</label><title>Spatial verification scores</title>
      <p id="d1e2465">To further investigate the model performances, we now turn our attention to the spatial verification scores, the FSS, and the MODE framework.
Figure <xref ref-type="fig" rid="Ch1.F4"/>a shows the model performance in terms of the FSS for a lead time of 60 min by the persistence model (used as the reference forecast). The FSS is computed based on different neighborhood sizes and thresholds of hourly precipitation rates. Specifically, the neighborhood scale <inline-formula><mml:math id="M87" display="inline"><mml:mi>s</mml:mi></mml:math></inline-formula> in kilometers varies along the <inline-formula><mml:math id="M88" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> axis. The FSS values for <inline-formula><mml:math id="M89" display="inline"><mml:mi>s</mml:mi></mml:math></inline-formula> attaining values of approximately 41, 69, 96, 152, and 290 km (square boxes of 3, 5, 7, 11, and 21 grid points, respectively) are plotted and marked as box–whisker plots for varying precipitation thresholds.
The boxes show the range of the first quartile (upper) to the third quartile (bottom) of the scores, and the whiskers are, respectively, the 95th percentile (upper) and 5th percentile (bottom).
As the threshold increases, FSS decreases, indicating that the persistence forecasts become increasingly imprecise for stronger precipitation events with a given spatial scale. For precipitation events exceeding <inline-formula><mml:math id="M90" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">Pr</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">8</mml:mn></mml:mrow></mml:math></inline-formula> mm h<inline-formula><mml:math id="M91" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, the persistence forecasts are considered useful <xref ref-type="bibr" rid="bib1.bibx40" id="paren.68"><named-content content-type="pre"><inline-formula><mml:math id="M92" display="inline"><mml:mrow><mml:mtext>FSS</mml:mtext><mml:mo>≥</mml:mo><mml:mn mathvariant="normal">0.5</mml:mn></mml:mrow></mml:math></inline-formula>; see, e.g.,</named-content></xref> for a neighborhood scale of <inline-formula><mml:math id="M93" display="inline"><mml:mrow><mml:mi>s</mml:mi><mml:mo>≈</mml:mo><mml:mn mathvariant="normal">69</mml:mn></mml:mrow></mml:math></inline-formula> km. Thus, the spatial accuracy of capturing these events is already fairly degraded, and the neighborhood scale of <inline-formula><mml:math id="M94" display="inline"><mml:mrow><mml:mi>s</mml:mi><mml:mo>≈</mml:mo><mml:mn mathvariant="normal">69</mml:mn></mml:mrow></mml:math></inline-formula> km (five grid points) is applied to compute the FSS for other models in the following.
Figure <xref ref-type="fig" rid="Ch1.F4"/>b compares the models' performance in terms of the FSS for heavy precipitation forecasting against the Eulerian persistence by illustrating the difference <inline-formula><mml:math id="M95" display="inline"><mml:mrow><mml:mi mathvariant="normal">Δ</mml:mi><mml:mtext>FSS</mml:mtext><mml:mo>=</mml:mo><mml:msub><mml:mtext>FSS</mml:mtext><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="normal">FSS</mml:mi><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. Here, <inline-formula><mml:math id="M96" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="normal">FSS</mml:mi><mml:mi mathvariant="normal">ref</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> denotes the reference persistence model, whereas <inline-formula><mml:math id="M97" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula> is used to denote the other competing models.
It is seen that all baseline models except from the ConvLSTM model can remarkably improve the spatial forecasting of such events, especially for longer lead times. Among them, CLGAN is superior to the others at all the lead times. DenseRotation performs well in the first lead hour, while PredRNN-v2 is promising for the further lead times.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F4" specific-use="star"><?xmltex \currentcnt{4}?><?xmltex \def\figurename{Figure}?><label>Figure 4</label><caption><p id="d1e2606"><bold>(a)</bold> FSS of persistence for different scales and intensity thresholds at 60 lead minutes for precipitation nowcasting. <bold>(b)</bold> Improvements in different models compared with persistence in terms of FSS with a threshold <inline-formula><mml:math id="M98" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">pr</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> of 8 mm h<inline-formula><mml:math id="M99" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and a neighborhood size of five. The boxes show the range of the first quartile (upper) to the third quartile (bottom) of the skill scores, and the whiskers denote the 95th percentile (upper) and 5th percentile (bottom), respectively.</p></caption>
            <?xmltex \igopts{width=483.69685pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f04.png"/>

          </fig>

</sec>
<?pagebreak page2746?><sec id="Ch1.S4.SS1.SSS3">
  <label>4.1.3</label><title>Object-based diagnostic evaluation</title>
      <p id="d1e2651">To fully access the performance of the models in predicting spatial precipitation attributes, i.e., area, location, and shape, the MODE verification framework is applied.
In the following, we present conditional quantile plots for the object area, for the location of the object centroid in east–west and north–south directions, for the aspect ratio, and for the orientation angle of the precipitation objects to show more details of predicted and observed precipitation objects. These plots visualize the joint distribution of the predictions and forecasts in a compact manner by applying a factorization into a conditional and marginal distribution <xref ref-type="bibr" rid="bib1.bibx34 bib1.bibx52" id="paren.69"/>. Figure <xref ref-type="fig" rid="Ch1.F5"/> illustrates the joint distribution in terms of the likelihood base-rate factorization. While the solid lines illustrate the forecasts conditioned on the observations for all models, the marginal distribution of the observations is plotted as a histogram.
Figure <xref ref-type="fig" rid="Ch1.F5"/>a shows the number of grid points of the observed and predicted precipitation objects, which represents the area of precipitation objects.
It can be seen that CLGAN and PredRNN-v2 are able to capture object area fairly well.
Only objects consisting of 90 to 130 grid points (<inline-formula><mml:math id="M100" display="inline"><mml:mrow><mml:mo>∼</mml:mo><mml:mn mathvariant="normal">16</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mn mathvariant="normal">000</mml:mn></mml:mrow></mml:math></inline-formula> km<inline-formula><mml:math id="M101" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula>) are slightly underestimated (see Fig. <xref ref-type="fig" rid="Ch1.F5"/>a). However, the other competing models perform remarkably worse.
Figure <xref ref-type="fig" rid="Ch1.F5"/>b and c show the distance between the centroid of the precipitation object and the western boundary and the southern boundary, which is again measured by the number of grid points. It is seen that the location of object centroids is generally well captured by all models.
Stronger deviations are visible near the lateral boundaries, especially in  the  east–west direction.
The aspect ratio and the orientation angle are used to assess the predicted precipitation shape in Fig. <xref ref-type="fig" rid="Ch1.F5"/>d and e.
Here, the aspect ratio is the ratio of the shorter to the longer edge of the precipitation objects. The orientation angle constitutes the angle between the precipitation objects and positive <inline-formula><mml:math id="M102" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> axis.
CLGAN shows slight improvements over the other models in that the central parts of the orientation angle and aspect ratio are well calibrated. However, larger deviations from the <inline-formula><mml:math id="M103" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> reference line are obtained near the tails of the conditional distributions. This indicates that further work on  the simulation of precipitation shape is required.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F5" specific-use="star"><?xmltex \currentcnt{5}?><?xmltex \def\figurename{Figure}?><label>Figure 5</label><caption><p id="d1e2711">Conditional quantile plots in terms of the likelihood base-rate factorization for <bold>(a)</bold> area, <bold>(b)</bold> east–west centroid and <bold>(c)</bold> north–south centroid locations, <bold>(d)</bold> aspect ratio, and <bold>(e)</bold> orientation angle at the lead time of 60 min. The solid black line is the <inline-formula><mml:math id="M104" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> reference line. The marginal distribution of the observations is presented as a histogram.</p></caption>
            <?xmltex \igopts{width=455.244094pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f05.png"/>

          </fig>

</sec>
</sec>
<sec id="Ch1.S4.SS2">
  <label>4.2</label><title>Case study</title>
      <p id="d1e2757">To gain further insight into the realism of our predictions, a heavy precipitation event occurring on 12 June 2019 is visualized as an example (see Fig. <xref ref-type="fig" rid="Ch1.F6"/>) to compare the model performance with an “eyeball” analysis.
Figure <xref ref-type="fig" rid="Ch1.F6"/>a shows the observed precipitation rates in millimeters per 10 min for every 20 min over the forecast period starting at 06:50 CST. It is seen that a fairly strong precipitation system moves from west to east while it further intensifies.
The predictions of the different models are presented as difference plots in Fig. <xref ref-type="fig" rid="Ch1.F6"/>b–f. For the first 60 min, persistence and DenseRotation show up with the smallest discrepancies. However, for longer lead times, clear dipole structures in the difference plots indicate that the movement of the system is not captured. Thus, the Lagrangian persistence framework is inaccurate for longer lead times, and more advanced models are required to capture the long-term dependence.</p>
      <p id="d1e2766">While the deep learning models also show increasing differences with longer lead times, they perform better in capturing the movement and the intensification of the precipitation system (see Fig. <xref ref-type="fig" rid="Ch1.F6"/>d–f). PredRNN-v2 tends to overestimate the precipitation intensity, which causes  large coherent areas of positive differences.
The averaged RMSE over the study area of the PredRNN-v2 forecasts at the lead time<?pagebreak page2747?> of 2 h is around 0.37 mm.
ConvLSTM and CLGAN perform even better with smaller discrepancies, with a lower RMSE of 0.33 and 0.34 mm, respectively.
Consistent with the quantitative evaluation results, ConvLSTM outperforms the CLGAN model in terms of RMSE in this given case, whereas CLGAN obtains a higher CSI and ETS.
The CSI of the CLGAN forecasts with a threshold of 8 mm h<inline-formula><mml:math id="M105" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> at the 2 h lead time is around 0.11, 0.01 higher than the ConvLSTM model.
The results demonstrate that the prediction by ConvLSTM has a lower bias, while CLGAN can capture more heavy precipitation grid points.
Compared to ConvLSTM, the difference plot of CLGAN contains more fine structures which indicates that CLGAN can generate more details of the precipitation system.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F6" specific-use="star"><?xmltex \currentcnt{6}?><?xmltex \def\figurename{Figure}?><label>Figure 6</label><caption><p id="d1e2785">A case study for a rain system moving from west to east while intensifying. <bold>(a)</bold> Observation (Obse). The predictions of all models are illustrated with difference plots: <bold>(b)</bold> persistence (Persi), <bold>(c)</bold> DenseRotation (Dense), <bold>(d)</bold> ConvLSTM (ConvL), <bold>(e)</bold> PredRNN-v2 (PredR), and <bold>(f)</bold> CLGAN.
The initial time of the prediction period is 06:50 CST on 12 June 2019.</p></caption>
          <?xmltex \igopts{width=483.69685pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f06.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS3">
  <label>4.3</label><title>Ablation study</title>
      <p id="d1e2821">As shown in Eq. (<xref ref-type="disp-formula" rid="Ch1.E3"/>), the loss function used in our CLGAN model consists of two terms: the adversarial loss <inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mi mathvariant="normal">GAN</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> and the reconstruction loss <inline-formula><mml:math id="M107" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula>. To assess the contribution of each loss term on the forecasts' performance, sensitivity experiments on the weight of the reconstruction loss in CLGAN were carried out. Larger weight of the adversarial loss (equals smaller weight of reconstruction loss) is equivalent to a stronger contribution made by the GAN component.
Figure <xref ref-type="fig" rid="Ch1.F7"/> presents the results of the CLGAN model with different weights <inline-formula><mml:math id="M108" display="inline"><mml:mi mathvariant="italic">λ</mml:mi></mml:math></inline-formula> assigned to the <inline-formula><mml:math id="M109" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss. It is seen that the RMSE is increased when reducing the weight of <inline-formula><mml:math id="M110" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss (Fig. <xref ref-type="fig" rid="Ch1.F7"/>a). However, the scores for dichotomous events and fuzzy verification framework reveal improvements for smaller <inline-formula><mml:math id="M111" display="inline"><mml:mi mathvariant="italic">λ</mml:mi></mml:math></inline-formula>. The model using a pure reconstruction loss (<inline-formula><mml:math id="M112" display="inline"><mml:mrow><mml:mi mathvariant="italic">λ</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> in Eq. <xref ref-type="disp-formula" rid="Ch1.E3"/>) performs significantly worse than the model applying an adversarial loss in terms of CSI and FSS (Fig. <xref ref-type="fig" rid="Ch1.F7"/>b and c). Similar results are obtained in terms of the OTS (see Fig. <xref ref-type="fig" rid="Ch1.F7"/>d).
The results of sensitivity experiments indicate that the adversarial training with the GAN component encourages the model to generate forecasts which are more similar to the ground truth data. Despite a slight increase in RMSE, a relatively stronger contribution of the GAN component helps to capture the statistical properties of the observed precipitation (on the tail, as well as their spatial attributes) which in turn improves the prediction of strong precipitation events.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F7" specific-use="star"><?xmltex \currentcnt{7}?><?xmltex \def\figurename{Figure}?><label>Figure 7</label><caption><p id="d1e2910">Mean scores (<bold>a</bold>: RMSE; <bold>b</bold>: CSI; <bold>c</bold>: FSS; and <bold>d</bold>: OTS) for all lead times over the verification period for different weights of the <inline-formula><mml:math id="M113" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss in CLGAN. The terms “weight1”, “weight999”, “weight99”, and “weight9” denote the weights of the <inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss component <inline-formula><mml:math id="M115" display="inline"><mml:mi mathvariant="italic">λ</mml:mi></mml:math></inline-formula> with the corresponding values of 1, 0.999, 0.99, and 0.9, respectively. As in Fig. <xref ref-type="fig" rid="Ch1.F3"/>, <inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">Pr</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">8</mml:mn></mml:mrow></mml:math></inline-formula> mm h<inline-formula><mml:math id="M117" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> is chosen for CSI, OTS, and FSS, together with setting <inline-formula><mml:math id="M118" display="inline"><mml:mi>s</mml:mi></mml:math></inline-formula> to five grid points for the latter. The area threshold <inline-formula><mml:math id="M119" display="inline"><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mi mathvariant="normal">A</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> of OTS is set to nine grid points.</p></caption>
          <?xmltex \igopts{width=483.69685pt}?><graphic xlink:href="https://gmd.copernicus.org/articles/16/2737/2023/gmd-16-2737-2023-f07.png"/>

        </fig>

</sec>
</sec>
<sec id="Ch1.S5" sec-type="conclusions">
  <label>5</label><title>Conclusion and discussion</title>
      <p id="d1e3019">A novel architecture CLGAN is proposed in this work which leverages the merits of U-Net, ConvLSTM, and GAN components to generate high-quality precipitation predictions up to 2 h over Guizhou, China.
The Eulerian persistence is used as the reference model to compare against the conventional optical flow method DenseRotation, as well as two competing video prediction models (ConvLSTM and PredRNN-v2). A Guizhou AWS_ML precipitation dataset is set up for the nowcasting task based on minute-level precipitation measurements of AWS gauges. The model performance is comprehensively evaluated by a series of domain-specific evaluation metrics, including point-by-point and object-based verification methods. The results demonstrate that DL-based video prediction models are generally superior to the conventional methods, especially for the lead times exceeding 60 min.
However, the use of grid-point-level losses (e.g., <inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> or <inline-formula><mml:math id="M121" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="script">L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> loss) diminishes their capability to capture heavy precipitation events. Since heavy precipitation events are strongly under-represented in the data during training, the models optimized solely on grid-point-level losses favor<?pagebreak page2749?> predicting weak precipitation rates to avoid large error contributions from strong precipitation events (double penalty problem).
By contrast, the GAN component of CLGAN encourages the generator to create predictions that share the statistical properties of observed precipitation, which makes it superior to the baseline and the competing models in dichotomous and spatial scores for heavy precipitation events.</p>
      <p id="d1e3044">Compared to the conventional methods, our results indicate that video prediction models with deep neural networks have a better capability of learning abstractions from data, which in turn can improve the prediction of complex evolving systems.
By learning the statistical dependency within the continuous sequence of precipitation data, video prediction models can simulate the precipitation patterns up to 2 h ahead fairly well. Since NWP models suffer from the spin-up issue in the first 6 h and the conventional approaches fail to capture long-term dependency, video prediction models show potential as a promising and reliable way for precipitation nowcasting.
However, a model performance degradation is expected for longer lead time, i.e., after 2 h, due to the error accumulation in the auto-regressive prediction task.
The quick evolution of the convective precipitation systems is furthermore challenging.
Our results also demonstrate that it is arduous to capture the shape of precipitation patterns by DL-based models as demonstrated by the MODE scores. Future work may try to integrate domain-specific evaluation metrics for spatial forecasts (e.g., FSS and MODE) as a loss function in DL-based models for precipitation nowcasting.
Additionally, we also see that a trade-off exists between evaluations on grid-point and object-based levels when the adversarial loss is varied. A grid search for the optimal combination of loss function coefficients is required to generate realistic forecasts with a low bias.</p>
      <p id="d1e3047">Beyond that, it is appealing to embed more predictors which could be retrieved from NWP models, e.g., the vertical velocity, water vapor, and thermal and other environmental conditions. The literature shows that the corresponding predictors and physical constraints can greatly improve the simulation of the targeted variable <xref ref-type="bibr" rid="bib1.bibx7 bib1.bibx16" id="paren.70"/>.
A careful selection of the predictors and an appropriate embedding solution are subject to our future work.
In addition, GAN models can easily be adapted to a probabilistic framework. By adding noise as an additional input, ensemble forecasts can be obtained from which a quantification of the forecast uncertainty can be deduced  <xref ref-type="bibr" rid="bib1.bibx33" id="paren.71"/>.
A probabilistic nowcasting system is appealing due to the strong inherent uncertainties in the dynamics of precipitation patterns. Furthermore, note that ensemble<?pagebreak page2750?> predictions corresponding to several future realizations provide the possibility to issue more strong precipitation events (cf. <xref ref-type="bibr" rid="bib1.bibx37" id="altparen.72"/>).
While this study focused on assessing the need for an adversarial loss formulation, future work will be directed towards a probabilistic nowcasting system.</p>
</sec>

      
      </body>
    <back><notes notes-type="codedataavailability"><title>Code and data availability</title>

      <p id="d1e3063">The Guizhou AWS_ML precipitation dataset and the exact version of the video prediction models used in this paper are archived on Zenodo: <ext-link xlink:href="https://doi.org/10.5281/zenodo.7278016" ext-link-type="DOI">10.5281/zenodo.7278016</ext-link> <xref ref-type="bibr" rid="bib1.bibx22" id="paren.73"/>. A frozen code repository can be obtained here: <uri>https://gitlab.jsc.fz-juelich.de/esde/machine-learning/ambs/-/tree/ambs_gmd_nowcasting_v1.0</uri> (last access: 25 June 2022). The dataset and scripts can help users to reproduce the results on their local machines or high-performance computers. By using these data and models, it is highly recommended to follow the README.md file of the code repository to run the end-to-end workflow.</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d1e3078">YJ, BG, and ML contributed equally to this work. YJ and BG contributed to the method development and performed the experiments. YJ wrote the manuscript draft, and all authors reviewed and edited the manuscript in several iterations.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d1e3084">The contact author has declared that none of the authors has any competing interests.</p>
  </notes><notes notes-type="disclaimer"><title>Disclaimer</title>

      <p id="d1e3090">Publisher's note: Copernicus Publications remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.</p>
  </notes><notes notes-type="sistatement"><title>Special issue statement</title>

      <p id="d1e3096">This article is part of the special issue “Benchmark datasets and machine learning algorithms for Earth system science data (ESSD/GMD inter-journal SI)”. It is not associated with a conference.</p>
  </notes><ack><title>Acknowledgements</title><p id="d1e3102">The authors acknowledge funding from  the DeepRain project under grant agreement 01 IS18047A from the Bundesministerium für Bildung und Forschung (BMBF), from the European Union H2020 MAELSTROM project (grant no. 955513, co-funding by BMBF), and from the ERC Advanced Grant IntelliAQ (grant no. 787576). We thank Dexuan Kong for preparing the datasets used in our research, as well as Martin G. Schultz for the helpful scientific discussions.</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d1e3107">This research has been supported by the Bundesministerium für Bildung und Forschung (grant nos. 955513 and 01 IS18047A) and the ERC Advanced Grant IntelliAQ (grant no. 787576).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d1e3113">This paper was edited by Nicola Bodini and reviewed by two anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><?xmltex \def\ref@label{{Austin and Bellon(1974)}}?><label>Austin and Bellon(1974)</label><?label austin1974use?><mixed-citation>Austin, G. and Bellon, A.: The use of digital weather radar records for
short-term precipitation forecasting, Q. J. Roy.
Meteor. Soc., 100, 658–664,
<ext-link xlink:href="https://doi.org/10.1002/qj.49710042612" ext-link-type="DOI">10.1002/qj.49710042612</ext-link>, 1974.</mixed-citation></ref>
      <ref id="bib1.bibx2"><?xmltex \def\ref@label{{Ayzel et~al.(2019)}}?><label>Ayzel et al.(2019)</label><?label ayzel2019optical?><mixed-citation>Ayzel, G., Heistermann, M., and Winterrath, T.: Optical flow models as an open benchmark for radar-based precipitation nowcasting (rainymotion v0.1), Geosci. Model Dev., 12, 1387–1402, <ext-link xlink:href="https://doi.org/10.5194/gmd-12-1387-2019" ext-link-type="DOI">10.5194/gmd-12-1387-2019</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx3"><?xmltex \def\ref@label{{Ayzel et~al.(2020)}}?><label>Ayzel et al.(2020)</label><?label ayzel2020rainnet?><mixed-citation>Ayzel, G., Scheffer, T., and Heistermann, M.: RainNet v1.0: a convolutional neural network for radar-based precipitation nowcasting, Geosci. Model Dev., 13, 2631–2644, <ext-link xlink:href="https://doi.org/10.5194/gmd-13-2631-2020" ext-link-type="DOI">10.5194/gmd-13-2631-2020</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx4"><?xmltex \def\ref@label{{Bowler et~al.(2004)}}?><label>Bowler et al.(2004)</label><?label bowler2004development?><mixed-citation>Bowler, N. E., Pierce, C. E., and Seed, A.: Development of a precipitation
nowcasting algorithm based upon optical flow techniques, J. Hydrol., 288, 74–91,
<ext-link xlink:href="https://doi.org/10.1016/j.jhydrol.2003.11.011" ext-link-type="DOI">10.1016/j.jhydrol.2003.11.011</ext-link>, 2004.</mixed-citation></ref>
      <ref id="bib1.bibx5"><?xmltex \def\ref@label{{Bowler et~al.(2006)}}?><label>Bowler et al.(2006)</label><?label bowler2006steps?><mixed-citation>Bowler, N. E., Pierce, C. E., and Seed, A. W.: STEPS: A probabilistic
precipitation forecasting scheme which merges an extrapolation nowcast with
downscaled NWP, Q. J. Roy. Meteor. Soc., 132, 2127–2155, <ext-link xlink:href="https://doi.org/10.1256/qj.04.100" ext-link-type="DOI">10.1256/qj.04.100</ext-link>,
2006.</mixed-citation></ref>
      <ref id="bib1.bibx6"><?xmltex \def\ref@label{{Davis et~al.(2006)}}?><label>Davis et al.(2006)</label><?label davis2006object?><mixed-citation>Davis, C., Brown, B., and Bullock, R.: Object-based verification of
precipitation forecasts. Part I: Methodology and application to mesoscale
rain areas, Mon. Weather Rev., 134, 1772–1784,
<ext-link xlink:href="https://doi.org/10.1175/MWR3145.1" ext-link-type="DOI">10.1175/MWR3145.1</ext-link>, 2006.</mixed-citation></ref>
      <ref id="bib1.bibx7"><?xmltex \def\ref@label{{Daw et~al.(2017)}}?><label>Daw et al.(2017)</label><?label daw2017physics?><mixed-citation>Daw, A., Karpatne, A., Watkins, W. D., Read, J. S., and Kumar, V.:
Physics-guided neural networks (pgnn): An application in lake temperature
modeling, in: Knowledge-Guided Machine Learning, 353–372, Chapman and
Hall/CRC, <ext-link xlink:href="https://doi.org/10.1201/9781003143376-15" ext-link-type="DOI">10.1201/9781003143376-15</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx8"><?xmltex \def\ref@label{{Dixon and Wiener(1993)}}?><label>Dixon and Wiener(1993)</label><?label dixon1993titan?><mixed-citation>Dixon, M. and Wiener, G.: TITAN: Thunderstorm identification, tracking,
analysis, and nowcasting – A radar-based methodology, J. Atmos. Ocean. Tech., 10, 785–797,
<ext-link xlink:href="https://doi.org/10.1175/1520-0426(1993)010&lt;0785:TTITAA&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0426(1993)010&lt;0785:TTITAA&gt;2.0.CO;2</ext-link>,
1993.</mixed-citation></ref>
      <ref id="bib1.bibx9"><?xmltex \def\ref@label{{Drozdzal et~al.(2016)}}?><label>Drozdzal et al.(2016)</label><?label drozdzal2016importance?><mixed-citation>Drozdzal, M., Vorontsov, E., Chartrand, G., Kadoury, S., and Pal, C.: The
importance of skip connections in biomedical image segmentation, in: Deep
learning and data labeling for medical applications, 179–187, Springer,
<ext-link xlink:href="https://doi.org/10.1007/978-3-319-46976-8_19" ext-link-type="DOI">10.1007/978-3-319-46976-8_19</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx10"><?xmltex \def\ref@label{{Ebert(2008)}}?><label>Ebert(2008)</label><?label ebert2008fuzzy?><mixed-citation>Ebert, E. E.: Fuzzy verification of high-resolution gridded forecasts: a review
and proposed framework, Meteorol. Appl., 15,
51–64, <ext-link xlink:href="https://doi.org/10.1002/met.25" ext-link-type="DOI">10.1002/met.25</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx11"><?xmltex \def\ref@label{{Ebert et~al.(2017)}}?><label>Ebert et al.(2017)</label><?label ebert2017self?><mixed-citation>
Ebert, F., Finn, C., Lee, A. X., and Levine, S.: Self-Supervised Visual Planning with Temporal Skip Connections, in: CoRL, arXiv preprint arXiv:1710.05268, 344–356, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx12"><?xmltex \def\ref@label{{Efron and Tibshirani(1994)}}?><label>Efron and Tibshirani(1994)</label><?label efron1994introduction?><mixed-citation>Efron, B. and Tibshirani, R. J.: An introduction to the bootstrap, CRC press,
<ext-link xlink:href="https://doi.org/10.1201/9780429246593" ext-link-type="DOI">10.1201/9780429246593</ext-link>, 1994.</mixed-citation></ref>
      <ref id="bib1.bibx13"><?xmltex \def\ref@label{{Ganguly and Bras(2003)}}?><label>Ganguly and Bras(2003)</label><?label ganguly2003distributed?><mixed-citation>Ganguly, A. R. and Bras, R. L.: Distributed quantitative precipitation
forecasting using information from radar and numerical weather prediction
models, J. Hydrometeorol., 4, 1168–1180,
<ext-link xlink:href="https://doi.org/10.1175/1525-7541(2003)004&lt;1168:DQPFUI&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1525-7541(2003)004&lt;1168:DQPFUI&gt;2.0.CO;2</ext-link>,
2003.</mixed-citation></ref>
      <ref id="bib1.bibx14"><?xmltex \def\ref@label{{Garcia-Garcia et~al.(2018)}}?><label>Garcia-Garcia et al.(2018)</label><?label garcia2018robotrix?><mixed-citation>Garcia-Garcia, A., Martinez-Gonzalez, P., Oprea, S., Cast<?pagebreak page2751?>ro-Vargas, J. A.,
Orts-Escolano, S., Garcia-Rodriguez, J., and Jover-Alvarez, A.: The robotrix:
An extremely photorealistic and very-large-scale indoor dataset of sequences
with robot trajectories and interactions, in: 2018 IEEE/RSJ International
Conference on Intelligent Robots and Systems (IROS), 6790–6797, IEEE,
<ext-link xlink:href="https://doi.org/10.1109/IROS.2018.8594495" ext-link-type="DOI">10.1109/IROS.2018.8594495</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx15"><?xmltex \def\ref@label{{Germann and Zawadzki(2002)}}?><label>Germann and Zawadzki(2002)</label><?label germann2002scale?><mixed-citation>Germann, U. and Zawadzki, I.: Scale-dependence of the predictability of
precipitation from continental radar images. Part I: Description of the
methodology, Mon. Weather Rev., 130, 2859–2873,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(2002)130&lt;2859:SDOTPO&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(2002)130&lt;2859:SDOTPO&gt;2.0.CO;2</ext-link>,
2002.</mixed-citation></ref>
      <ref id="bib1.bibx16"><?xmltex \def\ref@label{{Gong et~al.(2022)}}?><label>Gong et al.(2022)</label><?label gong2022temperature?><mixed-citation>Gong, B., Langguth, M., Ji, Y., Mozaffari, A., Stadtler, S., Mache, K., and Schultz, M. G.: Temperature forecasting by deep learning methods, Geosci. Model Dev., 15, 8931–8956, <ext-link xlink:href="https://doi.org/10.5194/gmd-15-8931-2022" ext-link-type="DOI">10.5194/gmd-15-8931-2022</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx17"><?xmltex \def\ref@label{{Goodfellow et~al.(2020)}}?><label>Goodfellow et al.(2020)</label><?label goodfellow2020generative?><mixed-citation>Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair,
S., Courville, A., and Bengio, Y.: Generative adversarial networks,
Communications of the ACM, 63, 139–144,
<ext-link xlink:href="https://doi.org/10.1145/3422622" ext-link-type="DOI">10.1145/3422622</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx18"><?xmltex \def\ref@label{{Grecu and Krajewski(2000)}}?><label>Grecu and Krajewski(2000)</label><?label grecu2000large?><mixed-citation>Grecu, M. and Krajewski, W.: A large-sample investigation of statistical
procedures for radar-based short-term quantitative precipitation forecasting,
J. Hydrol., 239, 69–84,
<ext-link xlink:href="https://doi.org/10.1016/S0022-1694(00)00360-7" ext-link-type="DOI">10.1016/S0022-1694(00)00360-7</ext-link>, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx19"><?xmltex \def\ref@label{{Harris et~al.(2022)}}?><label>Harris et al.(2022)</label><?label harris2022generative?><mixed-citation>Harris, L., McRae, A. T., Chantry, M., Dueben, P. D., and Palmer, T. N.: A
Generative Deep Learning Approach to Stochastic Downscaling of Precipitation
Forecasts, arXiv preprint arXiv:2204.02028,
<ext-link xlink:href="https://doi.org/10.1029/2022MS003120" ext-link-type="DOI">10.1029/2022MS003120</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx20"><?xmltex \def\ref@label{{Hu et~al.(2020)}}?><label>Hu et al.(2020)</label><?label hu2020probabilistic?><mixed-citation>Hu, A., Cotter, F., Mohan, N., Gurau, C., and Kendall, A.: Probabilistic future
prediction for video scene understanding, in: European Conference on Computer
Vision, 767–785, Springer,
<ext-link xlink:href="https://doi.org/10.1007/978-3-030-58517-4_45" ext-link-type="DOI">10.1007/978-3-030-58517-4_45</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx21"><?xmltex \def\ref@label{{Ji et~al.(2020)}}?><label>Ji et al.(2020)</label><?label ji2020multimodel?><mixed-citation>Ji, L., Zhi, X., Simmer, C., Zhu, S., and Ji, Y.: Multimodel ensemble forecasts
of precipitation based on an object-based diagnostic evaluation, Mon. Weather Rev., 148, 2591–2606,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-19-0266.1" ext-link-type="DOI">10.1175/MWR-D-19-0266.1</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx22"><?xmltex \def\ref@label{Ji et al.(2022)}?><label>Ji et al.(2022)</label><?label Ji2022data?><mixed-citation>Ji, Y., Gong, B., Langguth, M., Mozaffari, A., and Kong, D.: CLGAN: Guizhou ML-AWS precipitation dataset (1.0), Zenodo [data set], <ext-link xlink:href="https://doi.org/10.5281/zenodo.7278016" ext-link-type="DOI">10.5281/zenodo.7278016</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx23"><?xmltex \def\ref@label{{Johnson and Wang(2012)}}?><label>Johnson and Wang(2012)</label><?label johnson2012object?><mixed-citation>Johnson, A. and Wang, X.: Object-based evaluation of a storm-scale ensemble
during the 2009 NOAA Hazardous Weather Testbed Spring Experiment, Mon. Weather Rev., 141, 1079–1098,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-12-00140.1" ext-link-type="DOI">10.1175/MWR-D-12-00140.1</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx24"><?xmltex \def\ref@label{{Johnson et~al.(2013)}}?><label>Johnson et al.(2013)</label><?label johnson2013object?><mixed-citation>Johnson, A., Wang, X., Kong, F., and Xue, M.: Object-based evaluation of the
impact of horizontal grid spacing on convection-allowing forecasts, Mon. Weather Rev., 141, 3413–3425,
<ext-link xlink:href="https://doi.org/10.1175/MWR-D-13-00027.1" ext-link-type="DOI">10.1175/MWR-D-13-00027.1</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx25"><?xmltex \def\ref@label{{Johnson et~al.(1998)}}?><label>Johnson et al.(1998)</label><?label johnson1998storm?><mixed-citation>Johnson, J., MacKeen, P. L., Witt, A., Mitchell, E. D. W., Stumpf, G. J.,
Eilts, M. D., and Thomas, K. W.: The storm cell identification and tracking
algorithm: An enhanced WSR-88D algorithm, Weather Forecast., 13,
263–276,
<ext-link xlink:href="https://doi.org/10.1175/1520-0434(1998)013&lt;0263:TSCIAT&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0434(1998)013&lt;0263:TSCIAT&gt;2.0.CO;2</ext-link>,
1998.</mixed-citation></ref>
      <ref id="bib1.bibx26"><?xmltex \def\ref@label{{Kingma and Ba(2014)}}?><label>Kingma and Ba(2014)</label><?label kingma2014adam?><mixed-citation>Kingma, D. P. and Ba, J.: Adam: A method for stochastic optimization, arXiv
preprint arXiv:1412.6980, <ext-link xlink:href="https://doi.org/10.48550/arXiv.1412.6980" ext-link-type="DOI">10.48550/arXiv.1412.6980</ext-link>,
2014.</mixed-citation></ref>
      <ref id="bib1.bibx27"><?xmltex \def\ref@label{{Kroeger et~al.(2016)}}?><label>Kroeger et al.(2016)</label><?label kroeger2016fast?><mixed-citation>Kroeger, T., Timofte, R., Dai, D., and Van Gool, L.: Fast optical flow using
dense inverse search, in: European Conference on Computer Vision,
471–488, Springer, <ext-link xlink:href="https://doi.org/10.1007/978-3-319-46493-0_29" ext-link-type="DOI">10.1007/978-3-319-46493-0_29</ext-link>,
2016.</mixed-citation></ref>
      <ref id="bib1.bibx28"><?xmltex \def\ref@label{{Leinonen et~al.(2020)}}?><label>Leinonen et al.(2020)</label><?label leinonen2020stochastic?><mixed-citation>Leinonen, J., Nerini, D., and Berne, A.: Stochastic super-resolution for
downscaling time-evolving atmospheric fields with a generative adversarial
network, IEEE T. Geosci. Remote,
<ext-link xlink:href="https://doi.org/10.1109/TGRS.2020.3032790" ext-link-type="DOI">10.1109/TGRS.2020.3032790</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx29"><?xmltex \def\ref@label{{Li et~al.(2021)}}?><label>Li et al.(2021)</label><?label li2021msdm?><mixed-citation>Li, D., Liu, Y., and Chen, C.: MSDM v1.0: A machine learning model for precipitation nowcasting over eastern China using multisource data, Geosci. Model Dev., 14, 4019–4034, <ext-link xlink:href="https://doi.org/10.5194/gmd-14-4019-2021" ext-link-type="DOI">10.5194/gmd-14-4019-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx30"><?xmltex \def\ref@label{{Liu et~al.(2018)}}?><label>Liu et al.(2018)</label><?label liu2018future?><mixed-citation>Liu, W., Luo, W., Lian, D., and Gao, S.: Future frame prediction for anomaly
detection–a new baseline, in: Proceedings of the IEEE conference on computer
vision and pattern recognition, 6536–6545,
<ext-link xlink:href="https://doi.org/10.1109/CVPR.2018.00684" ext-link-type="DOI">10.1109/CVPR.2018.00684</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx31"><?xmltex \def\ref@label{{Mathieu et~al.(2015)}}?><label>Mathieu et al.(2015)</label><?label mathieu2015deep?><mixed-citation>Mathieu, M., Couprie, C., and LeCun, Y.: Deep multi-scale video prediction
beyond mean square error, arXiv preprint arXiv:1511.05440,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.1511.05440" ext-link-type="DOI">10.48550/arXiv.1511.05440</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx32"><?xmltex \def\ref@label{{Matsunobu et~al.(2022)}}?><label>Matsunobu et al.(2022)</label><?label matsunobu2022impact?><mixed-citation>Matsunobu, T., Keil, C., and Barthlott, C.: The impact of microphysical uncertainty conditional on initial and boundary condition uncertainty under varying synoptic control, Weather Clim. Dynam., 3, 1273–1289, <ext-link xlink:href="https://doi.org/10.5194/wcd-3-1273-2022" ext-link-type="DOI">10.5194/wcd-3-1273-2022</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx33"><?xmltex \def\ref@label{{Mordido et~al.(2018)}}?><label>Mordido et al.(2018)</label><?label mordido2018dropout?><mixed-citation>Mordido, G., Yang, H., and Meinel, C.: Dropout-gan: Learning from a dynamic
ensemble of discriminators, arXiv preprint arXiv:1807.11346,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.1807.11346" ext-link-type="DOI">10.48550/arXiv.1807.11346</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx34"><?xmltex \def\ref@label{{Murphy and Winkler(1987)}}?><label>Murphy and Winkler(1987)</label><?label murphy1987general?><mixed-citation>Murphy, A. H. and Winkler, R. L.: A general framework for forecast
verification, Mon. Weather Rev., 115, 1330–1338,
<ext-link xlink:href="https://doi.org/10.1175/1520-0493(1987)115&lt;1330:AGFFFV&gt;2.0.CO;2" ext-link-type="DOI">10.1175/1520-0493(1987)115&lt;1330:AGFFFV&gt;2.0.CO;2</ext-link>,
1987.</mixed-citation></ref>
      <ref id="bib1.bibx35"><?xmltex \def\ref@label{{Oprea et~al.(2020)}}?><label>Oprea et al.(2020)</label><?label oprea2020review?><mixed-citation>Oprea, S., Martinez-Gonzalez, P., Garcia-Garcia, A., Castro-Vargas, J. A.,
Orts-Escolano, S., Garcia-Rodriguez, J., and Argyros, A.: A review on deep
learning techniques for video prediction, IEEE T. Pattern
Anal.,  44, 2806–2826,
<ext-link xlink:href="https://doi.org/10.1109/TPAMI.2020.3045007" ext-link-type="DOI">10.1109/TPAMI.2020.3045007</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx36"><?xmltex \def\ref@label{{Price and Rasp(2022)}}?><label>Price and Rasp(2022)</label><?label price2022increasing?><mixed-citation>Price, I. and Rasp, S.: Increasing the accuracy and resolution of precipitation
forecasts using deep generative models, arXiv preprint arXiv:2203.12297,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.2203.12297" ext-link-type="DOI">10.48550/arXiv.2203.12297</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx37"><?xmltex \def\ref@label{{Ravuri et~al.(2021)}}?><label>Ravuri et al.(2021)</label><?label ravuri2021skillful?><mixed-citation>Ravuri, S. V., Lenc, K., Willson, M., Kangin, D., Lam, R. R., Mirowski, P. W., Fitzsimons, M., Athanassiadou, M., Kashem, S., Madge, S., Prudden, R., Mandhane, A., Clark, A., Brock, A., Simonyan, K., Hadsell, R., Robinson, N. H., Clancy, E., Arribas, A., and Mohamed, S.:  Skillful Precipitation Nowcasting using Deep Generative Models of Radar, arXiv preprint arXiv:2104.00954, <ext-link xlink:href="https://doi.org/10.1038/s41586-021-03854-z" ext-link-type="DOI">10.1038/s41586-021-03854-z</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx38"><?xmltex \def\ref@label{{Reichstein et~al.(2019)}}?><label>Reichstein et al.(2019)</label><?label reichstein2019deep?><mixed-citation>Reichstein, M., Camps-Valls, G., Stevens, B., Jung, M., Denzler, J., Carvalhais, N., and Prabhat: Deep learning and process understanding for data-driven Earth system science, Nature, 566, 195–204, <ext-link xlink:href="https://doi.org/10.1038/s41586-019-0912-1" ext-link-type="DOI">10.1038/s41586-019-0912-1</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx39"><?xmltex \def\ref@label{{Rinehart and Garvey(1978)}}?><label>Rinehart and Garvey(1978)</label><?label rinehart1978three?><mixed-citation>Rinehart, R. and Garvey, E.: Three-dimensional storm motion detection by
conventional weather radar, Nature, 273, 287–289,
<ext-link xlink:href="https://doi.org/10.1038/273287a0" ext-link-type="DOI">10.1038/273287a0</ext-link>, 1978.</mixed-citation></ref>
      <ref id="bib1.bibx40"><?xmltex \def\ref@label{{Roberts(2008)}}?><label>Roberts(2008)</label><?label roberts2008assessing?><mixed-citation>Roberts, N.: Assessing the spatial and temporal variation in the skill of
precipitation forecasts from an NWP model, Meteorol. Appl., 15, 163–169, <ext-link xlink:href="https://doi.org/10.1002/met.57" ext-link-type="DOI">10.1002/met.57</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx41"><?xmltex \def\ref@label{{Roberts and Lean(2008)}}?><label>Roberts and Lean(2008)</label><?label roberts2008scale?><mixed-citation>Roberts, N. M. and Lean, H. W.: Scale-selective ver<?pagebreak page2752?>ification of rainfall
accumulations from high-resolution forecasts of convective events, Mon. Weather Rev., 136, 78–97, <ext-link xlink:href="https://doi.org/10.1175/2007MWR2123.1" ext-link-type="DOI">10.1175/2007MWR2123.1</ext-link>,
2008.</mixed-citation></ref>
      <ref id="bib1.bibx42"><?xmltex \def\ref@label{{Ronneberger et~al.(2015)}}?><label>Ronneberger et al.(2015)</label><?label ronneberger2015u?><mixed-citation>Ronneberger, O., Fischer, P., and Brox, T.: U-net: Convolutional networks for
biomedical image segmentation, in: International Conference on Medical image
computing and computer-assisted intervention, 234–241, Springer,
<ext-link xlink:href="https://doi.org/10.1007/978-3-319-24574-4_28" ext-link-type="DOI">10.1007/978-3-319-24574-4_28</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx43"><?xmltex \def\ref@label{{Schultz et~al.(2021)}}?><label>Schultz et al.(2021)</label><?label schultz2021can?><mixed-citation>Schultz, M., Betancourt, C., Gong, B., Kleinert, F., Langguth, M., Leufen, L.,
Mozaffari, A., and Stadtler, S.: Can deep learning beat numerical weather
prediction?, Philos. T. Roy. Soc. A, 379,
20200097, <ext-link xlink:href="https://doi.org/10.1098/rsta.2020.0097" ext-link-type="DOI">10.1098/rsta.2020.0097</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx44"><?xmltex \def\ref@label{{Sha et~al.(2020)}}?><label>Sha et al.(2020)</label><?label sha2020deep?><mixed-citation>Sha, Y., Gagne II, D. J., West, G., and Stull, R.: Deep-learning-based gridded
downscaling of surface meteorological variables in complex terrain. Part II:
Daily precipitation, Appl. Meteorol. Climatol., 59,
2075–2092, <ext-link xlink:href="https://doi.org/10.1175/JAMC-D-20-0058.1" ext-link-type="DOI">10.1175/JAMC-D-20-0058.1</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx45"><?xmltex \def\ref@label{{Shi et~al.(2015)}}?><label>Shi et al.(2015)</label><?label shi2015convolutional?><mixed-citation>Shi, X., Chen, Z., Wang, H., Yeung, D.-Y., Wong, W.-K., and Woo, W.-C.: Convolutional LSTM network: A machine learning approach for precipitation nowcasting, in: Advances in neural information processing systems, 802–810,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.1506.04214" ext-link-type="DOI">10.48550/arXiv.1506.04214</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx46"><?xmltex \def\ref@label{{Shi et~al.(2017)}}?><label>Shi et al.(2017)</label><?label shi2017deep?><mixed-citation>Shi, X., Gao, Z., Lausen, L., Wang, H., Yeung, D.-Y., Wong, W.-k., and Woo,
W.-C.: Deep learning for precipitation nowcasting: A benchmark and a new
model, arXiv preprint arXiv:1706.03458,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.1706.03458" ext-link-type="DOI">10.48550/arXiv.1706.03458</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx47"><?xmltex \def\ref@label{{S{\o}nderby et~al.(2020)}}?><label>Sønderby et al.(2020)</label><?label sonderby2020metnet?><mixed-citation>Sønderby, C. K., Espeholt, L., Heek, J., Dehghani, M., Oliver, A., Salimans,
T., Agrawal, S., Hickey, J., and Kalchbrenner, N.: Metnet: A neural weather
model for precipitation forecasting, arXiv preprint arXiv:2003.12140,
<ext-link xlink:href="https://doi.org/10.48550/arXiv.2003.12140" ext-link-type="DOI">10.48550/arXiv.2003.12140</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx48"><?xmltex \def\ref@label{{Sun et~al.(2014)}}?><label>Sun et al.(2014)</label><?label sun2014use?><mixed-citation>Sun, J., Xue, M., Wilson, J. W., Zawadzki, I., Ballard, S. P., onvlee hooiMeyer, J., Joe, P. I., Barker, D. M., Li, P.-W., Golding, B., Xu, M., and Pinto, J. O.: Use of NWP for nowcasting convective precipitation: Recent progress and challenges, B. Am. Meteorol. Soc., 95, 409–426, <ext-link xlink:href="https://doi.org/10.1175/BAMS-D-11-00263.1" ext-link-type="DOI">10.1175/BAMS-D-11-00263.1</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx49"><?xmltex \def\ref@label{{Vasiloff et~al.(2007)}}?><label>Vasiloff et al.(2007)</label><?label vasiloff2007improving?><mixed-citation>Vasiloff, S. V., Seo, D.-J., Howard, K. W., Zhang, J., Kitzmiller, D., Mullusky, M. G., Krajewski, W. F., Brandes, E., Rabin, R. M., Berkowitz, D. S., Brooks, H., McGinley, J. A., Kuligowski, R. J., and Brown, B: Improving QPE and very short term QPF: An initiative for a community-wide integrated approach, B. Am. Meteorol. Soc., 88, 1899–1911, <ext-link xlink:href="https://doi.org/10.1175/BAMS-88-12-1899" ext-link-type="DOI">10.1175/BAMS-88-12-1899</ext-link>, 2007.
</mixed-citation></ref><?xmltex \hack{\newpage}?>
      <ref id="bib1.bibx50"><?xmltex \def\ref@label{{Wang et~al.(2017)}}?><label>Wang et al.(2017)</label><?label wang2017predrnn?><mixed-citation>Wang, Y., Long, M., Wang, J., Gao, Z., and Philip, S. Y.: Predrnn: Recurrent neural networks for predictive learning using
spatiotemporal lstms, in: Advances in Neural Information Processing Systems, 879–888, <uri>https://proceedings.neurips.cc/paper/2017/hash/e5f6ad6ce374177eef023bf5d0c018b6-Abstract.html</uri>
(last access: 12 December 2021), 2017.</mixed-citation></ref>
      <ref id="bib1.bibx51"><?xmltex \def\ref@label{{Wang et~al.(2021)}}?><label>Wang et al.(2021)</label><?label wang2021predrnn?><mixed-citation>Wang, Y., Wu, H., Zhang, J., Gao, Z., Wang, J., Yu, P. S., and Long, M.:
PredRNN: A Recurrent Neural Network for Spatiotemporal Predictive Learning,
arXiv preprint arXiv:2103.09504,
<ext-link xlink:href="https://doi.org/10.1109/TPAMI.2022.3165153" ext-link-type="DOI">10.1109/TPAMI.2022.3165153</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx52"><?xmltex \def\ref@label{{Wilks(2011)}}?><label>Wilks(2011)</label><?label wilks2011statistical?><mixed-citation>
Wilks, D. S.: Statistical methods in the atmospheric sciences, vol. 100, Academic Press, ISBN 9780123850225, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx53"><?xmltex \def\ref@label{{Wilson et~al.(2010)}}?><label>Wilson et al.(2010)</label><?label wilson2010nowcasting?><mixed-citation>Wilson, J. W., Feng, Y., Chen, M., and Roberts, R. D.: Nowcasting challenges
during the Beijing Olympics: Successes, failures, and implications for future
nowcasting systems, Weather Forecast., 25, 1691–1714,
<ext-link xlink:href="https://doi.org/10.1175/2010WAF2222417.1" ext-link-type="DOI">10.1175/2010WAF2222417.1</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx54"><?xmltex \def\ref@label{{Wolberg(1990)}}?><label>Wolberg(1990)</label><?label wolberg1990digital?><mixed-citation>
Wolberg, G.: Digital image warping, vol. 10662, IEEE computer society press Los
Alamitos, CA, 1990.</mixed-citation></ref>
      <ref id="bib1.bibx55"><?xmltex \def\ref@label{{Woo and Wong(2017)}}?><label>Woo and Wong(2017)</label><?label woo2017operational?><mixed-citation>Woo, W.-C. and Wong, W.-K.: Operational application of optical flow techniques
to radar-based rainfall nowcasting, Atmosphere, 8, 48,
<ext-link xlink:href="https://doi.org/10.3390/atmos8030048" ext-link-type="DOI">10.3390/atmos8030048</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx56"><?xmltex \def\ref@label{{Xie et~al.(2019)}}?><label>Xie et al.(2019)</label><?label xie2019improved?><mixed-citation>Xie, S., Wang, Y.-C., Lin, W., Ma, H.-Y., Tang, Q., Tang, S., Zheng, X., Golaz,
J.-C., Zhang, G. J., and Zhang, M.: Improved diurnal cycle of precipitation
in E3SM with a revised convective triggering function, J. Adv.
Model. Earth Sy., 11, 2290–2310,
<ext-link xlink:href="https://doi.org/10.1029/2019MS001702" ext-link-type="DOI">10.1029/2019MS001702</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx57"><?xmltex \def\ref@label{{Zahraei et~al.(2012)}}?><label>Zahraei et al.(2012)</label><?label zahraei2012quantitative?><mixed-citation>Zahraei, A., Hsu, K.-l., Sorooshian, S., Gourley, J., Lakshmanan, V., Hong, Y.,
and Bellerby, T.: Quantitative precipitation nowcasting: A Lagrangian
pixel-based approach, Atmos. Res., 118, 418–434,
<ext-link xlink:href="https://doi.org/10.1016/j.atmosres.2012.07.001" ext-link-type="DOI">10.1016/j.atmosres.2012.07.001</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx58"><?xmltex \def\ref@label{{Zahraei et~al.(2013)}}?><label>Zahraei et al.(2013)</label><?label zahraei2013short?><mixed-citation>Zahraei, A., Hsu, K.-l., Sorooshian, S., Gourley, J. J., Hong, Y., and
Behrangi, A.: Short-term quantitative precipitation forecasting using an
object-based approach, J. Hydrol., 483, 1–15,
<ext-link xlink:href="https://doi.org/10.1016/j.jhydrol.2012.09.052" ext-link-type="DOI">10.1016/j.jhydrol.2012.09.052</ext-link>, 2013.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>CLGAN: a generative adversarial network (GAN)-based video prediction model for precipitation nowcasting</article-title-html>
<abstract-html/>
<ref-html id="bib1.bib1"><label>Austin and Bellon(1974)</label><mixed-citation>
      
Austin, G. and Bellon, A.: The use of digital weather radar records for
short-term precipitation forecasting, Q. J. Roy.
Meteor. Soc., 100, 658–664,
<a href="https://doi.org/10.1002/qj.49710042612" target="_blank">https://doi.org/10.1002/qj.49710042612</a>, 1974.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Ayzel et al.(2019)</label><mixed-citation>
      
Ayzel, G., Heistermann, M., and Winterrath, T.: Optical flow models as an open benchmark for radar-based precipitation nowcasting (rainymotion v0.1), Geosci. Model Dev., 12, 1387–1402, <a href="https://doi.org/10.5194/gmd-12-1387-2019" target="_blank">https://doi.org/10.5194/gmd-12-1387-2019</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Ayzel et al.(2020)</label><mixed-citation>
      
Ayzel, G., Scheffer, T., and Heistermann, M.: RainNet v1.0: a convolutional neural network for radar-based precipitation nowcasting, Geosci. Model Dev., 13, 2631–2644, <a href="https://doi.org/10.5194/gmd-13-2631-2020" target="_blank">https://doi.org/10.5194/gmd-13-2631-2020</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Bowler et al.(2004)</label><mixed-citation>
      
Bowler, N. E., Pierce, C. E., and Seed, A.: Development of a precipitation
nowcasting algorithm based upon optical flow techniques, J. Hydrol., 288, 74–91,
<a href="https://doi.org/10.1016/j.jhydrol.2003.11.011" target="_blank">https://doi.org/10.1016/j.jhydrol.2003.11.011</a>, 2004.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Bowler et al.(2006)</label><mixed-citation>
      
Bowler, N. E., Pierce, C. E., and Seed, A. W.: STEPS: A probabilistic
precipitation forecasting scheme which merges an extrapolation nowcast with
downscaled NWP, Q. J. Roy. Meteor. Soc., 132, 2127–2155, <a href="https://doi.org/10.1256/qj.04.100" target="_blank">https://doi.org/10.1256/qj.04.100</a>,
2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Davis et al.(2006)</label><mixed-citation>
      
Davis, C., Brown, B., and Bullock, R.: Object-based verification of
precipitation forecasts. Part I: Methodology and application to mesoscale
rain areas, Mon. Weather Rev., 134, 1772–1784,
<a href="https://doi.org/10.1175/MWR3145.1" target="_blank">https://doi.org/10.1175/MWR3145.1</a>, 2006.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Daw et al.(2017)</label><mixed-citation>
      
Daw, A., Karpatne, A., Watkins, W. D., Read, J. S., and Kumar, V.:
Physics-guided neural networks (pgnn): An application in lake temperature
modeling, in: Knowledge-Guided Machine Learning, 353–372, Chapman and
Hall/CRC, <a href="https://doi.org/10.1201/9781003143376-15" target="_blank">https://doi.org/10.1201/9781003143376-15</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Dixon and Wiener(1993)</label><mixed-citation>
      
Dixon, M. and Wiener, G.: TITAN: Thunderstorm identification, tracking,
analysis, and nowcasting – A radar-based methodology, J. Atmos. Ocean. Tech., 10, 785–797,
<a href="https://doi.org/10.1175/1520-0426(1993)010&lt;0785:TTITAA&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0426(1993)010&lt;0785:TTITAA&gt;2.0.CO;2</a>,
1993.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Drozdzal et al.(2016)</label><mixed-citation>
      
Drozdzal, M., Vorontsov, E., Chartrand, G., Kadoury, S., and Pal, C.: The
importance of skip connections in biomedical image segmentation, in: Deep
learning and data labeling for medical applications, 179–187, Springer,
<a href="https://doi.org/10.1007/978-3-319-46976-8_19" target="_blank">https://doi.org/10.1007/978-3-319-46976-8_19</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Ebert(2008)</label><mixed-citation>
      
Ebert, E. E.: Fuzzy verification of high-resolution gridded forecasts: a review
and proposed framework, Meteorol. Appl., 15,
51–64, <a href="https://doi.org/10.1002/met.25" target="_blank">https://doi.org/10.1002/met.25</a>, 2008.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Ebert et al.(2017)</label><mixed-citation>
      
Ebert, F., Finn, C., Lee, A. X., and Levine, S.: Self-Supervised Visual Planning with Temporal Skip Connections, in: CoRL, arXiv preprint arXiv:1710.05268, 344–356, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Efron and Tibshirani(1994)</label><mixed-citation>
      
Efron, B. and Tibshirani, R. J.: An introduction to the bootstrap, CRC press,
<a href="https://doi.org/10.1201/9780429246593" target="_blank">https://doi.org/10.1201/9780429246593</a>, 1994.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Ganguly and Bras(2003)</label><mixed-citation>
      
Ganguly, A. R. and Bras, R. L.: Distributed quantitative precipitation
forecasting using information from radar and numerical weather prediction
models, J. Hydrometeorol., 4, 1168–1180,
<a href="https://doi.org/10.1175/1525-7541(2003)004&lt;1168:DQPFUI&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1525-7541(2003)004&lt;1168:DQPFUI&gt;2.0.CO;2</a>,
2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Garcia-Garcia et al.(2018)</label><mixed-citation>
      
Garcia-Garcia, A., Martinez-Gonzalez, P., Oprea, S., Castro-Vargas, J. A.,
Orts-Escolano, S., Garcia-Rodriguez, J., and Jover-Alvarez, A.: The robotrix:
An extremely photorealistic and very-large-scale indoor dataset of sequences
with robot trajectories and interactions, in: 2018 IEEE/RSJ International
Conference on Intelligent Robots and Systems (IROS), 6790–6797, IEEE,
<a href="https://doi.org/10.1109/IROS.2018.8594495" target="_blank">https://doi.org/10.1109/IROS.2018.8594495</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Germann and Zawadzki(2002)</label><mixed-citation>
      
Germann, U. and Zawadzki, I.: Scale-dependence of the predictability of
precipitation from continental radar images. Part I: Description of the
methodology, Mon. Weather Rev., 130, 2859–2873,
<a href="https://doi.org/10.1175/1520-0493(2002)130&lt;2859:SDOTPO&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(2002)130&lt;2859:SDOTPO&gt;2.0.CO;2</a>,
2002.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Gong et al.(2022)</label><mixed-citation>
      
Gong, B., Langguth, M., Ji, Y., Mozaffari, A., Stadtler, S., Mache, K., and Schultz, M. G.: Temperature forecasting by deep learning methods, Geosci. Model Dev., 15, 8931–8956, <a href="https://doi.org/10.5194/gmd-15-8931-2022" target="_blank">https://doi.org/10.5194/gmd-15-8931-2022</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Goodfellow et al.(2020)</label><mixed-citation>
      
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair,
S., Courville, A., and Bengio, Y.: Generative adversarial networks,
Communications of the ACM, 63, 139–144,
<a href="https://doi.org/10.1145/3422622" target="_blank">https://doi.org/10.1145/3422622</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Grecu and Krajewski(2000)</label><mixed-citation>
      
Grecu, M. and Krajewski, W.: A large-sample investigation of statistical
procedures for radar-based short-term quantitative precipitation forecasting,
J. Hydrol., 239, 69–84,
<a href="https://doi.org/10.1016/S0022-1694(00)00360-7" target="_blank">https://doi.org/10.1016/S0022-1694(00)00360-7</a>, 2000.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Harris et al.(2022)</label><mixed-citation>
      
Harris, L., McRae, A. T., Chantry, M., Dueben, P. D., and Palmer, T. N.: A
Generative Deep Learning Approach to Stochastic Downscaling of Precipitation
Forecasts, arXiv preprint arXiv:2204.02028,
<a href="https://doi.org/10.1029/2022MS003120" target="_blank">https://doi.org/10.1029/2022MS003120</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Hu et al.(2020)</label><mixed-citation>
      
Hu, A., Cotter, F., Mohan, N., Gurau, C., and Kendall, A.: Probabilistic future
prediction for video scene understanding, in: European Conference on Computer
Vision, 767–785, Springer,
<a href="https://doi.org/10.1007/978-3-030-58517-4_45" target="_blank">https://doi.org/10.1007/978-3-030-58517-4_45</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Ji et al.(2020)</label><mixed-citation>
      
Ji, L., Zhi, X., Simmer, C., Zhu, S., and Ji, Y.: Multimodel ensemble forecasts
of precipitation based on an object-based diagnostic evaluation, Mon. Weather Rev., 148, 2591–2606,
<a href="https://doi.org/10.1175/MWR-D-19-0266.1" target="_blank">https://doi.org/10.1175/MWR-D-19-0266.1</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Ji et al.(2022)</label><mixed-citation>
      
Ji, Y., Gong, B., Langguth, M., Mozaffari, A., and Kong, D.: CLGAN: Guizhou ML-AWS precipitation dataset (1.0), Zenodo [data set], <a href="https://doi.org/10.5281/zenodo.7278016" target="_blank">https://doi.org/10.5281/zenodo.7278016</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Johnson and Wang(2012)</label><mixed-citation>
      
Johnson, A. and Wang, X.: Object-based evaluation of a storm-scale ensemble
during the 2009 NOAA Hazardous Weather Testbed Spring Experiment, Mon. Weather Rev., 141, 1079–1098,
<a href="https://doi.org/10.1175/MWR-D-12-00140.1" target="_blank">https://doi.org/10.1175/MWR-D-12-00140.1</a>, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Johnson et al.(2013)</label><mixed-citation>
      
Johnson, A., Wang, X., Kong, F., and Xue, M.: Object-based evaluation of the
impact of horizontal grid spacing on convection-allowing forecasts, Mon. Weather Rev., 141, 3413–3425,
<a href="https://doi.org/10.1175/MWR-D-13-00027.1" target="_blank">https://doi.org/10.1175/MWR-D-13-00027.1</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Johnson et al.(1998)</label><mixed-citation>
      
Johnson, J., MacKeen, P. L., Witt, A., Mitchell, E. D. W., Stumpf, G. J.,
Eilts, M. D., and Thomas, K. W.: The storm cell identification and tracking
algorithm: An enhanced WSR-88D algorithm, Weather Forecast., 13,
263–276,
<a href="https://doi.org/10.1175/1520-0434(1998)013&lt;0263:TSCIAT&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0434(1998)013&lt;0263:TSCIAT&gt;2.0.CO;2</a>,
1998.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Kingma and Ba(2014)</label><mixed-citation>
      
Kingma, D. P. and Ba, J.: Adam: A method for stochastic optimization, arXiv
preprint arXiv:1412.6980, <a href="https://doi.org/10.48550/arXiv.1412.6980" target="_blank">https://doi.org/10.48550/arXiv.1412.6980</a>,
2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Kroeger et al.(2016)</label><mixed-citation>
      
Kroeger, T., Timofte, R., Dai, D., and Van Gool, L.: Fast optical flow using
dense inverse search, in: European Conference on Computer Vision,
471–488, Springer, <a href="https://doi.org/10.1007/978-3-319-46493-0_29" target="_blank">https://doi.org/10.1007/978-3-319-46493-0_29</a>,
2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Leinonen et al.(2020)</label><mixed-citation>
      
Leinonen, J., Nerini, D., and Berne, A.: Stochastic super-resolution for
downscaling time-evolving atmospheric fields with a generative adversarial
network, IEEE T. Geosci. Remote,
<a href="https://doi.org/10.1109/TGRS.2020.3032790" target="_blank">https://doi.org/10.1109/TGRS.2020.3032790</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Li et al.(2021)</label><mixed-citation>
      
Li, D., Liu, Y., and Chen, C.: MSDM v1.0: A machine learning model for precipitation nowcasting over eastern China using multisource data, Geosci. Model Dev., 14, 4019–4034, <a href="https://doi.org/10.5194/gmd-14-4019-2021" target="_blank">https://doi.org/10.5194/gmd-14-4019-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Liu et al.(2018)</label><mixed-citation>
      
Liu, W., Luo, W., Lian, D., and Gao, S.: Future frame prediction for anomaly
detection–a new baseline, in: Proceedings of the IEEE conference on computer
vision and pattern recognition, 6536–6545,
<a href="https://doi.org/10.1109/CVPR.2018.00684" target="_blank">https://doi.org/10.1109/CVPR.2018.00684</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Mathieu et al.(2015)</label><mixed-citation>
      
Mathieu, M., Couprie, C., and LeCun, Y.: Deep multi-scale video prediction
beyond mean square error, arXiv preprint arXiv:1511.05440,
<a href="https://doi.org/10.48550/arXiv.1511.05440" target="_blank">https://doi.org/10.48550/arXiv.1511.05440</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Matsunobu et al.(2022)</label><mixed-citation>
      
Matsunobu, T., Keil, C., and Barthlott, C.: The impact of microphysical uncertainty conditional on initial and boundary condition uncertainty under varying synoptic control, Weather Clim. Dynam., 3, 1273–1289, <a href="https://doi.org/10.5194/wcd-3-1273-2022" target="_blank">https://doi.org/10.5194/wcd-3-1273-2022</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Mordido et al.(2018)</label><mixed-citation>
      
Mordido, G., Yang, H., and Meinel, C.: Dropout-gan: Learning from a dynamic
ensemble of discriminators, arXiv preprint arXiv:1807.11346,
<a href="https://doi.org/10.48550/arXiv.1807.11346" target="_blank">https://doi.org/10.48550/arXiv.1807.11346</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Murphy and Winkler(1987)</label><mixed-citation>
      
Murphy, A. H. and Winkler, R. L.: A general framework for forecast
verification, Mon. Weather Rev., 115, 1330–1338,
<a href="https://doi.org/10.1175/1520-0493(1987)115&lt;1330:AGFFFV&gt;2.0.CO;2" target="_blank">https://doi.org/10.1175/1520-0493(1987)115&lt;1330:AGFFFV&gt;2.0.CO;2</a>,
1987.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Oprea et al.(2020)</label><mixed-citation>
      
Oprea, S., Martinez-Gonzalez, P., Garcia-Garcia, A., Castro-Vargas, J. A.,
Orts-Escolano, S., Garcia-Rodriguez, J., and Argyros, A.: A review on deep
learning techniques for video prediction, IEEE T. Pattern
Anal.,  44, 2806–2826,
<a href="https://doi.org/10.1109/TPAMI.2020.3045007" target="_blank">https://doi.org/10.1109/TPAMI.2020.3045007</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Price and Rasp(2022)</label><mixed-citation>
      
Price, I. and Rasp, S.: Increasing the accuracy and resolution of precipitation
forecasts using deep generative models, arXiv preprint arXiv:2203.12297,
<a href="https://doi.org/10.48550/arXiv.2203.12297" target="_blank">https://doi.org/10.48550/arXiv.2203.12297</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Ravuri et al.(2021)</label><mixed-citation>
      
Ravuri, S. V., Lenc, K., Willson, M., Kangin, D., Lam, R. R., Mirowski, P. W., Fitzsimons, M., Athanassiadou, M., Kashem, S., Madge, S., Prudden, R., Mandhane, A., Clark, A., Brock, A., Simonyan, K., Hadsell, R., Robinson, N. H., Clancy, E., Arribas, A., and Mohamed, S.:  Skillful Precipitation Nowcasting using Deep Generative Models of Radar, arXiv preprint arXiv:2104.00954, <a href="https://doi.org/10.1038/s41586-021-03854-z" target="_blank">https://doi.org/10.1038/s41586-021-03854-z</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Reichstein et al.(2019)</label><mixed-citation>
      
Reichstein, M., Camps-Valls, G., Stevens, B., Jung, M., Denzler, J., Carvalhais, N., and Prabhat: Deep learning and process understanding for data-driven Earth system science, Nature, 566, 195–204, <a href="https://doi.org/10.1038/s41586-019-0912-1" target="_blank">https://doi.org/10.1038/s41586-019-0912-1</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Rinehart and Garvey(1978)</label><mixed-citation>
      
Rinehart, R. and Garvey, E.: Three-dimensional storm motion detection by
conventional weather radar, Nature, 273, 287–289,
<a href="https://doi.org/10.1038/273287a0" target="_blank">https://doi.org/10.1038/273287a0</a>, 1978.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Roberts(2008)</label><mixed-citation>
      
Roberts, N.: Assessing the spatial and temporal variation in the skill of
precipitation forecasts from an NWP model, Meteorol. Appl., 15, 163–169, <a href="https://doi.org/10.1002/met.57" target="_blank">https://doi.org/10.1002/met.57</a>, 2008.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>Roberts and Lean(2008)</label><mixed-citation>
      
Roberts, N. M. and Lean, H. W.: Scale-selective verification of rainfall
accumulations from high-resolution forecasts of convective events, Mon. Weather Rev., 136, 78–97, <a href="https://doi.org/10.1175/2007MWR2123.1" target="_blank">https://doi.org/10.1175/2007MWR2123.1</a>,
2008.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>Ronneberger et al.(2015)</label><mixed-citation>
      
Ronneberger, O., Fischer, P., and Brox, T.: U-net: Convolutional networks for
biomedical image segmentation, in: International Conference on Medical image
computing and computer-assisted intervention, 234–241, Springer,
<a href="https://doi.org/10.1007/978-3-319-24574-4_28" target="_blank">https://doi.org/10.1007/978-3-319-24574-4_28</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Schultz et al.(2021)</label><mixed-citation>
      
Schultz, M., Betancourt, C., Gong, B., Kleinert, F., Langguth, M., Leufen, L.,
Mozaffari, A., and Stadtler, S.: Can deep learning beat numerical weather
prediction?, Philos. T. Roy. Soc. A, 379,
20200097, <a href="https://doi.org/10.1098/rsta.2020.0097" target="_blank">https://doi.org/10.1098/rsta.2020.0097</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>Sha et al.(2020)</label><mixed-citation>
      
Sha, Y., Gagne II, D. J., West, G., and Stull, R.: Deep-learning-based gridded
downscaling of surface meteorological variables in complex terrain. Part II:
Daily precipitation, Appl. Meteorol. Climatol., 59,
2075–2092, <a href="https://doi.org/10.1175/JAMC-D-20-0058.1" target="_blank">https://doi.org/10.1175/JAMC-D-20-0058.1</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Shi et al.(2015)</label><mixed-citation>
      
Shi, X., Chen, Z., Wang, H., Yeung, D.-Y., Wong, W.-K., and Woo, W.-C.: Convolutional LSTM network: A machine learning approach for precipitation nowcasting, in: Advances in neural information processing systems, 802–810,
<a href="https://doi.org/10.48550/arXiv.1506.04214" target="_blank">https://doi.org/10.48550/arXiv.1506.04214</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Shi et al.(2017)</label><mixed-citation>
      
Shi, X., Gao, Z., Lausen, L., Wang, H., Yeung, D.-Y., Wong, W.-k., and Woo,
W.-C.: Deep learning for precipitation nowcasting: A benchmark and a new
model, arXiv preprint arXiv:1706.03458,
<a href="https://doi.org/10.48550/arXiv.1706.03458" target="_blank">https://doi.org/10.48550/arXiv.1706.03458</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Sønderby et al.(2020)</label><mixed-citation>
      
Sønderby, C. K., Espeholt, L., Heek, J., Dehghani, M., Oliver, A., Salimans,
T., Agrawal, S., Hickey, J., and Kalchbrenner, N.: Metnet: A neural weather
model for precipitation forecasting, arXiv preprint arXiv:2003.12140,
<a href="https://doi.org/10.48550/arXiv.2003.12140" target="_blank">https://doi.org/10.48550/arXiv.2003.12140</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>Sun et al.(2014)</label><mixed-citation>
      
Sun, J., Xue, M., Wilson, J. W., Zawadzki, I., Ballard, S. P., onvlee hooiMeyer, J., Joe, P. I., Barker, D. M., Li, P.-W., Golding, B., Xu, M., and Pinto, J. O.: Use of NWP for nowcasting convective precipitation: Recent progress and challenges, B. Am. Meteorol. Soc., 95, 409–426, <a href="https://doi.org/10.1175/BAMS-D-11-00263.1" target="_blank">https://doi.org/10.1175/BAMS-D-11-00263.1</a>, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>Vasiloff et al.(2007)</label><mixed-citation>
      
Vasiloff, S. V., Seo, D.-J., Howard, K. W., Zhang, J., Kitzmiller, D., Mullusky, M. G., Krajewski, W. F., Brandes, E., Rabin, R. M., Berkowitz, D. S., Brooks, H., McGinley, J. A., Kuligowski, R. J., and Brown, B: Improving QPE and very short term QPF: An initiative for a community-wide integrated approach, B. Am. Meteorol. Soc., 88, 1899–1911, <a href="https://doi.org/10.1175/BAMS-88-12-1899" target="_blank">https://doi.org/10.1175/BAMS-88-12-1899</a>, 2007.


    </mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Wang et al.(2017)</label><mixed-citation>
      
Wang, Y., Long, M., Wang, J., Gao, Z., and Philip, S. Y.: Predrnn: Recurrent neural networks for predictive learning using
spatiotemporal lstms, in: Advances in Neural Information Processing Systems, 879–888, <a href="https://proceedings.neurips.cc/paper/2017/hash/e5f6ad6ce374177eef023bf5d0c018b6-Abstract.html" target="_blank"/>
(last access: 12 December 2021), 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Wang et al.(2021)</label><mixed-citation>
      
Wang, Y., Wu, H., Zhang, J., Gao, Z., Wang, J., Yu, P. S., and Long, M.:
PredRNN: A Recurrent Neural Network for Spatiotemporal Predictive Learning,
arXiv preprint arXiv:2103.09504,
<a href="https://doi.org/10.1109/TPAMI.2022.3165153" target="_blank">https://doi.org/10.1109/TPAMI.2022.3165153</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Wilks(2011)</label><mixed-citation>
      
Wilks, D. S.: Statistical methods in the atmospheric sciences, vol. 100, Academic Press, ISBN 9780123850225, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Wilson et al.(2010)</label><mixed-citation>
      
Wilson, J. W., Feng, Y., Chen, M., and Roberts, R. D.: Nowcasting challenges
during the Beijing Olympics: Successes, failures, and implications for future
nowcasting systems, Weather Forecast., 25, 1691–1714,
<a href="https://doi.org/10.1175/2010WAF2222417.1" target="_blank">https://doi.org/10.1175/2010WAF2222417.1</a>, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Wolberg(1990)</label><mixed-citation>
      
Wolberg, G.: Digital image warping, vol. 10662, IEEE computer society press Los
Alamitos, CA, 1990.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>Woo and Wong(2017)</label><mixed-citation>
      
Woo, W.-C. and Wong, W.-K.: Operational application of optical flow techniques
to radar-based rainfall nowcasting, Atmosphere, 8, 48,
<a href="https://doi.org/10.3390/atmos8030048" target="_blank">https://doi.org/10.3390/atmos8030048</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Xie et al.(2019)</label><mixed-citation>
      
Xie, S., Wang, Y.-C., Lin, W., Ma, H.-Y., Tang, Q., Tang, S., Zheng, X., Golaz,
J.-C., Zhang, G. J., and Zhang, M.: Improved diurnal cycle of precipitation
in E3SM with a revised convective triggering function, J. Adv.
Model. Earth Sy., 11, 2290–2310,
<a href="https://doi.org/10.1029/2019MS001702" target="_blank">https://doi.org/10.1029/2019MS001702</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Zahraei et al.(2012)</label><mixed-citation>
      
Zahraei, A., Hsu, K.-l., Sorooshian, S., Gourley, J., Lakshmanan, V., Hong, Y.,
and Bellerby, T.: Quantitative precipitation nowcasting: A Lagrangian
pixel-based approach, Atmos. Res., 118, 418–434,
<a href="https://doi.org/10.1016/j.atmosres.2012.07.001" target="_blank">https://doi.org/10.1016/j.atmosres.2012.07.001</a>, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Zahraei et al.(2013)</label><mixed-citation>
      
Zahraei, A., Hsu, K.-l., Sorooshian, S., Gourley, J. J., Hong, Y., and
Behrangi, A.: Short-term quantitative precipitation forecasting using an
object-based approach, J. Hydrol., 483, 1–15,
<a href="https://doi.org/10.1016/j.jhydrol.2012.09.052" target="_blank">https://doi.org/10.1016/j.jhydrol.2012.09.052</a>, 2013.

    </mixed-citation></ref-html>--></article>
