<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" dtd-version="3.0"><?xmltex \makeatother\@nolinetrue\makeatletter?>
  <front>
    <journal-meta>
<journal-id journal-id-type="publisher">AMT</journal-id>
<journal-title-group>
<journal-title>Atmospheric Measurement Techniques</journal-title>
<abbrev-journal-title abbrev-type="publisher">AMT</abbrev-journal-title>
<abbrev-journal-title abbrev-type="nlm-ta">Atmos. Meas. Tech.</abbrev-journal-title>
</journal-title-group>
<issn pub-type="epub">1867-8548</issn>
<publisher><publisher-name>Copernicus Publications</publisher-name>
<publisher-loc>Göttingen, Germany</publisher-loc>
</publisher>
</journal-meta>

    <article-meta>
      <article-id pub-id-type="doi">10.5194/amt-10-1557-2017</article-id><title-group><article-title>Data-driven clustering of rain events: microphysics information derived from
macro-scale observations</article-title>
      </title-group><?xmltex \runningtitle{Data-driven clustering of rain events}?><?xmltex \runningauthor{M. Djallel Dilmi et al.}?>
      <contrib-group>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Dilmi</surname><given-names>Mohamed Djallel</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Mallet</surname><given-names>Cécile</given-names></name>
          
        <ext-link>https://orcid.org/0000-0002-0943-3510</ext-link></contrib>
        <contrib contrib-type="author" corresp="yes" rid="aff1">
          <name><surname>Barthes</surname><given-names>Laurent</given-names></name>
          <email>laurent.barthes@latmos.ipsl.fr</email>
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Chazottes</surname><given-names>Aymeric</given-names></name>
          
        </contrib>
        <aff id="aff1"><institution>LATMOS-CNRS/UVSQ/UPSay, 11 boulevard d'Alembert, 78280 Guyancourt,
France</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Laurent Barthes (laurent.barthes@latmos.ipsl.fr)</corresp></author-notes><pub-date><day>25</day><month>April</month><year>2017</year></pub-date>
      
      <volume>10</volume>
      <issue>4</issue>
      <fpage>1557</fpage><lpage>1574</lpage>
      <history>
        <date date-type="received"><day>25</day><month>November</month><year>2016</year></date>
           <date date-type="rev-request"><day>29</day><month>November</month><year>2016</year></date>
           <date date-type="rev-recd"><day>21</day><month>March</month><year>2017</year></date>
           <date date-type="accepted"><day>29</day><month>March</month><year>2017</year></date>
      </history>
      <permissions>
<license license-type="open-access">
<license-p>This work is licensed under a Creative Commons Attribution 3.0 Unported License. To view a copy of this license, visit <ext-link ext-link-type="uri" xlink:href="http://creativecommons.org/licenses/by/3.0/">http://creativecommons.org/licenses/by/3.0/</ext-link></license-p>
</license>
</permissions><self-uri xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017.html">This article is available from https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017.html</self-uri>
<self-uri xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017.pdf">The full text article is available as a PDF file from https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017.pdf</self-uri>


      <abstract>
    <p>Rain time series records are generally studied using rainfall rate
or accumulation parameters, which are estimated for a fixed duration
(typically 1 min, 1 h or 1 day). In this study we use the concept of
“rain events”. The aim of the first part of this paper is to establish a
parsimonious characterization of rain events, using a minimal set of
variables selected among those normally used for the characterization of
these events. A methodology is proposed, based on the combined use of a
genetic algorithm (GA) and self-organizing maps (SOMs). It can be
advantageous to use an SOM, since it allows a high-dimensional data space to
be mapped onto a two-dimensional space while preserving, in an unsupervised
manner, most of the information contained in the initial space topology. The
2-D maps obtained in this way allow the relationships between variables to be
determined and redundant variables to be removed, thus leading to a minimal
subset of variables. We verify that such 2-D maps make it possible to
determine the characteristics of all events, on the basis of only five
features (the event duration, the peak rain rate, the rain event depth, the
standard deviation of the rain rate event and the absolute rain rate
variation of the order of 0.5). From this minimal subset of variables,
hierarchical cluster analyses were carried out. We show that clustering into
two classes allows the conventional convective and stratiform classes to be
determined, whereas classification into five classes allows this convective–stratiform
classification to be further refined. Finally, our study made
it possible to reveal the presence of some specific relationships between
these five classes and the microphysics of their associated rain events.</p>
  </abstract>
    </article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <title>Introduction</title>
      <p>The analysis of “precipitation events” or “rain events” can be used to
obtain information concerning the characteristics of precipitation at a
particular location and for a specific application. This is a convenient
way to summarize precipitation time series in the form of a small number of
characteristics that make sense for particular applications.</p>
      <p>The concept of a precipitation event is not new and has been used for many
years (Eagleson, 1970; Brown et al., 1984). A wide variety of definitions,
varying according to the context of each study, have been reported in the
literature (Larsen and Teves, 2015). Moreover, when a rain rate time series
(generally based on rain gauge records) is broken down into individual
rainfall events, a wide variety of their characteristics, such as average
rainfall rate, rain event duration and rainfall event distribution (known as
hydrological information), can be computed for each event. Our analysis of
the literature has led to the identification of 17 features used to
characterize rainfall, which makes it quite difficult to compare different
studies. The first goal of the present study is to select a reduced set of
features characterizing rainfall events, through the use of a data-driven
approach, without taking a priori knowledge of the field of application into
account, thereby characterizing rainfall events in the most parsimonious and
efficient manner.</p>
      <p>The second goal is to assess, without using any a priori criteria, whether the rain
events are still correctly clustered by the most relevant observed features.
Indeed, atmospheric process specialists distinguish between stratiform and
convective events, arguing that the physical processes involved in their
evolution are different. The goal here is to check that a small sample of
variables, derived from spot measurements to describe rain events, can allow
this distinction to be made and ultimately be used to refine it.
Hydrological (hereafter referred to as “macrophysical”) information makes
use of rain gauge measurements to characterize rain events. This information
is defined in order to characterize the features of global events but not
to provide any information concerning the raindrop microphysics of the
event. Nevertheless, in many applications such as remote sensing, knowledge
of the microphysics is essential. One key parameter in remote sensing is the
raindrop size distribution, noted as <inline-formula><mml:math id="M1" display="inline"><mml:mrow><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, which is defined by the number of
raindrops per unit volume and per unit raindrop diameter (<inline-formula><mml:math id="M2" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula>). Information
related to the raindrop size distribution is often derived from its proxies,
as explained in Sect. 5. Such features are not currently accessible
through rain gauge measurements, which provide macrophysical information
only. However, more expensive devices referred to as disdrometers can
provide both hydrological and microphysical information. There are currently
several tens of thousands of rain gauges operated throughout the world, in
locations equipped with a far smaller number (if any) of disdrometers. As
described later in this paper, it is possible to retrieve some microphysical
information from the hydrological data. As a consequence, rain gauge data
could provide valuable information in microphysics studies through the use
of a statistical approach to indirectly infer the missing microphysics
information. In the following, the terms “macrophysical” or “hydrological”
information are associated with characteristics related to rain rates or
rain accumulation, whereas the term “microphysical” is associated with the
characteristics of the raindrop size distribution.</p>
      <p>In the present study, we use a data-driven approach to study the
relationships between different rain properties. As disdrometers provide
drop size distributions, they allow 1-minute (or shorter) rain rates to be
estimated, and in the present study these can be used to derive the
hydrological information of interest, which is coherent with the data that
would be provided by standard rain gauges. Through the combined
interpretation of microphysical and hydrological information, we are also
able to analyze the microphysical properties of the rain event clusters
provided by our algorithm. This makes it possible to retrieve (unobservable)
microphysical information from rain gauge measurements.</p>
      <p>From a single rain rate time series, observed with a 1-minute time
resolution, we seek to answer the following questions:
<list list-type="bullet"><list-item>
      <p>Among the large number of hydrological variables described in the
literature, which are the most significant?</p></list-item><list-item>
      <p>Does the resulting description of rain events allow different types of
rain event to be discriminated?</p></list-item><list-item>
      <p>What (unobserved) microphysical properties of an event, or type of rain
event, can be inferred from its macrophysical description?</p></list-item></list>
Our paper is structured as follows. Section 2 presents the data used in our
study and lists various hydrological parameters that are commonly found in
the literature. Seventeen macrophysical variables are identified, requiring
appropriate normalization. Section 3 presents our methodology, which is
based on the use of a genetic algorithm (GA) implementing a self-organizing
map (SOM, also referred to as a topological map). This unsupervised approach
is used to select a small subset of variables from the 17 identified
variables, allowing a parsimonious characterization of rainfall events to be
applied. An exploratory statistical analysis of rainfall events is provided.
In Sect. 4, the rainfall events are grouped in clusters and are divided
into two classes. It is then shown that this grouping of the data set
corresponds to the standard convective–stratiform classification. We then
propose a five-subclass classification, which corresponds to a refinement of
the two initial groups. In Sect. 5 we include some additional
microphysical features of rainfall events, allowing the microphysical
properties of the five previously defined event classes to be studied. Our
conclusions are presented in Sect. 6.</p>
</sec>
<sec id="Ch1.S2">
  <title>The disdrometer data sets – data processing methodology</title>
      <p>This research relies on the analysis of raindrop measurements obtained with
a dual-beam spectropluviometer (DBS) disdrometer, first described by
Delahaye et al. (2006). This instrument allows the arrival time, diameter
and fall velocity of incoming drops to be recorded. As the capture area of
the sensor is 100 cm<inline-formula><mml:math id="M3" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula> its observations can be considered, in spatial
terms, to be “point-like”. In the present study the integration time
<inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">int</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> was set to 1 min, and the raindrop measurements were used to
estimate the corresponding 1-minute rain rate time series RR<inline-formula><mml:math id="M5" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mi>t</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. In
order to eliminate false raindrop detections that could be generated by dust
or insects, a threshold <inline-formula><mml:math id="M6" display="inline"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>=</mml:mo></mml:mrow></mml:math></inline-formula> 0.1 mm h<inline-formula><mml:math id="M7" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> was applied. Rain rates
lower than <inline-formula><mml:math id="M8" display="inline"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> are thus set to zero. This conventional threshold is also
chosen to ensure coherency with previous studies (Verrier et al., 2013;
Llasat et al., 2001). In the present study, we worked with two data sets
recorded during the period between July 2008 and July 2014, at the Site
Instrumental de Recherche par Télédétection Atmosphérique
(SIRTA <fn id="Ch1.Footn1"><p><uri>http://sirta.ipsl.fr/</uri></p></fn>)
in Palaiseau, France.</p>
<sec id="Ch1.S2.SS1">
  <title>Rain event definition</title>
      <p>In everyday life, it is common knowledge that rain starts at a certain
moment and stops some time later. However, due to its discreet nature, rain
(which generally consists of a very large number of raindrops) is not an
easy concept to define. Indeed, the exact definition of a rain event will
depend on the sensor's characteristics (specific surface capture, detection
threshold, instrumental noise) as well as the spatial or temporal
resolution chosen for the study. This definition may also depend on the
purpose of the study and thus on the scientific community behind it. There
is thus a wide range of criteria used to break down precipitation records
into rain events. For this reason, it is important to define and apply an
unambiguous definition of a “rain event”.</p>
      <p>In this study, the pattern produced by the 1-minute rain rate time series
RR<inline-formula><mml:math id="M9" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mi>t</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> can be simplified by grouping non-null rain rates into a set of
separate “primitive events” (Brown et al., 1985). On the basis of an
assigned minimum inter-event time (MIT; Coutinho et al., 2014) each rain
rate value, corresponding to a specific 1-minute period of observation, is
assigned to a given rainfall event, i.e., either the rainfall event in
progress or a subsequent event that is considered to be independent and
“new”. The MIT could also be defined as the duration of a dry period
<inline-formula><mml:math id="M10" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">dry</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> following which the next occurrence of non-null rainfall marks the
beginning of a new event (Driscoll et al., 1989). For dry periods shorter than the MIT, rain rates
from either side of this period are considered to belong to the same
“composite event”. Various authors have proposed different values of MIT
that ensure event independence. Llasat (2001) noted that “The definition of an episode is quite subjective. In this case it was felt
possible to distinguish between two different episodes, when the time which elapses between them without rainfall exceeds 1 h, which ensures
that the two episodes come from different “clouds”. Moussa and Bocquillon (1991) wrote, “the constant rain observations on less than 30 min represent
only 5 % of all the rainy periods.
The representative threshold of the discretization of the data is 30 min to an hour.”</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T1"><caption><p>Observation periods, availability of DBS observations and numbers
of rain events for the learning and test data sets.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.82}[.82]?><oasis:tgroup cols="4">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:thead>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">Observation period</oasis:entry>  
         <oasis:entry colname="col3">Availability</oasis:entry>  
         <oasis:entry colname="col4">Number of</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3">(%)</oasis:entry>  
         <oasis:entry colname="col4">rain events</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>  
         <oasis:entry colname="col1">Learning data set</oasis:entry>  
         <oasis:entry colname="col2">01/01/2013–12/31/2014</oasis:entry>  
         <oasis:entry colname="col3">96.4 %</oasis:entry>  
         <oasis:entry colname="col4">234</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">Test data set</oasis:entry>  
         <oasis:entry colname="col2">04/16/2008–01/31/2012</oasis:entry>  
         <oasis:entry colname="col3">60 %</oasis:entry>  
         <oasis:entry colname="col4">311</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

      <p>Dunkerley (2008a, b) carried out an analysis of the inter-event time (IET)
in order to check the influence of this variable on the definition of
rainfall events and its influence on the average rainfall rate. As
emphasized in this study, when determining a value for the MIT, it is
crucial to find an appropriate compromise between the independence of rain
events and the intra-event variability of rain rates. The choice of MIT thus
has a direct impact on the macrophysical characteristics that are ultimately
determined by the analysis. Other researchers have proposed to use MIT
values of 20 min, 1 h or even 1 day (see Dunkerley, 2008a, for a
detailed list). In the present study it was decided to set the MIT to 30 min. This is in agreement with the value used by Coutinho et al. (2014),
Haile et al. (2011), Dunkerley (2008a, b), Balme et al. (2006) and Cosgrove
and Garstang (1995).</p>
      <p>When applied to our data set, this choice leads to the identification of 545 rain events, which can be divided up into two subsets, i.e., one for learning
and the other for testing (Table 1). The learning data set is composed of
observations collected over a 2-year period between 2013 and 2014, with an
availability of 96.4 %, whereas the test data set collected during the 2008–2012 period contains periods with missing data due to a malfunction of
the recording device.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T2" specific-use="star"><caption><p>The 23 variables identified in the literature, used for the
characterization of rain events.</p></caption><oasis:table frame="topbot"><?xmltex \begin{scaleboxenv}{.96}[.96]?><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="left"/>
     <oasis:colspec colnum="4" colname="col4" align="left"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1">Number  (no.)</oasis:entry>  
         <oasis:entry colname="col2">Feature name</oasis:entry>  
         <oasis:entry colname="col3">Symbol</oasis:entry>  
         <oasis:entry colname="col4">Formula</oasis:entry>  
         <oasis:entry colname="col5">Normalization</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>  
         <oasis:entry colname="col1">1</oasis:entry>  
         <oasis:entry colname="col2">Event duration</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M11" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M12" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> [min]</oasis:entry>  
         <oasis:entry colname="col5">1</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4">With <inline-formula><mml:math id="M13" display="inline"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>: Event start time and <inline-formula><mml:math id="M14" display="inline"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>: Event end time</oasis:entry>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">2</oasis:entry>  
         <oasis:entry colname="col2">Mean event</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M15" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M16" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> [mm h<inline-formula><mml:math id="M17" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">rain rate</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">3</oasis:entry>  
         <oasis:entry colname="col2">Intra-dry</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M18" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M19" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:msub><mml:mi>I</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> [min]</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">duration</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4">With <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mfenced open="{" close=""><mml:mtable class="array" columnalign="center center"><mml:mtr><mml:mtd><mml:mn mathvariant="normal">1</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mtext>if</mml:mtext><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mo>[</mml:mo><mml:mi mathvariant="normal">mm</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn mathvariant="normal">0</mml:mn></mml:mtd><mml:mtd><mml:mtext>else</mml:mtext></mml:mtd></mml:mtr></mml:mtable></mml:mfenced></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">4</oasis:entry>  
         <oasis:entry colname="col2">First quartile</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M21" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">The 25th percentile [mm h<inline-formula><mml:math id="M22" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">5</oasis:entry>  
         <oasis:entry colname="col2">Median</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">the 50th percentile [mm h<inline-formula><mml:math id="M24" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">6</oasis:entry>  
         <oasis:entry colname="col2">Third quartile</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M25" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">The 75th percentile [mm h<inline-formula><mml:math id="M26" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">7</oasis:entry>  
         <oasis:entry colname="col2">Previous IET</oasis:entry>  
         <oasis:entry colname="col3">IET<inline-formula><mml:math id="M27" display="inline"><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">IET<inline-formula><mml:math id="M28" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub><mml:mfenced close=")" open="("><mml:mtext>current  event</mml:mtext></mml:mfenced><mml:mo>-</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub><mml:mfenced open="(" close=")"><mml:mtext>previous  event</mml:mtext></mml:mfenced><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> [min]</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">8</oasis:entry>  
         <oasis:entry colname="col2">Mean rain rate over</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M29" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">r</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M30" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">r</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> [mm h<inline-formula><mml:math id="M31" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">3</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">the rainy period</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">9</oasis:entry>  
         <oasis:entry colname="col2">Event rain</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M33" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:msqrt><mml:mrow><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:msup><mml:mo>)</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msqrt></mml:mrow></mml:math></inline-formula> [mm h<inline-formula><mml:math id="M34" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>]</oasis:entry>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">rate SD</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">10</oasis:entry>  
         <oasis:entry colname="col2">Mode</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M35" display="inline"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M36" display="inline"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub><mml:mo>=</mml:mo></mml:mrow></mml:math></inline-formula> the most frequent RR<inline-formula><mml:math id="M37" display="inline"><mml:msub><mml:mi/><mml:mi>t</mml:mi></mml:msub></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">11</oasis:entry>  
         <oasis:entry colname="col2">Rain rate peak</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M39" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mi mathvariant="normal">max</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">12</oasis:entry>  
         <oasis:entry colname="col2">Dry percentage</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi mathvariant="italic">%</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M41" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi mathvariant="italic">%</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">5</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">in event</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">13</oasis:entry>  
         <oasis:entry colname="col2">Rain event depth</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M42" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M43" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>∗</mml:mo><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mo>/</mml:mo><mml:mn mathvariant="normal">60</mml:mn></mml:mrow></mml:math></inline-formula> [mm]</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">14</oasis:entry>  
         <oasis:entry colname="col2">Absolute rain</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M44" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M45" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:munderover><mml:msup><mml:mfenced close="|" open="|"><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mfenced><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">6</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">15</oasis:entry>  
         <oasis:entry colname="col2">rate variation</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M46" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">For <inline-formula><mml:math id="M47" display="inline"><mml:mrow><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msub><mml:mi>c</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.5</mml:mn><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mn mathvariant="normal">1</mml:mn><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">3</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">16</oasis:entry>  
         <oasis:entry colname="col2">of order <inline-formula><mml:math id="M48" display="inline"><mml:mi>c</mml:mi></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M49" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">17</oasis:entry>  
         <oasis:entry colname="col2">Normalized absolute</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M50" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M51" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">3</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">18</oasis:entry>  
         <oasis:entry colname="col2">rain rate variation</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M52" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">For <inline-formula><mml:math id="M53" display="inline"><mml:mrow><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">2</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">19</oasis:entry>  
         <oasis:entry colname="col2">of order <inline-formula><mml:math id="M54" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M55" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">20</oasis:entry>  
         <oasis:entry colname="col2">Absolute rain</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M56" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>S</mml:mi><mml:mo>,</mml:mo><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M57" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>S</mml:mi><mml:mo>,</mml:mo><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:munderover><mml:msup><mml:mfenced open="|" close="|"><mml:mo>max⁡</mml:mo><mml:mfenced open="[" close="]"><mml:mfenced close=")" open="("><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>S</mml:mi></mml:mfenced><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mfenced><mml:mo>-</mml:mo><mml:mo>max⁡</mml:mo><mml:mfenced close="]" open="["><mml:mfenced open="(" close=")"><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:mi>S</mml:mi></mml:mfenced><mml:mo>,</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mfenced></mml:mfenced><mml:mi>C</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">6</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">rate variation of</oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4">With <inline-formula><mml:math id="M58" display="inline"><mml:mrow><mml:mi>S</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.3</mml:mn></mml:mrow></mml:math></inline-formula>    and    <inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:mi>C</mml:mi><mml:mo>=</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">order <inline-formula><mml:math id="M60" display="inline"><mml:mi>C</mml:mi></mml:math></inline-formula> and threshold <inline-formula><mml:math id="M61" display="inline"><mml:mi>S</mml:mi></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">21</oasis:entry>  
         <oasis:entry colname="col2"><inline-formula><mml:math id="M62" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">L</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> parameter</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M63" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:mi>R</mml:mi><mml:msub><mml:mi>R</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="italic">θ</mml:mi><mml:mfenced open="(" close=")"><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mfenced></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">begin</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi mathvariant="normal">end</mml:mi></mml:msub></mml:mrow></mml:munderover><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">5</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M65" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">For L<inline-formula><mml:math id="M66" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.3</mml:mn></mml:mrow></mml:math></inline-formula>,  1,  3 mm h<inline-formula><mml:math id="M67" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">22</oasis:entry>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4">With <inline-formula><mml:math id="M68" display="inline"><mml:mrow><mml:mi mathvariant="italic">θ</mml:mi><mml:mfenced close=")" open="("><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mfenced></mml:mrow></mml:math></inline-formula> is the Heaviside function defined as</oasis:entry>  
         <oasis:entry colname="col5">0</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">23</oasis:entry>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M69" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M70" display="inline"><mml:mrow><mml:mi mathvariant="italic">θ</mml:mi><mml:mfenced close=")" open="("><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mfenced><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mtext>if</mml:mtext><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>≥</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M71" display="inline"><mml:mrow><mml:mi mathvariant="italic">θ</mml:mi><mml:mfenced close=")" open="("><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mfenced><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mtext>if</mml:mtext><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi mathvariant="normal">RR</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>&lt;</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"/>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup><?xmltex \end{scaleboxenv}?></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S2.SS2">
  <title>Macrophysical description of rain events</title>
      <p>Rain events contain a wealth of information, which generally needs to be
condensed into a limited set of well-chosen features. However, there is no
conventional or commonly accepted list or specific set of macrophysical
features that can be used to accurately describe and summarize an event. In
the present study it was thus decided to consider a large number of
features, allowing the macrophysical rain event information described in the
literature to be correctly represented. Seventeen characteristics were
selected and identified (Llasat, 2001; Moussa and Bocquillon, 1991) and are
listed in Table 2. Some of these are parameter dependent, such as <inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, which uses three values
of the parameter <inline-formula><mml:math id="M73" display="inline"><mml:mi>c</mml:mi></mml:math></inline-formula>. These three values lead to three <inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> indices, namely <inline-formula><mml:math id="M75" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>,
<inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>. Finally, a total of 23 descriptors were defined and numbered
from 1 to 23 (column 1 in Table 2).</p>
      <p>Among the 23 indicators (hereafter referred to as variables) corresponding
to the previously defined features, some are very well known. These include
the event duration (<inline-formula><mml:math id="M78" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, the quartile (<inline-formula><mml:math id="M79" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, the mean event rain rate
(<inline-formula><mml:math id="M80" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and the standard rain rate deviation (<inline-formula><mml:math id="M81" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, as well as other, less
traditional parameters such as the parameter <inline-formula><mml:math id="M82" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">L</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (indicator for
the convective nature of the rain; see Llasat, 2001), the absolute rain
rate variation of order <inline-formula><mml:math id="M83" display="inline"><mml:mi>c</mml:mi></mml:math></inline-formula> (<inline-formula><mml:math id="M84" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>c</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> or the absolute rain rate variation
(<inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>s</mml:mi><mml:mo>,</mml:mo><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. Some variables that are usually used to describe time series,
such as the fractal dimension, multi-fractal parameters, trend, seasonality
and autocorrelation, require a long series of data and are not well suited
to an event-by-event analysis. This set of 17 features is not exhaustive,
and some other features could also be included, depending on the
application. One example is the case of hydrology, for which the positions
of the intensity peaks inside the event could be a relevant feature.
Although, for events comprising a very small number of samples (very low
value of variable <inline-formula><mml:math id="M86" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, the computation of some indicators (<inline-formula><mml:math id="M87" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M88" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is questionable, in the present study the 23 variables were
computed for each of the 545 rain events.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F1"><caption><p>PCA on the learning data set based on the 23 variables described
in Table 3. <bold>(a)</bold> Correlation circle on axes 1 and 2. <bold>(b)</bold> Correlation
circle on axes 1 and 3. All of the variables are normalized according to
the last column of Table 2.</p></caption>
          <?xmltex \igopts{width=184.942913pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f01.png"/>

        </fig>

</sec>
<sec id="Ch1.S2.SS3">
  <title>Principal component analysis (PCA) analysis and normalization step</title>
      <p>It is important to note that very few of these 23 variables are compatible
with the probabilistic assumptions generally associated with exploratory
statistical methods. They are often highly variable, with highly skewed
distributions, and therefore do not have normal distributions. It is thus
more difficult to make direct use of standard statistical methods with these
data, as it may lead to misleading interpretations (Daumas, 1982). It is
thus necessary to introduce an additional step in order to transform the
original distributions into tquasi normally distributed distributions. The most
suitable type of normalizing transformation for each of these variables was
selected empirically by testing seven different possible transformations (Table 3). For each variable, the retained transformation is that leading to a
distribution with the strongest similarity to a normal distribution, i.e.,
with a kurtosis close to 3 and a skewness close to 0. For each indicator,
the selected transformation is provided in the last column of Table 2.</p>
      <p>Following the normalization step, PCA was
carried out on the learning data set (see end of Sect. 2.1 and Table 1). It
follows that the two principal axes contain 73 % of the total information,
whereas the first five principal axes are needed to represent 90 % of the
total information. The IET<inline-formula><mml:math id="M89" display="inline"><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:math></inline-formula> variable (no. 7) is very well correlated
with axis 5, whereas the other variables are not. This means that there is
no linear relationship between IET<inline-formula><mml:math id="M90" display="inline"><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:math></inline-formula> and the other variables. For this
reason, this variable was not considered as a possible candidate in the
variable selection process during the remainder of the study. The results
obtained in Sect. 4.2 confirm that there is no relationship between this
variable and the other 22 variables. The correlation circle on axes 1 and 2
(Fig. 1a) shows that among the 23 variables, 16 are well correlated with
the axis (close to a unit circle) and are distributed in approximately  five
groups (hereafter referred to as PCA groups). The first PCA group (G<inline-formula><mml:math id="M91" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>
can be identified by the variables, which are grouped close to the first
axis and are well correlated with it. As an example, this is the case for
the variables <inline-formula><mml:math id="M92" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 9), <inline-formula><mml:math id="M93" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mi>N</mml:mi></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (nos. 17–18) and <inline-formula><mml:math id="M94" display="inline"><mml:mi mathvariant="italic">β</mml:mi></mml:math></inline-formula>
(nos. 21 to 23). A second PCA group (G<inline-formula><mml:math id="M95" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> comprises the variables
<inline-formula><mml:math id="M96" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 11) and <inline-formula><mml:math id="M97" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (no. 16), just above axis 1. The third PCA
group (G<inline-formula><mml:math id="M98" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mn mathvariant="normal">3</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is formed by the variable <inline-formula><mml:math id="M99" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (no. 15) only. The fourth
PCA group (G<inline-formula><mml:math id="M100" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mn mathvariant="normal">4</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> comprises the variables <inline-formula><mml:math id="M101" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>(no. 14) and <inline-formula><mml:math id="M102" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>s</mml:mi><mml:mo>,</mml:mo><mml:mi>c</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (no. 20)
and is well correlated with axis 2. The last PCA group (G<inline-formula><mml:math id="M103" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mn mathvariant="normal">5</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is formed
by the variable <inline-formula><mml:math id="M104" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 1). The correlation circle on axis 1 and 3
(Fig. 1b) shows that the variables <inline-formula><mml:math id="M105" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> (no. 4), <inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> (no. 5) and
<inline-formula><mml:math id="M107" display="inline"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> (no. 10) are quite well represented by these two axes. A similar
remark can be made for variables <inline-formula><mml:math id="M108" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 3) and <inline-formula><mml:math id="M109" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>
(no. 21) on axes 1 and 4 (not shown).</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T3" specific-use="star"><caption><p>Transformations used to normalize the variables listed in Table 2.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="4">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="left"/>
     <oasis:colspec colnum="4" colname="col4" align="left"/>
     <oasis:thead>
       <oasis:row>  
         <oasis:entry colname="col1">Transformation</oasis:entry>  
         <oasis:entry colname="col2">Transformation name</oasis:entry>  
         <oasis:entry colname="col3">Formula f(x)</oasis:entry>  
         <oasis:entry colname="col4">Notes</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1">number</oasis:entry>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>  
         <oasis:entry colname="col1">0</oasis:entry>  
         <oasis:entry colname="col2">Standardization</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M110" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mi>x</mml:mi><mml:mo>-</mml:mo><mml:mtext>mean</mml:mtext><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mi mathvariant="normal">SD</mml:mi><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">–</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">1</oasis:entry>  
         <oasis:entry colname="col2">Power</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M111" display="inline"><mml:mrow><mml:msup><mml:mi>x</mml:mi><mml:mi>n</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M112" display="inline"><mml:mrow><mml:mi>n</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.05</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">2</oasis:entry>  
         <oasis:entry colname="col2">Boxcox</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M113" display="inline"><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mo>(</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mi mathvariant="italic">γ</mml:mi></mml:msup><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>)</mml:mo></mml:mrow><mml:mi mathvariant="italic">γ</mml:mi></mml:mfrac></mml:mstyle></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:mi mathvariant="italic">γ</mml:mi><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.1</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">3</oasis:entry>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M115" display="inline"><mml:mrow><mml:mi mathvariant="italic">γ</mml:mi><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.2</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">4</oasis:entry>  
         <oasis:entry colname="col2"/>  
         <oasis:entry colname="col3"/>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:mi mathvariant="italic">γ</mml:mi><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">0.3</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">5</oasis:entry>  
         <oasis:entry colname="col2">Arc-sin of square</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M117" display="inline"><mml:mrow><mml:mi>arcsin⁡</mml:mi><mml:msqrt><mml:mi>x</mml:mi></mml:msqrt></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4">Data are between 0 and 1</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1">6</oasis:entry>  
         <oasis:entry colname="col2">Decimal Logarithm</oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M118" display="inline"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>+</mml:mo><mml:mi>c</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M119" display="inline"><mml:mrow><mml:mi>c</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.1</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p>Finally, PCA analysis clearly shows that, within each PCA group, many
variables are highly intercorrelated, i.e., linearly dependant on each
other. This means that several variables could be removed with no
substantial loss of information. This leads to the following question: which
variables can be removed in order to retain the most parsimonious subset of
variables representative of the full data set? The PCA extracts summary
variables, which are a linear combination of original variables, but does
not allow for the selection of variables. To answer to this question, we propose
a method for the global selection of variables that seeks to identify the
relevant variables in a data set. As it appears to be intuitively more
advantageous to select variables with a physical sense, rather than using
dimension reduction methods (e.g., PCA, which is more suitable for the
detection of linear relationships), the proposed method is based on the use
of a GA. The following section provides a brief introduction
into the concept of GAs and shows how they can be
advantageously used for the selection of variables in the context of the
present study.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <title>Variable selection using a genetic algorithm</title>
      <p>Computer-assisted variable selection is important for several reasons.
Indeed, the selection of a subset of variables in a high-dimensional space
can improve the performance of the model or its statistical properties, but
it
also provides more robust models and reduces their complexity. In practice,
it is not generally possible to try all potential combinations of variables
and to select the best of these, as a consequence of the enormous
computational cost associated with such an approach. Among the many
different variable-selection techniques described in the literature (Guyon
and Elisseeff, 2003), we chose to develop a model based on the use of
GAs to search for an optimal subset of variables. GAs (Holland, 1975) are stochastic optimization algorithms
based on the mechanics of natural selection and the genetics described by
Charles Darwin. In our study, a chromosome is defined as a subset made up
from our 23 variables. A first generation composed of a population of 60
potential chromosomes is arbitrarily chosen. The performance of each
chromosome (i.e., for each corresponding subset of 60 variables) is evaluated
through a fitness function f. This fitness function is defined in such a way
that the higher its value, the greater the fitness function's ability to
represent the full data set (of dimension 23), using the smallest possible
number of variables. On the basis of the performance of these 60
chromosomes, we create a new generation of 60 potential-solution
chromosomes, using classical evolutionary operators: selection, crossover
and mutation. The performance of this new generation is then evaluated. This
cycle is repeated until a predefined stop criterion is satisfied. The best
chromosome from the current generation then provides the optimal subset of
variables.</p>
<sec id="Ch1.S3.SS1">
  <title>Methodology</title>
      <p>We define by <inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> the chromosome number <inline-formula><mml:math id="M121" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>: <inline-formula><mml:math id="M122" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>=</mml:mo><mml:mfenced open="(" close=")"><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">23</mml:mn></mml:msub></mml:mfenced></mml:mrow></mml:math></inline-formula>. <inline-formula><mml:math id="M123" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> is a binary vector in
<inline-formula><mml:math id="M124" display="inline"><mml:mrow><mml:mo mathvariant="italic">{</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:msup><mml:mo mathvariant="italic">}</mml:mo><mml:mn mathvariant="normal">23</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> space such that each component has the following meaning:
            <disp-formula id="Ch1.E1" content-type="numbered"><mml:math id="M125" display="block"><mml:mrow><mml:mtext>for all</mml:mtext><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi>i</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mtext>in</mml:mtext><mml:mfenced close="}" open="{"><mml:mn mathvariant="normal">1</mml:mn><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">23</mml:mn></mml:mfenced><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mfenced close="" open="{"><mml:mtable class="array" columnalign="left left"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:mtd><mml:mtd><mml:mtext>The    variable    number</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mi>i</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mtext> in  Table 2  column  1</mml:mtext></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mtext>is  selected,</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:mtd><mml:mtd><mml:mtext>The    variable  number</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mi>i</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mtext>in  Table 2  column  1</mml:mtext></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mtext>is 
not  selected.</mml:mtext></mml:mtd></mml:mtr></mml:mtable></mml:mfenced></mml:mrow></mml:math></disp-formula>
          The word “selected” in Eq. (1) means that the corresponding variable will be
used, both in the learning step described in “Step 2” below and for
performance evaluation. Otherwise, if the corresponding variable is not
selected it will be used only for performance evaluation.</p>
      <p>As previously stated, the fitness function allows a measure to be provided
of how well a minimal subset of variables can represent the entire data
space (in dimension 23). The fitness function <inline-formula><mml:math id="M126" display="inline"><mml:mi>f</mml:mi></mml:math></inline-formula> is thus defined as follows:
            <disp-formula id="Ch1.E2" content-type="numbered"><mml:math id="M127" display="block"><mml:mrow><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mrow><mml:mi>n</mml:mi><mml:mfenced close=")" open="("><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mfenced><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
          where <inline-formula><mml:math id="M128" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> is chromosome <inline-formula><mml:math id="M129" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>, <inline-formula><mml:math id="M130" display="inline"><mml:mrow><mml:mi>n</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>) is the number of selected variables in chromosome
<inline-formula><mml:math id="M131" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M132" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>) is the topological error associated
with chromosome <inline-formula><mml:math id="M133" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>.</p>
      <p>As the aim of this approach is to minimize the number of selected variables
<inline-formula><mml:math id="M134" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula> and the topological error <inline-formula><mml:math id="M135" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mi>e</mml:mi></mml:mrow></mml:math></inline-formula>, we seek to maximize the fitness function. The
estimation of the topological error made from an SOM  is
somewhat complicated and requires some explanation. The notion of an
SOM, introduced by Kohonen (1982, 2001),
makes use of a popular clustering and visualization algorithm. SOM is a
neural network algorithm based on unsupervised learning, derived from the
technique of competitive learning (Kohonen, 1982, 2001; Vesanto and
Alhoniemi, 2000). It may be considered as a nonlinear generalization, which
has many advantages over the conventional feature extraction techniques such
as empirical orthogonal functions (EOF) or PCA
(e.g., Liu et al., 2006). SOM applications are becoming increasingly useful
in geosciences (e.g., Liu and Weisberg, 2011). As stated by Uriarte and
Martín (2008): “The SOM provides a nonlinear, ordered, smooth mapping
of high-dimensional input data manifolds onto the elements of a regular,
low-dimensional array. The main characteristic of the projection provided by
this algorithm is the preservation of neighborhood relationships; as far as
possible, nearby data vectors in the input space are mapped onto neighboring
locations in the output space.” This property makes it straightforward to
compute a topological error (see Uriarte and Martín, 2008, Eq. 2). For
each of the <inline-formula><mml:math id="M136" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> chromosomes, an SOM <inline-formula><mml:math id="M137" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is learned on
the learning data set. Only the selected variables are used during the
learning process. Finally, for each Map <inline-formula><mml:math id="M138" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></inline-formula> the topological error
<inline-formula><mml:math id="M139" display="inline"><mml:mrow><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> can be computed in accordance with Eq. (2) in
Uriarte and Martín (2008). Section 4 provides additional information
concerning SOMs.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F2" specific-use="star"><caption><p>Diagram for the selection of variables based on a genetic
algorithm associated with Kohonen maps.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f02.png"/>

        </fig>

      <p>The genetic algorithm is based on the following five steps
(Fig. 2.):</p>
      <p><list list-type="order">
            <list-item>

      <p>In the first step, initialization, a (initial) population <inline-formula><mml:math id="M140" display="inline"><mml:mrow><mml:mo mathvariant="italic">{</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:mi>k</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>,</mml:mo><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mn mathvariant="normal">60</mml:mn><mml:mo mathvariant="italic">}</mml:mo></mml:mrow></mml:math></inline-formula> of 60 chromosomes of dimension 23 is randomly generated.</p>
            </list-item>
            <list-item>

      <p>In the second step, evaluation, for each of the <inline-formula><mml:math id="M141" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> chromosomes, an SOM <inline-formula><mml:math id="M142" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is learned. Only the
selected variables are used for learning. Once the learning has been
completed, the test data set and the 23 variables are used on each of the 60
maps to estimate their topological error te(<inline-formula><mml:math id="M143" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, allowing their fitness score
<inline-formula><mml:math id="M144" display="inline"><mml:mrow><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> to be computed.</p>
            </list-item>
            <list-item>

      <p>In the third step, the best chromosome <inline-formula><mml:math id="M145" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> is selected from the full set of 60 chromosomes
according to the fitness score previously computed with the test data set. If
<inline-formula><mml:math id="M146" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula> remains unchanged over a period of 50 generations, the
procedure is  stopped and the most relevant variables are  selected, i.e., those for which the
corresponding components are equal to 1 in <inline-formula><mml:math id="M147" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>. Otherwise, go to step
4.</p>
            </list-item>
            <list-item>

      <p>In the fourth step, selection, a new population of 60 chromosomes is created from the current population by
randomly sampling with replacement chromosomes based on their probabilities,
determined using the formula
                  <disp-formula id="Ch1.E3" content-type="numbered"><mml:math id="M148" display="block"><mml:mrow><mml:mi>Pr⁡</mml:mi><mml:mfenced open="(" close=")"><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup></mml:mfenced><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mn mathvariant="normal">60</mml:mn></mml:munderover><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>i</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula></p>
            </list-item>
            <list-item>

      <p>The fifth step, reproduction, uses mutation and crossover possibilities in the new population.
Mutation consists in modifying (or not) certain
components of the chromosomes. The probability of mutation is in general
very low and is commonly set to <inline-formula><mml:math id="M149" display="inline"><mml:mrow><mml:mi>p</mml:mi><mml:mo>=</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mn mathvariant="normal">10</mml:mn><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">7</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. In the present case, the
number of generations needed to reach the objective is less than a few
hundred, such that the probability of a mutation is very low.
For crossover, in an initial step <inline-formula><mml:math id="M150" display="inline"><mml:mrow><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">60</mml:mn><mml:mn mathvariant="normal">2</mml:mn></mml:mfrac></mml:mstyle><mml:mo>=</mml:mo><mml:mn mathvariant="normal">30</mml:mn></mml:mrow></mml:math></inline-formula> pairs of
chromosomes are randomly drawn from the population. Then, for each pair
(<inline-formula><mml:math id="M151" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>k</mml:mi></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi>l</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> (called parents) one crossover point, noted <inline-formula><mml:math id="M152" display="inline"><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, is
randomly drawn over the range [1, 23], using a discrete uniform law. Two new
chromosomes (<inline-formula><mml:math id="M153" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mrow><mml:msup><mml:mi>k</mml:mi><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mrow><mml:msup><mml:mi>l</mml:mi><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> are created as follows:
                  <disp-formula id="Ch1.E4" content-type="numbered"><mml:math id="M154" display="block"><mml:mrow><mml:mfenced open="{" close=""><mml:mtable class="array" columnalign="center"><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mrow><mml:msup><mml:mi>k</mml:mi><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi>k</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi>k</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:mrow><mml:mi>k</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>l</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow><mml:mi>l</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">23</mml:mn><mml:mi>l</mml:mi></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mrow><mml:msup><mml:mi>l</mml:mi><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mfenced open="(" close=")"><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mi>l</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn><mml:mi>l</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow><mml:mi>l</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>k</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mrow><mml:msub><mml:mi>I</mml:mi><mml:mi>c</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow><mml:mi>k</mml:mi></mml:msubsup><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">…</mml:mi><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msubsup><mml:mi>x</mml:mi><mml:mn mathvariant="normal">23</mml:mn><mml:mi>k</mml:mi></mml:msubsup></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mfenced></mml:mrow></mml:math></disp-formula>
                Thus, from two parents, two children are generated, allowing a new
generation to be produced with the same number of chromosomes. Finally, the
algorithm returns to step 2.</p>
            </list-item>
          </list></p>
</sec>
<sec id="Ch1.S3.SS2">
  <title>Parsimonious description of a rain event</title>
      <p>The GA is applied to our data sets in order to obtain an
optimal subset of variables forming a subspace, which can (in a certain
sense) provide relatively accurate information concerning the global space,
whilst having the particularity of containing non-redundant information. At
the 187th generation the algorithm produces a subspace comprising five
variables, namely, event duration <inline-formula><mml:math id="M155" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 1), standard deviation
<inline-formula><mml:math id="M156" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 9), maximum rain rate during event <inline-formula><mml:math id="M157" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 11),
rain event depth <inline-formula><mml:math id="M158" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 13) and absolute rain rate variation <inline-formula><mml:math id="M159" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>
(no. 14).</p>
      <p>The three variables (<inline-formula><mml:math id="M160" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M161" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M162" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> selected using this
data-driven approach are commonly used in the study of hydrological
processes (Haile et al., 2011). Moreover, it should be noted that the
commonly used variable <inline-formula><mml:math id="M163" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, which is computed simply by dividing the rain
event depth (<inline-formula><mml:math id="M164" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> by the duration (<inline-formula><mml:math id="M165" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, was not selected by the
algorithm. This result could be expected, since it is correlated with the
latter variables, and the algorithm provides a parsimonious description.
Concerning the absolute rain rate variation (<inline-formula><mml:math id="M166" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, this variable was
proposed by Moussa and Bocquillon (1991). It tends to provide information on
the structure of the events, more specifically related to smooth events with
a small number of sharp peaks. In fact, this variable promotes low
variations of RR<inline-formula><mml:math id="M167" display="inline"><mml:msub><mml:mi/><mml:mi>t</mml:mi></mml:msub></mml:math></inline-formula> because <inline-formula><mml:math id="M168" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is in a certain sense a structure
function of order <inline-formula><mml:math id="M169" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> of the variable RR<inline-formula><mml:math id="M170" display="inline"><mml:msub><mml:mi/><mml:mi>t</mml:mi></mml:msub></mml:math></inline-formula> (see no. 14, column 4 in Table 2), with a low value for
the exponent (<inline-formula><mml:math id="M171" display="inline"><mml:mrow><mml:msub><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>=</mml:mo></mml:mrow></mml:math></inline-formula> 0.5). Finally, the standard
deviation variable (<inline-formula><mml:math id="M172" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, which is a second-order moment, is the
most commonly used indicator to describe the variability of the
precipitation rate within the rain event.</p>
</sec>
</sec>
<sec id="Ch1.S4">
  <title>SOM learned with the five selected variables</title>
      <p>An SOM is a topological map composed of neurons. In the present case, a
neuron is a vector of dimension 23 containing the 23 variables defined in
Table 2. Each neuron has six neighboring neurons. SOM is an unsupervised neural
network trained by a competitive learning strategy that performs two tasks:
vector quantization and vector projection. The SOM, which is different to
<inline-formula><mml:math id="M173" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula> means, uses the neighborhood interaction set to learn the topological
structure hidden in the data. In addition, in order to achieve optimal
referent vector (neuron) matching, its neighbors on the map are updated,
leading to the generation of regions in which neurons located in the same
neighborhood are very similar. The SOM can thus be considered as an
algorithm that maps a high-dimensional data space onto a two-dimensional
space called a map. A map can be used both to reduce the amount data by
means of clustering and to project the data in a nonlinear manner onto a
regular grid (the map grid).</p>
      <p>In the present study we used the toolbox developed by the SOM Toolbox
Team, which is available at the following site: <uri>http://www.cis.hut.fi/somtoolbox/</uri>. An SOM with 8 <inline-formula><mml:math id="M174" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 8 <inline-formula><mml:math id="M175" display="inline"><mml:mo>=</mml:mo></mml:math></inline-formula> 64 neurons
is considered here. This choice corresponds to a compromise, since a smaller
map would not be able to distinguish fine details whereas, in view of the
number of observations, and a larger map would not be meaningful.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F3"><caption><p>Distance matrix for the <inline-formula><mml:math id="M176" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>
map: the color of each neuron represents the average distance between
itself and its neighboring neurons. The value inside each neuron indicates
the number of rain events that it has captured inside the learning data set.
The black line separates the neurons into two classes, using the hierarchical
ascendant classification (see Sect. 4.1). The arrows represent the
gradients of the variables <inline-formula><mml:math id="M177" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M178" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M179" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.</p></caption>
        <?xmltex \igopts{width=236.157874pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f03.png"/>

      </fig>

      <?xmltex \floatpos{t}?><fig id="Ch1.F4" specific-use="star"><caption><p>Projection of the <inline-formula><mml:math id="M180" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> map
according to the 23 variables. The red-framed variables are those selected
by the GA algorithm. The last two variables <inline-formula><mml:math id="M181" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M182" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> are defined in Sect. 5.</p></caption>
        <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f04.png"/>

      </fig>

      <p>After learning by the GA algorithm described in the previous section, the
resulting map <inline-formula><mml:math id="M183" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> can be used to assign to any
event the best matching reference vector (neuron), in accordance with the
five
selected variables associated with the chromosome
<inline-formula><mml:math id="M184" display="inline"><mml:mrow><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup></mml:mrow></mml:math></inline-formula>. The <inline-formula><mml:math id="M185" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> map
obtained with this procedure can be considered as an optimal representation
of the initial data set.</p>
      <p>Figure 3 shows the distance matrix. For each neuron, the color indicates the
mean distance between a neuron and its neighbors. The value at the center of
each neuron represents the number of rain events of the learning data set,
captured by the corresponding neuron. All neurons capture rain events and
slightly more than half of these capture between three and five rain events, which
is close to the value that would be obtained (<inline-formula><mml:math id="M186" display="inline"><mml:mrow><mml:mn mathvariant="normal">234</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">64</mml:mn><mml:mo>≅</mml:mo><mml:mn mathvariant="normal">4</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> if the
rain events were uniformly distributed over the map.</p>
<sec id="Ch1.S4.SS1">
  <title>Projection of the selected and unlearned variables onto the SOM</title>
      <p>The five variables <inline-formula><mml:math id="M187" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M188" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M189" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M190" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M191" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>
used for learning are referred to as “selected” variables, whereas the
remaining 18 variables are referred to as “unlearned” variables. In order to
study the relationship between these variables, Fig. 4 shows the projections
for each of the variables in the <inline-formula><mml:math id="M192" display="inline"><mml:mrow><mml:mi>M</mml:mi><mml:mo>(</mml:mo><mml:msup><mml:mi mathvariant="bold-italic">x</mml:mi><mml:mi mathvariant="normal">Best</mml:mi></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> map
obtained with the aforementioned GA selection algorithm. The variables are
discussed individually, by considering their structure, as well as the
relationships between them. We note that the map is well structured for the
majority of variables. This advantageous structuration of most of the
variables confirms the ability of the selected variables to summarize all of
the significant characteristics of rain events. Only a small number of
characteristics are not adequately represented. It should be noted that
almost all variables are structured according to the first or second
diagonal. Among these, one may consider an initial subset comprising
variables that are more or less structured according to the first diagonal.
This is the case for the unlearned variable <inline-formula><mml:math id="M193" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, as well as for the
selected variables <inline-formula><mml:math id="M194" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M195" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. A second subset comprising
variables that are structured in approximate accordance with the second
diagonal can be identified. This is the case for the unlearned variables
<inline-formula><mml:math id="M196" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">r</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M197" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M198" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>,
which are very similar to the selected
variable <inline-formula><mml:math id="M199" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. The unlearned variables <inline-formula><mml:math id="M200" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M201" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>S</mml:mi><mml:mo>,</mml:mo><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> also belong to the second subset and
have a structure close to that of the selected variable <inline-formula><mml:math id="M202" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>.</p>
      <p>The map can be related to the previously implemented PCA (Fig. 1). As can be
seen in Fig. 4, the variables <inline-formula><mml:math id="M203" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (no. 16) and <inline-formula><mml:math id="M204" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (no. 11), which
have a similar structure, also belong to the same PCA group, namely group
G<inline-formula><mml:math id="M205" display="inline"><mml:msub><mml:mi/><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:math></inline-formula> (see Sect. 2.3, Fig. 1a). It is interesting to note that the
variables <inline-formula><mml:math id="M206" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>S</mml:mi><mml:mo>,</mml:mo><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> (no. 20) and <inline-formula><mml:math id="M207" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>(no. 11), which also have a similar
structure, do not belong to the same PCA group (groups <inline-formula><mml:math id="M208" display="inline"><mml:mrow><mml:msub><mml:mi>G</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M209" display="inline"><mml:mrow><mml:msub><mml:mi>G</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>
respectively) and are uncorrelated (they are orthogonal in Fig. 1a). This
remark means that the topological map reveals a relationship that cannot be
detected using PCA. As the rain event depth (<inline-formula><mml:math id="M210" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> depends on both the
duration and the intensity of the events, the corresponding map has a
top-down structure. Two distinct situations thus occur:
<list list-type="bullet"><list-item>
      <p>Those events which contribute the greatest quantities of water (Fig. 4.,
brown neuron at the bottom right of <inline-formula><mml:math id="M211" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> are among the longest (see
corresponding neuron of <inline-formula><mml:math id="M212" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, but they do not have an extremely high peak rain
rate (see corresponding neuron of <inline-formula><mml:math id="M213" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and are quite smooth (see
corresponding neuron of <inline-formula><mml:math id="M214" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M215" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>).</p></list-item><list-item>
      <p>Other events which contribute large amounts of water (but less than
previously; Fig. 4. red neuron at the bottom left of <inline-formula><mml:math id="M216" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> have short
durations (see corresponding neuron of <inline-formula><mml:math id="M217" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, but they are violent (see
corresponding neuron of <inline-formula><mml:math id="M218" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and less smooth (see corresponding
neuron of <inline-formula><mml:math id="M219" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M220" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>). The latter case reflects situations
that are typical of convective storms.</p></list-item></list>
The resulting map confirms the dependence structure of the two hydrological
variables, <inline-formula><mml:math id="M221" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M222" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, studied by Gargouri and Chebchoub (2010).</p>
      <p>Concerning the variable <inline-formula><mml:math id="M223" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="normal">IET</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (previous IET), the map is not
structured, reflecting the independence of the characteristics of a rain
event with respect to the drought period preceding the event. This
corroborates the results of several previous studies (Lavergnat and Gole,
1998, 2006; Akrour et al., 2015; de Montera et al., 2009) dealing with rain support simulations. When
studying temperate midlatitudes for relatively short periods, these authors
noticed that successive rain and no-rain periods are uncorrelated, such that
a rain time series could be considered as an independently drawn,
alternating series of rain events and periods without rain. This is
equivalent to an IET that does not characterize the rain
events. The same effects are not necessarily observed at other locations
and under different climatological conditions. Brown et al. (1983) also
investigated a possible correlation between IETs and the intra-event
characteristics and concluded that their data provided no evidence of this.</p>
      <p>The variable <inline-formula><mml:math id="M224" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mi mathvariant="normal">L</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (Llasat, 2001) is
considered to represent a measure of the convective nature of the rain, it
makes sense that the three variables <inline-formula><mml:math id="M225" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M226" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M227" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> are structured similarly, with the peak rain rate variable
<inline-formula><mml:math id="M228" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. This relationship is clearly visible on the maps.</p>
      <p>Several other relationships, which are not described in detail here, can be
observed. These include the correlation between the normalized absolute rain
rate variation (<inline-formula><mml:math id="M229" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and the standard deviation of the intensity
(<inline-formula><mml:math id="M230" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. We conclude that the combination of the
five selected variables provides a relatively accurate summary of the
information needed to describe the rain events. The poor structuring of some
variables is justified by the independence of these variables with respect
to the properties of the rain events; this is the case for the variable dry
percentage in event <inline-formula><mml:math id="M231" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> or the variable IET<inline-formula><mml:math id="M232" display="inline"><mml:msub><mml:mi/><mml:mi>P</mml:mi></mml:msub></mml:math></inline-formula>.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T4" specific-use="star"><caption><p>Coefficient of determination obtained on the learning and test
data sets. The values in bold correspond to the five
selected variables.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="13">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:colspec colnum="7" colname="col7" align="right"/>
     <oasis:colspec colnum="8" colname="col8" align="right"/>
     <oasis:colspec colnum="9" colname="col9" align="right"/>
     <oasis:colspec colnum="10" colname="col10" align="right"/>
     <oasis:colspec colnum="11" colname="col11" align="right"/>
     <oasis:colspec colnum="12" colname="col12" align="right"/>
     <oasis:colspec colnum="13" colname="col13" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1">Variables</oasis:entry>  
         <oasis:entry colname="col2"><inline-formula><mml:math id="M233" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M234" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M235" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"><inline-formula><mml:math id="M236" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col6"><inline-formula><mml:math id="M237" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col7"><inline-formula><mml:math id="M238" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col8">IET<inline-formula><mml:math id="M239" display="inline"><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col9"><inline-formula><mml:math id="M240" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mo>,</mml:mo><mml:mi mathvariant="normal">r</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col10"><inline-formula><mml:math id="M241" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col11"><inline-formula><mml:math id="M242" display="inline"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">0</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col12"><inline-formula><mml:math id="M243" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col13"><inline-formula><mml:math id="M244" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi mathvariant="italic">%</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M245" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> learning data set</oasis:entry>  
         <oasis:entry colname="col2"><bold>0.96</bold></oasis:entry>  
         <oasis:entry colname="col3">0.91</oasis:entry>  
         <oasis:entry colname="col4">0.58</oasis:entry>  
         <oasis:entry colname="col5">0.57</oasis:entry>  
         <oasis:entry colname="col6">0.48</oasis:entry>  
         <oasis:entry colname="col7">0.77</oasis:entry>  
         <oasis:entry colname="col8">0.31</oasis:entry>  
         <oasis:entry colname="col9">0.93</oasis:entry>  
         <oasis:entry colname="col10"><bold>0.97</bold></oasis:entry>  
         <oasis:entry colname="col11">0.50</oasis:entry>  
         <oasis:entry colname="col12"><bold>0.96</bold></oasis:entry>  
         <oasis:entry colname="col13">0.52</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M246" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> test data set</oasis:entry>  
         <oasis:entry colname="col2"><bold>0.93</bold></oasis:entry>  
         <oasis:entry colname="col3">0.84</oasis:entry>  
         <oasis:entry colname="col4">0.55</oasis:entry>  
         <oasis:entry colname="col5">0.57</oasis:entry>  
         <oasis:entry colname="col6">0.50</oasis:entry>  
         <oasis:entry colname="col7">0.74</oasis:entry>  
         <oasis:entry colname="col8">0.26</oasis:entry>  
         <oasis:entry colname="col9">0.84</oasis:entry>  
         <oasis:entry colname="col10"><bold>0.86</bold></oasis:entry>  
         <oasis:entry colname="col11">0.54</oasis:entry>  
         <oasis:entry colname="col12"><bold>0.82</bold></oasis:entry>  
         <oasis:entry colname="col13">0.50</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1">Variables</oasis:entry>  
         <oasis:entry colname="col2"><inline-formula><mml:math id="M247" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col3"><inline-formula><mml:math id="M248" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col4"><inline-formula><mml:math id="M249" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col5"><inline-formula><mml:math id="M250" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col6"><inline-formula><mml:math id="M251" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mi>N</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col7"><inline-formula><mml:math id="M252" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mi>N</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col8"><inline-formula><mml:math id="M253" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mi>N</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col9"><inline-formula><mml:math id="M254" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>S</mml:mi><mml:mo>,</mml:mo><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col10"><inline-formula><mml:math id="M255" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col11"><inline-formula><mml:math id="M256" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col12"><inline-formula><mml:math id="M257" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col13">–</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M258" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> learning data set</oasis:entry>  
         <oasis:entry colname="col2"><bold>0.97</bold></oasis:entry>  
         <oasis:entry colname="col3"><bold>0.97</bold></oasis:entry>  
         <oasis:entry colname="col4">0.93</oasis:entry>  
         <oasis:entry colname="col5">0.94</oasis:entry>  
         <oasis:entry colname="col6">0.91</oasis:entry>  
         <oasis:entry colname="col7">0.95</oasis:entry>  
         <oasis:entry colname="col8">0.78</oasis:entry>  
         <oasis:entry colname="col9">0.94</oasis:entry>  
         <oasis:entry colname="col10">0.70</oasis:entry>  
         <oasis:entry colname="col11">0.89</oasis:entry>  
         <oasis:entry colname="col12">0.96</oasis:entry>  
         <oasis:entry colname="col13"/>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M259" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> test data set</oasis:entry>  
         <oasis:entry colname="col2"><bold>0.94</bold></oasis:entry>  
         <oasis:entry colname="col3"><bold>0.99</bold></oasis:entry>  
         <oasis:entry colname="col4">0.91</oasis:entry>  
         <oasis:entry colname="col5">0.83</oasis:entry>  
         <oasis:entry colname="col6">0.83</oasis:entry>  
         <oasis:entry colname="col7">0.85</oasis:entry>  
         <oasis:entry colname="col8">0.71</oasis:entry>  
         <oasis:entry colname="col9">0.82</oasis:entry>  
         <oasis:entry colname="col10">0.61</oasis:entry>  
         <oasis:entry colname="col11">0.76</oasis:entry>  
         <oasis:entry colname="col12">0.89</oasis:entry>  
         <oasis:entry colname="col13">–</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <?xmltex \floatpos{t}?><fig id="Ch1.F5"><caption><p>The variable <inline-formula><mml:math id="M260" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> versus its
corresponding value, given by the best matching unit: from the learning
data set (circles) and the test data set (stars). The solid line corresponds
to the first diagonal.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f05.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS2">
  <title>Representation of rain events on SOM</title>
      <p>In an effort to provide additional information for validation of the map, we
compared each of the 23 variables with their corresponding value given by
the SOM, for the learning data set and the test data set. For each of the 311
events of the test data set, the best matching unit of the SOM, i.e., the
neuron that is the closest to the event, is determined with respect to the
five selected variables. As an example, for each event Fig. 5 shows the
current value of the unlearned variable <inline-formula><mml:math id="M261" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> as a function of the
corresponding value given by the best matching unit of the event. A spread
can be seen, in particular in the central zone, whereas the spread is
relatively small for values located near to the edges (which are more
numerous). A linear regression leads to a relatively good determination
coefficient (<inline-formula><mml:math id="M262" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula>; 0.96 and 0.89, respectively, for the learning and
test data sets). Table 4 lists the value of <inline-formula><mml:math id="M263" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> for the 23 variables
obtained with the learning and test data sets. As expected, the coefficient
of determination of the variable IET<inline-formula><mml:math id="M264" display="inline"><mml:msub><mml:mi/><mml:mi>p</mml:mi></mml:msub></mml:math></inline-formula> is very poor (<inline-formula><mml:math id="M265" display="inline"><mml:mrow><mml:mn mathvariant="normal">0.31</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">0.26</mml:mn></mml:mrow></mml:math></inline-formula>), since this
variable is not related to the five selected variables and as a consequence
cannot be well represented by the SOM (Fig. 4). The selected variables have
good determination coefficients, with both the learning and the test
data sets; this confirms the quality of the learning and the generalization
ability of the SOM. The quality of the learning step is confirmed by the
fact that the <inline-formula><mml:math id="M266" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> values of the selected variables obtained on the test
set are close to those obtained on the learning set. The <inline-formula><mml:math id="M267" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula>
corresponding to the unlearned variables obtained on the learning data set
emphasize the ability of the selected variables to provide the information
contained in the unlearned variables; in the case of the test data set it
denotes the ability of the SOM to derive all event characteristics from the
selected variables only.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F6" specific-use="star"><caption><p>Dendrogram obtained from the hierarchical cluster analysis of the
64 neurons in the SOM. The horizontal dashed line represents the threshold
between the two classes.</p></caption>
          <?xmltex \igopts{width=398.338583pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f06.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS3">
  <title>Hierarchical clustering of rain events</title>
      <p>We have shown that the distance matrix (Fig. 3) confirms the successful
deployment of the map. Based on the distance between neurons, it appears
that neurons can be grouped to obtain a limited number of classes, each with
its own characteristics. In order to group the 64 neurons into a small
number of classes, a hierarchical cluster analysis was carried out (Everitt,
1974). Only the five selected variables were used for the classification,
and a Euclidian distance was selected for the hierarchical algorithm. Figure 6
shows the resulting dendrogram, applied to the 64 neurons.</p>
      <p>Depending on the physical processes involved, experts tend to separate rain
events into two different classes: stratiform and convective events.
Although this classification is relatively crude, since stratiform and
convective events can sometimes exist inside the same rain event, it is very
commonly used. Concerning the time series, most authors use a very simple
scheme to distinguish between stratiform and convective rain types. For
reasons of simplicity, rain classification is sometimes defined using the
instantaneous rain rate and the standard deviation estimated over
consecutive samples. As an example, Bringi et al. (2003) defined stratiform
rain samples when the standard deviation of the rain rate, taken over five
consecutive 2 min samples, is less than 1.5 mm h<inline-formula><mml:math id="M268" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, the convective rain
samples are defined for a rain rate greater than or equal to 5 mm h<inline-formula><mml:math id="M269" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>,
and the standard deviation of the rain rate over five consecutive 2 min
samples is greater than 1.5 mm h<inline-formula><mml:math id="M270" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>.</p>
      <p>Firstly, we separate the dendrogram into two classes. The first class
contains 51 neurons and 79 % of the observations, whereas the second class
contains 13 neurons and 21 % of the observations. The solid black line in
Fig. 3 corresponds to the dividing line between these two classes. The first
class, containing the greatest number of neurons, is in most cases
characterized by relatively low rain rates. This can be seen by examining
the structure of the map, according to the mean rain rate variable
(<inline-formula><mml:math id="M271" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. Moreover, through analysis of the standard deviation (small values of
<inline-formula><mml:math id="M272" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, absolute rain rates <inline-formula><mml:math id="M273" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (high values of <inline-formula><mml:math id="M274" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and
low values of <inline-formula><mml:math id="M275" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> show that this class is more or less
characterized by quiet, homogeneous events. Our analysis of event durations
(<inline-formula><mml:math id="M276" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> shows that this class contains both short and long durations but
is dominated by the latter. These characterizations are relatively well
matched to a description involving stratiform and stable precipitations,
which are often the consequence of the slow, large-scale uprising of a large
mass of moist air which then condenses uniformly.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F7"><caption><p>Representation of the neurons in the <inline-formula><mml:math id="M277" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M278" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, and <inline-formula><mml:math id="M279" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M280" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> subspaces. The stars represent
neurons from group 1 (stratiform), and the squares correspond to neurons
from group 2 (convective). Dashed lines indicate the neuron no. 64.</p></caption>
          <?xmltex \igopts{width=241.848425pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f07.png"/>

        </fig>

      <p>The second group is characterized by a smaller number of neurons. This
corresponds to the higher values of the mean rain rates (<inline-formula><mml:math id="M281" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and peak
rain rates (<inline-formula><mml:math id="M282" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. The variables <inline-formula><mml:math id="M283" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M284" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>c</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> have the
opposite values with respect to those of the previous group. Most of the
event durations (<inline-formula><mml:math id="M285" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> in this group are short, with the exception of
neuron no. 64 (bottom right on the maps). This group fits well with the
definition of convective events resulting from the rapid rise of air masses
loaded with moisture for buoyancy. This convective moist air can lead to
the development of cumulus clouds up to an altitude in excess of 10 km and
to heavy rain.</p>
      <p>Our analysis of the structure of the variables <inline-formula><mml:math id="M286" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M287" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M288" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> in Fig. 4 confirms the previous interpretation of the two
groups. These three variables, which are representative of convective rain,
have high values for the neurons belonging to this group.</p>
      <p>Figure 7a and b show the neurons in the <inline-formula><mml:math id="M289" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M290" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M291" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> subspace. These three variables were not used in the learning step.
Nevertheless, the two classes are well separated, although an overlap does
occur in Fig. 7a due to neuron no. 64 (bottom right on the map, Fig. 4). Although it belongs to the convective class, this neuron
nevertheless has some characteristics of the stratiform class.</p>
      <p>The hypothesis that the two categories of precipitation events corresponding
to different dynamic regimes can be identified solely on the basis of
hydrometeorological variables is in agreement with the findings of Molini et al. (2011). These authors have shown that there is a strong agreement
between the hydro-meteorological classification (based on the duration and
extent of events from rain gauge network data) and dynamic classifications
(the convective adjustment timescale identified to distinguish between
equilibrium and non-equilibrium convection derived from ECMWF analysis). We
conclude that this unsupervised automatic clustering, based on the five
selected variables, makes it possible to correctly implement a
classification with these two well-known classes (stratiform and
convective). It should be noted that, unlike other classifications described
in the literature, this was established without making use of a priori information,
since it is produced by an unsupervised process.</p>
</sec>
<sec id="Ch1.S4.SS4">
  <title>Classification of events into several classes</title>
      <p>From the stratiform and convective classification described above, it is
interesting to refine the two classes into a set of subclasses. The synoptic
rainfall associated with midlatitude depressions provides an example of
stratiform precipitation, which forms in depressions in the vicinity of warm
and cold fronts. The very light type of rainfall (drizzle) associated with
stratus or stratocumulus is included in the class of stratiform
precipitation. This can occur under anticyclonic conditions,or in the warm
region of a depression. The associated rain depths (<inline-formula><mml:math id="M292" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> are minimal and
usually have no hydrological impact other than superficial wetting. In order
to identify relevant subclasses, our classification was broken down into a
number of unknown subclasses, such that <inline-formula><mml:math id="M293" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula> &gt; 2.</p>
      <p>An important step in hierarchical clustering is the selection of an optimal
number of partitions <inline-formula><mml:math id="M294" display="inline"><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>n</mml:mi><mml:mi mathvariant="normal">opt</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> in the data set (Grazioli et al., 2015).
Many indices can be used to evaluate each partition, from the point of view
of data similarity only. Most of these evaluate the scattering inside each
cluster, with respect to the distance between clusters, and assign
relatively favorable scores to partitions with compact and well-separated
clusters. Although different indices were tested, these did not provide the
same number of subclasses (between 2 and 32 with the indices tested in this
study). It should be noted that these did not take the physical meaning of
each class into account. Finally, we chose <inline-formula><mml:math id="M295" display="inline"><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi mathvariant="normal">opt</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">5</mml:mn></mml:mrow></mml:math></inline-formula>, since higher values
led to classes with the same physical sense. The new classification based on
the use of five subclasses is shown in Fig. 8.</p>

<?xmltex \floatpos{t}?><table-wrap id="Ch1.T5" specific-use="star"><caption><p>Summary of the rain event subclasses computed with the learning
data set.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="6">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right" colsep="1"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:thead>
       <oasis:row>  
         <oasis:entry colname="col1"/>  
         <oasis:entry rowsep="1" namest="col2" nameend="col3" align="center" colsep="1">Stratiform events </oasis:entry>  
         <oasis:entry rowsep="1" namest="col4" nameend="col6" align="center">Convective events </oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">  
         <oasis:entry colname="col1"/>  
         <oasis:entry colname="col2">Subclass 1</oasis:entry>  
         <oasis:entry colname="col3">Subclass 2</oasis:entry>  
         <oasis:entry colname="col4">Subclass 3</oasis:entry>  
         <oasis:entry colname="col5">Subclass 4</oasis:entry>  
         <oasis:entry colname="col6">Subclass 5</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>  
         <oasis:entry colname="col1">Variables</oasis:entry>  
         <oasis:entry colname="col2">Mean</oasis:entry>  
         <oasis:entry colname="col3">Mean</oasis:entry>  
         <oasis:entry colname="col4">Mean</oasis:entry>  
         <oasis:entry colname="col5">Mean</oasis:entry>  
         <oasis:entry colname="col6">Mean</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M296" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (min)</oasis:entry>  
         <oasis:entry colname="col2">321</oasis:entry>  
         <oasis:entry colname="col3">149</oasis:entry>  
         <oasis:entry colname="col4">464</oasis:entry>  
         <oasis:entry colname="col5">75</oasis:entry>  
         <oasis:entry colname="col6">49</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M297" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col2">0.36</oasis:entry>  
         <oasis:entry colname="col3">2.01</oasis:entry>  
         <oasis:entry colname="col4">3.62</oasis:entry>  
         <oasis:entry colname="col5">11.7</oasis:entry>  
         <oasis:entry colname="col6">9.64</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M298" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (mm h<inline-formula><mml:math id="M299" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>  
         <oasis:entry colname="col2">2.08</oasis:entry>  
         <oasis:entry colname="col3">10</oasis:entry>  
         <oasis:entry colname="col4">22</oasis:entry>  
         <oasis:entry colname="col5">52.7</oasis:entry>  
         <oasis:entry colname="col6">36.06</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M300" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (mm)</oasis:entry>  
         <oasis:entry colname="col2">1.99</oasis:entry>  
         <oasis:entry colname="col3">2.62</oasis:entry>  
         <oasis:entry colname="col4">11.24</oasis:entry>  
         <oasis:entry colname="col5">6.9</oasis:entry>  
         <oasis:entry colname="col6">2.72</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M301" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col2">75.7</oasis:entry>  
         <oasis:entry colname="col3">64.5</oasis:entry>  
         <oasis:entry colname="col4">193</oasis:entry>  
         <oasis:entry colname="col5">78.2</oasis:entry>  
         <oasis:entry colname="col6">40.94</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M302" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (mm h<inline-formula><mml:math id="M303" display="inline"><mml:mrow><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col2">0.37</oasis:entry>  
         <oasis:entry colname="col3">1.48</oasis:entry>  
         <oasis:entry colname="col4">2.35</oasis:entry>  
         <oasis:entry colname="col5">7.85</oasis:entry>  
         <oasis:entry colname="col6">7.11</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M304" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (min)</oasis:entry>  
         <oasis:entry colname="col2">80</oasis:entry>  
         <oasis:entry colname="col3">31</oasis:entry>  
         <oasis:entry colname="col4">75</oasis:entry>  
         <oasis:entry colname="col5">11</oasis:entry>  
         <oasis:entry colname="col6">1</oasis:entry>
       </oasis:row>
       <oasis:row>  
         <oasis:entry colname="col1"><inline-formula><mml:math id="M305" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>  
         <oasis:entry colname="col2">0.01</oasis:entry>  
         <oasis:entry colname="col3">0.42</oasis:entry>  
         <oasis:entry colname="col4">0.48</oasis:entry>  
         <oasis:entry colname="col5">0.89</oasis:entry>  
         <oasis:entry colname="col6">0.86</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <?xmltex \floatpos{t}?><fig id="Ch1.F8"><caption><p>Hierarchical clustering of the map into five subclasses. The
colors represent the subclass numbers:
subclass 1 is dark blue, subclass 2 is blue, subclass 3 is green, subclass 4 is orange, and subclass 5 is red.</p></caption>
          <?xmltex \igopts{width=170.716535pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f08.png"/>

        </fig>

      <p>From these five subclasses, two belong to the stratiform class and the other
three belong to the convective class. In the learning data set, the first
subclass represents 12 % of all events and 68, 1.2,
6.8 and 12 % for subclasses 2, 3, 4 and 5 respectively. The characteristics of these five
subclasses are summarized below and in Table 5. The five selected variables
are remarkably heterogeneous between classes, meaning the accuracy of these
variables for clustering:
<list list-type="bullet"><list-item>
      <p>Subclass 1 (drizzle and very light rain): the main feature of this class
is the very low mean value (<inline-formula><mml:math id="M306" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and standard deviation <inline-formula><mml:math id="M307" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> of
rain rate events, in addition to the features of the superclass. The mean
rain rate events lie in the range [0, 0.5] mm h<inline-formula><mml:math id="M308" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, with a mean value of
0.36 mm h<inline-formula><mml:math id="M309" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and <inline-formula><mml:math id="M310" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> in the range [0, 3] mm h<inline-formula><mml:math id="M311" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, with a
mean value equal to 0.1 mm h<inline-formula><mml:math id="M312" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. Although this event has a significant
duration, the corresponding subclass, which corresponds to drizzle, involves
only small quantities of water. It can also be noted that a low value of
<inline-formula><mml:math id="M313" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> is a good indicator (&lt; 0.01) for drizzle.</p></list-item><list-item>
      <p>Subclass 2 (“normal” events): this is a relatively broad class
containing 68 % of all events, with a mean event rain rate (<inline-formula><mml:math id="M314" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> in
the range [0.5 , 6] mm h<inline-formula><mml:math id="M315" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and a mean value of 1.48 mm h<inline-formula><mml:math id="M316" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. The
standard deviation <inline-formula><mml:math id="M317" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> lies in the range [1, 10] mm h<inline-formula><mml:math id="M318" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> with
a mean value of 2. This subclass is characterized by a significant relative
variation of some parameters (<inline-formula><mml:math id="M319" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M320" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>, for instance), together
with dry periods (<inline-formula><mml:math id="M321" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, which may be sufficiently long.</p></list-item></list></p>
      <p>The three remaining subclasses correspond to convective classes of events,
which are characterized by a strong temporal heterogeneity and significant
intensities. Depending on the depth of rain events, this convective class is
subdivided into three subclasses.
<list list-type="bullet"><list-item>
      <p>Subclass 3 contains relatively long events (<inline-formula><mml:math id="M322" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> with high values
for the rain event depth (<inline-formula><mml:math id="M323" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> variable and <inline-formula><mml:math id="M324" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>. This class
represents events with a very small likelihood of occurrence (1.2 %).</p></list-item><list-item>
      <p>Subclass 4 contains relatively short events (<inline-formula><mml:math id="M325" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> with peak rain
rate <inline-formula><mml:math id="M326" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> &gt; 50 mm h<inline-formula><mml:math id="M327" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, in addition to strong
heterogeneities (<inline-formula><mml:math id="M328" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M329" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M330" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula> are high) and large
values for the convective indicator (<inline-formula><mml:math id="M331" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">β</mml:mi><mml:mrow><mml:mi>L</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>.</p></list-item><list-item>
      <p>Subclass 5 contains events that  are characterized by relatively
low values for the rain event depth (<inline-formula><mml:math id="M332" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. This is due to the short
duration of the events (<inline-formula><mml:math id="M333" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. The variables <inline-formula><mml:math id="M334" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M335" display="inline"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mi>c</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>
remain high. Another feature of this subclass is that it includes continuous
events only, with no short, embedded dry periods (low values of <inline-formula><mml:math id="M336" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> in
Fig. 4 and Table 5).</p></list-item></list></p>
      <p>To conclude this section, this new classification allows the conventional
definition for stratiform events to be refined. The convective
classification can be subdivided into five different subclasses, each of
which is homogeneous. This classification is obtained for midlatitude
climates. As the data set used in this study is representative of only one
specific region and topography (i.e., the temperate climate encountered in
the Île-de-France region, France), its analysis cannot reveal information
related to different processes, i.e., those which are not sampled in the
data set. Such processes could lead to the identification of additional
specific clusters of events. In particular, there are no orographic rainfall
events or oceanic observations. The final step in this study involves
assessing whether the homogeneous character of each class is preserved at
the microphysics scale and attempting to identify any relationships between
the information present at the scale of both the microphysics and the
macrophysics of these events (hydrological information).</p>
</sec>
</sec>
<sec id="Ch1.S5">
  <title>Microphysical point of view</title>
      <p>Our study of the microphysical properties of rain is based on a
comprehensive analysis of its drop size distribution <inline-formula><mml:math id="M337" display="inline"><mml:mrow><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></inline-formula> corresponding to the
number of raindrops per unit volume and per interval of diameter <inline-formula><mml:math id="M338" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula>. The shape
of <inline-formula><mml:math id="M339" display="inline"><mml:mrow><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> reflects the microphysical processes involved. The identification of
various features of the drop size distribution, as well as the type of
precipitation, is very useful for many applications. As an example, this
information is used in the calculation of heating profiles in the
precipitation parameterization of atmospheric models to gain a more
detailed understanding of microphysical processes as well as for the
development of rain retrieval algorithms applied to remote sensing
observations. The microphysical characteristics of rainfall act as hidden
variables that affect the relationship between microwave remote sensing
measurements and the volume of water in a rainfall event (Ulaby et al., 1981;
Iguchi et al., 2009). It can thus be very useful to use conventional rain gauges to
determine the microphysical characteristics of rainfall events, thereby
improving the quality of active or passive remote sensing observations, and
the spatial properties of rainfall events in particular.</p>
      <p>A general expression for the drop size distribution defined by Testud et al. (2001) is commonly used in the literature. This allows a distinction to be
made between the stable shape function <inline-formula><mml:math id="M340" display="inline"><mml:mi>f</mml:mi></mml:math></inline-formula> and the variability induced by rain.
This variability is represented by two microphysical parameters, namely the
mass-weighted volume diameter (<inline-formula><mml:math id="M341" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and the parameter
<inline-formula><mml:math id="M342" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>. In some studies, the term
<inline-formula><mml:math id="M343" display="inline"><mml:mrow><mml:msub><mml:mi>N</mml:mi><mml:mi>w</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is used rather than <inline-formula><mml:math id="M344" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>. Not
all authors use exactly the same units; in particular, Bringi et al. (2003)
and Suh et al. (2016) use mm<inline-formula><mml:math id="M345" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> m<inline-formula><mml:math id="M346" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> for the units of <inline-formula><mml:math id="M347" display="inline"><mml:mrow><mml:msub><mml:mi>N</mml:mi><mml:mi>w</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> rather
than the unit m<inline-formula><mml:math id="M348" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, which is used in this study for
<inline-formula><mml:math id="M349" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>.

              <disp-formula id="Ch1.E5" content-type="numbered"><mml:math id="M350" display="block"><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mi>D</mml:mi><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>)</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mo>[</mml:mo><mml:msup><mml:mi mathvariant="normal">m</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup><mml:mo>]</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

        where <inline-formula><mml:math id="M351" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M352" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> are defined as

              <disp-formula id="Ch1.E6" content-type="numbered"><mml:math id="M353" display="block"><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mo>[</mml:mo><mml:mi mathvariant="normal">mm</mml:mi><mml:mo>]</mml:mo><mml:mo>,</mml:mo><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msup><mml:mn mathvariant="normal">4</mml:mn><mml:mn mathvariant="normal">4</mml:mn></mml:msup></mml:mrow><mml:mrow><mml:mi mathvariant="normal">Γ</mml:mi><mml:mo>(</mml:mo><mml:mn mathvariant="normal">4</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:mfrac></mml:mstyle><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msubsup><mml:mi>M</mml:mi><mml:mn mathvariant="normal">3</mml:mn><mml:mn mathvariant="normal">5</mml:mn></mml:msubsup></mml:mrow><mml:mrow><mml:msubsup><mml:mi>M</mml:mi><mml:mn mathvariant="normal">4</mml:mn><mml:mn mathvariant="normal">4</mml:mn></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mo>[</mml:mo><mml:msup><mml:mi mathvariant="normal">m</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup><mml:mo>]</mml:mo><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

        and <inline-formula><mml:math id="M354" display="inline"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is <inline-formula><mml:math id="M355" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>th-order moment of the drop size distribution <inline-formula><mml:math id="M356" display="inline"><mml:mrow><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>:
          <disp-formula id="Ch1.E7" content-type="numbered"><mml:math id="M357" display="block"><mml:mrow><mml:msub><mml:mi>M</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mrow><mml:mo>+</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow></mml:munderover><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo><mml:msup><mml:mi>D</mml:mi><mml:mi>i</mml:mi></mml:msup><mml:mi mathvariant="normal">d</mml:mi><mml:mi>D</mml:mi><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
        Rain samples are usually analyzed by computing the microphysical parameters
(<inline-formula><mml:math id="M358" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M359" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> for each rain sample
obtained over a given timescale. In the present study, <inline-formula><mml:math id="M360" display="inline"><mml:mrow><mml:mi>N</mml:mi><mml:mo>(</mml:mo><mml:mi>D</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is obtained by
considering the entire raindrop collection corresponding to each rain event
of (variable) duration <inline-formula><mml:math id="M361" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>. This approach leads to one pair (<inline-formula><mml:math id="M362" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi>m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M363" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> of microphysics variables per rain event, whereas most other
authors rely on values computed over a fixed timescale.</p>

      <?xmltex \floatpos{t}?><fig id="Ch1.F9" specific-use="star"><caption><p>Microphysical variable <inline-formula><mml:math id="M364" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> versus
<inline-formula><mml:math id="M365" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> for the five rainy event subclasses. The three
neurons corresponding to mixed events are circled. The dashed lines
correspond to borders <inline-formula><mml:math id="M366" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> &gt; 1.66 and
Log(<inline-formula><mml:math id="M367" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> &gt; 6.15.</p></caption>
        <?xmltex \igopts{width=369.885827pt}?><graphic xlink:href="https://amt.copernicus.org/articles/10/1557/2017/amt-10-1557-2017-f09.png"/>

      </fig>

      <p>Projections of the learned map, according to <inline-formula><mml:math id="M368" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M369" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>,
are shown in Fig. 4 (bottom right). It can be seen that the two maps are
well structured and that these two parameters have opposite influences on
the map projection. Although these two microphysical parameters were not
learned, the relationship between them is clearly accounted for by the
information used to structure the map (the five selected variables). Moreover,
the existence of a relationship between the microphysical and macrophysical
features of the rainfall is also confirmed in this figure, since both of the
macrophysical variables used to learn the SOM, i.e., <inline-formula><mml:math id="M370" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi mathvariant="normal">R</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and
<inline-formula><mml:math id="M371" display="inline"><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mi mathvariant="normal">max</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, have patterns similar to those revealed on the <inline-formula><mml:math id="M372" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> map.</p>
      <p>Many authors, including Atlas et al. (1999), Bringi et al. (2003), Marzuki et
al. (2013) and Suh et al. (2016), have endeavored to associate specific
microphysical properties with each type of precipitation (convective or
stratiform). In view of the maps shown in Fig. 4 and the
convective–stratiform classification developed in Sect. 4.3, we are able
to confirm that precipitation events classified as stratiform express small
values for <inline-formula><mml:math id="M373" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and large values for <inline-formula><mml:math id="M374" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>. In the case of the
convective class, the opposite trend is observed (i.e., larger values for
<inline-formula><mml:math id="M375" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and smaller values for <inline-formula><mml:math id="M376" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. Similar observations have
been reported by Testud et al. (2001). It can also be noticed that the two
microphysical variables are relatively homogeneous in the convective class,
whereas in the stratiform class they are characterized by a higher level of
variability.</p>
      <p>In order to improve our analysis of the microphysical information embedded
in the data set, we analyzed the relationship between the two microphysical
parameters using the reference vectors (neurons) from the map, which include
information related to the original rain events.</p>
      <p>Figure 9 shows the variable <inline-formula><mml:math id="M377" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> as a function of <inline-formula><mml:math id="M378" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> for the
64 neurons on the map. This relationship is indicated through the use of
distinct markers to identify the five subclasses defined in Sect. 4.4,
thus facilitating the discussion of the microphysics associated with
stratiform and convective rain. The two solid lines show the linear
regressions computed for these two classes.</p>
      <p>In the case of the stratiform subclasses (1 and 2) a clear relationship can
be observed between the two variables. The microphysics characteristics of
these two subclasses are clearly distinct. Indeed, subclass 1 (drizzle and
light rain) has the smallest <inline-formula><mml:math id="M379" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and the highest <inline-formula><mml:math id="M380" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and
varies over just a small range. Conversely, as in the case of the
macrophysical variables (see Sect. 4.4), the microphysical characteristics
of subclass 2 (normal events) are considerably more heterogeneous. Knowledge
of <inline-formula><mml:math id="M381" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> makes it straightforward to identify the corresponding subclass.
As a consequence, an event with <inline-formula><mml:math id="M382" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> lying in the range [0.5, 1]
millimeter belongs to subclass 1. Similarly, it is very likely that an event
with <inline-formula><mml:math id="M383" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> lying in the range [1, 1.7] millimeter belongs to subclass 2.</p>
      <p>For the convective events (subclasses 3, 4, 5), small differences can be
noticed with respect to <inline-formula><mml:math id="M384" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>. In the range [1.7, 2.5] mm, two
neurons belonging to subclass 4 are close to a neuron belonging to subclass
5, and they therefore have similar microphysics. Although they are located far
from all other subclass 2 neurons, three isolated neurons belonging to
subclass 2 (stratiform) can be noted. These are characterized by relatively
strong values of <inline-formula><mml:math id="M385" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> (2 mm) and low values of <inline-formula><mml:math id="M386" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula>. The
corresponding events are a mixture of stratiform and convective rain. A
typical case is given by convective rain associated with strong rain rates
occurring at the beginning of an event, whereas the remainder of the event
is stratiform with low rain rates and small variations.</p>
      <p>Following our classification, Fig. 9 indicates that there are real
relationships between the macrophysical and microphysical variables.
Nevertheless, knowledge of the variables (<inline-formula><mml:math id="M387" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M388" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> does not
allow the correct subclass to be determined in all cases.</p>
      <p>Researchers who study microphysical features and their association with
specific types of precipitation use simple schemes, based on rain rate
estimations over a fixed period of integration (a few minutes), in order to
separate stratiform and convective rain types. They also use these simple
schemes to label <inline-formula><mml:math id="M389" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M390" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> as
stratiform or convective (Testud et al., 2001). This approach is
significantly different to the method presented here, which assumes that all
of the samples in a given event belong to the same class. Our values for
<inline-formula><mml:math id="M391" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M392" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> are thus computed for
the timescale of a given event rather than for a fixed integration time.
Thus, although in the present study a good agreement is found for the range
of values covered by <inline-formula><mml:math id="M393" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, those determined for
<inline-formula><mml:math id="M394" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> do not cover the same range as in
the case of the previously cited studies.</p>
      <p>Many previous authors have observed that the drop size distribution is
closely related to processes controlling rainfall development mechanisms. In
the case of stratiform rainfall, the residence time of the drops is
relatively long and the raindrops grow by the accretion mechanism. In
convective rainfall, raindrops grow by the collision–coalescence mechanism,
associated with relatively strong vertical wind speeds. Numerous studies
have been published concerning the variability of
<inline-formula><mml:math id="M395" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M396" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi>m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>: Bringi et al. (2003)
studied rain samples from diverse climates and analyzed their variability in
stratiform and convective rainfall; Marzuki et al. (2013) investigated the
variability of the raindrop size distribution through a network of Parsivel
disdrometers in Indonesia; and Suh et al. (2016) investigated the raindrop
size distribution in Korea using a POSS disdrometer. In the case of
stratiform rain, all of these authors observe that
<inline-formula><mml:math id="M397" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M398" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> are nearly
log-linearly related, with a negative slope. This is consistent with the
trend shown in Fig. 9 for the two stratiform subclasses (1 and 2). Even the
three distinct neurons, which are isolated from the others, appear to be
governed by the same relationship.</p>
      <p>Marzuki et al. (2013) noted that during convective rain the increase in
value of <inline-formula><mml:math id="M399" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> with decreasing
<inline-formula><mml:math id="M400" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is nearly log linear, with a flatter slope. In the present case, the
dependence is also log linear, with a slope that is slightly flatter for
convective events than for stratiform events. In the aforementioned studies,
the data were aggregated over time, campaign or site, on the basis of a
criterion computed over a fixed period of time. We believe that this process
is weakly suited to determining the properties of convective events, as a
consequence of their strong variability and shorter characteristic time. In
this study we were able to retrieve the log-linear relationship between
<inline-formula><mml:math id="M401" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M402" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> without having to
learn it directly.</p>
      <p>When applying our algorithm to the various macroscopic properties by rain
event, we also take into account the variability of rain within an
individual rain event. Fig. 9 clearly shows that the spreading of parameters
<inline-formula><mml:math id="M403" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M404" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> inside each subclass
has the same magnitude as the distance between subclasses. This remark
confirms the hypothesis of Tapiador et al. (2010): the intra-event
variability can exceed the inter-event variability due to events arising
from different precipitation systems. It is thus preferable to examine the
properties of events with a more general approach rather than using
individual samples to study the distinction between stratiform and
convective processes. The three isolated neurons in subclass 2 described
above (circled in Fig. 9) have the same properties as the other events of
their subclass (i.e., the same slope for the log-linear relationship between
<inline-formula><mml:math id="M405" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M406" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. This example
confirms the ability of our methodology to preserve the macroscopic
information needed to cluster rain events, thus allowing the intra-event
variability as well as microphysical information to be (partially)
retrieved.</p>
      <p>Suh et al. (2016) also compare <inline-formula><mml:math id="M407" display="inline"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M408" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> probability density functions (pdf), for the case of stratiform and convective
samples over a 4-year period. On the basis of the <inline-formula><mml:math id="M409" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> pdf of both
stratiform and convective classes, they compute a threshold value for
<inline-formula><mml:math id="M410" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, such that when <inline-formula><mml:math id="M411" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> &gt; 1.66 mm the rainfall samples are
mainly convective and when <inline-formula><mml:math id="M412" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> &lt; 1.66 mm they are mainly
stratiform. This finding is consistent with the results of Atlas et al. (1999), who also found a threshold value for <inline-formula><mml:math id="M413" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>, distinguishing between
convective and stratiform rainfall. In Fig. 9, it can be seen that this
threshold is confirmed (vertical solid line), with <inline-formula><mml:math id="M414" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> smaller than 1.6 mm corresponding to stratiform events, whereas higher values correspond to
mainly convective events. When we consider the events analyzed in the
present study, there are also three neurons corresponding to a “mixed
event” beyond this threshold.</p>
      <p>Suh et al. (2016) show in Fig. 4c of their study that the pdf for convective
rainfall is higher than that corresponding to stratiform rainfall, when
<inline-formula><mml:math id="M415" display="inline"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> &gt; 6.2
(<inline-formula><mml:math id="M416" display="inline"><mml:mrow><mml:msub><mml:mi>N</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo>=</mml:mo></mml:mrow></mml:math></inline-formula> 3.2 in their figure). As described above, by considering the
data corresponding to rain events, rather than to samples recorded over
fixed periods of time, our range of values for
<inline-formula><mml:math id="M417" display="inline"><mml:mrow><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup></mml:mrow></mml:math></inline-formula> is smaller than that used in
other publications. In addition,
<inline-formula><mml:math id="M418" display="inline"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>N</mml:mi><mml:mn mathvariant="normal">0</mml:mn><mml:mo>*</mml:mo></mml:msubsup><mml:mo>)</mml:mo><mml:mo>&lt;</mml:mo><mml:mn mathvariant="normal">6.15</mml:mn></mml:mrow></mml:math></inline-formula> for all
neurons labeled as convective in our study, which is very close to the
value of 6.2 determined by Suh et al. (2016).</p>
      <p>In view of the generally satisfactory retrieval of microphysical information
from macrophysical parameters, we are of the opinion that the topological
map successfully restores some of the information implicitly embedded in the
data set. It is thus interesting to note that the macrophysical parameters of
rainfall are related to its microphysical properties. Firstly, the map
collects similar events, whilst ensuring, through the minimization of
topological errors, that the unfolding of the map is correct. A neuron is
thus closer to its neighbors than to any other neuron on the map. This
criterion ensures that the data space is optimally partitioned into
connected subparts, such that the neurons on the map can be related to the
underlying processes governing rainfall.</p>
</sec>
<sec id="Ch1.S6" sec-type="conclusions">
  <title>Conclusions</title>
      <p>Although the definition of a rain event is relatively subjective, this
study underlines the advantages of using event analysis rather than sample
analysis. This data-driven analysis of events shows that rain events exhibit
coherent features. As a consequence of the discrete and intermittent nature
of rainfall, some of the features commonly used to describe rain processes
are inadequate, in particular when they defined for a fixed duration.
Excessively long integration times (hours or days) can lead to the mixing of
observations that correspond to distinct physical processes as well as to the
mixing of rainy and clear air periods, within the same sample. An
excessively short integration time (seconds, minutes) leads to noisy data,
which are sensitive to the sensor's characteristics (sensor area, detection
threshold and noise). By analyzing entire rain events, rather than short
individual samples of fixed duration, it is possible to clearly identify
certain relationships between the different features of rain events, in
particular the influence of the microphysical properties of rain on its
macrophysical characteristics. This approach allows the intra-event
variability caused by measurement uncertainties to be reduced, thus
improving the accuracy with which physical processes can be identified.</p>
      <p>Once an event has been clearly identified, it is possible to choose a small
number of variables to describe it. We present a new data-driven approach,
which can be used to select the most relevant variables for this
characterization. This approach has generic properties and can be adapted to
many multivariate applications. A GA, when combined with
SOM clustering, can allow the unsupervised selection
of an optimal subset of five macrophysical variables. This is achieved by
minimizing a score function, which depends on the topology error of the SOM
and the number of variables. This score provides a parsimonious description
of the event, whilst preserving as much as possible the topology of the
initial space.</p>
      <p>Numerous variables derived mainly from rain rate recordings are used to
describe precipitation in the context of rain time series studies and a
wide variety of topics of interest, including hydrology, meteorology,
climate and weather forecasting. The algorithm proposed in this study
produces a subspace formed by only 5 of the 23 rain features described in
the literature. We show that these five features can be selected by the
algorithm in an unsupervised manner and, from the macrophysical point of
view, can provide an adequate description of the main characteristics of
rainfall events. These characteristics are the event duration, the peak
rain rate, the rain event depth, the standard deviation of the event rain
rate and the absolute rain rate variation of order 0.5.</p>
      <p>In order to confirm the relevance of the five selected features, we analyze
the corresponding SOM and are able to clearly reveal the presence of
relationships between these features. This approach also reveals the
independence of the inter-event time (IET<inline-formula><mml:math id="M419" display="inline"><mml:mrow><mml:msub><mml:mi/><mml:mi mathvariant="normal">p</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> characteristic and the
weak dependence of the dry percentage in event (<inline-formula><mml:math id="M420" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi mathvariant="normal">d</mml:mi><mml:mi mathvariant="italic">%</mml:mi><mml:mi mathvariant="normal">e</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> characteristic,
thus confirming that a rain time series can be considered as an alternating
series of independent rain events, interrupted by periods without rain.
Hierarchical clustering allows the well-known separation between stratiform
and convective events to be clearly identified. This dual classification is
then refined into a set of five relatively homogeneous subclasses. The
stratiform class is divided into two subclasses: a drizzle/very light rain
subclass and a normal event subclass. The convective class is divided into
three subclasses, characterized by a strong temporal heterogeneity and
significant rain rates.</p>
      <p>As this research was based on the analysis of observations made in
midlatitude plains in France, the relevance of this classification remains
to be confirmed through the analysis of data sets recorded in different
climatic zones and under different meteorological conditions, such as those
encountered in mountainous or coastal areas. If the SOM described in the
present study were learned with a more exhaustive data set, a larger map
would be produced, and this could reveal new types of rainfall behavior,
which remained undetected in the current data set. This point will be
addressed in future studies.</p>
      <p>The data-driven analysis of entire rain events (rather than the analysis of
fixed-length samples) is relevant to the study of interactions between the
macrophysical (based on the rain rate) and microphysical (based on raindrop)
properties of rain. In the present study, several strong relationships were
identified between these microphysical and macrophysical characteristics,
and we show that some of the five subclasses identified in this analysis
have specific microphysical characteristics. When a relationship between the
microphysical and macrophysical properties of rain is identified, this can
have many practical implications, especially for remote sensing. In the
context of weather radar applications, the microphysical properties of rain
are needed in order to estimate rain rates through the use of the Z–R
relationships. The estimation of microphysical rain characteristics, based
on easily observable rain gauge measurements, could play a significant role
in the development of the quantitative precipitation estimation (QPE).</p>
</sec>

      
      </body>
    <back><notes notes-type="dataavailability">

      <p>The dataset is available on request from the authors.</p>
  </notes><notes notes-type="competinginterests">

      <p>The authors declare that they have no conflict of interest.</p>
  </notes><ack><title>Acknowledgements</title><p>The authors wish to thank the team of SIRTA (Site Instrumental de Recherche par Télédétection Atmosphérique)
as well as ACTRIS-FR for financial support.
<?xmltex \hack{\newline}?><?xmltex \hack{\newline}?>
Edited by: G. Vulpiani<?xmltex \hack{\newline}?>
Reviewed by: D. Dunkerley, A. Parodi, and three anonymous referees</p></ack><ref-list>
    <title>References</title>

      <ref id="bib1.bib1"><label>1</label><mixed-citation>
Akrour, N., Chazottes, A., Verrier, S., Mallet, C., and Barthes, L.:
Simulation of yearly rainfall time series at microscale resolution with
actual properties: Intermittency, scale invariance, and rainfall
distribution, Water Resour. Res., 51, 7417–7435, 2015.</mixed-citation></ref>
      <ref id="bib1.bib2"><label>2</label><mixed-citation>
Atlas, D., Ulbrich, C. W., Marks, F. D., Amitai, E., and Williams, C. R.:
Systematic variation of drop size and radar-rainfall relations, J. Geophys.
Res.-Atmos., 104, 6155–6169, 1999.</mixed-citation></ref>
      <ref id="bib1.bib3"><label>3</label><mixed-citation>
Balme, M., Vischel, T., Lebel, T., Peugeot, C., and Galle, S.: Assessing
the water balance in the Sahel: impact of small scale rainfall variability
on runoff Part 1: rainfall variability analysis, J. Hydrol.,
331, 336–348, 2006.</mixed-citation></ref>
      <ref id="bib1.bib4"><label>4</label><mixed-citation>
Bringi, V. N., Chandrasekar, V., Hubbert, J., Gorgucci, E., Randeu, W. L.,
and Schoenhuber, M.: Raindrop size distribution in different climatic
regimes from disdrometer and dual-polarized radar analysis, J. Atmos. Sci.,
60, 354–365, 2003.</mixed-citation></ref>
      <ref id="bib1.bib5"><label>5</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Statistical analysis of
climatological data to characterize erosion potential: 1. Precipitation
Events in Western Oregon. Oregon Agricultural Experiment Station Spec. Rep.
No. 689, Oregon State University, 1983.</mixed-citation></ref>
      <ref id="bib1.bib6"><label>6</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Statistical analysis of
climatological data to characterize erosion potential: 4. Freezing events in
eastern Oregon/Washington. Oregon Agricultural Experiment Station Spec. Rep.
No. 689, Oregon State University, 1984.</mixed-citation></ref>
      <ref id="bib1.bib7"><label>7</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Exploratory Analysis of
Precipitation events with Implications for Stochastic Modeling, J. Clim. Appl. Meteorol., 57–67, 1985.</mixed-citation></ref>
      <ref id="bib1.bib8"><label>8</label><mixed-citation>
Cosgrove, C. M. and Garstang, M.: Simulation of rain events from rain-gauge
measurements, Int. J. Climatol., 15, 1021–1029, 1995.</mixed-citation></ref>
      <ref id="bib1.bib9"><label>9</label><mixed-citation>
Coutinho, J. V., Almeida, C. Das, N., Leal. A. M. F., and Barbarosa, L. R.:
Characterization of sub-daily rainfall properties in three rain gauges
located in northeast Brazil. Evolving Water Resources Systems:
Understanding, Predicting and Managing Water–Society Interactions
Proceedings of ICAR 2014, Bologna, Italy, 345–350, 2014.</mixed-citation></ref>
      <ref id="bib1.bib10"><label>10</label><mixed-citation>
Daumas, F.: Méthodes de normalisation de données, Revue de
statistique appliquée, 30, 23–38, 1982.</mixed-citation></ref>
      <ref id="bib1.bib11"><label>11</label><mixed-citation>
Delahaye, J.-Y., Barthès, L., Golé, P., Lavergnat J., and Vinson,
J. P.: a dual beam spectropluviometer concept, J. Hydrol.,
328, 110–120, 2006.</mixed-citation></ref>
      <ref id="bib1.bib12"><label>12</label><mixed-citation>
de Montera, L., Barthes, L., and Mallet, C.: The effect of rain-no rain
intermittency on the estimation of the Universal Multifractal model
parameters, J.  Hydrometeorol., 10,  493–506, 2009.</mixed-citation></ref>
      <ref id="bib1.bib13"><label>13</label><mixed-citation>
Driscoll, E. D., Palhegyi, G. E., Strecker, E. W., and Shelley, P. E.:
Analysis of storm events characteristics for selected rainfall
gauges throughout the United States, US Environmental Protection Agency, Washington, DC, 1989.</mixed-citation></ref>
      <ref id="bib1.bib14"><label>14</label><mixed-citation>
Dunkerley, D.: Rain event properties in nature and in rainfall simulation
experiments: a comparative review with recommendations for increasingly
systematic study and reporting, Hydrol. Process., 22, 4415–4435,
2008a.</mixed-citation></ref>
      <ref id="bib1.bib15"><label>15</label><mixed-citation>
Dunkerley, D.: Identifying individual rain events from pluviograph records:
a review with analysis of data from an Australian dryland site, Hydrol. Process., 22, 5024–5036, 2008b.</mixed-citation></ref>
      <ref id="bib1.bib16"><label>16</label><mixed-citation>
Eagleson, P. S.: Dynamic Hydrology, McGraw-Hill, 1970.</mixed-citation></ref>
      <ref id="bib1.bib17"><label>17</label><mixed-citation>
Everitt, B.: Cluster Analysis, London: Heinemann Educ. Books, 1974.</mixed-citation></ref>
      <ref id="bib1.bib18"><label>18</label><mixed-citation>
Gargouri, E. and Chebchoub, A.: Modélisation de la structure de
dépendance hauteur-durée d'événements pluvieux par la copule
de Gumbel, Hydrological Sciences-Journal-des Sciences Hydrologiques,
53, 802–817, 2010.</mixed-citation></ref>
      <ref id="bib1.bib19"><label>19</label><mixed-citation>Grazioli, J., Tuia, D., and Berne, A.: Hydrometeor classification from polarimetric radar measurements: a clustering approach,
Atmos. Meas. Tech., 8, 149–170, <ext-link xlink:href="http://dx.doi.org/10.5194/amt-8-149-2015" ext-link-type="DOI">10.5194/amt-8-149-2015</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bib20"><label>20</label><mixed-citation>
Guyon, I. and Elisseeff, A.: An Introduction to Variable and Feature
Selection, Kernel Machines Section, 3, 1157–1182, 2003.</mixed-citation></ref>
      <ref id="bib1.bib21"><label>21</label><mixed-citation>Haile, A. T., Rientjes, T. H. M., Habib, E., Jetten, V., and Gebremichael, M.: Rain event properties at the source of
the Blue Nile River, Hydrol. Earth Syst. Sci., 15, 1023–1034, <ext-link xlink:href="http://dx.doi.org/10.5194/hess-15-1023-2011" ext-link-type="DOI">10.5194/hess-15-1023-2011</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bib22"><label>22</label><mixed-citation>
Holland, J. H.: Adaptation In Natural And Artificial Systems, University of
Michigan Press, 1975.</mixed-citation></ref>
      <ref id="bib1.bib23"><label>23</label><mixed-citation>
Iguchi, T., Kozu, T., Kwiatkowski, J., Meneghini, R., Awaka, J., and
Okamoto, K.: Uncertainties in the rain profiling algorithm for the TRMM
precipitation radark, J. Meteorol. Soc. Jpn., 87A, 1–30, 2009.</mixed-citation></ref>
      <ref id="bib1.bib24"><label>24</label><mixed-citation>
Kohonen, T.: Self-organizing formation of topologically correct feature
maps, Biological Cybernetics, 46, 59–69, 1982.</mixed-citation></ref>
      <ref id="bib1.bib25"><label>25</label><mixed-citation>
Kohonen, T.: Self-Organizing Maps. Springer-Verlag, ISBN
3-540-67921-9, New York, Berlin, Heidelberg, 2001.</mixed-citation></ref>
      <ref id="bib1.bib26"><label>26</label><mixed-citation>Larsen, M. L. and Teves, J. B.: Identifying Individual Rain Events with a
Dense Disdrometer Network, Adv. Meteorol., 2015, ID582782, <ext-link xlink:href="http://dx.doi.org/10.1155/2015/582782" ext-link-type="DOI">10.1155/2015/582782</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bib27"><label>27</label><mixed-citation>
Lavergnat, J. and Golé, P.: A Stochastic Raindrop Time Distribution
Model, J. Appl. Meteorol., 37, 805–818, 1998.</mixed-citation></ref>
      <ref id="bib1.bib28"><label>28</label><mixed-citation>
Lavergnat, J. and Golé, P.: A stochastic model of raindrop release:
Application to the simulation of point rain observations, J. Hydrol., 328, 8–19, 2006.</mixed-citation></ref>
      <ref id="bib1.bib29"><label>29</label><mixed-citation>
Liu, Y. and Weisberg R. H.: A review of self-organizing map applications in
meteorology and oceanography, in: Self-Organizing Maps-Applications and
Novel Algorithm Design, 253–272, 2011.</mixed-citation></ref>
      <ref id="bib1.bib30"><label>30</label><mixed-citation>Liu, Y., Weisberg, R. H., and Mooers, C. N. K.: Performance evaluation of the
self- organizing map for feature extraction, J. Geophys.
Res., 111, C05018, <ext-link xlink:href="http://dx.doi.org/10.1029/2005JC003117" ext-link-type="DOI">10.1029/2005JC003117</ext-link>,  2006.</mixed-citation></ref>
      <ref id="bib1.bib31"><label>31</label><mixed-citation>Llasat, M. C.: An objective classification of rainfall events on the basis of
their convective features. Application to rainfall intensity in the north
east of Spain, Int. J. Climatol., 21, 1385–1400, 2001.
 </mixed-citation></ref><?xmltex \hack{\newpage}?>
      <ref id="bib1.bib32"><label>32</label><mixed-citation>
Marzuki, M., Hashiguchi, H., Yamamoto, M. K., Mori, S., and Yamanaka, M. D.:
Regional variability of raindrop size distribution over Indonesia, Ann.
Geophys., 31, 1941–1948, 2013.</mixed-citation></ref>
      <ref id="bib1.bib33"><label>33</label><mixed-citation>
Molini, L., Parodi, A., Rebora, N., and Craig, G. C.: Classifying severe
rainfall events over Italy by hydrometeorological and dynamical criteria,
Q. J. Roy. Meteorol. Soc., 137, 148–154,
2011.</mixed-citation></ref>
      <ref id="bib1.bib34"><label>34</label><mixed-citation>
Moussa, R. and Bocquillon, C.: Caractérisation fractale d'une série
chronologique d'intensité de pluie. Rencontres hydrologiques
Franco-Romaines, 363–370, 1991.</mixed-citation></ref>
      <ref id="bib1.bib35"><label>35</label><mixed-citation>Suh, S.-H., You, C.-H., and Lee, D.-I.: Climatological characteristics of raindrop size distributions in Busan,
Republic of Korea, Hydrol. Earth Syst. Sci., 20, 193–207, <ext-link xlink:href="http://dx.doi.org/10.5194/hess-20-193-2016" ext-link-type="DOI">10.5194/hess-20-193-2016</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bib36"><label>36</label><mixed-citation>Tapiador, F. J., Checa, R., and de Castro, M.: An experiment to measure the
spatial variability of rain drop size distribution using sixteen laser
disdrometers, Geophys. Res. Lett., 37, L16803, <ext-link xlink:href="http://dx.doi.org/10.1029/2010GL044120" ext-link-type="DOI">10.1029/2010GL044120</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bib37"><label>37</label><mixed-citation>
Testud, J. S., Oury, P., Amayenc, and Black, R. A.: The concept of
“normalized” distributions to describe raindrop spectra: A tool for cloud
physics and cloud remote sensing, J. Appl. Meteorol., 40, 1118–1140, 2001.</mixed-citation></ref>
      <ref id="bib1.bib38"><label>38</label><mixed-citation>
Ulaby, F. T., Moore, R. K., and Fung, A. K.: Microwave Remote Sensing:
Fundamentals and Radiometry, Vol. I. Artech House, 321-327, 1981.</mixed-citation></ref>
      <ref id="bib1.bib39"><label>39</label><mixed-citation>
Uriarte, E. A. and Martín, F. D., Topology Preservation in SOM, World
Academy of Science, Engineering and Technology, International Journal of
Computer, Electrical, Automation, Control and Information Engineering 2, 9,
2008.</mixed-citation></ref>
      <ref id="bib1.bib40"><label>40</label><mixed-citation>
Verrier, S., Barthès, L., and Mallet, C.: Theoretical and empirical scale
dependency of Z-R relationships: Evidence, impacts, and correction, J.
Geophys. Res.-Atmos., 118, 7435–7449, 2013.</mixed-citation></ref>
      <ref id="bib1.bib41"><label>41</label><mixed-citation>
Vesanto, J. and Alhoniemi, E.: Clustering of the self-organizing map, IEEE
Transactions on Neural Networks,  11, 586–600,  2000.</mixed-citation></ref>

  </ref-list><app-group content-type="float"><app><title/>

    </app></app-group></back>
    <!--<article-title-html>Data-driven clustering of rain events: microphysics information derived from macro-scale observations</article-title-html>
<abstract-html><p class="p">Rain time series records are generally studied using rainfall rate
or accumulation parameters, which are estimated for a fixed duration
(typically 1 min, 1 h or 1 day). In this study we use the concept of
<q>rain events</q>. The aim of the first part of this paper is to establish a
parsimonious characterization of rain events, using a minimal set of
variables selected among those normally used for the characterization of
these events. A methodology is proposed, based on the combined use of a
genetic algorithm (GA) and self-organizing maps (SOMs). It can be
advantageous to use an SOM, since it allows a high-dimensional data space to
be mapped onto a two-dimensional space while preserving, in an unsupervised
manner, most of the information contained in the initial space topology. The
2-D maps obtained in this way allow the relationships between variables to be
determined and redundant variables to be removed, thus leading to a minimal
subset of variables. We verify that such 2-D maps make it possible to
determine the characteristics of all events, on the basis of only five
features (the event duration, the peak rain rate, the rain event depth, the
standard deviation of the rain rate event and the absolute rain rate
variation of the order of 0.5). From this minimal subset of variables,
hierarchical cluster analyses were carried out. We show that clustering into
two classes allows the conventional convective and stratiform classes to be
determined, whereas classification into five classes allows this convective–stratiform
classification to be further refined. Finally, our study made
it possible to reveal the presence of some specific relationships between
these five classes and the microphysics of their associated rain events.</p></abstract-html>
<ref-html id="bib1.bib1"><label>1</label><mixed-citation>
Akrour, N., Chazottes, A., Verrier, S., Mallet, C., and Barthes, L.:
Simulation of yearly rainfall time series at microscale resolution with
actual properties: Intermittency, scale invariance, and rainfall
distribution, Water Resour. Res., 51, 7417–7435, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>2</label><mixed-citation>
Atlas, D., Ulbrich, C. W., Marks, F. D., Amitai, E., and Williams, C. R.:
Systematic variation of drop size and radar-rainfall relations, J. Geophys.
Res.-Atmos., 104, 6155–6169, 1999.
</mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>3</label><mixed-citation>
Balme, M., Vischel, T., Lebel, T., Peugeot, C., and Galle, S.: Assessing
the water balance in the Sahel: impact of small scale rainfall variability
on runoff Part 1: rainfall variability analysis, J. Hydrol.,
331, 336–348, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>4</label><mixed-citation>
Bringi, V. N., Chandrasekar, V., Hubbert, J., Gorgucci, E., Randeu, W. L.,
and Schoenhuber, M.: Raindrop size distribution in different climatic
regimes from disdrometer and dual-polarized radar analysis, J. Atmos. Sci.,
60, 354–365, 2003.
</mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>5</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Statistical analysis of
climatological data to characterize erosion potential: 1. Precipitation
Events in Western Oregon. Oregon Agricultural Experiment Station Spec. Rep.
No. 689, Oregon State University, 1983.
</mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>6</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Statistical analysis of
climatological data to characterize erosion potential: 4. Freezing events in
eastern Oregon/Washington. Oregon Agricultural Experiment Station Spec. Rep.
No. 689, Oregon State University, 1984.
</mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>7</label><mixed-citation>
Brown, B. G., Katz, R. W., and Murphy, A. H.: Exploratory Analysis of
Precipitation events with Implications for Stochastic Modeling, J. Clim. Appl. Meteorol., 57–67, 1985.
</mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>8</label><mixed-citation>
Cosgrove, C. M. and Garstang, M.: Simulation of rain events from rain-gauge
measurements, Int. J. Climatol., 15, 1021–1029, 1995.
</mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>9</label><mixed-citation>
Coutinho, J. V., Almeida, C. Das, N., Leal. A. M. F., and Barbarosa, L. R.:
Characterization of sub-daily rainfall properties in three rain gauges
located in northeast Brazil. Evolving Water Resources Systems:
Understanding, Predicting and Managing Water–Society Interactions
Proceedings of ICAR 2014, Bologna, Italy, 345–350, 2014.
</mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>10</label><mixed-citation>
Daumas, F.: Méthodes de normalisation de données, Revue de
statistique appliquée, 30, 23–38, 1982.
</mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>11</label><mixed-citation>
Delahaye, J.-Y., Barthès, L., Golé, P., Lavergnat J., and Vinson,
J. P.: a dual beam spectropluviometer concept, J. Hydrol.,
328, 110–120, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>12</label><mixed-citation>
de Montera, L., Barthes, L., and Mallet, C.: The effect of rain-no rain
intermittency on the estimation of the Universal Multifractal model
parameters, J.  Hydrometeorol., 10,  493–506, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>13</label><mixed-citation>
Driscoll, E. D., Palhegyi, G. E., Strecker, E. W., and Shelley, P. E.:
Analysis of storm events characteristics for selected rainfall
gauges throughout the United States, US Environmental Protection Agency, Washington, DC, 1989.
</mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>14</label><mixed-citation>
Dunkerley, D.: Rain event properties in nature and in rainfall simulation
experiments: a comparative review with recommendations for increasingly
systematic study and reporting, Hydrol. Process., 22, 4415–4435,
2008a.
</mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>15</label><mixed-citation>
Dunkerley, D.: Identifying individual rain events from pluviograph records:
a review with analysis of data from an Australian dryland site, Hydrol. Process., 22, 5024–5036, 2008b.
</mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>16</label><mixed-citation>
Eagleson, P. S.: Dynamic Hydrology, McGraw-Hill, 1970.
</mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>17</label><mixed-citation>
Everitt, B.: Cluster Analysis, London: Heinemann Educ. Books, 1974.
</mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>18</label><mixed-citation>
Gargouri, E. and Chebchoub, A.: Modélisation de la structure de
dépendance hauteur-durée d'événements pluvieux par la copule
de Gumbel, Hydrological Sciences-Journal-des Sciences Hydrologiques,
53, 802–817, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>19</label><mixed-citation>
Grazioli, J., Tuia, D., and Berne, A.: Hydrometeor classification from polarimetric radar measurements: a clustering approach,
Atmos. Meas. Tech., 8, 149–170, <a href="http://dx.doi.org/10.5194/amt-8-149-2015" target="_blank">doi:10.5194/amt-8-149-2015</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>20</label><mixed-citation>
Guyon, I. and Elisseeff, A.: An Introduction to Variable and Feature
Selection, Kernel Machines Section, 3, 1157–1182, 2003.
</mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>21</label><mixed-citation>
Haile, A. T., Rientjes, T. H. M., Habib, E., Jetten, V., and Gebremichael, M.: Rain event properties at the source of
the Blue Nile River, Hydrol. Earth Syst. Sci., 15, 1023–1034, <a href="http://dx.doi.org/10.5194/hess-15-1023-2011" target="_blank">doi:10.5194/hess-15-1023-2011</a>, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>22</label><mixed-citation>
Holland, J. H.: Adaptation In Natural And Artificial Systems, University of
Michigan Press, 1975.
</mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>23</label><mixed-citation>
Iguchi, T., Kozu, T., Kwiatkowski, J., Meneghini, R., Awaka, J., and
Okamoto, K.: Uncertainties in the rain profiling algorithm for the TRMM
precipitation radark, J. Meteorol. Soc. Jpn., 87A, 1–30, 2009.
</mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>24</label><mixed-citation>
Kohonen, T.: Self-organizing formation of topologically correct feature
maps, Biological Cybernetics, 46, 59–69, 1982.
</mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>25</label><mixed-citation>
Kohonen, T.: Self-Organizing Maps. Springer-Verlag, ISBN
3-540-67921-9, New York, Berlin, Heidelberg, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>26</label><mixed-citation>
Larsen, M. L. and Teves, J. B.: Identifying Individual Rain Events with a
Dense Disdrometer Network, Adv. Meteorol., 2015, ID582782, <a href="http://dx.doi.org/10.1155/2015/582782" target="_blank">doi:10.1155/2015/582782</a>, 2015.
</mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>27</label><mixed-citation>
Lavergnat, J. and Golé, P.: A Stochastic Raindrop Time Distribution
Model, J. Appl. Meteorol., 37, 805–818, 1998.
</mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>28</label><mixed-citation>
Lavergnat, J. and Golé, P.: A stochastic model of raindrop release:
Application to the simulation of point rain observations, J. Hydrol., 328, 8–19, 2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>29</label><mixed-citation>
Liu, Y. and Weisberg R. H.: A review of self-organizing map applications in
meteorology and oceanography, in: Self-Organizing Maps-Applications and
Novel Algorithm Design, 253–272, 2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>30</label><mixed-citation>
Liu, Y., Weisberg, R. H., and Mooers, C. N. K.: Performance evaluation of the
self- organizing map for feature extraction, J. Geophys.
Res., 111, C05018, <a href="http://dx.doi.org/10.1029/2005JC003117" target="_blank">doi:10.1029/2005JC003117</a>,  2006.
</mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>31</label><mixed-citation>
Llasat, M. C.: An objective classification of rainfall events on the basis of
their convective features. Application to rainfall intensity in the north
east of Spain, Int. J. Climatol., 21, 1385–1400, 2001.

</mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>32</label><mixed-citation>
Marzuki, M., Hashiguchi, H., Yamamoto, M. K., Mori, S., and Yamanaka, M. D.:
Regional variability of raindrop size distribution over Indonesia, Ann.
Geophys., 31, 1941–1948, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>33</label><mixed-citation>
Molini, L., Parodi, A., Rebora, N., and Craig, G. C.: Classifying severe
rainfall events over Italy by hydrometeorological and dynamical criteria,
Q. J. Roy. Meteorol. Soc., 137, 148–154,
2011.
</mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>34</label><mixed-citation>
Moussa, R. and Bocquillon, C.: Caractérisation fractale d'une série
chronologique d'intensité de pluie. Rencontres hydrologiques
Franco-Romaines, 363–370, 1991.
</mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>35</label><mixed-citation>
Suh, S.-H., You, C.-H., and Lee, D.-I.: Climatological characteristics of raindrop size distributions in Busan,
Republic of Korea, Hydrol. Earth Syst. Sci., 20, 193–207, <a href="http://dx.doi.org/10.5194/hess-20-193-2016" target="_blank">doi:10.5194/hess-20-193-2016</a>, 2016.
</mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>36</label><mixed-citation>
Tapiador, F. J., Checa, R., and de Castro, M.: An experiment to measure the
spatial variability of rain drop size distribution using sixteen laser
disdrometers, Geophys. Res. Lett., 37, L16803, <a href="http://dx.doi.org/10.1029/2010GL044120" target="_blank">doi:10.1029/2010GL044120</a>, 2010.
</mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>37</label><mixed-citation>
Testud, J. S., Oury, P., Amayenc, and Black, R. A.: The concept of
“normalized” distributions to describe raindrop spectra: A tool for cloud
physics and cloud remote sensing, J. Appl. Meteorol., 40, 1118–1140, 2001.
</mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>38</label><mixed-citation>
Ulaby, F. T., Moore, R. K., and Fung, A. K.: Microwave Remote Sensing:
Fundamentals and Radiometry, Vol. I. Artech House, 321-327, 1981.
</mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>39</label><mixed-citation>
Uriarte, E. A. and Martín, F. D., Topology Preservation in SOM, World
Academy of Science, Engineering and Technology, International Journal of
Computer, Electrical, Automation, Control and Information Engineering 2, 9,
2008.
</mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>40</label><mixed-citation>
Verrier, S., Barthès, L., and Mallet, C.: Theoretical and empirical scale
dependency of Z-R relationships: Evidence, impacts, and correction, J.
Geophys. Res.-Atmos., 118, 7435–7449, 2013.
</mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>41</label><mixed-citation>
Vesanto, J. and Alhoniemi, E.: Clustering of the self-organizing map, IEEE
Transactions on Neural Networks,  11, 586–600,  2000.
</mixed-citation></ref-html>--></article>
