<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "https://jats.nlm.nih.gov/nlm-dtd/publishing/3.0/journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0" article-type="research-article">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">HESS</journal-id><journal-title-group>
    <journal-title>Hydrology and Earth System Sciences</journal-title>
    <abbrev-journal-title abbrev-type="publisher">HESS</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Hydrol. Earth Syst. Sci.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1607-7938</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/hess-28-3367-2024</article-id><title-group><article-title>Regionalization of GR4J model parameters  for river flow prediction in Paraná, Brazil</article-title><alt-title>Regionalization of GR4J model parameters for river flow prediction in Paraná</alt-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Kuana</surname><given-names>Louise Akemi</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff2">
          <name><surname>Almeida</surname><given-names>Arlan Scortegagna</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="yes" rid="aff3 aff4">
          <name><surname>Mercuri</surname><given-names>Emílio Graciliano Ferreira</given-names></name>
          <email>emilio@ufpr.br</email>
        <ext-link>https://orcid.org/0000-0003-3101-7537</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff4">
          <name><surname>Noe</surname><given-names>Steffen Manfred</given-names></name>
          
        <ext-link>https://orcid.org/0000-0003-1514-1140</ext-link></contrib>
        <aff id="aff1"><label>1</label><institution>Programa de Pós-Graduação em Engenharia Ambiental, Universidade Federal do Paraná, Curitiba, Brazil</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>Sistema de Tecnologia e Monitoramento Ambiental do Paraná (Simepar), Curitiba, Brazil</institution>
        </aff>
        <aff id="aff3"><label>3</label><institution>Departamento de Engenharia Ambiental, Universidade Federal do Paraná, Curitiba, Brazil</institution>
        </aff>
        <aff id="aff4"><label>4</label><institution>Institute of Forestry and Engineering, Estonian University of Life Sciences, Tartu, Estonia</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Emílio Graciliano Ferreira Mercuri (emilio@ufpr.br)</corresp></author-notes><pub-date><day>29</day><month>July</month><year>2024</year></pub-date>
      
      <volume>28</volume>
      <issue>14</issue>
      <fpage>3367</fpage><lpage>3390</lpage>
      <history>
        <date date-type="received"><day>29</day><month>July</month><year>2023</year></date>
           <date date-type="rev-request"><day>21</day><month>August</month><year>2023</year></date>
           <date date-type="rev-recd"><day>31</day><month>May</month><year>2024</year></date>
           <date date-type="accepted"><day>10</day><month>June</month><year>2024</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2024 Louise Akemi Kuana et al.</copyright-statement>
        <copyright-year>2024</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024.html">This article is available from https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024.html</self-uri><self-uri xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024.pdf">The full text article is available as a PDF file from https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024.pdf</self-uri>
      <abstract><title>Abstract</title>

      <p id="d1e130">Regionalization methods dependent on hydrological models comprise techniques for transferring calibrated parameters in instrumented watersheds (donor basins) to non-instrumented watersheds (target basins). There is a lack of flow regionalization studies in regions with humid subtropical and hot temperate climates, and one of the main novelties of this research is to assess the regionalization of low flows in Paraná in the south of Brazil. In addition to filling this gap, this research presents innovative artificial-intelligence techniques for transferring parameters from hydrological models. This study aims to evaluate regionalization methods for transferring GR4J parameters and predicting river flow in catchments from the south of Brazil. We created a dataset for the state of Paraná with daily hydrological time series (precipitation, evapotranspiration, and river flow) and watershed physiographic and climatological indices for 126 catchments. Rigorous quality-controlling techniques were applied to recover data from 1979 to 2020. The regionalization methods compared in this study are based on simple spatial proximity, physiographic–climatic similarity, and regression by random forest techniques. Direct regression of <inline-formula><mml:math id="M1" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> was calculated using random forest techniques and compared with indirect methods, i.e. using regionalization of GR4J parameters. A set of 100 basins was used to train the regionalization models, and another 26 catchments (pseudo-non-instrumented) were used to evaluate and compare the performance of regionalizations. The GR4J model showed acceptable performances for the sample of 126 catchments, with 65 % of watersheds presenting a log-transformed Nash–Sutcliffe coefficient greater than 0.70 during the validation period. According to the evaluation carried out for the sample of 26 basins, regionalization based on physiographic–climatic similarity was shown to be the most robust method for the prediction of daily and <inline-formula><mml:math id="M2" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flow in basins from the state of Paraná. When increasing the number of donor basins, the method based on spatial proximity has comparable performance to the method based on physiographic–climatic similarity. Based on the physiographic–climatic characteristics of the basins, it was possible to classify six distinct groups of watersheds in Paraná. Each group shows similarities in forest cover, urban area, number of days with more than 150 mm of precipitation, and average duration of consecutive dry days. Although the physiographic–climatic similarity method obtained the best performance, the use of machine learning algorithms to regionalize the model parameters had good performance using climatic and physiographic indices as inputs. This research represents a proof of concept that basins without flow monitoring can have a good approximation of streamflow if physiographic–climatic information is provided.</p>
  </abstract>
    
<funding-group>
<award-group id="gs1">
<funding-source>Estonian Research Competency Council</funding-source>
<award-id>PRG1674</award-id>
</award-group>
<award-group id="gs2">
<funding-source>Horizon 2020</funding-source>
<award-id>871115</award-id>
</award-group>
<award-group id="gs3">
<funding-source>Coordenação de Aperfeiçoamento de Pessoal de Nível Superior</funding-source>
<award-id>PRINT</award-id>
</award-group>
</funding-group>
</article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d1e164">According to <xref ref-type="bibr" rid="bib1.bibx51" id="text.1"/>, regionalization methods dependent on rainfall–runoff models comprise techniques for transferring calibrated parameters in instrumented basins (donor basins) to non-instrumented basins (target basins). The study carried out by <xref ref-type="bibr" rid="bib1.bibx4" id="text.2"/> presents three techniques based on physical similarity, spatial proximity, and regression to estimate the parameters of three different hydrological models, with the purpose of predicting flows in watersheds that do not have monitoring. Although many advances have been made in this area of hydrology, there are still uncertainties in methods for estimating flows in ungauged basins <xref ref-type="bibr" rid="bib1.bibx22" id="paren.3"/>. Part of this is due to the uniqueness of each region across the globe, which concerns not only the uniqueness of each location but also the issue of availability of information (e.g. descriptive characteristics of basins and availability of hydrometeorological data). Additionally, hydrological systems are dependent on temporal and spatial scales with interactions between climate, vegetation, topography, and soil <xref ref-type="bibr" rid="bib1.bibx10 bib1.bibx26" id="paren.4"/> that make the task of estimating hydrological information in basins with little or no data challenging.</p>
      <p id="d1e179">The watersheds analysed in this research belong to the southern region of Brazil, with an area of approximately 199 315 km<inline-formula><mml:math id="M3" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula>. The hydrography of Paraná is composed mainly of the Iguazu River, Paraná River, Paranapanema River, Tibagi River, Ivaí River, and Piquiri River. The study of low flows and droughts is critical in the context of water availability in Brazil; river dams and reservoirs are used for power generation (70 % of Brazilian energy sector), to provide drinking water for the population, to irrigate crops, and to distribute water for industrial use <xref ref-type="bibr" rid="bib1.bibx17" id="paren.5"/>. The Paraná basin, a major hydroelectricity-producing region with 32 % (60 million people) of Brazil's population, experienced very severe drought in 2000 and 2014, compromising the water supply for 11 million people in São Paulo <xref ref-type="bibr" rid="bib1.bibx36" id="paren.6"/>. The state of Paraná faced one of the worst droughts in its history between 2020 and 2021 <xref ref-type="bibr" rid="bib1.bibx19 bib1.bibx28" id="paren.7"/>. There are few studies of flow regionalization in the south of Brazil <xref ref-type="bibr" rid="bib1.bibx29 bib1.bibx9" id="paren.8"/>; at the same time, the hydrological measurements and field work in the area are declining <xref ref-type="bibr" rid="bib1.bibx15 bib1.bibx35" id="paren.9"/>. Our work brings novel contributions for watersheds with similar climate and geography; also, it provides more information for governmental planning and management.</p>
      <p id="d1e207">In order to reveal research gaps and how our study goes beyond the existing literature, we highlight the following points: (i) the need to better understand regionalization techniques in a subtropical climate, which has very distinct and specific runoff generation mechanisms; (ii) a proof of concept that basins without flow monitoring can have a good approximation of streamflow if other physiographic–climatic indices are provided; and (iii) the fact that machine learning algorithms perform better with physiographic–climatic indices as inputs.</p>
      <p id="d1e210">The aim of this article is to improve the methodology for transferring parameters of the GR4J model calibrated in instrumented watersheds to predict daily flows in basins with little or no hydrological information. The performance of different regionalization methods are verified in Paraná basins that have a history of hydrometeorological data records. Other objectives are the following: (i) build a hydrological database for the state of Paraná, Brazil (the database consists of daily flow, precipitation, and evapotranspiration time series and catchment-related descriptive indices); (ii) develop and improve methods for transferring GR4J calibrated parameters through regionalization techniques based on spatial distance, physiographic–climatic similarity, and non-linear regression; and (iii) compare regionalization methods and random forest techniques to estimate the <inline-formula><mml:math id="M4" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flow.</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Data for state of Paraná</title>
      <p id="d1e233">The study area was delimited based on the hydrographic network of the state of Paraná, Brazil, which is available at Instituto Água e Terra (IAT) <xref ref-type="bibr" rid="bib1.bibx27" id="paren.10"/>, and on a rectangular polygon demarcated between latitudes of 22°15<inline-formula><mml:math id="M5" display="inline"><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup></mml:math></inline-formula>36<inline-formula><mml:math id="M6" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>′</mml:mo><mml:mo>′</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> and 26°54<inline-formula><mml:math id="M7" display="inline"><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup></mml:math></inline-formula>00<inline-formula><mml:math id="M8" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>′</mml:mo><mml:mo>′</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> S and longitudes of 48°00<inline-formula><mml:math id="M9" display="inline"><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup></mml:math></inline-formula>00<inline-formula><mml:math id="M10" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>′</mml:mo><mml:mo>′</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> and 54°42<inline-formula><mml:math id="M11" display="inline"><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup></mml:math></inline-formula>00<inline-formula><mml:math id="M12" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>′</mml:mo><mml:mo>′</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> W. Therefore, the study area includes the state of Paraná and extends into parts of the states of Santa Catarina and São Paulo, not completely covering the Paranapanema and Paraná river basins (Fig. <xref ref-type="fig" rid="Ch1.F1"/>).</p>

      <fig id="Ch1.F1" specific-use="star"><label>Figure 1</label><caption><p id="d1e328">River flow stations and watershed delineation.</p></caption>
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f01.png"/>

      </fig>

      <p id="d1e337">Time series of hydrometeorological observations were obtained at Agência Nacional de Águas e Saneamento Básico (ANA) via HidroWEB Portal, Instituto Nacional de Meteorologia (INMET), Sistema de Tecnologia e Monitoramento Ambiental do Paraná (Simepar), Água e Terra Institute (IAT), and Instituto de Desenvolvimento Rural do Paraná (IAPAR-EMATER).</p>
      <p id="d1e341">Although there are datasets at a national level, such as CAMELS-BR <xref ref-type="bibr" rid="bib1.bibx18" id="paren.11"/> and CABra <xref ref-type="bibr" rid="bib1.bibx3" id="paren.12"/>, the authors decided to construct a new dataset based on the hydrographic network of the state of Paraná. This network has a consistent topology and codification; its hierarchization was proposed by Otto Pfafstetter <xref ref-type="bibr" rid="bib1.bibx48" id="paren.13"/> and allows the extraction of information upstream and downstream of each river section <xref ref-type="bibr" rid="bib1.bibx55" id="paren.14"/>.</p>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Precipitation time series</title>
      <p id="d1e363">The <xref ref-type="bibr" rid="bib1.bibx33" id="text.15"/> quality control method was used to evaluate daily rainfall data series from 1389 stations, which are shown in Fig. <xref ref-type="fig" rid="Ch1.F2"/>. The method can be divided into four steps. First, stations coordinates, the period of operation, and the percentage of available data are checked. In the second stage, data that are not physically possible are identified and discarded, such as negative precipitation values and extreme events greater than 300 mm.</p>

      <fig id="Ch1.F2" specific-use="star"><label>Figure 2</label><caption><p id="d1e373">Location of pluviometric stations.</p></caption>
          <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f02.png"/>

        </fig>

      <p id="d1e382">The third stage consists of analysing the historical series of each station individually. For each year, a quality index is calculated, which depends on five factors, namely (i) the percentage of data available in each year of the series; (ii) the distribution of failures throughout the year, for which the penalty becomes greater for a series that has long continuous periods of failures; (iii) the probability that the series is formed by possible failures that have been padded with zeros, which penalizes the series that have monthly cumulative data equal to zero, indicating that “false zeros” are possible; (iv) the probability of systematic accumulation of two or more days of the week, which penalizes the station if a day of the week with a tendency for less rain than other days of the week is detected at the station; and (v) the probability that the series contains outliers. The quality index can range from 0 % to 100 %. Values equal to 100 % indicate absolute quality; values above 80 % are considered to be acceptable; and for values below 50 %, the quality is considered to be very low.</p>
      <p id="d1e386">The verification process between what was recorded at the station to be analysed (candidate station) and at neighbouring stations (auxiliary stations) is carried out in the fourth stage, also known as the relative quality control. At this stage, there are two indices that are relevant to the classification of daily values for the candidate post, which can be labelled as valid, “<inline-formula><mml:math id="M13" display="inline"><mml:mi>V</mml:mi></mml:math></inline-formula>”; doubtful, “<inline-formula><mml:math id="M14" display="inline"><mml:mi>D</mml:mi></mml:math></inline-formula>”; invalid, “<inline-formula><mml:math id="M15" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula>”; or insufficient information, “<inline-formula><mml:math id="M16" display="inline"><mml:mi>I</mml:mi></mml:math></inline-formula>”. The records identified as having insufficient information denote that there are fewer than two auxiliary stations in the region and in the same period to be properly evaluated. The first index, called the representativeness index, verifies the daily values for each station pair, candidate–auxiliary; this index considers the distance between stations, the altitude difference, and the correlation with measured data. A maximum distance of up to 50 km was defined between candidate and auxiliaries stations. The second index is used to analyse the monthly cumulative data. From the Simepar stations, maximum limits of monthly accumulation were established and were applied to evaluate the historical series of each candidate station.</p>
      <p id="d1e417">Validated daily data were spatialized on a 1 km <inline-formula><mml:math id="M17" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> 1 km grid using space–time kriging, where precipitation values were estimated on a daily scale using weighted averages between neighbourhood data. Then, a second spatialization method was applied to the most recent precipitation history. The method presented by <xref ref-type="bibr" rid="bib1.bibx16" id="text.16"/> was used to estimate precipitation spatially within the area of interest, where the Poisson equation was used to combine radar and satellite data with the records observed by telemetry precipitation stations. Finally, average rainfall was measured using the arithmetic mean of the grid points located within the drainage area of each watershed.</p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>River flow data and delineation of watersheds</title>
      <p id="d1e438">We constructed the river flow dataset in two ways: (i) by directly obtaining river flow time series, where observed water levels were previously transformed into flow by the agency responsible for operating the station, and (ii) by time series of water levels which still had not been transformed into flow; therefore, when available, the station's rating curves were obtained, and then the quota was transformed into flow. We have considered it to be acceptable to use stations with at least 5 years of river flow records for the application of regionalization methods. Time series from stations downstream of flow regularization (dams and reservoirs) were discarded or considered partially in the periods prior to the dams' construction.</p>
      <p id="d1e441">Conventional stations were obtained from the IAT and HidroWEB Portal databases. Although both banks preserve information from stations of different operators, it was accepted that information coming from the IAT bank would have priority over the ones from the HidroWEB Portal. The inventory provided by the technician responsible for the IAT informs us that there are 413 river flow stations in the study area with time series greater than 5 years. From the ANA metadata catalogue, only 15 different stations with at least 5 years of river flow records were identified. The telemetric series of 83 IAT stations and 57 Simepar stations were obtained from the Simepar database.</p>
      <p id="d1e444">Locations of the stations were checked through the manual procedure of hydro-referencing using the hydrographic network of the IAT. Finally, a quality control was carried out, in which non-consistent data were disregarded, such as sudden ruler changes clearly altering the base flow, series with large gaps alternating with short measurement periods, low-precision measurements, or measurements that presented constant values for long periods. In the end, a total of 284 river flow stations were obtained, with observations ranging from 1926 to 2020, as shown on the map in Fig. <xref ref-type="fig" rid="Ch1.F1"/>.</p>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>Potential evapotranspiration</title>
      <p id="d1e458">The FAO Penman–Monteith <xref ref-type="bibr" rid="bib1.bibx2" id="paren.17"/> equation was used to estimate potential evapotranspiration (ET). This method requires time series of air temperature, air relative humidity, wind speed, and solar radiation, which were obtained through Simepar telemetric stations with records ranging from 1997 to 2020.  ET was based on long-term average daily values, which means the same potential evapotranspiration series was repeated every year for each station. Subsequently, the punctual information was spatialized using a method of regression followed by interpolation, also known as regression kriging or a hybrid method of interpolation <xref ref-type="bibr" rid="bib1.bibx24" id="paren.18"/>.</p>
</sec>
<sec id="Ch1.S2.SS4">
  <label>2.4</label><title>Catchment descriptors</title>
      <p id="d1e475">Table <xref ref-type="table" rid="App1.Ch1.S1.T3"/> in the Appendix shows catchment descriptor statistics (mean, standard deviation, quartiles, minimum and maximum) for the 126 basins of the Paraná dataset. It has 39 descriptive indices divided into four categories: physiographic, climatological, land use or land cover, and soil type. Quantitative indices were used to describe the landscape, relief, climate, topology of the hydrography, land use, and soil type of the watershed. Physiographic indices were obtained for each geographic location and for the topography of the drainage networks for the selected basins. From the hydrographic network, areas and drainage sections were obtained, which were used as a basis for calculating the indices described in Table <xref ref-type="table" rid="App1.Ch1.S1.T3"/>. A digital elevation model (DEM) with a resolution of 30 m from NASA's Shuttle Radar Topography Mission (SRTM) was used to estimate slopes and altitudes. Land use and land cover maps for the year 2019, provided by MapBiomas <xref ref-type="bibr" rid="bib1.bibx56" id="paren.19"/>, were used for calculating the fractions of area that each class occupies in the basins and to determine the dominant class. The soil map was obtained from <xref ref-type="bibr" rid="bib1.bibx21" id="text.20"/> for the state of Paraná, with a scale of <inline-formula><mml:math id="M18" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">250</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mn mathvariant="normal">000</mml:mn></mml:mrow></mml:math></inline-formula>.</p>
      <p id="d1e503">The curve number (CN) method developed by <xref ref-type="bibr" rid="bib1.bibx54" id="text.21"/> relates soil and land use and land cover information to classify the region based on its storm water retention potential. The ANA metadata catalogue was used to estimate the CN in Paraná basins. Average precipitation series and potential evapotranspiration estimates, which were previously determined for each watershed, were used to calculate the indices related to precipitation and potential evapotranspiration. Furthermore, <xref ref-type="bibr" rid="bib1.bibx7" id="text.22"/> provided atlases of the state of Paraná with monthly average temperatures and average solar radiation for each season of the year. The atlases were produced based on measurements from the INMET, Simepar, and Instituto de Desenvolvimento Rural do Paraná (IDR-Paraná) stations during the period of 2006 to 2016. This information was used to compute average indices in basins located within the state, and for the catchments on borders or in other states, the average values of the nearest watershed were adopted.</p>
</sec>
<sec id="Ch1.S2.SS5">
  <label>2.5</label><title>Watershed selection</title>
      <p id="d1e520">The watershed selection consists of 126 river basins that have at least 15 years of flow data between 1979 and 2020, with each year counted having a maximum of 10 % of gaps. In addition, it was preferred that the historical series also had more recent data, which extended beyond the year 2010, and were limited to homogeneous historical series that passed the Pettitt test <xref ref-type="bibr" rid="bib1.bibx47" id="paren.23"/>. The non-parametric Pettitt test was calculated using the library pyHomogeneity in Python; it is able to indicate the year in which a sudden change in the temporal trend occurred. The selected 126 watersheds are depicted in Fig. <xref ref-type="fig" rid="Ch1.F3"/>. Figure <xref ref-type="fig" rid="App1.Ch1.S1.F12"/> shows the availability of data over the years; darker green indicates a greater amount of data being available in that year.</p>

      <fig id="Ch1.F3" specific-use="star"><label>Figure 3</label><caption><p id="d1e532">Location of selected watersheds in the state of Paraná.</p></caption>
          <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f03.png"/>

        </fig>

</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Methods</title>
      <p id="d1e550">The application of the physiographic–climatic similarity method implicitly considers two assumptions. The first assumption is that, if there is similarity between basins, there are similar hydrological responses. The second assumption is that the similarity between sets of calibrated parameters of the hydrological model between two or more river basins may reflect the similarity in terms of their behaviour in relation to the transformation of rainfall into flow <xref ref-type="bibr" rid="bib1.bibx41 bib1.bibx42 bib1.bibx44 bib1.bibx10" id="paren.24"/>. Spatial proximity assumes that neighbouring basins have similarities in climate, soil type, land use and cover, slope, altitude, and other characteristics <xref ref-type="bibr" rid="bib1.bibx4" id="paren.25"/>. Non-linear regression models seek to equate the relationship between dependent variables (e.g. hydrological model parameters) with different independent variables (e.g. descriptive characteristics; <xref ref-type="bibr" rid="bib1.bibx23" id="altparen.26"/>).</p>
      <p id="d1e562">After the dataset construction, calibration and validation of the GR4J model was performed in all basins. The hydrological model, GR4J (Génie Rural à 4 paramètres Journalier), proposed by <xref ref-type="bibr" rid="bib1.bibx46" id="text.27"/>, has been implemented in different countries, such as France  <xref ref-type="bibr" rid="bib1.bibx41 bib1.bibx42" id="paren.28"/>, Australia <xref ref-type="bibr" rid="bib1.bibx43" id="paren.29"/>, Brazil <xref ref-type="bibr" rid="bib1.bibx40" id="paren.30"/>, South Korea <xref ref-type="bibr" rid="bib1.bibx53" id="paren.31"/>, Mexico <xref ref-type="bibr" rid="bib1.bibx4" id="paren.32"/>, and Russia <xref ref-type="bibr" rid="bib1.bibx6" id="paren.33"/>. This parsimonious model has been showing promising results and stands out due to its dependence on only a few parameters and its use of two meteorological forcing variables on a daily scale; these variables are the total precipitation and potential evapotranspiration averaged at the basin scale, requiring historical series of observed flows for the adjustment of its four parameters. Three regionalization methods of the GR4J constants were tested; they are based on (i) physiographic–climatic similarity, (ii) simple spatial proximity, and (iii) non-linear regression. Estimates of <inline-formula><mml:math id="M19" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow using a machine learning algorithm based on the dataset were also compared with data. All these methods are explained in next subsections.</p>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Calibration and validation of the hydrological model</title>
      <p id="d1e605">GR4J model parameters are obtained through calibration, a process of making simulated flow be as close as possible to observed flow. Table <xref ref-type="table" rid="Ch1.T1"/> summarizes the minimum and maximum values used for searching each constant of the model.</p>

<table-wrap id="Ch1.T1" specific-use="star"><label>Table 1</label><caption><p id="d1e613">Descriptions and ranges for GR4J model parameters.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="3">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="left"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Parameters</oasis:entry>
         <oasis:entry colname="col2">Description</oasis:entry>
         <oasis:entry colname="col3">Interval</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2">Production tank capacity (mm)</oasis:entry>
         <oasis:entry colname="col3">0 to 6000</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M21" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2">Coefficient of underground exchanges (mm d<inline-formula><mml:math id="M22" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">20</mml:mn></mml:mrow></mml:math></inline-formula> to 10</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M24" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2">Propagation reservoir capacity (mm)</oasis:entry>
         <oasis:entry colname="col3">0 to 4000</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M25" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2">Unit hydrograph base time (days)</oasis:entry>
         <oasis:entry colname="col3">0.04 to 20</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d1e750">The simulation period was divided into three parts: warm-up, calibration, and validation. The first 5 years of the simulation were used as a warm-up to eliminate the uncertainties in the initial conditions <xref ref-type="bibr" rid="bib1.bibx20" id="paren.34"/>. The calibration and validation periods were defined as 70 % and 30 %, respectively, of the remaining time series after the warm-up.</p>
      <p id="d1e757">The differential evolution (DE) optimization method was used for GR4J calibration. This method was initially proposed by <xref ref-type="bibr" rid="bib1.bibx57" id="text.35"/> and is part of SciPy library in Python. DE is used in optimization problems that use a single objective function, as in our case. According to <xref ref-type="bibr" rid="bib1.bibx31" id="text.36"/> and <xref ref-type="bibr" rid="bib1.bibx38" id="text.37"/>, the use of the Nash–Sutcliffe logarithmic coefficient (<inline-formula><mml:math id="M26" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE) as an objective function is more influenced by low flows; therefore, this metric can be used to evaluate the performance of minimum-flow predictions. The <inline-formula><mml:math id="M27" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE can range from <inline-formula><mml:math id="M28" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mi mathvariant="normal">∞</mml:mi></mml:mrow></mml:math></inline-formula> (poor fit) to 1.0 (perfect fit) and is calculated as follows:
            <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M29" display="block"><mml:mrow><mml:mi>log⁡</mml:mi><mml:mi mathvariant="normal">NSE</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:msup><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>ln⁡</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">sim</mml:mi></mml:msubsup><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.001</mml:mn></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:mi>ln⁡</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.001</mml:mn></mml:mrow></mml:mfenced></mml:mrow></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:mi>ln⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.001</mml:mn></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:msubsup><mml:mover accent="true"><mml:mi>Q</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi>ln⁡</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup></mml:mrow></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>
          where <inline-formula><mml:math id="M30" display="inline"><mml:mrow><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">sim</mml:mi></mml:msubsup></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M31" display="inline"><mml:mrow><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup></mml:mrow></mml:math></inline-formula> correspond to simulated and observed flow on day <inline-formula><mml:math id="M32" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>, respectively. The average term <inline-formula><mml:math id="M33" display="inline"><mml:mrow><mml:msubsup><mml:mover accent="true"><mml:mi>Q</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi>ln⁡</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup></mml:mrow></mml:math></inline-formula> is calculated by <inline-formula><mml:math id="M34" display="inline"><mml:mrow><mml:msubsup><mml:mover accent="true"><mml:mi>Q</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mi>ln⁡</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mi>ln⁡</mml:mi><mml:mo>(</mml:mo><mml:msubsup><mml:mi>Q</mml:mi><mml:mi>i</mml:mi><mml:mi mathvariant="normal">obs</mml:mi></mml:msubsup><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.001</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>.</p>
      <p id="d1e1014">Other metrics used to evaluate the performance of regionalization methods are the Pearson correlation coefficient (<inline-formula><mml:math id="M35" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>); the Nash–Sutcliffe coefficient (NSE); and the Nash–Sutcliffe square root coefficient (sqrtNSE), where flow is transformed by the square root.</p>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Regionalization methods</title>
      <p id="d1e1032">In this work, classical regionalization techniques based on physiographic–climatic similarity, simple spatial proximity, and non-linear regression were used. Regionalization based on physiographic–climatic similarity starts by identifying and grouping the watersheds that have the greatest physical, climatic, and geographic similarities. The purpose of clustering is to identify homogeneous regions based on descriptive indexes. Regionalization based on simple spatial proximity considers the fact that the study region is homogeneous and, therefore, that nearby basins are similar based on climate, relief, vegetation, landscape, and soil type. Although both assume that physical similarities can be closely correlated with hydrological responses, if the region is heterogeneous, regionalization based on physiographic–climatic similarity transfers information between basins that are not necessarily geographically neighbours. A second assumption to be considered is that the similarity between parameters from two or more river basins may reflect on the similarity of their behaviour in relation to the transformation of rainfall into flow <xref ref-type="bibr" rid="bib1.bibx41 bib1.bibx42 bib1.bibx44 bib1.bibx10" id="paren.38"/>. On the other hand, regression methods consider the fact that hydrological model parameters may be related to some physical processes that occur in watersheds and, consequently, are associated with some descriptive characteristics <xref ref-type="bibr" rid="bib1.bibx4" id="paren.39"/>. In this way, it is possible to build a regression model for each parameter of the model.</p>
      <p id="d1e1041">The diagram in Fig. <xref ref-type="fig" rid="Ch1.F4"/> briefly summarizes the application of regionalization methods in this work. After calibrating the GR4J model for each of the 126 river basins, catchments were randomly divided into training and validation sets, with 80 % of the initial sample basins comprising the training set and 20 % forming the validation set.</p>
      <p id="d1e1046">The training set is formed by river basins considered to be possible donors of GR4J parameters and which were also used to train and build regionalization models. Basins of the validation set are considered to be pseudo non-instrumentalized (indicated by the blue arrow in the diagram of Fig. <xref ref-type="fig" rid="Ch1.F4"/>), even if it is known that these catchments have hydro-meteorological data and were calibrated. Each of the regionalization methods consists basically of different methodologies for selecting donor basins for transferring the parameters of the GR4J model to target basins.</p>

      <fig id="Ch1.F4" specific-use="star"><label>Figure 4</label><caption><p id="d1e1054">Diagram summarizing the regionalization methods.</p></caption>
          <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f04.png"/>

        </fig>

<sec id="Ch1.S3.SS2.SSS1">
  <label>3.2.1</label><title>Physiographic–climatic similarity</title>
      <p id="d1e1070">When applying methods of predictions in ungauged basins (PUBs), we must take into account the uniqueness of each region across the globe and all the available information in each dataset. Bearing in mind the uniqueness of each location, possible basin descriptors were carefully chosen so that they would synthesize different characteristics of river basins and would be capable of transmitting the diversity between catchments within the same sample. Thus, the following descriptors were initially selected: basin area, length of main river, altitude and average basin slope, latitude of the basin centroid, daily averages of precipitation and potential evapotranspiration, aridity index, average number of days with extreme precipitation events and fraction of area covered by forest, and agriculture and urbanization.</p>
      <p id="d1e1073">Descriptors were normalized so that the mean and standard deviation corresponded to 0 and 1, respectively. This procedure ensures that different variables share the same scale without significant loss of information and, thus, allows categories with different magnitudes to be compared equally. Then, characteristics that showed variability, i.e. that described the set of watersheds as being heterogeneous, were selected.</p>
      <p id="d1e1076">Table <xref ref-type="table" rid="App1.Ch1.S1.T3"/> shows the descriptive statistics adopted; it reveals that Paraná basins have diverse areas, main-river lengths, and average altitudes. On the other hand, the fraction of urban area showed little variation; despite this, it was preferable to keep this descriptor since urban infrastructure, as well as other anthropogenic activities, can seriously disturb the processes of the hydrological cycle.</p>
      <p id="d1e1081">High multicollinearity between the descriptors can lead clustering algorithms to make wrong decisions during the formation of groups <xref ref-type="bibr" rid="bib1.bibx11" id="paren.40"/>. Therefore, two analyses were performed to identify the correlation between descriptors. First, the Pearson correlation (<inline-formula><mml:math id="M36" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>) between each pair of descriptors was calculated. Second, the variance inflation factor (VIF) was determined to measure the degree of multicollinearity between descriptors. VIF ranges from 1 (when there is no multicollinearity) to infinity (when there is perfect multicollinearity); the threshold used in this work was below 5. Correlations between descriptors can be seen in Fig. <xref ref-type="fig" rid="App1.Ch1.S2.F13"/>. High correlations, with <inline-formula><mml:math id="M37" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula> values above 0.70, were found between the following pairs of descriptors: aridity index and days of monthly accumulated precipitation above 150 mm, average duration of days without rain and latitude of basin centroid, average slope of the basin and fraction of forest, fraction of agricultural area and fraction of forest, and annual potential evapotranspiration and average altitude of the basin. To reduce the dimensionality of data, the following descriptors were selected: area, forest fraction, urban area fraction, average duration of extreme events with high precipitation (days of monthly accumulated precipitation above 150 mm), and average duration of days without rain.</p>
      <p id="d1e1104">The Euclidean distance (dist), calculated using Eq. (<xref ref-type="disp-formula" rid="Ch1.E2"/>) below, is a metric that can express similarities (small distances) or differences (large distances) between <inline-formula><mml:math id="M38" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula> attributes of two basins (<inline-formula><mml:math id="M39" display="inline"><mml:mi>a</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M40" display="inline"><mml:mi>b</mml:mi></mml:math></inline-formula>) in an <inline-formula><mml:math id="M41" display="inline"><mml:mi>n</mml:mi></mml:math></inline-formula>-dimensional space of attributes <xref ref-type="bibr" rid="bib1.bibx59" id="paren.41"/>.

                  <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M42" display="block"><mml:mrow><mml:mstyle displaystyle="true" class="stylechange"/><mml:mi mathvariant="normal">dist</mml:mi><mml:mo>(</mml:mo><mml:mi>a</mml:mi><mml:mo>,</mml:mo><mml:mi>b</mml:mi><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msqrt><mml:mrow><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>k</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:msup><mml:mfenced close="]" open="["><mml:mrow><mml:msub><mml:mi mathvariant="normal">atrib</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>a</mml:mi><mml:mo>)</mml:mo><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="normal">atrib</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:mi>b</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:msqrt></mml:mrow></mml:math></disp-formula>

            Clusters were produced using the <inline-formula><mml:math id="M43" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-means method, which was implemented using the scikit-learn package. The application of the <inline-formula><mml:math id="M44" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-means algorithm involves the following: first, define the number of <inline-formula><mml:math id="M45" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> groups; second, for each group, initialize a centroid randomly within the range of each category; third, assign each point to the centroid that has the smallest Euclidean distance with respect to the point; four, compute a new location of <inline-formula><mml:math id="M46" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> centroids based on the average of all points assigned to it. The iterative process from the third to the fourth step is repeated until there are no more changes in the centroids <xref ref-type="bibr" rid="bib1.bibx60" id="paren.42"/>.</p>
      <p id="d1e1237">The value of <inline-formula><mml:math id="M47" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> directly affects how groups will be formed. Increasing the number <inline-formula><mml:math id="M48" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> leads to more groups, but consequently, each group will have fewer members (which brings homogeneity but does not guarantee representativeness). On the other hand, creating fewer groups generates groups with more members (which does not allow proper identification of the different groups). Two ways to evaluate if the appropriate number of clusters resulting from the agglomeration method is using the silhouette coefficient (Si) and the elbow method.</p>
      <p id="d1e1254">According to <xref ref-type="bibr" rid="bib1.bibx52" id="text.43"/>, the silhouette coefficient (Si) consists of calculating the average Euclidean distance (<inline-formula><mml:math id="M49" display="inline"><mml:mrow><mml:msub><mml:mi>a</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>) of a point <inline-formula><mml:math id="M50" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula>, with all points belonging to the same group. Then, the average distance (<inline-formula><mml:math id="M51" display="inline"><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>) of the point <inline-formula><mml:math id="M52" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula> with respect to all the points belonging to the nearest neighbouring group is calculated. Thus, the coefficient can be determined using the following equation:
              <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M53" display="block"><mml:mrow><mml:mi mathvariant="normal">Si</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi mathvariant="normal">max</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mi mathvariant="normal">p</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
            Si can vary between <inline-formula><mml:math id="M54" display="inline"><mml:mrow><mml:mo>[</mml:mo><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>,</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>]</mml:mo></mml:mrow></mml:math></inline-formula>, and the closer to 1 it is, the more distant the point <inline-formula><mml:math id="M55" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula> is from the neighbouring group. Values close to 0 indicate that the point <inline-formula><mml:math id="M56" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula> is close to the limit that divides both groups, and measurements close to <inline-formula><mml:math id="M57" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> indicate that the point <inline-formula><mml:math id="M58" display="inline"><mml:mi>p</mml:mi></mml:math></inline-formula> may have been associated with the wrong group.</p>
      <p id="d1e1390">The elbow method is a graphical tool for evaluating an optimal number of clusters. This technique involves calculating an agglomeration coefficient; in this work, the criterion used was the sum of the squared distances of each sample (<inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>) and the respective centroid (<inline-formula><mml:math id="M60" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>) of the grouping that the sample is part of, which can be expressed as the sum of squared errors (SSE):
              <disp-formula id="Ch1.E4" content-type="numbered"><label>4</label><mml:math id="M61" display="block"><mml:mrow><mml:mi mathvariant="normal">SSE</mml:mi><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:msup><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow></mml:mfenced><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
            As number of clusters grows, distances between samples and their respective centroids decrease. However, the number of groups and the clustering coefficient are expected to be small. Thus, from a graph with the agglomeration coefficient on the <inline-formula><mml:math id="M62" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> axis and the number of groups on the <inline-formula><mml:math id="M63" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> axis, it is possible to identify the point at which there is a sharp flattening or a rapid drop in this coefficient, suggesting an optimal number of clusters <xref ref-type="bibr" rid="bib1.bibx30" id="paren.44"/>.</p>
</sec>
<sec id="Ch1.S3.SS2.SSS2">
  <label>3.2.2</label><title>Simple spatial proximity</title>
      <p id="d1e1482">The distance between two points – in this case, the centroids of target and donor basins – that have known latitudes and longitudes can be calculated using the Haversine distance, <inline-formula><mml:math id="M64" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">H</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>: 
              <disp-formula id="Ch1.E5" content-type="numbered"><label>5</label><mml:math id="M65" display="block"><mml:mrow><mml:mtable rowspacing="0.2ex" class="split" displaystyle="true" columnalign="right left"><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">H</mml:mi></mml:msub><mml:mo>=</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mi>r</mml:mi><mml:mi>arcsin⁡</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd/><mml:mtd><mml:mrow><mml:mfenced close="]" open="["><mml:msqrt><mml:mrow><mml:msup><mml:mi>sin⁡</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mfenced close=")" open="("><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:mfrac></mml:mstyle></mml:mfenced><mml:mo>+</mml:mo><mml:mi>cos⁡</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:mi>cos⁡</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:mfenced><mml:msup><mml:mi>sin⁡</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mfenced open="(" close=")"><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow><mml:mn mathvariant="normal">2</mml:mn></mml:mfrac></mml:mstyle></mml:mfenced></mml:mrow></mml:msqrt></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
            where <inline-formula><mml:math id="M66" display="inline"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mi mathvariant="normal">H</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> refers to the distance in kilometres; <inline-formula><mml:math id="M67" display="inline"><mml:mi>r</mml:mi></mml:math></inline-formula> is the average radius of the Earth (approximately 6371 km); and <inline-formula><mml:math id="M68" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M69" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> are, respectively, the latitudes and longitudes of points 1 and 2.</p>
</sec>
<sec id="Ch1.S3.SS2.SSS3">
  <label>3.2.3</label><title>Regression</title>
      <p id="d1e1642">Multiple regression models, whether linear or non-linear, seek to find the best relationship between a dependent variable and independent variables; this is done by finding the minimum error given a target. In our case, the GR4J model parameters are dependent variables which will be calculated based on descriptive characteristics of the basins (independent variables). The non-linear regression method of random forests <xref ref-type="bibr" rid="bib1.bibx12" id="paren.45"/>, which was chosen for this work, is able to perform well when dealing with large datasets and is able to distribute weights for the independent variables according to their degree of importance. Thus, two types of regression methods were constructed: random forest I and random forest II.</p>
      <p id="d1e1648">Random forest I used 1000 decision trees and was trained using basin descriptors. For this technique, it is necessary to produce a regression model independently for each parameter (<inline-formula><mml:math id="M70" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M71" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M72" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, and <inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>); however, the parameters of a hydrological model generally present dependent relationships among themselves and sometimes cannot be observed independently. Thus, a second method, defined as random forest (RF) II, included the calibrated parameters of training basins as descriptors. The second method followed the following steps: (i) a correlation analysis between GR4J parameters was performed, and, thus, an ordered list of parameters from highest to lowest correlation index was created, and (ii) a first regression was done for the parameter with the lowest correlation index – in this case, only the descriptive characteristics were used to train the model. Then, regression was performed for the parameter with the second lowest correlation index; here, we used descriptive characteristics and the previous parameter that had the lowest correlation index. This process was followed until all the parameters had their regressions.</p>
</sec>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title><inline-formula><mml:math id="M74" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow estimate</title>
      <p id="d1e1716">Instituto Água e Terra (IAT), the environmental agency responsible for legal permissions for the use of water resources in the state of Paraná, uses the river flow with 95 % permanence (<inline-formula><mml:math id="M75" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) as a reference flow rate for permission licenses for water use <xref ref-type="bibr" rid="bib1.bibx1" id="paren.46"/>. We have proposed estimating <inline-formula><mml:math id="M76" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow through regression techniques based on basin information and, thus, comparing it with <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow calculated using regionalized simulated flows. The construction of permanence curves involved (i) ordering the flows <inline-formula><mml:math id="M78" display="inline"><mml:mi>Q</mml:mi></mml:math></inline-formula> in ascending order for <inline-formula><mml:math id="M79" display="inline"><mml:mi>N</mml:mi></mml:math></inline-formula> days; (ii) assigning to each ordered flow <inline-formula><mml:math id="M80" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi>m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> the corresponding ranking order <inline-formula><mml:math id="M81" display="inline"><mml:mi>m</mml:mi></mml:math></inline-formula>; (iii) computing the frequency or probability of the ordered flows <inline-formula><mml:math id="M82" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi>m</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> to be equalled or surpassed (<inline-formula><mml:math id="M83" display="inline"><mml:mrow><mml:mi>P</mml:mi><mml:mo>(</mml:mo><mml:mi>Q</mml:mi><mml:mo>≥</mml:mo><mml:msub><mml:mi>Q</mml:mi><mml:mi>m</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>), which can be calculated using the Weibull plot position shown in the following equation <xref ref-type="bibr" rid="bib1.bibx49" id="paren.47"/>:
            <disp-formula id="Ch1.E6" content-type="numbered"><label>6</label><mml:math id="M84" display="block"><mml:mrow><mml:mi>P</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mi>Q</mml:mi><mml:mo>≥</mml:mo><mml:msub><mml:mi>Q</mml:mi><mml:mi>m</mml:mi></mml:msub></mml:mrow></mml:mfenced><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>-</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mi>m</mml:mi><mml:mrow><mml:mi>N</mml:mi><mml:mo>+</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:mfrac></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
          After obtaining <inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flows for the training set, a transformation of units (from m<inline-formula><mml:math id="M86" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">3</mml:mn></mml:msup></mml:math></inline-formula> s<inline-formula><mml:math id="M87" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> to L s<inline-formula><mml:math id="M88" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> km<inline-formula><mml:math id="M89" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>) was performed, ensuring that the variable is not dependent on basin area. Then, another random forest regression method was trained and evaluated for the test set. This RF used 1000 decision trees and watershed descriptors that presented weights greater than 0.01.</p>
      <p id="d1e1918">As terminology may sound ambiguous, here, it is important to distinguish the training and test (or validation) sets used throughout this work. There are warm-up, calibration, and validation periods for river flow simulation, and there are also training and validation sets for machine learning performance evaluation. The 126 basins were divided into training and validation sets for regionalization evaluation. Also, for estimating GR4J parameters and <inline-formula><mml:math id="M90" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flows, additional training and test (or validation) sets were created for applying random forest regressions.</p>
</sec>
</sec>
<sec id="Ch1.S4">
  <label>4</label><title>Results</title>
<sec id="Ch1.S4.SS1">
  <label>4.1</label><title>Performance of the GR4J model</title>
      <p id="d1e1948">The GR4J model showed acceptable performances for the sample of 126 watersheds, as shown in Fig. <xref ref-type="fig" rid="Ch1.F5"/>, with about 65 % of Paraná watersheds presenting <inline-formula><mml:math id="M91" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE equal to or greater than 0.70 during the validation period. Basins located close to the Paraná coastline reached a lower efficiency when compared to other regions. Some river basins presented superior performances in the validation period when compared to the calibration period; however, inverse situations also occur. These phenomena may be associated with changes or improvements in measurement techniques, as well as being influenced by changes in land use and land cover.</p>

      <fig id="Ch1.F5" specific-use="star"><label>Figure 5</label><caption><p id="d1e1962">Performance of GR4J model during calibration and validation periods.</p></caption>
          <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f05.png"/>

        </fig>

</sec>
<sec id="Ch1.S4.SS2">
  <label>4.2</label><title>Performance of regionalization methods</title>
      <p id="d1e1979">The results of each regionalization method are described below.</p>
<sec id="Ch1.S4.SS2.SSS1">
  <label>4.2.1</label><title>Physiographic–climatic similarity</title>
      <p id="d1e1989">The regionalization method by physiographic–climatic similarity starts with defining the number <inline-formula><mml:math id="M92" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> of clusters used to group the basins. The elbow method indicated that <inline-formula><mml:math id="M93" display="inline"><mml:mrow><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">6</mml:mn></mml:mrow></mml:math></inline-formula> was appropriate, which is the point of abrupt slope change or curve flattening in Fig. <xref ref-type="fig" rid="Ch1.F6"/>. Accordingly, the silhouette coefficient (Si) was higher when the number of clusters <inline-formula><mml:math id="M94" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> was equal to 6, as shown in Fig. <xref ref-type="fig" rid="Ch1.F7"/>. After defining <inline-formula><mml:math id="M95" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>, we used 80 % of the 126 watersheds for training the <inline-formula><mml:math id="M96" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-means algorithm, which was used to group similar watersheds. The remaining 20 % were used to test and evaluate the clusters formed.</p>

      <fig id="Ch1.F6"><label>Figure 6</label><caption><p id="d1e2039">Elbow method for training set basins with <inline-formula><mml:math id="M97" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> ranging from 1 to 20.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f06.png"/>

          </fig>

      <fig id="Ch1.F7"><label>Figure 7</label><caption><p id="d1e2057">Silhouette coefficients for training set basins with <inline-formula><mml:math id="M98" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula> ranging from 1 to 20.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f07.png"/>

          </fig>

      <p id="d1e2074">The geospatial distribution of watersheds and clusters formed by the <inline-formula><mml:math id="M99" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-means algorithm can be seen in Fig. <xref ref-type="fig" rid="Ch1.F8"/> for basins in the training set (Fig. <xref ref-type="fig" rid="Ch1.F8"/>a) and validation set (Fig. <xref ref-type="fig" rid="Ch1.F8"/>b). Watershed location per group in the training set was similar to the basin spatial distribution in the validation set. Additionally, geographically close basins do not always belong to the same formed group.</p>

      <fig id="Ch1.F8" specific-use="star"><label>Figure 8</label><caption><p id="d1e2092">Clusters produced by the <inline-formula><mml:math id="M100" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-mean method for the 100 training basins <bold>(a)</bold> and 26 validation basins <bold>(b)</bold>.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f08.png"/>

          </fig>

      <p id="d1e2114">Descriptor distributions for each group in the training set are shown in boxplots in Fig. <xref ref-type="fig" rid="Ch1.F9"/>. Group 4 contains basins with the largest drainage areas, located in the second and third plateaus in the centre of the state of Paraná. Group 5 contains basins that have smaller drainage areas when compared to group 4, but these end up sharing similar characteristics to catchments in group 4. Groups 2 and 3 have a higher percentage of forests, but group 3 has a greater tendency to have more rainfall, smaller areas, and shorter periods of consecutive dry days. On the other hand, group 1 stands out for containing the basins that have the longest average duration of consecutive dry days. Finally, group 6 stands out from the others because it contains basins with the highest percentages of urban area and, therefore, may be more influenced by anthropogenic activities. Although the descriptors point to heterogeneities between the formed groups, it is still possible to see overlaps, mainly in relation to the calibrated parameters of the GR4J model, as shown in Fig. <xref ref-type="fig" rid="Ch1.F10"/>.</p>

      <fig id="Ch1.F9" specific-use="star"><label>Figure 9</label><caption><p id="d1e2123">Basin descriptor distributions for training set clusters.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f09.png"/>

          </fig>

      <fig id="Ch1.F10" specific-use="star"><label>Figure 10</label><caption><p id="d1e2135">GR4J parameter distribution for basins in training set groups.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f10.png"/>

          </fig>

      <p id="d1e2144">Basins from group 2, which are located on both the first and second plateaus of the state of Paraná, are similar in size to the basins of group 1 but have a higher percentage of forest as a distinct characteristic. Basins of group 3 are found mainly in the Paraná coastal region, near Serra do Mar, a long system of mountain ranges and escarpments, where orographic rain is more likely to occur. Group 5, which is present in greater quantity, contains hydrographic basins located in all plateaus of the state.</p>
      <p id="d1e2147">Looking at parameter distributions (Fig. <xref ref-type="fig" rid="Ch1.F10"/>) from a process perspective, we can find some relations with the catchment descriptor distributions (Fig. <xref ref-type="fig" rid="Ch1.F9"/>). Parameter <inline-formula><mml:math id="M101" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> represents the runoff-producing capacity of the watershed reservoir. Our result shows that group 2, which has a higher percentage of forests, also has the highest <inline-formula><mml:math id="M102" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> median and spread. This can support the hypothesis that more forest may improve the catchment capacity for generating runoff. Parameter <inline-formula><mml:math id="M103" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> represents the base time of the instantaneous unit hydrograph. The boxplots show that larger basins (group 4) have higher <inline-formula><mml:math id="M104" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> parameters; i.e. bigger watershed areas may increase the base time of a hydrograph. Parameter <inline-formula><mml:math id="M105" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> represents the propagation reservoir capacity. Our results show that group 6, which has more urban areas, also has the smaller <inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> parameter. This reflects the effect of city impermeabilization in terms of the flow propagation capacity; i.e. after a precipitation event, the watersheds with more urban area have a smaller propagation capacity or a fast response in terms of the flow peak.</p>
</sec>
<sec id="Ch1.S4.SS2.SSS2">
  <label>4.2.2</label><title>Simple spatial proximity</title>
      <p id="d1e2229">The simple spatial proximity regionalization method considers the fact that the region near the basin of interest is homogeneous and that it therefore has hydrological similarity. Assuming this hypothesis to be true, we have used the Haversine distance between pairs of receiving basins (pseudo-non-monitored) and donor basins (instrumented basins) to transfer parameters from the GR4J model. In both methods, namely physiographic–climatic similarity and simple spatial proximity, the receiving basins are all catchments within the validation set, and the possible parameter donor basins are those from the training set that reached <inline-formula><mml:math id="M107" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE equal to or greater than 0.70 during the validation period.</p>
      <p id="d1e2239">We have allowed more than one donor basin to transfer the GR4J parameters to target basins. When there is more than one donor catchment, the four parameters of each donor basin were used to estimate flow in the pseudo-non-instrumented target basin. Once the flows were simulated with the donor basin parameters, the averages of modelled flows were calculated and used for the target basin.</p>
      <p id="d1e2242">We compared the ability of both methods (spatial proximity and physiographic–climatic similarity) to generate good results in parameter regionalization. For this, we varied the number of donor catchments from 1 to 10 and evaluated the median <inline-formula><mml:math id="M108" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE of the receiving-catchment river flow simulations in the validation period. The analysis to identify the number of donor basins, shown in Fig. <xref ref-type="fig" rid="Ch1.F11"/>, indicates that the similarity method presented a maximum median <inline-formula><mml:math id="M109" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE for a total of one basin, and the proximity method presented better results using seven basins as donors.</p>

      <fig id="Ch1.F11" specific-use="star"><label>Figure 11</label><caption><p id="d1e2265">Evaluating the optimal number of donor basins based on <inline-formula><mml:math id="M110" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE medians during the validation period of receiving-basin simulations. The green and blue lines represent spatial proximity and physiographic–climatic similarity regionalization methods, respectively.</p></caption>
            <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f11.png"/>

          </fig>

</sec>
<sec id="Ch1.S4.SS2.SSS3">
  <label>4.2.3</label><title>Regression of GR4J parameters</title>
      <p id="d1e2289">To train the random forest regression model, known information about the watersheds in the training set was used, namely the descriptive characteristics (independent variables) and the calibrated parameters of the GR4J model (dependent variable). The descriptive characteristics that the random forest pointed out to be most relevant were the slope of the main river and the average slope of the basin for parameter <inline-formula><mml:math id="M111" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">1</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, the average altitude and the average radiation in winter for parameter <inline-formula><mml:math id="M112" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, the average radiation in winter for parameter <inline-formula><mml:math id="M113" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">3</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, and the fraction of Gleissol for parameter <inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:msub><mml:mi>X</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>. These reinforce that machine learning algorithms perform better with physiographic–climatic indices as inputs.</p>
</sec>
</sec>
<sec id="Ch1.S4.SS3">
  <label>4.3</label><title><inline-formula><mml:math id="M115" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow estimation</title>
      <p id="d1e2356">In order to compare the performance of estimating the <inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flow between direct (regression) and indirect (regionalization of parameters) techniques, the regression method of random forest, which was named random forest <inline-formula><mml:math id="M117" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>, was applied to directly regionalize <inline-formula><mml:math id="M118" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flow.</p>
      <p id="d1e2392">The construction of the random forest <inline-formula><mml:math id="M119" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> regression model used known information about the watersheds in the training set, namely the watershed descriptor (independent variables) and the <inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> (in L s<inline-formula><mml:math id="M121" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> km<inline-formula><mml:math id="M122" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>), estimated from the observed historical series (dependent variable). The most relevant characteristics identified by the random forest <inline-formula><mml:math id="M123" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> method were days of precipitation with monthly accumulation of 150 mm, basin centroid longitude, basin average slope, pasture fraction, and forest fraction. <inline-formula><mml:math id="M124" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> was calculated, for both observed and simulated flows, using calibration and validation periods. Thus, at least 15 years of fluviometric records were used to estimate the reference flow.</p>
      <p id="d1e2464">Correlations between observed <inline-formula><mml:math id="M125" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flows and those predicted by calibration and regionalization methods were calculated. Regionalizations with the highest performances were obtained by the physiographic–climatic similarity method, with a correlation (<inline-formula><mml:math id="M126" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>) of 0.973, and then the random forest <inline-formula><mml:math id="M127" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> method, with a correlation of 0.965. <inline-formula><mml:math id="M128" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> flows predicted by calibration of the GR4J model had a correlation of 0.9956, the regionalization method based on proximity had a correlation of 0.9386, and the random forest method reached a correlation equal to 0.9392.</p>
</sec>
</sec>
<sec id="Ch1.S5">
  <label>5</label><title>Discussion</title>
      <p id="d1e2517">In dry periods, the flows of rivers in Paraná are sustained basically by two mechanisms: baseflow and groundwater recharge. Even in periods with no precipitation, there can be movements of water from underground aquifers and saturated soil layers into surface waterbodies, such as rivers, lakes, or wetlands. All basins studied in this research drain into the Paraná River, beneath which there resides the Guarani aquifer, one of the largest sandstone aquifers in the world <xref ref-type="bibr" rid="bib1.bibx25" id="paren.48"/>. Karst terrains are also widespread throughout the Paraná basin, with the Açungui karst and non-carbonate karsts being the most important ones <xref ref-type="bibr" rid="bib1.bibx5 bib1.bibx58" id="paren.49"/>. The Açungui carbonate karst is characterized by large areas of horizontally bedded limestones and dolomites, which form extensive regions of little or no relief and are drained by low-gradient rivers <xref ref-type="bibr" rid="bib1.bibx5" id="paren.50"/>. Interbasin groundwater flow may also play an important role in the water balance during dry periods in karst catchments <xref ref-type="bibr" rid="bib1.bibx58" id="paren.51"/>.</p>
      <p id="d1e2532"><xref ref-type="bibr" rid="bib1.bibx8" id="text.52"/> identified that the rainfall season occurs in the months of December, January, and February in the south of Brazil. The basins under study in the state of Paraná are in a region of climatic transition, with reasonably well-distributed rainfall throughout the year. The region's seasonality is generally divided between the 6 months centred around summer, from October to March, which correspond to the wet period, and the remaining months, from April to September, which correspond to the dry period. However, the occurrence of cold fronts, low-pressure areas, and instability systems during the Brazilian winter can provoke large floods – even though this is the dry period – and interrupt the recession process of the hydrographs.</p>
      <p id="d1e2537">In Appendix <xref ref-type="sec" rid="App1.Ch1.S3"/>, we show the hydrographs by physiographic–climatic similarity group. The comparison of hydrographs separated by groups of similar watersheds show the seasonality and strength of smaller flow rates. In these climatic conditions, the predominance of low flows is expected from April to September. The slow release of groundwater volumes after the cessation of surface runoff causes a recession curve that is strongly influenced by river–aquifer interaction. This curve, which conceptual models try to capture through simple mathematical relationships, is influenced by various factors, namely soil properties, hydraulic characteristics and the extent of aquifers, the rate and amount of groundwater recharge, evaporation and evapotranspiration of the basin, and the spatial distribution of vegetation cover, among others <xref ref-type="bibr" rid="bib1.bibx39" id="paren.53"/>.</p>
      <p id="d1e2545">Due to the complexity of hydrological processes and the specificities existing in the river basins, representing the recession curve and simulating low flows using conceptual models is an arduous process, sometimes requiring a basin-by-basin hydrological analysis. Attempts to improve this representation in the design of hydrological models have resulted in an increase in parameters, as is the case of traditional Sacramento Soil Moisture Accounting (SAC-SMA) model, which uses two conceptual reservoirs to simulate low flows with an overlay effect that allows us to better capture low-flow variability for a wider range of river basins <xref ref-type="bibr" rid="bib1.bibx14" id="paren.54"/>. The model applied in our study has an improved version, the GR6J, dedicated to low flows, which uses two additional parameters to better represent exchanges between the river and groundwater <xref ref-type="bibr" rid="bib1.bibx50" id="paren.55"/>. In both cases, a better representation of low flows is achieved at the cost of increased model degrees of freedom, which is not ideal for regionalization issues.</p>
      <p id="d1e2555">In general, the vast majority of basins from the validation group presented results with <inline-formula><mml:math id="M129" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE greater than 0.50 for different regionalization methods, and only two basins within this group presented low performance. Another general behaviour was that basins with <inline-formula><mml:math id="M130" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE equal to or greater than 0.77 using calibrated parameters also achieved comparable performances with regionalized parameters. Additionally, in some cases where regionalization methods used more than one donor basin, they provided a diversified set of parameters. When combining this set of parameters, GR4J with the average of each parameter can result in superior performance compared to the use of calibrated parameters in a period prior to validation.</p>
      <p id="d1e2572">Results were evaluated using the Pearson correlation coefficient (<inline-formula><mml:math id="M131" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>), the Nash–Sutcliffe coefficient (NSE), and their variations: the flow transformed by the square root (sqrtNSE) and by the logarithm (<inline-formula><mml:math id="M132" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE). NSE gives more emphasis to the performance of higher flows, <inline-formula><mml:math id="M133" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE is more sensitive to low flows, and sqrtNSE provides an intermediate performance <xref ref-type="bibr" rid="bib1.bibx41" id="paren.56"/>.  Table <xref ref-type="table" rid="Ch1.T2"/> shows the median values of error statistics for estimated flows in the validation period of the 26 watersheds from the validation set. Among the regionalization methods, values that achieved the best results for each index are highlighted in bold. Thus, the physiographic–climatic similarity stands out positively by reaching <inline-formula><mml:math id="M134" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE and sqrtNSE equal to 0.736 and 0.726, respectively. Another point to be highlighted is that the spatial proximity method presents, in general, median results for the three coefficients. Table <xref ref-type="table" rid="Ch1.T2"/> also reveals that regionalization method performances – in particular for proximity and random forest – can reach median NSE values equal to or greater than when parameters were calibrated in a period prior to validation.</p>

<table-wrap id="Ch1.T2"><label>Table 2</label><caption><p id="d1e2614">Median values of error statistics calculated for validation period. In bold are the best results for each index among the regionalization methods.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="5">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="center"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:colspec colnum="5" colname="col5" align="center"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col5" align="center">Efficiency metrics in the validation period </oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2">Calibrated</oasis:entry>
         <oasis:entry colname="col3">Proximity</oasis:entry>
         <oasis:entry colname="col4">Similarity</oasis:entry>
         <oasis:entry colname="col5">Random</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4"/>
         <oasis:entry colname="col5">forest</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">NSE</oasis:entry>
         <oasis:entry colname="col2">0.621</oasis:entry>
         <oasis:entry colname="col3">0.635</oasis:entry>
         <oasis:entry colname="col4">0.602</oasis:entry>
         <oasis:entry colname="col5"><bold>0.643</bold></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M135" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE</oasis:entry>
         <oasis:entry colname="col2">0.758</oasis:entry>
         <oasis:entry colname="col3">0.702</oasis:entry>
         <oasis:entry colname="col4"><bold>0.736</bold></oasis:entry>
         <oasis:entry colname="col5">0.679</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">sqrtNSE</oasis:entry>
         <oasis:entry colname="col2">0.736</oasis:entry>
         <oasis:entry colname="col3">0.707</oasis:entry>
         <oasis:entry colname="col4"><bold>0.726</bold></oasis:entry>
         <oasis:entry colname="col5">0.713</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d1e2739">The review carried out by <xref ref-type="bibr" rid="bib1.bibx22" id="text.57"/> included the analysis of articles from different regions of the globe, which were recently published between 2013 and 2019 and in which the researchers also applied similar regionalization techniques (proximity, similarity, and regression). The same authors show evidence that regionalization methods based on distances (proximity and similarity) generally present superior performances compared to methods based on regression. The study carried out by <xref ref-type="bibr" rid="bib1.bibx32" id="text.58"/> in Europe explored the correlation between 16 different indices of hydrological-behaviour responses (e.g. baseline flow index, <inline-formula><mml:math id="M136" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">5</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M137" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) and 35 physical descriptors (e.g. area, slope, and aridity index), concluding that there are strong connections between the physical descriptors and the response rates of the hydrological behaviour.</p>
      <p id="d1e2770"><xref ref-type="bibr" rid="bib1.bibx37" id="text.59"/> explain that, due to the non-linear and multidimensional relationship between basin descriptive characteristics and model parameters, the application of regression methods that use machine learning techniques are becoming more common to extrapolate hydrological model parameters. In our study, random forest methods performed better when the average radiation in winter and the days of precipitation with monthly accumulation above 150 mm were used to inform the algorithm, which reveals key variables required for understanding regionalization techniques in humid subtropical and hot temperate climates.</p>
</sec>
<sec id="Ch1.S6" sec-type="conclusions">
  <label>6</label><title>Conclusions</title>
      <p id="d1e2783">In this study, three regionalization methods were developed and deployed with the purpose of estimating daily flows in basins of the state of Paraná. A set of hydrometeorological data was created and presented together with catchment descriptive indexes. The amount of collected data is greater than that of national-level datasets for the region since a higher density of fluviometric stations was used. GR4J was employed for 126 watersheds and achieved optimistic performances in the validation period (<inline-formula><mml:math id="M138" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE <inline-formula><mml:math id="M139" display="inline"><mml:mo>≥</mml:mo></mml:math></inline-formula> 0.70) for 65 % of the watersheds.</p>
      <p id="d1e2800">All regionalization methods showed positive performances. Median values of <inline-formula><mml:math id="M140" display="inline"><mml:mi>log⁡</mml:mi></mml:math></inline-formula>NSE in regionalizations were equal to 0.702, 0.736, and 0.679 for spatial proximity, physiographic–climatic similarity, and random forest methods, respectively. When comparing the median NSE between the three methods, random forest is slightly better. However, the median sqrtNSE was higher for the physiographic–climatic similarity method. The regionalization based on physiographic–climatic similarity proved to be the most robust method for predicting daily flow and <inline-formula><mml:math id="M141" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> reference flow. When increasing the number of donor basins, the method based on spatial proximity has comparable performance to the method based on physiographic–climatic similarity.</p>
      <p id="d1e2821">Based on the physiographic–climatic characteristics of the basins, it was possible to classify six distinct groups of watersheds in Paraná. Basins within each group showed similarities in their size, urban-area fraction, average duration of consecutive dry days, number of days with more than 150 mm of precipitation, and forest fraction. Interestingly, the last two descriptors were also relevant for the random forest <inline-formula><mml:math id="M142" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">95</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> model. The use of machine learning algorithms to regionalize streamflow had good performance using climatic and physiographic indices as inputs. This research represents a proof of concept that basins without flow monitoring can have a good approximation of streamflow if physiographic–climatic information is provided.</p>
      <p id="d1e2835">Our regionalization study showed that parameters are sensitive to basin physiographic characteristics and soil use, and this has a direct effect on the streamflow response, i.e. hydrograph peak time, hydrograph base time, production capacity, and propagation capacity. Urban impermeable areas produce a fast response in terms of the flow peak. Forests play a significant role in groundwater recharge and low-flow generation through various mechanisms: interception and slowing infiltration, enhancing soil structure and porosity, and reducing erosion through root system soil stabilization. Overall, forests act as natural sponges, slowing down the movement of water, enhancing infiltration, and promoting groundwater recharge. Protecting and maintaining forest ecosystems is essential for sustaining groundwater resources and ensuring water availability for both human and natural systems.</p>
      <p id="d1e2839">We recommend for future studies the use of stochastic optimization techniques for model calibration and the use of different hydrological models for parameter regionalizations. In addition, we suggest the estimation of confidence intervals for the regionalized parameters and the use of regionalization methods based on geostatistical techniques. Another recommendation is to include flow seasonality indices <xref ref-type="bibr" rid="bib1.bibx13 bib1.bibx45" id="paren.60"/> as descriptors to better characterize the physiographic–climatic similarity of the basins.</p>
</sec>

      
      </body>
    <back><app-group>

<app id="App1.Ch1.S1">
  <label>Appendix A</label><title>Data availability and physiographic–climatic indices</title>
      <p id="d1e2856">Figure <xref ref-type="fig" rid="App1.Ch1.S1.F12"/> shows the availability of data over the years; the darker the shade of green, the more data. Watershed descriptive characteristics are shown in Table <xref ref-type="table" rid="App1.Ch1.S1.T3"/>. Note that the region can be classified as having a humid subtropical climate <xref ref-type="bibr" rid="bib1.bibx34" id="paren.61"/>.</p><table-wrap id="App1.Ch1.S1.T3"><label>Table A1</label><caption><p id="d1e2870">Descriptive statistics for the Paraná dataset.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="8">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:colspec colnum="7" colname="col7" align="right"/>
     <oasis:colspec colnum="8" colname="col8" align="right"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Descriptors</oasis:entry>
         <oasis:entry colname="col2">Mean</oasis:entry>
         <oasis:entry colname="col3">Standard</oasis:entry>
         <oasis:entry colname="col4">Min</oasis:entry>
         <oasis:entry colname="col5">25 %</oasis:entry>
         <oasis:entry colname="col6">50 %</oasis:entry>
         <oasis:entry colname="col7">75 %</oasis:entry>
         <oasis:entry colname="col8">Max</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3">deviation</oasis:entry>
         <oasis:entry colname="col4"/>
         <oasis:entry colname="col5"/>
         <oasis:entry colname="col6"/>
         <oasis:entry colname="col7"/>
         <oasis:entry colname="col8"/>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Physiographic indices </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Mean altitude of the centroid (m)</oasis:entry>
         <oasis:entry colname="col2">743.16</oasis:entry>
         <oasis:entry colname="col3">210.64</oasis:entry>
         <oasis:entry colname="col4">62.00</oasis:entry>
         <oasis:entry colname="col5">611.25</oasis:entry>
         <oasis:entry colname="col6">773.50</oasis:entry>
         <oasis:entry colname="col7">892.25</oasis:entry>
         <oasis:entry colname="col8">1132.00</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Area (km<inline-formula><mml:math id="M143" display="inline"><mml:msup><mml:mi/><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">4474.15</oasis:entry>
         <oasis:entry colname="col3">7001.22</oasis:entry>
         <oasis:entry colname="col4">13.87</oasis:entry>
         <oasis:entry colname="col5">510.70</oasis:entry>
         <oasis:entry colname="col6">1523.48</oasis:entry>
         <oasis:entry colname="col7">4120.90</oasis:entry>
         <oasis:entry colname="col8">34 440.18</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average height of the basin (m)</oasis:entry>
         <oasis:entry colname="col2">804.33</oasis:entry>
         <oasis:entry colname="col3">171.96</oasis:entry>
         <oasis:entry colname="col4">262.85</oasis:entry>
         <oasis:entry colname="col5">666.36</oasis:entry>
         <oasis:entry colname="col6">835.74</oasis:entry>
         <oasis:entry colname="col7">920.45</oasis:entry>
         <oasis:entry colname="col8">1150.88</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average slope of the basin (m m<inline-formula><mml:math id="M144" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">0.16</oasis:entry>
         <oasis:entry colname="col3">0.06</oasis:entry>
         <oasis:entry colname="col4">0.06</oasis:entry>
         <oasis:entry colname="col5">0.12</oasis:entry>
         <oasis:entry colname="col6">0.14</oasis:entry>
         <oasis:entry colname="col7">0.18</oasis:entry>
         <oasis:entry colname="col8">0.33</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Strahler number</oasis:entry>
         <oasis:entry colname="col2">6.69</oasis:entry>
         <oasis:entry colname="col3">1.29</oasis:entry>
         <oasis:entry colname="col4">3.00</oasis:entry>
         <oasis:entry colname="col5">6.00</oasis:entry>
         <oasis:entry colname="col6">7.00</oasis:entry>
         <oasis:entry colname="col7">7.00</oasis:entry>
         <oasis:entry colname="col8">9.00</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Main river length (m)</oasis:entry>
         <oasis:entry colname="col2">185 335.43</oasis:entry>
         <oasis:entry colname="col3">164 551.04</oasis:entry>
         <oasis:entry colname="col4">7336.06</oasis:entry>
         <oasis:entry colname="col5">66 867.87</oasis:entry>
         <oasis:entry colname="col6">122 924.13</oasis:entry>
         <oasis:entry colname="col7">236 343.96</oasis:entry>
         <oasis:entry colname="col8">748 033.70</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Drainage density (km km<inline-formula><mml:math id="M145" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">2.55</oasis:entry>
         <oasis:entry colname="col3">1.04</oasis:entry>
         <oasis:entry colname="col4">0.72</oasis:entry>
         <oasis:entry colname="col5">1.81</oasis:entry>
         <oasis:entry colname="col6">2.33</oasis:entry>
         <oasis:entry colname="col7">3.28</oasis:entry>
         <oasis:entry colname="col8">5.59</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Main river slope (m m<inline-formula><mml:math id="M146" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">0.04</oasis:entry>
         <oasis:entry colname="col3">0.02</oasis:entry>
         <oasis:entry colname="col4">0.02</oasis:entry>
         <oasis:entry colname="col5">0.03</oasis:entry>
         <oasis:entry colname="col6">0.03</oasis:entry>
         <oasis:entry colname="col7">0.04</oasis:entry>
         <oasis:entry colname="col8">0.10</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Climatological indices </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Coefficient of variation of annual precipitation</oasis:entry>
         <oasis:entry colname="col2">0.17</oasis:entry>
         <oasis:entry colname="col3">0.02</oasis:entry>
         <oasis:entry colname="col4">0.13</oasis:entry>
         <oasis:entry colname="col5">0.16</oasis:entry>
         <oasis:entry colname="col6">0.17</oasis:entry>
         <oasis:entry colname="col7">0.18</oasis:entry>
         <oasis:entry colname="col8">0.21</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">July average temperature (°C)</oasis:entry>
         <oasis:entry colname="col2">15.55</oasis:entry>
         <oasis:entry colname="col3">0.58</oasis:entry>
         <oasis:entry colname="col4">14.56</oasis:entry>
         <oasis:entry colname="col5">15.16</oasis:entry>
         <oasis:entry colname="col6">15.42</oasis:entry>
         <oasis:entry colname="col7">15.98</oasis:entry>
         <oasis:entry colname="col8">17.02</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">January average temperature (°C)</oasis:entry>
         <oasis:entry colname="col2">22.96</oasis:entry>
         <oasis:entry colname="col3">0.46</oasis:entry>
         <oasis:entry colname="col4">22.26</oasis:entry>
         <oasis:entry colname="col5">22.56</oasis:entry>
         <oasis:entry colname="col6">22.89</oasis:entry>
         <oasis:entry colname="col7">23.31</oasis:entry>
         <oasis:entry colname="col8">24.07</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Precipitation days with monthly accumulation of 10 mm</oasis:entry>
         <oasis:entry colname="col2">152.86</oasis:entry>
         <oasis:entry colname="col3">17.45</oasis:entry>
         <oasis:entry colname="col4">112.79</oasis:entry>
         <oasis:entry colname="col5">142.02</oasis:entry>
         <oasis:entry colname="col6">151.56</oasis:entry>
         <oasis:entry colname="col7">162.55</oasis:entry>
         <oasis:entry colname="col8">213.14</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Precipitation days with monthly accumulation of 50 mm</oasis:entry>
         <oasis:entry colname="col2">146.07</oasis:entry>
         <oasis:entry colname="col3">17.55</oasis:entry>
         <oasis:entry colname="col4">104.38</oasis:entry>
         <oasis:entry colname="col5">133.40</oasis:entry>
         <oasis:entry colname="col6">145.44</oasis:entry>
         <oasis:entry colname="col7">156.40</oasis:entry>
         <oasis:entry colname="col8">208.45</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Precipitation days with monthly accumulation of 150 mm</oasis:entry>
         <oasis:entry colname="col2">82.84</oasis:entry>
         <oasis:entry colname="col3">17.43</oasis:entry>
         <oasis:entry colname="col4">51.38</oasis:entry>
         <oasis:entry colname="col5">70.54</oasis:entry>
         <oasis:entry colname="col6">81.52</oasis:entry>
         <oasis:entry colname="col7">93.96</oasis:entry>
         <oasis:entry colname="col8">159.12</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Annual potential evapotranspiration (mm)</oasis:entry>
         <oasis:entry colname="col2">1255.95</oasis:entry>
         <oasis:entry colname="col3">78.02</oasis:entry>
         <oasis:entry colname="col4">1139.88</oasis:entry>
         <oasis:entry colname="col5">1192.19</oasis:entry>
         <oasis:entry colname="col6">1243.22</oasis:entry>
         <oasis:entry colname="col7">1326.41</oasis:entry>
         <oasis:entry colname="col8">1423.14</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average annual precipitation (mm)</oasis:entry>
         <oasis:entry colname="col2">1678.61</oasis:entry>
         <oasis:entry colname="col3">216.08</oasis:entry>
         <oasis:entry colname="col4">1357.26</oasis:entry>
         <oasis:entry colname="col5">1511.72</oasis:entry>
         <oasis:entry colname="col6">1614.79</oasis:entry>
         <oasis:entry colname="col7">1828.01</oasis:entry>
         <oasis:entry colname="col8">2618.98</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average solar radiation in winter months (kwh m<inline-formula><mml:math id="M147" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">3.39</oasis:entry>
         <oasis:entry colname="col3">0.15</oasis:entry>
         <oasis:entry colname="col4">3.13</oasis:entry>
         <oasis:entry colname="col5">3.27</oasis:entry>
         <oasis:entry colname="col6">3.39</oasis:entry>
         <oasis:entry colname="col7">3.50</oasis:entry>
         <oasis:entry colname="col8">3.70</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average solar radiation in  summer months (kwh m<inline-formula><mml:math id="M148" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">5.53</oasis:entry>
         <oasis:entry colname="col3">0.19</oasis:entry>
         <oasis:entry colname="col4">5.16</oasis:entry>
         <oasis:entry colname="col5">5.37</oasis:entry>
         <oasis:entry colname="col6">5.54</oasis:entry>
         <oasis:entry colname="col7">5.71</oasis:entry>
         <oasis:entry colname="col8">5.86</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Aridity index</oasis:entry>
         <oasis:entry colname="col2">1.34</oasis:entry>
         <oasis:entry colname="col3">0.18</oasis:entry>
         <oasis:entry colname="col4">0.97</oasis:entry>
         <oasis:entry colname="col5">1.24</oasis:entry>
         <oasis:entry colname="col6">1.30</oasis:entry>
         <oasis:entry colname="col7">1.45</oasis:entry>
         <oasis:entry colname="col8">2.02</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Average daily precipitation (mm d<inline-formula><mml:math id="M149" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">4.60</oasis:entry>
         <oasis:entry colname="col3">0.59</oasis:entry>
         <oasis:entry colname="col4">3.72</oasis:entry>
         <oasis:entry colname="col5">4.14</oasis:entry>
         <oasis:entry colname="col6">4.42</oasis:entry>
         <oasis:entry colname="col7">5.00</oasis:entry>
         <oasis:entry colname="col8">7.17</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Frequency of days without rain (d yr<inline-formula><mml:math id="M150" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>)</oasis:entry>
         <oasis:entry colname="col2">208.28</oasis:entry>
         <oasis:entry colname="col3">17.94</oasis:entry>
         <oasis:entry colname="col4">148.24</oasis:entry>
         <oasis:entry colname="col5">197.65</oasis:entry>
         <oasis:entry colname="col6">209.87</oasis:entry>
         <oasis:entry colname="col7">219.21</oasis:entry>
         <oasis:entry colname="col8">249.17</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Average length of days without rain (d)</oasis:entry>
         <oasis:entry colname="col2">4.54</oasis:entry>
         <oasis:entry colname="col3">0.43</oasis:entry>
         <oasis:entry colname="col4">3.41</oasis:entry>
         <oasis:entry colname="col5">4.29</oasis:entry>
         <oasis:entry colname="col6">4.57</oasis:entry>
         <oasis:entry colname="col7">4.78</oasis:entry>
         <oasis:entry colname="col8">5.71</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Land use and land cover </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(1) Forest (%)</oasis:entry>
         <oasis:entry colname="col2">0.48</oasis:entry>
         <oasis:entry colname="col3">0.25</oasis:entry>
         <oasis:entry colname="col4">0.06</oasis:entry>
         <oasis:entry colname="col5">0.29</oasis:entry>
         <oasis:entry colname="col6">0.41</oasis:entry>
         <oasis:entry colname="col7">0.71</oasis:entry>
         <oasis:entry colname="col8">0.98</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(2) Agriculture (%)</oasis:entry>
         <oasis:entry colname="col2">0.27</oasis:entry>
         <oasis:entry colname="col3">0.21</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.07</oasis:entry>
         <oasis:entry colname="col6">0.26</oasis:entry>
         <oasis:entry colname="col7">0.42</oasis:entry>
         <oasis:entry colname="col8">0.79</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(3) Urban area (%)</oasis:entry>
         <oasis:entry colname="col2">0.03</oasis:entry>
         <oasis:entry colname="col3">0.07</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.01</oasis:entry>
         <oasis:entry colname="col7">0.02</oasis:entry>
         <oasis:entry colname="col8">0.53</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(4) Exposed soil (%)</oasis:entry>
         <oasis:entry colname="col2">0.00</oasis:entry>
         <oasis:entry colname="col3">0.00</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.00</oasis:entry>
         <oasis:entry colname="col8">0.01</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(5) Pasture (%)</oasis:entry>
         <oasis:entry colname="col2">0.23</oasis:entry>
         <oasis:entry colname="col3">0.13</oasis:entry>
         <oasis:entry colname="col4">0.01</oasis:entry>
         <oasis:entry colname="col5">0.14</oasis:entry>
         <oasis:entry colname="col6">0.19</oasis:entry>
         <oasis:entry colname="col7">0.29</oasis:entry>
         <oasis:entry colname="col8">0.80</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(6) Water (%)</oasis:entry>
         <oasis:entry colname="col2">0.00</oasis:entry>
         <oasis:entry colname="col3">0.01</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.00</oasis:entry>
         <oasis:entry colname="col8">0.08</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Curve number</oasis:entry>
         <oasis:entry colname="col2">77.37</oasis:entry>
         <oasis:entry colname="col3">4.68</oasis:entry>
         <oasis:entry colname="col4">57.93</oasis:entry>
         <oasis:entry colname="col5">75.90</oasis:entry>
         <oasis:entry colname="col6">77.88</oasis:entry>
         <oasis:entry colname="col7">80.39</oasis:entry>
         <oasis:entry colname="col8">87.91</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Soil type </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(1) Latosol (%)</oasis:entry>
         <oasis:entry colname="col2">0.24</oasis:entry>
         <oasis:entry colname="col3">0.18</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.08</oasis:entry>
         <oasis:entry colname="col6">0.25</oasis:entry>
         <oasis:entry colname="col7">0.35</oasis:entry>
         <oasis:entry colname="col8">0.78</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(2) Neosol (%)</oasis:entry>
         <oasis:entry colname="col2">0.20</oasis:entry>
         <oasis:entry colname="col3">0.18</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.02</oasis:entry>
         <oasis:entry colname="col6">0.15</oasis:entry>
         <oasis:entry colname="col7">0.33</oasis:entry>
         <oasis:entry colname="col8">0.70</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(3) Argisol (%)</oasis:entry>
         <oasis:entry colname="col2">0.14</oasis:entry>
         <oasis:entry colname="col3">0.17</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.11</oasis:entry>
         <oasis:entry colname="col7">0.21</oasis:entry>
         <oasis:entry colname="col8">0.91</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(4) Nitosol (%)</oasis:entry>
         <oasis:entry colname="col2">0.10</oasis:entry>
         <oasis:entry colname="col3">0.13</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.04</oasis:entry>
         <oasis:entry colname="col7">0.15</oasis:entry>
         <oasis:entry colname="col8">0.65</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(5) Cambisol (%)</oasis:entry>
         <oasis:entry colname="col2">0.20</oasis:entry>
         <oasis:entry colname="col3">0.23</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.02</oasis:entry>
         <oasis:entry colname="col6">0.11</oasis:entry>
         <oasis:entry colname="col7">0.31</oasis:entry>
         <oasis:entry colname="col8">0.95</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(6) Gleissol (%)</oasis:entry>
         <oasis:entry colname="col2">0.01</oasis:entry>
         <oasis:entry colname="col3">0.03</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.02</oasis:entry>
         <oasis:entry colname="col8">0.21</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(7) Organosol (%)</oasis:entry>
         <oasis:entry colname="col2">0.01</oasis:entry>
         <oasis:entry colname="col3">0.03</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.00</oasis:entry>
         <oasis:entry colname="col8">0.18</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(8) Spodosol (%)</oasis:entry>
         <oasis:entry colname="col2">0.00</oasis:entry>
         <oasis:entry colname="col3">0.00</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.00</oasis:entry>
         <oasis:entry colname="col8">0.00</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(9) Rocky outcrop (%)</oasis:entry>
         <oasis:entry colname="col2">0.02</oasis:entry>
         <oasis:entry colname="col3">0.04</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.01</oasis:entry>
         <oasis:entry colname="col8">0.27</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">(10) Urban area (%)</oasis:entry>
         <oasis:entry colname="col2">0.01</oasis:entry>
         <oasis:entry colname="col3">0.05</oasis:entry>
         <oasis:entry colname="col4">0.00</oasis:entry>
         <oasis:entry colname="col5">0.00</oasis:entry>
         <oasis:entry colname="col6">0.00</oasis:entry>
         <oasis:entry colname="col7">0.00</oasis:entry>
         <oasis:entry colname="col8">0.43</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

<fig id="App1.Ch1.S1.F12"><label>Figure A1</label><caption><p id="d1e4170">Availability of flow data by station. The darker the green colour is, the more data are available for that year.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f12.png"/>

      </fig>


</app>

<app id="App1.Ch1.S2">
  <label>Appendix B</label><title>Comparison of descriptors</title>
      <p id="d1e4191">Figure <xref ref-type="fig" rid="App1.Ch1.S2.F13"/> shows the Pearson correlation coefficients between watershed descriptors.</p>

      <fig id="App1.Ch1.S2.F13"><label>Figure B1</label><caption><p id="d1e4198">Pearson correlation coefficients between descriptors.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f13.png"/>

      </fig>

</app>

<app id="App1.Ch1.S3">
  <label>Appendix C</label><title>Hydrographs by physiographic–climatic similarity group</title>
      <p id="d1e4217">Below, we show the hydrographs separated by groups of watersheds with physiographic–climatic similarity. The hydrographs were produced based on flow records observed from 2009 to 2016, and the <inline-formula><mml:math id="M151" display="inline"><mml:mi>y</mml:mi></mml:math></inline-formula> axis was limited to 400 L s<inline-formula><mml:math id="M152" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> km<inline-formula><mml:math id="M153" display="inline"><mml:msup><mml:mi/><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> for comparison. Figures <xref ref-type="fig" rid="App1.Ch1.S3.F14"/>–<xref ref-type="fig" rid="App1.Ch1.S3.F19"/> show hydrographs from the catchments of groups 1, 2, 3, 4, 5, and 6, respectively.</p><fig id="App1.Ch1.S3.F14"><label>Figure C1</label><caption><p id="d1e4257">Hydrographs from basins of group 1.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f14.png"/>

      </fig>

      <fig id="App1.Ch1.S3.F15"><label>Figure C2</label><caption><p id="d1e4270">Hydrographs from basins of group 2.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f15.png"/>

      </fig>

      <fig id="App1.Ch1.S3.F16"><label>Figure C3</label><caption><p id="d1e4284">Hydrographs from basins of group 3.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f16.png"/>

      </fig>

<fig id="App1.Ch1.S3.F17"><label>Figure C4</label><caption><p id="d1e4298">Hydrographs from basins of group 4.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f17.png"/>

      </fig>

      <fig id="App1.Ch1.S3.F18"><label>Figure C5</label><caption><p id="d1e4311">Hydrographs from basins of group 5.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f18.png"/>

      </fig>

      <fig id="App1.Ch1.S3.F19"><label>Figure C6</label><caption><p id="d1e4324">Hydrographs from basins of group 6.</p></caption>
        
        <graphic xlink:href="https://hess.copernicus.org/articles/28/3367/2024/hess-28-3367-2024-f19.png"/>

      </fig>


</app>
  </app-group><notes notes-type="codeavailability"><title>Code availability</title>

      <p id="d1e4341">Available upon request from the authors.</p>
  </notes><notes notes-type="dataavailability"><title>Data availability</title>

      <p id="d1e4347">The dataset was obtained from Agência Nacional de Águas e Saneamento Básico (ANA), Instituto Nacional de Meteorologia (INMET), Sistema de Tecnologia e Monitoramento Ambiental do Paraná (Simepar), Água e Terra Institute (IAT), and Instituto de Desenvolvimento Rural do Paraná (IAPAR-EMATER).</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d1e4353">LAK and EGFM conceived and designed the study, performed the experiments, analysed the data, conducted the statistical analyses, and wrote the paper. ASA and SMF contributed to study design, critically reviewed the paper, and provided valuable feedback.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d1e4359">The contact author has declared that none of the authors has any competing interests.</p>
  </notes><notes notes-type="disclaimer"><title>Disclaimer</title>

      <p id="d1e4365">Publisher's note: Copernicus Publications remains neutral with regard to jurisdictional claims made in the text, published maps, institutional affiliations, or any other geographical representation in this paper. While Copernicus Publications makes every effort to include appropriate place names, the final responsibility lies with the authors.</p>
  </notes><ack><title>Acknowledgements</title><p id="d1e4371">This study was partially financed by the Estonian Research Council (grant no. PRG1674) and the European Union's Horizon 492 2020 Research and Innovation Programme ACTRIS IMP (grant no. 871115). The authors are grateful to the Erasmus<inline-formula><mml:math id="M154" display="inline"><mml:mo>+</mml:mo></mml:math></inline-formula> Eesti Maaülikool (EMÜ) staff mobility program and to the Network on Environmental Monitoring and Modeling (RESMA) project from the Federal University of Paraná (UFPR) – Coordination for the Improvement of Higher Education Personnel (CAPES) – Institutional Internationalization Program (PRINT) for facilitating the exchange of researchers between Brazil and Estonia. This work was carried out with the support of the Technology and Environmental Monitoring System of Paraná (Simepar) and the Sanitation Company of Paraná (Sanepar).</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d1e4384">This research has been supported by the Estonian Research Competency Council (grant no. PRG1674), the European Union's Horizon 2020 (grant no. 871115), and the Coordenação de Aperfeiçoamento de Pessoal de Nível Superior (PRINT).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d1e4390">This paper was edited by Frederiek Sperna Weiland and reviewed by Juraj Parajka and one anonymous referee.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>AGUASPARANÁ(2010)</label><mixed-citation>AGUASPARANÁ: Manual técnico de outorgas, i Edn., Estado do Paraná, <uri>https://www.iat.pr.gov.br/sites/agua-terra/arquivos_restritos/files/documento/2020-10/manual_outorgas_suderhsa_2006.pdf</uri> (last access: 17 July 2024), 2010.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Allen et al.(1998)Allen, Pereira, Raes, and Smith</label><mixed-citation> Allen, R. G., Pereira, L. S., Raes, D., and Smith, M.: Crop evapotranspiration: guidelines for computing crop water requirements, Food and Agriculture Organization of the United Nations, Rome, ISBN 9251042195, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Almagro et al.(2021)Almagro, Oliveira, Neto, Roy, and Troch</label><mixed-citation>Almagro, A., Oliveira, P. T. S., Neto, A. A. M., Roy, T., and Troch, P.: CABra: a novel large-sample dataset for Brazilian catchments, Hydrol. Earth Syst. Sci., 25, 3105–3135, <ext-link xlink:href="https://doi.org/10.5194/hess-25-3105-2021" ext-link-type="DOI">10.5194/hess-25-3105-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>Arsenault et al.(2019)Arsenault, Breton-Dufour, Poulin, Dallaire, and Romero-Lopez</label><mixed-citation>Arsenault, R., Breton-Dufour, M., Poulin, A., Dallaire, G., and Romero-Lopez, R.: Streamflow prediction in ungauged basins: analysis of regionalization methods in a hydrologically heterogeneous region of Mexico, Hydrolog. Sci. J., 64, 1297–1311, <ext-link xlink:href="https://doi.org/10.1080/02626667.2019.1639716" ext-link-type="DOI">10.1080/02626667.2019.1639716</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Auler and Farrant(1996)</label><mixed-citation> Auler, A. and Farrant, A.: A brief introduction to karst and caves in Brazil, Proceedings of the University of Bristol Spelaeological Society, 20, 187–200, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Ayzel et al.(2019)Ayzel, Varentsova, Erina, Sokolov, Kurochkina, and Moreydo</label><mixed-citation>Ayzel, G., Varentsova, N., Erina, O., Sokolov, D., Kurochkina, L., and Moreydo, V.: OpenForecast: The First Open-Source Operational Runoff Forecasting System in Russia, Water, 11, 1546, <ext-link xlink:href="https://doi.org/10.3390/w11081546" ext-link-type="DOI">10.3390/w11081546</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Barbieri et al.(2017)Barbieri, Costa, Olivieira, Jusevicius, and vila</label><mixed-citation>Barbieri, G. M. L., Costa, A. B. F., Olivieira, C., Jusevicius, M., and D'Ávila, V. C.: Atlas Solarimétrico Do Estado Do Paraná, Manuscrito não publicado, <uri>https://solar.copel.com/solar/atlas-solarimetrico-copel.pdf</uri> (last access: 17 July 2024), 2017.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Bartiko et al.(2019)Bartiko, Oliveira, Bonumá, and Chaffe</label><mixed-citation> Bartiko, D., Oliveira, D., Bonumá, N., and Chaffe, P.: Spatial and seasonal patterns of flood change across Brazil, Hydrolog. Sci. J., 64, 1071–1079, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Bazzo and Almeida(2016)</label><mixed-citation> Bazzo, J. P. V. and Almeida, R. C. d.: Regionalização de Vazões com o Emprego de Redes Neurais Artificiais RBF, in: I Simpósio de Métodos Numéricos em Engenharia, 30 November 2016, Curitiba, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>Blöschl et al.(2013)Blschl, Sivapalan, Wagener, Viglione, and Savenije</label><mixed-citation>Blöschl, G., Sivapalan, M., Wagener, T., Viglione, A., and Savenije, H.: Runoff Prediction in Ungauged Basins, Cambridge University Press, <ext-link xlink:href="https://doi.org/10.1017/cbo9781139235761" ext-link-type="DOI">10.1017/cbo9781139235761</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Boutsidis et al.(2014)Boutsidis, Zouzias, Mahoney, and Drineas</label><mixed-citation>Boutsidis, C., Zouzias, A., Mahoney, M. W., and Drineas, P.: Randomized Dimensionality Reduction for <inline-formula><mml:math id="M155" display="inline"><mml:mi>k</mml:mi></mml:math></inline-formula>-Means Clustering, IEEE T. Inf. Theory, 61, 1045–1062, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Breiman(2001)</label><mixed-citation> Breiman, L.: Random Forests, Mach. Learn., 45, 5–32, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Burn et al.(1997)Burn, Zrinji, and Kowalchuk</label><mixed-citation> Burn, D. H., Zrinji, Z., and Kowalchuk, M.: Regionalization of catchments for regional flood frequency analysis, J. Hydrol. Eng., 2, 76–82, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Burnash(1995)</label><mixed-citation>Burnash, R. J. C.: The NWS River Forecast System-catchment modeling, in: Computer models of watershed hydrology, 311–366, <uri>https://www.cabidigitallibrary.org/doi/full/10.5555/19961904770</uri> (last access: 1 February 2020), 1995.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Burt and McDonnell(2015)</label><mixed-citation> Burt, T. P. and McDonnell, J. J.: Whither field hydrology? The need for discovery science and outrageous hydrological hypotheses, Water Resour. Res., 51, 5919–5928, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Calvetti et al.(2017)Calvetti, Beneti, Neundorf, Inouye, dos Santos, Gomes, Herdies, and de Gonçalves</label><mixed-citation>Calvetti, L., Beneti, C., Neundorf, R. L. A., Inouye, R. T., dos Santos, T. N., Gomes, A. M., Herdies, D. L., and de Gonçalves, L. G. G.: Quantitative Precipitation Estimation Integrated by Poisson's Equation Using Radar Mosaic, Satellite, and Rain Gauge Network, J. Hydrol. Eng., 22, E5016003, <ext-link xlink:href="https://doi.org/10.1061/(asce)he.1943-5584.0001432" ext-link-type="DOI">10.1061/(asce)he.1943-5584.0001432</ext-link>, 2017. </mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Carneiro et al.(2020)Carneiro, Ostroski, and Mercuri</label><mixed-citation> Carneiro, L., Ostroski, A., and Mercuri, E. G. F.: Trophic state index for heavily impacted watersheds: modeling the influence of diffuse pollution in water bodies, Hydrolog. Sci. J., 65, 2548–2560, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>Chagas et al.(2020)Chagas, Chaffe, Addor, Fan, Fleischmann, Paiva, and Siqueira</label><mixed-citation>Chagas, V. B. P., Chaffe, P. L. B., Addor, N., Fan, F. M., Fleischmann, A. S., Paiva, R. C. D., and Siqueira, V. A.: CAMELS-BR: hydrometeorological time series and landscape attributes for 897 catchments in Brazil, Earth Syst. Sci. Data, 12, 2075–2096, <ext-link xlink:href="https://doi.org/10.5194/essd-12-2075-2020" ext-link-type="DOI">10.5194/essd-12-2075-2020</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Cunha et al.(2019)Cunha, Zeri, Deusdará Leal, Costa, Cuartas, Marengo, Tomasella, Vieira, Barbosa, Cunningham et al.</label><mixed-citation>Cunha, A. P. M. A., Zeri, M., Leal, K. D., Costa, L., Cuartas, L. A., Marengo, J. A., Tomasella, J., Vieira, R. M., Barbosa, A. A., Cunningham, C., Garcia, J. V. C., Broedel, E., Alvalá, R., and Ribeiro-Neto, G.: Extreme drought events over Brazil from 2011 to 2019, Atmosphere, 10, 642, <ext-link xlink:href="https://doi.org/10.3390/atmos10110642" ext-link-type="DOI">10.3390/atmos10110642</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Daggupati et al.(2015)Daggupati, Pai, Ale, Douglas-Mankin, andJ. Jeong, Parajuli, Saraswat, and Youssef</label><mixed-citation>Daggupati, P., Pai, N., Ale, S., Douglas-Mankin, K. R., andJ. Jeong, R. W. Z., Parajuli, P. B., Saraswat, D., and Youssef, M. A.: A Recommended Calibration and Validation Strategy for Hydrologic and Water Quality Models, Am. Soc. Agricult. Biol. Eng., 58, 1705–1719, <ext-link xlink:href="https://doi.org/10.13031/trans.58.10712" ext-link-type="DOI">10.13031/trans.58.10712</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Embrapa(2020)</label><mixed-citation>Embrapa: Mapa de solos do estado do Paraná, <uri>http://geoinfo.cnps.embrapa.br/layers/geonode:parana_solos_20201105</uri>, (last access: 5 July 2021), 2020.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Guo et al.(2020)Guo, Zhang, Zhang, and Wang</label><mixed-citation>Guo, Y., Zhang, Y., Zhang, L., and Wang, Z.: Regionalization of hydrological modeling for predicting streamflow in ungauged catchments: A comprehensive review, Wires Water, 8, e1487, <ext-link xlink:href="https://doi.org/10.1002/wat2.1487" ext-link-type="DOI">10.1002/wat2.1487</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>He et al.(2011)He, Bardossy, and Zehe</label><mixed-citation>He, Y., Bárdossy, A., and Zehe, E.: A review of regionalisation for continuous streamflow simulation, Hydrol. Earth Syst. Sci., 15, 3539–3553, <ext-link xlink:href="https://doi.org/10.5194/hess-15-3539-2011" ext-link-type="DOI">10.5194/hess-15-3539-2011</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Hengl et al.(2007)Hengl, Heuvelink, and G.Rossiter</label><mixed-citation> Hengl, T., Heuvelink, G. B. M., and Rossiter, G. D.: About regression-kriging: From equations to case studies, Comput. Geosci., 33, 1301–1315, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Hirata and Foster(2021)</label><mixed-citation>Hirata, R. and Foster, S.: The Guarani Aquifer System – from regional reserves to local use, Q. J. Eng. Geol. Hydrogeol., 54, qjegh2020-091, <ext-link xlink:href="https://doi.org/10.1144/qjegh2020-091" ext-link-type="DOI">10.1144/qjegh2020-091</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Hrachowitz et al.(2013)Hrachowitz, Savenije, Blschl, McDonnell, Sivapalan, Pomeroy, Arheimer, Blume, Clark, Ehret, Fenicia, Freer, Gelfan, Gupta, Hughes, Hut, Montanari, Pande, Tetzlaff, Troch, Uhlenbrook, Wagener, Winsemius, Woods, Zehe, and Cudennec</label><mixed-citation> Hrachowitz, M., Savenije, H. H. G., Blöschl, G., McDonnell, J. J., Sivapalan, M., Pomeroy, J. W., Arheimer, B., Blume, T., Clark, M. P., Ehret, U., Fenicia, F., Freer, J. E., Gelfan, A., Gupta, H. V., Hughes, D. A., Hut, R. W., Montanari, A., Pande, S., Tetzlaff, D., Troch, P. A., Uhlenbrook, S., Wagener, T., Winsemius, H. C., Woods, R. A., Zehe, E., and Cudennec, C.: A decade of Predictions in Ungauged Basins (PUB) – a review, Hydrolog. Sci. J., 58, 1–58, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>IAT(2020)</label><mixed-citation>IAT: Mapas e Dados Espaciais, <uri>http://www.iat.pr.gov.br/Pagina/Mapas-e-Dados-Espaciais</uri> (last access: 5 July 2021), 2020.</mixed-citation></ref>
      <ref id="bib1.bibx28"><label>Juliani et al.(2020)Juliani, de Campos, Almeida, and Leite</label><mixed-citation>Juliani, B. H. T., de Campos, A. L., Almeida, A. S., and Leite, E. A.: Estatísticas meteorológicas da seca de 2020 no estado do Paraná, in: Anais do II END – Encontro Nacional de Desastres da ABRHidro, ABRHidro, <uri>https://anais.abrhidro.org.br/job.php?Job=7358</uri> (last access: 17 July 2024), 2020.</mixed-citation></ref>
      <ref id="bib1.bibx29"><label>Kaviski et al.(2002)Kaviski, Rohn, and Mazer</label><mixed-citation> Kaviski, E., Rohn, M. d. C., and Mazer, W.: Projeto HG-171: Consistência e regionalização de dados hidrológicos, Centro de Hidráulica e Hidrologia Prof. Parigot de Souza, 2002.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Ketchen Junior and Shook(1996)</label><mixed-citation> Ketchen Junior, D. J. and Shook, C. L.: The application of cluster analysis in strategic management research: an analysis and critique, Strat. Manage. J., 17, 441–458, 1996.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Krause et al.(2005)Krause, Boyle, and Bäse</label><mixed-citation>Krause, P., Boyle, D. P., and Bäse, F.: Comparison of different efficiency criteria for hydrological model assessment, Adv. Geosci., 5, 89–97, <ext-link xlink:href="https://doi.org/10.5194/adgeo-5-89-2005" ext-link-type="DOI">10.5194/adgeo-5-89-2005</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Kuentz et al.(2017)Kuentz, Arheimer, Hundecha, and Wagener</label><mixed-citation>Kuentz, A., Arheimer, B., Hundecha, Y., and Wagener, T.: Understanding hydrologic variability across Europe through catchment classification, Hydrol. Earth Syst. Sci., 21, 2863–2879, <ext-link xlink:href="https://doi.org/10.5194/hess-21-2863-2017" ext-link-type="DOI">10.5194/hess-21-2863-2017</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Llabrés-Brustenga et al.(2019)Llabrs-Brustenga, Rius, Rodíguez-Sol, Casas-Castillo, and Redao</label><mixed-citation> Llabrés-Brustenga, A., Rius, A., Rodríguez-Sol, R., Casas-Castillo, M. C., and Redaño, A.: Quality control process of the daily rainfall series available in Catalonia from 1855 to the present, Theor. Appl. Climatol., 137, 2715–2729, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Matallo Junior(2001)</label><mixed-citation> Matallo Junior, H.: Indicadores de desertificação: histórico e perspectivas, Edições UNESCO Brasil, Brasília, DF, Brasil, ISBN 8587853279, 2001.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>Melo et al.(2020)Melo, Anache, Almeida, Coutinho, Ramos Filho, Rosalem, Pelinson, Ferreira, Schwamback, Calixto et al.</label><mixed-citation> Melo, D., Ramos, G., Ferreira, G., Schwamback, D., Siqueira, J., Duarte-Carvajalino, J., Jhunior, H., Nóbrega, J., Morita, A., Almeida, C., Coutinho, J., Leite, C., Guedes, A., Coelho, V. H., Anache, J., Pelinson, N., Rosalem, L., Calixto, K. G., and Wendland, E.: The big picture of field hydrology studies in Brazil, Hydrolog. Sci. J., 65, 1262–1280, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>Melo et al.(2016)Melo, Scanlon, Zhang, Wendland, and Yin</label><mixed-citation>Melo, D. D. C. D., Scanlon, B. R., Zhang, Z., Wendland, E., and Yin, L.: Reservoir storage and hydrologic responses to droughts in the Paraná River basin, south-eastern Brazil, Hydrol. Earth Syst. Sci., 20, 4673–4688, <ext-link xlink:href="https://doi.org/10.5194/hess-20-4673-2016" ext-link-type="DOI">10.5194/hess-20-4673-2016</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Mohamed et al.(2019)Mohamed, Ludovic, and Ribstein</label><mixed-citation>Mohamed, S., Ludovic, O., and Ribstein, P.: Random Forest Ability in Regionalizing Hourly Hydrological Model Parameters, Water, 11, 8, <ext-link xlink:href="https://doi.org/10.3390/w11081540" ext-link-type="DOI">10.3390/w11081540</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx38"><label>Muleta(2012)</label><mixed-citation>Muleta, M. K.: Model Performance Sensitivity to Objective Function during Automated Calibrations, J. Hydrol. Eng., 17, 756–767, <ext-link xlink:href="https://doi.org/10.1061/(asce)he.1943-5584.0000497" ext-link-type="DOI">10.1061/(asce)he.1943-5584.0000497</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx39"><label>Musy et al.(2014)Musy, Hingray, and Picouet</label><mixed-citation>Musy, A., Hingray, B., and Picouet, C.: Hydrology: a science for engineers, CRC Press, <ext-link xlink:href="https://doi.org/10.1201/b17169" ext-link-type="DOI">10.1201/b17169</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx40"><label>Neto et al.(2021)Neto, Vieira, and Matosinhos</label><mixed-citation>Neto, W. M. P., Vieira, F. R., and Matosinhos, C. C.: Avaliação da perfomance dos modelos GR4J, GR5J e GR6J na bacia hidrográfica do ribeirão São João, Minas Gerais, in: Base de Conhecimentos Gerados na Engenharia Ambiental e Sanitária 3, Atena, <ext-link xlink:href="https://doi.org/10.22533/at.ed.74521080423" ext-link-type="DOI">10.22533/at.ed.74521080423</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx41"><label>Oudin et al.(2008)Oudin, Andréassian, Perrin, Michel, and Moine</label><mixed-citation>Oudin, L., Andréassian, V., Perrin, C., Michel, C., and Moine, N. L.: Spatial proximity, physical similarity, regression and ungaged catchments: A comparison of regionalization approaches based on 913 French catchments, Water Resour. Res., 44, W03413, <ext-link xlink:href="https://doi.org/10.1029/2007wr006240" ext-link-type="DOI">10.1029/2007wr006240</ext-link>, 2008.</mixed-citation></ref>
      <ref id="bib1.bibx42"><label>Oudin et al.(2010)Oudin, Kay, Andréassian, and Perrin</label><mixed-citation>Oudin, L., Kay, A., Andréassian, V., and Perrin, C.: Are seemingly physically similar catchments truly hydrologically similar?, Water Resour. Res., 46, W11558, <ext-link xlink:href="https://doi.org/10.1029/2009wr008887" ext-link-type="DOI">10.1029/2009wr008887</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx43"><label>Pagano et al.(2010)Pagano, Hapuarachchi, and Wang</label><mixed-citation>Pagano, T., Hapuarachchi, P., and Wang, Q. J.: Continuous rainfall-runoff model comparison and short-term daily streamflow forecast skill evaluation, Tech. Rep., CSIRO, EP103545, <ext-link xlink:href="https://doi.org/10.4225/08/58542C672DD2C" ext-link-type="DOI">10.4225/08/58542C672DD2C</ext-link>, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx44"><label>Parajka et al.(2005)Parajka, Merz, and Blschl</label><mixed-citation>Parajka, J., Merz, R., and Blöschl, G.: A comparison of regionalisation methods for catchment model parameters, Hydrol. Earth Syst. Sci., 9, 157–171, <ext-link xlink:href="https://doi.org/10.5194/hess-9-157-2005" ext-link-type="DOI">10.5194/hess-9-157-2005</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx45"><label>Parajka et al.(2010)Parajka, Kohnová, Bálint, Barbuc, Borga, Claps, Cheval, Dumitrescu, Gaume, Hlavčová et al.</label><mixed-citation> Parajka, J., Kohnová, S., Bálint, G., Barbuc, M., Borga, M., Claps, P., Cheval, S., Dumitrescu, A., Gaume, E., Hlavčová, K., Merz, R., Pfaundler, M., Stancalie, G., Szolgay, J., and Blöschl, G.: Seasonal characteristics of flood regimes across the Alpine–Carpathian range, J. Hydrol., 394, 78–89, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx46"><label>Perrin et al.(2003)Perrin, Michel, and Andrassian</label><mixed-citation> Perrin, C., Michel, C., and Andréassian, V.: Improvement of a parsimonious model for streamflow simulation, J. Hydrol., 279, 275–289, 2003.</mixed-citation></ref>
      <ref id="bib1.bibx47"><label>Pettitt(1979)</label><mixed-citation>Pettitt, A. N.: A Non-Parametric Approach to the Change-Point Problem, Appl. Stat., 28, 126, <ext-link xlink:href="https://doi.org/10.2307/2346729" ext-link-type="DOI">10.2307/2346729</ext-link>, 1979.</mixed-citation></ref>
      <ref id="bib1.bibx48"><label>Pfafstetter(1989)</label><mixed-citation> Pfafstetter, O.: Classificação de bacias hidrográficas, manuscrito não publicado, DNOS – Departamento Nacional de Obras de Saneamento, 1989.</mixed-citation></ref>
      <ref id="bib1.bibx49"><label>Pugliese et al.(2014)Pugliese, Castellarin, and Brath</label><mixed-citation>Pugliese, A., Castellarin, A., and Brath, A.: Geostatistical prediction of flow–duration curves in an index-flow framework, Hydrol. Earth Syst. Sci., 18, 3801–3816, <ext-link xlink:href="https://doi.org/10.5194/hess-18-3801-2014" ext-link-type="DOI">10.5194/hess-18-3801-2014</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx50"><label>Pushpalatha et al.(2011)Pushpalatha, Perrin, Le Moine, Mathevet, and Andréassian</label><mixed-citation> Pushpalatha, R., Perrin, C., Le Moine, N., Mathevet, T., and Andréassian, V.: A downward structural sensitivity analysis of hydrological models to improve low-flow simulation, J. Hydrol., 411, 66–76, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx51"><label>Razavi and Coulibaly(2013)</label><mixed-citation>Razavi, T. and Coulibaly, P.: Streamflow Prediction in Ungauged Basins: Review of Regionalization Methods, J. Hydrol. Eng., 18, 958–975, <ext-link xlink:href="https://doi.org/10.1061/(asce)he.1943-5584.0000690" ext-link-type="DOI">10.1061/(asce)he.1943-5584.0000690</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx52"><label>Rousseeuw(1987)</label><mixed-citation>Rousseeuw, P. J.: Silhouettes: A graphical aid to the interpretation and validation of cluster analysis, J. Comput. Appl. Math., 20, 53–65, <ext-link xlink:href="https://doi.org/10.1016/0377-0427(87)90125-7" ext-link-type="DOI">10.1016/0377-0427(87)90125-7</ext-link>, 1987.</mixed-citation></ref>
      <ref id="bib1.bibx53"><label>Shin and Kim(2016)</label><mixed-citation>Shin, M.-J. and Kim, C.-S.: Assessment of the suitability of rainfall–runoff models by coupling performance statistics and sensitivity analysis, Hydrol. Res., 48, 1192–1213, <ext-link xlink:href="https://doi.org/10.2166/nh.2016.129" ext-link-type="DOI">10.2166/nh.2016.129</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx54"><label>Soil Conservation Service(1972)</label><mixed-citation>Soil Conservation Service: National engineering handbook, in: Chap. Seção 4, Hydrology, Department ofAgriculture, Washington, p. 762, <uri>https://books.google.com.br/books?id=sjOEf-5zjXgC</uri> (last access: 1 December 2023), 1972.</mixed-citation></ref>
      <ref id="bib1.bibx55"><label>Sousa et al.(2009)Sousa, Neto, Pacheco, and Barbosa</label><mixed-citation>Sousa, F. M. L., Neto, V. S. C., Pacheco, W. E., and Barbosa, S. A.: Sistema Nacional De Informações Sobre Recursos Hídricos: Sistematização Conceitual E Modelagem Funcional, in: Anais do XVIII Simpósio Brasileiro de Recursos Hídricos, Associação Brasileira de Recursos Hídricos, Campo Grande, <uri>https://anais.abrhidro.org.br/job.php?Job=10334</uri> (last access: 1 December 2023), 2009. </mixed-citation></ref>
      <ref id="bib1.bibx56"><label>Souza et al.(2020)Souza, Shimbo, Rosa, Parente, Alencar, Rudorff, Hasenack, Matsumoto, Ferreira, Souza-Filho, de Oliveira, Rocha, Fonseca, Marques, Diniz, Costa, Monteiro, Rosa, Vlez-Martin, Weber, Lenti, Paternost, Pareyn, Siqueira, Viera, Neto, Saraiva, Sales, Salgado, Vasconcelos, Galano, Mesquita, and Azevedo</label><mixed-citation>Souza, C. M., Shimbo, J. Z., Rosa, M. R., Parente, L. L., Alencar, A. A., Rudorff, B. F. T., Hasenack, H., Matsumoto, M., Ferreira, L. G., Souza-Filho, P. W. M., de Oliveira, S. W., Rocha, W. F., Fonseca, A. V., Marques, C. B., Diniz, C. G., Costa, D., Monteiro, D., Rosa, E. R., Vélez-Martin, E., Weber, E. J., Lenti, F. E. B., Paternost, F. F., Pareyn, F. G. C., Siqueira, J. V., Viera, J. L., Neto, L. C. F., Saraiva, M. M., Sales, M. H., Salgado, M. P. G., Vasconcelos, R., Galano, S., Mesquita, V. V., and Azevedo, T.: Reconstructing Three Decades of Land Use and Land Cover Changes in Brazilian Biomes with Landsat Archive and Earth Engine, Remote Sens., 12, 2735, <ext-link xlink:href="https://doi.org/10.3390/rs12172735" ext-link-type="DOI">10.3390/rs12172735</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx57"><label>Storn and Price(1997)</label><mixed-citation>Storn, R. and Price, K.: Differential Evolution – A Simple and Efficient Heuristic for Global Optimization over Continuous Spaces, J. Global Optimiz., 11, 341–359, <ext-link xlink:href="https://doi.org/10.1023/a:1008202821328" ext-link-type="DOI">10.1023/a:1008202821328</ext-link>, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx58"><label>Vestena and Kobiyama(2007)</label><mixed-citation> Vestena, L. R. and Kobiyama, M.: Water balance in karst: case study of the Ribeirão da Onça catchment in Colombo City, Paraná State-Brazil, Brazil. Arch. Biol. Technol., 50, 905–912, 2007.</mixed-citation></ref>
      <ref id="bib1.bibx59"><label>Viviroli et al.(2009)Viviroli, Mittelbach, Gurtz, and Weingartner</label><mixed-citation>Viviroli, D., Mittelbach, H., Gurtz, J., and Weingartner, R.: Continuous simulation for flood estimation in ungauged mesoscale catchments of Switzerland – Part II: Parameter regionalisation and flood estimation results, J. Hydrol., 377, 208–225, <ext-link xlink:href="https://doi.org/10.1016/j.jhydrol.2009.08.022" ext-link-type="DOI">10.1016/j.jhydrol.2009.08.022</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx60"><label>Wilks(2011)</label><mixed-citation>Wilks, D. S.: Statistical Methods in the Atmospheric Sciences, Academic Press, ISBN 0123850223, <uri>https://www.ebook.de/de/product/14751307/daniel_s_wilks_statistical_methods_in_the_atmospheric_sciences_100.html</uri> (last access: 1 December 2023), 2011.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Regionalization of GR4J model parameters  for river flow prediction in Paraná, Brazil</article-title-html>
<abstract-html/>
<ref-html id="bib1.bib1"><label>AGUASPARANÁ(2010)</label><mixed-citation>
      
AGUASPARANÁ: Manual técnico de outorgas, i Edn., Estado do Paraná, <a href="https://www.iat.pr.gov.br/sites/agua-terra/arquivos_restritos/files/documento/2020-10/manual_outorgas_suderhsa_2006.pdf" target="_blank"/> (last access: 17 July 2024), 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Allen et al.(1998)Allen, Pereira, Raes, and Smith</label><mixed-citation>
      
Allen, R. G., Pereira, L. S., Raes, D., and Smith, M.: Crop evapotranspiration: guidelines for computing crop water requirements, Food and Agriculture Organization of the United Nations, Rome, ISBN 9251042195, 1998.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Almagro et al.(2021)Almagro, Oliveira, Neto, Roy, and
Troch</label><mixed-citation>
      
Almagro, A., Oliveira, P. T. S., Neto, A. A. M., Roy, T., and Troch, P.:
CABra: a novel large-sample dataset for Brazilian catchments, Hydrol. Earth Syst. Sci., 25, 3105–3135, <a href="https://doi.org/10.5194/hess-25-3105-2021" target="_blank">https://doi.org/10.5194/hess-25-3105-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Arsenault et al.(2019)Arsenault, Breton-Dufour, Poulin, Dallaire, and Romero-Lopez</label><mixed-citation>
      
Arsenault, R., Breton-Dufour, M., Poulin, A., Dallaire, G., and Romero-Lopez,
R.: Streamflow prediction in ungauged basins: analysis of regionalization
methods in a hydrologically heterogeneous region of Mexico, Hydrolog. Sci. J., 64, 1297–1311, <a href="https://doi.org/10.1080/02626667.2019.1639716" target="_blank">https://doi.org/10.1080/02626667.2019.1639716</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Auler and Farrant(1996)</label><mixed-citation>
      
Auler, A. and Farrant, A.: A brief introduction to karst and caves in Brazil,
Proceedings of the University of Bristol Spelaeological Society, 20, 187–200, 1996.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Ayzel et al.(2019)Ayzel, Varentsova, Erina, Sokolov, Kurochkina, and Moreydo</label><mixed-citation>
      
Ayzel, G., Varentsova, N., Erina, O., Sokolov, D., Kurochkina, L., and Moreydo, V.: OpenForecast: The First Open-Source Operational Runoff Forecasting System in Russia, Water, 11, 1546, <a href="https://doi.org/10.3390/w11081546" target="_blank">https://doi.org/10.3390/w11081546</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Barbieri et al.(2017)Barbieri, Costa, Olivieira, Jusevicius, and
vila</label><mixed-citation>
      
Barbieri, G. M. L., Costa, A. B. F., Olivieira, C., Jusevicius, M., and
D'Ávila, V. C.: Atlas Solarimétrico Do Estado Do Paraná, Manuscrito não publicado, <a href="https://solar.copel.com/solar/atlas-solarimetrico-copel.pdf" target="_blank"/> (last access: 17 July 2024), 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Bartiko et al.(2019)Bartiko, Oliveira, Bonumá, and
Chaffe</label><mixed-citation>
      
Bartiko, D., Oliveira, D., Bonumá, N., and Chaffe, P.: Spatial and seasonal patterns of flood change across Brazil, Hydrolog. Sci. J., 64,
1071–1079, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Bazzo and Almeida(2016)</label><mixed-citation>
      
Bazzo, J. P. V. and Almeida, R. C. d.: Regionalização de Vazões com o Emprego de Redes Neurais Artificiais RBF, in: I Simpósio de Métodos Numéricos em Engenharia, 30 November 2016, Curitiba, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Blöschl et al.(2013)Blschl, Sivapalan, Wagener, Viglione, and
Savenije</label><mixed-citation>
      
Blöschl, G., Sivapalan, M., Wagener, T., Viglione, A., and Savenije, H.:
Runoff Prediction in Ungauged Basins, Cambridge University Press,
<a href="https://doi.org/10.1017/cbo9781139235761" target="_blank">https://doi.org/10.1017/cbo9781139235761</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Boutsidis et al.(2014)Boutsidis, Zouzias, Mahoney, and
Drineas</label><mixed-citation>
      
Boutsidis, C., Zouzias, A., Mahoney, M. W., and Drineas, P.: Randomized
Dimensionality Reduction for <i>k</i>-Means Clustering, IEEE T. Inf. Theory, 61, 1045–1062, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Breiman(2001)</label><mixed-citation>
      
Breiman, L.: Random Forests, Mach. Learn., 45, 5–32, 2001.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Burn et al.(1997)Burn, Zrinji, and
Kowalchuk</label><mixed-citation>
      
Burn, D. H., Zrinji, Z., and Kowalchuk, M.: Regionalization of catchments for
regional flood frequency analysis, J. Hydrol. Eng., 2, 76–82, 1997.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Burnash(1995)</label><mixed-citation>
      
Burnash, R. J. C.: The NWS River Forecast System-catchment modeling, in: Computer models of watershed hydrology, 311–366, <a href="https://www.cabidigitallibrary.org/doi/full/10.5555/19961904770" target="_blank"/> (last access: 1 February 2020), 1995.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Burt and McDonnell(2015)</label><mixed-citation>
      
Burt, T. P. and McDonnell, J. J.: Whither field hydrology? The need for
discovery science and outrageous hydrological hypotheses, Water Resour. Res., 51, 5919–5928, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Calvetti et al.(2017)Calvetti, Beneti, Neundorf, Inouye, dos Santos, Gomes, Herdies, and de Gonçalves</label><mixed-citation>
      
Calvetti, L., Beneti, C., Neundorf, R. L. A., Inouye, R. T., dos Santos, T. N., Gomes, A. M., Herdies, D. L., and de Gonçalves, L. G. G.: Quantitative Precipitation Estimation Integrated by Poisson's Equation Using Radar Mosaic, Satellite, and Rain Gauge Network, J. Hydrol. Eng., 22, E5016003, <a href="https://doi.org/10.1061/(asce)he.1943-5584.0001432" target="_blank">https://doi.org/10.1061/(asce)he.1943-5584.0001432</a>, 2017.


    </mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Carneiro et al.(2020)Carneiro, Ostroski, and
Mercuri</label><mixed-citation>
      
Carneiro, L., Ostroski, A., and Mercuri, E. G. F.: Trophic state index for
heavily impacted watersheds: modeling the influence of diffuse pollution in
water bodies, Hydrolog. Sci. J., 65, 2548–2560, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Chagas et al.(2020)Chagas, Chaffe, Addor, Fan, Fleischmann, Paiva,
and Siqueira</label><mixed-citation>
      
Chagas, V. B. P., Chaffe, P. L. B., Addor, N., Fan, F. M., Fleischmann, A. S., Paiva, R. C. D., and Siqueira, V. A.: CAMELS-BR: hydrometeorological time series and landscape attributes for 897 catchments in Brazil, Earth Syst. Sci. Data, 12, 2075–2096, <a href="https://doi.org/10.5194/essd-12-2075-2020" target="_blank">https://doi.org/10.5194/essd-12-2075-2020</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Cunha et al.(2019)Cunha, Zeri, Deusdará Leal, Costa, Cuartas,
Marengo, Tomasella, Vieira, Barbosa, Cunningham et al.</label><mixed-citation>
      
Cunha, A. P. M. A., Zeri, M., Leal, K. D., Costa, L., Cuartas, L. A., Marengo, J. A., Tomasella, J., Vieira, R. M., Barbosa, A. A., Cunningham, C., Garcia, J. V. C., Broedel, E., Alvalá, R., and Ribeiro-Neto, G.: Extreme drought events over Brazil from 2011 to 2019, Atmosphere, 10, 642, <a href="https://doi.org/10.3390/atmos10110642" target="_blank">https://doi.org/10.3390/atmos10110642</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Daggupati et al.(2015)Daggupati, Pai, Ale, Douglas-Mankin, andJ.
Jeong, Parajuli, Saraswat, and Youssef</label><mixed-citation>
      
Daggupati, P., Pai, N., Ale, S., Douglas-Mankin, K. R., andJ. Jeong, R. W. Z., Parajuli, P. B., Saraswat, D., and Youssef, M. A.: A Recommended Calibration and Validation Strategy for Hydrologic and Water Quality Models, Am. Soc. Agricult. Biol. Eng., 58, 1705–1719, <a href="https://doi.org/10.13031/trans.58.10712" target="_blank">https://doi.org/10.13031/trans.58.10712</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Embrapa(2020)</label><mixed-citation>
      
Embrapa: Mapa de solos do estado do Paraná,
<a href="http://geoinfo.cnps.embrapa.br/layers/geonode:parana_solos_20201105" target="_blank"/>,
(last access: 5 July 2021), 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Guo et al.(2020)Guo, Zhang, Zhang, and Wang</label><mixed-citation>
      
Guo, Y., Zhang, Y., Zhang, L., and Wang, Z.: Regionalization of hydrological
modeling for predicting streamflow in ungauged catchments: A comprehensive
review, Wires Water, 8, e1487, <a href="https://doi.org/10.1002/wat2.1487" target="_blank">https://doi.org/10.1002/wat2.1487</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>He et al.(2011)He, Bardossy, and Zehe</label><mixed-citation>
      
He, Y., Bárdossy, A., and Zehe, E.: A review of regionalisation for continuous streamflow simulation, Hydrol. Earth Syst. Sci., 15, 3539–3553, <a href="https://doi.org/10.5194/hess-15-3539-2011" target="_blank">https://doi.org/10.5194/hess-15-3539-2011</a>, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Hengl et al.(2007)Hengl, Heuvelink, and G.Rossiter</label><mixed-citation>
      
Hengl, T., Heuvelink, G. B. M., and Rossiter, G. D.: About regression-kriging: From equations to case studies, Comput. Geosci., 33, 1301–1315, 2007.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Hirata and Foster(2021)</label><mixed-citation>
      
Hirata, R. and Foster, S.: The Guarani Aquifer System – from regional reserves to local use, Q. J. Eng. Geol. Hydrogeol., 54, qjegh2020-091, <a href="https://doi.org/10.1144/qjegh2020-091" target="_blank">https://doi.org/10.1144/qjegh2020-091</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Hrachowitz et al.(2013)Hrachowitz, Savenije, Blschl, McDonnell,
Sivapalan, Pomeroy, Arheimer, Blume, Clark, Ehret, Fenicia, Freer, Gelfan,
Gupta, Hughes, Hut, Montanari, Pande, Tetzlaff, Troch, Uhlenbrook, Wagener,
Winsemius, Woods, Zehe, and Cudennec</label><mixed-citation>
      
Hrachowitz, M., Savenije, H. H. G., Blöschl, G., McDonnell, J. J., Sivapalan, M., Pomeroy, J. W., Arheimer, B., Blume, T., Clark, M. P., Ehret, U., Fenicia, F., Freer, J. E., Gelfan, A., Gupta, H. V., Hughes, D. A., Hut,
R. W., Montanari, A., Pande, S., Tetzlaff, D., Troch, P. A., Uhlenbrook, S.,
Wagener, T., Winsemius, H. C., Woods, R. A., Zehe, E., and Cudennec, C.: A
decade of Predictions in Ungauged Basins (PUB) – a review, Hydrolog. Sci. J., 58, 1–58, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>IAT(2020)</label><mixed-citation>
      
IAT: Mapas e Dados Espaciais,
<a href="http://www.iat.pr.gov.br/Pagina/Mapas-e-Dados-Espaciais" target="_blank"/> (last access: 5 July 2021), 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Juliani et al.(2020)Juliani, de Campos, Almeida, and
Leite</label><mixed-citation>
      
Juliani, B. H. T., de Campos, A. L., Almeida, A. S., and Leite, E. A.:
Estatísticas meteorológicas da seca de 2020 no estado do Paraná, in: Anais do II END – Encontro Nacional de Desastres da ABRHidro, ABRHidro, <a href="https://anais.abrhidro.org.br/job.php?Job=7358" target="_blank"/> (last access: 17 July 2024), 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Kaviski et al.(2002)Kaviski, Rohn, and Mazer</label><mixed-citation>
      
Kaviski, E., Rohn, M. d. C., and Mazer, W.: Projeto HG-171: Consistência e regionalização de dados hidrológicos, Centro de Hidráulica e Hidrologia Prof. Parigot de Souza, 2002.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Ketchen Junior and Shook(1996)</label><mixed-citation>
      
Ketchen Junior, D. J. and Shook, C. L.: The application of cluster analysis in strategic management research: an analysis and critique, Strat. Manage. J., 17, 441–458, 1996.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Krause et al.(2005)Krause, Boyle, and Bäse</label><mixed-citation>
      
Krause, P., Boyle, D. P., and Bäse, F.: Comparison of different efficiency criteria for hydrological model assessment, Adv. Geosci., 5, 89–97, <a href="https://doi.org/10.5194/adgeo-5-89-2005" target="_blank">https://doi.org/10.5194/adgeo-5-89-2005</a>, 2005.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Kuentz et al.(2017)Kuentz, Arheimer, Hundecha, and
Wagener</label><mixed-citation>
      
Kuentz, A., Arheimer, B., Hundecha, Y., and Wagener, T.: Understanding
hydrologic variability across Europe through catchment classification, Hydrol. Earth Syst. Sci., 21, 2863–2879, <a href="https://doi.org/10.5194/hess-21-2863-2017" target="_blank">https://doi.org/10.5194/hess-21-2863-2017</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Llabrés-Brustenga et al.(2019)Llabrs-Brustenga, Rius,
Rodíguez-Sol, Casas-Castillo, and Redao</label><mixed-citation>
      
Llabrés-Brustenga, A., Rius, A., Rodríguez-Sol, R., Casas-Castillo, M. C., and Redaño, A.: Quality control process of the daily rainfall series available in Catalonia from 1855 to the present, Theor. Appl.
Climatol., 137, 2715–2729, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Matallo Junior(2001)</label><mixed-citation>
      
Matallo Junior, H.: Indicadores de desertificação: histórico e perspectivas, Edições UNESCO Brasil, Brasília, DF, Brasil, ISBN 8587853279, 2001.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Melo et al.(2020)Melo, Anache, Almeida, Coutinho, Ramos Filho,
Rosalem, Pelinson, Ferreira, Schwamback, Calixto et al.</label><mixed-citation>
      
Melo, D., Ramos, G., Ferreira, G., Schwamback, D., Siqueira, J., Duarte-Carvajalino, J., Jhunior, H., Nóbrega, J., Morita, A., Almeida, C., Coutinho, J., Leite, C., Guedes, A., Coelho, V. H., Anache, J., Pelinson, N., Rosalem, L., Calixto, K. G., and Wendland, E.: The big picture of field hydrology studies in Brazil, Hydrolog. Sci. J., 65, 1262–1280, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Melo et al.(2016)Melo, Scanlon, Zhang, Wendland, and
Yin</label><mixed-citation>
      
Melo, D. D. C. D., Scanlon, B. R., Zhang, Z., Wendland, E., and Yin, L.: Reservoir storage and hydrologic responses to droughts in the Paraná River basin, south-eastern Brazil, Hydrol. Earth Syst. Sci., 20, 4673–4688, <a href="https://doi.org/10.5194/hess-20-4673-2016" target="_blank">https://doi.org/10.5194/hess-20-4673-2016</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Mohamed et al.(2019)Mohamed, Ludovic, and Ribstein</label><mixed-citation>
      
Mohamed, S., Ludovic, O., and Ribstein, P.: Random Forest Ability in
Regionalizing Hourly Hydrological Model Parameters, Water, 11, 8, <a href="https://doi.org/10.3390/w11081540" target="_blank">https://doi.org/10.3390/w11081540</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Muleta(2012)</label><mixed-citation>
      
Muleta, M. K.: Model Performance Sensitivity to Objective Function during
Automated Calibrations, J. Hydrol. Eng., 17, 756–767,
<a href="https://doi.org/10.1061/(asce)he.1943-5584.0000497" target="_blank">https://doi.org/10.1061/(asce)he.1943-5584.0000497</a>, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Musy et al.(2014)Musy, Hingray, and Picouet</label><mixed-citation>
      
Musy, A., Hingray, B., and Picouet, C.: Hydrology: a science for engineers, CRC Press, <a href="https://doi.org/10.1201/b17169" target="_blank">https://doi.org/10.1201/b17169</a>, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Neto et al.(2021)Neto, Vieira, and Matosinhos</label><mixed-citation>
      
Neto, W. M. P., Vieira, F. R., and Matosinhos, C. C.: Avaliação da
perfomance dos modelos GR4J, GR5J e GR6J na bacia hidrográfica do ribeirão São João, Minas Gerais, in: Base de Conhecimentos Gerados na Engenharia Ambiental e Sanitária 3, Atena, <a href="https://doi.org/10.22533/at.ed.74521080423" target="_blank">https://doi.org/10.22533/at.ed.74521080423</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>Oudin et al.(2008)Oudin, Andréassian, Perrin, Michel, and
Moine</label><mixed-citation>
      
Oudin, L., Andréassian, V., Perrin, C., Michel, C., and Moine, N. L.:
Spatial proximity, physical similarity, regression and ungaged catchments: A
comparison of regionalization approaches based on 913 French catchments,
Water Resour. Res., 44, W03413, <a href="https://doi.org/10.1029/2007wr006240" target="_blank">https://doi.org/10.1029/2007wr006240</a>, 2008.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>Oudin et al.(2010)Oudin, Kay, Andréassian, and
Perrin</label><mixed-citation>
      
Oudin, L., Kay, A., Andréassian, V., and Perrin, C.: Are seemingly
physically similar catchments truly hydrologically similar?, Water Resour. Res., 46, W11558, <a href="https://doi.org/10.1029/2009wr008887" target="_blank">https://doi.org/10.1029/2009wr008887</a>, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Pagano et al.(2010)Pagano, Hapuarachchi, and Wang</label><mixed-citation>
      
Pagano, T., Hapuarachchi, P., and Wang, Q. J.: Continuous rainfall-runoff model comparison and short-term daily streamflow forecast skill evaluation, Tech. Rep., CSIRO, EP103545, <a href="https://doi.org/10.4225/08/58542C672DD2C" target="_blank">https://doi.org/10.4225/08/58542C672DD2C</a>, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>Parajka et al.(2005)Parajka, Merz, and Blschl</label><mixed-citation>
      
Parajka, J., Merz, R., and Blöschl, G.: A comparison of regionalisation methods for catchment model parameters, Hydrol. Earth Syst. Sci., 9, 157–171, <a href="https://doi.org/10.5194/hess-9-157-2005" target="_blank">https://doi.org/10.5194/hess-9-157-2005</a>, 2005.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Parajka et al.(2010)Parajka, Kohnová, Bálint, Barbuc, Borga, Claps, Cheval, Dumitrescu, Gaume, Hlavčová
et al.</label><mixed-citation>
      
Parajka, J., Kohnová, S., Bálint, G., Barbuc, M., Borga, M., Claps, P., Cheval, S., Dumitrescu, A., Gaume, E., Hlavčová, K., Merz, R., Pfaundler, M., Stancalie, G., Szolgay, J., and Blöschl, G.: Seasonal characteristics of flood regimes across the Alpine–Carpathian range, J. Hydrol., 394, 78–89, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Perrin et al.(2003)Perrin, Michel, and Andrassian</label><mixed-citation>
      
Perrin, C., Michel, C., and Andréassian, V.: Improvement of a parsimonious model for streamflow simulation, J. Hydrol., 279, 275–289, 2003.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Pettitt(1979)</label><mixed-citation>
      
Pettitt, A. N.: A Non-Parametric Approach to the Change-Point Problem, Appl. Stat., 28, 126, <a href="https://doi.org/10.2307/2346729" target="_blank">https://doi.org/10.2307/2346729</a>, 1979.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>Pfafstetter(1989)</label><mixed-citation>
      
Pfafstetter, O.: Classificação de bacias hidrográficas, manuscrito não publicado, DNOS – Departamento Nacional de Obras de Saneamento, 1989.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>Pugliese et al.(2014)Pugliese, Castellarin, and Brath</label><mixed-citation>
      
Pugliese, A., Castellarin, A., and Brath, A.: Geostatistical prediction of flow–duration curves in an index-flow framework, Hydrol. Earth Syst. Sci., 18, 3801–3816, <a href="https://doi.org/10.5194/hess-18-3801-2014" target="_blank">https://doi.org/10.5194/hess-18-3801-2014</a>, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Pushpalatha et al.(2011)Pushpalatha, Perrin, Le Moine, Mathevet, and Andréassian</label><mixed-citation>
      
Pushpalatha, R., Perrin, C., Le Moine, N., Mathevet, T., and Andréassian,
V.: A downward structural sensitivity analysis of hydrological models to
improve low-flow simulation, J. Hydrol., 411, 66–76, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Razavi and Coulibaly(2013)</label><mixed-citation>
      
Razavi, T. and Coulibaly, P.: Streamflow Prediction in Ungauged Basins: Review of Regionalization Methods, J. Hydrol. Eng., 18, 958–975,
<a href="https://doi.org/10.1061/(asce)he.1943-5584.0000690" target="_blank">https://doi.org/10.1061/(asce)he.1943-5584.0000690</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Rousseeuw(1987)</label><mixed-citation>
      
Rousseeuw, P. J.: Silhouettes: A graphical aid to the interpretation and
validation of cluster analysis, J. Comput. Appl. Math., 20, 53–65, <a href="https://doi.org/10.1016/0377-0427(87)90125-7" target="_blank">https://doi.org/10.1016/0377-0427(87)90125-7</a>, 1987.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Shin and Kim(2016)</label><mixed-citation>
      
Shin, M.-J. and Kim, C.-S.: Assessment of the suitability of rainfall–runoff models by coupling performance statistics and sensitivity analysis, Hydrol. Res., 48, 1192–1213, <a href="https://doi.org/10.2166/nh.2016.129" target="_blank">https://doi.org/10.2166/nh.2016.129</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Soil Conservation Service(1972)</label><mixed-citation>
      
Soil Conservation Service: National engineering handbook, in: Chap. Seção 4, Hydrology, Department ofAgriculture, Washington, p. 762, <a href="https://books.google.com.br/books?id=sjOEf-5zjXgC" target="_blank"/> (last access: 1 December 2023), 1972.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>Sousa et al.(2009)Sousa, Neto, Pacheco, and Barbosa</label><mixed-citation>
      
Sousa, F. M. L., Neto, V. S. C., Pacheco, W. E., and Barbosa, S. A.: Sistema
Nacional De Informações Sobre Recursos Hídricos: Sistematização Conceitual E Modelagem Funcional, in: Anais do XVIII Simpósio Brasileiro de Recursos Hídricos, Associação Brasileira de Recursos Hídricos, Campo Grande, <a href="https://anais.abrhidro.org.br/job.php?Job=10334" target="_blank"/> (last access: 1 December 2023), 2009.


    </mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Souza et al.(2020)Souza, Shimbo, Rosa, Parente, Alencar, Rudorff,
Hasenack, Matsumoto, Ferreira, Souza-Filho, de Oliveira, Rocha, Fonseca,
Marques, Diniz, Costa, Monteiro, Rosa, Vlez-Martin, Weber, Lenti,
Paternost, Pareyn, Siqueira, Viera, Neto, Saraiva, Sales, Salgado,
Vasconcelos, Galano, Mesquita, and Azevedo</label><mixed-citation>
      
Souza, C. M., Shimbo, J. Z., Rosa, M. R., Parente, L. L., Alencar, A. A.,
Rudorff, B. F. T., Hasenack, H., Matsumoto, M., Ferreira, L. G., Souza-Filho,
P. W. M., de Oliveira, S. W., Rocha, W. F., Fonseca, A. V., Marques, C. B.,
Diniz, C. G., Costa, D., Monteiro, D., Rosa, E. R., Vélez-Martin, E., Weber, E. J., Lenti, F. E. B., Paternost, F. F., Pareyn, F. G. C., Siqueira, J. V., Viera, J. L., Neto, L. C. F., Saraiva, M. M., Sales, M. H., Salgado, M. P. G., Vasconcelos, R., Galano, S., Mesquita, V. V., and Azevedo, T.:
Reconstructing Three Decades of Land Use and Land Cover Changes in Brazilian
Biomes with Landsat Archive and Earth Engine, Remote Sens., 12, 2735,
<a href="https://doi.org/10.3390/rs12172735" target="_blank">https://doi.org/10.3390/rs12172735</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Storn and Price(1997)</label><mixed-citation>
      
Storn, R. and Price, K.: Differential Evolution – A Simple and Efficient
Heuristic for Global Optimization over Continuous Spaces, J. Global Optimiz., 11, 341–359, <a href="https://doi.org/10.1023/a:1008202821328" target="_blank">https://doi.org/10.1023/a:1008202821328</a>, 1997.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Vestena and Kobiyama(2007)</label><mixed-citation>
      
Vestena, L. R. and Kobiyama, M.: Water balance in karst: case study of the
Ribeirão da Onça catchment in Colombo City, Paraná State-Brazil, Brazil. Arch. Biol. Technol., 50, 905–912, 2007.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib59"><label>Viviroli et al.(2009)Viviroli, Mittelbach, Gurtz, and
Weingartner</label><mixed-citation>
      
Viviroli, D., Mittelbach, H., Gurtz, J., and Weingartner, R.: Continuous
simulation for flood estimation in ungauged mesoscale catchments of Switzerland – Part II: Parameter regionalisation and flood estimation results, J. Hydrol., 377, 208–225, <a href="https://doi.org/10.1016/j.jhydrol.2009.08.022" target="_blank">https://doi.org/10.1016/j.jhydrol.2009.08.022</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib60"><label>Wilks(2011)</label><mixed-citation>
      
Wilks, D. S.: Statistical Methods in the Atmospheric Sciences, Academic
Press, ISBN 0123850223,
<a href="https://www.ebook.de/de/product/14751307/daniel_s_wilks_statistical_methods_in_the_atmospheric_sciences_100.html" target="_blank"/>
(last access: 1 December 2023), 2011.

    </mixed-citation></ref-html>--></article>
