<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "https://jats.nlm.nih.gov/nlm-dtd/publishing/3.0/journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0" article-type="research-article">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">NHESS</journal-id><journal-title-group>
    <journal-title>Natural Hazards and Earth System Sciences</journal-title>
    <abbrev-journal-title abbrev-type="publisher">NHESS</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Nat. Hazards Earth Syst. Sci.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1684-9981</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/nhess-26-3417-2026</article-id><title-group><article-title>Deep Learning Emulation of Multivariate Climate Indices: A Case Study of the Fire Weather Index in the Iberian Peninsula</article-title><alt-title>Deep learning emulation of multivariate climate indices</alt-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Mirones</surname><given-names>Óscar</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff2 aff3">
          <name><surname>Bedia</surname><given-names>Joaquín</given-names></name>
          
        <ext-link>https://orcid.org/0000-0001-6219-4312</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff4">
          <name><surname>Soares</surname><given-names>Pedro M. M.</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Gutiérrez</surname><given-names>José M.</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="yes" rid="aff1 aff5">
          <name><surname>Baño-Medina</surname><given-names>Jorge</given-names></name>
          <email>bmedina@ifca.unican.es</email>
        <ext-link>https://orcid.org/0000-0003-3380-1579</ext-link></contrib>
        <aff id="aff1"><label>1</label><institution>Instituto de Física de Cantabria (IFCA), CSIC-Universidad de Cantabria, Santander, Spain</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>Dept. Matemática Aplicada y Ciencias de la Computación (MACC), Universidad de Cantabria, Santander, Spain</institution>
        </aff>
        <aff id="aff3"><label>3</label><institution>Grupo de Meteorología y Computación, Universidad de Cantabria, Unidad Asociada al CSIC, Santander, Spain</institution>
        </aff>
        <aff id="aff4"><label>4</label><institution>Instituto Dom Luiz (IDL) – Faculdade de Ciências da Universidade de Lisboa (FCUL),  Campo Grande Edifício C8, Piso 3, 1749-016 Lisboa, Portugal</institution>
        </aff>
        <aff id="aff5"><label>5</label><institution>Center for Western Weather and Water Extremes, Scripps Institution of Oceanography,  University of California San Diego, San Diego, CA, USA</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Jorge Baño-Medina (bmedina@ifca.unican.es)</corresp></author-notes><pub-date><day>22</day><month>July</month><year>2026</year></pub-date>
      
      <volume>26</volume>
      <issue>7</issue>
      <fpage>3417</fpage><lpage>3442</lpage>
      <history>
        <date date-type="received"><day>19</day><month>May</month><year>2025</year></date>
           <date date-type="rev-request"><day>10</day><month>July</month><year>2025</year></date>
           <date date-type="rev-recd"><day>18</day><month>June</month><year>2026</year></date>
           <date date-type="accepted"><day>10</day><month>July</month><year>2026</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2026 Óscar Mirones et al.</copyright-statement>
        <copyright-year>2026</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026.html">This article is available from https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026.html</self-uri><self-uri xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026.pdf">The full text article is available as a PDF file from https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026.pdf</self-uri>
      <abstract><title>Abstract</title>

      <p id="d2e148">The Fire Weather Index (FWI) is an essential multivariate climate index for assessing wildfire risk and the associated impacts of climate change, as it provides a quantitative measure of wildfire danger by integrating different critical near-surface fire-weather variables, namely air temperature, relative humidity, wind speed, and precipitation. FWI calculation depends on instantaneous data representing noon local standard times, which are often unavailable in many climate data repositories – particularly in climate projections. In these instances, a “Proxy” of actual FWI is often used, applying the same FWI formulation to daily aggregated values (mean, max, or min), despite known limitations in capturing extremes and temporal dynamics.</p>

      <p id="d2e151">This study investigates the use of deep learning (DL) models to emulate the reference FWI over the Iberian Peninsula – a predominantly Mediterranean and fire-prone region – using only daily inputs. The emulators are trained and evaluated using ERA5-Land data, which, while not observational ground truth, provides a consistent and high-resolution dataset suitable for controlled inter-comparison. The focus is not on validating FWI against observations, but on assessing the ability of DL models to reproduce the reference FWI more accurately than traditional proxy approaches, using the same input data source.</p>

      <p id="d2e154">Our results show substantial improvements in spatial accuracy, preservation of temporal sequences, and detection of extreme fire danger events when compared with the corresponding Proxy version. Furthermore, after evaluating different combinations of input variables for DL model training, we find that precipitation can be excluded without substantially affecting accuracy – especially for the upper part of the distribution – an important insight given the challenges climate models face in representing precipitation. These findings highlight the potential of deep learning tools to enhance the usability of FWI in contexts where sub-daily data are unavailable, and set the stage for the emulation of other multivariate climate indices, which are vital for climate impact studies, spatial planning and management, and adaptation decision-making.</p>
  </abstract>
    
<funding-group>
<award-group id="gs1">
<funding-source>Ministerio de Ciencia e Innovación</funding-source>
<award-id>PRE2021-100292</award-id>
</award-group>
<award-group id="gs2">
<funding-source>Ministerio de Ciencia e Innovación</funding-source>
<award-id>PID2023-149997OA-420 I00</award-id>
</award-group>
<award-group id="gs3">
<funding-source>Fundação para a Ciência e a Tecnologia</funding-source>
<award-id>2022.09185.PTDC</award-id>
</award-group>
<award-group id="gs4">
<funding-source>Fundação para a Ciência e a Tecnologia</funding-source>
<award-id>UID/50019/2025</award-id>
<award-id>LA/P/0068/2020</award-id>
</award-group>
</funding-group>
</article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d2e166">Fire is a global phenomenon that has shaped ecosystems since the emergence of land vegetation, influencing vegetation structure, the carbon cycle, and climate <xref ref-type="bibr" rid="bib1.bibx11" id="paren.1"/>. From a human perspective, wildfires can cause severe ecological damage, threaten lives and infrastructure, and degrade air and water quality, with far-reaching social and economic consequences <xref ref-type="bibr" rid="bib1.bibx12 bib1.bibx45" id="paren.2"/>.</p>
      <p id="d2e175">Fire danger indices are essential tools for wildfire risk assessment and climate impact studies. By integrating key atmospheric variables – such as temperature, humidity, wind speed, and precipitation – into a single, physically interpretable indicator <xref ref-type="bibr" rid="bib1.bibx24 bib1.bibx58" id="paren.3"/>, these indices support fire management agencies in resource allocation and early warning systems <xref ref-type="bibr" rid="bib1.bibx19" id="paren.4"/>. They also help researchers understand how climate change alters fire-weather patterns <xref ref-type="bibr" rid="bib1.bibx7 bib1.bibx35 bib1.bibx10 bib1.bibx50 bib1.bibx25" id="paren.5"/> informing adaptation strategies and policy decisions in fire-prone regions <xref ref-type="bibr" rid="bib1.bibx17 bib1.bibx18" id="paren.6"/>.</p>
      <p id="d2e190">Among the various fire-weather indices worldwide, the Canadian Fire Weather Index <xref ref-type="bibr" rid="bib1.bibx55" id="paren.7"><named-content content-type="pre">FWI,</named-content></xref> is one of the most widely adopted globally. In Europe, it underpins the European Forest Fire Information System <xref ref-type="bibr" rid="bib1.bibx48" id="paren.8"><named-content content-type="pre">EFFIS</named-content></xref> and has been extensively used for the estimation of fire danger in vulnerable regions <xref ref-type="bibr" rid="bib1.bibx15" id="paren.9"/>, such as Iberia <xref ref-type="bibr" rid="bib1.bibx50" id="paren.10"/>, as well as for the assessment of future climate change scenarios in the Euro-Mediterranean region <xref ref-type="bibr" rid="bib1.bibx5 bib1.bibx23" id="paren.11"/> and the Iberian Peninsula <xref ref-type="bibr" rid="bib1.bibx10" id="paren.12"/>.</p>
      <p id="d2e216">Although the adoption of the FWI presents clear advantages for effectively characterizing fire danger weather situations, it also poses some hurdles from an implementation point of view. One significant challenge is that the FWI was originally designed for noon local standard time (LST) conditions, representing the highest fire danger time in mid-latitude regions. This means it requires instantaneous noon time input data of near-surface air temperature, relative humidity, 10-m wind speed and last 24 h  accumulated precipitation <xref ref-type="bibr" rid="bib1.bibx37" id="paren.13"/>. However, such specific noon data are often missing from observational records. Furthermore, climate model databases quite often do not include sub-daily outputs due to the massive storage requirements <xref ref-type="bibr" rid="bib1.bibx16" id="paren.14"><named-content content-type="pre">e.g. the Earth System Grid Federation –ESGF–,</named-content></xref>, thus preventing the direct calculation of FWI from climate model simulations.</p>
      <p id="d2e228">To circumvent the lack of instantaneous data, daily mean aggregated inputs have been commonly used for FWI estimation, in particular to obtain future projections from model simulations. However, the resulting daily mean FWI proxies cannot be reliably transformed to match their instantaneous counterparts, leading to inconsistent future projections <xref ref-type="bibr" rid="bib1.bibx29" id="paren.15"/>. Consequently, approximations are often employed to minimize FWI distortion. One of the pioneering studies in this area utilized different daily <italic>proxy</italic> variables for FWI calculation from climate model outputs, comparing various combinations of minimum relative humidity and maximum temperature as surrogates for noon-time outputs. This approach aimed to minimize the distortion of the climate change signal in the reference FWI, using simulations from a regional model for both historical and a moderate future warming scenario <xref ref-type="bibr" rid="bib1.bibx5" id="paren.16"/>. Since then, other studies have adopted proxies to produce more accurate FWI projections <xref ref-type="bibr" rid="bib1.bibx1 bib1.bibx10 bib1.bibx40" id="paren.17"/>, as a practical solution given the unavailability of more precise data. Despite all these efforts, the approximate version of the FWI inevitably presents certain deviations from the original version that introduce added uncertainty <xref ref-type="bibr" rid="bib1.bibx5 bib1.bibx40" id="paren.18"/>.</p>
      <p id="d2e246">Emulation offers a promising alternative in this context. Recent advancements in machine learning techniques, particularly deep learning (DL), across various fields suggest that these sophisticated nonlinear models have the potential to effectively represent and capture the intricate structures and nonlinear dynamics inherent in Earth systems <xref ref-type="bibr" rid="bib1.bibx22" id="paren.19"/>, including numerical model climate predictions <xref ref-type="bibr" rid="bib1.bibx47" id="paren.20"/> and applications to downscaling <xref ref-type="bibr" rid="bib1.bibx2 bib1.bibx14 bib1.bibx38 bib1.bibx34 bib1.bibx51" id="paren.21"/>, among others. In this new framework, the emulation function is determined by the neural network's coefficients, allowing for varying topologies to produce different plausible fields <xref ref-type="bibr" rid="bib1.bibx3" id="paren.22"/>. This opens up a broad area of research where various techniques and architectures can be compared to evaluate their suitability for specific problems. It also introduces new challenges related to model interpretability, requiring methods of eXplainable Artificial Intelligence (XAI) that reveal how DL models make decisions. Such approaches play a key role in establishing the credibility of the outcomes among users <xref ref-type="bibr" rid="bib1.bibx21" id="paren.23"/> and for gaining valuable insights into the most influential variables and the physical interpretation of model's functioning <xref ref-type="bibr" rid="bib1.bibx57 bib1.bibx4" id="paren.24"/>. Among the variety of approaches available, saliency maps emerge as valuable tools aiding in the design and evaluation of DL climate downscaling approaches in general <xref ref-type="bibr" rid="bib1.bibx28" id="paren.25"/>, and have been successfully used to unravel DL-based FWI predictions in previous studies <xref ref-type="bibr" rid="bib1.bibx42" id="paren.26"/>.</p>
      <p id="d2e274">In this study, we explore the emulation of the Fire Weather Index (FWI) using DL models to approximate the index's behavior as closely as possible with commonly available input variables whose temporal frequency (daily) does not match the definition of the “reference FWI” (based on instantaneous noon local standard times). We build upon the reference FWI and the optimal Proxy FWI version presented in <xref ref-type="bibr" rid="bib1.bibx5" id="text.27"/>, which has been widely used in subsequent studies. We compare the emulation of the reference FWI using different combinations of input variables (hereafter predictor sets), including inadequate (daily mean), incomplete (removing some input variables), or suboptimal temporal frequency inputs (i.e. daily minimum relative humidity), which are more commonly available than the required instantaneous data. We test various DL model architectures and topologies and utilize XAI techniques, such as saliency maps, to provide end users and practitioners with a physical understanding of the emulation process.</p>
      <p id="d2e280">Our results indicate that DL emulators outperform the traditional Proxy approach in all validation aspects, including spatial representation, preservation of the temporal sequence, and detection of extreme fire danger conditions. The daily mean inputs, which are most widely available in public repositories, are adequate for accurately emulating the reference FWI using the developed DL methods. Furthermore, our findings suggest that precipitation can be omitted from the predictor set without significantly compromising FWI representation accuracy. Overall, we demonstrate that by leveraging simpler inputs, DL emulators can enhance the accessibility and applicability of FWI in impact studies, facilitating a more efficient and widespread use in climate impact studies and decision-making.</p>
      <p id="d2e283">In addition, we present a concise set of preliminary experiments on the emulation of the individual fuel-moisture codes (FFMC, DMC, DC). These initial results confirm that FFMC and DMC can be emulated with high fidelity using daily inputs – even when precipitation is excluded – whereas the DC shows a somewhat stronger sensitivity to the presence of precipitation in the predictor set due to its long-term moisture memory. Nonetheless, given that the primary objective of this first study is the emulation of the FWI itself, which integrates the combined effect of all sub-indices and is the most widely used fire-danger indicator in climate and impact research, we restrict our detailed analysis to the FWI while deferring a comprehensive component-level evaluation to a forthcoming companion work.</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Data and Methods</title>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Input Data</title>
      <p id="d2e301">All data used in this study were obtained from the ERA5-Land database <xref ref-type="bibr" rid="bib1.bibx44" id="paren.28"/>, distributed by the Climate Data Store (CDS) of the Copernicus Climate Change Service <xref ref-type="bibr" rid="bib1.bibx13" id="paren.29"/>. ERA5-Land is a high-resolution reanalysis dataset produced by the European Centre for Medium-Range Weather Forecasts (ECMWF). It offers detailed information on land surface variables, providing a consistent view of their evolution over several decades. ERA5-Land features a finer spatial resolution of approximately 9 km grid spacing, compared to ERA5 <xref ref-type="bibr" rid="bib1.bibx30" id="paren.30"><named-content content-type="pre"><inline-formula><mml:math id="M1" display="inline"><mml:mrow><mml:mo>∼</mml:mo><mml:mn mathvariant="normal">25</mml:mn></mml:mrow></mml:math></inline-formula> km</named-content></xref>. It covers data from January 1950 to the present, with an hourly temporal resolution. The dataset includes various land surface parameters (such as soil moisture, soil temperature, snow cover, and surface runoff) to control the simulated land fields, ensuring the data remains accurate and consistent. ERA5-Land uses atmospheric variables from ERA5, like air temperature, wind speed and humidity, and presents hourly temporal resolution, thus providing the necessary input variables for the calculation of both reference and Proxy FWI versions, as well as the different predictor sets used in this study (Sect. <xref ref-type="sec" rid="Ch1.S2.SS3"/>). </p>
</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>Fire Weather Index calculation</title>
      <p id="d2e335">The reference noon LST FWI is calculated from ERA5-Land daily records of four near-surface meteorological variables measured at 12 UTC (<inline-formula><mml:math id="M2" display="inline"><mml:mo lspace="0mm">∼</mml:mo></mml:math></inline-formula> noon time in the Iberian Peninsula): 24 h  accumulated precipitation, instantaneous wind speed, relative humidity, and temperature. These inputs are processed via a set of empirical equations that yield six intermediate components characterizing fuel moisture dynamics across different fuel layers <xref ref-type="bibr" rid="bib1.bibx55 bib1.bibx52" id="paren.31"/>, the influence of wind on fire spread, and the total fuel available for combustion. These components are then integrated to compute the FWI, a dimensionless index representing potential fire intensity under given meteorological conditions for a reference fuel type (mature pine stands).</p>
      <p id="d2e348">Beyond the four meteorological inputs required at noon LST, the Canadian FWI System relies on six sequential sub-indices that describe fuel moisture conditions and potential fire behavior <xref ref-type="bibr" rid="bib1.bibx55 bib1.bibx52 bib1.bibx37" id="paren.32"/>. Three of them (FFMC, DMC and DC) represent moisture contents at increasing depths of the forest floor and respond to atmospheric drivers at markedly different timescales.</p>
      <p id="d2e354">The <italic>Fine Fuel Moisture Code</italic> (FFMC) characterizes the moisture content of surface litter and fine fuels, which react on hourly–daily timescales to changes in temperature, relative humidity and wind speed, as well as to even small rainfall amounts due to their limited water-holding capacity. FFMC largely governs ignition likelihood and short-term fluctuations in fire danger.</p>
      <p id="d2e360">The <italic>Duff Moisture Code</italic> (DMC) represents the moisture content of moderately deep organic layers in the upper duff. Its evolution reflects medium-term drying (days to weeks), driven primarily by temperature and relative humidity, and decreases only when precipitation exceeds an effective infiltration threshold. The DMC modulates the potential for sustained fire spread once ignition occurs.</p>
      <p id="d2e367">The <italic>Drought Code</italic> (DC) reflects the long-term moisture content of deep, compact organic horizons that respond over weeks to months. The DC accumulates persistent atmospheric moisture deficits and only decreases when substantial or prolonged precipitation is able to percolate into deeper layers. As such, it is a key indicator of seasonal-scale drought and deep-burning potential. These different memory timescales explain the unequal sensitivity of the moisture codes to the meteorological drivers <xref ref-type="bibr" rid="bib1.bibx56" id="paren.33"/>.</p>
      <p id="d2e376">The three moisture codes are combined into two intermediate fire-behavior indices. The <italic>Initial Spread Index</italic> (ISI) combines FFMC with wind speed to estimate the expected rate of fire spread immediately after ignition. The <italic>Buildup Index</italic> (BUI) integrates DMC and DC to represent the total amount of fuel available for combustion, accounting for both medium- and long-term moisture deficits. Finally, the <italic>Fire Weather Index</italic> (FWI) combines ISI and BUI through a nonlinear formulation to estimate potential fire intensity for a standard fuel type.</p>
      <p id="d2e388">A practical consideration arises from the different response times and memory effects of these components. While the FWI system uses default initial values, leading to a short spin-up period, this transience is typically negligible during the fire season due to rapid stabilization and the minimal influence of snowmelt on moisture inputs <xref ref-type="bibr" rid="bib1.bibx8" id="paren.34"/>.</p>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>Predictors Sets</title>
      <p id="d2e402">The reference FWI is computed directly from instantaneous meteorological variables at 12:00 UTC (approximately noon local standard time in Iberia), as defined in the standard FWI formulation (see Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>). In contrast, the predictor sets used to train the DL emulators (P0, P1, P2) consist of different temporal aggregations and/or preprocessing of these same meteorological variables. Specifically, the emulators are provided with daily aggregated variables (daily means, daily minima/maxima, together with 24 h accumulated precipitation) rather than instantaneous noon values. This distinction is essential for interpreting emulator performance: the DL models learn to recover a noon-time fire weather signal from daily-aggregated inputs, which is an inherently different (and harder) task than the direct FWI computation. Reported errors reflect this added complexity and should not be interpreted as a simple approximation error.</p>
      <p id="d2e407">To emulate the reference FWI, we consider several predictor sets derived from ERA5-Land, summarized in Table <xref ref-type="table" rid="T1"/>. The initial experiments use P0, which builds on the same set of variables used in the optimal FWI approach from <xref ref-type="bibr" rid="bib1.bibx5" id="text.35"/> (hereafter Proxy FWI). This allows us to assess whether the DL models can more accurately replicate the FWI transfer function when provided with the same predictors as the Proxy FWI. Accordingly, P0 includes 24 h  accumulated precipitation, daily mean air temperature and wind speed, and minimum relative humidity. Proxy FWI is derived from the standard FWI formulation using the latter set of variables. The use of minimum relative humidity in P0 ensures the fairest possible comparison with Proxy FWI. However, in predictor sets P1 and P2, we intentionally replace minimum relative humidity with daily mean relative humidity. This is motivated by the fact that daily mean variables are more consistently provided by climate model outputs and reanalysis datasets, and we wanted to examine the performance of the emulator under such predictor-limited cases. Precipitation poses an additional challenge: it is one of the most difficult variables for climate models and reanalyses to represent reliably, due to its strong spatial and temporal variability, its dependence on complex physical processes, and the influence of multiple interacting factors. To examine whether robust performance can still be achieved in its absence, we exclude precipitation from predictor sets P1 and P2. Finally, in P2 we tried measuring the impact of using wind speed module versus wind speed zonal components, which could be relevant when determining fire danger. Other predictors sets have been tested, such as P1 using minimum relative humidity instead daily mean relative humidity or P0 including the 24 h precipitation, without any added value compared to the configurations mentioned in the manuscript. These experiment results are shown in Fig. <xref ref-type="fig" rid="FD2"/>.</p>

<table-wrap id="T1" specific-use="star"><label>Table 1</label><caption><p id="d2e420">Summary of the predictor sets assessed in this study. 12:00 UTC correspond to instantaneous model outputs verifying at that time. DM corresponds to Daily Mean values, and Max/Min to Daily Maximum/Minimum values. Precipitation is the daily (last 24 h) accumulated value, from 12:00 to 12:00 UTC. Cells marked with <inline-formula><mml:math id="M3" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula> indicate that the variable is not included in the corresponding predictor set. The “FWI” row indicates the instantaneous noon local standard time (LST) variables required by the standard FWI definition. The predictor sets P0, P1, P2 represent temporal aggregations or modifications of the noon LST variables, such as daily means, which are far more widely available across reanalysis and climate datasets. This motivates the emulation approach: rather than requiring instantaneous noon inputs, the emulators allow FWI to be derived directly from these more accessible variable forms.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="8">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="justify" colwidth="4cm" colsep="1"/>
     <oasis:colspec colnum="3" colname="col3" align="left"/>
     <oasis:colspec colnum="4" colname="col4" align="left"/>
     <oasis:colspec colnum="5" colname="col5" align="justify" colwidth="2.5cm"/>
     <oasis:colspec colnum="6" colname="col6" align="justify" colwidth="1.5cm"/>
     <oasis:colspec colnum="7" colname="col7" align="left"/>
     <oasis:colspec colnum="8" colname="col8" align="left"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col2" align="center" colsep="1">Predictors sets </oasis:entry>
         <oasis:entry namest="col3" nameend="col8" align="center">Variable </oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2"/>
         <oasis:entry colname="col3"/>
         <oasis:entry colname="col4"/>
         <oasis:entry colname="col5"/>
         <oasis:entry colname="col6"/>
         <oasis:entry colname="col7">Eastward</oasis:entry>
         <oasis:entry colname="col8">Northward</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1"/>
         <oasis:entry colname="col2">Description</oasis:entry>
         <oasis:entry colname="col3">Temp.</oasis:entry>
         <oasis:entry colname="col4">Rel. Hum.</oasis:entry>
         <oasis:entry colname="col5">Precipitation</oasis:entry>
         <oasis:entry colname="col6">Wind speed</oasis:entry>
         <oasis:entry colname="col7">wind</oasis:entry>
         <oasis:entry colname="col8">wind</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">FWI</oasis:entry>
         <oasis:entry colname="col2">Reference FWI definition  <xref ref-type="bibr" rid="bib1.bibx55" id="paren.36"/></oasis:entry>
         <oasis:entry colname="col3">12:00 UTC</oasis:entry>
         <oasis:entry colname="col4">12:00 UTC</oasis:entry>
         <oasis:entry colname="col5">24 h accumulated (12:00–12:00 UTC)</oasis:entry>
         <oasis:entry colname="col6">12:00 UTC</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M4" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col8"><inline-formula><mml:math id="M5" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">P0</oasis:entry>
         <oasis:entry colname="col2">“Best” Proxy FWI  <xref ref-type="bibr" rid="bib1.bibx5" id="paren.37"/></oasis:entry>
         <oasis:entry colname="col3">DM</oasis:entry>
         <oasis:entry colname="col4">Min</oasis:entry>
         <oasis:entry colname="col5">24 h accumulated (12:00–12:00 UTC)</oasis:entry>
         <oasis:entry colname="col6">DM</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M6" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col8"><inline-formula><mml:math id="M7" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">P1</oasis:entry>
         <oasis:entry colname="col2">Daily Mean inputs without precipitation</oasis:entry>
         <oasis:entry colname="col3">DM</oasis:entry>
         <oasis:entry colname="col4">DM</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M8" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6">DM</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M9" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col8"><inline-formula><mml:math id="M10" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">P2</oasis:entry>
         <oasis:entry colname="col2">Daily Mean inputs with wind components</oasis:entry>
         <oasis:entry colname="col3">DM</oasis:entry>
         <oasis:entry colname="col4">DM</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M11" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6">DM</oasis:entry>
         <oasis:entry colname="col7">DM</oasis:entry>
         <oasis:entry colname="col8">DM</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S2.SS4">
  <label>2.4</label><title>Deep Learning Methods</title>
      <p id="d2e689">In this section, we briefly describe the DL architectures trained in this study. A visual summary of their topologies is presented in Fig. <xref ref-type="fig" rid="F1"/>. The selected architectures include the Fully Connected Dense (FCD) model, chosen as benchmark for its relative simplicity, as well as DeepESD and U-Net, which have been recently introduced in the literature for climate downscaling <xref ref-type="bibr" rid="bib1.bibx2 bib1.bibx28 bib1.bibx27" id="paren.38"/>. Alternative architectures, such as Convolutional Long Short-Term Memory (ConvLSTM), were also evaluated. However, owing to their inferior performance relative to the selected architectures and their greater computational demands, they are not presented in this manuscript. The DL models are tasked to learn a mapping between different sets of FWI-related input variables (Table <xref ref-type="table" rid="T1"/>) and the reference FWI over the Iberian Peninsula as output, by minimizing a loss function (Sect. <xref ref-type="sec" rid="Ch1.S2.SS4.SSS4"/>). The model output is an emulated version of the reference ERA5-Land FWI at the same spatial resolution than its inputs. All models are trained with the same optimization parameters independently of the architecture, using the Adam optimizer with a learning rate set to 0.0001, a batch size of 64 and a maximum of 1000 epochs. An early stopping mechanism is set to prevent overfitting, which stops the training process if there is no improvement in the validation loss after 30 consecutive epochs. Furthermore, the input data are standardized as part of the model preprocessing:

            <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M12" display="block"><mml:mrow><mml:mi>x</mml:mi><mml:msub><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M13" display="inline"><mml:mrow><mml:mi>x</mml:mi><mml:msup><mml:mi/><mml:mo>′</mml:mo></mml:msup></mml:mrow></mml:math></inline-formula> represents the standardized value, <inline-formula><mml:math id="M14" display="inline"><mml:mi>x</mml:mi></mml:math></inline-formula> represents the raw (unstandardized) value, and <inline-formula><mml:math id="M15" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M16" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> represents the mean and standard deviation at gridpoint <inline-formula><mml:math id="M17" display="inline"><mml:mi>i</mml:mi></mml:math></inline-formula>. Parameters <inline-formula><mml:math id="M18" display="inline"><mml:mi mathvariant="italic">μ</mml:mi></mml:math></inline-formula> and <inline-formula><mml:math id="M19" display="inline"><mml:mi mathvariant="italic">σ</mml:mi></mml:math></inline-formula> have been computed relative to the training period 1979–2017 (per gridpoint).</p>

      <fig id="F1" specific-use="star"><label>Figure 1</label><caption><p id="d2e809">Schematic representation of the DL architectures used in this study, showcasing the different types of layers and their corresponding dimensions. The figure also highlights the distinct sets of predictors used as training inputs.</p></caption>
          <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f01.png"/>

        </fig>

<sec id="Ch1.S2.SS4.SSS1">
  <label>2.4.1</label><title>Fully Connected Dense</title>
      <p id="d2e825">The Fully Connected Dense (FCD) model is a DL architecture that relies exclusively on fully connected layers to process spatial input data. The input is first passed through two consecutive dense layers, each containing 50 neurons, with Rectified Linear Unit (ReLU) activation functions <xref ref-type="bibr" rid="bib1.bibx26" id="paren.39"/> applied to introduce non-linearity and enhance feature extraction. Following these hidden layers, the transformed feature representation is fed into a fully connected layer with 12 800 neurons, utilizing a linear activation function to produce the final output. This output is then reshaped to match the dimensions of the predictand, ensuring consistency with the target variable.</p>
</sec>
<sec id="Ch1.S2.SS4.SSS2">
  <label>2.4.2</label><title>DeepESD</title>
      <p id="d2e840">The DeepESD model <xref ref-type="bibr" rid="bib1.bibx2" id="paren.40"/> is designed as a combination of convolutional and dense layers envisaged to efficiently process spatial data. It consists of three convolutional layers with 50, 25, and 10 kernels, each using ReLU activation functions (Fig. <xref ref-type="fig" rid="F1"/> for more details). After the final convolutional operation, the resulting feature maps are flattened into a one-dimensional vector, which is subsequently passed through a fully connected dense layer comprising 12 800 neurons, reshaping to match the dimensions of the predictand.</p>
</sec>
<sec id="Ch1.S2.SS4.SSS3">
  <label>2.4.3</label><title>U-Net</title>
      <p id="d2e856">The U-Net model adopts an encoder-decoder architecture designed for spatial data processing. The encoder progressively extracts hierarchical features through a series of Convolutional Blocks (ConvBlocks), each containing convolutional layers with ReLU activation, followed by <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:mn mathvariant="normal">2</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:math></inline-formula> max pooling operations that reduce spatial dimensions while increasing the number of feature maps. The encoder expands from 64 to 512 channels, while reducing the spatial resolution from (80, 160) to (10, 20). In the decoder, the model reconstructs the original resolution using 2<inline-formula><mml:math id="M21" display="inline"><mml:mo>×</mml:mo></mml:math></inline-formula>2 transposed convolutions (deconvolutions) that progressively upsample feature maps <xref ref-type="bibr" rid="bib1.bibx46" id="paren.41"/>. This is followed by convolutional layers that refine the output. The final layer consists of a <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> convolution, followed by a linear activation function, producing an output of dimensions (1, 80, 160).</p>
</sec>
<sec id="Ch1.S2.SS4.SSS4">
  <label>2.4.4</label><title>Loss functions</title>
      <p id="d2e901">The DL models were fitted with two alternative loss functions: the mean squared error (MSE), which serves as a standard loss function, and an asymmetric loss function (ASYM), designed to better capture extreme events, as introduced by <xref ref-type="bibr" rid="bib1.bibx20" id="text.42"/>. The MSE is defined as the average of the squared differences between observed and predicted values. This metric penalizes larger errors more severely, making it useful for many regression tasks. However, MSE might not adequately emphasize errors associated with extreme events. To address this limitation, we implemented an asymmetric loss function defined as:

              <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M23" display="block"><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mi mathvariant="italic">θ</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mn mathvariant="normal">1</mml:mn><mml:mi>N</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mo>|</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>i</mml:mi></mml:msub><mml:mo>|</mml:mo><mml:mo>+</mml:mo><mml:msup><mml:mi mathvariant="italic">γ</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup><mml:mo>×</mml:mo><mml:mi mathvariant="normal">max</mml:mi><mml:mfenced open="(" close=")"><mml:mrow><mml:mn mathvariant="normal">0</mml:mn><mml:mo>,</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mover accent="true"><mml:mi>y</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfenced></mml:mrow></mml:math></disp-formula>

            where <inline-formula><mml:math id="M24" display="inline"><mml:mrow><mml:mi mathvariant="italic">γ</mml:mi><mml:mo>=</mml:mo><mml:mi>G</mml:mi><mml:mo>(</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, where <inline-formula><mml:math id="M25" display="inline"><mml:mi>G</mml:mi></mml:math></inline-formula> is the gamma cumulative distribution function (CDF) fitted to the time series of every grid point in the training dataset. This loss function weights the mean absolute error (MAE) by an amount proportional to the extremity of each value, thus assigning a higher penalty when the model underestimates extreme values. This tailored approach improves the evaluation and prediction of extreme events in the dataset.</p>
</sec>
</sec>
<sec id="Ch1.S2.SS5">
  <label>2.5</label><title>Model evaluation</title>
      <p id="d2e1025">In this study, the DL models are trained using daily data from the period 1979 to 2017. To evaluate the performance of the model, 20 % of these data is separated and treated as validation samples for model assessment during the training phase. The final results are presented for an independent test period spanning 2018–2021. We also evaluated longer test periods (2012–2021), which required shortening the training phase to 1979–2011. Since these tests produced robust and consistent results comparable to those for 2018–2021, they are not included in the text. To assess the performance of the DL models and the reproducibility of spatial and temporal patterns in the emulated results, several validation indices are used, each focusing on different spatial and temporal aspects of the predicted series (Table <xref ref-type="table" rid="T2"/>).</p>

<table-wrap id="T2" specific-use="star"><label>Table 2</label><caption><p id="d2e1033">Validation indices assessed in the study.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="2">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Metric</oasis:entry>
         <oasis:entry colname="col2">Description</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">MAE FWI</oasis:entry>
         <oasis:entry colname="col2">Mean Absolute Error for the FWI</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">MAE FWI95</oasis:entry>
         <oasis:entry colname="col2">Mean Absolute Error for the FWI percentile 95th</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Freq. FWI95 Relative Bias</oasis:entry>
         <oasis:entry colname="col2">Relative bias of the frequency of extreme FWI events (over 95th percentile)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Max Spell95 Bias</oasis:entry>
         <oasis:entry colname="col2">Inter-annual maximum spell length bias for extreme FWI events (over 95th percentile)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">ROC AUC</oasis:entry>
         <oasis:entry colname="col2">Receiver-operating characteristic (ROC) curve Area for extreme FWI events (over 95th percentile)</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1"><inline-formula><mml:math id="M26" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col2">Coefficient of determination</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d2e1124">At each grid point, validation indices are computed, and their spatial distributions are depicted in the maps presented in various figures in Sect. <xref ref-type="sec" rid="Ch1.S3"/>. These indices span different key characteristics of the predicted series to assess how models emulate FWI, such as the mean FWI and extreme FWI distribution, or the length of spells for extreme FWI events. Instead of using MAE (for FWI and FWI95) or bias (for Max Spell95), we assess relative bias with respect to the frequency of FWI95, given by the following formula:

            <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M27" display="block"><mml:mrow><mml:mtext>Relative bias</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mi>X</mml:mi><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mi>X</mml:mi><mml:mo stretchy="false" mathvariant="normal">^</mml:mo></mml:mover></mml:mrow><mml:mi>X</mml:mi></mml:mfrac></mml:mstyle><mml:mo>×</mml:mo><mml:mn mathvariant="normal">100</mml:mn></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M28" display="inline"><mml:mi>X</mml:mi></mml:math></inline-formula> represents the observed and <inline-formula><mml:math id="M29" display="inline"><mml:mover accent="true"><mml:mi>X</mml:mi><mml:mo mathvariant="normal" stretchy="false">^</mml:mo></mml:mover></mml:math></inline-formula> the estimated value.</p>
      <p id="d2e1175">We also analyze the Receiver Operating Characteristic (ROC) curve, and the area Under the Curve (AUC), a graphical tool for evaluating the performance of binary classification models <xref ref-type="bibr" rid="bib1.bibx31" id="paren.43"/>, in this case applied to assess the DL model's ability to discriminate dangerous FWI events. It plots the true positive rate (sensitivity) against the false positive rate at various decision thresholds and quantifies the overall discriminative ability of the model, with values ranging from 0.5 (random classification) to 1.0 (perfect classification). A higher AUC indicates better model performance in distinguishing between positive and negative cases, making the ROC AUC particularly useful in unbalanced classification problems where traditional accuracy measures may be misleading. In particular, the classification evaluated in the ROC AUC analysis pertains to FWI events classified as “extreme” (see Table <xref ref-type="table" rid="T3"/>), defined as those exceeding the 95th percentile of the reference FWI, according to the danger levels established by the Spanish Meteorological Agency (AEMET, <uri>https://www.aemet.es/documentos/es/datos_abiertos/Estadisticas/IM_riesgo_incendios/eimri_generalidades.pdf</uri>, last access: 17 February 2025), being thus impact-relevant for wildfire risk assessment in the Iberian Peninsula. In addition to the validation indices listed in Table <xref ref-type="table" rid="T2"/>, we also use scatter plots the assessment of the emulated FWI 95th percentile against reference FWI climatologies, where the <inline-formula><mml:math id="M30" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> coefficient of determination is used.</p>

<table-wrap id="T3"><label>Table 3</label><caption><p id="d2e1202">FWI Classes According to AEMET (Based on Percentiles). The right column shows the FWI reference spatial mean per category during the fire season.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="3">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Level</oasis:entry>
         <oasis:entry colname="col2">FWI Percentile Range</oasis:entry>
         <oasis:entry colname="col3">Spatial Mean Value</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Low</oasis:entry>
         <oasis:entry colname="col2">Below 40th percentile</oasis:entry>
         <oasis:entry colname="col3">17.57</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Moderate</oasis:entry>
         <oasis:entry colname="col2">40–65th percentile</oasis:entry>
         <oasis:entry colname="col3">42.27</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">High</oasis:entry>
         <oasis:entry colname="col2">65–85th percentile</oasis:entry>
         <oasis:entry colname="col3">55.67</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Very High</oasis:entry>
         <oasis:entry colname="col2">85–95th percentile</oasis:entry>
         <oasis:entry colname="col3">67.02</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Extreme</oasis:entry>
         <oasis:entry colname="col2">Above 95th percentile</oasis:entry>
         <oasis:entry colname="col3">80.70</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S2.SS6">
  <label>2.6</label><title>Model explainability</title>
      <p id="d2e1302">DL models have demonstrated remarkable predictive capabilities across various scientific fields, yet their complex architectures often limit interpretability. This challenge has motivated the growth of eXplainable Artificial Intelligence (XAI), which aims to provide tools for understanding the underlying relationships learned by machine learning models <xref ref-type="bibr" rid="bib1.bibx41" id="paren.44"/>. In this work, we apply saliency-based methods to investigate the internal logic behind our DL predictions. Specifically, we employ a variant of the Integrated Gradients (IG) technique <xref ref-type="bibr" rid="bib1.bibx53" id="paren.45"/>, which has been widely adopted in climate applications <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx36" id="paren.46"/> for its effectiveness and ease of interpretability.</p>
      <p id="d2e1314">Traditional IG computes feature attributions by integrating the gradients of the model output with respect to the input, along a linear path from a baseline input to the actual input. The method is formally expressed as:

            <disp-formula id="Ch1.E4" content-type="numbered"><label>4</label><mml:math id="M31" display="block"><mml:mrow><mml:msub><mml:mi>s</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:munderover><mml:mo movablelimits="false">∫</mml:mo><mml:mn mathvariant="normal">0</mml:mn><mml:mn mathvariant="normal">1</mml:mn></mml:munderover><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mo>∂</mml:mo><mml:mi>f</mml:mi><mml:mfenced close=")" open="("><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:mi>t</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mfenced></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mi mathvariant="normal">d</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the feature value, <inline-formula><mml:math id="M33" display="inline"><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi mathvariant="normal">b</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the baseline (typically zero), and <inline-formula><mml:math id="M34" display="inline"><mml:mi>f</mml:mi></mml:math></inline-formula> is the output of the model.</p>
      <p id="d2e1446">However, in this study we adopt a simplified approximation of IGs, where the integral is omitted and the feature relevance is estimated directly from the gradients at the input point. This results in a computationally simpler, yet informative, saliency map representation. The relevance of each input feature is therefore computed as:

            <disp-formula id="Ch1.E5" content-type="numbered"><label>5</label><mml:math id="M35" display="block"><mml:mrow><mml:msub><mml:mi>s</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>⋅</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:mo>∂</mml:mo><mml:mi>f</mml:mi><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>∂</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mstyle></mml:mrow></mml:math></disp-formula>

          This approach preserves the core idea of attributing the prediction to the input features based on the local sensitivity of the model, while avoiding the computational overhead of path integration. Although this approximation does not fully satisfy all axioms of the original IG method <xref ref-type="bibr" rid="bib1.bibx53" id="paren.47"/>, it has been shown to yield stable and interpretable attributions in practical applications <xref ref-type="bibr" rid="bib1.bibx54" id="paren.48"/>.</p>
      <p id="d2e1505">To address known XAI challenges <xref ref-type="bibr" rid="bib1.bibx39" id="paren.49"/>, the raw saliency maps are post-processed by: (1) taking the absolute value to focus on magnitude only, (2) zeroing all values below 10 % of the per-sample maximum to suppress gradient noise <xref ref-type="bibr" rid="bib1.bibx54" id="paren.50"/>, and (3) normalizing each map so its values sum to one, thus yielding relative contributions. These normalized maps are used for all subsequent explainability analyses.</p>
      <p id="d2e1515">This methodology enables us to quantify the relative importance of each predictor variable in the model’s decision-making process, providing valuable insights into the learned relationships and the adequacy of the different sets of predictors tested, aiding in the underlying physical interpretation of the results.</p>
</sec>
<sec id="Ch1.S2.SS7">
  <label>2.7</label><title>Software</title>
      <p id="d2e1526">Both FWI calculation and emulation are carried out using the R-based <italic>climate4R</italic> framework <xref ref-type="bibr" rid="bib1.bibx32" id="paren.51"/>. In particular, neural network emulators (Dense, DeepESD, U-Net) are defined and trained using the <italic>downscaleR.keras</italic> package <xref ref-type="bibr" rid="bib1.bibx2" id="paren.52"/>, an extension of the <italic>downscaleR</italic> package <xref ref-type="bibr" rid="bib1.bibx9" id="paren.53"/> that leverages Keras/TensorFlow.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Results and Discussion</title>
      <p id="d2e1558">Several DL models are evaluated through an intercomparison of different predictor sets (Table <xref ref-type="table" rid="T1"/>) and the loss functions optimized during DL model training (Sect. <xref ref-type="sec" rid="Ch1.S2.SS4.SSS4"/>). The results presented in this section focus on the main fire season over the Iberian Peninsula <xref ref-type="bibr" rid="bib1.bibx5" id="paren.54"><named-content content-type="pre">June to September, JJAS; see e.g.</named-content></xref> for the test period 2018–2021.</p>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Proxy FWI validation</title>
      <p id="d2e1577">Figure <xref ref-type="fig" rid="F2"/> presents a multi-panel visualization, where the rows correspond to the validation measures shown in Table <xref ref-type="table" rid="T2"/>. From left to right, the first two columns represent the reference FWI and the Proxy FWI climatologies (P0, Table <xref ref-type="table" rid="T1"/>). The last column indicates the MAE or bias between both – depending on the evaluated index, Table <xref ref-type="table" rid="T2"/>–. The FWI95 frequency for the reference FWI is not shown, because, by construction, each grid point in the reference data records exactly 5 % of days above its local 95th percentile threshold during the season. Therefore, the value in each grid across the spatial map is 0.05. It serves as the observed reference for error calculation: The 95th percentile of reference FWI is used as baseline to compute the frequency of Proxy FWI events exceeding this threshold (as well as emulated FWI in the following section).</p>

      <fig id="F2" specific-use="star"><label>Figure 2</label><caption><p id="d2e1590">Reference and Proxy Fire Weather Index (FWI) climatologies for the fire season (June–September) during the test period (2018–2021). These include mean FWI, 95th percentile of FWI (FWI95), FWI95 frequency (number of days exceeding the reference FWI95) and the mean maximum annual spell (the number of days exceeding reference FWI95 Max Spell95). The third column displays the mean absolute error (MAE) for the FWI Mean and FWI95, the bias for Max Spell95 and the relative bias for the FWI95 frequency. The spatial averaged values of MAE or absolute (relative) bias are displayed at the bottom right.</p></caption>
          <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f02.png"/>

        </fig>

      <p id="d2e1599">Focusing on the differences between the Proxy and the reference FWI, we note that there exists greater MAE in Southern Portugal, North-western Spain and North-Eastern Spain, along the Ebro Valley. Conversely, the lowest MAE values (close to zero) are found along the Cantabrian and Mediterranean coasts, northwestern Portugal, the Pyrenees, and the Balearic Islands. These low-MAE areas generally align with the lowest climatological FWI values (as depicted in first column, first row), except for the Balearic Islands, where climatological JJAS FWI remains relatively high (FWI <inline-formula><mml:math id="M36" display="inline"><mml:mo>≈</mml:mo></mml:math></inline-formula> 35).</p>
      <p id="d2e1610">Regarding extreme FWI events (above reference FWI 95th percentile, FWI95), the Proxy FWI95 exhibits the highest MAE in the central-eastern and southeastern continental regions, coincident with very high reference FWI95 climatological values. Although the Proxy FWI95 accurately represents southern Portugal, revealing low MAE values, overall, the most critical FWI95 regions (mostly over Coastal Mediterranean and Central and Iberian Massifs) tend to be underrepresented by the Proxy FWI.</p>
      <p id="d2e1613">The Proxy FWI exhibits a relatively high bias in the frequency of FWI95 events, over continental and coastal Mediterranean regions (negative) as well as the western atlantic façade (positive). This behavior arises from structural differences between the Proxy formulation (based on daily aggregated variables) and the reference FWI, which depends on instantaneous noon-LST conditions. Because extreme fire-weather situations typically occur near midday, the use of daily means and daily minimum/maximum values smooths the sharp sub-daily variations that drive threshold exceedances, leading to systematic over- or under-counting of FWI95 days. Despite this limitation, the Proxy FWI reproduces the main spatial gradients and climatological patterns of mean FWI, FWI95 and Max Spell95, confirming that it remains a useful and physically consistent baseline against which to assess the added value of the deep-learning emulators. Regarding the temporal sequence of extreme–event occurrences, the relative bias for the FWI95 frequency displays a spatial pattern similar to that of the Max Spell95, with overestimation of FWI95 days in Portugal, north–western Spain (Galicia), the Cantabrian mountain range, the Pyrenees and the Ebro Valley, and substantial underestimation across the rest of the Iberian Peninsula. Consistently, the spatial distribution of Max Spell95 shows that the Proxy FWI overestimates spell lengths in Galicia and Portugal (by up to 4 d) while in the continental interior it tends to underestimate them, remaining close to the reference values over most of south–western Spain.</p>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Emulated FWI validation</title>
      <p id="d2e1624">This section presents an intercomparison of the different DL models tested, by evaluating errors relative to the reference FWI and comparing them with the Proxy FWI results described in the previous section, which serves as a benchmark. For brevity, here we present the results using the predictor set P0 as input (Table <xref ref-type="table" rid="T1"/>), considering that the overall intercomparison results are consistent regardless of the predictor set tested. Since P0 includes the variables necessary for computing the Proxy FWI using the standard FWI formulation, our goal is to assess whether different DL architectures are able to capture the reference FWI more accurately using the same input information.</p>
<sec id="Ch1.S3.SS2.SSS1">
  <label>3.2.1</label><title>Prediction errors</title>
      <p id="d2e1636">The prediction errors for the Dense, DeepESD, and U-Net models relative to the reference FWI are presented in Fig. <xref ref-type="fig" rid="F3"/>. The columns correspond to the results of each DL model, and the rows display various validation metrics, including the mean absolute error (MAE) for FWI and FWI95 or the relative bias in the frequency of FWI95. Below the error maps, a scatter plot is displayed per DL model, including the Proxy FWI, comparing the predicted climatological values against the reference FWI95.</p>
      <p id="d2e1641">Before discussing the performance of the DL models across the validation indices, we first highlight how daily aggregation of the input data affects model performance. Figure <xref ref-type="fig" rid="FD1"/> in Appendix <xref ref-type="sec" rid="App1.Ch1.S4"/> presents the results from the Dense, DeepESD, and U-Net models trained using 12:00 UTC input variables (temperature, 24 h accumulated precipitation, relative humidity, and wind speed) to compute the FWI. This sensitivity experiment evaluates the models’ ability to learn the transfer function defining the FWI using inputs at 12:00 UTC, consistent with the temporal resolution of the reference index. This analysis complements our other model configurations by isolating the effect of temporal aggregation. Specifically, we compare performance when models are trained with instantaneous inputs (i.e., values at 12:00 UTC and 24 h precipitation) versus daily mean inputs. This framework allows us to separate the intrinsic biases of each model from the additional error introduced by using daily-aggregated predictors. Figure <xref ref-type="fig" rid="FD1"/> shows that models trained with instantaneous inputs exhibit lower bias and improved accuracy while maintaining a similar spatial error pattern. This confirms that part of the error observed in experiment P0 (Fig. <xref ref-type="fig" rid="F3"/>) stems from the mismatch in temporal resolution between the predictors and the reference FWI. However, some regions such as the Mediterranean areas for FWI MAE and the Cantabrian Mountains, the Pyrenees, and the Mediterranean coast for FWI MAE95 exhibit intrinsic errors even when instantaneous inputs are provided to the DL models. Moreover, the U-Net architecture demonstrates the lowest intrinsic bias in emulating the actual FWI function. It consistently shows the smallest biases in FWI and FWI95, as well as in the predicted frequency of FWI95 events and the mean annual maximum duration of FWI95 spells. These results suggest that U-Net offers enhanced generalization capabilities. Its ability to maintain low bias across both instantaneous and aggregated inputs indicates robustness to temporal variability, a key factor in modeling climate indices.</p>

      <fig id="F3" specific-use="star"><label>Figure 3</label><caption><p id="d2e1654">DL model results for various spatial validation indices presented in Fig. <xref ref-type="fig" rid="F2"/>. The results depict the differences relative to the reference FWI for the fire season (June–September) during the test period (2018–2021). The scatter plots compare climatology values between the Proxy FWI95 (in grey) and the corresponding DL model predictions (in blue), against reference FWI95. Linear fits are shown in grey for the Proxy FWI95 and in red for the DL model, with the corresponding <inline-formula><mml:math id="M37" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> values displayed in the top-left corner.</p></caption>
            <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f03.png"/>

          </fig>

      <p id="d2e1677">Focusing on the intercomparison of DL models trained with P0, all DL models exhibit similar performance across the validation indices compared to the Proxy FWI output. For each validation index, at least one DL model outperforms the Proxy FWI. While Dense reveals slightly worst mean FWI MAEs (Fig. <xref ref-type="fig" rid="F3"/>) than the Proxy FWI (Fig. <xref ref-type="fig" rid="F2"/>), DeepESD and U-Net have comparable performance (6.54 <italic>and</italic> 6.01  vs. 6.56). However, the DL models provide a smoother spatial distribution of MAE across both continental and coastal Mediterranean regions. In contrast, as discussed in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>, the Proxy FWI exhibits large MAE values in FWI-prone areas, such as the Ebro Valley, the Cantabrian mountain range, and southern Portugal (Fig. <xref ref-type="fig" rid="F2"/>), which are largely reduced by the DL emulated FWI (Fig. <xref ref-type="fig" rid="F3"/>).</p>
      <p id="d2e1694">Regarding FWI95 MAE, the U-Net model improves upon the Proxy FWI results (8.19 vs. 8.77), offering a smoother MAE distribution across the region and particularly low MAE values in some continental areas in Spain, where the Proxy FWI exhibits large errors. Except for Portugal, the Cantabrian and the Mediterranean coast, the U-Net model outperforms the Proxy FWI across Iberia. Meanwhile, the other DL models display considerable MAE values, especially in central and northeastern Portugal and the Cantabrian mountain range, where these errors are larger than those of the Proxy FWI.</p>
      <p id="d2e1697">For FWI95 frequency, all three DL models outperform the Proxy FWI, with the U-Net model achieving the lowest relative bias. The Dense and DeepESD models generally underestimate FWI95 frequency across the region, except in southeastern Iberia and some localized areas in northern Iberia. Moreover, the U-Net model performs slightly better than the other DL models, overestimating FWI95 frequency across northern Iberia.</p>
      <p id="d2e1700">To assess the temporal characteristics of the emulated FWI, we evaluate Max Spell95 w.r.t. reference FWI. Here, the U-Net model achieves the best average result, closely followed by DeepESD (bias of 0.64  vs. 0.69). The key difference between these models is that U-Net generally overestimates is some northern areas, with maximum exceedances of 1–2 d, whereas DeepESD tends to overestimate, with discrepancies of up to 1–2 d in specific areas across the south-western of the Iberian Peninsula.</p>
      <p id="d2e1703">Despite these local problems, overall the scatter plots depict that DL models consistently improve upon the Proxy FWI results, regardless of the specific DL model. In terms of <inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula>, the U-Net model achieves the best performance for FWI95, with a value of 0.982, compared to 0.934 for the Proxy FWI. Overall in the FWI95 scatter plots, the Proxy FWI underestimates higher values (above 60 for the reference FWI95), while the U-Net model, in particular, offers a significantly better representation of reference FWI.</p>
      <p id="d2e1717">Furthermore in Fig. <xref ref-type="fig" rid="F4"/>, we assess the DL model ability in discriminating extreme fire danger events (i.e., reference FWI values above the 95th percentile) using ROC-AUC histograms, where the Proxy FWI is represented in red and the DL models in blue. Overall, the DL models improve upon the results provided by the Proxy FWI. Regarding extreme event detection, DeepESD demonstrates the best performance, with a median ROC-AUC of 0.734, compared to 0.727 for U-Net and 0.7 for the Dense model, as shown in the histograms (Fig. <xref ref-type="fig" rid="F4"/>). In this regard, the DL models outperform the Proxy FWI, as indicated by the higher frequency of ROC-AUC values above 0.5 (indicating better-than-random classification). The only exception occurs for ROC-AUC values above 0.9, which correspond to the classification of grid points along the Cantabrian coast, where climatological FWI is the lowest Iberia (Fig. <xref ref-type="fig" rid="F2"/>) due to generally milder and moist conditions throughout the year, where Proxy FWI exhibits better FWI95 event discrimination than DL models.</p>

      <fig id="F4" specific-use="star"><label>Figure 4</label><caption><p id="d2e1729">DL prediction results corresponding to the fire season (June–September) for the test period (2018–2021). Histograms compare the Area Under the ROC Curve (AUC)for extreme FWI event classification (above the 95th percentile) at each grid point, between the Proxy FWI (in red) and the corresponding DL model (in blue). The median values of both are also indicated.</p></caption>
            <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f04.png"/>

          </fig>

      <p id="d2e1738">In addition to the spatial assessment, we provide an analysis of the monthly boxplots of the Reference FWI, Proxy FWI, and the DL-predicted FWI (Dense, DeepESD, and U-Net) for each climatological region, as shown in Fig. <xref ref-type="fig" rid="F5"/>. Each boxplot represents the distribution of FWI values for a given month over the 2018–2021 period, capturing both the median and the spread of the values across the climatological regions (ATL, COASM, and CONTM; see Fig. <xref ref-type="fig" rid="F2"/>). This visualization demonstrates that the DL models not only reproduce the spatial patterns of FWI but also effectively capture the temporal variability across the year and across different regions. The median and interquartile ranges of the DL predictions closely follow the Reference FWI throughout the seasonal cycle, including the high-risk summer months, whereas the traditional Proxy FWI tends to underestimate extremes during peak months (June–September). The results further show that the Proxy FWI departs from the reference, particularly during the summer when values are highest and variability is greatest. In contrast, the DL models consistently produce distributions that are much closer to the Reference FWI, reducing the discrepancies observed with the proxy. This improvement is shown in all three regions, confirming that the advantage of the DL models over the Proxy FWI is robust and not region-specific. The closer alignment of the DL predictions with the reference highlights their capacity to more accurately capture the seasonal dynamics of fire weather conditions without introducing significant seasonal biases.</p>

      <fig id="F5" specific-use="star"><label>Figure 5</label><caption><p id="d2e1747">Monthly boxplots for the reference FWI, proxy FWI, and the DL-predicted FWI (Dense, DeepESD, UNet) for CONTM (top-left), COASM (top-right) and ATL (bottom) regions. Each boxplot represents the distribution of FWI values for a given month over the test period, capturing both the median and the spread of the values.</p></caption>
            <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f05.png"/>

          </fig>

</sec>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title>Emulation Model Explainability</title>
      <p id="d2e1765">Here, we analyze the saliency attributed to the predictor variables, grouped by the same climatic regions and fire danger categories defined in Appendix <xref ref-type="sec" rid="App1.Ch1.S1"/>. The overall results indicate no major differences in the saliencies of the different DL models, underpinning their robustness and physical coherence. Therefore, here we present the explainability results of the U-Net model, since it outperforms Proxy FWI and attains similar validation results as DeepESD, while offering greater computational efficiency. To this aim, we compute the saliencies of each variable considering the time events corresponding to the different fire danger categories separately (Fig. <xref ref-type="fig" rid="F6"/>).</p>
      <p id="d2e1772">Focusing on the <italic>high</italic>, <italic>very high</italic>, and <italic>extreme</italic> fire danger categories, temperature and relative humidity emerge as the most influential variables in the model’s predictions. The relative importance of these two variables varies across climatic regions. In the Continental Mediterranean region, temperature and relative humidity show similar levels of saliency. For the <italic>high</italic> fire danger category, relative humidity slightly surpasses temperature, whereas for the <italic>very high</italic> and <italic>extreme</italic> categories, temperature becomes more dominant. Wind speed consistently ranks as the third most important variable in these categories, although its relevance diminishes as the fire danger increases. The relatively low saliency of wind speed for the most severe categories reflects both the structure of the training data and the role that wind plays within the Fire Weather Index (FWI) formulation at very high danger levels. In Iberia, extreme fire danger episodes are predominantly associated with persistent heat anomalies, very low relative humidity, and sustained fuel desiccation, rather than with short-lived wind maxima. Once fuels are critically dry, additional increases in wind speed exert a comparatively smaller influence on the final FWI class than thermodynamic controls, which leads the model to rely more strongly on temperature and relative humidity when discriminating between <italic>extreme</italic> and <italic>very extreme</italic> conditions. Moreover, wind exhibits higher temporal variability and weaker spatial coherence than the other predictors, particularly when daily aggregated inputs are considered.</p>
      <p id="d2e1800">This finding is consistent with the definition of the FWI itself, which relies primarily on temperature, relative humidity, and wind speed under dry conditions. The DL model thus reflects the structure and sensitivity of the FWI metric it is trained to emulate. Importantly, this does not mean that precipitation (or other inputs) is irrelevant for real-world fire risk; rather, it highlights that the predictand (FWI) gives limited weight to precipitation in high and extreme danger situations. Moreover, ERA5-Land predictor variables are not independent. Relative humidity, for instance, is derived from temperature and dew point, which are indirectly influenced by precipitation, while precipitation and wind are assimilated forcings. These inter-dependencies likely contribute to the model’s ability to achieve high predictive accuracy even when precipitation and wind receive lower attribution scores.</p>
      <p id="d2e1803">The Coastal Mediterranean region shows a similar pattern. However, in the <italic>high</italic> category, temperature becomes more salient than relative humidity. In contrast, for the <italic>extreme</italic> category, relative humidity overtakes temperature as the most relevant predictor.</p>
      <p id="d2e1813">Similarly, in the Atlantic region, temperature is the most significant variable for the <italic>high</italic> and <italic>very high</italic> categories. For the <italic>extreme</italic> category, however, relative humidity surpasses temperature. Notably, in this region, the difference in saliency between wind speed and the top-ranked variables is smaller compared to other regions, suggesting a higher reliance on this variable in this region. As with the other regions, the relevance of wind speed decreases with increasing fire danger, and precipitation remains consistently negligible.</p>
      <p id="d2e1825">For the <italic>medium</italic> and <italic>low</italic> fire danger categories, the importance of variables shifts more markedly and shows less consistency across climatic regions. In the Continental and Coastal Mediterranean regions, precipitation becomes the most relevant variable for the <italic>low</italic> category, followed by relative humidity, wind speed, and temperature. For the <italic>medium</italic> category, wind speed emerges as the most salient predictor, followed by temperature, relative humidity, and precipitation.</p>
      <p id="d2e1840">In contrast, in the Atlantic region, relative humidity is the most important variable for the <italic>low</italic> category, followed by precipitation, wind speed, and temperature. For the <italic>medium</italic> category, temperature takes the lead, followed by wind speed, relative humidity, and finally precipitation. These results suggest that precipitation is a key variable for <italic>low</italic> fire danger levels but loses significance as fire danger increases, becoming practically irrelevant in <italic>extreme</italic> conditions, when temperature and relative humidity gain prominence and become the most critical predictors. This is consistent with numerous studies that highlight the importance of high temperatures and low humidity in contributing to extreme fire danger <xref ref-type="bibr" rid="bib1.bibx33" id="paren.55"/>.</p>
      <p id="d2e1858">Given that precipitation has negligible saliency in high percentile fire danger cases, which are precisely the most important from the point of view of impacts, the next section evaluates model performance after removing precipitation from the predictor set. We also examine alternative predictor sets, including replacements for precipitation and wind speed, such as eastward and northward wind components, under the hypothesis that these may carry additional useful information for the emulators.</p>

      <fig id="F6"><label>Figure 6</label><caption><p id="d2e1863">Aggregated saliency of predictor variables by climatic region and fire danger category for the U-Net model. Columns represent the climatic regions defined in Fig. <xref ref-type="fig" rid="FA1"/>, while rows correspond to the fire danger categories described in Sect. <xref ref-type="sec" rid="Ch1.S3.SS2"/>. The displayed saliency values are derived from saliency maps, normalized at each grid point. Results for the JJAS season, test period (2018–2021).</p></caption>
          <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f06.png"/>

        </fig>

</sec>
<sec id="Ch1.S3.SS4">
  <label>3.4</label><title>Predictor set intercomparison</title>
      <p id="d2e1885">In this section, we intercompare different alternative input predictor sets, summarized in Table <xref ref-type="table" rid="T1"/>. For simplicity, we focus on the U-Net DL method as in Sect. <xref ref-type="sec" rid="Ch1.S3.SS3"/>. The U-Net trained with P0 and P1 yields better results than when trained with P2, except for the MAE FWI (Fig. <xref ref-type="fig" rid="F7"/>, analogous to Fig. <xref ref-type="fig" rid="F3"/> for comparability). The difference between U-Net P0, U-Net P1 and U-Net P2 is relatively small, with the most remarkable variations observed in the western Continental region. For the MAE of FWI95 events, U-Net P2 significantly overestimates across the entire domain, whereas U-Net P0 and U-Net P1 present similar MAE spatial distribution, standing P1 as the best approach. However, U-Net P1 shows a greater presence of errors in specific regions, such as north Portugal, the Pyrenees and the Cantabrian mountain ranges. Regarding the frequency of FWI95 events, U-Net P0 tends to overestimation across nearly the entire the north domain, while U-Net P1 tends to overestimate in the mid-south and east of the Peninsula and U-Net P2 exhibit a general underestimation. Among these, U-Net P0 appears to be the most effective in capturing this index. Examining the Max Spell95 results of the different experiments, U-Net P2 performs the worst, with overestimation up to 3 to 4 d in some areas, whereas the largest differences in P0 and P1 range between 2 and 3 d at most. U-Net P1 shows a slight improvement over U-Net P0, displaying more areas of overestimation, while U-Net P0 predominantly underestimates across most of the domain. Finally, the FWI95 scatter plots exhibit a high <inline-formula><mml:math id="M39" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> fit for all three predictor sets, outperforming the Proxy FWI. In terms of overall performance, U-Net P1 produces the best results, followed by  U-Net P0 and U-Net P2, though the differences among the latter remain relatively small.</p>

      <fig id="F7" specific-use="star"><label>Figure 7</label><caption><p id="d2e1909">U-Net model results for various spatial validation indices presented in Fig. <xref ref-type="fig" rid="F2"/> and different predictors sets assessed (see Table <xref ref-type="table" rid="T1"/>). The results depict the differences relative to reference FWI for the fire season (June–September) during the test period (2018–2021). Additionally, scatter plots compare climatology values of the Proxy FWI95 (in grey) and the U-Net model (in blue) against reference FWI95. Linear regression fit lines are shown in grey for the Proxy FWI95 and in red for the U-Net model (to enhance visualization), and the corresponding  <inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:msup><mml:mi>R</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msup></mml:mrow></mml:math></inline-formula> values displayed in the top-left corner.</p></caption>
          <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f07.png"/>

        </fig>

      <p id="d2e1933">Overall, in Fig. <xref ref-type="fig" rid="F8"/>, U-Net P1 yields the best results as classifiers for extreme events. These experiments improve upon the median results of the Proxy FWI, with U-Net P1 emerging as the most effective model in this context.</p>
      <p id="d2e1939">In summary, the U-Net model demonstrates superior performance over the Proxy FWI, with U-Net P1 generally excelling in overall performance. The inclusion of wind components as additional predictors in P2 has a deleterious effect on the temporal and distributional similarity of emulated and reference FWI. This suggests that, for emulating the FWI index itself, wind speed magnitude already captures the relevant wind-related information, and the addition of directional components introduces redundant or noisy predictors that can hamper model performance, despite the internal feature selection of deep learning methods. Consequently, a prior predictor screening may still be necessary to ensure optimal results. Importantly, as seen through the current section, the scope of this work is restricted to the statistical reconstruction and characterization of fire-weather conditions as represented by the FWI, rather than the analysis of wildfire occurrence, behavior, or impacts. The DL-based results are validated considering the FWI95, FWI95 frequency and the Max Spell95 (Figs. <xref ref-type="fig" rid="F3"/> and <xref ref-type="fig" rid="F4"/>) and the AUC of the ROC curve for extreme fire danger classification as noted in Figs. <xref ref-type="fig" rid="F7"/> and <xref ref-type="fig" rid="F8"/>. With this suite of validation metrics and experiments, we properly assess the ability of the DL-based results to statistically reproduce the most extreme events. Consequently, we do not investigate individual wildfire events or fuel-specific fire behavior, which lie outside the intended scope of this study.</p>
</sec>
<sec id="Ch1.S3.SS5">
  <label>3.5</label><title>Loss function assessment</title>
      <p id="d2e1958">We next present a summary of the validation results, including an intercomparison of alternative MSE and asymmetric loss functions in the DL models (Sect. <xref ref-type="sec" rid="Ch1.S2.SS4.SSS4"/>). Table <xref ref-type="table" rid="T4"/> presents the validation indices for different predictor sets and DL models, comparing their performance against the Proxy FWI baseline, where the ASYM loss function values are indicated in parentheses. For all indices, the DL models exhibit consistent improvements over the Proxy FWI, with at least one model outperforming it in each category.</p>
      <p id="d2e1965">Regarding the loss function, the ASYM loss generally improves performance in most validation indices. The ASYM-trained models tend to yield lower MAE than ASYM values, particularly for DeepESD in P1 (6.74 vs. 6.43) and P2 (6.24 vs. 7.12), showing better accuracy in estimating FWI. Similarly, in the FWI95 MAE, the ASYM-trained models consistently outperform MSE loss counterparts, as seen with DeepESD in P0 (9.28 vs. 6.82) and P2 (9.66 vs. 9.44). In terms of extreme event prediction, the Max Spell95 Bias shows inconclusive results. While the ASYM loss function improves the bias in some cases, such as U-Net in P2 (0.78 vs. 0.54), in other cases, it worsens bias, as observed with Dense in P2 (0.78 vs. 1.05). This suggests that ASYM does not consistently improve the temporal persistence of extreme events but may still provide benefits in specific cases. For the Relative Bias Frequency of FWI95, ASYM provides substantial benefits in reducing bias. The most notable case is U-Net in P1, where the ASYM-trained model achieves a significantly lower bias (0.358 vs. 0.269) than MSE, reinforcing ASYM's ability to enhance the reliability of extreme event frequency estimations.</p>

      <fig id="F8" specific-use="star"><label>Figure 8</label><caption><p id="d2e1970">U-Net model histograms results for the fire season (June–September) for the test period (2018–2021). Histograms compare the Area Under the Curve (AUC) of the Receiver Operating Characteristic (ROC) curve for extreme fire danger classification (above the 95th percentile) at each grid point between the Proxy FWI (in red) and the U-Net model (in blue). The median value is indicated for both results.</p></caption>
          <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f08.png"/>

        </fig>

      <p id="d2e1980">Overall, ASYM appears to be particularly beneficial in reducing error in FWI estimations (MAE and MAE FWI95) and in improving the classification of extreme fire weather danger, as evidenced by the lower relative bias for FWI95 than MSE. However, its benefits on temporal aspects, such as Max Spell95 Bias, are not consistent. These findings suggest that while ASYM is generally advantageous, its effects are model- and index-dependent, although sometimes overestimates low and moderate values in the distribution (see Fig. <xref ref-type="fig" rid="FC1"/>). Therefore, the ASYM loss function does not compromise the U-Net model performance in the lower quantiles, however for the Dense and DeepESD models low and moderate values for the FWI are worse than using the MSE as loss function. Therefore, the ASYM loss function can compromise the performance of low and moderate levels in some cases, and its selection should be carefully considered depending on the primary objective of the analysis.</p>

<table-wrap id="T4" specific-use="star"><label>Table 4</label><caption><p id="d2e1988">Validation results of DL FWI emulators for different predictor sets. Values in parentheses are for the ASYM loss function for model training (default values correspond to MSE loss). The first row presents the Proxy FWI results, as benchmark. Bold values highlight the best validation index within each predictor set, while underlined values denote the overall best across all indices.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="6">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:colspec colnum="5" colname="col5" align="center"/>
     <oasis:colspec colnum="6" colname="col6" align="center"/>
     <oasis:thead>
       <oasis:row>

         <oasis:entry colname="col1">Predictors</oasis:entry>

         <oasis:entry colname="col2"/>

         <oasis:entry colname="col3"/>

         <oasis:entry colname="col4">MAE</oasis:entry>

         <oasis:entry colname="col5">Max.</oasis:entry>

         <oasis:entry colname="col6">Rel. Bias</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col1">sets</oasis:entry>

         <oasis:entry colname="col2">DL Model</oasis:entry>

         <oasis:entry colname="col3">MAE</oasis:entry>

         <oasis:entry colname="col4">FWI95</oasis:entry>

         <oasis:entry colname="col5">Spell95 Bias</oasis:entry>

         <oasis:entry colname="col6">Freq. FWI95</oasis:entry>

       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row rowsep="1">

         <oasis:entry colname="col1">Proxy FWI</oasis:entry>

         <oasis:entry colname="col2"/>

         <oasis:entry colname="col3">6.56</oasis:entry>

         <oasis:entry colname="col4">8.77</oasis:entry>

         <oasis:entry colname="col5">0.94</oasis:entry>

         <oasis:entry colname="col6">0.657</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry rowsep="1" colname="col1" morerows="2">P0</oasis:entry>

         <oasis:entry colname="col2">Dense</oasis:entry>

         <oasis:entry colname="col3">7.01 (7.94)</oasis:entry>

         <oasis:entry colname="col4">10.72 (8.86)</oasis:entry>

         <oasis:entry colname="col5">0.73 (0.95)</oasis:entry>

         <oasis:entry colname="col6">0.337 (0.662)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col2">DeepESD</oasis:entry>

         <oasis:entry colname="col3">6.54 (6.81)</oasis:entry>

         <oasis:entry colname="col4">9.28 (<underline><bold>6.82</bold></underline>)</oasis:entry>

         <oasis:entry colname="col5">0.69 (0.6)</oasis:entry>

         <oasis:entry colname="col6">0.357 (0.401)</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col2">U-Net</oasis:entry>

         <oasis:entry colname="col3"><bold>6.01</bold> (6.09)</oasis:entry>

         <oasis:entry colname="col4">8.19 (8.51)</oasis:entry>

         <oasis:entry colname="col5"><bold>0.64</bold> (0.65)</oasis:entry>

         <oasis:entry colname="col6">0.331 (<bold>0.313</bold>)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry rowsep="1" colname="col1" morerows="2">P1</oasis:entry>

         <oasis:entry colname="col2">Dense</oasis:entry>

         <oasis:entry colname="col3">6.85 (7.33)</oasis:entry>

         <oasis:entry colname="col4">10.97 (9.83)</oasis:entry>

         <oasis:entry colname="col5">0.7 (0.92)</oasis:entry>

         <oasis:entry colname="col6">0.342 (0.514)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col2">DeepESD</oasis:entry>

         <oasis:entry colname="col3">6.74 (6.43)</oasis:entry>

         <oasis:entry colname="col4">8.43 (8.01)</oasis:entry>

         <oasis:entry colname="col5">1 (0.74)</oasis:entry>

         <oasis:entry colname="col6">0.666 (0.46)</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col2">U-Net</oasis:entry>

         <oasis:entry colname="col3">6.24 (<underline><bold>5.84</bold></underline>)</oasis:entry>

         <oasis:entry colname="col4"><bold>7.47</bold> (8.46)</oasis:entry>

         <oasis:entry colname="col5"><bold>0.62</bold> (0.63)</oasis:entry>

         <oasis:entry colname="col6">0.358 (<underline><bold>0.269</bold></underline>)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col1" morerows="2">P2</oasis:entry>

         <oasis:entry colname="col2">Dense</oasis:entry>

         <oasis:entry colname="col3">6.95 (7.58)</oasis:entry>

         <oasis:entry colname="col4">10.72 (9.8)</oasis:entry>

         <oasis:entry colname="col5">0.78 (1.05)</oasis:entry>

         <oasis:entry colname="col6">0.408 (0.588)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col2">DeepESD</oasis:entry>

         <oasis:entry colname="col3">6.24 (7.12)</oasis:entry>

         <oasis:entry colname="col4">9.66 (9.44)</oasis:entry>

         <oasis:entry colname="col5">0.69 (0.95)</oasis:entry>

         <oasis:entry colname="col6">0.362 (0.521)</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col2">U-Net</oasis:entry>

         <oasis:entry colname="col3"><bold>6.08</bold> (6.95)</oasis:entry>

         <oasis:entry colname="col4">10.5 (<bold>8.7</bold>)</oasis:entry>

         <oasis:entry colname="col5">0.78 (<underline><bold>0.54</bold></underline>)</oasis:entry>

         <oasis:entry colname="col6">0.42 (<bold>0.295</bold>)</oasis:entry>

       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="Ch1.S3.SS6">
  <label>3.6</label><title>Fuel Moisture Code sub-indices Assessment</title>
      <p id="d2e2296">The Canadian Forest Fire Weather Index System is structured around six interrelated components that describe both fuel moisture conditions and potential fire behavior. In particular, the three fuel moisture codes Fine Fuel Moisture Code (FFMC), Duff Moisture Code (DMC) and Drought Code (DC) represent distinct organic layers with characteristic response times ranging from hours (surface litter) to months (deep compacted layers). These components provide physically interpretable proxies of ignition potential, fire sustainability and drought build-up, and they underpin the behavior of the aggregated FWI. Given their distinct physical meaning and temporal memory, assessing the performance of the deep-learning emulators at the level of the individual moisture codes provides an additional and more stringent evaluation of the modeling framework. To explicitly evaluate rainfall sensitivity and temporal memory across scales, two predictor configurations are considered: P0, which includes precipitation, and P1, which excludes it. This design enables a controlled assessment of the extent to which daily precipitation information contributes to the emulation of fast-, intermediate-, and slow-response fuel layers. Beyond reproducing the aggregated FWI, this analysis examines whether the emulators are able to capture the multi-scale fuel-moisture dynamics embedded in the structure of the system. We emphasize that this analysis should be regarded as preliminary, and that a more comprehensive and detailed assessment is currently under development. Nevertheless, the present evaluation provides an initial characterization of the performance of the DL model originally designed to emulate FWI when the predictand is replaced by its individual fuel moisture components (DC, DMC, or FFMC) instead of the aggregated FWI index.</p>
      <p id="d2e2299">Table <xref ref-type="table" rid="T5"/> summaries the validation performance of the emulators for DC, DMC and FFMC under the two predictor configurations (P0, including precipitation; P1, excluding precipitation). The evaluation considers some of the validation metrics as the MAE, FWI95 MAE and relative bias in the frequency above the 95th percentile previously depicted in the manuscript. Moreover, we introduce an additional validation metric to explicitly assess the temporal consistency of the predicted subindex relative to the reference subindex. Specifically, we compute the Pearson correlation coefficient between the predicted and reference spatially aggregated time series. Prior to calculating the correlation, the seasonal cycle is removed from both time series to isolate anomalies and avoid inflating the correlation due to the shared annual cycle. To this end, the daily climatology (i.e., one value per Julian day of the year) is subtracted from each daily value in the time series. The daily climatological values are estimated over the full reference period using a centered moving average window of seven days, which smooths high-frequency noise while retaining the intra-seasonal structure of the climatology. This metric provides a domain-integrated measure of the emulator ability to reproduce the large-scale temporal variability of the fuel moisture codes, focusing explicitly on inter-daily to synoptic-scale fluctuations rather than on the mean seasonal evolution. As in the previous sections, these metrics are evaluated for the fire season (JJAS) in the test period (2018–2021), with the exception of the time series correlation, which is computed over the full annual cycle.</p>

<table-wrap id="T5" specific-use="star"><label>Table 5</label><caption><p id="d2e2307">Validation results of DL emulators for different predictor sets and FMC subindices; DC, DMC and FFMC.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="6">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="center"/>
     <oasis:colspec colnum="4" colname="col4" align="center"/>
     <oasis:colspec colnum="5" colname="col5" align="center"/>
     <oasis:colspec colnum="6" colname="col6" align="center"/>
     <oasis:thead>
       <oasis:row>

         <oasis:entry colname="col1"/>

         <oasis:entry colname="col2">Predictors</oasis:entry>

         <oasis:entry colname="col3"/>

         <oasis:entry colname="col4">MAE</oasis:entry>

         <oasis:entry colname="col5">Time series</oasis:entry>

         <oasis:entry colname="col6">Rel. Bias</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col1">Code</oasis:entry>

         <oasis:entry colname="col2">sets</oasis:entry>

         <oasis:entry colname="col3">MAE</oasis:entry>

         <oasis:entry colname="col4">FWI95</oasis:entry>

         <oasis:entry colname="col5">correlation</oasis:entry>

         <oasis:entry colname="col6">Freq. FWI95</oasis:entry>

       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>

         <oasis:entry rowsep="1" colname="col1" morerows="1">DC</oasis:entry>

         <oasis:entry colname="col2">P0</oasis:entry>

         <oasis:entry colname="col3">105.62</oasis:entry>

         <oasis:entry colname="col4">227.51</oasis:entry>

         <oasis:entry colname="col5">0.72</oasis:entry>

         <oasis:entry colname="col6">0.81</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col2">P1</oasis:entry>

         <oasis:entry colname="col3">118.17</oasis:entry>

         <oasis:entry colname="col4">225.74</oasis:entry>

         <oasis:entry colname="col5">0.61</oasis:entry>

         <oasis:entry colname="col6">0.82</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry rowsep="1" colname="col1" morerows="1">DMC</oasis:entry>

         <oasis:entry colname="col2">P0</oasis:entry>

         <oasis:entry colname="col3">50.63</oasis:entry>

         <oasis:entry colname="col4">173.32</oasis:entry>

         <oasis:entry colname="col5">0.75</oasis:entry>

         <oasis:entry colname="col6">0.93</oasis:entry>

       </oasis:row>
       <oasis:row rowsep="1">

         <oasis:entry colname="col2">P1</oasis:entry>

         <oasis:entry colname="col3">48.23</oasis:entry>

         <oasis:entry colname="col4">139.81</oasis:entry>

         <oasis:entry colname="col5">0.85</oasis:entry>

         <oasis:entry colname="col6">0.77</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col1" morerows="1">FFMC</oasis:entry>

         <oasis:entry colname="col2">P0</oasis:entry>

         <oasis:entry colname="col3">3.58</oasis:entry>

         <oasis:entry colname="col4">2.91</oasis:entry>

         <oasis:entry colname="col5">0.98</oasis:entry>

         <oasis:entry colname="col6">0.67</oasis:entry>

       </oasis:row>
       <oasis:row>

         <oasis:entry colname="col2">P1</oasis:entry>

         <oasis:entry colname="col3">3.36</oasis:entry>

         <oasis:entry colname="col4">2.16</oasis:entry>

         <oasis:entry colname="col5">0.97</oasis:entry>

         <oasis:entry colname="col6">0.52</oasis:entry>

       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d2e2498">For the DC, which reflects long-term deep fuel moisture and cumulative drought effects, correlations are moderate (0.72 for P0 and 0.61 for P1). This suggests a greater challenge in fully capturing low-frequency variability and long-memory dynamics. Absolute errors are comparatively large in magnitude (MAE <inline-formula><mml:math id="M41" display="inline"><mml:mo>∼</mml:mo></mml:math></inline-formula> 106–118), which is consistent with the broader dynamic range of the DC. P0 achieves both lower MAE (105.62 vs. 118.17) and comparable performance under extreme conditions (FWI95: 227.51 vs. 225.74 are nearly identical, with a slight advantage for P1 in extremes but a higher overall MAE). The relative bias in high-danger frequency is very similar between predictor sets (0.81–0.82), indicating comparable representation of extreme drought-related conditions despite differences in correlation strength. The modest improvement obtained when precipitation is included is consistent with the DC formulation, where only substantial or sustained rainfall events effectively reduce deep-layer dryness; although such events are infrequent during JJAS, their accumulated impact contributes to improved interannual and low-frequency variability representation.</p>
      <p id="d2e2508">The temporal evolution shown in Fig. <xref ref-type="fig" rid="FE1"/> provides additional context for interpreting the DC results. Despite the moderate correlations obtained for both predictor configurations, the DL reconstructions reproduce the main seasonal cycle and the timing of the major multi-month drought episodes throughout the 2018–2021 test period. Moreover, the two predictor sets generate remarkably similar DC trajectories, with only minor differences in amplitude and short-term fluctuations. This behavior is consistent with the validation metrics and indicates that the reduced skill observed for DC cannot be attributed solely to the exclusion of precipitation from the predictor set. Rather, it reflects the intrinsic difficulty of reproducing the long-memory dynamics and low-frequency variability embedded in the DC using architectures that do not explicitly model temporal dependencies, irrespective of whether precipitation is included. This limitation is further quantified in Appendix <xref ref-type="sec" rid="App1.Ch1.S5"/>, where the analysis of temporal persistence reveals a pronounced underestimation of memory for the DC. Together, these results demonstrate that the challenge lies primarily in the representation of persistence rather than in the availability of predictor variables. We note that this limitation is not intrinsic to the emulation framework itself, but to the specific class of architectures employed. Models explicitly designed to capture temporal dependencies, such as ConvLSTM networks, have been shown to better reproduce long-memory processes in a fire-weather prediction context <xref ref-type="bibr" rid="bib1.bibx42" id="paren.56"/>, albeit at a substantially higher computational cost, which limits their applicability in large-domain experiments such as the one presented here.</p>
      <p id="d2e2518">In the case of the DMC, representing intermediate-depth organic layers with response times of days to weeks, correlations are moderate to high (0.75 for P0 and 0.85 for P1). The removal of precipitation as a predictor in P1 markedly improves temporal correspondence. Overall MAE is slightly lower for P1 (48.23 vs. 50.63), and performance under extreme fire-weather conditions improves substantially (MAE FWI95: 139.81 vs. 173.32). The relative bias in high-danger frequency decreases from 0.93 (P0) to 0.77 (P1), indicating improved prediction of extreme-category occurrence under P1. This behavior indicates that during the Iberian fire season the DMC signal is largely governed by medium-term atmospheric drying driven by temperature and relative humidity, whereas rainfall events exceeding the effective infiltration threshold are comparatively rare at daily scale, thereby limiting the incremental value of explicit precipitation information for seasonal emulation.</p>
      <p id="d2e2521">Finally, for the FFMC, which governs short-term surface litter moisture and ignition potential, correlations are very high (0.98 for P0 and 0.97 for P1), demonstrating strong skill in reproducing rapid variability. Absolute errors are small (MAE <inline-formula><mml:math id="M42" display="inline"><mml:mo>∼</mml:mo></mml:math></inline-formula> 3.4–3.6; MAE FWI95: 2.16–2.91), with negligible differences in mean and extreme performance between predictor sets. The relative bias in extreme-day frequency decreases from 0.67 (P0) to 0.52 (P1), indicating improved representation of high-danger occurrences. Given the intrinsic sensitivity of FFMC to even light rainfall, the small differences between P0 and P1 suggest that the deep-learning architecture is able to reconstruct short-term surface moisture variability primarily from thermodynamic and wind predictors at daily scale, while the limited seasonal contribution of precipitation reflects its low frequency during JJAS rather than a lack of physical responsiveness.</p>
      <p id="d2e2531">In summary, the DL framework shows scale-dependent performance: very strong skill for the fast-response FFMC, moderate-to-high skill for the DMC (particularly under P1), and comparatively lower temporal fidelity for the slow-evolving DC. Despite these differences, biases in extreme-category frequency remain within a relatively narrow range, suggesting that the hierarchical structure of fuel-moisture dynamics is largely preserved. The differential sensitivity observed across predictor configurations is consistent with the physical formulation and temporal memory of each code, reinforcing the internal coherence of the emulation framework across fast, intermediate, and slow fuel-moisture regimes. Future work will also focus on the implementation of the emulator within the context of the PTI-Clima climate-service initiative (<uri>https://pti-clima.csic.es/</uri>, last access: 24 June 2026), alongside a more exhaustive validation against selected real-world wildfire events.</p>
</sec>
</sec>
<sec id="Ch1.S4" sec-type="conclusions">
  <label>4</label><title>Conclusions</title>
      <p id="d2e2546">In this study we addressed the challenge of emulating the reference Canadian Fire Weather Index (FWI), which requires instantaneous noon meteorological inputs that are rarely available in observational databases or climate model output. We proposed a deep-learning framework capable of reconstructing the reference FWI from commonly available daily variables and evaluated its performance over the Iberian Peninsula. Our results show that (i) all emulators reproduce the main spatial and temporal characteristics of the reference FWI, including high-percentile behavior (critical for extreme fire danger detection); (ii) they consistently outperform the “traditional” Proxy FWI – as introduced by <xref ref-type="bibr" rid="bib1.bibx5" id="text.57"/> for regional climate projections–, particularly in the detection of extreme fire-weather events; (iii) daily mean predictors are sufficient to recover the reference FWI signal with high fidelity; (iv) the emulator can approximate the FWI with limited accuracy loss even when precipitation is omitted from the predictors, a result that likely reflects the model's ability to indirectly recover precipitation-related information through covariation with other meteorological variables (temperature, humidity, wind) and learned climatological patterns, rather than indicating that precipitation is physically unimportant for the FWI system; and (v) the U-Net and DeepESD architectures offer the best overall balance across metrics. These findings highlight the potential of DL-based emulators to provide accurate fire-weather information when sub-daily inputs are unavailable, and to circumvent data limitations in applications where certain predictors may be missing or uncertain.</p>
      <p id="d2e2552">A key methodological distinction should be emphasized for the proper interpretation of these results: the reference FWI is computed from instantaneous noon LST meteorological variables as originally defined in the FWI system, whereas the DL emulators are trained using daily-aggregated predictors (daily means, minima, or accumulated values). The reported emulator skill should therefore be understood as the ability of DL models to recover a noon-time fire-weather signal from daily-aggregated inputs, a practically valuable capability precisely because such inputs are far more widely available than instantaneous noon observations.</p>
      <p id="d2e2555">Placed in the context of previous work, our results address a well known problem with FWI reconstruction from model outputs <xref ref-type="bibr" rid="bib1.bibx29" id="paren.58"/>, that can also be circumvented through the use of FWI-derived indicators <xref ref-type="bibr" rid="bib1.bibx40" id="paren.59"><named-content content-type="pre">e.g.:</named-content></xref>, posing a challenge in future climate impact studies and introducing additional uncertainty in the resulting projections. Our study extends earlier studies that relied on proxy variables to approximate the noon-LST FWI signal <xref ref-type="bibr" rid="bib1.bibx5" id="paren.60"/>, demonstrating that DL methods can recover the nonlinear relationships between daily predictors and instantaneous fire-weather conditions more effectively than proxy formulations. This improvement is particularly relevant for Mediterranean-type climates, where strong diurnal cycles amplify mismatches between daily means and noon conditions, and for climate-change applications relying on coarse temporal resolution model output. The study does, however, have limitations. The emulators were trained solely on ERA5-Land, which is not an observational ground truth; also, for simplicity, the analysis is restricted to the generic JJAS fire season (although the simulations are performed for the entire period), even though the fire season period varies across ecoregions within the Iberian Peninsula <xref ref-type="bibr" rid="bib1.bibx6" id="paren.61"><named-content content-type="pre">see e.g.</named-content></xref>. Additionally, while the precipitation-free experiments demonstrate the emulator's robustness under the specific conditions of the JJAS season, the results should not be interpreted as evidence that precipitation is physically unimportant for the FWI system. Rather, this finding illustrates how deep learning models can leverage statistical relationships and learned representations to approximate indices in data-sparse contexts, which may be valuable for applications but does not diminish the fundamental role of precipitation in fire-weather dynamics. Finally, the emulators are sensitive to the spatiotemporal characteristics of the training dataset and may require adaptation in environments with different climatic regimes. Our preliminary experiments on the individual fuel-moisture codes (FFMC, DMC, DC) show that FFMC and DMC can be emulated accurately from daily inputs, while the DC exhibits a stronger sensitivity to long-term memory processes, which are more challenging to reproduce with architectures that do not explicitly model temporal dependencies. This limitation could be alleviated by architectures explicitly designed to capture temporal dependencies, although their application at large spatial scales remains computationally demanding <xref ref-type="bibr" rid="bib1.bibx42" id="paren.62"/>. As this study focuses on the FWI – the final, integrated index most widely used in climate and impact applications – we leave a full component-level assessment to a companion work. Future research should also assess the practical implications of emulator performance through analyses of specific wildfire episodes, evaluating how accurately reconstructed fire-weather conditions translate into the detection and characterization of documented fire events. Such event-based validation would provide a complementary perspective to the statistical evaluation presented here and help further establish the applicability of the proposed framework in impact-oriented contexts. Ongoing research is extending this framework to other regions and seasons, incorporating additional environmental predictors, and analyzing the physical behavior of the individual FWI components in greater depth.</p>
</sec>

      
      </body>
    <back><app-group>

<app id="App1.Ch1.S1">
  <label>Appendix A</label><title>Climatological regions</title>
      <p id="d2e2589">Here, we present the division of the Iberian Peninsula into three regions – Atlantic (ATL), Continental Mediterranean (CONTM), and Coastal Mediterranean (COASM) – for the eXplainable Artificial Intelligence (XAI) analysis provided in the main manuscript (Fig. <xref ref-type="fig" rid="F6"/>).</p>

      <fig id="FA1"><label>Figure A1</label><caption><p id="d2e2596">Climatological regions within the Iberian Peninsula used for XAI analysis in Fig. <xref ref-type="fig" rid="F6"/>. Blue, green and red rectangles indicate ATL (Atlantic), COASM (Coastal Mediterranean), and CONTM (Continental Mediterranean) regions respectively.</p></caption>
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f09.png"/>

      </fig>

<fig id="FA2"><label>Figure A2</label><caption><p id="d2e2611">Monthly, spatially aggregated time series by climatological region for the reference FWI, the Proxy FWI, and the DL-based estimates over the 2018–2021 test period. The first, second and third row show the results for the Dense, DeepESD and U-Net models respectively.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f10.png"/>

      </fig>

</app>

<app id="App1.Ch1.S2">
  <label>Appendix B</label><title>ERA5-Land limitations</title>
      <p id="d2e2630">In this section, we highlight the existent limitations in using ERA5-Land data in our analysis due to the inherited biases in ERA5-Land compared with observation data. ERA5-Land, like other reanalysis products, is subject to inherent biases that arise from the limitations of the numerical models and data assimilation techniques used in its generation. These biases reflect systematic deviations from ground-based observations and can affect the reliability of the dataset for certain applications. Therefore, although ERA5-Land provides a valuable, spatially and temporally consistent climate dataset, its outputs should be used with caution and validated against local observations whenever possible.</p>
      <p id="d2e2634">In Fig. <xref ref-type="fig" rid="FB1"/>, we illustrate the ERA5-Land biases in some stations in Spain with respect to observation data provided by the Spanish Agency of Meteorology (AEMET).</p><fig id="FB1"><label>Figure B1</label><caption><p id="d2e2642">Biases between observational FWI data from the Spanish Agency of Meteorology (AEMET) and the FWI resulting from ERA5-Land computations. Top figure indicates de bias for the mean FWI, while bottom figure indicates it for the FWI95.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f11.png"/>

      </fig>

</app>

<app id="App1.Ch1.S3">
  <label>Appendix C</label><title>QQ-plot assessment</title>

      <fig id="FC1"><label>Figure C1</label><caption><p id="d2e2663">Quantile-Quantile (QQ) plots comparing the Proxy FWI (in grey) and DL model results (in blue). The results for the Coastal Mediterranean (COASM) (Fig. <xref ref-type="fig" rid="FA1"/> for the subregion boundaries). Dash-dotted vertical lines indicate different fire danger levels established by the Agencia Estatal de Meteorología (AEMET). Fire danger levels are defined as follows: <italic>low</italic> (below the green line), <italic>moderate</italic> (between the green and yellow lines), <italic>high</italic> (between the yellow and orange lines), <italic>very high</italic> (between the orange and red lines), and <italic>extreme</italic> (above the red line). The Mean Squared Error (MSE) for both results is displayed in the top right corner. A box in the bottom right provides an amplified view of QQ plot focused on values above the 95th percentile (extreme danger), along with the corresponding MSE values under the box. First row indicates the results for models trained with MSE, and the second provides the models with ASYM as loss function.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f12.png"/>

      </fig>


</app>

<app id="App1.Ch1.S4">
  <label>Appendix D</label><title>Sensitivity analysis of deep learning models and evaluation of other predictors sets</title>

      <fig id="FD1"><label>Figure D1</label><caption><p id="d2e2704">Results from the Dense, DeepESD and U-Net model trained with 12:00 UTC input variables (temperature, 24 h-accumulated precipitation, relative humidity and wind speed). The maps display differences relative to the reference FWI for the fire season (June–September) during the test period (2018–2021) for the validation indices. The MAE value inside the map represents the spatially aggregated mean absolute error of the deep learning predictions with respect to the FWI reference, while Bias denotes the spatially averaged bias in absolute value.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f13.png"/>

      </fig>

      <fig id="FD2"><label>Figure D2</label><caption><p id="d2e2717">Results from the U-Net model trained with the P1 predictors set (see Table <xref ref-type="table" rid="T1"/>), trained with P0 and adding the 24 h lagged precipitation, trained with P1 adding minimum relative humidity instead of daily mean relative humidity and trained with P1 adding precipitation. The maps display differences relative to the reference FWI for the fire season (June–September) during the test period for different validation indices. MAE represents the spatially aggregated mean absolute error of the deep learning predictions with respect to the FWI reference, while Bias denotes the spatially averaged bias in absolute value.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f14.png"/>

      </fig>


</app>

<app id="App1.Ch1.S5">
  <label>Appendix E</label><title>Temporal dynamics and Persistence of Fuel Moisture Codes</title>
      <p id="d2e2740">Figure <xref ref-type="fig" rid="FE1"/> highlights the ability of the DL-based emulators to reproduce the temporal evolution of the main fuel moisture codes under different predictor configurations for the CONTM region (see Fig. <xref ref-type="fig" rid="FA1"/>). In this Appendix, we evaluate the CONTM region because it is characterized by persistent and prolonged droughts across the Iberian Peninsula. Furthermore, the results obtained for the other regions exhibit similar behavior, making the CONTM region representative of the overall patterns. A clear scale-dependent behavior emerges, consistent with the intrinsic memory of each component. For the FFMC, which is characterized by short-term dynamics, the emulators show very high fidelity, accurately capturing both the amplitude and high-frequency variability of the reference signal. The DMC, with an intermediate memory, is also reasonably well reproduced, although with some local discrepancies in peak magnitude and timing. In contrast, the DC reveals a more pronounced limitation. While the emulators successfully reproduce the seasonal cycle and overall variability, they exhibit reduced temporal persistence compared to the reference, resulting in noisier time series and a weakened autocorrelation structure. This behavior reflects the difficulty of reproducing long-memory processes using architectures that rely primarily on instantaneous predictor–target mappings. These results suggest that, although the convolutional and dense architectures tested are highly effective for spatial reconstruction and short- to medium-term variability, they are less suited to capture the temporal coherence of slowly evolving indices such as the DC. Alternative architectures explicitly designed to model temporal dependencies, such as ConvLSTM networks <xref ref-type="bibr" rid="bib1.bibx42" id="paren.63"><named-content content-type="pre">see </named-content><named-content content-type="post">for an application to FWI downscaling</named-content></xref>, could potentially alleviate this limitation, although their computational cost remains a constraint in large-domain applications such as the one considered here.</p>
      <p id="d2e2754">Importantly, this limitation does not appear to be driven by the exclusion of precipitation from the predictor set. Both configurations considered (P0, including precipitation, and P1, excluding it) show a comparable loss of temporal persistence in the DC. This suggests that the reduced autocorrelation structure is not primarily due to missing information, but rather reflects the intrinsic difficulty of capturing long-memory processes with architectures that do not explicitly model temporal dependencies.</p>
      <p id="d2e2757">To complement this qualitative assessment of the temporal behavior, we next provide a quantitative evaluation of temporal persistence. The integral timescale (<inline-formula><mml:math id="M43" display="inline"><mml:mi mathvariant="italic">τ</mml:mi></mml:math></inline-formula>) is used as a measure of temporal persistence, computed here as the e-folding time of the autocorrelation function of anomaly time series (daily values with the corresponding monthly mean removed), up to a maximum lag of 30 d (<inline-formula><mml:math id="M44" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>). It represents the characteristic time over which the signal remains temporally correlated, with larger values indicating stronger memory and smoother temporal evolution.</p>
      <p id="d2e2778">Table <xref ref-type="table" rid="TE1"/> provides a quantitative assessment of the temporal persistence of the emulated series through the integral timescale <inline-formula><mml:math id="M45" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>. A clear scale-dependent behavior emerges across the different fuel moisture components. For the FFMC, characterized by short memory, the emulators reproduce the temporal structure with high fidelity, showing negligible bias in <inline-formula><mml:math id="M46" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>. In contrast, the DMC exhibits a systematic underestimation of persistence, and this effect becomes particularly pronounced for the DC, where <inline-formula><mml:math id="M47" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> is strongly underestimated in both predictor configurations. For the DC, the reference value (<inline-formula><mml:math id="M48" display="inline"><mml:mrow><mml:mi mathvariant="italic">τ</mml:mi><mml:mo>≈</mml:mo><mml:mn mathvariant="normal">16</mml:mn></mml:mrow></mml:math></inline-formula> d) is reduced by more than 50 % in P0 and by nearly 80 % in P1, indicating a substantial loss of long-term memory and an over-responsive temporal behavior. Importantly, this limitation is observed even when precipitation is included in the predictor set (P0), suggesting that it cannot be primarily attributed to missing inputs, but rather to the limited ability of the employed architectures to capture long-memory processes. These results are consistent with the qualitative analysis of the time series (Fig. <xref ref-type="fig" rid="FE1"/>) and confirm that the performance of the DL emulators decreases with the temporal persistence of the target variable. Moreover, the RMSE values shown in Fig. <xref ref-type="fig" rid="FE1"/> provide a complementary view of model performance. For the DC, errors are substantially larger than for the other components and increase when precipitation is excluded (P1), although they remain high even when precipitation is included (P0). In contrast, DMC and FFMC show much lower RMSE values, with negligible sensitivity to the inclusion of precipitation. This behavior is fully consistent with the persistence analysis based on <inline-formula><mml:math id="M49" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and supports the interpretation that the main limitation for DC arises from the representation of long-memory dynamics rather than from missing predictor information.</p>

<table-wrap id="TE1"><label>Table E1</label><caption><p id="d2e2848">Integral timescale (<inline-formula><mml:math id="M50" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) and bias by FWI component, considering the test period predictions of the region CONTM.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="6">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Component</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M51" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>  ERA5-Land</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M52" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> P0</oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M53" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">τ</mml:mi><mml:mn mathvariant="normal">30</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> P1</oasis:entry>
         <oasis:entry colname="col5">Bias P0</oasis:entry>
         <oasis:entry colname="col6">Bias P1</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">DC</oasis:entry>
         <oasis:entry colname="col2">16.28</oasis:entry>
         <oasis:entry colname="col3">7.64</oasis:entry>
         <oasis:entry colname="col4">3.10</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M54" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">8.65</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M55" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">13.18</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">DMC</oasis:entry>
         <oasis:entry colname="col2">9.28</oasis:entry>
         <oasis:entry colname="col3">3.36</oasis:entry>
         <oasis:entry colname="col4">4.16</oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M56" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">5.92</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6"><inline-formula><mml:math id="M57" display="inline"><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">5.12</mml:mn></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">FFMC</oasis:entry>
         <oasis:entry colname="col2">3.19</oasis:entry>
         <oasis:entry colname="col3">3.56</oasis:entry>
         <oasis:entry colname="col4">3.35</oasis:entry>
         <oasis:entry colname="col5">0.38</oasis:entry>
         <oasis:entry colname="col6">0.16</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

<fig id="FE1"><label>Figure E1</label><caption><p id="d2e3039">Daily, spatially aggregated time series for the reference, and the DL-based estimates considering the input sets P0 and P1 (see Table <xref ref-type="table" rid="T1"/>) over the CONTM region (see Fig. <xref ref-type="fig" rid="FA1"/>) during the 2018–2021 test period. The first, second and third row represents the Drought Code, the Duff Moisture Code and the Fine Fuel Moisture Code respectively.</p></caption>
        
        <graphic xlink:href="https://nhess.copernicus.org/articles/26/3417/2026/nhess-26-3417-2026-f15.png"/>

      </fig>

</app>
  </app-group><notes notes-type="codedataavailability"><title>Code and data availability</title>

      <p id="d2e3058">An illustrative example providing reproducible code and data via jupyter-notebook is provided in the open GitHub repository <uri>https://github.com/SantanderMetGroup/DeepFWI</uri> <xref ref-type="bibr" rid="bib1.bibx49" id="paren.64"/>, where access to the required open data curated in Zenodo is granted and software environment configuration is detailed. Please note that due to brevity and the significant computing infrastructure required for full analysis reproducibility, the examples provided are a small sample of function calls and model configurations. These examples are applied on a coarser-than-native ERA5-Land grid for a limited time period, making them suitable for running locally on a CPU and avoiding extensive computing times. Further details or complete training/test datasets are available upon request to the authors. The data underlying the results presented in this study are openly available in Zenodo at <ext-link xlink:href="https://doi.org/10.5281/zenodo.15075367" ext-link-type="DOI">10.5281/zenodo.15075367</ext-link> <xref ref-type="bibr" rid="bib1.bibx43" id="paren.65"/>.</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d2e3076">O.M., J.B. and J.B.-M. were responsible for the development and implementation of the code used in the experiments. O.M, J.B., J.M.G. and J.B.-M. contributed to the experimental framework design. All authors contributed to the analysis and interpretation of the results. All authors contributed to manuscript writing and they reviewed and approved the final manuscript.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d2e3084">The contact author has declared that none of the authors has any competing interests.</p>
  </notes><notes notes-type="disclaimer"><title>Disclaimer</title>

      <p id="d2e3090">Publisher's note: Copernicus Publications remains neutral with regard to jurisdictional claims made in the text, published maps, institutional affiliations, or any other geographical representation in this paper. The authors bear the ultimate responsibility for providing appropriate place names. Views expressed in the text are those of the authors and do not necessarily reflect the views of the publisher.</p>
  </notes><ack><title>Acknowledgements</title><p id="d2e3096">We thank our colleague, Dr. Jose Abad, for his insightful and fruitful discussions on deep learning model configuration and explainability approaches. We are also grateful to the three anonymous reviewers for their constructive and insightful comments, which have significantly contributed to improving the quality and clarity of the manuscript. We also sincerely thank Prof. C. Gouveia for her careful and detailed assessment, which has helped us to further refine the scope, assumptions, and interpretation of this work. O.M. has received research support from grant PRE2021-100292 funded by MCIN/AEI/10.13039/501100011033, as part of the R<inline-formula><mml:math id="M58" display="inline"><mml:mo>+</mml:mo></mml:math></inline-formula>D<inline-formula><mml:math id="M59" display="inline"><mml:mo>+</mml:mo></mml:math></inline-formula>i project CORDyS (PID2020-116595RB-I00) with funding from the Spanish Ministry of Science MCIN/AEI/10.13039/501100011033. J.M.G. and J.B. acknowledge funding by the Ministry for the Ecological Transition and the Demographic Challenge (MITECO) and the European Commission NextGenerationEU (Regulation EU 2020/2094), through CSIC's Interdisciplinary Thematic Platform Clima (PTI-Clima). J.B. has received research support from Grant PID2023-149997OA-I00 (PROTECT Project) funded by MICIU/AEI/10.13039/501100011033 and by ERDF/EU. P.M.M.S. acknowledges project DHEFEUS (<ext-link xlink:href="https://doi.org/10.54499/2022.09185.PTDC" ext-link-type="DOI">10.54499/2022.09185.PTDC</ext-link>) and UID/50019/2025 and LA/P/0068/2020 <ext-link xlink:href="https://doi.org/10.54499/LA/P/0068/2020" ext-link-type="DOI">10.54499/LA/P/0068/2020</ext-link>).</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d2e3121">This research has been supported by the Ministerio de Ciencia e Innovación (grant no. PRE2021-100292), the Ministerio de Ciencia e Innovación (grant no. PID2023-149997OA-420 I00), the Fundação para a Ciência e a Tecnologia (grant no. 2022.09185.PTDC), and the Fundação para a Ciência e a Tecnologia (grant nos. UID/50019/2025 and LA/P/0068/2020). J.M.G. and J.B. have been supported by the Ministry for the Ecological Transition and the Demographic Challenge (MITECO) and the European Commission NextGenerationEU (Regulation EU 2020/2094), through CSIC's Interdisciplinary Thematic Platform Clima (PTI-Clima).The article processing charges for this open-access publication were covered by the CSIC Open Access Publication Support Initiative through its Unit of Information Resources for Research (URICI).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d2e3132">This paper was edited by Joaquim G. Pinto and reviewed by Célia Gouveia and three anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>Abatzoglou et al.(2018)Abatzoglou, Williams, Boschetti, Zubkova, and Kolden</label><mixed-citation>Abatzoglou, J. T., Williams, A. P., Boschetti, L., Zubkova, M., and Kolden, C. A.: Global patterns of interannual climate–fire relationships, Global Change Biol., 24, 5164–5175, <ext-link xlink:href="https://doi.org/10.1111/gcb.14405" ext-link-type="DOI">10.1111/gcb.14405</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Baño-Medina et al.(2022)Baño-Medina, Manzanas, Cimadevilla, Fernández, González-Abad, Cofiño, and Gutiérrez</label><mixed-citation>Baño-Medina, J., Manzanas, R., Cimadevilla, E., Fernández, J., González-Abad, J., Cofiño, A. S., and Gutiérrez, J. M.: Downscaling multi-model climate projection ensembles with deep learning (DeepESD): contribution to CORDEX EUR-44, Geosci. Model Dev., 15, 6747–6758, <ext-link xlink:href="https://doi.org/10.5194/gmd-15-6747-2022" ext-link-type="DOI">10.5194/gmd-15-6747-2022</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Baño-Medina et al.(2024)Baño-Medina, Iturbide, Fernández, and Gutiérrez</label><mixed-citation>Baño-Medina, J., Iturbide, M., Fernández, J., and Gutiérrez, J. M.: Transferability and Explainability of Deep Learning Emulators for Regional Climate Model Projections: Perspectives for Future Applications, Artificial Intelligence for the Earth Systems, 3, <ext-link xlink:href="https://doi.org/10.1175/AIES-D-23-0099.1" ext-link-type="DOI">10.1175/AIES-D-23-0099.1</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>Baño-Medina et al.(2025)Baño-Medina, Sengupta, Doyle, Reynolds, Watson-Parris, and Monache</label><mixed-citation>Baño-Medina, J., Sengupta, A., Doyle, J. D., Reynolds, C. A., Watson-Parris, D., and Monache, L. D.: Are AI weather models learning atmospheric physics? A sensitivity analysis of cyclone Xynthia, npj Clim. Atmos. Sci., 8, 1–9, <ext-link xlink:href="https://doi.org/10.1038/s41612-025-00949-6" ext-link-type="DOI">10.1038/s41612-025-00949-6</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>Bedia et al.(2014a)Bedia, Herrera, Camia, Moreno, and Gutierrez</label><mixed-citation>Bedia, J., Herrera, S., Camia, A., Moreno, J. M., and Gutierrez, J. M.: Forest Fire Danger Projections in the Mediterranean using ENSEMBLES Regional Climate Change Scenarios, Clim. Change, 122, 185–199, <ext-link xlink:href="https://doi.org/10.1007/s10584-013-1005-z" ext-link-type="DOI">10.1007/s10584-013-1005-z</ext-link>, 2014a.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>Bedia et al.(2014b)Bedia, Herrera, and Gutierrez</label><mixed-citation>Bedia, J., Herrera, S., and Gutiérrez, J. M.: Assessing the predictability of fire occurrence and area burned across phytoclimatic regions in Spain, Nat. Hazards Earth Syst. Sci., 14, 53–66, <ext-link xlink:href="https://doi.org/10.5194/nhess-14-53-2014" ext-link-type="DOI">10.5194/nhess-14-53-2014</ext-link>, 2014b.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Bedia et al.(2015)Bedia, Herrera, Gutierrez, Benali, Brands, Mota, and Moreno</label><mixed-citation>Bedia, J., Herrera, S., Gutierrez, J., Benali, A., Brands, S., Mota, B., and Moreno, J.: Global patterns in the sensitivity of burned area to fire-weather: implications for climate change, Agr. Forest Meteorol., 214–215, 369–379, <ext-link xlink:href="https://doi.org/10.1016/j.agrformet.2015.09.002" ext-link-type="DOI">10.1016/j.agrformet.2015.09.002</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Bedia et al.(2018)Bedia, Golding, Casanueva, Iturbide, Buontempo, and Gutiérrez</label><mixed-citation>Bedia, J., Golding, N., Casanueva, A., Iturbide, M., Buontempo, C., and Gutiérrez, J. M.: Seasonal predictions of Fire Weather Index: Paving the way for their operational applicability in Mediterranean Europe, Clim. Services, 9, 101–110, <ext-link xlink:href="https://doi.org/10.1016/j.cliser.2017.04.001" ext-link-type="DOI">10.1016/j.cliser.2017.04.001</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Bedia et al.(2020)Bedia, Baño-Medina, Legasa, Iturbide, Manzanas, Herrera, Casanueva, San-Martín, Cofiño, and Gutiérrez</label><mixed-citation>Bedia, J., Baño-Medina, J., Legasa, M. N., Iturbide, M., Manzanas, R., Herrera, S., Casanueva, A., San-Martín, D., Cofiño, A. S., and Gutiérrez, J. M.: Statistical downscaling with the downscaleR package (v3.1.0): contribution to the VALUE intercomparison experiment, Geosci. Model Dev., 13, 1711–1735, <ext-link xlink:href="https://doi.org/10.5194/gmd-13-1711-2020" ext-link-type="DOI">10.5194/gmd-13-1711-2020</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>Bento et al.(2023)Bento, Lima, Santos, Lima, Russo, Nunes, DaCamara, Trigo, and Soares</label><mixed-citation>Bento, V. A., Lima, D. C. A., Santos, L. C., Lima, M. M., Russo, A., Nunes, S. A., DaCamara, C. C., Trigo, R. M., and Soares, P. M. M.: The future of extreme meteorological fire danger under climate change scenarios for Iberia, Weather Clim. Extrem., 42, 100623, <ext-link xlink:href="https://doi.org/10.1016/j.wace.2023.100623" ext-link-type="DOI">10.1016/j.wace.2023.100623</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Bowman et al.(2009)Bowman, Balch, Artaxo, Bond, Carlson, Cochrane, D'Antonio, DeFries, Doyle, Harrison, Johnston, Keeley, Krawchuk, Kull, Marston, Moritz, Prentice, Roos, Scott, Swetnam, van der Werf, and Pyne</label><mixed-citation>Bowman, D. M. J. S., Balch, J. K., Artaxo, P., Bond, W. J., Carlson, J. M., Cochrane, M. A., D'Antonio, C. M., DeFries, R. S., Doyle, J. C., Harrison, S. P., Johnston, F. H., Keeley, J. E., Krawchuk, M. A., Kull, C. A., Marston, J. B., Moritz, M. A., Prentice, I. C., Roos, C. I., Scott, A. C., Swetnam, T. W., van der Werf, G. R., and Pyne, S. J.: Fire in the Earth System, Science, 324, 481–484, <ext-link xlink:href="https://doi.org/10.1126/science.1163886" ext-link-type="DOI">10.1126/science.1163886</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Bowman et al.(2017)Bowman, Williamson, Abatzoglou, Kolden, Cochrane, and Smith</label><mixed-citation>Bowman, D. M. J. S., Williamson, G. J., Abatzoglou, J. T., Kolden, C. A., Cochrane, M. A., and Smith, A. M. S.: Human exposure and sensitivity to globally extreme wildfire events, Nat. Ecol. Evolut., 1, 0058, <ext-link xlink:href="https://doi.org/10.1038/s41559-016-0058" ext-link-type="DOI">10.1038/s41559-016-0058</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Buontempo et al.(2022)Buontempo, Burgess, Dee, Pinty, Thépaut, Rixen, Almond, Armstrong, Brookshaw, Alos, Bell, Bergeron, Cagnazzo, Comyn-Platt, Damasio-Da-Costa, Guillory, Hersbach, Horányi, Nicolas, Obregon, Ramos, Raoult, Muñoz-Sabater, Simmons, Soci, Suttie, Vamborg, Varndell, Vermoote, Yang, and Marcilla</label><mixed-citation>Buontempo, C., Burgess, S. N., Dee, D., Pinty, B., Thépaut, J.-N., Rixen, M., Almond, S., Armstrong, D., Brookshaw, A., Alos, A. L., Bell, B., Bergeron, C., Cagnazzo, C., Comyn-Platt, E., Damasio-Da-Costa, E., Guillory, A., Hersbach, H., Horányi, A., Nicolas, J., Obregon, A., Ramos, E. P., Raoult, B., Muñoz-Sabater, J., Simmons, A., Soci, C., Suttie, M., Vamborg, F., Varndell, J., Vermoote, S., Yang, X., and Marcilla, J. G. d.: The Copernicus Climate Change Service: Climate Science in Action, B. Am. Meteorol. Soc., 103, E2669–E2687, <ext-link xlink:href="https://doi.org/10.1175/BAMS-D-21-0315.1" ext-link-type="DOI">10.1175/BAMS-D-21-0315.1</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Bushenkova et al.(2024)Bushenkova, Soares, Johannsen, and Lima</label><mixed-citation>Bushenkova, A., Soares, P. M. M., Johannsen, F., and Lima, D. C. A.: Towards an improved representation of the urban heat island effect: A multi-scale  application of XGBoost for madrid, Urban Clim., 55, 101982,  <ext-link xlink:href="https://doi.org/10.1016/j.uclim.2024.101982" ext-link-type="DOI">10.1016/j.uclim.2024.101982</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Camia and Amatulli(2009)</label><mixed-citation>Camia, A. and Amatulli, G.: Weather Factors and Fire Danger in the Mediterranean, in: Earth Observation of Wildland Fires in Mediterranean Ecosystems, edited by: Chuvieco, E., pp. 71–82, Springer, Berlin, Heidelberg, <ext-link xlink:href="https://doi.org/10.1007/978-3-642-01754-4_6" ext-link-type="DOI">10.1007/978-3-642-01754-4_6</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Cinquini et al.(2014)Cinquini, Crichton, Mattmann, Harney, Shipman, Wang, Ananthakrishnan, Miller, Denvil, Morgan, Pobre, Bell, Doutriaux, Drach, Williams, Kershaw, Pascoe, Gonzalez, Fiore, and Schweitzer</label><mixed-citation>Cinquini, L., Crichton, D., Mattmann, C., Harney, J., Shipman, G., Wang, F., Ananthakrishnan, R., Miller, N., Denvil, S., Morgan, M., Pobre, Z., Bell, G. M., Doutriaux, C., Drach, R., Williams, D., Kershaw, P., Pascoe, S., Gonzalez, E., Fiore, S., and Schweitzer, R.: The Earth System Grid Federation: An open infrastructure for access to distributed geospatial data, Future Gener. Comp. Sy., 36, 400–417, <ext-link xlink:href="https://doi.org/10.1016/j.future.2013.07.002" ext-link-type="DOI">10.1016/j.future.2013.07.002</ext-link>, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Costa et al.(2020)Costa, De Rigo, Libertà, Houston Durrant, and San-Miguel-Ayanz</label><mixed-citation>Costa, H., De Rigo, D., Libertà, G., Houston Durrant, T., and San-Miguel-Ayanz, J.: European wildfire danger and vulnerability in a changing climate – Towards integrating risk dimensions – JRC PESETA IV project – Task 9 – forest fires, European Commission and Joint Research Centre, Publications Office of the European Union, <ext-link xlink:href="https://doi.org/10.2760/46951" ext-link-type="DOI">10.2760/46951</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>DaCamara et al.(2024)DaCamara, Bento, Nunes, Lemos, Soares, and Trigo</label><mixed-citation>DaCamara, C. C., Bento, V. A., Nunes, S. A., Lemos, G., Soares, P. M. M., and Trigo, R. M.: Impacts of fire prevention strategies in a changing climate: an assessment for Portugal, Environ. Res.-Clim., 3, 045002, <ext-link xlink:href="https://doi.org/10.1088/2752-5295/ad574f" ext-link-type="DOI">10.1088/2752-5295/ad574f</ext-link>,  2024.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>Di Giuseppe et al.(2020)Di Giuseppe, Vitolo, Krzeminski, Barnard, Maciel, and San-Miguel</label><mixed-citation>Di Giuseppe, F., Vitolo, C., Krzeminski, B., Barnard, C., Maciel, P., and San-Miguel, J.: Fire Weather Index: the skill provided by the European Centre for Medium-Range Weather Forecasts ensemble prediction system, Nat. Hazards Earth Syst. Sci., 20, 2365–2378, <ext-link xlink:href="https://doi.org/10.5194/nhess-20-2365-2020" ext-link-type="DOI">10.5194/nhess-20-2365-2020</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Doury et al.(2024)Doury, Somot, and Gadat</label><mixed-citation>Doury, A., Somot, S., and Gadat, S.: On the suitability of a convolutional neural network based RCM-emulator for fine spatio-temporal precipitation, Clim. Dynam., 62, 8587–8613, <ext-link xlink:href="https://doi.org/10.1007/s00382-024-07350-8" ext-link-type="DOI">10.1007/s00382-024-07350-8</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Dramsch et al.(2025)Dramsch, Kuglitsch, Fernández-Torres, Toreti, Albayrak, Nava, Ghaffarian, Cheng, Ma, Samek, Venguswamy, Koul, Muthuregunathan, and Hrast Essenfelder</label><mixed-citation>Dramsch, J. S., Kuglitsch, M. M., Fernández-Torres, M. A., Toreti, A., Albayrak, R. A., Nava, L., Ghaffarian, S., Cheng, X., Ma, J., Samek, W., Venguswamy, R., Koul, A., Muthuregunathan, R., and Hrast Essenfelder, A.: Explainability can foster trust in artificial intelligence in geoscience, Nat. Geosci., 18, 112–114, <ext-link xlink:href="https://doi.org/10.1038/s41561-025-01639-x" ext-link-type="DOI">10.1038/s41561-025-01639-x</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Duffy et al.(2023)Duffy, Vandal, Wang, Nemani, and Ganguly</label><mixed-citation>Duffy, K., Vandal, T. J., Wang, W., Nemani, R. R., and Ganguly, A. R.: A Framework for Deep Learning Emulation of Numerical Models With a Case Study in Satellite Remote Sensing, IEEE T. Neural Network. Learn. Syst., 34, 3345–3356, <ext-link xlink:href="https://doi.org/10.1109/TNNLS.2022.3169958" ext-link-type="DOI">10.1109/TNNLS.2022.3169958</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>Dupuy et al.(2020)Dupuy, Fargeon, Martin-StPaul, Pimont, Ruffault, Guijarro, Hernando, Madrigal, and Fernandes</label><mixed-citation>Dupuy, J.-l., Fargeon, H., Martin-StPaul, N., Pimont, F., Ruffault, J., Guijarro, M., Hernando, C., Madrigal, J., and Fernandes, P.: Climate change impact on future wildfire danger and activity in southern Europe: a review, Ann. For. Sci., 77, 1–24, <ext-link xlink:href="https://doi.org/10.1007/s13595-020-00933-5" ext-link-type="DOI">10.1007/s13595-020-00933-5</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>Fugioka et al.(2009)</label><mixed-citation>Fugioka, F. M., Gill, A., Viegas, D. X., and Wotton, B.: Fire Danger and  Fire Behavior Modeling Systems in Australia, Europe, and North  America, in: Developments in Environmental Science, edited by:  Bytnerowicz, A., Arbaugh, M., Riebau, A., and Andersen, C., Elsevier B.V.,  The Netherlands, <ext-link xlink:href="https://doi.org/10.1016/s1474-8177(08)00021-1" ext-link-type="DOI">10.1016/s1474-8177(08)00021-1</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Gincheva et al.(2024)Gincheva, Pausas, Torres-Vázquez, Bedia, Vicente-Serrano, Abatzoglou, Sánchez-Espigares, Chuvieco, Jerez, Provenzale, Trigo, and Turco</label><mixed-citation>Gincheva, A., Pausas, J. G., Torres-Vázquez, M. A., Bedia, J., Vicente-Serrano, S. M., Abatzoglou, J. T., Sánchez-Espigares, J. A., Chuvieco, E., Jerez, S., Provenzale, A., Trigo, R. M., and Turco, M.: The Interannual Variability of Global Burned Area Is Mostly Explained by Climatic Drivers, Earth's Fut., 12, e2023EF004334, <ext-link xlink:href="https://doi.org/10.1029/2023EF004334" ext-link-type="DOI">10.1029/2023EF004334</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Glorot and Bengio(2010)</label><mixed-citation>Glorot, X. and Bengio, Y.: Understanding the difficulty of training deep  feedforward neural networks, in: Proceedings of the Thirteenth  International Conference on Artificial Intelligence and Statistics, JMLR Workshop and Conference Proceedings, iSSN: 1938-7228, 249–256, <uri>https://proceedings.mlr.press/v9/glorot10a.html</uri> (last access: 17 July 2026), 2010.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>González-Abad and Gutiérrez(2024)</label><mixed-citation>González-Abad, J. and Gutiérrez, J. M.: Are Deep Learning Methods Suitable for Downscaling Global Climate Projections? Review and Intercomparison of Existing Models, arXiv:2411.05850 [physics], <ext-link xlink:href="https://doi.org/10.48550/arXiv.2411.05850" ext-link-type="DOI">10.48550/arXiv.2411.05850</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx28"><label>González-Abad et al.(2023)González-Abad, Baño-Medina, and Gutiérrez</label><mixed-citation>González-Abad, J., Baño-Medina, J., and Gutiérrez, J. M.: Using Explainability to Inform Statistical Downscaling Based on Deep Learning Beyond Standard Validation Approaches, J. Adv. Model. Earth Syst., 15, e2023MS003641, <ext-link xlink:href="https://doi.org/10.1029/2023MS003641" ext-link-type="DOI">10.1029/2023MS003641</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx29"><label>Herrera et al.(2013)Herrera, Bedia, Gutierrez, Fernandez, and Moreno</label><mixed-citation>Herrera, S., Bedia, J., Gutierrez, J. M., Fernandez, J., and Moreno, J. M.: On the projection of future fire danger conditions with various instantaneous/mean-daily data sources, Clim. Change, 118, 827–840, <ext-link xlink:href="https://doi.org/10.1007/s10584-012-0667-2" ext-link-type="DOI">10.1007/s10584-012-0667-2</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Hersbach et al.(2020)Hersbach, Bell, Berrisford, Hirahara, Horányi, Muñoz‐Sabater, Nicolas, Peubey, Radu, Schepers, Simmons, Soci, Abdalla, Abellan, Balsamo, Bechtold, Biavati, Bidlot, Bonavita, Chiara, Dahlgren, Dee, Diamantakis, Dragani, Flemming, Forbes, Fuentes, Geer, Haimberger, Healy, Hogan, Hólm, Janisková, Keeley, Laloyaux, Lopez, Lupu, Radnoti, Rosnay, Rozum, Vamborg, Villaume, and Thépaut</label><mixed-citation>Hersbach, H., Bell, B., Berrisford, P., Hirahara, S., Horányi, A., Muñoz‐Sabater, J., Nicolas, J., Peubey, C., Radu, R., Schepers, D., Simmons, A., Soci, C., Abdalla, S., Abellan, X., Balsamo, G., Bechtold, P., Biavati, G., Bidlot, J., Bonavita, M., Chiara, G., Dahlgren, P., Dee, D., Diamantakis, M., Dragani, R., Flemming, J., Forbes, R., Fuentes, M., Geer, A., Haimberger, L., Healy, S., Hogan, R. J., Hólm, E., Janisková, M., Keeley, S., Laloyaux, P., Lopez, P., Lupu, C., Radnoti, G., Rosnay, P., Rozum, I., Vamborg, F., Villaume, S., and Thépaut, J.: The ERA5 global reanalysis, Q. J. Roy.  Meteorol. Soc., 146, 1999–2049, <ext-link xlink:href="https://doi.org/10.1002/qj.3803" ext-link-type="DOI">10.1002/qj.3803</ext-link>, 2020.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Hogan and Mason(2011)</label><mixed-citation>Hogan, R. J. and Mason, I. B.: Deterministic Forecasts of Binary Events, chap. 3, pp. 31–59, John Wiley &amp; Sons, Ltd, ISBN 9781119960003, <ext-link xlink:href="https://doi.org/10.1002/9781119960003.ch3" ext-link-type="DOI">10.1002/9781119960003.ch3</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Iturbide et al.(2019)Iturbide, Bedia, Herrera, Baño-Medina, Fernández, Frías, Manzanas, San-Martín, Cimadevilla, Cofiño, and Gutiérrez</label><mixed-citation>Iturbide, M., Bedia, J., Herrera, S., Baño-Medina, J., Fernández, J., Frías, M., Manzanas, R., San-Martín, D., Cimadevilla, E., Cofiño, A., and Gutiérrez, J.: The R-based climate4R open framework for reproducible climate data access and post-processing, Environ. Modell. Softw., 111, 42–54, <ext-link xlink:href="https://doi.org/10.1016/j.envsoft.2018.09.009" ext-link-type="DOI">10.1016/j.envsoft.2018.09.009</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Jain et al.(2022)Jain, Castellanos-Acuna, Coogan, Abatzoglou, and Flannigan</label><mixed-citation>Jain, P., Castellanos-Acuna, D., Coogan, S. C. P., Abatzoglou, J. T., and Flannigan, M. D.: Observed increases in extreme fire weather driven by atmospheric humidity and temperature, Nat. Clim. Change, 12, 63–70, <ext-link xlink:href="https://doi.org/10.1038/s41558-021-01224-1" ext-link-type="DOI">10.1038/s41558-021-01224-1</ext-link>,  2022.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Johannsen et al.(2024)Johannsen, Soares, and Langendijk</label><mixed-citation>Johannsen, F., Soares, P. M. M., and Langendijk, G. S.: On the deep learning approach for improving the representation of urban climate: The Paris urban heat island and temperature extremes, Urban Clim., 56, 102039, <ext-link xlink:href="https://doi.org/10.1016/j.uclim.2024.102039" ext-link-type="DOI">10.1016/j.uclim.2024.102039</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>Jolly et al.(2015)Jolly, Cochrane, Freeborn, Holden, Brown, Williamson, and Bowman</label><mixed-citation>Jolly, W. M., Cochrane, M. A., Freeborn, P. H., Holden, Z. A., Brown, T. J., Williamson, G. J., and Bowman, D. M. J. S.: Climate-induced variations in global wildfire danger from 1979 to 2013, Nat. Commun., 6, 7537, <ext-link xlink:href="https://doi.org/10.1038/ncomms8537" ext-link-type="DOI">10.1038/ncomms8537</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>Kondylatos et al.(2022)Kondylatos, Prapas, Ronco, Papoutsis, Camps-Valls, Piles, Fernández-Torres, and Carvalhais</label><mixed-citation>Kondylatos, S., Prapas, I., Ronco, M., Papoutsis, I., Camps-Valls, G., Piles, M., Fernández-Torres, M.-A., and Carvalhais, N.: Wildfire Danger Prediction and Understanding With Deep Learning, Geophys. Res. Lett., 49, e2022GL099368, <ext-link xlink:href="https://doi.org/10.1029/2022GL099368" ext-link-type="DOI">10.1029/2022GL099368</ext-link> 2022.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Lawson and Armitage(2008)</label><mixed-citation>Lawson, B. D. and Armitage, O. B.: Weather Guide for the Canadian Forest Fire Danger Rating System, Tech. rep., Canadian Forest Service, <uri>https://publications.gc.ca/collections/collection_2009/nrcan/Fo134-8-2008E.pdf</uri> (last access: 17 July 2026), 2008.</mixed-citation></ref>
      <ref id="bib1.bibx38"><label>Li et al.(2025)Li, Wan, Yu, and Wang</label><mixed-citation>Li, S., Wan, H., Yu, Q., and Wang, X.: Downscaling of ERA5 reanalysis land surface temperature based on attention mechanism and Google Earth Engine, Sci. Rep., 15, 675, <ext-link xlink:href="https://doi.org/10.1038/s41598-024-83944-w" ext-link-type="DOI">10.1038/s41598-024-83944-w</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx39"><label>Mamalakis et al.(2023)Mamalakis, Barnes, and Hurrell</label><mixed-citation>Mamalakis, A., Barnes, E. A., and Hurrell, J. W.: Using Explainable Artificial Intelligence to Quantify “Climate Distinguishability” After Stratospheric Aerosol Injection, Geophys. Res. Lett., 50, e2023GL106137, <ext-link xlink:href="https://doi.org/10.1029/2023GL106137" ext-link-type="DOI">10.1029/2023GL106137</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx40"><label>Matteo et al.(2025)Matteo, Garnés-Morales, Moreno, Andreia, Azorín-Molina, Bedia, Giuseppe, Dunn, Herrera, Provenzale, Quilcaille, Vázquez, and Turco</label><mixed-citation>Matteo, A., Garnés-Morales, G., Moreno, A., Andreia, R., Azorín-Molina, C., Bedia, J., Giuseppe, F. D., Dunn, R. J. H., Herrera, S., Provenzale, A., Quilcaille, Y., Vázquez, M. A. T., and Turco, M.: Challenges in assessing Fire Weather changes in a warming climate, arXiv:2503.01818 [physics], <ext-link xlink:href="https://doi.org/10.48550/arXiv.2503.01818" ext-link-type="DOI">10.48550/arXiv.2503.01818</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx41"><label>McGovern et al.(2019)McGovern, Lagerquist, Gagne, Jergensen, Elmore, Homeyer, and Smith</label><mixed-citation>McGovern, A., Lagerquist, R., Gagne, D. J., Jergensen, G. E., Elmore, K. L., Homeyer, C. R., and Smith, T.: Making the Black Box More Transparent: Understanding the Physical Implications of Machine Learning, B. Am. Meteorol. Soc., 100, 2175–2199, <ext-link xlink:href="https://doi.org/10.1175/BAMS-D-18-0195.1" ext-link-type="DOI">10.1175/BAMS-D-18-0195.1</ext-link>,  2019.</mixed-citation></ref>
      <ref id="bib1.bibx42"><label>Mirones et al.(2025)Mirones, Baño-Medina, Brands, and Bedia</label><mixed-citation>Mirones, O., Baño-Medina, J., Brands, S., and Bedia, J.: Toward Spatio-Temporally Consistent Multi-Site Fire Danger Downscaling With Explainable Deep Learning, J. Geophys. Res.-Machine Learn. Comput., 2, e2024JH000331, <ext-link xlink:href="https://doi.org/10.1029/2024JH000331" ext-link-type="DOI">10.1029/2024JH000331</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx43"><label>Mirones et al.(2025)</label><mixed-citation>Mirones, Ó., Bedia Jiménez, J., and Baño-Medina, J.: Toy Dataset for Emulating the Fire Weather Index (FWI) Using Deep Learning Techniques, Zenodo [data set], <ext-link xlink:href="https://doi.org/10.5281/zenodo.15075367" ext-link-type="DOI">10.5281/zenodo.15075367</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx44"><label>Muñoz-Sabater et al.(2021)Muñoz-Sabater, Dutra, Agustí-Panareda, Albergel, Arduini, Balsamo, Boussetta, Choulga, Harrigan, Hersbach, Martens, Miralles, Piles, Rodríguez-Fernández, Zsoter, Buontempo, and Thépaut</label><mixed-citation>Muñoz-Sabater, J., Dutra, E., Agustí-Panareda, A., Albergel, C., Arduini, G., Balsamo, G., Boussetta, S., Choulga, M., Harrigan, S., Hersbach, H., Martens, B., Miralles, D. G., Piles, M., Rodríguez-Fernández, N. J., Zsoter, E., Buontempo, C., and Thépaut, J.-N.: ERA5-Land: a state-of-the-art global reanalysis dataset for land applications, Earth Syst. Sci. Data, 13, 4349–4383, <ext-link xlink:href="https://doi.org/10.5194/essd-13-4349-2021" ext-link-type="DOI">10.5194/essd-13-4349-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx45"><label>Pausas and Keeley(2021)</label><mixed-citation>Pausas, J. G. and Keeley, J. E.: Wildfires and global change, Front. Ecol. Environ., 19, 387–395, <ext-link xlink:href="https://doi.org/10.1002/fee.2359" ext-link-type="DOI">10.1002/fee.2359</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx46"><label>Prince(2023)</label><mixed-citation>Prince, S. J.: Understanding Deep Learning, The MIT Press,  <uri>http://udlbook.com</uri> (last access: March 2025), 2023.</mixed-citation></ref>
      <ref id="bib1.bibx47"><label>Rampal et al.(2024)Rampal, Hobeichi, Gibson, Abramowitz, Beucler, Gonz, Chapman, Harder, and Guti</label><mixed-citation>Rampal, N., Hobeichi, S., Gibson, P. B., Abramowitz, G., Beucler, T., Gonz, J., Chapman, W., Harder, P., and Guti, M.: Enhancing Regional Climate Downscaling through Advances in Machine Learning, IEEE T. Neural Netw. Learn. Syst., 3, <ext-link xlink:href="https://doi.org/10.1175/AIES-D-23-0066.1" ext-link-type="DOI">10.1175/AIES-D-23-0066.1</ext-link>,  2024. </mixed-citation></ref>
      <ref id="bib1.bibx48"><label>San-Miguel-Ayanz et al.(2013)San-Miguel-Ayanz, Schulte, Schmuck, and Camia</label><mixed-citation>San-Miguel-Ayanz, J., Schulte, E., Schmuck, G., and Camia, A.: The European Forest Fire Information System in the context of environmental policies of the European Union, For. Pol. Econom., 29, 19–25, <ext-link xlink:href="https://doi.org/10.1016/j.forpol.2011.08.012" ext-link-type="DOI">10.1016/j.forpol.2011.08.012</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx49"><label>SantanderMetGroup(2026)</label><mixed-citation>SantanderMetGroup: DeepFWI, GitHub [data set], <uri>https://github.com/SantanderMetGroup/DeepFWI</uri>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx50"><label>Santos et al.(2023)Santos, Lima, Bento, Nunes, DaCamara, Russo, Soares, and Trigo</label><mixed-citation>Santos, L. C., Lima, M. M., Bento, V. A., Nunes, S. A., DaCamara, C. C., Russo, A., Soares, P. M. M., and Trigo, R. M.: An Evaluation of the Atmospheric Instability Effect on Wildfire Danger Using ERA5 over the Iberian Peninsula, Fire, 6, 120, <ext-link xlink:href="https://doi.org/10.3390/fire6030120" ext-link-type="DOI">10.3390/fire6030120</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx51"><label>Soares et al.(2024)Soares, Johannsen, Lima, Lemos, Bento, and Bushenkova</label><mixed-citation>Soares, P. M. M., Johannsen, F., Lima, D. C. A., Lemos, G., Bento, V. A., and Bushenkova, A.: High-resolution downscaling of CMIP6 Earth system and global climate models using deep learning for Iberia, Geosci. Model Dev., 17, 229–259, <ext-link xlink:href="https://doi.org/10.5194/gmd-17-229-2024" ext-link-type="DOI">10.5194/gmd-17-229-2024</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx52"><label>Stocks et al.(1989)Stocks, Lawson, Alexander, Wagner, McAlpine, Lynham, and Dubé</label><mixed-citation>Stocks, B. J., Lawson, B. D., Alexander, M. E., Wagner, C. E. V., McAlpine, R. S., Lynham, T. J., and Dubé, D. E.: The Canadian Forest Fire Danger Rating System: An Overview, The Forestry Chronicle, 65, 450–457, <ext-link xlink:href="https://doi.org/10.5558/tfc65450-6" ext-link-type="DOI">10.5558/tfc65450-6</ext-link>, 1989.</mixed-citation></ref>
      <ref id="bib1.bibx53"><label>Sundararajan et al.(2017)</label><mixed-citation>Sundararajan, M., Taly, A., and Yan, Q.: Axiomatic attribution for deep  networks, in: International conference on machine learning, pp. 3319–3328,  PMLR, <uri>https://dl.acm.org/doi/10.5555/3305890.3306024</uri> (last access: 17 July 2026), 2017.</mixed-citation></ref>
      <ref id="bib1.bibx54"><label>Toms et al.(2021)Toms, Kashinath, Prabhat, and Yang</label><mixed-citation>Toms, B. A., Kashinath, K., Prabhat, and Yang, D.: Testing the reliability of interpretable neural networks in geoscience using the Madden–Julian oscillation, Geosci. Model Dev., 14, 4495–4508, <ext-link xlink:href="https://doi.org/10.5194/gmd-14-4495-2021" ext-link-type="DOI">10.5194/gmd-14-4495-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx55"><label>van Wagner(1987)</label><mixed-citation>van Wagner, C. E.: Development and structure of the Canadian Forest Fire Weather Index System, Tech. rep., Minister of Supply and Services Canada, Ottawa, <uri>https://ostrnrcan-dostrncan.canada.ca/handle/1845/228434</uri> (last access: 17 July 2026), 1987.</mixed-citation></ref>
      <ref id="bib1.bibx56"><label>Wotton(2009)</label><mixed-citation>Wotton, B. M.: Interpreting and using outputs from the Canadian Forest Fire Danger Rating System in research applications, Environ. Ecol. Stat., 16, 107–131, <ext-link xlink:href="https://doi.org/10.1007/s10651-007-0084-2" ext-link-type="DOI">10.1007/s10651-007-0084-2</ext-link>, 2009.</mixed-citation></ref>
      <ref id="bib1.bibx57"><label>Yang et al.(2024)Yang, Hu, Li, Mu, Yu, Xia, Li, Dasgupta, and Xiong</label><mixed-citation>Yang, R., Hu, J., Li, Z., Mu, J., Yu, T., Xia, J., Li, X., Dasgupta, A., and Xiong, H.: Interpretable machine learning for weather and climate prediction: A review, Atmos. Environ., 338, 120797, <ext-link xlink:href="https://doi.org/10.1016/j.atmosenv.2024.120797" ext-link-type="DOI">10.1016/j.atmosenv.2024.120797</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx58"><label>Yu et al.(2023)Yu, Feng, Wang, and Wright</label><mixed-citation>Yu, G., Feng, Y., Wang, J., and Wright, D. B.: Performance of Fire Danger Indices and Their Utility in Predicting Future Wildfire Danger Over the Conterminous United States, Earth's Fut., 11, e2023EF003823, <ext-link xlink:href="https://doi.org/10.1029/2023EF003823" ext-link-type="DOI">10.1029/2023EF003823</ext-link>, 2023.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Deep Learning Emulation of Multivariate Climate Indices: A Case Study of the Fire Weather Index in the Iberian Peninsula</article-title-html>
<abstract-html/>
<ref-html id="bib1.bib1"><label>Abatzoglou et al.(2018)Abatzoglou, Williams, Boschetti, Zubkova, and
Kolden</label><mixed-citation>
      
Abatzoglou, J. T., Williams, A. P., Boschetti, L., Zubkova, M., and Kolden,
C. A.: Global patterns of interannual climate–fire relationships, Global
Change Biol., 24, 5164–5175, <a href="https://doi.org/10.1111/gcb.14405" target="_blank">https://doi.org/10.1111/gcb.14405</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Baño-Medina et al.(2022)Baño-Medina, Manzanas, Cimadevilla,
Fernández, González-Abad, Cofiño, and
Gutiérrez</label><mixed-citation>
      
Baño-Medina, J., Manzanas, R., Cimadevilla, E., Fernández, J., González-Abad, J., Cofiño, A. S., and Gutiérrez, J. M.: Downscaling multi-model climate projection ensembles with deep learning (DeepESD): contribution to CORDEX EUR-44, Geosci. Model Dev., 15, 6747–6758, <a href="https://doi.org/10.5194/gmd-15-6747-2022" target="_blank">https://doi.org/10.5194/gmd-15-6747-2022</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Baño-Medina et al.(2024)Baño-Medina, Iturbide, Fernández, and
Gutiérrez</label><mixed-citation>
      
Baño-Medina, J., Iturbide, M., Fernández, J., and Gutiérrez, J. M.:
Transferability and Explainability of Deep Learning Emulators for
Regional Climate Model Projections: Perspectives for Future
Applications, Artificial Intelligence for the Earth Systems, 3,
<a href="https://doi.org/10.1175/AIES-D-23-0099.1" target="_blank">https://doi.org/10.1175/AIES-D-23-0099.1</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>Baño-Medina et al.(2025)Baño-Medina, Sengupta, Doyle, Reynolds,
Watson-Parris, and Monache</label><mixed-citation>
      
Baño-Medina, J., Sengupta, A., Doyle, J. D., Reynolds, C. A., Watson-Parris,
D., and Monache, L. D.: Are AI weather models learning atmospheric physics?
A sensitivity analysis of cyclone Xynthia, npj Clim. Atmos.
Sci., 8, 1–9, <a href="https://doi.org/10.1038/s41612-025-00949-6" target="_blank">https://doi.org/10.1038/s41612-025-00949-6</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>Bedia et al.(2014a)Bedia, Herrera, Camia, Moreno, and
Gutierrez</label><mixed-citation>
      
Bedia, J., Herrera, S., Camia, A., Moreno, J. M., and Gutierrez, J. M.: Forest
Fire Danger Projections in the Mediterranean using ENSEMBLES
Regional Climate Change Scenarios, Clim. Change, 122, 185–199,
<a href="https://doi.org/10.1007/s10584-013-1005-z" target="_blank">https://doi.org/10.1007/s10584-013-1005-z</a>, 2014a.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>Bedia et al.(2014b)Bedia, Herrera, and
Gutierrez</label><mixed-citation>
      
Bedia, J., Herrera, S., and Gutiérrez, J. M.: Assessing the predictability of fire occurrence and area burned across phytoclimatic regions in Spain, Nat. Hazards Earth Syst. Sci., 14, 53–66, <a href="https://doi.org/10.5194/nhess-14-53-2014" target="_blank">https://doi.org/10.5194/nhess-14-53-2014</a>, 2014b.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Bedia et al.(2015)Bedia, Herrera, Gutierrez, Benali, Brands, Mota,
and Moreno</label><mixed-citation>
      
Bedia, J., Herrera, S., Gutierrez, J., Benali, A., Brands, S., Mota, B., and
Moreno, J.: Global patterns in the sensitivity of burned area to
fire-weather: implications for climate change, Agr. Forest
Meteorol., 214–215, 369–379, <a href="https://doi.org/10.1016/j.agrformet.2015.09.002" target="_blank">https://doi.org/10.1016/j.agrformet.2015.09.002</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Bedia et al.(2018)Bedia, Golding, Casanueva, Iturbide, Buontempo, and
Gutiérrez</label><mixed-citation>
      
Bedia, J., Golding, N., Casanueva, A., Iturbide, M., Buontempo, C., and
Gutiérrez, J. M.: Seasonal predictions of Fire Weather Index: Paving
the way for their operational applicability in Mediterranean Europe,
Clim. Services, 9, 101–110, <a href="https://doi.org/10.1016/j.cliser.2017.04.001" target="_blank">https://doi.org/10.1016/j.cliser.2017.04.001</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Bedia et al.(2020)Bedia, Baño-Medina, Legasa, Iturbide, Manzanas,
Herrera, Casanueva, San-Martín, Cofiño, and
Gutiérrez</label><mixed-citation>
      
Bedia, J., Baño-Medina, J., Legasa, M. N., Iturbide, M., Manzanas, R., Herrera, S., Casanueva, A., San-Martín, D., Cofiño, A. S., and Gutiérrez, J. M.: Statistical downscaling with the downscaleR package (v3.1.0): contribution to the VALUE intercomparison experiment, Geosci. Model Dev., 13, 1711–1735, <a href="https://doi.org/10.5194/gmd-13-1711-2020" target="_blank">https://doi.org/10.5194/gmd-13-1711-2020</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>Bento et al.(2023)Bento, Lima, Santos, Lima, Russo, Nunes, DaCamara,
Trigo, and Soares</label><mixed-citation>
      
Bento, V. A., Lima, D. C. A., Santos, L. C., Lima, M. M., Russo, A., Nunes,
S. A., DaCamara, C. C., Trigo, R. M., and Soares, P. M. M.: The future of
extreme meteorological fire danger under climate change scenarios for
Iberia, Weather Clim. Extrem., 42, 100623,
<a href="https://doi.org/10.1016/j.wace.2023.100623" target="_blank">https://doi.org/10.1016/j.wace.2023.100623</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Bowman et al.(2009)Bowman, Balch, Artaxo, Bond, Carlson, Cochrane,
D'Antonio, DeFries, Doyle, Harrison, Johnston, Keeley, Krawchuk, Kull,
Marston, Moritz, Prentice, Roos, Scott, Swetnam, van der Werf, and
Pyne</label><mixed-citation>
      
Bowman, D. M. J. S., Balch, J. K., Artaxo, P., Bond, W. J., Carlson, J. M.,
Cochrane, M. A., D'Antonio, C. M., DeFries, R. S., Doyle, J. C., Harrison,
S. P., Johnston, F. H., Keeley, J. E., Krawchuk, M. A., Kull, C. A., Marston,
J. B., Moritz, M. A., Prentice, I. C., Roos, C. I., Scott, A. C., Swetnam,
T. W., van der Werf, G. R., and Pyne, S. J.: Fire in the Earth System,
Science, 324, 481–484, <a href="https://doi.org/10.1126/science.1163886" target="_blank">https://doi.org/10.1126/science.1163886</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Bowman et al.(2017)Bowman, Williamson, Abatzoglou, Kolden, Cochrane,
and Smith</label><mixed-citation>
      
Bowman, D. M. J. S., Williamson, G. J., Abatzoglou, J. T., Kolden, C. A.,
Cochrane, M. A., and Smith, A. M. S.: Human exposure and sensitivity to
globally extreme wildfire events, Nat. Ecol. Evolut., 1, 0058,
<a href="https://doi.org/10.1038/s41559-016-0058" target="_blank">https://doi.org/10.1038/s41559-016-0058</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Buontempo et al.(2022)Buontempo, Burgess, Dee, Pinty, Thépaut,
Rixen, Almond, Armstrong, Brookshaw, Alos, Bell, Bergeron, Cagnazzo,
Comyn-Platt, Damasio-Da-Costa, Guillory, Hersbach, Horányi, Nicolas,
Obregon, Ramos, Raoult, Muñoz-Sabater, Simmons, Soci, Suttie, Vamborg,
Varndell, Vermoote, Yang, and Marcilla</label><mixed-citation>
      
Buontempo, C., Burgess, S. N., Dee, D., Pinty, B., Thépaut, J.-N., Rixen, M.,
Almond, S., Armstrong, D., Brookshaw, A., Alos, A. L., Bell, B., Bergeron,
C., Cagnazzo, C., Comyn-Platt, E., Damasio-Da-Costa, E., Guillory, A.,
Hersbach, H., Horányi, A., Nicolas, J., Obregon, A., Ramos, E. P., Raoult,
B., Muñoz-Sabater, J., Simmons, A., Soci, C., Suttie, M., Vamborg, F.,
Varndell, J., Vermoote, S., Yang, X., and Marcilla, J. G. d.: The
Copernicus Climate Change Service: Climate Science in Action,
B. Am. Meteorol. Soc., 103, E2669–E2687,
<a href="https://doi.org/10.1175/BAMS-D-21-0315.1" target="_blank">https://doi.org/10.1175/BAMS-D-21-0315.1</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Bushenkova et al.(2024)Bushenkova, Soares, Johannsen, and
Lima</label><mixed-citation>
      
Bushenkova, A., Soares, P. M. M., Johannsen, F., and Lima, D. C. A.: Towards an improved representation of the urban heat island effect: A multi-scale  application of XGBoost for madrid, Urban Clim., 55, 101982,  <a href="https://doi.org/10.1016/j.uclim.2024.101982" target="_blank">https://doi.org/10.1016/j.uclim.2024.101982</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Camia and Amatulli(2009)</label><mixed-citation>
      
Camia, A. and Amatulli, G.: Weather Factors and Fire Danger in the
Mediterranean, in: Earth Observation of Wildland Fires in
Mediterranean Ecosystems, edited by: Chuvieco, E., pp. 71–82, Springer,
Berlin, Heidelberg, <a href="https://doi.org/10.1007/978-3-642-01754-4_6" target="_blank">https://doi.org/10.1007/978-3-642-01754-4_6</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Cinquini et al.(2014)Cinquini, Crichton, Mattmann, Harney, Shipman,
Wang, Ananthakrishnan, Miller, Denvil, Morgan, Pobre, Bell, Doutriaux, Drach,
Williams, Kershaw, Pascoe, Gonzalez, Fiore, and
Schweitzer</label><mixed-citation>
      
Cinquini, L., Crichton, D., Mattmann, C., Harney, J., Shipman, G., Wang, F.,
Ananthakrishnan, R., Miller, N., Denvil, S., Morgan, M., Pobre, Z., Bell,
G. M., Doutriaux, C., Drach, R., Williams, D., Kershaw, P., Pascoe, S.,
Gonzalez, E., Fiore, S., and Schweitzer, R.: The Earth System Grid
Federation: An open infrastructure for access to distributed geospatial
data, Future Gener. Comp. Sy., 36, 400–417,
<a href="https://doi.org/10.1016/j.future.2013.07.002" target="_blank">https://doi.org/10.1016/j.future.2013.07.002</a>, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Costa et al.(2020)Costa, De Rigo, Libertà, Houston Durrant, and
San-Miguel-Ayanz</label><mixed-citation>
      
Costa, H., De Rigo, D., Libertà, G., Houston Durrant, T., and
San-Miguel-Ayanz, J.: European wildfire danger and vulnerability in a
changing climate – Towards integrating risk dimensions – JRC PESETA IV
project – Task 9 – forest fires, European Commission and Joint Research
Centre, Publications Office of the European Union, <a href="https://doi.org/10.2760/46951" target="_blank">https://doi.org/10.2760/46951</a>,
2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>DaCamara et al.(2024)DaCamara, Bento, Nunes, Lemos, Soares, and
Trigo</label><mixed-citation>
      
DaCamara, C. C., Bento, V. A., Nunes, S. A., Lemos, G., Soares, P. M. M., and
Trigo, R. M.: Impacts of fire prevention strategies in a changing climate: an
assessment for Portugal, Environ. Res.-Clim., 3, 045002,
<a href="https://doi.org/10.1088/2752-5295/ad574f" target="_blank">https://doi.org/10.1088/2752-5295/ad574f</a>,  2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>Di Giuseppe et al.(2020)Di Giuseppe, Vitolo, Krzeminski, Barnard,
Maciel, and San-Miguel</label><mixed-citation>
      
Di Giuseppe, F., Vitolo, C., Krzeminski, B., Barnard, C., Maciel, P., and San-Miguel, J.: Fire Weather Index: the skill provided by the European Centre for Medium-Range Weather Forecasts ensemble prediction system, Nat. Hazards Earth Syst. Sci., 20, 2365–2378, <a href="https://doi.org/10.5194/nhess-20-2365-2020" target="_blank">https://doi.org/10.5194/nhess-20-2365-2020</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Doury et al.(2024)Doury, Somot, and Gadat</label><mixed-citation>
      
Doury, A., Somot, S., and Gadat, S.: On the suitability of a convolutional
neural network based RCM-emulator for fine spatio-temporal precipitation,
Clim. Dynam., 62, 8587–8613, <a href="https://doi.org/10.1007/s00382-024-07350-8" target="_blank">https://doi.org/10.1007/s00382-024-07350-8</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Dramsch et al.(2025)Dramsch, Kuglitsch, Fernández-Torres, Toreti,
Albayrak, Nava, Ghaffarian, Cheng, Ma, Samek, Venguswamy, Koul,
Muthuregunathan, and Hrast Essenfelder</label><mixed-citation>
      
Dramsch, J. S., Kuglitsch, M. M., Fernández-Torres, M. A., Toreti, A.,
Albayrak, R. A., Nava, L., Ghaffarian, S., Cheng, X., Ma, J., Samek, W.,
Venguswamy, R., Koul, A., Muthuregunathan, R., and Hrast Essenfelder, A.:
Explainability can foster trust in artificial intelligence in geoscience,
Nat. Geosci., 18, 112–114, <a href="https://doi.org/10.1038/s41561-025-01639-x" target="_blank">https://doi.org/10.1038/s41561-025-01639-x</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Duffy et al.(2023)Duffy, Vandal, Wang, Nemani, and
Ganguly</label><mixed-citation>
      
Duffy, K., Vandal, T. J., Wang, W., Nemani, R. R., and Ganguly, A. R.: A
Framework for Deep Learning Emulation of Numerical Models With
a Case Study in Satellite Remote Sensing, IEEE T.
Neural Network. Learn. Syst., 34, 3345–3356,
<a href="https://doi.org/10.1109/TNNLS.2022.3169958" target="_blank">https://doi.org/10.1109/TNNLS.2022.3169958</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Dupuy et al.(2020)Dupuy, Fargeon, Martin-StPaul, Pimont, Ruffault,
Guijarro, Hernando, Madrigal, and Fernandes</label><mixed-citation>
      
Dupuy, J.-l., Fargeon, H., Martin-StPaul, N., Pimont, F., Ruffault, J.,
Guijarro, M., Hernando, C., Madrigal, J., and Fernandes, P.: Climate change
impact on future wildfire danger and activity in southern Europe: a review,
Ann. For. Sci., 77, 1–24, <a href="https://doi.org/10.1007/s13595-020-00933-5" target="_blank">https://doi.org/10.1007/s13595-020-00933-5</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>Fugioka et al.(2009)</label><mixed-citation>
      
Fugioka, F. M., Gill, A., Viegas, D. X., and Wotton, B.: Fire Danger and  Fire Behavior Modeling Systems in Australia, Europe, and North  America, in: Developments in Environmental Science, edited by:  Bytnerowicz, A., Arbaugh, M., Riebau, A., and Andersen, C., Elsevier B.V.,  The Netherlands, <a href="https://doi.org/10.1016/s1474-8177(08)00021-1" target="_blank">https://doi.org/10.1016/s1474-8177(08)00021-1</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Gincheva et al.(2024)Gincheva, Pausas, Torres-Vázquez, Bedia,
Vicente-Serrano, Abatzoglou, Sánchez-Espigares, Chuvieco, Jerez, Provenzale,
Trigo, and Turco</label><mixed-citation>
      
Gincheva, A., Pausas, J. G., Torres-Vázquez, M. A., Bedia, J.,
Vicente-Serrano, S. M., Abatzoglou, J. T., Sánchez-Espigares, J. A.,
Chuvieco, E., Jerez, S., Provenzale, A., Trigo, R. M., and Turco, M.: The
Interannual Variability of Global Burned Area Is Mostly
Explained by Climatic Drivers, Earth's Fut., 12, e2023EF004334,
<a href="https://doi.org/10.1029/2023EF004334" target="_blank">https://doi.org/10.1029/2023EF004334</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Glorot and Bengio(2010)</label><mixed-citation>
      
Glorot, X. and Bengio, Y.: Understanding the difficulty of training deep  feedforward neural networks, in: Proceedings of the Thirteenth  International Conference on Artificial Intelligence and Statistics, JMLR Workshop and Conference Proceedings, iSSN: 1938-7228, 249–256, <a href="https://proceedings.mlr.press/v9/glorot10a.html" target="_blank"/> (last access: 17 July 2026), 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>González-Abad and Gutiérrez(2024)</label><mixed-citation>
      
González-Abad, J. and Gutiérrez, J. M.: Are Deep Learning Methods
Suitable for Downscaling Global Climate Projections? Review and
Intercomparison of Existing Models,
arXiv:2411.05850 [physics], <a href="https://doi.org/10.48550/arXiv.2411.05850" target="_blank">https://doi.org/10.48550/arXiv.2411.05850</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>González-Abad et al.(2023)González-Abad, Baño-Medina, and
Gutiérrez</label><mixed-citation>
      
González-Abad, J., Baño-Medina, J., and Gutiérrez, J. M.: Using
Explainability to Inform Statistical Downscaling Based on Deep
Learning Beyond Standard Validation Approaches, J. Adv. Model. Earth Syst., 15, e2023MS003641, <a href="https://doi.org/10.1029/2023MS003641" target="_blank">https://doi.org/10.1029/2023MS003641</a>,
2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Herrera et al.(2013)Herrera, Bedia, Gutierrez, Fernandez, and
Moreno</label><mixed-citation>
      
Herrera, S., Bedia, J., Gutierrez, J. M., Fernandez, J., and Moreno, J. M.: On
the projection of future fire danger conditions with various
instantaneous/mean-daily data sources, Clim. Change, 118, 827–840,
<a href="https://doi.org/10.1007/s10584-012-0667-2" target="_blank">https://doi.org/10.1007/s10584-012-0667-2</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Hersbach et al.(2020)Hersbach, Bell, Berrisford, Hirahara, Horányi,
Muñoz‐Sabater, Nicolas, Peubey, Radu, Schepers, Simmons, Soci, Abdalla,
Abellan, Balsamo, Bechtold, Biavati, Bidlot, Bonavita, Chiara, Dahlgren, Dee,
Diamantakis, Dragani, Flemming, Forbes, Fuentes, Geer, Haimberger, Healy,
Hogan, Hólm, Janisková, Keeley, Laloyaux, Lopez, Lupu, Radnoti, Rosnay,
Rozum, Vamborg, Villaume, and Thépaut</label><mixed-citation>
      
Hersbach, H., Bell, B., Berrisford, P., Hirahara, S., Horányi, A.,
Muñoz‐Sabater, J., Nicolas, J., Peubey, C., Radu, R., Schepers, D.,
Simmons, A., Soci, C., Abdalla, S., Abellan, X., Balsamo, G., Bechtold, P.,
Biavati, G., Bidlot, J., Bonavita, M., Chiara, G., Dahlgren, P., Dee, D.,
Diamantakis, M., Dragani, R., Flemming, J., Forbes, R., Fuentes, M., Geer,
A., Haimberger, L., Healy, S., Hogan, R. J., Hólm, E., Janisková, M.,
Keeley, S., Laloyaux, P., Lopez, P., Lupu, C., Radnoti, G., Rosnay, P.,
Rozum, I., Vamborg, F., Villaume, S., and Thépaut, J.: The ERA5 global
reanalysis, Q. J. Roy.  Meteorol. Soc., 146,
1999–2049, <a href="https://doi.org/10.1002/qj.3803" target="_blank">https://doi.org/10.1002/qj.3803</a>, 2020.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Hogan and Mason(2011)</label><mixed-citation>
      
Hogan, R. J. and Mason, I. B.: Deterministic Forecasts of Binary Events,
chap. 3, pp. 31–59, John Wiley &amp; Sons, Ltd, ISBN 9781119960003,
<a href="https://doi.org/10.1002/9781119960003.ch3" target="_blank">https://doi.org/10.1002/9781119960003.ch3</a>, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Iturbide et al.(2019)Iturbide, Bedia, Herrera, Baño-Medina,
Fernández, Frías, Manzanas, San-Martín, Cimadevilla, Cofiño,
and Gutiérrez</label><mixed-citation>
      
Iturbide, M., Bedia, J., Herrera, S., Baño-Medina, J., Fernández, J.,
Frías, M., Manzanas, R., San-Martín, D., Cimadevilla, E., Cofiño,
A., and Gutiérrez, J.: The R-based climate4R open framework for
reproducible climate data access and post-processing, Environ. Modell.
Softw., 111, 42–54, <a href="https://doi.org/10.1016/j.envsoft.2018.09.009" target="_blank">https://doi.org/10.1016/j.envsoft.2018.09.009</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Jain et al.(2022)Jain, Castellanos-Acuna, Coogan, Abatzoglou, and
Flannigan</label><mixed-citation>
      
Jain, P., Castellanos-Acuna, D., Coogan, S. C. P., Abatzoglou, J. T., and
Flannigan, M. D.: Observed increases in extreme fire weather driven by
atmospheric humidity and temperature, Nat. Clim. Change, 12, 63–70,
<a href="https://doi.org/10.1038/s41558-021-01224-1" target="_blank">https://doi.org/10.1038/s41558-021-01224-1</a>,  2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Johannsen et al.(2024)Johannsen, Soares, and
Langendijk</label><mixed-citation>
      
Johannsen, F., Soares, P. M. M., and Langendijk, G. S.: On the deep learning
approach for improving the representation of urban climate: The Paris
urban heat island and temperature extremes, Urban Clim., 56, 102039,
<a href="https://doi.org/10.1016/j.uclim.2024.102039" target="_blank">https://doi.org/10.1016/j.uclim.2024.102039</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>Jolly et al.(2015)Jolly, Cochrane, Freeborn, Holden, Brown,
Williamson, and Bowman</label><mixed-citation>
      
Jolly, W. M., Cochrane, M. A., Freeborn, P. H., Holden, Z. A., Brown, T. J.,
Williamson, G. J., and Bowman, D. M. J. S.: Climate-induced variations in
global wildfire danger from 1979 to 2013, Nat. Commun., 6, 7537,
<a href="https://doi.org/10.1038/ncomms8537" target="_blank">https://doi.org/10.1038/ncomms8537</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>Kondylatos et al.(2022)Kondylatos, Prapas, Ronco, Papoutsis,
Camps-Valls, Piles, Fernández-Torres, and
Carvalhais</label><mixed-citation>
      
Kondylatos, S., Prapas, I., Ronco, M., Papoutsis, I., Camps-Valls, G., Piles,
M., Fernández-Torres, M.-A., and Carvalhais, N.: Wildfire Danger
Prediction and Understanding With Deep Learning, Geophys.
Res. Lett., 49, e2022GL099368, <a href="https://doi.org/10.1029/2022GL099368" target="_blank">https://doi.org/10.1029/2022GL099368</a> 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Lawson and Armitage(2008)</label><mixed-citation>
      
Lawson, B. D. and Armitage, O. B.: Weather Guide for the Canadian Forest Fire Danger Rating System, Tech. rep., Canadian Forest Service, <a href="https://publications.gc.ca/collections/collection_2009/nrcan/Fo134-8-2008E.pdf" target="_blank"/> (last access: 17 July 2026), 2008.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Li et al.(2025)Li, Wan, Yu, and Wang</label><mixed-citation>
      
Li, S., Wan, H., Yu, Q., and Wang, X.: Downscaling of ERA5 reanalysis land
surface temperature based on attention mechanism and Google Earth
Engine, Sci. Rep., 15, 675, <a href="https://doi.org/10.1038/s41598-024-83944-w" target="_blank">https://doi.org/10.1038/s41598-024-83944-w</a>,
2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Mamalakis et al.(2023)Mamalakis, Barnes, and
Hurrell</label><mixed-citation>
      
Mamalakis, A., Barnes, E. A., and Hurrell, J. W.: Using Explainable Artificial
Intelligence to Quantify “Climate Distinguishability” After Stratospheric
Aerosol Injection, Geophys. Res. Lett., 50, e2023GL106137,
<a href="https://doi.org/10.1029/2023GL106137" target="_blank">https://doi.org/10.1029/2023GL106137</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Matteo et al.(2025)Matteo, Garnés-Morales, Moreno, Andreia,
Azorín-Molina, Bedia, Giuseppe, Dunn, Herrera, Provenzale, Quilcaille,
Vázquez, and Turco</label><mixed-citation>
      
Matteo, A., Garnés-Morales, G., Moreno, A., Andreia, R., Azorín-Molina, C.,
Bedia, J., Giuseppe, F. D., Dunn, R. J. H., Herrera, S., Provenzale, A.,
Quilcaille, Y., Vázquez, M. A. T., and Turco, M.: Challenges in assessing
Fire Weather changes in a warming climate, arXiv:2503.01818 [physics],
<a href="https://doi.org/10.48550/arXiv.2503.01818" target="_blank">https://doi.org/10.48550/arXiv.2503.01818</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>McGovern et al.(2019)McGovern, Lagerquist, Gagne, Jergensen, Elmore,
Homeyer, and Smith</label><mixed-citation>
      
McGovern, A., Lagerquist, R., Gagne, D. J., Jergensen, G. E., Elmore, K. L.,
Homeyer, C. R., and Smith, T.: Making the Black Box More Transparent:
Understanding the Physical Implications of Machine Learning,
B. Am. Meteorol. Soc., 100, 2175–2199,
<a href="https://doi.org/10.1175/BAMS-D-18-0195.1" target="_blank">https://doi.org/10.1175/BAMS-D-18-0195.1</a>,  2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>Mirones et al.(2025)Mirones, Baño-Medina, Brands, and
Bedia</label><mixed-citation>
      
Mirones, O., Baño-Medina, J., Brands, S., and Bedia, J.: Toward
Spatio-Temporally Consistent Multi-Site Fire Danger
Downscaling With Explainable Deep Learning, J. Geophys.
Res.-Machine Learn. Comput., 2, e2024JH000331,
<a href="https://doi.org/10.1029/2024JH000331" target="_blank">https://doi.org/10.1029/2024JH000331</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Mirones et al.(2025)</label><mixed-citation>
      
Mirones, Ó., Bedia Jiménez, J., and Baño-Medina, J.: Toy Dataset for Emulating the Fire Weather Index (FWI) Using Deep Learning Techniques, Zenodo [data set], <a href="https://doi.org/10.5281/zenodo.15075367" target="_blank">https://doi.org/10.5281/zenodo.15075367</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>Muñoz-Sabater et al.(2021)Muñoz-Sabater, Dutra, Agustí-Panareda,
Albergel, Arduini, Balsamo, Boussetta, Choulga, Harrigan, Hersbach, Martens,
Miralles, Piles, Rodríguez-Fernández, Zsoter, Buontempo, and
Thépaut</label><mixed-citation>
      
Muñoz-Sabater, J., Dutra, E., Agustí-Panareda, A., Albergel, C., Arduini, G., Balsamo, G., Boussetta, S., Choulga, M., Harrigan, S., Hersbach, H., Martens, B., Miralles, D. G., Piles, M., Rodríguez-Fernández, N. J., Zsoter, E., Buontempo, C., and Thépaut, J.-N.: ERA5-Land: a state-of-the-art global reanalysis dataset for land applications, Earth Syst. Sci. Data, 13, 4349–4383, <a href="https://doi.org/10.5194/essd-13-4349-2021" target="_blank">https://doi.org/10.5194/essd-13-4349-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Pausas and Keeley(2021)</label><mixed-citation>
      
Pausas, J. G. and Keeley, J. E.: Wildfires and global change, Front.
Ecol. Environ., 19, 387–395, <a href="https://doi.org/10.1002/fee.2359" target="_blank">https://doi.org/10.1002/fee.2359</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Prince(2023)</label><mixed-citation>
      
Prince, S. J.: Understanding Deep Learning, The MIT Press,  <a href="http://udlbook.com" target="_blank"/> (last access: March 2025), 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Rampal et al.(2024)Rampal, Hobeichi, Gibson, Abramowitz, Beucler,
Gonz, Chapman, Harder, and Guti</label><mixed-citation>
      
Rampal, N., Hobeichi, S., Gibson, P. B., Abramowitz, G., Beucler, T., Gonz, J.,
Chapman, W., Harder, P., and Guti, M.: Enhancing Regional Climate
Downscaling through Advances in Machine Learning, IEEE T.
Neural Netw. Learn. Syst., 3, <a href="https://doi.org/10.1175/AIES-D-23-0066.1" target="_blank">https://doi.org/10.1175/AIES-D-23-0066.1</a>,  2024.


    </mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>San-Miguel-Ayanz et al.(2013)San-Miguel-Ayanz, Schulte, Schmuck, and
Camia</label><mixed-citation>
      
San-Miguel-Ayanz, J., Schulte, E., Schmuck, G., and Camia, A.: The European
Forest Fire Information System in the context of environmental
policies of the European Union, For. Pol. Econom., 29, 19–25,
<a href="https://doi.org/10.1016/j.forpol.2011.08.012" target="_blank">https://doi.org/10.1016/j.forpol.2011.08.012</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>SantanderMetGroup(2026)</label><mixed-citation>
      
SantanderMetGroup: DeepFWI, GitHub [data set], <a href="https://github.com/SantanderMetGroup/DeepFWI" target="_blank"/>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Santos et al.(2023)Santos, Lima, Bento, Nunes, DaCamara, Russo,
Soares, and Trigo</label><mixed-citation>
      
Santos, L. C., Lima, M. M., Bento, V. A., Nunes, S. A., DaCamara, C. C., Russo,
A., Soares, P. M. M., and Trigo, R. M.: An Evaluation of the Atmospheric
Instability Effect on Wildfire Danger Using ERA5 over the
Iberian Peninsula, Fire, 6, 120, <a href="https://doi.org/10.3390/fire6030120" target="_blank">https://doi.org/10.3390/fire6030120</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Soares et al.(2024)Soares, Johannsen, Lima, Lemos, Bento, and
Bushenkova</label><mixed-citation>
      
Soares, P. M. M., Johannsen, F., Lima, D. C. A., Lemos, G., Bento, V. A., and Bushenkova, A.: High-resolution downscaling of CMIP6 Earth system and global climate models using deep learning for Iberia, Geosci. Model Dev., 17, 229–259, <a href="https://doi.org/10.5194/gmd-17-229-2024" target="_blank">https://doi.org/10.5194/gmd-17-229-2024</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Stocks et al.(1989)Stocks, Lawson, Alexander, Wagner, McAlpine,
Lynham, and Dubé</label><mixed-citation>
      
Stocks, B. J., Lawson, B. D., Alexander, M. E., Wagner, C. E. V., McAlpine,
R. S., Lynham, T. J., and Dubé, D. E.: The Canadian Forest Fire
Danger Rating System: An Overview, The Forestry Chronicle, 65,
450–457, <a href="https://doi.org/10.5558/tfc65450-6" target="_blank">https://doi.org/10.5558/tfc65450-6</a>, 1989.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Sundararajan et al.(2017)</label><mixed-citation>
      
Sundararajan, M., Taly, A., and Yan, Q.: Axiomatic attribution for deep  networks, in: International conference on machine learning, pp. 3319–3328,  PMLR, <a href="https://dl.acm.org/doi/10.5555/3305890.3306024" target="_blank"/> (last access: 17 July 2026), 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Toms et al.(2021)Toms, Kashinath, Prabhat, and
Yang</label><mixed-citation>
      
Toms, B. A., Kashinath, K., Prabhat, and Yang, D.: Testing the reliability of interpretable neural networks in geoscience using the Madden–Julian oscillation, Geosci. Model Dev., 14, 4495–4508, <a href="https://doi.org/10.5194/gmd-14-4495-2021" target="_blank">https://doi.org/10.5194/gmd-14-4495-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>van Wagner(1987)</label><mixed-citation>
      
van Wagner, C. E.: Development and structure of the Canadian Forest Fire
Weather Index System, Tech. rep., Minister of Supply and Services
Canada, Ottawa, <a href="https://ostrnrcan-dostrncan.canada.ca/handle/1845/228434" target="_blank"/> (last access: 17 July 2026), 1987.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Wotton(2009)</label><mixed-citation>
      
Wotton, B. M.: Interpreting and using outputs from the Canadian Forest
Fire Danger Rating System in research applications, Environ.
Ecol. Stat., 16, 107–131, <a href="https://doi.org/10.1007/s10651-007-0084-2" target="_blank">https://doi.org/10.1007/s10651-007-0084-2</a>, 2009.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Yang et al.(2024)Yang, Hu, Li, Mu, Yu, Xia, Li, Dasgupta, and
Xiong</label><mixed-citation>
      
Yang, R., Hu, J., Li, Z., Mu, J., Yu, T., Xia, J., Li, X., Dasgupta, A., and
Xiong, H.: Interpretable machine learning for weather and climate prediction:
A review, Atmos. Environ., 338, 120797,
<a href="https://doi.org/10.1016/j.atmosenv.2024.120797" target="_blank">https://doi.org/10.1016/j.atmosenv.2024.120797</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Yu et al.(2023)Yu, Feng, Wang, and Wright</label><mixed-citation>
      
Yu, G., Feng, Y., Wang, J., and Wright, D. B.: Performance of Fire Danger
Indices and Their Utility in Predicting Future Wildfire Danger
Over the Conterminous United States, Earth's Fut., 11,
e2023EF003823, <a href="https://doi.org/10.1029/2023EF003823" target="_blank">https://doi.org/10.1029/2023EF003823</a>, 2023.

    </mixed-citation></ref-html>--></article>
