<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing with OASIS Tables v3.0 20080202//EN" "https://jats.nlm.nih.gov/nlm-dtd/publishing/3.0/journalpub-oasis3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:oasis="http://docs.oasis-open.org/ns/oasis-exchange/table" xml:lang="en" dtd-version="3.0" article-type="research-article">
  <front>
    <journal-meta><journal-id journal-id-type="publisher">AMT</journal-id><journal-title-group>
    <journal-title>Atmospheric Measurement Techniques</journal-title>
    <abbrev-journal-title abbrev-type="publisher">AMT</abbrev-journal-title><abbrev-journal-title abbrev-type="nlm-ta">Atmos. Meas. Tech.</abbrev-journal-title>
  </journal-title-group><issn pub-type="epub">1867-8548</issn><publisher>
    <publisher-name>Copernicus Publications</publisher-name>
    <publisher-loc>Göttingen, Germany</publisher-loc>
  </publisher></journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.5194/amt-19-5659-2026</article-id><title-group><article-title>Machine learning-based emission rate estimates of global  methane super-emissions</article-title><alt-title>Machine learning-based emission rate estimates of global methane super-emissions</alt-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author" corresp="yes" rid="aff1">
          <name><surname>Roberts</surname><given-names>Clayton</given-names></name>
          <email>c.roberts@sron.nl</email>
        <ext-link>https://orcid.org/0000-0002-5184-7485</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Maasakkers</surname><given-names>Joannes D.</given-names></name>
          
        <ext-link>https://orcid.org/0000-0001-8118-0311</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>de Jong</surname><given-names>Tobias A.</given-names></name>
          
        <ext-link>https://orcid.org/0000-0002-5211-8081</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff2">
          <name><surname>Schuit</surname><given-names>Berend J.</given-names></name>
          
        <ext-link>https://orcid.org/0000-0003-1768-3592</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Sharma</surname><given-names>Shubham</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1">
          <name><surname>Huegens</surname><given-names>Theo</given-names></name>
          
        </contrib>
        <contrib contrib-type="author" corresp="no" rid="aff3">
          <name><surname>van den Berg</surname><given-names>Anne-Wil</given-names></name>
          
        <ext-link>https://orcid.org/0009-0000-1521-1570</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff4">
          <name><surname>Houweling</surname><given-names>Sander</given-names></name>
          
        <ext-link>https://orcid.org/0000-0002-6189-1009</ext-link></contrib>
        <contrib contrib-type="author" corresp="no" rid="aff1 aff4">
          <name><surname>Aben</surname><given-names>Ilse</given-names></name>
          
        </contrib>
        <aff id="aff1"><label>1</label><institution>SRON Space Research Organisation Netherlands, Leiden, the Netherlands</institution>
        </aff>
        <aff id="aff2"><label>2</label><institution>GHGSat Inc., Montreal, Canada</institution>
        </aff>
        <aff id="aff3"><label>3</label><institution>Meteorology and Air Quality group, Wageningen University &amp; Research, Wageningen, the Netherlands</institution>
        </aff>
        <aff id="aff4"><label>4</label><institution>Department of Earth Sciences, Vrije Universiteit Amsterdam, Amsterdam, the Netherlands</institution>
        </aff>
      </contrib-group>
      <author-notes><corresp id="corr1">Clayton Roberts (c.roberts@sron.nl)</corresp></author-notes><pub-date><day>8</day><month>September</month><year>2026</year></pub-date>
      
      <volume>19</volume>
      <issue>17</issue>
      <fpage>5659</fpage><lpage>5682</lpage>
      <history>
        <date date-type="received"><day>7</day><month>April</month><year>2026</year></date>
           <date date-type="rev-request"><day>19</day><month>May</month><year>2026</year></date>
           <date date-type="rev-recd"><day>21</day><month>July</month><year>2026</year></date>
           <date date-type="accepted"><day>30</day><month>July</month><year>2026</year></date>
      </history>
      <permissions>
        <copyright-statement>Copyright: © 2026 Clayton Roberts et al.</copyright-statement>
        <copyright-year>2026</copyright-year>
      <license license-type="open-access"><license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p></license></permissions><self-uri xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026.html">This article is available from https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026.html</self-uri><self-uri xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026.pdf">The full text article is available as a PDF file from https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026.pdf</self-uri>
      <abstract><title>Abstract</title>

      <p id="d2e176">Methane, the second most important anthropogenic greenhouse gas, has a global warming potential more than 80 times that of carbon dioxide over a 20-year period. Given its decadal atmospheric lifetime, reducing anthropogenic methane emissions is critical for limiting near-term warming. The TROPOspheric Monitoring Instrument (TROPOMI) provides daily global methane satellite observations, enabling rapid detection of super-emitters. Here, we develop ML-SPERE, a machine-learning framework based on a convolutional neural network trained on simulated TROPOMI methane observations and meteorological data to estimate emission rates for super-emitters. ML-SPERE outperforms the Integrated Mass Enhancement (IME) method on simulated plumes that incorporate real TROPOMI backgrounds and missing spatial data, reducing the median absolute percentage error from 42.4 % to 24.3 % for well-observed methane plumes. At low wind speeds, IME estimates exhibit a negative bias and ML-SPERE estimates do not. Applied to TROPOMI observations of a 200-d well blowout in Kazakhstan, ML-SPERE typically shows agreement with TROPOMI inverse modeling results and TROPOMI IME estimates. Compared to TROPOMI IME estimates, ML-SPERE estimates show improved agreement with IME estimates derived from high-resolution point-source imagers. Global spatial patterns of methane emissions inferred from ML-SPERE and the IME method for all super-emitters found by TROPOMI in 2021 are broadly consistent, with notable regional differences in northern Russia (where transient pipeline emissions may not be well characterized by either method), the Congo Basin (where area source emissions may not be well characterized by either method), and southeastern Australia (where IME estimates are potentially negatively biased owing to predominantly low wind speeds). Mean estimated emission rates for this dataset aggregated by estimated source sector remain similar between both methods. Overall, improved performance on simulated plumes and consistency with independent estimates for real-world observations demonstrate the utility of ML-SPERE for quantifying TROPOMI methane super-emitters.</p>
  </abstract>
    
<funding-group>
<award-group id="gs1">
<funding-source>European Space Agency</funding-source>
<award-id>SMART-CH4</award-id>
</award-group>
</funding-group>
</article-meta>
  </front>
<body>
      

<sec id="Ch1.S1" sec-type="intro">
  <label>1</label><title>Introduction</title>
      <p id="d2e188">Methane (<inline-formula><mml:math id="M1" display="inline"><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CH</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) is a powerful greenhouse gas, with a global warming potential more than 80 times greater than that of carbon dioxide (<inline-formula><mml:math id="M2" display="inline"><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CO</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula>) over a 20-year time horizon <xref ref-type="bibr" rid="bib1.bibx23" id="paren.1"/>. Due to its stronger radiative forcing and far shorter atmospheric lifetime compared to carbon dioxide, reducing anthropogenic methane emissions offers one of the most effective strategies for mitigating near-term climate change <xref ref-type="bibr" rid="bib1.bibx43" id="paren.2"/>. Super-emitters contribute disproportionately to total anthropogenic emissions, are frequently associated with correctable abnormal operating conditions, and can be highly transient <xref ref-type="bibr" rid="bib1.bibx65" id="paren.3"/>. Satellite observations have thus proven instrumental in detecting and abating methane emissions from super-emitters <xref ref-type="bibr" rid="bib1.bibx25 bib1.bibx61" id="paren.4"/>. The TROPOspheric Monitoring Instrument (TROPOMI), launched in 2017 aboard the Sentinel-5P satellite, provides daily global observations of atmospheric methane concentrations <xref ref-type="bibr" rid="bib1.bibx64 bib1.bibx22 bib1.bibx37 bib1.bibx38" id="paren.5"/>. With a spatial resolution down to <inline-formula><mml:math id="M3" display="inline"><mml:mrow><mml:mn mathvariant="normal">5.5</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">7</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:mrow></mml:math></inline-formula> at nadir, TROPOMI can be used to detect localized methane plumes associated with super-emitting point sources <xref ref-type="bibr" rid="bib1.bibx45 bib1.bibx34" id="paren.6"/>. These observations can be used to provide emission rate estimates and guide follow-up observations with high-resolution (<inline-formula><mml:math id="M4" display="inline"><mml:mo lspace="0mm">∼</mml:mo></mml:math></inline-formula>25 <inline-formula><mml:math id="M5" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula>) satellite instruments <xref ref-type="bibr" rid="bib1.bibx54" id="paren.7"/>, all of which can provide vital insight into global anthropogenic methane emissions on a near-real-time basis <xref ref-type="bibr" rid="bib1.bibx3" id="paren.8"/>. However, commonly used mass-balance and inverse modeling approaches for estimating emission rates for super-emitting methane plumes found in TROPOMI observations involve trade-offs between accuracy and computational cost, motivating the development of alternative methods that can efficiently quantify large numbers of plumes. We have thus developed a machine learning (ML)-based methodology for estimating super-emitter methane emission rates from TROPOMI observations.</p>
      <p id="d2e274">ML techniques have been applied to detect (but not to estimate emission rates for) methane plumes in TROPOMI data on a daily basis <xref ref-type="bibr" rid="bib1.bibx54" id="paren.9"/>. Such automated approaches are valuable given the impracticality of manually screening large datasets. <xref ref-type="bibr" rid="bib1.bibx54" id="text.10"/> trained a Convolutional Neural Network (CNN) to classify TROPOMI methane scenes as either likely or unlikely to contain a super-emitting plume, and employed a Support Vector Classifier (SVC) to reduce false positives prior to human verification. They also used the CNN to generate plume masks by combining the network's class activation map with a scene-specific methane threshold to delineate the plume. Applying this framework to all TROPOMI orbits in 2021 produced a catalog of nearly 3000 detections. Using bottom-up inventories, these detections were linked to the most likely underlying source sectors including urban areas, landfills, gas infrastructure, oil infrastructure, and coal mines.</p>
      <p id="d2e283"><xref ref-type="bibr" rid="bib1.bibx54" id="text.11"/> estimated methane emission rates for their detected TROPOMI plumes via the Integrated Mass Enhancement (IME) method, first developed for use with aircraft and high-resolution satellite observations of atmospheric methane concentrations <xref ref-type="bibr" rid="bib1.bibx11 bib1.bibx62" id="paren.12"/>. This mass-balance-based technique estimates emission rates by integrating the excess methane mass within a plume relative to the local background concentration and dividing it by the plume's residence time, which is typically inferred from the length of the plume and meteorological wind data. The accuracy of this method depends on several factors. Potential sources of uncertainty include the conversion of satellite-retrieved methane columns into plume enhancements and the subsequent delineation of plume pixels from background pixels <xref ref-type="bibr" rid="bib1.bibx62 bib1.bibx54" id="paren.13"/>. However, the dominant source of error arises from uncertainties in the wind field used to estimate the residence time. Wind datasets often show discrepancies of up to 50 % when compared to in-situ wind measurements at ground stations or airfields <xref ref-type="bibr" rid="bib1.bibx62" id="paren.14"/>. Moreover, the IME method requires the estimation of an effective wind speed that accounts not only for advection, but also for plume dispersion, diffusion processes, and plume rise. These effective wind speed calibrations are typically estimated using relationships with 10 <inline-formula><mml:math id="M6" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speeds, rely on simulated plumes, are instrument-specific, and show significant scatter <xref ref-type="bibr" rid="bib1.bibx62 bib1.bibx54" id="paren.15"/>. The IME method used in <xref ref-type="bibr" rid="bib1.bibx54" id="text.16"/> is an ensemble approach, and uncertainties are estimated by varying plume masking thresholds, background concentration estimates, and using three separate wind datasets to compute emission rates. Other mass-balance-based approaches like the cross-sectional flux (CSF) method <xref ref-type="bibr" rid="bib1.bibx63" id="paren.17"/> and variants of the IME method <xref ref-type="bibr" rid="bib1.bibx18" id="paren.18"/> can also be applied to estimate methane emission rates from TROPOMI plumes and have similar uncertainties. Atmospheric inversions <xref ref-type="bibr" rid="bib1.bibx39 bib1.bibx34" id="paren.19"><named-content content-type="pre">e.g.</named-content></xref> provide another avenue for estimating plume-level emission rates using atmospheric transport simulations. Compared with mass-balance approaches, Bayesian inversion methods provide emission-rate estimates that are more explicitly constrained by atmospheric transport physics and are therefore generally better suited to complex meteorological conditions and source configurations (e.g. overlapping sources), albeit at substantially greater computational cost and under the assumption that atmospheric transport can be modeled accurately.</p>
      <p id="d2e323">In response to the known limitations of mass-balance-based methods and the computational burden of atmospheric inversions, recent research has explored ML-based alternatives for estimating methane emission rates from point sources <xref ref-type="bibr" rid="bib1.bibx26" id="paren.20"/>. CNNs, a class of machine learning models well-suited for image-based tasks, have demonstrated strong performance in both classification and regression applications <xref ref-type="bibr" rid="bib1.bibx35 bib1.bibx36 bib1.bibx32 bib1.bibx60" id="paren.21"/>. In the context of methane remote sensing, CNNs have recently been applied to estimate point source emission rates directly from plume imagery (in addition to detection applications), offering the potential to reduce or eliminate dependence on meteorological datasets <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx47" id="paren.22"/>. Like other machine learning methods, CNNs are trained using input–output pairs, which in this case are methane plume images derived from satellite observations and corresponding known emission rates. Because real satellite observations never come with precise known emissions (except for controlled release experiments), training typically relies on simulated plumes <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx29 bib1.bibx48 bib1.bibx2 bib1.bibx47" id="paren.23"/>, where the emission rate is known for each generated image. CNN-based methods have shown promising results when applied to high-resolution satellite or airborne observations <xref ref-type="bibr" rid="bib1.bibx29" id="paren.24"/>. MethaNet <xref ref-type="bibr" rid="bib1.bibx28" id="paren.25"/> was able to estimate the emission rates of simulated methane plumes in high-resolution (1–5 <inline-formula><mml:math id="M7" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula>) aircraft observations with an average error of 29 %, which equals the performance of the IME method on a similar dataset reported in <xref ref-type="bibr" rid="bib1.bibx62" id="text.26"/> for source rates above 1.5 <inline-formula><mml:math id="M8" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, but without using wind data. Such CNNs can obtain comparable (or superior) performance to the IME method without using wind data because high-resolution imagery can resolve fine-scale turbulent structures within methane plumes, which implicitly encodes wind speed and direction <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx29" id="paren.27"/>. However, the relatively coarse spatial resolution of TROPOMI methane observations <xref ref-type="bibr" rid="bib1.bibx64" id="paren.28"/> limits their ability to capture such structures. ML-based approaches are also not without their disadvantages; regression dilution may see trained models exhibit estimates that are biased towards the mean of their training dataset <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx29" id="paren.29"/>, and models may fail to extrapolate beyond the geographic or emission domains of their training data <xref ref-type="bibr" rid="bib1.bibx2" id="paren.30"/>. Despite this, machine learning approaches such as CNNs may still extract meaningful patterns from TROPOMI methane data and meteorological datasets, learning complex, nonlinear relationships between observed methane enhancements and emission rates.</p>
      <p id="d2e387">In this study, we present an ML–based methodology for estimating emission rates of methane plumes from TROPOMI observations, and show that such a method can surpass the performance of the IME method. We refer to our methodology as the Machine Learning-Superemitting Plume Emission Rate Estimate, or ML-SPERE for short. Our approach leverages both TROPOMI methane plume observations and auxiliary meteorological data to produce emission rate estimates. In Sects. <xref ref-type="sec" rid="Ch1.S2.SS1"/>–<xref ref-type="sec" rid="Ch1.S2.SS4"/>, we detail the construction of our training dataset, our independent test set, model optimization and training procedures, and our methods of generating uncertainties for emission rate estimates. In Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>, we assess performance and generalization using the independent test set, and compare with IME estimates. Additionally, we contrast ML-SPERE and IME emission rate estimates for a TROPOMI plume dataset previously analyzed using both inverse modeling as well as IME estimates based on high-resolution satellite observations (Sect. <xref ref-type="sec" rid="Ch1.S3.SS2"/>; <xref ref-type="bibr" rid="bib1.bibx17" id="altparen.31"/>). We also contrast ML-SPERE and IME emission rates estimates for all 2021 TROPOMI super-emitter detections from <xref ref-type="bibr" rid="bib1.bibx55" id="text.32"/> (Sect. <xref ref-type="sec" rid="Ch1.S3.SS3"/>).</p>
</sec>
<sec id="Ch1.S2">
  <label>2</label><title>Data and methods</title>
      <p id="d2e415">In this section, we first describe the generation of synthetic training and validation datasets (Sect. <xref ref-type="sec" rid="Ch1.S2.SS1"/>), which provide the data to which ML-SPERE is exposed during model training. Next, we outline the creation of a separate and independent synthetic test dataset used to evaluate the performance of ML-SPERE (Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>). We then detail the optimization and training procedures for the model (Sect. <xref ref-type="sec" rid="Ch1.S2.SS3"/>), and describe the approach used to quantify uncertainties in the estimates produced by ML-SPERE (Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/>). Lastly, we provide a brief overview of the IME methodology used as a comparative benchmark when evaluating the performance of ML-SPERE, for both the synthetic test set and real TROPOMI methane plumes (Sect. <xref ref-type="sec" rid="Ch1.S2.SS5"/>).</p>
<sec id="Ch1.S2.SS1">
  <label>2.1</label><title>Training and validation dataset creation</title>
      <p id="d2e435">Supervised machine learning tasks require training data with accurate labels. As precisely known emission rates are not available for TROPOMI methane plume observations, we use WRF-Chem version 4.1.5 <xref ref-type="bibr" rid="bib1.bibx57 bib1.bibx15" id="paren.33"/> to simulate plumes with known emission rates. WRF-Chem is configured with three nested, centered domains, with the innermost domain containing <inline-formula><mml:math id="M9" display="inline"><mml:mrow><mml:mn mathvariant="normal">99</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">99</mml:mn></mml:mrow></mml:math></inline-formula> grid cells at a <inline-formula><mml:math id="M10" display="inline"><mml:mrow><mml:mn mathvariant="normal">4</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">4</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">km</mml:mi></mml:mrow></mml:mrow></mml:math></inline-formula> spatial resolution. Vertically, this domain consists of 38 pressure levels using a sigma coordinate system. For meteorological boundary conditions, we use the Global Data Assimilation System (GDAS) dataset at <inline-formula><mml:math id="M11" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="italic">°</mml:mi><mml:mo>×</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="italic">°</mml:mi></mml:mrow></mml:math></inline-formula> spatial and 6-hourly temporal resolution, provided by the National Centers for Environmental Protection (NCEP) <xref ref-type="bibr" rid="bib1.bibx42" id="paren.34"/>, covering the period 1 June–31 August 2019. We simulated plumes in seven different regions across the world that cover varied surface and meteorological conditions, shown in Fig. <xref ref-type="fig" rid="FA1"/>. Within each innermost domain, we initialize passive tracers at four different locations and two separate release heights (approximately 25 <inline-formula><mml:math id="M12" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> and 250 <inline-formula><mml:math id="M13" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula>, see Fig. <xref ref-type="fig" rid="FA1"/>). We choose these release heights to mitigate potential model overfitting to a single release height within the planetary boundary layer. Each tracer emits at a constant flux for the duration of the simulation. We sample WRF-Chem model output daily at 1400 <inline-formula><mml:math id="M14" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">h</mml:mi></mml:mrow></mml:math></inline-formula> local time, roughly aligning with the TROPOMI overpass. To match the pixel footprints of the corresponding TROPOMI overpass, we apply an area-weighted resampling scheme to project the simulation output onto the satellite's pixel footprints, followed by a pressure-weighted vertical averaging to compute the total column density. We train our model on the resulting resampled methane plume enhancement scenes that are divided into training and validation subsets using a random <inline-formula><mml:math id="M15" display="inline"><mml:mrow><mml:mn mathvariant="normal">90</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">10</mml:mn></mml:mrow></mml:math></inline-formula> split. We show in Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS5"/> that this choice of split is resilient to geographic overfitting by ML-SPERE.</p>
      <p id="d2e536">Data augmentation techniques, such as scaling, rotation, and translation, are crucial for training CNNs as they enhance model generalization, reduce overfitting, and, in the context of this work, improve robustness to variations in plume strength and plume orientation <xref ref-type="bibr" rid="bib1.bibx32 bib1.bibx56 bib1.bibx2" id="paren.35"/>. To increase the diversity and size of the training and validation datasets, we apply transformations to model output sampled from WRF-Chem before we resample to TROPOMI pixel footprints. First, the methane enhancements for each scene are linearly scaled to match a randomly drawn target emission rate <xref ref-type="bibr" rid="bib1.bibx28" id="paren.36"/>, sampled from a gamma distribution with <inline-formula><mml:math id="M16" display="inline"><mml:mrow><mml:mtext>location</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M17" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M18" display="inline"><mml:mrow><mml:mtext>scale</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">32.13</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M19" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, and <inline-formula><mml:math id="M20" display="inline"><mml:mrow><mml:mtext>shape</mml:mtext><mml:mo>=</mml:mo><mml:mn mathvariant="normal">1.38</mml:mn></mml:mrow></mml:math></inline-formula>. These parameters are chosen to match the distribution of IME-estimated emission rates from <xref ref-type="bibr" rid="bib1.bibx55" id="text.37"/>. The scaled plume is then randomly rotated by 0, 90, 180, or 270<inline-formula><mml:math id="M21" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">°</mml:mi></mml:mrow></mml:math></inline-formula>, followed by a random horizontal or vertical flip. After resampling to TROPOMI pixel footprints, a random crop is applied to produce a <inline-formula><mml:math id="M22" display="inline"><mml:mrow><mml:mn mathvariant="normal">32</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">32</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow></mml:mrow></mml:math></inline-formula> scene, the dimensions of which match the output of the pipeline of <xref ref-type="bibr" rid="bib1.bibx54" id="text.38"/>. The pixel containing the plume source remains within the central <inline-formula><mml:math id="M23" display="inline"><mml:mrow><mml:mn mathvariant="normal">28</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">28</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow></mml:mrow></mml:math></inline-formula> region of the cropped scene, which ensures that ML-SPERE does not require plumes to be centered within input scenes after training is complete. After data augmentation, there are 94 159 scenes in the training dataset and 10 829 scenes in the validation dataset.</p>
      <p id="d2e671">CNNs can be trained on multi-channel images, enabling the integration of additional meteorological data alongside methane data. While previous studies have trained CNNs on single-channel methane scenes for plume emission rate estimation, this approach has predominantly been applied to high-resolution instruments, where the fine spatial detail of the observations captures plume structures that convey information about the wind field <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx13 bib1.bibx29" id="paren.39"/>. Some high-resolution CNN approaches still explicitly incorporate wind information <xref ref-type="bibr" rid="bib1.bibx2" id="paren.40"/>. Our initial attempts to train a CNN for usage with TROPOMI observations without wind information exhibited over 30 % lower relative performance on our test dataset than the final model which incorporates wind channels. This is likely due to TROPOMI's coarser spatial resolution, which does not capture the same level of wind-related information in plumes as higher-resolution instruments. Consequently, the images in our training, validation, and test sets each contain seven channels, including data on the 10 <inline-formula><mml:math id="M24" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind field vector, pixel area, surface pressure, and the corresponding plume mask, shown in Fig. <xref ref-type="fig" rid="F1"/>. Channels for pixel area and surface pressure are included to provide context on methane mass contained within each pixel, which total column concentration alone does not provide. Plume masks are generated via the methodology described in <xref ref-type="bibr" rid="bib1.bibx54" id="text.41"/>. In the training and validation sets, we use the NCEP 10 <inline-formula><mml:math id="M25" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind fields, and in the test set (see Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>), we use the ERA5 10 <inline-formula><mml:math id="M26" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind fields <xref ref-type="bibr" rid="bib1.bibx21" id="paren.42"/>, i.e. the meteorology used in the respective simulations. Wind information is taken at the same timestamp as the methane observation. These extra channels undergo the same data augmentation steps as the methane channel to ensure consistency with the final augmented methane scene.</p>

      <fig id="F1" specific-use="star"><label>Figure 1</label><caption><p id="d2e718">Channels used for input to ML-SPERE. All channel dimensions are <inline-formula><mml:math id="M27" display="inline"><mml:mrow><mml:mn mathvariant="normal">32</mml:mn><mml:mo>×</mml:mo><mml:mn mathvariant="normal">32</mml:mn></mml:mrow></mml:math></inline-formula> TROPOMI pixels. <bold>(a)</bold> Methane channel. For training and validation data, resampled plume abundances are provided, but test dataset methane scenes must be background-reduced and masked to zero outside of the defined plume mask as described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/> and shown in Fig. <xref ref-type="fig" rid="F2"/>. <bold>(b–d)</bold> Wind vector information, split over three channels. <bold>(e)</bold> Surface pressure. <bold>(f)</bold>. Pixel area. <bold>(g)</bold> Binary plume mask. Prior to input to ML-SPERE, each of the seven channels is standardized independently using channel-wise statistics (mean and standard deviation) computed across all training images.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f01.png"/>

        </fig>

</sec>
<sec id="Ch1.S2.SS2">
  <label>2.2</label><title>Test dataset creation</title>
      <p id="d2e767">We evaluate our final trained model using a test set of synthetic TROPOMI methane plumes simulated with the HYSPLIT atmospheric transport model <xref ref-type="bibr" rid="bib1.bibx58 bib1.bibx8 bib1.bibx9 bib1.bibx7" id="paren.43"/>. These simulations use ERA5 global wind data <xref ref-type="bibr" rid="bib1.bibx20" id="paren.44"/> for atmospheric transport, provided at a spatial resolution of <inline-formula><mml:math id="M28" display="inline"><mml:mrow><mml:mn mathvariant="normal">0.25</mml:mn><mml:mi mathvariant="italic">°</mml:mi><mml:mo>×</mml:mo><mml:mn mathvariant="normal">0.25</mml:mn><mml:mi mathvariant="italic">°</mml:mi></mml:mrow></mml:math></inline-formula> and hourly temporal resolution. By constructing the test set using a different atmospheric transport code and meteorological inputs than those used for training, we assess the robustness of the model to changes in underlying plume transport representations. To further assess geographic generalization, we simulate test set plumes at multiple timestamps in 2019 at 224 global locations of known persistent methane emissions distinct from those used in training and validation, shown in Fig. <xref ref-type="fig" rid="FA2"/>. These “hotspot” locations are found using a wind-rotation methodology <xref ref-type="bibr" rid="bib1.bibx40" id="paren.45"/> and are reported to the International Methane Emissions Observatory <xref ref-type="bibr" rid="bib1.bibx24" id="paren.46"/>. The simulated plumes were scaled with emission rates matching the distribution described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS1"/> and processed identically to the training set, but without data augmentation (e.g. scaling or rotation). The final test set comprises 438 synthetic plumes.</p>

      <fig id="F2" specific-use="star"><label>Figure 2</label><caption><p id="d2e805"><bold>(a)</bold> simulated TROPOMI methane plume observation used in the test dataset, created as described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>. These simulated observations combine spatially resampled HYSPLIT plumes with real TROPOMI plume-free observations to create realistic synthetic observational data, complete with realistic spatial distributions of missing data. In <bold>(b)</bold>, the black outline shows the corresponding high-confidence plume mask. Valid pixels outside of this plume mask are used to estimate a background methane value for the scene. To produce the processed methane channel shown in <bold>(c)</bold>, each pixel within the high confidence plume mask is background-reduced, and pixels outside of the plume mask are set to 0. This scene is passed to the CNN along with other supplementary channels after standardization (i.e. <bold>c</bold> is used as the first channel in Fig. <xref ref-type="fig" rid="F1"/>). Background imagery in <bold>(a)</bold> relies on non-concurrent Sentinel-2 data (2022) adapted from Google Earth Engine <xref ref-type="bibr" rid="bib1.bibx14 bib1.bibx10" id="paren.47"/>.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f02.png"/>

        </fig>

      <p id="d2e836">To enhance the representativeness of the test set scenes and capture the challenges involved in estimating methane enhancements in TROPOMI data, we add a unique <inline-formula><mml:math id="M29" display="inline"><mml:mrow><mml:mn mathvariant="normal">32</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow><mml:mo>×</mml:mo><mml:mn mathvariant="normal">32</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">pixel</mml:mi></mml:mrow></mml:mrow></mml:math></inline-formula> TROPOMI observation as background methane to each scene <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx2" id="paren.48"/>. Using TROPOMI observations as backgrounds also introduces a realistic spatial distribution of missing data in the methane scene. These background scenes are selected from TROPOMI orbits from 2023 and 2024 using the method of <xref ref-type="bibr" rid="bib1.bibx54" id="text.49"/> to ensure they are highly unlikely to contain any real plumes. After combining a simulated plume with a randomly selected TROPOMI background scene, we pass the final scene back through the plume detection pipeline of <xref ref-type="bibr" rid="bib1.bibx54" id="text.50"/> to ensure that the synthetic plume is “detectable”. An example of a synthetic methane plume observation, complete with a real TROPOMI background methane and realistic missing spatial data, is shown in Fig. <xref ref-type="fig" rid="F2"/>a. As ML-SPERE is trained using methane plume enhancements, the background methane column must be removed from the test set scenes before being used as input. We perform background removal using the binary high-confidence plume mask that is generated as part of the plume detection pipeline of <xref ref-type="bibr" rid="bib1.bibx54" id="text.51"/>. First, the median methane value is calculated from pixels outside the plume mask and subtracted from all pixels in the plume mask. Then, pixels outside the plume mask are set to zero (Fig. <xref ref-type="fig" rid="F2"/>b and c). We report model performance on this test set as the primary benchmark before evaluating results on real TROPOMI observations.</p>
</sec>
<sec id="Ch1.S2.SS3">
  <label>2.3</label><title>Model optimization and training</title>
      <p id="d2e885">We develop and train ML-SPERE using TensorFlow <xref ref-type="bibr" rid="bib1.bibx1" id="paren.52"/>, a widely used machine learning framework in Python. Prior to training, both the training and validation sets are standardized on a per-channel basis relative to the training set, i.e. for each channel in an input image, values are transformed to have zero mean and unit variance based on the channel-wise mean and standard deviation of the training data. For hyperparameter optimization, we employ the <italic>Keras Tuner</italic> library <xref ref-type="bibr" rid="bib1.bibx44" id="paren.53"/> to perform a randomized grid search, fitting multiple models to the training data across a range of hyperparameter configurations. From this ensemble, we select the architecture and hyperparameters of the top-performing model, detailed in Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS2"/>. The selected model architecture is relatively shallow compared to standard benchmark CNNs such as ResNet and ImageNet <xref ref-type="bibr" rid="bib1.bibx19 bib1.bibx33" id="paren.54"/>, which is expected given the relatively small image size of the input scenes and is consistent with findings in previous work (e.g. <xref ref-type="bibr" rid="bib1.bibx54" id="altparen.55"/>), enabling efficient training that completes within a few hours on a standard desktop machine. We use the mean absolute percentage error (MAPE) as the loss function during training, chosen for its ability to reduce overfitting to high-emission scenes and provide balanced error sensitivity across the emission rate range <xref ref-type="bibr" rid="bib1.bibx6 bib1.bibx28 bib1.bibx29" id="paren.56"/>. Training is executed for a maximum of 100 epochs, halted early under two criteria: (1) if the validation MAPE does not improve for 10 consecutive epochs, or (2) if the ratio between the validation MAPE and training MAPE exceeds 1.05 for 10 epochs, indicating potential overfitting. ML-SPERE completed training after 18 epochs.</p>
</sec>
<sec id="Ch1.S2.SS4">
  <label>2.4</label><title>Emission rate uncertainty estimation</title>
      <p id="d2e918">An emission rate estimate is produced by passing a processed, seven-channel input image through the trained model, which outputs a single predicted emission rate. To estimate uncertainty in these predictions, we generate an ensemble of perturbed input images that reflect both variability in pre-processing steps known to introduce uncertainty as well as uncertainty on the input data itself. By passing this ensemble through the model and analyzing the resulting emission rate distribution, we obtain an estimate of the prediction uncertainty. Our perturbation ensemble incorporates uncertainty from three primary origins: the processing required to convert a TROPOMI methane scene into a plume abundance scene, uncertainty in the wind speed, and uncertainty associated with the trained model itself.</p>
      <p id="d2e921">There are two sources of uncertainty introduced during the processing of the methane channel: the method used to generate the plume mask, and the procedure for calculating methane enhancements above the local background (which is conditional on the definition of the plume mask). To address this, we first generate a binary plume mask following the procedure outlined in  <xref ref-type="bibr" rid="bib1.bibx54" id="text.57"/>. To account for uncertainty in this masking step, the methane scene standard deviation thresholding factor (ordinarily taken to be 1.8) is drawn from the normal distribution <inline-formula><mml:math id="M30" display="inline"><mml:mrow><mml:mi mathvariant="script">N</mml:mi><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1.8</mml:mn><mml:mo>,</mml:mo><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mn mathvariant="normal">0.2</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>. Once the plume mask is defined, background removal is applied to the methane channel as described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>. To incorporate uncertainty in the estimation of the local background methane level, we calculate the mean (<inline-formula><mml:math id="M31" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CH</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>) and standard deviation (<inline-formula><mml:math id="M32" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CH</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:msub></mml:mrow></mml:math></inline-formula>) of methane values in pixels outside the plume mask, excluding outliers beyond 1.5 times the interquartile range. A random background value is then sampled from the distribution <inline-formula><mml:math id="M33" display="inline"><mml:mrow><mml:mi mathvariant="script">N</mml:mi><mml:mo>(</mml:mo><mml:msub><mml:mi mathvariant="italic">μ</mml:mi><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CH</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mrow class="chem"><mml:msub><mml:mi mathvariant="normal">CH</mml:mi><mml:mn mathvariant="normal">4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, capturing scene-specific uncertainty in background removal. We adopt this background reduction approach because it is more conservative (i.e. it yields a larger estimate of background uncertainty) than defining background uncertainty as the standard error of the mean of background pixels or propagating the TROPOMI retrieval uncertainties on those pixels to the background mean.</p>
      <p id="d2e1011">To account for uncertainty in the wind speed channels, we first compute the average wind speed <inline-formula><mml:math id="M34" display="inline"><mml:mover accent="true"><mml:mi>w</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:math></inline-formula> across all pixels within the identified plume mask. If <inline-formula><mml:math id="M35" display="inline"><mml:mrow><mml:mover accent="true"><mml:mi>w</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>&gt;</mml:mo><mml:mn mathvariant="normal">3</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mrow></mml:math></inline-formula>, we assign a fractional wind speed uncertainty of <inline-formula><mml:math id="M36" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="false"><mml:mfrac style="text"><mml:mrow><mml:mn mathvariant="normal">1.5</mml:mn><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mrow><mml:mover accent="true"><mml:mi>w</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover></mml:mfrac></mml:mstyle></mml:mrow></mml:math></inline-formula>, using the standard deviation between in-situ airfield wind speed measurements and wind data shown in <xref ref-type="bibr" rid="bib1.bibx62" id="text.58"/>. For lower wind speeds (<inline-formula><mml:math id="M37" display="inline"><mml:mrow><mml:mover accent="true"><mml:mi>w</mml:mi><mml:mo mathvariant="normal">‾</mml:mo></mml:mover><mml:mo>&lt;</mml:mo><mml:mn mathvariant="normal">3</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mrow></mml:math></inline-formula>), we set a constant fractional uncertainty value of <inline-formula><mml:math id="M38" display="inline"><mml:mrow><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.5</mml:mn></mml:mrow></mml:math></inline-formula> (i.e. 50 % uncertainty). To incorporate this uncertainty into our ensemble of input images, we draw a wind scaling factor <inline-formula><mml:math id="M39" display="inline"><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mtext>scale</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> from a truncated normal distribution <inline-formula><mml:math id="M40" display="inline"><mml:mrow><mml:mi mathvariant="script">N</mml:mi><mml:mo>(</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mo>,</mml:mo><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msub><mml:mi mathvariant="italic">σ</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, and scale all pixel values in the wind magnitude channel by this factor. The usage of a truncated normal distribution ensures that <inline-formula><mml:math id="M41" display="inline"><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mtext>scale</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> is never negative.</p>
      <p id="d2e1179">A final source of uncertainty in our model's emission estimates stems from the model itself. To quantify this, we use Monte Carlo dropout as an approximation of Bayesian model uncertainty <xref ref-type="bibr" rid="bib1.bibx12" id="paren.59"/>. For each perturbed input image in the ensemble, we perform a single forward pass with dropout activated at inference time (dropout strength as detailed in Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS2"/>). This introduces model-based variability into the total uncertainty estimate for methane plume emissions. We repeat the entire procedure (including random generation of the plume mask, methane background threshold, wind scaling factor, and a forward pass with Monte Carlo dropout) 5000 times per scene to ensure that the ensemble is sufficiently sampled. This yields a posterior distribution of emission estimates, from which we compute summary statistics such as the mean, median, standard deviation, and 2.5th and 97.5th percentiles. Unless otherwise indicated, we report the uncertainty on the estimate as the standard deviation of the posterior distribution of emission estimates. To assess the contribution of each uncertainty source, we also generate posterior distributions while systematically omitting one of the four error components.</p>
</sec>
<sec id="Ch1.S2.SS5">
  <label>2.5</label><title>IME method overview</title>
      <p id="d2e1195">We compare the performance of ML-SPERE in estimating TROPOMI methane plume emission rates with that of the IME method. In the IME method <xref ref-type="bibr" rid="bib1.bibx62" id="paren.60"/>, the emission rate <inline-formula><mml:math id="M42" display="inline"><mml:mi>Q</mml:mi></mml:math></inline-formula> is estimated as

                <disp-formula id="Ch1.E1" content-type="numbered"><label>1</label><mml:math id="M43" display="block"><mml:mstyle displaystyle="true" class="stylechange"/><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:mi>Q</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:mfrac style="display"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mtext>eff</mml:mtext></mml:msub></mml:mrow><mml:mi>L</mml:mi></mml:mfrac></mml:mstyle><mml:mtext>IME</mml:mtext><mml:mspace linebreak="nobreak" width="1em"/><mml:mfenced close="]" open="["><mml:mrow class="unit"><mml:mi mathvariant="normal">kg</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

          where <inline-formula><mml:math id="M44" display="inline"><mml:mrow><mml:mi>L</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mfenced open="[" close="]"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:math></inline-formula> is the effective plume length (typically calculated as the square root of the area of the plume mask), <inline-formula><mml:math id="M45" display="inline"><mml:mrow><mml:mtext>IME</mml:mtext><mml:mspace width="0.125em" linebreak="nobreak"/><mml:mfenced open="[" close="]"><mml:mrow class="unit"><mml:mi mathvariant="normal">kg</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:math></inline-formula> is the total methane enhancement above the background summed over the plume mask, and <inline-formula><mml:math id="M46" display="inline"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mtext>eff</mml:mtext></mml:msub><mml:mspace linebreak="nobreak" width="0.125em"/><mml:mfenced open="[" close="]"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfenced></mml:mrow></mml:math></inline-formula> is the effective wind speed. Effective wind speed calibrations are specific to the instrument and can also vary according to which wind data products are used in the calibration. The effective wind speed calibration for TROPOMI as calculated from a 10 <inline-formula><mml:math id="M47" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speed (<inline-formula><mml:math id="M48" display="inline"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mn mathvariant="normal">10</mml:mn></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> is determined in <xref ref-type="bibr" rid="bib1.bibx54" id="text.61"/> as

                <disp-formula id="Ch1.E2" content-type="numbered"><label>2</label><mml:math id="M49" display="block"><mml:mstyle class="stylechange" displaystyle="true"/><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mi>U</mml:mi><mml:mtext>eff</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.59</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi>U</mml:mi><mml:mn mathvariant="normal">10</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.0</mml:mn><mml:mspace linebreak="nobreak" width="1em"/><mml:mfenced open="[" close="]"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfenced><mml:mo>,</mml:mo></mml:mrow></mml:math></disp-formula>

          whereas the effective wind speed calibration for a planetary boundary layer average wind speed (<inline-formula><mml:math id="M50" display="inline"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mtext>PBL</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula>) is given as

                <disp-formula id="Ch1.E3" content-type="numbered"><label>3</label><mml:math id="M51" display="block"><mml:mstyle class="stylechange" displaystyle="true"/><mml:mrow><mml:mstyle class="stylechange" displaystyle="true"/><mml:msub><mml:mi>U</mml:mi><mml:mtext>eff</mml:mtext></mml:msub><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.47</mml:mn><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msub><mml:mi>U</mml:mi><mml:mtext>PBL</mml:mtext></mml:msub><mml:mo>+</mml:mo><mml:mn mathvariant="normal">0.31</mml:mn><mml:mspace linebreak="nobreak" width="1em"/><mml:mfenced open="[" close="]"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfenced><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula></p>
      <p id="d2e1421"><xref ref-type="bibr" rid="bib1.bibx54" id="text.62"/> calculate <inline-formula><mml:math id="M52" display="inline"><mml:mi>Q</mml:mi></mml:math></inline-formula> using ERA5 10 <inline-formula><mml:math id="M53" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> winds <xref ref-type="bibr" rid="bib1.bibx21" id="paren.63"/> as well as GEOS FP 10 <inline-formula><mml:math id="M54" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> and planetary boundary layer winds <xref ref-type="bibr" rid="bib1.bibx41" id="paren.64"/>. They additionally estimate uncertainties using a grid-based ensemble approach, varying input parameters such as plume masking thresholds, background methane concentration, wind magnitudes, and effective wind speed coefficient values. IME estimates (and associated uncertainties) for synthetic and real methane observations in this work follow the methods of <xref ref-type="bibr" rid="bib1.bibx54" id="text.65"/> directly, using only the ERA5 10 <inline-formula><mml:math id="M55" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind product. We apply the IME method and <inline-formula><mml:math id="M56" display="inline"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mtext>eff</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> parameterization from <xref ref-type="bibr" rid="bib1.bibx54" id="text.66"/> without modification, allowing us to benchmark the performance of ML-SPERE against a published and operationally used implementation of the IME method that was specifically calibrated for TROPOMI methane observations.</p>
</sec>
</sec>
<sec id="Ch1.S3">
  <label>3</label><title>Results</title>
      <p id="d2e1490">After optimizing and training our model, we evaluate its performance on a test set of synthetic methane plume observations (Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>). Because the true emission rates are known, we can assess ML-SPERE under a range of scenarios and examine its performance. Next, we apply our model to a well-studied blowout, enabling evaluation of its estimates against both inverse modeling approaches and the IME method using data from TROPOMI and higher-resolution satellites (Sect. <xref ref-type="sec" rid="Ch1.S3.SS2"/>). Finally, we use ML-SPERE to quantify the emission rates of an entire year's worth of TROPOMI super-emitter detections (Sect. <xref ref-type="sec" rid="Ch1.S3.SS3"/>).</p>
<sec id="Ch1.S3.SS1">
  <label>3.1</label><title>Test set evaluation and comparison to the IME method</title>
      <p id="d2e1506">We evaluate ML-SPERE performance using the global test set of synthetic TROPOMI methane plume observations generated with HYSPLIT (Sect. <xref ref-type="sec" rid="Ch1.S2.SS2"/>). On this test set, ML-SPERE achieves a MAPE of 40.2 %, compared to 46.1 % for the IME method, as shown in Fig. <xref ref-type="fig" rid="F3"/>. Notably, the median absolute percentage error (MdAPE) for IME estimates is 40.5 %, while ML-SPERE yields a substantially lower MdAPE of 29.5 %. This difference arises because ML-SPERE estimates for this test set exhibit a heavy-tailed error distribution (arising from scenes where the plume is poorly observed, see following paragraph), and metrics such as MAPE are strongly influenced by outliers. Absolute errors produced by ML-SPERE generally scale with the true emission rate of test set plumes, but this relationship breaks down at the lowest emission rates, where an approximately constant absolute error leads to inflated relative errors. Similar behavior is observed for IME estimates. Additional performance metrics such as the Pearson correlation coefficient (<inline-formula><mml:math id="M57" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>), mean absolute error (MAE), median absolute error (MdAE), mean bias, and root mean squared error (RMSE) are summarized in Table <xref ref-type="table" rid="T1"/>. We find no significant change in bias for either method when applied only to plumes whose masks are truncated by the scene boundary. Across all metrics, ML-SPERE outperforms the IME method. In terms of computational efficiency, the IME method implementation from <xref ref-type="bibr" rid="bib1.bibx54" id="text.67"/> is faster than ML-SPERE. On a standard desktop machine, it takes approximately 15 <inline-formula><mml:math id="M58" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">s</mml:mi></mml:mrow></mml:math></inline-formula> to generate either an ensemble estimate of 43 923 members with the IME method or an ensemble estimate of 5000 members with ML-SPERE. These timings may vary depending on hardware. The time it takes ML-SPERE to produce an ensemble estimate scales linearly with the number of ensemble members.</p>

      <fig id="F3"><label>Figure 3</label><caption><p id="d2e1536">Estimated emission rates for the test set of HYSPLIT plumes via <bold>(a)</bold> the IME method and <bold>(b)</bold> ML-SPERE as a function of true plume emission rates. The colormaps represent the relative density of the visualised data. We visualise bootstrapped ordinary least squares (OLS) regression lines for emission rate estimates from both the IME method <bold>(c)</bold> and ML-SPERE <bold>(d)</bold>. The <inline-formula><mml:math id="M59" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> line is subtracted in <bold>(c)</bold> and <bold>(d)</bold> to show average bias for each method across the entire range of emission rates of test set plumes. The regression dilution apparent in panel <bold>(d)</bold> for ML-SPERE is partially alleviated when application is limited to well observed plumes, see Fig. <xref ref-type="fig" rid="FA4"/>d.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f03.png"/>

        </fig>

<table-wrap id="T1" specific-use="star"><label>Table 1</label><caption><p id="d2e1584">Comparison of performance metrics for the IME method and ML-SPERE emission rate estimation when evaluated on the test set (Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>). Best performance for each metric is shown in bold. Metrics shown are Mean Absolute Percentage Error (MAPE), Median Absolute Percentage Error (MdAPE), Pearson correlation coefficient (<inline-formula><mml:math id="M60" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>), mean absolute error (MAE), median absolute error (MdAE), mean bias, and root mean squared error (RMSE). In the first two rows, we compare results for the two methods across the entire test set (438 plumes), and in the second two rows, we compare results for the two methods for well-observed scenes only (264 plumes), testing only against plumes where the emission source location is contained within the estimated plume mask. We also include the performance of both methods only for plumes which are poorly observed in the last two rows (174 plumes).</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="8">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:colspec colnum="7" colname="col7" align="right"/>
     <oasis:colspec colnum="8" colname="col8" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Method</oasis:entry>
         <oasis:entry colname="col2">MAPE [<inline-formula><mml:math id="M61" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">%</mml:mi></mml:mrow></mml:math></inline-formula>]</oasis:entry>
         <oasis:entry colname="col3">MdAPE [<inline-formula><mml:math id="M62" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">%</mml:mi></mml:mrow></mml:math></inline-formula>]</oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M63" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5">MAE [<inline-formula><mml:math id="M64" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>]</oasis:entry>
         <oasis:entry colname="col6">MdAE [<inline-formula><mml:math id="M65" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>]</oasis:entry>
         <oasis:entry colname="col7">Bias [<inline-formula><mml:math id="M66" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>]</oasis:entry>
         <oasis:entry colname="col8">RMSE [<inline-formula><mml:math id="M67" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>]</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">All plumes </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">IME</oasis:entry>
         <oasis:entry colname="col2">46.1</oasis:entry>
         <oasis:entry colname="col3">40.5</oasis:entry>
         <oasis:entry colname="col4">0.65</oasis:entry>
         <oasis:entry colname="col5">24.4</oasis:entry>
         <oasis:entry colname="col6">15.0</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M68" display="inline"><mml:mo>-</mml:mo></mml:math></inline-formula>6.4</oasis:entry>
         <oasis:entry colname="col8">38.5</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">ML-SPERE</oasis:entry>
         <oasis:entry colname="col2"><bold>40.2</bold></oasis:entry>
         <oasis:entry colname="col3"><bold>29.5</bold></oasis:entry>
         <oasis:entry colname="col4"><bold>0.76</bold></oasis:entry>
         <oasis:entry colname="col5"><bold>19.1</bold></oasis:entry>
         <oasis:entry colname="col6"><bold>10.7</bold></oasis:entry>
         <oasis:entry colname="col7"><bold>2.4</bold></oasis:entry>
         <oasis:entry colname="col8"><bold>30.3</bold></oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Well-observed plumes only </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">IME</oasis:entry>
         <oasis:entry colname="col2">44.6</oasis:entry>
         <oasis:entry colname="col3">42.4</oasis:entry>
         <oasis:entry colname="col4">0.67</oasis:entry>
         <oasis:entry colname="col5">25.1</oasis:entry>
         <oasis:entry colname="col6">16.3</oasis:entry>
         <oasis:entry colname="col7"><inline-formula><mml:math id="M69" display="inline"><mml:mo>-</mml:mo></mml:math></inline-formula>11.8</oasis:entry>
         <oasis:entry colname="col8">37.6</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">ML-SPERE</oasis:entry>
         <oasis:entry colname="col2"><bold>33.3</bold></oasis:entry>
         <oasis:entry colname="col3"><bold>24.3</bold></oasis:entry>
         <oasis:entry colname="col4"><bold>0.79</bold></oasis:entry>
         <oasis:entry colname="col5"><bold>17.6</bold></oasis:entry>
         <oasis:entry colname="col6"><bold>10.0</bold></oasis:entry>
         <oasis:entry colname="col7"><bold>0.6</bold></oasis:entry>
         <oasis:entry colname="col8"><bold>27.9</bold></oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry namest="col1" nameend="col8">Poorly-observed plumes only </oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">IME</oasis:entry>
         <oasis:entry colname="col2"><bold>48.4</bold></oasis:entry>
         <oasis:entry colname="col3">38.4</oasis:entry>
         <oasis:entry colname="col4">0.65</oasis:entry>
         <oasis:entry colname="col5">23.4</oasis:entry>
         <oasis:entry colname="col6">13.1</oasis:entry>
         <oasis:entry colname="col7"><bold>1.8</bold></oasis:entry>
         <oasis:entry colname="col8">40.0</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">ML-SPERE</oasis:entry>
         <oasis:entry colname="col2">50.6</oasis:entry>
         <oasis:entry colname="col3"><bold>34.2</bold></oasis:entry>
         <oasis:entry colname="col4"><bold>0.70</bold></oasis:entry>
         <oasis:entry colname="col5"><bold>21.5</bold></oasis:entry>
         <oasis:entry colname="col6"><bold>12.8</bold></oasis:entry>
         <oasis:entry colname="col7">5.0</oasis:entry>
         <oasis:entry colname="col8"><bold>33.6</bold></oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

      <p id="d2e1962">The construction of plume masks is complicated by missing data and high methane background variability (various examples are shown in Fig. <xref ref-type="fig" rid="FA3"/>). We thus define a TROPOMI methane plume in the test set to be “well-observed” if the known emission source location is contained within the estimated plume mask. When filtering the test set to include only scenes where the plume is well-observed (resulting in 264 scenes), the performance gap between ML-SPERE and the IME method widens further (Table <xref ref-type="table" rid="T1"/>). MAPE and MdAPE of ML-SPERE decrease to 33.3 % and 24.3 %, respectively, whereas the corresponding IME errors remain largely unchanged at 44.6 % and 42.4 % (Fig. <xref ref-type="fig" rid="FA4"/>). Additional figures and analyses in Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS3"/> show that the improved performance observed when restricting the application of ML-SPERE to such plumes arises from the model's strong reliance on spatial gradients in methane abundance near the plume head or source region when inferring emission rates. This behavior lends ML-SPERE meaningful physical interpretability. Performance may degrade when the plume mask is compromised around the plume head, either by missing data or by methane enhancements that are weak relative to background variability. However, such conditions are generally visually identifiable, allowing for informed application of the model. Even when the source location is unknown, an analyst can inspect a scene to assess whether the estimated plume mask is compromised and, consequently, whether ML-SPERE is likely to provide a reliable emission estimate. For poorly-observed plumes in the test set (i.e. those that do not meet the criteria for being well-observed), ML-SPERE and the IME method exhibit broadly comparable performance. ML-SPERE performance degrades relative to its results on well-observed plumes, whereas IME performance remains similar across all subsets.</p>
      <p id="d2e1973">While ML-SPERE achieves a MAPE of 40.2 % on the full test set and 33.3 % on well-observed plumes, its performance on the validation set is substantially better, with a MAPE of 14.6 %. Significant differences between the validation (and training) set and the test set include the fact that the validation set contains no spatially varying backgrounds or missing data, and that the validation set was simulated using WRF-Chem with NCEP wind data, which matches the modeling framework used for training. The test set plumes are additionally simulated at a wide variety of locations around the globe which are distinct from the locations used for model training. We have already demonstrated that spatially varying backgrounds and missing data hamper the performance of ML-SPERE by complicating the estimation of accurate plume masks. In Sects. <xref ref-type="sec" rid="App1.Ch1.S1.SS5"/> and <xref ref-type="sec" rid="App1.Ch1.S1.SS6"/>, we demonstrate that further performance degradation on the test set (with respect to validation set performance) arises primarily from the use of HYSPLIT to generate test plumes, which exhibit morphologies distinct from those produced by WRF-Chem and encountered during training, rather than from limitations in geographic generalization.</p>

      <fig id="F4" specific-use="star"><label>Figure 4</label><caption><p id="d2e1982">Estimate bias (expressed as the residual between an emission rate estimate for a method and the true emission rate) for HYSPLIT test set plumes, estimated via the IME method <bold>(a)</bold> and ML-SPERE <bold>(b)</bold> as a function of 10 <inline-formula><mml:math id="M70" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speeds. Also shown are ordinary least square regressions for both datasets. Due to the linear calibration of the effective wind speed to the input 10 <inline-formula><mml:math id="M71" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speed, the IME estimate residuals show a strong negative bias at low wind speeds, however ML-SPERE yields unbiased estimates at wind speeds below 2 <inline-formula><mml:math id="M72" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f04.png"/>

        </fig>

      <p id="d2e2030">Figure <xref ref-type="fig" rid="F3"/> illustrates a systematic bias where ML-SPERE tends to underestimate emission rates for plumes with large emission rates above the mean of the training dataset, and overestimate emission rates for plumes with small emission rates below the mean. The IME method also tends to underestimate emissions for plumes with large emission rates, despite the fact that the IME method is not a mean-biased method. The tendency for ML-based predictions to regress toward the mean of the training dataset labels is a common issue reported in similar studies <xref ref-type="bibr" rid="bib1.bibx28 bib1.bibx29 bib1.bibx2" id="paren.68"/>. This regression dilution for ML-SPERE is reduced by <inline-formula><mml:math id="M73" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>/</mml:mo><mml:mn mathvariant="normal">3</mml:mn></mml:mrow></mml:math></inline-formula> when we restrict application to well-observed plumes (Fig. <xref ref-type="fig" rid="FA4"/>d), though no such improvement in estimate biases is found for the IME method (Fig. <xref ref-type="fig" rid="FA4"/>c). For scenes with wind speeds below 2 <inline-formula><mml:math id="M74" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, residuals between ML-SPERE emission rate estimates and true plume emission rates show no correlation with scene wind speed. In contrast, residuals of IME estimates in this wind speed regime exhibit a strong positive correlation, with the IME method significantly underestimating emissions for these plumes. Although low wind speeds aid in plume detection (48 % of the detections reported by <xref ref-type="bibr" rid="bib1.bibx54" id="text.69"/> are found at wind speeds below 2 <inline-formula><mml:math id="M75" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>), they complicate emission rate quantifications for mass-balance based approaches such as the IME method and the cross-sectional flux (CSF) method <xref ref-type="bibr" rid="bib1.bibx30 bib1.bibx31" id="paren.70"/>. Under such conditions, plumes often exhibit blob-like structures or irregularly rotated morphologies <xref ref-type="bibr" rid="bib1.bibx46" id="paren.71"/>, which complicate emission rate estimation. Equation (<xref ref-type="disp-formula" rid="Ch1.E1"/>) in conjunction with the 10 <inline-formula><mml:math id="M76" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> effective wind speed calibration from Eq. (<xref ref-type="disp-formula" rid="Ch1.E2"/>) implies that as <inline-formula><mml:math id="M77" display="inline"><mml:mrow><mml:msub><mml:mi>U</mml:mi><mml:mn mathvariant="normal">10</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> approaches zero, the IME-estimated emission rate <inline-formula><mml:math id="M78" display="inline"><mml:mi>Q</mml:mi></mml:math></inline-formula> must also approach zero. As a result, the IME method may be biased low and underestimate emissions from methane plumes detected by TROPOMI in low-wind-speed environments, for which the uncertainty on the wind speed is also relatively large. Figure <xref ref-type="fig" rid="F4"/> shows this trend clearly, and displays that the IME method systematically underestimates the emission rates for test set plumes with average wind speeds of less than 2 <inline-formula><mml:math id="M79" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. By contrast, ML-SPERE is an unbiased estimator for plumes at low wind speeds, despite utilizing the same input data as the IME method.</p>
      <p id="d2e2148">Ideally, a model that accurately estimates uncertainty would capture about 95 % of true emission values within the 2.5th–97.5th percentile range of its posterior distribution for each emission-rate estimate. Across the full test set, 61 % of true values fall within this interval for IME estimates, compared to 81 % for ML-SPERE. These results indicate that both methods underestimate predictive uncertainty, although ML-SPERE provides more reliable uncertainty estimates than the IME method. We retain the hyperparameters specified in the error ensemble described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/> (i.e. the assumed uncertainties in wind speeds, methane backgrounds, and plume masking strength) which are drawn from literature or informed by expert judgment. While increasing these values could widen the estimated uncertainty ranges (and thereby increase the proportion of true emission rates captured within them), doing so would risk masking model misspecification by attributing errors to input uncertainty alone. In Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS4"/>  and Fig. <xref ref-type="fig" rid="FA6"/>, we apply the probability integral transform to show that the ML-SPERE error ensemble generalizes better across simulated plume datasets than the IME approach.</p>
</sec>
<sec id="Ch1.S3.SS2">
  <label>3.2</label><title>Case study comparison: Kazakhstan's Karaturun East oil field 2023 blowout</title>
      <p id="d2e2165">In 2023, a well blowout took place in Kazakhstan's Karaturun East oilfield which lasted for more than 200 <inline-formula><mml:math id="M80" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:math></inline-formula> (days) <xref ref-type="bibr" rid="bib1.bibx17" id="paren.72"/>. This event was observed with multiple methane-sensing satellites, including TROPOMI and high-resolution (25–60 <inline-formula><mml:math id="M81" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula>) satellite instruments such as GHGSat <xref ref-type="bibr" rid="bib1.bibx63 bib1.bibx27" id="paren.73"/>, EMIT <xref ref-type="bibr" rid="bib1.bibx59" id="paren.74"/>, EnMAP <xref ref-type="bibr" rid="bib1.bibx53" id="paren.75"/>, and PRISMA <xref ref-type="bibr" rid="bib1.bibx16" id="paren.76"/>. For TROPOMI methane observations, inverse analysis estimates of emission rates were produced on a daily basis when observing conditions allowed. Separate IME estimates were also made using the observations of the point-source imagers when available. These quantifications showed that the daily estimated emission rate varied between 20–50 <inline-formula><mml:math id="M82" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> over the duration of the blowout, emitting an estimated total of <inline-formula><mml:math id="M83" display="inline"><mml:mrow><mml:mn mathvariant="normal">131</mml:mn><mml:mo>±</mml:mo><mml:mn mathvariant="normal">34</mml:mn></mml:mrow></mml:math></inline-formula> <inline-formula><mml:math id="M84" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">kt</mml:mi></mml:mrow></mml:math></inline-formula> of methane <xref ref-type="bibr" rid="bib1.bibx17" id="paren.77"/>. For dates during this blowout where a methane plume is detected in the TROPOMI data following the procedure of <xref ref-type="bibr" rid="bib1.bibx55" id="text.78"/>, we compare ML-SPERE emission rate estimates with those from <xref ref-type="bibr" rid="bib1.bibx17" id="text.79"/> as well as to IME emission rates estimates that we have produced using the TROPOMI data.</p>

      <fig id="F5" specific-use="star"><label>Figure 5</label><caption><p id="d2e2249"><bold>(a)</bold> Time series of daily estimated methane emission rates emission rates for the Karaturun East oilfield blowout. TROPOMI inverse modeling estimates and high-resolution satellite IME estimates are taken from <xref ref-type="bibr" rid="bib1.bibx17" id="text.80"/>. ML-SPERE data points are median values of the estimated posterior distribution of estimated emissions. Errorbars shown are <inline-formula><mml:math id="M85" display="inline"><mml:mrow><mml:mo>±</mml:mo><mml:mn mathvariant="normal">1</mml:mn><mml:mi mathvariant="italic">σ</mml:mi></mml:mrow></mml:math></inline-formula>, with the exception of error bars for ML-SPERE estimates, which extend from the 16th to the 84th percentile of the posterior distribution of emission estimates (all statistics are given in Table <xref ref-type="table" rid="TA3"/>). We choose to represent error bars for ML-SPERE with percentiles to show the skewed shape of the posterior distribution of the estimate (e.g. see Fig. <xref ref-type="fig" rid="FA6"/>a). <bold>(b)</bold> ML-SPERE estimates of the daily emission rate compared to the TROPOMI inverse modeling estimate for that same day. <bold>(c)</bold> IME estimates of the daily emission rate compared to the TROPOMI inverse modeling estimate for that same day.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f05.png"/>

        </fig>

      <p id="d2e2286">Figure <xref ref-type="fig" rid="F5"/>a shows a time series of the daily emission rate estimates from <xref ref-type="bibr" rid="bib1.bibx17" id="text.81"/> and the ML-SPERE estimates. In general, ML-SPERE estimates tend to be higher than the TROPOMI inverse modeling results presented in <xref ref-type="bibr" rid="bib1.bibx17" id="text.82"/>, although they typically agree within uncertainty and show slightly closer agreement than that between the TROPOMI inverse modeling and IME estimates. For three of the four dates in the time series where we have an IME emission rate estimate from a high-spatial resolution satellite, we find greater agreement with ML-SPERE estimates than we do with TROPOMI IME estimates. We also find general agreement between ML-SPERE estimates and TROPOMI IME estimates. Error bars in Fig. <xref ref-type="fig" rid="F5"/> for ML-SPERE estimates span from the 16th to the 84th percentile of the posterior distribution of emission rates, in order to show the skew of posterior estimates (e.g. see Fig. <xref ref-type="fig" rid="FA6"/>a). These error bars are typically skewed with a heavy tail toward higher emission values, since ML-SPERE will not produce estimates below zero but has no upper bound. For most days in the time series, the uncertainty in ML-SPERE predictions is dominated by uncertainty in the input wind data. The next largest contribution to the overall uncertainty arises from uncertainty in estimating the methane background level within each scene. In contrast, uncertainties associated with plume masking and the model itself contribute comparatively little to the total uncertainty (Table <xref ref-type="table" rid="TA3"/>).</p>
      <p id="d2e2305">In Fig. <xref ref-type="fig" rid="F5"/>b and c, we compare ML-SPERE estimates and TROPOMI IME estimates with the corresponding TROPOMI inverse modeling estimates. The ML-SPERE estimates exhibit a reduction in bias relative to the TROPOMI IME estimates when compared against the inverse modeling results, along with an improved <inline-formula><mml:math id="M86" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula> correlation and reduced chi-squared <inline-formula><mml:math id="M87" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">χ</mml:mi><mml:mi mathvariant="italic">ν</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup></mml:mrow></mml:math></inline-formula> (though the scatter between inverse analysis estimates and both ML-SPERE and TROPOMI IME estimates remains large). To investigate potential sources of bias between ML-SPERE and the inverse modeling estimates for these scenes, we apply ML-SPERE directly to the resampled simulated plumes from <xref ref-type="bibr" rid="bib1.bibx17" id="text.83"/> that were identified as the best matches to the observations. Using these simulated plumes (without missing data or background methane) and their corresponding meteorology as inputs, ML-SPERE shows improved agreement with the simulated emission rates (i.e. the results shown in <xref ref-type="bibr" rid="bib1.bibx17" id="altparen.84"/>), achieving a Pearson correlation of <inline-formula><mml:math id="M88" display="inline"><mml:mrow><mml:mi>R</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.91</mml:mn></mml:mrow></mml:math></inline-formula> and a substantially reduced mean bias of 15.4 <inline-formula><mml:math id="M89" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. In contrast, IME estimates applied to the same simulated plumes remain largely unchanged relative to the metrics shown in Fig. <xref ref-type="fig" rid="F5"/>c. The remaining positive bias of 15.4 <inline-formula><mml:math id="M90" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> in the ML-SPERE estimates could be attributable to plume morphologies resulting from complicated wind fields or conditions which violate the steady-state emission assumptions represented in the training dataset. This interpretation is supported by the fact that the IME method (which assumes steady-state emission scenarios) produces estimates that are twice as large as those of the ML-based estimates when applied to the same simulated plumes from <xref ref-type="bibr" rid="bib1.bibx17" id="text.85"/>. Furthermore, ML-SPERE does not show similar positive biases when applied to other simulated plume datasets with steady-state emissions (e.g. Figs. <xref ref-type="fig" rid="F3"/>, <xref ref-type="fig" rid="FA4"/>, <xref ref-type="fig" rid="FA7"/>, and <xref ref-type="fig" rid="FA8"/>). Applying ML-SPERE to the same simulated plumes while masking the abundances to reproduce the spatial pattern of missing data in the TROPOMI observations does not degrade performance. This robustness arises because the well blowout source remains clearly visible, and the absence of a spatially variable methane background allows the automatically generated plume masks in the ML-SPERE ensemble to reliably capture the source region. These results suggest that the remaining scatter between ML-SPERE and inverse modeling estimates for the real TROPOMI observations is likely driven by a combination of highly variable coastal methane backgrounds that complicate plume masking on certain dates, discrepancies between meteorological reanalysis products and true wind vectors, and conservative inverse modeling estimates where the forward modeled plume does not closely reproduce the TROPOMI observation. Although the actual underlying emission rates are not known for this super-emitting blowout, the improved agreement between ML-SPERE and inverse modeling estimates (compared to IME method vs. inverse modeling agreement) is notable as inverse modeling is a far more information-rich method, incorporating three-dimensional wind fields and atmospheric transport processes to produce plume emission rate estimates (albeit at a substantially higher computational cost than other methods).</p>
</sec>
<sec id="Ch1.S3.SS3">
  <label>3.3</label><title>Population study: cataloged 2021 TROPOMI methane plume detections</title>
      <p id="d2e2405"><xref ref-type="bibr" rid="bib1.bibx54" id="text.86"/> present an automated, ML–based algorithm for detecting large methane emission plumes in TROPOMI satellite data, with a corresponding catalog of detections for all of 2021 published in <xref ref-type="bibr" rid="bib1.bibx55" id="text.87"/>. This dataset comprises 2974 plume detections distributed globally, with all but 30 linked to dominant anthropogenic sources such as urban areas/landfills, oil, gas, or coal mining operations based on bottom-up inventories. We use this 2021 detection catalog as a benchmark dataset to compare the performance of ML-SPERE to the traditional IME approach. ML-SPERE and IME estimates exhibit a strong correlation (<inline-formula><mml:math id="M91" display="inline"><mml:mrow><mml:mi>R</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.79</mml:mn></mml:mrow></mml:math></inline-formula>, Fig. <xref ref-type="fig" rid="F6"/>a) and agree within errors (<inline-formula><mml:math id="M92" display="inline"><mml:mrow><mml:msubsup><mml:mi mathvariant="italic">χ</mml:mi><mml:mi mathvariant="italic">ν</mml:mi><mml:mn mathvariant="normal">2</mml:mn></mml:msubsup><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.94</mml:mn></mml:mrow></mml:math></inline-formula>, Fig. <xref ref-type="fig" rid="F6"/>a). ML-SPERE estimates tend to exceed those of the IME method at low IME-estimated emission rates (i.e. below 10 <inline-formula><mml:math id="M93" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>). This behavior may reflect the regression dilution discussed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and shown in Fig. <xref ref-type="fig" rid="F3"/>c. Additionally, it may indicate underestimation by the IME method for these plumes due to the wind-speed bias described in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>, as discussed further below. Emission estimates for TROPOMI plumes below 10 <inline-formula><mml:math id="M94" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> should be interpreted with caution, as this lies near the conventional detection threshold.</p>

      <fig id="F6" specific-use="star"><label>Figure 6</label><caption><p id="d2e2489"><bold>(a)</bold> TROPOMI 2021 methane plume detections quantified via the IME method and ML-SPERE. We quantify (and compare with the IME method) all but 44 of the 2974 scenes in <xref ref-type="bibr" rid="bib1.bibx55" id="text.88"/> with ML-SPERE. These 44 scenes had ML-SPERE ensemble members with empty plume masks, due to combinations of low plume enhancements and high masking thresholds. The colormap represents the relative density of the visualised data. <bold>(b)</bold> Residuals between IME emission rate estimates and ML-SPERE emission rate estimates as a function of scene 10 <inline-formula><mml:math id="M95" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> wind speed, for detected TROPOMI plumes for 2021 with scene wind speeds below 2 <inline-formula><mml:math id="M96" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. This trend in residuals shown here could be explained by unbiased estimates from ML-SPERE as a function of scene wind speed, and biased estimates from the IME method that grow increasingly negative with decreasing scene wind speed. <bold>(c)</bold> Plume emission rates averaged across estimated sectoral classification. Total number of plumes per estimated sector is printed above each grouping, and mean estimated emissions are printed within each bar. Errorbars are standard errors on the mean.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f06.png"/>

        </fig>

      <p id="d2e2534">As shown in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and Fig. <xref ref-type="fig" rid="F4"/>, ML-SPERE produced unbiased estimates with no correlation with wind speed for low-wind plumes in the synthetic test set, while the IME method showed increasing negative bias with decreasing wind speed. The residuals between the IME estimates and those of ML-SPERE for the real 2021 plume dataset exhibit a similar trend at lower wind speeds (Fig. <xref ref-type="fig" rid="F6"/>b), which is consistent with the findings from Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>. As nearly half of plumes are detected at wind speeds below 2 <inline-formula><mml:math id="M97" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, ML-SPERE may therefore provide less biased emission rate estimates than the IME method for these commonly detected plumes.</p>
      <p id="d2e2563">A global comparison of ML-SPERE and IME emission rate estimates for the full 2021 TROPOMI detection catalog reveals that although the two methods generally agree (Fig. <xref ref-type="fig" rid="F6"/>a), there are few regions where systematic differences emerge. These differences are statistically significant only in northern Russia and southwestern Australia, where ML-SPERE yields higher estimates than the IME method, and in the Congo Basin, where ML-SPERE yields lower estimates. A full analysis is given in Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS8"/> (with Figs. <xref ref-type="fig" rid="FA9"/> and <xref ref-type="fig" rid="FA10"/>) and summarized below.</p>
      <p id="d2e2574">In northern Russia, the largest discrepancies correspond to extremely compact, high-enhancement signals that coincide spatially with major gas pipeline infrastructure. These features resemble short-duration, transient emissions rather than the steady-state plumes on which ML-SPERE was trained, and their out-of-sample elevated methane enhancements likely contribute to the elevated emission estimates. Such transient, non steady-state emissions may also not be well estimated by the IME method (especially if the plume has detached from the source and is being transported downwind, as this violates the steady-state assumptions underlying the method). For these plumes, the total plume mass may provide a more relevant characterization than estimated instantaneous emission rate as emissions are no longer ongoing <xref ref-type="bibr" rid="bib1.bibx5" id="paren.89"/>. For ongoing or persistent events, the observed mass only provides a lower boundary for the total emitted mass, as mass is still being emitted at the time of observation. In contrast to pipeline emissions in northern Russia, methane enhancements in the Congo Basin have very large spatial extents and may be more representative of area-source emissions than (TROPOMI-scale) point-source emissions on which ML-SPERE was trained and the IME method was calibrated. Finally, for plumes in southeastern Australia, ML-SPERE estimates are systematically higher than IME estimates in a low–wind-speed regime, where we have already demonstrated that IME estimates exhibit a negative bias. Although transient emission events and spatially extended plumes can occur globally, analysis of this dataset suggests that temporal averaging limits statistically significant differences between total emissions estimated by the two methods to a small number of regions. These regional diagnostics highlight that ML-SPERE likely performs most accurately when the observed plume exhibits a coherent head and well-defined downwind structure, and that applying either method can benefit from human oversight in cases involving transient, discontinuous emissions or spatially fragmented plumes.</p>
      <p id="d2e2580">In Fig. <xref ref-type="fig" rid="F6"/>c, we compare average estimated emissions for the 2021 TROPOMI plume detections grouped by estimated source sector <xref ref-type="bibr" rid="bib1.bibx55" id="paren.90"/>. Sectoral emissions estimated for this dataset using the IME method and ML-SPERE largely agree, with the exception of those attributed to wetland emissions, which are underrepresented in this dataset. These plumes are spatially concentrated in the Congo Basin and, as discussed in the previous paragraph, have large spatial extents that may lead to overestimation by the IME method. The agreement across the remaining sectors likely reflects averaging over a broad distribution of wind speeds, which can mitigate wind-speed–dependent biases in the IME method. Taken together, the consistency between ML-SPERE and IME across most sectors supports the robustness of the sectoral emission estimates for this plume dataset.</p>
</sec>
</sec>
<sec id="Ch1.S4" sec-type="conclusions">
  <label>4</label><title>Conclusions</title>
      <p id="d2e2597">This study presents ML-SPERE, an ML-based method for quantifying emission rates of methane plumes detected in TROPOMI observations. We trained ML-SPERE using synthetic TROPOMI methane plumes, which were simulated using WRF-Chem and then resampled to TROPOMI pixel footprints. Wind data was also included in the training dataset. We additionally used HYSPLIT to simulate a test set of methane plumes at locations not included in the training data. This test set is global and includes real TROPOMI methane scenes as backgrounds, which introduces realistic challenges in reducing methane scenes to plume abundances. We evaluate ML-SPERE against this synthetic test set, and additionally use ML-SPERE to quantify the emission rates of real TROPOMI methane plumes observed during a well-studied blowout in Kazakhstan, as well as for an entire year's worth of automated super-emitter detections from 2021. At all stages of model evaluation, we benchmark the performance of ML-SPERE against that of the IME method.</p>
      <p id="d2e2600">When evaluating emission rate estimates produced by ML-SPERE for the full synthetic test set and comparing them to those from the IME method, we find that MAPE improves from 46.1 % to 40.2 %, MdAPE improves from 40.5 % to 29.5 %, and <inline-formula><mml:math id="M98" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula> improves from 0.65 to 0.76. ML-SPERE performance exceeds that of the IME method for every metric measured on the test set. ML-SPERE is found to have a regression bias across emission rates as is usually the case with ML regressors. When filtering test set scenes for plumes where we have spatially complete information around the source location and plume head, the performance gap between ML-SPERE and the IME method increases further, and metrics such as MAPE, MdAPE, <inline-formula><mml:math id="M99" display="inline"><mml:mi>R</mml:mi></mml:math></inline-formula>, and biases are all significantly improved for our model. We show empirically that this is because ML-SPERE has learned to use methane gradients around the plume head to inform emission rate estimates, which lends ML-SPERE clear physical interpretability. We also find that estimates from ML-SPERE are unbiased with respect to scene wind speed, but that IME estimates are not. We demonstrate that IME estimates can be systematically underestimated for methane plumes detected at wind speeds below 2 <inline-formula><mml:math id="M100" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, which are conditions where real TROPOMI methane plumes are commonly detected. This ability of ML-SPERE to provide unbiased estimates in this wind regime which is favorable for plume detection represents a significant improvement for emission rate quantification for TROPOMI methane plumes.</p>
      <p id="d2e2634">When using ML-SPERE to quantify emissions for the 2023 Karaturun East blowout in Kazakhstan, we find that ML-SPERE estimates typically agree within uncertainties with IME and inverse analysis estimates from TROPOMI observations. Compared with TROPOMI IME estimates, ML-SPERE estimates show improved agreement with the high-resolution IME estimates. ML-SPERE estimates show a slightly reduced bias relative to TROPOMI IME estimates when compared against inverse modeling though scatter remains. ML-SPERE was further evaluated using a global catalog of TROPOMI methane plume detections from 2021, providing a comprehensive real-world benchmark against the traditional IME method. Across this dataset, ML-SPERE and IME estimates show strong general agreement (<inline-formula><mml:math id="M101" display="inline"><mml:mrow><mml:mi>R</mml:mi><mml:mo>=</mml:mo><mml:mn mathvariant="normal">0.79</mml:mn></mml:mrow></mml:math></inline-formula>). Critically, the 2021 detections reproduce the wind-speed sensitivity found in our simulated test set: at low wind speeds (less than 2 <inline-formula><mml:math id="M102" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>), IME vs. ML-SPERE residuals become increasingly negatively biased, indicating that the implemented IME method may have a wind-speed dependent bias and that ML-SPERE does not show such bias. A global spatial comparison further reveals that, although agreement is generally strong, systematic regional differences arise in only a few locations and reflect scene-specific factors in addition to potential methodological limitations. In particular, scenarios in which neither the IME method nor ML-SPERE is well suited for application (e.g. violations of point-source or steady-state assumptions) likely require more careful consideration using situation-tailored approaches. Future work could examine re-training ML-SPERE using a larger dataset, with expanded geographic and seasonal diversity that focuses on regions and circumstances where we have shown that ML-SPERE and the IME method produce larger differences in emission rate estimates, though the results of Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS5"/> show that ML-SPERE generalizes beyond the spatial domain of training data. Average sectoral emissions for this dataset as estimated via the IME method or ML-SPERE are broadly unchanged, with the exception of wetland emissions, which are underrepresented in this dataset both spatially and numerically. Together, these diagnostics show that ML-SPERE performs robustly across diverse real-world conditions and particularly well in common low-wind regimes. When combined with human oversight, it offers a reliable and complementary alternative to the IME method for routine global quantification of methane emissions from TROPOMI plumes.</p>
</sec>

      
      </body>
    <back><app-group>

<app id="App1.Ch1.S1">
  <label>Appendix A</label><title>Additional material</title>
<sec id="App1.Ch1.S1.SS1">
  <label>A1</label><title>Simulated plume locations</title>
      <p id="d2e2687">In Fig. <xref ref-type="fig" rid="FA1"/> we show the locations of the WRF-Chem domain setups and tracer release location schemes. In Fig. <xref ref-type="fig" rid="FA2"/> we show the locations of the plumes simulated for the test set using HYSPLIT.</p>

      <fig id="FA1" specific-use="star"><label>Figure A1</label><caption><p id="d2e2696"><bold>(a)</bold> The inner domain of the nested WRF-Chem setups, shown with the yellow rectangles. For each of these inner domains, 8 passive tracers are initialized to emit over the duration of the simulations, at 4 different locations and 2 different release heights. Tracer locations are shown with the stars in the example domain shown in <bold>(b)</bold>, and at each star, tracers are initialized at approximately 25 and 250 <inline-formula><mml:math id="M103" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi></mml:mrow></mml:math></inline-formula> heights. Background imagery in <bold>(b)</bold> relies on non-concurrent Sentinel-2 data (2022) adapted from Google Earth Engine <xref ref-type="bibr" rid="bib1.bibx14 bib1.bibx10" id="paren.91"/>.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f07.png"/>

        </fig>

      <fig id="FA2" specific-use="star"><label>Figure A2</label><caption><p id="d2e2726">438 test set plumes simulated at 224 unique global locations with HYSPLIT.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f08.png"/>

        </fig>

</sec>
<sec id="App1.Ch1.S1.SS2">
  <label>A2</label><title>Model structure optimization</title>
      <p id="d2e2743">The class of model optimized is that of a “traditional” CNN, i.e. a series of convolutional blocks followed by a fully connected network and ending with a single output node. Each convolutional block is formed from a specified number of convolutional layers, followed by a max-pooling layer. Every convolutional layer within one block shares the same number of filters, and each convolutional layer within any block is followed by a rectified linear unit activation function. Additional model hyperparameters or design choices such as the learning rate, batch size, kernel size of the first convolutional layer, and dropout strength of the fully connected network were also optimized. Model hyperparameters were optimized by sampling hyperparameters from the choices or ranges shown in Table <xref ref-type="table" rid="TA1"/>, fitting 500 models to the training and validation data described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS1"/>, and choosing the best-performing model structure when evaluated on the validation dataset. Chosen hyperparameter values of the final optimized ML-SPERE are shown in the final column of Table <xref ref-type="table" rid="TA1"/>.</p>

<table-wrap id="TA1" specific-use="star"><label>Table A1</label><caption><p id="d2e2755">Sampled ranges/values for optimized hyperparameters of ML-SPERE. Hyperparameters of the final optimized model are shown in the right-most column.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="3">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="left"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Hyperparameter</oasis:entry>
         <oasis:entry colname="col2">Sampling range or possible choices</oasis:entry>
         <oasis:entry colname="col3">Value of best-performing model</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Total # convolutional blocks</oasis:entry>
         <oasis:entry colname="col2">[1, 2, 3]</oasis:entry>
         <oasis:entry colname="col3">2</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Block 1, # convolutional layers</oasis:entry>
         <oasis:entry colname="col2">[1, 2, 3]</oasis:entry>
         <oasis:entry colname="col3">3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Block 1, # filters per layer</oasis:entry>
         <oasis:entry colname="col2">[16, 32, 64]</oasis:entry>
         <oasis:entry colname="col3">64</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Block 2, # convolutional layers</oasis:entry>
         <oasis:entry colname="col2">[1, 2, 3]</oasis:entry>
         <oasis:entry colname="col3">3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Block 2, # filters per layer</oasis:entry>
         <oasis:entry colname="col2">[32, 64, 128]</oasis:entry>
         <oasis:entry colname="col3">32</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Kernel size, convolutional layer #1</oasis:entry>
         <oasis:entry colname="col2">[3, 5, 7]</oasis:entry>
         <oasis:entry colname="col3">3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Batch size</oasis:entry>
         <oasis:entry colname="col2">[16, 32, 64]</oasis:entry>
         <oasis:entry colname="col3">64</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Dropout strength, fully connected network</oasis:entry>
         <oasis:entry colname="col2">[0.1, 0.2, 0.3, 0.4, 0.5, 0.6]</oasis:entry>
         <oasis:entry colname="col3">0.5</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Learning rate</oasis:entry>
         <oasis:entry colname="col2">(<inline-formula><mml:math id="M104" display="inline"><mml:mrow><mml:msup><mml:mn mathvariant="normal">10</mml:mn><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>, <inline-formula><mml:math id="M105" display="inline"><mml:mrow><mml:msup><mml:mn mathvariant="normal">10</mml:mn><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>), continuously sampled in log space</oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M106" display="inline"><mml:mrow><mml:mn mathvariant="normal">6.29596</mml:mn><mml:mo>×</mml:mo><mml:msup><mml:mn mathvariant="normal">10</mml:mn><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">4</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula></oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="App1.Ch1.S1.SS3">
  <label>A3</label><title>ML-SPERE performance under ideal observing conditions</title>
      <p id="d2e2948">When restricting the evaluation of ML-SPERE on the test set to only include scenes where the emission source is contained within the estimated plume mask (e.g. Fig. <xref ref-type="fig" rid="FA3"/>a and b), we find that our average emission estimates are greatly improved (Table <xref ref-type="table" rid="T1"/>, Fig. <xref ref-type="fig" rid="FA4"/>). The enhanced performance we find for ML-SPERE when restricting the test set to plumes in which the true source pixel falls within the estimated plume mask suggests that ML-SPERE relies strongly on information near the plume head or source region when inferring emission rates. This hypothesis is supported by the fact that ML-SPERE is trained on resampled plume abundance scenes without complicating backgrounds or missing spatial data (e.g. Sect. <xref ref-type="sec" rid="Ch1.S2.SS1"/>, Fig. <xref ref-type="fig" rid="F1"/>), and information around the plume head is always available during training.</p>

      <fig id="FA3" specific-use="star"><label>Figure A3</label><caption><p id="d2e2963">A few examples of test set plumes and their derived plume masks outlined in black. Emission source locations are shown with a black x. In <bold>(a)</bold>, the plume is easily delineated. In <bold>(b)</bold>, the head of the plume is captured, but missing data prevents the rest of the plume from being included in the mask. In <bold>(c)</bold>, background variability results in only some of the plume being successfully delineated, and the source pixel is not included in the mask. Colorbar values are methane concentrations in <inline-formula><mml:math id="M107" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">ppbv</mml:mi></mml:mrow></mml:math></inline-formula>.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f09.png"/>

        </fig>

      <fig id="FA4"><label>Figure A4</label><caption><p id="d2e2991">Estimated emission rates for simulated HYSPLIT plumes via <bold>(a)</bold> the IME method and <bold>(b)</bold> ML-SPERE as a function of true plume emission rates, filtered to only include scenes from the test set where the emission source location is contained within the estimated plume mask. The colormaps represent the relative density of the visualised data. We visualise bootstrapped ordinary least squares (OLS) regression lines for emission rate estimates from both the IME method (in <bold>c</bold>) and ML-SPERE (in <bold>d</bold>). The <inline-formula><mml:math id="M108" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> line is subtracted in <bold>(c)</bold> and <bold>(d)</bold> to show average bias.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f10.png"/>

        </fig>

      <p id="d2e3032">We investigate this directly by visualizing the partial gradient of the ML-SPERE emission rate estimate with respect to the pixels in the methane channel of the input image, as shown in Fig. <xref ref-type="fig" rid="FA5"/>. The gradients in Fig. <xref ref-type="fig" rid="FA5"/>b indicate that increasing methane concentrations immediately upwind of the source pixel decreases the emission-rate estimate, whereas increasing concentrations at the source pixel or in pixels immediately downwind increases the estimate. This behavior follows directly from mass-balance considerations and shows that ML-SPERE has learned to use methane enhancement gradients aligned with the wind vector to generate emission-rate predictions. Information in the plume tail is less influential, as reflected by the weaker gradients in that region in Fig. <xref ref-type="fig" rid="FA5"/>b.</p>

      <fig id="FA5" specific-use="star"><label>Figure A5</label><caption><p id="d2e3043">A methane scene from the test set <bold>(a)</bold> and the corresponding gradients between the emission rate estimated by ML-SPERE and the pixels in the methane channel <bold>(b)</bold>. Gradients are strongest around the head of the plume.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f11.png"/>

        </fig>

</sec>
<sec id="App1.Ch1.S1.SS4">
  <label>A4</label><title>ML-SPERE and IME uncertainty estimates</title>
      <p id="d2e3066">The probability integral transform (PIT), originally formalized by <xref ref-type="bibr" rid="bib1.bibx4" id="text.92"/>, provides an estimate of posterior calibration by recording, for a set of samples, the percentile ranks of the true values within their respective predicted posterior distributions. In the context of this work, we can construct PIT histograms for both ML-SPERE and the IME method by evaluating the cumulative distribution function of each emission rate posterior distribution (e.g. for each plume in the simulated test set) at the corresponding true value. An example for a single plume's posterior emission rate estimate is shown in Fig. <xref ref-type="fig" rid="FA6"/>a. For a well-calibrated model, these percentiles should be uniformly distributed when taken for a large sample of plumes.</p>

      <fig id="FA6" specific-use="star"><label>Figure A6</label><caption><p id="d2e3076"><bold>(a)</bold> Examples of estimated posterior emission rate distributions for a single scene via both ML-SPERE and the IME method. The true emission rate for this plume (45.4 <inline-formula><mml:math id="M109" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>) lies at a posterior percentile of 0.57 for ML-SPERE, and 1 for the IME method. Calculating and binning such values for both methods and all plumes in the synthetic test set generates <bold>(b)</bold> the PIT for ML-SPERE and <bold>(c)</bold> the PIT for the IME method.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f12.png"/>

        </fig>

      <p id="d2e3110">The PIT histograms for ML-SPERE (Fig. <xref ref-type="fig" rid="FA6"/>b) and the IME method (Fig. <xref ref-type="fig" rid="FA6"/>c) reveal differing calibration characteristics between the two methods. For ML-SPERE, the distribution exhibits the largest values in the outermost bins. This indicates a tendency to generate narrow uncertainty bands that do not always capture the true emission rate of the plume, though as discussed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and shown in Fig. <xref ref-type="fig" rid="FA6"/>, the uncertainty ranges on ML-SPERE estimates capture true emission rates more often than those of the IME method. In contrast, the IME method exhibits a pronounced spike at the upper bound (100th percentile), alongside an otherwise roughly uniform distribution. This accumulation at the upper tail indicates a systematic tendency to underestimate emission rates, such that true emission rates frequently exceed the upper bound of the estimated uncertainty range. Together, these results suggest that while both methods underestimate uncertainty, ML-SPERE provides better-calibrated and less biased posterior estimates than the IME approach.</p>
      <p id="d2e3122">The synthetic test set plumes were simulated using HYSPLIT with ERA5 wind data while ML-SPERE was trained using plumes simulated with WRF-Chem and NCEP wind data. Similarly, the IME effective wind speed calibration shown in Eq. (<xref ref-type="disp-formula" rid="Ch1.E2"/>) and given in <xref ref-type="bibr" rid="bib1.bibx54" id="text.93"/> was estimated using a synthetic plume set also simulated by WRF-Chem using NCEP wind data. The results shown in Fig. <xref ref-type="fig" rid="FA6"/> therefore suggest that ML-SPERE (and the methods used for estimating uncertainties outlined in Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/>) extrapolates better across different simulated plume datasets than the IME method implemented here does.</p>
</sec>
<sec id="App1.Ch1.S1.SS5">
  <label>A5</label><title><inline-formula><mml:math id="M110" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold cross validation analysis</title>
      <p id="d2e3150">To assess model generalization prior to training ML-SPERE on the full training and validation dataset, we conducted a <inline-formula><mml:math id="M111" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold cross-validation across the seven used WRF-Chem domains shown in Fig. <xref ref-type="fig" rid="FA1"/>. In each fold, the model was re-trained while withholding plumes from one WRF-Chem domain as a cross-validation test set, with the remaining plumes partitioned into training and validation subsets. The results of this <inline-formula><mml:math id="M112" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold analysis are presented in Fig. <xref ref-type="fig" rid="FA7"/> and Table <xref ref-type="table" rid="TA2"/>. Model fitting converges to comparable performance levels across folds. For all domains except Uzbekistan, the withheld test set performance is slightly lower than validation set performance. This outcome is expected, as the test sets are constructed to be independent from the training and validation data along the relevant dimension of variability, which in this case is the geographic location of plumes.</p>

      <fig id="FA7" specific-use="star"><label>Figure A7</label><caption><p id="d2e3175">Visualizations of ML-SPERE estimated emission rates <inline-formula><mml:math id="M113" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>est.</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> vs. true emission rates <inline-formula><mml:math id="M114" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>true</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> for each of the <inline-formula><mml:math id="M115" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold test sets. Values for MAPE and MdAPE are shown in each panel, and the red dashed lines show the <inline-formula><mml:math id="M116" display="inline"><mml:mrow><mml:mn mathvariant="normal">1</mml:mn><mml:mo>:</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:math></inline-formula> line. The colormaps represent the relative density of the visualised data.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f13.png"/>

        </fig>

<table-wrap id="TA2"><label>Table A2</label><caption><p id="d2e3229">Test and validation set performance for each of the <inline-formula><mml:math id="M117" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold datasets created before final model fitting.</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="3">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:thead>
       <oasis:row>
         <oasis:entry colname="col1">Withheld WRF</oasis:entry>
         <oasis:entry colname="col2">Validation</oasis:entry>
         <oasis:entry colname="col3">Test set</oasis:entry>
       </oasis:row>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">domain</oasis:entry>
         <oasis:entry colname="col2">set MAPE</oasis:entry>
         <oasis:entry colname="col3">MAPE</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">Argentinia</oasis:entry>
         <oasis:entry colname="col2">13.89</oasis:entry>
         <oasis:entry colname="col3">15.18</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Azerbaijan</oasis:entry>
         <oasis:entry colname="col2">15.18</oasis:entry>
         <oasis:entry colname="col3">19.3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Iraq</oasis:entry>
         <oasis:entry colname="col2">15.48</oasis:entry>
         <oasis:entry colname="col3">18.09</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Kazakhstan</oasis:entry>
         <oasis:entry colname="col2">14.95</oasis:entry>
         <oasis:entry colname="col3">16.03</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Mexico</oasis:entry>
         <oasis:entry colname="col2">15.66</oasis:entry>
         <oasis:entry colname="col3">15.91</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Turkmenistan</oasis:entry>
         <oasis:entry colname="col2">14.5</oasis:entry>
         <oasis:entry colname="col3">15.69</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">Uzbekistan</oasis:entry>
         <oasis:entry colname="col2">16.83</oasis:entry>
         <oasis:entry colname="col3">14.44</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>


</sec>
<sec id="App1.Ch1.S1.SS6">
  <label>A6</label><title>Examining changes in ML-SPERE performance when estimating emission rates for HYSPLIT plumes vs. WRF-Chem plumes</title>
      <p id="d2e3375">We produced three synthetic plume sets for a cluster of four coal mines in Kazakhstan for the year 2022, temporally sampled with daily output at the TROPOMI overpass hour and spatially resampled to the TROPOMI pixel footprints of the corresponding TROPOMI orbit. In the first set, plumes were simulated with WRF-Chem using NCEP meteorology for boundary conditions, which corresponds to the settings used to create the training and validation set plumes used to train ML-SPERE. In the second plume set, plumes were simulated with WRF-Chem using ERA5 meteorology for boundary conditions. In the third plume set, plumes were simulated with HYSPLIT using ERA5 meteorology for particle transport, which mimics the global set of plumes used as a test set for ML-SPERE in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>. Across all three datasets, plumes were sampled at identical times, resampled to identical TROPOMI pixel orbit geometries, and use identical TROPOMI background data to produce a realistic scene.</p>
      <p id="d2e3380">Figure <xref ref-type="fig" rid="FA8"/> shows ML-SPERE estimates against true emission rates for each of these plume sets. ML-SPERE estimates for the plume sets simulated with WRF-Chem exhibit values of MAPE and MdAPE similar to those seen in the <inline-formula><mml:math id="M118" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold cross validation analysis of Sect. <xref ref-type="sec" rid="App1.Ch1.S1.SS5"/> (which did not include methane backgrounds or missing data), but the plume set simulated with HYSPLIT exhibits higher values of MAPE and MdAPE that are similar to the results of Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>. A visual inspection of these plume sets shows that the two plume sets simulated with WRF-Chem look very similar, despite using different meteorological datasets as boundary conditions, while the plumes from the HYSPLIT set can at times exhibit different morphologies, sometimes appearing more diffuse than the WRF-Chem plumes. This suggests that the difference in performance that ML-SPERE shows between the <inline-formula><mml:math id="M119" display="inline"><mml:mi>K</mml:mi></mml:math></inline-formula>-fold cross validation analysis and the test set results in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/> is driven by morphological differences between plumes simulated by WRF-Chem and HYSPLIT, and is not due to the change in meteorological data between the training data and the test data or the greater geographic range of plumes simulated in the test set. WRF-Chem and HYSPLIT can simulate plumes that are different from each other for a variety of reasons. WRF-Chem can only simulate “point sources” at a resolution equivalent to the native grid of the simulation, whereas HYSPLIT simulates particle emission at an actual point. Furthermore, WRF-Chem uses meteorology as boundary conditions for a spatial domain, whereas HYSPLIT uses meteorology to directly transport particles throughout a spatial domain and can be sensitive to diffusion parameterizations.</p>

      <fig id="FA8" specific-use="star"><label>Figure A8</label><caption><p id="d2e3408">ML-SPERE emission rate estimates (<inline-formula><mml:math id="M120" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>ML-SPERE</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula>) as a function of true plume emission rates (<inline-formula><mml:math id="M121" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>true</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula>) for three plume sets simulated for coal mines in Kazakhstan. Subpanel titles indicate the chemical transport code (either WRF-Chem or HYSPLIT) and the wind data (either NCEP or ERA5) used to simulate the plumes. ML-SPERE performance on plume sets simulated with WRF-Chem (panels <bold>a</bold> and <bold>b</bold>) is comparable to that of the validation plume set, whereas performance on the HYSPLIT plume set (panel <bold>c</bold>) is modestly degraded. This demonstrates that reduced test set performance arises primarily from differences in the transport models used to generate the plumes of the training and test sets, and not the global extrapolation. The colormaps represent the relative density of the visualised data.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f14.png"/>

        </fig>

</sec>
<sec id="App1.Ch1.S1.SS7">
  <label>A7</label><title>ML-SPERE estimates for  Karaturun East 2023 blowout</title>
      <p id="d2e3456">In Table <xref ref-type="table" rid="TA3"/>, we present the ML-SPERE emission rate estimates corresponding to the data from Fig. <xref ref-type="fig" rid="F5"/>. The table also reports the estimated contributions to the total uncertainty from plume masking, background removal, wind field error, and model error.</p>

<table-wrap id="TA3" specific-use="star"><label>Table A3</label><caption><p id="d2e3466">ML-SPERE emission estimate metrics by date for the Karaturun East 2023 blowout in Kazakhstan. <inline-formula><mml:math id="M122" display="inline"><mml:mi>Q</mml:mi></mml:math></inline-formula> refers to the posterior distribution of estimated emission rates obtained from the procedure described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/>. <inline-formula><mml:math id="M123" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>median</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula> is the median of the posterior distribution of estimated emission rates. <inline-formula><mml:math id="M124" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi mathvariant="italic">σ</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula> is the standard deviation of the posterior distribution of estimated emission rates. <inline-formula><mml:math id="M125" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">16</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M126" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">84</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula> are the 16th and 84th percentiles of the posterior distribution of estimated emission rates, respectively. Wind, masking, background, and model error are all estimated as described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/> (i.e. they do not sum to <inline-formula><mml:math id="M127" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi mathvariant="italic">σ</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula>). All table values are [<inline-formula><mml:math id="M128" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>].</p></caption><oasis:table frame="topbot"><oasis:tgroup cols="9">
     <oasis:colspec colnum="1" colname="col1" align="left"/>
     <oasis:colspec colnum="2" colname="col2" align="right"/>
     <oasis:colspec colnum="3" colname="col3" align="right"/>
     <oasis:colspec colnum="4" colname="col4" align="right"/>
     <oasis:colspec colnum="5" colname="col5" align="right"/>
     <oasis:colspec colnum="6" colname="col6" align="right"/>
     <oasis:colspec colnum="7" colname="col7" align="right"/>
     <oasis:colspec colnum="8" colname="col8" align="right"/>
     <oasis:colspec colnum="9" colname="col9" align="right"/>
     <oasis:thead>
       <oasis:row rowsep="1">
         <oasis:entry colname="col1">Date</oasis:entry>
         <oasis:entry colname="col2"><inline-formula><mml:math id="M129" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mtext>median</mml:mtext></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col3"><inline-formula><mml:math id="M130" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">16</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col4"><inline-formula><mml:math id="M131" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mn mathvariant="normal">84</mml:mn></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col5"><inline-formula><mml:math id="M132" display="inline"><mml:mrow><mml:msub><mml:mi>Q</mml:mi><mml:mi mathvariant="italic">σ</mml:mi></mml:msub></mml:mrow></mml:math></inline-formula></oasis:entry>
         <oasis:entry colname="col6">Wind error</oasis:entry>
         <oasis:entry colname="col7">Masking error</oasis:entry>
         <oasis:entry colname="col8">Background error</oasis:entry>
         <oasis:entry colname="col9">Model error</oasis:entry>
       </oasis:row>
     </oasis:thead>
     <oasis:tbody>
       <oasis:row>
         <oasis:entry colname="col1">21 June 2023</oasis:entry>
         <oasis:entry colname="col2">112.9</oasis:entry>
         <oasis:entry colname="col3">74.9</oasis:entry>
         <oasis:entry colname="col4">157.8</oasis:entry>
         <oasis:entry colname="col5">40.5</oasis:entry>
         <oasis:entry colname="col6">41.4</oasis:entry>
         <oasis:entry colname="col7">15.5</oasis:entry>
         <oasis:entry colname="col8">19.2</oasis:entry>
         <oasis:entry colname="col9">11.2</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">23 June 2023</oasis:entry>
         <oasis:entry colname="col2">81.2</oasis:entry>
         <oasis:entry colname="col3">53.3</oasis:entry>
         <oasis:entry colname="col4">111.4</oasis:entry>
         <oasis:entry colname="col5">28.8</oasis:entry>
         <oasis:entry colname="col6">25.9</oasis:entry>
         <oasis:entry colname="col7">11.0</oasis:entry>
         <oasis:entry colname="col8">18.6</oasis:entry>
         <oasis:entry colname="col9">7.6</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">26 June 2023</oasis:entry>
         <oasis:entry colname="col2">121.4</oasis:entry>
         <oasis:entry colname="col3">85.6</oasis:entry>
         <oasis:entry colname="col4">164.0</oasis:entry>
         <oasis:entry colname="col5">37.2</oasis:entry>
         <oasis:entry colname="col6">35.7</oasis:entry>
         <oasis:entry colname="col7">16.5</oasis:entry>
         <oasis:entry colname="col8">20.6</oasis:entry>
         <oasis:entry colname="col9">11.5</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">4 July 2023</oasis:entry>
         <oasis:entry colname="col2">96.8</oasis:entry>
         <oasis:entry colname="col3">61.3</oasis:entry>
         <oasis:entry colname="col4">139.9</oasis:entry>
         <oasis:entry colname="col5">40.1</oasis:entry>
         <oasis:entry colname="col6">37.6</oasis:entry>
         <oasis:entry colname="col7">14.6</oasis:entry>
         <oasis:entry colname="col8">24.2</oasis:entry>
         <oasis:entry colname="col9">9.5</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">11 July 2023</oasis:entry>
         <oasis:entry colname="col2">78.8</oasis:entry>
         <oasis:entry colname="col3">57.4</oasis:entry>
         <oasis:entry colname="col4">101.1</oasis:entry>
         <oasis:entry colname="col5">21.8</oasis:entry>
         <oasis:entry colname="col6">16.4</oasis:entry>
         <oasis:entry colname="col7">11.8</oasis:entry>
         <oasis:entry colname="col8">16.7</oasis:entry>
         <oasis:entry colname="col9">7.3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">12 July 2023</oasis:entry>
         <oasis:entry colname="col2">89.5</oasis:entry>
         <oasis:entry colname="col3">68.2</oasis:entry>
         <oasis:entry colname="col4">114.6</oasis:entry>
         <oasis:entry colname="col5">23.1</oasis:entry>
         <oasis:entry colname="col6">20.8</oasis:entry>
         <oasis:entry colname="col7">14.6</oasis:entry>
         <oasis:entry colname="col8">16.3</oasis:entry>
         <oasis:entry colname="col9">8.6</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">14 July 2023</oasis:entry>
         <oasis:entry colname="col2">89.0</oasis:entry>
         <oasis:entry colname="col3">64.2</oasis:entry>
         <oasis:entry colname="col4">117.0</oasis:entry>
         <oasis:entry colname="col5">26.8</oasis:entry>
         <oasis:entry colname="col6">25.7</oasis:entry>
         <oasis:entry colname="col7">12.5</oasis:entry>
         <oasis:entry colname="col8">17.5</oasis:entry>
         <oasis:entry colname="col9">8.7</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">6 August 2023</oasis:entry>
         <oasis:entry colname="col2">78.7</oasis:entry>
         <oasis:entry colname="col3">46.2</oasis:entry>
         <oasis:entry colname="col4">127.3</oasis:entry>
         <oasis:entry colname="col5">40.6</oasis:entry>
         <oasis:entry colname="col6">41.4</oasis:entry>
         <oasis:entry colname="col7">12.8</oasis:entry>
         <oasis:entry colname="col8">13.3</oasis:entry>
         <oasis:entry colname="col9">8.5</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">8 August 2023</oasis:entry>
         <oasis:entry colname="col2">68.8</oasis:entry>
         <oasis:entry colname="col3">48.0</oasis:entry>
         <oasis:entry colname="col4">98.1</oasis:entry>
         <oasis:entry colname="col5">25.8</oasis:entry>
         <oasis:entry colname="col6">23.2</oasis:entry>
         <oasis:entry colname="col7">10.9</oasis:entry>
         <oasis:entry colname="col8">14.9</oasis:entry>
         <oasis:entry colname="col9">7.0</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">9 August 2023</oasis:entry>
         <oasis:entry colname="col2">94.9</oasis:entry>
         <oasis:entry colname="col3">64.4</oasis:entry>
         <oasis:entry colname="col4">129.6</oasis:entry>
         <oasis:entry colname="col5">31.5</oasis:entry>
         <oasis:entry colname="col6">29.5</oasis:entry>
         <oasis:entry colname="col7">14.5</oasis:entry>
         <oasis:entry colname="col8">19.1</oasis:entry>
         <oasis:entry colname="col9">8.7</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">21 August 2023</oasis:entry>
         <oasis:entry colname="col2">83.9</oasis:entry>
         <oasis:entry colname="col3">44.6</oasis:entry>
         <oasis:entry colname="col4">132.2</oasis:entry>
         <oasis:entry colname="col5">43.8</oasis:entry>
         <oasis:entry colname="col6">39.5</oasis:entry>
         <oasis:entry colname="col7">17.2</oasis:entry>
         <oasis:entry colname="col8">20.0</oasis:entry>
         <oasis:entry colname="col9">8.5</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">31 August 2023</oasis:entry>
         <oasis:entry colname="col2">44.9</oasis:entry>
         <oasis:entry colname="col3">27.4</oasis:entry>
         <oasis:entry colname="col4">67.1</oasis:entry>
         <oasis:entry colname="col5">19.8</oasis:entry>
         <oasis:entry colname="col6">16.8</oasis:entry>
         <oasis:entry colname="col7">10.2</oasis:entry>
         <oasis:entry colname="col8">11.5</oasis:entry>
         <oasis:entry colname="col9">4.3</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">1 September 2023</oasis:entry>
         <oasis:entry colname="col2">46.4</oasis:entry>
         <oasis:entry colname="col3">26.6</oasis:entry>
         <oasis:entry colname="col4">70.9</oasis:entry>
         <oasis:entry colname="col5">21.7</oasis:entry>
         <oasis:entry colname="col6">19.5</oasis:entry>
         <oasis:entry colname="col7">7.3</oasis:entry>
         <oasis:entry colname="col8">13.8</oasis:entry>
         <oasis:entry colname="col9">4.6</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">3 September 2023</oasis:entry>
         <oasis:entry colname="col2">56.3</oasis:entry>
         <oasis:entry colname="col3">30.9</oasis:entry>
         <oasis:entry colname="col4">95.2</oasis:entry>
         <oasis:entry colname="col5">32.8</oasis:entry>
         <oasis:entry colname="col6">28.6</oasis:entry>
         <oasis:entry colname="col7">17.7</oasis:entry>
         <oasis:entry colname="col8">11.2</oasis:entry>
         <oasis:entry colname="col9">6.2</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">21 September 2023</oasis:entry>
         <oasis:entry colname="col2">23.9</oasis:entry>
         <oasis:entry colname="col3">19.2</oasis:entry>
         <oasis:entry colname="col4">29.4</oasis:entry>
         <oasis:entry colname="col5">5.5</oasis:entry>
         <oasis:entry colname="col6">5.0</oasis:entry>
         <oasis:entry colname="col7">3.5</oasis:entry>
         <oasis:entry colname="col8">3.8</oasis:entry>
         <oasis:entry colname="col9">2.0</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">25 September 2023</oasis:entry>
         <oasis:entry colname="col2">95.3</oasis:entry>
         <oasis:entry colname="col3">68.0</oasis:entry>
         <oasis:entry colname="col4">126.7</oasis:entry>
         <oasis:entry colname="col5">28.8</oasis:entry>
         <oasis:entry colname="col6">26.5</oasis:entry>
         <oasis:entry colname="col7">14.0</oasis:entry>
         <oasis:entry colname="col8">18.9</oasis:entry>
         <oasis:entry colname="col9">9.4</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">26 September 2023</oasis:entry>
         <oasis:entry colname="col2">94.3</oasis:entry>
         <oasis:entry colname="col3">68.3</oasis:entry>
         <oasis:entry colname="col4">120.6</oasis:entry>
         <oasis:entry colname="col5">24.9</oasis:entry>
         <oasis:entry colname="col6">25.9</oasis:entry>
         <oasis:entry colname="col7">12.7</oasis:entry>
         <oasis:entry colname="col8">13.9</oasis:entry>
         <oasis:entry colname="col9">8.7</oasis:entry>
       </oasis:row>
       <oasis:row>
         <oasis:entry colname="col1">27 September 2023</oasis:entry>
         <oasis:entry colname="col2">58.4</oasis:entry>
         <oasis:entry colname="col3">46.6</oasis:entry>
         <oasis:entry colname="col4">76.2</oasis:entry>
         <oasis:entry colname="col5">16.9</oasis:entry>
         <oasis:entry colname="col6">15.7</oasis:entry>
         <oasis:entry colname="col7">8.6</oasis:entry>
         <oasis:entry colname="col8">9.5</oasis:entry>
         <oasis:entry colname="col9">5.4</oasis:entry>
       </oasis:row>
     </oasis:tbody>
   </oasis:tgroup></oasis:table></table-wrap>

</sec>
<sec id="App1.Ch1.S1.SS8">
  <label>A8</label><title>Geographic trends in ML-SPERE and IME estimate residuals for the 2021 TROPOMI methane detection plume set</title>
      <p id="d2e4220">In Fig. <xref ref-type="fig" rid="FA9"/>, we show mean standardised logratios between ML-SPERE and IME estimates for the plumes catalogued in <xref ref-type="bibr" rid="bib1.bibx55" id="text.94"/> on a 15<inline-formula><mml:math id="M133" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">°</mml:mi></mml:mrow></mml:math></inline-formula> global grid. This allows us to examine the ratios of estimates obtained via ML-SPERE and the IME method in terms of standard deviations away from the mean ratio of the two methods across the entire dataset, and determine if there are any significant regional trends in the differences between estimates obtained via the two methods (defined as being either greater than 0.5 standard deviations or less than <inline-formula><mml:math id="M134" display="inline"><mml:mo>-</mml:mo></mml:math></inline-formula>0.5 standard deviations of the dataset-wide mean standardised logratio). The figure shows that statistically significant regional differences between ML-SPERE and IME emission-rate estimates occur primarily in northern Russia and southeastern Australia (where ML-SPERE estimates are systematically higher) and in the Congo region (where ML-SPERE estimates are systematically lower). For all these regions, we have a small sample size of plumes to examine.</p>

      <fig id="FA9" specific-use="star"><label>Figure A9</label><caption><p id="d2e4245">Global standardised logratios between emission rate estimates from ML-SPERE and the IME method for the plume set of <xref ref-type="bibr" rid="bib1.bibx55" id="text.95"/>, expressed as mean per grid cell. Number of plumes are plotted in each grid cell.</p></caption>
          <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f15.png"/>

        </fig>

<sec id="App1.Ch1.S1.SS8.SSS1">
  <label>A8.1</label><title>ML-SPERE estimates for plumes in northern Russia</title>
      <p id="d2e4264">We find that ML-SPERE estimates for detected plumes in northern Russia are on average higher than those from the IME method, with mean standardised logratios 0.5–0.75 standard deviations above the dataset-wide mean. Although the true emission rates for these plumes are unknown, the average IME estimate for plumes in this region are 34 <inline-formula><mml:math id="M135" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. Figure <xref ref-type="fig" rid="F3"/>d shows that ML-SPERE may overestimate emissions for plumes with true emission rates below 50 <inline-formula><mml:math id="M136" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace linebreak="nobreak" width="0.125em"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> due to regression dilution, if no human oversight is used to limit the application of ML-SPERE to plumes that are well-observed and have complete spatial information around the plume head.</p>

      <fig id="FA10" specific-use="star"><label>Figure A10</label><caption><p id="d2e4305"><bold>(a)</bold> Locations for the plumes in northern Russia for which ML-SPERE estimates exhibit the largest log-ratios when compared to the IME method, shown in red dots. Pipelines from the Global Energy Monitor are shown with grey lines. Examples of methane observations are shown in panels <bold>(b–e)</bold>, and colorbar values show methane concentrations in <inline-formula><mml:math id="M137" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">ppbv</mml:mi></mml:mrow></mml:math></inline-formula>. Latitude-longitude coordinates of the most upwind pixel within the plume masks are shown in the titles of panels <bold>(b–e)</bold>.</p></caption>
            <graphic xlink:href="https://amt.copernicus.org/articles/19/5659/2026/amt-19-5659-2026-f16.png"/>

          </fig>

      <p id="d2e4330">Mapping the locations of these plumes shows that they coincide with known natural gas pipeline infrastructure. Visual inspection shows that these scenes contain extremely compact, blob-like methane enhancements with very high concentrations (Fig. <xref ref-type="fig" rid="FA10"/>). Such features are characteristic of transient emissions (e.g. blowdowns at block valves) rather than the steady-state plumes on which ML-SPERE was trained. As shown in Sects. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and <xref ref-type="sec" rid="App1.Ch1.S1.SS3"/>, ML-SPERE relies heavily on methane enhancement gradients aligned with local wind vectors to infer emission rates. It is therefore plausible that ML-SPERE may overestimate emission rates for short-duration, highly concentrated emissions where mass-balance structure differs from that of steady-state plumes.</p>
      <p id="d2e4340">As discussed in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>, human oversight is essential when applying ML-SPERE (though this holds for all methods and not only for ML-based methods). Transient, non steady-state emissions may not be well estimated by either ML-SPERE or the IME method, and total plume mass may be more relevant. Finally, although ML-SPERE yields higher estimates than the IME method in these cases, the true emission rates are still unknown, and we cannot say with certainty that ML-SPERE estimates for these specific plumes were less accurate than IME estimates.</p>
</sec>
<sec id="App1.Ch1.S1.SS8.SSS2">
  <label>A8.2</label><title>ML-SPERE estimates in the Congo basin</title>
      <p id="d2e4354">As shown in Fig. <xref ref-type="fig" rid="FA9"/>, ML-SPERE estimates for plumes in the Congo Basin are on average lower than the corresponding IME estimates. Visual inspection of these scenes reveals very large plumes that are frequently spatially disrupted by cloud cover or by gaps in TROPOMI retrievals over water. As discussed in Sects. <xref ref-type="sec" rid="Ch1.S3.SS1"/> and <xref ref-type="sec" rid="App1.Ch1.S1.SS3"/>, ML-SPERE performs best when the region surrounding the plume head or emission source is well observed, and analyst judgment may be required when these conditions are not met. The Congo Basin plumes are also on average approximately 50 % larger than those in the training dataset, potentially reflecting more spatially diffusive emitters than those represented in the simulations. Consequently, ML-SPERE may be less well suited for application to these scenes, though it is important to recognize that the IME method was not designed to be applied to area-source emissions. Both methods may struggle to characterize emissions in this region.</p>
</sec>
<sec id="App1.Ch1.S1.SS8.SSS3">
  <label>A8.3</label><title>ML-SPERE estimates in Australia</title>
      <p id="d2e4371">For the 11 plumes in the grid cell in southwestern Australia in Fig. <xref ref-type="fig" rid="FA9"/>, ML-SPERE estimates are on average higher than IME estimates. Inspection of these plumes show that they have low wind speeds below 2 <inline-formula><mml:math id="M138" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">m</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">s</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> and average emission rates estimated via the IME method at 14 <inline-formula><mml:math id="M139" display="inline"><mml:mrow class="unit"><mml:mi mathvariant="normal">t</mml:mi><mml:mspace width="0.125em" linebreak="nobreak"/><mml:msup><mml:mi mathvariant="normal">h</mml:mi><mml:mrow><mml:mo>-</mml:mo><mml:mn mathvariant="normal">1</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula>. This suggests that although this may be an emission rate regime where ML-SPERE may overestimate emissions due to regression dilution, it is also possible that the IME method is underestimating emissions due to the low wind speed bias outlined in Sect. <xref ref-type="sec" rid="Ch1.S3.SS1"/>. This bias is hinted at operationally for the entire 2021 plume dataset in Fig. <xref ref-type="fig" rid="F6"/>b.</p>
</sec>
</sec>
</app>
  </app-group><notes notes-type="codedataavailability"><title>Code and data availability</title>

      <p id="d2e4420">The specific version of the TROPOMI data used in the Karaturun East case study is available at <uri>ftp://ftp.sron.nl/open-access-data-2/TROPOMI/tropomi/ch4/19_446/</uri> (last access: 14 August 2026). The dataset of detected plumes in 2021 TROPOMI data is available at  <ext-link xlink:href="https://doi.org/10.5281/zenodo.8087133" ext-link-type="DOI">10.5281/zenodo.8087133</ext-link> <xref ref-type="bibr" rid="bib1.bibx55" id="paren.96"/>. The WRF-Chem <xref ref-type="bibr" rid="bib1.bibx57" id="paren.97"/> code is available at <uri>https://github.com/wrf-model/WRF/releases/</uri>  (last access: 14 August 2026); in this work, version 4.1.5 was used. The HYSPLIT <xref ref-type="bibr" rid="bib1.bibx58 bib1.bibx8 bib1.bibx9 bib1.bibx7" id="paren.98"/> code is available at <uri>https://www.ready.noaa.gov/HYSPLIT.php</uri> (last access: 14 August 2026); in this work, version 5.2.3 was used. ERA5 data are available at <ext-link xlink:href="https://doi.org/10.24381/cds.bd0915c6" ext-link-type="DOI">10.24381/cds.bd0915c6</ext-link> <xref ref-type="bibr" rid="bib1.bibx20" id="paren.99"/> and <ext-link xlink:href="https://doi.org/10.24381/cds.adbb2d47" ext-link-type="DOI">10.24381/cds.adbb2d47</ext-link> <xref ref-type="bibr" rid="bib1.bibx21" id="paren.100"/>. NCEP data <xref ref-type="bibr" rid="bib1.bibx42" id="paren.101"/> are available at <uri>https://gdex.ucar.edu/datasets/d083002/</uri>  (last access: 14 August 2026). The trained ML-SPERE model is available at <ext-link xlink:href="https://doi.org/10.5281/zenodo.21786440" ext-link-type="DOI">10.5281/zenodo.21786440</ext-link> <xref ref-type="bibr" rid="bib1.bibx51" id="paren.102"/>. The codebase we use to implement ML-SPERE (i.e. generate emission rate estimates with full uncertainties as described in Sect. <xref ref-type="sec" rid="Ch1.S2.SS4"/>) can be found with release v1.0.4 at <ext-link xlink:href="https://doi.org/10.5281/zenodo.21934190" ext-link-type="DOI">10.5281/zenodo.21934190</ext-link> <xref ref-type="bibr" rid="bib1.bibx49" id="paren.103"/>. The test dataset of HYSPLIT plumes resampled to TROPOMI pixel footprints is available at  <ext-link xlink:href="https://doi.org/10.5281/zenodo.21792154" ext-link-type="DOI">10.5281/zenodo.21792154</ext-link> <xref ref-type="bibr" rid="bib1.bibx50" id="paren.104"/>. The training and validation datasets of WRF-Chem plumes resampled to TROPOMI pixel footprints is available at <ext-link xlink:href="https://doi.org/10.5281/zenodo.21808941" ext-link-type="DOI">10.5281/zenodo.21808941</ext-link> <xref ref-type="bibr" rid="bib1.bibx52" id="paren.105"/>.</p>
  </notes><notes notes-type="authorcontribution"><title>Author contributions</title>

      <p id="d2e4494">CR, JDM, and IA designed the study. CR wrote the code to develop ML-SPERE and produce the analysis for the paper. CR wrote the paper with contributions from all authors. TAdJ provided the HYSPLIT simulations to produce the test set. TH provided the TROPOMI backgrounds used in the creation of the test set. SH provided the WRF simulations used to train ML-SPERE. AvdB conducted an initial exploration into the feasibility of using WRF simulations as training data for this study.  BJS calculated all IME estimates in the paper. SS provided the simulations used in the Karaturun East blowout inversion analysis.</p>
  </notes><notes notes-type="competinginterests"><title>Competing interests</title>

      <p id="d2e4500">At least one of the (co-)authors is a member of the editorial board of <italic>Atmospheric Measurement Techniques</italic>. The peer-review process was guided by an independent editor, and the authors also have no other competing interests to declare.</p>
  </notes><notes notes-type="disclaimer"><title>Disclaimer</title>

      <p id="d2e4509">Publisher's note: Copernicus Publications remains neutral with regard to jurisdictional claims made in the text, published maps, institutional affiliations, or any other geographical representation in this paper. The authors bear the ultimate responsibility for providing appropriate place names. Views expressed in the text are those of the authors and do not necessarily reflect the views of the publisher.</p>
  </notes><ack><title>Acknowledgements</title><p id="d2e4515">We acknowledge and thank Matthieu Dogniaux for useful discussions and for providing feedback on the first draft version of this manuscript. We thank the team that realized the TROPOMI instrument and its data products, consisting of the partnership between Airbus Defence and Space Netherlands, KNMI, SRON, and TNO and commissioned by NSO and ESA. The Sentinel-5 Precursor is part of the EU Copernicus program, and Copernicus (modified) Sentinel-5P data (2021, 2023) have been used. The authors acknowledge the support of SURF Cooperative, as part of this work was carried out on the Dutch national e-infrastructure. We thank the referees for their time and expertise in reviewing the manuscript.</p></ack><notes notes-type="financialsupport"><title>Financial support</title>

      <p id="d2e4520">CR acknowledges funding from the NSO TROPOMI national program and the ESA Satellite Monitoring of Atmospheric Methane (SMART-CH4) project. TAdJ and SS acknowledge funding from the framework of UNEP's International Methane Emissions Observatory (IMEO).</p>
  </notes><notes notes-type="reviewstatement"><title>Review statement</title>

      <p id="d2e4526">This paper was edited by Andre Butz and reviewed by two anonymous referees.</p>
  </notes><ref-list>
    <title>References</title>

      <ref id="bib1.bibx1"><label>Abadi et al.(2015)</label><mixed-citation>Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M., Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., and Zheng, X.: TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems, <uri>https://www.tensorflow.org/</uri> (last access: 14 August 2026), 2015.</mixed-citation></ref>
      <ref id="bib1.bibx2"><label>Bruno et al.(2024)</label><mixed-citation>Bruno, J. H., Jervis, D., Varon, D. J., and Jacob, D. J.: U-Plume: automated algorithm for plume detection and source quantification by satellite point-source imagers, Atmos. Meas. Tech., 17, 2625–2636, <ext-link xlink:href="https://doi.org/10.5194/amt-17-2625-2024" ext-link-type="DOI">10.5194/amt-17-2625-2024</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx3"><label>Copernicus Atmospheric Monitoring Service(2025)</label><mixed-citation>Copernicus Atmospheric Monitoring Service: CAMS Methane Hotspot Explorer, <uri>https://atmosphere.copernicus.eu/ghg-services/cams-methane-hotspot-explorer</uri> (last access: 14 August 2026), 2025.</mixed-citation></ref>
      <ref id="bib1.bibx4"><label>David and Johnson(1948)</label><mixed-citation>David, F. N. and Johnson, N. L.: The probability integral transformation when parameters are estimated from the sample, Biometrika, 35, 182, <ext-link xlink:href="https://doi.org/10.2307/2332638" ext-link-type="DOI">10.2307/2332638</ext-link>, 1948.</mixed-citation></ref>
      <ref id="bib1.bibx5"><label>De Jong et al.(2025)</label><mixed-citation>De Jong, T. A., Maasakkers, J. D., Irakulis Loitxate, I., Randles, C. A., Tol, P., and Aben, I.: Daily global methane super emitter detection and source identification with sub daily tracking, Geophys. Res. Lett., 52, e2024GL111824, <ext-link xlink:href="https://doi.org/10.1029/2024GL111824" ext-link-type="DOI">10.1029/2024GL111824</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx6"><label>De Myttenaere et al.(2016)</label><mixed-citation>De Myttenaere, A., Golden, B., Le Grand, B., and Rossi, F.: Mean absolute percentage error for regression models, Neurocomputing, 192, 38–48, <ext-link xlink:href="https://doi.org/10.1016/j.neucom.2015.12.114" ext-link-type="DOI">10.1016/j.neucom.2015.12.114</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx7"><label>Draxler(1999)</label><mixed-citation> Draxler, R. R.: HYSPLIT4 Users's Guide, NOAA Tech. Memo. ERL ARL-230, 1999.</mixed-citation></ref>
      <ref id="bib1.bibx8"><label>Draxler and Hess(1997)</label><mixed-citation> Draxler, R. R. and Hess, G. D.: Description of the HYSPLIT4 Modeling System, NOAA Tech. Memo. ERL ARL-224, 1997.</mixed-citation></ref>
      <ref id="bib1.bibx9"><label>Draxler and Hess(1998)</label><mixed-citation> Draxler, R. R. and Hess, G. D.: An overview of the HYSPLIT_4 modelling system for trajectories, Aust. Meteorol. Mag., 47, 295–308, 1998.</mixed-citation></ref>
      <ref id="bib1.bibx10"><label>European Union(2026)</label><mixed-citation>European Union: Harmonized Sentinel-2 MSI: MultiSpectral Instrument, Level-2A (SR), accessed via Google Earth Engine: COPERNICUS/S2_SR_HARMONIZED, <ext-link xlink:href="https://doi.org/10.5270/S2_-znk9xsj" ext-link-type="DOI">10.5270/S2_-znk9xsj</ext-link>, 2026.</mixed-citation></ref>
      <ref id="bib1.bibx11"><label>Frankenberg et al.(2016)</label><mixed-citation>Frankenberg, C., Thorpe, A. K., Thompson, D. R., Hulley, G., Kort, E. A., Vance, N., Borchardt, J., Krings, T., Gerilowski, K., Sweeney, C., Conley, S., Bue, B. D., Aubrey, A. D., Hook, S., and Green, R. O.: Airborne methane remote measurements reveal heavy-tail flux distribution in Four Corners region, P. Natl. Acad. Sci. USA, 113, 9734–9739, <ext-link xlink:href="https://doi.org/10.1073/pnas.1605617113" ext-link-type="DOI">10.1073/pnas.1605617113</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx12"><label>Gal and Ghahramani(2016)</label><mixed-citation>Gal, Y. and Ghahramani, Z.: Dropout as a Bayesian approximation: representing model uncertainty in deep learning, in: International Conference on Machine Learning, 1050–1059, PMLR, <ext-link xlink:href="https://doi.org/10.48550/arXiv.1506.02142" ext-link-type="DOI">10.48550/arXiv.1506.02142</ext-link>, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx13"><label>Gao et al.(2023)</label><mixed-citation>Gao, S., Liu, X., Chen, Y., Jiang, J., Liu, Y., and Jiang, Y.: Atmospheric turbulence strength estimation using convolution neural network, IEEE Photonics J., 15, 1–7, <ext-link xlink:href="https://doi.org/10.1109/JPHOT.2023.3314833" ext-link-type="DOI">10.1109/JPHOT.2023.3314833</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx14"><label>Gorelick et al.(2017)</label><mixed-citation> Gorelick, N., Hancher, M., Dixon, M., Ilyushchenko, S., Thau, D., and Moore, R.: Google Earth Engine: planetary-scale geospatial analysis for everyone, Remote Sens. Environ., 202, 18–27, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx15"><label>Grell et al.(2005)</label><mixed-citation>Grell, G. A., Peckham, S. E., Schmitz, R., McKeen, S. A., Frost, G., Skamarock, W. C., and Eder, B.: Fully coupled “online” chemistry within the WRF model, Atmos. Environ., 39, 6957–6975, <ext-link xlink:href="https://doi.org/10.1016/j.atmosenv.2005.04.027" ext-link-type="DOI">10.1016/j.atmosenv.2005.04.027</ext-link>, 2005.</mixed-citation></ref>
      <ref id="bib1.bibx16"><label>Guanter et al.(2021)</label><mixed-citation>Guanter, L., Irakulis-Loitxate, I., Gorroño, J., Sánchez-García, E., Cusworth, D. H., Varon, D. J., Cogliati, S., and Colombo, R.: Mapping methane point emissions with the PRISMA spaceborne imaging spectrometer, Remote Sens. Environ., 265, 112671, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2021.112671" ext-link-type="DOI">10.1016/j.rse.2021.112671</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx17"><label>Guanter et al.(2024)</label><mixed-citation>Guanter, L., Roger, J., Sharma, S., Valverde, A., Irakulis-Loitxate, I., Gorroño, J., Zhang, X., Schuit, B. J., Maasakkers, J. D., Aben, I., Groshenry, A., Benoit, A., Peyle, Q., and Zavala-Araiza, D.: Multisatellite data depicts a record-breaking methane leak from a well blowout, Environ. Sci. Tech. Let., 11, 825–830, <ext-link xlink:href="https://doi.org/10.1021/acs.estlett.4c00399" ext-link-type="DOI">10.1021/acs.estlett.4c00399</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx18"><label>Hakkarainen et al.(2025)</label><mixed-citation>Hakkarainen, J., Ialongo, I., Varon, D. J., Kuhlmann, G., and Krol, M. C.: Linear integrated mass enhancement: a method for estimating hotspot emission rates from space-based plume observations, Remote Sens. Environ., 319, 114623, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2025.114623" ext-link-type="DOI">10.1016/j.rse.2025.114623</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx19"><label>He et al.(2016)</label><mixed-citation>He, K., Zhang, X., Ren, S., and Sun, J.: Deep residual learning for image recognition, in: 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), IEEE, Las Vegas, NV, USA, <ext-link xlink:href="https://doi.org/10.1109/CVPR.2016.90" ext-link-type="DOI">10.1109/CVPR.2016.90</ext-link>, 770–778, 2016.</mixed-citation></ref>
      <ref id="bib1.bibx20"><label>Hersbach et al.(2023a)</label><mixed-citation>Hersbach, H., Bell, B., Berrisford, P., Biavati, G., Horányi, A., Muñoz Sabater, J., Nicolas, J., Peubey, C., Radu, R., Rozum, I., Schepers, D., Simmons, A., Soci, C., Dee, D., and Thépaut, J.-N.: ERA5 hourly data on pressure levels from 1940 to present, Copernicus Climate Change Service (C3S) Climate Data Store (CDS) [data set], <ext-link xlink:href="https://doi.org/10.24381/cds.bd0915c6" ext-link-type="DOI">10.24381/cds.bd0915c6</ext-link>, 2023a.</mixed-citation></ref>
      <ref id="bib1.bibx21"><label>Hersbach et al.(2023b)</label><mixed-citation>Hersbach, H., Bell, B., Berrisford, P., Biavati, G., Horányi, A., Muñoz Sabater, J., Nicolas, J., Peubey, C., Radu, R., Rozum, I., Schepers, D., Simmons, A., Soci, C., Dee, D., and Thépaut, J.-N.: ERA5 hourly data on single levels from 1940 to present, Copernicus Climate Change Service (C3S) Climate Data Store (CDS) [data set], <ext-link xlink:href="https://doi.org/10.24381/cds.adbb2d47" ext-link-type="DOI">10.24381/cds.adbb2d47</ext-link>, 2023b.</mixed-citation></ref>
      <ref id="bib1.bibx22"><label>Hu et al.(2018)</label><mixed-citation>Hu, H., Landgraf, J., Detmers, R., Borsdorff, T., Aan De Brugh, J., Aben, I., Butz, A., and Hasekamp, O.: Toward global mapping of methane with TROPOMI: first results and intersatellite comparison to GOSAT, Geophys. Res. Lett., 45, 3682–3689, <ext-link xlink:href="https://doi.org/10.1002/2018GL077259" ext-link-type="DOI">10.1002/2018GL077259</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx23"><label>Intergovernmental Panel On Climate Change (IPCC)(2023)</label><mixed-citation>Intergovernmental Panel On Climate Change (IPCC): Climate Change 2021 – The Physical Science Basis: Working Group I Contribution to the Sixth Assessment Report of the Intergovernmental Panel on Climate Change, Cambridge University Press, 1st edn., <ext-link xlink:href="https://doi.org/10.1017/9781009157896" ext-link-type="DOI">10.1017/9781009157896</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx24"><label>International Methane Emissions Observatory(2026)</label><mixed-citation>International Methane Emissions Observatory: Eye on Methane data platform <inline-formula><mml:math id="M140" display="inline"><mml:mo>|</mml:mo></mml:math></inline-formula>  IMEO Eye on Methane, <uri>https://methanedata.unep.org</uri> (last access: 14 August 2026), 2026.</mixed-citation></ref>
      <ref id="bib1.bibx25"><label>Irakulis-Loitxate et al.(2022)</label><mixed-citation>Irakulis-Loitxate, I., Guanter, L., Maasakkers, J. D., Zavala-Araiza, D., and Aben, I.: Satellites detect abatable super-emissions in one of the world's largest methane hotspot regions, Environ. Sci. Technol., 56, 2143–2152, <ext-link xlink:href="https://doi.org/10.1021/acs.est.1c04873" ext-link-type="DOI">10.1021/acs.est.1c04873</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx26"><label>Jacob et al.(2022)</label><mixed-citation>Jacob, D. J., Varon, D. J., Cusworth, D. H., Dennison, P. E., Frankenberg, C., Gautam, R., Guanter, L., Kelley, J., McKeever, J., Ott, L. E., Poulter, B., Qu, Z., Thorpe, A. K., Worden, J. R., and Duren, R. M.: Quantifying methane emissions from the global scale down to point sources using satellite observations of atmospheric methane, Atmos. Chem. Phys., 22, 9617–9646, <ext-link xlink:href="https://doi.org/10.5194/acp-22-9617-2022" ext-link-type="DOI">10.5194/acp-22-9617-2022</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx27"><label>Jervis et al.(2021)</label><mixed-citation>Jervis, D., McKeever, J., Durak, B. O. A., Sloan, J. J., Gains, D., Varon, D. J., Ramier, A., Strupler, M., and Tarrant, E.: The GHGSat-D imaging spectrometer, Atmos. Meas. Tech., 14, 2127–2140, <ext-link xlink:href="https://doi.org/10.5194/amt-14-2127-2021" ext-link-type="DOI">10.5194/amt-14-2127-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx28"><label>Jongaramrungruang et al.(2022)</label><mixed-citation>Jongaramrungruang, S., Thorpe, A. K., Matheou, G., and Frankenberg, C.: MethaNet – an AI-driven approach to quantifying methane point-source emission from high-resolution 2-D plume imagery, Remote Sens. Environ., 269, 112809, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2021.112809" ext-link-type="DOI">10.1016/j.rse.2021.112809</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx29"><label>Joyce et al.(2023)</label><mixed-citation>Joyce, P., Ruiz Villena, C., Huang, Y., Webb, A., Gloor, M., Wagner, F. H., Chipperfield, M. P., Barrio Guilló, R., Wilson, C., and Boesch, H.: Using a deep neural network to detect methane point sources and quantify emissions from PRISMA hyperspectral satellite images, Atmos. Meas. Tech., 16, 2627–2640, <ext-link xlink:href="https://doi.org/10.5194/amt-16-2627-2023" ext-link-type="DOI">10.5194/amt-16-2627-2023</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx30"><label>Krings et al.(2011)</label><mixed-citation>Krings, T., Gerilowski, K., Buchwitz, M., Reuter, M., Tretner, A., Erzinger, J., Heinze, D., Pflüger, U., Burrows, J. P., and Bovensmann, H.: MAMAP – a new spectrometer system for column-averaged methane and carbon dioxide observations from aircraft: retrieval algorithm and first inversions for point source emission rates, Atmos. Meas. Tech., 4, 1735–1758, <ext-link xlink:href="https://doi.org/10.5194/amt-4-1735-2011" ext-link-type="DOI">10.5194/amt-4-1735-2011</ext-link>, 2011.</mixed-citation></ref>
      <ref id="bib1.bibx31"><label>Krings et al.(2013)</label><mixed-citation>Krings, T., Gerilowski, K., Buchwitz, M., Hartmann, J., Sachs, T., Erzinger, J., Burrows, J. P., and Bovensmann, H.: Quantification of methane emission rates from coal mine ventilation shafts using airborne remote sensing data, Atmos. Meas. Tech., 6, 151–166, <ext-link xlink:href="https://doi.org/10.5194/amt-6-151-2013" ext-link-type="DOI">10.5194/amt-6-151-2013</ext-link>, 2013.</mixed-citation></ref>
      <ref id="bib1.bibx32"><label>Krizhevsky et al.(2012)</label><mixed-citation> Krizhevsky, A., Sutskever, I., and Hinton, G. E.: Imagenet classification with deep convolutional neural networks, Adv. Neur. In., 25, 1097–1105, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx33"><label>Krizhevsky et al.(2017)</label><mixed-citation>Krizhevsky, A., Sutskever, I., and Hinton, G. E.: ImageNet classification with deep convolutional neural networks, Commun. ACM, 60, 84–90, <ext-link xlink:href="https://doi.org/10.1145/3065386" ext-link-type="DOI">10.1145/3065386</ext-link>, 2017.</mixed-citation></ref>
      <ref id="bib1.bibx34"><label>Lauvaux et al.(2022)</label><mixed-citation>Lauvaux, T., Giron, C., Mazzolini, M., d'Aspremont, A., Duren, R., Cusworth, D., Shindell, D., and Ciais, P.: Global assessment of oil and gas methane ultra-emitters, Science, 375, 557–561, <ext-link xlink:href="https://doi.org/10.1126/science.abj4351" ext-link-type="DOI">10.1126/science.abj4351</ext-link>, 2022.</mixed-citation></ref>
      <ref id="bib1.bibx35"><label>LeCun et al.(1989)</label><mixed-citation>LeCun, Y., Boser, B., Denker, J., Henderson, D., Howard, R., Hubbard, W., and Jackel, L.: Handwritten digit recognition with a back-propagation network, in: Advances in Neural Information Processing Systems, Vol. 2, Morgan-Kaufmann, <uri>https://papers.nips.cc/paper/1989/hash/53c3bce66e43be4f209556518c2fcb54-Abstract.html</uri> (last access: 14 August 2026), 1989.</mixed-citation></ref>
      <ref id="bib1.bibx36"><label>LeCun et al.(2010)</label><mixed-citation>LeCun, Y., Kavukcuoglu, K., and Farabet, C.: Convolutional networks and applications in vision, in: Proceedings of 2010 IEEE International Symposium on Circuits and Systems, <ext-link xlink:href="https://doi.org/10.1109/ISCAS.2010.5537907" ext-link-type="DOI">10.1109/ISCAS.2010.5537907</ext-link>, 253–256, 2010.</mixed-citation></ref>
      <ref id="bib1.bibx37"><label>Lorente et al.(2021)</label><mixed-citation>Lorente, A., Borsdorff, T., Butz, A., Hasekamp, O., aan de Brugh, J., Schneider, A., Wu, L., Hase, F., Kivi, R., Wunch, D., Pollard, D. F., Shiomi, K., Deutscher, N. M., Velazco, V. A., Roehl, C. M., Wennberg, P. O., Warneke, T., and Landgraf, J.: Methane retrieved from TROPOMI: improvement of the data product and validation of the first 2 years of measurements, Atmos. Meas. Tech., 14, 665–684, <ext-link xlink:href="https://doi.org/10.5194/amt-14-665-2021" ext-link-type="DOI">10.5194/amt-14-665-2021</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx38"><label>Lorente et al.(2023)</label><mixed-citation>Lorente, A., Borsdorff, T., Martinez-Velarte, M. C., and Landgraf, J.: Accounting for surface reflectance spectral features in TROPOMI methane retrievals, Atmos. Meas. Tech., 16, 1597–1608, <ext-link xlink:href="https://doi.org/10.5194/amt-16-1597-2023" ext-link-type="DOI">10.5194/amt-16-1597-2023</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx39"><label>Maasakkers et al.(2022a)</label><mixed-citation>Maasakkers, J. D., Omara, M., Gautam, R., Lorente, A., Pandey, S., Tol, P., Borsdorff, T., Houweling, S., and Aben, I.: Reconstructing and quantifying methane emissions from the full duration of a 38-day natural gas well blowout using space-based observations, Remote Sens. Environ., 270, 112755, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2021.112755" ext-link-type="DOI">10.1016/j.rse.2021.112755</ext-link>, 2022a.</mixed-citation></ref>
      <ref id="bib1.bibx40"><label>Maasakkers et al.(2022b)</label><mixed-citation>Maasakkers, J. D., Varon, D. J., Elfarsdóttir, A., McKeever, J., Jervis, D., Mahapatra, G., Pandey, S., Lorente, A., Borsdorff, T., Foorthuis, L. R., Schuit, B. J., Tol, P., Van Kempen, T. A., Van Hees, R., and Aben, I.: Using satellites to uncover large methane emissions from landfills, Science Advances, 8, eabn9683, <ext-link xlink:href="https://doi.org/10.1126/sciadv.abn9683" ext-link-type="DOI">10.1126/sciadv.abn9683</ext-link>, 2022b.</mixed-citation></ref>
      <ref id="bib1.bibx41"><label>Molod et al.(2012)</label><mixed-citation> Molod, A., Takacs, L., Suarez, M., Bacmeister, J., Song, I.-S., and Eichmann, A.: The GEOS-5 atmospheric general circulation model: Mean climate and development from MERRA to Fortuna, Tech. rep., Vol. 28, NASA/TM-2012-104606, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx42"><label>NCEP(2000)</label><mixed-citation>NCEP: NCEP FNL Operational Model Global Tropospheric Analyses, continuing from July 1999, NOAA National Centers for Environmental Prediction [data set], <ext-link xlink:href="https://doi.org/10.5065/D6M043C6" ext-link-type="DOI">10.5065/D6M043C6</ext-link>, 2000.</mixed-citation></ref>
      <ref id="bib1.bibx43"><label>Ocko et al.(2021)</label><mixed-citation>Ocko, I. B., Sun, T., Shindell, D., Oppenheimer, M., Hristov, A. N., Pacala, S. W., Mauzerall, D. L., Xu, Y., and Hamburg, S. P.: Acting rapidly to deploy readily available methane mitigation measures by sector can immediately slow global warming, Environ. Res. Lett., 16, 054042, <ext-link xlink:href="https://doi.org/10.1088/1748-9326/abf9c8" ext-link-type="DOI">10.1088/1748-9326/abf9c8</ext-link>, 2021.</mixed-citation></ref>
      <ref id="bib1.bibx44"><label>O'Malley et al.(2019)</label><mixed-citation>O'Malley, T., Bursztein, E., Long, J., Chollet, F., Jin, H., Invernizzi, L., and others: KerasTuner, GitHub [code], <uri>https://github.com/keras-team/keras-tuner</uri> (last access: 14 August 2026), 2019.</mixed-citation></ref>
      <ref id="bib1.bibx45"><label>Pandey et al.(2019)</label><mixed-citation>Pandey, S., Gautam, R., Houweling, S., Van Der Gon, H. D., Sadavarte, P., Borsdorff, T., Hasekamp, O., Landgraf, J., Tol, P., Van Kempen, T., Hoogeveen, R., Van Hees, R., Hamburg, S. P., Maasakkers, J. D., and Aben, I.: Satellite observations reveal extreme methane leakage from a natural gas well blowout, P. Natl. Acad. Sci. USA, 116, 26376–26381, <ext-link xlink:href="https://doi.org/10.1073/pnas.1908712116" ext-link-type="DOI">10.1073/pnas.1908712116</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx46"><label>Pandey et al.(2023)</label><mixed-citation>Pandey, S., van Nistelrooij, M., Maasakkers, J. D., Sutar, P., Houweling, S., Varon, D. J., Tol, P., Gains, D., Worden, J., and Aben, I.: Daily detection and quantification of methane leaks using Sentinel-3: a tiered satellite observation approach with Sentinel-2 and Sentinel-5p, Remote Sens. Environ., 296, 113716, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2023.113716" ext-link-type="DOI">10.1016/j.rse.2023.113716</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx47"><label>Plewa et al.(2025)</label><mixed-citation>Plewa, T., Butz, A., Frankenberg, C., Thorpe, A. K., and Marshall, J.: Improvements of AI-driven emission estimation for point sources applied to high resolution 2-D methane-plume imagery, Remote Sens. Environ., 331, 115002, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2025.115002" ext-link-type="DOI">10.1016/j.rse.2025.115002</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx48"><label>Radman et al.(2023)</label><mixed-citation>Radman, A., Mahdianpari, M., Varon, D. J., and Mohammadimanesh, F.: S2MetNet: A novel dataset and deep learning benchmark for methane point source quantification using Sentinel-2 satellite imagery, Remote Sens. Environ., 295, 113708, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2023.113708" ext-link-type="DOI">10.1016/j.rse.2023.113708</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx49"><label>Roberts(2026)</label><mixed-citation>Roberts, C.: ML-SPERE Codebase for TROPOMI Methane Plume Emission Rate Estimation (v1.0.4, Publication Release, for use within SRON), Zenodo [software], <ext-link xlink:href="https://doi.org/10.5281/zenodo.21934190" ext-link-type="DOI">10.5281/zenodo.21934190</ext-link>, 2026.</mixed-citation></ref>
      <ref id="bib1.bibx50"><label>Roberts et al.(2026a)</label><mixed-citation>Roberts, C., de Jong, T., Maasakkers, J., and Huegens, T.: Simulated HYSPLIT Plume Output Resampled to TROPOMI Pixel Footprints, Zenodo [data set], <ext-link xlink:href="https://doi.org/10.5281/zenodo.21792154" ext-link-type="DOI">10.5281/zenodo.21792154</ext-link>, 2026a.</mixed-citation></ref>
      <ref id="bib1.bibx51"><label>Roberts et al.(2026b)</label><mixed-citation>Roberts, C., Maasakkers, J., and Schuit, B.: ML-SPERE Trained CNN Model File and Channel-Wise Standardisation Parameters for TROPOMI Methane Super-Emitter Emission Estimation, Zenodo [code], <ext-link xlink:href="https://doi.org/10.5281/zenodo.21786440" ext-link-type="DOI">10.5281/zenodo.21786440</ext-link>, 2026b.</mixed-citation></ref>
      <ref id="bib1.bibx52"><label>Roberts et al.(2026c)</label><mixed-citation>Roberts, C., van den Berg, A.-W., de Jong, T., Houweling, S., and Maasakkers, J. D.: Simulated WRF-Chem Plume Output Resampled to TROPOMI Pixel Footprints, Zenodo [data set], <ext-link xlink:href="https://doi.org/10.5281/zenodo.21808941" ext-link-type="DOI">10.5281/zenodo.21808941</ext-link>, 2026c.</mixed-citation></ref>
      <ref id="bib1.bibx53"><label>Roger et al.(2024)</label><mixed-citation>Roger, J., Irakulis-Loitxate, I., Valverde, A., Gorroño, J., Chabrillat, S., Brell, M., and Guanter, L.: High-resolution methane mapping with the EnMAP satellite imaging spectroscopy mission, IEEE T. Geosci. Remote, 62, 1–12, <ext-link xlink:href="https://doi.org/10.1109/TGRS.2024.3352403" ext-link-type="DOI">10.1109/TGRS.2024.3352403</ext-link>, 2024.</mixed-citation></ref>
      <ref id="bib1.bibx54"><label>Schuit et al.(2023a)</label><mixed-citation>Schuit, B. J., Maasakkers, J. D., Bijl, P., Mahapatra, G., van den Berg, A.-W., Pandey, S., Lorente, A., Borsdorff, T., Houweling, S., Varon, D. J., McKeever, J., Jervis, D., Girard, M., Irakulis-Loitxate, I., Gorroño, J., Guanter, L., Cusworth, D. H., and Aben, I.: Automated detection and monitoring of methane super-emitters using satellite data, Atmos. Chem. Phys., 23, 9071–9098, <ext-link xlink:href="https://doi.org/10.5194/acp-23-9071-2023" ext-link-type="DOI">10.5194/acp-23-9071-2023</ext-link>, 2023a.</mixed-citation></ref>
      <ref id="bib1.bibx55"><label>Schuit et al.(2023b)</label><mixed-citation>Schuit, B. J., Maasakkers, J. D., Bijl, P., Mahapatra, G., Van den Berg, A.-W., Pandey, S., Lorente, A., Borsdorff, T., Houweling, S., Varon, D. J., McKeever, J., Jervis, D., Girard, M., Irakulis-Loitxate, I., Gorroño, J., Guanter, L., Cusworth, D. H., and Aben, I.: Dataset: all TROPOMI detected plumes for 2021. [Schuit et al. 2023: Automated detection and monitoring of methane super-emitters using satellite data], Zenodo [data set], <ext-link xlink:href="https://doi.org/10.5281/zenodo.8087133" ext-link-type="DOI">10.5281/zenodo.8087133</ext-link>, 2023b.</mixed-citation></ref>
      <ref id="bib1.bibx56"><label>Shorten and Khoshgoftaar(2019)</label><mixed-citation>Shorten, C. and Khoshgoftaar, T. M.: A survey on image data augmentation for deep learning, Journal of Big Data, 6, 60, <ext-link xlink:href="https://doi.org/10.1186/s40537-019-0197-0" ext-link-type="DOI">10.1186/s40537-019-0197-0</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx57"><label>Skamarock et al.(2019)</label><mixed-citation>Skamarock, W. C., Klemp, J., Dudhia, J., Gill, D., Liu, Z., Berner, J., Wang, W., Powers, J., Duda, M., Barker, D., and others: A description of the advanced research WRF model version 4 (Vol. 145), National Center for Atmospheric Research, <ext-link xlink:href="https://doi.org/10.5065/1dfh-6p97" ext-link-type="DOI">10.5065/1dfh-6p97</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx58"><label>Stein et al.(2015)</label><mixed-citation>Stein, A. F., Draxler, R. R., Rolph, G. D., Stunder, B. J. B., Cohen, M. D., and Ngan, F.: NOAA's HYSPLIT atmospheric transport and dispersion modeling system, B. Am. Meteorol. Soc., 96, 2059–2077, <ext-link xlink:href="https://doi.org/10.1175/BAMS-D-14-00110.1" ext-link-type="DOI">10.1175/BAMS-D-14-00110.1</ext-link>, 2015.</mixed-citation></ref>
      <ref id="bib1.bibx59"><label>Thorpe et al.(2023)</label><mixed-citation>Thorpe, A. K., Green, R. O., Thompson, D. R., Brodrick, P. G., Chapman, J. W., Elder, C. D., Irakulis-Loitxate, I., Cusworth, D. H., Ayasse, A. K., Duren, R. M., Frankenberg, C., Guanter, L., Worden, J. R., Dennison, P. E., Roberts, D. A., Chadwick, K. D., Eastwood, M. L., Fahlen, J. E., and Miller, C. E.: Attribution of individual methane and carbon dioxide emission sources using EMIT observations from space, Science Advances, 9, eadh2391, <ext-link xlink:href="https://doi.org/10.1126/sciadv.adh2391" ext-link-type="DOI">10.1126/sciadv.adh2391</ext-link>, 2023.</mixed-citation></ref>
      <ref id="bib1.bibx60"><label>Toshev and Szegedy(2014)</label><mixed-citation>Toshev, A. and Szegedy, C.: DeepPose: human pose estimation via deep neural networks, in: 2014 IEEE Conference on Computer Vision and Pattern Recognition, IEEE, Columbus, OH, USA, <ext-link xlink:href="https://doi.org/10.1109/CVPR.2014.214" ext-link-type="DOI">10.1109/CVPR.2014.214</ext-link>, 1653–1660, 2014.</mixed-citation></ref>
      <ref id="bib1.bibx61"><label>United Nations Environment Programme(2025)</label><mixed-citation>United Nations Environment Programme: An Eye on Methane 2025: From Measurement to Momentum, United Nations Environment Programme, <ext-link xlink:href="https://doi.org/10.59117/20.500.11822/48664" ext-link-type="DOI">10.59117/20.500.11822/48664</ext-link>, 2025.</mixed-citation></ref>
      <ref id="bib1.bibx62"><label>Varon et al.(2018)</label><mixed-citation>Varon, D. J., Jacob, D. J., McKeever, J., Jervis, D., Durak, B. O. A., Xia, Y., and Huang, Y.: Quantifying methane point sources from fine-scale satellite observations of atmospheric methane plumes, Atmos. Meas. Tech., 11, 5673–5686, <ext-link xlink:href="https://doi.org/10.5194/amt-11-5673-2018" ext-link-type="DOI">10.5194/amt-11-5673-2018</ext-link>, 2018.</mixed-citation></ref>
      <ref id="bib1.bibx63"><label>Varon et al.(2019)</label><mixed-citation>Varon, D. J., McKeever, J., Jervis, D., Maasakkers, J. D., Pandey, S., Houweling, S., Aben, I., Scarpelli, T., and Jacob, D. J.: Satellite discovery of anomalously large methane point sources from oil/gas production, Geophys. Res. Lett., 46, 13507–13516, <ext-link xlink:href="https://doi.org/10.1029/2019GL083798" ext-link-type="DOI">10.1029/2019GL083798</ext-link>, 2019.</mixed-citation></ref>
      <ref id="bib1.bibx64"><label>Veefkind et al.(2012)</label><mixed-citation>Veefkind, J., Aben, I., McMullan, K., Förster, H., De Vries, J., Otter, G., Claas, J., Eskes, H., De Haan, J., Kleipool, Q., Van Weele, M., Hasekamp, O., Hoogeveen, R., Landgraf, J., Snel, R., Tol, P., Ingmann, P., Voors, R., Kruizinga, B., Vink, R., Visser, H., and Levelt, P.: TROPOMI on the ESA Sentinel-5 precursor: a GMES mission for global observations of the atmospheric composition for climate, air quality and ozone layer applications, Remote Sens. Environ., 120, 70–83, <ext-link xlink:href="https://doi.org/10.1016/j.rse.2011.09.027" ext-link-type="DOI">10.1016/j.rse.2011.09.027</ext-link>, 2012.</mixed-citation></ref>
      <ref id="bib1.bibx65"><label>Zavala-Araiza et al.(2017)</label><mixed-citation>Zavala-Araiza, D., Alvarez, R. A., Lyon, D. R., Allen, D. T., Marchese, A. J., Zimmerle, D. J., and Hamburg, S. P.: Super-emitters in natural gas infrastructure are caused by abnormal process conditions, Nat. Commun., 8, 14012, <ext-link xlink:href="https://doi.org/10.1038/ncomms14012" ext-link-type="DOI">10.1038/ncomms14012</ext-link>, 2017.</mixed-citation></ref>

  </ref-list></back>
    <!--<article-title-html>Machine learning-based emission rate estimates of global  methane super-emissions</article-title-html>
<abstract-html/>
<ref-html id="bib1.bib1"><label>Abadi et al.(2015)</label><mixed-citation>
       Abadi, M.,
Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S.,
Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A.,
Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M.,
Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C.,
Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P.,
Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P.,
Wattenberg, M., Wicke, M., Yu, Y., and Zheng, X.: TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems, <a href="https://www.tensorflow.org/" target="_blank"/> (last access: 14 August 2026), 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib2"><label>Bruno et al.(2024)</label><mixed-citation>
       Bruno, J. H., Jervis, D., Varon, D. J., and Jacob, D. J.: U-Plume: automated algorithm for plume detection and source quantification by satellite point-source imagers, Atmos. Meas. Tech., 17, 2625–2636, <a href="https://doi.org/10.5194/amt-17-2625-2024" target="_blank">https://doi.org/10.5194/amt-17-2625-2024</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib3"><label>Copernicus Atmospheric Monitoring Service(2025)</label><mixed-citation>
       Copernicus Atmospheric Monitoring Service: CAMS Methane Hotspot Explorer, <a href="https://atmosphere.copernicus.eu/ghg-services/cams-methane-hotspot-explorer" target="_blank"/> (last access: 14 August 2026), 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib4"><label>David and Johnson(1948)</label><mixed-citation>
       David, F. N. and Johnson, N. L.: The probability integral transformation when parameters are estimated from the sample, Biometrika, 35, 182, <a href="https://doi.org/10.2307/2332638" target="_blank">https://doi.org/10.2307/2332638</a>, 1948.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib5"><label>De Jong et al.(2025)</label><mixed-citation>
       De Jong, T. A., Maasakkers, J. D., Irakulis Loitxate, I., Randles, C. A., Tol, P., and Aben, I.: Daily global methane super emitter detection and source identification with sub daily tracking, Geophys. Res. Lett., 52, e2024GL111824, <a href="https://doi.org/10.1029/2024GL111824" target="_blank">https://doi.org/10.1029/2024GL111824</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib6"><label>De Myttenaere et al.(2016)</label><mixed-citation>
       De Myttenaere, A., Golden, B., Le Grand, B., and Rossi, F.: Mean absolute percentage error for regression models, Neurocomputing, 192, 38–48, <a href="https://doi.org/10.1016/j.neucom.2015.12.114" target="_blank">https://doi.org/10.1016/j.neucom.2015.12.114</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib7"><label>Draxler(1999)</label><mixed-citation>
       Draxler, R. R.: HYSPLIT4 Users's Guide, NOAA Tech. Memo. ERL ARL-230, 1999.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib8"><label>Draxler and Hess(1997)</label><mixed-citation>
       Draxler, R. R. and Hess, G. D.: Description of the HYSPLIT4 Modeling System, NOAA Tech. Memo. ERL ARL-224, 1997.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib9"><label>Draxler and Hess(1998)</label><mixed-citation>
       Draxler, R. R. and Hess, G. D.: An overview of the HYSPLIT_4 modelling system for trajectories, Aust. Meteorol. Mag., 47, 295–308, 1998.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib10"><label>European Union(2026)</label><mixed-citation>
       European Union: Harmonized Sentinel-2 MSI: MultiSpectral Instrument, Level-2A (SR), accessed via Google Earth Engine: COPERNICUS/S2_SR_HARMONIZED, <a href="https://doi.org/10.5270/S2_-znk9xsj" target="_blank">https://doi.org/10.5270/S2_-znk9xsj</a>, 2026.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib11"><label>Frankenberg et al.(2016)</label><mixed-citation>
       Frankenberg, C., Thorpe, A. K., Thompson, D. R., Hulley, G., Kort, E. A., Vance, N., Borchardt, J., Krings, T., Gerilowski, K., Sweeney, C., Conley, S., Bue, B. D., Aubrey, A. D., Hook, S., and Green, R. O.: Airborne methane remote measurements reveal heavy-tail flux distribution in Four Corners region, P. Natl. Acad. Sci. USA, 113, 9734–9739, <a href="https://doi.org/10.1073/pnas.1605617113" target="_blank">https://doi.org/10.1073/pnas.1605617113</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib12"><label>Gal and Ghahramani(2016)</label><mixed-citation>
       Gal, Y. and Ghahramani, Z.: Dropout as a Bayesian approximation: representing model uncertainty in deep learning, in: International Conference on Machine Learning, 1050–1059, PMLR, <a href="https://doi.org/10.48550/arXiv.1506.02142" target="_blank">https://doi.org/10.48550/arXiv.1506.02142</a>, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib13"><label>Gao et al.(2023)</label><mixed-citation>
       Gao, S., Liu, X., Chen, Y., Jiang, J., Liu, Y., and Jiang, Y.: Atmospheric turbulence strength estimation using convolution neural network, IEEE Photonics J., 15, 1–7, <a href="https://doi.org/10.1109/JPHOT.2023.3314833" target="_blank">https://doi.org/10.1109/JPHOT.2023.3314833</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib14"><label>Gorelick et al.(2017)</label><mixed-citation>
       Gorelick, N., Hancher, M., Dixon, M., Ilyushchenko, S., Thau, D., and Moore, R.: Google Earth Engine: planetary-scale geospatial analysis for everyone, Remote Sens. Environ., 202, 18–27, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib15"><label>Grell et al.(2005)</label><mixed-citation>
       Grell, G. A., Peckham, S. E., Schmitz, R., McKeen, S. A., Frost, G., Skamarock, W. C., and Eder, B.: Fully coupled “online” chemistry within the WRF model, Atmos. Environ., 39, 6957–6975, <a href="https://doi.org/10.1016/j.atmosenv.2005.04.027" target="_blank">https://doi.org/10.1016/j.atmosenv.2005.04.027</a>, 2005.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib16"><label>Guanter et al.(2021)</label><mixed-citation>
       Guanter, L., Irakulis-Loitxate, I., Gorroño, J., Sánchez-García, E., Cusworth, D. H., Varon, D. J., Cogliati, S., and Colombo, R.: Mapping methane point emissions with the PRISMA spaceborne imaging spectrometer, Remote Sens. Environ., 265, 112671, <a href="https://doi.org/10.1016/j.rse.2021.112671" target="_blank">https://doi.org/10.1016/j.rse.2021.112671</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib17"><label>Guanter et al.(2024)</label><mixed-citation>
       Guanter, L., Roger, J., Sharma, S., Valverde, A., Irakulis-Loitxate, I., Gorroño, J., Zhang, X., Schuit, B. J., Maasakkers, J. D., Aben, I., Groshenry, A., Benoit, A., Peyle, Q., and Zavala-Araiza, D.: Multisatellite data depicts a record-breaking methane leak from a well blowout, Environ. Sci. Tech. Let., 11, 825–830, <a href="https://doi.org/10.1021/acs.estlett.4c00399" target="_blank">https://doi.org/10.1021/acs.estlett.4c00399</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib18"><label>Hakkarainen et al.(2025)</label><mixed-citation>
       Hakkarainen, J., Ialongo, I., Varon, D. J., Kuhlmann, G., and Krol, M. C.: Linear integrated mass enhancement: a method for estimating hotspot emission rates from space-based plume observations, Remote Sens. Environ., 319, 114623, <a href="https://doi.org/10.1016/j.rse.2025.114623" target="_blank">https://doi.org/10.1016/j.rse.2025.114623</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib19"><label>He et al.(2016)</label><mixed-citation>
       He, K., Zhang, X., Ren, S., and Sun, J.: Deep residual learning for image recognition, in: 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), IEEE, Las Vegas, NV, USA, <a href="https://doi.org/10.1109/CVPR.2016.90" target="_blank">https://doi.org/10.1109/CVPR.2016.90</a>, 770–778, 2016.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib20"><label>Hersbach et al.(2023a)</label><mixed-citation>
      
Hersbach, H., Bell, B., Berrisford, P., Biavati, G., Horányi, A., Muñoz Sabater, J., Nicolas, J., Peubey, C., Radu, R., Rozum, I., Schepers, D., Simmons, A., Soci, C., Dee, D., and Thépaut, J.-N.: ERA5 hourly data on pressure levels from 1940 to present, Copernicus Climate Change Service (C3S) Climate Data Store (CDS) [data set], <a href="https://doi.org/10.24381/cds.bd0915c6" target="_blank">https://doi.org/10.24381/cds.bd0915c6</a>, 2023a.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib21"><label>Hersbach et al.(2023b)</label><mixed-citation>
      
Hersbach, H., Bell, B., Berrisford, P., Biavati, G., Horányi, A., Muñoz Sabater, J., Nicolas, J., Peubey, C., Radu, R., Rozum, I., Schepers, D., Simmons, A., Soci, C., Dee, D., and Thépaut, J.-N.: ERA5 hourly data on single levels from 1940 to present, Copernicus Climate Change Service (C3S) Climate Data Store (CDS) [data set], <a href="https://doi.org/10.24381/cds.adbb2d47" target="_blank">https://doi.org/10.24381/cds.adbb2d47</a>, 2023b.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib22"><label>Hu et al.(2018)</label><mixed-citation>
       Hu, H., Landgraf, J., Detmers, R., Borsdorff, T., Aan De Brugh, J., Aben, I., Butz, A., and Hasekamp, O.: Toward global mapping of methane with TROPOMI: first results and intersatellite comparison to GOSAT, Geophys. Res. Lett., 45, 3682–3689, <a href="https://doi.org/10.1002/2018GL077259" target="_blank">https://doi.org/10.1002/2018GL077259</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib23"><label>Intergovernmental Panel On Climate Change (IPCC)(2023)</label><mixed-citation>
       Intergovernmental Panel On Climate Change (IPCC): Climate Change 2021 – The Physical Science Basis: Working Group I Contribution to the Sixth Assessment Report of the Intergovernmental Panel on Climate Change, Cambridge University Press, 1st edn., <a href="https://doi.org/10.1017/9781009157896" target="_blank">https://doi.org/10.1017/9781009157896</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib24"><label>International Methane Emissions Observatory(2026)</label><mixed-citation>
      
International Methane Emissions Observatory: Eye on Methane data platform |  IMEO Eye on Methane, <a href="https://methanedata.unep.org" target="_blank"/> (last access: 14 August 2026), 2026.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib25"><label>Irakulis-Loitxate et al.(2022)</label><mixed-citation>
       Irakulis-Loitxate, I., Guanter, L., Maasakkers, J. D., Zavala-Araiza, D., and Aben, I.: Satellites detect abatable super-emissions in one of the world's largest methane hotspot regions, Environ. Sci. Technol., 56, 2143–2152, <a href="https://doi.org/10.1021/acs.est.1c04873" target="_blank">https://doi.org/10.1021/acs.est.1c04873</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib26"><label>Jacob et al.(2022)</label><mixed-citation>
       Jacob, D. J., Varon, D. J., Cusworth, D. H., Dennison, P. E., Frankenberg, C., Gautam, R., Guanter, L., Kelley, J., McKeever, J., Ott, L. E., Poulter, B., Qu, Z., Thorpe, A. K., Worden, J. R., and Duren, R. M.: Quantifying methane emissions from the global scale down to point sources using satellite observations of atmospheric methane, Atmos. Chem. Phys., 22, 9617–9646, <a href="https://doi.org/10.5194/acp-22-9617-2022" target="_blank">https://doi.org/10.5194/acp-22-9617-2022</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib27"><label>Jervis et al.(2021)</label><mixed-citation>
       Jervis, D., McKeever, J., Durak, B. O. A., Sloan, J. J., Gains, D., Varon, D. J., Ramier, A., Strupler, M., and Tarrant, E.: The GHGSat-D imaging spectrometer, Atmos. Meas. Tech., 14, 2127–2140, <a href="https://doi.org/10.5194/amt-14-2127-2021" target="_blank">https://doi.org/10.5194/amt-14-2127-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib28"><label>Jongaramrungruang et al.(2022)</label><mixed-citation>
       Jongaramrungruang, S., Thorpe, A. K., Matheou, G., and Frankenberg, C.: MethaNet – an AI-driven approach to quantifying methane point-source emission from high-resolution 2-D plume imagery, Remote Sens. Environ., 269, 112809, <a href="https://doi.org/10.1016/j.rse.2021.112809" target="_blank">https://doi.org/10.1016/j.rse.2021.112809</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib29"><label>Joyce et al.(2023)</label><mixed-citation>
       Joyce, P., Ruiz Villena, C., Huang, Y., Webb, A., Gloor, M., Wagner, F. H., Chipperfield, M. P., Barrio Guilló, R., Wilson, C., and Boesch, H.: Using a deep neural network to detect methane point sources and quantify emissions from PRISMA hyperspectral satellite images, Atmos. Meas. Tech., 16, 2627–2640, <a href="https://doi.org/10.5194/amt-16-2627-2023" target="_blank">https://doi.org/10.5194/amt-16-2627-2023</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib30"><label>Krings et al.(2011)</label><mixed-citation>
       Krings, T., Gerilowski, K., Buchwitz, M., Reuter, M., Tretner, A., Erzinger, J., Heinze, D., Pflüger, U., Burrows, J. P., and Bovensmann, H.: MAMAP – a new spectrometer system for column-averaged methane and carbon dioxide observations from aircraft: retrieval algorithm and first inversions for point source emission rates, Atmos. Meas. Tech., 4, 1735–1758, <a href="https://doi.org/10.5194/amt-4-1735-2011" target="_blank">https://doi.org/10.5194/amt-4-1735-2011</a>, 2011.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib31"><label>Krings et al.(2013)</label><mixed-citation>
       Krings, T., Gerilowski, K., Buchwitz, M., Hartmann, J., Sachs, T., Erzinger, J., Burrows, J. P., and Bovensmann, H.: Quantification of methane emission rates from coal mine ventilation shafts using airborne remote sensing data, Atmos. Meas. Tech., 6, 151–166, <a href="https://doi.org/10.5194/amt-6-151-2013" target="_blank">https://doi.org/10.5194/amt-6-151-2013</a>, 2013.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib32"><label>Krizhevsky et al.(2012)</label><mixed-citation>
       Krizhevsky, A., Sutskever, I., and Hinton, G. E.: Imagenet classification with deep convolutional neural networks, Adv. Neur. In., 25, 1097–1105, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib33"><label>Krizhevsky et al.(2017)</label><mixed-citation>
       Krizhevsky, A., Sutskever, I., and Hinton, G. E.: ImageNet classification with deep convolutional neural networks, Commun. ACM, 60, 84–90, <a href="https://doi.org/10.1145/3065386" target="_blank">https://doi.org/10.1145/3065386</a>, 2017.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib34"><label>Lauvaux et al.(2022)</label><mixed-citation>
       Lauvaux, T., Giron, C., Mazzolini, M., d'Aspremont, A., Duren, R., Cusworth, D., Shindell, D., and Ciais, P.: Global assessment of oil and gas methane ultra-emitters, Science, 375, 557–561, <a href="https://doi.org/10.1126/science.abj4351" target="_blank">https://doi.org/10.1126/science.abj4351</a>, 2022.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib35"><label>LeCun et al.(1989)</label><mixed-citation>
       LeCun, Y., Boser, B., Denker, J., Henderson, D., Howard, R., Hubbard, W., and Jackel, L.: Handwritten digit recognition with a back-propagation network, in: Advances in Neural Information Processing Systems, Vol. 2, Morgan-Kaufmann, <a href="https://papers.nips.cc/paper/1989/hash/53c3bce66e43be4f209556518c2fcb54-Abstract.html" target="_blank"/> (last access: 14 August 2026), 1989.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib36"><label>LeCun et al.(2010)</label><mixed-citation>
       LeCun, Y., Kavukcuoglu, K., and Farabet, C.: Convolutional networks and applications in vision, in: Proceedings of 2010 IEEE International Symposium on Circuits and Systems, <a href="https://doi.org/10.1109/ISCAS.2010.5537907" target="_blank">https://doi.org/10.1109/ISCAS.2010.5537907</a>, 253–256, 2010.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib37"><label>Lorente et al.(2021)</label><mixed-citation>
       Lorente, A., Borsdorff, T., Butz, A., Hasekamp, O., aan de Brugh, J., Schneider, A., Wu, L., Hase, F., Kivi, R., Wunch, D., Pollard, D. F., Shiomi, K., Deutscher, N. M., Velazco, V. A., Roehl, C. M., Wennberg, P. O., Warneke, T., and Landgraf, J.: Methane retrieved from TROPOMI: improvement of the data product and validation of the first 2 years of measurements, Atmos. Meas. Tech., 14, 665–684, <a href="https://doi.org/10.5194/amt-14-665-2021" target="_blank">https://doi.org/10.5194/amt-14-665-2021</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib38"><label>Lorente et al.(2023)</label><mixed-citation>
       Lorente, A., Borsdorff, T., Martinez-Velarte, M. C., and Landgraf, J.: Accounting for surface reflectance spectral features in TROPOMI methane retrievals, Atmos. Meas. Tech., 16, 1597–1608, <a href="https://doi.org/10.5194/amt-16-1597-2023" target="_blank">https://doi.org/10.5194/amt-16-1597-2023</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib39"><label>Maasakkers et al.(2022a)</label><mixed-citation>
       Maasakkers, J. D., Omara, M., Gautam, R., Lorente, A., Pandey, S., Tol, P., Borsdorff, T., Houweling, S., and Aben, I.: Reconstructing and quantifying methane emissions from the full duration of a 38-day natural gas well blowout using space-based observations, Remote Sens. Environ., 270, 112755, <a href="https://doi.org/10.1016/j.rse.2021.112755" target="_blank">https://doi.org/10.1016/j.rse.2021.112755</a>, 2022a.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib40"><label>Maasakkers et al.(2022b)</label><mixed-citation>
       Maasakkers, J. D., Varon, D. J., Elfarsdóttir, A., McKeever, J., Jervis, D., Mahapatra, G., Pandey, S., Lorente, A., Borsdorff, T., Foorthuis, L. R., Schuit, B. J., Tol, P., Van Kempen, T. A., Van Hees, R., and Aben, I.: Using satellites to uncover large methane emissions from landfills, Science Advances, 8, eabn9683, <a href="https://doi.org/10.1126/sciadv.abn9683" target="_blank">https://doi.org/10.1126/sciadv.abn9683</a>, 2022b.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib41"><label>Molod et al.(2012)</label><mixed-citation>
       Molod, A., Takacs, L., Suarez, M., Bacmeister, J., Song, I.-S., and Eichmann, A.: The GEOS-5 atmospheric general circulation model: Mean climate and development from MERRA to Fortuna, Tech. rep., Vol. 28, NASA/TM-2012-104606, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib42"><label>NCEP(2000)</label><mixed-citation>
       NCEP: NCEP FNL Operational Model Global Tropospheric Analyses, continuing from July 1999, NOAA National Centers for Environmental Prediction [data set], <a href="https://doi.org/10.5065/D6M043C6" target="_blank">https://doi.org/10.5065/D6M043C6</a>, 2000.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib43"><label>Ocko et al.(2021)</label><mixed-citation>
       Ocko, I. B., Sun, T., Shindell, D., Oppenheimer, M., Hristov, A. N., Pacala, S. W., Mauzerall, D. L., Xu, Y., and Hamburg, S. P.: Acting rapidly to deploy readily available methane mitigation measures by sector can immediately slow global warming, Environ. Res. Lett., 16, 054042, <a href="https://doi.org/10.1088/1748-9326/abf9c8" target="_blank">https://doi.org/10.1088/1748-9326/abf9c8</a>, 2021.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib44"><label>O'Malley et al.(2019)</label><mixed-citation>
       O'Malley, T.,
Bursztein, E., Long, J., Chollet, F., Jin, H., Invernizzi, L., and
others: KerasTuner, GitHub [code],
<a href="https://github.com/keras-team/keras-tuner" target="_blank"/> (last access: 14 August 2026), 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib45"><label>Pandey et al.(2019)</label><mixed-citation>
       Pandey, S., Gautam, R., Houweling, S., Van Der Gon, H. D., Sadavarte, P., Borsdorff, T., Hasekamp, O., Landgraf, J., Tol, P., Van Kempen, T., Hoogeveen, R., Van Hees, R., Hamburg, S. P., Maasakkers, J. D., and Aben, I.: Satellite observations reveal extreme methane leakage from a natural gas well blowout, P. Natl. Acad. Sci. USA, 116, 26376–26381, <a href="https://doi.org/10.1073/pnas.1908712116" target="_blank">https://doi.org/10.1073/pnas.1908712116</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib46"><label>Pandey et al.(2023)</label><mixed-citation>
       Pandey, S., van Nistelrooij, M., Maasakkers, J. D., Sutar, P., Houweling, S., Varon, D. J., Tol, P., Gains, D., Worden, J., and Aben, I.: Daily detection and quantification of methane leaks using Sentinel-3: a tiered satellite observation approach with Sentinel-2 and Sentinel-5p, Remote Sens. Environ., 296, 113716, <a href="https://doi.org/10.1016/j.rse.2023.113716" target="_blank">https://doi.org/10.1016/j.rse.2023.113716</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib47"><label>Plewa et al.(2025)</label><mixed-citation>
       Plewa, T., Butz, A., Frankenberg, C., Thorpe, A. K., and Marshall, J.: Improvements of AI-driven emission estimation for point sources applied to high resolution 2-D methane-plume imagery, Remote Sens. Environ., 331, 115002, <a href="https://doi.org/10.1016/j.rse.2025.115002" target="_blank">https://doi.org/10.1016/j.rse.2025.115002</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib48"><label>Radman et al.(2023)</label><mixed-citation>
       Radman, A., Mahdianpari, M., Varon, D. J., and Mohammadimanesh, F.: S2MetNet: A novel dataset and deep learning benchmark for methane point source quantification using Sentinel-2 satellite imagery, Remote Sens. Environ., 295, 113708, <a href="https://doi.org/10.1016/j.rse.2023.113708" target="_blank">https://doi.org/10.1016/j.rse.2023.113708</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib49"><label>Roberts(2026)</label><mixed-citation>
      
Roberts, C.: ML-SPERE Codebase for TROPOMI Methane Plume Emission Rate Estimation (v1.0.4, Publication Release, for use within SRON), Zenodo [software], <a href="https://doi.org/10.5281/zenodo.21934190" target="_blank">https://doi.org/10.5281/zenodo.21934190</a>, 2026.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib50"><label>Roberts et al.(2026a)</label><mixed-citation>
      
Roberts, C., de Jong, T., Maasakkers, J., and Huegens, T.: Simulated HYSPLIT Plume Output Resampled to TROPOMI Pixel Footprints, Zenodo [data set], <a href="https://doi.org/10.5281/zenodo.21792154" target="_blank">https://doi.org/10.5281/zenodo.21792154</a>, 2026a.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib51"><label>Roberts et al.(2026b)</label><mixed-citation>
      
Roberts, C., Maasakkers, J., and Schuit, B.: ML-SPERE Trained CNN Model File and Channel-Wise Standardisation Parameters for TROPOMI Methane Super-Emitter Emission Estimation, Zenodo [code], <a href="https://doi.org/10.5281/zenodo.21786440" target="_blank">https://doi.org/10.5281/zenodo.21786440</a>, 2026b.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib52"><label>Roberts et al.(2026c)</label><mixed-citation>
      
Roberts, C., van den Berg, A.-W., de Jong, T., Houweling, S., and Maasakkers, J. D.: Simulated WRF-Chem Plume Output Resampled to TROPOMI Pixel Footprints, Zenodo [data set], <a href="https://doi.org/10.5281/zenodo.21808941" target="_blank">https://doi.org/10.5281/zenodo.21808941</a>, 2026c.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib53"><label>Roger et al.(2024)</label><mixed-citation>
       Roger, J., Irakulis-Loitxate, I., Valverde, A., Gorroño, J., Chabrillat, S., Brell, M., and Guanter, L.: High-resolution methane mapping with the EnMAP satellite imaging spectroscopy mission, IEEE T. Geosci. Remote, 62, 1–12, <a href="https://doi.org/10.1109/TGRS.2024.3352403" target="_blank">https://doi.org/10.1109/TGRS.2024.3352403</a>, 2024.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib54"><label>Schuit et al.(2023a)</label><mixed-citation>
       Schuit, B. J., Maasakkers, J. D., Bijl, P., Mahapatra, G., van den Berg, A.-W., Pandey, S., Lorente, A., Borsdorff, T., Houweling, S., Varon, D. J., McKeever, J., Jervis, D., Girard, M., Irakulis-Loitxate, I., Gorroño, J., Guanter, L., Cusworth, D. H., and Aben, I.: Automated detection and monitoring of methane super-emitters using satellite data, Atmos. Chem. Phys., 23, 9071–9098, <a href="https://doi.org/10.5194/acp-23-9071-2023" target="_blank">https://doi.org/10.5194/acp-23-9071-2023</a>, 2023a.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib55"><label>Schuit et al.(2023b)</label><mixed-citation>
      
Schuit, B. J., Maasakkers, J. D., Bijl, P., Mahapatra, G., Van den Berg, A.-W., Pandey, S., Lorente, A., Borsdorff, T., Houweling, S., Varon, D. J., McKeever, J., Jervis, D., Girard, M., Irakulis-Loitxate, I., Gorroño, J., Guanter, L., Cusworth, D. H., and Aben, I.: Dataset: all TROPOMI detected plumes for 2021. [Schuit et al. 2023: Automated detection and monitoring of methane super-emitters using satellite data], Zenodo [data set], <a href="https://doi.org/10.5281/zenodo.8087133" target="_blank">https://doi.org/10.5281/zenodo.8087133</a>, 2023b.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib56"><label>Shorten and Khoshgoftaar(2019)</label><mixed-citation>
       Shorten, C. and Khoshgoftaar, T. M.: A survey on image data augmentation for deep learning, Journal of Big Data, 6, 60, <a href="https://doi.org/10.1186/s40537-019-0197-0" target="_blank">https://doi.org/10.1186/s40537-019-0197-0</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib57"><label>Skamarock et al.(2019)</label><mixed-citation>
       Skamarock, W. C., Klemp, J., Dudhia, J., Gill, D., Liu, Z., Berner, J., Wang, W., Powers, J., Duda, M., Barker, D., and others: A description of the advanced research WRF model version 4 (Vol. 145), National Center for Atmospheric Research, <a href="https://doi.org/10.5065/1dfh-6p97" target="_blank">https://doi.org/10.5065/1dfh-6p97</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib58"><label>Stein et al.(2015)</label><mixed-citation>
       Stein, A. F., Draxler, R. R., Rolph, G. D., Stunder, B. J. B., Cohen, M. D., and Ngan, F.: NOAA's HYSPLIT atmospheric transport and dispersion modeling system, B. Am. Meteorol. Soc., 96, 2059–2077, <a href="https://doi.org/10.1175/BAMS-D-14-00110.1" target="_blank">https://doi.org/10.1175/BAMS-D-14-00110.1</a>, 2015.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib59"><label>Thorpe et al.(2023)</label><mixed-citation>
       Thorpe, A. K., Green, R. O., Thompson, D. R., Brodrick, P. G., Chapman, J. W., Elder, C. D., Irakulis-Loitxate, I., Cusworth, D. H., Ayasse, A. K., Duren, R. M., Frankenberg, C., Guanter, L., Worden, J. R., Dennison, P. E., Roberts, D. A., Chadwick, K. D., Eastwood, M. L., Fahlen, J. E., and Miller, C. E.: Attribution of individual methane and carbon dioxide emission sources using EMIT observations from space, Science Advances, 9, eadh2391, <a href="https://doi.org/10.1126/sciadv.adh2391" target="_blank">https://doi.org/10.1126/sciadv.adh2391</a>, 2023.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib60"><label>Toshev and Szegedy(2014)</label><mixed-citation>
       Toshev, A. and Szegedy, C.: DeepPose: human pose estimation via deep neural networks, in: 2014 IEEE Conference on Computer Vision and Pattern Recognition, IEEE, Columbus, OH, USA, <a href="https://doi.org/10.1109/CVPR.2014.214" target="_blank">https://doi.org/10.1109/CVPR.2014.214</a>, 1653–1660, 2014.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib61"><label>United Nations Environment Programme(2025)</label><mixed-citation>
       United Nations Environment Programme: An Eye on Methane 2025: From Measurement to Momentum, United Nations Environment Programme, <a href="https://doi.org/10.59117/20.500.11822/48664" target="_blank">https://doi.org/10.59117/20.500.11822/48664</a>, 2025.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib62"><label>Varon et al.(2018)</label><mixed-citation>
       Varon, D. J., Jacob, D. J., McKeever, J., Jervis, D., Durak, B. O. A., Xia, Y., and Huang, Y.: Quantifying methane point sources from fine-scale satellite observations of atmospheric methane plumes, Atmos. Meas. Tech., 11, 5673–5686, <a href="https://doi.org/10.5194/amt-11-5673-2018" target="_blank">https://doi.org/10.5194/amt-11-5673-2018</a>, 2018.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib63"><label>Varon et al.(2019)</label><mixed-citation>
       Varon, D. J., McKeever, J., Jervis, D., Maasakkers, J. D., Pandey, S., Houweling, S., Aben, I., Scarpelli, T., and Jacob, D. J.: Satellite discovery of anomalously large methane point sources from oil/gas production, Geophys. Res. Lett., 46, 13507–13516, <a href="https://doi.org/10.1029/2019GL083798" target="_blank">https://doi.org/10.1029/2019GL083798</a>, 2019.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib64"><label>Veefkind et al.(2012)</label><mixed-citation>
       Veefkind, J., Aben, I., McMullan, K., Förster, H., De Vries, J., Otter, G., Claas, J., Eskes, H., De Haan, J., Kleipool, Q., Van Weele, M., Hasekamp, O., Hoogeveen, R., Landgraf, J., Snel, R., Tol, P., Ingmann, P., Voors, R., Kruizinga, B., Vink, R., Visser, H., and Levelt, P.: TROPOMI on the ESA Sentinel-5 precursor: a GMES mission for global observations of the atmospheric composition for climate, air quality and ozone layer applications, Remote Sens. Environ., 120, 70–83, <a href="https://doi.org/10.1016/j.rse.2011.09.027" target="_blank">https://doi.org/10.1016/j.rse.2011.09.027</a>, 2012.

    </mixed-citation></ref-html>
<ref-html id="bib1.bib65"><label>Zavala-Araiza et al.(2017)</label><mixed-citation>
       Zavala-Araiza, D., Alvarez, R. A., Lyon, D. R., Allen, D. T., Marchese, A. J., Zimmerle, D. J., and Hamburg, S. P.: Super-emitters in natural gas infrastructure are caused by abnormal process conditions, Nat. Commun., 8, 14012, <a href="https://doi.org/10.1038/ncomms14012" target="_blank">https://doi.org/10.1038/ncomms14012</a>, 2017.

    </mixed-citation></ref-html>--></article>
