Repeated leak detection and repair surveys reduce methane emissions over scale of years
Bibliographic record
Abstract
Abstract Reducing methane emissions from the oil and gas industry is a critical climate action policy tool in Canada and the US. Optical gas imaging-based leak detection and repair (LDAR) surveys are commonly used to address fugitive methane emissions or leaks. Despite widespread use, there is little empirical measurement of the effectiveness of LDAR programs at reducing long-term leakage, especially over the scale of months to years. In this study, we measure the effectiveness of LDAR surveys by quantifying emissions at 36 unconventional liquids-rich natural gas facilities in Alberta, Canada. A representative subset of these 36 facilities were visited twice by the same detection team: an initial survey and a post-repair re-survey occurring ∼0.5–2 years after the initial survey. Overall, total emissions reduced by 44% after one LDAR survey, combining a reduction in fugitive emissions of 22% and vented emissions by 47%. Furthermore, >90% of the leaks found in the initial survey were not emitting in the re-survey, suggesting high repair effectiveness. However, fugitive emissions reduced by only 22% because of new leaks that occurred between the surveys. This indicates a need for frequent, effective, and low-cost LDAR surveys to target new leaks. The large reduction in vent emissions is associated with potentially stochastic changes to tank-related emissions, which contributed ∼45% of all emissions. Our data suggest a key role for tank-specific abatement strategies as an effective way to reduce oil and gas methane emissions. Finally, mitigation policies will also benefit from more definitive classification of leaks and vents.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".