Simplified Dark Data Analytics of Everyday Completions Data, and its Relation to Observed Induced Seismicity Events: An Unconventional Big Data Solution
Bibliographic record
Abstract
Abstract Diagnostic fracture injection tests (DFIT's), or "mini-fracs" are often used to gauge many reservoir and fracture design parameters. However, DFITs are not always conducted in conjunction with the main completions work. This paper proposes a novel workflow to determine the instantaneous shut-in pressure (ISIP) from readily available completions data. This is a valuable parameter in itself as related to the least principal in-situ stress states as demonstrated by the stress change relationships near faults in Lavoie et al. (2018). Directly using completions data from fracture stimulation operations, the authors have leveraged on the water-hammer signature in bottom-hole pressure data during completions to process the ISIP for each completions stage. Within this study, completions data from ~2100 stages from ~300 horizontal Montney formation wells were analyzed. A MATLAB script was used to automate the derived ISIP stress trends over the Montney formation and to deduce the ISIP in a consistent format. This novel workflow also validates the expected in-situ stress trends at depth, with a relationship of high ISIP gradients closer to fault zones similar to stress change behaviour as shown in Lavoie et al. (2018). Specifically, a positive spatial relationship was observed pertaining to local ISIP gradients, the lithostatic gradient, the minimum in-situ stress, and the propagation of hydraulic fractures that are prone to reactivation of critically stressed faults. Based on our real-time observations, field operators may allow to flowback a well for a short amount of time to deplete the anthropogenic reservoir pressure and stress shadowing prior to resuming fracture stimulation. Considering the continued push for higher fluid and sand loading in industry in the development of unconventional assets as an economic driver, there also exists a large and tangible corporate citizenship opportunity of mining real time completions dark data with the possibility of relating that live feed as a prescriptive tool to mitigate reactivation of critically stressed faults. This case study focuses on the Montney formation as a basis for processing easily available data from standard operations in an effort of systematically designating areas prone to seismicity risk in future hydraulic fracturing operations based on automated real-time analytics of dark data.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.009 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.003 | 0.003 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.004 | 0.004 |
| Open science | 0.002 | 0.004 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.002 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".