Interpreting Pressure and Flow Rate Data from Permanent Downhole Gauges with Convolution-Kernel-Based Data Mining Approaches
Notice bibliographique
Résumé
Abstract The paper describes a method of analyzing data from permanent downhole gauges, even in the presence of noise, gaps and outliers. The data mining approach developed allows for the revelation of the underlying data signal, and as a useful corollary also achieves deconvolution to compute the reservoir model – even with the existence of significant noise in the data. The convolution kernel was initially invented and applied in the domain of natural language machine learning. In the original linguistic study, the convolution kernel detected the relationship between words by decomposing words into parts, and evaluating the parts using a simple kernel function. The success of the convolution kernel method inspired us to apply it to data from permanent downhole gauges (PDG). In this study, the data mining process was conducted in two stages, namely learning and prediction processes. In the learning process, the PDG data that were decomposed into a series of pressure responses to the previous flow rate change events were used to train the convolution-kernel-based data mining algorithm until the convergence. After convergence, the reservoir model was obtained implicitly in the form of polynomials in the high-dimensional Hilbert space defined by the convolution kernel function. In the prediction process, a pressure prediction was made by the reservoir model (obtained in the learning process) to an arbitrary given flow rate history (usually a constant flow rate history for simplicity). This flow rate history and the corresponding pressure prediction revealed the reservoir model underlying the variable PDG data. In the previous work, a series of synthetic cases and real field cases have been used to test this approach. The method recovered the reservoir model successfully in all cases. In this paper, the method was tested under problematic data situations, including the existence of significant outliers and aberrant segments, incomplete production history, and unknown initial pressure. The results suggested that: 1) the method tolerated a moderate level of outliers and aberrant segments without any preprocessing; 2) the method could reveal the reservoir model with effective rate correction when the production history was incomplete; 3) the method could reveal the reservoir model and discover the appropriate initial pressure by using an optimization on initial pressure value when the initial pressure was unknown.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,005 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,003 | 0,002 |
| Études des sciences et des technologies | 0,000 | 0,001 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».