paper-with-me

홈 › Papers

Estimating oil and gas recovery factors via machine learning: Database-dependent accuracy and reliability

2022-10-22 · Alireza Roustazadeh, Behzad Ghanbarian, Mohammad B. Shadmand, Vahid Taslimitehrani, Larry W. Lake

With recent advances in artificial intelligence, machine learning (ML) approaches have become an attractive tool in petroleum engineering, particularly for reservoir characterizations. A key reservoir property is hydrocarbon recovery factor (RF) whose accurate estimation would provide decisive insights to drilling and production strategies. Therefore, this study aims to estimate the hydrocarbon RF for exploration from various reservoir characteristics, such as porosity, permeability, pressure, and water saturation via the ML. We applied three regression-based models including the extreme gradient boosting (XGBoost), support vector machine (SVM), and stepwise multiple linear regression (MLR) and various combinations of three databases to construct ML models and estimate the oil and/or gas RF. Using two databases and the cross-validation method, we evaluated the performance of the ML models. In each iteration 90 and 10% of the data were respectively used to train and test the models. The third independent database was then used to further assess the constructed models. For both oil and gas RFs, we found that the XGBoost model estimated the RF for the train and test datasets more accurately than the SVM and MLR models. However, the performance of all the models were unsatisfactory for the independent databases. Results demonstrated that the ML algorithms were highly dependent and sensitive to the databases based on which they were trained. Statistical tests revealed that such unsatisfactory performances were because the distributions of input features and target variables in the train datasets were significantly different from those in the independent databases (p-value < 0.05).

📄 PDF Abstract BibTeX arXiv:2210.12491

Code (0)

등록된 구현이 없습니다.

Tasks

regression

Methods 이 논문이 사용한 방법론

Test 설명 없음
SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…
Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Estimating oil recovery factor using machine learning: Applications of XGBoost classification

2022-10-28 · Alireza Roustazadeh, Behzad Ghanbarian, Frank Male, Mohammad B. Shadmand 외

In petroleum engineering, it is essential to determine the ultimate recovery factor, RF, particularly before exploitation and exploration. However, accurately estimating requires data that is not necessarily available or…

Feature Importance

On the Difficulty of Selecting Ising Models with Approximate Recovery

2016-02-11 · Jonathan Scarlett, Volkan Cevher

In this paper, we consider the problem of estimating the underlying graph associated with an Ising model given a number of independent and identically distributed samples. We adopt an \emph{approximate recovery} criterio…

On Correlation Detection and Alignment Recovery of Gaussian Databases

2022-11-02 · Ran Tamir

In this work, we propose an efficient two-stage algorithm solving a joint problem of correlation detection and partial alignment recovery between two Gaussian databases. Correlation detection is a hypothesis testing prob…

Iterative Alpha Expansion for estimating gradient-sparse signals from linear measurements

2019-05-15 · Sheng Xu, Zhou Fan

We consider estimating a piecewise-constant image, or a gradient-sparse signal on a general graph, from noisy linear measurements. We propose and study an iterative algorithm to minimize a penalized least-squares objecti…

compressed sensing

Fundamental limits for rank-one matrix estimation with groupwise heteroskedasticity

2021-06-22 · Joshua K. Behne, Galen Reeves

Low-rank matrix recovery problems involving high-dimensional and heterogeneous data appear in applications throughout statistics and machine learning. The contribution of this paper is to establish the fundamental limits…

ClusteringCommunity Detection