Machine learning-assisted production data analysis in liquid-rich Duvernay Formation

Background: With the advent of data analysis and machine learning, there is a growing impetus of analyzing and generating models on historic data. The data comes in numerous forms and shapes with an abundance of challenges. The most sorted form of data for analysis is the numerical data. With the plethora of algorithms and tools it is quite manageable to deal with such data. Another form of data is of categorical nature, which is subdivided into, ordinal (order wise) and nominal (number wise). This data can be broadly classified as Sequential and Non-Sequential. Sequential data analysis is easier to preprocess using algorithms. Objective: The challenge of applying machine learning algorithms on categorical data of nonsequential nature is dealt in this paper. Methods: Upon implementing several data analysis algorithms on such data, we end up getting a biased result, which makes it impossible to generate a reliable predictive model. In this paper, we will address this problem by walking through a handful of techniques which during our research helped us in dealing with a large categorical data of non-sequential nature. In subsequent sections, we will discuss the possible implementable solutions and shortfalls of these techniques. Results: The methods are applied to sample datasets available in public domain and the results with respect to accuracy of classification are satisfactory. Conclusion: The best pre-processing technique we observed in our research is one hot encoding, which facilitates breaking down the categorical features into binary and feeding it into an Algorithm to predict the outcome. The example that we took is not abstract but it is a real – time production services dataset, which had many complex variations of categorical features. Our Future work includes creating a robust model on such data and deploying it into industry standard applications.

Download Full-text

Measuring Engagement Level in Child-Robot Interaction Using Machine Learning Based Data Analysis

2020 International Conference on Data Analytics for Business and Industry: Way Towards a Sustainable Economy (ICDABI) ◽

10.1109/icdabi51230.2020.9325676 ◽

2020 ◽

Author(s):

George K. Sidiropoulos ◽

George A. Papakostas ◽

Chris Lytridis ◽

Christos Bazinas ◽

Vassilis G. Kaburlasos ◽

...

Keyword(s):

Machine Learning ◽

Data Analysis ◽

Robot Interaction

Download Full-text

Ensemble Machine Learning Assisted Reservoir Characterization Using Field Production Data–An Offshore Field Case Study

Energies ◽

10.3390/en14041052 ◽

2021 ◽

Vol 14 (4) ◽

pp. 1052

Author(s):

Baozhong Wang ◽

Jyotsna Sharma ◽

Jianhua Chen ◽

Patricia Persaud

Keyword(s):

Machine Learning ◽

Random Forest ◽

Reservoir Characterization ◽

Time Lapse ◽

Production Data ◽

Oil Saturation ◽

Ensemble Machine Learning ◽

Input Parameters ◽

Saturation Profiles ◽

Field Production

Estimation of fluid saturation is an important step in dynamic reservoir characterization. Machine learning techniques have been increasingly used in recent years for reservoir saturation prediction workflows. However, most of these studies require input parameters derived from cores, petrophysical logs, or seismic data, which may not always be readily available. Additionally, very few studies incorporate the production data, which is an important reflection of the dynamic reservoir properties and also typically the most frequently and reliably measured quantity throughout the life of a field. In this research, the random forest ensemble machine learning algorithm is implemented that uses the field-wide production and injection data (both measured at the surface) as the only input parameters to predict the time-lapse oil saturation profiles at well locations. The algorithm is optimized using feature selection based on feature importance score and Pearson correlation coefficient, in combination with geophysical domain-knowledge. The workflow is demonstrated using the actual field data from a structurally complex, heterogeneous, and heavily faulted offshore reservoir. The random forest model captures the trends from three and a half years of historical field production, injection, and simulated saturation data to predict future time-lapse oil saturation profiles at four deviated well locations with over 90% R-square, less than 6% Root Mean Square Error, and less than 7% Mean Absolute Percentage Error, in each case.

Download Full-text

Development of production data analysis models for multi-well gas condensate reservoirs

Journal of Petroleum Science and Engineering ◽

10.1016/j.petrol.2021.108552 ◽

2021 ◽

Vol 202 ◽

pp. 108552

Author(s):

Reza Jadidi ◽

Behnam Sedaee ◽

Shahab Gerami ◽

Ali Nakhaee

Keyword(s):

Data Analysis ◽

Gas Condensate ◽

Production Data ◽

Gas Condensate Reservoirs ◽

Analysis Models

Download Full-text

Machine learning based real-time vehicle data analysis for safe driving modeling

Proceedings of the 34th ACM/SIGAPP Symposium on Applied Computing - SAC '19 ◽

10.1145/3297280.3297584 ◽

2019 ◽

Cited By ~ 2

Author(s):

Pamul Yadav ◽

Sangsu Jung ◽

Dhananjay Singh

Keyword(s):

Machine Learning ◽

Data Analysis ◽

Real Time ◽

Safe Driving ◽

Vehicle Data

Download Full-text

Topologic Data Analysis and Machine Learning

JACC Cardiovascular Imaging ◽

10.1016/j.jcmg.2021.04.005 ◽

2021 ◽

Author(s):

Rebecca T. Hahn

Keyword(s):

Machine Learning ◽

Data Analysis

Download Full-text

Passenger data analysis of Titanic using machine learning approach in the context of chances of surviving the disaster

IOP Conference Series Materials Science and Engineering ◽

10.1088/1757-899x/1065/1/012042 ◽

2021 ◽

Vol 1065 (1) ◽

pp. 012042

Author(s):

Md Arfinul Haque ◽

G Shivaprasad ◽

G Guruprasad

Keyword(s):

Machine Learning ◽

Data Analysis ◽

Learning Approach ◽

Machine Learning Approach

Download Full-text

Classification of apatite structures via topological data analysis: a framework for a ‘Materials Barcode’ representation of structure maps

Scientific Reports ◽

10.1038/s41598-021-90070-4 ◽

2021 ◽

Vol 11 (1) ◽

Author(s):

Scott Broderick ◽

Ruhil Dongol ◽

Tianmu Zhang ◽

Krishna Rajan

Keyword(s):

Machine Learning ◽

Data Analysis ◽

Crystal Chemistry ◽

Persistent Homology ◽

Hierarchical Classification ◽

Topological Data Analysis ◽

Learning Tool ◽

Coordination Polyhedra ◽

Machine Learning Tool ◽

Topological Data

AbstractThis paper introduces the use of topological data analysis (TDA) as an unsupervised machine learning tool to uncover classification criteria in complex inorganic crystal chemistries. Using the apatite chemistry as a template, we track through the use of persistent homology the topological connectivity of input crystal chemistry descriptors on defining similarity between different stoichiometries of apatites. It is shown that TDA automatically identifies a hierarchical classification scheme within apatites based on the commonality of the number of discrete coordination polyhedra that constitute the structural building units common among the compounds. This information is presented in the form of a visualization scheme of a barcode of homology classifications, where the persistence of similarity between compounds is tracked. Unlike traditional perspectives of structure maps, this new “Materials Barcode” schema serves as an automated exploratory machine learning tool that can uncover structural associations from crystal chemistry databases, as well as to achieve a more nuanced insight into what defines similarity among homologous compounds.

Download Full-text

Machine learning-assisted production data analysis in liquid-rich Duvernay Formation

Production Data Analysis in Complex Fracture Network Horizontal Wells with SRV Effects

Bipolar Disorder and Oxidative Stress Injury Mechanism - Clinical Big Data Analysis Based on Machine Learning

Machine Learning Based Predictive Action on Categorical Non-Sequential Data

Measuring Engagement Level in Child-Robot Interaction Using Machine Learning Based Data Analysis

Ensemble Machine Learning Assisted Reservoir Characterization Using Field Production Data–An Offshore Field Case Study

Development of production data analysis models for multi-well gas condensate reservoirs

Machine learning based real-time vehicle data analysis for safe driving modeling

Topologic Data Analysis and Machine Learning

Passenger data analysis of Titanic using machine learning approach in the context of chances of surviving the disaster

Classification of apatite structures via topological data analysis: a framework for a ‘Materials Barcode’ representation of structure maps

Export Citation Format