<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="research-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Earth Sci.</journal-id>
<journal-title>Frontiers in Earth Science</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Earth Sci.</abbrev-journal-title>
<issn pub-type="epub">2296-6463</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1137645</article-id>
<article-id pub-id-type="doi">10.3389/feart.2023.1137645</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Earth Science</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Seismic prediction of porosity in tight reservoirs based on transformer</article-title>
<alt-title alt-title-type="left-running-head">Su et al.</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/feart.2023.1137645">10.3389/feart.2023.1137645</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Su</surname>
<given-names>Zhaodong</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/2078412/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Cao</surname>
<given-names>Junxing</given-names>
</name>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/1494631/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Xiang</surname>
<given-names>Tao</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/1629981/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Fu</surname>
<given-names>Jingcheng</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/1966366/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Shi</surname>
<given-names>Shaochen</given-names>
</name>
</contrib>
</contrib-group>
<aff>
<institution>State Key Laboratory of Oil and Gas Reservoir Geology and Exploitation</institution>, <institution>Chengdu University of Technology</institution>, <addr-line>Chengdu</addr-line>, <country>China</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1372560/overview">Alex Hay-Man Ng</ext-link>, Guangdong University of Technology, China</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1973342/overview">Chuang Xu</ext-link>, Guangdong University of Technology, China</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1687230/overview">Taki Hasan Rafi</ext-link>, Hanyang University, Republic of Korea</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Junxing Cao, <email>caojx@cdut.edu.cn</email>
</corresp>
</author-notes>
<pub-date pub-type="epub">
<day>09</day>
<month>06</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>11</volume>
<elocation-id>1137645</elocation-id>
<history>
<date date-type="received">
<day>04</day>
<month>01</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>31</day>
<month>05</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Su, Cao, Xiang, Fu and Shi.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Su, Cao, Xiang, Fu and Shi</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Porosity is a crucial index in reservoir evaluation. In tight reservoirs, the porosity is low, resulting in weak seismic responses to changes in porosity. Moreover, the relationship between porosity and seismic response is complex, making accurate porosity inversion prediction challenging. This paper proposes a Transformer-based seismic multi-attribute inversion prediction method for tight reservoir porosity to address this issue. The proposed method takes multiple seismic attributes as input data and porosity as output data. The Transformer mapping transformation network consists of an encoder, a multi-head attention layer, and a decoder and is optimized for training with a gating mechanism and a variable selection module. Applying this method to actual data from a tight sandstone gas exploration area in the Sichuan Basin yielded a porosity prediction coincidence rate of 95% with the well data.</p>
</abstract>
<kwd-group>
<kwd>tight reservoirs</kwd>
<kwd>porosity prediction</kwd>
<kwd>deep learning</kwd>
<kwd>multi-headed attention mechanism</kwd>
<kwd>Sichuan Basin</kwd>
</kwd-group>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Earth Geophysics</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="s1">
<title>1 Introduction</title>
<p>Unconventional oil and gas resources globally are abundant, with tight reservoirs being the primary focus of exploration, mainly referring to tight gas (<xref ref-type="bibr" rid="B39">Zou et al., 2014</xref>). Tight oil and gas research and development originated in North America (<xref ref-type="bibr" rid="B29">US Energy Information Administration, 2017</xref>; <xref ref-type="bibr" rid="B18">Hu et al., 2018</xref>). The porosity of reservoirs is a crucial parameter that determines oil and gas reserves and the productivity of reservoirs. Generally, the porosity of tight reservoirs is less than 10%. They are characterized by poor lateral continuity, strong vertical heterogeneity, complex lithology, and significant variations in physical properties, making it challenging to predict their porosity. Traditional interpretation methods for predicting porosity are slow and labor-intensive. This leads to a slow development of tight oil and gas exploration, emphasizing the critical need for new technologies and methods. Given the difficulty of predicting the physical properties of tight reservoirs, the accurate prediction of their porosity is particularly crucial during the development of new technologies.</p>
<p>Due to the complex nonlinear relationship between each parameter and porosity, conventional reservoir parameter prediction methods are not ideal. Traditionally, porosity prediction technology driven by a model obtains elastic properties by inversion first and then converts them into reservoir parameters such as porosity through a petrophysical model (<xref ref-type="bibr" rid="B3">Avseth et al., 2010</xref>; <xref ref-type="bibr" rid="B19">Johansen et al., 2013</xref>). However, this method is constrained by the shackles of linear equations. Tight sandstone reservoirs have problems such as low porosity and low permeability, poor physical properties, complex pore structure, and high irreducible water saturation, making it difficult to determine fluid properties and saturation. Some scholars have used the Bayesian formula&#x2019;s joint inversion of elastic and petrophysical properties to estimate reservoir properties (<xref ref-type="bibr" rid="B5">Bosch et al., 2009</xref>; <xref ref-type="bibr" rid="B10">de Figueiredo et al., 2018</xref>; <xref ref-type="bibr" rid="B33">Wang P. et al., 2020</xref>). This approach continuously propagates uncertainty from seismic data to reservoir properties by considering elasticity and reservoir properties. In recent years, for tight reservoirs, scholars have also carried out corresponding research on reservoir properties, such as porosity, with traditional methods (<xref ref-type="bibr" rid="B1">Adelinet et al., 2015</xref>; <xref ref-type="bibr" rid="B27">Pang et al., 2021</xref>). However, when combining the elastic properties through statistical rock physics, the process of joint inversion problems associated with reservoir properties is often limited by computational and time costs. Therefore, it is difficult to accurately identify the fluid type in tight reservoirs, especially to find oil-bearing reservoirs in the formation.</p>
<p>In recent years, the development of deep learning in geophysical exploration has been significant. Numerous experimental studies have confirmed that different data representations significantly impact the accuracy of task learning. A suitable data representation method can eliminate irrelevant factors to the learning objective while preserving intrinsically related information (<xref ref-type="bibr" rid="B24">Li et al., 2019a</xref>). Rapid and reliable porosity prediction based on seismic data is a critical issue in reservoir exploration and development. Several researchers have used deep learning for porosity prediction with some preliminary application results (<xref ref-type="bibr" rid="B31">Wang J. et al., 2020</xref>; <xref ref-type="bibr" rid="B28">Song et al., 2021</xref>). For instance, seismic inversion and feedforward neural networks have been used for three-dimensional porosity prediction (<xref ref-type="bibr" rid="B21">Leite, E. P. and Vidal A. C., 2011</xref>). Recurrent Neural Networks (RNNs) and Convolutional Neural Networks (CNNs) are commonly used deep learning techniques for geophysical data processing and interpretation. For example, based on Multilayer Long and Short Term Memory Network (MLSTM) developed based on traditional Long and Short Term Memory (LSTM) model for porosity prediction of logs (Wei Chen et al., 2020). Additionally, some researchers have directly predicted porosity using CNNs with full waveform inversion parameters (<xref ref-type="bibr" rid="B14">Feng, 2020</xref>) and have used the Gaussian Mixture Model Deep Neural Network (GMM-DNN) to invert porosity from seismic elastic parameters (<xref ref-type="bibr" rid="B32">Wang et al., 2022</xref>).</p>
<p>Most studies in seismic and well logging data processing have incorporated deep learning techniques, particularly the Transformer architecture, which is the best model for sequence data processing due to its attention mechanism (<xref ref-type="bibr" rid="B30">Vaswani et al., 2017</xref>). Today it is widely adopted in various fields, such as natural language processing (NLP), computer vision (CV), and speech processing. However, its flexible design has led to its adoption in many other fields. These include image classification (<xref ref-type="bibr" rid="B6">Chen et al., 2020</xref>; <xref ref-type="bibr" rid="B12">Dosovitskiy et al., 2020</xref>; <xref ref-type="bibr" rid="B26">Liu et al., 2021</xref>), object detection (<xref ref-type="bibr" rid="B8">Carion et al., 2020</xref>; <xref ref-type="bibr" rid="B37">Zheng et al., 2020</xref>; <xref ref-type="bibr" rid="B38">Zhu et al., 2020</xref>; <xref ref-type="bibr" rid="B26">Liu et al., 2021</xref>), speech-to-text translation (<xref ref-type="bibr" rid="B16">Han et al., 2021</xref>), and text-to-image generation (<xref ref-type="bibr" rid="B11">Ding et al., 2021</xref>; <xref ref-type="bibr" rid="B2">Ramesh et al., 2021</xref>). Among their salient benefits, Transformers enable modeling long dependencies between input sequence elements and support parallel processing of sequence as compared to recurrent networks, e.g., Long short-term memory (LSTM). Moreover, their simple design allows similar processing blocks for multiple modalities, including images, video, text, and speech, making them highly scalable for processing large volumes of data. Based on Transformer&#x2019;s implementation in different domains, these state-of-the-art results demonstrate Transformer&#x2019;s effectiveness. These advantages lay the foundation for using the Transformer network for seismic data applications. It is based on the attention mechanism (<xref ref-type="bibr" rid="B4">Bahdanau et al., 2014</xref>). In a sequence prediction task, it is essential to efficiently allocate resources to enhance the information of highly correlated sequence data while reducing the information of weakly correlated sequence data to improve prediction accuracy and reliability. The attention mechanism is a resource allocation mechanism that focuses on important features, dividing the degree of attention to the information in a feature-weighted manner and highlighting the impact of more important information. By mapping weights and learning parameter matrices, the attention mechanism (<xref ref-type="bibr" rid="B37">Zang H et al., 2020</xref>) reduces information loss and enhances the impact of important information, thus improving prediction accuracy and reliability. Transformer architecture, which abandons the usual recursion and convolution, has shown improved quality and superior parallel computing capabilities for processing large data (Brown et al., 2020; <xref ref-type="bibr" rid="B22">Lepikhin et al., 2021</xref>; <xref ref-type="bibr" rid="B41">Zu et al., 2022</xref>) compared to the Long Short-Term Memory network (LSTM) (<xref ref-type="bibr" rid="B17">Hocheriter et al., 1997</xref>), which overcomes the gradient vanishing and bursting problems in recurrent neural networks and can effectively process sequence data. Consequently, deep learning can solve problems with massive amounts of data. On this basis, this study proposes a Transformer architecture based on the attention mechanism. It establishes a new network TP (Transformer Prediction), which uses multi-attribute seismic and multi-scale data to predict the porosity of tight reservoirs. It has been shown that this network framework has better results in natural language translation, image processing, and data analysis processing. It is much faster than recurrent neural networks and convolutional neural networks regarding training speed. Even though the resolution of the prediction results is reduced compared with the logging data, the nonlinear inversion scheme can effectively reflect the complex relationship between rock properties and seismic data. It may obtain relatively more comprehensive and high-quality inversion results. As a regression process, apply deep learning to estimate tight reservoir porosity based on the output of a seismic inversion scheme. In order to learn the relationship between different features, a recurrent layer is used for local processing and a multi-head attention layer for long-term dependencies, allowing the network architecture to capture potential features better, ensuring better predictive performance. Finally, the method was successfully applied to the actual seismic data of a survey area in the Sichuan Basin and obtained a relatively good inversion result.</p>
</sec>
<sec sec-type="methods" id="s2">
<title>2 Methodology</title>
<p>We design the TP model to construct features efficiently for predicting porosity in dense reservoirs and, ultimately, the overall porosity. The main components of the TP are.<list list-type="simple">
<list-item>
<p>1) Gating mechanism: to skip unused components of the architecture to accommodate the input transmission of different seismic data.</p>
</list-item>
<list-item>
<p>2) Static covariates encoders: integration of static features into the network, conditioning the data by encoding the context vectors.</p>
</list-item>
<list-item>
<p>3) Variable selection module: Selects relevant variables for the input data.</p>
</list-item>
</list>
</p>
<p>
<xref ref-type="fig" rid="F1">Figure 1</xref> illustrates the overall structure of the TP, and the individual components are described in detail in the subsequent subsections.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>Transformer Prediction Network Architecture. GRN and LSTM are gated residual network and Long Short-Term Memory network, respectively.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g001.tif"/>
</fig>
<sec id="s2-1">
<title>2.1 Multi-headed attention mechanism</title>
<p>Seismic data is a type of sequence data, and the attention mechanism is inspired by the selective attention mechanism of human vision, which focuses on critical information while ignoring secondary information. It allows sifting through complex data to find helpful information for the present. The core idea of the attention mechanism is to selectively focus on input information by assigning weights that filter out less important information from a large amount of data. The multi-headed attention mechanism is the key component of the Transformer architecture, entirely based on the attention mechanism. It uses multiple copies of the single-headed attention mechanism to extract different information, and the outputs of these heads are concatenated and passed through a fully connected layer to produce the final output.</p>
<p>The multi-headed attention mechanism is defined as follows (<xref ref-type="bibr" rid="B30">Vaswani et al., 2017</xref>):<disp-formula id="e1">
<mml:math id="m1">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="[" close="]" separators="|">
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>K</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mi>P</mml:mi>
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:mi>K</mml:mi>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>z</mml:mi>
<mml:mi>t</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(1)</label>
</disp-formula>
<disp-formula id="e2">
<mml:math id="m2">
<mml:mrow>
<mml:mi>A</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>n</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>K</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi mathvariant="normal">s</mml:mi>
<mml:mi mathvariant="normal">o</mml:mi>
<mml:mi mathvariant="normal">f</mml:mi>
<mml:mi mathvariant="normal">t</mml:mi>
<mml:mi mathvariant="normal">m</mml:mi>
<mml:mi mathvariant="normal">a</mml:mi>
<mml:mi mathvariant="normal">x</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mfrac>
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:msup>
<mml:mi>K</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
</mml:mrow>
<mml:msqrt>
<mml:mi>D</mml:mi>
</mml:msqrt>
</mml:mfrac>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:math>
<label>(2)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf1">
<mml:math id="m3">
<mml:mrow>
<mml:mi>Q</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> denotes the information query set, <inline-formula id="inf2">
<mml:math id="m4">
<mml:mrow>
<mml:mi>K</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the corresponding similarity set, and <inline-formula id="inf3">
<mml:math id="m5">
<mml:mrow>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> denotes the value set. The matrices of <inline-formula id="inf4">
<mml:math id="m6">
<mml:mrow>
<mml:mi>Q</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, <inline-formula id="inf5">
<mml:math id="m7">
<mml:mrow>
<mml:mi>K</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> and <inline-formula id="inf6">
<mml:math id="m8">
<mml:mrow>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> are projected by the inputs in the attention module using the fully connected layer shown in Eq <xref ref-type="disp-formula" rid="e1">1</xref>. The scaling factor <inline-formula id="inf7">
<mml:math id="m9">
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>/</mml:mo>
<mml:msqrt>
<mml:mi>D</mml:mi>
</mml:msqrt>
</mml:mrow>
</mml:math>
</inline-formula> is applied to the parametric gradient, the softmax function such that the sum of <inline-formula id="inf8">
<mml:math id="m10">
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:msup>
<mml:mi>K</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mo>/</mml:mo>
<mml:msqrt>
<mml:mi>D</mml:mi>
</mml:msqrt>
</mml:mrow>
</mml:math>
</inline-formula> is equal to 1. The attention weight is denoted as <inline-formula id="inf9">
<mml:math id="m11">
<mml:mrow>
<mml:mi mathvariant="normal">W</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mi mathvariant="normal">s</mml:mi>
<mml:mi mathvariant="normal">o</mml:mi>
<mml:mi mathvariant="normal">f</mml:mi>
<mml:mi mathvariant="normal">t</mml:mi>
<mml:mi mathvariant="normal">m</mml:mi>
<mml:mi mathvariant="normal">a</mml:mi>
<mml:mi mathvariant="normal">x</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>Q</mml:mi>
<mml:msup>
<mml:mi>K</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mo>/</mml:mo>
<mml:msqrt>
<mml:mi>D</mml:mi>
</mml:msqrt>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula>, where <inline-formula id="inf10">
<mml:math id="m12">
<mml:mrow>
<mml:mi>W</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is based on the pairwise correlation between the query set <inline-formula id="inf11">
<mml:math id="m13">
<mml:mrow>
<mml:mi>Q</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> and the key set <inline-formula id="inf12">
<mml:math id="m14">
<mml:mrow>
<mml:mi>K</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, <inline-formula id="inf13">
<mml:math id="m15">
<mml:mrow>
<mml:mi>D</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> denotes the dimension of the sequence. In addition, the attention weights allow us to extract relevant features using the value set V. For more details on the Transformer architecture, refer to <xref ref-type="bibr" rid="B30">Vaswani et al. (2017)</xref>.</p>
<p>When using the LSTM network alone to input a long sequence for prediction, the gradient update may decay quickly, hindering the update of sequence data to some extent and making it challenging to represent the feature vector effectively. Additionally, subsequent input data also overwrites the previous input, resulting in the loss of details. However, by using the multi-head attention module and performing a linear transformation to generate <inline-formula id="inf14">
<mml:math id="m16">
<mml:mrow>
<mml:mi>Q</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, <inline-formula id="inf15">
<mml:math id="m17">
<mml:mrow>
<mml:mi>K</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, and <inline-formula id="inf16">
<mml:math id="m18">
<mml:mrow>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> matrices, after expanding the original low-dimensional information to a higher dimension, taking into account the differences between seismic data and natural language, and at the same time having the interaction capability of global information for the input of long sequence multiple seismic attribute data. It not only enhances the influence of strong signals of seismic attribute data but also focuses on weak signal anomalies, avoiding the problem of losing important feature information and reducing the interference of unnecessary information in the sequence. The joint LSTM can effectively learn the rich features and laws in the input data to obtain the porosity data in dense reservoirs. Therefore, the attention mechanism is auxiliary in extracting complex data features.</p>
</sec>
<sec id="s2-2">
<title>2.2 Variable selection and gating network</title>
<p>According to the Temporal Fusion Transformer architecture (<xref ref-type="bibr" rid="B25">Lim B et al., 2021</xref>), some modules were modified accordingly for the large seismic data volume. We added a gating and variable selection network to the Transformer for seismic data.</p>
<p>In geological studies, lithology typically varies with depth Logging data provides information about the formation rocks, while seismic attributes reflect different features of the formation. The corresponding values for porosity prediction should be a weighted sum of adjacent characteristic responses that have a certain correlation with each other. Therefore, when we establish the relationship between porosity and external input of various seismic attributes, we should consider the local correlation, the trend of change with depth, and the relationship of adjacent strata. However, since the relationships between external seismic attribute inputs are unknown and uninterpretable for porosity prediction, it is difficult to determine which variables are correlated and the degree of nonlinear processing required. In summary, we need a module for the nonlinear processing of the inputs in this network, so we added a gated residual network (GRN) as a building block for our network. The GRN would accept a primary input <inline-formula id="inf17">
<mml:math id="m19">
<mml:mrow>
<mml:mi mathvariant="bold">a</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> and context vector <inline-formula id="inf18">
<mml:math id="m20">
<mml:mrow>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, giving it the functionality to obtain more accurate and comprehensive features.<disp-formula id="e3">
<mml:math id="m21">
<mml:mrow>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mi>&#x3c9;</mml:mi>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi mathvariant="bold">a</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>L</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>y</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>m</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi mathvariant="bold">a</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>G</mml:mi>
<mml:mi>L</mml:mi>
<mml:msub>
<mml:mi>U</mml:mi>
<mml:mi>&#x3c9;</mml:mi>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(3)</label>
</disp-formula>
<disp-formula id="e4">
<mml:math id="m22">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">2</mml:mn>
</mml:msub>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">b</mml:mi>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
<label>(4)</label>
</disp-formula>
<disp-formula id="e5">
<mml:math id="m23">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">2</mml:mn>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>E</mml:mi>
<mml:mi>L</mml:mi>
<mml:mi>U</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">a</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>3</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">c</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">b</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(5)</label>
</disp-formula>
</p>
<p>Where ELU denotes the unit activation function (<xref ref-type="bibr" rid="B7">Clevert, Unterthiner, &#x26; Hochreiter, 2016</xref>), <inline-formula id="inf19">
<mml:math id="m24">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">1</mml:mn>
</mml:msub>
<mml:mo>&#x2208;</mml:mo>
<mml:msup>
<mml:mi mathvariant="double-struck">R</mml:mi>
<mml:msub>
<mml:mi>d</mml:mi>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>d</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>l</mml:mi>
</mml:mrow>
</mml:msub>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula>, <inline-formula id="inf20">
<mml:math id="m25">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3b7;</mml:mi>
<mml:mn mathvariant="bold">2</mml:mn>
</mml:msub>
<mml:mo>&#x2208;</mml:mo>
<mml:msup>
<mml:mi mathvariant="double-struck">R</mml:mi>
<mml:msub>
<mml:mi>d</mml:mi>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>d</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>l</mml:mi>
</mml:mrow>
</mml:msub>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> are intermediate layers, <inline-formula id="inf21">
<mml:math id="m26">
<mml:mrow>
<mml:mi>L</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>y</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>m</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is a standard layer normalization by Lei Ba, Kiros and Hinton (2016), <inline-formula id="inf22">
<mml:math id="m27">
<mml:mrow>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the metric indicating weight sharing, <inline-formula id="inf23">
<mml:math id="m28">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">b</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mo>&#xb7;</mml:mo>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2208;</mml:mo>
<mml:msup>
<mml:mi mathvariant="double-struck">R</mml:mi>
<mml:msub>
<mml:mi>d</mml:mi>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>d</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>l</mml:mi>
</mml:mrow>
</mml:msub>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> are the weights and biases. The ELU activation function will then function as a recognition function when <inline-formula id="inf24">
<mml:math id="m29">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">a</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>3</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">c</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">b</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is greater than zero, and ELU activation will produce a constant output when <inline-formula id="inf25">
<mml:math id="m30">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">a</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">W</mml:mi>
<mml:mrow>
<mml:mn>3</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mi mathvariant="bold">c</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi mathvariant="bold">b</mml:mi>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mi>&#x3c9;</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is less than zero. We use a component gating layer based on Gated Linear Units (GLUs) (<xref ref-type="bibr" rid="B9">Dauphin et al., 2017</xref>) to provide flexibility to suppress certain repetitions of unwanted parts of a particular seismic attribute dataset.</p>
<p>The variable selection also allows TP to remove any unnecessary noisy inputs that may negatively affect the prediction. However, identifying null or invalid values of the seismic attribute data manually in a large work area can be time-consuming and laborious. Hence, we introduced a variable selection module (<xref ref-type="bibr" rid="B15">Gal &#x26; Ghahramani, 2016</xref>) to automate this process. A linear transformation of the variables is also applied to convert each post-input variable into a d-dimensional vector, which matches the dimensions in the subsequent layers and is mainly used to achieve jump connections. The weights vary for different data (determined by the correlation of each attribute with the porosity). The variable selection network on the seismic data input is presented below, as shown in <xref ref-type="fig" rid="F2">Figure 2</xref>. Note that the variable selection network is the same for the other data set inputs.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>Variable selection network.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g002.tif"/>
</fig>
<p>We denote <inline-formula id="inf26">
<mml:math id="m31">
<mml:mrow>
<mml:msubsup>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
<mml:mo>&#x2208;</mml:mo>
<mml:msup>
<mml:mi mathvariant="double-struck">R</mml:mi>
<mml:msub>
<mml:mi>d</mml:mi>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>d</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>l</mml:mi>
</mml:mrow>
</mml:msub>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> as the <inline-formula id="inf27">
<mml:math id="m32">
<mml:mrow>
<mml:msub>
<mml:mi>j</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>h</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> variable input value in a time-domain earthquake sequence, where <inline-formula id="inf28">
<mml:math id="m33">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x39e;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:msup>
<mml:mrow>
<mml:mfenced open="[" close="]" separators="|">
<mml:mrow>
<mml:msubsup>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
<mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>T</mml:mi>
</mml:msup>
</mml:msubsup>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msubsup>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
<mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>m</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>T</mml:mi>
</mml:msup>
</mml:msubsup>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>T</mml:mi>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> is the flattened vector of which all have been entered. The weights for variable selection are then generated by feeding both <inline-formula id="inf29">
<mml:math id="m34">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x39e;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> and additional context vectors <inline-formula id="inf30">
<mml:math id="m35">
<mml:mrow>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> into GRN and then into the Softmax layer:<disp-formula id="e6">
<mml:math id="m36">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3c5;</mml:mi>
<mml:mrow>
<mml:mi>&#x3c7;</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mi mathvariant="normal">S</mml:mi>
<mml:mi mathvariant="normal">o</mml:mi>
<mml:mi mathvariant="normal">f</mml:mi>
<mml:mi mathvariant="normal">t</mml:mi>
<mml:mi mathvariant="normal">m</mml:mi>
<mml:mi mathvariant="normal">a</mml:mi>
<mml:mi mathvariant="normal">x</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mrow>
<mml:mi>&#x3c5;</mml:mi>
<mml:mi>&#x3c7;</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x39e;</mml:mi>
<mml:mi>t</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(6)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf31">
<mml:math id="m37">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3c5;</mml:mi>
<mml:mrow>
<mml:mi>&#x3c7;</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2208;</mml:mo>
<mml:msup>
<mml:mi mathvariant="double-struck">R</mml:mi>
<mml:msub>
<mml:mi>m</mml:mi>
<mml:mi>&#x3c7;</mml:mi>
</mml:msub>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> is the vector of variable selection weights and <inline-formula id="inf32">
<mml:math id="m38">
<mml:mrow>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is obtained from the static covariate encoder (<xref ref-type="sec" rid="s2-3">Section 2.3</xref>).</p>
<p>An additional non-linear processing layer is applied during the input process by passing each sequence element through its own GNR.<disp-formula id="e7">
<mml:math id="m39">
<mml:mrow>
<mml:msubsup>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mrow>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msubsup>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(7)</label>
</disp-formula>where <inline-formula id="inf33">
<mml:math id="m40">
<mml:mrow>
<mml:msubsup>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
</mml:mrow>
</mml:math>
</inline-formula> is the feature vector after processing the seismic attribute <inline-formula id="inf34">
<mml:math id="m41">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>. We note that each sequence has its own <inline-formula id="inf35">
<mml:math id="m42">
<mml:mrow>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mrow>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>, and the weights are shared globally. The processed features are then weighted and combined according to their variable selection weights as follows:<disp-formula id="e8">
<mml:math id="m43">
<mml:mrow>
<mml:msub>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mrow>
<mml:mstyle displaystyle="true">
<mml:munderover>
<mml:mo>&#x2211;</mml:mo>
<mml:mrow>
<mml:mi>j</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mi>m</mml:mi>
</mml:munderover>
</mml:mstyle>
<mml:msubsup>
<mml:mi mathvariant="bold">&#x3c5;</mml:mi>
<mml:mrow>
<mml:mi>&#x3c7;</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
</mml:mrow>
<mml:msubsup>
<mml:mover accent="true">
<mml:mi>X</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
</mml:mrow>
</mml:math>
<label>(8)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf36">
<mml:math id="m44">
<mml:mrow>
<mml:msubsup>
<mml:mi mathvariant="bold">&#x3c5;</mml:mi>
<mml:mrow>
<mml:mi>&#x3c7;</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:msubsup>
</mml:mrow>
</mml:math>
</inline-formula> is the <inline-formula id="inf37">
<mml:math id="m45">
<mml:mrow>
<mml:msub>
<mml:mi>j</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>h</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> element of vector <inline-formula id="inf38">
<mml:math id="m46">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="bold">&#x3c5;</mml:mi>
<mml:mrow>
<mml:mi>&#x3c7;</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>.</p>
<p>In summary, the correlation between seismic attribute data and porosity is unknown, and adding the GRN module provides the advantage of complementing the seismic attribute preference. Additionally, because the input of various seismic attribute data may contain many invalid values, the variable selection network is used to further optimize the input data and prevent significant errors in the prediction results.</p>
</sec>
<sec id="s2-3">
<title>2.3 Static covariant encoder</title>
<p>Due to the variation of rock properties with depth and the correlation between seismic response adjacencies, it is crucial to consider not only the local correlation between seismic attribute data but also the trend of seismic data with depth and adjacency information when establishing the mapping relationship between different attribute data and porosity. The TP model can integrate information from the logging porosity data. A separate GNR encoder is then used to generate a context vector. The context vectors are connected to different locations in the network architecture where static variables play an essential role. These include the selection of data for additional seismic attributes, the underlying processing of features, the use of logging curves to guide the identification of the context vector, and the enrichment of features for data adjacency.</p>
</sec>
<sec id="s2-4">
<title>2.4 Decoder</title>
<p>We have considered the use of the LSTM encoder-decoder as a building block for our prediction architecture, following the success of this approach in typical sequential coding problems (<xref ref-type="bibr" rid="B35">Wen et al., 2017</xref>; <xref ref-type="bibr" rid="B13">Fan et al., 2019</xref>; <xref ref-type="bibr" rid="B34">Wang et al., 2022</xref>). Due to the specificity of seismic data, the critical points in the data we apply are usually determined based on the surrounding values. Therefore, we propose to use a sequence-to-sequence layer to process the data naturally. Send <inline-formula id="inf39">
<mml:math id="m47">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>k</mml:mi>
<mml:mo>:</mml:mo>
<mml:mi>i</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> to the encoder and <inline-formula id="inf40">
<mml:math id="m48">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>k</mml:mi>
<mml:mo>:</mml:mo>
<mml:mi>i</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mi>&#x3c4;</mml:mi>
<mml:mi>max</mml:mi>
</mml:msub>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> to the decoder. A set of uniform features is generated as input to the decoder itself, using <inline-formula id="inf41">
<mml:math id="m49">
<mml:mrow>
<mml:mi>&#x3d5;</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2208;</mml:mo>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:mi>&#x3d5;</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>k</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:mi>&#x3d5;</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>&#x3c4;</mml:mi>
<mml:mi>max</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula>, and <inline-formula id="inf42">
<mml:math id="m50">
<mml:mrow>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is a position index. In addition, to allow static metadata to affect local processing, we use a context vector from a static covariate encoder to initialize the cell state and hidden state of the first LSTM in the layer, respectively. We also employ a gated jump connection on this layer at:<disp-formula id="e9">
<mml:math id="m51">
<mml:mrow>
<mml:mover accent="true">
<mml:mi>&#x3d5;</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>L</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>y</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>m</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>G</mml:mi>
<mml:mi>L</mml:mi>
<mml:msub>
<mml:mi>U</mml:mi>
<mml:mover accent="true">
<mml:mi>&#x3d5;</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>&#x3d5;</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(9)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf43">
<mml:math id="m52">
<mml:mrow>
<mml:mi>n</mml:mi>
<mml:mo>&#x2208;</mml:mo>
<mml:mrow>
<mml:mfenced open="[" close="]" separators="|">
<mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>k</mml:mi>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>&#x3c4;</mml:mi>
<mml:mi>max</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> is a location index, <inline-formula id="inf44">
<mml:math id="m53">
<mml:mrow>
<mml:mi>i</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> denotes the <inline-formula id="inf45">
<mml:math id="m54">
<mml:mrow>
<mml:msub>
<mml:mi>i</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>h</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> sequence data.</p>
<p>Since static covariates often have an enormous impact on features, we introduce a layer simultaneously to augment the corresponding features with static metadata. For a given location index <inline-formula id="inf46">
<mml:math id="m55">
<mml:mrow>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, the form of:<disp-formula id="e10">
<mml:math id="m56">
<mml:mrow>
<mml:mi>&#x3b8;</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mi>&#x3b8;</mml:mi>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mover accent="true">
<mml:mi>&#x3d5;</mml:mi>
<mml:mo>&#x223c;</mml:mo>
</mml:mover>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
<label>(10)</label>
</disp-formula>
</p>
<p>Where the whole layer shares the weights of <inline-formula id="inf47">
<mml:math id="m57">
<mml:mrow>
<mml:mi>G</mml:mi>
<mml:mi>R</mml:mi>
<mml:msub>
<mml:mi>N</mml:mi>
<mml:mi>&#x3b8;</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>, and <inline-formula id="inf48">
<mml:math id="m58">
<mml:mrow>
<mml:mi mathvariant="bold">c</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the context vector from the static covariate encoder. Finally, the attention layer, the decoder mask (<xref ref-type="bibr" rid="B30">Vaswani et al., 2017</xref>; <xref ref-type="bibr" rid="B23">Li et al., 2019b</xref>), is applied to the multi-head attention mask layer to ensure that each dimension notices its previous features.</p>
</sec>
</sec>
<sec id="s3">
<title>3 A case study</title>
<sec id="s3-1">
<title>3.1 Data preparation</title>
<p>Considering the completeness and universality, many seismic attributes need to be extracted, resulting in redundant information. Seismic attribute optimization is to preferably select the attributes with better correlation with the target reservoir parameters from a large number of attributes to eliminate the redundant information and thus improve the prediction accuracy. The preferred effect of seismic attributes is often reflected in oil and gas prediction results. In this study, the analysis of the rendezvous plots was first conducted using the Pearson correlation coefficient (PCC) as the criterion. The PCC reflects the degree of similarity of each unit change between two variables, and the closer the result is to 1 indicates the more significant correlation between the variables, which is calculated by the following equation:<disp-formula id="e11">
<mml:math id="m59">
<mml:mrow>
<mml:mi>P</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>C</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>v</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>y</mml:mi>
<mml:mo>,</mml:mo>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x5e;</mml:mo>
</mml:mover>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mrow>
<mml:msub>
<mml:mi>&#x3c3;</mml:mi>
<mml:mi>y</mml:mi>
</mml:msub>
<mml:msub>
<mml:mi>&#x3c3;</mml:mi>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x5e;</mml:mo>
</mml:mover>
</mml:msub>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
<label>(11)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf49">
<mml:math id="m60">
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>v</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>y</mml:mi>
<mml:mo>,</mml:mo>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x5e;</mml:mo>
</mml:mover>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> denotes the covariance of the two variables <inline-formula id="inf50">
<mml:math id="m61">
<mml:mrow>
<mml:mi>y</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> and <inline-formula id="inf51">
<mml:math id="m62">
<mml:mrow>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x5e;</mml:mo>
</mml:mover>
</mml:mrow>
</mml:math>
</inline-formula>, used to describe the linear relationship between the two variables, and <inline-formula id="inf52">
<mml:math id="m63">
<mml:mrow>
<mml:mi>&#x3c3;</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the standard deviation.</p>
<p>This study analyzed porosity separately with multiple attributes, and we used the attribute preference method to perform preliminary screening of each seismic attribute. The multiple seismic attributes shown in <xref ref-type="table" rid="T1">Table 1</xref> include post-stack seismic attributes and various derived post-stack attributes in order to avoid the multi-solution nature of single-parameter inversion to initially screen out the useless data with enormous influence. <xref ref-type="fig" rid="F1">Figure 1</xref> shows a representative rendezvous analysis of several attributes, and other rendezvous analyses are also performed in the same way we can select the corresponding seismic attributes. Finally, 12 seismic attribute bodies such as Q Factor, Fluid Factor, PG, P-impedance, and variance are selected as input according to PCC, as shown in <xref ref-type="table" rid="T2">Table 2</xref> and <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<table-wrap id="T1" position="float">
<label>TABLE 1</label>
<caption>
<p>Optional seismic attributes.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Index</th>
<th align="left">Attribute</th>
<th align="left">Index</th>
<th align="left">Attribute</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">0</td>
<td align="left">Post-stack</td>
<td align="left">15</td>
<td align="left">P-impedance</td>
</tr>
<tr>
<td align="left">1</td>
<td align="left">
<inline-formula id="inf53">
<mml:math id="m64">
<mml:mrow>
<mml:msub>
<mml:mi>V</mml:mi>
<mml:mi>p</mml:mi>
</mml:msub>
<mml:mo>/</mml:mo>
<mml:msub>
<mml:mi>V</mml:mi>
<mml:mi>s</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>
</td>
<td align="left">16</td>
<td align="left">Energy Half Time</td>
</tr>
<tr>
<td align="left">2</td>
<td align="left">Domain Amp</td>
<td align="left">17</td>
<td align="left">High Order Kurtosis</td>
</tr>
<tr>
<td align="left">3</td>
<td align="left">Domain Freq</td>
<td align="left">18</td>
<td align="left">Skew</td>
</tr>
<tr>
<td align="left">4</td>
<td align="left">First Peak Amp</td>
<td align="left">19</td>
<td align="left">Variance</td>
</tr>
<tr>
<td align="left">5</td>
<td align="left">First Peak Freq</td>
<td align="left">20</td>
<td align="left">Centroid Freq</td>
</tr>
<tr>
<td align="left">6</td>
<td align="left">Frequency Band</td>
<td align="left">21</td>
<td align="left">Twist</td>
</tr>
<tr>
<td align="left">7</td>
<td align="left">Peak Amp Above Average</td>
<td align="left">22</td>
<td align="left">GR</td>
</tr>
<tr>
<td align="left">8</td>
<td align="left">Per Frequency</td>
<td align="left">23</td>
<td align="left">Fluid Factor</td>
</tr>
<tr>
<td align="left">9</td>
<td align="left">Instant Amp</td>
<td align="left">24</td>
<td align="left">Q Factor</td>
</tr>
<tr>
<td align="left">10</td>
<td align="left">Instant Freq</td>
<td align="left">25</td>
<td align="left">RMS</td>
</tr>
<tr>
<td align="left">11</td>
<td align="left">Instant Phase</td>
<td align="left">26</td>
<td align="left">LamdaRho</td>
</tr>
<tr>
<td align="left">12</td>
<td align="left">Instant Weighted Freq</td>
<td align="left">27</td>
<td align="left">Coherence</td>
</tr>
<tr>
<td align="left">13</td>
<td align="left">P (AVO intercept)</td>
<td align="left">28</td>
<td align="left">Curvature</td>
</tr>
<tr>
<td align="left">14</td>
<td align="left">G (AVO intercept)</td>
<td align="left"/>
<td align="left"/>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap id="T2" position="float">
<label>TABLE 2</label>
<caption>
<p>Input seismic attribute data.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Index</th>
<th align="center">Attribute</th>
<th align="left">Index</th>
<th align="left">Attribute</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="center">6</td>
<td align="center">Frequency Band</td>
<td align="center">23</td>
<td align="center">Fluid Factor</td>
</tr>
<tr>
<td align="center">13</td>
<td align="center">P (AVO intercept)</td>
<td align="center">24</td>
<td align="center">Q Factor</td>
</tr>
<tr>
<td align="center">14</td>
<td align="center">G (AVO intercept)</td>
<td align="center">25</td>
<td align="center">RMS</td>
</tr>
<tr>
<td align="center">16</td>
<td align="center">Energy Half Time</td>
<td align="center">26</td>
<td align="center">LamdaRho</td>
</tr>
<tr>
<td align="center">19</td>
<td align="center">Variance</td>
<td align="center">27</td>
<td align="center">Coherence</td>
</tr>
<tr>
<td align="center">22</td>
<td align="center">GR</td>
<td align="center">28</td>
<td align="center">Curvature</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>Seismic attributes and porosity intersection map.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g003.tif"/>
</fig>
</sec>
<sec id="s3-2">
<title>3.2 Evaluation criteria</title>
<p>This study uses three primary metrics to evaluate the performance of porosity prediction, namely, mean square error (MSE), Pearson correlation coefficient (PCC), and coefficient of determination (<inline-formula id="inf54">
<mml:math id="m65">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula>), to quantitatively evaluate the inversion results.</p>
<p>Mean squared error. MSE is the sum of squares of the corresponding point errors between the predicted and actual data. The smaller the value, the better the fit between predicted and actual data. It is defined as shown in Eq <xref ref-type="disp-formula" rid="e12">12</xref>.<disp-formula id="e12">
<mml:math id="m66">
<mml:mrow>
<mml:mi>M</mml:mi>
<mml:mi>S</mml:mi>
<mml:mi>E</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mi>N</mml:mi>
</mml:mrow>
</mml:mfrac>
<mml:msubsup>
<mml:mo>&#x2211;</mml:mo>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mi>n</mml:mi>
</mml:msubsup>
<mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x2322;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
<label>(12)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf55">
<mml:math id="m67">
<mml:mrow>
<mml:msub>
<mml:mi>y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> and <inline-formula id="inf56">
<mml:math id="m68">
<mml:mrow>
<mml:msub>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x2322;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> are the predicted and actual values of the series, and <inline-formula id="inf57">
<mml:math id="m69">
<mml:mrow>
<mml:mi>N</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the number of sampling points.</p>
<p>Pearson Correlation Coefficient. PCC indicates the overall fit between the predicted and actual data. Its range is [-1,1]. The larger the value, the stronger the linear correlation. It is defined in Eq <xref ref-type="disp-formula" rid="e11">11</xref>.</p>
<p>Coefficient of determination. <inline-formula id="inf58">
<mml:math id="m70">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> is used to assess the fit between the variables and considers the mean square error between the predicted and actual data. It ranges from [0,1]. The larger the value, the better the fit between the variables. The following equation gives it:<disp-formula id="e13">
<mml:math id="m71">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:msup>
<mml:mrow>
<mml:msubsup>
<mml:mo>&#x2211;</mml:mo>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mi>n</mml:mi>
</mml:msubsup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#x2322;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
<mml:mrow>
<mml:msup>
<mml:mrow>
<mml:msubsup>
<mml:mo>&#x2211;</mml:mo>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mi>n</mml:mi>
</mml:msubsup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#xaf;</mml:mo>
</mml:mover>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
<label>(13)</label>
</disp-formula>
</p>
<p>Where <inline-formula id="inf59">
<mml:math id="m72">
<mml:mrow>
<mml:mover accent="true">
<mml:mi>y</mml:mi>
<mml:mo>&#xaf;</mml:mo>
</mml:mover>
</mml:mrow>
</mml:math>
</inline-formula> is the average value of <inline-formula id="inf60">
<mml:math id="m73">
<mml:mrow>
<mml:msub>
<mml:mi>y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mn>1,2</mml:mn>
<mml:mo>,</mml:mo>
<mml:mo>.</mml:mo>
<mml:mo>.</mml:mo>
<mml:mo>.</mml:mo>
<mml:mo>,</mml:mo>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula>.</p>
</sec>
<sec id="s3-3">
<title>3.3 Experiment and result analysis</title>
<p>In order to assess the validity of our proposed method, we conducted a method analysis experiment using logging data from two wells, X and Y, in a dense sandstone reservoir in an actual working area. The logging depth of well X was between 2750&#xa0;m and 2940m, while that of well Y was between 2714m and 2870&#xa0;m. The logging curves included porosity, density, spontaneous potential, p-wave velocity, natural gamma ray, deep lateral resistivity, compensated neutron, and others. After removing any outliers, we obtained two sets of logging data from wells X and Y for training and test datasets in the Transformer architecture. First, we screened the sensitive parameters of well X using Pearson correlation coefficients. Next, we used the Transformer model built with the logging data of well X to predict the porosity of well Y. To compare our model with other existing models, we also used bi-directional Long Short-Term Memory (BLSTM) and convolutional neural network (CNN) in this experiment.</p>
<p>The correlation between the porosity and other logging parameters was obtained by correlating the petrophysical logging parameters of the two test study wells, as shown in <xref ref-type="fig" rid="F4">Figure 4</xref>.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption>
<p>Logging parameters and porosity intersection map.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g004.tif"/>
</fig>
<p>The logging parameters selected for this study are density (DEN), spontaneous potential (SP), p-wave velocity (VP), natural gamma (GR), acoustic log(DT), and compensate neutron log(CNL). From the correlation graph shown in <xref ref-type="fig" rid="F4">Figure 4</xref>, it is evident that CNL, DEN, DT, and GR are highly correlated with porosity (POR), whereas natural potential and longitudinal velocity exhibit weaker correlations. Therefore, we have used the four input datasets with higher correlation for predicting porosity in tight reservoir logs. In this study, we have used Root Mean Square Error (RMSE), Mean Absolute Error (MAE), and Mean Absolute Percentage Error (MAPE) as evaluation metrics to assess the performance of the model. These metrics measure the deviation between the predicted and actual data, with lower values indicating better performance.</p>
<p>We trained the model on the preferred well X data and used it to predict the porosity of well Y. The results of the prediction are presented in <xref ref-type="fig" rid="F5">Figure 5</xref>. Using deep learning techniques and sensitive feature parameters, such as CNL, DEN, DT, and GR, we obtained good results in porosity prediction of wells. Three models, namely, Transformer, BLSTM, and CNN, were used for porosity prediction, and while errors existed, the overall fit to the actual porosity curve was satisfactory. As illustrated in <xref ref-type="fig" rid="F5">Figure 5</xref>, the predicted values of the Transformer model were closer to the actual values than those of the BLSTM and CNN models, especially in the areas of significant variation, where the results were more satisfactory. The prediction errors are summarized in <xref ref-type="table" rid="T3">Table 3</xref>, and it can be concluded that the Transformer model has more advantages.</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption>
<p>Results of porosity prediction across wells with different models (well Y).</p>
</caption>
<graphic xlink:href="feart-11-1137645-g005.tif"/>
</fig>
<table-wrap id="T3" position="float">
<label>TABLE 3</label>
<caption>
<p>Porosity prediction error for well Y.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Model</th>
<th align="left">RMSE</th>
<th align="left">MAE</th>
<th align="left">MAPE</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">Transformer</td>
<td align="left">0.512</td>
<td align="left">0.405</td>
<td align="left">18.458</td>
</tr>
<tr>
<td align="left">BLSTM</td>
<td align="left">0.562</td>
<td align="left">0.441</td>
<td align="left">20.664</td>
</tr>
<tr>
<td align="left">CNN</td>
<td align="left">0.592</td>
<td align="left">0.455</td>
<td align="left">20.397</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In this study, we applied the model to the actual data of an exploration area in eastern Sichuan for tight sandstone oil and gas exploration. It is difficult to describe the reservoir in this exploration area; the seismic response of the submerged river sand body is not obvious and variable, the reservoir thickness of the river sand body is large, but the porosity is small, the porosity prediction is difficult, and it is a dense gas reservoir. We used the logging data of 5 wells and optimized 12 types of seismic volume attribute data in this work area. We divided our dataset into three parts - a training set for learning, a validation set for hyperparameter tuning, and a test set reserved for performance evaluation. For hyperparameter optimization, we use the random search method. <xref ref-type="table" rid="T4">Table 4</xref> shows the entire search range of all hyperparameters as follows and the optimal model parameters.<list list-type="simple">
<list-item>
<p>&#x25cf; State Size&#x2013;10,20,40,80,160,240</p>
</list-item>
<list-item>
<p>&#x25cf; Dropout rate&#x2013;0.1, 0.2, 0.3, 0.4, 0.5, 0.7, 0.9</p>
</list-item>
<list-item>
<p>&#x25cf; Minibatch size&#x2013;64, 128, 256</p>
</list-item>
<list-item>
<p>&#x25cf; Learning rate&#x2013;0.0001, 0.001, 0.01, 0.1</p>
</list-item>
<list-item>
<p>&#x25cf; Max. gradient norm&#x2013;0.01, 1.0, 100.0</p>
</list-item>
<list-item>
<p>&#x25cf; Num. heads&#x2013;1, 4</p>
</list-item>
</list>
</p>
<table-wrap id="T4" position="float">
<label>TABLE 4</label>
<caption>
<p>Optimal network parameters.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="center">Network parameters</th>
<th align="left"/>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="center">Dropout rate</td>
<td align="center">0.3</td>
</tr>
<tr>
<td align="center">State size</td>
<td align="center">160</td>
</tr>
<tr>
<td align="center">Number of heads</td>
<td align="center">4</td>
</tr>
<tr>
<td colspan="2" align="center">Training Parameters</td>
</tr>
<tr>
<td align="center">&#x2003;Minibatch</td>
<td align="center">64</td>
</tr>
<tr>
<td align="center">&#x2003;Learning rate</td>
<td align="center">0.01</td>
</tr>
<tr>
<td align="center">&#x2003;Max gradient norm</td>
<td align="center">0.01</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>During the training process of our dataset, a significant amount of seismic data input is required. However, using the basic architecture of the Transformer model, the entire architecture can be deployed on a single GPU without consuming many computer resources. For instance, when we used NVIDIA Quadro GP100 GPU for the experimental data, our optimal parameter TP model only required less than 2&#xa0;hours to complete training. <xref ref-type="fig" rid="F6">Figure 6</xref> presents the comparison of well porosity prediction results, in which the red curve is the actual value, the blue curve is the CNN prediction value, the purple curve is the BLSTM prediction value and the green curve is the TP prediction value. With the CNN and BLSTM data were added to compare network effects in this study. From the comparison in the figure, we can find that the prediction result of TP is more accurate than CNN or BLSTM, and the mean square error (MSE), Pearson correlation coefficient (PCC), and are used as regression evaluation indicators. <xref ref-type="table" rid="T5">Table 5</xref> selects four wells (coded by A, B, C, and D), and it can also be proved from <xref ref-type="table" rid="T5">Table 5</xref> that our TP can obtain more accurate results.</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption>
<p>Comparison of well porosity prediction results.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g006.tif"/>
</fig>
<table-wrap id="T5" position="float">
<label>TABLE 5</label>
<caption>
<p>Comparison of the porosity prediction errors of different models; &#x201c;Means&#x201d; indicates the average value of the prediction errors of different methods; &#x201c;TP&#x201d; indicates that Transformer Prediction.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="center">Well name</th>
<th align="center">Method</th>
<th align="center">MSE</th>
<th align="center">PCC</th>
<th align="center">
<italic>R</italic>
<sup>2</sup>
</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td rowspan="2" align="center">A</td>
<td align="center">CNN</td>
<td align="center">0.2553</td>
<td align="center">0.9708</td>
<td align="center">0.8071</td>
</tr>
<tr>
<td align="center">BLSTM</td>
<td align="center">0.2029</td>
<td align="center">0.9642</td>
<td align="center">0.8245</td>
</tr>
<tr>
<td align="left"/>
<td align="center">TP</td>
<td align="center">0.1568</td>
<td align="center">0.9731</td>
<td align="center">0.8468</td>
</tr>
<tr>
<td rowspan="2" align="center">B</td>
<td align="center">CNN</td>
<td align="center">0.3367</td>
<td align="center">0.9561</td>
<td align="center">0.7533</td>
</tr>
<tr>
<td align="center">BLSTM</td>
<td align="center">0.2584</td>
<td align="center">0.9664</td>
<td align="center">0.8244</td>
</tr>
<tr>
<td align="left"/>
<td align="center">TP</td>
<td align="center">0.1286</td>
<td align="center">0.9859</td>
<td align="center">0.8958</td>
</tr>
<tr>
<td rowspan="2" align="center">C</td>
<td align="center">CNN</td>
<td align="center">0.2863</td>
<td align="center">0.9687</td>
<td align="center">0.7862</td>
</tr>
<tr>
<td align="center">BLSTM</td>
<td align="center">0.1397</td>
<td align="center">0.9462</td>
<td align="center">0.8423</td>
</tr>
<tr>
<td align="left"/>
<td align="center">TP</td>
<td align="center">0.0951</td>
<td align="center">0.9884</td>
<td align="center">0.9682</td>
</tr>
<tr>
<td rowspan="2" align="center">D</td>
<td align="center">CNN</td>
<td align="center">0.2794</td>
<td align="center">0.9694</td>
<td align="center">0.7955</td>
</tr>
<tr>
<td align="center">BLSTM</td>
<td align="center">0.1846</td>
<td align="center">0.9708</td>
<td align="center">0.8271</td>
</tr>
<tr>
<td align="left"/>
<td align="center">TP</td>
<td align="center">0.1007</td>
<td align="center">0.9862</td>
<td align="center">0.9441</td>
</tr>
<tr>
<td rowspan="2" align="center">Means</td>
<td align="center">CNN</td>
<td align="center">0.2894</td>
<td align="center">0.9663</td>
<td align="center">0.7855</td>
</tr>
<tr>
<td align="center">BLSTM</td>
<td align="center">0.1964</td>
<td align="center">0.9619</td>
<td align="center">0.8296</td>
</tr>
<tr>
<td align="left"/>
<td align="center">TP</td>
<td align="center">0.1203</td>
<td align="center">0.9834</td>
<td align="center">0.9137</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>
<xref ref-type="fig" rid="F7">Figure 7</xref> shows the porosity result profile predicted by the TP model in this study. Specifically, <xref ref-type="fig" rid="F7">Figure 7A</xref> presents the porosity result profile utilizing standard software for frequency division inversion, whereas <xref ref-type="fig" rid="F7">Figure 7B</xref> represents the porosity inversion of TP training. In the porosity result profile, the red area represents the distribution range of high porosity. The porosity curve of critical well location A in this work area is utilized to verify the prediction results. It is known that the porosity of tight reservoirs inverted by traditional methods is generally small. The important basis is the well-logging porosity, which is significantly affected by the logging data, so the overall resolution is low. <xref ref-type="fig" rid="F7">Figure 7A</xref> also proves that the overall porosity resolution inversion of the tight reservoir is not high, and the lateral continuity is poor. During porosity comparison with logging, the overall porosity in <xref ref-type="fig" rid="F7">Figure 7A</xref> is relatively small due to the complex characteristics of tight reservoirs. However, it can be found from the logging report that the actual porosity range is 0&#x2013;8%. Hence, the single inversion method still has limitations and does not match the porosity reservoir characteristics of the actual work area. In contrast, <xref ref-type="fig" rid="F7">Figure 7B</xref> illustrates the resulting porosity profile predicted by the TP model. It is observed that the high values are also concentrated near Well A compared to <xref ref-type="fig" rid="F7">Figure 7A</xref>. The overall porosity range also reaches 0&#x2013;8%, but <xref ref-type="fig" rid="F7">Figure 7B</xref> has a higher resolution. The comparison around 2058&#xa0;ms shows that the porosity distribution in <xref ref-type="fig" rid="F7">Figure 7B</xref> is more accurate in the lateral direction. <xref ref-type="fig" rid="F7">Figure 7A</xref> shows that the traditional method cannot accurately invert the pore distribution in the right half, far away from the well location, due to the limitation of logging data. <xref ref-type="table" rid="T6">Table 6</xref> shows the well and porosity profile prediction and the MSE, PCC, and <inline-formula id="inf61">
<mml:math id="m74">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> of logging and input porosity profile, respectively.</p>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption>
<p>Section of porosity results. <bold>(A)</bold>. Frequency division inversion porosity profile; <bold>(B)</bold>. TP predicted porosity profile.</p>
</caption>
<graphic xlink:href="feart-11-1137645-g007.tif"/>
</fig>
<table-wrap id="T6" position="float">
<label>TABLE 6</label>
<caption>
<p>Prediction of well and porosity profile and MSE, PCC and <inline-formula id="inf62">
<mml:math id="m75">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula>.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left"/>
<th align="left">MSE</th>
<th align="left">PCC</th>
<th align="left">
<inline-formula id="inf63">
<mml:math id="m76">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula>
</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">Training Well</td>
<td align="left">0.1203</td>
<td align="left">0.9834</td>
<td align="left">0.9137</td>
</tr>
<tr>
<td align="left">Test Well</td>
<td align="left">0.3285</td>
<td align="left">0.9422</td>
<td align="left">0.8861</td>
</tr>
<tr>
<td align="left">Profile</td>
<td align="left">0.1872</td>
<td align="left">0.9584</td>
<td align="left">0.9065</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>4 Discussion</title>
<p>In the previous experiment, the Transformer model&#x2019;s RMSE, MAE, and MAPE metrics for well X were 8.91%, 8.16%, and 10.68% lower than those of the BLSTM, respectively, and 13.51%, 10.98%, and 9.51% lower than those of the CNN. Although the overall effect was not as good as that of the well porosity prediction, the effect was more evident and reflected the generalization ability of the proposed method. Subsequent experiments comparing TP with several deep learning methods showed that the proposed method had high accuracy and strong continuity in predicting the porosity of dense reservoirs. The MSE, PCC, and <inline-formula id="inf64">
<mml:math id="m77">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> of predicted actual data porosity were 0.1203, 0.9834, and 0.9137, respectively. The MSE of predicted porosity was decreased by 0.1691 and 0.0761 when compared with the CNN-based and BLSTM-based methods, respectively, and the PCC was improved by 0.0171 and 0.0215, and <inline-formula id="inf65">
<mml:math id="m78">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> was improved by 0.1282 and 0.0841. The profile application also showed that TP accurately predicted porosity in the training work area (MSE &#x3d; 0.1872, PCC &#x3d; 0.9584, <inline-formula id="inf66">
<mml:math id="m79">
<mml:mrow>
<mml:msup>
<mml:mi>R</mml:mi>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> &#x3d; 0.9065). In summary, the TP model had certain advantages and effects regarding overall prediction accuracy and accuracy. Additionally, the TP model had an advantage in predicting the porosity of a tight reservoir, which was generally in line with the sedimentary characteristics of tight sandstone reservoirs. The experimental verification and accuracy analysis of the proposed method proved the effectiveness of the deep learning algorithm in predicting the porosity of tight reservoirs. By combining the particularity of tight reservoirs and the advantages of deep learning algorithms, this paper provided a new method for the porosity prediction of tight reservoirs, which significantly improved the accuracy of porosity prediction in tight reservoirs and provided valuable insights for reservoir exploration and exploitation. However, deep learning is a statistical model that needs to be improved in solving complex earthquake prediction problems in different work areas due to entanglement and multi-solution problems. Therefore, more targeted modules must be added to the model, which may lead to parameter redundancy and higher computational overhead costs that need to be weighed between accuracy and efficiency. To further improve the method model, the Transformer-based encoding-decoding model in porosity prediction may require the addition of appropriate constraints, a smoother loss function, or a more intelligent design, which requires more in-depth research.</p>
</sec>
<sec sec-type="conclusion" id="s5">
<title>5 Conclusion</title>
<p>This paper proposes the TP network model, a new attention-based method for predicting porosity in tight reservoirs, which improves prediction accuracy. During training, TP does not directly map seismic data to inversion parameters. Instead, the network utilizes specialized processing components to target large data of various seismic attributes, including (1) The self-attention mechanism, which enables global information interaction between data and captures deeper feature information, (2) Static covariate encoder, which integrates static metadata into the network and adjusts the data by encoding context vectors, (3) Gated network, which optimizes the transfer of data, and (4) Variable selection, which further optimizes data input. TP predicts porosity by learning the characteristics of logs as well as seismic attribute bodies. The proposed TP inversion network with high resolution, accuracy and horizontal continuity is verified in practical data experiments to effectively predict reservoir porosity in dense formations.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s6">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material further inquiries can be directed to the corresponding author.</p>
</sec>
<sec id="s7">
<title>Author contributions</title>
<p>ZS: Software, Methodology, Writing&#x2014;original draft; JC: Funding acquisition, Writing&#x2014;review and editing, Supervision; TX: Methodology, Formal analysis; JF and SS: Data curation, Visualization. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="s8">
<title>Funding</title>
<p>This work was supported by the National Natural Science Foundation of China (Grant Nos 42030812 and 41974160).</p>
</sec>
<sec sec-type="COI-statement" id="s9">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s10">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Adelinet</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ravalec</surname>
<given-names>M. L.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>Effective medium modeling: How to efficiently infer porosity from seismic data?</article-title> <source>Interpretation</source> <volume>3</volume> (<issue>4</issue>), <fpage>SAC1</fpage>&#x2013;<lpage>SAC7</lpage>. <pub-id pub-id-type="doi">10.1190/int-2015-0065.1</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Aditya</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Mikhail</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Gabriel</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Scott</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Chelsea</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Alec</surname>
<given-names>R.</given-names>
</name>
</person-group>, (<year>2021</year>). <article-title>Zero-shot text-to-image generation</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2102.12092">https://arxiv.org/abs/2102.12092</ext-link>.</citation>
</ref>
<ref id="B3">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Avseth</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Mukerji</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Mavko</surname>
<given-names>G.</given-names>
</name>
</person-group> (<year>2010</year>). <source>Quantitative seismic interpretation: Applying rock physics tools to reduce interpretation risk</source>. <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B4">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Bahdanau</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Cho</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Bengio</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>Neural machine translation by jointly learning to align and translate</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/1409.0473">https://arxiv.org/abs/1409.0473</ext-link>.</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bosch</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Carvajal</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Rodrigues</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Torres</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Aldana</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Sierra</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Petrophysical seismic inversion conditioned to well-log data: Methods and application to a gas reservoir</article-title>. <source>Geophysics</source> <volume>74</volume> (<issue>2</issue>), <fpage>O1</fpage>&#x2013;<lpage>O15</lpage>. <pub-id pub-id-type="doi">10.1190/1.3043796</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Radford</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2020</year>). &#x201c;<article-title>Generative pretraining from pixels</article-title>,&#x201d; in <conf-name>Proceedings of the ICML</conf-name>, <conf-loc>Baltimore, ML, USA</conf-loc>, <conf-date>June 2020</conf-date>.</citation>
</ref>
<ref id="B7">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Clevert</surname>
<given-names>D. A.</given-names>
</name>
<name>
<surname>Unterthiner</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Hochreiter</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>Fast and accurate deep network learning by exponential linear units (ELUs)</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/1511.07289">https://arxiv.org/abs/1511.07289</ext-link>.</citation>
</ref>
<ref id="B8">
<citation citation-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Cornia</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Stefanini</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Baraldi</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Cucchiara</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2020</year>). &#x201c;<article-title>Meshedmemory transformer for image captioning</article-title>,&#x201d; in <conf-name>Proceedings of the 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2020</conf-name>, <conf-loc>Seattle, WA, USA</conf-loc>, <conf-date>June 2020</conf-date>, <fpage>10575</fpage>&#x2013;<lpage>10584</lpage>.</citation>
</ref>
<ref id="B9">
<citation citation-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Dauphin</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Fan</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Auli</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2017</year>). &#x201c;<article-title>Language modeling with gated convolutional networks</article-title>,&#x201d; in <conf-name>Proceedings of the International conference on machine learning</conf-name>, <conf-loc>Baltimore, ML, USA</conf-loc>, <conf-date>December 2017</conf-date> (<publisher-name>PMLR</publisher-name>), <fpage>933</fpage>&#x2013;<lpage>941</lpage>.</citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>de Figueiredo</surname>
<given-names>L. P.</given-names>
</name>
<name>
<surname>Grana</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Bordignon</surname>
<given-names>F. L.</given-names>
</name>
<name>
<surname>Santos</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Roisenberg</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Rodrigues</surname>
<given-names>B. B.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Joint Bayesian inversion based on rock-physics prior modeling for the estimation of spatially correlated reservoir properties</article-title>. <source>Geophysics</source> <volume>83</volume> (<issue>5</issue>), <fpage>M49</fpage>&#x2013;<lpage>M61</lpage>. <pub-id pub-id-type="doi">10.1190/geo2017-0463.1</pub-id>
</citation>
</ref>
<ref id="B11">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Ding</surname>
<given-names>Ming</given-names>
</name>
<name>
<surname>Yang</surname>
<given-names>Zhuoyi</given-names>
</name>
<name>
<surname>Hong</surname>
<given-names>Wenyi</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Cogview: mastering text-to-image generation via transformers</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2105.13290">https://arxiv.org/abs/2105.13290</ext-link>.</citation>
</ref>
<ref id="B12">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Dosovitskiy</surname>
<given-names>Alexey</given-names>
</name>
<name>
<surname>Beyer</surname>
</name>
</person-group> (<year>2020</year>). <article-title>An image is worth 16x16 words: Transformers for image recognition at scale</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2010.11929">https://arxiv.org/abs/2010.11929</ext-link>.</citation>
</ref>
<ref id="B13">
<citation citation-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Fan</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Pan</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2019</year>). &#x201c;<article-title>Multi-horizon time series forecasting with temporal attention learning</article-title>,&#x201d; in <conf-name>Proceedings of the 25th ACM SIGKDD International conference on knowledge discovery and data mining</conf-name>, <conf-loc>Long Beach, CA, USA</conf-loc>, <conf-date>July 2019</conf-date>, <fpage>2527</fpage>&#x2013;<lpage>2535</lpage>.</citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Feng</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Estimation of reservoir porosity based on seismic inversion results using deep learning methods</article-title>. <source>J. Nat. Gas Sci. Eng.</source> <volume>77</volume>, <fpage>103270</fpage>. <pub-id pub-id-type="doi">10.1016/j.jngse.2020.103270</pub-id>
</citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gal</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Ghahramani</surname>
<given-names>Z.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>A theoretically grounded application of dropout in recurrent neural networks</article-title>. <source>Adv. neural Inf. Process. Syst.</source> <volume>29</volume>.</citation>
</ref>
<ref id="B16">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Han</surname>
<given-names>Chi</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Mingxuan</given-names>
</name>
<name>
<surname>Ji</surname>
<given-names>Heng</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>Lei</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Learning shared semantic space for speech-to-text translation</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2105.03095">https://arxiv.org/abs/2105.03095</ext-link>.</citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hocheriter</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Schmidhuber</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Cummins</surname>
<given-names>F.</given-names>
</name>
</person-group> (<year>1997</year>). <article-title>Long short-term memory</article-title>. <source>Neural comput</source>. <volume>9</volume> (<issue>8</issue>), <fpage>1735</fpage>&#x2013;<lpage>1780</lpage>. <pub-id pub-id-type="doi">10.1162/neco.1997.9.8.1735</pub-id>
</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hu</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Zhu</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Wu</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Profitable exploration and development of continental tight oil in China</article-title>. <source>Petroleum Explor. Dev.</source> <volume>45</volume> (<issue>4</issue>), <fpage>737</fpage>&#x2013;<lpage>748</lpage>. <pub-id pub-id-type="doi">10.11698/PED.2018.04.20</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Johansen</surname>
<given-names>T. A.</given-names>
</name>
<name>
<surname>Jensen</surname>
<given-names>E. H.</given-names>
</name>
<name>
<surname>Mavko</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Dvorkin</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Inverse rock physics modeling for reservoir quality prediction</article-title>. <source>Geophysics</source> <volume>78</volume> (<issue>2</issue>), <fpage>M1</fpage>&#x2013;<lpage>M18</lpage>. <pub-id pub-id-type="doi">10.1190/geo2012-0215.1</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Leite</surname>
<given-names>E. P.</given-names>
</name>
<name>
<surname>Vidal</surname>
<given-names>Alexandre C.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>3D porosity prediction from seismic inversion and neural networks</article-title>. <source>Comput. Geosciences37</source> <volume>8</volume>, <fpage>1174</fpage>&#x2013;<lpage>1180</lpage>. <pub-id pub-id-type="doi">10.1016/j.cageo.2010.08.001</pub-id>
</citation>
</ref>
<ref id="B22">
<citation citation-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Lepikhin</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>H. J.</given-names>
</name>
<name>
<surname>Xu</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2021</year>). &#x201c;<article-title>GShard: Scaling giant models with conditional computation and automatic sharding</article-title>,&#x201d; in <conf-name>Proceedings of the International Conference on Learning Representations</conf-name>, <conf-loc>Vienna, Austria</conf-loc>, <conf-date>May 2021</conf-date>.</citation>
</ref>
<ref id="B23">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Li</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2019b</year>). <article-title>Enhancing the locality and breaking the memory bottleneck of transformer on time series forecasting</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/1907.00235">https://arxiv.org/abs/1907.00235</ext-link>.</citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Li</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Liu</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Ren</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Yang</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Y.</given-names>
</name>
<etal/>
</person-group> (<year>2019a</year>). <article-title>Deep-learning inversion of seismic data</article-title>. <source>IEEE Trans. Geoscience Remote Sens.</source> <volume>58</volume> (<issue>3</issue>), <fpage>2135</fpage>&#x2013;<lpage>2149</lpage>. <pub-id pub-id-type="doi">10.1109/tgrs.2019.2953473</pub-id>
</citation>
</ref>
<ref id="B25">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lim</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Ar&#x131;k</surname>
<given-names>S. &#xd6;.</given-names>
</name>
<name>
<surname>Loeff</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Pfister</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Temporal Fusion Transformers for interpretable multi-horizon time series forecasting</article-title>. <source>Int. J. Forecast.</source> <volume>37</volume> (<issue>4</issue>), <fpage>1748</fpage>&#x2013;<lpage>1764</lpage>. <pub-id pub-id-type="doi">10.1016/j.ijforecast.2021.03.012</pub-id>
</citation>
</ref>
<ref id="B26">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Liu</surname>
<given-names>Ze</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Cao</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Swin transformer: Hierarchical vision transformer using shifted windows</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2103.14030">https://arxiv.org/abs/2103.14030</ext-link>.</citation>
</ref>
<ref id="B27">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Pang</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ba</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Carcione</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Ma</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Wei</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Seismic identification of tight-oil reservoirs by using 3D rock-physics templates</article-title>. <source>J. Petroleum Sci. Eng.</source> <volume>201</volume>, <fpage>108476</fpage>. <pub-id pub-id-type="doi">10.1016/j.petrol.2021.108476</pub-id>
</citation>
</ref>
<ref id="B28">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Song</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Feng</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Wu</surname>
<given-names>G.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Convolutional neural network, res&#x2010;unet&#x2b;&#x2b;, based dispersion curve picking from noise cross&#x2010;correlations</article-title>. <source>J. Geophys. Res. Solid Earth</source> <volume>126</volume> (<issue>11</issue>), <fpage>e2021JB022027</fpage>. <pub-id pub-id-type="doi">10.1029/2021JB022027</pub-id>
</citation>
</ref>
<ref id="B29">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Us Energy Information Administration(Eia)</surname>
</name>
</person-group> (<year>2013</year>). <source>Outlook For Shale Gas And Tight Oil Development In The Us</source>. <publisher-loc>Washington, DC, USA</publisher-loc>: <publisher-name>US Energy Information Administration</publisher-name>
</citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Vaswani</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Shazeer</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Parmar</surname>
<given-names>N.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Attention is all you need</article-title>. <source>Adv. neural Inf. Process. Syst.</source> <volume>30</volume>.</citation>
</ref>
<ref id="B31">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Cao</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>You</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2020b</year>). <article-title>Log reconstruction based on gated recurrent unit recurrent neural network</article-title>. <source>Seg. Glob. Meet. Abstr. Society of Exploration Geophysicists</source>, <fpage>91&#x2013;94</fpage>. <pub-id pub-id-type="doi">10.1190/iwmg2019_22.1</pub-id>
</citation>
</ref>
<ref id="B32">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Cao</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Yuan</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2022a</year>). <article-title>Deep learning reservoir porosity prediction method based on a spatiotemporal convolution bi-directional long short-term memory neural network model</article-title>. <source>Geomechanics Energy Environ.</source> <volume>32</volume>.<pub-id pub-id-type="publisher-id">100282</pub-id>
</citation>
</ref>
<ref id="B33">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>B.</given-names>
</name>
</person-group> (<year>2020a</year>). <article-title>Accurate porosity prediction for tight sandstone reservoir: A case study from North China</article-title>. <source>Geophysics</source> <volume>85</volume> (<issue>2</issue>), <fpage>B35</fpage>&#x2013;<lpage>B47</lpage>. <pub-id pub-id-type="doi">10.1190/geo2018-0852.1</pub-id>
</citation>
</ref>
<ref id="B34">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Niu</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Zhao</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>He</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>H.</given-names>
</name>
<etal/>
</person-group> (<year>2022b</year>). <article-title>Gaussian mixture model deep neural network and its application in porosity prediction of deep carbonate reservoir</article-title>. <source>Geophysics</source> <volume>87</volume> (<issue>2</issue>), <fpage>M59</fpage>&#x2013;<lpage>M72</lpage>. <pub-id pub-id-type="doi">10.1190/geo2020-0740.1</pub-id>
</citation>
</ref>
<ref id="B35">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Wen</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>A multi-horizon quantile recurrent forecaster</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/1711.11053">https://arxiv.org/abs/1711.11053</ext-link>.</citation>
</ref>
<ref id="B37">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Zheng</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gao</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>End-to-end object detection with adaptive clustering transformer</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2011.09315">https://arxiv.org/abs/2011.09315</ext-link>.</citation>
</ref>
<ref id="B38">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Zhu</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Su</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Lu</surname>
<given-names>L.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Deformable DETR: deformable transformers for end-to-end object detection</article-title>. <ext-link ext-link-type="uri" xlink:href="https://arxiv.org/abs/2010.04159">https://arxiv.org/abs/2010.04159</ext-link>.</citation>
</ref>
<ref id="B39">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Zou</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Tao</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Hou</surname>
<given-names>L.</given-names>
</name>
</person-group> (<year>2014</year>). <source>Unconventional petroleum geology</source>. <publisher-loc>Beijing, China</publisher-loc>: <publisher-name>Geological Publishing House</publisher-name>.</citation>
</ref>
<ref id="B41">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zu</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Ke</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Hou</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Cao</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>End-to-End deblending of simultaneous source data using transformer</article-title>. <source>IEEE Geoscience Remote Sens. Lett.</source> <volume>19</volume>, <fpage>1</fpage>&#x2013;<lpage>5</lpage>. <pub-id pub-id-type="doi">10.1109/lgrs.2022.3174106</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>