<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMC</journal-id>
<journal-id journal-id-type="nlm-ta">CMC</journal-id>
<journal-id journal-id-type="publisher-id">CMC</journal-id>
<journal-title-group>
<journal-title>Computers, Materials &#x0026; Continua</journal-title>
</journal-title-group>
<issn pub-type="epub">1546-2226</issn>
<issn pub-type="ppub">1546-2218</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">67952</article-id>
<article-id pub-id-type="doi">10.32604/cmc.2025.067952</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Prediction of Landslide Displacement Using a BiLSTM-RBF Model Based on a Hybrid Attention Mechanism</article-title>
<alt-title alt-title-type="left-running-head">Prediction of Landslide Displacement Using a BiLSTM-RBF Model Based on a Hybrid Attention Mechanism</alt-title>
<alt-title alt-title-type="right-running-head">Prediction of Landslide Displacement Using a BiLSTM-RBF Model Based on a Hybrid Attention Mechanism</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Chen</surname><given-names>Jiao</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Wang</surname><given-names>Xiao</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><email>xwang9@gzu.edu.cn</email></contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>He</surname><given-names>Zhiqin</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Chen</surname><given-names>Yi</given-names></name><xref ref-type="aff" rid="aff-2">2</xref></contrib>
<contrib id="author-5" contrib-type="author">
<name name-style="western"><surname>Ma</surname><given-names>Chao</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<aff id="aff-1"><label>1</label><institution>College of Electrical Engineering, Guizhou University</institution>, <addr-line>Guiyang, 550025</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>College of Computer Science and Technology, Guizhou University</institution>, <addr-line>Guiyang, 550025</addr-line>, <country>China</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Xiao Wang. Email: <email>xwang9@gzu.edu.cn</email></corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic">
<year>2025</year>
</pub-date>
<pub-date date-type="pub" publication-format="electronic">
<day>23</day><month>10</month><year>2025</year>
</pub-date>
<volume>85</volume>
<issue>3</issue>
<fpage>5423</fpage>
<lpage>5450</lpage>
<history>
<date date-type="received">
<day>16</day>
<month>5</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>15</day>
<month>8</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2025 The Authors.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Published by Tech Science Press.</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMC_67952.pdf"></self-uri>
<abstract>
<p>This research proposes an innovative solution to the inherent challenges faced by landslide displacement prediction models based on data-driven methods, such as the need for extensive historical datasets for training, the reliance on manual feature selection, and the difficulty in effectively utilizing landslide historical data. We have developed a dual-channel deep learning prediction model that integrates multimodal decomposition and an attention mechanism to overcome these challenges and improve prediction performance. The proposed methodology follows a three-stage framework: (1) Empirical Mode Decomposition (EMD) effectively segregates cumulative displacement and feature factors; (2) We have developed a Double Exponential Smoothing (DES) ensemble optimized through a Non-dominated Sorting Genetic Algorithm-II (NSGA-II) to enhance trend prediction; while employing a Bidirectional Long Short-Term Memory-Radial Basis Function (BiLSTM-RBF) network enhanced by a hybrid attention mechanism, which facilitates a global-local synergistic approach to hierarchical feature extraction, thereby improving the prediction of periodic displacements; (3) A bidirectional adaptive feature extraction mechanism aligns attention weights with BiLSTM propagation paths through spatial mapping, complemented by an innovative loss function incorporating Prediction Interval (PI) width optimization. In the comparative experiments of the Baishuihe landslide: the RMSE, MAE, and R<sup>2</sup> indexes of monitoring point ZG118 are improved by 19.8%, 35.2%, and 3.2% compared with the optimal baseline model (RBF-MIC); in the monitoring point ZG93, where the amount of data is less, the three indexes are even more improved by 52.1%, 32.3%, and 21.8% compared with the optimal baseline model (GRU-None). These results substantiate the model&#x2019;s capacity to overcome dual constraints of data paucity and feature engineering limitations in geohazard prediction.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Landslide displacement prediction</kwd>
<kwd>NSGA-II</kwd>
<kwd>BiLSTM</kwd>
<kwd>RBF</kwd>
<kwd>hybrid attention mechanism</kwd>
<kwd>PI</kwd>
</kwd-group>
<funding-group>
<award-group id="awg1">
<funding-source>Guizhou Province Science Technology</funding-source>
<award-id>[2024] General 007</award-id>
<award-id>[2022] General 264</award-id>
<award-id>[2023] General 096</award-id>
<award-id>[2023] General 412</award-id>
<award-id>[2023] General 409</award-id>
</award-group>
<award-group id="awg2">
<funding-source>National Natural Science Foundation of China</funding-source>
<award-id>61861007</award-id>
</award-group>
<award-group id="awg3">
<funding-source>Guizhou Province Science and Technology</funding-source>
<award-id>ZK [2021] General 303</award-id>
</award-group>
<award-group id="awg4">
<funding-source>GUIYANG HYDROPOWER INVESTIGATION DESIGN &#x0026; RESEARCH INSTITUTE CHECC</funding-source>
<award-id>YJ2022-12</award-id>
</award-group>
<award-group id="awg5">
<funding-source>Science and Technology Project of Power Construction Corporation of China</funding-source>
<award-id>DJ-ZDXM-2022-44</award-id>
</award-group>
</funding-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>As one of the important components of geomorphic hazards, landslides represent a significant danger to the safety of lives, the protection of property, and the realization of sustainable development goals. Records show that 1,470 landslides occurred in China between 1940 and 2020, resulting in 14,394 deaths [<xref ref-type="bibr" rid="ref-1">1</xref>]. The surface displacement and crack development of landslides not only intuitively reflect the geomorphic changes but are also significantly affected by geomorphic features and external factors. Meanwhile, they will change the mechanical properties and stability of slopes, which in turn affects the further evolution of the geomorphology. Therefore, developing an intelligent geohazard surveillance system is urgently needed for effective disaster prevention and mitigation. However, current landslide monitoring and prediction still face many challenges&#x2014;insufficient reliability of sensors, shortage of technical expertise and skilled labor, and low monitoring frequency have resulted in constrained availability of reliable landslide monitoring data [<xref ref-type="bibr" rid="ref-2">2</xref>,<xref ref-type="bibr" rid="ref-3">3</xref>]. Despite the large amount of historical and real-time data available, these data require extensive feature engineering to improve model performance [<xref ref-type="bibr" rid="ref-4">4</xref>], and it can be challenging to re-extract features and build models for different application scenarios [<xref ref-type="bibr" rid="ref-5">5</xref>]. Consequently, improving the accuracy of surface crack deformation prediction systems is an important way to develop potential disaster management strategies and enhance risk prevention protocols in dynamic geomorphologic environments [<xref ref-type="bibr" rid="ref-6">6</xref>].</p>
<p>Decomposing the cumulative displacement of landslides through time series can improve the prediction accuracy [<xref ref-type="bibr" rid="ref-7">7</xref>,<xref ref-type="bibr" rid="ref-8">8</xref>]. In addition, slight fluctuations in the external dynamic factors driving landslides (such as rainfall, reservoir water level, etc.) may destabilize the original system. Decomposing external factors helps to capture the fluctuating relationship between these factors and geomorphic changes [<xref ref-type="bibr" rid="ref-9">9</xref>,<xref ref-type="bibr" rid="ref-10">10</xref>].</p>
<p>The application of machine learning techniques to the prediction of natural environments, such as weather, rainfall, and wind speed, is very common [<xref ref-type="bibr" rid="ref-11">11</xref>&#x2013;<xref ref-type="bibr" rid="ref-13">13</xref>]. Jiang et al. [<xref ref-type="bibr" rid="ref-14">14</xref>] proposed Temporal Convolutional Networks (TCNs) for predicting landslides, which achieved higher prediction accuracy than traditional physical methods. Zhang et al. [<xref ref-type="bibr" rid="ref-15">15</xref>] confirmed that the Back Propagation (BP) neural network can predict regional landslide hazards with high accuracy. Among them, deep learning models, especially Long Short-Term Memory network (LSTM) models, outperform traditional methods in landslide prediction [<xref ref-type="bibr" rid="ref-16">16</xref>]. For example, the LSTM model outperforms the Support Vector Machine (SVM) [<xref ref-type="bibr" rid="ref-17">17</xref>,<xref ref-type="bibr" rid="ref-18">18</xref>], BP neural network [<xref ref-type="bibr" rid="ref-19">19</xref>], the Random Forest (RF), and the Autoregressive Integrated Moving Average (ARIMA) model [<xref ref-type="bibr" rid="ref-20">20</xref>]. In addition, the Bidirectional Long Short-Term Memory (BiLSTM) network has higher prediction accuracy compared to the LSTM [<xref ref-type="bibr" rid="ref-21">21</xref>]. Zhang et al. [<xref ref-type="bibr" rid="ref-22">22</xref>] proposed a DBi-LSTM neural network, which can effectively achieve stronger feature expression. Radial Basis Function (RBF) neural networks, as a shallow learning approach, offer distinct advantages for processing complex data with relatively small datasets. These networks typically demonstrate faster training times compared to other neural network architectures, which facilitate the rapid updating of models in real-time landslide monitoring systems and enhance prediction accuracy [<xref ref-type="bibr" rid="ref-23">23</xref>,<xref ref-type="bibr" rid="ref-24">24</xref>]. Furthermore, certain shapeless RBF have shown significant potential in pattern recognition applications [<xref ref-type="bibr" rid="ref-25">25</xref>]. Although introducing the attention mechanism to construct a BiLSTM combination model can effectively capture the correlation of input sequences [<xref ref-type="bibr" rid="ref-26">26</xref>,<xref ref-type="bibr" rid="ref-27">27</xref>], the global attention mechanism can comprehensively capture information [<xref ref-type="bibr" rid="ref-28">28</xref>]. Combining the global and local attention mechanisms can extract the global and local features of landslide displacement and more comprehensively capture complex spatio-temporal features [<xref ref-type="bibr" rid="ref-29">29</xref>]. In addition, Ge et al. [<xref ref-type="bibr" rid="ref-30">30</xref>] proposed a lightweight Transformer network that does not require manual feature selection, which significantly improving the learning efficiency of the model. Nevertheless, the weights of these attention mechanisms were not determined by directly utilizing the historical data on landslide displacement, which prevented them from making the most effective use of the data.</p>
<p>Existing studies quantify the uncertainty of landslide displacement prediction through Prediction Interval (PI) [<xref ref-type="bibr" rid="ref-31">31</xref>,<xref ref-type="bibr" rid="ref-32">32</xref>], and the change in their width can provide a reference for decision-making. However, relying solely on PI to quantify uncertainty makes it difficult to exploit its optimization potential during the model training process fully. To resolve the constraints of the current landslide monitoring data volume scarcity difficult to support model learning, model learning requires manual feature selection resulting in low prediction efficiency, and the model&#x2019;s adaptive feature extraction fails to capitalize on the historical information of landslides, this study proposes a BiLSTM-RBF landslide displacement prediction model with hybrid global-local attention enhancement, and the main contributions of this paper are as follows:
<list list-type="bullet">
<list-item>
<p>By using the weighted hidden states of BiLSTM to drive the RBF network, deep feature extraction is accomplished without the requirement for human feature selection. The raw time series data are directly input.</p></list-item>
<list-item>
<p>Combining the global attention and bidirectional local attention mechanisms, the former captures the global features, and the latter innovatively adopts the historical displacement data to compute the forward and reverse attentional scores for the BiLSTM network, respectively, and realizes bidirectional extraction of periodic fluctuating features.</p></list-item>
<list-item>
<p>In order to prevent the model from being overfitted due to the insufficient sample size, a composite loss function is introduced, and the improved width of the PI is used as a penalty term of the loss function.</p></list-item>
</list></p>
<p>In this paper, the BiLSTM-RBF fusion model is validated by RBF, BiLSTM, and Gated Recurrent Unit (GRU) baseline models. First, perform the feature factor decomposition and partial reconstruction on these baseline models; then, use the Pearson correlation coefficient and the Maximum Information Coefficient (MIC) to construct a feature dataset for the baseline models. However, the results indicate that BiLSTM-RBF maintains the highest prediction efficiency, despite the significantly reduced sample size. This is in contrast to the baseline model, which necessitates feature selection and limits the model&#x2019;s ability to be dynamically adjusted during training.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Study Areas and Data</title>
<p>The Baishuihe landslide, situated on the right bank of the Three Gorges Reservoir area, exhibits geomorphic characteristics of a composite concave slope structure with elevated eastern and western flanks and a gently inclined central section. Following significant deformation events during the June 2003 flood season, it was designated as a key monitoring target for geological hazards. Monitoring data from 2003 to 2012 reveal that the geomorphic evolution progressed sequentially through three distinct phases: a stable deformation stage, an accelerated deformation stage, and a slow deformation stage, each demonstrating markedly different displacement rates. The deformation characteristics and evolutionary process show distinct phased manifestations [<xref ref-type="bibr" rid="ref-33">33</xref>]. The synergistic interaction between reservoir water level fluctuations and intense rainfall events serves as the key controlling factor driving the staged evolution of landslide displacement rates.</p>
<p>This study utilizes observational data from monitoring stations ZG118 and ZG93 (selected from 11 GPS stations within the landslide area) spanning 2006&#x2013;2012 (see <xref ref-type="table" rid="table-1">Table 1</xref>, <xref ref-type="fig" rid="fig-1">Fig. 1</xref>), with variation trends of rainfall intensity, reservoir water level, and historical displacements illustrated in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>. (Basic characteristics and monitoring data (2006&#x2013;2012) of the Baishuihe landslide in Zigui County, Three Gorges Reservoir Area, Yangtze River from the Hubei Yangtze Three Gorges Landslide National Field Scientific Observation and Research Station).</p>
<table-wrap id="table-1">
<label>Table 1</label>
<caption>
<title>Overview of case study data for the Baishuihe</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Baishuihe</th>
<th>Monitoring time</th>
<th>Time step</th>
<th>Number of samples</th>
</tr>
</thead>
<tbody>
<tr>
<td>ZG118</td>
<td>2006/01&#x2013;2012/12</td>
<td>1 month</td>
<td>84</td>
</tr>
<tr>
<td>ZG93</td>
<td>2006/01&#x2013;2011/12</td>
<td>1 month</td>
<td>72</td>
</tr>
</tbody>
</table>
</table-wrap><fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Layout of monitoring sites for the Baishuihe landslide [<xref ref-type="bibr" rid="ref-34">34</xref>]</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-1.tif"/>
</fig><fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Changes in raw monitoring data for the Baishuihe landslide</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-2.tif"/>
</fig>
<p>As shown in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>, the displacement of the Baishuihe landslide increases monotonically with time and is seasonal. The obvious increase in displacement and the relative stabilization phase are from April to September and from October to April of the following year. These two time periods correspond to the time when the monthly cumulative rainfall is high and the reservoir level is low. Therefore, the change of landslide displacement has a periodicity of about one year, which is closely related to the seasonal heavy rainfall and fluctuating changes of the reservoir level.</p>
</sec>
<sec id="s3">
<label>3</label>
<title>Methodology</title>
<sec id="s3_1">
<label>3.1</label>
<title>Overall Working Framework</title>
<p>Decomposing the cumulative displacement into multiple components facilitates more accurate analysis and prediction. In this paper, we focus on the trend displacement and the periodic displacement, which are decomposed using Empirical Mode Decomposition (EMD) [<xref ref-type="bibr" rid="ref-35">35</xref>]. The maximum number of iterations for EMD is 500. Cubic spline interpolation is used, and the number of IMFs is not restricted. The endpoint effect is handled through mirror extension. The following are the stopping criteria:
<list list-type="simple">
<list-item><label>(1)</label><p>Main condition (normalized standard deviation threshold SD &#x003D; 0.05): When the SD values of two adjacent iterations are both below 0.05, the IMF is determined to have converged;</p></list-item>
<list-item><label>(2)</label><p>Auxiliary condition (normalized standard deviation threshold SD2 &#x003D; 0.5): If the single SD2 value suddenly drops below 0.5, do not stop temporarily until the main condition is met;</p></list-item>
<list-item><label>(3)</label><p>Global termination condition (residual energy TOL &#x003D; 0.05): When the residual energy is below 5% of the original signal energy, terminate the entire EMD process.</p></list-item>
</list></p>
<p>After decomposition, the component showing a trend change is extracted as the trend displacement, and the remaining components are reconstructed into the periodic displacement. These two types of displacement components are predicted separately and then superimposed in the time series to obtain the predicted result of the cumulative displacement. <xref ref-type="fig" rid="fig-3">Fig. 3</xref> shows the overall working framework of this study, and the main steps involved are as follows:</p>
<p><list list-type="bullet">
<list-item>
<p>Decompose cumulative displacement and its influencing factors, and construct a feature dataset for the baseline model using Pearson and MIC feature selection methods.</p></list-item>
<list-item>
<p>Use Non-dominated Sorting Genetic Algorithm-II (NSGA-II) to optimize <inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula> and <inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula> parameters in the Double Exponential Smoothing (DES) method for trend displacement prediction.</p></list-item>
<list-item>
<p>Design a bidirectional local attention mechanism based on historical periodic displacement calculation, and apply it to the model simultaneously with the global attention mechanism.</p></list-item>
<list-item>
<p>Construct a BiLSTM-RBF fusion model using the hidden state of BiLSTM.</p></list-item>
<list-item>
<p>Introduce the improved width of the PI as a penalty term in the loss function.</p></list-item>
<list-item>
<p>Conduct model evaluation, comparison, and analysis.</p></list-item>
</list></p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>Overall working framework diagram</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-3.tif"/>
</fig>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Data Preprocessing</title>
<p><list list-type="simple">
<list-item><label>(1)</label><p>Baseline model&#x2019;s data processing</p></list-item>
</list></p>
<p>Because landslide monitoring data are time-dependent, the monitoring time is added to the data set in a specific format, e.g., January 2006 as: 200601, December 2012: 201212. The related variables with displacement, rainfall, etc., are counted to obtain the original data set for the baseline model (see <xref ref-type="table" rid="table-2">Table 2</xref>). In addition, since the influencing factors (e.g., rainfall, reservoir level, etc.) also belong to seasonal external dynamics, their fluctuating changes may all have different degrees of influence on the periodic displacement. Therefore, this paper also uses EMD to decompose and partially reconstruct the monthly mean reservoir water level and monthly cumulative rainfall to obtain the extended dataset of periodic displacement.</p>
<table-wrap id="table-2">
<label>Table 2</label>
<caption>
<title>Original dataset for the baseline model</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Variable type</th>
<th>Feature index</th>
<th>Feature name</th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="10">Feature variable</td>
<td>1</td>
<td>Monitoring date</td>
</tr>
<tr>
<td>2</td>
<td><inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>X</mml:mi></mml:math></inline-formula> (mm)&#x002A;</td>
</tr>
<tr>
<td>3</td>
<td><inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>Y</mml:mi></mml:math></inline-formula> (mm)&#x002A;</td>
</tr>
<tr>
<td>4</td>
<td>Monthly variable of <inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>X</mml:mi></mml:math></inline-formula> (mm)</td>
</tr>
<tr>
<td>5</td>
<td>Monthly variable of <inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>Y</mml:mi></mml:math></inline-formula> (mm)</td>
</tr>
<tr>
<td>6</td>
<td>Monthly variable of <inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>F</mml:mi></mml:math></inline-formula> (mm)&#x002A;</td>
</tr>
<tr>
<td>7</td>
<td>Monthly cumulative rainfall (mm)</td>
</tr>
<tr>
<td>8</td>
<td>Monthly rainfall accumulation over the previous 2 months</td>
</tr>
<tr>
<td>9</td>
<td>Monthly rainfall accumulation during the initial 3 months</td>
</tr>
<tr>
<td>10</td>
<td>Monthly average reservoir level (m)</td>
</tr>
<tr>
<td>Target variable</td>
<td>24</td>
<td>Periodic displacement</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="table-2fn1" fn-type="other">
<p>Note: &#x002A;<inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>X</mml:mi></mml:math></inline-formula>, <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>Y</mml:mi></mml:math></inline-formula>, and <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>F</mml:mi></mml:math></inline-formula> are horizontal displacements in different directions. <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>F</mml:mi><mml:mo>=</mml:mo><mml:msqrt><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>X</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:mrow><mml:mi>Y</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:msqrt></mml:math></inline-formula>.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p><list list-type="simple">
<list-item><label>(2)</label><p>Data splitting</p>
</list-item>
</list></p>
<p>To explicitly model the temporal dependence of the landslide displacement sequence, before data splitting, the sliding-window recomposition strategy is first adopted to convert the original continuous time series into the &#x201C;input-output&#x201D; pairs required for supervised learning. That is, each sample takes the multi-dimensional features of the previous 2-time steps as the input and the displacement value of the next time step as the target. After recomposition, the data dimension is (number of samples, 2, number of features), which can be directly used for subsequent model training. The data splitting is as follows:</p>
<p><list list-type="simple">
<list-item><label>(a)</label><p>Data splitting method: Considering the time series characteristics of landslide displacement data, it is necessary to strictly follow the chronological order constraint of historical data for future prediction. Therefore, this paper adopts the standard holdout strategy [<xref ref-type="bibr" rid="ref-16">16</xref>] to split the data in chronological order. This can ensure the time coherence of the data during the model training process and avoid the false improvement of model performance caused by &#x201C;data leakage&#x201D;.</p></list-item>
<list-item><label>(b)</label><p>Data splitting ratio: Since this study decomposes the cumulative displacement into trend displacement and periodic displacement for separate prediction, and the prediction models for the two types of displacement are different, the data splitting ratio for the landslide monitoring points ZG118 and ZG93 are: Trend displacement: Training set: Test set &#x003D; 70%:30%; Periodic displacement: Training set: Validation set: Test set &#x003D; 80%:10%:10%.</p></list-item>
</list></p>
<p>In addition, MinMax normalization is adopted to ensure the quality and consistency of input data.</p>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Trend Displacement Prediction Model</title>
<sec id="s3_3_1">
<label>3.3.1</label>
<title>DES Method</title>
<p>In this paper, we choose the DES method that is adaptive to a small number of data samples, especially when the data has a trend, and can adapt to the changes in the time series data. The DES method, a time series forecasting technique, is particularly suitable for data with a trend but without seasonality. This method separately predicts level changes and trend changes using two smoothing equations. The updated formulas for the level and trend equations are shown in <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref>:
<disp-formula id="eqn-1"><label>(1)</label><mml:math id="mml-eqn-1" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<disp-formula id="eqn-2"><label>(2)</label><mml:math id="mml-eqn-2" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>the level component at time <italic>t</italic> is denoted as <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>, <inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula> represents the smoothing coefficient for the level, <inline-formula id="ieqn-14"><mml:math id="mml-ieqn-14"><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mspace width="0pt" /></mml:mrow></mml:msub></mml:math></inline-formula> corresponds to the actual observation at time <inline-formula id="ieqn-15"><mml:math id="mml-ieqn-15"><mml:mi>t</mml:mi></mml:math></inline-formula>, and <inline-formula id="ieqn-16"><mml:math id="mml-ieqn-16"><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula> indicates the trend component from the previous period. The trend component at time <italic>t</italic> is calculated as <inline-formula id="ieqn-17"><mml:math id="mml-ieqn-17"><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mspace width="0pt" /></mml:math></inline-formula>, with <inline-formula id="ieqn-18"><mml:math id="mml-ieqn-18"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula> being the smoothing coefficient for the trend. The forecasting formula for the test set is expressed as <xref ref-type="disp-formula" rid="eqn-3">Eq. (3)</xref>:
<disp-formula id="eqn-3"><label>(3)</label><mml:math id="mml-eqn-3" display="block"><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:math></disp-formula>where, <inline-formula id="ieqn-19"><mml:math id="mml-ieqn-19"><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula> is the predicted value of the test set.</p>
</sec>
<sec id="s3_3_2">
<label>3.3.2</label>
<title>Solving Trend Displacement Multi-Objective Optimization Problems Based on NSGA-II</title>
<p>NSGA-II is a multi-objective genetic algorithm proposed by Deb et al. in 2002. In multi-objective optimization, NSGA-II does not rely on explicit weighting or weighted combinations after normalization. Instead, it achieves an &#x201C;implicit balance&#x201D; among multiple objectives through the Pareto dominance relationship. As shown in <xref ref-type="fig" rid="fig-4">Fig. 4a</xref>, for any two solutions A and B, if both the RMSE and MAE of A are smaller than those of B, and the R<sup>2</sup> of A is greater than that of B, then A strictly dominates B, and B will be eliminated by the algorithm. If A is better in some indicators (e.g., lower RMSE) but worse in other indicators (e.g., higher MAE), then A and B do not dominate each other and will be jointly retained in the Pareto front. Solutions like A and B are Pareto optimal solutions. The goal of the multi-objective optimization algorithm is to find these Pareto optimal solutions [<xref ref-type="bibr" rid="ref-36">36</xref>].</p>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Main schematic diagram of NSGA-II. (<bold>a</bold>) Pareto optimal solution schematic; (<bold>b</bold>) Crowding distance comparison</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-4.tif"/>
</fig>
<p>As shown in <xref ref-type="fig" rid="fig-4">Fig. 4b</xref>, when combining the generated offspring population with the parent population to form a new population, in addition to considering the non-dominated rank, the crowding distance also needs to be considered within each non-dominated layer. Each objective function will be normalized before calculating the crowding distance. The calculation of all crowding distance values is the sum of the absolute values of the differences between adjacent individuals of each individual across all objective functions. Individuals with larger crowding distances are more likely to be retained, such as <inline-formula id="ieqn-20"><mml:math id="mml-ieqn-20"><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula>, <inline-formula id="ieqn-21"><mml:math id="mml-ieqn-21"><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula>, and the trimmed <inline-formula id="ieqn-22"><mml:math id="mml-ieqn-22"><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula>. This can prevent the individuals in the population from being overly concentrated in the objective space, ensuring that the algorithm evenly searches for the optimal solution among multiple objectives, thereby achieving a balance among multiple objectives.</p>
<p><bold>Solving for the Pareto optimal solution for trend displacement</bold>. Using NSGA-II to optimize the variables <inline-formula id="ieqn-23"><mml:math id="mml-ieqn-23"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula> and <inline-formula id="ieqn-24"><mml:math id="mml-ieqn-24"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula> of the DES method. To enhance the model&#x2019;s performance on new datasets, objective functions are established to minimize three evaluation metrics on the test set: <inline-formula id="ieqn-25"><mml:math id="mml-ieqn-25"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>R</mml:mi><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi></mml:math></inline-formula>, <inline-formula id="ieqn-26"><mml:math id="mml-ieqn-26"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:mi>E</mml:mi></mml:math></inline-formula>, and <inline-formula id="ieqn-27"><mml:math id="mml-ieqn-27"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:msup><mml:mi>R</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. The decision variables are the two parameters, <inline-formula id="ieqn-28"><mml:math id="mml-ieqn-28"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula> and <inline-formula id="ieqn-29"><mml:math id="mml-ieqn-29"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula>, in the DES model. The multi-objective problem is described as follows:
<disp-formula id="eqn-4"><label>(4)</label><mml:math id="mml-eqn-4" display="block"><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign="left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mi>O</mml:mi><mml:mi>b</mml:mi><mml:mi>j</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>v</mml:mi><mml:msub><mml:mi>e</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>O</mml:mi><mml:mi>b</mml:mi><mml:mi>j</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>v</mml:mi><mml:msub><mml:mi>e</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>O</mml:mi><mml:mi>b</mml:mi><mml:mi>j</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>v</mml:mi><mml:msub><mml:mi>e</mml:mi><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable><mml:mo fence="true" stretchy="true" symmetric="true"></mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-5"><label>(5)</label><mml:math id="mml-eqn-5" display="block"><mml:mi>s</mml:mi><mml:mo>.</mml:mo><mml:mi>t</mml:mi><mml:mo>.</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign="left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mo>&#x003C;</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>&#x003C;</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mo>&#x003C;</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>&#x003C;</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr></mml:mtable><mml:mo fence="true" stretchy="true" symmetric="true"></mml:mo></mml:mrow></mml:math></disp-formula>where, &#x201C;<inline-formula id="ieqn-30"><mml:math id="mml-ieqn-30"><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:math></inline-formula>&#x201D; means minimizing the function,
<disp-formula id="eqn-6"><label>(6)</label><mml:math id="mml-eqn-6" display="block"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>R</mml:mi><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:msqrt><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:munderover><mml:mo>&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:msqrt><mml:mo>=</mml:mo><mml:msqrt><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:munderover><mml:mo>&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:msqrt><mml:mo>,</mml:mo></mml:math></disp-formula>
<disp-formula id="eqn-7"><label>(7)</label><mml:math id="mml-eqn-7" display="block"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:munderover><mml:mo>&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:munderover><mml:mo>&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:math></disp-formula>
<disp-formula id="eqn-8"><label>(8)</label><mml:math id="mml-eqn-8" display="block"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:msup><mml:mi>R</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mfrac><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mover><mml:mi>Y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mfrac><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mover><mml:mi>Y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:math></disp-formula>where <inline-formula id="ieqn-31"><mml:math id="mml-ieqn-31"><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the actual value, <inline-formula id="ieqn-32"><mml:math id="mml-ieqn-32"><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>+</mml:mo><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the predicted value, <inline-formula id="ieqn-33"><mml:math id="mml-ieqn-33"><mml:mover><mml:mi>Y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover></mml:math></inline-formula> is the mean of the actual values, and <inline-formula id="ieqn-34"><mml:math id="mml-ieqn-34"><mml:mi>n</mml:mi></mml:math></inline-formula> is the sample size, all of which belong to the test set. The hyperparameter settings for NSGA-II are shown in <xref ref-type="table" rid="table-3">Table 3</xref>.</p>
<table-wrap id="table-3">
<label>Table 3</label>
<caption>
<title>NSGA-II hyperparameter configuration</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Parameter name</th>
<th>Values range</th>
</tr>
</thead>
<tbody>
<tr>
<td>Initial population size</td>
<td>30</td>
</tr>
<tr>
<td>Number of iterations</td>
<td>100</td>
</tr>
<tr>
<td>Crossover probability</td>
<td>0.3</td>
</tr>
<tr>
<td>Mutation ratio</td>
<td>0.7</td>
</tr>
<tr>
<td>Individual mutation probability</td>
<td>0.2</td>
</tr>
<tr>
<td>Individual mutation range</td>
<td>(1 &#x00D7; 10<sup>&#x2212;6</sup>, 0.999999)</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec id="s3_4">
<label>3.4</label>
<title>Periodic Displacement Prediction Model</title>
<p>The periodic displacement prediction employs a BiLSTM-RBF fusion model enhanced by a global-local hybrid attention mechanism. Specifically, the rate of change in historical periodic landslide displacement is utilized as weights for the bidirectional local attention mechanism. Additionally, the width of the PI serves as a penalty term in the loss function.</p>
<sec id="s3_4_1">
<label>3.4.1</label>
<title>Baseline Models for Periodic Displacement Prediction</title>
<p>By manually selecting features for three baseline models (RBF neural network, BiLSTM neural network, GRU neural network), a feature dataset for periodic displacement is constructed to prepare for the model validation of the proposed network, which does not require manual feature selection.</p>
<p><list list-type="simple">
<list-item><label>(1)</label><p>Feature selection methods</p></list-item>
</list></p>
<p>In order to enhance the baseline model&#x2019;s prediction accuracy, we evaluated the linear and dependency relationships between influencing factors and periodic displacements in the extended dataset using the MIC and Pearson correlation coefficient [<xref ref-type="bibr" rid="ref-37">37</xref>]. As the final dataset for periodic displacements, factors with Pearson&#x2019;s correlation coefficient absolute values and MIC values of at least 0.12 and 0.4 were chosen.
<list list-type="simple">
<list-item><label>(2)</label><p>BiLSTM/GRU/RBF network architecture</p></list-item></list></p>
<p><bold>BiLSTM/GRU neural network.</bold> BiLSTM networks, which are generated from enhancements to LSTM networks. In BiLSTM, the LSTM network is constructed as two oppositely oriented LSTM layers (<xref ref-type="fig" rid="fig-5">Fig. 5a</xref>), which deal with the forward and reverse directions of the sequence data. LSTM is a model for optimizing Recurrent Neural Networks (RNN) [<xref ref-type="bibr" rid="ref-38">38</xref>], with a total of three computational mechanisms, namely, forgetting gates, input gates, and output gates (<xref ref-type="fig" rid="fig-5">Fig. 5b</xref>). Each LSTM recurrent unit has an external state of the previous moment and the current moment&#x2019;s input as the unit input. The internal structure of GRU is more concise than LSTM (<xref ref-type="fig" rid="fig-5">Fig. 5c</xref>). It consists of an update gate that controls the retention of old information and the introduction of new information, and a reset gate that determines how much of the old information has been forgotten (<xref ref-type="fig" rid="fig-5">Fig. 5d</xref>).</p>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>(<bold>a</bold>) BiLSTM neural network structure; (<bold>b</bold>) LSTM recurrent unit internal structure; (<bold>c</bold>) GRU neural network structure; (<bold>d</bold>) GRU recurrent unit internal structure</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-5.tif"/>
</fig>
<p><bold>RBF neural network.</bold> The RBF neural network is a 3-layer feed-forward network with a single hidden layer and its action function is a Gaussian basis function [<xref ref-type="bibr" rid="ref-39">39</xref>]. The <italic>j</italic>-th neuron of the hidden layer is calculated in the following manner:
<disp-formula id="eqn-9"><label>(9)</label><mml:math id="mml-eqn-9" display="block"><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>exp</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mfrac><mml:mrow><mml:mo fence="false" stretchy="false">&#x2016;</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:msup><mml:mo fence="false" stretchy="false">&#x2016;</mml:mo><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>2</mml:mn><mml:msubsup><mml:mi>b</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo><mml:mspace width="thinmathspace" /><mml:mspace width="thinmathspace" /><mml:mspace width="thinmathspace" /><mml:mspace width="thinmathspace" /><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo></mml:math></disp-formula>where, <inline-formula id="ieqn-35"><mml:math id="mml-ieqn-35"><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mi>j</mml:mi><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mi>j</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>]</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula id="ieqn-36"><mml:math id="mml-ieqn-36"><mml:mi>b</mml:mi><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mo>[</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>]</mml:mo></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> represent the center vector and the width vector of the <italic>j</italic>-th hidden layer neuron of the RBF network. The output of the RBF neural network is given by <xref ref-type="disp-formula" rid="eqn-10">Eq. (10)</xref>.
<disp-formula id="eqn-10"><label>(10)</label><mml:math id="mml-eqn-10" display="block"><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msub><mml:mi>&#x03C9;</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>&#x03C9;</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>+</mml:mo><mml:msub><mml:mi>&#x03C9;</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:math></disp-formula></p>
</sec>
<sec id="s3_4_2">
<label>3.4.2</label>
<title>BiLSTM-RBF-Attention for Periodic Displacement Prediction</title>
<p><list list-type="simple">
<list-item><label>(1)</label><p>Attention mechanism</p></list-item>
</list></p>
<p><bold>Global Attention Mechanism.</bold> The global attention technique is implemented by using the popular Keras core layer (Dense). The specific steps include: input transformation, weight calculation, normalization, and application of weights. The Dense layer takes all the transformed hidden states of BiLSTM as input, calculates a raw attention score for each time step, and then converts it into normalized attention weights through the softmax function. This approach dynamically generates independent attention weights <inline-formula id="ieqn-37"><mml:math id="mml-ieqn-37"><mml:mi>G</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> for the hidden states at each time step by attending to the entire input sequence data.</p>
<p><bold>Local Attention Mechanism.</bold> As shown in the shaded part of <xref ref-type="fig" rid="fig-6">Fig. 6a</xref>,<xref ref-type="fig" rid="fig-6">b</xref>, the cumulative displacement and periodic displacement of monitoring stations ZG118 and ZG93 synchronously show an upward trend, and the change is obvious. While the cumulative displacement maintains a relatively stable stage, the periodic displacement shows a decreasing trend. From this analysis, the short-term rapid increase of landslide cumulative displacement primarily originates from the increase in periodic displacement, so we must draw the model&#x2019;s attention to the increasing trend while predicting the periodic displacement. As per the principle of change, this paper proposes a two-way local attention mechanism applied to BiLSTM, designing local attention weights for the forward and reverse sequences of the periodic displacement.</p>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Cumulative displacement decomposition diagram. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-6.tif"/>
</fig>
<p>(a) Window size for the local attention mechanism</p>
<p>The window size of the local attention mechanism is fixed to 2-time steps (forward sequence: focus only on the current month displacement and the preceding month displacement in the forward periodic displacement sequence; reverse sequence: focus only on the current month displacement and the next month displacement in the reverse periodic displacement sequence), and the window sliding step size is one month.</p>
<p>(b) Weights of the local attention mechanism</p>
<p>Since the periodic displacement shows a notable increasing trend in the short term and to prevent data leakage, the historical rate of change of the periodic displacement is adopted as the local attention weights. To convert the final attention scores to the range between 0 and 1 and focus the local attention mechanism on periodic displacement sequences with an increasing trend, ending with the application of the sigmoid function. The attention weights for the forward and reverse periodic displacement sequences are denoted as <inline-formula id="ieqn-38"><mml:math id="mml-ieqn-38"><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-39"><mml:math id="mml-ieqn-39"><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>, as shown in <xref ref-type="disp-formula" rid="eqn-11">Eqs. (11)</xref> and <xref ref-type="disp-formula" rid="eqn-12">(12)</xref>:
<disp-formula id="eqn-11"><label>(11)</label><mml:math id="mml-eqn-11" display="block"><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign="left left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mo>,</mml:mo></mml:mtd><mml:mtd><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mtext>sigmoid</mml:mtext></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mspace width="negativethinmathspace" /><mml:mrow><mml:mo>/</mml:mo></mml:mrow><mml:mn>1</mml:mn><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd><mml:mtd><mml:mi>i</mml:mi><mml:mo>&#x2260;</mml:mo><mml:mn>0</mml:mn></mml:mtd></mml:mtr></mml:mtable><mml:mo fence="true" stretchy="true" symmetric="true"></mml:mo></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:mi>n</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-12"><label>(12)</label><mml:math id="mml-eqn-12" display="block"><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign="left left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mo>,</mml:mo></mml:mtd><mml:mtd><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mi>n</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mtext>sigmoid</mml:mtext></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mspace width="negativethinmathspace" /><mml:mrow><mml:mo>/</mml:mo></mml:mrow><mml:mn>1</mml:mn><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd><mml:mtd><mml:mi>i</mml:mi><mml:mo>&#x2260;</mml:mo><mml:mi>n</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr></mml:mtable><mml:mo fence="true" stretchy="true" symmetric="true"></mml:mo></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:mi>n</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>where, <inline-formula id="ieqn-40"><mml:math id="mml-ieqn-40"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-41"><mml:math id="mml-ieqn-41"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> represent the forward and reverse periodic displacements at the time step <inline-formula id="ieqn-42"><mml:math id="mml-ieqn-42"><mml:mi>i</mml:mi></mml:math></inline-formula>.</p>
<p>Apply the forward attention weights to the hidden states of the forward LSTM layer and the reverse attention weights to the hidden states of the reverse LSTM layer (<xref ref-type="disp-formula" rid="eqn-13">Eqs. (13)</xref> and <xref ref-type="disp-formula" rid="eqn-14">(14)</xref>). Then, concatenate the weighted hidden states of the forward and reverse directions at corresponding time steps to obtain the hidden states of the BiLSTM after the application of bidirectional local attention weights, as shown in <xref ref-type="disp-formula" rid="eqn-15">Eq. (15)</xref>, <xref ref-type="fig" rid="fig-7">Fig. 7</xref>:<disp-formula id="eqn-13"><label>(13)</label><mml:math id="mml-eqn-13" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<disp-formula id="eqn-14"><label>(14)</label><mml:math id="mml-eqn-14" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<disp-formula id="eqn-15"><label>(15)</label><mml:math id="mml-eqn-15" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:mi>L</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>F</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2295;</mml:mo><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
<fig id="fig-7">
<label>Figure 7</label>
<caption>
<title>Schematic diagram of bidirectional local attention mechanism</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-7.tif"/>
</fig>
<p>In summary, the global-local hybrid attention mechanism is simultaneously applied to the weighted concatenated hidden states following the BiLSTM hidden states, as shown in <xref ref-type="disp-formula" rid="eqn-16">Eq. (16)</xref>:
<disp-formula id="eqn-16"><label>(16)</label><mml:math id="mml-eqn-16" display="block"><mml:mi>b</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>G</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:mi>L</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></disp-formula></p>
<p><list list-type="simple">
<list-item><label>(2)</label><p>Loss function</p></list-item>
</list></p>
<p>The PI is an important concept in statistics, which provides a range to estimate the possible values of future observations. To construct a PI, the mean of the predicted values is used as the center of the interval, and the standard deviation of the residuals is considered to reflect the variability of the data. In the study, the PI is built on the foundation of the normal distribution, with a confidence level set at 95%. The mathematical formula for the PI is <xref ref-type="disp-formula" rid="eqn-17">Eq. (17)</xref>:
<disp-formula id="eqn-17"><label>(17)</label><mml:math id="mml-eqn-17" display="block"><mml:mi>P</mml:mi><mml:mi>I</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>&#x00B1;</mml:mo><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:math></disp-formula>where, <inline-formula id="ieqn-43"><mml:math id="mml-ieqn-43"><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula> is the model&#x2019;s predicted value, <inline-formula id="ieqn-44"><mml:math id="mml-ieqn-44"><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the critical value of the normal distribution, and <inline-formula id="ieqn-45"><mml:math id="mml-ieqn-45"><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the standard error of the prediction. From <xref ref-type="disp-formula" rid="eqn-17">Eq. (17)</xref>, the width of the PI <inline-formula id="ieqn-46"><mml:math id="mml-ieqn-46"><mml:mi>W</mml:mi><mml:mi>P</mml:mi><mml:msup><mml:mi>I</mml:mi><mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> is derived as:
<disp-formula id="eqn-18"><label>(18)</label><mml:math id="mml-eqn-18" display="block"><mml:mi>W</mml:mi><mml:mi>P</mml:mi><mml:msup><mml:mi>I</mml:mi><mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>+</mml:mo><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:math></disp-formula></p>
<p>Construct a loss function that uses the width of the PI as a penalty term, to penalize the width when the model&#x2019;s predicted value deviates from the observed value. Therefore, this paper replaces the model&#x2019;s predicted standard error <inline-formula id="ieqn-47"><mml:math id="mml-ieqn-47"><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> with the batch Mean Absolute Error (<inline-formula id="ieqn-48"><mml:math id="mml-ieqn-48"><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:msub><mml:mi>E</mml:mi><mml:mrow><mml:mi>b</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>c</mml:mi><mml:mi>h</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>), and the improved width of the PI is:
<disp-formula id="eqn-19"><label>(19)</label><mml:math id="mml-eqn-19" display="block"><mml:mi>W</mml:mi><mml:mi>P</mml:mi><mml:mi>I</mml:mi><mml:mo>=</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>Z</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:msub><mml:mi>E</mml:mi><mml:mrow><mml:mi>b</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>c</mml:mi><mml:mi>h</mml:mi></mml:mrow></mml:msub></mml:math></disp-formula></p>
<p>The final loss function is as in <xref ref-type="disp-formula" rid="eqn-20">Eq. (20)</xref>:<disp-formula id="eqn-20"><label>(20)</label><mml:math id="mml-eqn-20" display="block"><mml:mi>L</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>s</mml:mi><mml:mo>=</mml:mo><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi><mml:mo>+</mml:mo><mml:mi>&#x03BB;</mml:mi><mml:mo>&#x22C5;</mml:mo><mml:mi>W</mml:mi><mml:mi>P</mml:mi><mml:mi>I</mml:mi><mml:mspace width="thinmathspace" /><mml:mo>,</mml:mo></mml:math></disp-formula>where, <inline-formula id="ieqn-49"><mml:math id="mml-ieqn-49"><mml:mi>&#x03BB;</mml:mi></mml:math></inline-formula> is a hyperparameter that balances the weights of the <inline-formula id="ieqn-50"><mml:math id="mml-ieqn-50"><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi></mml:math></inline-formula> and <inline-formula id="ieqn-51"><mml:math id="mml-ieqn-51"><mml:mi>W</mml:mi><mml:mi>P</mml:mi><mml:mi>I</mml:mi></mml:math></inline-formula>.
<list list-type="simple">
<list-item><label>(3)</label>
<p>Periodic displacement prediction based on BiLSTM-RBF-Attention</p></list-item>
</list></p>
<p><xref ref-type="fig" rid="fig-8">Fig. 8</xref> shows the structure of the proposed model, which directly takes the original time series data as its input (ZG118: reservoir level and historical displacement; ZG93: monthly cumulative rainfall, reservoir level, and historical displacement). All of the weighted hidden states <inline-formula id="ieqn-52"><mml:math id="mml-ieqn-52"><mml:mi>b</mml:mi><mml:msub><mml:mi>h</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> of BiLSTM are used as inputs to the RBF neural network. The &#x201C;expand_dims&#x201D; operation in TensorFlow is used to add the necessary dimensions to the input of the RBF to integrate the weighted hidden states of the BiLSTM. The specific process includes: First, obtain the weighted hidden states of the BiLSTM as the input tensor &#x201C;inputs&#x201D;, whose shape is (<italic>batch_size</italic>, <italic>time_steps</italic>, <italic>2LSTM_units</italic>). Second, add a single dimension after the feature dimension through &#x201C;tf.expand_dims (inputs, -1)&#x201D; to convert it to &#x201C;(<italic>batch_size</italic>, <italic>time_steps</italic>, <italic>2LSTM_units</italic>, 1)&#x201D;. Then, broadcast the RBF center tensor &#x201C;centers&#x201D; with the shape of &#x201C;(<italic>2LSTM_units</italic>, <italic>RBF_neurons</italic>)&#x201D; to match the shape of the input tensor, broadcasting it to (1, 1, <italic>2LSTM_units</italic>, <italic>RBF_neurons</italic>). After broadcasting, the shapes of the two are compatible, enabling element-wise difference calculation. By using the method of taking the output of BiLSTM as the input of RBF, the effective integration of the two models is achieved.</p>
<fig id="fig-8">
<label>Figure 8</label>
<caption>
<title>Structure of BiLSTM-RBF-Attention model</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-8.tif"/>
</fig>
<p><bold>Model parameter settings:</bold> the model was trained in CPU mode on Intel (R) Iris (R) Xe Graphics integrated graphics card, implemented using TensorFlow 2.15.0 and keras framework, and the programming language was Python (Python 3.9). After experimental testing, the training parameters of BiLSTM-RBF-Attention were finally set as follows: <italic>epochs</italic> &#x003D; 200, <italic>learning rate</italic> (Adam optimizer) &#x003D; 0.001, <italic>batch</italic> &#x003D; 2, <italic>LSTM_units (ZG118)</italic> &#x003D; 64, <italic>LSTM_units (ZG93)</italic> &#x003D; 32, <italic>RBF_units (ZG118)</italic> &#x003D; 96, <italic>RBF_units (ZG93)</italic> &#x003D; 64, and the loss function penalty coefficients <inline-formula id="ieqn-53"><mml:math id="mml-ieqn-53"><mml:mi>&#x03BB;</mml:mi></mml:math></inline-formula> are <inline-formula id="ieqn-54"><mml:math id="mml-ieqn-54"><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>5</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and <inline-formula id="ieqn-55"><mml:math id="mml-ieqn-55"><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. Setting independent random seeds for the data characteristics of different landslide monitoring points, at ZG118: (NumPy: 0, TensorFlow: 1); at ZG93: (NumPy: 6, TensorFlow: 8).</p>
</sec>
</sec>
<sec id="s3_5">
<label>3.5</label>
<title>Evaluation Metrics</title>
<p>In this paper, the evaluation metrics shown in <xref ref-type="table" rid="table-4">Table 4</xref> are selected to compare and analyze the landslide displacement prediction models objectively. Model predictive fidelity exhibits an inverse proportionality to residual error magnitudes, as quantified by diminishing root mean square error (RMSE) and mean absolute error (MAE) values. The coefficient of determination (R&#x00B2;), bounded between 0 and 1, serves as a critical diagnostic metric, with values closer to 1 signifying enhanced congruence between simulated surface deformation and observational datasets.</p>
<table-wrap id="table-4">
<label>Table 4</label>
<caption>
<title>Model evaluation metrics</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Metrics type</th>
<th>Metrics definition</th>
</tr>
</thead>
<tbody>
<tr>
<td>RMSE</td>
<td><inline-formula id="ieqn-56"><mml:math id="mml-ieqn-56"><mml:mi>R</mml:mi><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:msqrt><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:msqrt></mml:math></inline-formula></td>
</tr>
<tr>
<td>MAE</td>
<td><inline-formula id="ieqn-57"><mml:math id="mml-ieqn-57"><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac></mml:mstyle><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow></mml:math></inline-formula></td>
</tr>
<tr>
<td>R<sup>2</sup></td>
<td><inline-formula id="ieqn-58"><mml:math id="mml-ieqn-58"><mml:msup><mml:mi>R</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x2212;</mml:mo><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mfrac><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:mover><mml:mi>y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula></td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="table-4fn1" fn-type="other">
<p>Note: where, <inline-formula id="ieqn-59"><mml:math id="mml-ieqn-59"><mml:mi>n</mml:mi></mml:math></inline-formula>: sample size; <inline-formula id="ieqn-60"><mml:math id="mml-ieqn-60"><mml:msub><mml:mi>y</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>: the <italic>i</italic>-th observed value; <inline-formula id="ieqn-61"><mml:math id="mml-ieqn-61"><mml:msub><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>: the <italic>i</italic>-th predicted value; <inline-formula id="ieqn-62"><mml:math id="mml-ieqn-62"><mml:mover><mml:mi>y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover></mml:math></inline-formula>: the average of observed values.</p>
</fn>
</table-wrap-foot>
</table-wrap>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Results</title>
<p><xref ref-type="fig" rid="fig-6">Fig. 6</xref> illustrates the cumulative displacement decomposition results of the Baishuihe landslide monitoring points ZG118 and ZG93. After EMD is completed, the mean values of the residuals of ZG118 and ZG93 are <inline-formula id="ieqn-63"><mml:math id="mml-ieqn-63"><mml:mn>6.16</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>15</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and <inline-formula id="ieqn-64"><mml:math id="mml-ieqn-64"><mml:mo>&#x2212;</mml:mo><mml:mn>1.92</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>14</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> respectively, and the slopes of the linear trends are <inline-formula id="ieqn-65"><mml:math id="mml-ieqn-65"><mml:mo>&#x2212;</mml:mo><mml:mn>3.88</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>16</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> and <inline-formula id="ieqn-66"><mml:math id="mml-ieqn-66"><mml:mo>&#x2212;</mml:mo><mml:mn>8.20</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>17</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, respectively. Moreover, the histograms show a symmetric distribution (<xref ref-type="fig" rid="fig-9">Fig. 9a</xref>,<xref ref-type="fig" rid="fig-9">b</xref>), which meets the theoretical requirements of EMD for zero-mean and trend-free residuals, verifying the completeness of the decomposition.</p>
<fig id="fig-9">
<label>Figure 9</label>
<caption>
<title>Residual graph and residual distribution graph based on EMD. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-9.tif"/>
</fig>
<p>In <xref ref-type="fig" rid="fig-6">Fig. 6</xref>, the shaded area represents the change rule of the cumulative displacement, which is characterized by a distinct upward trend as the periodic displacement increases. This is equivalent to the calculation principle of <xref ref-type="disp-formula" rid="eqn-11">Eqs. (11)</xref> and <xref ref-type="disp-formula" rid="eqn-12">(12)</xref>.</p>
<sec id="s4_1">
<label>4.1</label>
<title>Trend Displacement Prediction Results</title>
<p>From the obtained Pareto front (<xref ref-type="fig" rid="fig-10">Fig. 10a</xref>,<xref ref-type="fig" rid="fig-10">b</xref>), the most suitable Pareto solution (highlighted in red) was selected by comprehensively evaluating the performance of RMSE, MAE, and R&#x00B2; on both the training and test sets. The Pareto decision variables for ZG118 and ZG93 are (<inline-formula id="ieqn-67"><mml:math id="mml-ieqn-67"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula>: 0.96637, <inline-formula id="ieqn-68"><mml:math id="mml-ieqn-68"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula>: 0.00035) and (<inline-formula id="ieqn-69"><mml:math id="mml-ieqn-69"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula>: 0.86971, <inline-formula id="ieqn-70"><mml:math id="mml-ieqn-70"><mml:mi>&#x03B2;</mml:mi></mml:math></inline-formula>: 0.00052), respectively.</p>
<fig id="fig-10">
<label>Figure 10</label>
<caption>
<title>Pareto front of DES. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-10.tif"/>
</fig>
<p><xref ref-type="fig" rid="fig-11">Fig. 11a</xref>,<xref ref-type="fig" rid="fig-11">b</xref> displays the trend displacement prediction results for the Baishuihe landslide at monitoring locations ZG118 and ZG93. The Pareto optimal solution produced RMSE, MAE, and R<sup>2</sup> values of 0.35, 0.29 and 0.99 mm for ZG118; for ZG93, the corresponding metrics were 1.02, 0.93 and 0.99 mm.</p>
<fig id="fig-11">
<label>Figure 11</label>
<caption>
<title>Results of the trend displacement prediction. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-11.tif"/>
</fig>
</sec>
<sec id="s4_2">
<label>4.2</label>
<title>Periodic Displacement Prediction Results</title>
<sec id="s4_2_1">
<label>4.2.1</label>
<title>Feature Selection for Baseline Models</title>
<p><xref ref-type="fig" rid="fig-12">Fig. 12a</xref>,<xref ref-type="fig" rid="fig-12">b</xref> shows the decomposition results of the reservoir level and monthly cumulative rainfall. The extended dataset for the baseline model is obtained from these results (see <xref ref-type="table" rid="table-5">Table 5</xref>).</p>
<fig id="fig-12">
<label>Figure 12</label>
<caption>
<title>Decomposition and partial reconstruction results of features. (<bold>a</bold>) Reservoir level; (<bold>b</bold>) monthly cumulative rainfall</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-12.tif"/>
</fig><table-wrap id="table-5">
<label>Table 5</label>
<caption>
<title>Extended dataset for the baseline model</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
</colgroup>
<thead>
<tr>
<th align="center">Variable type</th>
<th align="center">Feature index</th>
<th align="center">Feature name</th>
</tr>
</thead>
<tbody>
<tr>
<td></td>
<td>1&#x223C;10</td>
<td>See <xref ref-type="table" rid="table-2">Table 2</xref></td>
</tr>
<tr>
<td/>
<td>11</td>
<td>Reservoir level IMF1</td>
</tr>
<tr>
<td/>
<td>12</td>
<td>Reservoir level IMF2</td>
</tr>
<tr>
<td/>
<td>13</td>
<td>Reservoir level IMF3</td>
</tr>
<tr>
<td/>
<td>14</td>
<td>Reservoir level IMF (1, 2, 3) reconstruction</td>
</tr>
<tr>
<td/>
<td>15</td>
<td>Reservoir level IMF (1, 2) reconstruction</td>
</tr>
<tr>
<td/>
<td>16</td>
<td>Reservoir level IMF (2, 3) reconstruction</td>
</tr>
<tr>
<td>Feature variable</td>
<td>17</td>
<td>Reservoir level IMF4</td>
</tr>
<tr>
<td/>
<td>18</td>
<td>Rainfall IMF1</td>
</tr>
<tr>
<td/>
<td>19</td>
<td>Rainfall IMF2</td>
</tr>
<tr>
<td/>
<td>20</td>
<td>Rainfall IMF3</td>
</tr>
<tr>
<td/>
<td>21</td>
<td>Rainfall IMF4</td>
</tr>
<tr>
<td/>
<td>22</td>
<td>Rainfall IMF (1, 2, 3, 4) reconstruction</td>
</tr>
<tr>
<td/>
<td>23</td>
<td>Rainfall IMF5</td>
</tr>
<tr>
<td>Target variable</td>
<td>24</td>
<td>Periodic displacement</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The feature selection processes using Pearson and MIC are presented in <xref ref-type="fig" rid="fig-13">Figs. 13a</xref>,<xref ref-type="fig" rid="fig-13">b</xref> and <xref ref-type="fig" rid="fig-14">14a</xref>,<xref ref-type="fig" rid="fig-14">b</xref> for the extended datasets for monitoring points ZG118 and ZG93. In summary, the final dataset index for monitoring point ZG118 were determined as Pearson-selected features: 9, 15, 12, 6, 18, 13 and MIC-selected features: 3, 2, 23, 17, 1, 10, 13. while for ZG93, the corresponding index were Pearson: 5, 4, 6, 2, 3, 9, 21, 22, 7, 18, 20, 13 and MIC: 17, 13. Consequently, the surface displacement and deformation of the ZG118/ZG93 landslide monitoring site are influenced to varying degrees by the fluctuation components of the characterisation variables. For the ZG118/ZG93, feature factor 13 (Reservoir level IMF3) is chosen in both feature selection techniques, demonstrating both a complex nonlinear connection and a high linear correlation with the periodic displacement. It has more stability as a baseline model feature factor.</p>
<fig id="fig-13">
<label>Figure 13</label>
<caption>
<title>Feature selection heatmap of Pearson. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-13.tif"/>
</fig><fig id="fig-14">
<label>Figure 14</label>
<caption>
<title>Feature selection heatmap of MIC. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-14.tif"/>
</fig>
</sec>
<sec id="s4_2_2">
<label>4.2.2</label>
<title>Model Evaluation</title>
<p>The results of the periodic displacement prediction for the ZG118 and ZG93 (<xref ref-type="fig" rid="fig-15">Fig. 15a</xref>,<xref ref-type="fig" rid="fig-15">b</xref>) demonstrate that the BiLSTM-RBF model proposed in this study, which integrates a global-local hybrid attention mechanism, significantly outperforms baseline models requiring manual feature selection. As shown in <xref ref-type="fig" rid="fig-16">Fig. 16a</xref>,<xref ref-type="fig" rid="fig-16">b</xref>, the BiLSTM-RBF model achieved the lowest RMSE (ZG118: 10.45 mm; ZG93: 10.16 mm) and MAE (ZG118: 7.46 mm; ZG93: 9.15 mm) values and the highest R&#x00B2; (ZG118: 0.91; ZG93: 0.90) scores at both monitoring points (the &#x201C;None&#x201D; option in feature selection methods indicates the direct use of monthly cumulative rainfall, reservoir water levels, and historical displacement as inputs). Notably, after feature selection based on the Pearson and MIC, the RBF neural network exhibited superior predictive performance compared to predictions using raw data inputs.</p>
<fig id="fig-15">
<label>Figure 15</label>
<caption>
<title>Results of the periodic displacement prediction. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-15.tif"/>
</fig><fig id="fig-16">
<label>Figure 16</label>
<caption>
<title>Model validation for periodic displacement prediction. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-16.tif"/>
</fig>
<p>Among baseline models, the BiLSTM and RBF models showed better prediction accuracy for ZG118 than for ZG93 after feature decomposition. Furthermore, analysis of the ZG118 prediction results revealed that although RBF-MIC and BiLSTM-Pearson demonstrated higher accuracy, their evaluation metrics remained inferior to those of the proposed model. For ZG93 with a smaller sample size, baseline models exhibited suboptimal predictive performance even after feature selection, while the proposed model consistently achieved the best and most stable results. Moreover, the PIs of ZG118 and ZG93 both meet the expected coverage requirements for the actual data.</p>
<p>As shown in <xref ref-type="fig" rid="fig-17">Fig. 17a</xref>,<xref ref-type="fig" rid="fig-17">b</xref>, at the ZG118 and ZG93 monitoring points, compared with the RBF, BiLSTM, and GRU series models, the median error of BiLSTM-RBF is significantly lower and the dispersion is smaller (in the box-plot, the box of BiLSTM-RBF is the shortest and the whiskers are the narrowest), which proves that its performance advantage is not due to random fluctuations but a stable effect brought by the model structure (hybrid attention, bidirectional time-series modeling, and RBF fusion).</p>
<fig id="fig-17">
<label>Figure 17</label>
<caption>
<title>Box plot of the Wilcoxon test. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-17.tif"/>
</fig>
</sec>
<sec id="s4_2_3">
<label>4.2.3</label>
<title>Ablation Experiment</title>
<p>In order to verify the effectiveness and role of the introduced components, as well as the impact of different activation functions on local attention, ablation experiments were conducted on the Baishuihe ZG118 and ZG93 datasets. To ensure that the comparison between ablation experiments is only caused by the difference of the target variable, and all experiments use the same random seed as the main model. In <xref ref-type="table" rid="table-6">Table 6</xref>, &#x201C;&#x2713;&#x201D; and &#x201C;&#x274C;&#x201D; respectively indicate whether the component is enabled, and each row gives the corresponding indicator.</p>
<table-wrap id="table-6">
<label>Table 6</label>
<caption>
<title>Ablation experiment results of Baishuihe landslide ZG118 and ZG93</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
</colgroup>
<thead>
<tr>
<th align="center">Baishuihe</th>
<th align="center">Methods</th>
<th align="center">Global attention</th>
<th align="center">Local attention</th>
<th align="center">Sigmoid</th>
<th align="center">Tanh</th>
<th align="center">RMSE (mm)</th>
<th align="center">MAE (mm)</th>
<th align="center">R<sup>2</sup></th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="6">ZG118</td>
<td>BiLSTM-RBF-None</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>14.69</td>
<td>11.40</td>
<td>0.82</td>
</tr>
<tr>
<td>BiLSTM-RBF-Global</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>13.19</td>
<td>10.40</td>
<td>0.85</td>
</tr>
<tr>
<td>BiLSTM-RBF-Local1</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>12.35</td>
<td>9.80</td>
<td>0.87</td>
</tr>
<tr>
<td>BiLSTM-RBF-Local2</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>12.70</td>
<td>9.89</td>
<td>0.86</td>
</tr>
<tr>
<td>BiLSTM-RBF-Global-Local2</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>11.95</td>
<td>9.40</td>
<td>0.88</td>
</tr>
<tr>
<td><bold>BiLSTM-RBF (Ours)</bold></td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td><bold>10.45</bold></td>
<td><bold>7.46</bold></td>
<td><bold>0.91</bold></td>
</tr>
<tr>
<td rowspan="6">ZG93</td>
<td>BiLSTM-RBF-None</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>16.37</td>
<td>14.33</td>
<td>0.73</td>
</tr>
<tr>
<td>BiLSTM-RBF-Global</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>&#x274C;</td>
<td>14.94</td>
<td>12.81</td>
<td>0.78</td>
</tr>
<tr>
<td>BiLSTM-RBF-Local1</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>13.81</td>
<td>12.32</td>
<td>0.81</td>
</tr>
<tr>
<td>BiLSTM-RBF-Local2</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>19.32</td>
<td>18.06</td>
<td>0.63</td>
</tr>
<tr>
<td>BiLSTM-RBF-Global-Local2</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td>&#x2713;</td>
<td>18.75</td>
<td>17.74</td>
<td>0.65</td>
</tr>
<tr>
<td><bold>BiLSTM-RBF (Ours)</bold></td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x274C;</td>
<td><bold>10.16</bold></td>
<td><bold>9.15</bold></td>
<td><bold>0.90</bold></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The results in <xref ref-type="table" rid="table-6">Table 6</xref> indicate that:
<list list-type="simple">
<list-item><label>(1)</label><p>Whether attention is introduced or not, the fusion of BiLSTM and RBF is superior to single BiLSTM or RBF, proving the effectiveness of this structure itself.</p>
</list-item>
<list-item><label>(2)</label><p>In ZG118, after introducing global and local attention separately or jointly, compared with BiLSTM-RBF-None without attention, RMSE decreases by 10.2%&#x2013;28.9%, MAE decreases by 8.8%&#x2013;34.6%, and R&#x00B2; increases by 3.7%&#x2013;11.0%, verifying the effectiveness and complementarity of the two types of attention. In addition, by comparing the two groups of models, BiLSTM-RBF-Local1 and BiLSTM-RBF-Local2, and BiLSTM-RBF-Global-Local2 and BiLSTM-RBF (Ours), it is verified that using sigmoid in local attention is more stable than tanh.</p></list-item>
<list-item><label>(3)</label><p>In ZG93, both global attention (BiLSTM-RBF-Global) and local attention using sigmoid (BiLSTM-RB-Local1) are superior to BiLSTM-RBF-None without attention, verifying the effectiveness of each component again. In addition, when the local attention uses the tanh activation function, the R&#x00B2; values of BiLSTM-RBF-Local2 and BiLSTM-RBF-Global-Local2 are 0.63 and 0.69 respectively, which are not only lower than that of the method proposed in this paper (0.90), but even lower than that of BiLSTM-RBF-None without attention (0.73), further proving the advantages of sigmoid in terms of stability and performance.</p></list-item>
</list></p>
<p>In summary, it shows that the global attention mechanism effectively captures the long-term dependency relationship of landslide displacement, while the bidirectional local attention mechanism accurately extracts short-term local features.</p>
</sec>
</sec>
<sec id="s4_3">
<label>4.3</label>
<title>Cumulative Displacement Prediction Results</title>
<p><xref ref-type="fig" rid="fig-18">Fig. 18a</xref>,<xref ref-type="fig" rid="fig-18">b</xref> shows the cumulative displacement prediction results corresponding to the test set of trend and periodic displacements for the two landslide monitoring sites, and the evaluation metrics are shown in <xref ref-type="fig" rid="fig-19">Fig. 19a</xref>,<xref ref-type="fig" rid="fig-19">b</xref>, where the experimental results show that the overall prediction accuracies of the BiLSTM-RBF model are all better than the baseline model.</p>
<fig id="fig-18">
<label>Figure 18</label>
<caption>
<title>Results of the cumulative displacement prediction. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-18.tif"/>
</fig><fig id="fig-19">
<label>Figure 19</label>
<caption>
<title>Model validation for cumulative displacement prediction. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-19.tif"/>
</fig>
<p>As shown in <xref ref-type="table" rid="table-7">Table 7</xref>, the model in this paper is compared with other models with larger sample sizes under the most restricted condition with the smallest sample size. It shows that the RMSE of this paper&#x2019;s model is the lowest in both monitoring points ZG118 and ZG93. The BiLSTM-RBF prediction model is confirmed to be superior by the performance comparison of the aforementioned models. The improved performance of this paper&#x2019;s proposed model mostly depends on:</p>
<p><list list-type="simple">
<list-item><label>(1)</label><p>The historical periodic displacement data are used to construct a bidirectional local attention mechanism for BiLSTM, which enhances the model&#x2019;s bidirectional adaptive feature extraction from the original input.</p></list-item>
<list-item><label>(2)</label><p>The global attention acts on the combined hidden state of BiLSTM, and the local attention acts on the hidden state of the forward and reverse layers of BiLSTM, and the hybrid global-local attention mechanism enhances the model&#x2019;s feature extraction capability.</p></list-item>
<list-item><label>(3)</label><p>By fully utilizing the hidden states of BiLSTM to drive RBF neural network for nonlinear feature mapping, the model can extract the features of the weighted hidden state more efficiently, thus capturing the key features of the data at multiple levels under the condition of limited data, and further improving the performance of the model.</p></list-item>
</list></p>
<table-wrap id="table-7">
<label>Table 7</label>
<caption>
<title>Performance of various prediction models</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col align="center"/>
<col align="center"/>
<col/>
</colgroup>
<thead>
<tr>
<th>Baishuihe</th>
<th>Reference</th>
<th>Prediction model</th>
<th>Monitoring time</th>
<th align="center">Number of samples</th>
<th align="center">RMSE (mm)</th>
<th>R<sup>2</sup></th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="4">ZG118</td>
<td>Ref. [<xref ref-type="bibr" rid="ref-40">40</xref>]</td>
<td>CNN-BiGRU-Attention</td>
<td>2003/06&#x2013;2016/12</td>
<td>163</td>
<td>28.94</td>
<td>0.944</td>
</tr>
<tr>
<td>Ref. [<xref ref-type="bibr" rid="ref-41">41</xref>]</td>
<td>LSTNet</td>
<td>2007/01&#x2013;2012/12</td>
<td>72</td>
<td>12.90</td>
<td>0.95</td>
</tr>
<tr>
<td>Ref. [<xref ref-type="bibr" rid="ref-42">42</xref>]</td>
<td>Bootstrap-KELM-BPNN</td>
<td>2004/07&#x2013;2013/12</td>
<td>114</td>
<td>12.18</td>
<td>0.96</td>
</tr>
<tr>
<td><bold>Ours</bold></td>
<td><bold>BiLSTM-RBF</bold></td>
<td><bold>2006/01&#x2013;2012/12</bold></td>
<td><bold>84</bold></td>
<td><bold>10.42</bold></td>
<td><bold>0.96</bold></td>
</tr>
<tr>
<td rowspan="4">ZG93</td>
<td>Ref. [<xref ref-type="bibr" rid="ref-43">43</xref>]</td>
<td>ELM</td>
<td>2005/06&#x2013;2016/12</td>
<td>139</td>
<td>17.41</td>
<td>0.968</td>
</tr>
<tr>
<td>Ref. [<xref ref-type="bibr" rid="ref-42">42</xref>]</td>
<td>Bootstrap-KELM-BPNN</td>
<td>2004/07&#x2013;2013/12</td>
<td>114</td>
<td>11.47</td>
<td>0.96</td>
</tr>
<tr>
<td>Ref. [<xref ref-type="bibr" rid="ref-44">44</xref>]</td>
<td>OVMD-GWO-KELM</td>
<td>2006/06&#x2013;2016/12</td>
<td>127</td>
<td>16.365</td>
<td>0.99</td>
</tr>
<tr>
<td><bold>Ours</bold></td>
<td><bold>BiLSTM-RBF</bold></td>
<td><bold>2006/01&#x2013;2011/12</bold></td>
<td><bold>72</bold></td>
<td><bold>10.37</bold></td>
<td><bold>0.95</bold></td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Discussions</title>
<p>BiLSTM constitutes an efficient deep learning model by combining forward and reverse information. However, the traditional global attention mechanism is directly applied to its concatenated hidden states, which makes the process of successfully extracting bidirectional characteristics from sequence data challenging. For this reason, this study proposes a bi-directional local attention mechanism based on historical displacement data to enhance bi-directional feature recognition and trend-capturing capabilities. Additionally, an RBF neural network is included to improve the model&#x2019;s effectiveness in recognizing data patterns.
<list list-type="simple">
<list-item><label>(1)</label><p>When optimizing the smoothing coefficients of the DES using NSGA-II, constraining the range of individual mutation is crucial to obtain the optimal solution, which can dynamically adjust the smoothing coefficients in accordance with the number of samples, and the optimized coefficients are better than the current widely used parameters of the DES [<xref ref-type="bibr" rid="ref-34">34</xref>]. In periodic displacement prediction, although batch-level optimization is achieved by using the width of the PI as a penalty term in the loss function, independent optimization for each individual time step has not been realized. Additionally, the penalty coefficient <inline-formula id="ieqn-71"><mml:math id="mml-ieqn-71"><mml:mi>&#x03BB;</mml:mi></mml:math></inline-formula> needs to be adjusted according to the change in the sample size. However, the impact of the <inline-formula id="ieqn-72"><mml:math id="mml-ieqn-72"><mml:mi>&#x03BB;</mml:mi></mml:math></inline-formula> value on the model performance is nonlinear, which makes it a challenge to balance the prediction accuracy and the coverage rate of the prediction interval. These limitations will be the focus of future improvements to the loss function structure.</p></list-item>
<list-item><label>(2)</label><p>From <xref ref-type="fig" rid="fig-16">Fig. 16</xref>, the improvement effect of feature decomposition on the neural network relies upon the combination of the model and the feature selection method. The degree of improvement effect is different in ZG118 and ZG93 monitoring points, and even performance degradation occurs in BiLSTM-MIC and GRU-MIC. However, in the RBF-MIC and BiLSTM-Pearson at monitoring site ZG118, and the RBF-Pearson and RBF-MIC baseline models at monitoring site ZG93, the RMSE, MAE, and R<sup>2</sup> were all better than those of using the raw temporal input data (monthly cumulative rainfall, reservoir level, and historical periodic displacement) directly. Therefore, it is crucial to perform feature factor decomposition before manual feature selection, but it is necessary to choose the appropriate feature selection method and neural network model.</p></list-item>
<list-item><label>(3)</label><p>To further verify the robustness of the feature decomposition method, this study compared the impacts of EMD, Variational Mode Decomposition (VMD), and Complete Ensemble Empirical Mode Decomposition with Adaptive Noise (CEEMDAN) on the model prediction performance. For each decomposition method, the same processing procedure as that of EMD was adopted: the component showing a trend change was extracted as the trend displacement, and the remaining components were reconstructed into the periodic displacement. Meanwhile, using the same model parameters as those based on EMD decomposition, the periodic displacement and trend displacement of the monitoring points ZG118 and ZG93 of Baishuihe landslide were predicted and analyzed, respectively. By comparing the prediction results (see <xref ref-type="table" rid="table-8">Tables 8</xref> and <xref ref-type="table" rid="table-9">9</xref>), it can be seen that the VMD performed the worst in both the prediction of trend displacement and periodic displacement, while the performance of CEEMDAN was significantly better than that of VMD. It is worth noting that under the condition of the same model parameters, the three model evaluation indicators for the prediction of trend displacement and periodic displacement based on EMD were still the best, indicating that under the model framework combined with the hybrid attention mechanism in this paper, the boundary effect of EMD can be significantly improved through data-driven feature extraction. Overall, the prediction results of the model based on EMD proposed in this paper have good robustness.</p>
</list-item>
</list></p>
<table-wrap id="table-8">
<label>Table 8</label>
<caption>
<title>Trend displacement prediction indicators of BiLSTM-RBF based on different modal decompositions</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Baishuihe</th>
<th>Decomposition method</th>
<th>RMSE (mm)</th>
<th>MAE (mm)</th>
<th>R<sup>2</sup></th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="3">ZG118</td>
<td><bold>EMD</bold></td>
<td><bold>0.35</bold></td>
<td><bold>0.29</bold></td>
<td><bold>0.99</bold></td>
</tr>
<tr>
<td>VMD</td>
<td>10.23</td>
<td>9.12</td>
<td>0.98</td>
</tr>
<tr>
<td>CEEMDAN</td>
<td>8.05</td>
<td>7.91</td>
<td>0.99</td>
</tr>
<tr>
<td rowspan="3">ZG93</td>
<td><bold>EMD</bold></td>
<td><bold>1.02</bold></td>
<td><bold>0.93</bold></td>
<td><bold>0.99</bold></td>
</tr>
<tr>
<td>VMD</td>
<td>10.74</td>
<td>9.42</td>
<td>0.97</td>
</tr>
<tr>
<td>CEEMDAN</td>
<td>2.40</td>
<td>0.97</td>
<td>0.99</td>
</tr>
</tbody>
</table>
</table-wrap><table-wrap id="table-9">
<label>Table 9</label>
<caption>
<title>Periodic displacement prediction indicators of BiLSTM-RBF based on different modal decompositions</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Baishuihe</th>
<th>Decomposition method</th>
<th>RMSE (mm)</th>
<th>MAE (mm)</th>
<th>R<sup>2</sup></th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="3">ZG118</td>
<td><bold>EMD</bold></td>
<td><bold>10.45</bold></td>
<td><bold>7.46</bold></td>
<td><bold>0.91</bold></td>
</tr>
<tr>
<td>VMD</td>
<td>51.34</td>
<td>49.40</td>
<td>&#x2212;3.47</td>
</tr>
<tr>
<td>CEEMDAN</td>
<td>14.18</td>
<td>12.57</td>
<td>0.84</td>
</tr>
<tr>
<td rowspan="3">ZG93</td>
<td><bold>EMD</bold></td>
<td><bold>10.16</bold></td>
<td><bold>9.15</bold></td>
<td><bold>0.90</bold></td>
</tr>
<tr>
<td>VMD</td>
<td>16.57</td>
<td>14.52</td>
<td>0.39</td>
</tr>
<tr>
<td>CEEMDAN</td>
<td>20.46</td>
<td>14.12</td>
<td>0.70</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><list list-type="simple">
<list-item><label>(4)</label><p>In the BiLSTM-RBF model for periodic displacement prediction of the Baishuihe landslide, when compared with the best baseline model, the RMSE, MAE, and R&#x00B2; of ZG118 improved by 2.72, 4.25 and 0.06 mm; In ZG93, the three metrics improved by 11.36, 4.48 and 0.36 mm. which confirmed the model&#x2019;s ability to predict the displacement of landslides accurately. <xref ref-type="table" rid="table-10">Table 10</xref> compares the total running time and the number of model parameters for a single run (epochs &#x003D; 200) in predicting periodic displacement. These include the BiLSTM-RBF model with raw time series data as input and the baseline model under different feature selection methods (the running time of the baseline model excludes the time needed for manual feature selection). This shows that model type, number of neurons, number of features, and number of samples all affect the model run time and number of model parameters to varying degrees. Specifically, the number of neurons has the most impact on model parameters.</p>
</list-item>
</list></p>
<table-wrap id="table-10">
<label>Table 10</label>
<caption>
<title>Efficiency evaluation of various models run once (epochs &#x003D; 200)</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
</colgroup>
<thead>
<tr>
<th>Baishuihe</th>
<th>Model</th>
<th align="center">Number of neurons</th>
<th align="center">Number of features</th>
<th align="center">Total number of parameters</th>
<th align="center">Running time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="10">ZG118</td>
<td>RBF (None)</td>
<td>64</td>
<td>2</td>
<td>259</td>
<td>17.34</td>
</tr>
<tr>
<td>RBF (Pearson)</td>
<td>64</td>
<td>6</td>
<td>519</td>
<td>15.37</td>
</tr>
<tr>
<td>RBF (MIC)</td>
<td>64</td>
<td>7</td>
<td>649</td>
<td>15.77</td>
</tr>
<tr>
<td>BiLSTM (None)</td>
<td>64</td>
<td>2</td>
<td>34433</td>
<td>23.34</td>
</tr>
<tr>
<td>BiLSTM (Pearson)</td>
<td>64</td>
<td>6</td>
<td>36481</td>
<td>23.71</td>
</tr>
<tr>
<td>BiLSTM (MIC)</td>
<td>32</td>
<td>7</td>
<td>10561</td>
<td>20.51</td>
</tr>
<tr>
<td>GRU (None)</td>
<td>64</td>
<td>2</td>
<td>13121</td>
<td>18.80</td>
</tr>
<tr>
<td>GRU (Pearson)</td>
<td>64</td>
<td>6</td>
<td>13889</td>
<td>18.31</td>
</tr>
<tr>
<td>GRU (MIC)</td>
<td>64</td>
<td>7</td>
<td>14273</td>
<td>18.72</td>
</tr>
<tr>
<td><bold>BiLSTM-RBF (None)</bold></td>
<td></td>
<td><bold>2</bold></td>
<td><bold>47176</bold></td>
<td><bold>27.19</bold></td>
</tr>
<tr>
<td rowspan="10">ZG93</td>
<td>RBF (None)</td>
<td>64</td>
<td>3</td>
<td>324</td>
<td>15.71</td>
</tr>
<tr>
<td>RBF (Pearson)</td>
<td>64</td>
<td>12</td>
<td>974</td>
<td>14.90</td>
</tr>
<tr>
<td>RBF (MIC)</td>
<td>64</td>
<td>2</td>
<td>324</td>
<td>14.54</td>
</tr>
<tr>
<td>BiLSTM (None)</td>
<td>64</td>
<td>3</td>
<td>34945</td>
<td>23.32</td>
</tr>
<tr>
<td>BiLSTM (Pearson)</td>
<td>32</td>
<td>12</td>
<td>11841</td>
<td>20.20</td>
</tr>
<tr>
<td>BiLSTM (MIC)</td>
<td>32</td>
<td>2</td>
<td>9281</td>
<td>20.57</td>
</tr>
<tr>
<td>GRU (None)</td>
<td>32</td>
<td>3</td>
<td>3585</td>
<td>13.21</td>
</tr>
<tr>
<td>GRU (Pearson)</td>
<td>32</td>
<td>12</td>
<td>4545</td>
<td>13.34</td>
</tr>
<tr>
<td>GRU (MIC)</td>
<td>32</td>
<td>2</td>
<td>3585</td>
<td>18.62</td>
</tr>
<tr>
<td><bold>BiLSTM-RBF (None)</bold></td>
<td></td>
<td><bold>3</bold></td>
<td><bold>13640</bold></td>
<td><bold>24.45</bold></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As in <xref ref-type="fig" rid="fig-20">Fig. 20</xref>, to quantify the model efficiency, the &#x201C;time consumption per parameter&#x201D; (i.e., the running time divided by the number of parameters) is introduced to reflect the slope <italic>k</italic> of the linear fitting of the relationship curve between the running time and the number of parameters. A steeper slope (larger <italic>k</italic>) means that a small increase in the number of parameters will lead to a significant increase in the running time, indicating that the model is more &#x201C;inefficient&#x201D; in terms of computing resource utilization. A gentler slope (smaller <italic>k</italic>) indicates that parameter expansion has a small impact on the running time, resulting in higher computational efficiency. Therefore, the model run at monitoring point ZG118 showed the greatest prediction efficiency. At monitoring point ZG93, the run efficiency was second only to the baseline models BiLSTM (None) and BiLSTM (Pearson), but compared to the most efficient BiLSTM (None), the model parameters were reduced by about 2.5 times, while the run time increased by only 4.85%, indicating that the model memory footprint was significantly reduced without sacrificing real-time performance.</p>
<fig id="fig-20">
<label>Figure 20</label>
<caption>
<title>Visual analytics for the efficiency evaluation. (<bold>a</bold>) ZG118; (<bold>b</bold>) ZG93</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_67952-fig-20.tif"/>
</fig>
<p><list list-type="simple">
<list-item><label>(5)</label><p>In <xref ref-type="table" rid="table-7">Table 7</xref>, this study verifies the effectiveness of the proposed model in overcoming data scarcity and making full use of historical data through model comparison under different sample sizes. Except for the LSTNet model in Reference [<xref ref-type="bibr" rid="ref-41">41</xref>], the number of monitoring samples used in this study (84 and 72 samples for ZG118 and ZG93, respectively) is less than that of other literature models. Especially in References [<xref ref-type="bibr" rid="ref-40">40</xref>] and [<xref ref-type="bibr" rid="ref-43">43</xref>], the sample size used in this study is almost half of theirs, while the proposed model still achieves the lowest RMSE. However, each model is trained and tested on time series of different lengths. This difference directly leads to limitations in model performance evaluation. Future research can conduct model training and comparison on a dataset with a unified time span and sampling frequency to ensure the comparability and rigor of the results.</p>
</list-item>
</list>
<list list-type="simple">
<list-item><label>(6)</label><p>This model constructs a more efficient prediction framework. However, to further verify its adaptability and robustness in current mainstream time series prediction models and for different geological landslides, this paper selects the Bazimen landslide, which has significant differences in geological background, deformation mechanism, and inducing factors from the Baishuihe landslide, to further verify the generalization of the model and compares it with Transformer. The selected monitoring points and time period are: ZG111, 2007/01&#x2013;2012/12, with 72 samples. The data preprocessing strictly follows the established process of the Baishuihe landslide. Two periodic displacement components, IMF1 and IMF2, of the Bazimen landslide are decomposed through EMD. Model parameter settings: The same model parameters for IMF1 and IMF2 are: epochs &#x003D; 200, learning rate &#x003D; 0.001, batch&#x003D;2. The data splitting in chronological order (Training set: Validation set: Test set &#x003D; 60%: 10%: 30%). Other model parameters: LSTM_units (IMF1) &#x003D; 5, LSTM_units (IMF2) &#x003D; 20, RBF_units (IMF1) &#x003D; 20, RBF_units (IMF2) &#x003D; 100. In IMF1 and IMF2, the penalty coefficients of the loss function <inline-formula id="ieqn-73"><mml:math id="mml-ieqn-73"><mml:mi>&#x03BB;</mml:mi></mml:math></inline-formula> are 0.1 and 0.01 respectively, the random seeds are respectively: (NumPy: 1, TensorFlow: 5) and (NumPy: 1, TensorFlow: 2). The experimental results in <xref ref-type="table" rid="table-11">Table 11</xref> show that the model proposed in this paper demonstrates the optimal performance in the displacement prediction of the Bazimen landslide across regions and when compared with the current mainstream Transformer models. This fully verifies that the BiLSTM-RBF model with a hybrid attention mechanism has good migration ability under different geological conditions.</p>
</list-item>
</list></p>
<table-wrap id="table-11">
<label>Table 11</label>
<caption>
<title>Model validation based on the Transformer and across landslide areas</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Landslide area</th>
<th align="center" colspan="2">Monitoring point</th>
<th>Model</th>
<th>RMSE (mm)</th>
<th>MAE (mm)</th>
<th>R<sup>2</sup></th>
</tr>
</thead>
<tbody>
<tr>
<td rowspan="4">Baishuihe</td>
<td align="center" rowspan="2">ZG118</td>
<td/>
<td><bold>BiLSTM-RBF</bold></td>
<td><bold>10.45</bold></td>
<td><bold>7.46</bold></td>
<td><bold>0.91</bold></td>
</tr>
<tr>
<td/>   
<td>Transformer</td>
<td>21.34</td>
<td>18.59</td>
<td>0.61</td>
</tr>
<tr>
<td align="center" rowspan="2">ZG93</td>
<td><bold>BiLSTM-RBF</bold></td>
<td/>
<td><bold>10.16</bold></td>
<td><bold>9.15</bold></td>
<td><bold>0.90</bold></td>
</tr>
<tr>
<td/>   
<td>Transformer</td>
<td>18.71</td>
<td>15.14</td>
<td>0.65</td>
</tr>
<tr>
<td align="center" rowspan="4">Bazimen</td>
<td align="center" rowspan="4">ZG111</td>
<td>IMF1</td>
<td><bold>BiLSTM-RBF</bold></td>
<td><bold>18.75</bold></td>
<td><bold>16.11</bold></td>
<td><bold>0.32</bold></td>
</tr>
<tr>
<td></td>
<td>Transformer</td>
<td>21.70</td>
<td>17.03</td>
<td>0.09</td>
</tr>
<tr>
<td>IMF2</td>
<td><bold>BiLSTM-RBF</bold></td>
<td><bold>8.58</bold></td>
<td><bold>7.31</bold></td>
<td><bold>0.96</bold></td>
</tr>
<tr>
<td></td>
<td>Transformer</td>
<td>20.29</td>
<td>17.16</td>
<td>0.77</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s6">
<label>6</label>
<title>Conclusion</title>
<p>In this study, we propose a BiLSTM-RBF model with global and bidirectional local attention mechanisms, which can directly take raw time series data as inputs and is appropriate for landslide displacement prediction with a limited sample size. By designing an interpretable attention mechanism, all hidden states of BiLSTM are fully applied to realize pattern recognition and the full extraction of landslide monitoring data features. From results of the cumulative displacement prediction for the Baishuihe landslide (<xref ref-type="fig" rid="fig-19">Fig. 19</xref>): at monitoring point ZG118, the evaluation metrics of BiLSTM-RBF, RMSE, MAE, and R<sup>2</sup>, are all improved by 2.57, 4.05 and 0.03 mm compared with the best-performing baseline model (RBF-MIC). At monitoring point ZG93, all three metrics are improved by 11.27, 4.52 and 0.17 mm compared with the best-performing baseline model (GRU-None). Thus, the proposed model still performs consistently on fewer datasets and can significantly improve model performance. The proposed model eliminates the requirement for manual feature selection while achieving enhanced computational efficiency in prediction tasks compared to the baseline model.</p>
<p>Overall, this study presents the BiLSTM-RBF model enhanced with a hybrid global-local attentional mechanism, which aims to fully utilize the existing available data for the precise and efficient prediction of landslide displacement in the case of insufficient landslide monitoring data. The model introduces the working principle of the interpretable attention mechanism and successfully builds the BiLSTM-RBF network model. Both the improvement of the model structure and the design of the weights of the attention mechanism focus on the historical data and deeply excavate the features, which provides an intelligent analysis method with interpretability, high accuracy, and high efficiency for geohazard prediction. The proposed BiLSTM-RBF model with a hybrid attention mechanism has the potential to significantly improve the accuracy of landslide displacement prediction, which is crucial for enhancing disaster prevention and risk mitigation strategies in landslide-prone areas.</p>
</sec>
</body>
<back>
<ack>
<p>Not applicable.</p>
</ack>
<sec>
<title>Funding Statement</title>
<p>This work was supported in part by the Guizhou Province Science Technology Support Plan ([2024] General 007, [2022] General 264, [2023] General 096, [2023] General 412, and [2023] General 409); in part by the National Natural Science Foundation of China (Grant No. 61861007); in part by the Guizhou Province Science and Technology Planning Project (ZK [2021] General 303); in part by the Project of GUIYANG HYDROPOWER INVESTIGATION DESIGN &#x0026; RESEARCH INSTITUTE CHECC (YJ2022-12); in part by the Science and Technology Project of Power Construction Corporation of China, Ltd. (DJ-ZDXM-2022-44).</p>
</sec>
<sec>
<title>Author Contributions</title>
<p>The authors confirm contribution to the paper as follows: Conceptualization, Jiao Chen, Xiao Wang and Zhiqin He; methodology, Jiao Chen, Xiao Wang and Yi Chen; software, Jiao Chen; validation, Jiao Chen, Chao Ma; formal analysis, Zhiqin He, Yi Chen; investigation, Zhiqin He, Chao Ma; resources, Xiao Wang; data curation, Jiao Chen, Zhiqin He, Chao Ma; writing&#x2014;original draft preparation, Jiao Chen; writing&#x2014;review and editing, Jiao Chen, Xiao Wang and Yi Chen; visualization, Jiao Chen; supervision, Xiao Wang; project administration, Xiao Wang, Zhiqin He; funding acquisition, Xiao Wang. All authors reviewed the results and approved the final version of the manuscript.</p>
</sec>
<sec sec-type="data-availability">
<title>Availability of Data and Materials</title>
<p>The dataset is provided by National Cryosphere Desert Data Center. (<ext-link ext-link-type="uri" xlink:href="http://www.ncdc.ac.cn">http://www.ncdc.ac.cn</ext-link>). (Deformation monitoring data of Baishuihe landslide in Zigui County, Three Gorges Reservoir area (2006): <ext-link ext-link-type="uri" xlink:href="https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.db1668.2022">https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.db1668.2022</ext-link>; Basic characteristics and monitoring data of Baishuihe landslide in Zigui County, Three Gorges Reservoir area (2007&#x2013;2012): <ext-link ext-link-type="uri" xlink:href="https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.2020.71">https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.2020.71</ext-link>; Basic characteristics and monitoring data of Bazimen landslide in Zigui County, Three Gorges Reservoir area (2007&#x2013;2012): <ext-link ext-link-type="uri" xlink:href="https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.2020.70">https://cstr.cn/CSTR:11738.11.ncdc.Sanxia.2020.70</ext-link>). (Websites above are accessed on 14 August 2025).</p>
</sec>
<sec>
<title>Ethics Approval</title>
<p>Not applicable.</p>
</sec>
<sec sec-type="COI-statement">
<title>Conflicts of Interest</title>
<p>The authors declare that they have no conflicts of interest to report regarding the present study.</p>
</sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname> <given-names>S</given-names></string-name>, <string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Peng</surname> <given-names>J</given-names></string-name>, <string-name><surname>Zhou</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>S</given-names></string-name>, <string-name><surname>Chen</surname> <given-names>Y</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Fatal landslides in China from 1940 to 2020: occurrences and vulnerabilities</article-title>. <source>Landslides</source>. <year>2023</year>;<volume>20</volume>(<issue>6</issue>):<fpage>1243</fpage>&#x2013;<lpage>64</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10346-023-02034-6</pub-id>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Poudel</surname> <given-names>N</given-names></string-name>, <string-name><surname>Mani Dixit</surname> <given-names>A</given-names></string-name>, <string-name><surname>Shiga</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Cao</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Shaw</surname> <given-names>R</given-names></string-name></person-group>. <article-title>Big data challenges and opportunities for disaster early warning system</article-title>. <source>Prevent Treat Natural Dis</source>. <year>2024</year>;<volume>3</volume>(<issue>1</issue>):<fpage>155</fpage>&#x2013;<lpage>64</lpage>. doi:<pub-id pub-id-type="doi">10.54963/ptnd.v3i1.283</pub-id>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Xu</surname> <given-names>W</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>H</given-names></string-name>, <string-name><surname>Chen</surname> <given-names>J</given-names></string-name>, <string-name><surname>Kang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Pu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Ye</surname> <given-names>Y</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Combining numerical simulation and deep learning for landslide displacement prediction: an attempt to expand the deep learning dataset</article-title>. <source>Sustainability</source>. <year>2022</year>;<volume>14</volume>(<issue>11</issue>):<fpage>6908</fpage>. doi:<pub-id pub-id-type="doi">10.3390/su14116908</pub-id>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ge</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>J</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>C</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>X</given-names></string-name>, <string-name><surname>Deng</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Integrating feature selection with machine learning for accurate reservoir landslide displacement prediction</article-title>. <source>Water</source>. <year>2024</year>;<volume>16</volume>(<issue>15</issue>):<fpage>2152</fpage>. doi:<pub-id pub-id-type="doi">10.3390/w16152152</pub-id>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname> <given-names>WF</given-names></string-name>, <string-name><surname>Yu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Zhao</surname> <given-names>QY</given-names></string-name>, <string-name><surname>Cheng</surname> <given-names>G</given-names></string-name>, <string-name><surname>Hou</surname> <given-names>XB</given-names></string-name>, <string-name><surname>He</surname> <given-names>SQ</given-names></string-name></person-group>. <article-title>Time series forecasting fusion network model based on prophet and improved LSTM</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>74</volume>(<issue>2</issue>):<fpage>3199</fpage>&#x2013;<lpage>219</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2023.032595</pub-id>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Gatto</surname> <given-names>MPA</given-names></string-name>, <string-name><surname>Montrasio</surname> <given-names>L</given-names></string-name></person-group>. <article-title>X-SLIP: a SLIP-based multi-approach algorithm to predict the spatial-temporal triggering of rainfall-induced shallow landslides over large areas</article-title>. <source>Comput Geotech</source>. <year>2023</year>;<volume>154</volume>(<issue>1</issue>):<fpage>105175</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.compgeo.2022.105175</pub-id>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Dong</surname> <given-names>J</given-names></string-name>, <string-name><surname>Lu</surname> <given-names>G</given-names></string-name>, <string-name><surname>Yan</surname> <given-names>F</given-names></string-name></person-group>. <article-title>Landslide displacement prediction based on CEEMDAN-LSTM</article-title>. <source>Communicat Sci Technol Heilongjiang</source>. <year>2024</year>;<volume>47</volume>(<issue>5</issue>):<fpage>158</fpage>&#x2013;<lpage>61</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.16402/j.cnki.issn1008-3383.2024.05.014</pub-id>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Jin</surname> <given-names>A</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>S</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>X</given-names></string-name></person-group>. <article-title>Landslide displacement prediction based on time series and long short-term memory networks</article-title>. <source>Bull Eng Geol Environ</source>. <year>2024</year>;<volume>83</volume>(<issue>7</issue>):<fpage>264</fpage>. doi:<pub-id pub-id-type="doi">10.1007/s10064-024-03714-w</pub-id>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Guo</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Chen</surname> <given-names>L</given-names></string-name>, <string-name><surname>Gui</surname> <given-names>L</given-names></string-name>, <string-name><surname>Du</surname> <given-names>J</given-names></string-name>, <string-name><surname>Yin</surname> <given-names>K</given-names></string-name>, <string-name><surname>Do</surname> <given-names>HM</given-names></string-name></person-group>. <article-title>Landslide displacement prediction based on variational mode decomposition and WA-GWO-BP model</article-title>. <source>Landslides</source>. <year>2020</year>;<volume>17</volume>(<issue>3</issue>):<fpage>567</fpage>&#x2013;<lpage>83</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10346-019-01314-4</pub-id>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Lu</surname> <given-names>G</given-names></string-name>, <string-name><surname>Dong</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Prediction of landslide displacement with step-like curve using variational mode decomposition and periodic neural network</article-title>. <source>Bull Eng Geol Environ</source>. <year>2021</year>;<volume>80</volume>(<issue>5</issue>):<fpage>3783</fpage>&#x2013;<lpage>99</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10064-021-02136-2</pub-id>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Shaiba</surname> <given-names>H</given-names></string-name>, <string-name><surname>Marzouk</surname> <given-names>R</given-names></string-name>, <string-name><surname>Nour</surname> <given-names>MK</given-names></string-name>, <string-name><surname>Negm</surname> <given-names>N</given-names></string-name>, <string-name><surname>Hilal</surname> <given-names>AM</given-names></string-name>, <string-name><surname>Mohamed</surname> <given-names>A</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Weather forecasting prediction using ensemble machine learning for big data applications</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>73</volume>(<issue>2</issue>):<fpage>3367</fpage>&#x2013;<lpage>82</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2022.030067</pub-id>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ganapathy</surname> <given-names>GP</given-names></string-name>, <string-name><surname>Srinivasan</surname> <given-names>K</given-names></string-name>, <string-name><surname>Datta</surname> <given-names>D</given-names></string-name>, <string-name><surname>Chang</surname> <given-names>CY</given-names></string-name>, <string-name><surname>Purohit</surname> <given-names>O</given-names></string-name>, <string-name><surname>Zaalishvili</surname> <given-names>V</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Rainfall forecasting using machine learning algorithms for localized events</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>71</volume>(<issue>3</issue>):<fpage>6333</fpage>&#x2013;<lpage>50</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2022.023254</pub-id>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Alhussan</surname> <given-names>AA</given-names></string-name>, <string-name><surname>El-kenawy</surname> <given-names>ES</given-names></string-name>, <string-name><surname>Aleisa</surname> <given-names>HN</given-names></string-name>, <string-name><surname>El-said</surname> <given-names>M</given-names></string-name>, <string-name><surname>Ward</surname> <given-names>SA</given-names></string-name>, <string-name><surname>Khafaga</surname> <given-names>DS</given-names></string-name></person-group>. <article-title>Optimization ensemble weights model for wind forecasting system</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>73</volume>(<issue>2</issue>):<fpage>2619</fpage>&#x2013;<lpage>35</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2022.030445</pub-id>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Jiang</surname> <given-names>W</given-names></string-name>, <string-name><surname>Leng</surname> <given-names>X</given-names></string-name>, <string-name><surname>Lin</surname> <given-names>X</given-names></string-name>, <string-name><surname>Feng</surname> <given-names>L</given-names></string-name>, <string-name><surname>Jiang</surname> <given-names>H</given-names></string-name></person-group>. <article-title>Landslide displacement prediction based on time series and temporal convolutional network</article-title>. <source>Sci Technol Eng</source>. <year>2023</year>;<volume>23</volume>(<issue>9</issue>):<fpage>3672</fpage>&#x2013;<lpage>9</lpage>. (In Chinese).</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>P</given-names></string-name>, <string-name><surname>Lin</surname> <given-names>J</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>J</given-names></string-name>, <string-name><surname>Xiang</surname> <given-names>C</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Earthquake-triggered landslide susceptibility prediction in Jiuzhaigou based on BP neural network</article-title>. <source>J Eng Geol</source>. <year>2024</year>;<volume>32</volume>(<issue>1</issue>):<fpage>133</fpage>&#x2013;<lpage>45</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.13544/j.cnki.jeg.2022-0013</pub-id>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ebrahim</surname> <given-names>KMP</given-names></string-name>, <string-name><surname>Fares</surname> <given-names>A</given-names></string-name>, <string-name><surname>Faris</surname> <given-names>N</given-names></string-name>, <string-name><surname>Zayed</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Exploring time series models for landslide prediction: a literature review</article-title>. <source>Geoenviron Disasters</source>. <year>2024</year>;<volume>11</volume>(<issue>1</issue>):<fpage>25</fpage>. doi:<pub-id pub-id-type="doi">10.1186/s40677-024-00288-3</pub-id>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yang</surname> <given-names>B</given-names></string-name>, <string-name><surname>Yin</surname> <given-names>K</given-names></string-name>, <string-name><surname>Lacasse</surname> <given-names>S</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>Time series analysis and long short-term memory neural network to predict landslide displacement</article-title>. <source>Landslides</source>. <year>2019</year>;<volume>16</volume>(<issue>4</issue>):<fpage>677</fpage>&#x2013;<lpage>94</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10346-018-01127-x</pub-id>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Cai</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>W</given-names></string-name>, <string-name><surname>Meng</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Shi</surname> <given-names>C</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>R</given-names></string-name></person-group>. <article-title>Prediction of landslide displacement based on GA-LSSVM with multiple factors</article-title>. <source>Bull Eng Geol Environ</source>. <year>2016</year>;<volume>75</volume>(<issue>2</issue>):<fpage>637</fpage>&#x2013;<lpage>46</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10064-015-0804-z</pub-id>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Dai</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Dai</surname> <given-names>W</given-names></string-name>, <string-name><surname>Yu</surname> <given-names>W</given-names></string-name>, <string-name><surname>Bai</surname> <given-names>D</given-names></string-name></person-group>. <article-title>Determination of landslide displacement warning thresholds by applying DBA-LSTM and numerical simulation algorithms</article-title>. <source>Appl Sci</source>. <year>2022</year>;<volume>12</volume>(<issue>13</issue>):<fpage>6690</fpage>. doi:<pub-id pub-id-type="doi">10.3390/app12136690</pub-id>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Filipovi&#x0107;</surname> <given-names>N</given-names></string-name>, <string-name><surname>Brdar</surname> <given-names>S</given-names></string-name>, <string-name><surname>Mimi&#x0107;</surname> <given-names>G</given-names></string-name>, <string-name><surname>Marko</surname> <given-names>O</given-names></string-name>, <string-name><surname>Crnojevi&#x0107;</surname> <given-names>V</given-names></string-name></person-group>. <article-title>Regional soil moisture prediction system based on Long Short-Term Memory network</article-title>. <source>Biosyst Eng</source>. <year>2022</year>;<volume>213</volume>(<issue>10</issue>):<fpage>30</fpage>&#x2013;<lpage>8</lpage>. doi:<pub-id pub-id-type="doi">10.1016/j.biosystemseng.2021.11.019</pub-id>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname> <given-names>M</given-names></string-name>, <string-name><surname>Li</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wen</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>An updated approach to predict landslide displacement by combining variational mode decomposition with bidirectional long short-term memory neural network model</article-title>. <source>Mount Res</source>. <year>2021</year>;<volume>39</volume>(<issue>6</issue>):<fpage>855</fpage>&#x2013;<lpage>66</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.16089/j.cnki.1008-2786.000644</pub-id>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname> <given-names>M</given-names></string-name>, <string-name><surname>Han</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>P</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>C</given-names></string-name></person-group>. <article-title>Landslide displacement prediction based on optimized empirical mode decomposition and deep bidirectional long short-term memory network</article-title>. <source>J Mt Sci</source>. <year>2023</year>;<volume>20</volume>(<issue>3</issue>):<fpage>637</fpage>&#x2013;<lpage>56</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s11629-022-7638-5</pub-id>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Huang</surname> <given-names>L</given-names></string-name>, <string-name><surname>Hao</surname> <given-names>J</given-names></string-name>, <string-name><surname>Li</surname> <given-names>W</given-names></string-name>, <string-name><surname>Zhou</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Jia</surname> <given-names>P</given-names></string-name></person-group>. <article-title>Landslide susceptibility assessment by the coupling method of RBF neural network and information value: a case study in Min Xian, Gansu Province</article-title>. <source>Chin J Geolog Hazard Cont</source>. <year>2021</year>;<volume>32</volume>(<issue>6</issue>):<fpage>116</fpage>&#x2013;<lpage>26</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.16031/j.cnki.issn.1003-8035.2021.06-14</pub-id>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhao</surname> <given-names>X</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>F</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>H</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Study on improved learning vector quantization landslide vulnerability evaluation model</article-title>. <source>Sci Surv Mapp</source>. <year>2023</year>;<volume>48</volume>(<issue>5</issue>):<fpage>239</fpage>&#x2013;<lpage>46</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.16251/j.cnki.1009-2307.2023.05.028</pub-id>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Tavaen</surname> <given-names>S</given-names></string-name>, <string-name><surname>Kaennakham</surname> <given-names>S</given-names></string-name></person-group>. <article-title>Numerical comparison of shapeless radial basis function networks in pattern recognition</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>74</volume>(<issue>2</issue>):<fpage>4081</fpage>&#x2013;<lpage>98</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2023.032329</pub-id>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Tang</surname> <given-names>F</given-names></string-name>, <string-name><surname>Tang</surname> <given-names>T</given-names></string-name>, <string-name><surname>Zhu</surname> <given-names>H</given-names></string-name>, <string-name><surname>Hu</surname> <given-names>C</given-names></string-name>, <string-name><surname>Ma</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Li</surname> <given-names>X</given-names></string-name></person-group>. <article-title>Rainfall landslide deformation prediction based on attention mechanism and Bi-LSTM</article-title>. <source>Bull Surv Mapp</source>. <year>2022</year>;<volume>9</volume>:<fpage>74</fpage>&#x2013;<lpage>9</lpage>. (In Chinese) doi:<pub-id pub-id-type="doi">10.13474/j.cnki.11-2246.2022.0267</pub-id>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Chen</surname> <given-names>H</given-names></string-name>, <string-name><surname>Feng</surname> <given-names>X</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhao</surname> <given-names>H</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Guo</surname> <given-names>L</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Research on predicting surface displacement of landslides based on CNN-BiLSTM-Attention in the Three Gorges reservoir area</article-title>. <source>Sediment Geol Tethyan Geol</source>. <year>2024</year>;<volume>44</volume>(<issue>3</issue>):<fpage>572</fpage>&#x2013;<lpage>81</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.19826/j.cnki.1009-3850.2024.08006</pub-id>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Jiang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zheng</surname> <given-names>L</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Lu</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>Deformation mechanism-assisted deep learning architecture for predicting step-like displacement of reservoir landslide</article-title>. <source>Int J Appl Earth Obs Geoinf</source>. <year>2024</year>;<volume>133</volume>:<fpage>104121</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.jag.2024.104121</pub-id>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Xu</surname> <given-names>M</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>D</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>Y</given-names></string-name></person-group>. <article-title>An adaptive spatial-temporal prediction model for landslide displacement based on decomposition architecture</article-title>. <source>Eng Appl Artif Intell</source>. <year>2024</year>;<volume>137</volume>(<issue>B</issue>):<fpage>109215</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.engappai.2024.109215</pub-id>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ge</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>X</given-names></string-name>, <string-name><surname>Deng</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>K</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>H</given-names></string-name></person-group>. <article-title>LiteTransNet: an interpretable approach for landslide displacement prediction using transformer model with attention mechanism</article-title>. <source>Eng Geol</source>. <year>2024</year>;<volume>331</volume>(<issue>7</issue>):<fpage>107446</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.enggeo.2024.107446</pub-id>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>L</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhou</surname> <given-names>T</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>M</given-names></string-name></person-group>. <article-title>Data-driven combination-interval prediction for landslide displacement based on copula and VMD-WOA-KELM method</article-title>. <source>J Earth Sci</source>. <year>2025</year>;<volume>36</volume>(<issue>1</issue>):<fpage>291</fpage>&#x2013;<lpage>306</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s12583-021-1555-3</pub-id>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ge</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name>, <string-name><surname>Lacasse</surname> <given-names>S</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>H</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>Data-augmented landslide displacement prediction using generative adversarial network</article-title>. <source>J Rock Mech Geotechnical Eng</source>. <year>2024</year>;<volume>16</volume>(<issue>10</issue>):<fpage>4017</fpage>&#x2013;<lpage>33</lpage>. doi:<pub-id pub-id-type="doi">10.1016/j.jrmge.2024.01.003</pub-id>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Xie</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>G</given-names></string-name>, <string-name><surname>Cao</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Miao</surname> <given-names>F</given-names></string-name></person-group>. <article-title>Fractal characteristics of displacement and cracks in the Baishuihe landslide in the Three Gorges Reservoir Area</article-title>. <source>Bullet Geolog Sci Technol</source>. <year>2024</year>;<volume>43</volume>(<issue>4</issue>):<fpage>244</fpage>&#x2013;<lpage>51</lpage>. (In Chinese). doi:<pub-id pub-id-type="doi">10.19509/j.cnki.dzkq.tb20230166</pub-id>.</mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Nava</surname> <given-names>L</given-names></string-name>, <string-name><surname>Carraro</surname> <given-names>E</given-names></string-name>, <string-name><surname>Reyes-Carmona</surname> <given-names>C</given-names></string-name>, <string-name><surname>Puliero</surname> <given-names>S</given-names></string-name>, <string-name><surname>Bhuyan</surname> <given-names>K</given-names></string-name>, <string-name><surname>Rosi</surname> <given-names>A</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Landslide displacement forecasting using deep learning and monitoring data across selected sites</article-title>. <source>Landslides</source>. <year>2023</year>;<volume>20</volume>(<issue>10</issue>):<fpage>2111</fpage>&#x2013;<lpage>29</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10346-023-02104-9</pub-id>.</mixed-citation></ref>
<ref id="ref-35"><label>[35]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Meng</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Qin</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Cai</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Tian</surname> <given-names>B</given-names></string-name>, <string-name><surname>Yuan</surname> <given-names>C</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>X</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Dynamic forecast model for landslide displacement with step-like deformation by applying GRU with EMD and error correction</article-title>. <source>Bull Eng Geol Environ</source>. <year>2023</year>;<volume>82</volume>(<issue>6</issue>):<fpage>211</fpage>. doi:<pub-id pub-id-type="doi">10.1007/s10064-023-03247-8</pub-id>.</mixed-citation></ref>
<ref id="ref-36"><label>[36]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Shi</surname> <given-names>F</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>H</given-names></string-name>, <string-name><surname>Yu</surname> <given-names>L</given-names></string-name>, <string-name><surname>Hu</surname> <given-names>F</given-names></string-name></person-group>. <source>MATLAB intelligent algorithms: 30 case studies</source>. <publisher-loc>Beijing, China</publisher-loc>: <publisher-name>Beijing University of Aeronautics and Astronautics Press</publisher-name>; <year>2011</year>. <fpage>89</fpage> p.</mixed-citation></ref>
<ref id="ref-37"><label>[37]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wang</surname> <given-names>H</given-names></string-name>, <string-name><surname>Long</surname> <given-names>G</given-names></string-name>, <string-name><surname>Shao</surname> <given-names>P</given-names></string-name>, <string-name><surname>Lv</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Gan</surname> <given-names>F</given-names></string-name>, <string-name><surname>Liao</surname> <given-names>J</given-names></string-name></person-group>. <article-title>A DES-BDNN based probabilistic forecasting approach for step-like landslide displacement</article-title>. <source>J Clean Prod</source>. <year>2023</year>;<volume>394</volume>(<issue>3</issue>):<fpage>136281</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.jclepro.2023.136281</pub-id>.</mixed-citation></ref>
<ref id="ref-38"><label>[38]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Hochreiter</surname> <given-names>S</given-names></string-name>, <string-name><surname>Schmidhuber</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Long short-term memory</article-title>. <source>Neural Comput</source>. <year>1997</year>;<volume>9</volume>(<issue>8</issue>):<fpage>1735</fpage>&#x2013;<lpage>80</lpage>. doi:<pub-id pub-id-type="doi">10.1162/neco.1997.9.8.1735</pub-id>; <pub-id pub-id-type="pmid">9377276</pub-id></mixed-citation></ref>
<ref id="ref-39"><label>[39]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Liu</surname> <given-names>J</given-names></string-name></person-group>. <source>Intelligent control</source>. <edition>4th ed</edition>. <publisher-loc>Beijing, China</publisher-loc>: <publisher-name>Publishing House of Electronics Industry</publisher-name>; <year>2017</year>. <fpage>135</fpage> p. (In Chinese)</mixed-citation></ref>
<ref id="ref-40"><label>[40]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Meng</surname> <given-names>S</given-names></string-name>, <string-name><surname>Shi</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Peng</surname> <given-names>M</given-names></string-name>, <string-name><surname>Li</surname> <given-names>G</given-names></string-name>, <string-name><surname>Zheng</surname> <given-names>H</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>L</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Landslide displacement prediction with step-like curve based on convolutional neural network coupled with bi-directional gated recurrent unit optimized by attention mechanism</article-title>. <source>Eng Appl Artif Intell</source>. <year>2024</year>;<volume>133</volume>(<issue>A</issue>):<fpage>108078</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.engappai.2024.108078</pub-id>.</mixed-citation></ref>
<ref id="ref-41"><label>[41]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Bai</surname> <given-names>D</given-names></string-name>, <string-name><surname>Lu</surname> <given-names>G</given-names></string-name>, <string-name><surname>Zhu</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Zhu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Tao</surname> <given-names>C</given-names></string-name>, <string-name><surname>Fang</surname> <given-names>J</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Prediction interval estimation of landslide displacement using bootstrap, variational mode decomposition, and long and short-term time-series network</article-title>. <source>Remote Sens</source>. <year>2022</year>;<volume>14</volume>(<issue>22</issue>):<fpage>5808</fpage>. doi:<pub-id pub-id-type="doi">10.3390/rs14225808</pub-id>.</mixed-citation></ref>
<ref id="ref-42"><label>[42]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Miao</surname> <given-names>F</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>L</given-names></string-name>, <string-name><surname>Xue</surname> <given-names>Y</given-names></string-name></person-group>. <article-title>Landslide displacement interval prediction based on different Bootstrap methods and KELM-BPNN model</article-title>. <source>Chin J Rock Mech Eng</source>. <year>2019</year>;<volume>38</volume>(<issue>5</issue>):<fpage>912</fpage>&#x2013;<lpage>26</lpage>. (In Chinese) doi:<pub-id pub-id-type="doi">10.13722/j.cnki.jrme.2018.1380</pub-id>.</mixed-citation></ref>
<ref id="ref-43"><label>[43]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>D</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Yin</surname> <given-names>K</given-names></string-name>, <string-name><surname>Miao</surname> <given-names>F</given-names></string-name>, <string-name><surname>Glade</surname> <given-names>T</given-names></string-name>, <string-name><surname>Leo</surname> <given-names>C</given-names></string-name></person-group>. <article-title>Displacement characteristics and prediction of Baishuihe landslide in the Three Gorges Reservoir</article-title>. <source>J Mt Sci</source>. <year>2019</year>;<volume>16</volume>(<issue>9</issue>):<fpage>2203</fpage>&#x2013;<lpage>14</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s11629-019-5470-3</pub-id>.</mixed-citation></ref>
<ref id="ref-44"><label>[44]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Li</surname> <given-names>B</given-names></string-name>, <string-name><surname>Miao</surname> <given-names>F</given-names></string-name>, <string-name><surname>Deng</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>Adaptive hybrid machine learning model for forecasting the step-like displacement of reservoir colluvial landslides: a case study in the Three Gorges reservoir area</article-title>. <source>China Stoch Environ Res Risk Assess</source>. <year>2023</year>;<volume>37</volume>(<issue>3</issue>):<fpage>903</fpage>&#x2013;<lpage>23</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s00477-022-02322-y</pub-id>.</mixed-citation></ref>
</ref-list>
</back></article>