<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">IASC</journal-id>
<journal-id journal-id-type="nlm-ta">IASC</journal-id>
<journal-id journal-id-type="publisher-id">IASC</journal-id>
<journal-title-group>
<journal-title>Intelligent Automation &#x0026; Soft Computing</journal-title>
</journal-title-group>
<issn pub-type="epub">2326-005X</issn>
<issn pub-type="ppub">1079-8587</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">22812</article-id>
<article-id pub-id-type="doi">10.32604/iasc.2022.022812</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Fine-Grained Bandwidth Estimation for Smart Grid Communication Network</article-title><alt-title alt-title-type="left-running-head">Fine-Grained Bandwidth Estimation for Smart Grid Communication Network</alt-title><alt-title alt-title-type="right-running-head">Fine-Grained Bandwidth Estimation for Smart Grid Communication Network</alt-title>
</title-group>
<contrib-group content-type="authors">
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Luo</surname><given-names>Jingtang</given-names></name>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib id="author-2" contrib-type="author">
<name name-style="western"><surname>Liao</surname><given-names>Jingru</given-names></name>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Zhang</surname><given-names>Chenlin</given-names></name>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Wang</surname><given-names>Ziqi</given-names></name>
<xref ref-type="aff" rid="aff-4">4</xref>
</contrib>
<contrib id="author-5" contrib-type="author">
<name name-style="western"><surname>Zhang</surname><given-names>Yuhang</given-names></name>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib id="author-6" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Xu</surname><given-names>Jie</given-names></name>
<xref ref-type="aff" rid="aff-2">2</xref><email>xuj@uestc.edu.cn</email>
</contrib>
<contrib id="author-7" contrib-type="author">
<name name-style="western"><surname>Huang</surname><given-names>Zhengwen</given-names></name>
<xref ref-type="aff" rid="aff-5">5</xref>
</contrib>
<aff id="aff-1"><label>1</label><institution>State Grid Sichuan Economic Research Institute</institution>, <addr-line>Chengdu, 610041</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>School of Information and Communication Engineering, University of Electronic Science and Technology of China</institution>, <addr-line>Chengdu, 611731</addr-line>, <country>China</country></aff>
<aff id="aff-3"><label>3</label><institution>Science and Technology on Security Communication Laboratory, Institute of Southwestern Communication</institution>, <addr-line>Chengdu, 610093</addr-line>, <country>China</country></aff>
<aff id="aff-4"><label>4</label><institution>State Grid Economic and Technological Research Institute CO., LTD</institution>, <addr-line>Beijing, 100055</addr-line>, <country>China</country></aff>
<aff id="aff-5"><label>5</label><institution>Department of Electronic and Computer Engineering, Brunel University</institution>, <addr-line>Uxbridge, UB8 3PH</addr-line>, <country>United Kingdom</country></aff>
</contrib-group><author-notes><corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Jie Xu. Email: <email>xuj@uestc.edu.cn</email></corresp></author-notes>
<pub-date pub-type="epub" date-type="pub" iso-8601-date="2021-11-3"><day>3</day>
<month>11</month>
<year>2021</year></pub-date>
<volume>32</volume>
<issue>2</issue>
<fpage>1225</fpage>
<lpage>1239</lpage>
<history>
<date date-type="received"><day>19</day><month>8</month><year>2021</year></date>
<date date-type="accepted"><day>11</day><month>10</month><year>2021</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2022 Luo et al.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Luo et al.</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_IASC_22812.pdf"></self-uri>
<abstract>
<p>Accurate estimation of communication bandwidth is critical for the sensing and controlling applications of smart grid. Different from public network, the bandwidth requirements of smart grid communication network must be accurately estimated in prior to the deployment of applications or even the building of communication network. However, existing methods for smart grid usually model communication nodes in coarse-grained ways, so their estimations become inaccurate in scenarios where the same type of nodes have very different bandwidth requirements. To solve this issue, we propose a fine-grained estimation method based on multivariate nonlinear fitting. Firstly, we use linear fitting to calculate the convergence weights of each node. Then, we use correlation to select the important characteristics. Finally, we use multivariate nonlinear fitting to learn the nonlinear relationship between characteristics and convergence weight, and complete the fine-grained bandwidth estimation. Our method exploits multiple node characteristics to reveal how different nodes affect bandwidth requirements differently, and it can learn multivariate estimation parameters from present network without human interference. We use NS2 to simulate a real-world regional smart grid. Simulation shows that our method outperforms existing works by up to 56.5&#x0025; higher estimation accuracy.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Bandwidth estimation</kwd>
<kwd>fine-grained</kwd>
<kwd>multivariate nonlinear fitting</kwd>
<kwd>smart grid communication network</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>Smart grid empowers modern society by creating the foundation necessary for electric transportation, energy efficiency, emissions reductions, and new energy technologies. Private communication networks are widely used by smart grid to deliver massive sensing and controlling data for critical applications like power metering, environment monitoring, and power dispatching. Different from the applications in public networks (e.g., social media), the applications in smart gird usually have very stringent communication QoS (Quality of Service) requirements. For instance, dispatching application demands that transmission delay must be lower than 100&#x2005;ms and transmission error rate must be lower than 10<sup>&#x2212;8</sup>. As a result, to meet these applications&#x2019; QoS demands, the bandwidth requirements of each communication node must be accurately estimated in prior to the deployment of applications or even the building of communication network.</p>
<p>The most reasonable idea for bandwidth estimation is using present networks&#x2019; bandwidth consumption information to estimate new networks&#x2019; bandwidth requirements. Based on this idea, [<xref ref-type="bibr" rid="ref-1">1</xref>] and [<xref ref-type="bibr" rid="ref-2">2</xref>] propose elastic coefficient method, which has been widely used in practice due its ease of use. However, because elastic coefficient method assumes that all data are uploaded to a few core nodes (namely, the dispatching centers of smart grid), it often results in significant overestimation of bandwidth demands.</p>
<p>To solve the above problem, some works like [<xref ref-type="bibr" rid="ref-3">3</xref>&#x2013;<xref ref-type="bibr" rid="ref-7">7</xref>] exploit importance recognition methods to reveal the influences of different nodes on bandwidth estimation. However, they mainly identify important nodes based on the physical topology of the network, such as node centrality, K-shell, structure hole, PageRank. But to accurately analyze node importance, the applications on each node should also be considered explicitly.</p>
<p>There are some other ways to improve bandwidth estimation. For instance, [<xref ref-type="bibr" rid="ref-8">8</xref>] proposes a new method of optimizing bandwidth calculation. In [<xref ref-type="bibr" rid="ref-9">9</xref>], the bandwidth of each application is estimated and accumulated to obtain the bandwidth of a single node. Although these works increase bandwidth estimation accuracy to some extent, they still have some shortcomings such as ignoring the characteristics of different nodes and relying on human experience for parameter selection.</p>
<p>In this paper, we propose a novel fine-grained bandwidth estimation method for smart grid. Compared to present works, our method achieves up to 56.5&#x0025; higher estimation accuracy. Such performance is mainly due to the following two novelties:<list list-type="simple"><list-item><label>1)</label>
<p>Our method divides the characteristics and studying the influence of different characteristics of each node. Our method explicitly reveals how data converge from outer nodes to core nodes in smart grid, and how such convergence is affected by each node&#x2019;s characteristics (e.g., number and type of applications). As a result, our method can provide fine-grained bandwidth estimation for different nodes;</p></list-item><list-item><label>2)</label>
<p>The parameter setting in our method requires no human interference. The parameters are learned through multiple iterations. Our method exploits multivariate nonlinear fitting to learn parameter settings from present network. Since our method can learn multiple node characteristics as well as the nonlinear relationships among these characteristics without human interference, it achieves higher estimation accuracy especially in heterogeneous networks where nodes have highly diverse bandwidth requirements.</p></list-item></list></p>
<p>The rest of this paper is organized as follows. Section 2 introduces the existing researches related to our work. Section 3 introduces two important features of smart grid communication network. Section 4 proposes a fine-grained bandwidth estimation method. Section 5 proposes a multivariate nonlinear fitting scheme to learn estimation parameters. Section 6 exploits simulations to validate the accuracy of our estimation method. Section 7 concludes this paper.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Related Works</title>
<p>The most popular methods estimate bandwidth according to node&#x2019;s voltage level [<xref ref-type="bibr" rid="ref-1">1</xref>&#x2013;<xref ref-type="bibr" rid="ref-2">2</xref>,<xref ref-type="bibr" rid="ref-10">10</xref>,<xref ref-type="bibr" rid="ref-11">11</xref>]. In general, these methods assume that the grid nodes with the same voltage level (e.g., 220&#x2005;kV substations) have very close bandwidth demands for uploading data to their upper nodes. While these methods have very simple estimation process, they often make significant overestimation of bandwidth demands especially for the core nodes like dispatching centers, due to the fact that nodes with the same voltage level probably have very different bandwidth requirements.</p>
<p>Recognizing the importance of different nodes based on network topology [<xref ref-type="bibr" rid="ref-12">12</xref>] is an effective way to improve bandwidth estimation. Reference [<xref ref-type="bibr" rid="ref-13">13</xref>] introduces several types of node centralities, such as degree centrality, close centrality, intermediate centrality, and eigenvector centrality. Furthermore, several centrality indicators may be used together to comprehensively analyze the importance of a node. In [<xref ref-type="bibr" rid="ref-14">14</xref>], a K-shell algorithm is used to calculate the influence of nodes in the network. In [<xref ref-type="bibr" rid="ref-15">15</xref>], an E-Burt algorithm based on structural holes is proposed, which sets the weight of the edge as the edge connection. Reference [<xref ref-type="bibr" rid="ref-16">16</xref>] uses PageRank algorithm to obtain the node weight to replace the node degree matrix in the centrality, and determines the importance of nodes in the network through the improved centrality. Reference [<xref ref-type="bibr" rid="ref-4">4</xref>] improves the traditional calculation method by evaluating the importance of power communication network nodes based on node strength and node tightness.</p>
<p>Some works analyze how different applications affect bandwidth requirements. Reference [<xref ref-type="bibr" rid="ref-8">8</xref>] improves estimation accuracy in tree-structured networks through selecting concurrent proportions for different applications. In reference [<xref ref-type="bibr" rid="ref-9">9</xref>], the bandwidth of each service is estimated and accumulated to obtain the bandwidth of each node. This work uses no machine learning technologies, and it focuses on estimating bandwidth of single node rather than whole network. Reference [<xref ref-type="bibr" rid="ref-17">17</xref>] proposes a passive capacity and available bandwidth measurement method for the data plane, employing packet dispersion and autocorrelation.</p>
<p>However, as far as we know, the existing works only consider one or two node characteristics (e.g., voltage level, topology, applications, etc.), which makes their estimations coarse-grained and thereby inaccurate especially in heterogeneous networks with highly diverse nodes. Moreover, since many of the existing works rely on expert experience to select and configure estimation parameters, they are less adaptive to rapidly developing smart grids with more advanced applications like demand response [<xref ref-type="bibr" rid="ref-18">18</xref>], integrating renewable energy [<xref ref-type="bibr" rid="ref-19">19</xref>], and cyber security [<xref ref-type="bibr" rid="ref-20">20</xref>].</p>
<p>Machine learning is one of today&#x2019;s most rapidly growing technical fields [<xref ref-type="bibr" rid="ref-21">21</xref>&#x2013;<xref ref-type="bibr" rid="ref-23">23</xref>]. Traditional machine learning models, such as logistic regression [<xref ref-type="bibr" rid="ref-24">24</xref>], support vector machine [<xref ref-type="bibr" rid="ref-25">25</xref>], and decision tree [<xref ref-type="bibr" rid="ref-26">26</xref>], are based on statistical learning theories. These models have high interpretability (i.e., a human can easily understand the models&#x2019; behaviors) [<xref ref-type="bibr" rid="ref-27">27</xref>,<xref ref-type="bibr" rid="ref-28">28</xref>] and are relatively simple to train. In recent years, deep learning models based on artificial neural networks have achieved outstanding performances for many difficult tasks like computer vision [<xref ref-type="bibr" rid="ref-29">29</xref>&#x2013;<xref ref-type="bibr" rid="ref-32">32</xref>], medical diagnosis [<xref ref-type="bibr" rid="ref-33">33</xref>], translation [<xref ref-type="bibr" rid="ref-34">34</xref>], path planning [<xref ref-type="bibr" rid="ref-35">35</xref>] and semantic understanding [<xref ref-type="bibr" rid="ref-36">36</xref>&#x2013;<xref ref-type="bibr" rid="ref-39">39</xref>]. However, deep learning models still lack sufficient interpretability until now [<xref ref-type="bibr" rid="ref-27">27</xref>]. In this paper, we exploit traditional logistic regression model (namely, nonlinear fitting) to estimate bandwidth requirements, because power grid is a highly regulated domain where the interpretability of decisions is mandatory. In fact, our nonlinear fitting method is able to provide rather accurate estimations in complex smart grid scenarios, as will be proved by simulations later.</p>
</sec>
<sec id="s3">
<label>3</label>
<title>Features of Smart Grid Communication Network</title>
<p>Unlike public network, communication network in smart grid is built according to the structure and the applications of smart grid, thus it has the following two distinct features, as <xref ref-type="fig" rid="fig-1">Fig. 1</xref> illustrates.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Hierarchical structure and converged data flow of smart grid communication network</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-1.png"/>
</fig>
<p>Firstly, communication nodes of smart grid are usually built on electricity substations, and communication links are usually built along electricity cables. As a result, smart grid communication network has a hierarchal tree structure, where lower-voltage nodes connect to higher-voltage nodes, and the latter connect to dispatching centers.</p>
<p>Secondly, as lower-voltage substations generate application data, some of these data are aggregated to higher-voltage substations (these higher-voltage substations may also generate some data to upload), and eventually aggregated to dispatching centers. Hence, bandwidth requirements hierarchically converge from lower-voltage substations to dispatching centers.</p>
<p>For a new smart grid communication network, we often just know the number and the bandwidth demands of the applications on each node. The bandwidth requirements from lower-voltage nodes to higher-voltage node are unknown and need to be estimated, as discussed in the next section.</p>
</sec>
<sec id="s4">
<label>4</label>
<title>Fine-Grained Bandwidth Estimation Method</title>
<p>Based on the aforementioned features, we propose a fine-grained method for estimating bandwidth requirements of communication nodes in smart grid. We divide and study the characteristics of each node to obtain a more accurate bandwidth estimation method. <xref ref-type="table" rid="table-1">Tab. 1</xref> summarizes the variables used in this paper.</p>
<table-wrap id="table-1"><label>Table 1</label>
<caption>
<title>Summary of variables</title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Variable notation</th>
<th align="left">Definition</th>
<th align="left">Type</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left"><inline-formula id="ieqn-5">
<mml:math id="mml-ieqn-5"><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow></mml:math>
</inline-formula></td>
<td align="left">Estimated converged bandwidth of higher-voltage node.</td>
<td align="left">Estimation result</td>
</tr>
<tr>
<td align="left"><italic>y</italic></td>
<td align="left">Converged bandwidth of higher-voltage node.</td>
<td align="left">Known parameter (only for learning)</td>
</tr>
<tr>
<td align="left"><italic>N</italic></td>
<td align="left">Number of lower-voltage nodes.</td>
<td align="left">Known parameter</td>
</tr>
<tr>
<td align="left"><italic>B</italic><sub><italic>i</italic></sub></td>
<td align="left">Bandwidth requirement of lower-voltage node.</td>
<td align="left">Known parameter</td>
</tr>
<tr>
<td align="left"><italic>&#x03C9;</italic><sub><italic>i</italic></sub></td>
<td align="left">Convergence weight of lower-voltage node.</td>
<td align="left">Coefficient to learn</td>
</tr>
<tr>
<td align="left"><inline-formula id="ieqn-6">
<mml:math id="mml-ieqn-6"><mml:msub><mml:mrow><mml:mover><mml:mi>&#x03C9;</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:mi>i</mml:mi></mml:msub></mml:math>
</inline-formula></td>
<td align="left">Estimated Convergence weight of lower-voltage node.</td>
<td align="left">Coefficient to learn</td>
</tr>
<tr>
<td align="left"><italic>X</italic><sub><italic>j</italic></sub></td>
<td align="left">Node characteristic that may affect bandwidth convergence.</td>
<td align="left">Coefficient to learn</td>
</tr>
<tr>
<td align="left"><italic>&#x03B2;</italic><sub><italic>j</italic></sub></td>
<td align="left">Coefficient relating node characteristics to node convergence weight.</td>
<td align="left">Coefficient to learn</td>
</tr>
<tr>
<td align="left"><italic>k</italic></td>
<td align="left">Number of key node characteristics</td>
<td align="left">Hyperparameter</td>
</tr>
<tr>
<td align="left"><italic>n</italic></td>
<td align="left">Power of nonlinear fitting from key characteristics to convergence weight.</td>
<td align="left">Hyperparameter</td>
</tr>
<tr>
<td align="left"><italic>&#x03B1;</italic><sub>1</sub></td>
<td align="left">Learning rate of gradient descent of linear fitting.</td>
<td align="left">Hyperparameter</td>
</tr>
<tr>
<td align="left"><italic>&#x03B1;</italic><sub>2</sub></td>
<td align="left">Learning rate of gradient descent of nonlinear fitting.</td>
<td align="left">Hyperparameter</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As shown by <xref ref-type="fig" rid="fig-2">Fig. 2</xref>, our basic idea is using lower-voltage nodes&#x2019; bandwidth requirements (which are derived from application requirements) to estimate the bandwidth requirements of higher-voltage nodes, and then use these bandwidth estimations to further estimate the bandwidth requirements of even higher-voltage nodes, and eventually estimate the bandwidth requirements of dispatching centers.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Fine-grained bandwidth estimation method</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-2.png"/>
</fig>
<p>Specifically, the bandwidth of an upper node (a higher-voltage node or a dispatching center) can be estimated as follows:<disp-formula id="eqn-1"><label>(1)</label>
<mml:math id="mml-eqn-1" display="block"><mml:mover><mml:mi>y</mml:mi><mml:mo accent="false">&#x00AF;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mrow><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:msub><mml:mi>B</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:mspace width="1em" /><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x2208;</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:mspace width="thickmathspace" /><mml:mn>1</mml:mn><mml:mo stretchy="false">]</mml:mo></mml:math>
</disp-formula>where <inline-formula id="ieqn-1">
<mml:math id="mml-ieqn-1"><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow></mml:math>
</inline-formula> is the estimated converged bandwidth of the upper node, <italic>B<sub>1</sub>, B<sub>2</sub>, &#x2026;, B<sub>N</sub></italic> are the total bandwidth requirements of <italic>N</italic> lower nodes (i.e., data uploaded by even lower nodes plus data generated by the node itself), <italic>w<sub>1</sub>, w<sub>2</sub>, &#x2026;, w<sub>N</sub></italic> are the convergence weights of the <italic>N</italic> nodes (i.e., the ratio of data to upload).</p>
<p>The convergence weight <italic>w<sub>i</sub></italic> of node <italic>i</italic> must lie in [0,1] because a lower-voltage node can never transmit data more than its own bandwidth. The value of <italic>w<sub>i</sub></italic> can be derived from <italic>k</italic> characteristics of node <italic>i</italic>:<disp-formula id="eqn-2"><label>(2)</label>
<mml:math id="mml-eqn-2" display="block"><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>k</mml:mi></mml:munderover><mml:mrow><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mi>X</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>&#x2026;</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:math>
</disp-formula>where <italic>X</italic><sub>1</sub>, <italic>X</italic><sub>2</sub>, &#x2026;, <italic>X<sub>k</sub></italic> are node characteristics, <italic>f</italic>(&#x2009;&#x22C5;&#x2009;) is a multinomial function of single characteristic, <italic>g</italic>(&#x2009;&#x22C5;&#x2009;) is a multinomial function of multiple characteristics, <italic>&#x03B2;</italic><sub><italic>j</italic></sub> and <italic>&#x03B2;</italic><sub><italic>k</italic>&#x002B;1</sub> are coefficients relating node characteristics to node convergence weight.</p>
<p><xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref> reveal the bandwidth convergence of lower-voltage nodes to higher-voltage nodes (and dispatching centers), and how such convergence is affected by multiple node characteristics. This way, we achieve fine-grained estimations corresponding to the differences among nodes.</p>
<p>Notably, we have not determined which node characteristics should be considered, and how these characteristics are related to node convergence weight. These two issues will be solved by multivariate nonlinear fitting in the next section.</p>
</sec>
<sec id="s5">
<label>5</label>
<title>Multivariate Nonlinear Fitting Scheme</title>
<p>In this section, we propose a multivariate nonlinear fitting scheme to relating node characteristics to node bandwidth. In other words, our scheme will learn all the undetermined coefficients in <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref> from measured node bandwidth and characteristics.</p>
<p>As <xref ref-type="fig" rid="fig-3">Fig. 3</xref> illustrates, our fitting scheme has three major steps. The first step is linear fitting, which takes converged bandwidth as target value to construct the loss function. Linear fitting uses the gradient descent method to derive convergence weight for each node. The second step is correlation coefficient calculation, which is using correlation coefficient to find key characteristics with more significant impacts on convergence weight across network. The last step is multivariate nonlinear fitting. This step reveals the nonlinear relationship among multiple characteristics and the convergence weight of each node, and therefore relates node characteristics to node bandwidth.</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>Multivariate nonlinear fitting scheme relating node characteristics to node bandwidth</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-3.png"/>
</fig>
<sec id="s5_1">
<label>5.1</label>
<title>Linear Learning</title>
<p>Here we utilize gradient descent to learn node convergence weight from real bandwidth data. In brief, we iteratively calculate the cost between the estimated bandwidth (which is derived from node convergence weight) and the actual bandwidth, and update convergence weight with gradient descent [<xref ref-type="bibr" rid="ref-40">40</xref>], until the cost becomes minimal.</p>
<p>First, we initialize node convergence weights as small positive numbers that are randomly generated within (0, 1), and substitute these weights and the known bandwidth requirements of lower nodes into <xref ref-type="disp-formula" rid="eqn-1">Eq. (1)</xref> to derive the estimated bandwidth of the upper node <inline-formula id="ieqn-2">
<mml:math id="mml-ieqn-2"><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow></mml:math>
</inline-formula>.</p>
<p>Then we calculate the <italic>cost</italic> for gradient descent as follows:<disp-formula id="eqn-3"><label>(3)</label>
<mml:math id="mml-eqn-3" display="block"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mo>=</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>y</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:msup><mml:mo stretchy="false">)</mml:mo><mml:mn>2</mml:mn></mml:msup></mml:math>
</disp-formula>where <italic>y</italic> is the actual bandwidth of the converged node.</p>
<p>The idea of gradient descent is to minimize <italic>cost</italic> by gradually adjusting node convergence weights. Specifically, the gradient of the cost function can be computed as:<disp-formula id="eqn-4"><label>(4)</label>
<mml:math id="mml-eqn-4" display="block"><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mrow><mml:mfrac><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo>=</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>y</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>y</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:mo stretchy="false">)</mml:mo><mml:mo>&#x00D7;</mml:mo><mml:msub><mml:mi>B</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mstyle></mml:math>
</disp-formula>where <italic>w<sub>i</sub></italic> and <italic>B<sub>i</sub></italic> are the convergence weight and the bandwidth of the node <italic>i</italic>, respectively.</p>
<p>We gradually adjust the convergence weight of each node as follows:<disp-formula id="eqn-5"><label>(5)</label>
<mml:math id="mml-eqn-5" display="block"><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy="false">&#x2190;</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>&#x03B1;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>&#x00D7;</mml:mo><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mrow><mml:mfrac><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle></mml:math>
</disp-formula>where <italic>&#x03B1;</italic><sub>1</sub> is the learning rate of gradient descent. Its value will be determined by experiments later.</p>
<p><xref ref-type="disp-formula" rid="eqn-4">Eqs. (4)</xref> and <xref ref-type="disp-formula" rid="eqn-5">(5)</xref> will be executed iteratively until the cost function converges to its minimum. The obtained convergence weight <italic>w<sub>i</sub></italic> essentially reflects the radio of node <italic>i</italic>&#x2019;s bandwidth that converges to its upper node.</p>
</sec>
<sec id="s5_2">
<label>5.2</label>
<title>Correlation Coefficient Calculation</title>
<p>We use correlation coefficient calculation to decide which node characteristics (e.g., number or type of application) have more significant impacts on convergence weight. Only these key node characteristics will be concerned for bandwidth estimation. This is to simplify the acquirement process of node characteristics as well as the subsequent non-linear multivariate learning.</p>
<p>Specifically, for all <italic>N</italic> nodes, we calculate correlation coefficient between each candidate characteristic and the set of convergence weights as follows:<disp-formula id="eqn-6"><label>(6)</label>
<mml:math id="mml-eqn-6" display="block"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>r</mml:mi><mml:mi>r</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi mathvariant="bold">w</mml:mi></mml:mrow><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mrow><mml:mi mathvariant="bold">X</mml:mi></mml:mrow><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mrow><mml:mfrac><mml:mrow><mml:mrow><mml:mi>cov</mml:mi></mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi mathvariant="bold">w</mml:mi></mml:mrow><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mrow><mml:mi mathvariant="bold">X</mml:mi></mml:mrow><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mrow><mml:msqrt><mml:mrow><mml:mi>var</mml:mi></mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi mathvariant="bold">w</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo><mml:mo>&#x22C5;</mml:mo><mml:mrow><mml:mi>var</mml:mi></mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mrow><mml:mi mathvariant="bold">X</mml:mi></mml:mrow><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:msqrt></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle></mml:math>
</disp-formula>where <bold>w</bold> is the set of the convergence weights of all nodes (obtained in Section 5.1), <bold>X</bold><italic><sub>j</sub></italic> is the set of the values of a candidate characteristic for all nodes, cov(<bold>w</bold>, <bold>X</bold><italic><sub>j</sub></italic>) is the covariance of <bold>w</bold> and <bold>X</bold><italic><sub>j</sub></italic>, and var(<bold>w</bold>) and var(<bold>X</bold><italic><sub>j</sub></italic>) are the variance of <bold>w</bold> and <bold>X</bold><italic><sub>j</sub></italic>, respectively.</p>
<p>After calculating the correlation coeffecients for all candidate characteristics, we choose the key characteristics with higher correlation coeffecients for the non-linear multivariate learning in the next subsection. The number of key characteristics should be carefuly decided to balance between implemation complexity and bandwidth estimation accuracy. This will be further disccussed in Section 6.</p>
</sec>
<sec id="s5_3">
<label>5.3</label>
<title>Nonlinear Multivariate Fitting</title>
<p>At last, we exploit nonlinear multivariate fitting to reveal how a node&#x2019;s convergence weight is affected by its characteristics. Here, &#x201C;nonlinear&#x201D; is to reflect the complexity of such relationship, and &#x201C;multivariate&#x201D; is to reflect the combined impacts of multiple node characteristics. Multivariate nonlinear fitting means using mathematical model to express the nonlinear relationship between different characteristics and the convergence weight.</p>
<p>Since we have derived the convergence weight in Section 5.1 and the key characteristics in Section 5.2, we only need to determine the rest unknown parameters in <xref ref-type="disp-formula" rid="eqn-2">Eq. (2)</xref>, namely, the multinomial functions <italic>f</italic>(&#x2009;&#x22C5;&#x2009;) and <italic>g</italic>(&#x2009;&#x22C5;&#x2009;), and the coefficients <italic>&#x03B2;</italic><sub><italic>j</italic></sub> and <italic>&#x03B2;</italic><sub><italic>k</italic>&#x002B;1</sub>. This is accomplished through fitting <xref ref-type="disp-formula" rid="eqn-2">Eq. (2)</xref> to the convergence weight obtained in <xref ref-type="disp-formula" rid="eqn-5">Eq. (5)</xref>.</p>
<p>First, suppose that we have decided there are <italic>k</italic> key characteristics and the power of <italic>f</italic>(&#x2009;&#x22C5;&#x2009;) is <italic>n</italic> (we will discuss how to derive them later). The nonlinear influence of each characteristic <italic>X<sub>j</sub></italic> on the convergence weight can be expressed by:<disp-formula id="eqn-7"><label>(7)</label>
<mml:math id="mml-eqn-7" display="block"><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:msubsup><mml:mi>X</mml:mi><mml:mi>i</mml:mi><mml:mn>2</mml:mn></mml:msubsup><mml:mo>+</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>+</mml:mo><mml:msub><mml:mi>a</mml:mi><mml:mi>n</mml:mi></mml:msub><mml:msubsup><mml:mi>X</mml:mi><mml:mi>j</mml:mi><mml:mi>n</mml:mi></mml:msubsup></mml:math>
</disp-formula></p>
<p>and the nonlinearly correlated influence of the <italic>k</italic> key characteristics on the convergence weight can be expressed by:</p>
<p><disp-formula id="eqn-8"><label>(8)</label>
<mml:math id="mml-eqn-8" display="block"><mml:mi>g</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mi>X</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mi>X</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>3</mml:mn><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>3</mml:mn></mml:msub><mml:mo>&#x2026;</mml:mo><mml:mo>+</mml:mo><mml:msub><mml:mi>b</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:mi>k</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:msub><mml:mi>X</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>&#x2026;</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math>
</disp-formula>where the coefficients <italic>a</italic><sub>1</sub>, &#x2026;, <italic>a<sub>n</sub></italic>, <italic>b</italic><sub>(1,2)</sub>, &#x2026;, <italic>b</italic><sub>(1,2,&#x2026;,<italic>k</italic>)</sub> will be learned later. It is noteworthy that the complexity of <italic>g</italic>(&#x2009;&#x22C5;&#x2009;) is <italic>o</italic>(2<italic><sup>k</sup></italic>), which is why we should keep the number of node characteristics <italic>k</italic> as small as possible.</p>
<p>Second, we substitute <xref ref-type="disp-formula" rid="eqn-7">Eqs. (7)</xref> and <xref ref-type="disp-formula" rid="eqn-8">(8)</xref> into <xref ref-type="disp-formula" rid="eqn-2">Eq. (2)</xref> to express the fitted convergence weight <inline-formula id="ieqn-3">
<mml:math id="mml-ieqn-3"><mml:msub><mml:mrow><mml:mover><mml:mi>w</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:mi>i</mml:mi></mml:msub></mml:math>
</inline-formula> for node <italic>i</italic> as follows:<disp-formula id="eqn-9"><label>(9)</label>
<mml:math id="mml-eqn-9" display="block"><mml:msub><mml:mrow><mml:mover><mml:mi>w</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:mi>i</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>k</mml:mi></mml:munderover><mml:mrow><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mspace width="thickmathspace" /><mml:msub><mml:mi>X</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>&#x2026;</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:math>
</disp-formula></p>
<p>Third, we compare the fitted convergence weight <inline-formula id="ieqn-4">
<mml:math id="mml-ieqn-4"><mml:msub><mml:mrow><mml:mover><mml:mi>w</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow><mml:mi>i</mml:mi></mml:msub></mml:math>
</inline-formula> in <xref ref-type="disp-formula" rid="eqn-9">Eq. (9)</xref> to the convergence weight <italic>w<sub>i</sub></italic> derived by <xref ref-type="disp-formula" rid="eqn-5">Eq. (5)</xref>, i.e.,<disp-formula id="eqn-10"><label>(10)</label>
<mml:math id="mml-eqn-10" display="block"><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>s</mml:mi><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:mrow><mml:msup><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>w</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mrow><mml:mover><mml:mi>w</mml:mi><mml:mo stretchy="false">&#x00AF;</mml:mo></mml:mover></mml:mrow></mml:mrow><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:math>
</disp-formula></p>
<p>Now we can learn the undetermined parameters via multivariate gradient descent [<xref ref-type="bibr" rid="ref-10">10</xref>]:<disp-formula id="eqn-11"><label>(11)</label>
<mml:math id="mml-eqn-11" display="block"><mml:mi>&#x03B8;</mml:mi><mml:mo>=</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x2190;</mml:mo><mml:msub><mml:mi>&#x03B1;</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mrow><mml:mfrac><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mi mathvariant="normal">&#x2202;</mml:mi><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle></mml:math>
</disp-formula>where <italic>&#x03B1;</italic><sub>2</sub> is the learning rate, and its value will be determined by experiments later. <italic>&#x03B8;</italic> represents multiple variables <italic>a</italic><sub>1</sub>, &#x2026;, <italic>a<sub>n</sub></italic>, <italic>b</italic><sub>(1,2)</sub>, &#x2026;, <italic>b</italic><sub>(1,2,&#x2026;,<italic>k</italic>)</sub>, <italic>&#x03B2;</italic><sub>1</sub>, &#x2026;, <italic>&#x03B2;</italic><sub><italic>k</italic>&#x002B;1</sub>.</p>
<p>The above <xref ref-type="disp-formula" rid="eqn-10">Eqs. (10)</xref> and <xref ref-type="disp-formula" rid="eqn-11">(11)</xref> will be executed iteratively until the loss function <xref ref-type="disp-formula" rid="eqn-10">Eq. (10)</xref> converges to its minimum. By repeating this process on all node convergence weights, the values of <italic>a</italic><sub>1</sub>, &#x2026;, <italic>a<sub>n</sub></italic>, <italic>b</italic><sub>(1,2)</sub>, &#x2026;, <italic>b</italic><sub>(1,2,&#x2026;,<italic>k</italic>)</sub>, <italic>&#x03B2;</italic><sub>1</sub>, &#x2026;, <italic>&#x03B2;</italic><sub><italic>k</italic>&#x002B;1</sub> are learned.</p>
</sec>
<sec id="s5_4">
<label>5.4</label>
<title>Determining Hyperparameters</title>
<p>Finally, we determine the hyperparameters that should be set before learning, namely, the learning rates <italic>&#x03B1;</italic><sub>1</sub> in <xref ref-type="disp-formula" rid="eqn-5">Eq. (5)</xref> and <italic>&#x03B1;</italic><sub>2</sub> in <xref ref-type="disp-formula" rid="eqn-11">Eq. (11)</xref>, the number of the key node characteristics <italic>n</italic>, and the power of nonlinear fitting <italic>k</italic> in <xref ref-type="disp-formula" rid="eqn-7">Eq. (7)</xref>.</p>
<p>We begin with the learning rates <italic>&#x03B1;</italic><sub>1</sub> and <italic>&#x03B1;</italic><sub>2</sub>. We initially set them to very small values, e.g., 10<sup>&#x2212;8</sup>, and observe that whether the cost in <xref ref-type="disp-formula" rid="eqn-3">Eq. (3)</xref> or the lost in <xref ref-type="disp-formula" rid="eqn-10">Eq. (10)</xref> steadily decreases as we perform linear fitting in <xref ref-type="disp-formula" rid="eqn-4">Eq. (4)</xref> or nonlinear fitting in <xref ref-type="disp-formula" rid="eqn-11">Eq. (11)</xref>, respectively. If the decrement is too slow, we gradually increase <italic>&#x03B1;</italic><sub>1</sub> or <italic>&#x03B1;</italic><sub>2</sub> to learn the unknown parameters more drastically. On the other hand, if the decrement is unstable, we gradually decrease <italic>&#x03B1;</italic><sub>1</sub> or <italic>&#x03B1;</italic><sub>2</sub> to learn more cautiously. We keep adjusting <italic>&#x03B1;</italic><sub>1</sub> and <italic>&#x03B1;</italic><sub>2</sub> until the decrement is stable and notable. The resulting <italic>&#x03B1;</italic><sub>1</sub> and <italic>&#x03B1;</italic><sub>2</sub> will be used for learning later.</p>
<p>Afterwards, we determine the number of the key node characteristics <italic>n</italic>, and the power of nonlinear fitting <italic>k</italic>. For practical considerations, we should set them as small as possible (otherwise, there are too many variables to learn in <xref ref-type="disp-formula" rid="eqn-7">Eqs. (7)</xref> and <xref ref-type="disp-formula" rid="eqn-8">(8)</xref>). Therefore, we gradually increase them from <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>1 and <italic>k&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>1. For each pair of (<italic>n</italic>, <italic>k</italic>), we use <xref ref-type="disp-formula" rid="eqn-6">Eq. (6)</xref> to find the <italic>k</italic> key node characteristics, use <xref ref-type="disp-formula" rid="eqn-9">Eqs. (9)</xref>&#x2013;<xref ref-type="disp-formula" rid="eqn-11">(11)</xref> to derive the other unknown parameters in <xref ref-type="disp-formula" rid="eqn-7">Eqs. (7)</xref> and <xref ref-type="disp-formula" rid="eqn-8">(8)</xref>, and use <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref> to obtain bandwidth estimations for all upper nodes. This increment process stops as the bandwidth estimation accuracy has become reasonably high and the accuracy increment has become marginal. The pair of (<italic>n</italic>, <italic>k</italic>) that has the highest bandwidth estimation accuracy will be used for learning later.</p>
<p>After determining the hyperparameters <italic>&#x03B1;</italic><sub>1</sub>, <italic>&#x03B1;</italic><sub>2</sub>, <italic>n</italic>, and <italic>k</italic>, we can learn all parameters&#x2019; values in <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref> from a present network with <xref ref-type="disp-formula" rid="eqn-3">Eqs. (3)</xref>&#x2013;<xref ref-type="disp-formula" rid="eqn-11">(11)</xref>. Then we can use <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> and <xref ref-type="disp-formula" rid="eqn-2">(2)</xref> to estimate the bandwidth requirements of new networks.</p>
</sec>
</sec>
<sec id="s6">
<label>6</label>
<title>Simulation Results</title>
<p>We perform NS2 and TCL simulations to test our estimation method. Among them, NS2 is an open-source simulation platform for network technology. TCL is the script language on NS2.The simulated networks are based on a real-world regional smart grid in China. The network topologies are set as <xref ref-type="fig" rid="fig-4">Figs. 4</xref> and <xref ref-type="fig" rid="fig-6">6</xref>, and the applications are configured and deployed according to [<xref ref-type="bibr" rid="ref-1">1</xref>].</p>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Network for learning estimation parameters</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-4.png"/>
</fig>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Network for estimating bandwidth</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-6.png"/>
</fig>
<sec id="s6_1">
<label>6.1</label>
<title>Learning from Present Network</title>
<p>We learn the estimation parameters from the network in <xref ref-type="fig" rid="fig-4">Fig. 4</xref>. This network is based on the regional power grid of a moderate-sized city in China, which has 2 regional dispatching centers, 9 220&#x2005;kV substations, and 1 110&#x2005;kV substation. It represents a &#x201C;present network&#x201D; where the nodes&#x2019; bandwidth consumptions have been known. During learning, Z2 node is used for validating, and the rest nodes are used for fitting.</p>
<p>We first determine the linear learning rate <italic>&#x03B1;</italic><sub>1</sub> and the nonlinear learning rate <italic>&#x03B1;</italic><sub>2</sub>. We increase <italic>&#x03B1;</italic><sub>1</sub> and <italic>&#x03B1;</italic><sub>2</sub> from 10<sup>&#x2212;8</sup> to 10<sup>&#x2212;4</sup>, respectively, and find that the fitting costs of <xref ref-type="disp-formula" rid="eqn-4">Eqs. (4)</xref> and <xref ref-type="disp-formula" rid="eqn-11">(11)</xref> stably decrease only when <italic>&#x03B1;</italic><sub>1&#x2009;&#x2264;&#x2009;</sub>10<sup>&#x2212;6</sup> and <italic>&#x03B1;</italic><sub>2&#x2009;&#x2264;&#x2009;</sub>10<sup>&#x2212;6</sup>. Since larger learning rates lead to quicker learning, we choose <italic>&#x03B1;</italic><sub>1&#x2009;</sub>&#x003D;<sub>&#x2009;</sub>10<sup>&#x2212;6</sup> and <italic>&#x03B1;</italic><sub>2&#x2009;</sub>&#x003D;<sub>&#x2009;</sub>10<sup>&#x2212;6</sup>.</p>
<p>Next, we investigate 8 node characteristics that maybe related to convergence weight, as listed in <xref ref-type="table" rid="table-2">Tab. 2</xref>.</p>
<table-wrap id="table-2"><label>Table 2</label>
<caption>
<title>Definitions of node characteristics</title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Node characteristics</th>
<th align="left">Definition</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">Node&#x2019;s total bandwidth</td>
<td align="left">The total bandwidth of the data uploaded from the lower nodes and the data generated by the node itself</td>
</tr>
<tr>
<td align="left">Number of real-time applications</td>
<td align="left">The number of real-time applications (e.g., controlling signals of dispatching system) carried by the node.</td>
</tr>
<tr>
<td align="left">Node strength</td>
<td align="left">The number of other nodes connected to the node.</td>
</tr>
<tr>
<td align="left">Number of normal applications</td>
<td align="left">The number of non-real-time applications (e.g., consumer information of energy meters) carried by the node.</td>
</tr>
<tr>
<td align="left">Bandwidth of real-time applications</td>
<td align="left">The bandwidth requirements of the real-time applications on the node.</td>
</tr>
<tr>
<td align="left">Bandwidth of normal applications</td>
<td align="left">The bandwidth requirements of the non-real-time applications on the node.</td>
</tr>
<tr>
<td align="left">Node&#x2019;s voltage level</td>
<td align="left">The node&#x2019;s voltage level in power grid (e.g., a 220&#x2005;kV substation).</td>
</tr>
<tr>
<td align="left">Node capacity</td>
<td align="left">The node&#x2019;s transmission channel capacity to its upper nodes.</td>
</tr>
<tr>
<td align="left">Node distance</td>
<td align="left">The length of the shortest path between the node and the dispatching center</td>
</tr>
<tr>
<td align="left">Node centrality</td>
<td align="left">The average length of the shortest path between the node and all other nodes</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>We calculate correlation coefficient as <xref ref-type="disp-formula" rid="eqn-6">Eq. (6)</xref> to find that there are <italic>k&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3 node characteristics closely related to convergence weight, which are: node&#x2019;s total bandwidth&#x2009;&#x003E;&#x2009;number of real-time applications&#x2009;&#x003E;&#x2009;node strength.</p>
<p>We show 4 node characteristics with the highest correlation coefficient values in <xref ref-type="table" rid="table-3">Tab. 3</xref>, which are node&#x2019;s total bandwidth, number of real-time applications, node strength, and number of normal applications. It can be seen that while the first 3 characteristics have correlation coefficients larger than 0.6, the fourth characteristic (i.e., number of normal applications) drastically drops to 0.36. Such result indicates that number of normal applications (and the characteristics after it) has statistically ignorable impacts on convergence weight.</p>
<table-wrap id="table-3"><label>Table 3</label>
<caption>
<title>The first 4 node characteristics most related to convergence weight</title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/><col align="left"/><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left"><italic>&#x00A0;</italic></th>
<th align="center">Node bandwidth</th>
<th align="left">Number of real-time applications</th>
<th align="left">Node strength</th>
<th>Number of normal applications</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">Convergence weight</td>
<td align="left">0.78</td>
<td align="left">0.73</td>
<td align="left">0.63</td>
<td>0.36</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The above result is reasonable. Firstly, a node&#x2019;s total bandwidth basically reflects how important it is for smart grid (as complex applications tend to be deployed in critical grid sites), which means that a node with higher bandwidth requirement often has proportionally more data to upload. Secondly, most real-time applications need to communicate with dispatching center, so a node with more real-time applications implies that it has more data to upload. Thirdly, a node with higher node strength means that it is connected by many other nodes, so it tends to have more data to upload. Later we will show that our estimation based on these three characteristics indeed achieves rather high accuracy.</p>
<p>Afterwards, we determine the power of <xref ref-type="disp-formula" rid="eqn-7">Eq. (7)</xref>, <italic>n</italic>. According to <xref ref-type="disp-formula" rid="eqn-9">Eq. (9)</xref>, the value of <italic>n</italic> directly affects the fitting performance from node characteristics to convergence weight. To demonstrate this, we draw the fitting curves relating the 3 node characteristics (for conciseness, we only show each node&#x2019;s total bandwidth on the x-axis) to the 11 nodes&#x2019; convergence weights in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>. Observe that the fitting curve can barely match itself to all the points when <italic>n</italic> &#x2264; 2, which means that the relationship between the node characteristics and the convergence weight is too complicated for these values of <italic>n</italic> to capture. When <italic>n &#x2265;</italic> 3, the fitting curve is able to reach most of the points, and thereby the convergence weight is well related to the node characteristics.</p>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>The fitting curves relating node characteristics to convergence weight for different values of n. Note that Z2 is excluded here for it is used for validating (a) <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>1 (b) <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>2 (c) <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3 (d) <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>4</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_22812-fig-5.png"/>
</fig>
<p>Nevertheless, according to machine learning theory, <italic>n</italic> being too large will lead to overfitting, that is, our method can achieve high accuracy during learning, but its accuracy will drop if it is applied to nodes excluded by learning process (i.e., Z2 in this case). This statement is verified by <xref ref-type="table" rid="table-4">Tab. 4</xref>, where we compute the average bandwidth estimation accuracy across the network for different values of <italic>n</italic>. Observe that although the fitting accuracy keeps increasing as <italic>n</italic> grows, the validating accuracy for Z2 reaches the maximum at <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3, and decreases as <italic>n</italic> becomes larger. This clearly indicates that overfitting occurs for <italic>n&#x2009;</italic>&#x003E;<italic>&#x2009;</italic>3. Combing this result with <xref ref-type="fig" rid="fig-5">Fig. 5</xref>, we let <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3.</p>
<table-wrap id="table-4"><label>Table 4</label>
<caption>
<title>Average accuracy for different values of <italic>n</italic></title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/><col align="left"/><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left"><italic>n</italic></th>
<th align="left">1 (&#x0025;)</th>
<th align="left">2 (&#x0025;)</th>
<th align="left">3 (&#x0025;)</th>
<th align="left">4 (&#x0025;)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">Fitting accuracy</td>
<td align="left">78.5</td>
<td align="left">91.9</td>
<td align="left">93.7</td>
<td align="left">93.8</td>
</tr>
<tr>
<td align="left">Validating accuracy</td>
<td align="left">71.5</td>
<td align="left">89.3</td>
<td align="left">90.6</td>
<td align="left">89.6</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Finally, with <italic>k&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3 and <italic>n&#x2009;</italic>&#x003D;<italic>&#x2009;</italic>3, we can further derive the values of <italic>a</italic><sub>1</sub>, &#x2026;, <italic>a<sub>n</sub></italic>, <italic>b</italic><sub>(1,2)</sub>, &#x2026;, <italic>b</italic><sub>(1,2,&#x2026;,<italic>k</italic>)</sub>, <italic>&#x03B2;</italic><sub>1</sub>, &#x2026;, <italic>&#x03B2;</italic><sub><italic>k</italic>&#x002B;1</sub>. The deriving process and the final results are omitted here.</p>
</sec>
<sec id="s6_2">
<label>6.2</label>
<title>Estimating on New Network</title>
<p>Now we use the results in Section 6.1 to estimate the network bandwidth in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>. This network is based on a small city&#x2019;s power grid, which has 1 regional dispatching centers, 1 220&#x2005;kV substations, 5 110&#x2005;kV substation, and 1 35&#x2005;kV substation. This represents a &#x201C;new network&#x201D; where only applications&#x2019; bandwidth requirements are known. Note that the network is highly heterogeneous with 4 different types of nodes, which is difficult for bandwidth estimation.</p>
<p>We are mostly interested in the estimation accuracy of the dispatching center, because it is the most important node in smart grid&#x2019;s control system, and its estimation accuracy basically depends on the estimation accuracies of all the other nodes. <xref ref-type="table" rid="table-5">Tab. 5</xref> compares our method with 3 recent works in [<xref ref-type="bibr" rid="ref-2">2</xref>,<xref ref-type="bibr" rid="ref-4">4</xref>,<xref ref-type="bibr" rid="ref-6">6</xref>], where <italic>t</italic>, <italic>a</italic>, and <italic>b</italic> mean that network topology, applications, and node&#x2019;s total bandwidth are considered in estimation, respectively.</p>
<table-wrap id="table-5"><label>Table 5</label>
<caption>
<title>Estimated bandwidth of dispatching center</title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/><col align="left"/><col align="left"/><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Method</th>
<th align="left"><italic>b</italic></th>
<th align="left"><italic>t</italic></th>
<th align="left"><italic>a</italic></th>
<th align="left">Estimation (Mbps)</th>
<th align="left">Accuracy (&#x0025;)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">PCCM [<xref ref-type="bibr" rid="ref-2">2</xref>]</td>
<td align="left">&#x2013;</td>
<td align="left">&#x2013;</td>
<td align="left">&#x221A;</td>
<td align="left">1190.10</td>
<td align="left">43.1</td>
</tr>
<tr>
<td align="left">CA [<xref ref-type="bibr" rid="ref-4">4</xref>]</td>
<td align="left">&#x2013;</td>
<td align="left">&#x221A;</td>
<td align="left">&#x2013;</td>
<td align="left">662.27</td>
<td align="left">87.3</td>
</tr>
<tr>
<td align="left">CBT [<xref ref-type="bibr" rid="ref-6">6</xref>]</td>
<td align="left">&#x2013;</td>
<td align="left">&#x221A;</td>
<td align="left">&#x221A;</td>
<td align="left">660.81</td>
<td align="left">87.0</td>
</tr>
<tr>
<td align="left">Our method (<italic>t</italic>)</td>
<td align="left">&#x2013;</td>
<td align="left">&#x221A;</td>
<td align="left">&#x2013;</td>
<td align="left">807.53</td>
<td align="left">93.6</td>
</tr>
<tr>
<td align="left">Our method (<italic>t &#x002B; a</italic>)</td>
<td align="left">&#x2013;</td>
<td align="left">&#x221A;</td>
<td align="left">&#x221A;</td>
<td align="left">802.58</td>
<td align="left">94.2</td>
</tr>
<tr>
<td align="left">Our method (<italic>t &#x002B; a &#x002B; b</italic>)</td>
<td align="left">&#x221A;</td>
<td align="left">&#x221A;</td>
<td align="left">&#x221A;</td>
<td align="left">761.59</td>
<td align="left">99.6</td>
</tr>
<tr>
<td align="left">Actual bandwidth</td>
<td align="left">&#x2013;</td>
<td align="left">&#x2013;</td>
<td align="left">&#x2013;</td>
<td align="left">759.56</td>
<td align="left">&#x2013;</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Observe that our method (<italic>t &#x002B; a &#x002B; b</italic>) is just slightly higher than the actual bandwidth requirement of 759.56 Mbps, which outperforms the alternative methods (PCCM [<xref ref-type="bibr" rid="ref-2">2</xref>], CA [<xref ref-type="bibr" rid="ref-4">4</xref>], and CBT [<xref ref-type="bibr" rid="ref-6">6</xref>]) by 12.6&#x0025; to 56.5&#x0025; higher accuracy. The major reason is that our method accomplishes finer-grained estimation via considering network topology (node strength), applications (number of real-time applications), and node bandwidth as key node characteristics (see Section 6.1). In contrast, the alternative methods only consider one or two of these characteristics, thus they can hardly differentiate the convergence impacts of different lower-voltage nodes on the dispatching center.</p>
<p>In fact, <xref ref-type="table" rid="table-5">Tab. 5</xref> also shows that our method achieves higher accuracy as it takes more characteristics into consideration, i.e., <italic>t</italic> &#x003C; (<italic>t &#x002B; a</italic>) &#x003C; (<italic>t &#x002B; a &#x002B; b</italic>). Such result further proves that finer-grained estimation leads to higher accuracy.</p>
<p>We further investigate how the selection of node characteristics affects convergence weight learning, and eventually affects bandwidth estimation. This can be clearly illustrated by the ranking of learned convergence weights of different methods in <xref ref-type="table" rid="table-6">Tab. 6</xref>.</p>
<table-wrap id="table-6"><label>Table 6</label>
<caption>
<title>Ranking of learned convergence weights</title></caption>
<table frame="hsides"><colgroup><col align="left"/><col align="left"/><col align="left"/><col align="left"/><col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">PCCM</th>
<th align="left">CA</th>
<th align="left">CBT</th>
<th align="left">Our method<break/>(<italic>t &#x002B; a &#x002B; b</italic>)</th>
<th align="left">Actual</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S7</td>
<td align="left">S7</td>
<td align="left">S7</td>
<td align="left">S7</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S2</td>
<td align="left">S2</td>
<td align="left">S2</td>
<td align="left">S2</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S6</td>
<td align="left">S6</td>
<td align="left">S6</td>
<td align="left">S6</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S4</td>
<td align="left">S0</td>
<td align="left">S4</td>
<td align="left">S4</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S0</td>
<td align="left">S1</td>
<td align="left">S0</td>
<td align="left">S0</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S1</td>
<td align="left">S4</td>
<td align="left">S1</td>
<td align="left">S1</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S8</td>
<td align="left">S5</td>
<td align="left">S3</td>
<td align="left">S3</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S5</td>
<td align="left">S3</td>
<td align="left">S5</td>
<td align="left">S5</td>
</tr>
<tr>
<td align="left">&#x2013;</td>
<td align="left">S3</td>
<td align="left">S8</td>
<td align="left">S8</td>
<td align="left">S8</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>By considering node strength (<italic>t</italic>), number of real-time applications (<italic>a</italic>), and node bandwidth (<italic>b</italic>), our method derives the same ranking as the actual network does. As the matter of fact, for the dispatching center (S7), the ranking of the convergence weights of all nodes should be: itself (S7), the nodes linked by lower-voltage nodes (S2, S6, and S4), the nodes without lower-voltage nodes (S0, S1, S3), and the lowest-voltage nodes (S5, S8). Our method&#x2019;s fine-grained recognition of different nodes is the key to accurate bandwidth estimation.</p>
<p>On the other hand, because PCCM ignores network topology and node bandwidth, it cannot fully recognize the differences among various nodes to learn convergence weights in a heterogeneous network like <xref ref-type="fig" rid="fig-6">Fig. 6</xref>. This explains why PCCM has rather low estimation accuracy in <xref ref-type="table" rid="table-5">Tab. 5</xref>.</p>
<p>Both CA and CBT have considered topology for estimation, so they can roughly infer how the nodes converges the dispatching center, and hence make partially correct rankings in <xref ref-type="table" rid="table-6">Tab. 6</xref>. However, note that both CA and CBT incorrectly rank S3 after S5 due to their neglection of node bandwidth. This explains their relatively low estimation accuracy in <xref ref-type="table" rid="table-5">Tab. 5</xref>.</p>
</sec>
</sec>
<sec id="s7">
<label>7</label>
<title>Conclusion</title>
<p>In this paper, we propose a novel fine-grained bandwidth estimation method for smart grid communication network. The method achieves fine-grained estimations through explicitly considering how bandwidth requirements of different nodes converges to upper nodes, and it exploits multivariate nonlinear learning to derive multiple convergence parameters from present network without needing human experience. Due to these two novelties, our method outperforms existing methods by up to 56.5&#x0025; higher estimation accuracy. Through the comparison of different characteristics, we find that the fitting accuracy of the three characteristics selected in this paper is higher, which can reach 99.6&#x0025;. In future, we will collect transmission data from other industrial Internet to train this model. So that this method can be applied to other Industrial Internets, such as the communication networks of railway or oil pipeline. Furthermore, we will study how to directly estimate bandwidth requirement and predict long-term development based on current network information using deep learning and big data technologies.</p>
</sec>
</body>
<back><fn-group>
<fn fn-type="other">
<p><bold>Funding Statement:</bold> This work was supported by Natural Science Foundation of China (Grant No.62071098); Sichuan Application and Basic Research Funds (Grant No. 2021YJ0313); Sichuan Science and Technology Program (Grant No. 2021YFG0307).</p>
</fn>
<fn fn-type="conflict">
<p><bold>Conflicts of Interest:</bold> The authors declare that they have no conflicts of interest to report regarding the present study.</p>
</fn>
</fn-group>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Zhao</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Fan</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Jin</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Information flow model and service flow calculation of electric power backbone communication network</article-title>,&#x201D; <source>Telecommunications Science</source>, vol. <volume>33</volume>, no. <issue>5</issue>, pp. <fpage>153</fpage>&#x2013;<lpage>163</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Peng</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>C.</given-names> <surname>Xu</surname></string-name></person-group>, &#x201C;<article-title>A study on the change law and trend of the elasticity coefficient of jiangsu electric power</article-title>,&#x201D; <source>Statistical Science and Practice</source>, vol. <volume>10</volume>, pp. <fpage>13</fpage>&#x2013;<lpage>16</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Mankad</surname></string-name> and <string-name><given-names>G.</given-names> <surname>Michailidis</surname></string-name></person-group>, &#x201C;<article-title>Discovery of path-important nodes using structured semi-nonnegative matrix factorization</article-title>,&#x201D; in <conf-name>IEEE Int. Workshop on Computational Advances in Multi-Sensor Adaptive Processing (CAMSAP)</conf-name>, Saint Martin, France, pp. <fpage>288</fpage>&#x2013;<lpage>291</lpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Ji</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Meng</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Bao</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Qi</surname></string-name></person-group>, &#x201C;<article-title>A node importance evaluation method for electric power communication networks with weighted bandwidth</article-title>,&#x201D; <source>Power Information and Communication Technology</source>, vol. <volume>12</volume>, no. <issue>5</issue>, pp. <fpage>25</fpage>&#x2013;<lpage>29</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Zeng</surname></string-name></person-group>. &#x201C;<article-title>Evaluation of node importance and invulnerability simulation analysis in complex load-network</article-title>,&#x201D; <source>Neurocomputing</source>, vol. <volume>416</volume>, no. <issue>1</issue>, pp. <fpage>158</fpage>&#x2013;<lpage>164</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>B.</given-names> <surname>Fan</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Zheng</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Tang</surname></string-name></person-group>, &#x201C;<article-title>Risk assessment of power communication network based on node importance</article-title>,&#x201D; in <conf-name>Conf. IEEE Advanced Information Management, Communicates, Electronic and Automation Control (IMCEC)</conf-name>, <conf-loc>Chongqing, China</conf-loc>, pp. <fpage>818</fpage>&#x2013;<lpage>821</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>Q.</given-names> <surname>Xiong</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Shi</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Shi</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Wang</surname></string-name></person-group>, &#x201C;<article-title>Evaluating the importance of nodes in complex networks</article-title>,&#x201D; <source>Physica a: Statistical Mechanics and its Applications</source>, vol. <volume>452</volume>, pp. <fpage>209</fpage>&#x2013;<lpage>219</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Gao</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Sun</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Zhou</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Chen</surname></string-name></person-group>, &#x201C;<article-title>Study on the prediction and analysis method of electric power communication service</article-title>,&#x201D; <source>Electric Power Information and Communication Technology</source>, vol. <volume>14</volume>, no. <issue>7</issue>, pp. <fpage>108</fpage>&#x2013;<lpage>112</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Zhu</surname></string-name>, <string-name><given-names>X.</given-names> <surname>You</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Zheng</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Fan</surname></string-name></person-group>, &#x201C;<article-title>Power communication website bandwidth estimation method based on position weighting</article-title>,&#x201D; <source>Telecommunication Science</source>, vol. <volume>33</volume>, no. <issue>10</issue>, pp. <fpage>150</fpage>&#x2013;<lpage>156</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>F.</given-names> <surname>Phillipson</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Worm</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Neumann</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Sangers</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Wiarda</surname></string-name></person-group>, &#x201C;<article-title>Estimating bandwidth coverage using geometric models</article-title>,&#x201D; in <conf-name>European Conf. Networks and Optical Communications (NOC), Lisbon</conf-name>, <conf-loc>Portugal</conf-loc>, pp. <fpage>152</fpage>&#x2013;<lpage>156</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Takano</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Mutoh</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Oguchi</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Abe</surname></string-name></person-group>, &#x201C;<article-title>Estimating available bandwidth in mobile networks by correlation coefficient</article-title>,&#x201D; in <conf-name>Conf. Asia-Pacific Network Operations and Management Symposium (APNOMS)</conf-name>, <conf-loc>Kanazawa, Japan</conf-loc>, pp. <fpage>1</fpage>&#x2013;<lpage>4</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Q.</given-names> <surname>Huang</surname></string-name></person-group>, &#x201C;<article-title>A novel important node discovery algorithm based on local community aggregation and recognition in complex networks</article-title>,&#x201D; <source>International Journal of Wireless Information Networks</source>, vol. <volume>27</volume>, pp. <fpage>253</fpage>&#x2013;<lpage>260</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>F.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Xiao</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Xue</surname></string-name></person-group>, &#x201C;<article-title>Complex network node centrality measurement based on multiple attributes</article-title>,&#x201D; in <conf-name>Int. Conf. on Modelling, Identification and Control (ICMIC)</conf-name>, <conf-loc>Guiyang, China</conf-loc>, pp. <fpage>1</fpage>&#x2013;<lpage>5</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Liang</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Li</surname></string-name></person-group>, &#x201C;<article-title>Identifying and ranking influential spreaders in complex networks</article-title>,&#x201D; in <conf-name>Int. Computer Conf. on Wavelet Active Media Technology and Information Processing (ICCWAMTIP)</conf-name>, <conf-loc>Chengdu, China</conf-loc>, pp. <fpage>393</fpage>&#x2013;<lpage>396</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Hu</surname></string-name> and <string-name><given-names>T.</given-names> <surname>Mei</surname></string-name></person-group>, &#x201C;<article-title>Ranking influential nodes in complex networks with structural holes</article-title>,&#x201D; <source>Physica a: Statistical Mechanics and Its Applications</source>, vol. <volume>490</volume>, pp. <fpage>624</fpage>&#x2013;<lpage>631</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>T.</given-names> <surname>Agryzkov</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Oliver</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Tortosa</surname></string-name></person-group>, &#x201C;<article-title>A new betweenness centrality measure based on an algorithm for ranking the nodes of a network</article-title>,&#x201D; <source>Applied Mathematics and Computation</source>, vol. <volume>244</volume>, pp. <fpage>467</fpage>&#x2013;<lpage>478</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>N. S.</given-names> <surname>Kagami</surname></string-name>, <string-name><given-names>R. I. T.</given-names> <surname>da Costa Filho</surname></string-name> and <string-name><given-names>L. P.</given-names> <surname>Gaspary</surname></string-name></person-group>, &#x201C;<article-title>CAPEST: Offloading network capacity and available bandwidth estimation to programmable data planes</article-title>,&#x201D; <source>IEEE Transactions on Network and Service Management</source>, vol. <volume>17</volume>, no. <issue>1</issue>, pp. <fpage>175</fpage>&#x2013;<lpage>189</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>X.</given-names> <surname>Yan</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Ozturk</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Hu</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Song</surname></string-name></person-group>, &#x201C;<article-title>A review on price-driven residential demand response</article-title>,&#x201D; <source>Renewable and Sustainable Energy Reviews</source>, vol. <volume>96</volume>, pp. <fpage>411</fpage>&#x2013;<lpage>419</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Ourahou</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Ayrir</surname></string-name>, <string-name><given-names>B. E.</given-names> <surname>Hassouni</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Haddi</surname></string-name></person-group>, &#x201C;<article-title>Review on smart grid control and reliability in presence of renewable energies: Challenges and prospects</article-title>,&#x201D; <source>Mathematics and Computers in Simulation</source>, vol. <volume>167</volume>, pp. <fpage>19</fpage>&#x2013;<lpage>31</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. Z.</given-names> <surname>Gunduz</surname></string-name> and <string-name><given-names>R.</given-names> <surname>Das</surname></string-name></person-group>, &#x201C;<article-title>Cyber-security on smart grid: Threats and potential solutions</article-title>,&#x201D; <source>Computer Networks</source>, vol. <volume>169</volume>, pp. 107094, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. I.</given-names> <surname>Jordan</surname></string-name> and <string-name><given-names>T. M.</given-names> <surname>Mitchell</surname></string-name></person-group>, &#x201C;<article-title>Machine learning: Trends, perspectives, and prospects</article-title>,&#x201D; <source>Science</source>, vol. <volume>349</volume>, no. <issue>6245</issue>, pp. <fpage>255</fpage>&#x2013;<lpage>260</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>H. Y.</given-names> <surname>Wang</surname></string-name> and <string-name><given-names>J. H.</given-names> <surname>Guo</surname></string-name></person-group>, &#x201C;<article-title>Multi-model ensemble with rich spatial information for object detection</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>99</volume>, pp. 107098, 2020.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Lin</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Sun</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Yang</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Hybrid malware detection approach with feedback-directed machine learning</article-title>,&#x201D; <source>Science China Information Sciences</source>, vol. <volume>63</volume>, no. <issue>3</issue>, pp. <fpage>139103</fpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Sperandei</surname></string-name></person-group>, &#x201C;<article-title>Understanding logistic regression analysis</article-title>,&#x201D; <source>Biochemia Medica</source>, vol. <volume>24</volume>, no. <issue>1</issue>, pp. <fpage>12</fpage>&#x2013;<lpage>18</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Cervantes</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Garcia-Lamont</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Rodr&#x00ED;guez-Mazahua</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Lopez</surname></string-name></person-group>, &#x201C;<article-title>A comprehensive survey on support vector machine classification: Applications, challenges and trends</article-title>,&#x201D; <source>Neurocomputing</source>, vol. <volume>408</volume>, pp. <fpage>189</fpage>&#x2013;<lpage>215</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Song</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Ying</surname></string-name></person-group>, &#x201C;<article-title>Decision tree methods: Applications for classification and prediction</article-title>,&#x201D; <source>Shanghai Archives of Psychiatry</source>, vol. <volume>27</volume>, no. <issue>2</issue>, pp. <fpage>130</fpage>&#x2013;<lpage>135</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>D. V.</given-names> <surname>Carvalho</surname></string-name>, <string-name><given-names>E. M.</given-names> <surname>Pereira</surname></string-name> and <string-name><given-names>J. S.</given-names> <surname>Cardoso</surname></string-name></person-group>, &#x201C;<article-title>Machine learning interpretability: A survey on methods and metrics</article-title>,&#x201D; <source>Electronics</source>, vol. <volume>8</volume>, no. <issue>8</issue>, pp. <fpage>832</fpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>K.</given-names> <surname>Ma</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Hu</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Yang</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Resource allocation for smart grid communication based on a multi-swarm artificial bee colony algorithm with cooperative learning</article-title>,&#x201D; <source>Engineering Applications of Artificial Intelligence</source>, vol. <volume>81</volume>, pp. <fpage>29</fpage>&#x2013;<lpage>36</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Luo</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Cao</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Ma</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Zhang</surname></string-name> and <string-name><given-names>R.</given-names> <surname>He</surname></string-name></person-group>, &#x201C;<article-title>FA-Gan: Face augmentation GAN for deformation-invariant face recognition</article-title>,&#x201D; <source>IEEE Transactions on Information Forensics and Security</source>, vol. <volume>16</volume>, pp. <fpage>2341</fpage>&#x2013;<lpage>2355</lpage>, <year>2021</year>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Song</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Yang</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Yuan</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Bufford</surname></string-name></person-group>, &#x201C;<article-title>Deep 3d-multiscale densenet for hyperspectral image classification based on spatial-spectral information</article-title>,&#x201D; <source>Intelligent Automation &#x0026; Soft Computing</source>, vol. <volume>26</volume>, no. <issue>6</issue>, pp. <fpage>1441</fpage>&#x2013;<lpage>1458</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Song</surname></string-name>, <string-name><given-names>H. L.</given-names> <surname>Wei</surname></string-name>, <string-name><given-names>J. H.</given-names> <surname>Guo</surname></string-name>, <string-name><given-names>Y. F.</given-names> <surname>Zhou</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>A fast human action recognition network based on spatio-temporal features</article-title>,&#x201D; <source>Neurocomputing</source>, vol. <volume>441</volume>, pp. <fpage>350</fpage>&#x2013;<lpage>358</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Wei</surname></string-name>, <string-name><given-names>T.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Hou</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Online multi-expert learning for visual tracking</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>29</volume>, pp. <fpage>934</fpage>&#x2013;<lpage>946</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S. H.</given-names> <surname>Park</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Han</surname></string-name></person-group>, &#x201C;<article-title>Methodologic guide for evaluating clinical performance and effect of artificial intelligence technology for medical diagnosis and prediction</article-title>,&#x201D; <source>Radiology</source>, vol. <volume>286</volume>, no. <issue>3</issue>, pp. <fpage>800</fpage>&#x2013;<lpage>809</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Vaswani</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Shazeer</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Parmar</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Uszkoreit</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Jones</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Attention is all you need</article-title>,&#x201D; in <conf-name>Proc. of Int. Conf. on Neural Information Processing Systems</conf-name>, <conf-loc>NY, USA</conf-loc>, pp. <fpage>6000</fpage>&#x2013;<lpage>6010</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-35"><label>[35]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>X.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Wan</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Wang</surname></string-name></person-group>, &#x201C;<article-title>A study of unmanned path planning based on a double-twin rbm-bp deep neural network</article-title>,&#x201D; <source>Intelligent Automation &#x0026; Soft Computing</source>, vol. <volume>26</volume>, no. <issue>6</issue>, pp. <fpage>1531</fpage>&#x2013;<lpage>1548</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-36"><label>[36]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>T. B.</given-names> <surname>Brown</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Mann</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Ryder</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Subbiah</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Kaplan</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Language models are few-shot learner</article-title>,&#x201D; in <conf-name>Proc. of Int. Conf. on Neural Information Processing Systems</conf-name>, Virtual-only Conference, pp. <fpage>1877</fpage>&#x2013;<lpage>1901</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-37"><label>[37]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Niu</surname></string-name> and <string-name><given-names>Z.</given-names> <surname>Gao</surname></string-name></person-group>, &#x201C;<article-title>A novel scene text recognition method based on deep learning</article-title>,&#x201D; <source>Computers, Materials &#x0026; Continua</source>, vol. <volume>60</volume>, no. <issue>2</issue>, pp. <fpage>781</fpage>&#x2013;<lpage>794</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-38"><label>[38]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Wan</surname></string-name> and <string-name><given-names>W.</given-names> <surname>Chen</surname></string-name></person-group>, &#x201C;<article-title>Mii: A novel text classification model combining deep active learning with bert</article-title>,&#x201D; <source>Computers, Materials &#x0026; Continua</source>, vol. <volume>63</volume>, no. <issue>3</issue>, pp. <fpage>1499</fpage>&#x2013;<lpage>1514</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-39"><label>[39]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Huh</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Otgonchimeg</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Seo</surname></string-name></person-group>, &#x201C;<article-title>Advanced metering infrastructure design and test bed experiment using intelligent agents: Focusing on the PLC network base technology for smart grid system</article-title>,&#x201D; <source>Supercomput</source>, vol. <volume>72</volume>, no. <issue>5</issue>, pp. <fpage>1862</fpage>&#x2013;<lpage>1877</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-40"><label>[40]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>I.</given-names> <surname>Goodfellow</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Bengio</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Courville</surname></string-name></person-group>, &#x201C;<chapter-title>in Deep learning</chapter-title>,&#x201D; Cambridge, MA, USA, MIT Press, <year>2016</year>.</mixed-citation></ref>
</ref-list>
</back>
</article>