<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="1.0">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMC</journal-id>
<journal-id journal-id-type="nlm-ta">CMC</journal-id>
<journal-id journal-id-type="publisher-id">CMC</journal-id>
<journal-title-group>
<journal-title>Computers, Materials &#x0026; Continua</journal-title>
</journal-title-group>
<issn pub-type="epub">1546-2226</issn>
<issn pub-type="ppub">1546-2218</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">15504</article-id>
<article-id pub-id-type="doi">10.32604/cmc.2021.015504</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>An Optimized Approach to Vehicle-Type Classification Using a Convolutional Neural Network</article-title>
<alt-title alt-title-type="left-running-head">An Optimized Approach to Vehicle-Type Classification Using a Convolutional Neural Network</alt-title>
<alt-title alt-title-type="right-running-head">An Optimized Approach to Vehicle-Type Classification Using a Convolutional Neural Network</alt-title>
</title-group>
<contrib-group content-type="authors">
<contrib id="author-1" contrib-type="author">
<name name-style="western">
<surname>Habib</surname>
<given-names>Shabana</given-names>
</name>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western">
<surname>Khan</surname>
<given-names>Noreen Fayyaz</given-names>
</name>
<xref ref-type="aff" rid="aff-2">2</xref><email>noreen.fayyaz@fu.edu.pk</email></contrib>
<aff id="aff-1"><label>1</label><institution>Department of Information Technology, College of Computer, Qassim University</institution>, <addr-line>Buraidah, 51452</addr-line>, <country>Saudi Arabia</country></aff>
<aff id="aff-2"><label>2</label><institution>Department of Computer Science, Islamia College University</institution>, <addr-line>Peshawar</addr-line>, <country>Pakistan</country></aff>
</contrib-group>
<author-notes><corresp id="cor1">&#x002A;Corresponding Author: Noreen Fayyaz Khan. Email: <email>noreen.fayyaz@fu.edu.pk</email></corresp></author-notes>
<pub-date pub-type="epub" date-type="pub" iso-8601-date="2021-08-23">
<day>23</day>
<month>8</month>
<year>2021</year>
</pub-date>
<volume>69</volume>
<issue>3</issue>
<fpage>3321</fpage>
<lpage>3335</lpage>
<history>
<date date-type="received">
<day>25</day>
<month>11</month>
<year>2020</year>
</date>
<date date-type="accepted">
<day>2</day>
<month>3</month>
<year>2021</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2021 Habib and Khan</copyright-statement>
<copyright-year>2021</copyright-year>
<copyright-holder>Habib and Khan</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMC_15504.pdf"></self-uri>
<abstract>
<p>Vehicle type classification is considered a central part of an intelligent traffic system. In recent years, deep learning had a vital role in object detection in many computer vision tasks. To learn high-level deep features and semantics, deep learning offers powerful tools to address problems in traditional architectures of handcrafted feature-extraction techniques. Unlike other algorithms using handcrated visual features, convolutional neural network is able to automatically learn good features of vehicle type classification. This study develops an optimized automatic surveillance and auditing system to detect and classify vehicles of different categories. Transfer learning is used to quickly learn the features by recording a small number of training images from vehicle frontal view images. The proposed system employs extensive data-augmentation techniques for effective training while avoiding the problem of data shortage. In order to capture rich and discriminative information of vehicles, the convolutional neural network is fine-tuned for the classification of vehicle types using the augmented data. The network extracts the feature maps from the entire dataset and generates a label for each object (vehicle) in an image, which can help in vehicle-type detection and classification. Experimental results on a public dataset and our own dataset demonstrated that the proposed method is quite effective in detection and classification of different types of vehicles. The experimental results show that the proposed model achieves 96.04% accuracy on vehicle type classification.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Vehicle classification</kwd>
<kwd>convolutional neural network</kwd>
<kwd>deep learning</kwd>
<kwd>surveillance</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>Surveillance systems have achieved good results in terms of security. Image analysis, such as detecting a moving vehicle in an image, is a challenging task that can be solved by analyzing the foreground [<xref ref-type="bibr" rid="ref-1">1</xref>]. Dramatic improvements have been observed in the areas of speech recognition and document recognition genomics for automation technologies [<xref ref-type="bibr" rid="ref-2">2</xref>]. Major issues in surveillance systems include brightness, lighting, occlusion of shadows, and fragmentation, and all have a negative impact on objects to be detected [<xref ref-type="bibr" rid="ref-3">3</xref>,<xref ref-type="bibr" rid="ref-4">4</xref>]. Much research has been done on vehicle-type recognition systems that include handcrafted feature-extraction techniques such as speeded-up robust features (Surf), local binary patterns (LBPs), histogram of gradients (HOG) [<xref ref-type="bibr" rid="ref-5">5</xref>], annular coil, radar detection radio wave or infrared contour scanning, vehicle weight, and laser sensor measurement [<xref ref-type="bibr" rid="ref-6">6</xref>,<xref ref-type="bibr" rid="ref-7">7</xref>]. Automatic detection and classification of vehicle types using a convolutional neural network (CNN) is an unsolved problem. Deep CNNs [<xref ref-type="bibr" rid="ref-8">8</xref>,<xref ref-type="bibr" rid="ref-9">9</xref>], as well as extensively annotated datasets (e.g., ImageNet [<xref ref-type="bibr" rid="ref-10">10</xref>]), have brought remarkable progress in image recognition. Deep learning approaches are useful for feature extraction, and selection without prior knowledge has been investigated [<xref ref-type="bibr" rid="ref-11">11</xref>]. The most popular form of deep learning model, the CNN, consists of a series of convolutional layers followed by pooling layers and fully connected layers. The convolutional layer forms feature maps within which each unit is associated with a set of weights called filters. The pooling layer computes the sampling feature maps by summarizing the presence of feature maps. The fully connected layers are used for classification. All the weights are updated by gradient-based learning.</p>
<p>In this research, we employed the AlexNet CNN architecture and customized its network layers and options according to our classification objectives. Image features are the primary elements for any object detection. The model extracts the features from the training dataset through convolutional and pooling layers. The proposed model works on backpropagation gradients that help to reduce the discrepancy between the correct output and that produced by the system. The CNN helps learns the semantics of the categories of vehicle images so as to produce accurate detection and classification results.</p>
<p>The key contributions of this work follow.</p>
<list list-type="bullet">
<list-item><p>Vehicles are automatically detected and classified irrespective of brightness, lighting, occlusion of shadows, and fragmentation.</p></list-item>
<list-item><p>The AlexNet model is used to classify vehicle types by customizing layers according to the problem domain.</p></list-item>
<list-item><p>Deep learning helps to automatically learn features through filters in the convolutional layer.</p></list-item>
<list-item><p>The system helps in vehicle-tracking systems in commercial parking areas and assists in counting the number of vehicles on a road.</p></list-item>
</list>
</sec>
<sec id="s2">
<label>2</label>
<title>Related Work</title>
<p>The classification of vehicles is a challenging problem in the field of vision-based surveillance. There is huge within-class variability in vision-based vehicle detection systems, as vehicles may differ in color, size, and shape, illumination can vary, and background can be cluttered. Furthermore, the appearance of vehicles depends on their posture and might be affected by neighboring objects [<xref ref-type="bibr" rid="ref-12">12</xref>]. Shallow classification models are used by traditional image classification systems, such as support vector machine (SVM) [<xref ref-type="bibr" rid="ref-13">13</xref>], Bayesian [<xref ref-type="bibr" rid="ref-14">14</xref>], random forest (RF) [<xref ref-type="bibr" rid="ref-15">15</xref>], and boosting [<xref ref-type="bibr" rid="ref-16">16</xref>], to extract features for classification, such as local binary patterns (LBP), histogram of oriented gradients (HOG) [<xref ref-type="bibr" rid="ref-17">17</xref>], and scale-invariant feature transform (SIFT) [<xref ref-type="bibr" rid="ref-18">18</xref>]. These methods rely on hand-designed features. Shallow models are trained by original training data with limited fitness in representation learning [<xref ref-type="bibr" rid="ref-8">8</xref>]. Fu et al. [<xref ref-type="bibr" rid="ref-19">19</xref>] proposed a hierarchical multi-SVM method for vehicle classification. Other methods include a real-time system for multiple vehicle detection and classification using a Gaussian mixture model with hole filling algorithm (GMMHF), Gabor kernel for feature extraction, and multi-class vehicle classification [<xref ref-type="bibr" rid="ref-20">20</xref>]. Use of a CNN to classify images marks a massive revolution. Some deep learning techniques surpass humans on tasks like face recognition and image classification [<xref ref-type="bibr" rid="ref-21">21</xref>&#x2013;<xref ref-type="bibr" rid="ref-23">23</xref>]. Machine learning techniques have seen successful application, and are considered the best choice compared to neural networks and support vector machine for vehicle detection and classification [<xref ref-type="bibr" rid="ref-24">24</xref>]. Roadside LiDAR sensors [<xref ref-type="bibr" rid="ref-25">25</xref>], frequency-modulated continuous-wave (FMCW) radar signals [<xref ref-type="bibr" rid="ref-26">26</xref>], sensors for vehicle classification and counting [<xref ref-type="bibr" rid="ref-27">27</xref>], and distributed optical sensing technology in vibration-based vehicle classification systems [<xref ref-type="bibr" rid="ref-28">28</xref>] also play a significant role in vehicle classification.</p>
</sec>
<sec id="s3">
<label>3</label>
<title>Proposed System</title>
<p>Increased traffic has become an issue in many towns and cities, causing serious traffic congestion problems. This paper develops a simple and efficient vehicle-type recognition system using a CNN model. The framework, as shown in <xref ref-type="fig" rid="fig-1">Fig. 1</xref>, includes three steps: a) preparing the dataset; b) feature extraction by CNN; and c) classification of test data.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Proposed framework for vehicle-type detection and classification</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-1.png"/>
</fig>
<sec id="s3_1">
<label>3.1</label>
<title>Preparing the Dataset</title>
<p>For effective learning patterns, a huge amount of data is required for a deep learning-based approach. The effective deployment of deep learning models requires abundant high-quality data [<xref ref-type="bibr" rid="ref-13">13</xref>]. To attain the desired accuracy, we apply four data-augmentation techniques to extend the dataset. <xref ref-type="table" rid="table-1">Tab. 1</xref> shows the extended dataset. Data augmentation includes flipping, skewness, rotation, and translations for geometric transformation invariance. The second column in <xref ref-type="table" rid="table-1">Tab. 1</xref> lists techniques, and the third column shows their invariance parameters.</p>
<table-wrap id="table-1">
<label>Table 1</label>
<caption>
<title>Different techniques of data augmentation with respective parameters</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>No.</th>
<th>Techniques of data augmentation</th>
<th>Parameters</th>
<th></th>
</tr>
</thead>
<tbody>
<tr>
<td>1.</td>
<td>Flip</td>
<td>TopLeftBottomRight</td>
<td/>
</tr>
<tr>
<td>2.</td>
<td>Shear</td>
<td>Along X-axis at 10<inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula>Along Y-axis at 10<inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula></td>
<td/>
</tr>
<tr>
<td>3.</td>
<td>Rotation</td>
<td>90<inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula>60<inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula>45<inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> &#x2212;45<inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> &#x2212;60<inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> &#x2212;90<inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mo>&#x2218;</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula></td>
<td/>
</tr>
<tr>
<td>4.</td>
<td>Skewness</td>
<td>LeftRightForwardBackward</td>
<td/>
</tr>
</tbody>
</table>
</table-wrap>
<p>The four augmentation techniques and 16 parameters extend each sample to form 16 samples. <xref ref-type="table" rid="table-2">Tab. 2</xref> shows eight classes of vehicles: bike, bus, car, horse buggy, jeep, rickshaw, truck, and van. The third column presents the number of images of each vehicle type before and after augmentation. The purpose of augmentation is the effective deployment of the deep learning model. It also helps to avoid overfitting and memorizes the targeted details of the training images. All of the images in the dataset are preprocessed to the size of 227 (width) <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula> 227 (height) <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula> 3 (color channels) to prepare for training.</p>
<table-wrap id="table-2">
<label>Table 2</label>
<caption>
<title>Statistics of vehicle type data set before augmentation and after augmentation</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>No.</th>
<th>Vehicle type</th>
<th colspan="2">Number of images</th>
</tr>
<tr>
<th></th>
<th></th>
<th>Before augmentation</th>
<th>After augmentation</th>
</tr>
</thead>
<tbody>
<tr>
<td>1</td>
<td>Bike</td>
<td>40</td>
<td>640</td>
</tr>
<tr>
<td>2</td>
<td>Bus</td>
<td>37</td>
<td>592</td>
</tr>
<tr>
<td>3</td>
<td>Car</td>
<td>37</td>
<td>592</td>
</tr>
<tr>
<td>4</td>
<td>Horse buggy</td>
<td>47</td>
<td>752</td>
</tr>
<tr>
<td>5</td>
<td>Jeep</td>
<td>35</td>
<td>560</td>
</tr>
<tr>
<td>6</td>
<td>Riksha</td>
<td>50</td>
<td>800</td>
</tr>
<tr>
<td>7</td>
<td>Truck</td>
<td>45</td>
<td>720</td>
</tr>
<tr>
<td>8</td>
<td>Van</td>
<td>51</td>
<td>816</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Feature Extraction by CNN</title>
<p>In the proposed system, the AlexNet CNN architecture is fine-tuned [<xref ref-type="bibr" rid="ref-9">9</xref>], as shown in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>. AlexNet has eight layers. Five are convolutional (conv) layers, where conv1, conv2, conv3, and conv4 are followed by the max-pooling layer, and the last three are fully connected layers. Dataset features are extracted by applying a deep CNN model containing manifold convolutional layers. A feature map (fmap) that represents a higher-level abstraction of the input data is generated by each convolutional layer. The fmaps of the starting convolutional layers extract low-level features such as color, shape, corners, and edges, and the fmap of the last convolutional layer includes the high-level features, which are forwarded to the fully connected layers for classification. <xref ref-type="table" rid="table-3">Tab. 3</xref> provides an architectural analysis of each layer of the fine-tuned AlexNet model. The first convolutional layer is produced by applying a filter of size <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:mn>11</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>11</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>3</mml:mn></mml:math></inline-formula> on an image of size <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:mn>227</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>227</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>3</mml:mn></mml:math></inline-formula>. Images are convolved with their respective filters. The convolutional layers detect the same features at different locations in an image. Layers learn all the edges and blobs of the images in the dataset by learning the 34,944 parameters.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Architecture of fine-tuned Alex net CNN mode</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-2.png"/>
</fig>
 
<table-wrap id="table-3">
<label>Table 3</label>
<caption>
<title>LayerWise analysis of CNN architecture of Alex net model</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Layer name</th>
<th>Size of layer</th>
<th>Weights</th>
<th>Biases</th>
<th>Parameters</th>
</tr>
</thead>
<tbody>
<tr>
<td>Input image</td>
<td><inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:mn>227</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>227</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>3</mml:mn></mml:math></inline-formula></td>
<td>0</td>
<td>0</td>
<td>0</td>
</tr>
<tr>
<td>Conv-1</td>
<td><inline-formula id="ieqn-14"><mml:math id="mml-ieqn-14"><mml:mn>55</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>55</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>96</mml:mn></mml:math></inline-formula></td>
<td>34,848</td>
<td>96</td>
<td>34,944</td>
</tr>
<tr>
<td>MaxPool-1</td>
<td><inline-formula id="ieqn-15"><mml:math id="mml-ieqn-15"><mml:mn>27</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>27</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>96</mml:mn></mml:math></inline-formula></td>
<td>0</td>
<td>0</td>
<td>0</td>
</tr>
<tr>
<td>Conv-2</td>
<td><inline-formula id="ieqn-16"><mml:math id="mml-ieqn-16"><mml:mn>27</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>27</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula></td>
<td>614,400</td>
<td>256</td>
<td>614,656</td>
</tr>
<tr>
<td>MaxPool-2</td>
<td><inline-formula id="ieqn-17"><mml:math id="mml-ieqn-17"><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula></td>
<td>0</td>
<td>0</td>
<td>0</td>
</tr>
<tr>
<td>Conv-3</td>
<td><inline-formula id="ieqn-18"><mml:math id="mml-ieqn-18"><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>384</mml:mn></mml:math></inline-formula></td>
<td>884,736</td>
<td>384</td>
<td>885,120</td>
</tr>
<tr>
<td>Conv-4</td>
<td><inline-formula id="ieqn-19"><mml:math id="mml-ieqn-19"><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>384</mml:mn></mml:math></inline-formula></td>
<td>1,327,104</td>
<td>384</td>
<td>1,327,488</td>
</tr>
<tr>
<td>Conv-5</td>
<td><inline-formula id="ieqn-20"><mml:math id="mml-ieqn-20"><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>13</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula></td>
<td>884,736</td>
<td>256</td>
<td>884,992</td>
</tr>
<tr>
<td>MaxPool-3</td>
<td><inline-formula id="ieqn-21"><mml:math id="mml-ieqn-21"><mml:mn>6</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>6</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula></td>
<td>0</td>
<td>0</td>
<td>0</td>
</tr>
<tr>
<td>FC-1</td>
<td><inline-formula id="ieqn-22"><mml:math id="mml-ieqn-22"><mml:mn>4096</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula></td>
<td>37,748,736</td>
<td>4096</td>
<td>37,752,832</td>
</tr>
<tr>
<td>FC-2</td>
<td><inline-formula id="ieqn-23"><mml:math id="mml-ieqn-23"><mml:mn>4096</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula></td>
<td>16,777,216</td>
<td>4096</td>
<td>16,781,312</td>
</tr>
<tr>
<td>FC-3</td>
<td><inline-formula id="ieqn-24"><mml:math id="mml-ieqn-24"><mml:mn>8</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula></td>
<td>32,768</td>
<td>8</td>
<td>32,776</td>
</tr>
<tr>
<td>Output</td>
<td><inline-formula id="ieqn-25"><mml:math id="mml-ieqn-25"><mml:mn>8</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula></td>
<td>0</td>
<td>0</td>
<td>0</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The number of parameters of the convolutional layer is formulated as</p>
<p><disp-formula id="eqn-1">
<label>(1)</label>
<!--<tex-math id="tex-eqn-1"><![CDATA[$$\begin{equation}W_{c}=K^{2}\times C\times N,
 \label{eqn-1} \end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-1" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mi>K</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x00D7;</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>N</mml:mi><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula></p>
<p><disp-formula id="eqn-2">
<label>(2)</label>
<!--<tex-math id="tex-eqn-2"><![CDATA[$$\begin{equation}B_{c}=N,
 \label{eqn-2} \end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-2" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>N</mml:mi><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula></p>
<disp-formula id="eqn-3">
<label>(3)</label>
<!--<tex-math id="tex-eqn-3"><![CDATA[$$\begin{equation}P_{c}=W_{c}+B_{c},
 \label{eqn-3}\end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-3" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula>
<p><italic>where</italic></p>
<p><inline-formula id="ieqn-26"><mml:math id="mml-ieqn-26"><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>w</mml:mi><mml:mi>e</mml:mi><mml:mi>i</mml:mi><mml:mi>g</mml:mi><mml:mi>h</mml:mi><mml:mi>t</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>o</mml:mi><mml:mi>l</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-27"><mml:math id="mml-ieqn-27"><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>b</mml:mi><mml:mi>a</mml:mi><mml:mi>s</mml:mi><mml:mi>e</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>o</mml:mi><mml:mi>l</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-28"><mml:math id="mml-ieqn-28"><mml:msub><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>p</mml:mi><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>m</mml:mi><mml:mi>e</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>o</mml:mi><mml:mi>l</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mo>,</mml:mo></mml:math></inline-formula></p>
<p><inline-formula id="ieqn-29"><mml:math id="mml-ieqn-29"><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:mi>s</mml:mi><mml:mi>i</mml:mi><mml:mi>z</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>k</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>u</mml:mi><mml:mi>s</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>o</mml:mi><mml:mi>l</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mo>,</mml:mo></mml:math></inline-formula></p>
<p><inline-formula id="ieqn-30"><mml:math id="mml-ieqn-30"><mml:mi>N</mml:mi><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>k</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mi>s</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-31"><mml:math id="mml-ieqn-31"><mml:mi>C</mml:mi><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>h</mml:mi><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mi>p</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>i</mml:mi><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>g</mml:mi><mml:mi>e</mml:mi></mml:mstyle></mml:math></inline-formula>.</p>
<p>The size (O) of the output tensor (image) of the maximum pool layer is formulated as</p>
<p><disp-formula id="eqn-4">
<label>(4)</label>
<!--<tex-math id="tex-eqn-4"><![CDATA[$$\begin{equation}
O=\frac{I-P_{s}}{S}+1,
 \label{eqn-4}
\end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-4" display="block"><mml:mi>O</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>I</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>S</mml:mi></mml:mrow></mml:mfrac><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo></mml:math></disp-formula></p>
<p><italic>where</italic></p>
<p><inline-formula id="ieqn-32"><mml:math id="mml-ieqn-32"><mml:mi>O</mml:mi><mml:mo>=</mml:mo><mml:mi>s</mml:mi><mml:mi>i</mml:mi><mml:mi>z</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>o</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>p</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>i</mml:mi><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>g</mml:mi><mml:mi>e</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-33"><mml:math id="mml-ieqn-33"><mml:mi>I</mml:mi><mml:mo>=</mml:mo><mml:mi>s</mml:mi><mml:mi>i</mml:mi><mml:mi>z</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mi>p</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>i</mml:mi><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>g</mml:mi><mml:mi>e</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-34"><mml:math id="mml-ieqn-34"><mml:mi>S</mml:mi><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>o</mml:mi><mml:mi>l</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>o</mml:mi><mml:mi>p</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><italic>P<sub>s</sub></italic> = <italic>pool</italic>  <italic>size</italic>.</p>
<p><inline-formula id="ieqn-35"><mml:math id="mml-ieqn-35"><mml:mstyle class="text"><mml:mtext>The&#x00A0;number&#x00A0;of&#x00A0;parameters&#x00A0;</mml:mtext></mml:mstyle><mml:mstyle class="text"><mml:mtext>is&#x00A0;formulated&#x00A0;</mml:mtext></mml:mstyle></mml:math></inline-formula>as</p>
<p><disp-formula id="eqn-5">
<label>(5)</label>
<!--<tex-math id="tex-eqn-5"><![CDATA[$$\begin{equation}W_{ff}=F_{-1}\times F,
 \label{eqn-5} \end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-5" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mo lspace='0pt' rspace='0pt'>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>&#x00D7;</mml:mo><mml:mi>F</mml:mi><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula></p>
<p><disp-formula id="eqn-6">
<label>(6)</label>
<!--<tex-math id="tex-eqn-6"><![CDATA[$$\begin{equation}B_{ff}=F,
 \label{eqn-6} \end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-6" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>F</mml:mi><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula></p>
<disp-formula id="eqn-7">
<label>(7)</label>
<!--<tex-math id="tex-eqn-7"><![CDATA[$$\begin{equation}P_{ff}=W_{ff}+B_{ff},
 \label{eqn-7}\end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-7" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula>
<p><italic>where</italic></p>
<p><inline-formula id="ieqn-36"><mml:math id="mml-ieqn-36"><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>w</mml:mi><mml:mi>e</mml:mi><mml:mi>i</mml:mi><mml:mi>g</mml:mi><mml:mi>h</mml:mi><mml:mi>t</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>w</mml:mi><mml:mi>h</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>h</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-37"><mml:math id="mml-ieqn-37"><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>b</mml:mi><mml:mi>i</mml:mi><mml:mi>a</mml:mi><mml:mi>s</mml:mi><mml:mi>e</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>w</mml:mi><mml:mi>h</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>h</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-38"><mml:math id="mml-ieqn-38"><mml:msub><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>p</mml:mi><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>m</mml:mi><mml:mi>e</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>w</mml:mi><mml:mi>h</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>h</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-39"><mml:math id="mml-ieqn-39"><mml:mi>F</mml:mi><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>u</mml:mi><mml:mi>r</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>,</p>
<p><inline-formula id="ieqn-40"><mml:math id="mml-ieqn-40"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mo lspace='0pt' rspace='0pt'>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>u</mml:mi><mml:mi>m</mml:mi><mml:mi>b</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>o</mml:mi><mml:mi>f</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>n</mml:mi><mml:mi>e</mml:mi><mml:mi>u</mml:mi><mml:mi>r</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>v</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>u</mml:mi><mml:mi>s</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>F</mml:mi><mml:mi>C</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>y</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi></mml:mstyle></mml:math></inline-formula>.</p>
<p><xref ref-type="fig" rid="fig-3">Fig. 3</xref> shows the features extracted from an image by the deep CNN model. In the first convolutional layer, conv1 extracts the edges and blobs, conv2 and conv3 extract the texture, conv4 and conv5 extract the object parts, and the last fully connected layer detects the object classes.</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>Features extraction of the 5 convolutional layers 
 
</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-3.png"/>
</fig>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Random images of vehicles from training dataset 
 
</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-4.png"/>
</fig>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Vehicle Type Classification</title>
<p>We discuss the classification of multiple categories of vehicles through several experiments on the dataset. The parameters used in the AlexNet model were customized for optimal results. The network was trained by splitting the dataset into 70% for training and 30% for validation. The activations of the pre-trained model learned patterns in different datasets through transfer learning. All the layers of the pre-trained network were extracted, except the last three layers, which were configured for 1000 classes. We fine-tuned these three layers for our classification problem. Training was optimized by setting the minibatch size to 7, maximum epochs to 35, and learning rate to 1e &#x2212;5. Experiments were evaluated using MATLAB with a deep learning toolbox. The proposed model was trained on a GPU. The elapsed time was 47 s. The model took random images from the validation dataset and accurately labeled them according to their type. <xref ref-type="fig" rid="fig-4">Fig. 4</xref> shows random images from the training dataset, classified with their class labels from the testing (validation) dataset. Images were correctly classified according to their type, except the one in the third row and first column; this means that the model needed more data for training of that particular image to give better results.</p>
</sec>
<sec id="s3_4">
<label>3.4</label>
<title>Evolution Method</title>
<p>We present accuracy as the evaluation method. This is computed with the help of a confusion matrix, as shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>, which shows the performance of the algorithm for each class of vehicle. Accuracy shows the percentage of the correctly predicted class in the entire testing dataset, and is formulated as</p>
<p><disp-formula id="eqn-8">
<label>(8)</label>
<!--<tex-math id="tex-eqn-8"><![CDATA[$$\begin{equation}
\mathit{Accuracy}=\frac{\mathit{correctly}~\mathit{predicted}~\mathit{class}}
{\mathit{total}~\mathit{testing}~\mathit{class}}\times 100.
 \label{eqn-8}
\end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-8" display="block"><mml:mstyle mathvariant="italic"><mml:mi>A</mml:mi><mml:mi>c</mml:mi><mml:mi>c</mml:mi><mml:mi>u</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>c</mml:mi><mml:mi>y</mml:mi></mml:mstyle><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>r</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>l</mml:mi><mml:mi>y</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>s</mml:mi><mml:mi>s</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mstyle mathvariant="italic"><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mi>t</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>t</mml:mi><mml:mi>e</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mi>g</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>s</mml:mi><mml:mi>s</mml:mi></mml:mstyle></mml:mrow></mml:mfrac><mml:mo>&#x00D7;</mml:mo><mml:mn>100</mml:mn><mml:mo>.</mml:mo></mml:math></disp-formula></p>
<fig id="fig-5">
<label>Figure 5</label> 
<caption>
<title>Confusion matrix of the optimized vehicle classification approach 
 
</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-5.png"/>
</fig>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Results and Discussion</title>
<p>We discuss the experimental assessment for the detection and classification of vehicle types. We evaluated the dataset with the stochastic gradient descent with momentum (SGDM) and adaptive moment estimation (Adam) algorithms with epochs ranging from 10 to 40. The optimizers were used to change the weights of CNN to reduce the losses. During training, optimization is a key component that helps the model to adjust weights during backpropagation [<xref ref-type="bibr" rid="ref-29">29</xref>,<xref ref-type="bibr" rid="ref-30">30</xref>], which is formulated as:</p>
<p><disp-formula id="eqn-9">
<label>(9)</label>
<!--<tex-math id="tex-eqn-9"><![CDATA[$$\begin{equation}
\mathit{repeat}~\mathit{until}~ \mathit{convergence}\{\theta _{j}:=\theta _{j}-\propto \frac{\partial }{\partial \theta _{j}}j \left(\theta _{0},\,\theta _{1}\right)\}
 \label{eqn-9}
\end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-9" display="block"><mml:mstyle mathvariant="italic"><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>p</mml:mi><mml:mi>e</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>u</mml:mi><mml:mi>n</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>l</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi><mml:mi>e</mml:mi><mml:mi>r</mml:mi><mml:mi>g</mml:mi><mml:mi>e</mml:mi><mml:mi>n</mml:mi><mml:mi>c</mml:mi><mml:mi>e</mml:mi></mml:mstyle><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mo>&#x221D;</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x2202;</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x2202;</mml:mi><mml:msub><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac><mml:mi>j</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mspace width="0.3em"/><mml:msub><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mo>}</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>where j = 0, 1 represents the feature index number.</p>
<sec id="s4_1">
<label>4.1</label>
<title>Results with SGDM Optimizer</title>
<p>The gradient vectors were accelerated in the right direction, leading to fast convergence with the help of SGD with momentum (SGDM) [<xref ref-type="bibr" rid="ref-31">31</xref>]. We trained the model with an SGDM optimizer by applying different epochs. The SGDM algorithm is formulated as:</p>
<p><disp-formula id="eqn-10">
<label>(10)</label>
<!--<tex-math id="tex-eqn-10"><![CDATA[$$\begin{align}
&\vartheta =\theta -\alpha . \Delta _{J}(\theta . h \left(i\right);\,h \left(j\right))
 \label{eqn-10} \\
&h \left(i\right);\,h \left(j\right)~ are~ the~ \mathit{training}~ data
\nonumber \\
&\vartheta =\text{updated weight}
\nonumber \\
&\alpha =\text{learning rate}
\nonumber \\
&\Delta _{J}=cost~ \mathit{function}.
\nonumber
\end{align}$$]]></tex-math>-->
<mml:math id="mml-eqn-10" display="block"><mml:mtable columnalign="right left" columnspacing="1pt"><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mi>&#x03D1;</mml:mi><mml:mo>=</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>-</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo>.</mml:mo><mml:msub><mml:mrow><mml:mtext>&#x0394;</mml:mtext></mml:mrow><mml:mrow><mml:mi>J</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>&#x03B8;</mml:mi><mml:mo>.</mml:mo><mml:mi>h</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>;</mml:mo><mml:mspace width="0.3em"/><mml:mi>h</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mi>h</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>;</mml:mo><mml:mspace width="0.3em"/><mml:mi>h</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mi>g</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>d</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>a</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mi>&#x03D1;</mml:mi><mml:mo>=</mml:mo><mml:mstyle><mml:mtext>updated&#x00A0;weight</mml:mtext></mml:mstyle></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mi>&#x03B1;</mml:mi><mml:mo>=</mml:mo><mml:mstyle><mml:mtext>learning&#x00A0;rate</mml:mtext></mml:mstyle></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:msub><mml:mrow><mml:mtext>&#x0394;</mml:mtext></mml:mrow><mml:mrow><mml:mi>J</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>f</mml:mi><mml:mi>u</mml:mi><mml:mi>n</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi></mml:mstyle><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
<p>The model first trained with 10 epochs, giving 94.06% accuracy with a collapse time of 15 s. For better performance, the model was trained further with 15, 20, 25, 30, 35, and 40 epochs. The number of iterations is directly proportional to the number of epochs. It is observed that the system showed the best result on epoch 35, with 95.02% accuracy. System performance then started to decline due to overfitting.</p>
</sec>
<sec id="s4_2">
<label>4.2</label>
<title>Results with ADAM Optimizer</title>
<p>The Adam optimizer iteratively updated the network weights in the training dataset [<xref ref-type="bibr" rid="ref-29">29</xref>]. After training the model with an SGDM optimizer, it was assessed with an Adam optimizer for improved accuracy. The Adam algorithm is formulated as:</p>
<p><disp-formula id="eqn-11">
<label>(11)</label>
<!--<tex-math id="tex-eqn-11"><![CDATA[$$\begin{equation}\hat{m}_{t}=\frac{m_{t}}{1-\beta _{1}^{t}}
 \label{eqn-11} \end{equation}$$]]></tex-math>-->
<mml:math id="mml-eqn-11" display="block"><mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msubsup></mml:mrow></mml:mfrac></mml:mrow><mml:mrow></mml:mrow></mml:math>
</disp-formula></p>
<p><disp-formula id="eqn-12">
<label>(12)</label>
<!--<tex-math id="tex-eqn-12"><![CDATA[$$\begin{align}&\hat{v}_{t}=\frac{v_{t}}{1-\beta _{2}^{t}}
 \label{eqn-12} \\
& \text{w}\text{h}\text{e}\text{r}\text{e}\nonumber \end{align}$$]]></tex-math>-->
<mml:math id="mml-eqn-12" display="block"><mml:mtable columnalign="right left" columnspacing="1pt"><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msubsup></mml:mrow></mml:mfrac></mml:mstyle></mml:mtd><mml:mtd></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mstyle><mml:mtext>w</mml:mtext></mml:mstyle><mml:mstyle><mml:mtext>h</mml:mtext></mml:mstyle><mml:mstyle><mml:mtext>e</mml:mtext></mml:mstyle><mml:mstyle><mml:mtext>r</mml:mtext></mml:mstyle><mml:mstyle><mml:mtext>e</mml:mtext></mml:mstyle></mml:mtd><mml:mtd></mml:mtd></mml:mtr><mml:mtr><mml:mtd></mml:mtd><mml:mtd><mml:mi>m</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mi>d</mml:mi><mml:mspace width=".3em" /><mml:mi>v</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mi>t</mml:mi><mml:mi>h</mml:mi><mml:mi>e</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>fi</mml:mi><mml:mi>r</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>e</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>m</mml:mi><mml:mi>e</mml:mi><mml:mi>n</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mi>a</mml:mi><mml:mi>n</mml:mi><mml:mi>d</mml:mi><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>s</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>d</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>e</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>m</mml:mi><mml:mi>e</mml:mi><mml:mi>n</mml:mi><mml:mi>t</mml:mi></mml:mstyle><mml:mspace width=".3em" /><mml:mstyle mathvariant="italic"><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>s</mml:mi><mml:mi>p</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>v</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mi>y</mml:mi></mml:mstyle><mml:mo>.</mml:mo></mml:mtd><mml:mtd></mml:mtd></mml:mtr></mml:mtable></mml:math>
</disp-formula></p>
<p>The model was trained with the same number of epochs as previously described. It is observed that the best result obtained with the Adam optimizer was at 35 epochs, with 90.10% accuracy. The literature shows that SGDM performs better than Adam in reducing the loss [<xref ref-type="bibr" rid="ref-31">31</xref>]. The experimental results in <xref ref-type="table" rid="table-4">Tabs. 4</xref> and <xref ref-type="table" rid="table-5">5</xref> show that the accuracy of the model with SGDM was 95.02% with 46 s training time, while the accuracy with Adam was 90.10%, with a training time of 50 s. Therefore, SGDM provided better accuracy than Adam.</p>
<table-wrap id="table-4">
<label>Table 4</label>
<caption>
<title>Model accuracy with SGDM optimizer</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Optimizer</th>
<th>Epochs</th>
<th>Iteration</th>
<th>Iteration/epoch</th>
<th>Validation accuracy (%)</th>
<th>Training time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td>SGDM</td>
<td>10</td>
<td>70</td>
<td>7</td>
<td>94.06</td>
<td>15</td>
</tr>
<tr>
<td>SGDM</td>
<td>15</td>
<td>105</td>
<td>7</td>
<td>94.06</td>
<td>20</td>
</tr>
<tr>
<td>SGDM</td>
<td>20</td>
<td>140</td>
<td>7</td>
<td>94.06</td>
<td>26</td>
</tr>
<tr>
<td>SGDM</td>
<td>25</td>
<td>175</td>
<td>7</td>
<td>94.06</td>
<td>33</td>
</tr>
<tr>
<td>SGDM</td>
<td>30</td>
<td>210</td>
<td>7</td>
<td>94.08</td>
<td>41</td>
</tr>
<tr>
<td><bold>SGDM</bold></td>
<td><bold>35</bold></td>
<td><bold>245</bold></td>
<td><bold>7</bold></td>
<td><bold>95.02</bold></td>
<td><bold>46 </bold></td>
</tr>
<tr>
<td>SGDM</td>
<td>40</td>
<td>280</td>
<td>7</td>
<td>88.42</td>
<td>54</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap id="table-5">
<label>Table 5</label>
<caption>
<title>Model accuracy with ADAM optimizer</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Optimizer</th>
<th>Epochs</th>
<th>Iteration</th>
<th>Iteration/epoch</th>
<th>Validation accuracy (%)</th>
<th>Training time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td>ADAM</td>
<td>10</td>
<td>70</td>
<td>7</td>
<td>84.16</td>
<td>15</td>
</tr>
<tr>
<td>ADAM</td>
<td>15</td>
<td>105</td>
<td>7</td>
<td>85.12</td>
<td>20</td>
</tr>
<tr>
<td>ADAM</td>
<td>20</td>
<td>140</td>
<td>7</td>
<td>87.11</td>
<td>29</td>
</tr>
<tr>
<td>ADAM</td>
<td>25</td>
<td>175</td>
<td>7</td>
<td>88.15</td>
<td>30</td>
</tr>
<tr>
<td>ADAM</td>
<td>30</td>
<td>210</td>
<td>7</td>
<td>89.12</td>
<td>43</td>
</tr>
<tr>
<td><bold>ADAM</bold></td>
<td><bold>35</bold></td>
<td><bold>245</bold></td>
<td><bold>7</bold></td>
<td><bold>90.10</bold></td>
<td><bold>50 </bold></td>
</tr>
<tr>
<td>ADAM</td>
<td>40</td>
<td>280</td>
<td>7</td>
<td>87.13</td>
<td>58</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s4_3">
<label>4.3</label>
<title>Model Accuracy with SGDM Optimizer and Different Learning Rates</title>
<p>For more convincing results, the model was evaluated with different learning rates with 35 epochs, which gives the best results in <xref ref-type="table" rid="table-4">Tabs. 4</xref> and <xref ref-type="table" rid="table-5">5</xref>. The learning rate determines the step size at each iteration in an optimization algorithm. It is a tuning parameter that minimizes the loss function [<xref ref-type="bibr" rid="ref-32">32</xref>,<xref ref-type="bibr" rid="ref-33">33</xref>]. The model was trained again with SGDM and 35 epochs, with learning rates ranging from 1e &#x2212;3 to 1e &#x2212;6. The learning rate affected the quick convergence of the model toward local minima. It improved the performance of our model from 95.02% to 96.04% by adjusting the learning rate <inline-formula id="ieqn-41"><mml:math id="mml-ieqn-41"><mml:mi>&#x03B1;</mml:mi><mml:mo>=</mml:mo></mml:math></inline-formula> 1e &#x2212;5, as shown in <xref ref-type="table" rid="table-6">Tab. 6</xref>.</p>
<table-wrap id="table-6">
<label>Table 6</label>
<caption>
<title>Model accuracy with SGDM optimizer with different learning rates</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>Optimizer</th>
<th>Epochs</th>
<th>Learning rate</th>
<th>Training time (s)</th>
<th>Validation accuracy (%)</th>
</tr>
</thead>
<tbody>
<tr>
<td>SGDM</td>
<td>35</td>
<td>1e<sup>&#x2212;6</sup></td>
<td>52</td>
<td>88.12</td>
</tr>
<tr>
<td><bold>SGDM</bold></td>
<td><bold>35</bold></td>
<td><bold>1e</bold><inline-formula id="ieqn-42"><mml:math id="mml-ieqn-42"><mml:msup><mml:mrow></mml:mrow><mml:mrow><mml:mstyle class="text"><mml:mtext class="textbf" mathvariant="bold">&#x2013;5</mml:mtext></mml:mstyle></mml:mrow></mml:msup></mml:math></inline-formula></td>
<td><bold>47</bold></td>
<td><bold>96.04</bold></td>
</tr>
<tr>
<td>SGDM</td>
<td>35</td>
<td>1e<sup>&#x2212;4</sup></td>
<td>46</td>
<td>95.02</td>
</tr>
<tr>
<td>SGDM</td>
<td>35</td>
<td>1e<sup>&#x2212;3</sup></td>
<td>45</td>
<td>76.3</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><xref ref-type="fig" rid="fig-6">Figs. 6</xref> and <xref ref-type="fig" rid="fig-7">7</xref> show the training progress of the model used in this study. The upper graphs show the accuracy of the model, which is measured by the performance estimated on a set of samples from the test data. The lower graph shows the loss of the model. It is computed by the gradient of the loss function concerning the parameters [<xref ref-type="bibr" rid="ref-34">34</xref>], and it shows the gap between the actual and expected output scores. During the first epoch, the rate of accuracy increased from 20% to 80% due to backpropagating gradients that updated the maximum weights of the filters in our classification task. For this, SGDM was set at a learning rate of 0.00001. This allowed for fine-tuning to make progress in the remaining epochs. The accuracy fluctuated between 80% and 95% because the maximum number of weights had been trained. Similarly, during the first epoch, the loss decreased dramatically from 2.5 to 0.5 in each iteration; the model minimized the error by updating the weights. By the completion of 35 epochs, the proposed model had 96.04% accuracy on the validation dataset.</p>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Best result with SGDM optimizer 
 
</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-6.png"/>
</fig>
<fig id="fig-7">
<label>Figure 7</label>
<caption>
<title>Best result with Adam optimizer 
 
</title>
</caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CMC_15504-fig-7.png"/>
</fig>
</sec>
<sec id="s4_4">
<label>4.4</label>
<title>Comparative Analysis</title>
<p>We evaluated the proposed method against state-of-the-art deep learning methods [<xref ref-type="bibr" rid="ref-35">35</xref>]. Dong used CNN for automatic feature extraction by classifying four categories of vehicles, with 89.4% accuracy. Another method, based on a comparative analysis of ANN, SVM, and logistic regression, classified small vehicles and big vehicles with 93.4% accuracy. Huttuman automatically extracted features using deep neural networks, and classified four classes of vehicles with 97% accuracy. Adu Gyamfi used a deep CNN to classify 13 vehicle classes with 89% accuracy. The last two methods, as mentioned in <xref ref-type="table" rid="table-7">Tab. 7</xref>, used LeNet, AlexNet, VGG-16, and an inception module for vehicle classification, with respective accuracies of 80% and 80.3% which are much lower than those of our proposed system. Our method achieved better accuracy than all of the above methods except the Huttuman approach, whose accuracy was high, but could classify only four classes of vehicles while the proposed method classified eight classes of vehicles. <xref ref-type="table" rid="table-7">Tab. 7</xref> shows a comparative analysis of the proposed method.</p>
<table-wrap id="table-7">
<label>Table 7</label>
<caption>
<title>Comparative analysis of the proposed method</title>
</caption>
<table>
<colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>No.</th>
<th>Publications</th>
<th>Accuracy (%)</th>
<th>Key features</th>
<th>Vehicle classes</th>
<th>Vehicle categories</th>
</tr>
</thead>
<tbody>
<tr>
<td>1</td>
<td>Dong et al. TITS&#x2019; 15 [<xref ref-type="bibr" rid="ref-36">36</xref>]</td>
<td>89.4</td>
<td>Automatic feature extraction using CNN and Softmax classifier for multi-task learning</td>
<td>Truck, minivan, passenger car, sedan</td>
<td>4</td>
</tr>
<tr>
<td>2</td>
<td>Denis et al.&#x2019; 15 [<xref ref-type="bibr" rid="ref-24">24</xref>]</td>
<td>93.4</td>
<td>ANN, SVM, Logistic regression</td>
<td>Light motor vehicle, high motor vehicle</td>
<td>2</td>
</tr>
<tr>
<td>3</td>
<td>Huttuman et al.&#x2019; 16 [<xref ref-type="bibr" rid="ref-37">37</xref>]</td>
<td>97</td>
<td>Automatically extracted features using DNN</td>
<td>Bus, truck, van, small car</td>
<td>4</td>
</tr>
<tr>
<td>4</td>
<td>Adu_Gyamfi et al.&#x2019; 17 [<xref ref-type="bibr" rid="ref-38">38</xref>]</td>
<td>89</td>
<td>Deep convolutional neural network</td>
<td>13 vehicle classes</td>
<td>13</td>
</tr>
<tr>
<td>5</td>
<td>Audebert et al.&#x2019; 17 [<xref ref-type="bibr" rid="ref-39">39</xref>]</td>
<td>80</td>
<td>LeNet, AlexNet, VGG-16 for vehicle classification</td>
<td>Sedans, vans, pickups, trucks</td>
<td>4</td>
</tr>
<tr>
<td>6</td>
<td>Tan et al., 18 [<xref ref-type="bibr" rid="ref-40">40</xref>]</td>
<td>80.3</td>
<td>AlexNet and inception module for vehicle classification</td>
<td>Sedans, vans, pickups, trucks</td>
<td>4</td>
</tr>
<tr>
<td>7</td>
<td><bold>Proposed system</bold></td>
<td><bold>96.04</bold></td>
<td><bold>CNN with SGDM and ADAM optimizer</bold></td>
<td><bold>Bus, van, truck, bike, car, Jeep, Horse Buggy</bold></td>
<td><bold>8</bold></td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Conclusion</title>
<p>Our method detects and classifies multiple classes of vehicles through a deep learning model. It can help in vehicle tracking systems for surveillance in big parking slots where security is a concern. It can help solve traffic issues by directing large vehicles to one side of a road and keep the traffic moving by knowing what vehicles are ahead in a queue. This research will help by classifying the vehicles in parking zones and automatically allocate tickets according to vehicle types. The accuracy of the model can be improved by increasing the sample size.</p>
</sec>
</body>
<back>
<ack><p>We acknowledge the overall paper editing support by Dr. Sheroz Khan and Dr. Muhammad Islam, Department of Electrical Engineering and Renewable Engineering, College of Engineering &#x0026; Information Technology, Onaizah Colleges, Al-Qassim, Saudi Arabia;2053, Saudi Arabia. We thank LetPub (www.letpub.com) for its linguistic assistance during the preparation of this manuscript.</p></ack>
<fn-group><fn fn-type="other"><p><bold>Funding Statement:</bold> This work is supported by the Information Technology Department, College of Computer, Qassim University, 6633, Buraidah 51452, Saudi Arabia.</p></fn>
<fn fn-type="conflict"><p><bold>Conflicts of Interest:</bold> The authors declare that they have no conflicts of interest to report regarding the present study.</p></fn></fn-group>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Y. L.</given-names> <surname>Tian</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Lu</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Hampapur</surname></string-name></person-group>, &#x201C;<article-title>Robust and efficient foreground analysis for real-time video surveillance</article-title>,&#x201D; in <conf-name>IEEE Computer Society Conf. on Computer Vision and Pattern Recognition</conf-name>, vol. <volume>1</volume>, pp. <fpage>1182</fpage>&#x2013;<lpage>1187</lpage>, <year>2005</year>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Sun</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Wang</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Zhang</surname></string-name></person-group>, &#x201C;<article-title>Deep learning for plant identification in natural environment</article-title>,&#x201D; <source>Computational Intelligence and Neuroscience</source>, vol. <volume>2017</volume>, no. <issue>4</issue>, pp. <fpage>1</fpage>&#x2013;<lpage>6</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Maja</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Pentland</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Vinciarelli</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Cucchiara</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Daoudi</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Joint ACM workshop on human gesture and behavior understanding</article-title>,&#x201D; in <conf-name> Proc. of the 19th ACM Int. Conf. on Multimedia</conf-name>, vol. <volume>19</volume>, pp. <fpage>615</fpage>&#x2013;<lpage>616</lpage>, <year>2011</year>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y. L.</given-names> <surname>Tian</surname></string-name>, <string-name><given-names>R. S.</given-names> <surname>Feris</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Hampapur</surname></string-name> and <string-name><given-names>M. T.</given-names> <surname>Sun</surname></string-name></person-group>, &#x201C;<article-title>Robust detection of abandoned and removed objects in complex surveillance videos</article-title>,&#x201D; <source>IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews) 41</source>, vol. <volume>5</volume>, pp. <fpage>565</fpage>&#x2013;<lpage>576</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Deng</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Dong</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Socher</surname></string-name>, <string-name><given-names>L. J.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Li</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Imagenet: A large-scale hierarchical image database</article-title>,&#x201D; in <conf-name>IEEE Conf. on Computer Vision and Pattern Recognition</conf-name>, Florida, USA, pp. <fpage>248</fpage>&#x2013;<lpage>255</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>W.</given-names> <surname>Zhan</surname></string-name> and <string-name><given-names>Q.</given-names> <surname>Wan</surname></string-name></person-group>, &#x201C;<article-title>Real-time and automatic vehicle type recognition system design and its application</article-title>,&#x201D; in <conf-name>Int. Symp. on Intelligence Computation and Applications</conf-name>, Wuhan, China, pp. <fpage>200</fpage>&#x2013;<lpage>208</lpage>, <year>2012</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Xiong</surname></string-name></person-group>, <source>Research of Automobile Classifying Method Based on Inductive Loop</source>. <publisher-loc>Hunan, China</publisher-loc>: <publisher-name>ChangSha University of Science and Technology</publisher-name>, pp. <fpage>2</fpage>&#x2013;<lpage>5</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Krizhevsky</surname></string-name>, <string-name><given-names>I.</given-names> <surname>Sutskever</surname></string-name> and <string-name><given-names>G. E.</given-names> <surname>Hinton</surname></string-name></person-group>, &#x201C;<article-title>Imagenet classification with deep convolutional neural networks</article-title>,&#x201D; <source>Communications of the ACM 60</source>, vol. <volume>6</volume>, no. <issue>6</issue>, pp. <fpage>84</fpage>&#x2013;<lpage>90</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>V.</given-names> <surname>Sze</surname></string-name>, <string-name><given-names>Y. H.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>T. J.</given-names> <surname>Yang</surname></string-name> and <string-name><given-names>J. S.</given-names> <surname>Emer</surname></string-name></person-group>, &#x201C;<article-title>Efficient processing of deep neural networks: A tutorial and survey</article-title>,&#x201D; <source>Proc. of the IEEE</source>, vol. <volume>12</volume>, pp. <fpage>2295</fpage>&#x2013;<lpage>2329</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>O.</given-names> <surname>Russakovsky</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Deng</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Su</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Krause</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Satheesh</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Imagenet large scale visual recognition challenge</article-title>,&#x201D; <source>International Journal of Computer Vision 115</source>, vol. <volume>3</volume>, no. <issue>3</issue>, pp. <fpage>211</fpage>&#x2013;<lpage>252</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H. C.</given-names> <surname>Shin</surname></string-name>, <string-name><given-names>H. R.</given-names> <surname>Roth</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Gao</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Lu</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Xu</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Deep convolutional neural networks for computer-aided detection: CNN architectures, dataset characteristics and transfer learning</article-title>,&#x201D; <source>IEEE Transactions on Medical Imaging 35</source>, vol. <volume>5</volume>, no. <issue>5</issue>, pp. <fpage>1285</fpage>&#x2013;<lpage>1298</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>X.</given-names> <surname>Wen</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Shao</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Xue</surname></string-name> and <string-name><given-names>W.</given-names> <surname>Fang</surname></string-name></person-group>, &#x201C;<article-title>A rapid learning algorithm for vehicle classification</article-title>,&#x201D; <source>Information Sciences</source>, vol. <volume>295</volume>, no. <issue>3</issue>, pp. <fpage>395</fpage>&#x2013;<lpage>406</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>D.</given-names> <surname>Zhao</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Chen</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Lv</surname></string-name></person-group>, &#x201C;<article-title>Deep reinforcement learning with visual attention for vehicle classification</article-title>,&#x201D; <source>IEEE Transactions on Cognitive and Developmental Systems 9</source>, vol. <volume>4</volume>, pp. <fpage>356</fpage>&#x2013;<lpage>367</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>D. P.</given-names> <surname>Pietro</surname></string-name> and <string-name><given-names>F.</given-names> <surname>Hristea</surname></string-name></person-group>, &#x201C;<article-title>Unsupervised word sense disambiguation with N-gram features</article-title>,&#x201D; <source>Artificial Intelligence Review 41</source>, vol. <volume>2</volume>, no. <issue>2</issue>, pp. <fpage>241</fpage>&#x2013;<lpage>260</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Liaw</surname></string-name> and <string-name><given-names>M.</given-names> <surname>Wiener</surname></string-name></person-group>, &#x201C;<article-title>Classification and regression by random forest</article-title>,&#x201D; <source>R News</source>, vol. <volume>2</volume>, no. <issue>3</issue>, pp. <fpage>18</fpage>&#x2013;<lpage>22</lpage>, <year>2002</year>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Shotton</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Winn</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Rother</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Criminisi</surname></string-name></person-group>, &#x201C;<article-title>Textonboost for image understanding: Multi-class object recognition and segmentation by jointly modeling texture, layout, and context</article-title>,&#x201D; <source>International Journal of Computer Vision</source>, vol. <volume>81</volume>, no. <issue>1</issue>, pp. <fpage>2</fpage>&#x2013;<lpage>23</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>N.</given-names> <surname>Dalal</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Triggs</surname></string-name></person-group>, &#x201C;<article-title>Histograms of oriented gradients for human detection</article-title>,&#x201D; in <conf-name>2005 IEEE Computer Society Conf. on Computer Vision and Pattern Recognition</conf-name>, San Diego, CA, USA, <publisher-name>IEEE</publisher-name>, <year>2005</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>D. G.</given-names> <surname>Lowe</surname></string-name></person-group>, &#x201C;<article-title>Distinctive image features from scale-invariant keypoints</article-title>,&#x201D; <source>International Journal of Computer Vision</source>, vol. <volume>60</volume>, no. <issue>2</issue>, pp. <fpage>91</fpage>&#x2013;<lpage>110</lpage>, <year>2004</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Fu</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Ma</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name> and <string-name><given-names>D.</given-names> <surname>Lu</surname></string-name></person-group>, &#x201C;<article-title>A vehicle classification system based on hierarchical multi-SVMs in crowded traffic scenes</article-title>,&#x201D; <source>Neuro Computing</source>, vol. <volume>211</volume>, pp. <fpage>182</fpage>&#x2013;<lpage>190</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Nurhadiyatna</surname></string-name>, <string-name><given-names>A. L.</given-names> <surname>Latifah</surname></string-name> and <string-name><given-names>D.</given-names> <surname>Fryantoni</surname></string-name></person-group>, &#x201C;<article-title>Gabor filtering for feature extraction in real time vehicle classification system</article-title>,&#x201D; in <conf-name>9th Int. Symp. on Image and Signal Processing and Analysis</conf-name>, Zagreb, Croatia, <publisher-name>IEEE</publisher-name>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Kaiming</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Ren</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Sun</surname></string-name></person-group>, &#x201C;<article-title>Deep residual learning for image recognition</article-title>,&#x201D; in <conf-name>Proc. of the IEEE Conf. on Computer Vision and Pattern Recognition</conf-name>, Las Vegas, NV, USA, pp. <fpage>770</fpage>&#x2013;<lpage>778</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Pierre</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Eigen</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Mathieu</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Fergus</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Overfeat: Integrated recognition, localization and detection using convolutional networks</article-title>,&#x201D; <source>ArXiv</source>, vol. <volume>13</volume>, no. <issue>12</issue>, pp. <fpage>62</fpage>&#x2013;<lpage>92</lpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Sun</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Wang</surname></string-name> and <string-name><given-names>X.</given-names> <surname>Tang</surname></string-name></person-group>, &#x201C;<article-title>Deep learning face representation from predicting 10,000 classes</article-title>,&#x201D; in <conf-name>Proc. of the IEEE Conf. on Computer Vision and Pattern Recognition</conf-name>, Columbus, OH, USA, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>K.</given-names> <surname>Denis</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Hostettler</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Birk</surname></string-name> and <string-name><surname>Evgeny Osipov</surname></string-name></person-group>, &#x201C;<article-title>Comparison of machine learning techniques for vehicle classification using road side sensors</article-title>,&#x201D; in <conf-name>IEEE 18th Int. Conf. on Intelligent Transportation Systems</conf-name>, Rio de Janeiro, Brazil, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Wu</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zheng</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Lv</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Automatic vehicle classification using roadside LiDAR data</article-title>,&#x201D; <source>Transportation Research Record</source>, vol. <volume>2673</volume>, no. <issue>6</issue>, pp. <fpage>153</fpage>&#x2013;<lpage>164</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Samuele</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Facheris</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Cuccoli</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Marinai</surname></string-name></person-group>, &#x201C;<article-title>Vehicle classification based on convolutional networks applied to FMCW radar signals</article-title>,&#x201D; in <conf-name>Italian Conf. for the Traffic Police</conf-name>, <publisher-loc>Berlin, Germany</publisher-loc>, <publisher-name>Springer</publisher-name>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>W.</given-names> <surname>Balid</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Tafish</surname></string-name> and <string-name><given-names>H. H.</given-names> <surname>Refai</surname></string-name></person-group>, &#x201C;<article-title>Intelligent vehicle counting and classification sensor for real-time traffic surveillance</article-title>,&#x201D; <source>IEEE Transactions on Intelligent Transportation Systems</source>, vol. <volume>19</volume>, no. <issue>6</issue>, pp. <fpage>1784</fpage>&#x2013;<lpage>1794</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Zhao</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Wu</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Zeng</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Zhong</surname></string-name></person-group>, &#x201C;<article-title>A vibration-based vehicle classification system using distributed optical sensing technology</article-title>,&#x201D; <source>Transportation Research Record</source>, vol. <volume>2672</volume>, no. <issue>43</issue>, pp. <fpage>12</fpage>&#x2013;<lpage>23</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>E. M.</given-names> <surname>Dogo</surname></string-name>, <string-name><given-names>O. J.</given-names> <surname>Afolabi</surname></string-name>, <string-name><given-names>N. I.</given-names> <surname>Nwulu</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Twala</surname></string-name></person-group>, &#x201C;<article-title>A comparative analysis of gradient descent-based optimization algorithms on convolutional neural networks</article-title>,&#x201D; in <conf-name>2018 Int. Conf. on Computational Techniques, Electronics and Mechanical Systems</conf-name>, <publisher-loc>New York, US</publisher-loc>, <publisher-name>IEEE</publisher-name>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Bottou</surname></string-name></person-group>, &#x201C;<chapter-title>Stochastic gradient descent tricks</chapter-title>,&#x201D; in <source>Neural Networks: Tricks of the Trade</source>, <publisher-loc>Berlin, Germany</publisher-loc>: <publisher-name>Springer</publisher-name>, pp. <fpage>421</fpage>&#x2013;<lpage>436</lpage>, <year>2012</year>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Bottou</surname></string-name></person-group>, &#x201C;<article-title>Large-scale machine learning with stochastic gradient descent</article-title>,&#x201D; in <conf-name>Proc. of COMPSTAT&#x2019;2010</conf-name>, France, pp. <fpage>177</fpage>&#x2013;<lpage>186</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Benard</surname></string-name> and <string-name><given-names>S. Y.</given-names> <surname>Arnob</surname></string-name></person-group>, &#x201C;<article-title>Discriminator-actor-critic: Addressing sample inefficiency and reward bias in adversarial im-itation learning</article-title>,&#x201D; <source>ICLR Reproducibility-Challenge</source>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Ozbulak</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Aytar</surname></string-name> and <string-name><given-names>H. K.</given-names> <surname>Ekenel</surname></string-name></person-group>, &#x201C;<article-title>How transferable are CNN-based features for age and gender classification</article-title>,&#x201D; in <conf-name>Int. Conf. of the Biometrics Special Interest Group</conf-name>, Darmstadt, Germany, <publisher-name>IEEE</publisher-name>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Yann</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Bottou</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Bengio</surname></string-name> and <string-name><given-names>P.</given-names> <surname>Haffner</surname></string-name></person-group>, &#x201C;<article-title>Gradient-based learning applied to document recognition</article-title>,&#x201D; <source>Proc. of the IEEE</source>, vol. <volume>86</volume>, no. <issue>11</issue>, pp. <fpage>2278</fpage>&#x2013;<lpage>2324</lpage>, <year>1998</year>.</mixed-citation></ref>
<ref id="ref-35"><label>[35]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Won</surname></string-name></person-group>, &#x201C;<article-title>Intelligent traffic monitoring systems for vehicle classification: A survey</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>8</volume>, pp. <fpage>73340</fpage>&#x2013;<lpage>73358</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-36"><label>[36]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Dong</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Wu</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Pei</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Jia</surname></string-name></person-group>, &#x201C;<article-title>Vehicle type classification using a semi-supervised convolutional neural network</article-title>,&#x201D; <source>IEEE Transactions on Intelligent Transportation Systems</source>, vol. <volume>16</volume>, no. <issue>4</issue>, pp. <fpage>2247</fpage>&#x2013;<lpage>2256</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-37"><label>[37]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Huttunen</surname></string-name>, <string-name><given-names>F. S.</given-names> <surname>Yancheshmeh</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Chen</surname></string-name></person-group>, &#x201C;<article-title>Car type recognition with deep neural networks</article-title>,&#x201D; in <conf-name>IEEE Intelligent Vehicles Symp.</conf-name>, Gothenburg, Sweden, <publisher-name>IEEE</publisher-name>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-38"><label>[38]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Gyamfi</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Okyere</surname></string-name>, <string-name><given-names>S. K.</given-names> <surname>Asare</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Sharma</surname></string-name> and <string-name><given-names>T.</given-names> <surname>Titus</surname></string-name></person-group>, &#x201C;<article-title>Automated vehicle recognition with deep convolutional neural networks</article-title>,&#x201D; <source>Transportation Research Record</source>, vol. <volume>2645</volume>, no. <issue>1</issue>, pp. <fpage>113</fpage>&#x2013;<lpage>122</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-39"><label>[39]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>N.</given-names> <surname>Audebert</surname></string-name>, <string-name><given-names>B. Le</given-names> <surname>Saux</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Lef&#x00E8;vre</surname></string-name></person-group>, &#x201C;<article-title>Segment-before-detect: Vehicle detection and classification through semantic segmentation of aerial images</article-title>,&#x201D; <source>Remote Sensing</source>, vol. <volume>9</volume>, no. <issue>4</issue>, pp. <fpage>368</fpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-40"><label>[40]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Tan</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Das</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Chaudhry</surname></string-name></person-group>, &#x201C;<article-title>Vehicle detection and classification in aerial imagery</article-title>,&#x201D; in <conf-name>25th IEEE Int. Conf. on Image Processing</conf-name>, Athens, Greece, <publisher-name>IEEE</publisher-name>, <year>2018</year>.</mixed-citation></ref>
</ref-list>
</back>
</article>