<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CSSE</journal-id>
<journal-id journal-id-type="nlm-ta">CSSE</journal-id>
<journal-id journal-id-type="publisher-id">CSSE</journal-id>
<journal-title-group>
<journal-title>Computer Systems Science &#x0026; Engineering</journal-title>
</journal-title-group>
<issn pub-type="ppub">0267-6192</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">25972</article-id>
<article-id pub-id-type="doi">10.32604/csse.2023.025972</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Suicide Ideation Detection of Covid Patients Using Machine Learning Algorithm</article-title><alt-title alt-title-type="left-running-head">Suicide Ideation Detection of Covid Patients Using Machine Learning Algorithm</alt-title><alt-title alt-title-type="right-running-head">Suicide Ideation Detection of Covid Patients Using Machine Learning Algorithm</alt-title>
</title-group>
<contrib-group content-type="authors">
<contrib id="author-1" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Punithavathi</surname><given-names>R.</given-names></name>
<xref ref-type="aff" rid="aff-1">1</xref><email>r.punithavathi@gmail.com</email>
</contrib>
<contrib id="author-2" contrib-type="author">
<name name-style="western"><surname>Thenmozhi</surname><given-names>S.</given-names></name>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Jothilakshmi</surname><given-names>R.</given-names></name>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Ellappan</surname><given-names>V.</given-names></name>
<xref ref-type="aff" rid="aff-4">4</xref>
</contrib>
<contrib id="author-5" contrib-type="author">
<name name-style="western"><surname>Ul</surname><given-names>Islam Md Tahzib</given-names></name>
<xref ref-type="aff" rid="aff-5">5</xref>
</contrib>
<aff id="aff-1"><label>1</label><institution>M Kumarasamy College of Engineering</institution>, <addr-line>Karur, Tamil Nadu</addr-line>, <country>India</country></aff>
<aff id="aff-2"><label>2</label><institution>Dayananda Sagar College of Engineering</institution>, <addr-line>Bangalore, Karnataka</addr-line>, <country>India</country></aff>
<aff id="aff-3"><label>3</label><institution>R.M.D Engineering College</institution>, <addr-line>Kavaraipettai, Tamil Nadu</addr-line>, <country>India</country></aff>
<aff id="aff-4"><label>4</label><institution>Adama Science and Technology University</institution>, <addr-line>Adama</addr-line>, <country>Ethiopia</country></aff>
<aff id="aff-5"><label>5</label><institution>Dhaka International University</institution>, <addr-line>Dhaka</addr-line>, <country>Bangaladesh</country></aff>
</contrib-group><author-notes><corresp id="cor1"><label>&#x002A;</label>Corresponding Author: R. Punithavathi. Email: <email>r.punithavathi@gmail.com</email></corresp></author-notes>
<pub-date pub-type="epub" date-type="pub" iso-8601-date="2022-08-04"><day>04</day>
<month>08</month>
<year>2022</year></pub-date>
<volume>45</volume>
<issue>1</issue>
<fpage>247</fpage>
<lpage>261</lpage>
<history>
<date date-type="received"><day>11</day><month>12</month><year>2021</year></date>
<date date-type="accepted"><day>10</day><month>3</month><year>2022</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2023 Punithavathi et al.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Punithavathi et al.</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CSSE_25972.pdf"></self-uri>
<abstract>
<p>During Covid pandemic, many individuals are suffering from suicidal ideation in the world. Social distancing and quarantining, affects the patient emotionally. Affective computing is the study of recognizing human feelings and emotions. This technology can be used effectively during pandemic for facial expression recognition which automatically extracts the features from the human face. Monitoring system plays a very important role to detect the patient condition and to recognize the patterns of expression from the safest distance. In this paper, a new method is proposed for emotion recognition and suicide ideation detection in COVID patients. This helps to alert the nurse, when patient emotion is fear, cry or sad. The research presented in this paper has introduced Image Processing technology for emotional analysis of patients using Machine learning algorithm. The proposed Convolution Neural Networks (CNN) architecture with DnCNN preprocessing enhances the performance of recognition. The system can analyze the mood of patients either in real time or in the form of video files from CCTV cameras. The proposed method accuracy is more when compared to other methods. It detects the chances of suicide attempt based on stress level and emotional recognition.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>HOG</kwd>
<kwd>ACO-CS</kwd>
<kwd>optimized KNN</kwd>
<kwd>PCA</kwd>
<kwd>emotion detection</kwd>
<kwd>covid</kwd>
<kwd>face recognition</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>In past few years, Coronavirus and its variants have affected the world socially and economically. Emotion recognition systems analyse vocal tone, facial expressions using machine learning and deep learning algorithms. The Emotions such as fear, happiness, sadness, angry and surprise can be detected using the recent methods. The Conventional facial emotion recognition (FER) system consists of stages namely: image acquisition, pre-processing, feature extraction, classification or regression. <xref ref-type="fig" rid="fig-1">Fig. 1</xref> represents a conventional FER system. The first step is pre-processing where the input image obtained from several peoples are processed. The pre-processed data will be send to detection unit. The detection will be made on intensity, smoothing and data augmentation (DA).</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Conventional FER system</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-1.png"/>
</fig>
<p>In second step, pre-processed features will be extracted by Actions Units (AUs). In final step, features are transferred to the classifier. The classification unit consists of Support Machine Vectors (SVMs) or Convolutional Neural Networks (CNNs). The proposed method detects the patient&#x2019;s emotion and suicide ideation based on the stress and emotional variations on Temperamental Dysregulation. The stress level is categorized based on different classes.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Literature Survey</title>
<p>Wang et al. 2018 analyzed the depression patient&#x2019;s videos and facial features are extracted. From the located facial features, the classification is done using a SVM method [<xref ref-type="bibr" rid="ref-1">1</xref>]. With the use of Image processing technology Ketcham et al. 2018 recognized emotions of patients in real time by the use of video files from the CCTV camera [<xref ref-type="bibr" rid="ref-2">2</xref>]. Shojaeilangari et al. 2015 presented an extreme sparse learning method [<xref ref-type="bibr" rid="ref-3">3</xref>] which uses the dictionary and nonlinear classification model. This method gives accurate results even when the input data is affected by noise signals. Compound facial emotion analysis method was presented in literature by the use of iCV-MEFED data set [<xref ref-type="bibr" rid="ref-4">4</xref>]. Yan et al. 2016 described an emotion recognition method from facial expression and sparse kernel reduced-rank regression (SKRRR) fusion technique [<xref ref-type="bibr" rid="ref-5">5</xref>]. Here an openSMILE algorithm is used for extracting features from speech. A scale invariant feature transform for facial analysis is employed. Facial expressions are synthesized by Mumenthaler et al. 2020 [<xref ref-type="bibr" rid="ref-6">6</xref>]. Zhang et al. 2019 developed an algorithm based on deep learning framework named as spatial-temporal recurrent neural network (STRNN) [<xref ref-type="bibr" rid="ref-7">7</xref>] which uses multidirectional recurrent neural network (RNN) for emotion recognition. Chakraborty et al. 2009 made use of external stimulus to excite certain emotions in human [<xref ref-type="bibr" rid="ref-8">8</xref>] such as eye opening, mouth opening, eyebrow length etc., The image was analyzed by segmentation method which divides the image into separate regions of interest. A hybrid cascaded Gaussian mixture model and deep neural network based classifier [<xref ref-type="bibr" rid="ref-9">9</xref>] is presented by Shahin et al. 2019, in which the emotions are contrasted with support vector machine and multilayer perceptron. Tzirakis et al. 2017 described an emotion recognition method with audio and video systems [<xref ref-type="bibr" rid="ref-10">10</xref>]. The convolutional neural network is used for recognizing audio and deep residual network for video recognition. In gaming environment, the player&#x2019;s emotions are detected and recognized with their heart beat. The use of Bidirectional long- and short-term memory (Bi-LSTM) network for heart beat recognition and CNN for facial recognition was done [<xref ref-type="bibr" rid="ref-11">11</xref>]. Further Du et al. 2020 presented a SOM-BP network which combines the heart rate and facial features to detect the player&#x2019;s emotion. Li et al. 2019 [<xref ref-type="bibr" rid="ref-12">12</xref>] presents a maximization algorithm to express compound or mixture of emotions. For this multi-expression analysis, a deep CNN is used. It increases the detection capability. Hossain et al. 2017 described the usage of Bandlet transform to analyze the features in images for facial expression detection [<xref ref-type="bibr" rid="ref-13">13</xref>]. The method is applied on images captured from the camera. For feature selection, Kruskal-Wallis method is used and the dominant bins are sent to Gaussian model for further classification of emotions. A sampling approach method based on human perceptual psychology [<xref ref-type="bibr" rid="ref-14">14</xref>] is proposed by Cruz et al. 2014 which extracts the temporal dynamics of facial emotions. It is tested in Audio/Visual Emotion Challenge event and found that there is considerable increase in accuracy of detection. Wang et al. 2015 developed a tensor independent color space (TICS) [<xref ref-type="bibr" rid="ref-15">15</xref>] for identification of micro-expressions. The Region of interests is identified, and dynamic texture histograms are estimated for every ROI. It is found that CIELab and CIELuv are useful for identification of micro-expressions. Xia et al. 2017 described a Deep Belief network (DBN) for acoustic emotion recognition [<xref ref-type="bibr" rid="ref-16">16</xref>]. Features are extracted and classified with the use of support vector machine. Tariq et al. 2012 addressed an emotion recognition method [<xref ref-type="bibr" rid="ref-17">17</xref>] based on subject dependent and subject-independent emotion recognition. Emotions recognition are person-specific and person-independent and that can be done either manually or automatically. A light weight neutral emotion classification [<xref ref-type="bibr" rid="ref-18">18</xref>] is presented by Chiranjeevi et al. 2015, using statistical texture model. The neutral appearance is identified for emotion recognition. It decreases the computational complexity. Zhang et al. 2016 developed a emotion recognition method [<xref ref-type="bibr" rid="ref-19">19</xref>] for extracting multiscale features using biorthogonal wavelet. For classification, fuzzy multiclass support vector machine is used. Wang et al. 2010 applied four different methods [<xref ref-type="bibr" rid="ref-20">20</xref>] namely Eigen face approach, fisher face approach, Active Appearance model (AAM), and AAM based LDA method on visible and infrared facial expression database for emotion recognition. The relationship between facial temperature and emotion are studied using statistical methods. A new CNN model is presented by Zhang et al. that to integrate the content information of the deep network [<xref ref-type="bibr" rid="ref-21">21</xref>]. An audio-visual emotion recognition based on machine learning is presented by Seng et al. 2018. For visual path BDPCA and LSLDA [<xref ref-type="bibr" rid="ref-22">22</xref>] are used. Prosodic features and spectral features are used to recognize emotions. FER system based on hierarchical deep learning scheme [<xref ref-type="bibr" rid="ref-23">23</xref>] is developed by Kim et al. 2019. Here the features extracted are fused with the geometric features. This method employs autoencoder technique for emotion detection. The emotion recognition was carried out with the absence of sequence data. Wang et al. presents a method to analyze the multimodal spontaneous facial expression database which contains natural visible and infrared facial expressions (NVIE). The method can recognize emotions [<xref ref-type="bibr" rid="ref-24">24</xref>] with respect to their gender and can recognize the existence of multiple emotions/expressions. It is found by Wang et al. that there is a difference in emotion recognition for various environment. Zhang et al. 2019 described a method on the basis on CNN in which the facial image is first normalized. The edge of the faces is extracted to get implicit features [<xref ref-type="bibr" rid="ref-25">25</xref>] by the use of maximum pooling method. Finally, the image is classified by the use of Softmax classifier. Mistry et al. 2017 uses evolutionary particle swarm optimization (PSO) based feature optimization technique [<xref ref-type="bibr" rid="ref-26">26</xref>] for facial emotion recognition. He also presented a novel method of feature extraction using mGA-embedded PSO algorithm. Finally for recognition of the facial expression, different classifiers are employed. Mel&#x00E9;ndez et al. 2020 done a study on variation of emotion of adults due to covid-19 [<xref ref-type="bibr" rid="ref-27">27</xref>]. The study focuses on finding the basic emotions of two groups of adults one confined and other unconfined during covid-19. The results show that there is a high variation in emotions with increasing sadness, depression and reduced happiness.</p>
</sec>
<sec id="s3">
<label>3</label>
<title>Background Methodology</title>
<p>In medical records, it&#x2019;s been identified that the possible propositions for suicide during Covid situation depends on the factors like insufficient integration, stress, frustration, extreme social regulation, enmeshed social regulations, thwarted belongingness, loneliness etc,. <xref ref-type="fig" rid="fig-2">Fig. 2</xref> presents various factors inducing suicide ideations during Covid times. Some of the methods used in clinical diagnosis for suicide ideations detection is shown in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Factors and detection methods for suicide ideation</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-2.png"/>
</fig>
<sec id="s3_1">
<label>3.1</label>
<title>Pre-Processing</title>
<p>The image quality and amount of noise present determines the face recognition rate. But noise can be removed when pre-processing is carried out before extraction process. During preprocessing the image is converted into gray color to obtain a normalized uniform intensity so that color grading process can be carried out. The Effective color grading processes such as cropping is utilized in face detection. After that images are filtered with low pass filter. In this step, unwanted regions in image are eliminated. The Target regions are highlighted using Viola&#x2013;Jones algorithm.</p>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Viola and Jones Based Face Detection</title>
<p>The Viola-Jones algorithm scans the sub window to detect the faces in given input image [<xref ref-type="bibr" rid="ref-28">28</xref>]. In few cases, image rescaling of given input with fixed size is done. Scale invariant detectors are used to rescale the images in minimum time. The Rescaling and feature extraction are done using Haar wavelets.</p>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Feature Extraction and Dimensionality Reduction</title>
<sec id="s3_3_1">
<label>3.3.1</label>
<title>Histograms of Oriented Gradients (HOG)</title>
<p>Navneet Dalal and Bill Triggs (2005) proposed a framed histogram based oriented gradient to extract dense features from all regions present in the image [<xref ref-type="bibr" rid="ref-29">29</xref>]. For object detection, the proposed feature extraction does not use segmentation techniques to extract the features. HOG attempts to capture the pattern by extracting data based on gradients. The pattern divides the cells into 8 &#x00D7; 8 pixels. Each cell consists of gradients to normalize the histogram. This represents 1-D array of histogram known as descriptor.</p>
</sec>
<sec id="s3_3_2">
<label>3.3.2</label>
<title>Dimensionality Reduction</title>
<p>Even though the Data with larger dimension can provide better object description, the recognition speed becomes slow. To speedup the recognition process, the data/dimensionality reduction is done. Dimensionality reduction methods such as PCA or Bag of Words are utilized. In facial recognition application, PCA performs an ortho-normal transformation for converting correlated variables into linearly uncorrelated variables. The Principal components are represented as symmetric covariance matrix with Eigen values. For analysis, ORL image database are utilized. The wavelet transform algorithm has been utilized to reduce noises in the input images.</p>
</sec>
<sec id="s3_3_3">
<label>3.3.3</label>
<title>Principal Component Analysis: PCA</title>
<p>The Principal Component Analysis (PCA) has the capability to reduce the data quantity, but it does not minimize the accuracy [<xref ref-type="bibr" rid="ref-30">30</xref>]. Each sample variance is increased between each covariance. Karkpearson described visual recognition with size reduction for description vector in SIFT and SURF. The Eigen space describes the declaration syntax of input samples with the training samples.</p>
</sec>
</sec>
<sec id="s3_4">
<label>3.4</label>
<title>Classification</title>
<p>In Classification stage, K-Nearest Neighbors (KNN), and Support Vector Machine (SVM) classifiers are utilized.</p>
<sec id="s3_4_1">
<label>3.4.1</label>
<title>K-Nearest Neighbors (KNN)</title>
<p>Each classifier performs its training with random images. The K-Nearest Neighbors classifiers are utilized in machine learning for predicting the test sample. This Classifier encloses a nonparametric technique to classify unknown objects by identifying its closest neighbors [<xref ref-type="bibr" rid="ref-31">31</xref>]. It is most common classifier utilized in pattern recognition for classification. Based on k value, the KNN classifier evaluation gets varies. The Training samples are collected when it&#x2019;s nearest to test sample. In next stage, average operation will be performed on them.</p>
</sec>
<sec id="s3_4_2">
<label>3.4.2</label>
<title>Multi-Class SVM Problem</title>
<p>To solve multi class&#x2019;s analysis, One-vs-One approach, One-vs-All approach and Weston and Watkins&#x2019; approaches are utilized. The classes correspond to different facial expressions. There are 10 facial expression classes in ADFES database, 8 expressions in TFEID database, and 7 expressions in MUG, KDEF, WSEFEP, and JAFFE database. The One-vs-All approach performs correlation of one class with other classes. Class-1 is defined to represent the positive objects while the other &#x201C;not Class_n&#x201D; is used to represent the negative objects. The &#x2018;n&#x2019; represents the expressions number used in the classifier. At last, the classifier selects an effective class for its relevant tested samples.</p>
</sec>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Proposed Emotion Detection Using Deep Learning</title>
<p>Human psychological stress and emotion are very much interconnected. In computational psychology, the relationship between stress and emotions is used to understand the human behavior. A CNN is trained to detect, recognize and classify facial expressions and discrete emotion categories (Anger, Disgust, Neutral, Fear, Sad, Happy and Surprise). Further, logarithmic regression is applied to evaluate stress as a function of deciphered emotions. We performed experiments on Facial Expression Recognition (FER2013) dataset to evaluate our architecture. The feature extraction no longer required for CNN method. The method works well with images. Deep Learning understands the variables, relations and performs extraction according to its characteristics. Sometimes it is complicated to extract high level abstract features from raw data. So that Convolutional Neural Networks (CNN) is utilized for image detection and classification. This is the main advantage of the existing method, but the preprocessing stage should be enhanced.</p>
<sec id="s4_1">
<label>4.1</label>
<title>Proposed Deep learning&#x2019;s Model for Facial Emotion Detection</title>
<p>The CNN extracts the features of the image data sets at different levels. The Outputs of convolutional layer are known as feature maps. The Feature map are based on the number of filters it utilizes. Then it will be passed to non-linear gating function. The proposed depression detection is shown in <xref ref-type="fig" rid="fig-3">Fig. 3</xref>.</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>Architectural diagram for the proposed &#x2018;Depression Detection&#x2019; system</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-3.png"/>
</fig>
<p>Consider xi and yj represents i-th input and j-th output of feature map. <xref ref-type="fig" rid="fig-4">Fig. 4</xref> shows the block diagram of the proposed architecture.</p>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Block diagram of proposed architecture</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-4.png"/>
</fig>
<sec id="s4_1_1">
<label>4.1.1</label>
<title>DnCNN for Preprocessing</title>
<p>The Deep CNN architecture with denoising models based on MLP and CSF are shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>. The original mapping function to predict ECG image is done through residual learning by DCNN. The Convolutional filters dimension and the receptive field of Denoising Convolutional Neural Network (DnCNN) are initialized based on depth value. Proper depth values will optimize the performance and efficiency for larger images. Considering the Gaussian denoising, the receptive field size of DnCNN is initialized to 35 &#x00D7; 35 with the corresponding depth of 17. Based on the depth, the DnCNN has three types of layers as shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>. Batch normalization is added between convolution and ReLU units. The DnCNN works on the residual learning formulation and batch normalization is used to speed up training. This integration of residual learning and batch normalization also increases the denoising performance.</p>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>Proposed DnCNN architecture for image Pre- processing</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-5.png"/>
</fig>
</sec>
<sec id="s4_1_2">
<label>4.1.2</label>
<title>Proposed Deep CNN Structure</title>
<p>This Network has large number of convolutional networks mixed with nonlinear and pooling layers (<xref ref-type="fig" rid="fig-6">Fig. 6</xref>). If image gets passes on one convolution layer, the first layer output acts as input for the second layer. This process gets repeated for entire layers. After completing convolution operations, the output of the convolution layer will be added with nonlinear layer. The Nonlinear layer performs nonlinear properties on obtained result. If nonlinear properties are not involved, the network fails to detect and analysis the response variable. The Pooling layer follows the nonlinear layer. In this layer, the down sampling operations are performed with respect to image dimensions. Finally, image volume is minimized. After the entire process, the convolutional, nonlinear, and pooling layers will be linked with fully connected layer. The Fully connected layer extracts output data from convolutional networks. Fully connected layer ouput is said to be N dimensional vector. In this architecture N &#x003D; 10. Where N denotes the number of classes&#x2019;.</p>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Proposed deep learning architecture for emotion detection and stress level analysis</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-6.png"/>
</fig>
</sec>
<sec id="s4_1_3">
<label>4.1.3</label>
<title>Convolution Layers</title>
<p>The Hidden layer learns the local patterns in small 2D windows. This Layer detects the visual features of images. It also extracts the property of an image at certain point which is used to recognize horizontal edges. Each output result is sum and product of an original image. The Output will be 4 &#x00D7; 4 dimensions. The Input image of 6 &#x00D7; 6 dimensions is applied to 3 &#x00D7; 3 filter at 4 &#x00D7; 4 position. The Matrix evaluates the image at upper left and then moves towards right. If row is progressed, another row will be progressed from below. Multiple filters can be utilized to detect multiple features.</p>
</sec>
<sec id="s4_1_4">
<label>4.1.4</label>
<title>Pooling Layers</title>
<p>The Pooling layers combines with neurons to extract the input data properties. An &#x201C;average pooling&#x201D; extracts the average value from group of neurons. The &#x201C;Max pooling&#x201D; selects the maximum value from each group. In pooling layer when neurons are grouped, entries will be minimized. This Minimization is based on group&#x2019;s parameters dimensions. For example, in <xref ref-type="fig" rid="fig-7">Fig. 7</xref>, a single size is transformed into a group with lesser dimension. Even though there is image transformation, spatial relationship exists within it.</p>
<fig id="fig-7">
<label>Figure 7</label>
<caption>
<title>Pooling example</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-7.png"/>
</fig>
</sec>
<sec id="s4_1_5">
<label>4.1.5</label>
<title>Reducing Overfitting</title>
<p>The Architecture with several layers may lead to overfit the training data. To eliminate those issues, dropout is applied in neural network as represented in <xref ref-type="fig" rid="fig-8">Fig. 8</xref>. The Dropping unit eliminates the neurons and their weight contacts from the network. In Dropout, fixed probability is utilized but dropping choice of units is random. The Forward and back ward propagation networks use dropout of units, So that it leads to thin network.</p>
<fig id="fig-8">
<label>Figure 8</label>
<caption>
<title>Scheme of dropout performance</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-8.png"/>
</fig>
<p>The Deep learning is performed on neural networks. In that, image will be passed into classifier. The Classifier alters the image size to 48 &#x00D7; 48. These Images consists of 3 channels for each color (RGB). The Matrix form for each image is 48 &#x00D7; 48 &#x00D7; 3. When number of neurons are larger evaluation time will increase. To avoid those issues ReLU activation function must be utilized.</p>
</sec>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Results and Discussion</title>
<p>The proposed algorithm for covid patient emotion detection is performed and its results are noted. This experiment is carried out in MATLAB software.</p>
<sec id="s5_1">
<label>5.1</label>
<title>Dataset</title>
<p>The Datasets with different emotions and models are needed to train and test the images. The Data contains 48 &#x00D7; 48 pixel grayscale images [<xref ref-type="bibr" rid="ref-31">31</xref>]. The Database is generated with Google image search API and automatic registration processes are set, so that face gets centred. It maintains same amount of space for each image. Each picture categorized under 7 major categories given in <xref ref-type="fig" rid="fig-9">Fig. 9</xref>.</p>
<fig id="fig-9">
<label>Figure 9</label>
<caption>
<title>FER example</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-9.png"/>
</fig>
<p><bold>Case 1: Train the Model with Lesser Dataset</bold></p>
<p>In Case1 (<xref ref-type="fig" rid="fig-10">Figs. 10</xref>,<xref ref-type="fig" rid="fig-11">11</xref>), FER dataset consists of 7 emotions. Additionally, 3 emotions are included. Angry, cry, disappointed, disgust, fear, happy, laugh, neutral, sad and surprise emotions are evaluated. Before execution, the data has to be organized within a set and labelling has to be finished. <xref ref-type="table" rid="table-1">Tab. 1</xref> lists the 10 facial expressions and labels. <xref ref-type="table" rid="table-2">Tab. 2</xref> lists the number of training and validation dataset utilized under each emotion.</p>
<fig id="fig-10">
<label>Figure 10</label>
<caption>
<title>Number of images in training dataset</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-10.png"/>
</fig>
<fig id="fig-11">
<label>Figure 11</label>
<caption>
<title>Number of images in testing dataset</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-11.png"/>
</fig>
<table-wrap id="table-1"><label>Table 1</label>
<caption>
<title>Ten facial expression categories and the corresponded labels</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th></th>
<th>Angry</th>
<th>Cry</th>
<th>Disappointed</th>
<th>Disgust</th>
<th>Fear</th>
<th>Happy</th>
<th>Laugh</th>
<th>Neutral</th>
<th>Sad</th>
<th>Surprise</th>
</tr>
</thead>
<tbody>
<tr>
<td>Label</td>
<td>0</td>
<td>1</td>
<td>2</td>
<td>3</td>
<td>4</td>
<td>5</td>
<td>6</td>
<td>7</td>
<td>8</td>
<td>9</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap id="table-2"><label>Table 2</label>
<caption>
<title>Number of training and validation images used</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th></th>
<th>Angry</th>
<th>Cry</th>
<th>Disappointed</th>
<th>Disgust</th>
<th>Fear</th>
<th>Happy</th>
<th>Laugh</th>
<th>Neutral</th>
<th>Sad</th>
<th>Surprise</th>
</tr>
</thead>
<tbody>
<tr>
<td>Training</td>
<td>100</td>
<td>100</td>
<td>104</td>
<td>100</td>
<td>100</td>
<td>100</td>
<td>100</td>
<td>100</td>
<td>100</td>
<td>100</td>
</tr>
<tr>
<td>Validation</td>
<td>24</td>
<td>30</td>
<td>30</td>
<td>27</td>
<td>21</td>
<td>100</td>
<td>25</td>
<td>24</td>
<td>33</td>
<td>35</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><bold>Case 2: Train the Model with Larger Dataset</bold></p>
<p>In case 2 (<xref ref-type="fig" rid="fig-12">Figs. 12</xref>,<xref ref-type="fig" rid="fig-13">13</xref>), FER dataset has several images categorized into 7 emotions. <xref ref-type="table" rid="table-3">Tab. 3</xref> lists the number of training and validation Dataset utilized in each emotion.</p>
<fig id="fig-12">
<label>Figure 12</label>
<caption>
<title>Number of images in training dataset</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-12.png"/>
</fig>
<fig id="fig-13">
<label>Figure 13</label>
<caption>
<title>Number of images in test dataset</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-13.png"/>
</fig>
<table-wrap id="table-3"><label>Table 3</label>
<caption>
<title>Number of training and validation images used</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th></th>
<th>Angry</th>
<th>Disgust</th>
<th>Fear</th>
<th>Happy</th>
<th>Neutral</th>
<th>Sad</th>
<th>Surprise</th>
</tr>
</thead>
<tbody>
<tr>
<td>Training</td>
<td>3995</td>
<td>436</td>
<td>4097</td>
<td>7215</td>
<td>4965</td>
<td>4830</td>
<td>3171</td>
</tr>
<tr>
<td>Validation</td>
<td>958</td>
<td>111</td>
<td>1024</td>
<td>1774</td>
<td>1233</td>
<td>1247</td>
<td>831</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s5_2">
<label>5.2</label>
<title>CNN Training Program</title>
<p>The Database must fulfil both quality and quantity (<xref ref-type="fig" rid="fig-10">Fig. 10</xref>). The Quantity related failures are solved in 2 approaches. 1) Reducing the complexity of failures, 2) Increasing the number of datasets. The Learning rate, batch size, epochs count and drop out parameter are initiated with 1e&#x2212;4, 20, 60 and 50%. The experiment results show that the small size input images has utilized minimum timing in training process. <xref ref-type="fig" rid="fig-14">Figs. 14a</xref> and <xref ref-type="fig" rid="fig-15">15a</xref> represents the overall training accuracy results of datasets. Initially training accuracy obtained was 20% but finally it was improved to 1. When training reaches it, the accuracy will get stabilized by minor fluctuations. The Minor fluctuation was the new feature found by the neural network at learning. The Case 2 provides a better good accuracy when compared with case1. <xref ref-type="fig" rid="fig-14">Figs. 14b</xref> and <xref ref-type="fig" rid="fig-15">15b</xref> represents the loss occurred in training phase and these losses are reduced to 0 at final stage. <xref ref-type="table" rid="table-5">Tab. 5</xref> shows the performance of existing and proposed method.</p>
<fig id="fig-14">
<label>Figure 14</label>
<caption>
<title>Case 1 (a) overall training accuracy (b) overall training loss of the datasets</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-14.png"/>
</fig>
<fig id="fig-15">
<label>Figure 15</label>
<caption>
<title>Case 2 (a) overall training accuracy (b) overall training loss of the datasets</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-15.png"/>
</fig>
<table-wrap id="table-4"><label>Table 4</label>
<caption>
<title>Performance analysis</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>S.No.</th>
<th>Image</th>
<th>Stress level (0&#x2013;9)</th>
<th>Emotional recognition</th>
<th>Chances of sucide attempt (%)</th>
</tr>
</thead>
<tbody>
<tr>
<td>1.</td>
<td>Test 1</td>
<td>4</td>
<td>Low Temperamental Dysregulation</td>
<td>35</td>
</tr>
<tr>
<td>2.</td>
<td>Test 2</td>
<td>6</td>
<td>Severely Dysregulated Depressive nature</td>
<td>70</td>
</tr>
<tr>
<td>3.</td>
<td>Test 3</td>
<td>5</td>
<td>Moderate Temperamental Dysregulation</td>
<td>57</td>
</tr>
<tr>
<td>4.</td>
<td>Test 4</td>
<td>7</td>
<td>Severely Dysregulated Depressive and Affective Temperaments nature</td>
<td>75</td>
</tr>
<tr>
<td>5.</td>
<td>Test 5</td>
<td>2</td>
<td>Low Temperamental Dysregulation</td>
<td>10</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap id="table-5"><label>Table 5</label>
<caption>
<title>Performance analysis with existing methods</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th>S. No.</th>
<th>Preprocessing</th>
<th>Classification</th>
<th>Accuracy (%)</th>
</tr>
</thead>
<tbody>
<tr>
<td>1.</td>
<td>Convnets</td>
<td>LSTM</td>
<td>76</td>
</tr>
<tr>
<td>2.</td>
<td>Time and F</td>
<td>SVM</td>
<td>79</td>
</tr>
<tr>
<td>3.</td>
<td>Proposed</td>
<td>CNN</td>
<td>82</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s5_3">
<label>5.3</label>
<title>Clinical Outcomes and Discussion</title>
<p>The investigation shows that the covid affected patients&#x2019; facial emotion processing represents a vulnerability factor for depression. The confusion matrix of emotion detection is shown in <xref ref-type="fig" rid="fig-16">Fig. 16</xref>. The levels of depression can be low, moderate, and severely dysregulated depressive profile. The <xref ref-type="table" rid="table-4">Tab. 4</xref> shows a 5 data from the pool to classify between low, moderate, and severely dysregulated depressive profiles. The suicide ideation can be confirmed through the depressive symptoms. These results are consistent with literature. The affective temperaments with stress level above 5 shown a high depression and suicidal behavior. These behaviors can be strongly related to anxious temperament. The clinical history can be verified that the presence of cyclothymic&#x2013;depressive&#x2013;irritable&#x2013;anxious temperamental constellation pattern in the patients. The suicidal ideation follows from a dysphoric&#x2013;dysregulated depressive profile. It&#x2019;s a challenging task to differentiate the low temperamental and severely dysregulated depressive component. The challenge also lies on the recognition of neutral facial expressions by severe affective temperament dysregulation. Further research is required to accurately classify the neutral facial expressions. Nevertheless, the results of the present study did not cover the interpretation of neutral facial expressions and problems arising due to psychological or social risk factors which also contributes to the development of depressive symptoms and disorders.</p>
<fig id="fig-16">
<label>Figure 16</label>
<caption>
<title>Confusion matrix of emotion detection</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="CSSE_25972-fig-16.png"/>
</fig>
</sec>
</sec>
<sec id="s6">
<label>6</label>
<title>Conclusion</title>
<p>In this paper, the recognition of human feelings using image processing technology has been presented. The COVID patient&#x2019;s emotional disorder which leads to suicide ideation are monitored. The monitoring system based on face detection and emotional analysis are used to identify and solve those disorders in patients and doctors. Here CNN model with DnCNN pre-processing establishes efficient accuracy. This Algorithm is progressed on real time data sets. During covid times large number of individuals is suffering from suicidal ideation in the world due to different issues. The detection through self-expression, emotion release, and personal interaction, will find the suicidal thoughts in them. This work uses facial expression recognition technique that automatically extracts the features on human face to recognize the patterns of suicide ideation. The proposed emotion recognition system monitors the emotion of COVID patients and alerts the nurse, when patient emotion is fear, cry or sad. This research has introduced Image Processing technology for emotional analysis of patients using Machine learning algorithm. Proposed Convolutional Neural Networks (CNN) architecture with DnCNN pre-processing have enhanced the performance. The system can analyze the mood of patients either in real time or in the form of video files from CCTV cameras.</p>
<p>The Hardware for the alert sending is completed using a Raspberry pi hardware unit. The program is done in python and the embedded unit interacts through the GSM modules &#x2013;SIM 900 with the mobile phone. Since this paper presents the Image processing methods and results, the hardware unit is to be presented in a future paper. The work can be extended to detect the infection level in the patients along with the emotion detection</p>
</sec>
</body>
<back><fn-group>
<fn fn-type="other">
<p><bold>Acknowledgement &#x0026; Declarations:</bold> We acknowledge our work place/colleges for giving time for the work.</p>
</fn>
<fn fn-type="other">
<p><bold>Funding Statement:</bold> The authors received no specific funding for this study.</p>
</fn>
<fn fn-type="conflict">
<p><bold>Conflicts of Interest:</bold> The authors declare that they have no conflicts of interest to report regarding the present study.</p>
</fn>
</fn-group>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Q.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Yang</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Yu</surname></string-name></person-group>, &#x201C;<article-title>&#x2018;Facial expression video analysis for depression detection in chinese patients</article-title>,&#x201D; <source>Journal of Visual Communication and Image Representation</source>, vol. <volume>57</volume>, no. <issue>10</issue>, pp. <fpage>228</fpage>&#x2013;<lpage>233</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Ketcham</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Piyaneeranart</surname></string-name> and <string-name><given-names>T.</given-names> <surname>Ganokratanaa</surname></string-name></person-group>, &#x201C;<article-title>Emotional detection of patients major depressive disorder in medical diagnosis</article-title>,&#x201D; in <conf-name>14th Int. Conf. on Signal-Image Technology &#x0026; Internet-Based Systems (SITIS)</conf-name>, <publisher-loc>Las Palmas de Gran Canaria, Spain</publisher-loc>, pp. <fpage>332</fpage>&#x2013;<lpage>338</lpage>, <year>2018</year>. </mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Shojaeilangari</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Yau</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Nandakumar</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>E. K.</given-names> <surname>Teoh</surname></string-name></person-group>, &#x201C;<article-title>Robust representation and recognition of facial emotions using extreme sparse learning</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>24</volume>, no. <issue>7</issue>, pp. <fpage>2140</fpage>&#x2013;<lpage>2152</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Guo</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Lei</surname></string-name>, <string-name><given-names>J.</given-names> <surname>wan</surname></string-name>, <string-name><given-names>E.</given-names> <surname>Avots</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Hajarolasvadi</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Dominant and complementary emotion recognition from still images of faces</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>6</volume>, pp. <fpage>26391</fpage>&#x2013;<lpage>26403</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Yan</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Zheng</surname></string-name>, <string-name><given-names>Q.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Lu</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Li</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Sparse kernel reduced-rank regression for bimodal emotion recognition from facial expression and speech</article-title>,&#x201D; <source>IEEE Transactions on Multimedia</source>, vol. <volume>18</volume>, no. <issue>7</issue>, pp. <fpage>1319</fpage>&#x2013;<lpage>1329</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Mumenthaler</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Sander</surname></string-name> and <string-name><given-names>A. S. R.</given-names> <surname>Manstead</surname></string-name></person-group>, &#x201C;<article-title>Emotion recognition in simulated social interactions</article-title>,&#x201D; <source>IEEE Transactions on Affective Computing</source>, vol. <volume>11</volume>, no. <issue>2</issue>, pp. <fpage>308</fpage>&#x2013;<lpage>312</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>T.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>W.</given-names> <surname>Zheng</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Cui</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zong</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Li</surname></string-name></person-group>, &#x201C;<article-title>Spatial-temporal recurrent neural network for emotion recognition</article-title>,&#x201D; <source>IEEE Transactions on Cybernetics</source>, vol. <volume>49</volume>, no. <issue>3</issue>, pp. <fpage>839</fpage>&#x2013;<lpage>847</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Chakraborty</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Konar</surname></string-name>, <string-name><given-names>U. K.</given-names> <surname>Chakraborty</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Chatterjee</surname></string-name></person-group>, &#x201C;<article-title>Emotion recognition from facial expressions and its control using fuzzy logic</article-title>,&#x201D; <source>IEEE Transactions on Systems, Man, and Cybernetics - Part A: Systems and Humans</source>, vol. <volume>39</volume>, no. <issue>4</issue>, pp. <fpage>726</fpage>&#x2013;<lpage>743</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>I.</given-names> <surname>Shahin</surname></string-name>, <string-name><given-names>A. B.</given-names> <surname>Nassif</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Hamsa</surname></string-name></person-group>, &#x201C;<article-title>Emotion recognition using hybrid gaussian mixture model and deep neural network</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>7</volume>, pp. <fpage>26777</fpage>&#x2013;<lpage>26787</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Tzirakis</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Trigeorgis</surname></string-name>, <string-name><given-names>M. A.</given-names> <surname>Nicolaou</surname></string-name>, <string-name><given-names>B. W.</given-names> <surname>Schuller</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Zafeiriou</surname></string-name></person-group>, &#x201C;<article-title>End-to-End multimodal emotion recognition using deep neural networks</article-title>,&#x201D; <source>IEEE Journal of Selected Topics in Signal Processing</source>, vol. <volume>11</volume>, no. <issue>8</issue>, pp. <fpage>1301</fpage>&#x2013;<lpage>1309</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Du</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Long</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Yuan</surname></string-name></person-group>, &#x201C;<article-title>Non-contact emotion recognition combining heart rate and facial expression for interactive gaming environments</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>8</volume>, pp. <fpage>11896</fpage>&#x2013;<lpage>11906</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>W.</given-names> <surname>Deng</surname></string-name></person-group>, &#x201C;<article-title>Reliable crowdsourcing and deep locality-preserving learning for unconstrained facial expression recognition</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>28</volume>, no. <issue>1</issue>, pp. <fpage>356</fpage>&#x2013;<lpage>370</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. S.</given-names> <surname>Hossain</surname></string-name> and <string-name><given-names>G.</given-names> <surname>Muhammad</surname></string-name></person-group>, &#x201C;<article-title>An emotion recognition system for mobile applications</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>5</volume>, pp. <fpage>2281</fpage>&#x2013;<lpage>2287</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Cruz</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Bhanu</surname></string-name> and <string-name><given-names>N. S.</given-names> <surname>Thakoor</surname></string-name></person-group>, &#x201C;<article-title>Vision and attention theory based sampling for continuous facial emotion recognition</article-title>,&#x201D; <source>IEEE Transactions on Affective Computing</source>, vol. <volume>5</volume>, no. <issue>4</issue>, pp. <fpage>418</fpage>&#x2013;<lpage>431</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Wang</surname></string-name></person-group>, &#x201C;<article-title>Micro-expression recognition using color spaces</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>24</volume>, no. <issue>12</issue>, pp. <fpage>6034</fpage>&#x2013;<lpage>6047</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Xia</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name></person-group>, &#x201C;<article-title>A multi-task learning framework for emotion recognition using 2D continuous space</article-title>,&#x201D; <source>IEEE Transactions on Affective Computing</source>, vol. <volume>8</volume>, no. <issue>1</issue>, pp. <fpage>3</fpage>&#x2013;<lpage>14</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>U.</given-names> <surname>Tariq</surname></string-name>, <string-name><given-names>K. H.</given-names> <surname>Lin</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Zhou</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Wang</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Recognizing emotions from an ensemble of features</article-title>,&#x201D; <source>IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics)</source>, vol. <volume>42</volume>, no. <issue>4</issue>, pp. <fpage>1017</fpage>&#x2013;<lpage>1026</lpage>, <year>2012</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Chiranjeevi</surname></string-name>, <string-name><given-names>V.</given-names> <surname>Gopalakrishnan</surname></string-name> and <string-name><given-names>P.</given-names> <surname>Moogi</surname></string-name></person-group>, &#x201C;<article-title>Neutral face classification using personalized appearance models for fast and robust emotion detection</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>24</volume>, no. <issue>9</issue>, pp. <fpage>2701</fpage>&#x2013;<lpage>2711</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>Z. J.</given-names> <surname>Yang</surname></string-name>, <string-name><given-names>H. M.</given-names> <surname>Lu</surname></string-name>, <string-name><given-names>X. X.</given-names> <surname>Zhou</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Phillips</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Facial emotion recognition based on bi-orthogonal wavelet entropy, fuzzy support vector machine, and stratified cross validation</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>4</volume>, pp. <fpage>8375</fpage>&#x2013;<lpage>8385</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Lv</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Lv</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Wu</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>A natural visible and infrared facial expression database for expression recognition and emotion inference</article-title>,&#x201D; <source>IEEE Transactions on Multimedia</source>, vol. <volume>12</volume>, no. <issue>7</issue>, pp. <fpage>682</fpage>&#x2013;<lpage>691</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>W.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>X.</given-names> <surname>He</surname></string-name> and <string-name><given-names>W.</given-names> <surname>Lu</surname></string-name></person-group>, &#x201C;<article-title>Exploring discriminative representations for image emotion recognition with CNNs</article-title>,&#x201D; <source>IEEE Transactions on Multimedia</source>, vol. <volume>22</volume>, no. <issue>2</issue>, pp. <fpage>515</fpage>&#x2013;<lpage>523</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>K. P.</given-names> <surname>Seng</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Ang</surname></string-name> and <string-name><given-names>C. S.</given-names> <surname>Ooi</surname></string-name></person-group>, &#x201C;<article-title>A combined rule-based &#x0026; machine learning audio-visual emotion recognition approach</article-title>,&#x201D; <source>IEEE Transactions on Affective Computing</source>, vol. <volume>9</volume>, no. <issue>1</issue>, pp. <fpage>3</fpage>&#x2013;<lpage>13</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Kim</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Kim</surname></string-name>, <string-name><given-names>P. P.</given-names> <surname>Roy</surname></string-name> and <string-name><given-names>D.</given-names> <surname>Jeong</surname></string-name></person-group>, &#x201C;<article-title>Efficient facial expression recognition algorithm based on hierarchical deep neural network structure</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>7</volume>, pp. <fpage>41273</fpage>&#x2013;<lpage>41285</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names><surname>Wang</surname></string-name>, <string-name><given-names>Z.</given-names><surname>Liu</surname></string-name>, <string-name><given-names>Z.</given-names><surname>Wang</surname></string-name>, <string-name><given-names>G.</given-names><surname> Wu</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Shen</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Analyses of a multimodal spontaneous facial expression database</article-title>,&#x201D; <source>IEEE Transactions on Affective Computing</source>, vol. <volume>4</volume>, no. <issue>1</issue>, pp. <fpage>34</fpage>&#x2013;<lpage>46</lpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Jolfaei</surname></string-name> and <string-name><given-names>M.</given-names> <surname>Alazab</surname></string-name></person-group>, &#x201C;<article-title>A face emotion recognition method using convolutional neural network and image edge computing</article-title>,&#x201D; <source>IEEE Access</source>, vol. <volume>7</volume>, pp. <fpage>159081</fpage>&#x2013;<lpage>159089</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>K.</given-names> <surname>Mistry</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>S. C.</given-names> <surname>Neoh</surname></string-name>, <string-name><given-names>C. P.</given-names> <surname>Lim</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Fielding</surname></string-name></person-group>, &#x201C;<article-title>A micro-GA embedded PSO feature selection approach to intelligent facial emotion recognition</article-title>,&#x201D; <source>IEEE Transactions on Cybernetics</source>, vol. <volume>47</volume>, no. <issue>6</issue>, pp. <fpage>1496</fpage>&#x2013;<lpage>1509</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J. C.</given-names> <surname>Mel&#x00E9;ndez</surname></string-name>, <string-name><given-names>E.</given-names> <surname>Satorres</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Reyes-Olmedo</surname></string-name>, <string-name><given-names>I.</given-names> <surname>Delhom</surname></string-name>, <string-name><given-names>E.</given-names> <surname>Real</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Emotion recognition changes in a confinement situation due to COVID-19</article-title>,&#x201D; <source>Journal of Environmental Psychology</source>, vol. <volume>72</volume>, pp. 101518, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>W.</given-names> <surname>Lu</surname></string-name> and <string-name><given-names>M.</given-names> <surname>Yang</surname></string-name></person-group>, &#x201C;<article-title>Face detection based on viola-jones algorithm applying composite features</article-title>,&#x201D; in <conf-name>Int. Conf. on Robots &#x0026; Intelligent System (ICRIS)</conf-name>, <publisher-loc>Haikou, China</publisher-loc>, pp. <fpage>82</fpage>&#x2013;<lpage>85</lpage>, <year>2019</year>. </mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>N.</given-names> <surname>Dalal</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Triggs</surname></string-name></person-group>, &#x201C;<article-title>Histograms of oriented gradients for human detection</article-title>,&#x201D; <source>IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR&#x0027;05)</source>, vol. <volume>1</volume>, pp. <fpage>886</fpage>&#x2013;<lpage>893</lpage>, <year>2005</year>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>A. L.</given-names> <surname>Ramadhani</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Musa</surname></string-name> and <string-name><given-names>E. P.</given-names> <surname>Wibowo</surname></string-name></person-group>, &#x201C;<article-title>Human face recognition application using PCA and Eigen face approach</article-title>,&#x201D; in <conf-name>Second Int. Conf. on Informatics and Computing (ICIC)</conf-name>, Jayapura, Indonesia, pp. <fpage>1</fpage>&#x2013;<lpage>5</lpage>, <year>2017</year>. </mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="other"><uri>https://www.kaggle.com/msambare/fer2013</uri>.</mixed-citation></ref>
</ref-list>
</back>
</article>