<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMC</journal-id>
<journal-id journal-id-type="nlm-ta">CMC</journal-id>
<journal-id journal-id-type="publisher-id">CMC</journal-id>
<journal-title-group>
<journal-title>Computers, Materials &#x0026; Continua</journal-title>
</journal-title-group>
<issn pub-type="epub">1546-2226</issn>
<issn pub-type="ppub">1546-2218</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">35313</article-id>
<article-id pub-id-type="doi">10.32604/cmc.2023.035313</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Tackling Faceless Killers: Toxic Comment Detection to Maintain a Healthy Internet Environment</article-title>
<alt-title alt-title-type="left-running-head">Tackling Faceless Killers: Toxic Comment Detection to Maintain a Healthy Internet Environment</alt-title>
<alt-title alt-title-type="right-running-head">Tackling Faceless Killers: Toxic Comment Detection to Maintain a Healthy Internet Environment</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Park</surname><given-names>Semi</given-names></name></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Lee</surname><given-names>Kyungho</given-names></name><email>kevinlee@korea.ac.kr</email></contrib>
<aff id="aff-1"><institution>School of Cybersecurity, Korea University</institution>, <addr-line>Seoul, 02841</addr-line>, <country>Korea</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Kyungho Lee. Email: <email>kevinlee@korea.ac.kr</email></corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic">
<year>2023</year></pub-date>
<pub-date date-type="pub" publication-format="electronic"><day>09</day>
<month>6</month>
<year>2023</year></pub-date>
<volume>76</volume>
<issue>1</issue>
<fpage>813</fpage>
<lpage>826</lpage>
<history>
<date date-type="received"><day>16</day><month>8</month><year>2022</year></date>
<date date-type="accepted"><day>28</day><month>9</month><year>2022</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2023 Park and Lee</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Park and Lee</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMC_35313.pdf"></self-uri>
<abstract>
<p>According to BBC News, online hate speech increased by 20&#x0025; during the COVID-19 pandemic. Hate speech from anonymous users can result in psychological harm, including depression and trauma, and can even lead to suicide. Malicious online comments are increasingly becoming a social and cultural problem. It is therefore critical to detect such comments at the national level and detect malicious users at the corporate level. To achieve a healthy and safe Internet environment, studies should focus on institutional and technical topics. The detection of toxic comments can create a safe online environment. In this study, to detect malicious comments, we used approximately 9,400 examples of hate speech from a Korean corpus of entertainment news comments. We developed toxic comment classification models using supervised learning algorithms, including decision trees, random forest, a support vector machine, and K-nearest neighbors. The proposed model uses random forests to classify toxic words, achieving an F1-score of 0.94. We analyzed the trained model using the permutation feature importance, which is an explanatory machine learning method. Our experimental results confirmed that the toxic comment classifier properly classified hate words used in Korea. Using this research methodology, the proposed method can create a healthy Internet environment by detecting malicious comments written in Korean.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Toxic comments</kwd>
<kwd>toxic text classification</kwd>
<kwd>machine learning</kwd>
<kwd>healthy internet environment</kwd>
</kwd-group>
<funding-group>
<award-group id="awg1">
<funding-source>grant-in-aid of HANWHASYSTEMS</funding-source>
</award-group>
</funding-group>
</article-meta>
</front>
<body>
<sec id="s1"><label>1</label><title>Introduction</title>
<p>Because of the highly contagious coronavirus, the proportion of people working from home has increased 35.4 fold [<xref ref-type="bibr" rid="ref-1">1</xref>]. As their time at home increases, people are spending more time online than offline. The time spent consuming online content has increased by more than 40&#x0025; compared to the pre-COVID-19 period [<xref ref-type="bibr" rid="ref-2">2</xref>]. However, consequently, toxic comments and cyberbullying have also increased. The anonymity of the Internet has enabled the propagation of hate speech and cyberbullying in the form of offensive, inappropriate, and toxic comments. Cyberbullying causes psychological harm such as depression, trauma, and suicidal tendencies [<xref ref-type="bibr" rid="ref-3">3</xref>&#x2013;<xref ref-type="bibr" rid="ref-5">5</xref>]. Therefore, online comments are an important topic of research for creating a healthy Internet environment [<xref ref-type="bibr" rid="ref-6">6</xref>&#x2013;<xref ref-type="bibr" rid="ref-9">9</xref>].</p>
<p>Although hate speech has different legal definitions in different countries, the Cambridge Dictionary defines it as public remarks expressing hate or promoting violence against an individual or group based on race, religion, gender, or sexual orientation. From legal, political, philosophical, and cultural perspectives, hate speech and freedom of expression have some overlapping features. Expressing hate is considered freedom of expression in the United States and some other countries. However, in Germany, the incitement of hatred is punishable under the German Criminal Code. Moreover, in Korea, a person can be punished under criminal law if someone defames, abuses, or insults another individual. Owing to its unique cultural characteristics, all countries have different views on hate speech. Waldron [<xref ref-type="bibr" rid="ref-10">10</xref>] argued that the expression of hate should be regulated because it has severe consequences on the lives, dignity, and reputations of members of minority groups.</p>
<p>Owing to the recent Korean wave, including the drama series Squid Games and the K-pop boy band BTS, the monetary value of exported Korean content (K-content) exceeded more than 10 billion dollars [<xref ref-type="bibr" rid="ref-11">11</xref>]. This Korean wave resulted in the dissemination of quickly translated Korean content to the world, but malicious comments and fake news were also exported. Malicious comments are increasingly becoming a social and cultural concern; therefore, the detection of offensive comments in advance at the national level and malicious users at the corporate level is critical. However, Parekh et al. [<xref ref-type="bibr" rid="ref-12">12</xref>] found that studies on other languages are lacking compared to studies on English. Current technology used to detect malicious comments is focused on the English language [<xref ref-type="bibr" rid="ref-13">13</xref>]. To detect harmful information on the Internet in advance at the national and corporate levels, it is critical to study malicious comments or hate expressions among Korean words. Machine or deep learning methods have been used to detect malicious content in many social media posts or news comments.</p>
<p>In a previous study, a novel approach to automatically classifying toxic comments and preventing them from being posted was proposed [<xref ref-type="bibr" rid="ref-14">14</xref>&#x2013;<xref ref-type="bibr" rid="ref-17">17</xref>]. Machine learning methods have been used in several studies on detecting harmful comments containing profanity, hate speech, toxic speech, and extremism [<xref ref-type="bibr" rid="ref-18">18</xref>&#x2013;<xref ref-type="bibr" rid="ref-20">20</xref>]. Researchers have been using conventional machine learning algorithms, such as decision tree (DT), logistic regression, and support vector machine (SVM) models. Furthermore, neural network algorithms, such as convolutional neural networks (CNN) and long short-term memory (LSTM) have been developed. Through the Kaggle challenge, Google is developing tools for detecting toxic comments and ensuring healthy online conversations, including the development of a multi-label classifier used to detect harmful comments, including threats, obscenities, and profanity [<xref ref-type="bibr" rid="ref-21">21</xref>]. However, no tool can reliably discriminate between a low false-positive rate and high accuracy [<xref ref-type="bibr" rid="ref-18">18</xref>].</p>
<p>Toxic comments are intended to maliciously demean people. This study contributes to research on natural language processing (NLP) in Korean for the classification of toxic comments. In this study, we used a supervised learning-based classifier to detect hateful words in Korean. In addition, using a Korean online news comment corpus, we developed a classifier to determine whether a comment is toxic. We used a random forest algorithm and achieved an F1-score of 0.94. We also used the feature importance to attain model interpretability, thereby contributing to the development of a healthy Internet culture.</p>
<p>The remainder of this paper is organized as follows. In Section 2, related works on the detection of toxic comments are reviewed. The proposed methodology and data preparation process are described in Section 3. Section 4 presents some experimental results. Finally, Section 5 provides some concluding remarks and areas of future research.</p>
</sec>
<sec id="s2"><label>2</label><title>Related Work</title>
<p>Machine and deep learning methods have been used to detect malicious content in many social media posts or news comments. Using a perspective approach, Google&#x2019;s Jigsaw conducted a study on the automatic detection of harmful language on social media platforms using machine learning [<xref ref-type="bibr" rid="ref-21">21</xref>]. A study was conducted to neutralize the toxicity detection system through hostile cases in which the perspective API and the project output are deceived. The perspective score can be lowered by modifying the toxic phrase through the use of space. Duplicating characters and periods can also help lower the toxicity score. The Wikipedia dataset provided by Kaggle during the Toxic Comment Classification Challenge contains 159,571 records for classifying malicious comments. This dataset has a class imbalance, and most of it was written in English because it consists of data collected from English Wikipedia. High accuracy was achieved in classifying the toxicity in this dataset through the use of linear regression (LR), a CNN, an LSTM, a gated recurrent unit, and CNN &#x002B; LSTM [<xref ref-type="bibr" rid="ref-18">18</xref>&#x2013;<xref ref-type="bibr" rid="ref-20">20</xref>].</p>
<p>Yin et al. [<xref ref-type="bibr" rid="ref-14">14</xref>] created a dataset of 24,783 tweets using the Twitter API. They classified these tweets into offensive, clean, and hated classes. Most of the dataset consisted of offensive Twitter messages (tweets). Various detection tools have been developed to automatically identify harmful messages [<xref ref-type="bibr" rid="ref-15">15</xref>,<xref ref-type="bibr" rid="ref-16">16</xref>]. Nobata et al. [<xref ref-type="bibr" rid="ref-16">16</xref>] proposed a methodology for detecting hate speech that is subtler than profanity in various settings by improving the level of knowledge using a deep learning method. Chen et al. [<xref ref-type="bibr" rid="ref-17">17</xref>] proposed a lexical syntactic feature architecture for detecting offensive language on social media and achieved an accuracy of 98.24&#x0025; in detecting sentence attacks and 77.9&#x0025; accuracy in detecting user attacks.</p>
<p>Although many toxicity detection studies have been conducted in English, few studies have focused on the Korean language; therefore, harmful language detection in related Korean datasets is required. A total of 9,400 online news comments in the Korean Hate Speech Dataset were collected and manually labeled [<xref ref-type="bibr" rid="ref-22">22</xref>]. Park et al. [<xref ref-type="bibr" rid="ref-23">23</xref>] developed a dataset of 100,000 people by collecting unlabeled data from public datasets. online communities, and news portal comments. The model was trained using a Bi-LSTM neural network model and generative pre-training, and by efficiently learning an encoder (ELECTRA) that accurately classifies the token replacements, and achieved excellent performance with an F1-score of 0.963 and a mean squared error of 0.029. Park et al. [<xref ref-type="bibr" rid="ref-24">24</xref>] proposed the Korean Text Offensiveness Analysis System (KOAS) to measure profanity and implicit aggression for the analysis of insults in Korean. A total of 46,853 Korean sentences from three domains were collected and classified as positive, neutral, or hostile. The KOAS revealed proficiency detection accuracy of more than 90&#x0025; and an emotion analysis accuracy of more than 80&#x0025;. We propose a machine learning based classifier that understands the characteristics of Korean culture and analyzes Korean characteristics for detecting malicious comments.</p>
</sec>
<sec id="s3"><label>3</label><title>Methodology</title>
<p>As of 2020, Korea has an Internet usage rate of 96.5&#x0025;, and the majority of people use the Internet [<xref ref-type="bibr" rid="ref-25">25</xref>]. Koreans have created diverse online communities to share news and small stories. Although an online community is a place of connection that provides useful information and allows people to comfort each other, it also becomes a platform for disseminating slander, insults, and false information online under the guise of freedom of expression. In Korea, defamation and insults on the Internet are stipulated in the Criminal Act, allowing a complaint to be filed, or an accusation to be leveled against an offending party. Platform operators typically share the personal information of users who have posted malicious comments on the community to assist with an investigation into such cases. Thus, personal information is provided to investigative agencies, and the user churn rate for the platforms increases. As personal information is provided to investigative agencies, platform users will leave the community. To prevent this, operators have introduced a function for automatically expressing swear words in an alternative form when users post a comment, or a way to detect and delete malicious comments. In addition, after some celebrities committed suicide following the posting of malicious comments, Naver and Daum, the largest Korean online platforms, prohibited users from writing comments on entertainment news articles.</p>
<p>This study was conducted to detect toxic comments in Korean, and the detailed methodology is shown in <xref ref-type="fig" rid="fig-1">Fig. 1</xref>. The Korean hate speech dataset (KHSD) was created by collecting Korean news comments. This dataset is the first reliable, manually annotated corpus dataset collected in the Korean language. In this study, only nouns were extracted using the kind Korean morpheme analyzer (KKMA) preprocessing library, and an exploratory data analysis was conducted using a holdout validation. Subsequently, vectorization was conducted using the term frequency-inverse document frequency (TF-IDF) for NLP. A data analysis model was created for evaluating classification algorithms, including K-nearest neighbor (K-NN) and random forest (RF).</p>
<fig id="fig-1"><label>Figure 1</label><caption><title>Flowchart of research methodology</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-1.tif"/></fig>
<sec id="s3_1"><label>3.1</label><title>Dataset</title>
<p>The Korean Hate Speech Dataset is a Korean online news comment corpus that detects prejudice, hate speech, and insults. The criteria for judging malicious comments differ considerably for each person and extracting their harmful elements is difficult. This dataset was created to introduce automation when many insults related to malicious comments were introduced [<xref ref-type="bibr" rid="ref-22">22</xref>].</p>
<p>In the dataset, 32 annotators participated in the pilot study and final tagging processes to investigate whether each comment expressed social prejudice and hatred. Social prejudice was categorized into opinions on gender, other opinions, and none, and hate prejudice was categorized into hate, aggression, or none. This dataset was manually annotated and consisted of 7,896 training data, 471 validation data, and 974 test data. We checked three conditions to classify the data as hateful. First, we checked whether a misrepresentation or insults are content that expresses hostility toward a particular group. Second, we checked whether the content is an expression of hostility toward an individual or a statement that seriously undermines one&#x2019;s social status. Third, we checked whether hate speech or insults are present. We also classified a comment as offensive if it was a sarcastic, cynical, crosstalk, or inhumane remark that offended the target of the comment or any third parties who viewed it.</p>
<p>To create a toxic comment classification model, this study defined all offensive expressions as toxic.</p>
</sec>
<sec id="s3_2"><label>3.2</label><title>Preprocessing</title>
<sec id="s3_2_1"><label>3.2.1</label><title>Korean Language Preprocessing (KoNLPy)</title>
<p>In this study, a different method was used to preprocess English words because the KHSD, which is a Korean corpus, was applied. Because different morphemes and suffixes can be connected to the same word in Korean, if word-based counting used in English or other languages is attempted, it is not treated as the same word. Although the meaning is the same, it is vectorized into a different word, which makes natural language processing in Korean ineffective. For this reason, it must be preprocessed differently from English. Thus, strings containing unnecessary information were removed using regular expressions that are almost similar. In the case of Hangul, the range of consonants is from &#x201C;<inline-graphic xlink:href="CMC_35313-inline-1.tif"/>&#x201D; to &#x201C;<inline-graphic xlink:href="CMC_35313-inline-2.tif"/>,&#x201D; and the range of vowels is from &#x201C;<inline-graphic xlink:href="CMC_35313-inline-3.tif"/>&#x201D; to &#x201C;<inline-graphic xlink:href="CMC_35313-inline-4.tif"/>.&#x201D; To combine these characters and generally leave only Hangul characters, the range of Hangeul is from &#x201C;<inline-graphic xlink:href="CMC_35313-inline-5.tif"/>&#x201D; to &#x201C;<inline-graphic xlink:href="CMC_35313-inline-6.tif"/>.&#x201D; In this study, the rest of the range, other than Hangul characters, was replaced with spaces. Similar to the abbreviation &#x201C;lol&#x201D; in English, Korean also contains shortened phrases consisting of consonants only, such as &#x201C;<inline-graphic xlink:href="CMC_35313-inline-7.tif"/>&#x201D; and &#x201C;<inline-graphic xlink:href="CMC_35313-inline-8.tif"/>.&#x201D; However, to the best of our knowledge, no practical benefit exists, and even if analyzed, these consonants exist in most comments and are all replaced with blanks because no significant use exists in detecting malicious expressions. Thereafter, only nouns were extracted using the KoNLPy tagging class through a Korean preprocessing library called KKMA [<xref ref-type="bibr" rid="ref-26">26</xref>].</p>
<p>The KKMA used in this study works well regardless of spacing errors. KKMA uses dynamic programming to handle situations in which one or more possible morphemes appear from a single syllable. Considering the number of cases in which the length of a word is increased from one syllable, possible combinations of morphemes at a specific length are created and stored in memory. Then, by increasing the length sequentially again, it is possible to consider all possible combinations of morphemes by generating a new result based on the previously stored results. A dictionary-based morpheme combination is used to determine whether a morpheme combination created using dynamic programming is suitable. Part-of-speech combination, phonemic contention, and form combination are example conditions. A probabilistic model is used to select the candidate group that meets such conditions. Through the above method, KKMA generates several candidate morpheme combinations through dynamic programming and an adjacency condition check, and finally separates the stems and endings through a probabilistic model.</p>
<p>KoNLPy is a Python package for representative Korean NLP. The Korean language is the 13th most spoken language in the world, and various tools, such as KoNLPy, have been developed to extract valuable characteristics from texts based on their complexity and subtlety. KoNLPy is open-source software that can adopt five tagging methods internally. Tagging denotes analyzing of morphemes in Korean, which refers to grasping the structure of various linguistic attributes, such as morphemes, roots, prefixes, suffixes, and parts of speech. As shown in <xref ref-type="fig" rid="fig-2">Figs. 2</xref> and <xref ref-type="fig" rid="fig-3">3</xref>, based on the documentation of the KoNLPy project, KKMA exhibits slower loading and execution time compared with other tagging classes.</p>
<fig id="fig-2"><label>Figure 2</label><caption><title>Loading time of KoNLPy tagging class</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-2.tif"/></fig><fig id="fig-3"><label>Figure 3</label><caption><title>Execution time of KoNLPy tagging class</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-3.tif"/></fig>
<p>If the dataset size is large, another tagging class should be considered because of the time problem of the KKMA. Because the dataset used in our experiment is not large, we performed the experiment using the KKMA tagging class. Compared with Hannanum, Komoran, Mecab, and Okt, the tagged results of this study were used to obtain the performance data we wanted in KKMA; therefore, we conducted an experiment using the data shown in <xref ref-type="fig" rid="fig-4">Fig. 4</xref>.</p>
<fig id="fig-4"><label>Figure 4</label><caption><title>Results of KKMA tagging class</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-4.tif"/></fig>
<p>The noun used in KKMA matches the tagging result of another tagging class as displayed in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>.</p>
<fig id="fig-5"><label>Figure 5</label><caption><title>None tagging rule</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-5.tif"/></fig>
</sec>
<sec id="s3_2_2"><label>3.2.2</label><title>Sentence Vectorization (TF-IDF)</title>
<p>TF-IDF is a method of weighing the importance of each word using word and inverse document frequencies [<xref ref-type="bibr" rid="ref-27">27</xref>]. TF-IDF can not only statistically express how often a specific word appears in a document but also compare the weights of the words in the document and express similarity between documents using the cosine similarity method [<xref ref-type="bibr" rid="ref-28">28</xref>,<xref ref-type="bibr" rid="ref-29">29</xref>].
<disp-formula id="eqn-1"><label>(1)</label><mml:math id="mml-eqn-1" display="block"><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mtext>d</mml:mtext></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:mtext>t</mml:mtext></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-2"><label>(2)</label><mml:math id="mml-eqn-2" display="block"><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-3"><label>(3)</label><mml:math id="mml-eqn-3" display="block"><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>d</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>log</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mfrac><mml:mi>n</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>+</mml:mo><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-4"><label>(4)</label><mml:math id="mml-eqn-4" display="block"><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mstyle displaystyle="false" scriptlevel="0"><mml:mtext>&#x2013;</mml:mtext></mml:mstyle><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mo>=</mml:mo><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mtext>d</mml:mtext></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:mtext>t</mml:mtext></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x2217;</mml:mo></mml:mrow><mml:mtext>&#x00A0;</mml:mtext><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>d</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mtext>d</mml:mtext></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:mtext>t</mml:mtext></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x2217;</mml:mo></mml:mrow><mml:mtext>&#x00A0;</mml:mtext><mml:mrow><mml:mtext>log</mml:mtext></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mfrac><mml:mi>n</mml:mi><mml:mrow><mml:mn>1</mml:mn><mml:mo>+</mml:mo><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p><xref ref-type="disp-formula" rid="eqn-1">Eq. (1)</xref> represents the number of occurrences of a specific word t in a specific document <italic>d</italic>; thus, a value representing the frequency of occurrence of each word in each document can be expressed. <xref ref-type="disp-formula" rid="eqn-2">Eq. (2)</xref> represents the number of documents in which a specific word t appears. As the number of documents appearing for a word containing a specific topic is tracked, this number is a critical factor in determining the rarity and importance of a specific word <italic>t</italic>. The <inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi></mml:math></inline-formula> value is inversely proportional to <xref ref-type="disp-formula" rid="eqn-3">Eq. (3)</xref>; therefore, this device prevents the weights from overflowing in the model for technical terms, slang, and misspelled words. When the value of <inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> becomes zero, the value of <inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>d</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> in <xref ref-type="disp-formula" rid="eqn-4">Eq. (4)</xref> can also be ignored. In <xref ref-type="disp-formula" rid="eqn-4">Eq. (4)</xref>, the final value is determined by multiplying the value of <inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>d</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> by the value of <inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>d</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> by reflecting the aforementioned series of processes [<xref ref-type="bibr" rid="ref-30">30</xref>].</p>
<p>Because of the usefulness as discussed, TF-IDF was applied as a baseline model instead of a one-hot encoding method that is discontinuously expressed when conducting NLP classification tasks.</p>
</sec>
</sec>
<sec id="s3_3"><label>3.3</label><title>Data Analysis Model</title>
<p>TF-IDF was used for vectorization, and a classification algorithm was used for classifying online news comments. Representative classification algorithms include decision tree, random forest, support vector machine, and K-NN algorithms, as shown in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>.</p>
<fig id="fig-6"><label>Figure 6</label><caption><title>Decision tree, random forest, SVM, and K-NN</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-6.tif"/></fig>
<p>A DT is a tree that sorts and categorizes objects based on the shape values. Each node represents the shape of the object to be categorized, and each branch represents a possible value for the node. It starts with the root node, and their feature values are arranged into an instance. A DT is commonly used to categorize data in a variety of computational domains. Tree paths or rules are mutually exclusive and complete, which is an interesting and important property of decision trees and rule sets.</p>
<p>RF is a classification algorithm composed of several decision trees and is used to determine the most appropriate classification in a subset. Many trees are created and adjusted to solve the overfitting problem of the decision trees; therefore, they do not considerably affect the prediction. Because each of its nodes reflects the shape of an instance to be classed and each branch indicates a value that the node can assume, a DT sorts and categorizes instances based on the shape values. According to the feature values, instances are sorted, starting with the root node. DTs are primarily used in various computational fields to classify data. DT learning algorithms are widely accepted because they can be applied to numerous problems. The fact that tree routes or rules are mutually exclusive and complete is an exciting and essential property of decision trees and rule sets [<xref ref-type="bibr" rid="ref-31">31</xref>].</p>
<p>SVMs are supervised learning algorithms that can be used for classification, regression, and outlier detection. Because an SVM has a high-dimensional space, a subset of the training points is employed in the decision function for specifying the memory efficiency and final kernel functions for the decision function. Although a standard kernel is provided, a custom kernel can be specified. Most real-world problems involve indivisible data, i.e., data in which no hyperplane separates positive from negative instances in the training dataset. This separability problem can be resolved by mapping the data to a high-dimensional space and defining a split hyperplane. The modified feature space, in contrast to the input space filled in by the training instance, has a large number of dimensions [<xref ref-type="bibr" rid="ref-32">32</xref>].</p>
<p>A K-NN is a final example. When there is little or no prior knowledge of the data distribution, K-NN classification is one of the most extensively used methods for classifying objects. When reliable parameter estimates of the probability density are unknown or difficult to determine, a K-NN is a viable choice for achieving a discriminant analysis. A K-NN is a supervised learning technique that classifies the results of a new instance query using the majority of the k-parameter neighbor categories. The main task of the algorithm is classifying new entities using attributes and training data. In this case, a majority vote among k items is used for the classification.</p>
<p>We compared the performances of the DT, RF, SVM, and K-NN algorithms. Scikit-learn was used to implement the four models, and the default values provided by scikit-learn were applied as the hyperparameters, as shown in <xref ref-type="fig" rid="fig-7">Fig. 7</xref>.</p>
<fig id="fig-7"><label>Figure 7</label><caption><title>Code of RF classifier</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-7.tif"/></fig>
</sec>
<sec id="s3_4"><label>3.4</label><title>Exploratory Data Analysis</title>
<p>The data were constructed using the holdout validation. Although the test set should be used to evaluate the model performance, the test set of the dataset used in our experiments cannot be applied because it is used in Kaggle competitions [<xref ref-type="bibr" rid="ref-22">22</xref>]. Therefore, the validation set was used as the test set for validation and the training set for model training. The dataset consisted of 6,664 normal (false) news comments and 1,232 malicious (true) comments, as shown in <xref ref-type="fig" rid="fig-8">Fig. 8</xref>.</p>
<fig id="fig-8"><label>Figure 8</label><caption><title>Label distribution of dataset</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-8.tif"/></fig>
<p><xref ref-type="fig" rid="fig-9">Fig. 9</xref> shows the distribution of the dataset after the first step of preprocessing, i.e., after string removal using the regular expressions. Most of the comments were within 40 characters.</p>
<fig id="fig-9"><label>Figure 9</label><caption><title>Length distribution of the dataset</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-9.tif"/></fig>
<p>After extracting the noun part, which is the second step of the preprocessing, the distribution by the label was evaluated, as displayed in the graph in <xref ref-type="fig" rid="fig-10">Fig. 10</xref>. For the words in these comments, the difference in the distribution between malicious and normal comments could not be confirmed. The experiment revealed that maliciousness should be classified according to the actual content of the comment rather than the statistical characteristics, including sentence length and number of words.</p>
<fig id="fig-10"><label>Figure 10</label><caption><title>Boxplot grouped by label comments <italic>vs.</italic> number of nouns</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-10.tif"/></fig>
<p>This technique was applied to the test set in the same manner as the method applied in the data preprocessing of the training set. A TF-IDF vectorizer was used to vectorize the dataset. The vectorized results are displayed in <xref ref-type="fig" rid="fig-11">Fig. 11</xref>.</p>
<fig id="fig-11"><label>Figure 11</label><caption><title>Vectorized comments using TF-IDF</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-11.tif"/></fig>
</sec>
</sec>
<sec id="s4"><label>4</label><title>Results</title>
<sec id="s4_1"><label>4.1</label><title>Model Validation</title>
<p>The trained model predicted the test data and evaluated its accuracy. The precision, recall, accuracy, and F1-score were used as representative metrics. We used these metrics shown in <xref ref-type="fig" rid="fig-12">Fig. 12</xref> for model validation. The precision is the ratio of true negatives <italic>vs.</italic> true positives, whereas the recall is the ratio of false positives <italic>vs.</italic> true positives. The accuracy is intuitively related to the model performance. The F1-score is a harmonic average value that considers both precision and recall.</p>
<fig id="fig-12"><label>Figure 12</label><caption><title>Confusion matrix</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-12.tif"/></fig>
<p>These values can be expressed as follows:
<disp-formula id="eqn-5"><label>(5)</label><mml:math id="mml-eqn-5" display="block"><mml:mrow><mml:mtext mathvariant="italic">Precision</mml:mtext></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi>F</mml:mi><mml:mi>P</mml:mi></mml:mrow></mml:mfrac></mml:math></disp-formula>
<disp-formula id="eqn-6"><label>(6)</label><mml:math id="mml-eqn-6" display="block"><mml:mrow><mml:mtext mathvariant="italic">Recall</mml:mtext></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi>F</mml:mi><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:math></disp-formula>
<disp-formula id="eqn-7"><label>(7)</label><mml:math id="mml-eqn-7" display="block"><mml:mrow><mml:mtext mathvariant="italic">Accuracy</mml:mtext></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi>T</mml:mi><mml:mi>N</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi>T</mml:mi><mml:mi>N</mml:mi><mml:mo>+</mml:mo><mml:mi>F</mml:mi><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi>F</mml:mi><mml:mi>N</mml:mi></mml:mrow></mml:mfrac></mml:math></disp-formula>
<disp-formula id="eqn-8"><label>(8)</label><mml:math id="mml-eqn-8" display="block"><mml:mi>F</mml:mi><mml:mn>1</mml:mn><mml:mstyle displaystyle="false" scriptlevel="0"><mml:mtext>&#x2013;</mml:mtext></mml:mstyle><mml:mrow><mml:mtext mathvariant="italic">Score</mml:mtext></mml:mrow><mml:mo>=</mml:mo><mml:mn>2</mml:mn><mml:mrow><mml:mo>&#x2217;</mml:mo></mml:mrow><mml:mspace width="thinmathspace" /><mml:mfrac><mml:mrow><mml:mrow><mml:mtext mathvariant="italic">Precision</mml:mtext></mml:mrow><mml:mo>&#x2217;</mml:mo><mml:mrow><mml:mtext mathvariant="italic">Recall</mml:mtext></mml:mrow></mml:mrow><mml:mrow><mml:mrow><mml:mtext mathvariant="italic">Precision</mml:mtext></mml:mrow><mml:mo>+</mml:mo><mml:mrow><mml:mtext mathvariant="italic">Recall</mml:mtext></mml:mrow></mml:mrow></mml:mfrac></mml:math></disp-formula></p>
<p>We used classification_report of scikit-learn to obtain the results of the four metrics in <xref ref-type="table" rid="table-1">Table 1</xref>. As a results of training the DF, RM, SVM, and K-NN models for predicting the test data, the RF model was shown to be excellent based on all scores.</p>
<table-wrap id="table-1"><label>Table 1</label><caption><title>Evaluation results of different classifiers</title></caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Classifier</th>
<th align="left">Precision</th>
<th align="left">Recall</th>
<th align="left">Accuracy</th>
<th align="left">F1-score</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">Decision tree</td>
<td align="left">0.90</td>
<td align="left">0.90</td>
<td align="left">0.90</td>
<td align="left">0.90</td>
</tr>
<tr>
<td align="left">Random forest</td>
<td align="left">0.94</td>
<td align="left">0.94</td>
<td align="left">0.94</td>
<td align="left">0.94</td>
</tr>
<tr>
<td align="left">SVM</td>
<td align="left">0.92</td>
<td align="left">0.92</td>
<td align="left">0.92</td>
<td align="left">0.91</td>
</tr>
<tr>
<td align="left">K-NN</td>
<td align="left">0.78</td>
<td align="left">0.86</td>
<td align="left">0.86</td>
<td align="left">0.79</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s4_2"><label>4.2</label><title>Model Description</title>
<p>Even if the trained model achieves an excellent performance, evaluating whether an overfitting problem exists is crucial. Therefore, the concepts of explanatory AI (XAI) and permutation feature importance, which can be used to identify which features are important, have emerged [<xref ref-type="bibr" rid="ref-33">33</xref>]. The degree of error of the model is measured by removing each feature and replacing the removed feature with random noise to eliminate the relationship with the resulting value. We can confirm that if there is a difference in error, the feature is important when the model makes a prediction. The importance of the permutation function is calculated as follows:</p>
<p>The inputs are the fitted predictive model m and tabular dataset <italic>D</italic>. The training dataset is input into D in our experiment. The reference score of model <italic>m</italic> on data <italic>D</italic> is computed as <italic>s</italic>. For example, the score is calculated using the accuracy for a classifier or <inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:msup><mml:mi>R</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> for a regressor. For each feature <italic>j</italic> (column in <inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:mi>D</mml:mi></mml:math></inline-formula>) and for each repetition <italic>K</italic> of <inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:mo stretchy="false">(</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x2026;</mml:mo><mml:mo>,</mml:mo><mml:mi>K</mml:mi></mml:math></inline-formula>), a corrupted version of data <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:mrow><mml:mover><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>&#x007E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula> is generated by randomly shuffling the column <italic>j</italic> of dataset <italic>D</italic>. The score <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:msub><mml:mi>s</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is computed for model <italic>m</italic> on the corrupted data, <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:mrow><mml:mover><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>&#x007E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula>. <xref ref-type="disp-formula" rid="eqn-9">Eq. (9)</xref> is the formula for computing importance <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:msub><mml:mi>i</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> for feature <inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:msub><mml:mi>f</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> [<xref ref-type="bibr" rid="ref-34">34</xref>]:
<disp-formula id="eqn-9"><label>(9)</label><mml:math id="mml-eqn-9" display="block"><mml:msub><mml:mi>i</mml:mi><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>K</mml:mi></mml:mfrac><mml:msubsup><mml:mrow><mml:mo>&#x2211;</mml:mo></mml:mrow><mml:mrow><mml:mi>k</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>K</mml:mi></mml:mrow></mml:msubsup><mml:msub><mml:mi>S</mml:mi><mml:mrow><mml:mi>k</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:math></disp-formula></p>
<p>This study confirmed that the RF achieved the best performance among the other approaches. The RF provides interpretability through its feature importance. The results of XAI using permutation feature importance are displayed in <xref ref-type="fig" rid="fig-13">Fig. 13</xref>.</p>
<fig id="fig-13"><label>Figure 13</label><caption><title>Feature importance results</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_35313-fig-13.tif"/></fig>
<p>Here, &#x201C;<inline-graphic xlink:href="CMC_35313-inline-10.tif"/>&#x201D; (pigheaded), &#x201C;<inline-graphic xlink:href="CMC_35313-inline-11.tif"/>&#x201D; (fangirl), &#x201C;<inline-graphic xlink:href="CMC_35313-inline-12.tif"/>&#x201D; (Korean men), and &#x201C;<inline-graphic xlink:href="CMC_35313-inline-13.tif"/>&#x201D; (Korean girl) displayed in <xref ref-type="fig" rid="fig-13">Fig. 13</xref> are representative of hate speech words used in Korea. As a result of feature importance, the learning was conducted well because the corresponding hate expressions appeared as actual weights in the model.</p>

</sec>
</sec>
<sec id="s5"><label>5</label><title>Conclusion</title>
<p>Online trolls claim the right to freedom of expression, make light of hate speech, and post controversial comments on the Internet. Troll comments propagate form individuals to society, and campaigns have been conducted worldwide. The No Hate Comments Day campaign was held with the participation of teenagers from ten countries, including the United States and Australia. The Kind Comments campaign was created to protect youth from cyberbullying and prevent suicides by celebrities resulting from mean comments. Although we did not directly participate in this campaign, we had a small hand in creating a healthy Internet environment. Thus, we studied the detection of offensive comments using machine learning.</p>
<p>In this study, a harmful comment classification model was generated using the KHSD, which is a Korean news comment corpus. Nouns were extracted using KoNLPy that Korean preprocessing library and were subsequently vectorized using TF-IDF. We proposed a toxicity detection model using DT, RF, SVM, and K-NN for classification. As a result of the experiment, the F1-score of the RF algorithm was 0.94, which was the highest performance achieved.</p>
<p>This study obtained the decision criteria in which toxic words were classified by the proposed model through XAI. The feature importance of the trained model was analyzed to prove that hate expressions commonly used in Korea have a high weight in our experiment.</p>
<p>The limitations of our study were the use of only two algorithms and one dataset because reliable Korean comment datasets have not been sufficiently developed. In the future, we plan to apply various classification algorithms to various datasets. Moreover, we expect that researchers will use deep learning to detect new Korean hate words.</p>
</sec>
</body>
<back>
<sec><title>Funding Statement</title>
<p>This study was supported by a grant-in-aid of HANWHASYSTEMS.</p></sec>
<sec sec-type="COI-statement"><title>Conflicts of Interest</title>
<p>The authors declare that they have no conflicts of interest to report regarding the present study.</p></sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="web"><person-group person-group-type="author"><collab>Shiftee</collab></person-group>, &#x201C;<article-title>Shiftee reveals WFH big data analysis of 2020 and 2021</article-title>,&#x201D; <year>2022</year>. [Online]. Available: <ext-link ext-link-type="uri" xlink:href="https://shiftee.io/en/blog/article/shifteeWorkFromHomeBigDataNews">https://shiftee.io/en/blog/article/shifteeWorkFromHomeBigDataNews</ext-link></mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="web"><person-group person-group-type="author"><collab>Korea Research</collab></person-group>, &#x201C;<article-title>[Plan] malicious comments, whether regulation and blocking are the best&#x2013;A study on the perception of comments</article-title>,&#x201D; <year>2021</year>. [Online]. Available: <ext-link ext-link-type="uri" xlink:href="https://hrcopinion.co.kr/archives/17398">https://hrcopinion.co.kr/archives/17398</ext-link></mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Hinduja</surname></string-name> and <string-name><given-names>J. W.</given-names> <surname>Patchin</surname></string-name></person-group>, &#x201C;<article-title>Bullying, cyberbullying, and suicide</article-title>,&#x201D; <source>Archives of Suicide Research</source>, vol. <volume>14</volume>, no. <issue>3</issue>, pp. <fpage>206</fpage>&#x2013;<lpage>221</lpage>, <year>2010</year>; <pub-id pub-id-type="pmid">20658375</pub-id></mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Alhabash</surname></string-name>, <string-name><given-names>A. R.</given-names> <surname>McAlister</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Lou</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Hagerstrom</surname></string-name></person-group>, &#x201C;<article-title>From clicks to behaviors: The mediating effect of intentions to like, share, and comment on the relationship between message evaluations and offline behavioral intentions</article-title>,&#x201D; <source>Journal of Interactive Advertising</source>, vol. <volume>15</volume>, no. <issue>2</issue>, pp. <fpage>82</fpage>&#x2013;<lpage>96</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Hsueh</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Yogeeswaran</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Malinen</surname></string-name></person-group>, &#x201C;<article-title>Leave your comment below: Can biased online comments influence our own prejudicial attitudes and behaviors?</article-title>,&#x201D; <source>Human Communication Research</source>, vol. <volume>41</volume>, no. <issue>4</issue>, pp. <fpage>557</fpage>&#x2013;<lpage>576</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Koutamanis</surname></string-name>, <string-name><given-names>H. G.</given-names> <surname>Vossen</surname></string-name> and <string-name><given-names>P. M.</given-names> <surname>Valkenburg</surname></string-name></person-group>, &#x201C;<article-title>Adolescents&#x2019; comments in social media: Why do adolescents receive negative feedback and who is most at risk?</article-title>,&#x201D; <source>Computers in Human Behavior</source>, vol. <volume>53</volume>, pp. <fpage>486</fpage>&#x2013;<lpage>494</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. J.</given-names> <surname>Lee</surname></string-name> and <string-name><given-names>J. W.</given-names> <surname>Chun</surname></string-name></person-group>, &#x201C;<article-title>Reading others&#x2019; comments and public opinion poll results on social media: Social judgment and spiral of empowerment</article-title>,&#x201D; <source>Computers in Human Behavior</source>, vol. <volume>65</volume>, pp. <fpage>479</fpage>&#x2013;<lpage>487</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Rim</surname></string-name> and <string-name><given-names>D.</given-names> <surname>Song</surname></string-name></person-group>, &#x201C;<article-title>How negative becomes less negative: Understanding the effects of comment valence and response sidedness in social media</article-title>,&#x201D; <source>Journal of Communication</source>, vol. <volume>66</volume>, no. <issue>3</issue>, pp. <fpage>475</fpage>&#x2013;<lpage>495</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>R&#x00F6;sner</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Winter</surname></string-name> and <string-name><given-names>N. C.</given-names> <surname>Kr&#x00E4;mer</surname></string-name></person-group>, &#x201C;<article-title>Dangerous minds? Effects of uncivil online comments on aggressive cognitions, emotions, and behavior</article-title>,&#x201D; <source>Computers in Human Behavior</source>, vol. <volume>58</volume>, pp. <fpage>461</fpage>&#x2013;<lpage>470</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Waldron</surname></string-name></person-group>, &#x201C;<chapter-title>Approaching hate speech</chapter-title>,&#x201D; in <source>The Harm in Hate Speech</source>, <publisher-loc>Cambridge, MA and London, England</publisher-loc>: <publisher-name>Harvard University Press</publisher-name>, pp. <fpage>1</fpage>&#x2013;<lpage>17</lpage>, <year>2012</year>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="web"><person-group person-group-type="author"><collab>The Financial News</collab></person-group>, &#x201C;<article-title>K-content exports exceed 14 trillion won in the &#x2018;Korean Wave&#x2019;</article-title>,&#x201D; <year>2022</year>. [Online]. Available: <ext-link ext-link-type="uri" xlink:href="https://www.fnnews.com/news/202201240917248024">https://www.fnnews.com/news/202201240917248024</ext-link></mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Parekh</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Patel</surname></string-name></person-group>, &#x201C;<article-title>Toxic comment tools: A case study</article-title>,&#x201D; <source>International Journal of Advanced Research Computer Science</source>, vol. <volume>8</volume>, no. <issue>5</issue>, pp. <fpage>964</fpage>&#x2013;<lpage>967</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>J. A.</given-names> <surname>Leite</surname></string-name>, <string-name><given-names>D. F.</given-names> <surname>Silva</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Bontcheva</surname></string-name> and <string-name><given-names>C.</given-names> <surname>Scarton</surname></string-name></person-group>, &#x201C;<article-title>Toxic language detection in social media for Brazilian Portuguese: New dataset and multilingual analysis</article-title>,&#x201D; arXiv preprint arXiv:2010.04543, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>D.</given-names> <surname>Yin</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Xue</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Hong</surname></string-name>, <string-name><given-names>B. D.</given-names> <surname>Davison</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Kontostathis</surname></string-name> <etal>et al.,</etal></person-group> &#x201C;<article-title>Detection of harassment on web 2.0</article-title>,&#x201D; in <conf-name>Proc. Content Analysis in the WEB</conf-name>, <conf-loc>Madrid, Spain</conf-loc>, vol. <volume>2</volume>, pp. <fpage>1</fpage>&#x2013;<lpage>7</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>van Hee</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Jacobs</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Emmery</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Desmet</surname></string-name>, <string-name><given-names>E.</given-names> <surname>Lefever</surname></string-name> <etal>et al.,</etal></person-group> &#x201C;<article-title>Automatic detection of cyberbullying in social media text</article-title>,&#x201D; <source>PLoS One</source>, vol. <volume>13</volume>, no. <issue>10</issue>, pp. <fpage>e0203794</fpage>, <year>2018</year>; <pub-id pub-id-type="pmid">30296299</pub-id></mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Nobata</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Tetreault</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Thomas</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Mehdad</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Chang</surname></string-name></person-group>, &#x201C;<article-title>Abusive language detection in online user content</article-title>,&#x201D; in <conf-name>Proc. 25th Int. Conf. World Wide Web</conf-name>, <conf-loc>Montreal, Canada</conf-loc>, pp. <fpage>145</fpage>&#x2013;<lpage>153</lpage>, <year>2016</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zhou</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Zhu</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Xu</surname></string-name></person-group>, &#x201C;<article-title>Detecting offensive language in social media to protect adolescent online safety</article-title>,&#x201D; in <conf-name>2012 Int. Conf. Privacy, Security, Risk and Trust and 2012 Int. Conf. on Social Computing</conf-name>, <conf-loc>Amsterdam, Netherlands</conf-loc>, pp. <fpage>71</fpage>&#x2013;<lpage>80</lpage>, <year>2012</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Zaheri</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Leath</surname></string-name> and <string-name><given-names>D.</given-names> <surname>Stroud</surname></string-name></person-group>, &#x201C;<article-title>Toxic comment classification</article-title>,&#x201D; <source>SMU Data Science Review</source>, vol. <volume>3</volume>, no. <issue>1</issue>, pp. <fpage>13</fpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>M. A.</given-names> <surname>Saif</surname></string-name>, <string-name><given-names>A. N.</given-names> <surname>Medvedev</surname></string-name>, <string-name><given-names>M. A.</given-names> <surname>Medvedev</surname></string-name> and <string-name><given-names>T.</given-names> <surname>Atanasova</surname></string-name></person-group>, &#x201C;<article-title>Classification of online toxic comments using the logistic regression and neural networks models</article-title>,&#x201D; in <conf-name>AIP Conf. Proc.</conf-name>, <conf-loc>Sozopol, Bulgaria</conf-loc>, vol. <volume>2048</volume>, no. <issue>1</issue>, pp. <fpage>060011</fpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>S. V.</given-names> <surname>Georgakopoulos</surname></string-name>, <string-name><given-names>S. K.</given-names> <surname>Tasoulis</surname></string-name>, <string-name><given-names>A. G.</given-names> <surname>Vrahatis</surname></string-name> and <string-name><given-names>V. P.</given-names> <surname>Plagianakos</surname></string-name></person-group>, &#x201C;<article-title>Convolutional neural networks for toxic comment classification</article-title>,&#x201D; in <conf-name>Proc. 10th Hellenic Conf. Artificial Intelligence</conf-name>, <conf-loc>Patras, Greece</conf-loc>, pp. <fpage>1</fpage>&#x2013;<lpage>6</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Hosseini</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Kannan</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Zhang</surname></string-name> and <string-name><given-names>R.</given-names> <surname>Poovendran</surname></string-name></person-group>, &#x201C;<article-title>Deceiving google&#x2019;s perspective API built for detecting toxic comments</article-title>,&#x201D; arXiv preprint arXiv:1702.08138, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Moon</surname></string-name>, <string-name><given-names>W. I.</given-names> <surname>Cho</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Lee</surname></string-name></person-group>, &#x201C;<article-title>BEEP! Korean corpus of online news comments for toxic speech detection</article-title>,&#x201D; arXiv preprint arXiv:2005.12503, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>J. W.</given-names> <surname>Park</surname></string-name>, <string-name><given-names>Y. Y.</given-names> <surname>Na</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Park</surname></string-name></person-group>, &#x201C;<article-title>A new dataset for Korean toxic comment detection</article-title>,&#x201D; in <conf-name>Proc. Korea Information Processing Society Conf.</conf-name>, <conf-loc>Yeosu, Korea</conf-loc>, pp. <fpage>606</fpage>&#x2013;<lpage>609</lpage>, <year>2021</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>S. H.</given-names> <surname>Park</surname></string-name>, <string-name><given-names>K. M.</given-names> <surname>Kim</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Cho</surname></string-name>, <string-name><given-names>J. J.</given-names> <surname>Park</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Park</surname></string-name> <etal>et al.,</etal></person-group> &#x201C;<article-title>KOAS: Korean text offensiveness analysis system</article-title>,&#x201D; in <conf-name>Proc. 2021 Conf. Empirical Methods in Natural Language Processing: System Demonstrations</conf-name>, <conf-loc>Punta Cana, Dominican Republic</conf-loc>, pp. <fpage>72</fpage>&#x2013;<lpage>78</lpage>, <year>2021</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="web"><person-group person-group-type="author"><collab>International Telecommunication Union</collab></person-group>, &#x201C;<article-title>ITU releases 2021 global and regional ICT estimates</article-title>,&#x201D; <year>2021</year>. [Online]. Available: <ext-link ext-link-type="uri" xlink:href="https://www.itu.int/en/ITU-D/Statistics/Pages/stat/default.aspx">https://www.itu.int/en/ITU-D/Statistics/Pages/stat/default.aspx</ext-link></mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>E. L.</given-names> <surname>Park</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Cho</surname></string-name></person-group>, &#x201C;<article-title>KoNLPy: Korean natural language processing in Python</article-title>,&#x201D; in <conf-name>Annu. Conf. Human and Language Technology</conf-name>, <conf-loc>Gangwon-do, Korea</conf-loc>, pp. <fpage>133</fpage>&#x2013;<lpage>136</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>K. S.</given-names> <surname>Jones</surname></string-name></person-group>, &#x201C;<article-title>A statistical interpretation of term specificity and its application in retrieval</article-title>,&#x201D; <source>Journal of Documentation</source>, vol. <volume>28</volume>, no. <issue>1</issue>, pp. <fpage>11</fpage>&#x2013;<lpage>21</lpage>, <year>1972</year>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H. C.</given-names> <surname>Wu</surname></string-name>, <string-name><given-names>R. W. P.</given-names> <surname>Luk</surname></string-name>, <string-name><given-names>K. F.</given-names> <surname>Wong</surname></string-name> and <string-name><given-names>K. L.</given-names> <surname>Kwok</surname></string-name></person-group>, &#x201C;<article-title>Interpreting TF-IDF term weights as making relevance decisions</article-title>,&#x201D; <source>ACM Transactions on Information Systems (TOIS)</source>, vol. <volume>26</volume>, no. <issue>3</issue>, pp. <fpage>1</fpage>&#x2013;<lpage>37</lpage>, <year>2008</year>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Tata</surname></string-name> and <string-name><given-names>J. M.</given-names> <surname>Patel</surname></string-name></person-group>, &#x201C;<article-title>Estimating the selectivity of TF-IDF based cosine similarity predicates</article-title>,&#x201D; <source>ACM Sigmod Record</source>, vol. <volume>36</volume>, no. <issue>2</issue>, pp. <fpage>7</fpage>&#x2013;<lpage>12</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Vajjala</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Majumder</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Gupta</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Surana</surname></string-name></person-group>, &#x201C;<article-title>Text representation</article-title>,&#x201D; in <conf-name>Practical Natural Language Processing: A Comprehensive Guide to Building Real-World NLP Systems</conf-name>, <conf-loc>Sebastopol, CA, United States of America, O&#x2019;Reilly Media</conf-loc>, pp. <fpage>90</fpage>&#x2013;<lpage>92</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S. B.</given-names> <surname>Kotsiantis</surname></string-name>, <string-name><given-names>I.</given-names> <surname>Zaharakis</surname></string-name> and <string-name><given-names>P.</given-names> <surname>Pintelas</surname></string-name></person-group>, &#x201C;<article-title>Supervised machine learning: A review of classification techniques</article-title>,&#x201D; in <source>Emerging Artificial Intelligence Applications in Computer Engineering</source>, <conf-loc>Amsterdam, Netherlands, IOS Press</conf-loc>, vol. <volume>160</volume>, no. <issue>1</issue>, pp. <fpage>3</fpage>&#x2013;<lpage>24</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>I.</given-names> <surname>Muhammad</surname></string-name> and <string-name><given-names>Z.</given-names> <surname>Yan</surname></string-name></person-group>, &#x201C;<article-title>Supervised machine learning approaches: A survey</article-title>,&#x201D; <source>ICTACT Journal on Soft Computing</source>, vol. <volume>5</volume>, no. <issue>3</issue>, pp. <fpage>946</fpage>&#x2013;<lpage>952</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Altmann</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Tolo&#x015F;i</surname></string-name>, <string-name><given-names>O.</given-names> <surname>Sander</surname></string-name> and <string-name><given-names>T.</given-names> <surname>Lengauer</surname></string-name></person-group>, &#x201C;<article-title>Permutation importance: A corrected feature importance measure</article-title>,&#x201D; <source>Bioinformatics</source>, vol. <volume>26</volume>, no. <issue>10</issue>, pp. <fpage>1340</fpage>&#x2013;<lpage>1347</lpage>, <year>2010</year>; <pub-id pub-id-type="pmid">20385727</pub-id></mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="web"><person-group person-group-type="author"><collab>Scikit-learn</collab></person-group>, &#x201C;<article-title>4.2. Permutation feature importance</article-title>,&#x201D; <year>2022</year>. [Online]. Available: <ext-link ext-link-type="uri" xlink:href="https://scikit-learn.org/stable/modules/permutation_importance.html#permutation-feature-importance">https://scikit-learn.org/stable/modules/permutation_importance.html#permutation-feature-importance</ext-link></mixed-citation></ref>
</ref-list>
</back>
</article>