<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="review-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMES</journal-id>
<journal-id journal-id-type="nlm-ta">CMES</journal-id>
<journal-id journal-id-type="publisher-id">CMES</journal-id>
<journal-title-group>
<journal-title>Computer Modeling in Engineering &#x0026; Sciences</journal-title>
</journal-title-group>
<issn pub-type="epub">1526-1506</issn>
<issn pub-type="ppub">1526-1492</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">31513</article-id>
<article-id pub-id-type="doi">10.32604/cmes.2023.031513</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Review</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>A Survey of Knowledge Graph Construction Using Machine Learning</article-title>
<alt-title alt-title-type="left-running-head">A Survey of Knowledge Graph Construction Using Machine Learning</alt-title>
<alt-title alt-title-type="right-running-head">A Survey of Knowledge Graph Construction Using Machine Learning</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Zhao</surname><given-names>Zhigang</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Luo</surname><given-names>Xiong</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><xref ref-type="aff" rid="aff-2">2</xref><xref ref-type="aff" rid="aff-3">3</xref><email>xluo@ustb.edu.cn</email></contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Chen</surname><given-names>Maojian</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><xref ref-type="aff" rid="aff-2">2</xref><xref ref-type="aff" rid="aff-3">3</xref></contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Ma</surname><given-names>Ling</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<aff id="aff-1"><label>1</label><institution>School of Computer and Communication Engineering, University of Science and Technology Beijing</institution>, <addr-line>Beijing, 100083</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>Shunde Innovation School, University of Science and Technology Beijing</institution>, <addr-line>Foshan, 528399</addr-line>, <country>China</country></aff>
<aff id="aff-3"><label>3</label><institution>Beijing Key Laboratory of Knowledge Engineering for Materials Science</institution>, <addr-line>Beijing, 100083</addr-line>, <country>China</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Xiong Luo. Email: <email>xluo@ustb.edu.cn</email></corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic">
<year>2023</year></pub-date>
<pub-date date-type="pub" publication-format="electronic"><day>30</day>
<month>12</month>
<year>2023</year></pub-date>
<volume>139</volume>
<issue>1</issue>
<fpage>225</fpage>
<lpage>257</lpage>
<history>
<date date-type="received">
<day>25</day>
<month>6</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>13</day>
<month>9</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2024 Zhao et al.</copyright-statement>
<copyright-year>2024</copyright-year>
<copyright-holder>Zhao et al.</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMES_31513.pdf"></self-uri>
<abstract>
<p>Knowledge graph (KG) serves as a specialized semantic network that encapsulates intricate relationships among real-world entities within a structured framework. This framework facilitates a transformation in information retrieval, transitioning it from mere string matching to far more sophisticated entity matching. In this transformative process, the advancement of artificial intelligence and intelligent information services is invigorated. Meanwhile, the role of machine learning method in the construction of KG is important, and these techniques have already achieved initial success. This article embarks on a comprehensive journey through the last strides in the field of KG via machine learning. With a profound amalgamation of cutting-edge research in machine learning, this article undertakes a systematical exploration of KG construction methods in three distinct phases: entity learning, ontology learning, and knowledge reasoning. Especially, a meticulous dissection of machine learning-driven algorithms is conducted, spotlighting their contributions to critical facets such as entity extraction, relation extraction, entity linking, and link prediction. Moreover, this article also provides an analysis of the unresolved challenges and emerging trajectories that beckon within the expansive application of machine learning-fueled, large-scale KG construction.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Knowledge graph (KG)</kwd>
<kwd>semantic network</kwd>
<kwd>relation extraction</kwd>
<kwd>entity linking</kwd>
<kwd>knowledge reasoning</kwd>
</kwd-group>
<funding-group>
<award-group id="awg1">
<funding-source>Beijing Natural Science Foundation</funding-source>
<award-id>L211020</award-id>
<award-id>M21032</award-id>
</award-group>
<award-group id="awg2">
<funding-source>National Natural Science Foundation of China</funding-source>
<award-id>U1836106</award-id>
<award-id>62271045</award-id>
</award-group>
<award-group id="awg3">
<funding-source>Scientific and Technological Innovation Foundation of Foshan</funding-source>
<award-id>BK21BF001</award-id>
<award-id>BK20BF010</award-id>
</award-group>
</funding-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>The continuous development of information technologies brings significant convenience to human life, paralleled by an exponential surge in information proliferation. Massive amounts of data are collected and studied in many fields such as social network, biomedical engineering, security science, and many others. Under this background, search engine has become an indispensable instrument, facilitating people&#x2019;s quest for knowledge and information online. Traditionally, a search engine has a user input a query term, whereupon it furnishes hyperlinks directing to the most relevant web pages corresponding to the provided keyword [<xref ref-type="bibr" rid="ref-1">1</xref>].</p>
<p>In May 2012, the emergence of knowledge graph (KG) brought a novel paradigm for enhancing search engines. Within this framework, user search results transcend the realm of single web page links, encompassing instead a tapestry of structured entity information closely related to the search query. This transformative approach even delves into the realm of potential hidden knowledge within the KG. The intelligent optimization of search answers through KG can effectively improve the functions of future search engines in three aspects: refining responses, nurturing interactive dialogues, and bolstering predictive capabilities [<xref ref-type="bibr" rid="ref-2">2</xref>]. This multifaceted augmentation leads into an era of heightened search engine functionality. Furthermore, the scope of KG&#x2019;s influence extends considerably into domains beyond search, encompassing intelligent question answering, knowledge engineering, data mining, and digital library.</p>
<p>Recent years have witnessed a remarkable surge of interest from various disciplines engineering and science. As depicted in <xref ref-type="fig" rid="fig-1">Fig. 1</xref> and supported by data from the Web of Science, the number of published papers with &#x201C;Knowledge Graph&#x201D; in their title has exhibited a steady rise up until the end of 2022. Additionally, <xref ref-type="fig" rid="fig-2">Figs. 2</xref>&#x2013;<xref ref-type="fig" rid="fig-4">4</xref> provide insight into the volume of published papers across diverse research areas, publication resources, and institutions. In the field of academic research, Computer Science has emerged as a primary hub for the propagation of KG, closely followed by Mathematics and Engineering. This observation highlights the substantial contribution and enthusiasm originating from the Computer Science towards the exploration, development, and advancement of KG-related topics. Notably, the significant research output in Mathematics and Engineering emphasizes the interdisciplinary essence of KGs, signifying their influence across domain that extend beyond the realm of computer. Shifting the focus to the publication platforms, it is evident that <italic>Lecture Notes in Computer Science</italic> has emerged as the leading platform for KG-centric research, closely followed by <italic>Lecture Notes in Artificial Intelligence</italic> and <italic>IEEE Access</italic>. This indicates that <italic>Lecture Notes in Computer Science</italic> has been the preferred choice for researchers to share their findings and advancements in the field of KGs. Lastly, the Chinese Academy of Sciences, University of Chinese Academy of Science, and Rluk Research Libraries UK have emerged as the leading contributors in terms of publishing papers related to KG. These institutions have demonstrated a robust presence and active participation in KG research, highlighting their expertise and dedicated contributions to the advancement of this field. Their significant published works attest to the valuable role played by these institutions in the exploration and evolution of KG-related topics.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Number of published papers with &#x201C;Knowledge Graph&#x201D; in the title in recent years</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-1.tif"/>
</fig><fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>The top 10 research areas ranked by the number of publication related to KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-2.tif"/>
</fig><fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>The top 10 publication resources ranked by the number of publication in KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-3.tif"/>
</fig><fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>The top 10 institutions ranked by the number of published papers related to KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-4.tif"/>
</fig>
<p>Generally speaking, the establishment and utilization of large-scale KG necessitate the synergy of diverse intelligent information processing technologies. In recent years, machine learning methods have as a pivotal force within the KG construction. Therefore, this article conducts a survey in relation to this field. We first introduce the development history, fundamental principles and technical framework of KG. Second, we summarize the machine learning-based key technologies integral to the actualization, dissected across three pivotal dimensions. Finally, we expound upon the current challenges and future forthcoming trends poised to guide the construction of large-scale KG. The main contributions are as follows:
<list list-type="bullet">
<list-item>
<p>This article offers an extensive and up-to-date review of existing research and literature in the field of KG construction, specifically focusing on methodologies driven by machine learning techniques. Meanwhile, it provides a structured categorization and classification of diverse machine learning-driven approaches utilized for constructing KGs. This categorization could help readers understand the landscape and taxonomy of methods in this field.</p></list-item>
<list-item>
<p>This article delves into the various machine learning methodologies employed in the construction of KGs. Meanwhile, we conduct the comparative evaluation and analysis of different machine learning techniques, showcasing their respective performance, scalability, and suitability under various conditions.</p></list-item>
<list-item>
<p>This article identifies and discusses challenges and open research questions within the domain of KG construction using machine learning, highlighting potential avenues for further exploration and innovation.</p></list-item>
</list></p>
<p>The organization of this article is arranged as follows. In <xref ref-type="sec" rid="s2">Section 2</xref>, we present some fundamental concepts and traditional technical architecture of KG. <xref ref-type="sec" rid="s3">Section 3</xref> focuses on a comprehensive exploration of KG design propelled by the prowess of machine learning methods, meticulously partitioned into three parts, i.e., entity learning, ontology learning, and knowledge reasoning. In <xref ref-type="sec" rid="s4">Section 4</xref>, we discuss the prospective research directions and challenges of large-scale KG construction technologies. Through machine learning methods, we dissect this discussion into three distinct segments, relation extraction, link prediction, and construction of industrial KG. <xref ref-type="sec" rid="s5">Section 5</xref> provides a conclusion and reflections.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Knowledge Graph</title>
<p>In this section, we simply introduce the essence and the foundational framework of KG.</p>
<sec id="s2_1">
<label>2.1</label>
<title>The Development of Knowledge Graph</title>
<p>With the development of the Internet, Web technology has gone through the &#x201C;Web 1.0&#x201D; era characterized by the web of documents and the &#x201C;Web 2.0&#x201D; era characterized by the web of data. Today, the trajectory points towards the &#x201C;Web 3.0&#x201D; era characterized by the web of knowledge [<xref ref-type="bibr" rid="ref-3">3</xref>] and even anticipates the &#x201C;Web 4.0&#x201D; era defined by the Metaverse paradigm [<xref ref-type="bibr" rid="ref-4">4</xref>]. Driven by the continuous growth of user-generated content and open-linked data on the Internet, the quest for knowledge interconnected aligning with the ever-evolving network information resources becomes imperative. This quest takes a fresh perspective in accordance with the principles of knowledge organization in the big data environment, aiming to reveal deeper cognitive insights [<xref ref-type="bibr" rid="ref-5">5</xref>]. In the midst of this dynamic context, Google introduced KG in May 2012. Its goal is to enhance search outcomes, describe the various entities and concepts inherent to the real world, and illuminate their relationships. By these merits, KG emerges as a substantial stride forward from prevailing semantic web technologies. Illustrated in <xref ref-type="fig" rid="fig-5">Fig. 5</xref> are pivotal milestones making the history of KG across different years. For instance, conception of the semantic network as a vehicle for knowledge representation was proposed in 1960. Furthermore, the philosophical concept of &#x201C;ontology&#x201D; was integrated into the KG in 1980, facilitating the structured and formalized description of knowledge.</p>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>Milestones in the development of KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-5.tif"/>
</fig>
<p>The origin of designing KG comes from a series of practical applications, spanning fields such as semantic search, machine question answering, information retrieval, online learning and others. With the exploration of KG advances, various structured KGs have been developed by both academic researchers and industry practitioners. Currently, a tableau of prominent and expansive large-scale open knowledge bases associated with KG exists globally, as enumerated in <xref ref-type="table" rid="table-1">Table 1</xref>. Here, large-scale knowledge bases like Freebase [<xref ref-type="bibr" rid="ref-6">6</xref>], DBPedia [<xref ref-type="bibr" rid="ref-7">7</xref>], and Wikidata [<xref ref-type="bibr" rid="ref-8">8</xref>] take center stage using Wikipedia as a foundational source. Notably, Freebase differentiates itself by its user-generated content, open accessibility, and structured data, which supports all its entries.</p>
<table-wrap id="table-1">
<label>Table 1</label>
<caption>
<title>The sizes of some prominent large-scale knowledge bases</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Knowledge graph</th>
<th>Number of entities</th>
<th>Number of relation types</th>
<th>Number of facts</th>
</tr>
</thead>
<tbody>
<tr>
<td>Freebase [<xref ref-type="bibr" rid="ref-6">6</xref>]</td>
<td>40 M</td>
<td>35000</td>
<td>637 M</td>
</tr>
<tr>
<td>DBpedia [<xref ref-type="bibr" rid="ref-7">7</xref>]</td>
<td>5 M</td>
<td>1367</td>
<td>538 M</td>
</tr>
<tr>
<td>Wikidata [<xref ref-type="bibr" rid="ref-8">8</xref>]</td>
<td>18 M</td>
<td>1632</td>
<td>66 M</td>
</tr>
<tr>
<td>YAGO2 [<xref ref-type="bibr" rid="ref-9">9</xref>]</td>
<td>10 M</td>
<td>114</td>
<td>447 M</td>
</tr>
<tr>
<td>Google KG [<xref ref-type="bibr" rid="ref-7">7</xref>]</td>
<td>570 M</td>
<td>35000</td>
<td>18000 M</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In recent years, an array of research results on the Chinese KG has grown vigorously. For example, Sogou established &#x201C;Knowledge Cube&#x201D;, marking the inception of the foremost knowledge base search product in the domestic search engine industry. Through effectively integrating fragmented Internet knowledge, Baidu founded &#x201C;Baidu Zhixin&#x201D; and brought forth a next-generation search engine product. Contributions extend to academia as well, Tsinghua University built &#x201C;XLore&#x201D;, a pioneering large-scale Chinese-English cross-language KG. The Institute of Computing Technology of Chinese Academy of Sciences established a prototype system termed &#x201C;People Cube, Work Cube, Knowledge Cube&#x201D; based on an open knowledge network OpenKN. Shanghai Jiao Tong University designed &#x2018;Zhishi.me&#x2019;, a dedicated research platform for Chinese KG. Additionally, the GDM Lab at Fudan University launched the Chinese KG project. Generally, these products and projects have given rise to expansive knowledge bases spanning diverse fields, providing users with intelligent search and question-and-answer services.</p>
</sec>
<sec id="s2_2">
<label>2.2</label>
<title>The Definition of Knowledge Graph</title>
<p>KG is a special semantic network composed of nodes and directed edges, and it is also known as a heterogeneous information network or semantic knowledge base. In the KG, each node represents an entity in the real world, while directed edges interlink these nodes to denote the intricate relationships between those entities. Facts are generally represented in the form of triples (<inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:mtext mathvariant="italic">subject</mml:mtext></mml:math></inline-formula>, <inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:mtext mathvariant="italic">predicate</mml:mtext></mml:math></inline-formula>, <inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:mtext mathvariant="italic">object</mml:mtext></mml:math></inline-formula>) (SPO), where <inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:mtext mathvariant="italic">subject</mml:mtext></mml:math></inline-formula> and <inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:mtext mathvariant="italic">object</mml:mtext></mml:math></inline-formula> signify entities, and <inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:mtext mathvariant="italic">predicate</mml:mtext></mml:math></inline-formula> represents the relation between them [<xref ref-type="bibr" rid="ref-10">10</xref>]. For example, the textual data &#x201C;Chao Deng is an actor who played the character Tailang Xu in the comedy movie Duckweed&#x201D; can be expressed via the following set of SPO triples exemplified in <xref ref-type="table" rid="table-2">Table 2</xref>. The transformed version of <xref ref-type="table" rid="table-2">Table 2</xref> into a KG is described in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>. This KG encapsulates the interrelationships between various entities, allowing for a structured representation of the given information.</p>
<table-wrap id="table-2">
<label>Table 2</label>
<caption>
<title>An example of SPO triples extracted from text data</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Subject</th>
<th>Predicate</th>
<th>Object</th>
</tr>
</thead>
<tbody>
<tr>
<td>Chao Deng</td>
<td>Occupation</td>
<td>Actor</td>
</tr>
<tr>
<td>Chao Deng</td>
<td>Starred in</td>
<td>Duckweed</td>
</tr>
<tr>
<td>Chao Deng</td>
<td>Played</td>
<td>Tailang Xu</td>
</tr>
<tr>
<td>Tailang Xu</td>
<td>Character in</td>
<td>Duckweed</td>
</tr>
<tr>
<td>Duckweed</td>
<td>Genre</td>
<td>Comedy movie</td>
</tr>
</tbody>
</table>
</table-wrap><fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>The updated version of <xref ref-type="table" rid="table-2">Table 2</xref> after translating it into a KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-6.tif"/>
</fig>
</sec>
<sec id="s2_3">
<label>2.3</label>
<title>The Technical Architecture of Knowledge Graph</title>
<p>The structure of KG includes two aspects: a logical structure and a technical structure. The former includes the data layer and the pattern layer, while the latter refers to the technological process involved in KG construction. This involves a sequence of stages, including data acquisition, entity learning, ontology learning, knowledge reasoning, and knowledge update.</p>
<p>The data layer, alternatively known as the entity layer, functions as a repository for knowledge housing information in the form of facts. These facts are succinctly conveyed through triples as the basic expression of facts, where a graph database is chosen as a storage medium.</p>
<p>Above the data layer, the pattern layer, commonly referred to as the ontology layer, represents and stores refined concepts and knowledge. The pattern layer leverages ontology constructs, effectively serving as an embodiment of the KG. It plays an important role in defining and organizing entities, properties, classes, and relationships within a KG. Typically, ontology is represented using formal languages like the Web Ontology Language (OWL) [<xref ref-type="bibr" rid="ref-11">11</xref>], which is grounded in description logic. OWL&#x2019;s expressive capabilities empower ontologies to accurately define semantic relationships among entities, properties, classes, and relationships. Moreover, ontology-based reasoning facilitates the identification of missing information in the KG, the discovery of hidden relationships between entities, and the ability to address intricate semantic queries. Furthermore, ontology enables seamless interaction and knowledge sharing among different KGs. This fosters the construction of larger, more comprehensive KG, and promotes the reuse of knowledge, thereby augmenting the overall value of knowledge graph. Generally, the axioms, rules, and constraints in the ontology base are used to standardize the entities, the types and the attributes of entities, and the relationship between the entities, so that the KG has a strong structure and less redundancy [<xref ref-type="bibr" rid="ref-12">12</xref>].</p>
<p>The technical architecture of KG construction can be classified into two principal paradigms: top-down and bottom-up. In the former, the pattern layer is first defined, and the construction begins from the top-level concept. It then proceeds to progressively refine, layer, and generate instances downward. The process is depicted in <xref ref-type="fig" rid="fig-7">Fig. 7</xref>. Conversely, the latter starts from the underlying entities, extracts entities, and gradually abstracts them upwards to form upper-level concepts and knowledge. This architectural construct is illustrated in <xref ref-type="fig" rid="fig-8">Fig. 8</xref>. It starts from the original semi-structured data or unstructured data and adopts a series of technologies to extract knowledge. Then, it integrates with the structural data, with the ontology layer contributing to the enrichment of upper-level concepts. Finally, a complete KG is generated. Furthermore, the comprehensive KG is continuously updated and augmented. New knowledge is continuously extracted to promote the KG refinement. Within this framework, each iteration generally includes data acquisition, entity learning, ontology learning, knowledge reasoning, and knowledge updating [<xref ref-type="bibr" rid="ref-13">13</xref>].</p>
<fig id="fig-7">
<label>Figure 7</label>
<caption>
<title>The technical architecture of the top-down KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-7.tif"/>
</fig><fig id="fig-8">
<label>Figure 8</label>
<caption>
<title>The technical architecture of bottom-up KG</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-8.tif"/>
</fig>
</sec>
</sec>
<sec id="s3">
<label>3</label>
<title>Knowledge Graph Construction Using Machine Learning</title>
<p>A typical machine learning approach is usually operated on a structured data matrix, where each row in the matrix corresponds to an object characterized by an attribute eigenvector. The main task of machine learning is to achieve the mapping from these eigenvectors to various forms of output through learning. Additionally, unsupervised learning can facilitate clustering and factor analysis. Then, according to the bottom-up KG construction process described in <xref ref-type="fig" rid="fig-8">Fig. 8</xref>, we summarize the realization of KG driven by machine learning from the following three parts, each bearing its unique set of implementation techniques and challenges.</p>
<sec id="s3_1">
<label>3.1</label>
<title>Entity Learning</title>
<p>Entity learning refers to the intricate construction process of entity layer within the KG. From bottom to top, it includes three modules: entity extraction, relationship extraction, and entity linking.</p>
<sec id="s3_1_1">
<label>3.1.1</label>
<title>Entity Extraction</title>
<p>Entity extraction stands as the initial and pivotal step in the knowledge extraction, involving the automatic identification of named entities from the original corpus. This foundational process relies on the automatic detection and categorization of named entities within a given corpus. <xref ref-type="fig" rid="fig-9">Fig. 9</xref> illustrates various categories to which these named entities can be ascribed, such as Person, Country, City, and more. This process enables the identification of important entities, laying the foundation for subsequent knowledge extraction and analysis tasks.</p>
<fig id="fig-9">
<label>Figure 9</label>
<caption>
<title>Examples of different types of entities</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-9.tif"/>
</fig>
<p>Generally, there are three typical methods in this field, and they are the rule-based, the traditional machine learning-based, and deep learning-based extraction methods. <xref ref-type="fig" rid="fig-10">Fig. 10</xref> displays the classification results obtained from these methods, showcasing the effectiveness and performance of each strategy in identifying and categorizing named entities. Here, rule extraction is an early pattern implemented by manually designing rules primarily toward proper nouns within a specific domain&#x2019;s text. It is based on painstakingly handcrafted patterns, necessitating a lot of human efforts, culminating in limited extraction capacity and constrained scalability. Meanwhile, the related extraction methods based on machine learning models have witnessed remarkable progress. These methods integrate machine learning algorithms into entity extraction to achieve automatic or semi-automatic entity identification. Finally, with the development of artificial neural networks, some deep learning-based methods have been proposed to attain heightened proficiency in entity extraction tasks with reduced human intervention. This progression signifies a remarkable leap forward in the field.</p>
<fig id="fig-10">
<label>Figure 10</label>
<caption>
<title>Classification results of entity extraction methods</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-10.tif"/>
</fig>
<p>For the traditional machine learning-based entity extraction, the inception involved the utilization of supervised machine learning algorithms. For example, the decision trees and the conditional random field (CRF) model [<xref ref-type="bibr" rid="ref-14">14</xref>] were employed to realize entity recognition for Telugu-English code-mixed social media data. In addition, Sykes et al. [<xref ref-type="bibr" rid="ref-15">15</xref>] conducted a comprehensive comparison between rule-based and machine learning-based extraction methods, effectively emphasizing the latter&#x2019;s superior adaptability and versatility.</p>
<p>Furthermore, with the progressive evolution of web technologies, the combination of open-linked data with machine learning algorithms frequently yields more compelling results. Within this framework, the basic idea revolves around employing machine learning to extract entities with similar contextual features from the web page, subsequently achieving entity classification or clustering. For example, Whitelaw et al. [<xref ref-type="bibr" rid="ref-16">16</xref>] proposed an iterative approach for expanding the entity corpus in a network environment. This approach depended on the construction of feature models grounded in known entities, enabling the processing of massive datasets. With this mechanism, they effectively modeled new entities to achieve continuous iterative expansion of the entity. Furthermore, employing unsupervised learning algorithms, Jain and Pennacchiotti [<xref ref-type="bibr" rid="ref-17">17</xref>] successfully extracted newly emerging named entities from the server logs of search engines. Then, this method found practical application in search engine technology, allowing for automatic information completion based on user-input keywords. Additionally, while constructing KG, an application of intelligent corpus annotation for entity extraction was presented [<xref ref-type="bibr" rid="ref-18">18</xref>].</p>
<p>Deep learning-based methods offer the advantage of automated text feature selection for entity extraction, which reduces incompleteness and manual work. They have shown promising results in this area. One such method employs convolutional neural networks (CNN) to automatically learn features from the input text data. For example, a deep learning-based method [<xref ref-type="bibr" rid="ref-19">19</xref>&#x2013;<xref ref-type="bibr" rid="ref-21">21</xref>] was proposed for entity extraction that utilized a CNN to learn contextual features from the input text data. Those models were trained on a large dataset of annotated text datasets and exhibited satisfactory performance on several benchmark datasets. Similarly, Cho et al. [<xref ref-type="bibr" rid="ref-22">22</xref>] presented a deep learning-based strategy for entity recognition in biomedical texts, combining a CNN with a long short-term memory (LSTM) network. Their model acquired complex features from the input text data, attaining remarkable accuracy on a challenging biomedical entity recognition task. Overall, deep learning-based methods have the potential to advance entity extraction, thereby constructing more precise and comprehensive KG. Hence, this enhancement leads to more powerful applications in many fields, such as natural language processing (NLP) and information retrieval. Notably, <xref ref-type="table" rid="table-3">Table 3</xref> showcases the top 10 entity extraction models from 2019 to 2022, ranked by F1-score on various open-source datasets (CoNLL 2003 [<xref ref-type="bibr" rid="ref-23">23</xref>], ACE 2005 [<xref ref-type="bibr" rid="ref-24">24</xref>], and Ontonotes 2005 [<xref ref-type="bibr" rid="ref-25">25</xref>]) This information is sourced from the Papers with Code website (<ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com">https://paperswithcode.com</ext-link>).</p>
<table-wrap id="table-3">
<label>Table 3</label>
<caption>
<title>The top 10 entity extraction models from 2019 to 2022</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Datasets</th>
<th>Model</th>
<th>F1-score (%)</th>
<th>Year</th>
<th>Platform</th>
</tr>
</thead>
<tbody>
<tr>
<td>CoNLL 2003</td>
<td>ACE&#x002B;document-context [<xref ref-type="bibr" rid="ref-26">26</xref>]</td>
<td>94.60</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Co-regularized LUKE [<xref ref-type="bibr" rid="ref-27">27</xref>]</td>
<td>94.22</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>ASP&#x002B;T5-3B [<xref ref-type="bibr" rid="ref-28">28</xref>]</td>
<td>94.10</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>FLERT XLM-R [<xref ref-type="bibr" rid="ref-29">29</xref>]</td>
<td>94.09</td>
<td>2020</td>
<td>Github/HuggingFace</td>
</tr>
<tr>
<td></td>
<td>PL-Marker [<xref ref-type="bibr" rid="ref-30">30</xref>]</td>
<td>94.00</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>LUKE [<xref ref-type="bibr" rid="ref-31">31</xref>]</td>
<td>93.91</td>
<td>2020</td>
<td>Github/HuggingFace</td>
</tr>
<tr>
<td></td>
<td>CL-KL [<xref ref-type="bibr" rid="ref-32">32</xref>]</td>
<td>93.85</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>XLNet-GCN [<xref ref-type="bibr" rid="ref-33">33</xref>]</td>
<td>93.82</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>ASP&#x002B;flan-T5-large [<xref ref-type="bibr" rid="ref-28">28</xref>]</td>
<td>93.80</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>InferNER [<xref ref-type="bibr" rid="ref-34">34</xref>]</td>
<td>93.76</td>
<td>2021</td>
<td>&#x2013;</td>
</tr>
<tr>
<td>ACE 2005</td>
<td>PURE [<xref ref-type="bibr" rid="ref-35">35</xref>]</td>
<td>90.90</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PromptNER[RoBERTa-large] [<xref ref-type="bibr" rid="ref-36">36</xref>]</td>
<td>88.26</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PIQN [<xref ref-type="bibr" rid="ref-37">37</xref>]</td>
<td>87.42</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PromptNER[BERT-large] [<xref ref-type="bibr" rid="ref-36">36</xref>]</td>
<td>87.21</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>DiffusionNER [<xref ref-type="bibr" rid="ref-38">38</xref>]</td>
<td>86.93</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>BERT-MRC [<xref ref-type="bibr" rid="ref-39">39</xref>]</td>
<td>86.88</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Locate and Label [<xref ref-type="bibr" rid="ref-40">40</xref>]</td>
<td>86.67</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>BoningKnife [<xref ref-type="bibr" rid="ref-41">41</xref>]</td>
<td>85.46</td>
<td>2021</td>
<td>&#x2013;</td>
</tr>
<tr>
<td></td>
<td>Biaffine-NER [<xref ref-type="bibr" rid="ref-42">42</xref>]</td>
<td>85.40</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Second-best learning and decoding [<xref ref-type="bibr" rid="ref-43">43</xref>]</td>
<td>84.34</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td>Ontonotes 2005</td>
<td>BERT-MRC&#x002B;DSC [<xref ref-type="bibr" rid="ref-44">44</xref>]</td>
<td>92.07</td>
<td>2019</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PL-Marker [<xref ref-type="bibr" rid="ref-30">30</xref>]</td>
<td>91.90</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Baseline&#x002B;BS [<xref ref-type="bibr" rid="ref-45">45</xref>]</td>
<td>91.74</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Biaffine-NER [<xref ref-type="bibr" rid="ref-42">42</xref>]</td>
<td>91.30</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>BERT-MRC [<xref ref-type="bibr" rid="ref-39">39</xref>]</td>
<td>91.11</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PIQN [<xref ref-type="bibr" rid="ref-37">37</xref>]</td>
<td>90.96</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>HGN [<xref ref-type="bibr" rid="ref-46">46</xref>]</td>
<td>90.92</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>Syn-LSTM&#x002B;BERT [<xref ref-type="bibr" rid="ref-47">47</xref>]</td>
<td>90.85</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>DiffusionNER [<xref ref-type="bibr" rid="ref-38">38</xref>]</td>
<td>90.66</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>W2NER [<xref ref-type="bibr" rid="ref-48">48</xref>]</td>
<td>90.50</td>
<td>2022</td>
<td>Github</td>
</tr>
</tbody>
</table>
<table-wrap-foot><fn><p>Notes: For more detailed and up-to-date information about the models than what is presented in the article, readers can refer to the following links: <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/named-entity-recognition-ner-on-conll-2003">https://paperswithcode.com/sota/named-entity-recognition-ner-on-conll-2003</ext-link>; <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/named-entity-recognition-on-ace-2005">https://paperswithcode.com/sota/named-entity-recognition-on-ace-2005</ext-link>; <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/named-entity-recognition-ner-on-ontonotes-v5">https://paperswithcode.com/sota/named-entity-recognition-ner-on-ontonotes-v5</ext-link>.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>For the CoNLL 2003 dataset, several remarkable models have garnered high F1-score, highlighting their prowess in entity extraction. The Automated Concatenation of Embeddings (ACE) combined with document-context [<xref ref-type="bibr" rid="ref-26">26</xref>] stands out at an impressive 94.60%. Close on its heels, the Co-regularized Language Understanding with Knowledge-based Embeddings (LUKE) [<xref ref-type="bibr" rid="ref-27">27</xref>] achieves a commendable 94.22%, while the Autoregressive Structured Prediction (ASP) fused with Text-to-Text Transfer Transformer (T5)-3B [<xref ref-type="bibr" rid="ref-28">28</xref>] achieves a noteworthy 94.10% (the &#x201C;3B&#x201D; represents 3 billion parameters). These models have demonstrated outstanding performance in extracting entities from the CoNLL 2003 dataset.</p>
<p>In the ACE 2005 dataset, the Princeton University Relation Extraction system (PURE) [<xref ref-type="bibr" rid="ref-35">35</xref>] emerges triumphant with an F1-score of 90.90%. Not far behind, PromptNER [<xref ref-type="bibr" rid="ref-36">36</xref>] has a second position with 88.26%, followed by Parallel Instance Query Network (PIQN) [<xref ref-type="bibr" rid="ref-37">37</xref>] with 87.42%. These models have shown strong performance in entity extraction from the ACE 2005 dataset.</p>
<p>For the Ontonotes 2005 dataset, the Bidirectional Encoder Representations from Transformers (BERT)-Machine Reading Comprehension (MRC)&#x002B;dice coefficient (DSC) [<xref ref-type="bibr" rid="ref-44">44</xref>] attains the pinnacle with the highest F1-score of 92.07%. Packed Levitated (PL)-Marker [<xref ref-type="bibr" rid="ref-30">30</xref>] follows closely behind with 91.90%, closely trailed by Baseline&#x002B;Boundary Smoothing (BS) [<xref ref-type="bibr" rid="ref-45">45</xref>] achieving 91.74%. These models have demonstrated their efficacy in entity extraction from the Ontonotes 2005 dataset.</p>
<p>These state-of-the-art models serve as prime examples of the evolutions made in entity extraction techniques, leveraging various approaches including deep learning, Prompt, and co-regularization methods. With remarkable F1-score on their respective datasets, these models indicate their effectiveness in extracting entities from text.</p>
</sec>
<sec id="s3_1_2">
<label>3.1.2</label>
<title>Relation Extraction</title>
<p>Relation extraction is an important subtask of knowledge extraction. In the past, this task entailed manual rule construction, followed by pattern-matching techniques to extract corresponding relation instances from text. However, the advent of machine learning has revolutionized relation extraction methods. It leverages lexical and syntactic attributes for model training, effectively transmuting relation extraction challenges into classification or clustering problems. According to the extent of human involvement and dependence on labeled corpus, machine learning-based relation extraction approaches can be divided into the supervised learning-based relation extraction, the semi-supervised learning-based relation extraction, and the unsupervised learning-based relation extraction.</p>
<p><bold>(1) Relation Extraction with Supervised Learning Method</bold></p>
<p>Supervised learning-based relation extraction is an automatic mode on the basis of meticulously labeled training data. Through ongoing learning from these training samples, the classification and predictions are performed on datasets. Within this paradigm, binary relation extraction is treated as a classification problem. As shown in the following definition, the triple (<inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:math></inline-formula>) indicates that there is a semantic relationship <inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math></inline-formula> between the head entity <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:math></inline-formula> and the tail entity <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:math></inline-formula>, and function <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mo>&#x22C5;</mml:mo><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula> represents the relationship classifier employed in the context:</p>
<p><disp-formula id="ueqn-1"><mml:math id="mml-ueqn-1" display="block"><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign="left left" rowspacing=".2em" columnspacing="1em" displaystyle="false"><mml:mtr><mml:mtd><mml:mn>1</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mtext>if the triple</mml:mtext></mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mspace width="thinmathspace" /><mml:mrow><mml:mtext>exists</mml:mtext></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mrow><mml:mtext>otherwise.</mml:mtext></mml:mrow></mml:mtd></mml:mtr></mml:mtable><mml:mo fence="true" stretchy="true" symmetric="true"></mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>Supervised learning employs two primary categories of relationship classification methods: feature vector-based and kernel function-based. In the first category, features are extracted from the training samples and represented as sequence feature vectors, using the results of part-of-speech tagging and syntax parsing. Here, the prominent methods are support vector machine (SVM) [<xref ref-type="bibr" rid="ref-49">49</xref>] and maximum entropy (ME) [<xref ref-type="bibr" rid="ref-50">50</xref>]. For example, a classifier system employing SVM integrated lexical features from polarity lexicons and lists of offensive/profane words to identify and classify offensive language in social media [<xref ref-type="bibr" rid="ref-49">49</xref>]. The second category of methods effectively avoids the challenges of dimensionality caused by nonlinear transformations. Recent years have witnessed widespread utilization of kernel function-based approaches in many fields [<xref ref-type="bibr" rid="ref-51">51</xref>]. In relation extraction using kernel functions, such as convolution kernel [<xref ref-type="bibr" rid="ref-52">52</xref>], tree kernel [<xref ref-type="bibr" rid="ref-53">53</xref>], subsequence kernel [<xref ref-type="bibr" rid="ref-54">54</xref>] and some improved kernel methods [<xref ref-type="bibr" rid="ref-55">55</xref>] play a pivotal role. The key of these methods involves projecting the implicit feature vector of a sentence into the feature space using the kernel function and calculating the inner product between these projections, so as to assess the similarity of the relationship between entities. For example, an innovative tree kernel, termed feature-enriched tree kernel (FTK) was proposed [<xref ref-type="bibr" rid="ref-53">53</xref>], while achieving a 5.4% enhancement in F-measure over the traditional convolution tree kernel.</p>
<p>Supervised learning-based relation extraction methods yield excellent experimental results, but their effectiveness heavily depends on the classification features provided by part-of-speech tagging and syntactic parsing. To address this problem, the use of supervised relation extraction has witnessed a surge of interest driven by the deep learning model [<xref ref-type="bibr" rid="ref-56">56</xref>]. A recurrent neural network (RNN)-based relation extraction model was proposed by Socher et al. [<xref ref-type="bibr" rid="ref-57">57</xref>]. This approach involved vectorizing each node of the syntactic tree through syntactic analysis. Guided by the syntactic structure, it iterated continuously from the lowest word vector of the tree. Finally, the vector representation of the sentence was attained and employed as the foundation for relationship classification. This method effectively utilized syntactic structure information, but overlooked the position information of words. Then, convolutional neural network (CNN) took center stage. Here, the word vector was treated as an initialization parameter, engaging in convolution training with dynamic optimization during the learning process, culminating in classification [<xref ref-type="bibr" rid="ref-58">58</xref>]. To complement the local dependencies captured by piecewise CNNs, a self-attention mechanism was proposed to capture rich contextual dependencies [<xref ref-type="bibr" rid="ref-59">59</xref>]. The experiments were performed on the NYT dataset and the experimental results demonstrated that the model provided a new benchmark in the area under curve (AUC) metric. Expanding on the CNN-based methodology presented [<xref ref-type="bibr" rid="ref-57">57</xref>], a refined iteration emerged [<xref ref-type="bibr" rid="ref-60">60</xref>]. This advancement involved inputting both word vectors and word positional vectors, with the sentence representation being obtained through the learning of convolutional, pooling, and nonlinear layers. This method fully considered the entity location information and other related lexical features, which achieved good relation extraction results. On the standard SemEval-2018 Task 7 dataset, the CNN method achieved superior performance when compared to alternative relation extraction methods [<xref ref-type="bibr" rid="ref-61">61</xref>]. Furthermore, except for deep learning models, deep reinforcement learning has also been used in the relation extraction [<xref ref-type="bibr" rid="ref-62">62</xref>]. This approach casts the relation extraction as a two-step decision-making game, employing the Q-Learning algorithm with value function approximation to learn control policy. The experiments were conducted on the ACE 2005 corpus, and they showed that the deep reinforcement learning model achieved a state-of-the-art performance in relation extraction tasks.</p>
<p><bold>(2) Relation Extraction with Semi-supervised Learning Method</bold></p>
<p>Semi-supervised learning-based relation extraction aims to realize the binary relation classification with limited training samples, thus circumventing the constraints imposed by manual annotation of extensive training data. It mainly adopts the bootstrapping method and some other methods.</p>
<p>The idea of bootstrapping is to artificially construct a small set of initial relation instances as a seed. This seed set serves as the foundation for model training, and through iterative expansion, it gradually augments to encompass a more extensive collection of relation instances, and finally completes the relation extraction task. Here, the entity alignment technique was improved to reduce the data noise [<xref ref-type="bibr" rid="ref-63">63</xref>]. However, this method operates under the assumption that a single entity pair corresponds to just one relationship. To address this limitation, a multi-instance multi-label (MIML) method was proposed to model the relationship extraction. This methodology describes the situation where an entity pair may have multiple relationships [<xref ref-type="bibr" rid="ref-64">64</xref>]. Moreover, the integration of a Bayesian network with MIML was explored for relation extraction [<xref ref-type="bibr" rid="ref-65">65</xref>], further expanding and capabilities of the approach.</p>
<p>Although the bootstrapping method is intuitive and effective, it may introduce a large number of noisy instances during seed expansion, leading to semantic drift. To address it, a deep co-learning was proposed [<xref ref-type="bibr" rid="ref-66">66</xref>], and it was a semi-supervised end-to-end deep learning method for evaluating the credibility of Arabic Blogs. A coupled semi-supervised learning method was used to establish constraints between different categories of extraction templates [<xref ref-type="bibr" rid="ref-67">67</xref>]. This strategic implementation effectively curbed the generation of false templates, thereby bolstering the precision of relation extraction. Meanwhile, a method was proposed by combining matrix-vector recursive neural network (MV-RNN) with bootstrapping [<xref ref-type="bibr" rid="ref-68">68</xref>]. Here, through the tree structure in MV-RNN, the semantic information of the entire sentence could be extracted as relation classifier features, and it greatly improved the accuracy of the results avoiding the problem of requiring a large amount of corpus in MV-RNN. More recently, the integration of transfer learning, a popular machine learning strategy, into the semi-supervised learning, yielded a novel framework [<xref ref-type="bibr" rid="ref-69">69</xref>]. Applied within the context of low-resource entity and relation extraction in the scientific domain, this framework demonstrated satisfactory performance, underscoring its potential and versatility.</p>
<p><bold>(3) Relation Extraction with Unsupervised Learning Method</bold></p>
<p>Unsupervised learning-based relation extraction assumes that pairs of entities with the same semantic relation have similar contexts, and it transforms the relation extraction task into a clustering problem. Hence, it does not require manual corpus annotation, but the accuracy rate is relatively low.</p>
<p>Generally, it is implemented using various clustering algorithms. For example, the large pre-trained language model was used for adaptive clustering on contextualized relational features to improve computational performance in relation classification [<xref ref-type="bibr" rid="ref-70">70</xref>]. After taking the entity set in the Wikipedia entry as the object, and using the dependency features and shallow grammar templates, all semantic relationship instances corresponding to entities were extracted in a large-scale corpus by pattern clustering [<xref ref-type="bibr" rid="ref-71">71</xref>]. Meanwhile, the templates were extracted and aggregated from search engine summaries, and they were clustered to discover implicit semantic relationships represented by entity pairs [<xref ref-type="bibr" rid="ref-72">72</xref>]. To further elevate the efficacy of relational templates&#x2019; clustering, a co-clustering algorithm was used, leveraging the dual nature of the dual of relational instances and relational templates. Moreover, the integration of a logistic regression model played a pivotal role in filtering representative extraction templates from the clustering results of relational templates [<xref ref-type="bibr" rid="ref-73">73</xref>].</p>
<p>For the above three basic types of machine learning, <xref ref-type="table" rid="table-4">Table 4</xref> offers a comprehensive comparison and analysis of the algorithms used in relation extraction. The table highlights the distinctions between classical algorithms, extraction ideas, human intervention levels, and extraction performance across different methods. Specifically, supervised and semi-supervised methods extract relationships by classifying learning sample data labels, while unsupervised-based methods cluster data and group related entities together to achieve relationship extraction. A noteworthy observation is that the performance of the extraction model improves with an increased corpus size and manual intervention. Taking into account the trade-off between performance and computational efforts, the semi-supervised learning methods emerge as favorable choices for relation extraction. It not only ensures better extraction performance but also avoids the limitation of requiring an extensive amount of manually annotated corpus. Moreover, <xref ref-type="table" rid="table-5">Table 5</xref> provides a comprehensive overview of the performance exhibited by various excellent relation extraction models on two famous datasets (NYT [<xref ref-type="bibr" rid="ref-74">74</xref>] and WebNLG [<xref ref-type="bibr" rid="ref-75">75</xref>]) in recent years. The F1-score achieved by these models indicate their effectiveness in relation extraction.</p>
<table-wrap id="table-4">
<label>Table 4</label>
<caption>
<title>Comparison of three kinds of machine learning-based relation extraction methods</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Extraction method</th>
<th>Examples of learning algorithms</th>
<th>The idea of extraction</th>
<th>Manual intervention</th>
<th>Extraction performance</th>
</tr>
</thead>
<tbody>
<tr>
<td>Supervised learning-based ones</td>
<td>SVM/ME/CRF /RNN/CNN</td>
<td>Classifying</td>
<td>More</td>
<td>High</td>
</tr>
<tr>
<td>Semi-supervised learning -based ones</td>
<td>Bootstrapping /Co-learning/MV-RNN</td>
<td>Classifying</td>
<td>Less</td>
<td>Medium</td>
</tr>
<tr>
<td>Unsupervised learning-based ones</td>
<td>Hierarchical Clustering /Co-clustering</td>
<td>Clustering</td>
<td>None</td>
<td>Low</td>
</tr>
</tbody>
</table>
</table-wrap><table-wrap id="table-5">
<label>Table 5</label>
<caption>
<title>The top 10 relation extraction models from 2020 to 2022</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Datasets</th>
<th>Model</th>
<th>F1-score (%)</th>
<th>Year</th>
<th>Platform</th>
</tr>
</thead>
<tbody>
<tr>
<td>NYT</td>
<td>UniRel [<xref ref-type="bibr" rid="ref-76">76</xref>]</td>
<td>93.7</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>REBEL [<xref ref-type="bibr" rid="ref-77">77</xref>]</td>
<td>93.4</td>
<td>2021</td>
<td>Github/HuggingFace</td>
</tr>
<tr>
<td></td>
<td>DIRECT [<xref ref-type="bibr" rid="ref-78">78</xref>]</td>
<td>92.5</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PFN [<xref ref-type="bibr" rid="ref-79">79</xref>]</td>
<td>92.4</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>SPN [<xref ref-type="bibr" rid="ref-80">80</xref>]</td>
<td>92.5</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>TDEER [<xref ref-type="bibr" rid="ref-81">81</xref>]</td>
<td>92.5</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>RIFRE [<xref ref-type="bibr" rid="ref-82">82</xref>]</td>
<td>92.0</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>TPLinker [<xref ref-type="bibr" rid="ref-83">83</xref>]</td>
<td>91.9</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PCNN&#x002B;RL&#x002B;HME [<xref ref-type="bibr" rid="ref-84">84</xref>]</td>
<td>90.0</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>CasRel [<xref ref-type="bibr" rid="ref-85">85</xref>]</td>
<td>89.6</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td>WebNLG</td>
<td>UniRel [<xref ref-type="bibr" rid="ref-76">76</xref>]</td>
<td>94.7</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>PFN [<xref ref-type="bibr" rid="ref-79">79</xref>]</td>
<td>93.6</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>SPN [<xref ref-type="bibr" rid="ref-80">80</xref>]</td>
<td>93.4</td>
<td>2023</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>TDEER [<xref ref-type="bibr" rid="ref-81">81</xref>]</td>
<td>93.1</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>RIFRE [<xref ref-type="bibr" rid="ref-82">82</xref>]</td>
<td>92.6</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>TPLinker [<xref ref-type="bibr" rid="ref-83">83</xref>]</td>
<td>91.9</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>CasRel [<xref ref-type="bibr" rid="ref-85">85</xref>]</td>
<td>91.8</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>RIN [<xref ref-type="bibr" rid="ref-86">86</xref>]</td>
<td>90.1</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>CGT [<xref ref-type="bibr" rid="ref-87">87</xref>]</td>
<td>83.4</td>
<td>2021</td>
<td>&#x2013;</td>
</tr>
<tr>
<td></td>
<td>JointER [<xref ref-type="bibr" rid="ref-88">88</xref>]</td>
<td>83.1</td>
<td>2020</td>
<td>Github</td>
</tr>
</tbody>
</table>
<table-wrap-foot><fn><p>Notes: For more detailed and up-to-date information about the models than what is presented in the article, readers can refer to the following links: <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/relation-extraction-on-nyt">https://paperswithcode.com/sota/relation-extraction-on-nyt</ext-link>; <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/relation-extraction-on-webnlg">https://paperswithcode.com/sota/relation-extraction-on-webnlg</ext-link>.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>Focusing on the NYT dataset, UniRel [<xref ref-type="bibr" rid="ref-76">76</xref>] achieved the highest F1-score of 93.7%, followed by Relation Extraction By End-to-end Language generation (REBEL) [<xref ref-type="bibr" rid="ref-77">77</xref>] with 93.4%, and Djacency lIst oRiented rElational faCT (DIRECT) [<xref ref-type="bibr" rid="ref-78">78</xref>] at 92.5%. These models have exhibited robust competence in extracting relations from the NYT dataset.</p>
<p>For the WebNLG dataset, UniRel [<xref ref-type="bibr" rid="ref-78">78</xref>] maintains its lead, achieving the highest F1-score of 94.7%, followed by Partition Filter Network (PFN) [<xref ref-type="bibr" rid="ref-79">79</xref>] with 93.6%, and Set Prediction Networks (SPN) [<xref ref-type="bibr" rid="ref-80">80</xref>] with 93.4%. These models have exhibited remarkable performance in relation extraction from the WebNLG dataset. Furthermore, other models such as Translating Decoding Schema for Joint Extraction of Entities and Relations (TDEER) [<xref ref-type="bibr" rid="ref-81">81</xref>] and Representation Iterative Fusion based on Heterogeneous Graph Neural Network for Joint Entity and Relation Extraction (RIFRE) [<xref ref-type="bibr" rid="ref-82">82</xref>] also achieved comparable performance on both datasets, further highlighting the consistent excellence of these methodologies.</p>
<p>These excellent models utilize various techniques, including RNNs, partition-based methods, and joint learning frameworks. The impressive F1-score attained by these models underscores their ability to effectively extract relations from text across different datasets. This analysis showcases the advancements in relation extraction models and their potential applications in numerous NLP tasks.</p>
</sec>
<sec id="s3_1_3">
<label>3.1.3</label>
<title>Entity Linking</title>
<p>Due to the diversity of information expression, entity ambiguity remains a frequent and formidable hurdle in natural language understanding. For example, &#x201C;apple&#x201D;, can signify either a fruit or the renowned technology company. However, when contextual information such as &#x201C;Steve Jobs&#x201D; is provided, it becomes evident that in the given text, the entity &#x201C;apple&#x201D; refers to the company, as illustrated in <xref ref-type="fig" rid="fig-11">Fig. 11</xref>. The presence of contextual cues helps disambiguate the intended meaning of entities, aiding in enhancing the precision and comprehension of text.</p>
<fig id="fig-11">
<label>Figure 11</label>
<caption>
<title>An example of entity linking</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-11.tif"/>
</fig>
<p>Entity linking is an effective disambiguation method. By linking entity mentions to the corresponding entities in the knowledge bases, this method conducts precise entity annotation within documents. This enables computers to attain a more profound grasp of the semantic information of the text, effectively addressing the problems of synonym and polysemy. Commonly, entity linking leverages well-established knowledge bases such as Wikipedia, DBPedia, and Freebase. Specifically, open knowledge bases like Baidu Encyclopedia and Interactive Encyclopedia also find prominent applications in Chinese entity linking. Generally speaking, entity linking includes two subtasks: entity recognition and entity disambiguation. By skillfully addressing both aspects, this technique contributes to a more comprehensive and nuanced understanding of textual content.</p>
<p><bold>(1) Entity Recognition</bold></p>
<p>Entity recognition aims to identify fragments of text, and it may link to specific entries in the knowledge bases, including a specific word or phrase. Typical entity types include place names, person names, institution names, times, dates, percentages, and amounts. With the development of Internet information, novel entity categories have surfaced in recent years, including movie titles and product names, reflecting the expanding landscape of entity recognition [<xref ref-type="bibr" rid="ref-89">89</xref>,<xref ref-type="bibr" rid="ref-90">90</xref>].</p>
<p>Currently, entity recognition technology is mainly based on statistical machine learning methods, treating the task as a sequence labeling problem. There are three main solutions: hidden Markov model (HMM), maximum entropy Markov model (MEMM), and CRF model. An overview of these models, their key evaluation criteria, and comparisons are delineated in <xref ref-type="table" rid="table-6">Table 6</xref>. Over recent years, numerous studies have utilized these three models. For instance, a novel generative model was proposed [<xref ref-type="bibr" rid="ref-91">91</xref>], linking it to HMM while proving its generically identifiable nature without any observed training labels. However, HMM exhibits the tag bias problem by assuming that the current tag solely depends on the previous one. Additionally, its use of local normalization to compute the probability of observation series makes calculating the function complex, leading to reduced efficiency. MEMM addresses the tag bias problem but introduces the tag inconsistency issue. Furthermore, it often relies on manually designed features, which are pivotal for the model&#x2019;s performance and generalization ability, making feature selection a challenging task. CRF model is different from HMM and MEMM. By employing global modeling, CRF simultaneously considers the observation and labeling of the entire sequence when calculating the probability. This enables the avoidance of label bias and local normalization problems found in HMM and MEMM. Moreover, the CRF model resolves the label inconsistency problem present in MEMM by modeling dependencies between observations and label sequences concurrently. Additionally, CRF utilizes complex and rich feature representations that capture dependencies between observed and labeled sequences, thereby enhancing model performance and generalization. Unlike MEMM, CRF model is less dependent on manually selected features and can learn feature weights that suit the task, thereby streamlining feature engineering. By overcoming some of the main shortcomings of HMM and MEMM models, CRF achieves superior performance in sequence labeling and has been widely used in the entity recognition field [<xref ref-type="bibr" rid="ref-92">92</xref>].</p>
<table-wrap id="table-6">
<label>Table 6</label>
<caption>
<title>Comparative analysis of three named entity recognition methods</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Model</th>
<th>Decision condition</th>
<th>Feature selection</th>
<th>Whether to mark the deviation</th>
<th>Computational complexity</th>
</tr>
</thead>
<tbody>
<tr>
<td>HMM</td>
<td>Mutually information probability</td>
<td>Limited</td>
<td>No</td>
<td>Low</td>
</tr>
<tr>
<td>MEMM</td>
<td>Conditional probability</td>
<td>Flexible</td>
<td>Yes</td>
<td>Low</td>
</tr>
<tr>
<td>CRF</td>
<td>Global probability</td>
<td>Flexible</td>
<td>No</td>
<td>High</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><bold>(2) Entity Disambiguation</bold></p>
<p>In the context of a specific entity mentioned in the text, entity disambiguation is mainly used to analyze the semantic information and select the corresponding entity from the candidate entities. This process depends on the candidate entity and the contextual information. Usually, given the ambiguity in natural language, there are numerous candidate entities vying for consideration. Methods of entity disambiguation are mainly based on supervised learning and unsupervised learning.</p>
<p>Most works adopt supervised learning methods for disambiguation and use training data to automatically design ranking models. The requisite training data for entity linking comprises an ordered list of all candidate entities associated with the target mentioned in a given context. In this list, the first entity is usually the one that the mention refers to in this context. By linearly combining features like entity popularity, semantic similarity, and connection between entities, a maximum margin-based data were employed to train feature weights, while achieving entity disambiguation through a ranking model [<xref ref-type="bibr" rid="ref-93">93</xref>]. Moreover, two machine learning sorting-based methods using listwise and pairwise were presented to implement entity disambiguation, outperforming traditional disambiguation methods [<xref ref-type="bibr" rid="ref-94">94</xref>]. The listwise method is from the LiNet algorithm, which uses the ordered list as a training instance to obtain a sorting model. The essence of the pairwise method is to transform the sorting problem into a classification problem. It combines the items in the ordered list into pairs and constructs training instances according to the relative positional relationship between the items to develop a sorting perception. Furthermore, recent developments have seen the combination of supervised learning with graph theory to address the entity disambiguation task [<xref ref-type="bibr" rid="ref-95">95</xref>], showcasing the innovative fusion of established techniques to enhance disambiguation accuracy.</p>
<p>On the other side, unsupervised learning algorithms have been used in entity disambiguation. For example, a clustering-based personal name disambiguation system was proposed to extract personal attributes and social relations between entities from text, subsequently mapping them onto an undirected weighted graph [<xref ref-type="bibr" rid="ref-96">96</xref>]. Clustering algorithms were then used to cluster these graphs, each cluster contained all web pages that directed a person. Significant models have been shown in <xref ref-type="table" rid="table-7">Table 7</xref> in the field of entity disambiguation, where datasets are ACE2004 [<xref ref-type="bibr" rid="ref-97">97</xref>] and AIDA-CoNLL [<xref ref-type="bibr" rid="ref-98">98</xref>]. These models focus on disambiguating entities, which is crucial for accurately identifying the intended meaning of ambiguous terms.</p>
<table-wrap id="table-7">
<label>Table 7</label>
<caption>
<title>The top 5 entity disambiguation models</title>
</caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th>Datasets</th>
<th>Model</th>
<th>F1-score (%)</th>
<th>Year</th>
<th>Platform</th>
</tr>
</thead>
<tbody>
<tr>
<td>ACE 2004</td>
<td>KBED [<xref ref-type="bibr" rid="ref-99">99</xref>]</td>
<td>93.4</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>LUKE[confidence-order] [<xref ref-type="bibr" rid="ref-100">100</xref>]</td>
<td>91.9</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>ReFinED [<xref ref-type="bibr" rid="ref-101">101</xref>]</td>
<td>91.6</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>NER4EL [<xref ref-type="bibr" rid="ref-102">102</xref>]</td>
<td>91.3</td>
<td>2021</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>GENRE [<xref ref-type="bibr" rid="ref-103">103</xref>]</td>
<td>90.1</td>
<td>2020</td>
<td>&#x2013;</td>
</tr>
<tr>
<td>AIDA-CoNLL</td>
<td>LUKE[confidence-order] [<xref ref-type="bibr" rid="ref-100">100</xref>]</td>
<td>95.0</td>
<td>2022</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>DCA-SL&#x002B;Triples [<xref ref-type="bibr" rid="ref-104">104</xref>]</td>
<td>94.9</td>
<td>2020</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>DeepType [<xref ref-type="bibr" rid="ref-105">105</xref>]</td>
<td>94.9</td>
<td>2018</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>NTEE [<xref ref-type="bibr" rid="ref-106">106</xref>]</td>
<td>94.7</td>
<td>2017</td>
<td>Github</td>
</tr>
<tr>
<td></td>
<td>DCA-SL [<xref ref-type="bibr" rid="ref-107">107</xref>]</td>
<td>94.6</td>
<td>2019</td>
<td>Github</td>
</tr>
</tbody>
</table>
<table-wrap-foot><fn><p>Notes: For more detailed and up-to-date information about the models than what is presented in the article, readers can refer to the following links: <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/entity-disambiguation-on-ace2004">https://paperswithcode.com/sota/entity-disambiguation-on-ace2004</ext-link>; <ext-link ext-link-type="uri" xlink:href="https://paperswithcode.com/sota/entity-disambiguation-on-aida-conll">https://paperswithcode.com/sota/entity-disambiguation-on-aida-conll</ext-link>.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>For the ACE 2004 dataset, Entity Disambiguation by Reasoning over a Knowledge Base (KBED) [<xref ref-type="bibr" rid="ref-99">99</xref>] stands out with the highest F1-score of 93.4%, followed by LUKE[confidence-order] [<xref ref-type="bibr" rid="ref-100">100</xref>] with 91.9%, and Representation and Fine-grained typing for Entity Disambiguation (ReFinED) [<xref ref-type="bibr" rid="ref-101">101</xref>] at 91.6%. These models have shown strong performance in entity disambiguation on the ACE 2004 dataset.</p>
<p>Regarding the AIDA-CoNLL dataset, LUKE[confidence-order] [<xref ref-type="bibr" rid="ref-100">100</xref>] maintains its prominence with the highest F1-score of 95.0%, closely followed by Dynamic Context Augmentation (DCA)-Supervised Learning (SL)&#x002B;Triples [<xref ref-type="bibr" rid="ref-104">104</xref>] and DeepType [<xref ref-type="bibr" rid="ref-105">105</xref>] with F1-score of 94.9%. These models have demonstrated excellent performance in disambiguating entities in the AIDA-CoNLL dataset. It is worth mentioning that LUKE[confidence-order] [<xref ref-type="bibr" rid="ref-100">100</xref>] is the only model present in both datasets, indicating its robustness and effectiveness across different evaluation scenarios.</p>
<p>The top-performing models employ diverse techniques such as knowledge-based methods, confidence ordering, and deep learning approaches. Their high F1-score emphasize their proficiency in accurately disambiguating entities and determining their intended meanings in different contexts.</p>
</sec>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Ontology Learning</title>
<p>An ontology is a formalized specification of the shared conceptual model, and it defines the pattern layer of KG. The composition of an ontology as <italic>O</italic>, entails the components (<italic>C</italic>, <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:mi>r</mml:mi><mml:mi>o</mml:mi><mml:mi>o</mml:mi><mml:mi>t</mml:mi></mml:math></inline-formula>, <italic>R</italic>). Here, <italic>C</italic> is the set of upper-level concepts, the <inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:mi>r</mml:mi><mml:mi>o</mml:mi><mml:mi>o</mml:mi><mml:mi>t</mml:mi></mml:math></inline-formula> is the root identifier, and <italic>R</italic> is the binary relationship on <italic>C</italic>, including the synonymy relation and the hyponymy relation as <xref ref-type="fig" rid="fig-12">Fig. 12</xref> [<xref ref-type="bibr" rid="ref-108">108</xref>]. The purpose of an ontology lies in establishing an organized framework, thereby facilitating the organization and categorization of concepts within the KG, enabling efficient retrieval and navigation of knowledge stored within. Generally speaking, ontology construction has three ways: manual construction, automatic construction, and semi-automatic construction. Manual construction method necessitates the participation of domain experts and entails the utilization of dedicated ontology editing tools. However, this approach tends to demand significant human and material resources, leading to scalability issues. Hence, it cannot keep up with the rapid development and update of Internet data. Ontology automatic construction or ontology learning includes extraction of concept, synonymy relation, and hyponymy relation from bottom to top. This process is predominantly automated, often relying on data-driven or cross-language knowledge linkage techniques that are grounded in machine learning principles. For instance, in order to address the ontology automation issues in the semantic web, ontology learning was achieved through the presentation of automatic or semi-automatic, aimed at either generating new ontology resources or repurposing existing ones [<xref ref-type="bibr" rid="ref-109">109</xref>].</p>
<fig id="fig-12">
<label>Figure 12</label>
<caption>
<title>The flowchart of ontology modeling</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-12.tif"/>
</fig>
<p>Concept is the most basic unit of human comprehension. The common methods of concept extraction include linguistic methods [<xref ref-type="bibr" rid="ref-110">110</xref>], statistical methods [<xref ref-type="bibr" rid="ref-111">111</xref>], and machine learning methods. Within the domain of machine learning-based concept extraction, prevalent methods center around SVM [<xref ref-type="bibr" rid="ref-49">49</xref>], HMM [<xref ref-type="bibr" rid="ref-91">91</xref>], bootstrapping [<xref ref-type="bibr" rid="ref-112">112</xref>], and clustering. These techniques entail the extraction of pertinent categorical attributes from the dataset. For example, bootstrapping was used to automatically extract domain vocabulary from large-scale, unlabeled real corpus [<xref ref-type="bibr" rid="ref-112">112</xref>]. A bootstrapping-based seed expansion mechanism was developed to realize the automatic extraction of domain seed words [<xref ref-type="bibr" rid="ref-113">113</xref>]. Using the concept of clusters, the multi-topic extraction algorithm was introduced to acquire multiple topics by clustering concepts [<xref ref-type="bibr" rid="ref-114">114</xref>].</p>
<p>Relationship extraction between concepts mainly refers to the synonymy and the hyponymy relation. Synonymy relation extraction is to examine the degree of probability that any two entities belong to the same conceptual level. For example, entities like &#x201C;Beijing&#x201D; and &#x201C;Shanghai&#x201D;, both denoting city names, exhibit a synonymy relation. Hyponymy relation extraction gauges the probability that any two entities establish a hierarchical relation, where one serves as a subtype of the other. For example, &#x201C;Beijing&#x201D; is the hyponymy of &#x201C;city&#x201D;. For the supervised learning method, a novel approach was proposed for synonym identification using the principle of distributional similarity [<xref ref-type="bibr" rid="ref-115">115</xref>]. Compared to the traditional similarity models, the experimental results showed that a satisfactory performance was achieved while increasing by over 120% on the <inline-formula id="ieqn-14"><mml:math id="mml-ieqn-14"><mml:msub><mml:mi>F</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:math></inline-formula> metric. A syntactic parser was used to construct a syntax tree, and the contextual features of concepts were used as concept attributes to generate concept lattices [<xref ref-type="bibr" rid="ref-116">116</xref>]. This led to the establishment of a partial order relationship of concept lattices, which subsequently formed the conceptual hierarchy of ontology. For the unsupervised learning method, a method was proposed to learn taxonomy from a collection of text documents, each dedicated to describing a distinct concept [<xref ref-type="bibr" rid="ref-117">117</xref>]. Specifically, with the continuous development of online encyclopedias, the machine learning method using the linked data has gradually become an efficient strategy for hyponymy relation extraction. Machine learning techniques were used to explicitly represent the meaning of any text as a weighted vector of Wikipedia-based concepts. The cosine of the angle between these vectors was then calculated to measure the similarity between concepts or texts, effectively fostering the extraction of hyponymy relations [<xref ref-type="bibr" rid="ref-118">118</xref>].</p>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Knowledge Reasoning</title>
<p>Building upon the foundation of existing entities and relationships, knowledge reasoning amis to mine implicit connections between entities through a sophisticated reasoning mechanism. The ultimate goal is to enhance and amplify the original KG by unveiling implicit relationships. For instance, as illustrated in <xref ref-type="fig" rid="fig-13">Fig. 13</xref>, if the KG contains the information &#x201C;Lion is-a animal&#x201D; and &#x201C;Animal can run&#x201D;, it becomes possible to infer the knowledge that &#x201C;Lion can run&#x201D; through logical reasoning. This process stands as a testament to the potential of knowledge reasoning, allowing the enrichment and augmentation of the KG through the generation of novel insights from preexisting information. This iterative process culminates in heightened completeness and a more profound grasp of semantic understanding.</p>
<fig id="fig-13">
<label>Figure 13</label>
<caption>
<title>An example of knowledge reasoning</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-13.tif"/>
</fig>
<p>Traditional knowledge reasoning is mainly based on logical reasoning methods, including predicate logic, description logic, rule-based reasoning, and others. The predicate logic method is generally designed for simple entity relations. This method takes propositions as the fundamental units of reasoning, where atomic propositions are generally decomposed into two parts: individual words and predicates. Description logic can be used for complex relationships between entities. The typical path ranking algorithm plays a pivotal role in establishing rule-based reasoning. Distinctive relationship paths were used as one-dimensional features and the classification feature vector of the relationship was constructed by counting a large number of relationship paths in the graph [<xref ref-type="bibr" rid="ref-119">119</xref>]. However, it is important to acknowledge that logical reasoning necessitates the formulation of rules, a task that often proves computationally onerous and encounters challenges posed by data sparsity.</p>
<p>Link prediction is a new type of knowledge reasoning method under the statistical machine learning framework. It is to predict the possibility of the linked relationship between two unlinked nodes through the known nodes and link information in the KG, while discovering the implicit relationship between entities. The link prediction in the KG is generally realized by the representation learning method on the basis of the triples <inline-formula id="ieqn-15"><mml:math id="mml-ieqn-15"><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula> used to constitute the knowledge. For the representation learning method, the semantic information of an entity is represented as dense low-dimensional real-valued vectors. Within this space, gauging the semantic similarity between objects is facilitated by mathematical methods such as cosine distance and Euclidean distance [<xref ref-type="bibr" rid="ref-120">120</xref>].</p>
<p>Recently, a large number of works on the representation learning-based link prediction have been proposed. One illustrative instance is the structured embedding (SE) method. This method projects the entity vectors <inline-formula id="ieqn-16"><mml:math id="mml-ieqn-16"><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-17"><mml:math id="mml-ieqn-17"><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:math></inline-formula> through the two relation matrices of the inter-entity relation <inline-formula id="ieqn-18"><mml:math id="mml-ieqn-18"><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math></inline-formula> to the corresponding space of <inline-formula id="ieqn-19"><mml:math id="mml-ieqn-19"><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math></inline-formula> [<xref ref-type="bibr" rid="ref-121">121</xref>]. Then, the distance between two projection vectors on this space was calculated to gauge the confidence of <inline-formula id="ieqn-20"><mml:math id="mml-ieqn-20"><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math></inline-formula>. The single layer model (SLM) was proposed through the application of a single-layer neural network [<xref ref-type="bibr" rid="ref-122">122</xref>]. This network employs nonlinear operation to define a scoring function for each triple. fostering the synergistic representation of the semantic connection between entities and relationships to reason about unknown relationships. While the accuracy of results exhibited significant enhancement over traditional methods, the adoption of nonlinear operations inevitably led to heightened computational complexity [<xref ref-type="bibr" rid="ref-122">122</xref>]. Another innovative contribution is the semantic matching energy (SME) model, which relies on low-dimensional vectors to represent entities and relationships [<xref ref-type="bibr" rid="ref-123">123</xref>]. Multiple projection matrices are employed to represent the connections between entities and relationships. The latent factor model (LFM), delves into a relationship-centered bilinear transformation. This transformation encapsulates the semantic connection between entities and relationships [<xref ref-type="bibr" rid="ref-124">124</xref>]. The RESACL model was presented as a typical knowledge representation method through matrix decomposition [<xref ref-type="bibr" rid="ref-125">125</xref>]. In this study, all triples <inline-formula id="ieqn-21"><mml:math id="mml-ieqn-21"><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula> were represented as a large tensor, wherein the presence or absence of a triple dictates the value at the corresponding tensor position. Through the tensor decomposition algorithm, the tensor value corresponding to each triple in the tensor could be decomposed into entity and relation representations. Based on the characteristic of translation invariance of the word vector space, TransE model presents a pioneering perspective [<xref ref-type="bibr" rid="ref-126">126</xref>]. As depicted in <xref ref-type="fig" rid="fig-14">Fig. 14</xref>, in this model, the semantic relation <inline-formula id="ieqn-22"><mml:math id="mml-ieqn-22"><mml:msub><mml:mi>r</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:math></inline-formula> between entities <inline-formula id="ieqn-23"><mml:math id="mml-ieqn-23"><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-24"><mml:math id="mml-ieqn-24"><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:math></inline-formula> was regarded as some kind of translation vector, and it actually was the translation from the head entity <inline-formula id="ieqn-25"><mml:math id="mml-ieqn-25"><mml:msub><mml:mi>e</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:math></inline-formula> to the tail entity <inline-formula id="ieqn-26"><mml:math id="mml-ieqn-26"><mml:msub><mml:mi>e</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:math></inline-formula>. Demonstrating substantial improvements in establishing intricate semantic connections within vast, sparse KGs, TransE has solidified its status as a pivotal model in this domain. Recent developments have exhibited a growing interest in the TransE model, exemplified by the increased attention it has garnered in studies [<xref ref-type="bibr" rid="ref-127">127</xref>,<xref ref-type="bibr" rid="ref-128">128</xref>].</p>
<fig id="fig-14">
<label>Figure 14</label>
<caption>
<title>A brief description of the TransE model</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-14.tif"/>
</fig>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>The Future Challenges of Constructing a Large-Scale Knowledge Graph</title>
<p>Amidst the era of big data, the KG provides a new learning paradigm to efficiently organize, manage and understand massive amounts of information, while presenting this data in a manner closely aligned with human cognition. Hence, it promotes rapid advancements across many fields such as information retrieval, knowledge recommendation, and others. Although significant improvements have been achieved in the research of KG, the recently developed machine learning technologies and methods in big data analysis still cannot effectively match the demands of exploiting and using KG due to the complexity of real world application scenarios. In effect, a myriad of technical challenges persist, necessitating adept resolution to release the full potential of large-scale KG as a powerful methodology for intelligent information service. Here, we summarize the future research trends and challenges of machine learning-driven large-scale KG construction from three aspects.</p>
<sec id="s4_1">
<label>4.1</label>
<title>Relation Extraction</title>
<p>On one hand, the current focus of relation extraction predominantly revolves around the monolingual text. However, in the real-world, factual knowledge finds its repository in diverse sources, such as multilingual texts, pictures, audio, and video. Expanding the horizons of relation extraction to encompass these various sources stands as a promising avenue for future exploration, heralding the potential to broaden the spectrum of extracted relations and extend the scope of knowledge coverage [<xref ref-type="bibr" rid="ref-56">56</xref>]. On the other hand, for the deep learning-based relation extraction methods, the integration of syntactic trees via neural network models yields the effective amalgamation of syntactic information. However, it also leads to the introduction of a large amount of noise, which poses an impact on the accuracy of the model [<xref ref-type="bibr" rid="ref-57">57</xref>,<xref ref-type="bibr" rid="ref-60">60</xref>]. Constructing multiple possible syntactic trees of sentences and fusing them for relation extraction may be a development prospect. Furthermore, the open field-based relationship extraction is constantly updated and iterative. How to introduce deep learning models to achieve rapid learning of new relationships and knowledge is also a problem that needs to be explored [<xref ref-type="bibr" rid="ref-129">129</xref>].</p>
</sec>
<sec id="s4_2">
<label>4.2</label>
<title>Link Prediction</title>
<p>Link prediction plays an important role in the design of KG, serving as a critical component in inferring absent relationships. For example, as shown in <xref ref-type="fig" rid="fig-15">Fig. 15</xref>, where known relationships between entities A and B, B and C, and C and D, the link prediction method enables the inference of the relationship type between entities A and D. By leveraging the existing relationships within the KG, this technique enables the identification and prediction of unobserved connections, thereby enhancing the overall comprehension and knowledge extraction from the graph. Generally, the implementation of link prediction is mainly based on triples used to constitute knowledge. However, the types of knowledge are rich and diverse, and some complex knowledge cannot be directly represented by triples. Hence, different knowledge representation methods need to be set up for different scenarios. For instance, considering the temporal dynamism of factual values, a compound value type structure has been introduced, involving auxiliary nodes to represent <inline-formula id="ieqn-27"><mml:math id="mml-ieqn-27"><mml:mi>n</mml:mi></mml:math></inline-formula>-order relations and temporal attributes for facts [<xref ref-type="bibr" rid="ref-130">130</xref>]. Multiple <inline-formula id="ieqn-28"><mml:math id="mml-ieqn-28"><mml:mi>n</mml:mi></mml:math></inline-formula>-order relations can be represented by a single <inline-formula id="ieqn-29"><mml:math id="mml-ieqn-29"><mml:mo stretchy="false">(</mml:mo><mml:mi>n</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula>-order tensor, which is solved by higher-order tensor decomposition using the RESACL model [<xref ref-type="bibr" rid="ref-131">131</xref>]. The representation of learning-based link prediction is still in the initial stage of exploration [<xref ref-type="bibr" rid="ref-132">132</xref>]. It achieves unsatisfactory performance on large-scale KG with strong sparsity and the representation of low-frequency entities and relationships. It is urgent to design a more efficient online learning scheme for KG. Concurrently, the domain of network embedding-based algorithms, including those grounded in graph neural networks (GNN), has showcased compelling computational prowess in task completion. Specifically, the attention mechanism-based heterogeneous GNN is conducive to capturing information of various semantics in KGs [<xref ref-type="bibr" rid="ref-133">133</xref>]. Additionally, the introduction of the multi-scale dynamic convolutional network (M-DCN) has provided a framework for representing KG embeddings [<xref ref-type="bibr" rid="ref-134">134</xref>]. Therefore, how to creatively investigate those algorithms in the achievement of link prediction for KG is also an interesting direction [<xref ref-type="bibr" rid="ref-135">135</xref>,<xref ref-type="bibr" rid="ref-136">136</xref>].</p>
<fig id="fig-15">
<label>Figure 15</label>
<caption>
<title>Node roles for link prediction</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMES_31513-fig-15.tif"/>
</fig>
</sec>
<sec id="s4_3">
<label>4.3</label>
<title>Construction of Industrial Knowledge Graph</title>
<p>In a broader context, the construction of knowledge graphs (KGs) has spanned general domains, and the corresponding theoretical investigations have followed suit. However, the construction of KG for a specific field, especially in industry, has attracted less attention in recent years. General KGs emphasize the entity layer and it is difficult to generate a global ontology pattern. Notably, there are many differences between industrial KG and general KG [<xref ref-type="bibr" rid="ref-137">137</xref>]. The industrial KG has a clear industry background, while the entities have rich data patterns. Furthermore, the industrial KGs need comprehensive consideration of personnel at various hierarchical levels. This has led to new diverse challenges for researchers focusing on the design of a large-scale industrial KG. Thus, it needs to be explored in this direction, while presenting some new machine learning-driven methods in these fields [<xref ref-type="bibr" rid="ref-138">138</xref>,<xref ref-type="bibr" rid="ref-139">139</xref>].</p>
<p>Over the last year, significant advancements have been made in NLP tasks with the emergence of large language models (LLMs) like ChatGPT [<xref ref-type="bibr" rid="ref-140">140</xref>], Dolly [<xref ref-type="bibr" rid="ref-141">141</xref>], and LLaMA [<xref ref-type="bibr" rid="ref-142">142</xref>]. Despite their achievements, LLMs are black-box models, lacking transparency in tracking their search process. On the other hand, KG offers a high level of professional feasibility, providing more accurate answers and better interpretability compared to LLMs. Leveraging the respective advantages of LLMs and KG, we can utilize LLMs for data annotation and data enhancement in the future, facilitating the rapid application of industrial KGs [<xref ref-type="bibr" rid="ref-143">143</xref>]. Meanwhile, industrial KG can serve as an external knowledge base, enabling the introduction of specified constraints to LLMs [<xref ref-type="bibr" rid="ref-144">144</xref>]. This allows for controlled content generation and enhances the adaptability of LLMs within specific industrial fields.</p>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Conclusion</title>
<p>The centrality of KGs within the realm of next-generation search engines has established them as a pivotal focal point within the sphere of intelligent information processing, coinciding with the advent of the big data era. With the ever-increasing demand for processing performance in consideration of complex application scenarios, there are many exploratory works to be explored in the KG construction. In this article, we comprehensively review the key technologies in KG construction from the perspective of machine learning-related implementation methods. We are concerned with the core construction algorithms of KG in three aspects, including entity learning, ontology learning, and knowledge reasoning. Especially, the machine learning-driven algorithms for entity extraction, relation extraction, entity linking, and link prediction are deeply analyzed. In addition, considering the current development level of machine learning methods, we summarize some key problems and possible research trends in the construction of large-scale KG to serve as an impetus for researchers to work in the future.</p>
</sec>
</body>
<back>
<ack>
<p>None.</p>
</ack>
<sec><title>Funding Statement</title>
<p>This work was supported in part by the Beijing Natural Science Foundation under Grants L211020 and M21032, in part by the National Natural Science Foundation of China under Grants U1836106 and 62271045, and in part by the Scientific and Technological Innovation Foundation of Foshan under Grants BK21BF001 and BK20BF010.</p>
</sec>
<sec><title>Author Contributions</title>
<p>The authors confirm contribution to the paper as follows: study conception and design: Z. Zhao, X. Luo; data collection: Z. Zhao, M. Chen, L. Ma; analysis and interpretation of results: Z. Zhao, X. Luo, M. Chen; draft manuscript preparation: Z. Zhao, X. Luo, L. Ma. All authors reviewed the results and approved the final version of the manuscript.</p>
</sec>
<sec sec-type="data-availability"><title>Availability of Data and Materials</title>
<p>The presented data are open-sourced and can be accessed through the &#x201C;Web of Science&#x201D; and &#x201C;Papers with Code website&#x201D;.</p>
</sec>
<sec sec-type="COI-statement"><title>Conflicts of Interest</title>
<p>The authors declare that they have no conflicts of interest to report regarding the present study.</p>
</sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>1.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ragavan</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Rubavathi</surname>, <given-names>C. Y.</given-names></string-name></person-group> (<year>2022</year>). <article-title>A novel big data storage reduction model for drill down search</article-title>. <source>Computer Systems Science and Engineering</source><italic>,</italic> <volume>41</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>373</fpage>&#x2013;<lpage>387</lpage>. <pub-id pub-id-type="doi">10.32604/csse.2022.020452</pub-id></mixed-citation></ref>
<ref id="ref-2"><label>2.</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Singha</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Official Google blog: Introducing the knowledge graph: Things not strings</article-title>. <ext-link ext-link-type="uri" xlink:href="http://googleblog.blogspot.pt/2012/05/introducing-knowledge-graph-things-not.html">http://googleblog.blogspot.pt/2012/05/introducing-knowledge-graph-things-not.html</ext-link></mixed-citation></ref>
<ref id="ref-3"><label>3.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Kshetri</surname>, <given-names>N.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Web 3.0 and the metaverse shaping organizations&#x2019; brand and product strategies</article-title>. <source>IT Professional</source><italic>,</italic> <volume>24</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>11</fpage>&#x2013;<lpage>15</lpage>.</mixed-citation></ref>
<ref id="ref-4"><label>4.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Greenbaum</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2022</year>). <article-title>The virtual worlds of the metaverse</article-title>. <source>Science</source><italic>,</italic> <volume>377</volume><italic>(</italic><issue>6604</issue><italic>),</italic> <fpage>377</fpage>.</mixed-citation></ref>
<ref id="ref-5"><label>5.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zamini</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Reza</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Rabiei</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2022</year>). <article-title>A review of knowledge graph completion</article-title>. <source>Information</source><italic>, </italic><volume>13</volume><italic>(</italic><issue>8</issue><italic>),</italic> <fpage>396</fpage>.</mixed-citation></ref>
<ref id="ref-6"><label>6.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Miller</surname>, <given-names>G. A.</given-names></string-name></person-group> (<year>1995</year>). <article-title>WordNet: A lexical database for English</article-title>. <source>Communications of the ACM</source><italic>,</italic> <volume>38</volume><italic>(</italic><issue>11</issue><italic>),</italic> <fpage>39</fpage>&#x2013;<lpage>41</lpage>.</mixed-citation></ref>
<ref id="ref-7"><label>7.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Nickel</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Murphy</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Tresp</surname>, <given-names>V.</given-names></string-name>, <string-name><surname>Gabrilovich</surname>, <given-names>E.</given-names></string-name></person-group> (<year>2015</year>). <article-title>A review of relational machine learning for knowledge graphs</article-title>. <source>Proceedings of the IEEE</source><italic>,</italic> <volume>104</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>11</fpage>&#x2013;<lpage>33</lpage>.</mixed-citation></ref>
<ref id="ref-8"><label>8.</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Sowa</surname>, <given-names>J. F.</given-names></string-name></person-group> (<year>2014</year>). <source>Principles of semantic networks: Explorations in the representation of knowledge</source>. <publisher-loc>San Mateo, CA, USA</publisher-loc>: <publisher-name>Morgan Kaufmann</publisher-name>.</mixed-citation></ref>
<ref id="ref-9"><label>9.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hoffart</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Suchanek</surname>, <given-names>F. M.</given-names></string-name>, <string-name><surname>Berberich</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Lewis-Kelham</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>De Melo</surname>, <given-names>G.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2011</year>). <article-title>YAGO2: Exploring and querying world knowledge in time, space, context, and many languages</article-title>. <conf-name>Proceedings of the 20th International Conference Companion on World Wide Web</conf-name>, pp. <fpage>229</fpage>&#x2013;<lpage>232</lpage>. <publisher-loc>New York, NY, USA</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-10"><label>10.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Opdahl</surname>, <given-names>A. L.</given-names></string-name>, <string-name><surname>Al-Moslmi</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Dang-Nguyen</surname>, <given-names>D. T.</given-names></string-name>, <string-name><surname>Gallofr&#x00E9; Oca&#x00F1;a</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Tessem</surname>, <given-names>B.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>Semantic knowledge graphs for the news: A review</article-title>. <source>ACM Computing Surveys</source><italic>,</italic> <volume>55</volume><italic>(</italic><issue>7</issue><italic>),</italic> <fpage>1</fpage>&#x2013;<lpage>38</lpage>.</mixed-citation></ref>
<ref id="ref-11"><label>11.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>McGuinness</surname>, <given-names>D. L.</given-names></string-name>, <string-name><surname>van Harmelen</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Deborah</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Frank</surname></string-name></person-group> (<year>2004</year>). <article-title>OWL web ontology language overview</article-title>. <source>W3C Recommendation</source><italic>,</italic> <volume>10</volume><italic>(</italic><issue>10</issue><italic>),</italic> <fpage>1</fpage>&#x2013;<lpage>12</lpage>.</mixed-citation></ref>
<ref id="ref-12"><label>12.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Lygerakis</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Kampelis</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Kolokotsa</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Knowledge graphs&#x2019; ontologies and applications for energy efficiency in buildings: A review</article-title>. <source>Energies</source><italic>,</italic> <volume>15</volume><italic>(</italic><issue>20</issue><italic>),</italic> <fpage>7520</fpage>.</mixed-citation></ref>
<ref id="ref-13"><label>13.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yu</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Mao</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Cai</surname>, <given-names>Q.</given-names></string-name></person-group> (<year>2021</year>). <article-title>A domain knowledge graph construction method based on Wikipedia</article-title>. <source>Journal of Information Science</source><italic>,</italic> <volume>47</volume><italic>(</italic><issue>6</issue><italic>),</italic> <fpage>783</fpage>&#x2013;<lpage>793</lpage>.</mixed-citation></ref>
<ref id="ref-14"><label>14.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Srirangam</surname>, <given-names>V. K.</given-names></string-name>, <string-name><surname>Reddy</surname>, <given-names>A. A.</given-names></string-name>, <string-name><surname>Singh</surname>, <given-names>V.</given-names></string-name>, <string-name><surname>Shrivastava</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2019</year>). <article-title>Corpus creation and analysis for named entity recognition in Telugu-English code-mixed social media data</article-title>. <conf-name>Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics: Student Research Workshop</conf-name>, pp. <fpage>183</fpage>&#x2013;<lpage>189</lpage>. <publisher-loc>Florence, Italy</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-15"><label>15.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Sykes</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Grivas</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Grover</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Tobin</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Sudlow</surname>, <given-names>C.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Comparison of rule-based and neural network models for negation detection in radiology reports</article-title>. <source>Natural Language Engineering</source><italic>,</italic> <volume>27</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>203</fpage>&#x2013;<lpage>224</lpage>.</mixed-citation></ref>
<ref id="ref-16"><label>16.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Whitelaw</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Kehlenbeck</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Petrovic</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Ungar</surname>, <given-names>L.</given-names></string-name></person-group> (<year>2008</year>). <article-title>Web-scale named entity recognition</article-title>. <conf-name>Proceedings of the 17th ACM Conference on Information and Knowledge Management</conf-name>, pp. <fpage>123</fpage>&#x2013;<lpage>132</lpage>. <publisher-loc>Singapore</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-17"><label>17.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Jain</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Pennacchiotti</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Open entity extraction from web search query logs</article-title>. <conf-name>Proceedings of the 23rd International Conference on Computational Linguistics</conf-name>, pp. <fpage>510</fpage>&#x2013;<lpage>518</lpage>. <conf-loc>Beijing, China</conf-loc>.</mixed-citation></ref>
<ref id="ref-18"><label>18.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Fan</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Ma</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Yang</surname>, <given-names>Z.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Research on application of intelligent corpus annotation of entity extraction with construction of knowledge graph</article-title>. <source>Mathematical Problems in Engineering</source><italic>,</italic> <volume>2022</volume><italic>,</italic> <comment>2552331</comment>.</mixed-citation></ref>
<ref id="ref-19"><label>19.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Collobert</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Weston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Bottou</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Karlen</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Kavukcuoglu</surname>, <given-names>K.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2011</year>). <article-title>Natural language processing (almost) from scratch</article-title>. <source>Journal of Machine Learning Research</source><italic>,</italic> <volume>12</volume><italic>,</italic> <fpage>2493</fpage>&#x2013;<lpage>2537</lpage>.</mixed-citation></ref>
<ref id="ref-20"><label>20.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhao</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Yang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>L.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2016</year>). <article-title>ML-CNN: A novel deep learning based disease named entity recognition architecture</article-title>. <conf-name>Proceedings of the 2016 IEEE International Conference on Bioinformatics and Biomedicine</conf-name>, pp. <fpage>794</fpage>. <publisher-loc>Shenzhen, China</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-21"><label>21.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Gui</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Ma</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Q.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>Y. G.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2019</year>). <article-title>CNN-based Chinese NER with lexicon rethinking</article-title>. <conf-name>Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence</conf-name>, pp. <fpage>4982</fpage>&#x2013;<lpage>4988</lpage>. <publisher-loc>Macao, China</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-22"><label>22.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Cho</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Ha</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Park</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Park</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2020</year>). <article-title>Combinatorial feature embedding based on CNN and LSTM for biomedical named entity recognition</article-title>. <source>Journal of Biomedical Informatics</source><italic>,</italic> <volume>103</volume><italic>,</italic> <fpage>103381</fpage>; <pub-id pub-id-type="pmid">32004641</pub-id></mixed-citation></ref>
<ref id="ref-23"><label>23.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Tjong Kim Sang</surname>, <given-names>E. F.</given-names></string-name>, <string-name><surname>de Meulder</surname>, <given-names>F.</given-names></string-name></person-group> (<year>2003</year>). <article-title>Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition</article-title>. <conf-name>Proceedings of the Seventh Conference on Natural Language Learning at HLT-NAACL</conf-name>, pp. <fpage>142</fpage>&#x2013;<lpage>147</lpage>. <publisher-loc>Edmonton, AB, Canada</publisher-loc>, <publisher-name>Computational Linguistics in Flanders</publisher-name>.</mixed-citation></ref>
<ref id="ref-24"><label>24.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Walker</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Strassel</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Medero</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Maeda</surname>, <given-names>K.</given-names></string-name></person-group> (<year>2006</year>). <source>ACE 2005 multilingual training corpus</source>. <publisher-loc>Philadelphia, USA</publisher-loc>: <source>Linguistic Data Consortium</source>, <publisher-name>University of Pennsylvania</publisher-name>.</mixed-citation></ref>
<ref id="ref-25"><label>25.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hovy</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Marcus</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Palmer</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Ramshaw</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Weischedel</surname>, <given-names>R.</given-names></string-name></person-group> (<year>2006</year>). <article-title>OntoNotes: The 90% solution</article-title>. <conf-name>Proceedings of the Human Language Technology Conference of the NAACL, Companion Volume: Short Papers</conf-name>, pp. <fpage>57</fpage>&#x2013;<lpage>60</lpage>. <publisher-loc>Boulder, CO, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-26"><label>26.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Bach</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Huang</surname>, <given-names>Z.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Automated concatenation of embeddings for structured prediction</article-title>. <conf-name>Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing</conf-name>, pp. <fpage>2643</fpage>&#x2013;<lpage>2660</lpage>. <publisher-loc>Bangkok, Thailand</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-27"><label>27.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhou</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Learning from noisy labels for entity-centric information extraction</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>5381</fpage>&#x2013;<lpage>5392</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-28"><label>28.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Liu</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>Y. E.</given-names></string-name>, <string-name><surname>Monath</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Cotterell</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Sachan</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Autoregressive structured prediction with language models</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>993</fpage>&#x2013;<lpage>1005</lpage>. <publisher-loc>Abu Dhabi, United Arab Emirates</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-29"><label>29.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Schweter</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Akbik</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2020</year>). <article-title>FLERT: Document-level features for named entity recognition</article-title>. <source>arXiv preprint</source> <comment>arXiv:2011.06993</comment>.</mixed-citation></ref>
<ref id="ref-30"><label>30.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Ye</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Lin</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Packed levitated marker for entity and relation extraction</article-title>. <conf-name>Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>4904</fpage>&#x2013;<lpage>4917</lpage>. <publisher-loc>Dublin, Ireland</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-31"><label>31.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yamada</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>Asai</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Shindo</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Takeda</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Matsumoto</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2020</year>). <article-title>LUKE: Deep contextualized entity representations with entity-aware self-attention</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>6442</fpage>&#x2013;<lpage>6454</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-32"><label>32.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Bach</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Huang</surname>, <given-names>Z.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Improving named entity recognition by external context retrieving and cooperative learning</article-title>. <conf-name>Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing</conf-name>, pp. <fpage>1800</fpage>&#x2013;<lpage>1812</lpage>. <publisher-loc>Bangkok, Thailand</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-33"><label>33.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hanh</surname>, <given-names>T. T. H.</given-names></string-name>, <string-name><surname>Doucet</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Sidere</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Moreno</surname>, <given-names>J. G.</given-names></string-name>, <string-name><surname>Pollak</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Named entity recognition architecture combining contextual and global features</article-title>. <conf-name>Proceedings of the 23rd International Conference on Asia-Pacific Digital Libraries</conf-name>, pp. <fpage>264</fpage>&#x2013;<lpage>276</lpage>. <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-34"><label>34.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Shahzad</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Amin</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Esteves</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Ngonga Ngomo</surname>, <given-names>A. C.</given-names></string-name></person-group> (<year>2021</year>). <article-title>InferNER: An attentive model leveraging the sentence-level information for named entity recognition in Microblogs</article-title>. <conf-name>Proceedings of the 34th International Florida Artificial Intelligence Research Society Conference</conf-name>, <publisher-loc>Miami, FL, USA</publisher-loc>, <publisher-name>Library Press</publisher-name>.</mixed-citation></ref>
<ref id="ref-35"><label>35.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhong</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2021</year>). <article-title>A frustratingly easy approach for entity and relation extraction</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>50</fpage>&#x2013;<lpage>61</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-36"><label>36.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Shen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>R.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2023</year>). <article-title>PromptNER: Prompt locating and typing for named entity recognition</article-title>. <conf-name>Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics Volume 1: Long Papers</conf-name>, pp. <fpage>12492</fpage>&#x2013;<lpage>12507</lpage>. <conf-loc>Toronto, Canada</conf-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-37"><label>37.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Shen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Xie</surname>, <given-names>P.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>Parallel instance query network for named entity recognition</article-title>. <conf-name>Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>947</fpage>&#x2013;<lpage>961</lpage>. <publisher-loc>Dublin, Ireland</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-38"><label>38.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Shen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Song</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Lu</surname>, <given-names>W.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2023</year>). <article-title>DiffusionNER: Boundary diffusion for named entity recognition</article-title>. <conf-name>Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics Volume 1: Long Papers</conf-name>, pp. <fpage>3875</fpage>&#x2013;<lpage>3890</lpage>. <conf-loc>Toronto, Canada</conf-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-39"><label>39.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Feng</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Meng</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Han</surname>, <given-names>Q.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>F.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>A unified MRC framework for named entity recognition</article-title>. <conf-name>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>5849</fpage>&#x2013;<lpage>5859</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-40"><label>40.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Shen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Ma</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>W.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Locate and label: A two-stage identifier for nested named entity recognition</article-title>. <conf-name>Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing</conf-name>, pp. <fpage>2782</fpage>&#x2013;<lpage>2794</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-41"><label>41.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Jiang</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Karlsson</surname>, <given-names>B. F.</given-names></string-name></person-group> (<year>2021</year>). <article-title>BoningKnife: Joint entity mention detection and typing for nested NER via prior boundary knowledge</article-title>. <source>arXiv preprint</source> <comment>arXiv:2107.09429</comment>.</mixed-citation></ref>
<ref id="ref-42"><label>42.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Bohnet</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Poesio</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2020</year>). <article-title>Named entity recognition as dependency parsing</article-title>. <conf-name>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>6470</fpage>&#x2013;<lpage>6476</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-43"><label>43.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Shibuya</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Hovy</surname>, <given-names>E.</given-names></string-name></person-group> (<year>2020</year>). <article-title>Nested named entity recognition via second-best sequence learning and decoding</article-title>. <source>Transactions of the Association for Computational Linguistics</source><italic>,</italic> <volume>8</volume><italic>,</italic> <fpage>605</fpage>&#x2013;<lpage>620</lpage>.</mixed-citation></ref>
<ref id="ref-44"><label>44.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Meng</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Liang</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>F.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>Dice loss for data-imbalanced NLP tasks</article-title>. <conf-name>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>465</fpage>&#x2013;<lpage>476</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-45"><label>45.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhu</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>J.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Boundary smoothing for named entity recognition</article-title>. <conf-name>Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics</conf-name>, vol. <volume>1</volume>, pp. <fpage>7096</fpage>&#x2013;<lpage>7108</lpage>. <publisher-loc>Dublin, Ireland</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-46"><label>46.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Shen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Wan</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Chang</surname>, <given-names>T. H.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Hero-Gang neural model for named entity recognition</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>1924</fpage>&#x2013;<lpage>1936</lpage>. <publisher-loc>Seattle, WA, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-47"><label>47.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Xu</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Jie</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Lu</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Bing</surname>, <given-names>L.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Better feature integration for named entity recognition</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>3457</fpage>&#x2013;<lpage>3469</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-48"><label>48.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Fei</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>M.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>Unified named entity recognition as word-word relation classification</article-title>. <conf-name>Proceedings of the 36th AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>10965</fpage>&#x2013;<lpage>10973</lpage>. <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-49"><label>49.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Plaza-del Arco</surname>, <given-names>F. M.</given-names></string-name>, <string-name><surname>Molina-Gonz&#x00E1;lez</surname>, <given-names>M. D.</given-names></string-name>, <string-name><surname>Martin</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Ure&#x00F1;a-L&#x00F3;pez</surname>, <given-names>L. A.</given-names></string-name></person-group> (<year>2019</year>). <article-title>SINAI at SemEval-2019 task 6: Incorporating lexicon knowledge into SVM learning to identify and categorize offensive language in social media</article-title>. <conf-name>Proceedings of the 13th International Workshop on Semantic Evaluation</conf-name>, pp. <fpage>735</fpage>&#x2013;<lpage>738</lpage>. <publisher-loc>Minneapolis, MN, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-50"><label>50.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Lv</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Pan</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>Z.</given-names></string-name></person-group> (<year>2021</year>). <article-title>A novel Chinese entity relationship extraction method based on the bidirectional maximum entropy Markov model</article-title>. <source>Complexity</source><italic>,</italic> <volume>2021</volume><italic>,</italic> <fpage>1</fpage>&#x2013;<lpage>8</lpage>.</mixed-citation></ref>
<ref id="ref-51"><label>51.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Ma</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>W.</given-names></string-name></person-group> (<year>2023</year>). <article-title>A hybrid deep transfer learning model with kernel metric for COVID-19 pneumonia classification using chest CT images</article-title>. <source>IEEE/ACM Transactions on Computational Biology and Bioinformatics</source><italic>,</italic> <volume>20</volume><italic>(</italic><issue>4</issue><italic>),</italic> <fpage>2506</fpage>&#x2013;<lpage>2517</lpage>; <pub-id pub-id-type="pmid">36279353</pub-id></mixed-citation></ref>
<ref id="ref-52"><label>52.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhang</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Hou</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Xia</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2012</year>). <article-title>A novel convolution kernel model for Chinese relation extraction based on semantic feature and instances partition</article-title>. <conf-name>Proceedings of the 5th International Symposium on Computational Intelligence and Design</conf-name>, pp. <fpage>411</fpage>&#x2013;<lpage>414</lpage>. <publisher-loc>Hangzhou, China</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-53"><label>53.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Sun</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Han</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2014</year>). <article-title>A feature-enriched tree kernel for relation extraction</article-title>. <conf-name>Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>61</fpage>&#x2013;<lpage>67</lpage>. <publisher-loc>Baltimore, Maryland</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-54"><label>54.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Sobhana</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Ghosh</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Mitra</surname>, <given-names>P.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Entity relation extraction from geological text using conditional random fields and subsequence kernels</article-title>. <conf-name>Proceedings of the Annual IEEE India Conference</conf-name>, pp. <fpage>832</fpage>&#x2013;<lpage>840</lpage>. <publisher-loc>Kochi, Kerala, India</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-55"><label>55.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Luo</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>W.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2018</year>). <article-title>Short-term wind speed forecasting via stacked extreme learning machine with generalized correntropy</article-title>. <source>IEEE Transactions on Industrial Informatics</source><italic>,</italic> <volume>14</volume><italic>(</italic><issue>11</issue><italic>),</italic> <fpage>4963</fpage>&#x2013;<lpage>4971</lpage>.</mixed-citation></ref>
<ref id="ref-56"><label>56.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Kambar</surname>, <given-names>M. E. Z. N.</given-names></string-name>, <string-name><surname>Esmaeilzadeh</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Heidari</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2022</year>). <article-title>A survey on deep learning techniques for joint named entities and relation extraction</article-title>. <conf-name>Proceedings of the IEEE World AI IoT Congress</conf-name>, pp. <fpage>218</fpage>&#x2013;<lpage>224</lpage>. <publisher-loc>Seattle, WA, USA</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-57"><label>57.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Socher</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Huval</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Manning</surname>, <given-names>C. D.</given-names></string-name>, <string-name><surname>Ng</surname>, <given-names>A. Y.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Semantic compositionality through recursive matrix-vector spaces</article-title>. <conf-name>Proceedings of the Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning</conf-name>, pp. <fpage>1201</fpage>&#x2013;<lpage>1211</lpage>. <publisher-loc>Jeju Island, Korea</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-58"><label>58.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Kim</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2014</year>). <article-title>Convolutional neural networks for sentence classification</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>1746</fpage>&#x2013;<lpage>1751</lpage>. <publisher-loc>Doha, Qatar</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-59"><label>59.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Long</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Shen</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Zhou</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Yao</surname>, <given-names>L.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>Self-attention enhanced selective gate with entity-aware embedding for distantly supervised relation extraction</article-title>. <conf-name>Proceedings of the 34th AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>8269</fpage>&#x2013;<lpage>8276</lpage>. <publisher-loc>New York, USA</publisher-loc>, <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-60"><label>60.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>dos Santos</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Xiang</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Zhou</surname>, <given-names>B.</given-names></string-name></person-group> (<year>2015</year>). <article-title>Classifying relations by ranking with convolutional neural networks</article-title>. <conf-name>Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing</conf-name>, pp. <fpage>626</fpage>&#x2013;<lpage>634</lpage>. <publisher-loc>Beijing, China</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-61"><label>61.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yin</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Yin</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>Z.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2019</year>). <article-title>Relation classification in scientific papers based on convolutional neural network</article-title>. <conf-name>Proceedings of the 8th CCF International Conference on Natural Language Processing and Chinese Computing</conf-name>, pp. <fpage>242</fpage>&#x2013;<lpage>253</lpage>. <publisher-loc>Dunhuang, China</publisher-loc>, <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-62"><label>62.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Feng</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Hao</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Jin</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2017</year>). <article-title>Relation extraction with deep reinforcement learning</article-title>. <source>IEICE Transactions on Information and Systems</source><italic>,</italic> <volume>100</volume><italic>(</italic><issue>8</issue><italic>),</italic> <fpage>1893</fpage>&#x2013;<lpage>1902</lpage>.</mixed-citation></ref>
<ref id="ref-63"><label>63.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Takamatsu</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Sato</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>Nakagawa</surname>, <given-names>H.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Reducing wrong labels in distant supervision for relation extraction</article-title>. <conf-name>Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>721</fpage>&#x2013;<lpage>729</lpage>. <publisher-loc>Jeju Island, Korea</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-64"><label>64.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hoffmann</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Ling</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Zettlemoyer</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Weld</surname>, <given-names>D. S.</given-names></string-name></person-group> (<year>2011</year>). <article-title>Knowledge-based weak supervision for information extraction of overlapping relations</article-title>. <conf-name>Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>541</fpage>&#x2013;<lpage>550</lpage>. <publisher-loc>Portland, OR, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-65"><label>65.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Surdeanu</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Tibshirani</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Nallapati</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Manning</surname>, <given-names>C. D.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Multi-instance multi-label learning for relation extraction</article-title>. <conf-name>Proceedings of the Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning</conf-name>, pp. <fpage>455</fpage>&#x2013;<lpage>465</lpage>. <publisher-loc>Jeju Island, Korea</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-66"><label>66.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Helwe</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Elbassuoni</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Al Zaatari</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>El-Hajj</surname>, <given-names>W.</given-names></string-name></person-group> (<year>2019</year>). <article-title>Assessing Arabic weblog credibility via deep co-learning</article-title>. <conf-name>Proceedings of the Fourth Arabic Natural Language Processing Workshop</conf-name>, pp. <fpage>130</fpage>&#x2013;<lpage>136</lpage>. <publisher-loc>Florence, Italy</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-67"><label>67.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Carlson</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Betteridge</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>R. C.</given-names></string-name>, <string-name><surname>Hruschka</surname>, <given-names>E. R.</given-names>
<suffix>Jr</suffix></string-name>, <string-name><surname>Mitchell</surname>, <given-names>T. M.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Coupled semi-supervised learning for information extraction</article-title>. <conf-name>Proceedings of the 3rd ACM International Conference on Web Search and Data Mining</conf-name>, pp. <fpage>101</fpage>&#x2013;<lpage>110</lpage>. <publisher-loc>New York, NY, USA</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-68"><label>68.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Jianshu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Guang</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Chunyun</surname>, <given-names>Z.</given-names></string-name></person-group> (<year>2014</year>). <article-title>A bootstrapping and MV-RNN mixed method for relation extraction</article-title>. <conf-name>Proceedings of the 4th IEEE International Conference on Network Infrastructure and Digital Content</conf-name>, pp. <fpage>117</fpage>&#x2013;<lpage>120</lpage>. <publisher-loc>Beijing, China</publisher-loc>, <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref-69"><label>69.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Mao</surname>, <given-names>X. L.</given-names></string-name>, <string-name><surname>Heyan</surname>, <given-names>H.</given-names></string-name></person-group> (<year>2022</year>). <article-title>A semi-supervised transfer learning framework for low resource entity and relation extraction in scientific domain</article-title>. <conf-name>Proceedings of the 3rd Workshop on Extraction and Evaluation of Knowledge Entities from Scientific Documents</conf-name>, pp. <fpage>41</fpage>&#x2013;<lpage>47</lpage>. <publisher-loc>Cologne, Germany</publisher-loc>, <publisher-name>CEUR-WS</publisher-name>.</mixed-citation></ref>
<ref id="ref-70"><label>70.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hu</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Wen</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Yu</surname>, <given-names>S.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>SelfORE: Self-supervised relational feature learning for open relation extraction</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>3673</fpage>&#x2013;<lpage>3682</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-71"><label>71.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yan</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Okazaki</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Matsuo</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Yang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Ishizuka</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2009</year>). <article-title>Unsupervised relation extraction by mining Wikipedia texts using information from the web</article-title>. <conf-name>Proceedings of the Joint Conference of the 47th Annual Meeting of the ACL and the 4th International Joint Conference on Natural Language Processing of the AFNLP</conf-name>, pp. <fpage>1021</fpage>&#x2013;<lpage>1029</lpage>. <publisher-loc>Singapore</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-72"><label>72.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Bollegala</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Matsuo</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Ishizuka</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2009</year>). <article-title>Measuring the similarity between implicit semantic relations using web search engines</article-title>. <conf-name>Proceedings of the 2nd ACM International Conference on Web Search and Data Mining</conf-name>, pp. <fpage>104</fpage>&#x2013;<lpage>113</lpage>. <publisher-loc>Barcelona, Spain</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-73"><label>73.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Bollegala</surname>, <given-names>D. T.</given-names></string-name>, <string-name><surname>Matsuo</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Ishizuka</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Relational duality: Unsupervised extraction of semantic relations between entities on the web</article-title>. <conf-name>Proceedings of the 19th International Conference of World Wide Web</conf-name>, pp. <fpage>151</fpage>&#x2013;<lpage>160</lpage>. <publisher-loc>Raleigh, NC, USA</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-74"><label>74.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Riedel</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Yao</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>McCallum</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Modeling relations and their mentions without labeled text</article-title>. <conf-name>Proceedings of the European Conference on Machine Learning and Knowledge Discovery in Databases</conf-name>, pp. <fpage>148</fpage>&#x2013;<lpage>163</lpage>. <publisher-loc>Barcelona, Spain</publisher-loc>, <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-75"><label>75.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Gardent</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Shimorina</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Narayan</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Perez-Beltrachini</surname>, <given-names>L.</given-names></string-name></person-group> (<year>2017</year>). <article-title>Creating training corpora for NLG micro-planners</article-title>. <conf-name>Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics</conf-name>, vol. <volume>1</volume>, pp. <fpage>179</fpage>&#x2013;<lpage>188</lpage>. <publisher-loc>Vancouver, Canada</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-76"><label>76.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Tang</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Mao</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>Y.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>UniRel: Unified representation and interaction for joint relational triple extraction</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>7087</fpage>&#x2013;<lpage>7099</lpage>. <publisher-loc>Abu Dhabi, United Arab Emirates</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-77"><label>77.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Huguet Cabot</surname>, <given-names>P. L.</given-names></string-name>, <string-name><surname>Navigli</surname>, <given-names>R.</given-names></string-name></person-group> (<year>2021</year>). <article-title>REBEL: Relation extraction by end-to-end language generation</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>2370</fpage>&#x2013;<lpage>2381</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-78"><label>78.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhao</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Kang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Adjacency list oriented relational fact extraction via adaptive multi-task learning</article-title>. <conf-name>Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021</conf-name>, pp. <fpage>3075</fpage>&#x2013;<lpage>3087</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-79"><label>79.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yan</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Fu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Q.</given-names></string-name>, <string-name><surname>Wei</surname>, <given-names>Z.</given-names></string-name></person-group> (<year>2021</year>). <article-title>A partition filter network for joint entity and relation extraction</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>185</fpage>&#x2013;<lpage>197</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-80"><label>80.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Sui</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Zeng</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>J.</given-names></string-name></person-group> (<year>2023</year>). <article-title>Joint entity and relation extraction with set prediction networks</article-title>. <source>IEEE Transactions on Neural Networks and Learning Systems</source><italic>,</italic> pp. <fpage>1</fpage>&#x2013;<lpage>12</lpage>. <pub-id pub-id-type="doi">10.1109/TNNLS.2023.3264735</pub-id>; <pub-id pub-id-type="pmid">37067968</pub-id></mixed-citation></ref>
<ref id="ref-81"><label>81.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Dong</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Yang</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Luan</surname>, <given-names>B.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>TDEER: An efficient translating decoding schema for joint extraction of entities and relations</article-title>, <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>8055</fpage>&#x2013;<lpage>8064</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-82"><label>82.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhao</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Cheng</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Gao</surname>, <given-names>K.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Representation iterative fusion based on heterogeneous graph neural network for joint entity and relation extraction</article-title>. <source>Knowledge-Based Systems</source><italic>,</italic> <volume>219</volume><italic>,</italic> <fpage>106888</fpage>.</mixed-citation></ref>
<ref id="ref-83"><label>83.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Yu</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Zhu</surname>, <given-names>H.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>TPLinker: Single-stage joint extraction of entities and relations through token pair linking</article-title>. <conf-name>Proceedings of the 28th International Conference on Computational Linguistics</conf-name>, pp. <fpage>1572</fpage>&#x2013;<lpage>1582</lpage>. <publisher-loc>Barcelona, Spain</publisher-loc>, <publisher-name>International Committee on Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-84"><label>84.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>J.</given-names></string-name></person-group> (<year>2020</year>). <article-title>RH-Net: Improving neural relation extraction via reinforcement learning and hierarchical relational searching</article-title>. <source>arXiv preprint</source> <comment>arXiv:2010.14255</comment>.</mixed-citation></ref>
<ref id="ref-85"><label>85.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Wei</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Su</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Tian</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Chang</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2020</year>). <article-title>A novel cascade binary tagging framework for relational triple extraction</article-title>. <conf-name>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics</conf-name>, pp. <fpage>1476</fpage>&#x2013;<lpage>1488</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-86"><label>86.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Sun</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Mensah</surname>, <given-names>R. Z. S.</given-names></string-name>, <string-name><surname>Mao</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2020</year>). <article-title>Recurrent interaction network for jointly extracting entities and classifying relations</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>3722</fpage>&#x2013;<lpage>3732</lpage>. <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-87"><label>87.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Ye</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Deng</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>C.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Contrastive triple extraction with generative transformer</article-title>. <conf-name>Proceedings of the Thirty-Fifth AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>14257</fpage>&#x2013;<lpage>14265</lpage>. <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-88"><label>88.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yu</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Shu</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>Y.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>Joint extraction of entities and relations based on a novel decomposition strategy</article-title>. <conf-name>Proceedings of the 24th European Conference on Artificial Intelligence</conf-name>, pp. <fpage>2282</fpage>&#x2013;<lpage>2289</lpage>. <publisher-loc>Santiago de Compostela, Spain</publisher-loc>, <publisher-name>IOS Press</publisher-name>.</mixed-citation></ref>
<ref id="ref-89"><label>89.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Han</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>C.</given-names></string-name></person-group> (<year>2020</year>). <article-title>A survey on deep learning for named entity recognition</article-title>. <source>IEEE Transactions on Knowledge and Data Engineering</source><italic>,</italic> <volume>34</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>50</fpage>&#x2013;<lpage>70</lpage>.</mixed-citation></ref>
<ref id="ref-90"><label>90.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Cheng</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Xia</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>L.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>A review of Chinese named entity recognition</article-title>. <source>KSII Transactions on Internet &#x0026; Information Systems</source><italic>,</italic> <volume>15</volume><italic>(</italic><issue>6</issue><italic>),</italic> <fpage>2012</fpage>&#x2013;<lpage>2030</lpage>.</mixed-citation></ref>
<ref id="ref-91"><label>91.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Safranchik</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Bach</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2020</year>). <article-title>Weakly supervised sequence tagging from noisy rules</article-title>. <conf-name>Proceedings of the 34th AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>5570</fpage>&#x2013;<lpage>5578</lpage>. <publisher-loc>New York, NY, USA</publisher-loc>, <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-92"><label>92.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Khan</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Daud</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Shahzad</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Amjad</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Banjar</surname>, <given-names>A.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>Named entity recognition using conditional random fields</article-title>. <source>Applied Sciences</source><italic>,</italic> <volume>12</volume><italic>(</italic><issue>13</issue><italic>),</italic> <fpage>6391</fpage>.</mixed-citation></ref>
<ref id="ref-93"><label>93.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Shen</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Linden: Linking named entities with knowledge base via semantic knowledge</article-title>. <conf-name>Proceedings of the 21st International Conference on World Wide Web</conf-name>, pp. <fpage>449</fpage>&#x2013;<lpage>458</lpage>. <publisher-loc>Lyon, France</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-94"><label>94.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zheng</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Huang</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Zhu</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Learning to link entities with knowledge base</article-title>. <conf-name>Proceedings of the Annual Conference of the North American Chapter of the Association of Computational Linguistics</conf-name>, pp. <fpage>483</fpage>&#x2013;<lpage>491</lpage>. <publisher-loc>Los Angeles, CA, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-95"><label>95.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Mihaljevi&#x0107;</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Santamar&#x00ED;a</surname>, <given-names>L.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Disambiguation of author entities in ADS using supervised learning and graph theory methods</article-title>. <source>Scientometrics</source><italic>,</italic> <volume>126</volume><italic>(</italic><issue>5</issue><italic>),</italic> <fpage>3893</fpage>&#x2013;<lpage>3917</lpage>.</mixed-citation></ref>
<ref id="ref-96"><label>96.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Emami</surname>, <given-names>H.</given-names></string-name></person-group> (<year>2019</year>). <article-title>A graph-based approach to person name disambiguation in web</article-title>. <source>ACM Transactions on Management Information Systems</source><italic>,</italic> <volume>10</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>1</fpage>&#x2013;<lpage>25</lpage>.</mixed-citation></ref>
<ref id="ref-97"><label>97.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Mitchell</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Strassel</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Huang</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Zakhary</surname>, <given-names>R.</given-names></string-name></person-group> (<year>2005</year>). <source>ACE 2004 multilingual training corpus</source>. <publisher-loc>Philadelphia, USA</publisher-loc>: <publisher-name>Linguistic Data Consortium, University of Pennsylvania</publisher-name>.</mixed-citation></ref>
<ref id="ref-98"><label>98.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hoffart</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Yosef</surname>, <given-names>M. A.</given-names></string-name>, <string-name><surname>Bordino</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>F&#x00FC;rstenau</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Pinkal</surname>, <given-names>M.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2011</year>). <article-title>Robust disambiguation of named entities in text</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>782</fpage>&#x2013;<lpage>792</lpage>. <publisher-loc>Edinburgh, Scotland, UK</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-99"><label>99.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Ayoola</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Fisher</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Pierleoni</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Improving entity disambiguation by reasoning over a knowledge base</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>2899</fpage>&#x2013;<lpage>2912</lpage>. <publisher-loc>Seattle, WA, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-100"><label>100.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yamada</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>Washio</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Shindo</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Matsumoto</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Global entity disambiguation with BERT</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies</conf-name>, pp. <fpage>3264</fpage>&#x2013;<lpage>3271</lpage>. <publisher-loc>Seattle, WA, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-101"><label>101.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Ayoola</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Tyagi</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Fisher</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Christodoulopoulos</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Pierleoni</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2022</year>). <article-title>ReFinED: An efficient zero-shot-capable approach to end-to-end entity linking</article-title>. <conf-name>Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies: Industry Track</conf-name>, pp. <fpage>209</fpage>&#x2013;<lpage>220</lpage>. <publisher-loc>Seattle, WA, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-102"><label>102.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Tedeschi</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Conia</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Cecconi</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Navigli</surname>, <given-names>R.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Named entity recognition for entity linking: What works and what&#x2019;s next</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing</conf-name>, pp. <fpage>2584</fpage>&#x2013;<lpage>2596</lpage>. <publisher-loc>Punta Cana, Dominican Republic</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-103"><label>103.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>de Cao</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Izacard</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Riedel</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Petroni</surname>, <given-names>F.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Autoregressive entity retrieval</article-title>. <conf-name>Proceedings of the 9th International Conference on Learning Representations</conf-name>, pp. <fpage>1</fpage>&#x2013;<lpage>20</lpage>.</mixed-citation></ref>
<ref id="ref-104"><label>104.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Mulang&#x2019;</surname>, <given-names>I. O.</given-names></string-name>, <string-name><surname>Singh</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Prabhu</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Nadgeri</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Hoffart</surname>, <given-names>J.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2020</year>). <article-title>Evaluating the impact of knowledge graph context on entity disambiguation models</article-title>. <conf-name>Proceedings of the 29th ACM International Conference on Information &#x0026; Knowledge Management</conf-name>, pp. <fpage>2157</fpage>&#x2013;<lpage>2160</lpage>. <publisher-loc>Ireland</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-105"><label>105.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Raiman</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Raiman</surname>, <given-names>O.</given-names></string-name></person-group> (<year>2018</year>). <article-title>DeepType: Multilingual entity linking by neural type system evolution</article-title>. <conf-name>Proceedings of the 32nd AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>5406</fpage>&#x2013;<lpage>5413</lpage>. <publisher-loc>New Orleans, LA, USA</publisher-loc>, <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-106"><label>106.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yamada</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>Shindo</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Takeda</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Takefuji</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2017</year>). <article-title>Learning distributed representations of texts and entities from knowledge base</article-title>. <source>Transactions of the Association for Computational Linguistics</source><italic>,</italic> <volume>5</volume><italic>,</italic> <fpage>397</fpage>&#x2013;<lpage>411</lpage>.</mixed-citation></ref>
<ref id="ref-107"><label>107.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Gu</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Lin</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Tang</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Zhuang</surname>, <given-names>Y.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2019</year>). <article-title>Learning dynamic context augmentation for global entity linking</article-title>. <conf-name>Proceedings of the Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing</conf-name>, pp. <fpage>271</fpage>&#x2013;<lpage>281</lpage>. <publisher-loc>Hong Kong, China</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-108"><label>108.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Gruber</surname>, <given-names>T. R.</given-names></string-name></person-group> (<year>1993</year>). <article-title>A translation approach to portable ontology specifications</article-title>. <source>Knowledge Acquisition</source><italic>,</italic> <volume>5</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>199</fpage>&#x2013;<lpage>220</lpage>.</mixed-citation></ref>
<ref id="ref-109"><label>109.</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Liu</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2019</year>). <source>Enhancing ontology learning with machine learning and natural language processing techniques. (Ph.D. Thesis)</source>. <publisher-name>Rensselaer Polytechnic Institute Troy</publisher-name>, <publisher-loc>New York, USA</publisher-loc>.</mixed-citation></ref>
<ref id="ref-110"><label>110.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Shamsfard</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Barforoush</surname>, <given-names>A. A.</given-names></string-name></person-group> (<year>2004</year>). <article-title>Learning ontologies from natural language texts</article-title>. <source>International Journal of Human-Computer Studies</source><italic>,</italic> <volume>60</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>17</fpage>&#x2013;<lpage>63</lpage>.</mixed-citation></ref>
<ref id="ref-111"><label>111.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Cimiano</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Staab</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2005</year>). <article-title>Learning concept hierarchies from text with a guided hierarchical clustering algorithm</article-title>. <conf-name>Proceedings of the ICML Workshop on Learning and Extending Lexical Ontologies with Machine Learning Methods</conf-name>, pp. <fpage>1</fpage>&#x2013;<lpage>10</lpage>. <publisher-loc>New York, NY, USA</publisher-loc>, <publisher-name>CiteSeer</publisher-name>.</mixed-citation></ref>
<ref id="ref-112"><label>112.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Chen</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Zhu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Yao</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2003</year>). <article-title>Automatic learning field words by bootstrapping</article-title>. <conf-name>Proceedings of China National Conference on Computational Linguistics</conf-name>, pp. <fpage>67</fpage>&#x2013;<lpage>72</lpage>. <publisher-loc>Beijing, China</publisher-loc>, <publisher-name>Tsinghua University Press</publisher-name>.</mixed-citation></ref>
<ref id="ref-113"><label>113.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ji</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Zhao</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Xiao</surname>, <given-names>G.</given-names></string-name></person-group> (<year>2009</year>). <article-title>Chinese document re-ranking based on automatically acquired term resource</article-title>. <source>Language Resources and Evaluation</source><italic>,</italic> <volume>43</volume><italic>(</italic><issue>4</issue><italic>),</italic> <fpage>385</fpage>&#x2013;<lpage>406</lpage>.</mixed-citation></ref>
<ref id="ref-114"><label>114.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ma</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Yongjun</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Zhijian</surname>, <given-names>W.</given-names></string-name></person-group> (<year>2015</year>). <article-title>Multi-topic extraction algorithm based on concept clusters</article-title>. <source>CAAI Transactions on Intelligent Systems</source><italic>,</italic> <volume>10</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>261</fpage>&#x2013;<lpage>266</lpage>.</mixed-citation></ref>
<ref id="ref-115"><label>115.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Hagiwara</surname>, <given-names>M.</given-names></string-name></person-group> (<year>2008</year>). <article-title>A supervised learning approach to automatic synonym identification based on distributional features</article-title>. <conf-name>Proceedings of the 46th Annual Meeting of the Association for Computational Linguistics on Human Language Technologies: Student Research Workshop</conf-name>, pp. <fpage>1</fpage>&#x2013;<lpage>6</lpage>. <publisher-loc>Columbus, Ohio, USA</publisher-loc>, <publisher-name>Association for Computational Linguistics</publisher-name>.</mixed-citation></ref>
<ref id="ref-116"><label>116.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Cimiano</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Hotho</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Stumme</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Tane</surname>, <given-names>J.</given-names></string-name></person-group> (<year>2004</year>). <article-title>Conceptual knowledge processing with formal concept analysis and ontologies</article-title>, <conf-name>Proceedings of the 2nd International Conference on Formal Concept Analysi</conf-name>, pp. <fpage>189</fpage>&#x2013;<lpage>207</lpage>. <publisher-loc>Sydney, Australia</publisher-loc>, <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-117"><label>117.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Paukkeri</surname>, <given-names>M. S.</given-names></string-name>, <string-name><surname>Garc&#x00ED;a-Plaza</surname>, <given-names>A. P.</given-names></string-name>, <string-name><surname>Fresno</surname>, <given-names>V.</given-names></string-name>, <string-name><surname>Unanue</surname>, <given-names>R. M.</given-names></string-name>, <string-name><surname>Honkela</surname>, <given-names>T.</given-names></string-name></person-group> (<year>2012</year>). <article-title>Learning a taxonomy from a set of text documents</article-title>. <source>Applied Soft Computing</source><italic>,</italic> <volume>12</volume><italic>(</italic><issue>3</issue><italic>),</italic> <fpage>1138</fpage>&#x2013;<lpage>1148</lpage>.</mixed-citation></ref>
<ref id="ref-118"><label>118.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Gabrilovich</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Markovitch</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2007</year>). <article-title>Computing semantic relatedness using Wikipedia-based explicit semantic analysis</article-title>. <conf-name>Proceedings of the 20th International Joint Conference on Artifical Intelligence</conf-name>, pp. <fpage>1606</fpage>&#x2013;<lpage>1611</lpage>. <publisher-loc>Hyderabad, India</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-119"><label>119.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Lao</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Cohen</surname>, <given-names>W. W.</given-names></string-name></person-group> (<year>2010</year>). <article-title>Relational retrieval using a combination of path-constrained random walks</article-title>. <source>Machine Learning</source><italic>,</italic> <volume>81</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>53</fpage>&#x2013;<lpage>67</lpage>.</mixed-citation></ref>
<ref id="ref-120"><label>120.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Lin</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Xie</surname>, <given-names>R.</given-names></string-name></person-group> (<year>2016</year>). <article-title>Knowledge representation learning: A review</article-title>. <source>Journal of Computer Research and Development</source><italic>,</italic> <volume>53</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>247</fpage>&#x2013;<lpage>261</lpage>.</mixed-citation></ref>
<ref id="ref-121"><label>121.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Bordes</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Weston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Collobert</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Bengio</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2011</year>). <article-title>Learning structured embeddings of knowledge bases</article-title>, <conf-name>Proceedings of the 25th AAAI Conference on Artificial Intelligence</conf-name>, pp. <fpage>301</fpage>&#x2013;<lpage>306</lpage>. <publisher-loc>San Francisco, CA, USA</publisher-loc>, <publisher-name>Association for the Advance of Artificial Intelligence</publisher-name>.</mixed-citation></ref>
<ref id="ref-122"><label>122.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Socher</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Manning</surname>, <given-names>C. D.</given-names></string-name>, <string-name><surname>Ng</surname>, <given-names>A.</given-names></string-name></person-group> (<year>2013</year>). <article-title>Reasoning with neural tensor networks for knowledge base completion</article-title>. <conf-name>Proceedings of the 27th Annual Conference on Neural Information Processing Systems</conf-name>, pp. <fpage>926</fpage>&#x2013;<lpage>934</lpage>. <publisher-loc>Lake Tahoe, NV, USA</publisher-loc>, <publisher-name>Curran Associates, Inc</publisher-name>.</mixed-citation></ref>
<ref id="ref-123"><label>123.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Bordes</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Glorot</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Weston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Bengio</surname>, <given-names>Y.</given-names></string-name></person-group> (<year>2014</year>). <article-title>A semantic matching energy function for learning with multi-relational data: Application to word-sense disambiguation</article-title>. <source>Machine Learning</source><italic>,</italic> <volume>94</volume><italic>(</italic><issue>2</issue><italic>),</italic> <fpage>233</fpage>&#x2013;<lpage>259</lpage>.</mixed-citation></ref>
<ref id="ref-124"><label>124.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Jenatton</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Roux</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Bordes</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Obozinski</surname>, <given-names>G. R.</given-names></string-name></person-group> (<year>2012</year>). <article-title>A latent factor model for highly multi-relational data</article-title>. <conf-name>Proceedings of the 26th Annual Conference on Neural Information Processing Systems</conf-name>, pp. <fpage>3167</fpage>&#x2013;<lpage>3175</lpage>. <publisher-loc>Lake Tahoe, NV, USA</publisher-loc>, <publisher-name>Curran Associates, Inc</publisher-name>.</mixed-citation></ref>
<ref id="ref-125"><label>125.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Nickel</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Tresp</surname>, <given-names>V.</given-names></string-name>, <string-name><surname>Kriegel</surname>, <given-names>H. P.</given-names></string-name>, <string-name><surname>et</surname> <given-names>al.</given-names></string-name></person-group> (<year>2011</year>). <article-title>A three-way model for collective learning on multi-relational data</article-title>. <conf-name>Proceedings of the 28th International Conference on Machine Learning</conf-name>, pp. <fpage>809</fpage>&#x2013;<lpage>816</lpage>. <publisher-loc>Bellevue, WA, USA</publisher-loc>, <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-126"><label>126.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Bordes</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Usunier</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Garcia-Duran</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Weston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Yakhnenko</surname>, <given-names>O.</given-names></string-name></person-group> (<year>2013</year>). <article-title>Translating embeddings for modeling multi-relational data</article-title>. <conf-name>Proceedings of the 27th Annual Conference on Neural Information Processing Systems</conf-name>, pp. <fpage>2787</fpage>&#x2013;<lpage>2795</lpage>. <publisher-loc>Lake Tahoe, NV, USA</publisher-loc>, <publisher-name>Curran Associates, Inc</publisher-name>.</mixed-citation></ref>
<ref id="ref-127"><label>127.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Karetnikov</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Ehrlinger</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Geist</surname>, <given-names>V.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Enhancing TransE to predict process behavior in temporal knowledge graphs</article-title>. <conf-name>Proceedings of the 33rd International Conference on Database and Expert Systems Applications Workshops</conf-name>, pp. <fpage>369</fpage>&#x2013;<lpage>374</lpage>. <publisher-loc>Vienna, Austria</publisher-loc>, <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-128"><label>128.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yang</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>J.</given-names></string-name></person-group> (<year>2021</year>). <article-title>Knowledge graph representation learning as groupoid: Unifying TransE, RotatE, QuatE, ComplEx</article-title>. <conf-name>Proceedings of the 30th ACM International Conference on Information &#x0026; Knowledge Management</conf-name>, pp. <fpage>2311</fpage>&#x2013;<lpage>2320</lpage>. <publisher-name>Association for Computing Machinery</publisher-name>.</mixed-citation></ref>
<ref id="ref-129"><label>129.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>El-Gohary</surname>, <given-names>N.</given-names></string-name></person-group> (<year>2023</year>). <article-title>Deep learning-based relation extraction and knowledge graph-based representation of construction safety requirements</article-title>. <source>Automation in Construction</source><italic>,</italic> <volume>147</volume><italic>,</italic> <fpage>104696</fpage>.</mixed-citation></ref>
<ref id="ref-130"><label>130.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Krompa&#x00DF;</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Nickel</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Tresp</surname>, <given-names>V.</given-names></string-name></person-group> (<year>2014</year>). <article-title>Probabilistic latent-factor database models</article-title>. <conf-name>Proceedings of the 1st Workshop on Linked Data for Knowledge Discovery Co-located with European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases</conf-name>, pp. <fpage>74</fpage>&#x2013;<lpage>83</lpage>. <publisher-loc>Nancy</publisher-loc>, <publisher-name>France, CEUR-WS</publisher-name>.</mixed-citation></ref>
<ref id="ref-131"><label>131.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ji</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Cassidy</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>Q.</given-names></string-name>, <string-name><surname>Tamang</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2014</year>). <article-title>Tackling representation, annotation and classification challenges for temporal knowledge base population</article-title>. <source>Knowledge and Information Systems</source><italic>,</italic> <volume>41</volume><italic>(</italic><issue>3</issue><italic>),</italic> <fpage>611</fpage>&#x2013;<lpage>646</lpage>.</mixed-citation></ref>
<ref id="ref-132"><label>132.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Peng</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Lu</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Shang</surname>, <given-names>X.</given-names></string-name></person-group> (<year>2020</year>). <article-title>A survey of network representation learning methods for link prediction in biological network</article-title>. <source>Current Pharmaceutical Design</source><italic>,</italic> <volume>26</volume><italic>(</italic><issue>26</issue><italic>),</italic> <fpage>3076</fpage>&#x2013;<lpage>3084</lpage>; <pub-id pub-id-type="pmid">31951161</pub-id></mixed-citation></ref>
<ref id="ref-133"><label>133.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Xiong</surname>, <given-names>N. N.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Learning knowledge graph embedding with heterogeneous relation attention networks</article-title>. <source>IEEE Transactions on Neural Networks and Learning Systems</source><italic>,</italic> <volume>33</volume><italic>(</italic><issue>8</issue><italic>),</italic> <fpage>3961</fpage>&#x2013;<lpage>3973</lpage>; <pub-id pub-id-type="pmid">33606639</pub-id></mixed-citation></ref>
<ref id="ref-134"><label>134.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Liu</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Xiong</surname>, <given-names>N. N.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Multi-scale dynamic convolutional network for knowledge graph embedding</article-title>. <source>IEEE Transactions on Knowledge and Data Engineering</source><italic>,</italic> <volume>34</volume><italic>(</italic><issue>5</issue><italic>),</italic> <fpage>2335</fpage>&#x2013;<lpage>2347</lpage>.</mixed-citation></ref>
<ref id="ref-135"><label>135.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wu</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Song</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Ge</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Ge</surname>, <given-names>T.</given-names></string-name></person-group> (<year>2022</year>). <article-title>Link prediction on complex networks: An experimental survey</article-title>. <source>Data Science and Engineering</source><italic>,</italic> <volume>7</volume><italic>(</italic><issue>3</issue><italic>),</italic> <fpage>253</fpage>&#x2013;<lpage>278</lpage>; <pub-id pub-id-type="pmid">35754861</pub-id></mixed-citation></ref>
<ref id="ref-136"><label>136.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Yang</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Hu</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Sun</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>Y.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2021</year>). <article-title>Inductive link prediction with interactive structure learning on attributed graph</article-title>. <conf-name>Proceedings of the European Conference on Machine Learning and Knowledge Discovery in Database</conf-name>, pp. <fpage>383</fpage>&#x2013;<lpage>398</lpage>. <publisher-loc>Bilbao, Spain</publisher-loc>, <publisher-name>Springer</publisher-name>.</mixed-citation></ref>
<ref id="ref-137"><label>137.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wang</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Zhang</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Gao</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2022</year>). <article-title>A novel knowledge graph development for industry design: A case study on indirect coal liquefaction process</article-title>. <source>Computers in Industry</source><italic>,</italic> <volume>139</volume><italic>,</italic> <fpage>103647</fpage>.</mixed-citation></ref>
<ref id="ref-138"><label>138.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yuan</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Li</surname>, <given-names>H.</given-names></string-name></person-group> (<year>2023</year>). <article-title>Research on the standardization model of data semantics in the knowledge graph construction of oil &#x0026; gas industry</article-title>. <source>Computer Standards &#x0026; Interfaces</source><italic>,</italic> <volume>84</volume><italic>,</italic> <fpage>103705</fpage>.</mixed-citation></ref>
<ref id="ref-139"><label>139.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yin</surname>, <given-names>Z.</given-names></string-name>, <string-name><surname>Shi</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Yuan</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Tan</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Xu</surname>, <given-names>S.</given-names></string-name></person-group> (<year>2023</year>). <article-title>A study on a knowledge graph construction method of safety reports for process industries</article-title>. <source>Processes</source><italic>,</italic> <volume>11</volume><italic>(</italic><issue>1</issue><italic>),</italic> <fpage>146</fpage>.</mixed-citation></ref>
<ref id="ref-140"><label>140.</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Ouyang</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Wu</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Jiang</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Almeida</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Wainwright</surname>, <given-names>C.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2022</year>). <article-title>Training language models to follow instructions with human feedback</article-title>. <conf-name>Advances in Neural Information Processing Systems</conf-name>, pp. <fpage>27730</fpage>&#x2013;<lpage>27744</lpage>. <publisher-loc>New Orleans, USA</publisher-loc>, <publisher-name>Curran Associates, Inc</publisher-name>.</mixed-citation></ref>
<ref id="ref-141"><label>141.</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Conover</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Hayes</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Mathur</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Meng</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Xie</surname>, <given-names>J.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2023</year>). <article-title>Free Dolly: Introducing the world&#x2019;s first truly open instruction-tuned LLM</article-title>. <ext-link ext-link-type="uri" xlink:href="https://www.databricks.com/blog/2023/04/12/dolly-first-open-commercially-viable-instruction-tuned-llm">https://www.databricks.com/blog/2023/04/12/dolly-first-open-commercially-viable-instruction-tuned-llm</ext-link></mixed-citation></ref>
<ref id="ref-142"><label>142.</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Touvron</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Lavril</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Izacard</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Martinet</surname>, <given-names>X.</given-names></string-name>, <string-name><surname>Lachaux</surname>, <given-names>M. A.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2023</year>). <article-title>LLaMA: Open and efficient foundation language models</article-title>. <source>arXiv preprint</source> <comment>arXiv:2302.13971</comment>.</mixed-citation></ref>
<ref id="ref-143"><label>143.</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Trajanoska</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Stojanov</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Trajanov</surname>, <given-names>D.</given-names></string-name></person-group> (<year>2023</year>). <article-title>Enhancing knowledge graph construction using large language models</article-title>. <source>arXiv preprint</source> <comment>arXiv:2305.04676</comment>.</mixed-citation></ref>
<ref id="ref-144"><label>144.</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Pan</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Luo</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Chen</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Wang</surname>, <given-names>J.</given-names></string-name> <etal>et al.</etal></person-group> (<year>2023</year>). <article-title>Unifying large language models and knowledge graphs: A roadmap</article-title>. <source>arXiv preprint</source> <comment>arXiv:2306.08302</comment>.</mixed-citation></ref>
</ref-list>
</back></article>