<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CSSE</journal-id>
<journal-id journal-id-type="nlm-ta">CSSE</journal-id>
<journal-id journal-id-type="publisher-id">CSSE</journal-id>
<journal-title-group>
<journal-title>Computer Systems Science &#x0026; Engineering</journal-title>
</journal-title-group>
<issn pub-type="ppub">0267-6192</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">38234</article-id>
<article-id pub-id-type="doi">10.32604/csse.2023.038234</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Radon CLF: A Novel Approach for Skew Detection Using Radon Transform</article-title>
<alt-title alt-title-type="left-running-head">Radon CLF: A Novel Approach for Skew Detection Using Radon Transform</alt-title>
<alt-title alt-title-type="right-running-head">Radon CLF: A Novel Approach for Skew Detection Using Radon Transform</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Chen</surname><given-names>Yuhang</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Bahaghighat</surname><given-names>Mahdi</given-names></name><xref ref-type="aff" rid="aff-2">2</xref><email>Bahaghighat@eng.ikiu.ac.ir</email></contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Kelishomi</surname><given-names>Aghil Esmaeili</given-names></name><xref ref-type="aff" rid="aff-3">3</xref></contrib>
<contrib id="author-4" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Du</surname><given-names>Jingyi</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><email>du-jingyi@xust.edu.cn</email></contrib>
<aff id="aff-1"><label>1</label><institution>Laboratory for Control Engineering, Xi&#x2019;an University of Science &#x0026;Technology</institution>, <addr-line>Xi&#x2019;an</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>Computer Engineering Department, Imam Khomeini International University</institution>, <addr-line>Qazvin</addr-line>, <country>Iran</country></aff>
<aff id="aff-3"><label>3</label><institution>MOE Key Laboratory for Intelligent and Network Security, Xi&#x2019;an Jiaotong University</institution>, <addr-line>Xi&#x2019;an</addr-line>, <country>China</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Authors: Mahdi Bahaghighat. Email: <email>Bahaghighat@eng.ikiu.ac.ir</email>; Jingyi Du. Email: <email>du-jingyi@xust.edu.cn</email> </corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic"><year>2023</year></pub-date>
<pub-date date-type="pub" publication-format="electronic"><day>26</day><month>5</month><year>2023</year></pub-date>
<volume>47</volume>
<issue>1</issue>
<fpage>675</fpage>
<lpage>697</lpage>
<history>
<date date-type="received"><day>03</day><month>12</month><year>2022</year></date>
<date date-type="accepted"><day>3</day><month>3</month><year>2023</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2023 Chen et al.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Chen et al.</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CSSE_38234.pdf"></self-uri>
<abstract>
<p>In the digital world, a wide range of handwritten and printed documents should be converted to digital format using a variety of tools, including mobile phones and scanners. Unfortunately, this is not an optimal procedure, and the entire document image might be degraded. Imperfect conversion effects due to noise, motion blur, and skew distortion can lead to significant impact on the accuracy and effectiveness of document image segmentation and analysis in Optical Character Recognition (OCR) systems. In Document Image Analysis Systems (DIAS), skew estimation of images is a crucial step. In this paper, a novel, fast, and reliable skew detection algorithm based on the Radon Transform and Curve Length Fitness Function (CLF), so-called Radon CLF, was proposed. The Radon CLF model aims to take advantage of the properties of Radon spaces. The Radon CLF explores the dominating angle more effectively for a 1D signal than it does for a 2D input image due to an innovative fitness function formulation for a projected signal of the Radon space. Several significant performance indicators, including Mean Square Error (MSE), Mean Absolute Error (MAE), Peak Signal-to-Noise Ratio (PSNR), Structural Similarity Measure (SSIM), Accuracy, and run-time, were taken into consideration when assessing the performance of our model. In addition, a new dataset named DSI5000 was constructed to assess the accuracy of the CLF model. Both two- dimensional image signal and the Radon space have been used in our simulations to compare the noise effect. Obtained results show that the proposed method is more effective than other approaches already in use, with an accuracy of roughly 99.87&#x0025; and a run-time of 0.048 (s). The introduced model is far more accurate and time-efficient than current approaches in detecting image skew.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Document image analysis</kwd>
<kwd>skew detection</kwd>
<kwd>Radon transform</kwd>
<kwd>pattern recognition</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1"><label>1</label><title>Introduction</title>
<p>Numerous printed and handwritten documents have been converted to digital format in the digital age utilizing a variety of devices, including mobile phones and dedicated scanners. Unfortunately, this procedure is far from ideal, and the entire document image may suffer from degradations including skew distortion, motion blur, and noise. The accuracy and efficiency of document image segmentation and analysis in Optical Character Recognition (OCR) systems can be directly impacted by these affecting factors. <xref ref-type="fig" rid="fig-1">Fig. 1</xref> depicts the operational procedures of an OCR system ([<xref ref-type="bibr" rid="ref-1">1</xref>]). In all OCR systems, the preprocessing steps are fundamental tasks that can affect the system&#x2019;s performance directly [<xref ref-type="bibr" rid="ref-2">2</xref>&#x2013;<xref ref-type="bibr" rid="ref-7">7</xref>]. The following is a list of some of the most fundamental and significant preprocessing methods used in Document Image Analysis (DIA):
<list list-type="bullet">
<list-item><p>Binarization</p></list-item>
<list-item><p>Skew Detection</p></list-item>
<list-item><p>Skew Correction</p></list-item>
<list-item><p>Noise Removal</p></list-item>
<list-item><p>Image Quality Enhancement</p></list-item>
<list-item><p>Dual-page Splitting</p></list-item>
<list-item><p>Straighten Curved Text Lines</p></list-item>
<list-item><p>Baseline Detection &#x0026; Extraction</p></list-item>
</list></p>
<fig id="fig-1"><label>Figure 1</label><caption><title>The different steps of any optical character recognition (OCR) system [<xref ref-type="bibr" rid="ref-1">1</xref>]</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-1.tif"/></fig>
<p>In general, skew detection is a primary step that plays a critical role in obtaining a high accuracy DIA system. The skewness in a document image can degrade exceedingly subsequent document processing algorithms in an OCR system.</p>
<p>In this work, the skew issue in input document images is the main focus. In this regard, a new dataset named DSI5000 (5000 Directional Synthetic Images) was created and developed by us. It includes about 5696 Directional Synthetic Images (DSI) with different intensities, angles, and frequencies to analyze the accuracy of the proposed Radon Transform Curve Length Fitness Function (CLF) algorithm so-called the Radon CLF. The Radon CLF model tries to benefit Radon space characteristics. Based on an innovative fitness function definition for a projected signal of the Radon space, the Radon CLF explores the dominant angle more efficiently for a 1D signal rather than the 2D input image. In this study, we have concentrated on the speed, robustness (against the noise effect), and the accuracy of the proposed model for the skew detection problem.</p>
<p>The many sections of this paper are arranged as follows: Section 2 introduces and discusses related works. Section 3 describes the proposed approach for skew detection. Experimental results and comparative analysis are the subject matter of Section 4. Finally, Section 5 includes the conclusion and future studies.</p>
</sec>
<sec id="s2"><label>2</label><title>Related Works</title>
<p>Many approaches based on projection profile, topline, and scanline methods, have been used in the past to conduct skew detection on images. These are the most straightforward and often used approaches to identify document skew; however, the majority of them involve slow algorithms with low accuracy. These techniques generally rely on the fonts and grammatical structure of a given language as well. They do not perform well enough with manuscripts that either incorporate multiple languages or different fonts, in practice. For example, an approach using the horizontal projection histogram for just Arabic text was presented by [<xref ref-type="bibr" rid="ref-8">8</xref>]. They present a method that was based entirely on polygonal approximated skeleton processing.</p>
<p>In Signal &#x0026; Image Processing and Computer Vision theory, there are many utilitarian directional filters and transforms, for example: the Hough Transform (HT), the Gabor Wavelet Transform (GWT), and Directional Median Filters (DMF). These directional filters can be used in the analysis of the directional patterns ([<xref ref-type="bibr" rid="ref-9">9</xref>&#x2013;<xref ref-type="bibr" rid="ref-13">13</xref>]). For example, [<xref ref-type="bibr" rid="ref-14">14</xref>] used Hough Transform and Run-Length Encoding (RLE) algorithms for the skew detection problem. In 2010, an algorithm was proposed by [<xref ref-type="bibr" rid="ref-15">15</xref>]. It has three steps: firstly, the projection of the vertical and horizontal graphics of the image was eliminated. A binary image was applied to the dilation operation. In the last phase, the skew angle was achieved with the help of the Hough Transform. The proposed algorithm can detect the skew angle in the range between &#x2212;90<sup>&#x25E6;</sup>, and &#x002B;90<sup>&#x25E6;</sup>, with high precision.</p>
<p>Reference [<xref ref-type="bibr" rid="ref-16">16</xref>] deployed the Fast Hough Transform (FHT) to detect skew angles. In this approach, there is no need for a binary representation of the text image. The Fast Hough Transform was applied on both vertical and horizontal lines. This technique reduced the computational cost; while, the calculated error was around 0.547. Skew detection approaches found on Hough Transform usually impose a high computational cost. Reference [<xref ref-type="bibr" rid="ref-17">17</xref>] improved an algorithm based on least squares to handle a multi-skew problem. In [<xref ref-type="bibr" rid="ref-18">18</xref>] and [<xref ref-type="bibr" rid="ref-19">19</xref>], they used a binary text document dataset to evaluate the skew angle with the help of linear regression algorithms. The time cost of these methods was smaller than other algorithms based on the Hough Transform. The algorithm performed text lines classification. To increase the accuracy, the variation of skew angle was calculated in the &#x00B1;10<sup>&#x25E6;</sup> range. With increasing the angle, the accuracy was reducing to the same ratio.</p>
<p>In [<xref ref-type="bibr" rid="ref-20">20</xref>], they introduced a method that inscribes the text in a document image found on an arbitrary polygon and derivation of the baseline from the polygon&#x2019;s centroid. It had proven that their algorithm was suitable to apply to documents written in different fonts. In [<xref ref-type="bibr" rid="ref-21">21</xref>], an algorithm was presented to perform skew angle correction for handwritten text documents. The algorithm used the Hough Transform and the restricted box technique. In this algorithm, linear regression functions have been used to compute the skew angle in a manuscript. The algorithm&#x2019;s main point is that it displays parallel rectangles in the binary image with the lowest number of pixels in line with horizontal and vertical orientations. Then, the linear regression function would be calculated for the skew angle. In [<xref ref-type="bibr" rid="ref-22">22</xref>], morphological functions were used to detect skew angles. These functions can be analyzed in the frequency domain found on Fourier Transform (FT). The Fourier Transform could detect text distortion in different languages, including English, Hindi, and Punjabi. In [<xref ref-type="bibr" rid="ref-23">23</xref>], a nearest-neighbor-chain or NNC algorithm was introduced with language-independent capability. Size restriction was the main challenge to the detection of nearest-neighbors (NN), before the skew detection.</p>
<p>In [<xref ref-type="bibr" rid="ref-24">24</xref>], the Principle Component Analysis (PCA) and its wide applications in Image &#x0026; Signal Processing, were presented. In [<xref ref-type="bibr" rid="ref-6">6</xref>], finding the skew angle has done using the PCA approach. After converting the input document into a binary picture, the algorithm utilized Sobel and Gaussian filters. It helps to find edges and reduce noise. The PCA-based method gets the covariance matrix, after which it produces the Eigen values and Eigen vectors, and then calculates the unit vector for the principal component. Later, the document&#x2019;s skew angle can be determined using the principal components. The introduced algorithm could get an accuracy of about 90&#x0025; but with a high time cost.</p>
<p>In [<xref ref-type="bibr" rid="ref-25">25</xref>], they presented an adaptive Skew Correction technique for document images. It uses image&#x2019;s layout features and classification to detect the type of document image. Three different classes, named Text image, Form image, and Complex image, were considered in the classification problem. Then, based on the type of document image, one of three proposed algorithms: Morphological Clustering (MC), Piecewise Projection Profile (PPP), and Skeleton Line Detection (SKLD), should be used to correct the skew of a document image.</p>
<p>Transferring the input image to Radon space and using a strong feature extraction method like the suggested Radon CLF, is an appropriate alternative to working with highly loaded 2D data. It can benefit high-speed processing algorithms in a one-dimensional signal space rather than image space. Radon CLF can also lead to improved results in terms of computational time and accuracy. The subsequent sections will go through this idea.</p>
</sec>
<sec id="s3"><label>3</label><title>Methodology</title>
<p>Today, enormous amounts of information are stored in printed documents. The primary and important step in the processing of paper-based documents is to convert them into digital records. In practice, the main problem is that the document may be rotated unaptly on a flatbed scanner at a random angle. In this situation, the scanned image may be skewed. The skew is considered a surplus distortion which can degrade the image quality. The skew can impose severe challenges in digital image analysis and deteriorate the overall performance of any OCR system. In this section, the methodology and our proposed approach to detect skew is discussed in more detail. In <xref ref-type="fig" rid="fig-2">Fig. 2</xref>, a general scheme of our proposed approach is presented.</p>
<fig id="fig-2"><label>Figure 2</label><caption><title>The proposed Radon CLF algorithm for skew detection in document images</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-2.tif"/></fig>
<sec id="s3_1"><label>3.1</label><title>The Radon Transformation</title>
<p>The Radon function is calculated from an image matrix such as <italic>f(x, y)</italic> in specific directions. The Radon function accounts for linear integrals over different paths or certain paths (from different sources/beams) in a specific direction [<xref ref-type="bibr" rid="ref-9">9</xref>,<xref ref-type="bibr" rid="ref-26">26</xref>&#x2013;<xref ref-type="bibr" rid="ref-32">32</xref>]. This function takes multiple parallel-beam projections of the input image from different angles by rotating the source around the center of the image.</p>
<p>In <xref ref-type="fig" rid="fig-3">Fig. 3</xref>, a single image at a specified rotation angle of the Radon transform is illustrated. Besides, <xref ref-type="fig" rid="fig-4">Fig. 4</xref> indicates the geometry of the Radon Transform.</p>
<fig id="fig-3"><label>Figure 3</label><caption><title>The parallel-beam projection at rotation angle [<xref ref-type="bibr" rid="ref-33">33</xref>]</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-3.tif"/></fig><fig id="fig-4"><label>Figure 4</label><caption><title>The Geometry of the Radon Transform [<xref ref-type="bibr" rid="ref-37">37</xref>]</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-4.tif"/></fig>
<p>The computation of projections can be down from any angle. Generally, the Radon transform (<inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>) of the function <italic>f(x, y)</italic> is the linear integral of it, in parallel to <italic>y&#x2019;</italic>-axis [<xref ref-type="bibr" rid="ref-34">34</xref>&#x2013;<xref ref-type="bibr" rid="ref-36">36</xref>] (see <xref ref-type="disp-formula" rid="eqn-1">Eqs. (1)</xref> to <xref ref-type="disp-formula" rid="eqn-4">(4)</xref>):
<disp-formula id="eqn-1"><label>(1)</label><mml:math id="mml-eqn-1" display="block"><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msubsup><mml:mo>&#x222B;</mml:mo><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mi mathvariant="normal">&#x221E;</mml:mi></mml:mrow><mml:mrow><mml:mo>+</mml:mo><mml:mi mathvariant="normal">&#x221E;</mml:mi></mml:mrow></mml:msubsup><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mrow><mml:msup><mml:mi>&#x03C1;</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:msup><mml:mi>&#x03C1;</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mi>d</mml:mi><mml:mi>y</mml:mi></mml:math></disp-formula>
<disp-formula id="eqn-2"><label>(2)</label><mml:math id="mml-eqn-2" display="block"><mml:msub><mml:mrow><mml:msup><mml:mi>&#x03C1;</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>&#x2212;</mml:mo><mml:msup><mml:mi>y</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-3"><label>(3)</label><mml:math id="mml-eqn-3" display="block"><mml:msub><mml:mrow><mml:msup><mml:mi>&#x03C1;</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:msup><mml:mi>y</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-4"><label>(4)</label><mml:math id="mml-eqn-4" display="block"><mml:mrow><mml:mo>[</mml:mo><mml:mtable columnalign="left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msup><mml:mi>y</mml:mi><mml:mrow><mml:msup><mml:mi></mml:mi><mml:mo>&#x2032;</mml:mo></mml:msup></mml:mrow></mml:msup></mml:mtd></mml:mtr></mml:mtable><mml:mo>]</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mtable columnalign="left left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>&#x2212;</mml:mo><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable><mml:mo>]</mml:mo></mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mtable columnalign="left" rowspacing="4pt" columnspacing="1em"><mml:mtr><mml:mtd><mml:mi>x</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>y</mml:mi></mml:mtd></mml:mtr></mml:mtable><mml:mo>]</mml:mo></mml:mrow></mml:math></disp-formula>where the pair (<italic>x&#x2019;, y&#x2019;</italic>) is the new place of the (<italic>x, y</italic>) after rotating with the angle <italic>theta</italic> in a two-dimensional Cartesian coordinate system.</p>
</sec>
<sec id="s3_2"><label>3.2</label><title>Skew Detection</title>
<p>This section conceptualizes our new skew detection approach so-called Radon CLF (Radon Curve Length Fitness Function). Usually, the local region in a document image has a consistent orientation and frequency [<xref ref-type="bibr" rid="ref-38">38</xref>]. So, it can be modeled as a surface wave characterized entirely by the dominant orientation and frequency pattern. This approximation model is practical enough for our purpose of evaluating the performance of the Radon Transform for the skew estimation problem. According to <xref ref-type="disp-formula" rid="eqn-5">Eq. (5)</xref>, a local region of the image can be modeled as a surface wave [<xref ref-type="bibr" rid="ref-39">39</xref>]:
<disp-formula id="eqn-5"><label>(5)</label><mml:math id="mml-eqn-5" display="block"><mml:mi>I</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mtext>&#x00A0;</mml:mtext><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>A</mml:mi><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mn>2</mml:mn><mml:mi>&#x03C0;</mml:mi><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mtext>&#x00A0;</mml:mtext><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>s</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mi>y</mml:mi><mml:mtext>&#x00A0;</mml:mtext><mml:mi>s</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>where <italic>f</italic> is the frequency, the <italic>theta</italic> is the dominant orientation, and <italic>A</italic> is the amplitude of the <italic>cosine</italic> function. <italic>A</italic> is the intensity adjustment parameter of the synthesized image <italic>I(x, y)</italic>. An example image and its projection by the Radon Transform are shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>. As it can be seen in the projection function in <xref ref-type="fig" rid="fig-5">Fig. 5b</xref>, providing that it has been projected in the actual orientation, which is parallel to the local orientation of the input image, it can be approximately treated as a semi-sinusoidal plane wave. Besides, the noisy version of the image in <xref ref-type="fig" rid="fig-5">Fig. 5a</xref> and its Radon projection are depicted in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>. Although, the noise power is utterly high (the Gaussian noise with the standard deviation (&#x03C3;) about 20), the comparison between <xref ref-type="fig" rid="fig-5">Figs. 5b</xref>, and <xref ref-type="fig" rid="fig-6">6b</xref> shows that the semi-sinusoidal structure of the Radon projection is still satisfied with a few distortions. <xref ref-type="fig" rid="fig-7">Figs. 7</xref> and <xref ref-type="fig" rid="fig-8">8</xref> are depicted with a new angle. Now, two new Radon transform maps are compared between a noise-free directional image and its noisy version with &#x03C3;&#x2009;&#x003D;&#x2009;9. With comparing two <xref ref-type="fig" rid="fig-7">Figs. 7a</xref> and <xref ref-type="fig" rid="fig-8">8a</xref> in the image signal space, it can be viewed that the noise effect is quite eye-catching. On the other side, when we make comparison between <xref ref-type="fig" rid="fig-7">Figs. 7b</xref> and <xref ref-type="fig" rid="fig-8">8b</xref> in the Radon space, it indicates that the presence of noise does not impact the projected pattern in this space, considerably.</p>
<fig id="fig-5"><label>Figure 5</label><caption><title>(a) A well-defined 150&#x2009;&#x00D7;&#x2009;150 synthetic image (x). (b) The Radon Transform: R(x)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-5.tif"/></fig><fig id="fig-6"><label>Figure 6</label><caption><title>(a) A noisy synthetic image (Gaussian Noise with &#x03C3;&#x2009;&#x003D;&#x2009;20). (b) The Radon Transform: R(x)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-6.tif"/></fig><fig id="fig-7"><label>Figure 7</label><caption><title>(a) A noise-free synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;30 and &#x03C3;&#x2009;&#x003D;&#x2009;0. (b) The Radon transform map of the image</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-7.tif"/></fig><fig id="fig-8"><label>Figure 8</label><caption><title>(a) A noisy synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;30 and &#x03C3;&#x2009;&#x003D;&#x2009;9. (b) The Radon transform map of the image</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-8.tif"/></fig>
<p>These comparisons show that the proposed skew detection algorithm can abundantly tolerate the additive noise more efficiently in the Radon space rather than the signal space, providing that an appropriate feature extraction method is available for the Radon space. This observed phenomenon should be evaluated further in the following sections.</p>
</sec>
<sec id="s3_3"><label>3.3</label><title>Curve Length of a Function</title>
<p>In <xref ref-type="fig" rid="fig-9">Fig. 9</xref>, <italic>f(x)</italic> is shown as an example of an one-dimensional continuous function. The Arc length <italic>L</italic> of a function such as <italic>y&#x2009;&#x003D;&#x2009;f(x)</italic> between <italic>a</italic> and <italic>b</italic> (from the point (<italic>a, f (a)</italic>) to the point (<italic>b, f (b)</italic>)) can be derived using <xref ref-type="disp-formula" rid="eqn-6">Eq. (6)</xref>:
<disp-formula id="eqn-6"><label>(6)</label><mml:math id="mml-eqn-6" display="block"><mml:mi>L</mml:mi><mml:mo>=</mml:mo><mml:msubsup><mml:mo>&#x222B;</mml:mo><mml:mrow><mml:mi>a</mml:mi></mml:mrow><mml:mrow><mml:mi>b</mml:mi></mml:mrow></mml:msubsup><mml:msqrt><mml:mn>1</mml:mn><mml:mo>+</mml:mo><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:mfrac><mml:mrow><mml:mi>d</mml:mi><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi><mml:mi>x</mml:mi></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:msqrt><mml:mtext>&#x00A0;</mml:mtext><mml:mi>d</mml:mi><mml:mi>x</mml:mi></mml:math></disp-formula>where <inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mfrac><mml:mrow><mml:mi>d</mml:mi><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi><mml:mi>x</mml:mi></mml:mrow></mml:mfrac></mml:mstyle></mml:math></inline-formula> denotes the first derivative of the function <italic>f(x)</italic>. Supposing that <italic>C</italic> would be a curve in Euclidean (or, generally, a metric) space <italic>X&#x2009;&#x003D;&#x2009;R<sup>n</sup></italic>, so <italic>C</italic> is a continuous function of an image where <italic>f:</italic> [<italic>a, b</italic>] <inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:mo stretchy="false">&#x2192;</mml:mo></mml:math></inline-formula> <italic>X</italic> of the interval [<italic>a, b</italic>] into <italic>X</italic>. From a partition <italic>a</italic>&#x2009;<italic>&#x003D;</italic>&#x2009;<italic>x<sub>0</sub></italic>&#x2009;&#x003C;&#x2009;<italic>&#x2026;</italic>&#x2009;&#x003C;&#x2009;<italic>x<sub>n&#x2212;1</sub></italic>&#x2009;&#x003C;&#x2009;<italic>x<sub>n</sub></italic>&#x2009;<italic>&#x003D;</italic>&#x2009;<italic>b</italic> of the interval [<italic>a, b</italic>], there is a finite collection of points <italic>f</italic>(<italic>x<sub>0</sub></italic>), <italic>f</italic>(<italic>x<sub>1</sub></italic>),<italic>&#x2026;</italic>, <italic>f</italic>(<italic>x<sub>n&#x2212;1</sub></italic>), <italic>and f</italic>(<italic>x<sub>n</sub></italic>), which can be used to calculate the length of the line segment connecting the two points. According to <xref ref-type="disp-formula" rid="eqn-7">Eq. (7)</xref>, the arc length <italic>L</italic> of <italic>C</italic> is then defined to be <italic>L(C)</italic>:
<disp-formula id="eqn-7"><label>(7)</label><mml:math id="mml-eqn-7" display="block"><mml:mi>L</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>C</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msubsup><mml:mrow><mml:mo>&#x2211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msubsup><mml:mi>d</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">)</mml:mo><mml:mo>,</mml:mo><mml:mi>f</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mo>+</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula>where the supermom is calculated of all possible sections of [<italic>a, b</italic>] and n is unbounded. This definition of the arc length does not require that <italic>C</italic> be defined by a differentiable function. Generally, the notion of differentiability is not defined in a metric space.</p>
<fig id="fig-9"><label>Figure 9</label><caption><title>An example of 1D function f(x)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-9.tif"/></fig>
</sec>
<sec id="s3_4"><label>3.4</label><title>Definition of the Proposed Fitness Function</title>
<p>In the case of the discreet Radon function, <inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula>, this article defines the curve length of the function, <inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>, as <xref ref-type="disp-formula" rid="eqn-8">Eq. (8)</xref>:
<disp-formula id="eqn-8"><label>(8)</label><mml:math id="mml-eqn-8" display="block"><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:msub><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>Thus, the estimated skew of an input image, <inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula>, can be obtained the using proposed <xref ref-type="disp-formula" rid="eqn-9">Eq. (9)</xref>:
<disp-formula id="eqn-9"><label>(9)</label><mml:math id="mml-eqn-9" display="block"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mtext mathvariant="italic">Argmax</mml:mtext></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>R</mml:mi><mml:mrow><mml:mi>&#x03B8;</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>If the orientation is the same as the local skew, the projection function will result in a semi-sinusoidal plane wave. If not, the resultant pattern can be an erratic, non-sinusoidal signal with a smaller amplitude. This fact is shown in <xref ref-type="fig" rid="fig-10">Fig. 10</xref>. <xref ref-type="fig" rid="fig-10">Fig. 10</xref> makes a comparison between Radon projection patterns on the actual orientation at &#x03B8;&#x2009;&#x003D;&#x2009;70 (the red line with error&#x2009;&#x003D;&#x2009;0) and some incorrect orientations, such as 71 (error&#x2009;&#x003D;&#x2009;&#x002B;&#x2009;1<sup>&#x25E6;</sup>), and 60 (error&#x2009;&#x003D;&#x2009;&#x2212;10<sup>&#x25E6;</sup>). This study proposes the curve length of the projected Radon pattern as a fitness function for skew estimation, and call it the Radon CLF algorithm.</p>
<fig id="fig-10"><label>Figure 10</label><caption><title>Comparison between Radon projection patterns on the actual orientation at &#x03B8;&#x2009;&#x003D;&#x2009;70 (the red line with error&#x2009;&#x003D;&#x2009;0), and some incorrect ones such as 71 (error&#x2009;&#x003D;&#x2009;&#x002B;&#x2009;1<sup>&#x25E6;</sup>) &#x0026; 60 (error&#x2009;&#x003D;&#x2009;&#x2212;10<sup>&#x25E6;</sup>)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-10.tif"/></fig>
<p>It will be shown that in the correct orientation, such as &#x03B8;&#x2009;&#x003D;&#x2009;70 in <xref ref-type="fig" rid="fig-10">Fig. 10</xref>, the length of the curve would be well over the curve lengths of other incorrect orientations.</p>
</sec>
<sec id="s3_5"><label>3.5</label><title>Performance Evaluation Metrics</title>
<p>Performance metrics are a vital part of every algorithm analysis. To evaluate the performance of our proposed method, we have used the Mean Squared Error (MSE) (for both the skew estimator and image comparison), the Mean Absolute Error (MAE) (for the skew estimator), Accuracy (for the skew estimator), Peak Signal-to-Noise Ratio (PSNR) (for the images comparison), and Structural Similarity Measure (SSIM) (for the images comparison) along with the Computational Time.</p>
<p>In statistics, the MSE is considered the Mean Squared Deviation (MSD) of an estimator. It can measure the average of the square of the errors or the average squared difference between the actual/ground-truth value and the estimated value ([<xref ref-type="bibr" rid="ref-40">40</xref>]).
<disp-formula id="eqn-10"><label>(10)</label><mml:math id="mml-eqn-10" display="block"><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:msubsup><mml:mrow><mml:mo>&#x2211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:msubsup><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:msub><mml:mi>&#x03B8;</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:math></disp-formula></p>
<p>In <xref ref-type="disp-formula" rid="eqn-10">Eq. (10)</xref>, <inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:msub><mml:mi>&#x03B8;</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the true value of the angle for the <inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:msup><mml:mi>i</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mi>h</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> sample, and <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:msub><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> is the estimated angle for it. The <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:msub><mml:mi>&#x03B8;</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mtext>&#x00A0;</mml:mtext></mml:math></inline-formula> is the error signal for the <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:msup><mml:mi>i</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mi>h</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> sample, and should be calculated for all <italic>n</italic> available samples.</p>
<p>In comparison to the <italic>MSE</italic>, the Mean Absolute Error or <italic>MAE</italic> is the absolute average of the difference between the ground-truth and the predicted value [<xref ref-type="bibr" rid="ref-40">40</xref>] (see <xref ref-type="disp-formula" rid="eqn-11">Eq. (11)</xref>).
<disp-formula id="eqn-11"><label>(11)</label><mml:math id="mml-eqn-11" display="block"><mml:mi>M</mml:mi><mml:mi>A</mml:mi><mml:mi>E</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>n</mml:mi></mml:mfrac><mml:msubsup><mml:mrow><mml:mo>&#x2211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:msubsup><mml:mrow><mml:mo>|</mml:mo><mml:msub><mml:mi>&#x03B8;</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>|</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>In addition to <italic>MSE</italic> and <italic>MAE</italic>, the <italic>Accuracy</italic> is defined using <xref ref-type="disp-formula" rid="eqn-12">Eqs. (12)</xref> to <xref ref-type="disp-formula" rid="eqn-14">(14)</xref>:
<disp-formula id="eqn-12"><label>(12)</label><mml:math id="mml-eqn-12" display="block"><mml:mrow><mml:mo>|</mml:mo><mml:mi>T</mml:mi><mml:mo>|</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>c</mml:mi><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>d</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>T</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mtext>&#x00A0;</mml:mtext><mml:mo stretchy="false">&#x2194;</mml:mo><mml:mtext>&#x00A0;</mml:mtext><mml:mi>T</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mi mathvariant="normal">&#x2200;</mml:mi><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:mi>D</mml:mi><mml:mi>B</mml:mi><mml:mo fence="false" stretchy="false">|</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>=</mml:mo><mml:mn>0</mml:mn><mml:mo>}</mml:mo></mml:mrow><mml:mo>}</mml:mo></mml:mrow></mml:math></disp-formula>where <italic>&#x007C;T&#x007C;</italic> is used for the cardinal of the set <italic>T</italic>, and it means the total number of error-free predictions (true predictions) in the whole dataset (Card (DB)). Similarly, the <italic>&#x007C;F&#x007C;</italic> is used for the cardinal of the set <italic>F</italic> which means the total number of false predictions (predictions with error) in the entire dataset.
<disp-formula id="eqn-13"><label>(13)</label><mml:math id="mml-eqn-13" display="block"><mml:mrow><mml:mo>|</mml:mo><mml:mi>F</mml:mi><mml:mo>|</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mi>c</mml:mi><mml:mi>a</mml:mi><mml:mi>r</mml:mi><mml:mi>d</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>F</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mtext>&#x00A0;</mml:mtext><mml:mo stretchy="false">&#x2194;</mml:mo><mml:mtext>&#x00A0;</mml:mtext><mml:mi>F</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mi mathvariant="normal">&#x2200;</mml:mi><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:mi>D</mml:mi><mml:mi>B</mml:mi><mml:mo fence="false" stretchy="false">|</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>&#x2260;</mml:mo><mml:mn>0</mml:mn><mml:mo>}</mml:mo></mml:mrow><mml:mo>}</mml:mo></mml:mrow></mml:math></disp-formula>
<disp-formula id="eqn-14"><label>(14)</label><mml:math id="mml-eqn-14" display="block"><mml:mrow><mml:mtext mathvariant="italic">Accuracy</mml:mtext></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>T</mml:mi><mml:mo>|</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo>|</mml:mo><mml:mi>T</mml:mi><mml:mo>|</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mrow><mml:mo>|</mml:mo><mml:mi>F</mml:mi><mml:mo>|</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:math></disp-formula></p>
<p>The 1D MSE formula can be extended to the 2D space ([<xref ref-type="bibr" rid="ref-41">41</xref>,<xref ref-type="bibr" rid="ref-42">42</xref>]): <italic>MSE<sup>2D</sup></italic>. In this study, it is needed to compare two images as two matrices. <xref ref-type="disp-formula" rid="eqn-15">Eq. (15)</xref> represents the <italic>MSE<sup>2D</sup></italic> for two images:
<disp-formula id="eqn-15"><label>(15)</label><mml:math id="mml-eqn-15" display="block"><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:msup><mml:mi>E</mml:mi><mml:mrow><mml:mn>2</mml:mn><mml:mi>D</mml:mi></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mi>m</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>n</mml:mi></mml:mrow></mml:mfrac><mml:msubsup><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msubsup><mml:msubsup><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msubsup><mml:mo stretchy="false">(</mml:mo><mml:mi>I</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>I</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:msup><mml:mrow><mml:mo>(</mml:mo><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:math></disp-formula>where <italic>I</italic>(<italic>x, y</italic>) is an <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:mi>m</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>n</mml:mi></mml:math></inline-formula> original image while the <inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:mrow><mml:mover><mml:mi>I</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:math></inline-formula> would be whether a noisy image or a disturbing version of the original image. The <italic>MSE<sup>2D</sup></italic> measures differences between two images and shows the quality degradation. The zero <italic>MSE<sup>2D</sup></italic> means that the two images are the identical (the perfect similarity).</p>
<p>The <italic>PSNR</italic> or Peak Signal-to-Noise Ratio also represents a measure of the image error. Both the <italic>PSNR</italic>, and <italic>MSE<sup>2D</sup></italic> are usually used to measure an image&#x2019;s quality after its variation ([<xref ref-type="bibr" rid="ref-41">41</xref>,<xref ref-type="bibr" rid="ref-42">42</xref>]). The <italic>PSNR</italic> can be calculated directly found on the <italic>MSE<sup>2D</sup></italic> ([<xref ref-type="bibr" rid="ref-40">40</xref>]), according to <xref ref-type="disp-formula" rid="eqn-16">Eq. (16)</xref>:
<disp-formula id="eqn-16"><label>(16)</label><mml:math id="mml-eqn-16" display="block"><mml:mi>P</mml:mi><mml:mi>S</mml:mi><mml:mi>N</mml:mi><mml:mi>R</mml:mi><mml:mo>=</mml:mo><mml:mn>10</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo>(</mml:mo><mml:mfrac><mml:msup><mml:mn>255</mml:mn><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mrow><mml:mi>M</mml:mi><mml:mi>S</mml:mi><mml:msup><mml:mi>E</mml:mi><mml:mrow><mml:mn>2</mml:mn><mml:mi>D</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:mfrac><mml:mo>)</mml:mo></mml:mrow></mml:math></disp-formula></p>
<p>In addition to <italic>PSNR</italic> and <italic>MSE<sup>2D</sup></italic>, Structural Similarity Measure or <italic>SSIM</italic> have been deployed in this paper using <xref ref-type="disp-formula" rid="eqn-17">Eq. (17)</xref> ([<xref ref-type="bibr" rid="ref-42">42</xref>]). In a 2D space, the <italic>MSE<sup>2D</sup></italic> will calculate distance as the mean of the square error between each corresponding pixel for the two target images. In contrast, the <italic>SSIM</italic> tries to do the opposite, and looks for similarities within pixels ([<xref ref-type="bibr" rid="ref-42">42</xref>&#x2013;<xref ref-type="bibr" rid="ref-44">44</xref>]). To remedy some of the issues associated with MSE for image processing, <italic>SSIM</italic> have used. In <xref ref-type="disp-formula" rid="eqn-17">Eq. (17)</xref>, <italic>&#x03BC;<sub>x</sub></italic> and <italic>&#x03BC;<sub>y</sub></italic> are the average of <italic>x</italic> and <italic>y</italic> while <italic>&#x03C3;<sub> x</sub></italic> and <italic>&#x03C3;<sub> y</sub></italic> are the standard deviation of them, respectively. Similarly, <italic>&#x03BC;<sub>x</sub></italic><sup>2</sup> and <italic>&#x03BC;<sub>y</sub></italic><sup>2</sup> denote the variances and <italic>&#x03C3;<sub> xy</sub></italic> is the covariance of <italic>x</italic> and <italic>y</italic>. The <italic>c<sub>1</sub></italic> and <italic>c<sub>2</sub></italic> are two adjustable constant parameters.
<disp-formula id="eqn-17"><label>(17)</label><mml:math id="mml-eqn-17" display="block"><mml:mi>S</mml:mi><mml:mi>S</mml:mi><mml:mi>I</mml:mi><mml:mi>M</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mn>2</mml:mn><mml:msub><mml:mi>&#x03BC;</mml:mi><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mi>&#x03BC;</mml:mi><mml:mrow><mml:mi>y</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:mn>2</mml:mn><mml:msub><mml:mi>&#x03C3;</mml:mi><mml:mrow><mml:mi>x</mml:mi><mml:mi>y</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03BC;</mml:mi></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>+</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03BC;</mml:mi></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mrow><mml:mo>(</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03C3;</mml:mi></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>+</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x03C3;</mml:mi></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>+</mml:mo><mml:msub><mml:mi>c</mml:mi><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:math></disp-formula></p>
<p>It is worth mentioning that the <italic>SSIM</italic> can vary between &#x2212;1 and 1; where <italic>SSIM</italic>&#x2009;&#x003D;&#x2009;1 indicates the perfect similarity ([<xref ref-type="bibr" rid="ref-40">40</xref>,<xref ref-type="bibr" rid="ref-43">43</xref>,<xref ref-type="bibr" rid="ref-44">44</xref>]).</p>
</sec>
<sec id="s3_6"><label>3.6</label><title>Skew Correction</title>
<p>When the skew angle is detected by an algorithm, the next step would be Skew Correction. Technically, it is just a simple rotation procedure for a 2D image. So far, different methods such as contour-oriented projection, direct/indirect based method, and others, have been introduced to correct skewed images. In our simulation, the rotation of an input image is done through the Affine Transformation (AT) using <xref ref-type="disp-formula" rid="eqn-18">Eqs. (18)</xref> and <xref ref-type="disp-formula" rid="eqn-19">(19)</xref> ([<xref ref-type="bibr" rid="ref-45">45</xref>]):
<disp-formula id="eqn-18"><label>(18)</label><mml:math id="mml-eqn-18" display="block"><mml:msup><mml:mi>x</mml:mi><mml:mrow><mml:mo>&#x2217;</mml:mo></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mi>x</mml:mi><mml:mo>+</mml:mo><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mi>y</mml:mi></mml:math></disp-formula>
<disp-formula id="eqn-19"><label>(19)</label><mml:math id="mml-eqn-19" display="block"><mml:msup><mml:mi>y</mml:mi><mml:mrow><mml:mo>&#x2217;</mml:mo></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mi>sin</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mi>x</mml:mi><mml:mo>+</mml:mo><mml:mi>cos</mml:mi><mml:mo>&#x2061;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mi>y</mml:mi></mml:math></disp-formula></p>
<p>Here, the (<italic>x, y</italic>) is the coordinates of a pixel in the skewed input image, <inline-formula id="ieqn-14"><mml:math id="mml-ieqn-14"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula> denotes the detected angle, and (<italic>x</italic>&#x002A;, <italic>y</italic>&#x002A;) is a new place of the pixel.</p>
</sec>
</sec>
<sec id="s4"><label>4</label><title>Experimental Results</title>
<p>The introduced method was implemented in the Python programming environment using an Intel(R) Core(TM) i7-7700HQ 2.80&#x2005;GHz CPU. To evaluate the model, a dataset called DSI5000 was created and developed by us. It has about 5696 Directional Synthetic Images (DSI) with different amplitudes (intensities), orientations (dominant directions), and frequencies (repetitive line patterns). In addition to DSI5000, many real-world scanned image documents were also included in our simulations. Both handwriting and printed samples were gathered at various resolutions from different Persian, Arabic, English, and multilingual resources.</p>
<p>In the first experiment, the Radon projections were calculated for an input image with an actual orientation of &#x03B8;&#x2009;&#x003D;&#x2009;70. In <xref ref-type="fig" rid="fig-11">Fig. 11</xref>, the Radon projections were illustrated for angles between 60 and 80 (including the true angle at &#x03B8;&#x2009;&#x003D;&#x2009;70). It is a semi-sinusoidal pattern with well-defined harmonics and the highest amplitude for the projection angle&#x2009;&#x003D;&#x2009;70 (the correct orientation). <italic>&#x03B8;<sub>0</sub></italic> refers to all other projection angles, such as 60, 65, 68, 69, 71, 72, 75, and 80. These angels have noticeably lower amplitudes. For <italic>&#x03B8;<sub>0</sub></italic>, it can be seen the more significant gap between a signal&#x2019;s peaks compared to the peaks of the signal associated with the actual angle (&#x03B8;). Then, <xref ref-type="fig" rid="fig-12">Fig. 12</xref> plots the Fitness Function <italic>vs.</italic> &#x03B8;. It illustrates that the fitness function of a synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 in the orientation ranges from 1 to 180 degrees has just one global minimum at &#x03B8;&#x2009;&#x003D;&#x2009;70. Therefore, the skew of the image can be detected precisely and uniquely by the proposed fitness function. Our simulations indicate that the proposed feature extraction approach based on the curve length of the projected signal has the potential to discriminate these variations among projected signals, finely.</p>
<fig id="fig-11"><label>Figure 11</label><caption><title>Radon Transform output for different projection angles (The correct orientation: &#x03B8;&#x2009;&#x003D;&#x2009;70)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-11.tif"/></fig><fig id="fig-12"><label>Figure 12</label><caption><title>Fitness Function <italic>vs.</italic> &#x03B8;. It has a global minimum at &#x03B8;&#x2009;&#x003D;&#x2009;70 (the dominant angel)</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-12.tif"/></fig>
<sec id="s4_1"><label>4.1</label><title>Evaluation of the Noise Effect</title>
<p>One of the most important things related to any algorithm is its robustness against noise effects. This section, hankers for drawing a broad picture of the noise tolerance of our model. <xref ref-type="fig" rid="fig-13">Fig. 13</xref> represents an input directional image with its noisy version (An additive Gaussian noise with &#x03C3;&#x2009;&#x003D;&#x2009;10). The Error image, which is defined as a differential image between two images, also is depicted in this experiment. The Error energy shows that the degrading effect of the additive Gaussian noise is so high. To make an appropriate compassion to <xref ref-type="fig" rid="fig-13">Fig. 13</xref>, the experiment is repeated; but this time, the image space is replaced by the Radon space.</p>
<fig id="fig-13"><label>Figure 13</label><caption><title>Image signal space: Input image <italic>vs.</italic> Noisy image <italic>vs.</italic> Error image</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-13.tif"/></fig>
<p>In <xref ref-type="fig" rid="fig-14">Fig. 14</xref>, the Radon projection maps are depicted for the input image, its noisy version, and the Error image in the Radon space. When the noisy image in the signal space in <xref ref-type="fig" rid="fig-13">Fig. 13</xref> is compared with the noisy image in the Radon space in <xref ref-type="fig" rid="fig-14">Fig. 14</xref>, it shows that in the Radon space, the information related to the orientation can still be extracted from the noisy image. In contrast, its corresponding noisy image in the signal space has almost no helpful information about the actual direction. It indeed means that the noise effect in the Radon space is much lower, and an appropriate estimator can spot a dominant orientation even with a high power noise.</p>
<fig id="fig-14"><label>Figure 14</label><caption><title>Radon signal space: Radon Transform of the input image <italic>vs</italic>. Noisy image <italic>vs.</italic> Error image</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-14.tif"/></fig>
<p>We compare the <italic>MSE<sup>2D</sup></italic>, <italic>PSNR</italic>, and <italic>SSIM</italic> for the input image and the noisy image in both image signal space and Radon space to perform a more thorough analysis. <xref ref-type="fig" rid="fig-15">Fig. 15</xref> shows the <italic>MSE<sup>2D</sup></italic> for two different spaces in various noise powers. In our implementations, the &#x03C3; was increased from 0 to 65.</p>
<fig id="fig-15"><label>Figure 15</label><caption><title>Comparing the MSE between the image signal space, and the Radon space</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-15.tif"/></fig>
<p>The figure indicates that the MSE<sup>2D&#x2212;SS</sup> (the MSE in the signal space) is growing up exponentially with rising noise power. On the opposite side, MSE<sup>2D&#x2212;RS</sup> (the MSE in the Radon space) is well below the corresponding values in the signal space with little variations of the MSE between &#x03C3;&#x2009;&#x003D;&#x2009;0 to &#x03C3;&#x2009;&#x003D;&#x2009;65.</p>
<p>In addition to the MSE, the PSNR is also illustrated in <xref ref-type="fig" rid="fig-16">Fig. 16</xref> for two different spaces. The achieved results show that there is a big gap between the PSNR in the signal space (PSNR<sup>SS</sup>) (the red line) and the PSNR in the Radon space (PSNR<sup>RS</sup>) (the blue line). Indeed, the PSNR<sup>SS</sup> is dominated by the PSNR<sup>RS</sup>. This indicates the Radon space keeps up the image quality rather than the signal space. Then, in order to remedy some of the issues associated with MSE and PSNR, SSIM is deployed. As mentioned, the SSIM value can vary between &#x2212;1 and 1, where 1 indicates the perfect similarity. <xref ref-type="fig" rid="fig-17">Fig. 17</xref> compares the SSIM between the signal space and the Radon space for different values of the &#x03C3;. The results unveil that SSIM<sup>RS</sup> (SSIM in the Radon space) decreased gradually from 1 to about 0.8 with increasing of the &#x03C3; while SSIM<sup>SS</sup> (SSIM in the signal space) was diving sharply from 1 to almost 0 during similar noise conditions. This indicates the perfect similarity is more achievable in the Radon space rather than signal space, providing that an appropriate feature extraction procedure is available. It is worth mentioning that all inputs were normalized using <xref ref-type="disp-formula" rid="eqn-20">Eq. (20)</xref> before computing the Error Image, <italic>MSE<sup>2D</sup></italic>, <italic>PSNR</italic>, and <italic>SSIM</italic>.
<disp-formula id="eqn-20"><label>(20)</label><mml:math id="mml-eqn-20" display="block"><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>m</mml:mi><mml:mi>a</mml:mi><mml:mi>x</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>m</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:math></disp-formula></p>
<fig id="fig-16"><label>Figure 16</label><caption><title>Comparing the PSNR between the image signal space and the Radon space</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-16.tif"/></fig><fig id="fig-17"><label>Figure 17</label><caption><title>Comparing the SSIM between the image signal space and the Radon space</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-17.tif"/></fig>
<p>In <xref ref-type="disp-formula" rid="eqn-20">Eq. (20)</xref>, <italic>x<sub>S</sub></italic> is the scaled version (the normalized version), and <italic>x</italic> is the input image (either in the signal space or the Radon space). Alongside, <italic>x<sub>min</sub></italic> and <italic>x<sub>max</sub></italic> are the minimum and maximum of the <italic>x</italic>, respectively.</p>
<p>In image processing problems where feature detection is the only need, mapping of an original image from image space to corresponding feature space via a useful transform, with subsequent processing in lower dimension feature space, would be an appropriate. Radon domain, when properly executed, can lead to minimum entropy or maximum sparseness. High-resolution Radon Transform methods can efficiently remove random or correlated noise, improve signal clarity, by utilizing the move-out or curvature of the signal of interest. This article has deployed 2D Mean Square Error (<italic>MSE<sup>2D</sup>), Peak Signal-to-Noise Ratio (PSNR</italic>), and Structural Similarity Measure (<italic>SSIM</italic>) to evaluate the accuracy of the proposed feature extraction algorithm. There are noticeable gaps between Red &#x0026; Blue lines, corresponding to Image &#x0026; Radon spaces for all three evaluation metrics. This indicates the proposed feature detection method is strong enough to benefit the Radon space potential characteristics, such as getting the least entropy or the most sparseness.</p>
<p>In a new scenario, the aim is to evaluate the noise power impact on the fitness function variations. In <xref ref-type="fig" rid="fig-18">Fig. 18</xref>, several fitness functions are drawn for noisy images with &#x03C3;&#x2009;&#x003D;&#x2009;0 (no noise) to &#x03C3;&#x2009;&#x003D;&#x2009;11 with &#x03B8;&#x2009;&#x003D;&#x2009;80. In almost all experiments, the fitness function has a global minimum of about &#x03B8;&#x2009;&#x003D;&#x2009;80 (the perfect estimation). The second extremum is also highlighted in this figure. When the noise power increases sharply, a local minima may change and even be converted to a fake global minimum. This phenomenon will be investigated in the succeeding scenario. In <xref ref-type="fig" rid="fig-19">Fig. 19</xref>, the skew and frequency of the synthetic input image are upgraded. In almost all curves for noisy images with &#x03C3;&#x2009;&#x003D;&#x2009;0 (no noise) to &#x03C3;&#x2009;&#x003D;&#x2009;9, the fitness function has a global minimum of about &#x03B8;&#x2009;&#x003D;&#x2009;30. This means the estimator is doing magnificently even with a relatively high noise power. This indicates that the proposed approach tolerates noise effects marvelously. Now, the noise power increases, manifestly. <xref ref-type="fig" rid="fig-20">Fig. 20</xref>, shows that when &#x03C3; rises from 10 to 45, there is no longer a regular pattern. Besides, for the very high noise powers such as &#x03C3;&#x2009;&#x003D;&#x2009;30,&#x2002;40, and 45 the real extremums were flipped, and replaced by other false local minimums. As a result, in presence of the very high power additive Gaussian noise, the error can be grown sharply.</p>
<fig id="fig-18"><label>Figure 18</label><caption><title>Comparing the fitness function <italic>vs.</italic> &#x03B8; for noisy images from &#x03C3;&#x2009;&#x003D;&#x2009;0 (no noise) to &#x03C3;&#x2009;&#x003D;&#x2009;11 with the estimated <inline-formula id="ieqn-15"><mml:math id="mml-ieqn-15"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula>&#x2009;&#x003D;&#x2009;80</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-18.tif"/></fig><fig id="fig-19"><label>Figure 19</label><caption><title>Comparing the fitness function <italic>vs.</italic> &#x03B8; for noisy images with &#x03C3;&#x2009;&#x003D;&#x2009;0 (no noise) to &#x03C3;&#x2009;&#x003D;&#x2009;9 with the estimated <inline-formula id="ieqn-16"><mml:math id="mml-ieqn-16"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula>&#x2009;&#x003D;&#x2009;30</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-19.tif"/></fig><fig id="fig-20"><label>Figure 20</label><caption><title>Comparing the fitness function <italic>vs.</italic> &#x03B8; for noisy images with &#x03C3;&#x2009;&#x003D;&#x2009;0 (no noise) to &#x03C3;&#x2009;&#x003D;&#x2009;45</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-20.tif"/></fig>
</sec>
<sec id="s4_2"><label>4.2</label><title>Computational Time</title>
<p>To reduce the run-time, the image size can be reduced by a scale factor &#x03B1;. This step can potentially speed up the processing time but it may lead to a reduction in the accuracy. As a result, in this section, the aim is to analyze the proposed re-scaling procedure effects on not only the run-time but also some critical algorithm&#x2019;s performance measures such as the Accuracy, MSE, and MAE. For this purpose, a 65&#x2009;&#x00D7;&#x2009;65 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 is considered in <xref ref-type="fig" rid="fig-21">Fig. 21</xref>. In this new experiment, the scaling factor (&#x03B1;) is selected from the set [1.0, 0.9, 0.7, 0.5], and &#x03B1;&#x2009;&#x003D;&#x2009;1 means there is no re-scaling procedure. <xref ref-type="fig" rid="fig-21">Fig. 21</xref> demonstrates that the global minimum of the fitness function has no drift due to the re-scaling procedure. Therefore, the estimator can detect the skew accurately for even &#x03B1;&#x2009;&#x003D;&#x2009;0.5 for a very tiny input image with the original size of 65&#x2009;&#x00D7;&#x2009;65. Then, the experiment is extended for a 600&#x2009;&#x00D7;&#x2009;600 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 at several scale factors, such as 1.0, 0.9, 0.7, 0.5, 0.4, 0.3, 0.15, and also 0.1. Similarly, the results are very satisfying even for &#x03B1;&#x2009;&#x003D;&#x2009;0.1 for the recent example (See <xref ref-type="fig" rid="fig-22">Fig. 22</xref>).</p>
<fig id="fig-21"><label>Figure 21</label><caption><title>The fitness function for a 65&#x2009;&#x00D7;&#x2009;65 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 at different scale factors</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-21.tif"/></fig><fig id="fig-22"><label>Figure 22</label><caption><title>The fitness function for a 600&#x2009;&#x00D7;&#x2009;600 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 at different scale factors</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-22.tif"/></fig>
<p>To have more discussions, the outcomes of some new experiments are reported in both <xref ref-type="table" rid="table-1">Tables 1</xref>, and <xref ref-type="table" rid="table-2">2</xref>. In these tables, &#x03B1; is the scale factor, Image size<sup>&#x002A;</sup> is the size of the scaled image, <inline-formula id="ieqn-17"><mml:math id="mml-ieqn-17"><mml:mi>&#x03B8;</mml:mi></mml:math></inline-formula> is the dominant orientation of an input image, <inline-formula id="ieqn-18"><mml:math id="mml-ieqn-18"><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula> represents the estimated skew, <italic>Error</italic> <inline-formula id="ieqn-19"><mml:math id="mml-ieqn-19"><mml:mo>=</mml:mo><mml:mi>&#x03B8;</mml:mi><mml:mo>&#x2212;</mml:mo><mml:mrow><mml:mover><mml:mi>&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula>, while <italic>Time</italic> shows the run-time in the second. In <xref ref-type="table" rid="table-1">Table 1</xref>, for all scale factors that range between 0.1, and 1, the estimation error would be precisely zero, while the computation time decreases from 3.16612 s for &#x03B1;&#x2009;&#x003D;&#x2009;1 (Image size<sup>&#x002A;</sup>&#x2009;&#x003D;&#x2009;600&#x2009;&#x00D7;&#x2009;600) to just 0.08803 s for &#x03B1;&#x2009;&#x003D;&#x2009;0.1 (Image size<sup>&#x002A;</sup>&#x2009;&#x003D;&#x2009;60&#x2009;&#x00D7;&#x2009;60). This means the proposed Radon CLF is not only an accurate algorithm but also can reduce the run-time blatantly. In <xref ref-type="table" rid="table-2">Table 2</xref>, the error would be zero except for &#x03B1;&#x2009;&#x003D;&#x2009;0.3, with a remarkably tiny image including only 20 rows and 20 columns of pixels. In this table, the estimation error is still zero for any Image size<sup>&#x002A;</sup> greater than 20&#x2009;&#x00D7;&#x2009;20. This implies that any more reduction of the input size can increase the probability of the error. Furthermore, <xref ref-type="table" rid="table-2">Table 2</xref> denotes that with re-scaling the input image from its original size of 65&#x2009;&#x00D7;&#x2009;65 to 26&#x2009;&#x00D7;&#x2009;26, the run-time falls from 0.09996 s to 0.04103, and at the same time, there is still no error.</p>
<table-wrap id="table-1"><label>Table 1</label><caption><title>The results of the proposed Radon CLF algorithm for a 600&#x2009;&#x00D7;&#x2009;600 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;70 and different scale factors (&#x03B1;)</title></caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Row</th>
<th align="left">Scale factor</th>
<th align="left">Image size<sup>&#x002A;</sup></th>
<th align="left">&#x03B8;</th>
<th align="left"><inline-formula id="ieqn-20"><mml:math id="mml-ieqn-20"><mml:mrow><mml:mover><mml:mi mathvariant="bold-italic">&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula></th>
<th align="left">Error</th>
<th align="left">Time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">1</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;1.00</td>
<td align="left">600&#x2009;&#x00D7;&#x2009;600</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">3.1661</td>
</tr>
<tr>
<td align="left">2</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.90</td>
<td align="left">540&#x2009;&#x00D7;&#x2009;540</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">2.4341</td>
</tr>
<tr>
<td align="left">3</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.70</td>
<td align="left">420&#x2009;&#x00D7;&#x2009;420</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">1.4540</td>
</tr>
<tr>
<td align="left">4</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.50</td>
<td align="left">300&#x2009;&#x00D7;&#x2009;300</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">0.7380</td>
</tr>
<tr>
<td align="left">5</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.40</td>
<td align="left">240&#x2009;&#x00D7;&#x2009;240</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">0.5046</td>
</tr>
<tr>
<td align="left">6</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.30</td>
<td align="left">180&#x2009;&#x00D7;&#x2009;180</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">0.3119</td>
</tr>
<tr>
<td align="left">7</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.15</td>
<td align="left">090&#x2009;&#x00D7;&#x2009;090</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">0.1220</td>
</tr>
<tr>
<td align="left">8</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.10</td>
<td align="left">060&#x2009;&#x00D7;&#x2009;060</td>
<td align="left">70</td>
<td align="left">70</td>
<td align="left">0</td>
<td align="left">0.0880</td>
</tr>
</tbody>
</table>
</table-wrap><table-wrap id="table-2"><label>Table 2</label><caption><title>The results of the proposed Radon CLF algorithm for a 65&#x2009;&#x00D7;&#x2009;65 synthetic image with &#x03B8;&#x2009;&#x003D;&#x2009;10 and different scale factors (&#x03B1;)</title></caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Row</th>
<th align="left">Scale factor</th>
<th align="left">Image size<sup>&#x002A;</sup></th>
<th align="left">&#x03B8;</th>
<th align="left"><inline-formula id="ieqn-21"><mml:math id="mml-ieqn-21"><mml:mrow><mml:mover><mml:mi mathvariant="bold-italic">&#x03B8;</mml:mi><mml:mo stretchy="false">&#x005E;</mml:mo></mml:mover></mml:mrow></mml:math></inline-formula></th>
<th align="left">Error</th>
<th align="left">Time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">1</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;1.00</td>
<td align="left">65&#x2009;&#x00D7;&#x2009;65</td>
<td align="left">10</td>
<td align="left">010</td>
<td align="left">0</td>
<td align="left">0.09996</td>
</tr>
<tr>
<td align="left">2</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.90</td>
<td align="left">58&#x2009;&#x00D7;&#x2009;58</td>
<td align="left">10</td>
<td align="left">010</td>
<td align="left">0</td>
<td align="left">0.07100</td>
</tr>
<tr>
<td align="left">3</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.70</td>
<td align="left">46&#x2009;&#x00D7;&#x2009;46</td>
<td align="left">10</td>
<td align="left">010</td>
<td align="left">0</td>
<td align="left">0.06199</td>
</tr>
<tr>
<td align="left">4</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.50</td>
<td align="left">32&#x2009;&#x00D7;&#x2009;32</td>
<td align="left">10</td>
<td align="left">010</td>
<td align="left">0</td>
<td align="left">0.05199</td>
</tr>
<tr>
<td align="left">5</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.40</td>
<td align="left">26&#x2009;&#x00D7;&#x2009;26</td>
<td align="left">10</td>
<td align="left">010</td>
<td align="left">0</td>
<td align="left">0.04600</td>
</tr>
<tr>
<td align="left">6</td>
<td align="left">&#x03B1;&#x2009;&#x003D;&#x2009;0.30</td>
<td align="left">20&#x2009;&#x00D7;&#x2009;20</td>
<td align="left">10</td>
<td align="left">167</td>
<td align="left">&#x2212;157</td>
<td align="left">0.04103</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><xref ref-type="table" rid="table-3">Table 3</xref> shows the achieved results due to running the Radon CLF algorithm on about 5696 images in the DSI5000 dataset. The dataset has been divided into two parts named DSI5000-p1 and DSI5000-p2. Each part has an equal number of samples, around 2848 images. According to <xref ref-type="table" rid="table-3">Table 3</xref>, the DSI5000-p1 includes images with a lower size (65&#x2009;&#x00D7;&#x2009;65). In comparison with the DSI5000-p2, the DSI5000-p1 has a lower run-time together with a lower accuracy.</p>
<table-wrap id="table-3"><label>Table 3</label><caption><title>Comparing the achieved results of the proposed Radon CLF algorithm for all samples in the dataset</title></caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">DB</th>
<th align="left">Image size</th>
<th align="left">MSE</th>
<th align="left">MAE</th>
<th align="left">Accuracy</th>
<th align="left">Time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">DSI5000-p1</td>
<td align="left">065&#x2009;&#x00D7;&#x2009;065</td>
<td align="left">0.0014060</td>
<td align="left">0.0014071</td>
<td align="left">99.85</td>
<td align="left">0.044</td>
</tr>
<tr>
<td align="left">DSI5000-p2</td>
<td align="left">120&#x2009;&#x00D7;&#x2009;120</td>
<td align="left">0.0000062</td>
<td align="left">0.0000093</td>
<td align="left">99.99</td>
<td align="left">0.087</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In addition to DSI5000, many real-world scanned image documents were also included in our simulations. Both handwriting, and printed samples, at various resolutions, were gathered from different Persian/Arabic &#x0026; English multilingual resources such as books, booklets, letters, and newspapers. <xref ref-type="fig" rid="fig-23">Fig. 23</xref> shows some samples and the result of the Radon CLF algorithm. Our method can accurately detect the skew in real photos, according to experimental results.</p>
<fig id="fig-23"><label>Figure 23</label><caption><title>Examination of the proposed algorithm on real documents</title></caption><graphic mimetype="image" mime-subtype="tif" xlink:href="CSSE_38234-fig-23.tif"/></fig>
<p>Finally, <xref ref-type="table" rid="table-4">Table 4</xref> draws a comparison between the proposed approach and other available algorithms. Experimental results show that our algorithm is capable of skew compensating for large documents far faster than well-known existing methods, with a run-time of about 0.048 s and an Accuracy of 99.87&#x0025; for DSI5000 dataset.</p>
<table-wrap id="table-4"><label>Table 4</label><caption><title>The performance comparison among different algorithms</title></caption>
<table frame="hsides">
<colgroup>
<col align="left"/>
<col align="left"/>
<col align="left"/>
<col align="left"/>
</colgroup>
<thead>
<tr>
<th align="left">Row</th>
<th align="left">Algorithm</th>
<th align="left">Accuracy</th>
<th align="left">Average Time (s)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">1</td>
<td align="left">Radon CLF (the proposed)</td>
<td align="left">99.87&#x0025;</td>
<td align="left">0.048</td>
</tr>
<tr>
<td align="left">2</td>
<td align="left">Al-Shatnawi et al.[<xref ref-type="bibr" rid="ref-20">20</xref>]</td>
<td align="left">87.00&#x0025;</td>
<td align="left">0.390</td>
</tr>
<tr>
<td align="left">3</td>
<td align="left">Yu et al. [<xref ref-type="bibr" rid="ref-46">46</xref>]</td>
<td align="left">97.38&#x0025;</td>
<td align="left">0.420</td>
</tr>
<tr>
<td align="left">4</td>
<td align="left">Chethan et al. [<xref ref-type="bibr" rid="ref-15">15</xref>]</td>
<td align="left">99.73&#x0025;</td>
<td align="left">1.710</td>
</tr>
<tr>
<td align="left">5</td>
<td align="left">Le et al. [<xref ref-type="bibr" rid="ref-38">38</xref>]</td>
<td align="left">99.66&#x0025;</td>
<td align="left">2.330</td>
</tr>
<tr>
<td align="left">6</td>
<td align="left">Sarfraz et al. [<xref ref-type="bibr" rid="ref-47">47</xref>]</td>
<td align="left">99.66&#x0025;</td>
<td align="left">2.330</td>
</tr>
<tr>
<td align="left">7</td>
<td align="left">Narasimha et al. [<xref ref-type="bibr" rid="ref-7">7</xref>]</td>
<td align="left">99.20&#x0025;</td>
<td align="left">17.90</td>
</tr>
<tr>
<td align="left">8</td>
<td align="left">Ravikumar et al. [<xref ref-type="bibr" rid="ref-48">48</xref>]</td>
<td align="left">98.00&#x0025;</td>
<td align="left">1.500</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec id="s5"><label>5</label><title>Conclusion and Future Works</title>
<p>In this paper, we proposed a novel, fast, and reliable skew detection algorithm for text images based on the Radon Transform and Curve Length Fitness Function (CLF). In addition, approximately 5696 synthetic images were incorporated into a new dataset called DSI5000. Many real image documents were also included in our simulations along with synthetic images. From various Persian, Arabic, English, and multilingual sources, random handwriting and printed samples of some books, booklets, letters, and newspapers, were collected at different resolutions.</p>
<p>The resilience of signal and image processing algorithms against noise effects is one of the most crucial issues. Through the utilization of many performance indicators, such as accuracy, MSE, MAE, PSNR, SSIM, and error signal comparison in both the signal space and the Radon space, we have created a detailed picture of the noise tolerance of our model in this study. Our approach is superior to other existing methods in terms of accuracy as well as timing efficiency, as shown by the results with the Accuracy of about 99.87&#x0025; &#x0026; run-time of around 0.048 (s) for DSI5000 dataset. For multilingual manuscripts with various font types, sizes, and styles, the suggested Radon CLF approach could find skews between 0<sup>&#x00B0;</sup> and 90<sup>&#x00B0;</sup>.</p>
<p>Machine Learning (ML) is a fast-growing and interesting field of applied research with high demands in scientific communities and advanced technologies. Deep Learning (DL) is a branch of ML that makes use of Artificial Neural Networks (ANN) to simulate how the human brain learns [<xref ref-type="bibr" rid="ref-49">49</xref>&#x2013;<xref ref-type="bibr" rid="ref-51">51</xref>]. In the future, we intend to utilize DL models. They can be used to develop Radon CLF method for other computer vision applications, such as Camera Rotations Automatic Recovery, Rotation estimation in the urban environment, Fingerprint Recognition etc., which are particularly sensitive to directional patterns. Directional patterns have two main attributes: Dominant Orientation and Frequency. In our future studies, we will focus more on the joint estimation of both features using deep learning and Radon CLF.</p>
</sec>
</body>
<back>
<sec><title>Funding Statement</title>
<p>There are no sources of funding that have supported this work.</p></sec>
<sec sec-type="COI-statement"><title>Conflicts of Interest</title>
<p>The authors declare that they have no conflicts of interest to report regarding the present study.</p></sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Chaudhuri</surname></string-name>, <string-name><given-names>K.</given-names><surname>Mandaviya</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Badelia</surname></string-name> and <string-name><given-names>S. K.</given-names> <surname>Ghosh</surname></string-name></person-group>, <source> Optical Character Recognition Systems for Different Languages with Soft Computing</source>, vol. <volume>352</volume>. <publisher-loc>New York,
USA</publisher-loc>: <publisher-name>Springer</publisher-name>, pp. <fpage>9</fpage>&#x2013;<lpage>41</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Khorsheed</surname></string-name></person-group>, &#x201C;<article-title>Offline recognition of omni font Arabic text using the HMM ToolKit (HTK)</article-title>,&#x201D; <source>Pattern Recognition Letters</source>, vol. <volume>28</volume>, no. <issue>12</issue>, pp. <fpage>1563</fpage>&#x2013;<lpage>1571</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>B. V.</given-names> <surname>Dhandra</surname></string-name>, <string-name><given-names>V. S.</given-names> <surname>Malemath</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Mallikarjun</surname></string-name></person-group>, &#x201C;<article-title>Skew detection in binary image documents based on image dilation and region labeling approach</article-title>,&#x201D; in <conf-name>18th Int. Conf. on Pattern Recognition (ICPR&#x2019;06)</conf-name>,<conf-loc>Hong Kong, China</conf-loc>, pp. <fpage>954</fpage>&#x2013;<lpage>957</lpage>, <year>2006</year>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Dehbovid</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Razzazi</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Alirezaii</surname></string-name></person-group>, &#x201C;<article-title>A novel method for de-warping in Persian document images captured by cameras</article-title>,&#x201D; in <conf-name>2010 Int. Conf. on Computer Information Systems and Industrial Management Applications (CISIM)</conf-name>, <conf-loc>Krakow, Poland</conf-loc>, pp. <fpage>614</fpage>&#x2013;<lpage>619</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>T.</given-names> <surname>Nawaz</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Naqvi</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Rehman</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Faiz</surname></string-name></person-group>, &#x201C;<article-title>Optical character recognition system for Urdu (Naskh font) using pattern matching technique</article-title>,&#x201D; <source>International Journal of Image Processing (IJIP)</source>, vol. <volume>3</volume>, no. <issue>3</issue>, pp. <fpage>92</fpage>&#x2013;<lpage>104</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Salagar</surname></string-name></person-group>, &#x201C;<chapter-title>Analysis of PCA usage to detect and correct skew in document images</chapter-title>,&#x201D; in <source>Information and Communication Technology for Competitive Strategies</source>, vol.<volume>191</volume>. <publisher-loc>New York, USA</publisher-loc>: <publisher-name>Springer</publisher-name>, pp. <fpage>687</fpage>&#x2013;<lpage>695</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>R. S.</given-names> <surname>Narasimha</surname></string-name> and <string-name><given-names>S. D.</given-names> <surname>Parag</surname></string-name></person-group>, &#x201C;<article-title>A novel local skew correction and segmentation approach for printed multilingual Indian documents</article-title>,&#x201D; <source>Alexandria Engineering Journal</source>, vol. <volume>57</volume>, no. <issue>3</issue>, pp. <fpage>1609</fpage>&#x2013;<lpage>1618</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Pechwitz</surname></string-name> and <string-name><given-names>V.</given-names> <surname>Margner</surname></string-name></person-group>, &#x201C;<article-title>Baseline estimation for Arabic handwritten words</article-title>,&#x201D; in <conf-name>Proc. Eighth Int. Workshop on Frontiers in Handwriting Recognition</conf-name>, <conf-loc>Niagra-on-the-Lake, ON, Canada</conf-loc>, pp. <fpage>479</fpage>&#x2013;<lpage>484</lpage>, <year>2002</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Huang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Jin</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Cao</surname></string-name></person-group>, &#x201C;<article-title>Hough transform-based multi-object autofocusing compressive holography</article-title>,&#x201D; <source>Applied Optics</source>, vol. <volume>62</volume>, pp. <fpage>23</fpage>&#x2013;<lpage>30</lpage>, <year>2023</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Kumar</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Karibasappa</surname></string-name></person-group>, &#x201C;<article-title>An approach for brain tumour detection based on dual-tree complex Gabor wavelet transform and neural network using Hadoop big data analysis</article-title>,&#x201D; <source>Multimedia Tools and Applications</source>, vol. <volume>81</volume>, pp. <fpage>39251</fpage>&#x2013;<lpage>39274</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Mukhopadhyay</surname></string-name> and <string-name><given-names>B. B.</given-names> <surname>Chaudhuri</surname></string-name></person-group>, &#x201C;<article-title>A survey of Hough transform</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>48</volume>,no. <issue>3</issue>, pp. <fpage>993</fpage>&#x2013;<lpage>1010</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Chen</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Zhang</surname></string-name></person-group>, &#x201C;<article-title>Multi-stage directional median filter</article-title>,&#x201D; <source>International Journal of Signal Processing</source>, vol. <volume>5</volume>, no. <issue>4</issue>, pp. <fpage>249</fpage>&#x2013;<lpage>252</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Bahaghighat</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Abedini</surname></string-name>, <string-name><given-names>Q.</given-names> <surname>Xin</surname></string-name>, <string-name><given-names>M. M.</given-names> <surname>Zanjireh</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Mirjalili</surname></string-name></person-group>, &#x201C;<article-title>Using machine learning and computer vision to estimate the angular velocity of wind turbines in smart grids remotely</article-title>,&#x201D; <source>Energy Reports</source>, vol. <volume>7</volume>, pp. <fpage>8561</fpage>&#x2013;<lpage>8576</lpage>, <year>2021</year>.</mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y. Y.</given-names> <surname>Tang</surname></string-name>, <string-name><given-names>S. W.</given-names> <surname>Lee</surname></string-name> and <string-name><given-names>C. Y.</given-names> <surname>Suen</surname></string-name></person-group>, &#x201C;<article-title>Automatic document processing: A survey</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>29</volume>, no. <issue>12</issue>, pp. <fpage>1931</fpage>&#x2013;<lpage>1952</lpage>, <year>1996</year>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H. K.</given-names> <surname>Chethan</surname></string-name> and <string-name><given-names>G. H.</given-names> <surname>Kumar</surname></string-name></person-group>, &#x201C;<article-title>Graphics separation and skew correction for mobile captured documents and comparative analysis with existing methods</article-title>,&#x201D; <source>International Journal of Computer Applications</source>, vol. <volume>7</volume>, no. <issue>3</issue>, pp. <fpage>42</fpage>&#x2013;<lpage>47</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>P. V.</given-names> <surname>Bezmaternykh</surname></string-name> and <string-name><given-names>D. P.</given-names> <surname>Nikolaev</surname></string-name></person-group>, &#x201C;<article-title>A document skew detection method using fast Hough transform</article-title>,&#x201D; in <conf-name>Twelfth Int. Conf. on Machine Vision (ICMV 2019)</conf-name>, <conf-loc>Amsterdam, Netherlands</conf-loc>, vol. pp. <fpage>11433</fpage>, pp. <fpage>132</fpage>&#x2013;<lpage>137</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>L. Y.</given-names> <surname>Chiu</surname></string-name>, <string-name><given-names>Y. Y.</given-names> <surname>Tang</surname></string-name> and <string-name><given-names>C. Y.</given-names> <surname>Suen</surname></string-name></person-group>, &#x201C;<article-title>Document skew detection based on the fractal and least squares method</article-title>,&#x201D; in <conf-name>Proc. of 3rd Int. Conf. on Document Analysis and Recognition</conf-name>, <conf-loc>Montreal, QC, Canada</conf-loc>, vol. <volume>12</volume>, pp. <fpage>1149</fpage>&#x2013;<lpage>1152</lpage>, <year>1995</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Shivakumara</surname></string-name>, <string-name><given-names>D. S.</given-names> <surname>Guru</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Hemantha</surname></string-name> and <string-name><given-names>P.</given-names> <surname>Nagabhushan</surname></string-name></person-group>, &#x201C;<article-title>A novel technique for estimation of skew in binary text document images based on linear regression analysis</article-title>,&#x201D; <source>Sadhana</source>, vol. <volume>30</volume>, no. <issue>1</issue>, pp. <fpage>69</fpage>&#x2013;<lpage>85</lpage>, <year>2005</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>P.</given-names> <surname>Shivakumara</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Hemantha</surname></string-name>, <string-name><given-names>D. S.</given-names> <surname>Guru</surname></string-name> and <string-name><given-names>P.</given-names> <surname>Nagabhushan</surname></string-name></person-group>, &#x201C;<article-title>Skew estimation of binary document images using static and dynamic thresholds useful for document image mosaicking</article-title>,&#x201D; in <conf-name>National Workshop on IT Services and Applications (WITSA)</conf-name>, <conf-loc>New Delhi, India</conf-loc>, pp. <fpage>27</fpage>&#x2013;<lpage>28</lpage>, <year>2003</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Al-Shatnawi</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Omar</surname></string-name></person-group>, &#x201C;<article-title>Skew detection and correction technique for Arabic document images based on centre of gravity</article-title>,&#x201D; <source>Journal of Computer Science</source>, vol. <volume>5</volume>, no. <issue>5</issue>, pp. <fpage>363</fpage>&#x2013;<lpage>368</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>T.</given-names> <surname>Jundale</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Hegadi</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Ravindra</surname></string-name></person-group>, &#x201C;<article-title>Skew detection and correction of Devanagari script using Hough transform</article-title>,&#x201D; <source>Procedia Computer Science</source>, vol. <volume>45</volume>, pp. <fpage>305</fpage>&#x2013;<lpage>311</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Singh</surname></string-name> and <string-name><given-names>R.</given-names> <surname>Kaur</surname></string-name></person-group>, &#x201C;<article-title>Improved skew detection and correction approach using Discrete Fourier algorithm</article-title>,&#x201D; <source>International Journal of Soft Computing and Engineering</source>, vol. <volume>3</volume>, no. <issue>4</issue>, pp. <fpage>5</fpage>&#x2013;<lpage>7</lpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y.</given-names> <surname>Lu</surname></string-name> and <string-name><given-names>C. L.</given-names> <surname>Tan</surname></string-name></person-group>, &#x201C;<article-title>A nearest-neighbor chain based approach to skew estimation in document images</article-title>,&#x201D; <source>Pattern Recognition Letters</source>, vol. <volume>24</volume>, no. <issue>14</issue>, pp. <fpage>2315</fpage>&#x2013;<lpage>2323</lpage>, <year>2003</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Chang</surname></string-name></person-group>, &#x201C;<article-title>Application of principal component analysis in image signal processing</article-title>,&#x201D; in <conf-name>Proc. Int. Conf. on Image, Signal Processing, and Pattern Recognition (ISPP 2022)</conf-name>, <conf-loc>Guilin, China</conf-loc>, pp. <fpage>12</fpage>&#x2013;<lpage>24</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>W.</given-names> <surname>Bao</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Yang</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Wen</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Zeng</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Guo</surname></string-name> <etal>et al.,</etal></person-group> &#x201C;<article-title>A novel adaptive deskewing algorithm for document images</article-title>,&#x201D; <source>Sensors</source>, vol. <volume>22</volume>, no. <issue>20</issue>, pp. <fpage>7944</fpage>&#x2013;<lpage>7962</lpage>, <year>2022</year>; <pub-id pub-id-type="pmid">36298294</pub-id></mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>Y. L.</given-names> <surname>Chaitra</surname></string-name> and <string-name><given-names>R.</given-names> <surname>Dinesh</surname></string-name></person-group>, &#x201C;<chapter-title>An impact of radon transforms and filtering techniques for text localization in natural scene text images</chapter-title>,&#x201D; in <source>ICT with Intelligent Applications</source>, vol. <volume>248</volume>. <publisher-loc>New York, USA</publisher-loc>: <publisher-name>Springer</publisher-name>, pp. <fpage>563</fpage>&#x2013;<lpage>573</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>P. C.</given-names> <surname>Theofanopoulos</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Sakr</surname></string-name> and <string-name><given-names>G. C.</given-names> <surname>Trichopoulos</surname></string-name></person-group>, &#x201C;<article-title>Multistatic terahertz imaging using the Radon transform</article-title>,&#x201D; <source>IEEE Transactions on Antennas and Propagation</source>, vol. <volume>67</volume>, no. <issue>4</issue>, pp. <fpage>2700</fpage>&#x2013;<lpage>2709</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L. J.</given-names> <surname>Nelson</surname></string-name> and <string-name><given-names>R. A.</given-names> <surname>Smith</surname></string-name></person-group>, &#x201C;<article-title>Fibre direction and stacking sequence measurement in carbon fibre composites using Radon transforms of ultrasonic data</article-title>,&#x201D; <source>Composites Part A: Applied Science and Manufacturing</source>, vol. <volume>118</volume>, pp. <fpage>1</fpage>&#x2013;<lpage>8</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Beylkin</surname></string-name></person-group>, &#x201C;<article-title>Discrete radon transform</article-title>,&#x201D; <source>IEEE Transactions on Acoustics, Speech, and Signal Processing</source>, vol. <volume>35</volume>, no. <issue>2</issue>, pp. <fpage>162</fpage>&#x2013;<lpage>172</lpage>, <year>1987</year>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Aftab</surname></string-name>, <string-name><given-names>S. F.</given-names> <surname>Ali</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Mahmood</surname></string-name> and <string-name><given-names>U.</given-names> <surname>Suleman</surname></string-name></person-group>, &#x201C;<article-title>A boosting framework for human posture recognition using Spatio-temporal features along with radon transform</article-title>,&#x201D; <source>Multimedia Tools and Applications</source>, vol. <volume>81</volume>, pp. <fpage>42325</fpage>&#x2013;<lpage>42351</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Grathwohl</surname></string-name>, <string-name><given-names>P. C.</given-names> <surname>Kunstmann</surname></string-name>, <string-name><given-names>E. T.</given-names> <surname>Quinto</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Rieder</surname></string-name></person-group>, &#x201C;<article-title>Imaging with the elliptic Radon transform in three dimensions from an analytical and numerical perspective</article-title>,&#x201D; <source>SIAM Journal on Imaging Sciences</source>, vol. <volume>13</volume>, no. <issue>4</issue>, pp. <fpage>2250</fpage>&#x2013;<lpage>2280</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Chelbi</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Mekhmoukh</surname></string-name></person-group>, &#x201C;<article-title>Features based image registration using cross-correlation and Radon transform</article-title>,&#x201D; <source>Alexandria Engineering Journal</source>, vol. <volume>57</volume>, no. <issue>4</issue>, pp. <fpage>2313</fpage>&#x2013;<lpage>2318</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="thesis"><person-group person-group-type="author"><string-name><given-names>Q.</given-names> <surname>Zhang</surname></string-name></person-group>, &#x201C;<article-title>Automated road network extraction from high spatial resolution multi-spectral imagery</article-title>,&#x201D; <source>Ph.D. Dissertation</source>, <publisher-name>University of Calgary</publisher-name>, <publisher-loc>Canada</publisher-loc>, <year>2006</year>.</mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>J. S.</given-names> <surname>Seo</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Haitsma</surname></string-name>, <string-name><given-names>T.</given-names> <surname>Kalker</surname></string-name> and <string-name><given-names>C. D.</given-names> <surname>Yoo</surname></string-name></person-group>, &#x201C;<article-title>A robust image fingerprinting system using the Radon transform</article-title>,&#x201D; <source>Signal Processing: Image Communication</source>, vol. <volume>19</volume>, no. <issue>4</issue>, pp. <fpage>325</fpage>&#x2013;<lpage>339</lpage>, <year>2004</year>.</mixed-citation></ref>
<ref id="ref-35"><label>[35]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>V.</given-names> <surname>Kiani</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Pourreza</surname></string-name> and <string-name><given-names>H. R.</given-names> <surname>Pourreza</surname></string-name></person-group>, &#x201C;<article-title>Offline signature verification using local radon transform and support vector machines</article-title>, &#x201D; <source>International Journal of Image Processing</source>, vol. <volume>3</volume>, no. <issue>5</issue>, pp. <fpage>184</fpage>&#x2013;<lpage>194</lpage>, <year>2009</year>.</mixed-citation></ref>
<ref id="ref-36"><label>[36]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Beckmann</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Bhandari</surname></string-name> and <string-name><given-names>F.</given-names> <surname>Krahmer</surname></string-name></person-group>, &#x201C;<article-title>The modulo Radon transform: Theory, algorithms, and applications</article-title>,&#x201D; <source>SIAM Journal on Imaging Sciences</source>, vol. <volume>15</volume>, no. <issue>2</issue>, pp. <fpage>455</fpage>&#x2013;<lpage>490</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-37"><label>[37]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S. H.</given-names> <surname>Mohana</surname></string-name> and <string-name><given-names>C. J.</given-names> <surname>Prabhakar</surname></string-name></person-group>, &#x201C;<article-title>Stem-Calyx recognition of an apple using shape descriptors</article-title>,&#x201D; <source>Signal &#x0026; Image Processing : An International Journal (SIPIJ)</source>, vol. <volume>5</volume>, no. <issue>6</issue>, pp. <fpage>17</fpage>&#x2013;<lpage>31</lpage>, <year>2014</year>.</mixed-citation></ref>
<ref id="ref-38"><label>[38]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>D. S.</given-names> <surname>Le</surname></string-name>, <string-name><given-names>G. R.</given-names> <surname>Thoma</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Wechsler</surname></string-name></person-group>, &#x201C;<article-title>Automated page orientation and skew angle detection for binary document images</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>27</volume>, no. <issue>10</issue>, pp. <fpage>1325</fpage>&#x2013;<lpage>1344</lpage>, <year>1994</year>.</mixed-citation></ref>
<ref id="ref-39"><label>[39]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Chikkerur</surname></string-name>, <string-name><given-names>A. N.</given-names> <surname>Cartwright</surname></string-name> and <string-name><given-names>V.</given-names> <surname>Govindaraju</surname></string-name></person-group>, &#x201C;<article-title>Fingerprint enhancement using STFT analysis</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>40</volume>, no. <issue>1</issue>, pp. <fpage>198</fpage>&#x2013;<lpage>211</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-40"><label>[40]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>U.</given-names> <surname>Sara</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Akter</surname></string-name> and <string-name><given-names>M. S.</given-names> <surname>Uddin</surname></string-name></person-group>, &#x201C;<article-title>Image quality assessment through FSIM, SSIM, MSE and PSNR&#x2014;A comparative study</article-title>,&#x201D; <source>Journal of Computer and Communications</source>, vol. <volume>7</volume>, no. <issue>3</issue>, pp. <fpage>8</fpage>&#x2013;<lpage>18</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-41"><label>[41]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Z.</given-names> <surname>Cui</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Gan</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Tang</surname></string-name>, <string-name><given-names>F.</given-names> <surname>Liu</surname></string-name> and <string-name><given-names>X.</given-names> <surname>Zhu</surname></string-name></person-group>, &#x201C;<article-title>Image signature based mean square error for image quality assessment</article-title>,&#x201D; <source>Chinese Journal of Electronics</source>, vol. <volume>24</volume>, no. <issue>4</issue>, pp. <fpage>755</fpage>&#x2013;<lpage>760</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-42"><label>[42]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Bahaghighat</surname></string-name>, <string-name><given-names>S. A.</given-names> <surname>Motamedi</surname></string-name> and <string-name><given-names>Q.</given-names> <surname>Xin</surname></string-name></person-group>, &#x201C;<article-title>Image transmission over cognitive radio networks for smart grid applications</article-title>,&#x201D; <source>Applied Sciences</source>, vol. <volume>9</volume>, no. <issue>24</issue>, pp. <fpage>5498</fpage>&#x2013;<lpage>5529</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-43"><label>[43]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Corona</surname></string-name>, <string-name><given-names>O.</given-names> <surname>Maciel-Castillo</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Morales-Casta&#x00F1;eda</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Gonzalez</surname></string-name></person-group>, &#x201C;<article-title>A new method to solve rotated template matching using metaheuristic algorithms and the structural similarity index</article-title>,&#x201D; <source>Mathematics and Computers in Simulation</source>, vol. <volume>206</volume>, pp. <fpage>130</fpage>&#x2013;<lpage>146</lpage>, <year>2023</year>.</mixed-citation></ref>
<ref id="ref-44"><label>[44]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Bhatt</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Naik</surname></string-name> and <string-name><given-names>V. K.</given-names> <surname>Subramanian</surname></string-name></person-group>, &#x201C;<article-title>SSIM compliant modeling framework with denoising and deblurring applications</article-title>,&#x201D; <source>IEEE Transactions on Image Processing</source>, vol. <volume>30</volume>, pp. <fpage>2611</fpage>&#x2013;<lpage>2626</lpage>, <year>2021</year>; <pub-id pub-id-type="pmid">33502978</pub-id></mixed-citation></ref>
<ref id="ref-45"><label>[45]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Bafjaish</surname></string-name>, <string-name><given-names>M. S.</given-names> <surname>Azmi</surname></string-name>, <string-name><given-names>M. N.</given-names> <surname>Al-Mhiqani</surname></string-name>, <string-name><given-names>A. R.</given-names> <surname>Radzid</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Mahdin</surname></string-name></person-group>, &#x201C;<article-title>Skew detection and correction of Mushaf Al-Quran script using hough transform</article-title>,&#x201D; <source>International Journal of Advanced Computer Science and Applications</source>, vol. <volume>9</volume>, no. <issue>8</issue>, pp. <fpage>402</fpage>&#x2013;<lpage>409</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-46"><label>[46]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>B.</given-names> <surname>Yu</surname></string-name> and <string-name><given-names>A. K.</given-names> <surname>Jain</surname></string-name></person-group>, &#x201C;<article-title>A robust and fast skew detection algorithm for generic documents</article-title>,&#x201D; <source>Pattern Recognition</source>, vol. <volume>29</volume>, no. <issue>10</issue>, pp. <fpage>1599</fpage>&#x2013;<lpage>1629</lpage>, <year>1996</year>.</mixed-citation></ref>
<ref id="ref-47"><label>[47]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Sarfraz</surname></string-name>, <string-name><given-names>S. A.</given-names> <surname>Mahmoud</surname></string-name> and <string-name><given-names>Z.</given-names> <surname>Rasheed</surname></string-name></person-group>, &#x201C;<article-title>On skew estimation and correction of text, computer graphics</article-title>,&#x201D; in <conf-name>Proc. of Computer Graphics, Imaging and Visualization Conf. (CGIV 2007)</conf-name>, <conf-loc>Bangkok, Thailand</conf-loc>, pp. <fpage>308</fpage>&#x2013;<lpage>313</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-48"><label>[48]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M.</given-names> <surname>Ravikumar</surname></string-name> and <string-name><given-names>O. A.</given-names> <surname>Boraik</surname></string-name></person-group>, &#x201C;<article-title>Estimation and correction of multiple skew Arabic handwritten document images</article-title>,&#x201D; <source>International Conference on Innovative Computing and Communications</source>, vol. <volume>1387</volume>, pp. <fpage>553</fpage>&#x2013;<lpage>564</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-49"><label>[49]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><given-names>A.</given-names> <surname>Hajikarimi</surname></string-name> and <string-name><given-names>M.</given-names> <surname>Bahaghighat</surname></string-name></person-group>, &#x201C;<chapter-title>Optimum outlier detection in internet of things industries using autoencoder</chapter-title>,&#x201D; in <source>Frontiers in Nature-Inspired Industrial Optimization</source>, <publisher-loc>New York, USA</publisher-loc>: <publisher-name>Springer</publisher-name>, pp. <fpage>77</fpage>&#x2013;<lpage>92</lpage>, <year>2022</year>.</mixed-citation></ref>
<ref id="ref-50"><label>[50]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>A. M.</given-names> <surname>Hilal</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Al-Rasheed</surname></string-name>, <string-name><given-names>J. S.</given-names> <surname>Alzahrani</surname></string-name>, <string-name><given-names>M. M.</given-names> <surname>Eltahir</surname></string-name> and <string-name><given-names>M. A.</given-names> <surname>Duhayyim</surname></string-name></person-group>, &#x201C;<article-title>Competitive multi-verse optimization with deep learning based sleep stage classification</article-title>,&#x201D; <source>Computer Systems Science and Engineering</source>, vol. <volume>45</volume>, no. <issue>2</issue>, pp. <fpage>1249</fpage>&#x2013;<lpage>1263</lpage>, <year>2023</year>.</mixed-citation></ref>
<ref id="ref-51"><label>[51]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. J.</given-names> <surname>Umer</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Sharif</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Alhaisoni</surname></string-name>, <string-name><given-names>U.</given-names> <surname>Tariq</surname></string-name> and <string-name><given-names>Y. J.</given-names> <surname>Kim</surname></string-name></person-group>, &#x201C;<article-title>A framework of deep learning and selection-based breast cancer detection from histopathology images</article-title>,&#x201D; <source>Computer Systems Science and Engineering</source>, vol. <volume>45</volume>, no. <issue>2</issue>, pp. <fpage>1001</fpage>&#x2013;<lpage>1016</lpage>, <year>2023</year>.</mixed-citation></ref>
</ref-list>
</back>
</article>