<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">IASC</journal-id>
<journal-id journal-id-type="nlm-ta">IASC</journal-id>
<journal-id journal-id-type="publisher-id">IASC</journal-id>
<journal-title-group>
<journal-title>Intelligent Automation &#x0026; Soft Computing</journal-title>
</journal-title-group>
<issn pub-type="epub">2326-005X</issn>
<issn pub-type="ppub">1079-8587</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">20032</article-id>
<article-id pub-id-type="doi">10.32604/iasc.2022.020032</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Generating Synthetic Trajectory Data Using GRU</article-title><alt-title alt-title-type="left-running-head">Generating Synthetic Trajectory Data Using GRU</alt-title><alt-title alt-title-type="right-running-head">Generating Synthetic Trajectory Data Using GRU</alt-title>
</title-group>
<contrib-group content-type="authors">
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Liu</surname><given-names>Xinyao</given-names></name>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Cui</surname><given-names>Baojiang</given-names></name>
<xref ref-type="aff" rid="aff-1">1</xref><email>cuibj@bupt.edu.cn</email>
</contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Xing</surname><given-names>Lantao</given-names></name>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<aff id="aff-1"><label>1</label><institution>Beijing University of Posts and Telecommunications</institution>, <addr-line>Beijing, 100876</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>Nanyang Technological University</institution>, <addr-line>Nanyang Avenue, 639798</addr-line>, <country>Singapore</country></aff>
</contrib-group>
<author-notes><corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Baojiang Cui. Email: <email>cuibj@bupt.edu.cn</email></corresp></author-notes>
<pub-date pub-type="epub" date-type="pub" iso-8601-date="2022-04-13"><day>13</day>
<month>04</month>
<year>2022</year></pub-date>
<volume>34</volume>
<issue>1</issue>
<fpage>295</fpage>
<lpage>305</lpage>
<history>
<date date-type="received"><day>06</day><month>5</month><year>2021</year></date>
<date date-type="accepted"><day>15</day><month>6</month><year>2021</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2022 Liu, Cui and Xing</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Liu, Cui and Xing</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_IASC_20032.pdf"></self-uri>
<abstract>
<p>With the rise of mobile network, user location information plays an increasingly important role in various mobile services. The analysis of mobile users&#x2019; trajectories can help develop many novel services or applications, such as targeted advertising recommendations, location-based social networks, and intelligent navigation. However, privacy issues limit the sharing of such data. The release of location data resulted in disclosing users&#x2019; privacy, such as home addresses, medical records, and other living habits. That promotes the development of trajectory generators, which create synthetic trajectory data by simulating moving objects. At current, there are some disadvantages in the process of generation. The prediction of the following position in the trajectory generation is very dependent on the historical location data, but the relationship between trajectory positions tends to be ignored. Most commonly used methods only adopt the probability distribution of users&#x2019; positions to generate synthetic data. On the one hand, this type of statistical method is too rough, and on the other hand, it cannot bring more benefits in availability by increasing data volume. We propose a new trajectory generation method in this paper&#x2013;Trajectory Generation Model with RNNs(TGMRNN), to address the deficiencies above. It adopts the RNN model to replace the traditional Markov model to generate trajectory data with higher availability. Meanwhile, it solves the problem that RNNs are unsuitable for continuous location data by representing trajectories as discretized data with the grid method. We have conducted experiments in a real data set. Compared with the Markov model, the results of TGMRNN demonstrate that it is superior to some existing methods.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Synthesis trajectory</kwd>
<kwd>GRU</kwd>
<kwd>grid method</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>Current intelligent services are data-driven, and the provision and optimization of service are based on massive user data. And when service providers use user&#x2019; data, they will inevitably violate the user&#x2019;s privacy. With the development of communication technology, positioning technology has been more widely applied. Mobile phones, computers, cars, and other devices, used in our daily life are all equipped with positioning technology. Newzoo Global Mobile Market Report, published in 2020, shows that The number of global smartphone users will reach 3.5 billion in 2020, up 6.7% year-on-year, and the number will reach 4.1 billion in 2023. Research on location-based services has made significant progress. Three main factors promote its rapid development. The first factor is that people are aware of the importance of human mobility research, which is very important to guide social construction, such as disease transmission, urban planning, and personal location services. The second factor is the rapid development of sensor devices. Our ability to collect location data information has been significantly improved with the rapid popularization and development of sensor devices (such as smartphones, wearable devices, onboard sensors, etc.). The third factor is the emergence of premium location services. More and more location-based services and applications, such as AutoNavi, Didi, Meituan, and others, collect users&#x2019; locations to understand their spatiotemporal mobility patterns. And they are based on location to provide users with convenient life and rich entertainment [<xref ref-type="bibr" rid="ref-1">1</xref>]. But, some malicious applications and unreliable third parties, which are mounted on these devices and provide a variety of location-based services (LBS), may violate users&#x2019; location privacy by acquiring and analyzing the subscriber&#x2019;s location information [<xref ref-type="bibr" rid="ref-2">2</xref>,<xref ref-type="bibr" rid="ref-3">3</xref>]. The lawbreakers can infer the user&#x2019;s sensitive information, such as a family address, medical records, religion, and other privacy. They may commit a crime based on location privacy. Therefore, a reliable and effective location privacy protection mechanism is urgently needed by users.</p>
<p>Trajectory generators can well solve the above problems. Namely, on the premise of guaranteeing the users&#x2019; location privacy, it can publish the highly available synthetic trajectory data [<xref ref-type="bibr" rid="ref-4">4</xref>&#x2013;<xref ref-type="bibr" rid="ref-6">6</xref>]. The trajectory generators extract features and model the moving pattern of users&#x2019; to train the model. The synthetic data generated by the trained model should be highly similar to the original data, and it also could be used in data mining. At present, the main research content in trajectory generation is how to improve the availability of generated data under the premise of privacy. Researchers invest in the analysis and mining of mobile objects&#x2019; location information and produce many research results. But most of the researches focus on analyzing the historical trajectory of moving objects and mining exciting information, and there are a few types of research on trajectory generation technology.</p>
<p>At present, the mainstream trajectory generation method is to generate synthetic trajectory data through the statistical distribution of the data set, constructing prefix tree or the Markov model etc. Compared with the availability, researchers pay more attention to protecting the privacy of the generated data. Hence, more research on the availability of the generated data needs to be invested [<xref ref-type="bibr" rid="ref-7">7</xref>,<xref ref-type="bibr" rid="ref-8">8</xref>]. The Markov model used in the current study has some shortcomings. First, the Markov model is not enough to process some complex data, and the accuracy of the Markov model is relatively low. Second, the simple low-order Markov model works better than the high-order Markov model in the location prediction. Still, low-order Markov can not obtain the context of the high dimensional sequence [<xref ref-type="bibr" rid="ref-9">9</xref>].</p>
<p>To address these problems, we have presented the Trajectory Generation Model with RNNs (TGMRNN), which analyzes the historical location data of moving objects to generate synthetic trajectory data. Recursive neural networks can learn and model these hidden patterns contained in the original trajectory dataset. Trajectory data is sequential data having long-term temporal dependence. In practice, we need to analyze the trajectory data for complex and hidden movement patterns. The position sequence is sequential, and there is a context relation between the positions. The model should be able to continuously learn the characteristics of the position before and after the current position. This paper is to generate the high availability data set by processing and analyzing the historical tracks. Recurrent neural networks (RNN) can store memory through connections with feedback nerves. Besides, RNN networks have a built-in advantage in learning and generating data with long-term time dependence. Therefore, we choose to use RNNs to achieve the generation of forged data sets.</p>
<p>The contributions in this paper are summarized as follow:</p>
<p>(1) We transform location data to trajectory sequence by grid method. The location information mainly consisted of latitude and longitude, which is continuous. The grid method divides the geographic space into different cells and then takes the corresponding cell number as the number of points falling into the corresponding cell. The transformation from continuous data to discrete data is realized. And we convert the problem of continuous location generation to that of sequence generation.</p>
<p>(2) We adopt the GRU model to generate the trajectory, ensuring that the trajectory generator could process high-dimensional sequences and preserve the relationship between the positions. Markov model is only effective in processing low-order (1-order or 2-order) sequences relative to the GRU model. We design the training and generating process using the many-to-one pattern, combined with the N-gram method, which can customize the length of the processing sequence.</p>
<p>(3) We conduct experiments on a real dataset, demonstrating the effectiveness of TGMRNN compared with the mainstream model.</p>
<p>The rest of this paper is organized as follows: we review relevant related work in Section 2. Section 3 describes the Trajectory Generation Model using RNNs in detail. The experimental results and performance analysis are presented in Section 4. Section 5 concludes this paper.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Related Work</title>
<p>This section describes related work on trajectory data generation based on Markov models and RNN models.</p>
<p>Mechanisms for generating trajectory data need to model the sequential data. Markov model is a standard sequence generation method. It models the past sequence behaviour of users to generate the next behaviour of users. Some studies use the Markov model to predict the next place for trajectory generation. Simmons et al. [<xref ref-type="bibr" rid="ref-10">10</xref>] trans the Hidden Markov Model (HMM) for each user to indicate which place the user will. Liao et al. [<xref ref-type="bibr" rid="ref-11">11</xref>] train a hierarchical Markov model on subscribers&#x2019; daily trajectories. The methods above train the model and predict the future location based on a single object&#x2019;s historical trajectory. In this case, when the user reaches an area that has not been reached, the model will not predict. Xue et al. [<xref ref-type="bibr" rid="ref-12">12</xref>] propose that the collective-pattern-based method divides group tracks into sub trajectories to overcome the shortcomings of modelling individuals. These sub trajectories are combined into growth tracks, and then the Markov model is used to predict the location. However, there are also drawbacks to the collective-pattern-based method. The collective-pattern-based approach is too coarse-grained in predicting the next area, such as using a first-order Markov model that predicts the same next place for all users in the same location. In some estimation generation work based on privacy protection, Xue et al. [<xref ref-type="bibr" rid="ref-13">13</xref>] conduct a prefix-tree using individual tracks and then construct a Markov model to predict the users&#x2019; next locations. Chen et al. [<xref ref-type="bibr" rid="ref-14">14</xref>] utilize a hybrid-granularity prefix tree structure, which is efficient data-dependent yet differentially private, to generate trajectories. Chen et al. [<xref ref-type="bibr" rid="ref-15">15</xref>] proposed a novel approach, which extracts the essential information of a sequential database in terms of a set of variable-length n-grams with large counts to generate and release trajectory data. He et al. [<xref ref-type="bibr" rid="ref-16">16</xref>] took the prefix tree structure and a hierarchical reference system to model users&#x2019; moving pattern at different velocities. Wang et al. [<xref ref-type="bibr" rid="ref-17">17</xref>] proposed a private trajectories calibration and publication system (PTCP), which generates synthetic trajectories by the build noise-enhanced prefix tree. Xu et al. [<xref ref-type="bibr" rid="ref-18">18</xref>] proposed DP-LTOP, which divides the original trajectory sequence into different sub-segments, and then selects the appropriate locations and segments to form a synthetic trajectory. Wang et al. [<xref ref-type="bibr" rid="ref-19">19</xref>] proposed DP-PSP, which addresses the heterogeneity of trajectories by the anchor point clustering and road segment mechanism to better the synthetic trajectories&#x2019; usability. Gursoy et al. [<xref ref-type="bibr" rid="ref-20">20</xref>] proposed Adatrace, which consist of four steps, features extraction, feature extraction, synopsis learning, privacy and utility preserving noise injection, and generation of differentially private synthetic location traces and generates synthetic trajectory by the Markov model with noise. This paper points out three threats, Bayesian inference threat, partial sniffing threat, and outlier leakage threat which are important but many studies have overlooked. Gursoy et al. [<xref ref-type="bibr" rid="ref-21">21</xref>] proposed DP-star, similar work to Adatrace. DP-star takes the Minimum Description Length metric to normalize the trajectory data to benefit the model training process. Ou et al. [<xref ref-type="bibr" rid="ref-22">22</xref>] adopted to extract reliable segments from sub-trajectories, build an exploration tree, and generate synthetic trajectories. Ghane et al. [<xref ref-type="bibr" rid="ref-23">23</xref>] proposed TGM, which could generate trajectories with arbitrary length, could encode trajectory data into the graphical generative model effectively.</p>
<p>Recurrent Neural Networks (RNNs) has been not only successfully applied in Nature Language Processing but also achieve outstanding performance for sequential data, such as text generation, image captioning [<xref ref-type="bibr" rid="ref-24">24</xref>], and location prediction [<xref ref-type="bibr" rid="ref-25">25</xref>]. RNN iteratively reads the sequence, iterating through each element of the sequence and updating its representation based on the input and the previous state. The connection between the hidden units and their respective projections is preserved. Gating units are often used in RNN models to transform the information flow in a more structured manner. A critical factor in determining the applicability of RNN is the size of the data set because RNN has poor generalization performance over small data volumes [<xref ref-type="bibr" rid="ref-26">26</xref>]. Based on the application of the current research results of RNN and the characteristics of the sequence generation problem, we adopt the RNN model to achieve trajectory generation. RNNs network is practised to process sequence data with long-term space-time dependencies, and it is ideal for achieving the inherent attributes of the continuous position.</p>
</sec>
<sec id="s3">
<label>3</label>
<title>TGMRNN Overview and Core Components</title>
<p>This section presents the architecture of TGMRNN.</p>

<sec id="s3_1">
<label>3.1</label>
<title>Preliminaries</title>
<p>RNN is suitable for processing time-sequential data, and it can transfer the output and state of the current moment to the next moment as input. Therefore, this kind of string structure can maintain the relationship between moments. And there are problems of gradient disappearance and gradient explosion. Hence it is difficult for RNN to maintain long-term dependence. Researchers further create many excellent evolution models based on RNN, such as Long Short-Term Memory (LSTM) and GRU, to solve these problems. These models solve long-term dependence by adding memory units and avoiding gradient explosions by gating units. Compared with LSTM, GRU has fewer parameters and faster training.</p>
<p>The proliferation of digital mobile data, such as GPS tracking, wireless communication records, and social media location records, coupled with the superior predictive power of artificial intelligence, has sparked rapid development in the field of human mobility research. Trajectory generation technology is an essential part of human mobility research. We want to take feature extraction from a data set containing mobile users&#x2019; actual location trajectories to build a generation model and achieve the purpose of maintaining statistical utility. The research on human mobility mainly includes three aspects: the next location prediction, the people flow forecast, and trajectory generation. Mobile communication networks, GPS, and social networks are the primary source of location data. For example, the mobile communication network has developed into the fifth generation of mobile communication technology. The mobile communication network can almost cover all the range of human activities by various heterogeneous networks (such as satellite networks). The mobile communication devices will interact with the base stations when we use mobile phones and other mobile communication devices to send and receive information. In this interaction process, the users&#x2019; geographical location will be acquired by the mobile communication network.</p>
<p>Mobility data describes the movement of a group of individuals during the observation period, usually collected by sensors, and is stored as a spatiotemporal trajectory or flow of movement. An individual trajectory is a group of records, typically each containing the identification, the geographic location, and the timestamp of the location.</p>
<p>The format of a trajectory is defined as follows:</p>
<p><bold>Definition 1 (Trajectory).</bold> Let <inline-formula id="ieqn-1">
<mml:math id="mml-ieqn-1"><mml:mi>i</mml:mi></mml:math>
</inline-formula> denote a unit and a trajectory of <inline-formula id="ieqn-2">
<mml:math id="mml-ieqn-2"><mml:mi>i</mml:mi></mml:math>
</inline-formula> is <inline-formula id="ieqn-3">
<mml:math id="mml-ieqn-3"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mi>S</mml:mi><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mi>S</mml:mi><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mi>S</mml:mi><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>, which is composed of <inline-formula id="ieqn-4">
<mml:math id="mml-ieqn-4"><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> sequential locations visited by the unit <inline-formula id="ieqn-5">
<mml:math id="mml-ieqn-5"><mml:mi>i</mml:mi></mml:math>
</inline-formula>. A spatiotemporal location, denoted by <inline-formula id="ieqn-6">
<mml:math id="mml-ieqn-6"><mml:mi>S</mml:mi><mml:mi>T</mml:mi></mml:math>
</inline-formula>, contains the geographic location and the timestamp, and <inline-formula id="ieqn-7">
<mml:math id="mml-ieqn-7"><mml:mi>S</mml:mi><mml:mi>T</mml:mi><mml:mo>=</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mo>,</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:math>
</inline-formula>. Let <inline-formula id="ieqn-8">
<mml:math id="mml-ieqn-8"><mml:mi>l</mml:mi><mml:mo>=</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>i</mml:mi><mml:mi>t</mml:mi><mml:mi>u</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mo>,</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>g</mml:mi><mml:mi>i</mml:mi><mml:mi>t</mml:mi><mml:mi>u</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:math>
</inline-formula> because the geographic location of a unit is usually represented by latitude and longitude. And <inline-formula id="ieqn-9">
<mml:math id="mml-ieqn-9"><mml:mi>t</mml:mi></mml:math>
</inline-formula> represents the timestamp of the corresponding location.</p>
<p>We need to discrete the continuous two-dimensional(latitude and longitude) spatial data by embedding technology in trajectory generation. The embedding technology is to divide the two-dimensional space into a limited number of independent regions and then map the location points to labels. The grid method is defined as follows:</p>
<p><bold>Definition 2 (Grid method).</bold> For the geographic region <inline-formula id="ieqn-10">
<mml:math id="mml-ieqn-10"><mml:mi>A</mml:mi></mml:math>
</inline-formula>, a grid method <inline-formula id="ieqn-11">
<mml:math id="mml-ieqn-11"><mml:mi>G</mml:mi></mml:math>
</inline-formula> contains <inline-formula id="ieqn-12">
<mml:math id="mml-ieqn-12"><mml:mi>n</mml:mi></mml:math>
</inline-formula> independent regions, represented as <inline-formula id="ieqn-13">
<mml:math id="mml-ieqn-13"><mml:mi>G</mml:mi><mml:mo>=</mml:mo><mml:munderover><mml:mo movablelimits="false">&#x2211;</mml:mo><mml:mrow><mml:mi>I</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>n</mml:mi></mml:munderover><mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi mathvariant="normal">g</mml:mi></mml:mrow><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mrow></mml:math>
</inline-formula>, and it satisfies <inline-formula id="ieqn-14">
<mml:math id="mml-ieqn-14"><mml:mi mathvariant="normal">&#x2200;</mml:mi><mml:mi>i</mml:mi><mml:mo>&#x2208;</mml:mo><mml:mi>n</mml:mi><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>g</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x2260;</mml:mo><mml:mi mathvariant="normal">&#x2205;</mml:mi></mml:math>
</inline-formula> and <inline-formula id="ieqn-15">
<mml:math id="mml-ieqn-15"><mml:mi mathvariant="normal">&#x2200;</mml:mi><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi><mml:mo>&#x2208;</mml:mo><mml:mi>n</mml:mi><mml:mo>,</mml:mo><mml:mi>i</mml:mi><mml:mo>&#x2260;</mml:mo><mml:mi>j</mml:mi></mml:math>
</inline-formula> <inline-formula id="ieqn-16">
<mml:math id="mml-ieqn-16"><mml:mrow><mml:msub><mml:mi>g</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x2229;</mml:mo><mml:mrow><mml:msub><mml:mi>g</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow><mml:mo>=</mml:mo><mml:mi mathvariant="normal">&#x2205;</mml:mi></mml:math>
</inline-formula>. And each region <inline-formula id="ieqn-17">
<mml:math id="mml-ieqn-17"><mml:mi>G</mml:mi></mml:math>
</inline-formula> has its own label. Therefore <inline-formula id="ieqn-18">
<mml:math id="mml-ieqn-18"><mml:mi>G</mml:mi></mml:math>
</inline-formula> could map geographic locations to labels.</p>
<p>Our goal is to extract and learn the mobility patterns of users to generate synthetic data.</p>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>TGMRNN Model</title>
<p>Consider a dataset of accurate location trajectories, denoted by <inline-formula id="ieqn-19">
<mml:math id="mml-ieqn-19"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mrow><mml:mi mathvariant="normal">a</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula>. We want to design a generative model for learning and modelling the original location data and generating synthetic trajectory data set, denoted by <inline-formula id="ieqn-20">
<mml:math id="mml-ieqn-20"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>s</mml:mi><mml:mi>y</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula>. The synthetic trajectories <inline-formula id="ieqn-21">
<mml:math id="mml-ieqn-21"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>s</mml:mi><mml:mi>y</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> need a high degree of similarity to <inline-formula id="ieqn-22">
<mml:math id="mml-ieqn-22"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mrow><mml:mi mathvariant="normal">a</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula>.</p>
<p>We design and develop TGMRNN, a trajectory dataset generator. <xref ref-type="fig" rid="fig-1">Fig. 1</xref> illustrates the system architecture of TGMRNN.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>TGMRNN system architecture</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_20032-fig-1.png"/>
</fig>
<sec id="s3_2_1">
<label>3.2.1</label>
<title>Data Pre-Processing</title>
<p>TGMRNN preprocesses the input data set <inline-formula id="ieqn-23">
<mml:math id="mml-ieqn-23"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula>. Our expected trajectory data should be collected by the same sampling rate and vehicle from users, and track locations are distributed within the specified geographic area. Although there is no unified format for the publicly available trajectory data sets, they generally contain user identification, time, and geographical location, which is sufficient for training. Therefore, before we start training the model, we need to preprocess the input data set <inline-formula id="ieqn-24">
<mml:math id="mml-ieqn-24"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> and transform it into our expected format, the standard format. In this process, we take to use the Grid method to transform the data, and its details are processing process is as follows:</p>
<p>A. Transform the input data into a location sequence. Let <inline-formula id="ieqn-25">
<mml:math id="mml-ieqn-25"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> denote a piece of data in the data set <inline-formula id="ieqn-26">
<mml:math id="mml-ieqn-26"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> and <inline-formula id="ieqn-27">
<mml:math id="mml-ieqn-27"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi mathvariant="normal">P</mml:mi></mml:mrow><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi mathvariant="normal">P</mml:mi></mml:mrow><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>n</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:math>
</inline-formula>. The information <inline-formula id="ieqn-28">
<mml:math id="mml-ieqn-28"><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> contains geographic location and time, and some also include user identification, transportation, and other information. We require that the user&#x2019;s location should be collected at the same sampling rate. Therefore the time interval between adjacent locations in a trajectory is the same. We divide the data set D into multiple trajectories. If the time interval between two adjacent locations is greater than the time interval corresponding to the sampling rate, the front and back should be divided into two trajectories from the current location. <inline-formula id="ieqn-29">
<mml:math id="mml-ieqn-29"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> is transformed to <inline-formula id="ieqn-30">
<mml:math id="mml-ieqn-30"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup></mml:math>
</inline-formula> by this way. There are trajectories in <inline-formula id="ieqn-31">
<mml:math id="mml-ieqn-31"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup></mml:math>
</inline-formula>. Let <inline-formula id="ieqn-32">
<mml:math id="mml-ieqn-32"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> denote a trajectory, and <inline-formula id="ieqn-33">
<mml:math id="mml-ieqn-33"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> consists of the sequence of locations, which could be represented as <inline-formula id="ieqn-34">
<mml:math id="mml-ieqn-34"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo><mml:mo>,</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mrow><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>, where <inline-formula id="ieqn-35">
<mml:math id="mml-ieqn-35"><mml:mo stretchy="false">(</mml:mo><mml:mi>l</mml:mi><mml:mi>a</mml:mi><mml:mrow><mml:msub><mml:mi>t</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:math>
</inline-formula> is the latitude and longitude of the location. <inline-formula id="ieqn-36">
<mml:math id="mml-ieqn-36"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup></mml:math>
</inline-formula> consists of multiple trajectories, which is represented as <inline-formula id="ieqn-37">
<mml:math id="mml-ieqn-37"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mi>n</mml:mi></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>. <inline-formula id="ieqn-38">
<mml:math id="mml-ieqn-38"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup></mml:math>
</inline-formula> preserves the locations of the moving object and the context between the locations in this way.</p>
<p>B. Map the geographic location to the grid cell identity. TGMRNN discretizes continuous two-dimensional (latitude and longitude) spatial data by the grid method. As shown in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>, for example, there is a trajectory <inline-formula id="ieqn-39">
<mml:math id="mml-ieqn-39"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow></mml:math>
</inline-formula> in the area and <inline-formula id="ieqn-40">
<mml:math id="mml-ieqn-40"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>3</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>5</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mn>6</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula> where <inline-formula id="ieqn-41">
<mml:math id="mml-ieqn-41"><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> stands for a location of the trajectory. Then, TGMRNN maps the locations of the trajectories to the cell identifications of the grid to which the locations belong to. TGMRNN transforms the location sequence <inline-formula id="ieqn-42">
<mml:math id="mml-ieqn-42"><mml:mrow><mml:msub><mml:mi>T</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow></mml:math>
</inline-formula> to cell identification sequence <inline-formula id="ieqn-43">
<mml:math id="mml-ieqn-43"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow></mml:math>
</inline-formula>, where <inline-formula id="ieqn-44">
<mml:math id="mml-ieqn-44"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>=</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>5</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>8</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>8</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>9</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>. <inline-formula id="ieqn-45">
<mml:math id="mml-ieqn-45"><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> is short for <inline-formula id="ieqn-46">
<mml:math id="mml-ieqn-46"><mml:mi>C</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mi>l</mml:mi><mml:mi mathvariant="normal">&#x005F;</mml:mi><mml:mn>4</mml:mn></mml:math>
</inline-formula> in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>. TGMRNN transform <inline-formula id="ieqn-47">
<mml:math id="mml-ieqn-47"><mml:msubsup><mml:mi>D</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>w</mml:mi></mml:mrow><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msubsup></mml:math>
</inline-formula> to <inline-formula id="ieqn-48">
<mml:math id="mml-ieqn-48"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>j</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mi>r</mml:mi><mml:mi>y</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> by mapping <inline-formula id="ieqn-49">
<mml:math id="mml-ieqn-49"><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> to <inline-formula id="ieqn-50">
<mml:math id="mml-ieqn-50"><mml:mrow><mml:msub><mml:mi>S</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:math>
</inline-formula> one by one, and we denote the format <inline-formula id="ieqn-51">
<mml:math id="mml-ieqn-51"><mml:mrow><mml:msub><mml:mi>D</mml:mi><mml:mrow><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>j</mml:mi><mml:mi>e</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi><mml:mi>o</mml:mi><mml:mi>r</mml:mi><mml:mi>y</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
</inline-formula> as the standard format.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Grid construction. Divide maps into different levels of granularity on the different scales of the grid method</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_20032-fig-2.png"/>
</fig>
</sec>
<sec id="s3_2_2">
<label>3.2.2</label>
<title>Training</title>

<p>As shown in <xref ref-type="fig" rid="fig-3">Fig. 3</xref>, we used the RNN model to train and learn the Sequence data. The system&#x2019;s core function is the RNN model. The TGMRNN configures the network model to learn the convolution sequence and extend it to the space-time domain for trajectory generation. This paper adopts the GRU model. Compared with LSTM, which can remember the information in the past and selectively forget some unimportant tips to model the long-term context and other relations, GRU reduces the gradient disappearance problem while retaining the long-term sequence information. TGMRNN divides the trajectory sequence dataset into sample sequences. For each input sequence, the corresponding output contains the same length of the sequence but moves one character to the right. For example, if the sample sequence is <inline-formula id="ieqn-52">
<mml:math id="mml-ieqn-52"><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>3</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>5</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>, the input is <inline-formula id="ieqn-53">
<mml:math id="mml-ieqn-53"><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>3</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>, and the output is <inline-formula id="ieqn-54">
<mml:math id="mml-ieqn-54"><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>3</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow><mml:mo>,</mml:mo><mml:mrow><mml:msub><mml:mi>C</mml:mi><mml:mn>5</mml:mn></mml:msub></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:math>
</inline-formula>. The problem can then be thought of as a standard classification problem: given the previous RNN state and the input of this timestamp, predict the class of the following sequence unit. We then added the optimizer and loss function, defined the appropriate epoch, and trained.</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>The process of model training</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_20032-fig-3.png"/>
</fig>
</sec>
<sec id="s3_2_3">
<label>3.2.3</label>
<title>Generation Algorithm</title>
<p><xref ref-type="fig" rid="fig-4">Fig. 4</xref> illustrates the basic cyclic network generation architecture. The process of sequence generation using the trained RNN model is as follows:</p>
<p>(1) Set the start position, initialize the RNN state, and set the trajectory&#x2019;s length to be generated. The starting position and RNN state are used to obtain the predicted distribution of the next position.</p>
<p>(2) The classification distribution is used to calculate the predicted location index, which is then used as the next input to the model.</p>
<p>(3) The returned state is fed back to the model. The model has more context to learn from than just one location. After the next location is predicted, the changed RNN state is fed back into the model. The model learns by always getting more context from the previously indicated place and continuously generates new sequences until the termination item is triggered.</p>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>The process of trajectory sequence generation</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_20032-fig-4.png"/>
</fig>
</sec>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Experiment</title>
<p>In this section, we report experiments conducted to assess the effectiveness of our generation algorithm.</p>
<sec id="s4_1">
<label>4.1</label>
<title>Settings and Data Set</title>
<p>We implemented the experiment on Colaboratory of Google, which provided GPU for TGMRNN. We built TGMRNN by PyTorch 1.8.1 and Python 3.7.10. We compare it with the Markov model. We choose the T-Drive trajectory data sample [<xref ref-type="bibr" rid="ref-27">27</xref>], coming from Microsoft T-Drive project. It contains trajectory data of more than 10,000 taxis in Beijing for a week in 2008, which contains 15 million coordinate points and tracks over a total distance of more than 9 million kilometres. There is enough data to allow us to set up experiments of different sizes for a comprehensive comparison. We preprocessed the data set to form tracks. After eliminating the abnormal data such as vacant position, single-point position, and error information, more than 1.8 million tracks with a length greater than one were generated in total. We compared the 1-order Markov model with the TGMRNN model during the deployment experiment because research shows that low-order Markov and high-order Markov performance is similar. To keep the training time reasonable, we set 100 epochs to train the model.</p>
<p>We set different numbers of mobile users, other grid partition measurements as a control experiment. We used different numbers of moving objects, 10, 100, and 1000 respectively, for training, and the corresponding number of tracks are 3215, 20306, and 172419, respectively.</p>
<p>We set the Markov model and TGMRNN model to learn data, respectively, generate trajectory data as the same amount of the training data, and then carry out statistical analysis.</p>
</sec>
<sec id="s4_2">
<label>4.2</label>
<title>Evaluation Metrics</title>
<p>We adopt Kullback&#x2013;Leibler divergence (KL-Divergence) to evaluate the generated trajectory dataset&#x2019;s effectiveness and the real user movement trajectory dataset. KL-divergence is equivalent to the difference of Shannon entropy between two probability distributions. KL-divergence could evaluate the degree of similarity among different data sets.</p>
<p>Supposed <inline-formula id="ieqn-55">
<mml:math id="mml-ieqn-55"><mml:mi>P</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:math>
</inline-formula> and <inline-formula id="ieqn-56">
<mml:math id="mml-ieqn-56"><mml:mi>Q</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:math>
</inline-formula> are two probability distributions on the random variable <inline-formula id="ieqn-57">
<mml:math id="mml-ieqn-57"><mml:mi>x</mml:mi></mml:math>
</inline-formula>. In the case of discrete random variables, the definition of relative entropy is :</p>
<p><disp-formula id="eqn-1"><label>(1)</label>

<mml:math id="mml-eqn-1" display="block"><mml:mi>K</mml:mi><mml:mi>L</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>P</mml:mi><mml:mo>&#x2225;</mml:mo><mml:mi>Q</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:mo>&#x2211;</mml:mo><mml:mi>P</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mi>log</mml:mi><mml:mstyle displaystyle="true" scriptlevel="0"><mml:mrow><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mrow><mml:mi>Q</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle></mml:math>
</disp-formula></p>
<p>The KL-divergence is nonnegative: <inline-formula id="ieqn-58">
<mml:math id="mml-ieqn-58"><mml:mi>K</mml:mi><mml:mi>L</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>P</mml:mi><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy="false">|</mml:mo></mml:mrow><mml:mi>Q</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>&#x2265;</mml:mo><mml:mn>0</mml:mn></mml:math>
</inline-formula>, equal to <inline-formula id="ieqn-59">

</inline-formula> when <inline-formula id="ieqn-60">
<mml:math id="mml-ieqn-60"><mml:mi>P</mml:mi><mml:mo>=</mml:mo><mml:mi>Q</mml:mi></mml:math>
</inline-formula>. The closer a KL-Divergence value is to zero, the closer the generated trajectory data set&#x2019;s location distribution is to the entire data set.</p>
</sec>
<sec id="s4_3">
<label>4.3</label>
<title>Evaluation of TGMRNN</title>
<p>We evaluate the performance of the TGMRNN and Markov models.</p>

<p>We evaluate the different influence on TGMRNN and Markov model by adopting the grid measurement 25 &#x00D7; 25, 50 &#x00D7; 50, 75 &#x00D7; 75, 100 &#x00D7; 100. The experimental results are shown in <xref ref-type="table" rid="table-1">Tabs. 1</xref> and <xref ref-type="table" rid="table-2">2</xref>. With the increase of the value of the grid, from 25 &#x00D7; 25 to 100 &#x00D7; 100, the map is more finely divided, and the effectiveness of synthetic trajectory data, generated by TGMRNN, gradually strengthens, where KL-divergence declines from 0.223 to 0.204 in 100 objects and from 0.173 to 0.133 in 100 objects. It is clear that the result for 1000 objects is better than the result for 100 objects under the same gird size, and the best value occurs in 1000 objects and 75 &#x00D7; 75. However, changes in the grid size have little impact on the results of Markov, where the values of KL-divergence stabilize around 0.233 in 100 objects and 0.324 in 1000 objects. And as shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>, the changes in the grid size do not make a significant difference in the availability of the generated data of the Markov model. We also evaluate the effect of a change in the number of objects on the trajectory generation, with 10, 100, and 1000. In the experiment of 10 objects, the experimental results are not representative because the number of objects is too small. With the increase of the number of objects, the availability of data generated by Markov declines significantly from 100 objects to 1000 objects, whose average of KL-divergence is 0.235 and 0.324, respectively, indicating that with the rise in the number of data, the Markov model is challenging to model the trajectory data. On the contrary, the availability of TGMRNN&#x2019;s generated data significantly rise from 100 objects to 1000 objects. TGMRNN has better data availability than the Markov model of the same grid size and the same number of users, 100 objects and 1000 objects.</p>
<table-wrap id="table-1"><label>Table 1</label>
<caption>
<title>The experiment result of the Markov model</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th></th>
<th>Grid 25 &#x00D7; 25</th>
<th>Grid 50 &#x00D7; 50</th>
<th>Grid 75 &#x00D7; 75</th>
<th>Grid 100 &#x00D7; 100</th>
</tr>
</thead>
<tbody>
<tr>
<td>10 Objects</td>
<td>0.153</td>
<td>0.180</td>
<td>0.106</td>
<td>0.145</td>
</tr>
<tr>
<td>100 Objects</td>
<td>0.230</td>
<td>0.230</td>
<td>0.231</td>
<td>0.230</td>
</tr>
<tr>
<td>1000 Objects</td>
<td>0.325</td>
<td>0.319</td>
<td>0.329</td>
<td>0.324</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap id="table-2"><label>Table 2</label>
<caption>
<title>The experiment result of the TGMRNN</title></caption>
<table><colgroup>
<col/>
<col/>
<col/>
<col/>
<col/>
</colgroup>
<thead>
<tr>
<th></th>
<th>Grid 25 &#x00D7; 25</th>
<th>Grid 50 &#x00D7; 50</th>
<th>Grid 75 &#x00D7; 75</th>
<th>Grid 100 &#x00D7; 100</th>
</tr>
</thead>
<tbody>
<tr>
<td>10 Objects</td>
<td>0.21</td>
<td>0.203</td>
<td>0.177</td>
<td>0.170</td>
</tr>
<tr>
<td>100 Objects</td>
<td>0.223</td>
<td>0.214</td>
<td>0.205</td>
<td>0.204</td>
</tr>
<tr>
<td>1000 Objects</td>
<td>0.173</td>
<td>0.163</td>
<td>0.128</td>
<td>0.133</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>KL-divergence with different grid-scale and a different number of moving objects</title></caption>
<graphic mimetype="image" mime-subtype="png" xlink:href="IASC_20032-fig-5.png"/>
</fig>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Conclusion</title>
<p>We propose TGMRNN to generate synthetic trajectory data. It could produce trajectory data for mobile service providers, who need a good deal of user data to improve the quality of their products and prevent the disclosure of users&#x2019; privacy. TGMRNN could transform continuous location data to discrete sequence data and preserve the relationship between the location of high-dimensional sequences. TGMRNN adopts the many-to-one pattern to train the model, which has fewer parameters and faster training and take the trained model to generate synthetic trajectory by predicting the next location. And experiments show that TGMRNN is effective.</p>
<p>In the future, trajectory generation is a hot research field. There are still many problems to be solved, such as (1) we should solve the geographic data sparsity to improve the availability of the trained model; (2) we need to improve the efficiency and effectiveness of high-dimensional sequence execution; (3) we need more measures for the availability of the trajectory data; (4) we should introduce stricter privacy mechanism in trajectory generation, such differential privacy; (5) we should eliminate some privacy threats, such as Bayesian inference threat, partial sniffing threat, and outlier leakage threat.</p>
</sec>
</body>
<back>
<ack>
<p>My deepest gratitude goes first and foremost to Dr Cui, my supervisor, for his constant encouragement and guidance.</p>
</ack><fn-group>
<fn fn-type="other">
<p><bold>Funding Statement:</bold> The work was supported by National Natural Science Foundation of China (61941114), National Natural Science Foundation of China (Grant No. 61802025), National Natural Science Foundation of China (No. 62001055), Beijing Natural Science Foundation (4204107), Funds of &#x201C;YinLing&#x201D; (No. A02B01C03-201902D0).</p>
</fn>
<fn fn-type="conflict">
<p><bold>Conflicts of Interest:</bold> The authors declare that they have no conflicts of interest to report regarding the present study.</p>
</fn>
</fn-group>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Moreira-Matias</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Gama</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Ferreira</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Mendes-Moreira</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Damas</surname></string-name></person-group>, &#x201C;<article-title>Predicting taxi-passenger demand using streaming data</article-title>,&#x201D; <source>IEEE Transactions on Intelligent Transportation Systems</source>, vol. <volume>14</volume>, no. <issue>3</issue>, pp. <fpage>1393</fpage>&#x2013;<lpage>1402</lpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>F.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Tu</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Fu</surname></string-name> <etal>et al.</etal></person-group><italic>,</italic> &#x201C;<article-title>Trajectory recovery from ash: User privacy is not preserved in aggregated mobility data</article-title>,&#x201D; in <conf-name>Proc. 26th Int. Conf. World Wide Web</conf-name>, <conf-loc>Perth, Australia</conf-loc>, pp. <fpage>1241</fpage>&#x2013;<lpage>1250</lpage>, <year>2017</year>. </mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>Y. A.</given-names> <surname>Montjoye</surname></string-name>, <string-name><given-names>C. A.</given-names> <surname>Hidalgo</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Verleysen</surname></string-name> and <string-name><given-names>V. D.</given-names> <surname>Blondel</surname></string-name></person-group>, &#x201C;<article-title>Unique in the crowd: The privacy bounds of human mobility</article-title>,&#x201D; <source>Scientific Report</source>, vol. <volume>3</volume>, no. <issue>1</issue>, pp. <fpage>193</fpage>, <year>2013</year>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>B.</given-names> <surname>Hu</surname></string-name> and <string-name><given-names>J.</given-names> <surname>Wang</surname></string-name></person-group>, &#x201C;<article-title>Deep learning for distinguishing computer generated images and natural images: A survey</article-title>,&#x201D; <source>Journal of Information Hiding and Privacy Protection</source>, vol. <volume>2</volume>, no. <issue>2</issue>, pp. <fpage>37</fpage>&#x2013;<lpage>47</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Cui</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Ma</surname></string-name> and <string-name><given-names>C.</given-names> <surname>Wu</surname></string-name></person-group>, &#x201C;<article-title>Frequent itemset mining of user&#x2019;s multi-attribute under local differential privacy</article-title>,&#x201D; <source>Computers, Materials &#x0026; Continua</source>, vol. <volume>65</volume>, no. <issue>1</issue>, pp. <fpage>369</fpage>&#x2013;<lpage>385</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Sun</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Ge</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Huang</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Wu</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Gao</surname></string-name></person-group>, &#x201C;<article-title>Differentially private real-time streaming data publication based on sliding window under exponential decay</article-title>,&#x201D; <source>Computers, Materials &#x0026; Continua</source>, vol. <volume>58</volume>, no. <issue>1</issue>, pp. <fpage>61</fpage>&#x2013;<lpage>78</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Yuan</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zheng</surname></string-name>, <string-name><given-names>C. Y.</given-names> <surname>Zhang</surname></string-name>, <string-name><given-names>W. L.</given-names> <surname>Xie</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Xie</surname></string-name> <etal>et al</etal>.</person-group>, &#x201C;<article-title>T-drive: Driving directions based on taxi trajectories</article-title>,&#x201D; <comment>in Proc. of the 18th SIGSPATIAL Int. Conf. on Advances in Geographic Information Systems</comment>, <publisher-loc>New York, NY, USA</publisher-loc>, pp. <fpage>99</fpage>&#x2013;<lpage>108</lpage>, <year>2010</year>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><given-names>V.</given-names> <surname>Kulkarni</surname></string-name>, <string-name><given-names>N.</given-names> <surname>Tagasovska</surname></string-name>, <string-name><given-names>T.</given-names> <surname>Vatter</surname></string-name> and <string-name><given-names>B.</given-names> <surname>Garbinato</surname></string-name></person-group>, &#x201C;<article-title>Generating Synthetic Mobility Traffic using Recurrent Neural Networks</article-title>,&#x201D; <comment><italic>ACM SIGSPATIAL Workshop on Artificial Intelligence and Deep Learning for Geographic Knowledge Discovery</italic></comment>, <publisher-loc>Redondo Beach, California, USA</publisher-loc>, pp. <fpage>1</fpage>&#x2013;<lpage>4</lpage>. <year>2017</year>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Song</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Kotz</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Jain</surname></string-name> and <string-name><given-names>X.</given-names> <surname>He</surname></string-name></person-group>, &#x201C;<article-title>Evaluating next-cell predictors with extensive Wi-Fi mobility data</article-title>,&#x201D; <source>IEEE Transactions on Mobile Computing</source>, vol. <volume>5</volume>, no. <issue>12</issue>, pp. <fpage>1633</fpage>&#x2013;<lpage>1649</lpage>, <year>2006</year>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Simmons</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Browning</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zhang</surname></string-name> and <string-name><given-names>V.</given-names> <surname>Sadekar</surname></string-name></person-group>, &#x201C;<article-title>Learning to predict driver route and destination intent</article-title>,&#x201D; in <conf-name>IEEE Intelligent Transportation Systems Conf.</conf-name>, <publisher-loc>Toronto, ON, Canada</publisher-loc>, pp. <fpage>127</fpage>&#x2013;<lpage>132</lpage>, <year>2006</year>. </mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Liao</surname></string-name>, <string-name><given-names>D. J.</given-names> <surname>Patterson</surname></string-name>, <string-name><given-names>D.</given-names> <surname>Fox</surname></string-name> and <string-name><given-names>H.</given-names> <surname>Kautz</surname></string-name></person-group>, &#x201C;<article-title>Learning and inferring transportation routines</article-title>,&#x201D; in <source>Artificial Intelligence, <italic>Amsterdam: Elsevier</italic></source>, vol. <volume>171</volume>, no. <issue>5</issue>, pp. <fpage>311</fpage>&#x2013;<lpage>331</lpage>, <year>2007</year>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>A. Y.</given-names> <surname>Xue</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Rui</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Yu</surname></string-name>, <string-name><given-names>X.</given-names> <surname>Xing</surname></string-name> and <string-name><given-names>Z.</given-names> <surname>Xu</surname></string-name></person-group>, &#x201C;<article-title>Destination prediction by sub-trajectory synthesis and privacy protection against such prediction</article-title>,&#x201D; in <conf-name>IEEE Int. Conf. on Data Engineering</conf-name>, <conf-loc>Arlington, VA, USA</conf-loc>, pp. <fpage>254</fpage>&#x2013;<lpage>265</lpage>, <year>2012</year>. </mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>G.</given-names> <surname>Xue</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Li</surname></string-name>, <string-name><given-names>H.</given-names> <surname>Zhu</surname></string-name> and <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name></person-group>, &#x201C;<article-title>Traffic-known urban vehicular route prediction based on partial mobility patterns</article-title>,&#x201D; in <conf-name>Int. Conf. on Parallel &#x0026; Distributed Systems</conf-name>, <publisher-loc>Shenzhen, China</publisher-loc>, pp. <fpage>369</fpage>&#x2013;<lpage>375</lpage>, <year>2010</year>. </mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>B. C. M.</given-names> <surname>Fung</surname></string-name> and <string-name><given-names>B. C.</given-names> <surname>Desai</surname></string-name></person-group>, &#x201C;<article-title>Differentially private transit data publication: A case study on the montreal transportation system</article-title>,&#x201D; in <conf-name>Proc. of the 18th ACM SIGKDD Int. Conf. on Knowledge Discovery and Data Mining</conf-name>, <publisher-loc>New York, NY, USA</publisher-loc>, pp. <fpage>213</fpage>&#x2013;<lpage>221</lpage>, <year>2012</year>. </mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>R.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Acs</surname></string-name> and <string-name><given-names>C.</given-names> <surname>Castelluccia</surname></string-name></person-group>, &#x201C;<article-title>Differentially private sequential data publication via variable-length n-grams</article-title>,&#x201D; in <conf-name>Proc. of the 2012 ACM Conf. on Computer and Communications Security</conf-name>, <publisher-loc>New York, NY, USA</publisher-loc>, pp. <fpage>638</fpage>&#x2013;<lpage>649</lpage>, <year>2012</year>. </mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>X.</given-names> <surname>He</surname></string-name>, <string-name><given-names>G.</given-names> <surname>Cormode</surname></string-name> and <string-name><given-names>A.</given-names> <surname>Machanavajjhala</surname></string-name></person-group>, &#x201C;<article-title>DPT: Differentially private trajectory synthesis using hierarchical reference systems</article-title>,&#x201D; <source>Proceedings of the VLDB Endowment</source>, vol. <volume>8</volume>, no. <issue>11</issue>, pp. <fpage>1154</fpage>&#x2013;<lpage>1165</lpage>, <year>2015</year>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Wang</surname></string-name> and <string-name><given-names>R. O.</given-names> <surname>Sinnott</surname></string-name></person-group>, &#x201C;<article-title>Protecting personal trajectories of social media users through differential privacy</article-title>,&#x201D; <source>Computers &#x0026; Security</source>, vol. <volume>67</volume>, no. <issue>12</issue>, pp. <fpage>142</fpage>&#x2013;<lpage>163</lpage>, <year>2017</year>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>C.</given-names> <surname>Xu</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Zhu</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Guan</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Yu</surname></string-name></person-group>, &#x201C;<article-title>DP-LTOD: Differential privacy latent trajectory community discovering services over location-based social networks</article-title>,&#x201D; <source>IEEE Transactions on Services Computing</source>, vol. <volume>1</volume>, pp. <fpage>1</fpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Wang</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Sinnott</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Nepal</surname></string-name></person-group>, &#x201C;<article-title>Privacy-protected statistics publication over social media user trajectory streams</article-title>,&#x201D; <source>Future Generation Computer Systems</source>, vol. <volume>87</volume>, no. <issue>1</issue>, pp. <fpage>792</fpage>&#x2013;<lpage>802</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>M. E.</given-names> <surname>Gursoy</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Liu</surname></string-name> and <string-name><given-names>S.</given-names> <surname>Truex</surname></string-name></person-group>, &#x201C;<article-title>Utility-aware synthesis of differentially private and attack-resilient location traces</article-title>,&#x201D; in <conf-name>Proc. of the 2018 ACM SIGSAC Conf. on Computer and Communications Security</conf-name>, <publisher-loc>New York, NY, USA</publisher-loc>, pp. <fpage>196</fpage>&#x2013;<lpage>211</lpage>, <year>2018</year>. </mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>M. E.</given-names> <surname>Gursoy</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Truex</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Yu</surname></string-name></person-group>, &#x201C;<article-title>Differentially private and utility preserving publication of trajectory data</article-title>,&#x201D; <source>IEEE Transactions on Mobile Computing</source>, vol. <volume>18</volume>, no. <issue>10</issue>, pp. <fpage>2315</fpage>&#x2013;<lpage>2329</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>L.</given-names> <surname>Ou</surname></string-name>, <string-name><given-names>Z.</given-names> <surname>Qin</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Liao</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Hong</surname></string-name> and <string-name><given-names>X.</given-names> <surname>Jia</surname></string-name></person-group>, &#x201C;<article-title>Releasing correlated trajectories: Towards high utility and optimal differential privacy</article-title>,&#x201D; <source>IEEE Transactions on Dependable and Secure Computing</source>, vol. <volume>17</volume>, no. <issue>5</issue>, pp. <fpage>1109</fpage>&#x2013;<lpage>1123</lpage>, <year>2018</year>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>S.</given-names> <surname>Ghane</surname></string-name>, <string-name><given-names>L.</given-names> <surname>Kulik</surname></string-name> and <string-name><given-names>K.</given-names> <surname>Ramamohanarao</surname></string-name></person-group>, &#x201C;<article-title>TGM: A generative mechanism for publishing trajectories with differential privacy</article-title>,&#x201D; <source>IEEE Internet of Things Journal</source>, vol. <volume>7</volume>, no. <issue>4</issue>, pp. <fpage>2611</fpage>&#x2013;<lpage>2621</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>H.</given-names> <surname>Chen</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Li</surname></string-name> and <string-name><given-names>Z.</given-names> <surname>Zhang</surname></string-name></person-group>, &#x201C;<article-title>A differential privacy based (k-&#x03C8;)-anonymity method for trajectory data publishing</article-title>,&#x201D; <source>Computers, Materials &#x0026; Continua</source>, vol. <volume>65</volume>, no. <issue>3</issue>, pp. <fpage>2665</fpage>&#x2013;<lpage>2685</lpage>, <year>2020</year>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>Q.</given-names> <surname>Liu</surname></string-name>, <string-name><given-names>S.</given-names> <surname>Wu</surname></string-name> and <string-name><given-names>L.</given-names> <surname>Wang</surname></string-name></person-group>, &#x201C;<article-title>Predicting the next location: A recurrent model with spatial and temporal contexts</article-title>,&#x201D; in <conf-name>Thirtieth AAAI Conf. on Artificial Intelligence</conf-name>, <conf-loc>Phoenix, Arizona USA</conf-loc>, pp. <fpage>194</fpage>&#x2013;<lpage>200</lpage>, <year>2016</year>. </mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><given-names>I.</given-names> <surname>You</surname></string-name>, <string-name><given-names>C.</given-names> <surname>Choi</surname></string-name>, <string-name><given-names>V.</given-names> <surname>Sharma</surname></string-name>, <string-name><given-names>I.</given-names> <surname>Woungang</surname></string-name> and <string-name><given-names>B. K.</given-names> <surname>Bhargava</surname></string-name></person-group>, &#x201C;<article-title>Advances in security and privacy technologies for forthcoming smart systems, services, computing, and networks</article-title>,&#x201D; <source>Intelligent Automation &#x0026; Soft Computing</source>, vol. <volume>25</volume>, no. <issue>1</issue>, pp. <fpage>117</fpage>&#x2013;<lpage>119</lpage>, <year>2019</year>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><given-names>J.</given-names> <surname>Yuan</surname></string-name>, <string-name><given-names>Y.</given-names> <surname>Zheng</surname></string-name> and <string-name><given-names>X.</given-names> <surname>Xie</surname></string-name></person-group>, &#x201C;<article-title>Driving with knowledge from the physical world</article-title>,&#x201D; in <conf-name>Proc. of the 17th ACM SIGKDD Int. Conf. on Knowledge Discovery and Data Mining</conf-name>, <publisher-loc>New York, NY, USA</publisher-loc>, pp. <fpage>316</fpage>&#x2013;<lpage>324</lpage>, <year>2011</year>. </mixed-citation></ref>
</ref-list>
</back>
</article>