SearcharxivSearch

arXiv subjects

Aref Darzi

Publications and source records attributed to Aref Darzi.

10 recordsLinked to original sources

Using Geographically Weighted Models to Explore Temporal and Spatial Varying Impacts on Commute Trip Change Due to Covid-19

COVID-19 has deeply affected daily life and travel behaviors. Understanding these changes is crucial, prompting an investigation into socio-demographic and socio-economic factors. This study used large-scale mobile device location data in Washington, D.C., Maryland, and Virginia (DMV area) to unveil the impacts of these variables on commute trip changes. It reflected short and long-term impacts through linear regression and geographically weighted regression models. Findings indicated that counties with a higher percentage of people using walking and biking during the initial phase of COVID-19 experienced greater reductions in commute trips. For the long-term effect in November, the impact of active modes became insignificant, and individuals using public modes showed more significant trip reductions. Positive correlations were observed between median income levels and reduced commute trips. Sectors requiring ongoing outdoor operations during the pandemic showed substantial negative correlations. In the DMV area, counties with a higher proportion of Democratic voters experienced less trip reduction. Applying Geographically Weighted Regression models captured local spatial relationships, showing the emergence of local correlations as the pandemic evolved, suggesting a geographical impact pattern. Initially global, the pandemic's impact on commuting behaviors became more influenced by spatial factors over time, showing localized effects.

stat.AP

An elaborated pattern-based method of identifying data oscillations from mobile device location data

In recent years, passively collected GPS data have been popularly applied in various transportation studies, such as highway performance monitoring, travel behavior analysis, and travel demand estimation. Despite multiple advantages, one of the issues is data oscillations (aka outliers or data jumps), which are unneglectable since they may distort mobility patterns and lead to wrongly or biased conclusions. For transportation studies driven by GPS data, assuring the data quality by removing noises caused by data oscillations is undoubtedly important. Most GPS-based studies simply remove oscillations by checking the high speed. However, this method can mistakenly identify normal points as oscillations. Some other studies specifically discuss the removal of outliers in GPS data, but they all have limitations and do not fit passively collected GPS data. Many studies are well developed for addressing the ping-pong phenomenon in cellular data, or cellular tower data, but the oscillations in passively collected GPS data are very different for having much more various and complicated patterns and being more uncertain. Current methods are insufficient and inapplicable to passively collected GPS data. This paper aims to address the oscillated points in passively collected GPS data. A set of heuristics are proposed by identifying the abnormal movement patterns of oscillations. The proposed heuristics well fit the features of passively collected GPS data and are adaptable to studies of different scales, which are also computationally cost-effective in comparison to current methods.

cs.DB

A Big Data Driven Framework for Duplicate Device Detection from Multi-sourced Mobile Device Location Data

Mobile Device Location Data (MDLD) has been popularly utilized in various fields. Yet its large-scale applications are limited because of either biased or insufficient spatial coverage of the data from individual data vendors. One approach to improve the data coverage is to leverage the data from multiple data vendors and integrate them to build a more representative dataset. For data integration, further treatments on the multi-sourced dataset are required due to several reasons. First, the possibility of carrying more than one device could result in duplicated observations from the same data subject. Additionally, when utilizing multiple data sources, the same device might be captured by more than one data provider. Our paper proposes a data integration methodology for multi-sourced data to investigate the feasibility of integrating data from several sources without introducing additional biases to the data. By leveraging the uniqueness of travel pattern of each device, duplicate devices are identified. The proposed methodology is shown to be cost-effective while it achieves the desired accuracy level. Our findings suggest that devices sharing the same imputed home location and the top five most-visited locations during a month can represent the same user in the MDLD. It is shown that more than 99.6% of the sample devices having the aforementioned attribute in common are observed at the same location simultaneously. Finally, the proposed algorithm has been successfully applied to the national-level MDLD of 2020 to produce the national passenger origin-destination data for the NextGeneration National Household Travel Survey (NextGen NHTS) program.

cs.CY

Constructing Evacuation Evolution Patterns and Decisions Using Mobile Device Location Data: A Case Study of Hurricane Irma

Understanding individuals' behavior during hurricane evacuation is of paramount importance for local, state, and government agencies hoping to be prepared for natural disasters. Complexities involved with human decision-making procedures and lack of data for such disasters are the main reasons that make hurricane evacuation studies challenging. In this paper, we utilized a large mobile phone Location-Based Services (LBS) data to construct the evacuation pattern during the landfall of Hurricane Irma. By employing our proposed framework on more than 11 billion mobile phone location sightings, we were able to capture the evacuation decision of 807,623 smartphone users who were living within the state of Florida. We studied users' evacuation decisions, departure and reentry date distribution, and destination choice. In addition to these decisions, we empirically examined the influence of evacuation order and low-lying residential areas on individuals' evacuation decisions. Our analysis revealed that 57.92% of people living in mandatory evacuation zones evacuated their residences while this ratio was 32.98% and 33.68% for people living in areas with no evacuation order and voluntary evacuation order, respectively. Moreover, our analysis revealed the importance of the individuals' mobility behavior in modeling the evacuation decision choice. Historical mobility behavior information such as number of trips taken by each individual and the spatial area covered by individuals' location trajectory estimated significant in our choice model and improve the overall accuracy of the model significantly.

cs.LG

A Data-Driven Travel Mode Share Estimation Framework based on Mobile Device Location Data

Mobile device location data (MDLD) contains abundant travel behavior information to support travel demand analysis. Compared to traditional travel surveys, MDLD has larger spatiotemporal coverage of population and its mobility. However, ground truth information such as trip origins and destinations, travel modes, and trip purposes are not included by default. Such important attributes must be imputed to maximize the usefulness of the data. This paper tends to study the capability of MDLD on estimating travel mode share at aggregated levels. A data-driven framework is proposed to extract travel behavior information from the MDLD. The proposed framework first identifies trip ends with a modified Spatiotemporal Density-based Spatial Clustering of Applications with Noise (ST-DBSCAN) algorithm. Then three types of features are extracted for each trip to impute travel modes using machine learning models. A labeled MDLD dataset with ground truth information is used to train the proposed models, resulting in 95% accuracy in identifying trip ends and 93% accuracy in imputing five travel modes (drive, rail, bus, bike and walk) with a Random Forest (RF) classifier. The proposed framework is then applied to two large-scale MDLD datasets, covering the Baltimore-Washington metropolitan area and the United States, respectively. The estimated trip distance, trip time, trip rate distribution, and travel mode share are compared against travel surveys at different geographies. The results suggest that the proposed framework can be readily applied in different states and metropolitan regions with low cost in order to study multimodal travel demand, understand mobility trends, and support decision making.

cs.CY

A Data-Driven Analytical Framework of Estimating Multimodal Travel Demand Patterns using Mobile Device Location Data

While benefiting people's daily life in so many ways, smartphones and their location-based services are generating massive mobile device location data that has great potential to help us understand travel demand patterns and make transportation planning for the future. While recent studies have analyzed human travel behavior using such new data sources, limited research has been done to extract multimodal travel demand patterns out of them. This paper presents a data-driven analytical framework to bridge the gap. To be able to successfully detect travel modes using the passively collected location information, we conduct a smartphone-based GPS survey to collect ground truth observations. Then a jointly trained single-layer model and deep neural network for travel mode imputation is developed. Being "wide" and "deep" at the same time, this model combines the advantages of both types of models. The framework also incorporates the multimodal transportation network in order to evaluate the closeness of trip routes to the nearby rail, metro, highway and bus lines and therefore enhance the imputation accuracy. To showcase the applications of the introduced framework in answering real-world planning needs, a separate mobile device location data is processed through trip end identification and attribute generation, in a way that the travel mode imputation can be directly applied. The estimated multimodal travel demand patterns are then validated against typical household travel surveys in the same Washington D.C. and Baltimore Metropolitan Regions.

cs.LG

How different age groups responded to the COVID-19 pandemic in terms of mobility behaviors: a case study of the United States

The rapid spread of COVID-19 has affected thousands of people from different socio-demographic groups all over the country. A decisive step in preventing or slowing the outbreak is the use of mobility interventions, such as government stay-at-home orders. However, different socio-demographic groups might have different responses to these orders and regulations. In this paper, we attempt to fill the current gap in the literature by examining how different communities with different age groups performed social distancing by following orders such as the national emergency declaration on March 13, as well as how fast they started changing their behavior after the regulations were imposed. For this purpose, we calculated the behavior changes of people in different mobility metrics, such as percentage of people staying home during the study period (March, April, and May 2020), in different age groups in comparison to the days before the pandemic (January and February 2020), by utilizing anonymized and privacy-protected mobile device data. Our study indicates that senior communities outperformed younger communities in terms of their behavior change. Senior communities not only had a faster response to the outbreak in comparison to young communities, they also had better performance consistency during the pandemic.

cs.CY

COVID-19 and income profile: How communities in the United States responded to mobility restrictions in the pandemic's early stages

Mobility interventions in communities play a critical role in containing a pandemic at an early stage. The real-world practice of social distancing can enlighten policymakers and help them implement more efficient and effective control measures. A lack of such research using real-world observations initiates this article. We analyzed the social distancing performance of 66,149 census tracts from 3,142 counties in the United States with a specific focus on income profile. Six daily mobility metrics, including a social distancing index, stay-at-home percentage, miles traveled per person, trip rate, work trip rate, and non-work trip rate, were produced for each census tract using the location data from over 100 million anonymous devices on a monthly basis. Each mobility metric was further tabulated by three perspectives of social distancing performance: "best performance", "effort", and "consistency". We found that for all 18 indicators, high-income communities demonstrated better social distancing performance. Such disparities between communities of different income levels are presented in detail in this article. The comparisons across scenarios also raise other concerns for low-income communities, such as employment status, working conditions, and accessibility to basic needs. This article lays out a series of facts extracted from real-world data and offers compelling perspectives for future discussions.

physics.soc-ph

Quarantine Fatigue: first-ever decrease in social distancing measures after the COVID-19 outbreak before reopening United States

By the emergence of the novel coronavirus disease (COVID-19) in Wuhan, China, and its rapid outbreak worldwide, the infectious illness has changed our everyday travel patterns. In this research, our team investigated the changes in the daily mobility pattern of people during the pandemic by utilizing an integrated data panel. To incorporate various aspects of human mobility, the team focused on the Social Distancing Index (SDI) which was calculated based on five basic mobility measures. The SDI patterns showed a plateau stage in the beginning of April that lasted for about two weeks. This phenomenon then followed by a universal decline of SDI, increased number of trips and reduction in percentage of people staying at home. We called the observation Quarantine Fatigue. The Rate of Change (ROC) method was employed to trace back the start date of quarantine fatigue which was indicated to be April 15th. Our analysis showed that despite the existence of state-to-state variations, most states started experiencing a quarantine fatigue phenomenon during the same period. This observation became more important by knowing that none of the states had officially announced the reopening until late April showing that people decided to loosen up their social distancing practices before the official reopening announcement. Moreover, our analysis indicated that official reopening led to a rapid decline in SDI, raising the concern of a second wave of outbreak. The synchronized trend among states also emphasizes the importance of a more nationwide decision-making attitude for the future as the condition of each state depends on the nationwide behavior.

cs.CY

Quantifying human mobility behavior changes in response to non-pharmaceutical interventions during the COVID-19 outbreak in the United States

Ever since the first case of the novel coronavirus disease (COVID-19) was confirmed in Wuhan, China, social distancing has been promoted worldwide, including the United States. It is one of the major community mitigation strategies, also known as non-pharmaceutical interventions. However, our understanding is remaining limited in how people practice social distancing. In this study, we construct a Social Distancing Index (SDI) to evaluate people's mobility pattern changes along with the spread of COVID-19. We utilize an integrated dataset of mobile device location data for the contiguous United States plus Alaska and Hawaii over a 100-day period from January 1, 2020 to April 9, 2020. The major findings are: 1) the declaration of the national emergency concerning the COVID-19 outbreak greatly encouraged social distancing and the mandatory stay-at-home orders in most states further strengthened the practice; 2) the states with more confirmed cases have taken more active and timely responses in practicing social distancing; 3) people in the states with fewer confirmed cases did not pay much attention to maintaining social distancing and some states, e.g., Wyoming, North Dakota, and Montana, already began to practice less social distancing despite the high increasing speed of confirmed cases; 4) some counties with the highest infection rates are not performing much social distancing, e.g., Randolph County and Dougherty County in Georgia, and some counties began to practice less social distancing right after the increasing speed of confirmed cases went down, e.g., in Blaine County, Idaho, which may be dangerous as well.

cs.CY