SearcharxivSearch

arXiv subjects

Shengxuan Ding

Publications and source records attributed to Shengxuan Ding.

10 recordsLinked to original sources

Ozone: A Unified Platform for Transportation Research

Intelligent Transportation Systems increasingly depend on heterogeneous data from roadside cameras, UAV imagery, LiDAR, and in-vehicle sensors, yet the lack of unified data standards, model interfaces, and evaluation protocols across these sources hampers reproducibility, cross-dataset benchmarking, and cross-region transferability of research findings. Existing trajectory datasets follow incompatible conventions for coordinate systems, object representations, and metadata fields, forcing researchers to build custom preprocessing pipelines for each dataset and simulator combination. To address these challenges, we propose Ozone, a unified platform for transportation research organized around five interconnected layers -- Hardware, Data, Model, Evaluation, and Prototype -- each with standardized schemas, automated conversion pipelines, and interoperable interfaces. In the first release, the data schema unifies four trajectory datasets -- NGSIM, highD, CitySim, and UTE -- into a canonical format with oriented bounding boxes, kinematic variables, and pre-computed surrogate safety measures. Digital-twin maps in CARLA and calibrated traffic models provide integrated benchmarking environments. Case studies in human-factor research, traffic scene generation, and safety-critical modeling demonstrate that Ozone reduces experiment setup time by 85%, achieves 91% cross-city transfer efficiency for safety models, and improves cross-dataset reproducibility to within 3% variance. The source code and datasets are publicly available.

cs.DB

Exploratory analysis of injury severity under different levels of driving automation (SAE Level 2-5) using multi-source data

Vehicles equipped with automated driving capabilities have shown potential to improve safety and operations. Advanced driver assistance systems (ADAS) and automated driving systems (ADS) have been widely developed to support vehicular automation. Although the studies on the injury severity outcomes that involve automated vehicles are ongoing, there is limited research investigating the difference between injury severity outcomes for the ADAS and ADS equipped vehicles. To ensure a comprehensive analysis, a multi-source dataset that includes 1,001 ADAS crashes (SAE Level 2 vehicles) and 548 ADS crashes (SAE Level 4 vehicles) is used. Two random parameters multinomial logit models with heterogeneity in the means of random parameters are considered to gain a better understanding of the variables impacting the crash injury severity outcomes for the ADAS (SAE Level 2) and ADS (SAE Level 4) vehicles. It was found that while 67 percent of crashes involving the ADAS equipped vehicles in the dataset took place on a highway, 94 percent of crashes involving ADS took place in more urban settings. The model estimation results also reveal that the weather indicator, driver type indicator, differences in the system sophistication that are captured by both manufacture year and high/low mileage as well as rear and front contact indicators all play a role in the crash injury severity outcomes. The results offer an exploratory assessment of safety performance of the ADAS and ADS equipped vehicles using the real-world data and can be used by the manufacturers and other stakeholders to dictate the direction of their deployment and usage.

stat.AP

RF-Enhanced Road Infrastructure for Intelligent Transportation

The EPC GEN 2 communication protocol for Ultra-high frequency Radio Frequency Identification (RFID) has offered a promising avenue for advancing the intelligence of transportation infrastructure. With the capability of linking vehicles to RFID readers to crowdsource information from RFID tags on road infrastructures, the RF-enhanced road infrastructure (REI) can potentially transform data acquisition for urban transportation. Despite its potential, the broader adoption of RFID technologies in building intelligent roads has been limited by a deficiency in understanding how the GEN 2 protocol impacts system performance under different transportation settings. This paper fills this knowledge gap by presenting the system architecture and detailing the design challenges associated with REI. Comprehensive real-world experiments are conducted to assess REI's effectiveness across various urban contexts. The results yield crucial insights into the optimal design of on-vehicle RFID readers and on-road RFID tags, considering the constraints imposed by vehicle dynamics, road geometries, and tag placements. With the optimized designs of encoding schemes for reader-tag communication and on-vehicle antennas, REI is able to fulfill the requirements of traffic sign inventory management and environmental monitoring while falling short of catering to the demand for high-speed navigation. In particular, the Miller 2 encoding scheme strikes the best balance between reading performance (e.g., throughput) and noise tolerance for the multipath effect. Additionally, we show that the on-vehicle antenna should be oriented to maximize the available time for reading on-road tags, although it may reduce the received power by the tags in the forward link.

cs.RO

ChatGPT is on the Horizon: Could a Large Language Model be Suitable for Intelligent Traffic Safety Research and Applications?

ChatGPT embarks on a new era of artificial intelligence and will revolutionize the way we approach intelligent traffic safety systems. This paper begins with a brief introduction about the development of large language models (LLMs). Next, we exemplify using ChatGPT to address key traffic safety issues. Furthermore, we discuss the controversies surrounding LLMs, raise critical questions for their deployment, and provide our solutions. Moreover, we propose an idea of multi-modality representation learning for smarter traffic safety decision-making and open more questions for application improvement. We believe that LLM will both shape and potentially facilitate components of traffic safety research.

cs.CL

ChatGPT for Shaping the Future of Dentistry: The Potential of Multi-Modal Large Language Model

The ChatGPT, a lite and conversational variant of Generative Pretrained Transformer 4 (GPT-4) developed by OpenAI, is one of the milestone Large Language Models (LLMs) with billions of parameters. LLMs have stirred up much interest among researchers and practitioners in their impressive skills in natural language processing tasks, which profoundly impact various fields. This paper mainly discusses the future applications of LLMs in dentistry. We introduce two primary LLM deployment methods in dentistry, including automated dental diagnosis and cross-modal dental diagnosis, and examine their potential applications. Especially, equipped with a cross-modal encoder, a single LLM can manage multi-source data and conduct advanced natural language reasoning to perform complex clinical operations. We also present cases to demonstrate the potential of a fully automatic Multi-Modal LLM AI system for dentistry clinical application. While LLMs offer significant potential benefits, the challenges, such as data privacy, data quality, and model bias, need further study. Overall, LLMs have the potential to revolutionize dental diagnosis and treatment, which indicates a promising avenue for clinical application and research in dentistry.

cs.CL

TrafficSafetyGPT: Tuning a Pre-trained Large Language Model to a Domain-Specific Expert in Transportation Safety

Large Language Models (LLMs) have shown remarkable effectiveness in various general-domain natural language processing (NLP) tasks. However, their performance in transportation safety domain tasks has been suboptimal, primarily attributed to the requirement for specialized transportation safety expertise in generating accurate responses [1]. To address this challenge, we introduce TrafficSafetyGPT, a novel LLAMA-based model, which has undergone supervised fine-tuning using TrafficSafety-2K dataset which has human labels from government produced guiding books and ChatGPT-generated instruction-output pairs. Our proposed TrafficSafetyGPT model and TrafficSafety-2K train dataset are accessible at https://github.com/ozheng1993/TrafficSafetyGPT.

cs.CL

Traffic flow clustering framework using drone video trajectories to identify surrogate safety measures

The utilization of traffic conflict indicators is crucial for assessing traffic safety, especially when the crash data is unavailable. To identify traffic conflicts based on traffic flow characteristics across various traffic states, we propose a framework that utilizes unsupervised learning to automatically establish surrogate safety measures (SSM) thresholds. Different traffic states and corresponding transitions are identified with the three-phase traffic theory using high-resolution trajectory data. Meanwhile, the SSMs are mapped to the corresponding traffic states from the perspectives of time, space, and deceleration. Three models, including k-means, GMM, and Mclust, are investigated and compared to optimize the identification of traffic conflicts. It is observed that Mclust outperforms the others based on the evaluation metrics. According to the results, there is a variation in the distribution of traffic conflicts among different traffic states, wide moving jam (phase J) has the highest conflict risk, followed by synchronous flow (phase S), and free flow (phase F). Meanwhile, the thresholds of traffic conflicts cannot be fully represented by the same value through different traffic states. It reveals that the heterogeneity of thresholds is exhibited across traffic state transitions, which justifies the necessity of dynamic thresholds for traffic conflict analysis.

physics.soc-ph

Emergency Resource Layout with Multiple Objectives under Complex Disaster Scenarios

Effective placement of emergency rescue resources, particularly with joint suppliers in complex disaster scenarios, is crucial for ensuring the reliability, efficiency, and quality of emergency rescue activities. However, limited research has considered the interaction between different disasters and material classification, which are highly vital to the emergency rescue. This study provides a novel and practical framework for reliable strategies of emergency rescue under complex disaster scenarios. The study employs a scenario-based approach to represent complex disasters, such as earthquakes, mudslides, floods, and their interactions. In optimizing the placement of emergency resources, the study considers government-owned suppliers, framework agreement suppliers, and existing suppliers collectively supporting emergency rescue materials. To determine the selection of joint suppliers and their corresponding optimal material quantities under complex disaster scenarios, the research proposes a multi-objective model that integrates cost, fairness, emergency efficiency, and uncertainty into a facility location problem. Finally, the study develops an NSGA-II-XGB algorithm to solve a disaster-prone province example and verify the feasibility and effectiveness of the proposed multi-objective model and solution methods. The results show that the methodology proposed in this paper can greatly reduce emergency costs, rescue time, and the difference between demand and suppliers while maximizing the coverage of rescue resources. More importantly, it can optimize the scale of resources by determining the location and number of materials provided by joint suppliers for various kinds of disasters simultaneously. This research represents a promising step towards making informed configuration decisions in emergency rescue work.

stat.AP

Accessibility optimization of public transportation in historical districts:a study of Belin District, Xian

With the continuous improvement of urbanization and motorization, travel demand in historical blocks is higher than before. The contradiction between supply of transportation facilities and environmental protection is more serious. Traditional public transport planning methods aim to improve mobility. However, several existing studies do not put travelers in the first place, which ignore quantitative description of land uses and the characteristics of travelers. This paper intends to improve individual accessibility of public transportation in historical districts. Based on POI data, the calculation of accessibility combines the utility model and spatial interaction model. To optimize public transportation, the goal of improving individual accessibility is transformed into reducing residents' travel negative utility. Using a real-world dataset in Beilin District of Xian, the performance of the proposed model is evaluated. This paper calculates the present accessibility and uses ant colony algorithm to optimize public transportation. The results demonstrate that the calculation and optimization of public transportation accessibility are practical, which are also valuable to public transportation planning and organization in historical blocks.

physics.soc-ph

AVOID: Autonomous Vehicle Operation Incident Dataset Across the Globe

Crash data of autonomous vehicles (AV) or vehicles equipped with advanced driver assistance systems (ADAS) are the key information to understand the crash nature and to enhance the automation systems. However, most of the existing crash data sources are either limited by the sample size or suffer from missing or unverified data. To contribute to the AV safety research community, we introduce AVOID: an open AV crash dataset. Three types of vehicles are considered: Advanced Driving System (ADS) vehicles, Advanced Driver Assistance Systems (ADAS) vehicles, and low-speed autonomous shuttles. The crash data are collected from the National Highway Traffic Safety Administration (NHTSA), California Department of Motor Vehicles (CA DMV) and incident news worldwide, and the data are manually verified and summarized in ready-to-use format. In addition, land use, weather, and geometry information are also provided. The dataset is expected to accelerate the research on AV crash analysis and potential risk identification by providing the research community with data of rich samples, diverse data sources, clear data structure, and high data quality.

cs.RO