SearcharxivSearch

arXiv subjects

Ilia Smirnov

Publications and source records attributed to Ilia Smirnov.

5 recordsLinked to original sources

Generalized Multi-hop Traffic Pressure for Heterogeneous Traffic Perimeter Control

Perimeter control (PC) prevents loss of traffic network capacity due to congestion in urban areas. Homogeneous PC allows all access points to a protected region to have identical permitted inflow. However, homogeneous PC performs poorly when the congestion in the protected region is heterogeneous (e.g., imbalanced demand) since the homogeneous PC does not consider specific traffic conditions around each perimeter intersection. When the protected region has spatially heterogeneous congestion, one needs to modulate the perimeter inflow rate to be higher near low-density regions and vice versa for high-density regions. A naïve approach is to leverage 1-hop traffic pressure to measure traffic condition around perimeter intersections, but such metric is too spatially myopic for PC. To address this issue, we formulate multi-hop downstream pressure grounded on Markov chain theory, which ``looks deeper'' into the protected region beyond perimeter intersections. In addition, we formulate a two-stage hierarchical control scheme that can leverage this novel multi-hop pressure to redistribute the total permitted inflow provided by a pre-trained deep reinforcement learning homogeneous control policy. Experimental results show that our heterogeneous PC approaches leveraging multi-hop pressure significantly outperform homogeneous PC in scenarios where the origin-destination flows are highly imbalanced with high spatial heterogeneity. Moveover, our approach is shown to be robust against turning ratio uncertainties by a sensitivity analysis.

cs.LG

Multi-hop Upstream Anticipatory Traffic Signal Control with Deep Reinforcement Learning

Coordination in traffic signal control is crucial for managing congestion in urban networks. Existing pressure-based control methods focus only on immediate upstream links, leading to suboptimal green time allocation and increased network delays. However, effective signal control inherently requires coordination across a broader spatial scope, as the effect of upstream traffic should influence signal control decisions at downstream intersections, impacting a large area in the traffic network. Although agent communication using neural network-based feature extraction can implicitly enhance spatial awareness, it significantly increases the learning complexity, adding an additional layer of difficulty to the challenging task of control in deep reinforcement learning. To address the issue of learning complexity and myopic traffic pressure definition, our work introduces a novel concept based on Markov chain theory, namely \textit{multi-hop upstream pressure}, which generalizes the conventional pressure to account for traffic conditions beyond the immediate upstream links. This farsighted and compact metric informs the deep reinforcement learning agent to preemptively clear the multi-hop upstream queues, guiding the agent to optimize signal timings with a broader spatial awareness. Simulations on synthetic and realistic (Toronto) scenarios demonstrate controllers utilizing multi-hop upstream pressure significantly reduce overall network delay by prioritizing traffic movements based on a broader understanding of upstream congestion.

cs.LG

How to manipulate nanoparticle morphology with vacancies

Stacking defects in noble metal nanoparticles significantly impact their optical, catalytic, and electrical properties. While some mechanisms behind their formation have been studied, the ability to deliberately manipulate nanoparticle bulk morphology remains largely unexplored. In this work, we introduce a pioneering mechanism - vacancy-driven twinning - that enables the transformation of face-centered cubic (fcc) gold into locally hexagonal close-packed (hcp) structures. This innovative approach, demonstrated through computational simulations, facilitates the creation of realistic , randomly multi-twinned nanoparticle models. By employing a recently developed multidomain X-ray diffraction method (MDXRD), we quantitatively assess the degree of twinning. It is a crucial step in transferring theoretical studies into practical applications. Our work aims to develop tools for modifying and controlling the bulk structure of fcc nanoparticles

cond-mat.mtrl-sci

SECRM-2D: RL-Based Efficient and Comfortable Route-Following Autonomous Driving with Analytic Safety Guarantees

Over the last decade, there has been increasing interest in autonomous driving systems. Reinforcement Learning (RL) shows great promise for training autonomous driving controllers, being able to directly optimize a combination of criteria such as efficiency comfort, and stability. However, RL- based controllers typically offer no safety guarantees, making their readiness for real deployment questionable. In this paper, we propose SECRM-2D (the Safe, Efficient and Comfortable RL- based driving Model with Lane-Changing), an RL autonomous driving controller (both longitudinal and lateral) that balances optimization of efficiency and comfort and follows a fixed route, while being subject to hard analytic safety constraints. The aforementioned safety constraints are derived from the criterion that the follower vehicle must have sufficient headway to be able to avoid a crash if the leader vehicle brakes suddenly. We evaluate SECRM-2D against several learning and non-learning baselines in simulated test scenarios, including freeway driving, exiting, merging, and emergency braking. Our results confirm that representative previously-published RL AV controllers may crash in both training and testing, even if they are optimizing a safety objective. By contrast, our controller SECRM-2D is successful in avoiding crashes during both training and testing, improves over the baselines in measures of efficiency and comfort, and is more faithful in following the prescribed route. In addition, we achieve a good theoretical understanding of the longitudinal steady-state of a collection of SECRM-2D vehicles.

cs.RO

Nanopowder Diffraction

As in the available literature there are still misconceptions about powder diffraction phenomena observed for small nanocrystals ($D<10$ nm), we propose here a systematic and concise review of the involved issues that can be approached by atomistic simulations. Most of phenomenological tools of powder diffraction can be now verified constructing realistic atomistic models, following their thermodynamics and impact on the diffraction pattern. The models concern small cuts of the perfect lattice as well as relaxed nanocrystals also with typical strain and faults, proven by experiments to approximate the real nanocrystals. The discussed examples concern metal nanocrystals. We describe the origin of peak shifts that for cuts of the perfect lattice are mostly due to multiplying of broad profiles by steep slope factors -- atomic scattering and Lorentz. The Lorentz factor embedded in the Debye summation is discussed in more detail. For models relaxed with realistic force fields and for the real nanocrystals the peak shift additionally includes contribution from surface relaxation which enables development of an experimental {\it{in situ }} method sensitive to the state of the surface. Such results are briefly reviewed. For small fcc nanocrystals we discuss the importance and effect of multiple (111) cross twinning on the peak shift and height. The strain and size effects for the perfect multitwinned clusters -- decahedra and icosahedra, are explained and visualised. The manuscript proposes new methods to interpret powder diffraction patterns of real, defected (twinned) fcc nanoparticles and points to unsuitability of the Rietveld method in this case.

cond-mat.mtrl-sci