SearcharxivSearch

arXiv subjects

Hailong Liu

Publications and source records attributed to Hailong Liu.

At least 19 recordsLinked to original sources

KFTD: Koopman-Fourier Time-Differentiable Network for Continuous Ocean Spatiotemporal Forecasting

Accurate oceanic forecasting is critical for climate monitoring and disaster early warning. However, ocean spatiotemporal forecasting encounters the double challenges of modeling complex dynamical systems and ensuring computational efficiency. We present Koopman Fourier Time-Differentiable (KFTD) Network, a time continuous twostage paradigm that decouples interpolation from prediction to achieve efficient and scalable spatiotemporal modeling. We map complex nonlinear dynamics into the Koopman linear space and exploit Fourier analysis to enable continuous time interpolation at arbitrary sub-steps. A lightweight residual network consumes the high fidelity intermediate states to yield the final forecast. Unlike diffusion models, KFTD eliminates multi step noise sampling and directly evolves the system in continuous time, yielding a 4 computational speedup. We further introduce a DPP Loss that supports arbitrary PDE constraints in an endtoend manner, breaking the physical consistency bottleneck of pure data-driven approaches. Empirical results on four ocean datasets confirm that our continuous time framework reduces MSE by an average of 5.6% (up to 12.7% for SST) and improves efficiency over MCVD by 76.25%.

physics.ao-ph

High-order thermodynamic nonequilibrium in three-dimensional compressible flows: Kinetic moment closure and multigradient coupling

High-order thermodynamic nonequilibrium (TNE) in three-dimensional compressible flows reflects the breakdown of low-order kinetic moment closure in strong-gradient regions. Using Chapman-Enskog analysis, we identify the kinetic moment constraints required to describe third-order TNE. The analysis yields the third-order constitutive relations and evolution equations for the viscous stress and heat flux, together with second-order expressions for their associated higher-order fluxes. These constraints enable the construction of a three-dimensional super-Burnett-level discrete Boltzmann model with 91 discrete velocities. The resulting D3V91 model reproduces shock-tube wave structures and resolves high-order TNE contributions that lower-order DBMs do not capture reliably. These results demonstrate that high-order TNE has a multigradient, rather than single-gradient, origin. For the four TNE quantities considered here, odd-order central moments, including the heat flux and the viscous-stress flux , are primarily governed by temperature gradients, whereas even-order central moments, including the viscous stress and the heat-flux-related flux , are dominated by velocity gradients. These leading-gradient dependences are not exclusive; they are substantially modified by density gradients, secondary gradients and transition-layer widths through higher-order derivative terms, gradient products and cross-couplings. When the secondary contributions become comparable to the leading-gradient terms, the nonequilibrium response transitions from a near-linear regime to an approximately exponential regime. This work establishes a super-Burnett-level DBM framework that treats kinetic moment closure and multigradient coupling consistently, providing a basis for resolving and interpreting high-order TNE in three-dimensional compressible flows.

physics.flu-dyn

Hanger Reflex Based Driving Assistance for Drivers with Peripheral Visual Field Defects

Drivers with peripheral visual field defects may fail to notice pedestrians in their peripheral visual field, leading to delayed hazard awareness and increased collision risk. This study explores hanger reflex cue (HRC) as a driving assistance method for drivers with peripheral visual field defects, in which mechanical pressure is applied to specific regions of the head to facilitate anticipatory orientation toward potentially risky pedestrians and support safer driving. In a driving simulator experiment with 15 participants, we compared driving behavior with and without HRC during pedestrian encounters under simulated peripheral visual field defect. The results showed that HRC significantly shifted drivers' modal head rotation angle toward the risky pedestrian and significantly increased gaze duration toward that pedestrian. Collision occurrence was lower in the w/ HRC condition than in the w/o HRC condition, although the direct effect of HRC on collision occurrence showed only a marginal trend. A piecewise structural equation modeling analysis further suggested that HRC may contribute to collision reduction through a sequential pathway from head rotation to gaze allocation and then to collision occurrence. These findings provide preliminary evidence that HRC can support anticipatory attention allocation toward peripheral hazards and may offer a promising driving assistance method for drivers with visual field impairment.

cs.HC

A discrete Boltzmann model with state-dependent power-law relaxation time for nonequilibrium transport in compressible flows

Thermodynamic nonequilibrium effects play a central role in momentum and energy transport in compressible flows. In conventional BGK kinetic models, the relaxation time $\tau$ is taken as a constant, which neglects the dependence of the relaxation process on local macroscopic states. To overcome this limitation, we develop a discrete Boltzmann model with a density- and temperature-dependent power-law relaxation time, termed DTRT-DBM, in which $\tau=\tau_0(\rho/\rho_0)^a(T/T_0)^b$. This formulation extends the discrete Boltzmann framework to flows with spatially varying nonequilibrium intensity. The model is validated by the Sod shock tube and by analytical solutions for viscous stress and heat flux, demonstrating accurate recovery of both macroscopic wave structures and nonequilibrium quantities across shock waves, rarefaction waves, and contact discontinuities. On this basis, phase diagrams of viscous stress and heat flux are constructed to examine how these quantities depend on the power-law exponents $a$ and $b$. The extrema of these quantities depend exponentially on the model parameters and exhibit regime-dependent behaviour. The roles of $a$ and $b$ are not symmetric: the nonequilibrium response is more sensitive to $a$ when density gradients dominate, but more sensitive to $b$ when temperature gradients dominate. Within the parameter range and flow configurations examined here, higher-order viscous stress increases the growth rate of the total viscous-stress extremum, whereas higher-order heat flux reduces the growth rate of the total heat-flux extremum. These results show that the proposed model can capture different higher-order nonequilibrium responses in compressible flows and provides a framework for the modelling and analysis of multiscale nonequilibrium processes.

physics.flu-dyn

An eHMI Presenting Request-to-Intervene and Takeover Status of Level 3 Automated Vehicles to Support Surrounding Traffic Safety

Level 3 automated vehicles (AVs) issue a request to intervene (RtI) when the automated driving system approaches its system limitations. Although this takeover transition is safety-critical, it is usually invisible to surrounding manually driven vehicle (MV) drivers. This study proposes an external human-machine interface (eHMI) called eHMI C+O that externalizes the RtI-related takeover status of a Level~3 AV using cyan and orange light bars. A driving-simulator experiment with 40 participants examined whether the proposed eHMI supports surrounding MV drivers during AV takeover scenarios. The results showed that, compared with the ADS-status-only eHMI condition, which is similar to ``Automated Driving Marker Lights,'' and the no-eHMI condition, the proposed eHMI C+O significantly improved participants' understanding of the AV's driving intention, their prediction of its behavior, and their perceived sufficiency of the information presented by the AV. It also reduced hesitation, increased confidence, and promoted earlier and larger increases in time headway after the RtI was issued. In the AV accident scenario, eHMI C+O significantly reduced the odds of accident involvement for the following MV compared with the no-eHMI condition, corresponding to a 76.8% reduction in accident odds. Exploratory path analysis suggested that the safety benefit of the proposed eHMI C+O may be associated with improved situation awareness and earlier defensive driving responses. These findings indicate that externalizing RtI-related takeover status can help surrounding drivers better understand Level 3 AVs and respond more safely during safety-critical takeover transitions.

cs.HC

How Do People Accept Robot in Public Space? A Comparative Study between Germany and Japan

With the increasing deployment of robots in public spaces, encounters between robots and incidentally copresent persons (InCoPs) are becoming more frequent. However, InCoPs remain largely underexplored in the literature, particularly from a cross-cultural perspective. Therefore, the present study investigates differences in InCoPs' existence acceptance (EA) of autonomous cleaning robots in public spaces among Japanese and German participants. Online survey results revealed that Germans showed significantly higher EA. Social Norms and Trust were the strongest positive EA predictors across cultures. More specifically, for Germans, EA was directly influenced by Usefulness, Interest and Anger, showing a functional-affective pattern where functional perceptions boost EA and anger suppresses it. For Japanese participants, Trust, Surprise and Fear were the direct associational factors, forming a trust-emotion pattern. These findings suggest that the cognitive and emotional drivers of public robot acceptance may vary across countries, emphasizing the need for adaptive robot design.

cs.HC

An Educational Human Machine Interface Providing Request-to-Intervene Trigger and Reason Explanation for Enhancing the Driver's Comprehension of ADS's System Limitations

Level 3 automated driving systems (ADS) have attracted significant attention and are being commercialized. A level 3 ADS prompts the driver to take control by issuing a request to intervene (RtI) when its operational design domains (ODD) are exceeded. However, complex traffic situations can cause drivers to perceive multiple potential triggers of RtI simultaneously, causing hesitation or confusion during take-over. Therefore, drivers need to clearly understand the ADS's system limitations to ensure safe take-over. This study proposes a voice-based educational human machine interface~(HMI) for providing RtI trigger cues and reason to help drivers understand ADS's system limitations. The results of a between-group experiment using a driving simulator showed that incorporating effective trigger cues and reason into the RtI was related to improved driver comprehension of the ADS's system limitations. Moreover, most participants, instructed via the proposed method, could proactively take over control of the ADS in cases where RtI fails; meanwhile, their number of collisions was lower compared with the other RtI HMI conditions. Therefore, using the proposed method to continually enhance the driver's understanding of the system limitations of ADS through the proposed method is associated with safer and more effective real-time interactions with ADS.

cs.HC

Driving Path Indication Reduces Motion Sickness and Influences Head Motion of Passengers in Autonomous Personal Mobility Vehicle

Autonomous personal mobility vehicles (APMVs) are novel smart mobility devices designed to provide automated individual transportation in indoor or mixed-traffic environments. However, in such environments, frequent pedestrian avoidance maneuvers may cause rapid steering adjustments and passive postural responses from passengers, thereby increasing the risk of motion sickness. This study investigated whether indicating the future driving path could mitigate motion sickness in APMV passengers. A mixed-design experiment was conducted with 40 participants under two self-reported genders as a between-subject factor (male and female), two driving paths as a between-subject factor (irregular and regular) and three driving conditions as a within-subject factor (manual driving (MD), automated driving without path indication (AD w/o path), and automated driving with path indication (AD w/ path)). Motion sickness was evaluated using the Motion Illness Symptom Classification (MISC), and head motion was assessed by calculating the delay time of participants' head yaw rate relative to APMV's yaw rate in the turning direction. The results showed that driving condition was the only factor that significantly affected both motion sickness and head-motion delay. Compared with the AD w/o path condition, both the MD and AD w/ path conditions were associated with lower motion sickness severity, longer motion sickness onset latency, and earlier head motion relative to vehicle motion. Notably, the AD w/ path condition achieved motion sickness levels comparable to those in the MD condition. Furthermore, repeated-measures correlation analysis showed significant associations between head-motion delay and all MISC metrics but the underlying physiological mechanism remains to be elucidated. These findings suggest that presenting information about future driving path can mitigate motion sickness in APMV passengers.

cs.HC

Kinetic study of compressible Rayleigh-Taylor instability with time-varying acceleration

Rayleigh-Taylor (RT) instability commonly arises in compressible systems with time-dependent acceleration in practical applications. To capture the complex dynamics of such systems, a two-component discrete Boltzmann method is developed to systematically investigate the compressible RT instability driven by variable acceleration. Specifically, the effects of different acceleration periods, amplitudes, and phases are systematically analyzed. The simulation results are interpreted from three key perspectives: the density gradient, which characterizes the spatial variation in density; the thermodynamic non-equilibrium strength, which quantifies the system's deviation from local thermodynamic equilibrium; and the fraction of non-equilibrium regions, which captures the spatial distribution of non-equilibrium behaviors. Notably, the fluid system exhibits rich and diverse dynamic patterns resulting from the interplay of multiple competing physical mechanisms, including time-dependent acceleration, RT instability, diffusion, and dissipation effects. These findings provide deeper insights into the evolution and regulation of compressible RT instability under complex driving conditions.

physics.flu-dyn

Multispectral radiation temperature inversion based on Transformer-LSTM-SVM

The key challenge in multispectral radiation thermometry is accurately measuring emissivity. Traditional constrained optimization methods often fail to meet practical requirements in terms of precision, efficiency, and noise resistance. However, the continuous advancement of neural networks in data processing offers a potential solution to this issue. This paper presents a multispectral radiation thermometry algorithm that combines Transformer, LSTM (Long Short-Term Memory), and SVM (Support Vector Machine) to mitigate the impact of emissivity, thereby enhancing accuracy and noise resistance. In simulations, compared to the BP neural network algorithm, GIM-LSTM, and Transformer-LSTM algorithms, the Transformer-LSTM-SVM algorithm demonstrates an improvement in accuracy of 1.23%, 0.46% and 0.13%, respectively, without noise. When 5% random noise is added, the accuracy increases by 1.39%, 0.51%, and 0.38%, respectively. Finally, experiments confirmed that the maximum temperature error using this method is less than 1%, indicating that the algorithm offers high accuracy, fast processing speed, and robust noise resistance. These characteristics make it well-suited for real-time high-temperature measurements with multi-wavelength thermometry equipment.

physics.optics

Where Do Passengers Gaze? Impact of Passengers' Personality Traits on Their Gaze Pattern Toward Pedestrians During APMV-Pedestrian Interactions with Diverse eHMIs

Autonomous Personal Mobility Vehicles (APMVs) are designed to address the ``last-mile'' transportation challenge for everyone. When an APMV encounters a pedestrian, it uses an external Human-Machine Interface (eHMI) to negotiate road rights. Through this interaction, passengers also engage with the process. This study examines passengers' gaze behavior toward pedestrians during such interactions, focusing on whether different eHMI designs influence gaze patterns based on passengers' personality traits. The results indicated that when using a visual-based eHMI, passengers often struggled to perceive the communication content. Consequently, passengers with higher Neuroticism scores, who were more sensitive to communication details, might seek cues from pedestrians' reactions. In addition, a multimodal eHMI (visual and voice) using neutral voice did not significantly affect the gaze behavior of passengers toward pedestrians, regardless of personality traits. In contrast, a multimodal eHMI using affective voice encouraged passengers with high Openness to Experience scores to focus on pedestrians' heads. In summary, this study revealed how different eHMI designs influence passengers' gaze behavior and highlighted the effects of personality traits on their gaze patterns toward pedestrians, providing new insights for personalized eHMI designs.

cs.HC

Data-driven Causal Discovery for Pedestrians-Autonomous Personal Mobility Vehicle Interactions with eHMIs: From Psychological States to Walking Behaviors

Autonomous personal mobility vehicle (APMV) is a new type of small smart vehicle designed for mixed-traffic environments, including interactions with pedestrians. To enhance the interaction experience between pedestrians and APMVs and to prevent potential risks, it is crucial to investigate pedestrians' walking behaviors when interacting with APMVs and to understand the psychological processes underlying these behaviors. This study aims to investigate the causal relationships between subjective evaluations of pedestrians and their walking behaviors during interactions with an APMV equipped with an external human-machine interface (eHMI). An experiment of pedestrian-APMV interaction was conducted with 42 pedestrian participants, in which various eHMIs on the APMV were designed to induce participants to experience different levels of subjective evaluations and generate the corresponding walking behaviors. Based on the hypothesized model of the pedestrian's cognition-decision-behavior process, the results of causal discovery align with the previously proposed model. Furthermore, this study further analyzes the direct and total causal effects of each factor and investigates the causal processes affecting several important factors in the field of human-vehicle interaction, such as situation awareness, trust in vehicle, risk perception, hesitation in decision making, and walking behaviors.

cs.HC

Technical Report: Competition Solution For Modelscope-Sora

This report presents the approach adopted in the Modelscope-Sora challenge, which focuses on fine-tuning data for video generation models. The challenge evaluates participants' ability to analyze, clean, and generate high-quality datasets for video-based text-to-video tasks under specific computational constraints. The provided methodology involves data processing techniques such as video description generation, filtering, and acceleration. This report outlines the procedures and tools utilized to enhance the quality of training data, ensuring improved performance in text-to-video generation models.

cs.CV

A Digital Human Model for Symptom Progression of Vestibular Motion Sickness based on Subjective Vertical Conflict Theory

Digital human models of motion sickness have been actively developed, among which models based on subjective vertical conflict (SVC) theory are the most actively studied. These models facilitate the prediction of motion sickness in various scenarios such as riding in a car. Most SVC theory models predict the motion sickness incidence (MSI), which is defined as the percentage of people who would vomit with the given specific motion stimulus. However, no model has been developed to describe milder forms of discomfort or specific symptoms of motion sickness, even though predicting milder symptoms is important for applications in automobiles and daily use vehicles. Therefore, the purpose of this study was to build a computational model of symptom progression of vestibular motion sickness based on SVC theory. We focused on a model of vestibular motion sickness with six degrees-of-freedom (6DoF) head motions. The model was developed by updating the output part of the state-of-the-art SVC model, termed the 6DoF-SVC (IN1) model, from MSI to the MIsery SCale (MISC), which is a subjective rating scale for symptom progression. We conducted an experiment to measure the progression of motion sickness during a straight fore-aft motion. It was demonstrated that our proposed method, with the parameters of the output parts optimized by the experimental results, fits well with the observed MISC.

cs.HC

Enhancing Hybrid Eye Typing Interfaces with Word and Letter Prediction: A Comprehensive Evaluation

Eye typing interfaces enable a person to enter text into an interface using only their own eyes. But despite the inherent advantages of touchless operation and intuitive design, such eye-typing interfaces often suffer from slow typing speeds, resulting in slow words per minute (WPM) counts. In this study, we add word and letter prediction to the eye-typing interface and investigate users' typing performance as well as their subjective experience while using the interface. In experiment 1, we compared three typing interfaces with letter prediction (LP), letter+word prediction (L+WP), and no prediction (NoP), respectively. We found that the interface with L+WP achieved the highest average text entry speed (5.48 WPM), followed by the interface with LP (3.42 WPM), and the interface with NoP (3.39 WPM). Participants were able to quickly understand the procedural design for word prediction and perceived this function as very helpful. Compared to LP and NoP, participants needed more time to familiarize themselves with L+WP in order to reach a plateau regarding text entry speed. Experiment 2 explored training effects in L+WP interfaces. Two moving speeds were implemented: slow (6.4°/s same speed as in experiment 1) and fast (10°/s). The study employed a mixed experimental design, incorporating moving speeds as a between-subjects factor, to evaluate its influence on typing performance throughout 10 consecutive training sessions. The results showed that the typing speed reached 6.17 WPM for the slow group and 7.35 WPM for the fast group after practice. Overall, the two experiments show that adding letter and word prediction to eye-typing interfaces increases typing speeds. We also find that more extended training is required to achieve these high typing speeds.

cs.HC

Is Silent eHMI Enough? A Passenger-Centric Study on Effective eHMI for Autonomous Personal Mobility Vehicles in the Field

Autonomous Personal Mobility Vehicle (APMV) is a miniaturized autonomous vehicle designed to provide short-distance mobility to everyone in pedestrian-rich environments. By the characteristic of the open design, passengers on the APMV are exposed to the communication between the eHMI deployed on APMVs and pedestrians. Therefore, to ensure an optimal passenger experience, eHMI designs for APMVs must consider the potential impact of APMV-pedestrian communications on passengers' subjective feelings. To this end, this study discussed three external human-machine interface (eHMI) designs, i.e., 1) graphical user interface (GUI)-based eHMI with text message (eHMI-T), 2) multimodal user interface (MUI)-based eHMI with neutral voice (eHMI-NV), and 3) MUI-based eHMI with affective voice (eHMI-AV), from the perspective of APMV passengers in the communication between APMV and pedestrians. In the riding field experiment (N=24), we found that eHMI-T may be less suitable for APMVs. This conclusion was drawn based on passengers' feedback, as they expressed an awkward feeling during the "silent time" when the eHMI-T provided information only to pedestrians but not to passengers. Additionally, these two MUI-based eHMIs with voice cues had their own advantages, i.e., eHMI-NV has an advantage in pragmatic quality, while eHMI-AV has an advantage in hedonic quality. The study also highlights the necessity of considering passengers' personalities when desig

cs.HC

GaVe: A Webcam-Based Gaze Vending Interface Using One-Point Calibration

Even before the Covid-19 pandemic, beneficial use cases for hygienic, touchless human-machine interaction have been explored. Gaze input, i.e., information input via eye-movements of users, represents a promising method for contact-free interaction in human-machine systems. In this paper, we present the GazeVending interface (GaVe), which lets users control actions on a display with their eyes. The interface works on a regular webcam, available on most of today's laptops, and only requires a one-point calibration before use. GaVe is designed in a hierarchical structure, presenting broad item cluster to users first and subsequently guiding them through another selection round, which allows the presentation of a large number of items. Cluster/item selection in GaVe is based on the dwell time of fixations, i.e., the time duration that users look at a given Cluster/item. A user study (N=22) was conducted to test optimal dwell time thresholds and comfortable human-to-display distances. Users' perception of the system, as well as error rates and task completion time were registered. We found that all participants were able to use the system with a short time training, and showed good performance during system usage, selecting a target item within a group of 12 items in 6.76 seconds on average. Participants were able to quickly understand and know how to interact with the interface. We provide design guidelines for GaVe and discuss the potentials of the system.

cs.HC

Enhancing the Driver's Comprehension of ADS's System Limitations: An HMI for Providing Request-to-Intervene Trigger Information

Level 3 automated driving systems (ADS) have attracted significant attention and are being commercialized. A Level 3 ADS prompts the driver to take control by requesting to intervene (RtI) when its operational design domain (ODD) or system limitations are exceeded. However, complex traffic situations may lead drivers to perceive multiple potential triggers of RtI simultaneously, causing hesitation or confusion during take-over. Therefore, drivers must clearly understand the ADS's system limitations to understand the triggers of RtI and ensure safe take-over. In this study, we propose a voice-based HMI for providing RtI trigger cues to help drivers understand ADS's system limitations. The results of a between-group experiment using a driving simulator showed that incorporating effective trigger cues into the RtI enabled drivers to comprehend the ADS's system limitations better and reduce collisions. It also improved the subjective evaluations of drivers, such as the comprehensibility of system limitations, hesitation in response to RtI, and acceptance of ADS behaviors when encountering RtI while using the ADS. Therefore, enhanced comprehension resulting from trigger cues is essential for promoting a safer and better user experience using ADS during RtI.

cs.HC