SearcharxivSearch

arXiv subjects

Gerhard P. Hancke

Publications and source records attributed to Gerhard P. Hancke.

3 recordsLinked to original sources

EgoCS-400K: An Egocentric Gameplay Dataset for World Models

The shift from video generation to interactive world modeling places new demands on data: beyond captioned videos, world models require temporally aligned video-action-language trajectories grounded in the actions, camera motion, states, and events that drive future scene changes. However, such data is difficult to obtain at scale. Web video datasets offer broad visual coverage but lack executable actions and reliable states; robotic datasets provide action and state supervision but are costly and limited in scene diversity; and existing simulators often lack large-scale human-driven interaction trajectories. In this paper, we introduce EgoCS-400K, a large-scale replay-grounded egocentric Counter-Strike dataset for world models, built from public professional CS and CS2 match demos that preserve human gameplay trajectories and enable parsing, replaying, rendering, and temporal alignment. We extract player states, view directions, movements, keyboard/button inputs, view-angle changes, weapon usage, game events, and round-level context, and render clean first-person videos from the same trajectories. EgoCS-400K contains over 400,000 first-person videos and 10,000 hours of gameplay from more than 1,000 matches and 40,000 rounds, covering 13 maps and 10 player viewpoints per round. It supports a range of interactive visual modeling tasks, including action-conditioned future prediction, state- and event-aware scene rollout, replay-grounded captioning, and agent egocentric action understanding. By connecting visual observations with human actions, camera motion, game states, and events at scale, EgoCS-400K serves as a practical bridge between passive web videos, controllable game simulation, and costly real-world embodied data.

cs.CV

A Review of Hydrogen-Enabled Resilience Enhancement for Multi-Energy Systems

Ensuring resilience in multi-energy systems (MESs) has become increasingly urgent and challenging due to the growing frequency and severity of extreme events, such as natural disasters, extreme weather, and cyber-physical attacks. Among the various approaches to enhancing MES resilience, hydrogen integration offers significant potential in cross-temporal, cross-spatial, and cross-sector flexibility, as well as black-start capability. Although considerable efforts have been devoted to this area, a systematic review of resilience enhancement in hydrogen-enabled MESs is still lacking. To address this gap, this paper presents a comprehensive review of hydrogen-enabled MES resilience enhancement. First, advantages, vulnerabilities, and challenges related to hydrogen-enabled MES resilience enhancement are summarized. Next, a resilience enhancement framework for hydrogen-enabled MESs is proposed, based on which existing resilience metrics and event-oriented contingency models are reviewed and discussed. Planning measures are then classified according to the types of hydrogen-related facilities, together with uncertainty handling methods, scenario generation methods, and planning problem formulation frameworks. In addition, operational enhancement measures are categorized into three response stages: prevention, emergency response, and restoration. Finally, research gaps are identified and future directions are discussed, including comprehensive resilience metric design, advanced extreme-event scenario generation, spatiotemporal cyber-physical contingency modeling under compound extreme events, coordinated planning and operation across multiple networks and timescales, low-carbon resilient planning and operation, and large language model-assisted whole-process resilience enhancement.

eess.SY

Deformable Object Tracking with Gated Fusion

The tracking-by-detection framework receives growing attentions through the integration with the Convolutional Neural Networks (CNNs). Existing tracking-by-detection based methods, however, fail to track objects with severe appearance variations. This is because the traditional convolutional operation is performed on fixed grids, and thus may not be able to find the correct response while the object is changing pose or under varying environmental conditions. In this paper, we propose a deformable convolution layer to enrich the target appearance representations in the tracking-by-detection framework. We aim to capture the target appearance variations via deformable convolution, which adaptively enhances its original features. In addition, we also propose a gated fusion scheme to control how the variations captured by the deformable convolution affect the original appearance. The enriched feature representation through deformable convolution facilitates the discrimination of the CNN classifier on the target object and background. Extensive experiments on the standard benchmarks show that the proposed tracker performs favorably against state-of-the-art methods.

cs.CV