SearcharxivSearch

arXiv subjects

Kaifeng Wang

Publications and source records attributed to Kaifeng Wang.

4 recordsLinked to original sources

Autonomous Reshaping of Expression Landscapes by DNA Methylation

DNA methylation is usually treated as an epigenetic memory mark: transcriptional history is written into regulatory DNA and later stabilizes a chosen cell identity. This picture explains persistence, but it makes memory passive. Here we show that the same promoter-level coupling required for methylation memory can instead turn methylation into an internal control variable for regulatory dynamics. Transcription-factor occupancy protects regulatory DNA from methylation, while methylation shifts later transcription-factor binding thresholds. Under time-scale separation, this reciprocal loop separates into fast expression dynamics conditioned on methylation and a slow methylation flow written by expression. Minimal promoter, self-activation, and fate-toggle models show that this feedback does more than preserve a past state: it autonomously reshapes the expression landscape. In a methylation-coupled toggle, the preferred expression state can move continuously through single-well drift, allowing commitment without first entering a multiwell regime. Stochastic simulations further show that evolving methylation reduces fate reversals relative to a frozen landscape, making weak early expression bias more predictive of later fate. These results recast DNA methylation from a downstream stabilizer of cell identity into a slow dynamical coordinate that can help determine how regulatory states are chosen.

q-bio.MN

Red-Team Multi-Agent Reinforcement Learning for Emergency Braking Scenario

Current research on decision-making in safety-critical scenarios often relies on inefficient data-driven scenario generation or specific modeling approaches, which fail to capture corner cases in real-world contexts. To address this issue, we propose a Red-Team Multi-Agent Reinforcement Learning framework, where background vehicles with interference capabilities are treated as red-team agents. Through active interference and exploration, red-team vehicles can uncover corner cases outside the data distribution. The framework uses a Constraint Graph Representation Markov Decision Process, ensuring that red-team vehicles comply with safety rules while continuously disrupting the autonomous vehicles (AVs). A policy threat zone model is constructed to quantify the threat posed by red-team vehicles to AVs, inducing more extreme actions to increase the danger level of the scenario. Experimental results show that the proposed framework significantly impacts AVs decision-making safety and generates various corner cases. This method also offers a novel direction for research in safety-critical scenarios.

cs.LG

Dynamic Residual Safe Reinforcement Learning for Multi-Agent Safety-Critical Scenarios Decision-Making

In multi-agent safety-critical scenarios, traditional autonomous driving frameworks face significant challenges in balancing safety constraints and task performance. These frameworks struggle to quantify dynamic interaction risks in real-time and depend heavily on manual rules, resulting in low computational efficiency and conservative strategies. To address these limitations, we propose a Dynamic Residual Safe Reinforcement Learning (DRS-RL) framework grounded in a safety-enhanced networked Markov decision process. It's the first time that the weak-to-strong theory is introduced into multi-agent decision-making, enabling lightweight dynamic calibration of safety boundaries via a weak-to-strong safety correction paradigm. Based on the multi-agent dynamic conflict zone model, our framework accurately captures spatiotemporal coupling risks among heterogeneous traffic participants and surpasses the static constraints of conventional geometric rules. Moreover, a risk-aware prioritized experience replay mechanism mitigates data distribution bias by mapping risk to sampling probability. Experimental results reveal that the proposed method significantly outperforms traditional RL algorithms in safety, efficiency, and comfort. Specifically, it reduces the collision rate by up to 92.17%, while the safety model accounts for merely 27% of the main model's parameters.

cs.RO

A Novel Energy-Efficient Salicide-Enhanced Tunnel Device Technology Based on 300mm Foundry Platform Towards AIoT Applications

This work demonstrates a novel energy-efficient tunnel FET (TFET)-CMOS hybrid foundry platform for ultralow-power AIoT applications. By utilizing the proposed monolithic integration process, the novel complementary n and p-type Si TFET technology with dopant segregated source junction and self-aligned drain underlap design is successfully integrated into a 300mm CMOS baseline process without CMOS performance penalty and any new materials, experimentally demonstrating the large Ion and record high Ion/Ioff ratio of 10^7 among TFETs by industry-manufacturers. The device performance and variability are also co-optimized for high-volume production. Further circuit-level implementations are presented based on the calibrated compact model. The proposed TFET-CMOS hybrid logic and SRAM topologies show significant energy efficiency improvement with comparable operation speed compared with standard CMOS circuits, indicating its great potential for power-constraint AIoT applications.

cond-mat.mes-hall