SearcharxivSearch

arXiv subjects

Dandan Ma

Publications and source records attributed to Dandan Ma.

8 recordsLinked to original sources

LLMs for High-Frequency Decision-Making: Normalized Action Reward-Guided Consistency Policy Optimization

While Large Language Models (LLMs) form the cornerstone of sequential decision-making agent development, they have inherent limitations in high-frequency decision tasks. Existing research mainly focuses on discrete embodied decision scenarios with low-frequency and significant semantic differences in state space (e.g., household planning). These methods suffer from limited performance in high-frequency decision-making tasks, since high-precision numerical state information in such tasks undergoes frequent updates with minimal fluctuations, and exhibiting policy misalignment between the learned sub-tasks and composite tasks. To address these issues, this paper proposes Normalized Action Reward guided Consistency Policy Optimization (NAR-CP). 1) Our method first acquires predefined dense rewards from environmental feedback of candidate actions via reward functions, then completes reward shaping through normalization, and theoretically verifies action reward normalization does not impair optimal policy. 2) To reduce policy misalignment in composite tasks, we use LLMs to infer sub-observation candidate actions and generate joint policies, with consistency loss ensuring precise alignment between global semantic policies and sub-semantic policies. Experiments on UAV pursuit, a typical high-frequency task, show our method delivers superior performance on independent and composite tasks with excellent generalization to unseen tasks.

cs.AI

SAGE-LLM: Towards Safe and Generalizable LLM Controller with Fuzzy-CBF Verification and Graph-Structured Knowledge Retrieval for UAV Decision

In UAV dynamic decision, complex and variable hazardous factors pose severe challenges to the generalization capability of algorithms. Despite offering semantic understanding and scene generalization, Large Language Models (LLM) lack domain-specific UAV control knowledge and formal safety assurances, restricting their direct applicability. To bridge this gap, this paper proposes a train-free two-layer decision architecture based on LLMs, integrating high-level safety planning with low-level precise control. The framework introduces three key contributions: 1) A fuzzy Control Barrier Function verification mechanism for semantically-augmented actions, providing provable safety certification for LLM outputs. 2) A star-hierarchical graph-based retrieval-augmented generation system, enabling efficient, elastic, and interpretable scene adaptation. 3) Systematic experimental validation in pursuit-evasion scenarios with unknown obstacles and emergent threats, demonstrating that our SAGE-LLM maintains performance while significantly enhancing safety and generalization without online training. The proposed framework demonstrates strong extensibility, suggesting its potential for generalization to broader embodied intelligence systems and safety-critical control domains.

cs.RO

In silico modeling for personalized stenting in aortic coarctation

Stent intervention is a recommended therapy to reduce the pressure gradient and restore blood flow for patients with coarctation of the aorta (CoA). In this work, we developed a framework for personalized stent intervention in CoA using in silico modeling, combining computational fluid dynamics (CFD) and image-based prediction of the geometry of the aorta after stent intervention. Firstly, the blood flow in the aorta, whose geometry was reconstructed from magnetic resonance imaging (MRI) data, was numerically modeled using the lattice Boltzmann method (LBM). Both large eddy simulation (LES) and direct numerical simulation (DNS) were considered to adequately resolve the turbulent hemodynamics, with boundary conditions extracted from phase-contrast flow MRI. By comparing the results from CFD and 4D-Flow MRI in 3D-printed flow phantoms, we concluded that the LBM based LES is capable of obtaining accurate aortic flow with acceptable computational cost. In silico stent implantation for a patient with CoA was then performed by predicting the deformed geometry after stent intervention and predicting the blood flow. By evaluating the pressure drop and maximum wall shear stress, an optimal stent can be selected.

physics.med-ph

Co-contributorship Network and Division of Labor in Individual Scientific Collaborations

Collaborations are pervasive in current science. Collaborations have been studied and encouraged in many disciplines. However, little is known how a team really functions from the detailed division of labor within. In this research, we investigate the patterns of scientific collaboration and division of labor within individual scholarly articles by analyzing their co-contributorship networks. Co-contributorship networks are constructed by performing the one-mode projection of the author-task bipartite networks obtained from 138,787 papers published in PLoS journals. Given a paper, we define three types of contributors: Specialists, Team-players, and Versatiles. Specialists are those who contribute to all their tasks alone; team-players are those who contribute to every task with other collaborators; and versatiles are those who do both. We find that team-players are the majority and they tend to contribute to the five most common tasks as expected, such as "data analysis" and "performing experiments". The specialists and versatiles are more prevalent than expected by a random-graph null model. Versatiles tend to be senior authors associated with funding and supervisions. Specialists are associated with two contrasting roles: the supervising role as team leaders or marginal and specialized contributions.

cs.SI

Unidirectional and controllable higher-order diffraction by a Rydberg electromagnetically induced grating

A method for diffracting the weak probe beam into unidirectional and higher-order directions is proposed via a novel Rydberg electromagnetically induced grating, providing a new way for the implementations of quantum devices with cold Rydberg atoms. The proposed scheme utilizes a suitable and position-dependent adjustment to the two-photon detuning besides the modulation of the standing-wave coupling field, bringing a in-phase modulation which can change the parity of the dispersion. We observe that when the modulation amplitude is appropriate, a perfect unidirectional diffraction grating can be realized. In addition, due to the mutual effect between the van der Waals (vdWs) interaction and the atom-field interaction length that deeply improves the dispersion of the medium, the probe energy can be counter-intuitively transferred into higher-order diffractions as increasing the vdWs interaction, leading to the realization of a controllable higher-order diffraction grating via strong blockade.

physics.atom-ph

Properties of collective Rabi oscillations with two Rydberg atoms

Motivated by experimental advances [e.g. A. Ga{ë}tan {\it et.al.} Nat. Phys. 5 115 (2009)] that the collective excitation of two Rydberg atoms was observed, we provide an elaborate theoretical study for the dynamical behavior of two-atom Rabi oscillations. In the large-intermediate-detuning case, the two-photon Rabi oscillation is found to be significantly affected by the strength of the interatomic van der Waals interaction. With a careful comparison of the exact numbers and values of the oscillation frequency, we propose a new way to determine the strength of excitation blockade, well agreeing with the previous universal criterion for full, partial and none blockade regions. In the small-intermediate-detuning case we find a blockade-like effect, but the collective enhancement factor is smaller than $\sqrt{2}$ due to the quantum interference of double optical transitions involving the intermediate state. Moreover, a fast two-photon Rabi oscillation in $ns$ timescale is manifested by employing intense lasers with an intensity of $\sim$MW/cm$^2$, offering a possibility of ultrafast control of quantum dynamics with Rydberg atoms.

physics.atom-ph

Robust switching of superposition-states via a coherent double stimulated Raman adiabatic passage

Coherent manipulation of quantum states is of crucial importance in accurate control of a quantum system. A fundamental goal is coherently transferring the population of a desired state with near-unit fidelity. For this propose, we theoretically demonstrate a novel coherent double-stimulated Raman adiabatic passage (STIRAP) in a three-level $Λ$-type system for realizing the switch of unequal initial preparations on two ground states. This operation uses single optical pulse sequence accomplishing both bright-STIRAP and dark-STIRAP process, in which the intermediate-level detuning and the pulse delay must be given an optimal adjustment. Besides, owing to the imperfection of double-STIRAP transition, the sensitivity of the switch fidelity with respect to the spontaneous loss from the intermediate state, to the pulse amplitude and to the population difference, are also discussed. This work suggests a simple and experimentally-feasible all-optical approach to switch the superposition quantum state, serving as one-step closer to the goal of coherent manipulation of quantum systems.

physics.atom-ph

Anomalous excitation enhancement with Rydberg-dressed atoms

We develop the research achievement of recent work [M. Gärttner, et.al., Phys. Rev. Letts. 113, 233002 (2014)], in which an anomalous excitation enhancement is observed in a three-level Rydberg-atom ensemble with many-body coherence. In our novel theoretical analysis, this effect is ascribed to the existence of a quasi-dark state as well as its avoided crossings to nearby Rydberg-dressed states. Moreover, we show that with an appropriate control of the optical detuning to the intermediate state, the enhancement can evoke a direct facilitation to atom-light coupling that even breaks through the conventional $\sqrt{N}$ limit of strong-blockaded ensembles. As a consequence, the intensity of the probe laser for intermediate transition can be reduced considerably, increasing the feasibility of experiments with Rydberg-dressed atoms.

physics.atom-ph