Searcharxiv⌕ Search

arXiv subjects

Cheng Wu

Publications and source records attributed to Cheng Wu.

33 records · Page 2Linked to original sources

Multi Pseudo Q-learning Based Deterministic Policy Gradient for Tracking Control of Autonomous Underwater Vehicles

This paper investigates trajectory tracking problem for a class of underactuated autonomous underwater vehicles (AUVs) with unknown dynamics and constrained inputs. Different from existing policy gradient methods which employ single actor-critic but cannot realize satisfactory tracking control accuracy and stable learning, our proposed algorithm can achieve high-level tracking control accuracy of AUVs and stable learning by applying a hybrid actors-critics architecture, where multiple actors and critics are trained to learn a deterministic policy and action-value function, respectively. Specifically, for the critics, the expected absolute Bellman error based updating rule is used to choose the worst critic to be updated in each time step. Subsequently, to calculate the loss function with more accurate target value for the chosen critic, Pseudo Q-learning, which uses sub-greedy policy to replace the greedy policy in Q-learning, is developed for continuous action spaces, and Multi Pseudo Q-learning (MPQ) is proposed to reduce the overestimation of action-value function and to stabilize the learning. As for the actors, deterministic policy gradient is applied to update the weights, and the final learned policy is defined as the average of all actors to avoid large but bad updates. Moreover, the stability analysis of the learning is given qualitatively. The effectiveness and generality of the proposed MPQ-based Deterministic Policy Gradient (MPQ-DPG) algorithm are verified by the application on AUV with two different reference trajectories. And the results demonstrate high-level tracking control accuracy and stable learning of MPQ-DPG. Besides, the results also validate that increasing the number of the actors and critics will further improve the performance.

cs.LG↗

Directing liquid crystalline self-organization of rod-like particles through tunable attractive single tips

Dispersions of rodlike colloidal particles exhibit a plethora of liquid crystalline states, including nematic, smectic A, smectic B, and columnar phases. This phase behavior can be explained by presuming the predominance of hard-core volume exclusion between the particles. We show here how the self-organization of rodlike colloids can be controlled by introducing a weak and highly localized directional attractive interaction between one of the ends of the particles. This has been performed by functionalizing the tips of filamentous viruses by means of regioselectively grafting fluorescent dyes onto them, resulting in a hydrophobic patch whose attraction can be tuned by varying the number of bound dye molecules. We show, in agreement with our computer simulations, that increasing the single tip attraction stabilizes the smectic phase at the expense of the nematic phase, leaving all other liquid crystalline phases invariant. For a sufficiently strong tip attraction, the nematic state may be suppressed completely to get a direct isotropic liquid-to-smectic phase transition. Our findings provide insights into the rational design of building blocks for functional structures formed at low densities.

cond-mat.soft↗

Rod-Like Virus-Based Multiarm Colloidal Molecules

We report on the construction of multiarm colloidal molecules by tip-linking filamentous bacteriophages, functionalized either by biological engineering or chemical conjugation. The affinity for streptavidin of a genetically modified vector phage displaying Strep-tags fused to one end of the viral particle is measured by determining the dissociation constant, Kd. In order to improve both the colloidal stability and the efficiency of the self-assembly process, a biotinylation protocol having a chemical yield higher than 90% is presented to regioselectively functionalize the cystein residues located at one end of the bacteriophages. For both viral systems, a theoretical comparison is performed by developing a quantitative model of the self-assembly and interaction of the modified viruses with streptavidin compounds, which accurately accounts for our experimental results. Multiarm colloidal structures of different valencies are then produced by conjugation of these tip-functionalized viruses with streptavidin activated nanoparticles. We succeed to form stable virus based colloidal molecules, whose number of arms, called valency, is solely controlled by tuning the molar excess. Thanks to a fluorescent labeling of the viral arms, the dynamics of such systems is also presented in real time by fluorescence microscopy.

cond-mat.soft↗

Testing Local Realism into the Past without Detection and Locality Loopholes

Inspired by the recent remarkable progress in the experimental test of local realism, we report here such a test that achieves an efficiency greater than (78%)^2 for entangled photon pairs separated by 183 m. Further utilizing the randomness in cosmic photons from pairs of stars on the opposite sides of the sky for the measurement setting choices, we not only close the locality and detection loopholes simultaneously, but also test the null hypothesis against local hidden variable mechanisms for events that took place 11 years ago (13 orders of magnitude longer than previous experiments). After considering the bias in measurement setting choices, we obtain an upper bound on the p value of 7.87 * 10^-4, which clearly indicates the rejection with high confidence of potential local hidden variable models. One may further push the time constraint on local hidden variable mechanisms deep into the cosmic history by taking advantage of the randomness in photon emissions from quasars with large aperture telescopes.

quant-ph↗

Device independent quantum random number generation

Randomness is critical for many information processing applications, including numerical modeling and cryptography. Device-independent quantum random number generation (DIQRNG) based on the loophole free violation of Bell inequality produces unpredictable genuine randomness without any device assumption and is therefore an ultimate goal in the field of quantum information science. However, due to formidable technical challenges, there were very few reported experimental studies of DIQRNG, which were vulnerable to the adversaries. Here we present a fully functional DIQRNG against the most general quantum adversaries. We construct a robust experimental platform that realizes Bell inequality violation with entangled photons with detection and locality loopholes closed simultaneously. This platform enables a continuous recording of a large volume of data sufficient for security analysis against the general quantum side information and without assuming independent and identical distribution. Lastly, by developing a large Toeplitz matrix (137.90 Gb $\times$ 62.469 Mb) hashing technique, we demonstrate that this DIQRNG generates $6.2469\times 10^7$ quantum-certified random bits in 96 hours (or 181 bits/s) with uniformity within $10^{-5}$. We anticipate this DIQRNG may have profound impact on the research of quantum randomness and information-secured applications.

quant-ph↗

Experimental test of measurement dependent local Bell inequality with human free will

A Bell test can rule out local realistic models, and has potential applications in communications and information tasks. For example, a Bell inequality violation can certify the presence of intrinsic randomness in measurement outcomes, which then can be used to generate unconditional randomness. A Bell test requires, however, measurements that are chosen independently of other physical variables in the test, as would be the case if the measurement settings were themselves unconditionally random. This situation seems to create a "bootstrapping problem" that was recently addressed in The BIG Bell Test, a collection of Bell tests and related tests using human setting choices. Here we report in detail our experimental methods and results within the BIG Bell Test. We perform a experimental test of a special type of Bell inequality - the measurement dependent local inequality. With this inequality, even a small amount of measurement independence makes it possible to disprove local realistic models. The experiment uses human-generated random numbers to select the measurement settings in real time, and implements the measurement setting with space-like separation from the distant measurement. The experimental result shows a Bell inequality violation that cannot be explained by local hidden variable models with independence parameter (as defined in [Putz et al. Phys. Rev. Lett. 113, 190402 (2014).] ) l > 0.10 +/- 0.05. This result quantifies the degree to which a hidden variable model would need to constrain human choices, if it is to reproduce the experimental results.

quant-ph↗

Depth Control of Model-Free AUVs via Reinforcement Learning

In this paper, we consider depth control problems of an autonomous underwater vehicle (AUV) for tracking the desired depth trajectories. Due to the unknown dynamical model of the AUV, the problems cannot be solved by most of model-based controllers. To this purpose, we formulate the depth control problems of the AUV as continuous-state, continuous-action Markov decision processes (MDPs) under unknown transition probabilities. Based on deterministic policy gradient (DPG) and neural network approximation, we propose a model-free reinforcement learning (RL) algorithm that learns a state-feedback controller from sampled trajectories of the AUV. To improve the performance of the RL algorithm, we further propose a batch-learning scheme through replaying previous prioritized trajectories. We illustrate with simulations that our model-free method is even comparable to the model-based controllers as LQI and NMPC. Moreover, we validate the effectiveness of the proposed RL algorithm on a seafloor data set sampled from the South China Sea.

cs.RO↗

Random number generation with cosmic photons

Random numbers are indispensable for a variety of applications ranging from testing physics foundation to information encryption. In particular, nonlocality tests provide a strong evidence to our current understanding of nature -- quantum mechanics. All the random number generators (RNG) used for the existing tests are constructed locally, making the test results vulnerable to the freedom-of-choice loophole. We report an experimental realization of RNGs based on the arrival time of cosmic photons. The measurement outcomes (raw data) pass the standard NIST statistical test suite. We present a realistic design to employ these RNGs in a Bell test experiment, which addresses the freedom-of-choice loophole.

quant-ph↗

How to Stop Consensus Algorithms, locally?

This paper studies problems on locally stopping distributed consensus algorithms over networks where each node updates its state by interacting with its neighbors and decides by itself whether certain level of agreement has been achieved among nodes. Since an individual node is unable to access the states of those beyond its neighbors, this problem becomes challenging. In this work, we first define the stopping problem for generic distributed algorithms. Then, a distributed algorithm is explicitly provided for each node to stop consensus updating by exploring the relationship between the so-called local and global consensus. Finally, we show both in theory and simulation that its effectiveness depends both on the network size and the structure.

cs.DC↗

Distributed Random-Fixed Projected Algorithm for Constrained Optimization Over Digraphs

This paper is concerned with a constrained optimization problem over a directed graph (digraph) of nodes, in which the cost function is a sum of local objectives, and each node only knows its local objective and constraints. To collaboratively solve the optimization, most of the existing works require the interaction graph to be balanced or "doubly-stochastic", which is quite restrictive and not necessary as shown in this paper. We focus on an epigraph form of the original optimization to resolve the "unbalanced" problem, and design a novel two-step recursive algorithm with a simple structure. Under strongly connected digraphs, we prove that each node asymptotically converges to some common optimal solution. Finally, simulations are performed to illustrate the effectiveness of the proposed algorithms.

cs.DC↗

Distributed Convex Optimization with Inequality Constraints over Time-varying Unbalanced Digraphs

This paper considers a distributed convex optimization problem with inequality constraints over time-varying unbalanced digraphs, where the cost function is a sum of local objectives, and each node of the graph only knows its local objective and inequality constraints. Although there is a vast literature on distributed optimization, most of them require the graph to be balanced, which is quite restrictive and not necessary. Very recently, the unbalanced problem has been resolved only for either time-invariant graphs or unconstrained optimization. This work addresses the unbalancedness by focusing on an epigraph form of the constrained optimization. A striking feature is that this novel idea can be easily used to study time-varying unbalanced digraphs. Under local communications, a simple iterative algorithm is then designed for each node. We prove that if the graph is uniformly jointly strongly connected, each node asymptotically converges to some common optimal solution.

cs.DC↗

Experimental quantum data locking

Classical correlation can be locked via quantum means--quantum data locking. With a short secret key, one can lock an exponentially large amount of information, in order to make it inaccessible to unauthorized users without the key. Quantum data locking presents a resource-efficient alternative to one-time pad encryption which requires a key no shorter than the message. We report experimental demonstrations of quantum data locking scheme originally proposed by DiVincenzo et al. [Phys. Rev. Lett. 92, 067902 (2004)] and a loss-tolerant scheme developed by Fawzi, Hayde, and Sen [J. ACM. 60, 44 (2013)]. We observe that the unlocked amount of information is larger than the key size in both experiments, exhibiting strong violation of the incremental proportionality property of classical information theory. As an application example, we show the successful transmission of a photo over a lossy channel with quantum data (un)locking and error correction.

quant-ph↗

Likelihood Ratio Based Scheduler for Secure Detection in Cyber Physical Systems

This paper is concerned with a binary detection problem over a non-secure network. To satisfy the communication rate constraint and against possible cyber attacks, which are modeled as deceptive signals injected to the network, a likelihood ratio based (LRB) scheduler is designed in the sensor side to smartly select sensor measurements for transmission. By exploring the scheduler, some sensor measurements are successfully retrieved from the attacked data at the decision center. We show that even under a moderate communication rate constraint of secure networks, an optimal LRB scheduler can achieve a comparable asymptotic detection performance to the standard N-P test using the full set of measurements, and is strictly better than the random scheduler. For non-secure networks, the LRB scheduler can also maintain the detection functionality but suffers graceful performance degradation under different attack intensities. Finally, we perform simulations to validate our theoretical results.

eess.SY↗

3-D Velocity Regulation for Nonholonomic Source Seeking Without Position Measurement

We consider a three-dimensional problem of steering a nonholonomic vehicle to seek an unknown source of a spatially distributed signal field without any position measurement. In the literature, there exists an extremum seeking-based strategy under a constant forward velocity and tunable pitch and yaw velocities. Obviously, the vehicle with a constant forward velocity may exhibit certain overshoots in the seeking process and can not slow down even it approaches the source. To resolve this undesired behavior, this paper proposes a regulation strategy for the forward velocity along with the pitch and yaw velocities. Under such a strategy, the vehicle slows down near the source and stays within a small area as if it comes to a full stop, and controllers for angular velocities become succinct. We prove the local exponential convergence via the averaging technique. Finally, the theoretical results are illustrated with simulations.

cs.RO↗

A Real-Time Detecting Algorithm for Tracking Community Structure of Dynamic Networks

In this paper a simple but efficient real-time detecting algorithm is proposed for tracking community structure of dynamic networks. Community structure is intuitively characterized as divisions of network nodes into subgroups, within which nodes are densely connected while between which they are sparsely connected. To evaluate the quality of community structure of a network, a metric called modularity is proposed and many algorithms are developed on optimizing it. However, most of the modularity based algorithms deal with static networks and cannot be performed frequently, due to their high computing complexity. In order to track the community structure of dynamic networks in a fine-grained way, we propose a modularity based algorithm that is incremental and has very low computing complexity. In our algorithm we adopt a two-step approach. Firstly we apply the algorithm of Blondel et al for detecting static communities to obtain an initial community structure. Then, apply our incremental updating strategies to track the dynamic communities. The performance of our algorithm is measured in terms of the modularity. We test the algorithm on tracking community structure of Enron Email and three other real world datasets. The experimental results show that our algorithm can keep track of community structure in time and outperform the well known CNM algorithm in terms of modularity.

cs.SI↗