SearcharxivSearch

arXiv subjects

Muzaffar Qureshi

Publications and source records attributed to Muzaffar Qureshi.

13 recordsLinked to original sources

Safe Output-Feedback Adaptive Optimal Control of Input-Constrained Control-Affine Nonlinear Systems

In this paper, a novel online, safe output-feedback, critic-only, adaptive optimal control framework is developed for safety-critical control of partially observable systems. The developed framework ensures system stability and safety, regardless of the lack of full-state measurements, while learning and implementing a near-optimal controller. The approach leverages linear matrix inequality-based observer design methods to efficiently search for observer gains for effective state estimation. Then, approximate dynamic programming is used to develop an approximate controller that uses simulated experiences to guarantee the safety and stability of the closed-loop system. Safety is enforced by adding a recentered robust Lyapunov-like barrier function to the cost function that effectively enforces safety constraints, even in the presence of state estimation errors. Lyapunov-based stability analysis is used to guarantee uniform ultimate boundedness of the trajectories of the closed-loop system and ensure safety. Simulation studies are performed to demonstrate the effectiveness of the developed method through two real-world safety-critical scenarios, specifically one ensuring that the state trajectories of a given system remain within a given set, and the other ensuring that the system avoids an obstacle.

eess.SY

A Hough transform approach to safety-aware scalar field mapping using Gaussian Processes

This paper presents a framework for mapping unknown scalar fields using a sensor-equipped autonomous robot operating in unsafe environments. The unsafe regions are defined as regions of high-intensity, where the field value exceeds a predefined safety threshold. For safe and efficient mapping of the scalar field, the sensor-equipped robot must avoid high-intensity regions during the measurement process. In this paper, the scalar field is modeled as a sample from a Gaussian process (GP), which enables Bayesian inference and provides closed-form expressions for both the predictive mean and the uncertainty. Concurrently, the spatial structure of the high-intensity regions is estimated in real-time using the Hough transform (HT), leveraging the evolving GP posterior. A safe sampling strategy is then employed to guide the robot towards safe measurement locations, using probabilistic safety guarantees on the evolving GP posterior. The estimated high-intensity regions also facilitate the design of safe motion plans for the robot. The effectiveness of the approach is verified through two numerical simulation studies and an indoor experiment for mapping a light-intensity field using a wheeled mobile robot.

cs.RO

Adaptive Control with Sparse Identification of Nonlinear Dynamics

This paper develops a sparsity-promoting integral concurrent learning (SP-ICL) adaptation law for a linearly parametrized uncertain nonlinear control-affine system. The unknown parameters are learned using ICL with sparsity-promoting $\ell_1$ regularization. The use of $\ell_1$ regularization for sparsity promotion is common in system identification and machine learning; however, unlike existing approaches, this paper develops an online parameter update law that integrates the regularization penalty with ICL via sliding modes. Using the SP-ICL update law, we show via non-smooth Lyapunov analysis that the trajectories of the closed-loop system are ultimately bounded. Simulations verify the effectiveness of the sparsity penalty in the SP-ICL update law on recovering sparse dynamics during trajectory tracking.

math.OC

Decentralized Scalar Field Mapping using Gaussian Process

Decentralized Gaussian process (GP) methods offer a scalable framework for multi-agent scalar-field estimation by replacing a centralized global model with multiple local models maintained by individual agents. A team of agents operates through overlapping domains; neighboring agents generally produce inconsistent distributions over shared regions. This paper investigates whether these inter-agent posterior discrepancies can be systematically exploited to improve team-level predictive performance and answers this question positively through a novel decentralized intersection data-sharing and assimilation protocol. Specifically, each agent constructs neighbor-specific packets from its local GP together with the geometry of the overlap between subdomains and selectively assimilates information received from neighboring agents to improve consistency of its posterior over the shared regions. The proposed architecture preserves locality in both computation and communication, supports decentralized neighbor-to-neighbor data assimilation, and allows local GP models to evolve cooperatively across the network without requiring the exchange full packet exchange or centralized inference.

eess.SY

A Taylor Series Approach to Correct Localization Errors in Robotic Field Mapping using Gaussian Processes

Gaussian Processes (GPs) are powerful non-parametric Bayesian models for regression of scalar fields, formulated under the assumption that measurement locations are perfectly known and the corresponding field measurements have Gaussian noise. However, many real-world scalar field mapping applications rely on sensor-equipped mobile robots to collect field measurements, where imperfect localization introduces state uncertainty. Such discrepancies between the estimated and true measurement locations degrade GP mean and covariance estimates. To address this challenge, we propose a method for updating the GP models when improved estimates become available. Leveraging the differentiability of the kernel function, a second-order correction algorithm is developed using the precomputed Jacobians and Hessians of the GP mean and covariance functions for real-time refinement based on measurement location discrepancy data. Simulation results demonstrate improved prediction accuracy and computational efficiency compared to full model retraining.

cs.RO

Safe Adaptive Feedback Control via Barrier States

This paper presents a safe feedback control framework for nonlinear control-affine systems with parametric uncertainty by leveraging adaptive dynamic programming (ADP) with barrier-state augmentation. The developed ADP-based controller enforces control invariance by optimizing a value function that explicitly penalizes the barrier state, thereby embedding safety directly into the Bellman structure. The near-optimal control policy computed using model-based reinforcement learning is combined with a concurrent learning estimator to identify the unknown parameters and guarantee uniform convergence without requiring persistency of excitation. Using a barrier-state Lyapunov function, we establish boundedness of the barrier dynamics and prove closed-loop stability and safety. Numerical simulations on an optimal obstacle-avoidance problem validate the effectiveness of the developed approach.

math.OC

A Switched Systems Approach to Image-Based Feature Tracking for Autonomous Satellite Inspection

This paper presents an information-based guidance and control architecture for an autonomous deputy spacecraft tasked with inspecting a chief satellite in orbit. The primary objective is for the deputy spacecraft to maximize information gain while tracking features on the chief satellite. The deputy spacecraft needs to respect various constraints such as illumination, field-of-view (FOV), fuel limitations, and avoidance regions. Additionally, the absence of GPS information poses a significant challenge for relative localization within the space environment. To learn the structure of the chief satellite while achieving relative self-localization and maximizing the information gain, this paper integrates a memory regression extension (MRE)-based distance observer with an information-maximizing adaptive controller. The distance observer utilizes feedback from a camera. A switched systems approach is used to determine the minimum dwell time required for a feature to remain within the FOV of the camera to ensure accurate estimation. A k-means clustering algorithm acts as a high-level planner to intermittently generate goal locations that guide the deputy spacecraft toward the nearest cluster of uninspected points on the chief satellite, subject to illumination and FOV constraints. A Lyapunov-based stability analysis is conducted to analyze the developed architecture, and simulation results validate the theoretical results of the paper.

eess.SY

Safe Output-Feedback Adaptive Optimal Control of Affine Nonlinear Systems

In this paper, we develop a safe control synthesis method that integrates state estimation and parameter estimation within an adaptive optimal control (AOC) and control barrier function (CBF)-based control architecture. The developed approach decouples safety objectives from the learning objectives using a CBF-based guarding controller where the CBFs are robustified to account for the lack of full-state measurements. The coupling of this guarding controller with the AOC-based stabilizing control guarantees safety and regulation despite the lack of full state measurement. The paper leverages recent advancements in deep neural network-based adaptive observers to ensure safety in the presence of state estimation errors. Safety and convergence guarantees are provided using a Lyapunov-based analysis, and the effectiveness of the developed controller is demonstrated through simulation under mild excitation conditions.

eess.SY

A Taylor Series Approach to Correction of Input Errors in Gaussian Process Regression

Gaussian Processes (GPs) are widely recognized as powerful non-parametric models for regression and classification. Traditional GP frameworks predominantly operate under the assumption that the inputs are either accurately known or subject to zero-mean noise. However, several real-world applications such as mobile sensors have imperfect localization, leading to inputs with biased errors. These biases can typically be estimated through measurements collected over time using, for example, Kalman filters. To avoid recomputation of the entire GP model when better estimates of the inputs used in the training data become available, we introduce a technique for updating a trained GP model to incorporate updated estimates of the inputs. By leveraging the differentiability of the mean and covariance functions derived from the squared exponential kernel, a second-order correction algorithm is developed to update the trained GP models. Precomputed Jacobians and Hessians of kernels enable real-time refinement of the mean and covariance predictions. The efficacy of the developed approach is demonstrated using two simulation studies, with error analyses revealing improvements in both predictive accuracy and uncertainty quantification.

eess.SY

Improved Dwell-times for Switched Nonlinear Systems using Memory Regression Extension

This paper presents a switched systems approach for extending the dwell-time of an autonomous agent during GPS-denied operation by leveraging memory regressor extension (MRE) techniques. To maintain accurate trajectory tracking despite unknown dynamics and environmental disturbances, the agent periodically acquires access to GPS, allowing it to correct accumulated state estimation errors. The motivation for this work arises from the limitations of existing switched system approaches, where increasing estimation errors during GPS-denied intervals and overly conservative dwell-time conditions restrict the operational efficiency of the agent. By leveraging MRE techniques during GPS-available intervals, the developed method refines the estimates of unknown system parameters, thereby enabling longer and more reliable operation in GPS-denied environments. A Lyapunov-based switched-system stability analysis establishes that improved parameter estimates obtained through concurrent learning allow extended operation in GPS-denied intervals without compromising closed-loop system stability. Simulation results validate the theoretical findings, demonstrating dwell-time extensions and enhanced trajectory tracking performance.

eess.SY

Gaussian Process-Based Scalar Field Estimation in GPS-Denied Environments

This paper presents a methodology for an autonomous agent to map an unknown scalar field in GPS-denied regions. To reduce localization errors, the agent alternates between GPS-enabled and GPS-denied areas while collecting measurements. User-defined error bounds determine the dwell time in each region. A switching trajectory is then designed to ensure field measurements in GPS-denied regions remain within the specified error limits. A Lyapunov-based stability analysis guarantees bounded error trajectories while tracking the desired path. The effectiveness of the proposed methodology is demonstrated through simulations, with an error analysis comparing the GP-predicted scalar field model to the actual field.

eess.SY

Scalar Field Mapping with Adaptive High-Intensity Region Avoidance

This research is motivated by a scenario where a group of UAVs is assigned to map an unknown scalar field, with the imperative of maintaining a safe distance from the sources of the field to evade detection or damage. The location of the sources is unknown a priori, so the UAVs rely on measurements of the field intensity to gauge safety. The UAVs estimate the unknown scalar field using Gaussian process (GP) regression and use the estimate to generate a map of high-intensity regions using Hough transform (HT), updated online based on the field measurements. A convergence analysis shows the boundedness of the error between the actual scalar field and the learned scalar field. The effectiveness of the method is evaluated through simulations, showcasing its ability to accurately learn scalar fields with multiple high-intensity regions while reducing the number of measurements taken inside the high-intensity regions.

eess.SY

An adaptive optimal control approach to monocular depth observability maximization

This paper presents an integral concurrent learning (ICL)-based observer for a monocular camera to accurately estimate the Euclidean distance to features on a stationary object, under the restriction that state information is unavailable. Using distance estimates, an infinite horizon optimal regulation problem is solved, which aims to regulate the camera to a goal location while maximizing feature observability. Lyapunov-based stability analysis is used to guarantee exponential convergence of depth estimates and input-to-state stability of the goal location relative to the camera. The effectiveness of the proposed approach is verified in simulation, and a table illustrating improved observability is provided.

eess.SY