arXiv · 2208.00065
An Actor Critic Method for Free Terminal Time Optimal Control
Abstract
Optimal control problems with free terminal time present many challenges including nonsmooth and discontinuous control laws, irregular value functions, many local optima, and the curse of dimensionality. To overcome these issues, we propose an adaptation of the model-based actor-critic paradigm from the field of Reinforcement Learning via an exponential transformation to learn an approximate feedback control and value function pair. We demonstrate the algorithm's effectiveness on prototypical examples featuring each of the main pathological issues present in problems of this type.
Explore related subjects
Keep this discovery
Evan Burton, Tenavi Nakamura-Zimmerer, Qi Gong, Wei Kang. 2022-07-29. An Actor Critic Method for Free Terminal Time Optimal Control. https://arxiv.org/abs/2208.00065
Cite the original work for its findings. Save a collection to share your selection of sources.