SearcharxivSearch

arXiv subjects

Maurice Poot

Publications and source records attributed to Maurice Poot.

7 recordsLinked to original sources

Feedforward Control in the Presence of Input Nonlinearities: A Learning-based Approach

Advanced feedforward control methods enable mechatronic systems to perform varying motion tasks with extreme accuracy and throughput. The aim of this paper is to develop a data-driven feedforward controller that addresses input nonlinearities, which are common in typical applications such as semiconductor back-end equipment. The developed method consists of parametric inverse-model feedforward that is optimized for tracking error reduction by exploiting ideas from iterative learning control. Results on a simulated set-up indicate improved performance over existing identification methods for systems with nonlinearities at the input.

eess.SY

Cross-Coupled Iterative Learning Control for Complex Systems: A Monotonically Convergent and Computationally Efficient Approach

Cross-coupled iterative learning control (ILC) can achieve high performance for manufacturing applications in which tracking a contour is essential for the quality of a product. The aim of this paper is to develop a framework for norm-optimal cross-coupled ILC that enables the use of exact contour errors that are calculated offline, and iteration- and time-varying weights. Conditions for the monotonic convergence of this iteration-varying ILC algorithm are developed. In addition, a resource-efficient implementation is proposed in which the ILC update law is reframed as a linear quadratic tracking problem, reducing the computational load significantly. The approach is illustrated on a simulation example.

eess.SY

Is Vanilla Policy Gradient Overlooked? Analyzing Deep Reinforcement Learning for Hanabi

In pursuit of enhanced multi-agent collaboration, we analyze several on-policy deep reinforcement learning algorithms in the recently published Hanabi benchmark. Our research suggests a perhaps counter-intuitive finding, where Proximal Policy Optimization (PPO) is outperformed by Vanilla Policy Gradient over multiple random seeds in a simplified environment of the multi-agent cooperative card game. In our analysis of this behavior we look into Hanabi-specific metrics and hypothesize a reason for PPO's plateau. In addition, we provide proofs for the maximum length of a perfect game (71 turns) and any game (89 turns). Our code can be found at: https://github.com/bramgrooten/DeepRL-for-Hanabi

cs.LG

Gaussian Process Position-Dependent Feedforward: With Application to a Wire Bonder

Mechatronic systems have increasingly stringent performance requirements for motion control, leading to a situation where many factors, such as position-dependency, cannot be neglected in feedforward control. The aim of this paper is to compensate for position-dependent effects by modeling feedforward parameters as a function of position. A framework to model and identify feedforward parameters as a continuous function of position is developed by combining Gaussian processes and feedforward parameter learning techniques. The framework results in a fully data-driven approach, which can be readily implemented for industrial control applications. The framework is experimentally validated and shows a significant performance increase on a commercial wire bonder.

eess.SY

Learning nonlinear feedforward: a Gaussian Process Approach Applied to a Printer with Friction

Feedforward control is essential to achieving good tracking performance in positioning systems. The aim of this paper is to develop an identification strategy for inverse models of systems with nonlinear dynamics of unknown structure using input-output data, which directly delivers feedforward signals for a-priori unknown tasks. To this end, inverse systems are regarded as noncausal nonlinear finite impulse response (NFIR) systems and modeled as a Gaussian Process with a stationary kernel function that imposes properties such as smoothness and periodicity. The approach is validated experimentally on a consumer printer with friction and shown to lead to improved tracking performance with respect to linear feedforward.

eess.SY

Position-Dependent Snap Feedforward: A Gaussian Process Framework

Mechatronic systems have increasingly high performance requirements for motion control. The low-frequency contribution of the flexible dynamics, i.e. the compliance, should be compensated for by means of snap feedforward to achieve high accuracy. Position-dependent compliance, which often occurs in motion systems, requires the snap feedforward parameter to be modeled as a function of position. Position-dependent compliance is compensated for by using a Gaussian process to model the snap feedforward parameter as a continuous function of position. A simulation of a flexible beam shows that a significant performance increase is achieved when using the Gaussian process snap feedforward parameter to compensate for position-dependent compliance.

eess.SY

On the Role of Models in Learning Control: Actor-Critic Iterative Learning Control

Learning from data of past tasks can substantially improve the accuracy of mechatronic systems. Often, for fast and safe learning a model of the system is required. The aim of this paper is to develop a model-free approach for fast and safe learning for mechatronic systems. The developed actor-critic iterative learning control (ACILC) framework uses a feedforward parameterization with basis functions. These basis functions encode implicit model knowledge and the actor-critic algorithm learns the feedforward parameters without explicitly using a model. Experimental results on a printer setup demonstrate that the developed ACILC framework is capable of achieving the same feedforward signal as preexisting model-based methods without using explicit model knowledge.

eess.SY