非線形パラメータ変動システムのための強化学習を用いた動的ニューラルネットワークベース制御手法とモーフィング航空機への応用
Abstract:
This article proposes a dynamic neural network (DNN)-based control method to realize the optimal control of nonlinear parameter-varying (NPV) systems. Specifically, a DNN-based control policy (DNN-CP) composed of static shared layers and a parameter-related dynamic layer is constructed to improve the generalization and adaptability. An extreme learning machine (ELM)-based weight prediction model is established to fit the relationship between the dynamic weights and the system parameters. The shared layers are updated by solving the constrained multiobjective problem to reduce performance conflicts among different systems, and the weight prediction model is tuned by maximizing parameter-related objectives to achieve optimal control of each system. To improve data efficiency and adaptability, a supervised learning-based pretraining and reinforcement learning (RL)-based fine-tuning algorithm is developed. Finally, the control performance of the DNN-CP is verified on morphing aircraft. We demonstrate that the designed DNN-CP and training algorithm can achieve generalization capabilities, and DNN-CP can be immediately generalized to any system within the parameter space without sample collection or fine-tuning. Compared with other methods, DNN-CP has better control performance on the system with continuously varying parameters.
関連する概念動画
Feedback control systems
Linear feedback systems are theoretical models that simplify analysis and design. These systems operate under the principle that their output is directly proportional to their input within certain ranges. For instance, an amplifier in a control system behaves linearly as long as the input signal remains within a specific range. However, most physical systems exhibit inherent nonlinearity...
Multi-input and Multi-variable systems
In the absence of...
PD Controller: Design
Designing a continuous-data controller requires selecting and linking components like adders and integrators, which are fundamental in Proportional,...
Neural Control of Respiration
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
Linear Approximation in Time Domain
For a simple pendulum with a mass evenly distributed along its length and the center of mass located at half the pendulum's length,...
Time-Domain Interpretation of PD Control
Consider the example of control of motor torque. Initially, a positive...


