Fixed-final-time optimal control of nonlinear systems with terminal constraints

Ali Heydari1, S N Balakrishnan

  • 1Mechanical & Aerospace Engineering Department, Missouri University of Science and Technology, United States.

Summary

A novel model-based reinforcement learning algorithm solves fixed-final-time optimal control problems for nonlinear systems. The neurocontroller demonstrates versatility for various conditions and constraints.

Related Concept Videos

Linear Approximation in Time Domain01:21

Linear Approximation in Time Domain

Nonlinear systems often require sophisticated approaches for accurate modeling and analysis, with state-space representation being particularly effective. This method is especially useful for systems where variables and parameters vary with time or operating conditions, such as in a simple pendulum or a translational mechanical system with nonlinear springs.
For a simple pendulum with a mass evenly distributed along its length and the center of mass located at half the pendulum's length, the...
Feedback control systems01:26

Feedback control systems

Feedback control systems are categorized in various ways based on their design, analysis, and signal types.
Linear feedback systems are theoretical models that simplify analysis and design. These systems operate under the principle that their output is directly proportional to their input within certain ranges. For instance, an amplifier in a control system behaves linearly as long as the input signal remains within a specific range. However, most physical systems exhibit inherent nonlinearity...
Time and frequency -Domain Interpretation of Phase-lead Control01:24

Time and frequency -Domain Interpretation of Phase-lead Control

Phase-lead controllers are commonly used in various control systems to enhance response speed and stability. Adjusting the brightness on a television screen offers a practical example of phase-lead control. When contrast is enhanced, a phase-lead controller is employed. Mathematically, phase-lead control is identified when the first parameter is smaller than the second.
The design of phase-lead control involves the strategic placement of poles and zeros to balance steady-state error and system...
Time-Domain Interpretation of PD Control01:07

Time-Domain Interpretation of PD Control

Proportional-Derivative (PD) control is a widely used control method in various engineering systems to enhance stability and performance. In a system with only proportional control, common issues include high maximum overshoot and oscillation, observed in both the error signal and its rate of change. This behavior can be divided into three distinct phases: initial overshoot, subsequent undershoot, and gradual stabilization.
Consider the example of control of motor torque. Initially, a positive...
Time and frequency -Domain Interpretation of PI Control01:27

Time and frequency -Domain Interpretation of PI Control

Proportional-Integral (PI) controllers are essential in many control systems to improve stability and performance. They are commonly used in everyday devices like thermostats to enhance system damping and reduce steady-state error. When the zero in the controller's transfer function is optimally placed, the system benefits significantly in terms of stability and accuracy.
Acting as a low-pass filter, the PI controller slows the system's response and extends settling times. This requires careful...
Linear Approximation in Frequency Domain01:26

Linear Approximation in Frequency Domain

Linear systems are characterized by two main properties: superposition and homogeneity. Superposition allows the response to multiple inputs to be the sum of the responses to each individual input. Homogeneity ensures that scaling an input by a scalar results in the response being scaled by the same scalar.
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear.