Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Open and closed-loop control systems01:17

Open and closed-loop control systems

Control systems are foundational elements in automation and engineering. They are broadly categorized into open-loop and closed-loop systems. These classifications hinge on the presence or absence of feedback mechanisms, significantly influencing the system's performance, complexity, and application.
An open-loop control system operates without feedback from the output. It consists of two primary elements: the controller and the controlled process. The controller receives an input signal and...
Multi-input and Multi-variable systems01:22

Multi-input and Multi-variable systems

Cruise control systems in cars are designed as multi-input systems to maintain a driver's desired speed while compensating for external disturbances such as changes in terrain. The block diagram for a cruise control system typically includes two main inputs: the desired speed set by the driver and any external disturbances, such as the incline of the road. By adjusting the engine throttle, the system maintains the vehicle's speed as close to the desired value as possible.
In the absence of...
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving01:29

Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving

Mechanistic models play a crucial role in algorithms for numerical problem-solving, particularly in nonlinear mixed effects modeling (NMEM). These models aim to minimize specific objective functions by evaluating various parameter estimates, leading to the development of systematic algorithms. In some cases, linearization techniques approximate the model using linear equations.
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
Parameters Affecting Nonlinear Elimination: Zero-Order Input, First-Order Absorption and Two-Compartment Model01:13

Parameters Affecting Nonlinear Elimination: Zero-Order Input, First-Order Absorption and Two-Compartment Model

Drugs administered through various routes can lead to nonlinear elimination, resulting in complex pharmacokinetic behaviors crucial to understanding efficacious drug dosing.
When a drug is administered through a constant intravenous infusion and eliminated via nonlinear pharmacokinetics, it follows zero-order input. For example, oral drugs undergo first-order absorption upon administration and are eliminated through nonlinear pharmacokinetics.
In the case of subcutaneously administered drugs,...
Modeling with Differential Equations01:25

Modeling with Differential Equations

Population dynamics can be described mathematically by considering the population size P(t) as a function of time. The rate of change of the population is then represented by the derivative of P(t). A simple assumption is that the rate of growth is proportional to the size of the population itself. This leads to an exponential growth model, where the population increases rapidly without bound. While this is a useful first approximation, it does not reflect realistic long-term...
Lagrange Multipliers: Problem Solving01:30

Lagrange Multipliers: Problem Solving

A silo with a cylindrical base, flat bottom, and hemispherical roof is a common design in agricultural and industrial storage due to its structural efficiency and ease of construction. Optimizing its dimensions to maximize storage capacity for a given amount of material—i.e., a fixed surface area—is a classic problem in applied calculus and engineering design. The key parameters are the radius r of the base and the height h of the cylindrical section.The total volume of the silo is obtained by...

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Anti-Disturbance Intermediate Observer-Based Fault Estimation and Fault-Tolerant Control for Markovian Jump Systems.

IEEE transactions on cybernetics·2026
Same author

A novel dynamic signal Lemma for predefined-time stabilization of high-order nonlinear systems with dynamic uncertainties.

Neural networks : the official journal of the International Neural Network Society·2026
Same author

Observer-Based Fuzzy Secure Control for High-Order MASs Against Communication Delays Under Jointly Connected Switching Topology.

IEEE transactions on cybernetics·2026
Same author

TSFA: A Two-Stage Feature Alignment Method for Unsupervised Open-Set Domain Adaptation in Time-Series Classification.

IEEE transactions on neural networks and learning systems·2026
Same author

Optimal Containment Control for Stochastic Multiagent Systems via Simplified ADP Under Secure Communication.

IEEE transactions on cybernetics·2026
Same author

Constraint-Based Fuzzy Adaptive Security Formation Control for Nonlinear Multiagent Systems Against Deception Attacks.

IEEE transactions on cybernetics·2025

Related Experiment Video

Updated: Jun 18, 2026

The Modular Design and Production of an Intelligent Robot Based on a Closed-Loop Control Strategy
11:53

The Modular Design and Production of an Intelligent Robot Based on a Closed-Loop Control Strategy

Published on: October 14, 2017

Model-Free Algorithms for Cooperative Output Regulation of Discrete-Time Multiagent Systems via Q-Learning Method.

Huaguang Zhang, Tianbiao Wang, Dazhong Ma

    IEEE Transactions on Cybernetics
    |March 27, 2025
    PubMed
    Summary

    A novel model-free Q-learning algorithm enables cooperative output regulation for multiagent systems with unknown parameters. This data-driven approach ensures policy stability and avoids system model requirements.

    More Related Videos

    WheelCon: A Wheel Control-Based Gaming Platform for Studying Human Sensorimotor Control
    08:18

    WheelCon: A Wheel Control-Based Gaming Platform for Studying Human Sensorimotor Control

    Published on: August 15, 2020

    Large Scale Energy Efficient Sensor Network Routing Using a Quantum Processor Unit
    05:30

    Large Scale Energy Efficient Sensor Network Routing Using a Quantum Processor Unit

    Published on: September 8, 2023

    Related Experiment Videos

    Last Updated: Jun 18, 2026

    The Modular Design and Production of an Intelligent Robot Based on a Closed-Loop Control Strategy
    11:53

    The Modular Design and Production of an Intelligent Robot Based on a Closed-Loop Control Strategy

    Published on: October 14, 2017

    WheelCon: A Wheel Control-Based Gaming Platform for Studying Human Sensorimotor Control
    08:18

    WheelCon: A Wheel Control-Based Gaming Platform for Studying Human Sensorimotor Control

    Published on: August 15, 2020

    Large Scale Energy Efficient Sensor Network Routing Using a Quantum Processor Unit
    05:30

    Large Scale Energy Efficient Sensor Network Routing Using a Quantum Processor Unit

    Published on: September 8, 2023

    Area of Science:

    • Control Systems Engineering
    • Artificial Intelligence
    • Robotics

    Background:

    • Cooperative output regulation is crucial for multiagent systems.
    • Unknown system parameters pose a significant challenge in practical applications.
    • Existing methods often require complete system models, limiting their applicability.

    Purpose of the Study:

    • To develop a model-free Q-learning algorithm for cooperative output regulation in discrete-time multiagent systems.
    • To address the challenge of unknown system parameters.
    • To ensure policy stability and convergence in learning algorithms.

    Main Methods:

    • A model-free Q-learning algorithm is proposed, operating independently of system parameters.
    • An immediate cost formulation eliminates the need for solving regulator equations.
    • A data-driven algorithm is introduced to compute initial stable gains for unstable initial policies.
    • The stability of algorithm iterations and a unique Q-function matrix condition are formally derived.

    Main Results:

    • The proposed Q-learning algorithm achieves a streamlined structure for direct optimal policy determination.
    • Formal stability analysis confirms the convergence of each algorithm iteration.
    • The data-driven approach successfully ensures convergence to stability even with unstable initial policies.
    • Demonstration that distributed observers and excitation noise do not introduce bias.

    Conclusions:

    • The model-free Q-learning approach offers an effective solution for cooperative output regulation in multiagent systems with unknown parameters.
    • The developed algorithm ensures stability and convergence, enhancing practical applicability.
    • Simulation examples validate the efficacy and robustness of the proposed method.