Video Experimental Relacionado
Updated: Feb 13, 2026

14:13
A Training Program Using an Agility Ladder for Community-Dwelling Older Adults
Published on: March 7, 2020
11.5K
SR-GRAT: Aprendizaje de refuerzo guiado por respuesta simétrica con orientación adaptativa para el control ágil de
IEEE transactions on cybernetics
|February 11, 2026
Resumen
Este estudio introduce un nuevo marco de aprendizaje por refuerzo (RL) para el control del piloto automático de aeronaves. Utiliza curvas de aprendizaje adaptativas y objetivos dinámicos para mejorar la estabilidad y acelerar el aprendizaje para misiones aéreas complejas.
Área de la Ciencia:
- Ingeniería Aeroespacial Ingeniería Aeroespacial.
- La inteligencia artificial es inteligencia artificial.
- Sistemas de control de los sistemas de control.
Sus antecedentes:
- Los métodos tradicionales de control de piloto automático a menudo usan progresos de aprendizaje estático.
- Esto puede conducir a un rendimiento subóptimo, con agentes abrumados o estancados.
- Existe la necesidad de estrategias de control adaptativas en entornos aéreos dinámicos.
Objetivo del estudio:
- Presentar un nuevo marco curricular para el control de piloto automático de aeronaves de ala fija.
- Mejorar la eficiencia y la estabilidad del aprendizaje utilizando mecanismos adaptativos y dinámicos.
- Demostrar la efectividad del marco en escenarios complejos de misiones aéreas.
Principales métodos:
- Se introdujo el aprendizaje por refuerzo guiado por respuesta simétrica (RL).
- Implementó una curva de aprendizaje bidireccional adaptativa que se ajusta en función de las tendencias de recompensa.
- Utilizó un mecanismo dinámico de generación de objetivos dentro de los episodios.
Principales resultados:
- Se logró una convergencia más rápida y se redujo el exceso en las tareas de seguimiento de trayectoria.
- Demostró un rendimiento de seguimiento más preciso, incluso bajo turbulencias.
- Demostró superioridad sobre los métodos de referencia en la navegación de puntos de ruta y la búsqueda dinámica.
Conclusiones:
- El marco RL propuesto ofrece una mayor estabilidad y una convergencia más rápida para el control de aeronaves.
- La generación dinámica de objetivos mejora la eficiencia del aprendizaje sin la configuración manual de recompensas.
- El controlador es robusto y aplicable a misiones complejas como el combate aéreo autónomo.
Más Videos Relacionados
Videos de Conceptos Relacionados
Cells of the Adaptive Immune Response
9.1K
The T and B lymphocytes of the adaptive immune system develop from common lymphoid progenitor cells in the bone marrow. These progenitors give rise to precursors that eventually develop into both T and B lymphocytes. As these precursors mature, they gain the ability to detect and respond to foreign antigens in the body, a process known as immunocompetence. Additionally, these precursors acquire self-tolerance, a process that ensures they do not react to self-antigens. This intricate system...
9.1K
Conservation of Mass in Fixed, Nondeforming Control Volume
1.6K
The principle of conservation of mass is fundamental in fluid dynamics and is crucial for analyzing flow within fixed control volumes, such as pipes or ducts. This principle states that the total mass within a control volume remains constant unless altered by the inflow or outflow of mass through the control surfaces. This results in a vital relationship for steady, incompressible flow where the mass entering a system equals the mass leaving it.
In the case of a sewer pipe, which can be modeled...
In the case of a sewer pipe, which can be modeled...
1.6K
Symmetric Member in Bending
618
In the study of the mechanics of materials, analyzing the behavior of prismatic members under opposing couples is crucial for understanding internal stress distributions, which are essential for structural design. When subjected to couples, a prismatic member experiences internal forces that maintain equilibrium. A couple, characterized by two equal and opposite forces, creates a moment but no resultant force. The internal forces at any section cut of the member must balance these external...
618
Target Cell Response to Hormones
5.8K
Hormones intricately bind to receptors on the surface or within target cells, initiating a cascade of cellular responses.
Notably, the cellular response can be regulated by altering the number of receptors expressed in the cell. For example, prolonged exposure to elevated hormone levels results in a gradual decline or down-regulation in the number of receptors for that specific hormone on the cell surface. Conversely, in response to low hormone levels, cells may use up-regulation, producing an...
Notably, the cellular response can be regulated by altering the number of receptors expressed in the cell. For example, prolonged exposure to elevated hormone levels results in a gradual decline or down-regulation in the number of receptors for that specific hormone on the cell surface. Conversely, in response to low hormone levels, cells may use up-regulation, producing an...
5.8K
Fixed Action Patterns
17.7K
A fixed action pattern (FAP) is a specific, hard-wired sequence of behaviors that occurs in response to an external stimulus, called a sign stimulus. The behavior is “fixed” because it is essentially unchangeable—proceeding similarly across individuals of a species every time it occurs.
17.7K
Reinforcement
952
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
952

