Related Experiment Videos
MaskCtrl: Training mask networks as self-explainable and performant controllers via deep reinforcement learning
Shi Peng1, Si Liu2, Dapeng Zhi1
1Shanghai Key Laboratory of Trustworthy Computing, East China Normal University, Shanghai, China.
Abstract:
Explainability is critical for ensuring safe and accountable control in deep reinforcement learning (DRL). Most explanation approaches train explanation models (e.g., mask networks) offline from system trajectories, rendering the training disconnected from environmental feedback on perturbed states. This disconnection frequently leads to feature importance misalignment, i.e., critical features are overlooked, and irrelevant ones are incorrectly highlighted. To address this problem, we propose MaskCtrl, a novel DRL framework for training mask networks as self-explainable controllers. A trained self-explainable controller achieves two key capabilities: (1) decision-making performance comparable to the original neural controller from which it is derived, and (2) accurate identification of state features critical to the decision-making. These dual strengths stem from our DRL-based training mechanism, which leverages system rewards as environmental feedback to correct feature misalignment during training. Empirical results demonstrate that our self-explainable controllers outperform offline-trained explanation models by up to 50.58% in critical feature identification fidelity. Specifically, this enhanced feature alignment between explanations and decision-making enables 25.2% greater effectiveness in adversarial attacks and a more pronounced robustness gain, with only an 8.7% drop in reward against significant adversarial perturbations.
Related Concept Videos
Masking and Demasking Agents
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on the metal...
Neural Control of Respiration
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
Observational Learning
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Modeling in Therapy
Participant Modeling
Participant modeling involves therapists demonstrating calm and effective behaviors in situations...