基于多头自我注意和多代理深度强化学习的认知无线电网络的多用户机会主义频谱访问
Weiwei Bai1, Guoqiang Zheng1, Weibing Xia2
1College of Information Engineering, Henan University of Science and Technology, Luoyang 471023, China.
Sensors (Basel, Switzerland)
|April 12, 2025
概括
本研究介绍了一种先进的方法,用于在认知无线电网络中获得机会性频谱访问. 新的方法通过优化通道选择和功率控制,显著提高了二级用户的整体吞吐量.
科学领域:
- 无线通信无线通信
- 人工智能的人工智能
- 网络优化 网络优化
背景情况:
- 认知无线电网络可以实现动态频谱访问,但多用户机会模式往往遭受低总量吞吐量.
- 有效的资源配置对于最大限度地利用频谱和网络性能至关重要.
研究的目的:
- 提出一种新的多用户机会性频谱接入 (MOSA) 方法,以提高认知无线电网络的总吞吐量.
- 解决动态环境的联合通道选择和功率控制现有方法的局限性.
主要方法:
- 一种多用户的机会性频谱访问方法,集成多头自我注意和多代理深度强化学习 (MADRL).
- 开发一个优化模型,用于联合道选择和功率控制,使用一个集中训练与分散执行 (CTDE) 框架.
- 设计一个多约束的动态比例奖励函数,以指导代理人的决策.
- 在批评网络中整合一个多头自我注意机制,以改善联合行动价值估计.
主要成果:
- 拟议的方法证明了有效的收性质.
- 与基线方法相比,对二级用户的总吞吐量有显著的改善.
- 在机会性频谱接入场景中增强动态性能.
结论:
- 多头自我注意和MADRL的整合为多用户机会性频谱访问提供了有效的解决方案.
- 开发的方法成功优化了频道选择和功率控制,从而提高了网络吞吐量.
- 这种方法为未来的认知无线电网络研究和开发提供了一个有希望的方向.
相关概念视频
Multi-input and Multi-variable systems
547
Cruise control systems in cars are designed as multi-input systems to maintain a driver's desired speed while compensating for external disturbances such as changes in terrain. The block diagram for a cruise control system typically includes two main inputs: the desired speed set by the driver and any external disturbances, such as the incline of the road. By adjusting the engine throttle, the system maintains the vehicle's speed as close to the desired value as possible.
In the absence of...
In the absence of...
547
Associative Learning
2.1K
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
2.1K
Cognitive Learning
1.6K
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
1.6K
Observational Learning
1.5K
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
1.5K


