Related Experiment Video
Updated: Jul 19, 2025

Large Scale Energy Efficient Sensor Network Routing Using a Quantum Processor Unit
Published on: September 8, 2023
Joint Optimization of Bandwidth and Power Allocation in Uplink Systems with Deep Reinforcement Learning
Chongli Zhang1, Tiejun Lv1, Pingmu Huang2
1School of Information and Communication Engineering, Beijing University of Posts and Telecommunications (BUPT), Beijing 100876, China.
This study introduces a joint-priority-based reinforcement learning (JPRL) approach to optimize wireless resource allocation, significantly improving system throughput and reducing interference in multi-user systems.
Area of Science:
- Wireless communication systems
- Artificial intelligence in telecommunications
- Resource management in networks
Background:
- Increasing numbers of users in multi-cell systems cause explosive interference, degrading communication quality.
- Inter-cell interference is a major challenge in optimizing wireless resource utilization.
- Existing methods struggle to balance throughput maximization with quality of service constraints.
Purpose of the Study:
- To propose a novel approach for joint optimization of bandwidth and transmit power allocation.
- To enhance system throughput while suppressing co-channel interference.
- To guarantee quality of service (QoS) constraints in wireless networks.
Main Methods:
- Developed a joint-priority-based reinforcement learning (JPRL) approach.
- Decoupled the joint problem into bandwidth assignment and power allocation sub-problems.
- Utilized multi-agent double deep Q network (MADDQN) for bandwidth allocation and prioritized multi-agent deep deterministic policy gradient (P-MADDPG) for power allocation.
Main Results:
- The JPRL method demonstrated accelerated model training.
- Achieved superior system throughput compared to alternative methods.
- Average throughput was 10.4-15.5% higher than homogeneous-learning benchmarks and 17.3% higher than genetic algorithms.
Conclusions:
- The proposed JPRL approach effectively optimizes wireless resource utilization.
- JPRL significantly improves system throughput and mitigates interference.
- This method offers a promising solution for future wireless communication systems.
More Related Videos
05:28Author Spotlight: Enhancing Upper Limb Rehabilitation in Stroke Patients Through Advanced Robotic and Neuromodulation Technologies
Published on: October 11, 2024
11:19Dorsal Column Steerability with Dual Parallel Leads using Dedicated Power Sources: A Computational Model
Published on: February 10, 2011
Related Concept Videos
Maximum Power Transfer
By substituting the entire circuit with...
Maximum Power Flow and Line Loadability
Reinforcement Schedules
Once a behavior is learned,...
Reducing Line Loss
With a step-up transformer at the source, the voltage is increased, thereby reducing the current in the transmission lines since power loss...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning