Related Experiment Video
Updated: Sep 11, 2025

10:15
Integration of 5G Experimentation Infrastructures into a Multi-Site NFV Ecosystem
Published on: February 3, 2021
3.9K
Intelligent Priority-Aware Spectrum Access in 5G Vehicular IoT: A Reinforcement Learning Approach
Adeel Iqbal1, Tahir Khurshaid2, Yazdan Ahmad Qadri1
1School of Computer Science and Engineering, Yeungnam University, Gyeongsan-si 38541, Republic of Korea.
Sensors (Basel, Switzerland)
|August 14, 2025
Summary
This study introduces a novel reinforcement learning framework for intelligent spectrum management in vehicular networks. It balances performance metrics like throughput, delay, and fairness for diverse traffic needs.
Area of Science:
- Wireless communication networks
- Vehicular Internet of Things (V-IoT)
- Next-generation cellular networks
Background:
- Efficient spectrum access is critical for Vehicular Internet of Things (V-IoT) systems.
- Next-generation cellular networks require dynamic Quality of Service (QoS) management.
- Existing spectrum management solutions struggle with the dynamic nature of vehicular environments.
Purpose of the Study:
- To propose a novel reinforcement learning (RL)-based priority-aware spectrum management (RL-PASM) framework.
- To dynamically allocate spectrum resources for high-priority (HP), low-priority (LP), and best-effort (BE) traffic classes.
- To evaluate the performance of different RL algorithms in a centralized RSU-based control system.
Main Methods:
- Modeling the spectrum management environment as a discrete-time Markov Decision Process (MDP).
- Utilizing a context-sensitive reward function for fairness-preserving decisions (access, preemption, coexistence, hand-off).
- Comparing four RL algorithms: Q-Learning, Double Q-Learning, Deep Q-Network (DQN), and Actor-Critic (AC).
Main Results:
- RL-PASM effectively balances throughput, latency, fairness, and energy efficiency.
- DQN achieved the highest average throughput, while Q-Learning offered the lowest average delay and highest energy efficiency.
- Double Q-Learning and Actor-Critic maintained high fairness and low interruption probability.
Conclusions:
- RL-PASM provides a robust and adaptable solution for intelligent, priority-aware spectrum access in vehicular networks.
- The framework is suitable for scalable and resource-constrained deployments, particularly edge-constrained vehicular environments.
- The choice of RL algorithm allows for tailored optimization based on specific network priorities (e.g., throughput vs. energy efficiency).
Keywords:
5GInternet of Thingspriority-aware spectrum managementreinforcement learningresource allocationspectrum accessMore Related Videos
Related Concept Videos
Reinforcement
341
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
341
Reinforcement Schedules
242
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
Once a behavior is learned,...
242
Cognitive Learning
519
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
519
Observational Learning
312
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
312
Rolling Resistance: Problem Solving
451
Rolling resistance, also known as rolling friction, is the force that resists the motion of a rolling object, such as a wheel, tire, or ball, when it moves over a surface. It is caused by the deformation of the object and the surface in contact with each other, as well as other factors like internal friction, hysteresis, and energy losses within the materials. Rolling resistance opposes the object's motion, requiring additional energy to overcome it and maintain movement. In practical...
451
Associative Learning
575
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
575

