Related Experiment Video
Updated: Jul 10, 2025

Tactile Vibrating Toolkit and Driving Simulation Platform for Driving-Related Research
Published on: December 18, 2020
Beam management optimization for V2V communications based on deep reinforcement learning
1Huazhong University of Science and Technology, Wuhan, 430074, China.
Abstract:
Intelligent connected vehicles have garnered significant attention from both academia and industry in recent years as they form the backbone of intelligent transportation and smart cities. Vehicular networks now exchange a range of mixed information types, including safety, sensing, and multimedia, due to advancements in communication and vehicle technology. Accordingly, performance requirements have also evolved, prioritizing higher spectral efficiencies while maintaining low latency and high communication reliability. To address the trade-off between communication spectral efficiency, delay, and reliability, the 3rd Generation Partnership Project (3GPP) recommends the 5G NR FR2 frequency band (24 GHz to 71 GHz) for vehicle-to-everything communications (V2X) in the Release 17 standard. However, wireless transmissions at such high frequencies pose challenges such as high path loss, signal processing complexity, long pre-access phase, unstable network structure, and fluctuating channel conditions. To overcome these issues, this paper proposes a deep reinforcement learning (DRL)-assisted intelligent beam management method for vehicle-to-vehicle (V2V) communication. By utilizing DRL, the optimal control of beam management (i.e., beam alignment and tracking) is achieved, enabling a trade-off among spectral efficiency, delay, and reliability in complex and fluctuating communication scenarios at the 5G NR FR2 band. Simulation results demonstrate the superiority of our method over the 5G standard-based beam management method in communication delay, and the extended Kalman Filter (EKF)-based beam management method in reliability and spectral efficiency.
More Related Videos
05:41A Step-by-Step Implementation of DeepBehavior, Deep Learning Toolbox for Automated Behavior Analysis
Published on: February 6, 2020
09:09Radio Frequency Identification and Motion-sensitive Video Efficiently Automate Recording of Unrewarded Choice Behavior by Bumblebees
Published on: November 15, 2014
Related Concept Videos
Rolling Resistance: Problem Solving
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Reducing Line Loss
With a step-up transformer at the source, the voltage is increased, thereby reducing the current in the transmission lines since power loss...
Reinforcement Schedules
Once a behavior is learned,...
Distributed Loads: Problem Solving