数据优先级意识资源分配在互联网上使用多代理深度强化学习的车辆的数据优先级.
Cong Wang1, Yingshan Guan2, Sancheng Peng3
1School of Computer and Communication Engineering, Northeastern University at Qinhuangdao, Qinhuangdao, 066004, Hebei, China; Hebei Key Laboratory of Marine Perception Network and Data Processing, Qinhuangdao, 066004, Hebei, China.
概括
智能运输系统面临资源限制. 新的NL-MAPPO框架优化了车辆互联网的频谱和功率,提高了通信效率并减少了延迟.
科学领域:
- 计算机科学 计算机科学
- 电气工程 电气工程
- 运输工程 运输工程
背景情况:
- 智能运输系统 (ITS) 和汽车互联网 (IoV) 面临频谱资源限制和实时通信需求.
- 在 IoV 中有效的资源配置是具有挑战性的,特别是考虑到数据优先级和动态车辆环境.
- 优化传输功率和频谱分配对于最大限度地提高IoV性能至关重要.
研究的目的:
- 设计一种基于时间序列的多代理深度强化学习框架 (NL-MAPPO),用于 IoV 中的资源配置.
- 解决 IoV 通信中动态车辆特征和多样化数据优先级的挑战.
- 为了尽量减少传输延迟和能源消耗,同时最大限度地提高车辆对车辆 (V2V) 链路容量.
主要方法:
- 制定了资源分配问题作为一个多代理马尔科夫决策过程.
- 开发了一个多代理资源分配算法,利用共享关键机制进行全球道信息共享.
- 引入了一个基于时间序列的通道信息提取机制,以捕获时间动态.
主要成果:
- 拟议的NL-MAPPO框架有效优化了频谱分配和传输功率.
- 在最大限度地减少传输延迟和能源消耗方面取得了显著的改善.
- 实现了车辆对车辆 (V2V) 链路总容量的最大化.
结论:
- 与现有的方法相比,NL-MAPPO框架为IOV中的资源配置提供了更好的解决方案.
- 该方法有效地平衡了性能指标,如延迟,能源消耗和链路容量.
- 基于时间序列的多代理深度强化学习是未来IoV研究的一个有希望的方向.
相关概念视频
Distributed Loads: Problem Solving
745
Beams are structural elements commonly employed in engineering applications requiring different load-carrying capacities. The first step in analyzing a beam under a distributed load is to simplify the problem by dividing the load into smaller regions, which allows one to consider each region separately and calculate the magnitude of the equivalent resultant load acting on each portion of the beam. The magnitude of the equivalent resultant load for each region can be determined by calculating...
745
Reinforcement
354
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
354
Reinforcement Schedules
243
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
Once a behavior is learned,...
243
Rolling Resistance: Problem Solving
466
Rolling resistance, also known as rolling friction, is the force that resists the motion of a rolling object, such as a wheel, tire, or ball, when it moves over a surface. It is caused by the deformation of the object and the surface in contact with each other, as well as other factors like internal friction, hysteresis, and energy losses within the materials. Rolling resistance opposes the object's motion, requiring additional energy to overcome it and maintain movement. In practical...
466
Observational Learning
321
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
321
Multi-input and Multi-variable systems
152
Cruise control systems in cars are designed as multi-input systems to maintain a driver's desired speed while compensating for external disturbances such as changes in terrain. The block diagram for a cruise control system typically includes two main inputs: the desired speed set by the driver and any external disturbances, such as the incline of the road. By adjusting the engine throttle, the system maintains the vehicle's speed as close to the desired value as possible.
In the absence...
In the absence...
152


