Dynamic multi objective task scheduling in cloud computing using reinforcement learning for energy and cost

Xiaomo Yu1,2, Jie Mi2, Ling Tang3

  • 1Guangxi Colleges and Universities Key Laboratory of Intelligent Logistics Technology, Nanning Normal University, Nanning, 530001, Guangxi, China.

Scientific Reports
|November 26, 2025
PubMed
Summary

This study introduces a Reinforcement Learning-Driven Multi-Objective Task Scheduling (RL-MOTS) framework using Deep Q-Network (DQN) for efficient cloud task allocation. RL-MOTS significantly reduces energy consumption and costs while ensuring Quality of Service (QoS).

Related Concept Videos

Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
438
Distributed Loads: Problem Solving01:21

Distributed Loads: Problem Solving

Beams are structural elements commonly employed in engineering applications requiring different load-carrying capacities. The first step in analyzing a beam under a distributed load is to simplify the problem by dividing the load into smaller regions, which allows one to consider each region separately and calculate the magnitude of the equivalent resultant load acting on each portion of the beam. The magnitude of the equivalent resultant load for each region can be determined by calculating...
1.1K
Maxwell-Boltzmann Distribution: Problem Solving01:20

Maxwell-Boltzmann Distribution: Problem Solving

Individual molecules in a gas move in random directions, but a gas containing numerous molecules has a predictable distribution of molecular speeds, which is known as the Maxwell-Boltzmann distribution, f(v).
This distribution function f(v) is defined by saying that the expected number N (v1,v2) of particles with speeds between v1 and v2 is given by
2.8K
Energy Budgets00:51

Energy Budgets

Organisms must balance energy intake with the energy required for growth, maintenance and reproduction. These trade-offs result in a variety of survivorship and reproductive strategies, including semelparity and iteroparity. Semelparous species, like annual plants, have only one reproductive episode in their lifetimes and consequently have short lifespans. Iteroparous species, by contrast, have many reproductive events during their lifetimes but have relatively few offspring. These two...
10.5K
Maximum Power Flow and Line Loadability01:23

Maximum Power Flow and Line Loadability

The maximum power flow for lossy transmission lines is derived using ABCD parameters in phasor form. These parameters create a matrix relationship between the sending-end and receiving-end voltages and currents, allowing the determination of the receiving-end current. This relationship facilitates calculating the complex power delivered to the receiving end, from which real and reactive power components are derived.
578
Work and Energy for Variable Forces01:10

Work and Energy for Variable Forces

When an object is acted upon by a variable force, the amount of work done and the change in energy of the object can be more complex to calculate compared to when a constant force is applied. Work is the product of force and displacement, while energy is the capacity of a system to do work. When a constant force is applied to an object, the work done can be calculated as the product of the force and the distance moved in the direction of the force. However, when a variable force is applied, the...
5.6K