Related Experiment Videos
An optimized resource allocation in cloud using prediction enabled reinforcement learning
S Kayalvili1, R Senthilkumar2, S Yasotha3
1Kongu Engineering College, Erode, Tamil Nadu, India. skayalvili10@gmail.com.
Abstract:
Due to its many applications, cloud computing has gained popularity in recent years. It is simple and fast to access shared resources at any time from any location. Cloud-based package facilities need adaptive resource allocation (RA) to provide Quality-of-Service (QoS) while lowering resource prices owing to workloads and service demands that change over time. As a result of the constantly shifting system states, resource allocation presents enormous challenges. The old methods often require specialist knowledge, which may result in poor adaptability. Additionally, it aims for environments with set workloads; hence, it cannot be used successfully in real-world contexts with fluctuating workloads. This research therefore proposes a Prediction-enabled feedback system to solve these significant problems with the reinforcement learning-based RA (PCRA) framework. Firstly, this research creates a more accurate Q-value prediction to forecast management value processes at various scheme conditions, using Q-values as the basis. For accurate Q-value prediction, the model makes use of several prediction learners using the Q-learning method. Also, an improved optimization-based algorithm is utilized to discover impartial resource allocations called the Feature Selection Whale Optimization Algorithm (FSWOA). Simulations based on practical scenarios using CloudStack and RUBiS benchmarks demonstrate the effectiveness of PCRA for real-time RA. Simulations demonstrate that the PCRA framework achieves a 94.7% Q-value prediction accuracy and reduces SLA violations and resource cost by 17.4% compared to traditional round-robin scheduling.
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Reinforcement Schedules
Once a behavior is learned,...
Observational Learning
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Predicting Reaction Outcomes
Cognitive Learning
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...