Improving energy autonomy of positive energy districts using multi-agent deep reinforcement learning

Jernej Hribar1,2, Mihael Mohorčič3, Andrej Čampa3,4

  • 1Jozef Stefan Institute, Jamova cesta 39, 1000, Ljubljana, Slovenia. jernej.hribar@ijs.si.

Scientific Reports
|July 30, 2025
PubMed
Abstract

Related Concept Videos

Reinforcement01:23

Reinforcement

Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
343
Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
242
Observational Learning01:12

Observational Learning

Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
314
Distributed Loads: Problem Solving01:21

Distributed Loads: Problem Solving

Beams are structural elements commonly employed in engineering applications requiring different load-carrying capacities. The first step in analyzing a beam under a distributed load is to simplify the problem by dividing the load into smaller regions, which allows one to consider each region separately and calculate the magnitude of the equivalent resultant load acting on each portion of the beam. The magnitude of the equivalent resultant load for each region can be determined by calculating...
738
Avoidance Learning and Learned Helplessness01:14

Avoidance Learning and Learned Helplessness

Avoidance learning and learned helplessness are critical concepts in understanding behavioral responses to negative stimuli.
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
1.9K
Energy to Drive Translocation01:37

Energy to Drive Translocation

Mitochondrial protein import is powered by two distinct energy sources: ATP hydrolysis and electrochemical potential across the inner membrane. Newly synthesized precursors are bound by cytosolic chaperones of the Hsp70 family, which guide them to the import receptors on the mitochondrial surface. Utilizing the energy of ATP hydrolysis, Hsp70 chaperones transfer these precursors to the TOM receptors on the mitochondrial outer membrane.
Generally, polypeptides are unfolded by two distinct...
2.1K