Jove
Visualize
联系我们
JoVE
x logofacebook logolinkedin logoyoutube logo
关于 JoVE
概览领导团队博客JoVE 帮助中心
作者
出版流程编辑委员会范围与政策同行评审常见问题投稿
图书馆员
用户评价订阅访问资源图书馆顾问委员会常见问题
研究
JoVE JournalMethods CollectionsJoVE Encyclopedia of Experiments存档
教育
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab Manual教师资源中心教师网站
使用条款与条件
隐私政策
政策

相关概念视频

Reinforcement01:23

Reinforcement

353
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
353
Reinforcement Schedules01:24

Reinforcement Schedules

243
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
243
Observational Learning01:12

Observational Learning

321
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
321
Primary and Secondary Reinforcers01:23

Primary and Secondary Reinforcers

422
In psychology, reinforcement is a key concept in behavior modification. B.F. Skinner demonstrated this with his experiments involving rats in what is known as a Skinner box. The rats learned to press a lever to receive food, a primary reinforcer that fulfilled their innate need for nourishment.
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
422
Avoidance Learning and Learned Helplessness01:14

Avoidance Learning and Learned Helplessness

1.9K
Avoidance learning and learned helplessness are critical concepts in understanding behavioral responses to negative stimuli.
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
1.9K
Associative Learning01:27

Associative Learning

605
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
605

您也可能阅读

相关文章

通过共同作者、期刊和引用图与本文相关的文章。

排序
Same author

T-DNA Mutagenesis Reveals FpPer1 as a Dual-Function Regulator of Virulence and Fungicide Resistance in <i>Fusarium pseudograminearum</i>.

Journal of fungi (Basel, Switzerland)·2025
Same author

Rapid Switching of Reagent Ions in Dopant-Assisted Positive Photoionization Ion Mobility Spectrometry for Highly Selective and Sensitive Ammonia Detection in Atmospheric Environments.

Analytical chemistry·2025
Same author

Triglyceride-Glucose Index: a novel prognostic predictor for postoperative cerebral infarction in off-pump coronary artery bypass grafting - insights from a nationwide multicentre study.

Open heart·2025
Same author

LOXL1-AS1 suppresses ferroptosis in cervical cancer through m6A-dependent regulation of TFRC.

Journal of molecular histology·2025
Same author

Harnessing glycolysis- and cholesterol synthesis-related genes for prognostic modeling in lung adenocarcinoma.

Computer methods in biomechanics and biomedical engineering·2025
Same author

BMI as a Mediator in the Relationship Between Dietary Trace Elements and Type 2 Diabetes Mellitus: Findings from a Rural Cohort.

Nutrients·2025

相关实验视频

Updated: Sep 18, 2025

The HoneyComb Paradigm for Research on Collective Human Behavior
06:48

The HoneyComb Paradigm for Research on Collective Human Behavior

Published on: January 19, 2019

9.5K

游戏中的多代理强化学习:研究和应用.

Haiyang Li1, Ping Yang1, Weidong Liu1

  • 1High-Tech Institute of Xi'an, Xi'an 710038, China.

Biomimetics (Basel, Switzerland)
|June 25, 2025
PubMed
概括

这项研究结合了多代理强化学习 (MARL) 和游戏理论,灵感来自生物系统. 它增强了集体智能,以便在智能城市等动态环境中进行复杂的决策.

科学领域:

  • 人工智能的人工智能
  • 计算游戏理论 计算游戏理论
  • 生物启发的计算 生物启发的计算

背景情况:

  • 生物系统表现出自我组织的智能.
  • 在复杂的系统中,将游戏理论理性和多代理适应性联系起来至关重要.
  • 现有的框架缺乏将生物灵感原则与集体决策的先进人工智能相结合.

研究的目的:

  • 系统地审查多代理强化学习 (MARL) 和游戏理论的融合.
  • 阐明这种集成范式在动态开放环境中集体智能决策的潜力.
  • 确定技术突破和绘制开发路径,以加强多代理系统.

主要方法:

  • 利用随机游戏和广泛形式的游戏理论框架.
  • 建立基于价值函数优化,政策梯度学习和在线搜索规划的方法分类.
  • 结合生物灵感优化方法,包括进化计算和基于人口的学习.

主要成果:

  • 开发了一种方法学分类,阐明了MARL和游戏理论中的算法进步.
  • 识别了用于智慧城市场景的MARL应用中的技术突破 (例如,智能交通,无人机调度).
  • 强调了生物灵感机制对动态战略生成和勘探效率的有效性.
关键词:
进化计算是一种进化计算.游戏理论的游戏理论.多种代理强化学习的多种代理强化学习随机游戏 随机游戏 随机游戏

更多相关视频

Combining Computer Game-Based Behavioural Experiments With High-Density EEG and Infrared Gaze Tracking
13:40

Combining Computer Game-Based Behavioural Experiments With High-Density EEG and Infrared Gaze Tracking

Published on: December 16, 2010

16.8K
Investigating Motor Skill Learning Processes with a Robotic Manipulandum
07:52

Investigating Motor Skill Learning Processes with a Robotic Manipulandum

Published on: February 12, 2017

8.8K

相关实验视频

Last Updated: Sep 18, 2025

The HoneyComb Paradigm for Research on Collective Human Behavior
06:48

The HoneyComb Paradigm for Research on Collective Human Behavior

Published on: January 19, 2019

9.5K
Combining Computer Game-Based Behavioural Experiments With High-Density EEG and Infrared Gaze Tracking
13:40

Combining Computer Game-Based Behavioural Experiments With High-Density EEG and Infrared Gaze Tracking

Published on: December 16, 2010

16.8K
Investigating Motor Skill Learning Processes with a Robotic Manipulandum
07:52

Investigating Motor Skill Learning Processes with a Robotic Manipulandum

Published on: February 12, 2017

8.8K

结论:

  • 马尔和游戏理论的整合,通过生物启发的计算来增强,为集体智能提供了巨大的潜力.
  • 这种跨学科的方法为开发能够在复杂,动态环境中进行最佳决策的先进多代理系统提供了路线图.
  • 调查结果揭示了集团决策的核心原则,并绘制了技术发展路径.