Jove
Visualize
联系我们
JoVE
x logofacebook logolinkedin logoyoutube logo
关于 JoVE
概览领导团队博客JoVE 帮助中心
作者
出版流程编辑委员会范围与政策同行评审常见问题投稿
图书馆员
用户评价订阅访问资源图书馆顾问委员会常见问题
研究
JoVE JournalMethods CollectionsJoVE Encyclopedia of Experiments存档
教育
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab Manual教师资源中心教师网站
使用条款与条件
隐私政策
政策

相关概念视频

Hindsight Biases01:12

Hindsight Biases

3.4K
Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now? 
3.4K
Predicting Reaction Outcomes02:24

Predicting Reaction Outcomes

8.5K
Kinetics describes the rate and path by which a reaction occurs. In contrast, thermodynamics deals with state functions and describes the properties, behavior, and components of a system. It is not concerned with the path taken by the process and cannot address the rate at which a reaction occurs. Although it does provide information about what can happen during a reaction process, it does not describe the detailed steps of what appears on an atomic or a molecular level. On the other hand,...
8.5K
The Representativeness Heuristic02:13

The Representativeness Heuristic

15.8K
The representative heuristic describes a biased way of thinking, in which you unintentionally stereotype someone or something. For example, you may assume that your professors spend their free time reading books and engaging in intellectual conversation, because the idea of them spending their time playing volleyball or visiting an amusement park does not fit in with your stereotypes of professors.
15.8K
Fundamental Attribution Error01:14

Fundamental Attribution Error

12.9K
According to some social psychologists, people tend to overemphasize internal factors as explanations—or attributions—for the behavior of other people. They tend to assume that the behavior of another person is a trait of that person, and to underestimate the power of the situation on the behavior of others. They tend to fail to recognize when the behavior of another is due to situational variables, and thus to the person’s state. This erroneous assumption is...
12.9K
The Availability Heuristic01:08

The Availability Heuristic

6.0K
A heuristic is a general problem-solving framework (Tversky & Kahneman, 1974). You can think of these as mental shortcuts that are used to solve problems. Different types of heuristics are used in different types of situations, and the impulse to use a heuristic occurs when one of five conditions is met (Pratkanis, 1989):
6.0K
Decision Making: Traditional Method01:14

Decision Making: Traditional Method

4.1K
The process of hypothesis testing based on the traditional method includes calculating the critical value, testing the value of the test statistic using the sample data, and interpreting these values.
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
4.1K

您也可能阅读

相关文章

通过共同作者、期刊和引用图与本文相关的文章。

排序
Same author

A multifaceted suite of metrics for comparative myoelectric prosthesis controller research.

PloS one·2024
Same author

Prediction, Knowledge, and Explainability: Examining the Use of General Value Functions in Machine Knowledge.

Frontiers in artificial intelligence·2022
Same author

Preliminary testing of eye gaze interfaces for controlling a haptic system intended to support play in children with physical impairments: Attentive versus explicit interfaces.

Journal of rehabilitation and assistive technologies engineering·2022
Same author

Examining the Use of Temporal-Difference Incremental Delta-Bar-Delta for Real-World Predictive Knowledge Architectures.

Frontiers in robotics and AI·2021

相关实验视频

Updated: Jul 27, 2025

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
05:47

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems

Published on: June 13, 2025

348

一个好的预测是什么? 评估代理人的知识的挑战.

Alex Kearney1,2, Anna J Koop1, Patrick M Pilarski1,2,3,4

  • 1Department of Computing Science, University of Alberta, Edmonton, AB, Canada.

Adaptive behavior
|June 7, 2023
PubMed
概括

评估人工智能 (AI) 模型需要的不仅仅是准确性. 本研究引入了一种新的方法,通过检查内部学习过程来评估AI知识,重点关注更好的预测任务的功能相关性.

科学领域:

  • 人工智能的人工智能
  • 机器学习 机器学习
  • 强化学习是一种强化学习.

背景情况:

  • 通过任务独立的世界模型获得一般知识对人工智能代理至关重要.
  • 当前的评估方法通常依赖于估计器的准确性,这可能不反映真正的有用性.
  • 评估所学知识的实用性仍然是人工智能研究的一个重大挑战.

研究的目的:

  • 突出使用估计器准确度来评估AI知识的局限性.
  • 为持续学习环境中人工智能模型提出一种新的评估方法.
  • 引入一种评估预测知识有用性的方法.

主要方法:

  • 在Minecraft中使用思想实验和经验示例证明了准确性和有用性之间的冲突.
  • 使用通用价值函数 (GVF) 框架进行分析.
  • 建议通过检查内部学习过程来评估AI知识,特别是特征相关性.

主要成果:

  • 一个模型的准确性并不总是与它对代理商的实际有用性相关.
  • 提出了一种基于特征与预测任务相关性的新评估指标.
  • 该研究为评估在线持续学习中的知识实用性提供了一个框架.
关键词:
强化学习是一种强化学习.知识代理知识代理知识代理一般的价值函数一般的价值函数.

更多相关视频

Multimodal Protocol for Assessing Metacognition and Self-Regulation in Adults with Learning Difficulties
12:55

Multimodal Protocol for Assessing Metacognition and Self-Regulation in Adults with Learning Difficulties

Published on: September 27, 2020

8.5K
Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

632

相关实验视频

Last Updated: Jul 27, 2025

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
05:47

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems

Published on: June 13, 2025

348
Multimodal Protocol for Assessing Metacognition and Self-Regulation in Adults with Learning Difficulties
12:55

Multimodal Protocol for Assessing Metacognition and Self-Regulation in Adults with Learning Difficulties

Published on: September 27, 2020

8.5K
Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

632

结论:

  • 重新思考人工智能模型评估超出了简单的准确性是必不可少的.
  • 根据它们的实用性和内部学习过程来评估预测,可以提供更强大的评估.
  • 这项工作开创了对人工智能预测知识的"通过使用评估"的探索.