Jove
Visualize
联系我们
JoVE
x logofacebook logolinkedin logoyoutube logo
关于 JoVE
概览领导团队博客JoVE 帮助中心
作者
出版流程编辑委员会范围与政策同行评审常见问题投稿
图书馆员
用户评价订阅访问资源图书馆顾问委员会常见问题
研究
JoVE JournalMethods CollectionsJoVE Encyclopedia of Experiments存档
教育
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab Manual教师资源中心教师网站
使用条款与条件
隐私政策
政策

相关概念视频

Visual Agnosia01:12

Visual Agnosia

981
Visual agnosia is a condition characterized by the inability to recognize visually presented objects despite having normal vision. For instance, a person with visual agnosia can describe the shape and color of an object but cannot identify or name it. This impairment does not affect their visual field, acuity, color vision, brightness discrimination, language, or memory. An example of this condition in a social setting is someone at a dinner party asking for "that silver thing with a round...
981
Multi-input and Multi-variable systems01:22

Multi-input and Multi-variable systems

395
Cruise control systems in cars are designed as multi-input systems to maintain a driver's desired speed while compensating for external disturbances such as changes in terrain. The block diagram for a cruise control system typically includes two main inputs: the desired speed set by the driver and any external disturbances, such as the incline of the road. By adjusting the engine throttle, the system maintains the vehicle's speed as close to the desired value as possible.
In the absence of...
395

您也可能阅读

相关文章

通过共同作者、期刊和引用图与本文相关的文章。

排序
Same author

Further evidence for embodied cognition: The link between spontaneous respiratory oscillations and number processing.

PloS one·2026
Same author

Comment on "Integrating large language models and active inference to understand eye movements in reading and dyslexia" by Donnarumma et al.

Physics of life reviews·2026
Same author

Biomarkers of stroke recovery using EEG-based resting-state functional connectivity: a systematic review.

Journal of neuroengineering and rehabilitation·2026
Same author

Obesity accelerates brain ageing: a multimodal imaging study.

Brain communications·2025
Same author

Unveiling the sensorimotor basis of numerical processing: A functional near-infrared spectroscopy (fNIRS) study.

NeuroImage·2025
Same author

Temporal momentum: an online replication and beyond.

Psychological research·2025

相关实验视频

Updated: Jan 18, 2026

Constructing and Visualizing Models using Mime-based Machine-learning Framework
06:19

Constructing and Visualizing Models using Mime-based Machine-learning Framework

Published on: July 22, 2025

2.3K

视觉计数仍然是多式联通生成AI的挑战.

Alberto Testolin1, Kuinan Hou2, Marco Zorzi3,4

  • 1Department of General Psychology and Department of Mathematics, University of Padova, Padova, Italy.

PloS one
|September 12, 2025
PubMed
概括

当前的人工智能模型在视觉计数方面扎,无法准确地计算图像中的对象或生成特定数量的图像. 这表明,仅仅增加人工智能模型大小并不能发展出强大的计数技能.

科学领域:

  • 人工智能的人工智能
  • 认知科学 认知科学
  • 计算机视觉 计算机视觉

背景情况:

  • 人类和动物具有天生的数值能力,而当前的AI系统表现出有限的视觉计数技能.
  • 评估AI的数值感对于推进多式联络基础模型至关重要.

研究的目的:

  • 引入两个认知科学启发的基准任务,用于精确评估人工智能模型中的视觉计数能力.
  • 提供对AI的数感和计数能力的客观测量.

主要方法:

  • 评估了流行的视觉问答 (VQA),图像到文本和文本到图像的人工智能模型.
  • 利用基准任务来评估AI命名对象数量和生成具有特定数量的图像的能力.

主要成果:

  • 先进的AI模型在命名对象计数和生成具有目标数字的图像方面都表现出较低的准确性.
  • 对于超出分类范围的数字,模型性能显著降低,错误通常取决于对象类别.
  • 人工智能模型表现出明显的错误,即使是小数量,与人类行为不同.

结论:

  • 开发直观的视觉数字理解仍然是人工智能面临的重大挑战.
  • 仅仅增加人工智能模型大小不太可能培养系统的计数技能.

更多相关视频

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

1.0K

相关实验视频

Last Updated: Jan 18, 2026

Constructing and Visualizing Models using Mime-based Machine-learning Framework
06:19

Constructing and Visualizing Models using Mime-based Machine-learning Framework

Published on: July 22, 2025

2.3K
Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

1.0K
  • 该研究发布了基准代码,以帮助未来的人工智能计数技能评估.