视觉模型需要一些行动行动.
Constantin Rothkopf1,2,3,4, Frank Bremmer3,4,5, Katja Fiehler3,4,6
1Centre for Cognitive Science, Technical University of Darmstadt, Darmstadt, Germany constantin.rothkopf@cogsci.tu-darmstadt.de.
The Behavioral and brain sciences
|December 6, 2023
概括
这项研究批评将大脑数据与深度神经网络进行对象识别的比较. 它强调,目前的研究忽视了行动和互动在感知中的关键作用.
科学领域:
- 认知神经科学 认知神经科学
- 计算机视觉 计算机视觉
- 神经科学是一个神经科学.
背景情况:
- 研究经常将腹部流数据与深度神经网络进行对象识别的比较.
- 目前的基准测试计划面临着关于方法的合理批评.
研究的目的:
- 确定当前研究中的局限性,比较对象识别的神经和计算模型.
- 为了强调在感知研究中被忽视的行动和相互作用的重要性.
主要方法:
- 批判性地分析关于腹腔流研究和深度神经网络的现有文献.
- 确定当前比较方法的基本局限性.
主要成果:
- 当前比较行为/大脑数据与深度神经网络的研究存在重大局限性.
- 行动和相互作用在感知中的关键作用在很大程度上被忽视了.
结论:
- 专注于比较静态数据忽略了感知的动态性质.
- 未来的研究必须整合行动和相互作用,以更好地了解腹部流和物体识别.
相关概念视频
Vision
53.5K
Vision is the result of light being detected and transduced into neural signals by the retina of the eye. This information is then further analyzed and interpreted by the brain. First, light enters the front of the eye and is focused by the cornea and lens onto the retina—a thin sheet of neural tissue lining the back of the eye. Because of refraction through the convex lens of the eye, images are projected onto the retina upside-down and reversed.
53.5K
Color Vision
586
Color perception begins in the retina, the light-sensitive layer at the back of the eye. Two main theories explain how colors are seen: the trichromatic theory and the opponent-process theory. The trichromatic theory, proposed by Thomas Young in 1802 and extended by Hermann von Helmholtz in 1852, suggests that color vision is based on three types of cone receptors in the retina. These cones are sensitive to different but overlapping ranges of wavelengths corresponding to red, blue, and green.
586
Visual System
588
Light enters the eye through the cornea, a transparent, dome-shaped surface covering the surface of the eyeball that helps to direct and focus incoming light. This light is then channeled toward the pupil, an adjustable opening whose size is controlled by the iris. The iris, a pigmented muscle, regulates the amount of light entering the eye by contracting or dilating the pupil, thereby ensuring optimal light levels for clear vision.
Once through the pupil, the light passes through the lens, a...
Once through the pupil, the light passes through the lens, a...
588
Depth Perception and Spatial Vision
673
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
673
Parallel Processing
155
The brain processes sensory information rapidly due to parallel processing, which involves sending data across multiple neural pathways at the same time. This method allows the brain to manage various sensory qualities, such as shapes, colors, movements, and locations, all concurrently. For instance, when observing a forest landscape, the brain simultaneously processes the movement of leaves, the shapes of trees, the depth between them, and the various shades of green. This enables a quick and...
155
Steps in the Modeling Process
212
Albert Bandura's theory of observational learning identifies four critical processes: attention, retention, motor reproduction, and reinforcement or motivation.
Attention is the first necessary component for observational learning. It involves focusing on what the model is doing and saying. For example, if you decide to take a drawing class to enhance your skills, you need to pay close attention to the instructor's words and hand movements. The characteristics of the model significantly...
Attention is the first necessary component for observational learning. It involves focusing on what the model is doing and saying. For example, if you decide to take a drawing class to enhance your skills, you need to pay close attention to the instructor's words and hand movements. The characteristics of the model significantly...
212


