在人类类型的立体视觉中实现强大的匹配的新原则
Ming Xie1, Tingfeng Lai1, Yuhui Fang1
1School of Mechanical and Aerospace Engineering, Nanyang Technological University, Singapore 639798, Singapore.
Biomimetics (Basel, Switzerland)
|July 28, 2023
概括
这项研究引入了一个新的机器人立体视觉匹配原理,使用先进的图像采样和受限制库伦能量神经网络. 这种方法增强了智能系统的机器感知和识别能力.
科学领域:
- 机器人和人工智能 机器人和人工智能
- 计算机视觉 计算机视觉
- 机器学习 机器学习
背景情况:
- 智能机器人和机器人需要类似人类的视觉感知才能在现实世界中互动.
- 立体视觉匹配仍然是实现机器对动态环境的强大理解的一个重大挑战.
研究的目的:
- 介绍一个强大的立体视觉匹配的新原则.
- 通过视觉感知增强智能机器的认知和识别能力.
主要方法:
- 整合一个自上而下的图像采样策略.
- 混合特征提取技术.混合特征提取技术.
- 限制库伦能量 (RCE) 神经网络用于增量学习和识别的应用.
主要成果:
- 建议解决方案的初步版本已成功实施和测试.
- 该系统在现实场景中展示了强大的立体视觉匹配的潜力,并与海事机器人X挑战数据进行了验证.
结论:
- 拟议的立体视觉匹配原则提供了一个新的研究方向.
- 这项工作可能会为未来的智能机器人,车辆和机器开发先进的立体视觉系统.
相关概念视频
Depth Perception and Spatial Vision
730
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
730
Sign Test for Matched Pairs
163
The sign test for matched pairs offers a robust method for comparing two paired samples, often for the effects of an intervention in one of them. This method is very useful in situations where the underlying distribution of the data is unknown. The test compares two related samples—often pre- and post-treatment measurements on the same subjects—to determine if there are significant differences in their median values.
To conduct the sign test, we first calculate the differences in...
To conduct the sign test, we first calculate the differences in...
163
Modeling and Similitude
291
Scaled modeling is a fundamental technique in engineering, enabling the study of large and complex systems by creating smaller, manageable replicas that recreate critical characteristics of the original. In hydrology and civil infrastructure, for example, scaled models of dams help analyze water flow, turbulence, and pressure. This method allows for accurate predictions of real-world behavior within a controlled environment, significantly reducing the cost and time involved in full-scale...
291
Prosopagnosia
210
Prosopagnosia, also known as face blindness, is the inability to recognize faces. In severe cases, individuals with prosopagnosia may not recognize close family members, including parents and spouses, by their faces. For instance, someone with prosopagnosia might walk past their child in a crowd, only realizing their mistake upon noticing their child's distinctive backpack or favorite jacket. Prosopagnosia specifically impairs facial recognition, while the recognition of other objects or...
210
Vision
53.6K
Vision is the result of light being detected and transduced into neural signals by the retina of the eye. This information is then further analyzed and interpreted by the brain. First, light enters the front of the eye and is focused by the cornea and lens onto the retina—a thin sheet of neural tissue lining the back of the eye. Because of refraction through the convex lens of the eye, images are projected onto the retina upside-down and reversed.
53.6K
Stereotype Content Model
14.8K
The Stereotype Content Model (SCM) was first proposed by Susan Fiske and her colleagues (Fiske, Cuddy, Glick & Xu, 2002; see also Fiske, 2012 and Fiske, 2017). The SCM specifies that when someone encounters a new group, they will stereotype them based on two metrics: warmth—or that group’s perceived intent, and how likely they are to provide help or inflict harm—and competence—or their ability to carry out that objective. Depending on the warmth-competence...
14.8K


