导航到现实世界的对象
Theophile Gervet1, Soumith Chintala2, Dhruv Batra2,3
1Carnegie Mellon University, Pittsburgh, PA, USA.
Science robotics
|June 28, 2023
概括
模块化学习在移动机器人的现实世界语义导航中表现出色,成功率达到90%. 端到端的学习由于模拟到真实的差距而扎,突出了模块化.
科学领域:
- 机器人技术 机器人技术 机器人技术
- 人工智能的人工智能
- 计算机视觉 计算机视觉
背景情况:
- 经典的空间导航管道缺乏对现实世界机器人部署的语义理解.
- 基于学习的方法,包括端到端和模块化方法,旨在增强机器人导航.
- 之前对视觉导航政策的评估是有限的,主要是在模拟中.
研究的目的:
- 在现实世界不受控制的环境中实证地比较语义视觉导航方法.
- 评估移动机器人的经典,模块化和端到端学习方法的性能.
- 识别机器人导航当前模拟基准中的挑战.
主要方法:
- 一项大规模的实证研究,比较语义视觉导航方法.
- 测试的方法包括经典,模块化学习和端到端学习方法.
- 在没有事先绘制地图或仪器的情况下,在六个家庭进行评估.
主要成果:
- 模块化学习在现实世界的语义导航中取得了90%的成功率.
- 端到端学习显示了显著的绩效下降,从模拟中的77%降至现实世界的23%.
- 端到端故障的主要原因被确定为模拟和现实之间的大型图像域差距.
结论:
- 模块化学习是一种可靠的方法,用于现实世界中的机器人导航到对象,从而实现有效的sim-to-real传输.
- 目前的模拟器是不可靠的基准,原因是图像和错误模式中的sim-to-real差距很大.
- 提供了改善模拟器和推进语义视觉导航研究的建议.
相关概念视频
Design Example: Identifying the Locations of Monuments in the Field Using Global Positioning System Device
121
Surveyors use Global Positioning System (GPS) technology to measure the precise location and elevation of points on Earth. In a recent survey, GPS receivers were used to determine the coordinates and elevations of two park monuments. The process involved careful mission planning, data collection, and correction to ensure accuracy. The survey began with mission planning to identify optimal satellite visibility and minimize Position Dilution of Precision (PDOP). A geodetic control point...
121
Collisions in Multiple Dimensions: Problem Solving
4.3K
In multiple dimensions, the conservation of momentum applies in each direction independently. Hence, to solve collisions in multiple dimensions, we should write down the momentum conservation in each direction separately. To help understand collisions in multiple dimensions, consider an example.
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
4.3K
Schemas
11.7K
A schema is a mental construct consisting of a cluster or collection of related concepts (Bartlett, 1932). There are many different types of schemata, and they all have one thing in common: schemata are a method of organizing information that allows the brain to work more efficiently. When a schema is activated, the brain makes immediate assumptions about the person or object being observed.
11.7K
Three-Dimensional Force System:Problem Solving
696
A three-dimensional force system refers to a scenario in which three forces act simultaneously in three different directions. This type of problem is commonly encountered in physics and engineering, where it is necessary to calculate the resultant force on the system, which can then be used to predict or analyze the behavior of the object or structure under consideration.
To solve a three-dimensional force system, first resolve each force into its respective scalar components. Do this using...
To solve a three-dimensional force system, first resolve each force into its respective scalar components. Do this using...
696
Inertial Frames of Reference
7.2K
Newton’s first law is usually considered to be a statement about reference frames. It provides a method for identifying a special type of reference frame: the inertial reference frame. In principle, we can make the net force on a body zero. If its velocity relative to a given frame is constant, then that frame is said to be inertial. So, by definition, an inertial reference frame is a reference frame where Newton's first law holds valid. Newton's first law applies to objects with...
7.2K
Depth Perception and Spatial Vision
745
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
745


