无突出点和审美意识的全景视频导航
概括
这项研究引入了全景视频导航的新"有意义驱动"方法,超越了传统的基于突出性的方法. 这种新的技术为沉浸式内容产生了更相关,更美观的导航路径.
科学领域:
- 计算机视觉 计算机视觉
- 人与计算机的交互
- 多媒体系统 多媒体系统
背景情况:
- 目前的全景视频导航严重依赖于突出性驱动的方法,经常使用现成工具.
- 这些以突出性为基础的方法不充分代表视频内容,并导致低美感的导航路径.
- 需要对 Saliency 适用于全景视频导航的适用性进行批判性重新评估.
研究的目的:
- 为全景视频提出一种新的,没有突出性的导航范式,优先考虑内容意义.
- 开发一个无监督的学习方案,在导航路径内产生高美感的视角.
- 通过更适当的内容覆盖和愉快的观看来增强用户体验.
主要方法:
- 一个新的导航范式训练在眼睛固定,但由感知内容的意义驱动.
- 一个无监督的学习计划,以确保局部视图的美学质量.
- 开发定量评估方案,包括客观和主观的用户研究.
主要成果:
- 与以突出为导向的方法相比,拟议的方法产生了更好的内容覆盖率的导航路径.
- 这种方法成功地产生了具有高度美学视角的导航路径,增强了用户体验.
- 定量评估表明了新范式的有效性和潜力.
结论:
- Saliency 驱动的方法对于有效的全景视频导航是不够的.
- 一个"有意义的驱动"的方法,加上无监督的美学优化,提供了一个优越的替代方案.
- 这项研究为全景视频导航系统的新方向奠定了基础.
相关概念视频
Depth Perception and Spatial Vision
508
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
508
Vision
52.9K
Vision is the result of light being detected and transduced into neural signals by the retina of the eye. This information is then further analyzed and interpreted by the brain. First, light enters the front of the eye and is focused by the cornea and lens onto the retina—a thin sheet of neural tissue lining the back of the eye. Because of refraction through the convex lens of the eye, images are projected onto the retina upside-down and reversed.
52.9K


