Related Experiment Video
Updated: Jul 12, 2025

Author Spotlight: An Automated Method for Assessing Visual Acuity in Infants and Toddlers Using an Eye-Tracking System
Published on: March 17, 2023
Scene Uyghur Recognition Based on Visual Prediction Enhancement
Yaqi Liu1,2,3, Fanjie Kong1,2,3, Miaomiao Xu1,2,3
1College of Information Science and Engineering, Xinjang University, No. 777 Huarui Street, Urumqi 830017, China.
Abstract:
Aiming at the problems of Uyghur oblique deformation, character adhesion and character similarity in scene images, this paper proposes a scene Uyghur recognition model with enhanced visual prediction. First, the content-aware correction network TPS++ is used to perform feature-level correction for skewed text. Then, ABINet is used as the basic recognition network, and the U-Net structure in the vision model is improved to aggregate horizontal features, suppress multiple activation phenomena, better describe the spatial characteristics of character positions, and alleviate the problem of character adhesion. Finally, a visual masking semantic awareness (VMSA) module is added to guide the vision model to consider the language information in the visual space by masking the corresponding visual features on the attention map to obtain more accurate visual prediction. This module can not only alleviate the correction load of the language model, but also distinguish similar characters using the language information. The effectiveness of the improved method is verified by ablation experiments, and the model is compared with common scene text recognition methods and scene Uyghur recognition methods on the self-built scene Uyghur dataset.
More Related Videos
04:48Application of Deep Learning-Based Medical Image Segmentation via Orbital Computed Tomography
Published on: November 30, 2022
07:12Development of a Gaze-Contingent Display Framework Designed for Perceptual and Oculomotor Research with Simulated Central Vision Loss
Published on: April 11, 2025