An image caption model based on attention mechanism and deep reinforcement learning.

Tong Bai1, Sen Zhou2, Yu Pang1

  • 1School of Optoelectronic Engineering, Chongqing University of Posts and Telecommunications, Chongqing, China.

Frontiers in Neuroscience
|October 23, 2023
PubMed
Summary

This study introduces a novel guided decoding network for image captioning, enhancing visual information processing and description generation. The improved model achieves better performance on standard datasets and evaluation metrics.