,

Yue Zhang1, Wanshu Fan2, Peixi Peng1

  • 1National and Local Joint Engineering Laboratory of Computer Aided Design, School of Software Engineering, Dalian University, Dalian, 116622, Liaoning, China.

概括

本研究介绍了机器人手术的视觉问题答案 (VQA) 模型,该模型在回答问题时精确定位相关的图像区域. 新的双模态提示模型增强了多模态交互,提高了手术环境中的准确性.