Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image Analysis

Summary

Med3DVLM, a novel 3D vision-language model (VLM), enhances 3D medical image analysis with efficient spatial feature extraction and improved image-text alignment. This model achieves superior performance in retrieval, report generation, and visual question answering tasks.