CRViT-YOLO:畳み込み再構築ビジョントランスフォーマーを用いた多形態血球検出法
Yaning Du1, Yuliang Ma2, Qingshan She3
1School of Automation, Hangzhou Dianzi University, Hangzhou, Zhejiang 310018, China; Jinan Key Laboratory of Rehabilitation and Evaluation of Motor Dysfunction, The People's Hospital of Huaiyin, Jinan, Shandong 250100, China.
Abstract:
Complete blood cell counting plays a critical role in medical diagnostics; however, conventional manual examination is time-consuming and prone to errors due to variations in data sources, image quality, cell morphology, and staining characteristics. Deep learning has emerged as a promising solution to enhance both the accuracy and efficiency of blood cell detection. In this study, we present CRViT-YOLO, a novel detection framework built upon the YOLOv9 architecture. The proposed framework incorporates a Convolutional-Reconstructed Vision Transformer (CRViT) module to improve feature extraction by effectively capturing both local and global contextual information. Furthermore, a Feature Enhancement Module (FEM) is introduced to refine local feature representations, while the integration of the EIoU loss function enhances localization accuracy, particularly for densely packed or overlapping cells across diverse scales and types, and demonstrates robust performance in detecting polymorphic, healthy, and pathological cells. Extensive experiments conducted on four publicly available datasets-BCCD, BCDD, LISC, and BBBC041-validate the effectiveness and generalizability of the proposed approach, achieving mean average precision (mAP@50) scores of 93.9 %, 99.4 %, 98.8 %, and 76.0 %, respectively, in multi-class blood cell detection tasks.
関連する概念動画
Blood Transfusion and Agglutination
History
The history of blood transfusion dates back to the 17th century, when early attempts were made in animals. In 1818 James Blundell, a British doctor, performed the first successful human blood transfusion. Later in 1900, Karl...
Flow Cytometry
In...

