Scale-Invariant Feature Matching Network for V-D-T Few-Shot Semantic Segmentation

Summary

This study introduces a novel network for multi-modal few-shot semantic segmentation, improving accuracy by effectively fusing visible, depth, and thermal images. The proposed scale-invariant feature matching network enhances object detection across various sizes.