Disentangled Cross-Modal Transformer for RGB-D Salient Object Detection and Beyond

Summary

This study introduces a novel Disentangled Feature Pyramid (DFP) module for RGB-D salient object detection (SOD). DFP reduces fusion ambiguity by disentangling cross-modal contexts and representations, improving performance and adaptability.