A modified transformer based on adaptive frequency enhanced attention, large kernel convolution, and multiscale
Xiao Chang1, Shaobin Cai2,3, Wanchen Cai4
1College of Information Engineering, Huzhou University, Huzhou, China.
Abstract:
Bearing fault diagnosis has attracted increasing attention due to its critical role in monitoring the health of rotating machinery. Data-driven models based on deep learning (DL) have demonstrated strong capabilities in feature extraction. However, their performance often degrades under strong noise interference, which limits their applicability in real-world industrial scenarios. To address this issue, this paper proposes a novel attention-enhanced Transformer model that integrates large-kernel convolution and multiscale CNN structures for robust fault diagnosis. The proposed framework effectively combines spatiotemporal feature modeling with adaptive frequency-domain enhancement, enabling it to suppress noise and highlight informative diagnostic features. Experimental results on the Paderborn University and Case Western Reserve University datasets show that the proposed method achieves superior recognition accuracy under various signal-to-noise ratios, outperforming several state-of-the-art models. Furthermore, ablation studies and visualization analyses validate the effectiveness and soundness of the proposed architecture.
Related Concept Videos
Transformers with Off-Nominal Turns Ratios
Transformers
The iron core has a substantial relative permeability. Therefore, the magnetic field lines generated due to the current in one winding are almost entirely confined within the core, such that the same magnetic flux permeates each turn of both...
Energy Losses in Transformers
There are four main reasons for energy losses in transformers.
The first cause can be the high resistance of the...

