paper-with-me

Papers

Joint Attention-Guided Feature Fusion Network for Saliency Detection of Surface Defects

2024-02-05 · Xiaoheng Jiang, Feng Yan, Yang Lu, Ke Wang, Shuai Guo, Tianzhu Zhang, Yanwei Pang, Jianwei Niu, Mingliang Xu

Surface defect inspection plays an important role in the process of industrial manufacture and production. Though Convolutional Neural Network (CNN) based defect inspection methods have made huge leaps, they still confront a lot of challenges such as defect scale variation, complex background, low contrast, and so on. To address these issues, we propose a joint attention-guided feature fusion network (JAFFNet) for saliency detection of surface defects based on the encoder-decoder network. JAFFNet mainly incorporates a joint attention-guided feature fusion (JAFF) module into decoding stages to adaptively fuse low-level and high-level features. The JAFF module learns to emphasize defect features and suppress background noise during feature fusion, which is beneficial for detecting low-contrast defects. In addition, JAFFNet introduces a dense receptive field (DRF) module following the encoder to capture features with rich context information, which helps detect defects of different scales. The JAFF module mainly utilizes a learned joint channel-spatial attention map provided by high-level semantic features to guide feature fusion. The attention map makes the model pay more attention to defect features. The DRF module utilizes a sequence of multi-receptive-field (MRF) units with each taking as inputs all the preceding MRF feature maps and the original input. The obtained DRF features capture rich context information with a large range of receptive fields. Extensive experiments conducted on SD-saliency-900, Magnetic tile, and DAGM 2007 indicate that our method achieves promising performance in comparison with other state-of-the-art methods. Meanwhile, our method reaches a real-time defect detection speed of 66 FPS.

📄 PDF Abstract BibTeX arXiv:2402.02797

Code (0)

등록된 구현이 없습니다.

Tasks

Defect DetectionSaliency Detection

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Relevance-guided Audio Visual Fusion for Video Saliency Prediction

2024-11-18 · Li Yu, Xuanzhe Sun, Pan Gao, Moncef Gabbouj

Audio data, often synchronized with video frames, plays a crucial role in guiding the audience's visual attention. Incorporating audio information into video saliency prediction tasks can enhance the prediction of human …

PredictionSaliency PredictionVideo Saliency Prediction

GazeFusion: Saliency-Guided Image Generation

2024-03-16 · Yunxiang Zhang, Nan Wu, Connor Z. Lin, Gordon Wetzstein 외

Diffusion models offer unprecedented image generation power given just a text prompt. While emerging approaches for controlling diffusion models have enabled users to specify the desired spatial layouts of the generated …

Image Generation

Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex Degradations

2026-08-01 · Axi Niu, Jieheng Li, Kang Zhang, Qingsen Yan 외 arxiv

Infrared-visible image fusion under realistic degradation scenarios is a challenging task, as degradations not only cause a loss of reliable modality-specific information in observed images but also hinder the fusion pro…

EEG-Driven Image Reconstruction with Saliency-Guided Diffusion Models

2025-10-30 · Igor Abramov, Ilya Makarov arxiv

Existing EEG-driven image reconstruction methods often overlook spatial attention mechanisms, limiting fidelity and semantic coherence. To address this, we propose a dual-conditioning framework that combines EEG embeddin…

Image ReconstructionImage Generation

Data Augmentation via Latent Diffusion for Saliency Prediction

2024-09-11 · Bahar Aydemir, Deblina Bhattacharjee, Tong Zhang, Mathieu Salzmann 외

Saliency prediction models are constrained by the limited diversity and quantity of labeled data. Standard data augmentation techniques such as rotating and cropping alter scene composition, affecting saliency. We propos…

Data AugmentationDiversityPredictionSaliency Prediction