paper-with-me

홈 › Papers

Spatial-wise Dynamic Distillation for MLP-like Efficient Visual Fault Detection of Freight Trains

2023-12-10 · Yang Zhang, Huilin Pan, Mingying Li, An Wang, Yang Zhou, Hongliang Ren

Despite the successful application of convolutional neural networks (CNNs) in object detection tasks, their efficiency in detecting faults from freight train images remains inadequate for implementation in real-world engineering scenarios. Existing modeling shortcomings of spatial invariance and pooling layers in conventional CNNs often ignore the neglect of crucial global information, resulting in error localization for fault objection tasks of freight trains. To solve these problems, we design a spatial-wise dynamic distillation framework based on multi-layer perceptron (MLP) for visual fault detection of freight trains. We initially present the axial shift strategy, which allows the MLP-like architecture to overcome the challenge of spatial invariance and effectively incorporate both local and global cues. We propose a dynamic distillation method without a pre-training teacher, including a dynamic teacher mechanism that can effectively eliminate the semantic discrepancy with the student model. Such an approach mines more abundant details from lower-level feature appearances and higher-level label semantics as the extra supervision signal, which utilizes efficient instance embedding to model the global spatial and semantic information. In addition, the proposed dynamic teacher can jointly train with students to further enhance the distillation efficiency. Extensive experiments executed on six typical fault datasets reveal that our approach outperforms the current state-of-the-art detectors and achieves the highest accuracy with real-time detection at a lower computational cost. The source code will be available at \url{https://github.com/MVME-HBUT/SDD-FTI-FDet}.

📄 PDF Abstract BibTeX arXiv:2312.05832

Code (1)

mvme-hbut/sdd-fti-fdet 공식 구현 pytorch

Tasks

Fault Detectionobject-detectionObject Detection

Similar Papers 제목 키워드 기반

ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation

2025-03-08 · Qizhen Lan, Qing Tian

Dense visual prediction tasks, such as detection and segmentation, are crucial for time-critical applications (e.g., autonomous driving and video surveillance). While deep models achieve strong performance, their efficie…

Autonomous Drivingfeature selectionKnowledge DistillationModel Compression+3

M^2C-EvDet: Multi-Domain Multi-Order Cross-Modal Knowledge Distillation for Event-based Object Detection

2026-06-23 · Wei Bao, Siqi Li, Shouan Pan, Yi Xie 외 arxiv

Event-based object Detection (EvDet), as a biologically inspired visual perception paradigm, demonstrates superior performance in scenarios demanding high temporal resolution and a wide dynamic range. Nevertheless, the i…

Knowledge DistillationObject Detection

LIX: Implicitly Infusing Spatial Geometric Prior Knowledge into Visual Semantic Segmentation for Autonomous Driving

2024-03-13 · Sicen Guo, Ziwei Long, Zhiyuan Wu, Qijun Chen 외

Despite the impressive performance achieved by data-fusion networks with duplex encoders for visual semantic segmentation, they become ineffective when spatial geometric data are not available. Implicitly infusing the sp…

Autonomous DrivingKnowledge DistillationSemantic Segmentation

SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning

2016-11-17 · CVPR 2017 7 · Long Chen, Hanwang Zhang, Jun Xiao, Liqiang Nie 외

Visual attention has been successfully applied in structural prediction tasks such as visual captioning and question answering. Existing visual attention models are generally spatial, i.e., the attention is modeled as sp…

Image CaptioningSentence

DMKD: Improving Feature-based Knowledge Distillation for Object Detection Via Dual Masking Augmentation

2023-09-06 · Guang Yang, Yin Tang, Zhijian Wu, Jun Li 외

Recent mainstream masked distillation methods function by reconstructing selectively masked areas of a student network from the feature map of its teacher counterpart. In these methods, the masked regions need to be prop…

Knowledge Distillationobject-detectionObject Detection