paper-with-me

Papers

MLMA-Net: multi-level multi-attentional learning for multi-label object detection in textile defect images

2021-01-31 · Bing Wei, Kuangrong Hao, Lei Gao

For the sake of recognizing and classifying textile defects, deep learning-based methods have been proposed and achieved remarkable success in single-label textile images. However, detecting multi-label defects in a textile image remains challenging due to the coexistence of multiple defects and small-size defects. To address these challenges, a multi-level, multi-attentional deep learning network (MLMA-Net) is proposed and built to 1) increase the feature representation ability to detect small-size defects; 2) generate a discriminative representation that maximizes the capability of attending the defect status, which leverages higher-resolution feature maps for multiple defects. Moreover, a multi-label object detection dataset (DHU-ML1000) in textile defect images is built to verify the performance of the proposed model. The results demonstrate that the network extracts more distinctive features and has better performance than the state-of-the-art approaches on the real-world industrial dataset.

📄 PDF Abstract BibTeX arXiv:2102.00376

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learningobject-detectionObject Detection

Similar Papers 제목 키워드 기반

MLMA: Towards Multilingual ASR With Mamba-based Architectures

2025-10-21 · Mohamed Nabih Ali, Daniele Falavigna, Alessio Brutti arxiv

Multilingual automatic speech recognition (ASR) remains a challenging task, especially when balancing performance across high- and low-resource languages. Recent advances in sequence modeling suggest that architectures b…

Speech Recognition

Multi-Level Matching and Aggregation Network for Few-Shot Relation Classification

2019-06-16 · ACL 2019 7 · Zhi-Xiu Ye, Zhen-Hua Ling

This paper presents a multi-level matching and aggregation network (MLMAN) for few-shot relation classification. Previous studies on this topic adopt prototypical networks, which calculate the embedding vector of a query…

Few-Shot Relation ClassificationGeneral ClassificationRelationRelation Classification

AttnGAN: Fine-Grained Text to Image Generation with Attentional Generative Adversarial Networks

2017-11-28 · CVPR 2018 6 · Tao Xu, Pengchuan Zhang, Qiuyuan Huang, Han Zhang 외

In this paper, we propose an Attentional Generative Adversarial Network (AttnGAN) that allows attention-driven, multi-stage refinement for fine-grained text-to-image generation. With a novel attentional generative networ…

Generative Adversarial NetworkImage GenerationImage-text matchingText Matching+2

Weakly-Supervised Multi-Level Attentional Reconstruction Network for Grounding Textual Queries in Videos

2020-03-16 · Yijun Song, Jingwen Wang, Lin Ma, Zhou Yu 외

The task of temporally grounding textual queries in videos is to localize one video segment that semantically corresponds to the given query. Most of the existing approaches rely on segment-sentence pairs (temporal annot…

Sentence

Hierarchical Attention Models for Multi-Relational Graphs

2024-04-14 · Roshni G. Iyer, Wei Wang, Yizhou Sun

We present Bi-Level Attention-Based Relational Graph Convolutional Networks (BR-GCN), unique neural network architectures that utilize masked self-attentional layers with relational graph convolutions, to effectively ope…

Graph AttentionLink PredictionNode ClassificationRelation