paper-with-me

홈 › Papers

Slot Based Image Augmentation System for Object Detection

2019-07-19 · Yingwei Zhou

Object Detection has been a significant topic in computer vision. As the continuous development of Deep Learning, many advanced academic and industrial outcomes are established on localising and classifying the target objects, such as instance segmentation, video tracking and robotic vision. As the core concept of Deep Learning, Deep Neural Networks (DNNs) and associated training are highly integrated with task-driven modelling, having great effects on accurate detection. The main focus of improving detection performance is proposing DNNs with extra layers and novel topological connections to extract the desired features from input data. However, training these models can be computationally expensive and laborious progress as the complicated model architecture and enormous parameters. Besides, the dataset is another reason causing this issue and low detection accuracy, because of insufficient data samples or difficult instances. To address these training difficulties, this thesis presents two different approaches to improve the detection performance in the relatively light-weight way. As the intrinsic feature of data-driven in deep learning, the first approach is "slot-based image augmentation" to enrich the dataset with extra foreground and background combinations. Instead of the commonly used image flipping method, the proposed system achieved similar mAP improvement with less extra images which decrease training time. This proposed augmentation system has extra flexibility adapting to various scenarios and the performance-driven analysis provides an alternative aspect of conducting image augmentation

📄 PDF Abstract BibTeX arXiv:1907.12900

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningImage AugmentationInstance SegmentationObjectobject-detectionObject DetectionSemantic Segmentation

Similar Papers 제목 키워드 기반

Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning

2023-10-13 · Jinwoo Kim, Janghyuk Choi, Jaehyun Kang, Changyeon Lee 외

The binding problem in artificial neural networks is actively explored with the goal of achieving human-level recognition skills through the comprehension of the world in terms of symbol-like entities. Especially in the …

Image AugmentationObject

Automatic Vision-Based Parking Slot Detection and Occupancy Classification

2023-08-16 · Ratko Grbić, Brando Koch

Parking guidance information (PGI) systems are used to provide information to drivers about the nearest parking lots and the number of vacant parking slots. Recently, vision-based solutions started to appear as a cost-ef…

HIT-SCIR at MMNLU-22: Consistency Regularization for Multilingual Spoken Language Understanding

2023-01-05 · Bo Zheng, Zhouyang Li, Fuxuan Wei, Qiguang Chen 외

Multilingual spoken language understanding (SLU) consists of two sub-tasks, namely intent detection and slot filling. To improve the performance of these two sub-tasks, we propose to use consistency regularization based …

Data AugmentationIntent Detectionslot-fillingSlot Filling+1

SlotDiffusion: Object-Centric Generative Modeling with Diffusion Models

2023-05-18 · NeurIPS 2023 11 · Ziyi Wu, Jingyu Hu, Wuyue Lu, Igor Gilitschenski 외

Object-centric learning aims to represent visual data with a set of object entities (a.k.a. slots), providing structured representations that enable systematic generalization. Leveraging advanced architectures like Trans…

Image GenerationObjectObject DiscoverySemantic Segmentation+3

Speech2Slot: An End-to-End Knowledge-based Slot Filling from Speech

2021-05-10 · Pengwei Wang, Xin Ye, Xiaohuan Zhou, Jinghui Xie 외

In contrast to conventional pipeline Spoken Language Understanding (SLU) which consists of automatic speech recognition (ASR) and natural language understanding (NLU), end-to-end SLU infers the semantic meaning directly …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModellingNatural Language Understanding+7