paper-with-me

Papers

Exploring the Applications of Faster R-CNN and Single-Shot Multi-box Detection in a Smart Nursery Domain

2018-08-27 · Somnuk Phon-Amnuaisuk, Ken T. Murata, Praphan Pavarangkoon, Kazunori Yamamoto, Takamichi Mizuhara

The ultimate goal of a baby detection task concerns detecting the presence of a baby and other objects in a sequence of 2D images, tracking them and understanding the semantic contents of the scene. Recent advances in deep learning and computer vision offer various powerful tools in general object detection and can be applied to a baby detection task. In this paper, the Faster Region-based Convolutional Neural Network and the Single-Shot Multi-Box Detection approaches are explored. They are the two state-of-the-art object detectors based on the region proposal tactic and the multi-box tactic. The presence of a baby in the scene obtained from these detectors, tested using different pre-trained models, are discussed. This study is important since the behaviors of these detectors in a baby detection task using different pre-trained models are still not well understood. This exploratory study reveals many useful insights into the applications of these object detectors in the smart nursery domain.

📄 PDF Abstract BibTeX arXiv:1808.08675

Code (0)

등록된 구현이 없습니다.

Tasks

Objectobject-detectionObject DetectionRegion Proposal

Similar Papers 제목 키워드 기반

XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model

2024-06-07 · Edresson Casanova, Kelly Davis, Eren Gölge, Görkem Göknar 외

Most Zero-shot Multi-speaker TTS (ZS-TTS) systems support only a single language. Although models like YourTTS, VALL-E X, Mega-TTS 2, and Voicebox explored Multilingual ZS-TTS they are limited to just a few high/medium r…

text-to-speechText to SpeechVoice CloningZero-Shot Multi-Speaker TTS

FROST: Faster and more Robust One-shot Semi-supervised Training

2020-11-18 · Helena E. Liu, Leslie N. Smith

Recent advances in one-shot semi-supervised learning have lowered the barrier for deep learning of new applications. However, the state-of-the-art for semi-supervised learning is slow to train and the performance is sens…

Multi-student Diffusion Distillation for Better One-step Generators

2024-10-30 · Yanke Song, Jonathan Lorraine, Weili Nie, Karsten Kreis 외

Diffusion models achieve high-quality sample generation at the cost of a lengthy multistep inference procedure. To overcome this, diffusion distillation techniques produce student generators capable of matching or surpas…

Image Generation

Exploring the Zero-Shot Capabilities of LLMs Handling Multiple Problems at once

2024-06-16 · Zhengxiang Wang, Jordan Kodner, Owen Rambow

Recent studies have proposed placing multiple problems in a single prompt to improve input token utilization for a more efficient LLM inference. We call this MPP, in contrast to conventional SPP that prompts an LLM with …

A Single-shot Object Detector with Feature Aggragation and Enhancement

2019-02-08 · Weiqiang Li, Guizhong Liu

For many real applications, it is equally important to detect objects accurately and quickly. In this paper, we propose an accurate and efficient single shot object detector with feature aggregation and enhancement (FAEN…