paper-with-me

홈 › Papers

A comprehensive overview of deep learning models for object detection from videos/images

2026-01-21 · Sukana Zulfqar, Sadia Saeed, M. Azam Zia, Anjum Ali, Faisal Mehmood, Abid Ali arxiv

Object detection in video and image surveillance is a well-established yet rapidly evolving task, strongly influenced by recent deep learning advancements. This review summarises modern techniques by examining architectural innovations, generative model integration, and the use of temporal information to enhance robustness and accuracy. Unlike earlier surveys, it classifies methods based on core architectures, data processing strategies, and surveillance specific challenges such as dynamic environments, occlusions, lighting variations, and real-time requirements. The primary goal is to evaluate the current effectiveness of semantic object detection, while secondary aims include analysing deep learning models and their practical applications. The review covers CNN-based detectors, GAN-assisted approaches, and temporal fusion methods, highlighting how generative models support tasks such as reconstructing missing frames, reducing occlusions, and normalising illumination. It also outlines preprocessing pipelines, feature extraction progress, benchmarking datasets, and comparative evaluations. Finally, emerging trends in low-latency, efficient, and spatiotemporal learning approaches are identified for future research.

📄 PDF Abstract BibTeX arXiv:2601.14677

Code (0)

등록된 구현이 없습니다.

Tasks

Object Detection

Similar Papers 제목 키워드 기반

Few-Shot Object Detection: Research Advances and Challenges

2024-04-07 · Zhimeng Xin, Shiming Chen, Tianxu Wu, Yuanjie Shao 외

Object detection as a subfield within computer vision has achieved remarkable progress, which aims to accurately identify and locate a specific object from images or videos. Such methods rely on large-scale labeled train…

Few-Shot LearningFew-Shot Object DetectionObjectobject-detection+2

The Endoscapes Dataset for Surgical Scene Segmentation, Object Detection, and Critical View of Safety Assessment: Official Splits and Benchmark

2023-12-19 · Aditya Murali, Deepak Alapatt, Pietro Mascagni, Armine Vardazaryan 외

This technical report provides a detailed overview of Endoscapes, a dataset of laparoscopic cholecystectomy (LC) videos with highly intricate annotations targeted at automated assessment of the Critical View of Safety (C…

AnatomyInstance Segmentationobject-detectionObject Detection+3

Quality Assessment of DIBR-synthesized views: An Overview

2019-11-16 · Shishun Tian, Lu Zhang, Wenbin Zou, Xia Li 외

The Depth-Image-Based-Rendering (DIBR) is one of the main fundamental technique to generate new views in 3D video applications, such as Multi-View Videos (MVV), Free-Viewpoint Videos (FVV) and Virtual Reality (VR). Howev…

Survey

A Comprehensive Study on Object Detection Techniques in Unconstrained Environments

2023-04-11 · Hrishitva Patel

Object detection is a crucial task in computer vision that aims to identify and localize objects in images or videos. The recent advancements in deep learning and Convolutional Neural Networks (CNNs) have significantly i…

Objectobject-detectionObject Detection

Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing

2024-09-26 · Huthaifa I. Ashqar, Ahmed Jaber, Taqwa I. Alhadidi, Mohammed Elhenawy

This study aims to comprehensively review and empirically evaluate the application of multimodal large language models (MLLMs) and Large Vision Models (VLMs) in object detection for transportation systems. In the first f…

Event DetectionObjectobject-detectionObject Detection+1