paper-with-me

Papers

SpectraSentinel: LightWeight Dual-Stream Real-Time Drone Detection, Tracking and Payload Identification

2025-07-30 · Shahriar Kabir, Istiak Ahmmed Rifti, H. M. Shadman Tabib, Mushfiqur Rahman, Sadatul Islam Sadi, Hasnaen Adil, Ahmed Mahir Sultan Rumi, Ch Md Rakin Haider arxiv

The proliferation of drones in civilian airspace has raised urgent security concerns, necessitating robust real-time surveillance systems. In response to the 2025 VIP Cup challenge tasks - drone detection, tracking, and payload identification - we propose a dual-stream drone monitoring framework. Our approach deploys independent You Only Look Once v11-nano (YOLOv11n) object detectors on parallel infrared (thermal) and visible (RGB) data streams, deliberately avoiding early fusion. This separation allows each model to be specifically optimized for the distinct characteristics of its input modality, addressing the unique challenges posed by small aerial objects in diverse environmental conditions. We customize data preprocessing and augmentation strategies per domain - such as limiting color jitter for IR imagery - and fine-tune training hyperparameters to enhance detection performance under conditions of heavy noise, low light, and motion blur. The resulting lightweight YOLOv11n models demonstrate high accuracy in distinguishing drones from birds and in classifying payload types, all while maintaining real-time performance. This report details the rationale for a dual-modality design, the specialized training pipelines, and the architectural optimizations that collectively enable efficient and accurate drone surveillance across RGB and IR channels.

📄 PDF Abstract BibTeX arXiv:2507.22650

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Equivalent Transformation and Dual Stream Network Construction for Mobile Image Super-Resolution

2023-01-01 · CVPR 2023 1 · Jiahao Chao, Zhou Zhou, Hongfan Gao, Jiali Gong 외

In recent years, there has been an increasing demand for real-time super-resolution networks on mobile devices. To address this issue, many lightweight super-resolution models have been proposed. However, these model…

Image Super-ResolutionSuper-Resolution

Ultralight Polarity-Split Neuromorphic SNN for Event-Stream Super-Resolution

2025-08-05 · Chuanzhi Xu, Haoxian Zhou, Langyi Chen, Yuk Ying Chung 외 arxiv

Event cameras offer unparalleled advantages such as high temporal resolution, low latency, and high dynamic range. However, their limited spatial resolution poses challenges for fine-grained perception tasks. In this wor…

Deja Vu in Plots: Leveraging Cross-Session Evidence with Retrieval-Augmented LLMs for Live Streaming Risk Assessment

2026-01-22 · Yiran Qiao, Xiang Ao, Jing Chen, Yang Liu 외 arxiv

The rise of live streaming has transformed online interaction, enabling massive real-time engagement but also exposing platforms to complex risks such as scams and coordinated malicious behaviors. Detecting these risks i…

Online Hybrid Lightweight Representations Learning: Its Application to Visual Tracking

2022-05-23 · Ilchae Jung, Minji Kim, Eunhyeok Park, Bohyung Han

This paper presents a novel hybrid representation learning framework for streaming data, where an image frame in a video is modeled by an ensemble of two distinct deep neural networks; one is a low-bit quantized network …

Representation LearningVisual Tracking

DualSep: A Light-weight dual-encoder convolutional recurrent network for real-time in-car speech separation

2024-09-13 · Ziqian Wang, Jiayao Sun, Zihan Zhang, Xingchen Li 외

Advancements in deep learning and voice-activated technologies have driven the development of human-vehicle interaction. Distributed microphone arrays are widely used in in-car scenarios because they can accurately captu…

CPUSpeech Separation