Enabling Collaborative Video Sensing at the Edge through Convolutional Sharing
While Deep Neural Network (DNN) models have provided remarkable advances in machine vision capabilities, their high computational complexity and model sizes present a formidable roadblock to deployment in AIoT-based sensing applications. In this paper, we propose a novel paradigm by which peer nodes in a network can collaborate to improve their accuracy on person detection, an exemplar machine vision task. The proposed methodology requires no re-training of the DNNs and incurs minimal processing latency as it extracts scene summaries from the collaborators and injects back into DNNs of the reference cameras, on-the-fly. Early results show promise with improvements in recall as high as 10% with a single collaborator, on benchmark datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Human DetectionSimilar Papers 제목 키워드 기반
Color When It Counts: Grayscale-Guided Online Triggering for Always-On Streaming Video Sensing
Always-on sensing is essential for next-generation edge/wearable AI systems, yet continuous high-fidelity RGB video capture remains prohibitively expensive for resource-constrained mobile and edge platforms. We present a…
IoV-Oriented Integrated Sensing, Computation, and Communication: System Design and Resource Allocation
—In future mobile communication systems, the deep integration of communication, sensing, and computation has become a new trend. This article paper designs an integrated sensing, computation, and communication (ISCC) sy…
Deep Reinforcement LearningEdge-computingIntegrated sensing and communicationISACUniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models
The domain gap between remote sensing imagery and natural images has recently received widespread attention and Vision-Language Models (VLMs) have demonstrated excellent generalization performance in remote sensing multi…
Question AnsweringScene ClassificationVisual Question AnsweringA Survey on Video Analytics in Cloud-Edge-Terminal Collaborative Systems
The explosive growth of video data has driven the development of distributed video analytics in cloud-edge-terminal collaborative (CETC) systems, enabling efficient video processing, real-time inference, and privacy-pres…
Autonomous DrivingEdge-computingPrivacy PreservingScheduling+1CONVERGE: A Multi-Agent Vision-Radio Architecture for xApps
Telecommunications and computer vision have evolved independently. With the emergence of high-frequency wireless links operating mostly in line-of-sight, visual data can help predict the channel dynamics by detecting obs…