paper-with-me

홈 › Papers

StreamTinyNet: video streaming analysis with spatial-temporal TinyML

2024-07-22 · Hazem Hesham Yousef Shalby, Massimo Pavan, Manuel Roveri

Tiny Machine Learning (TinyML) is a branch of Machine Learning (ML) that constitutes a bridge between the ML world and the embedded system ecosystem (i.e., Internet of Things devices, embedded devices, and edge computing units), enabling the execution of ML algorithms on devices constrained in terms of memory, computational capabilities, and power consumption. Video Streaming Analysis (VSA), one of the most interesting tasks of TinyML, consists in scanning a sequence of frames in a streaming manner, with the goal of identifying interesting patterns. Given the strict constraints of these tiny devices, all the current solutions rely on performing a frame-by-frame analysis, hence not exploiting the temporal component in the stream of data. In this paper, we present StreamTinyNet, the first TinyML architecture to perform multiple-frame VSA, enabling a variety of use cases that requires spatial-temporal analysis that were previously impossible to be carried out at a TinyML level. Experimental results on public-available datasets show the effectiveness and efficiency of the proposed solution. Finally, StreamTinyNet has been ported and tested on the Arduino Nicla Vision, showing the feasibility of what proposed.

📄 PDF Abstract BibTeX arXiv:2407.17524

Code (0)

등록된 구현이 없습니다.

Tasks

Edge-computing

Similar Papers 제목 키워드 기반

Streaming Video Model

2023-03-30 · CVPR 2023 1 · Yucheng Zhao, Chong Luo, Chuanxin Tang, Dongdong Chen 외

Video understanding tasks have traditionally been modeled by two separate architectures, specially tailored for two distinct tasks. Sequence-based video tasks, such as action recognition, use a video backbone to directly…

Action RecognitionDecodermodelMultiple Object Tracking+2

Streaming Video Diffusion: Online Video Editing with Diffusion Models

2024-05-30 · Feng Chen, Zhen Yang, Bohan Zhuang, Qi Wu

We present a novel task called online video editing, which is designed to edit \textbf{streaming} frames while maintaining temporal consistency. Unlike existing offline video editing assuming all frames are pre-establish…

Video Editing

Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training

2026-03-12 · Fangfu Liu, Diankun Wu, Jiawei Chi, Yimo Cai 외 arxiv

Humans perceive and understand real-world spaces through a stream of visual observations. Therefore, the ability to streamingly maintain and update spatial evidence from potentially unbounded video streams is essential f…

Adaptive 3D Gaussian Splatting Video Streaming: Visual Saliency-Aware Tiling and Meta-Learning-Based Bitrate Adaptation

2025-07-19 · Han Gong, Qiyue Li, Jie Li, Zhi Liu arxiv

3D Gaussian splatting video (3DGS) streaming has recently emerged as a research hotspot in both academia and industry, owing to its impressive ability to deliver immersive 3D video experiences. However, research in this …

OmniStream: Mastering Perception, Reconstruction and Action in Continuous Streams

2026-03-12 · Yibin Yan, Jilan Xu, Shangzhe Di, Haoning Wu 외 arxiv

Modern visual agents require representations that are general, causal, and physically structured to operate in real-time streaming environments. However, current vision foundation models remain fragmented, specializing n…

Representation LearningSpatial Reasoning