paper-with-me

Papers

NEMO: enabling neural-enhanced video streaming on commodity mobile devices

2020-09-21 · Annual International Conference on Mobile Computing and Networking 2020 9 · Hyunho Yeo, Chan Ju Chong, Youngmok Jung, Juncheol Ye, Dongsu Han

The demand for mobile video streaming has experienced tremendous growth over the last decade. However, existing methods of video delivery fall short of delivering high-quality video. Recent advances in neural super-resolution have opened up the possibility of enhancing video quality by leveraging client-side computation. Unfortunately, mobile devices cannot benefit from this because it is too expensive in computation and power-hungry. To overcome the limitation, we present NEMO, a system that enables real-time video super-resolution on mobile devices. NEMO applies neural super-resolution to a few select frames and transfers the outputs to benefit the remaining frames. The frames to which super-resolution is applied are carefully chosen to maximize the overall quality gains. NEMO leverages fine-grained dependencies using information from the video codec and strives to provide guarantees in the quality degradation compared to per-frame super-resolution. Our evaluation using a full system implementation on Android shows NEMO improves the overall processing throughput by x11.5, reduces energy consumption by 88.6%, and maintains device temperatures at acceptable levels compared to per-frame super-resolution, while ensuring high video quality. Overall, this leads to a 31.2% improvement in quality of experience for mobile users.

📄 PDF Abstract BibTeX

Code (1)

kaist-ina/nemo tf

Tasks

Super-ResolutionVideo Super-Resolution

Similar Papers 제목 키워드 기반

Ultra Flash: Scaling Real-Time Streaming Video Generation to High Resolutions

2026-06-08 · Luxury, Jie Huang, Zihao Fan, Xiaoxiao Ma 외 arxiv

While recent autoregressive video diffusion models achieve remarkable streaming quality, they remain confined to low resolutions (e.g., 480P), leaving efficient, scalable, real-time high-resolution video generation a fun…

Video Generation

VoLUT: Efficient Volumetric streaming enhanced by LUT-based super-resolution

2025-02-17 · Chendong Wang, Anlan Zhang, Yifan Yang, Lili Qiu 외

3D volumetric video provides immersive experience and is gaining traction in digital media. Despite its rising popularity, the streaming of volumetric video content poses significant challenges due to the high data bandw…

Super-Resolution

NeMo: Needle in a Montage for Video-Language Understanding

2025-09-29 · Zi-Yuan Hu, Shuo Liang, Duo Zheng, Yanyang Li 외 arxiv

Recent advances in video large language models (VideoLLMs) call for new evaluation protocols and benchmarks for complex temporal reasoning in video-language understanding. Inspired by the needle in a haystack test widely…

Video Question Answering

VideoLLM-online: Online Video Large Language Model for Streaming Video

2024-06-17 · CVPR 2024 1 · Joya Chen, Zhaoyang Lv, Shiwei Wu, Kevin Qinghong Lin 외

Recent Large Language Models have been enhanced with vision capabilities, enabling them to comprehend images, videos, and interleaved vision-language content. However, the learning methods of these large multimodal model…

GPULanguage ModelingLanguage ModellingLarge Language Model+1

video-SALMONN S: Memory-Enhanced Streaming Audio-Visual LLM

2025-10-13 · Guangzhi Sun, Yixuan Li, Xiaodong Wu, Yudong Yang 외 arxiv

Long-duration streaming video understanding is fundamental for future AI agents, yet remains limited by ineffective long-term memory. We introduce video-SALMONN S, a memory-enhanced streaming audio-visual large language …