paper-with-me

Papers

Hysia: Serving DNN-Based Video-to-Retail Applications in Cloud

2020-06-09 · Huaizheng Zhang, Yuanming Li, Qiming Ai, Yong Luo, Yonggang Wen, Yichao Jin, Nguyen Binh Duong Ta

Combining \underline{v}ideo streaming and online \underline{r}etailing (V2R) has been a growing trend recently. In this paper, we provide practitioners and researchers in multimedia with a cloud-based platform named Hysia for easy development and deployment of V2R applications. The system consists of: 1) a back-end infrastructure providing optimized V2R related services including data engine, model repository, model serving and content matching; and 2) an application layer which enables rapid V2R application prototyping. Hysia addresses industry and academic needs in large-scale multimedia by: 1) seamlessly integrating state-of-the-art libraries including NVIDIA video SDK, Facebook faiss, and gRPC; 2) efficiently utilizing GPU computation; and 3) allowing developers to bind new models easily to meet the rapidly changing deep learning (DL) techniques. On top of that, we implement an orchestrator for further optimizing DL model serving performance. Hysia has been released as an open source project on GitHub, and attracted considerable attention. We have published Hysia to DockerHub as an official image for seamless integration and deployment in current cloud environments.

📄 PDF Abstract BibTeX arXiv:2006.05117

Code (2)

cap-ntu/Video-to-Retail-Platform 공식 구현 tf
NeosXu/benchmarking_pytorchvideo_and_mmaction2 pytorch

Tasks

GPUVideo RetrievalVideo-to-Shop

Similar Papers 제목 키워드 기반

PhysiAgent: An Embodied Agent Framework in Physical World

2025-09-29 · Zhihao Wang, Jianxiong Li, Jinliang Zheng, Wencong Zhang 외 arxiv

Vision-Language-Action (VLA) models have achieved notable success but often struggle with limited generalizations. To address this, integrating generalized Vision-Language Models (VLMs) as assistants to VLAs has emerged …

Scene Understanding

Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis

2025-06-17 · Varun Mannam, Zhenyu Shi

Accurate video annotation plays a vital role in modern retail applications, including customer behavior analysis, product interaction detection, and in-store activity recognition. However, conventional annotation methods…

Activity Recognition

A Serverless Cloud-Fog Platform for DNN-Based Video Analytics with Incremental Learning

2021-02-05 · Huaizheng Zhang, Meng Shen, Yizheng Huang, Yonggang Wen 외

DNN-based video analytics have empowered many new applications (e.g., automated retail). Meanwhile, the proliferation of fog devices provides developers with more design options to improve performance and save cost. To t…

Incremental LearningManagement

MARLIN: A Cloud Integrated Robotic Solution to Support Intralogistics in Retail

2024-07-02 · Dennis Mronga, Andreas Bresser, Fabian Maas, Adrian Danzglock 외

In this paper, we present the service robot MARLIN and its integration with the K4R platform, a cloud system for complex AI applications in retail. At its core, this platform contains so-called semantic digital twins, a …

Autonomous NavigationTask Planning

Cloud-Native Generative AI for Automated Planogram Synthesis: A Diffusion Model Approach for Multi-Store Retail Optimization

2026-01-02 · Ravi Teja Pagidoju, Shriya Agarwal arxiv

Planogram creation is a significant challenge for retail, requiring an average of 30 hours per complex layout. This paper introduces a cloud-native architecture using diffusion models to automatically generate store-spec…