paper-with-me

홈 › Papers

AIVD: Adaptive Edge-Cloud Collaboration for Accurate and Efficient Industrial Visual Detection

2026-01-08 · Yunqing Hu, Zheming Yang, Chang Zhao, Qi Guo, Meng Gao, Pengcheng Li, Wen Ji arxiv

Multimodal large language models (MLLMs) demonstrate exceptional capabilities in semantic understanding and visual reasoning, yet they still face challenges in precise object localization and resource-constrained edge-cloud deployment. To address this, this paper proposes the AIVD framework, which achieves unified precise localization and high-quality semantic generation through the collaboration between lightweight edge detectors and cloud-based MLLMs. To enhance the cloud MLLM's robustness against edge cropped-box noise and scenario variations, we design an efficient fine-tuning strategy with visual-semantic collaborative augmentation, significantly improving classification accuracy and semantic consistency. Furthermore, to maintain high throughput and low latency across heterogeneous edge devices and dynamic network conditions, we propose a heterogeneous resource-aware dynamic scheduling algorithm. Experimental results demonstrate that AIVD substantially reduces resource consumption while improving MLLM classification performance and semantic generation quality. The proposed scheduling strategy also achieves higher throughput and lower latency across diverse scenarios.

📄 PDF Abstract BibTeX arXiv:2601.04734

Code (0)

등록된 구현이 없습니다.

Tasks

Object LocalizationVisual Reasoning

Similar Papers 제목 키워드 기반

CE-CoLLM: Efficient and Adaptive Large Language Models Through Cloud-Edge Collaboration

2024-11-05 · Hongpeng Jin, Yanzhao Wu

Large Language Models (LLMs) exhibit remarkable human-like predictive capabilities. However, it is challenging to deploy LLMs to provide efficient and adaptive inference services at the edge. This paper proposes a novel …

Collaborative InferenceLarge Language Model

LAECIPS: Large Vision Model Assisted Adaptive Edge-Cloud Collaboration for IoT-based Embodied Intelligence System

2024-04-16 · Shijing Hu, Zhihui Lu, Xin Xu, Ruijun Deng 외

Embodied intelligence (EI) enables manufacturing systems to flexibly perceive, reason, adapt, and operate within dynamic shop floor environments. In smart manufacturing, a representative EI scenario is robotic visual ins…

Autonomous DrivingContinual LearningIndustrial RobotsSemantic Segmentation

PRISM: Privacy-Aware Routing for Adaptive Cloud-Edge LLM Inference via Semantic Sketch Collaboration

2025-11-27 · Junfei Zhan, Haoxun Shen, Zheng Lin, Tengjiao He arxiv

Large Language Models (LLMs) demonstrate impressive capabilities in natural language understanding and generation, but incur high communication overhead and privacy risks in cloud deployments, while facing compute and me…

Natural Language Understanding

DECICE: Device-Edge-Cloud Intelligent Collaboration Framework

2023-05-04 · Julian Kunkel, Christian Boehme, Jonathan Decker, Fabrizio Magugliani 외

DECICE is a Horizon Europe project that is developing an AI-enabled open and portable management framework for automatic and adaptive optimization and deployment of applications in computing continuum encompassing from I…

Management

edgeVLM: Cloud-edge Collaborative Real-time VLM based on Context Transfer

2025-08-18 · Chen Qian, Xinran Yu, Zewen Huang, Danyang Li 외 arxiv

Vision-Language Models (VLMs) are increasingly deployed in real-time applications such as autonomous driving and human-computer interaction, which demand fast and reliable responses based on accurate perception. To meet …

Autonomous DrivingVisual Grounding