paper-with-me

Papers

Talk Less, Fly Lighter: Autonomous Semantic Compression for UAV Swarm Communication via LLMs

2025-08-16 · Fei Lin, Tengchao Zhang, Qinghua Ni, Jun Huang, Siji Ma, Yonglin Tian, Yisheng Lv, Naiqi Wu arxiv

The rapid adoption of Large Language Models (LLMs) in unmanned systems has significantly enhanced the semantic understanding and autonomous task execution capabilities of Unmanned Aerial Vehicle (UAV) swarms. However, limited communication bandwidth and the need for high-frequency interactions pose severe challenges to semantic information transmission within the swarm. This paper explores the feasibility of LLM-driven UAV swarms for autonomous semantic compression communication, aiming to reduce communication load while preserving critical task semantics. To this end, we construct four types of 2D simulation scenarios with different levels of environmental complexity and design a communication-execution pipeline that integrates system prompts with task instruction prompts. On this basis, we systematically evaluate the semantic compression performance of nine mainstream LLMs in different scenarios and analyze their adaptability and stability through ablation studies on environmental complexity and swarm size. Experimental results demonstrate that LLM-based UAV swarms have the potential to achieve efficient collaborative communication under bandwidth-constrained and multi-hop link conditions.

📄 PDF Abstract BibTeX arXiv:2508.12043

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Baseline for the Commands For Autonomous Vehicles Challenge

2020-04-20 · Simon Vandenhende, Thierry Deruyttere, Dusan Grujicic

The Commands For Autonomous Vehicles (C4AV) challenge requires participants to solve an object referral task in a real-world setting. More specifically, we consider a scenario where a passenger can pass free-form natural…

Autonomous VehiclesObject

DAVD-Net: Deep Audio-Aided Video Decompression of Talking Heads

2020-06-01 · CVPR 2020 6 · Xi Zhang, Xiaolin Wu, Xinliang Zhai, Xianye Ben 외

Close-up talking heads are among the most common and salient object in video contents, such as face-to-face conversations in social media, teleconferences, news broadcasting, talk shows, etc. Due to the high sensitivity …

Video CompressionVideo Reconstruction

FairFly: A Fair Motion Planner for Fleets of Autonomous UAVs in Urban Airspace

2020-08-21

We present a solution to the problem of fairly planning a fleet of Unmanned Aerial Vehicles (UAVs) that have different missions and operators, such that no one operator unfairly gets to finish its missions early at the e…

Fairness

Multi-modality Deep Restoration of Extremely Compressed Face Videos

2021-07-05 · Xi Zhang, Xiaolin Wu

Arguably the most common and salient object in daily video communications is the talking head, as encountered in social media, virtual classrooms, teleconferences, news broadcasting, talk shows, etc. When communication b…

Quantization

Spotlighter: Revisiting Prompt Tuning from a Representative Mining View

2025-08-31 · Yutong Gao, Maoyuan Shao, Xinyang Huang, Chuang Zhu 외 arxiv

CLIP's success has demonstrated that prompt tuning can achieve robust cross-modal semantic alignment for tasks ranging from open-domain recognition to fine-grained classification. However, redundant or weakly relevant fe…