paper-with-me

홈 › Papers

Rethinking Air-Ground Collaboration: A Progressive Cross-Task Benchmark and Socialized Learning Framework

2026-06-17 · Zhoupeng Guo, Yunqi Zhu, Zhihe Fan, Xinjie Yao, Ruipu Zhao, Boan Tao, Yiming Sun, Zhen Wang, Pengfei Zhu arxiv

Air-ground collaborative perception is crucial for robust visual understanding in real-world dynamic environments. However, existing studies typically formulate collaboration as single-task cross-view fusion, overlooking the functional dependencies among localization, target association, and fine-grained parsing. In addition, the heterogeneous nature of aerial and ground views introduces substantial geometric, scale, and occlusion discrepancies, making uniform feature sharing vulnerable to negative transfer. To tackle these issues, we model air-ground perception as a progressive cross-task collaboration task and construct the Air-Ground Progressive Collaboration (AGPC) benchmark, a spatio-temporally aligned benchmark comprising more than 745K raw video frames. Built upon this benchmark, we propose Socialized Co-Perception (SCP), a coarse-to-fine framework that organizes collaboration progressively from aerial global localization to ground target association and identity-aware parsing. Its core module, the Dual-Layer Router (DLR), decouples input-side multi-scale expert selection from output-side task-conditioned modulation, enabling selective cross-view and cross-task interaction while suppressing harmful interference. Extensive experiments demonstrate the effectiveness of SCP. It achieves a 3.73\% coevolutionary gain and a 7.86\% improvement in average downstream performance. These results show that task-conditioned collaboration is more effective than uniform fusion for heterogeneous air-ground perception. The code is available at https://github.com/g1136639260-spec/AGSCP.

📄 PDF Abstract BibTeX arXiv:2606.18841

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance

2025-09-28 · Ky Dan Nguyen, Hoang Lam Tran, Anh-Dung Dinh, Daochang Liu 외 arxiv

Autoregressive (AR) models based on next-scale prediction have emerged as a powerful tool for image generation, but they face a critical weakness: information inconsistencies between patches across timesteps introduced b…

Text-to-Image Generation

Alignment-Process-Outcome: Rethinking How AIs and Humans Collaborate

2026-03-09 · Haichang Li, Anjun Zhu, Arpit Narechania arxiv

In real-world collaboration, alignment, process structure, and outcome quality do not exhibit a simple linear or one-to-one correspondence: similar alignment may accompany either rapid convergence or extensive multi-bran…

Combining Progressive Rethinking and Collaborative Learning: A Deep Framework for In-Loop Filtering

2020-01-16 · Dezhao Wang, Sifeng Xia, Wenhan Yang, Jiaying Liu

In this paper, we aim to address issues of (1) joint spatial-temporal modeling and (2) side information injection for deep-learning based in-loop filter. For (1), we design a deep network with both progressive rethinking…

Socialized Division and Collaboration: Rethinking Class-Incremental Learning under Optimization Conflicts

2026-08-21 · Xinjie Yao, Zhihe Fan, Yunqi Zhu, Jiaqi Zhou 외 arxiv

Class-incremental learning is commonly instantiated as a single-model paradigm, where a unified model sequentially adapts to an unbounded stream of sessions. While effective under mild distributional shifts, this formula…

class-incremental learningContinual Learning

Progressively Dual Prior Guided Few-shot Semantic Segmentation

2022-11-20 · Qinglong Cao, Yuntian Chen, Xiwen Yao, Junwei Han

Few-shot semantic segmentation task aims at performing segmentation in query images with a few annotated support samples. Currently, few-shot segmentation methods mainly focus on leveraging foreground information without…

Few-Shot Semantic SegmentationSegmentationSemantic Segmentation