paper-with-me

홈 › Papers

CoSense3D: an Agent-based Efficient Learning Framework for Collective Perception

2024-04-29 · Yunshuang Yuan, Monika Sester

Collective Perception has attracted significant attention in recent years due to its advantage for mitigating occlusion and expanding the field-of-view, thereby enhancing reliability, efficiency, and, most crucially, decision-making safety. However, developing collective perception models is highly resource demanding due to extensive requirements of processing input data for many agents, usually dozens of images and point clouds for a single frame. This not only slows down the model development process for collective perception but also impedes the utilization of larger models. In this paper, we propose an agent-based training framework that handles the deep learning modules and agent data separately to have a cleaner data flow structure. This framework not only provides an API for flexibly prototyping the data processing pipeline and defining the gradient calculation for each agent, but also provides the user interface for interactive training, testing and data visualization. Training experiment results of four collective object detection models on the prominent collective perception benchmark OPV2V show that the agent-based training can significantly reduce the GPU memory consumption and training time while retaining inference performance. The framework and model implementations are available at \url{https://github.com/YuanYunshuang/CoSense3D}

📄 PDF Abstract BibTeX arXiv:2404.18617

Code (1)

yuanyunshuang/cosense3d 공식 구현 pytorch

Tasks

Data VisualizationDecision MakingGPUobject-detectionObject Detection

Similar Papers 제목 키워드 기반

DiscoSense: Commonsense Reasoning with Discourse Connectives

2022-10-22 · Prajjwal Bhargava, Vincent Ng

We present DiscoSense, a benchmark for commonsense reasoning via understanding a wide variety of discourse connectives. We generate compelling distractors in DiscoSense using Conditional Adversarial Filtering, an extensi…

Sentence Completion

StreamLTS: Query-based Temporal-Spatial LiDAR Fusion for Cooperative Object Detection

2024-07-04 · Yunshuang Yuan, Monika Sester

Cooperative perception via communication among intelligent traffic agents has great potential to improve the safety of autonomous driving. However, limited communication bandwidth, localization errors and asynchronized c…

Autonomous DrivingObjectobject-detectionObject Detection

CoSense-LLM: Semantics at the Edge with Cost- and Uncertainty-Aware Cloud-Edge Cooperation

2025-10-22 · Hasan Akgul, Mari Eplik, Javier Rojas, Aina Binti Abdullah 외 arxiv

We present CoSense-LLM, an edge-first framework that turns continuous multimodal sensor streams (for example Wi-Fi CSI, IMU, audio, RFID, and lightweight vision) into compact, verifiable semantic tokens and coordinates w…

When Multi-Robot Systems Meet Agentic AI:Towards Embodied Collective Intelligence

2026-06-26 · Yuxuan Yan, Yuanyuan Jia, Qianqian Yang arxiv

Embodied AI is increasingly becoming agentic, shifting robots from perception--control pipelines towards closed-loop systems that can retrieve context, deliberate during execution, monitor feedback, and refine future beh…

Learning Collective Dynamics of Multi-Agent Systems using Event-based Vision

2024-11-11 · Minah Lee, Uday Kamal, Saibal Mukhopadhyay

This paper proposes a novel problem: vision-based perception to learn and predict the collective dynamics of multi-agent systems, specifically focusing on interaction strength and convergence time. Multi-agent systems ar…

Event-based vision