paper-with-me

홈 › Papers

TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks

2026-04-08 · Xiangyu Wang, Jin Wu, Haoran Shi, Wei Xia, Jiarui Yu, Chanjin Zheng arxiv

Recently, multi-Large Language Model (LLM) frameworks have been proposed to solve contextualized tasks. However, these frameworks do not explicitly emulate human team role division, which may lead to a single perspective, thereby weakening performance on multi-step contextualized tasks. To address this issue, we propose TeamLLM, a human-like Team-Oriented Multi-LLM Collaboration Framework. TeamLLM adopts four team roles with distinct division and employs a three-phase multi-LLM collaboration for multi-step contextualized tasks. To evaluate the effectiveness of TeamLLM on multi-step contextualized tasks, we propose Contextually-Grounded and Procedurally-Structured tasks (CGPST) and construct the CGPST benchmark. This benchmark has four core features: contextual grounding, procedural structure, process-oriented evaluation and multi-dimensional assessment. We evaluate ten popular LLMs on CGPST at overall-level, step-level, and dimension-level. Results show that TeamLLM substantially improves performance on CGPST. We release the benchmark with scenarios, full-process responses and human scores from ten LLMs. The code and data are available at https://anonymous.4open.science/r/TeamLLM-anonymous-C50E/.

📄 PDF Abstract BibTeX arXiv:2604.06765

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RedTeamLLM: an Agentic AI framework for offensive security

2025-05-11 · Brian Challita, Pierre Parrend

From automated intrusion testing to discovery of zero-day attacks before software launch, agentic AI calls for great promises in security engineering. This strong capability is bound with a similar threat: the security a…

Management

Multi-Agent Collaboration via Cross-Team Orchestration

2024-06-13 · Zhuoyun Du, Chen Qian, Wei Liu, Zihao Xie 외

Large Language Models (LLMs) have significantly impacted various domains, especially through organized LLM-driven autonomous agents. A representative scenario is in software development, where agents can collaborate in a…

Story Generation

Societal AI Research Has Become Less Interdisciplinary

2025-06-10 · Dror Kris Markus, Fabrizio Gilardi, Daria Stetsenko

As artificial intelligence (AI) systems become deeply embedded in everyday life, calls to align AI development with ethical and societal values have intensified. Interdisciplinary collaboration is often championed as a k…

FairnessMisinformation

Deep neural networks for collaborative learning analytics: Evaluating team collaborations using student gaze point prediction

2020-10-16 · Zang Guo, Roghayeh Barmaki

Automatic assessment and evaluation of team performance during collaborative tasks is key to the learning analytics and computer-supported cooperative work research. There is a growing interest in the use of gaze-oriente…

Anatomy

Mixed-Initiative Human-Robot Teaming under Suboptimality with Online Bayesian Adaptation

2024-03-24 · Manisha Natarajan, Chunyue Xue, Sanne van Waveren, Karen Feigh 외

For effective human-agent teaming, robots and other artificial intelligence (AI) agents must infer their human partner's abilities and behavioral response patterns and adapt accordingly. Most prior works make the unreali…

Decision MakingSequential Decision Making