paper-with-me

Papers

Adversarial Multi-Agent Evaluation of Large Language Models through Iterative Debates

2024-10-07 · Chaithanya Bandi, Abir Harrasse

This paper explores optimal architectures for evaluating the outputs of large language models (LLMs) using LLMs themselves. We propose a novel framework that interprets LLMs as advocates within an ensemble of interacting agents, allowing them to defend their answers and reach conclusions through a judge and jury system. This approach offers a more dynamic and comprehensive evaluation process compared to traditional human-based assessments or automated metrics. We discuss the motivation behind this framework, its key components, and comparative advantages. We also present a probabilistic model to evaluate the error reduction achieved by iterative advocate systems. Finally, we outline experiments to validate the effectiveness of multi-advocate architectures and discuss future research directions.

📄 PDF Abstract BibTeX arXiv:2410.04663

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

2026-06-26 · Ting Ma, Xiufeng Huang, Benlei Cui, Xiaowen Xu 외 arxiv

As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We argue that the essence of safety is adversarial: many failures arise no…

Adversarial RobustnessReinforcement Learning

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation

2026-08-04 · Saqib Shouqi, Abdullah Nazly, Januki Wanniarachchi, Ravisha De Alwis arxiv

Role-Playing Language Agents (RPLAs) are increasingly deployed in high-stakes applications such as healthcare assistance, customer support, and education, where maintaining consistent personas, ethical constraints, and b…

VoiceAgentBench: Are Voice Assistants ready for agentic tasks?

2025-10-09 · Dhruv Jain, Harshit Shukla, Gautam Rajeev, Ashish Kulkarni 외 arxiv

Large scale Speech Language Models have enabled voice assistants capable of understanding natural spoken queries and performing complex tasks. However, existing speech benchmarks largely focus on isolated capabilities su…

Adversarial RobustnessQuestion AnsweringVoice Conversion

Dissecting Adversarial Robustness of Multimodal LM Agents

2024-06-18 · Chen Henry Wu, Rishi Shah, Jing Yu Koh, Ruslan Salakhutdinov 외

As language models (LMs) are used to build autonomous agents in real environments, ensuring their adversarial robustness becomes a critical challenge. Unlike chatbots, agents are compound systems with multiple components…

Adversarial RobustnessAdversarial Text

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

2026-04-15 · Jiachen Qian, Zhaolu Kang arxiv

The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their adversarial robustness remains underexplored. We identify Visual Domi…

Adversarial RobustnessAdversarial Attack