paper-with-me

Papers

GAIA: A General Agency Interaction Architecture for LLM-Human B2B Negotiation & Screening

2025-11-09 · Siming Zhao, Qi Li arxiv

Organizations are increasingly exploring delegation of screening and negotiation tasks to AI systems, yet deployment in high-stakes B2B settings is constrained by governance: preventing unauthorized commitments, ensuring sufficient information before bargaining, and maintaining effective human oversight and auditability. Prior work on large language model negotiation largely emphasizes autonomous bargaining between agents and omits practical needs such as staged information gathering, explicit authorization boundaries, and systematic feedback integration. We propose GAIA, a governance-first framework for LLM-human agency in B2B negotiation and screening. GAIA defines three essential roles - Principal (human), Delegate (LLM agent), and Counterparty - with an optional Critic to enhance performance, and organizes interactions through three mechanisms: information-gated progression that separates screening from negotiation; dual feedback integration that combines AI critique with lightweight human corrections; and authorization boundaries with explicit escalation paths. Our contributions are fourfold: (1) a formal governance framework with three coordinated mechanisms and four safety invariants for delegation with bounded authorization; (2) information-gated progression via task-completeness tracking (TCI) and explicit state transitions that separate screening from commitment; (3) dual feedback integration that blends Critic suggestions with human oversight through parallel learning channels; and (4) a hybrid validation blueprint that combines automated protocol metrics with human judgment of outcomes and safety. By bridging theory and practice, GAIA offers a reproducible specification for safe, efficient, and accountable AI delegation that can be instantiated across procurement, real estate, and staffing workflows.

📄 PDF Abstract BibTeX arXiv:2511.06262

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding

2026-08-12 · Fan Zhang, Guangming Yao, Jinyang Wu, Hao Wu 외 hf

Video understanding is a fundamental task for evaluating the capabilities of multimodal large language models (MLLMs). However, existing leading models have already achieved approximately 90% accuracy on the Video-MME le…

Video Question Answering

GAIA: a benchmark for General AI Assistants

2023-11-21 · Grégoire Mialon, Clémentine Fourrier, Craig Swift, Thomas Wolf 외

We introduce GAIA, a benchmark for General AI Assistants that, if solved, would represent a milestone in AI research. GAIA proposes real-world questions that require a set of fundamental abilities such as reasoning, mult…

Philosophy

Intent-aligned AI systems deplete human agency: the need for agency foundations research in AI safety

2023-05-30 · Catalin Mitelut, Ben Smith, Peter Vamplew

The rapid advancement of artificial intelligence (AI) systems suggests that artificial general intelligence (AGI) systems may soon arrive. Many researchers are concerned that AIs and AGIs will harm humans via intentional…

Transforming Agency. On the mode of existence of Large Language Models

2024-07-15 · Xabier E. Barandiaran, Lola S. Almendros

This paper investigates the ontological characterization of Large Language Models (LLMs) like ChatGPT. Between inflationary and deflationary accounts, we pay special attention to their status as agents. This requires exp…

GAIA -- A Large Language Model for Advanced Power Dispatch

2024-08-07 · Yuheng Cheng, Huan Zhao, Xiyuan Zhou, Junhua Zhao 외

Power dispatch is essential for providing stable, cost-effective, and eco-friendly electricity to society. However, traditional methods falter as power systems grow in scale and complexity, struggling with multitasking, …

Decision MakingLanguage ModelingLanguage ModellingLarge Language Model+1