paper-with-me

홈 › Papers

AI Agents for Inventory Control: Human-LLM-OR Complementarity

2026-02-13 · Jackie Baek, Yaopeng Fu, Will Ma, Tianyi Peng arxiv

Inventory control is a fundamental operations problem in which ordering decisions are traditionally guided by theoretically grounded operations research (OR) algorithms. However, such algorithms often rely on rigid modeling assumptions and can perform poorly when demand distributions shift or relevant contextual information is unavailable. Recent advances in large language models (LLMs) have generated interest in AI agents that can reason flexibly and incorporate rich contextual signals, but it remains unclear how best to incorporate LLM-based methods into traditional decision-making pipelines. We study how OR algorithms, LLMs, and humans can interact and complement each other in a multi-period inventory control setting. We construct InventoryBench, a benchmark of over 1,000 inventory instances spanning both synthetic and real-world demand data, designed to stress-test decision rules under demand shifts, seasonality, and uncertain lead times. Through this benchmark, we find that OR-augmented LLM methods outperform either method in isolation, suggesting that these methods are complementary rather than substitutes. We further investigate the role of humans through a controlled classroom experiment that embeds LLM recommendations into a human-in-the-loop decision pipeline. Contrary to prior findings that human-AI collaboration can degrade performance, we show that, on average, human-AI teams achieve higher profits than either humans or AI agents operating alone. Beyond this population-level finding, we formalize an individual-level complementarity effect and derive a distribution-free lower bound on the fraction of individuals who benefit from AI collaboration; empirically, we find this fraction to be substantial.

📄 PDF Abstract BibTeX arXiv:2602.12631

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AIM-Bench: Evaluating Decision-making Biases of Agentic LLM as Inventory Manager

2025-08-15 · Xuhua Zhao, Yuxuan Xie, Caihua Chen, Yuxiang Sun arxiv

Recent advances in mathematical reasoning and the long-term planning capabilities of large language models (LLMs) have precipitated the development of agents, which are being increasingly leveraged in business operations…

Mathematical Reasoning

Controllable Complementarity: Subjective Preferences in Human-AI Collaboration

2025-03-07 · Chase McDonald, Cleotilde Gonzalez

Research on human-AI collaboration often prioritizes objective performance. However, understanding human subjective preferences is essential to improving human-AI complementarity and human experiences. We investigate hum…

Collaborative Causal Sensemaking: Closing the Complementarity Gap in Human-AI Decision Support

2025-12-08 · Raunak Jain arxiv

LLM-based agents are increasingly deployed for expert decision support, yet human-AI teams in high-stakes settings do not yet reliably outperform the best individual. We argue this complementarity gap reflects a fundamen…

Tree-Based Formalization of Multi-Agent Complementarity in Human-AI Interactions

2026-06-03 · Andrea Ferrario arxiv

Complementarity is the case in which a human--AI interaction (HAI) outperforms the best prediction benchmark available among its members. Although this idea is central in HAI research, formal work on complementarity rema…

Binary Classification

Applying Psychometrics to Large Language Model Simulated Populations: Recreating the HEXACO Personality Inventory Experiment with Generative Agents

2025-08-01 · Sarah Mercer, Daniel P. Martin, Phil Swatton arxiv

Generative agents powered by Large Language Models demonstrate human-like characteristics through sophisticated natural language interactions. Their ability to assume roles and personalities based on predefined character…