Cooperative Multi-Agent Reinforcement Learning for Inventory Management
With Reinforcement Learning (RL) for inventory management (IM) being a nascent field of research, approaches tend to be limited to simple, linear environments with implementations that are minor modifications of off-the-shelf RL algorithms. Scaling these simplistic environments to a real-world supply chain comes with a few challenges such as: minimizing the computational requirements of the environment, specifying agent configurations that are representative of dynamics at real world stores and warehouses, and specifying a reward framework that encourages desirable behavior across the whole supply chain. In this work, we present a system with a custom GPU-parallelized environment that consists of one warehouse and multiple stores, a novel architecture for agent-environment dynamics incorporating enhanced state and action spaces, and a shared reward specification that seeks to optimize for a large retailer's supply chain needs. Each vertex in the supply chain graph is an independent agent that, based on its own inventory, able to place replenishment orders to the vertex upstream. The warehouse agent, aside from placing orders from the supplier, has the special property of also being able to constrain replenishment to stores downstream, which results in it learning an additional allocation sub-policy. We achieve a system that outperforms standard inventory control policies such as a base-stock policy and other RL-based specifications for 1 product, and lay out a future direction of work for multiple products.
Code (0)
등록된 구현이 없습니다.
Tasks
GPUManagementMulti-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
MARLIM: Multi-Agent Reinforcement Learning for Inventory Management
Maintaining a balance between the supply and demand of products by optimizing replenishment decisions is one of the most important challenges in the supply chain industry. This paper presents a novel reinforcement learni…
ManagementMulti-agent Reinforcement Learningreinforcement-learningReinforcement LearningA Versatile Multi-Agent Reinforcement Learning Benchmark for Inventory Management
Multi-agent reinforcement learning (MARL) models multiple agents that interact and learn within a shared environment. This paradigm is applicable to various industrial scenarios such as autonomous driving, quantitative t…
Autonomous DrivingManagementMulti-agent Reinforcement Learningreinforcement-learning+1InvAgent: A Large Language Model based Multi-Agent System for Inventory Management in Supply Chains
Supply chain management (SCM) involves coordinating the flow of goods, information, and finances across various entities to deliver products efficiently. Effective inventory management is crucial in today's volatile and …
Decision MakingLanguage ModelingLanguage ModellingLarge Language Model+2An Analysis of Multi-Agent Reinforcement Learning for Decentralized Inventory Control Systems
Most solutions to the inventory management problem assume a centralization of information that is incompatible with organisational constraints in real supply chain networks. The inventory management problem is a well-kno…
ManagementMulti-agent Reinforcement Learningreinforcement-learningMulti-Agent Reinforcement Learning with Shared Resources for Inventory Management
In this paper, we consider the inventory management (IM) problem where we need to make replenishment decisions for a large number of stock keeping units (SKUs) to balance their supply and demand. In our setting, the cons…
ManagementMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1