Strategic bidding in freight transport using deep reinforcement learning
This paper presents a multi-agent reinforcement learning algorithm to represent strategic bidding behavior in freight transport markets. Using this algorithm, we investigate whether feasible market equilibriums arise without any central control or communication between agents. Studying behavior in such environments may serve as a stepping stone towards self-organizing logistics systems like the Physical Internet. We model an agent-based environment in which a shipper and a carrier actively learn bidding strategies using policy gradient methods, posing bid- and ask prices at the individual container level. Both agents aim to learn the best response given the expected behavior of the opposing agent. A neutral broker allocates jobs based on bid-ask spreads. Our game-theoretical analysis and numerical experiments focus on behavioral insights. To evaluate system performance, we measure adherence to Nash equilibria, fairness of reward division and utilization of transport capacity. We observe good performance both in predictable, deterministic settings (~95% adherence to Nash equilibria) and highly stochastic environments (~85% adherence). Risk-seeking behavior may increase an agent's reward share, as long as the strategies are not overly aggressive. The results suggest a potential for full automation and decentralization of freight transport markets.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement LearningFairnessMulti-agent Reinforcement LearningPolicy Gradient Methodsreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
STraM: A strategic network design model for national freight transport decarbonization
National freight transport models are valuable tools for assessing the impact of various policies and investments on achieving decarbonization targets under different future scenarios. However, these models struggle to a…
Smart Containers With Bidding Capacity: A Policy Gradient Algorithm for Semi-Cooperative Learning
Smart modular freight containers -- as propagated in the Physical Internet paradigm -- are equipped with sensors, data storage capability and intelligence that enable them to route themselves from origin to destination w…
An exact and two heuristic strategies for truthful bidding in combinatorial transport auctions
To support a freight carrier in a combinatorial transport auction, we proposes an exact and two heuristic strategies for bidding on subsets of requests. The exact bidding strategy is based on the concept of elementary re…
Systemic Decarbonization of Road Freight Transport: A Comprehensive Total Cost of Ownership Model
The decarbonization of road freight transport is crucial for reducing greenhouse gas emissions (GHG) and achieving climate neutrality goals. This study develops a comprehensive Total Cost of Ownership (TCO) model to eval…
Real-Time Bidding with Multi-Agent Reinforcement Learning in Display Advertising
Real-time advertising allows advertisers to bid for each impression for a visiting user. To optimize specific goals such as maximizing revenue and return on investment (ROI) led by ad placements, advertisers not only nee…
ClusteringMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1