paper-with-me

홈 › Papers

End-to-end Deep Reinforcement Learning for Stochastic Multi-objective Optimization in C-VRPTW

2025-12-01 · Abdo Abouelrous, Laurens Bliek, Yaoxin Wu, Yingqian Zhang arxiv

In this work, we consider learning-based applications in routing to solve a Vehicle Routing variant characterized by stochasticity and multiple objectives. Such problems are representative of practical settings where decision-makers have to deal with uncertainty in the operational environment as well as multiple conflicting objectives due to different stakeholders. We specifically consider travel time uncertainty. We also consider two objectives, total travel time and route makespan, that jointly target operational efficiency and labor regulations on shift length, although different objectives could be incorporated. Learning-based methods offer earnest computational advantages as they can repeatedly solve problems with limited interference from the decision-maker. We specifically focus on end-to-end deep learning models that leverage the attention mechanism and multiple solution trajectories. These models have seen several successful applications in routing problems. However, since travel times are not a direct input to these models due to the large dimensions of the travel time matrix, accounting for uncertainty is a challenge, especially in the presence of multiple objectives. In turn, we propose a model that simultaneously addresses stochasticity and multi-objectivity and provide a refined training mechanism for this model through scenario clustering to reduce training time. Our results show that our model is capable of constructing a Pareto Front of good quality within acceptable run times compared to three baselines.

📄 PDF Abstract BibTeX arXiv:2512.01518

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

MODRL/D-EL: Multiobjective Deep Reinforcement Learning with Evolutionary Learning for Multiobjective Optimization

2021-07-16 · Yongxin Zhang, Jiahai Wang, Zizhen Zhang, Yalan Zhou

Learning-based heuristics for solving combinatorial optimization problems has recently attracted much academic attention. While most of the existing works only consider the single objective problem with simple constraint…

Combinatorial OptimizationDeep Reinforcement LearningMultiobjective Optimizationreinforcement-learning+1

Multiobjective Vehicle Routing Optimization with Time Windows: A Hybrid Approach Using Deep Reinforcement Learning and NSGA-II

2024-07-18 · Rixin Wu, Ran Wang, Jie Hao, Qiang Wu 외

This paper proposes a weight-aware deep reinforcement learning (WADRL) approach designed to address the multiobjective vehicle routing problem with time windows (MOVRPTW), aiming to use a single deep reinforcement learni…

Deep Reinforcement LearningMultiobjective Optimization

Deep Reinforcement Learning for Electric Vehicle Routing Problem with Time Windows

2020-10-05 · Bo Lin, Bissan Ghaddar, Jatin Nathwani

The past decade has seen a rapid penetration of electric vehicles (EV) in the market, more and more logistics and transportation companies start to deploy EVs for service provision. In order to model the operations of a …

Deep Reinforcement LearningGraph Embeddingreinforcement-learningReinforcement Learning (RL)

The Static and Stochastic VRPTW with both random Customers and Reveal Times: algorithms and recourse strategies

2017-08-10 · Michael Saint-Guillain, Christine Solnon, Yves Deville

Unlike its deterministic counterpart, static and stochastic vehicle routing problems (SS-VRP) aim at modeling and solving real-life operational problems by considering uncertainty on data. We consider the SS-VRPTW-CR int…

Data-Driven Stochastic VRP: Integration of Forecast Duration into Optimization for Utility Workforce Management

2026-01-12 · Matteo Garbelli arxiv

This paper investigates the integration of machine learning forecasts of intervention durations into a stochastic variant of the Capacitated Vehicle Routing Problem with Time Windows (CVRPTW). In particular, we exploit t…