Supply chain / Operations research2025
Reinforcement Learning for Multi-Echelon Supply Chain Optimization
RL agents that match classical optimal stock policies on fixed lead-times and outperform them when lead-times vary.
- Optimality gap vs Clark-Scarf
- 0.02%
- Lift vs alternatives (varying LT)
- +4.1%