· 1 min read
Optimizing Call Center Operations with Reinforcement Learning: PPO vs. Value Iteration
PPO beats classical Value Iteration at call routing in a discrete-event call centre simulation: shortest customer wait, least agent idle time, highest RL reward.