A Reinforcement Learning Approach for Initialization of Column Generation with Application to Aircraft Recovery Problem
10th International Conference on Cyber Security and Information Engineering (ICCSIE)
This paper formulates the generation of initial columns as a sequential decision-making problem. The policy combines a Graph Attention Network with a Pointer Network and PPO to encode flight topology and generate recovery plans that improve the starting point for column generation.