Learning and interpreting policies for simultaneous entanglement requests in quantum networks
This work forms a Markov Decision Process for the problem and uses double deep Q-networks with Message Passing Neural Networks, experience replay buffers, and curriculum training to obtain policies, indicating a promising method for interpretable policy extraction for large quantum networks, where direct training becom...