Berto, F., Hua, C., Park, J., Luttmann, L., Ma, Y., Bu, F., Wang, J., et al. (2025). RL4CO: An Extensive Reinforcement Learning for Combinatorial Optimization Benchmark.
Proceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining, 5278-5289. New York: Association of Computing Machinery.
https://doi.org/10.1145/3711896.3737433