[1]
A. Levine and S. Feizi, “Goal-Conditioned Q-learning as Knowledge Distillation”, AAAI, vol. 37, no. 7, pp. 8500–8509, Jun. 2023.