[1]
H. Hoang, T. Mai, and P. Varakantham, “Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning”, AAAI, vol. 38, no. 11, pp. 12439–12447, Mar. 2024.