(1)
Hoang, H.; Mai, T.; Varakantham, P. Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning. AAAI 2024, 38, 12439-12447.