Hoang, H., Mai, T., & Varakantham, P. (2024). Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning. Proceedings of the AAAI Conference on Artificial Intelligence, 38(11), 12439–12447. https://doi.org/10.1609/aaai.v38i11.29136