1.
Hoang H, Mai T, Varakantham P. Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning. AAAI [Internet]. 2024 Mar. 24 [cited 2026 Jul. 26];38(11):12439-47. Available from: https://ojs.aaai.org/index.php/AAAI/article/view/29136