[1]
T. He, W. Zhao, and C. Liu, “AutoCost: Evolving Intrinsic Cost for Zero-Violation Reinforcement Learning”, AAAI, vol. 37, no. 12, pp. 14847–14855, Jun. 2023.