(1)
Enhancing Stable Behavioral Imitation through Adaptive Reward Weighting in TD3-SAC-GAIL. JECIR 2026, 4 (2), 57-74. https://doi.org/10.63544/tvffe878.