[1]
“Enhancing Stable Behavioral Imitation through Adaptive Reward Weighting in TD3-SAC-GAIL”, JECIR, vol. 4, no. 2, pp. 57–74, Aug. 2026, doi: 10.63544/tvffe878.