“Enhancing Stable Behavioral Imitation through Adaptive Reward Weighting in TD3-SAC-GAIL” (2026) Journal of Engineering and Computational Intelligence Review, 4(2), pp. 57–74. doi:10.63544/tvffe878.