Enhancing Stable Behavioral Imitation through Adaptive Reward Weighting in TD3-SAC-GAIL. (2026). Journal of Engineering and Computational Intelligence Review, 4(2), 57-74. https://doi.org/10.63544/tvffe878