“Enhancing Stable Behavioral Imitation through Adaptive Reward Weighting in TD3-SAC-GAIL”. 2026. Journal of Engineering and Computational Intelligence Review 4 (2): 57-74. https://doi.org/10.63544/tvffe878.