Chen, Z., & Shao, X. (2026). Reinforcement Learning Exploration Strategy Based on Performance Feedback: Asymptotic Convergence Proof and Experimental Validation. Mathematics (2227-7390), 14(9), 1487. https://doi.org/10.3390/math14091487
Chicago Style (17th ed.) CitationChen, Zheng, and Xinhui Shao. "Reinforcement Learning Exploration Strategy Based on Performance Feedback: Asymptotic Convergence Proof and Experimental Validation." Mathematics (2227-7390) 14, no. 9 (2026): 1487. https://doi.org/10.3390/math14091487.
MLA (9th ed.) CitationChen, Zheng, and Xinhui Shao. "Reinforcement Learning Exploration Strategy Based on Performance Feedback: Asymptotic Convergence Proof and Experimental Validation." Mathematics (2227-7390), vol. 14, no. 9, 2026, p. 1487, https://doi.org/10.3390/math14091487.