
www.rsisinternational.org
INTERNATIONAL JOURNAL OF LATEST TECHNOLOGY IN ENGINEERING,
MANAGEMENT & APPLIED SCIENCE (IJLTEMAS)
ISSN 2278-2540 | DOI: 10.51583/IJLTEMAS | Volume XV, Issue VI, June 2026
ACKNOWLEDGEMENTS
The authors acknowledge the developers of the OhioT1DM dataset for making the clinical data publicly
available for research. The authors also appreciate the constructive comments and suggestions from anonymous
reviewers, which helped improve the quality of this manuscript.
REFERENCES
1. Bolland, A., Lambrechts, G., & Ernst, D. (2024). Off-policy maximum entropy rl with future state and
action visitation measures. arXiv preprint arXiv:2412.06655.
2. Dénes-Fazakas, L., Szilágyi, L., Kovács, L., De Gaetano, A., & Eigner, G. (2024). Reinforcement learning:
a paradigm shift in personalized blood glucose management for diabetes. Biomedicines, 12(9), 2143.
3. Elsayed, N. A., Aleppo, G., Bannuru, R. R., Bruemmer, D., Collins, B. S., Ekhlaspour, L., & American
Diabetes Association Professional Practice Committee. (2024). 16. Diabetes Care in the Hospital:
Standards of Care in Diabetes—2024. Diabetes Care, 47.
4. Haarnoja, T., Zhou, A., Abbeel, P., & Levine, S. (2018). Soft Actor-Critic: Off-Policy Maximum Entropy
Deep Reinforcement Learning with a Stochastic Actor. Proceedings of the 35th International Conference
on Machine Learning. https://doi.org/10.48550/arXiv.1801.01290
5. Lei, J., Sun, X., Li, Y., Li, K., Zhang, S., Zeng, H., & Zhang, Y. (2024, August). An Improved Adaptive
Glucose Control Approach for Type 1 Diabetes with Temporal Dependence. In 2024 IEEE 9th
International Conference on Computational Intelligence and Applications (ICCIA) (pp. 209-214). IEEE.
6. Manas, S., Pillai, G. N. & Gupta, M. K. (2023). Improved Soft Actor-Critic: Reducing Bias and Estimation
Error for Fast Learning. IEEE International Student’s Conference on Electrical, Electronics and Computer
Science (SCEECS), 1 - 9,2023,doi:10.1109/SCEECS57921.
7. Milton T. & Lieck R. (2024). Fully-Automated Patient-Agnostic Diabetes Management with Deep
Reinforcement Learning. IEEE International Conference on Bioinformatics and Biomedicine (BIBM),
1085-1091.
8. Mnih, V. , Adria, P. B. , M. Mehdi, G. Alex, H. Tim, P. L. Timothy, S. David & K. Koray, (2016).
Asynchronous Methods for Deep Reinforcement Learning. Proceedings of the 33
rd
International
Conference on Machine Learning, New York. NY USA. JLMR. W & CP, 48,
doi:10.48550/arXiv.1602.01783.
9. Parveen, A. (2021). A Personalized Deep Learning Approach for Blood Glucose Prediction in People with
T1DM (Master's thesis, Stevens Institute of Technology).
10. Schulman, J., Wolski, F., Dhariwal, P., Radford, A., & Klimov, O. (2017). Proximal Policy Optimization
Algorithms. arXiv preprint arXiv:1707.06347.
11. Singh, R., & Raj R. R. (2023). Optimizing Glycemic Control in Type 1 Diabetic Patients using a Deep
Learning-Based Artificial Pancreas with a Secure Glucagon and Insulin Delivery System. bioRxiv, 12. doi:
https://doi.org/10.1101/2023.12.07.566476.
12. Tuomas, H., Aurick, Z. Pieter, A. & Sergey, L. (2018). Soft Actor-Critic: Off - policy Maximum Entropy
Deep reinforcement Learning with a Stochastic Actor. International Journal of Research and Innovation
in Social Sciences, doi:10.48550/arXiv.1801.01290,
https://www.researchgate.net/publication/322306636_Soft_Actor-Critic_Off-
policy_Maximum_Entropy_Deep_Reinforcement_Learning_with_a_Stochastic_Actor
13. Zhao, X., Ding, S., An, Y., & Jia, W. (2019). Applications of asynchronous deep reinforcement learning
based on dynamic updating weights: X. Zhao et al. Applied Intelligence, 49(2), 581-591.
14. Zheng, M., Zhang, J., Zhan, C., Ren, X., & Lü, S. (2025). Proximal policy optimization with reward-based
prioritization. Expert Systems with Applications, 283, 127659.