Constraint-Aware Machine Learning for Ensuring Feasible Predictions in Operational Data Science

Main Article Content

  Lau Meng Cheng
  Adolf Asih Suprianto

Abstract

Background: Large-scale machine learning models require substantial computational resources during training, particularly in GPU-intensive and distributed computing environments. Although modern training pipelines achieve high predictive performance, they often rely on static resource allocation strategies that do not adapt to evolving learning dynamics. This leads to inefficient GPU utilization, increased training time, and unnecessary computational cost, limiting the scalability and sustainability of large-scale AI systems.
Aims: This study aims to improve training efficiency by proposing a Dynamic Resource Allocation (DRA) framework that integrates real-time learning signals into computational resource management. The framework dynamically adjusts GPU allocation based on convergence behavior, enabling efficient alignment between computational demand and model training stages.
Methods: The proposed framework employs a descriptive, analytical, and comparative experimental design, using secondary, confidential machine learning training logs comprising 1,500 training runs. The system integrates a training-monitoring module, a convergence-analysis mechanism, and an adaptive resource controller. Performance evaluation is conducted by comparing static allocation and dynamic allocation strategies using metrics such as training time, GPU utilization, model accuracy, and computational efficiency.
Results: Experimental results demonstrate that the proposed framework significantly improves training efficiency. The dynamic allocation strategy reduces training time by approximately 30–32%, increases GPU utilization by up to 17%, and improves overall computational efficiency without degrading model accuracy. Furthermore, the convergence analysis shows that the proposed method achieves faster, more stable convergence than static allocation strategies.
Conclusion: The findings confirm that integrating training-aware resource allocation into machine learning pipelines significantly enhances both efficiency and sustainability. By dynamically aligning computational resources with learning behavior, the proposed framework reduces wasteful computation while maintaining predictive performance. This approach provides a scalable solution for efficient large-scale model training in cloud and high-performance computing environments.

Article Details

How to Cite
Meng Cheng, L., & Suprianto, A. A. (2026). Constraint-Aware Machine Learning for Ensuring Feasible Predictions in Operational Data Science. International Journal of Advances in Artificial Intelligence and Machine Learning, 3(2), 99–112. https://doi.org/10.58723/ijaaiml.v3i2.796
Section
Articles

References

Abdykadyrov, A., Mussapirova, G., Smailov, N., & Seissenbiyeva, Z. (2026). AI-Driven Dynamic Resource Allocation for Energy-Efficient Optical Fiber Communication Networks : Modeling , Algorithms , and Performance Evaluation. J.. Sens. Actuator Netw, 15(2), 1–26. https://doi.org/10.3390/jsan15020028

Alam, S., Atif, M., & Muhammad, R. (2026). AI-driven resource allocation in cloud computing: a systematic review revealing critical sustainability and evaluation gaps. Computing, 108(67). https://doi.org/10.1007/s00607-026-01655-8

Alla, H., Madhura, K., & Aladakatti, S. S. (2026). A carbon aware job scheduling framework for data center sustainability using deep learning training. Artificial Intelligence, 6(813). https://doi.org/10.1007/s44163-026-02022-4

Bolón-canedo, V., Morán-fernández, L., Cancela, B., Alonso-betanzos, A., & Ai, G. (2024). Neurocomputing A review of green artificial intelligence : Towards a more sustainable future. Neurocomputing, 599(December 2023), 128096. https://doi.org/10.1016/j.neucom.2024.128096

Cheng, Y., Sun, Z., & Turner, M. (2026). Distributed Training Strategies for Reducing Carbon Footprint in Large Scale Model Development. Computer Life, 14(2), 8–15. https://doi.org/10.54097/YHPPK428

Dayeh, S. B., & Mohammed, B. Y. (2023). A Comparative Analytical Descriptive Study. Proceedings of the First International Conference on Legal Sciences: Intellectual Property - Contemporary Problems & Legal Solutions (ICLS-22), 1–14. https://doi.org/10.2139/SSRN.4492812

Duan, J., Zhang, S., Wang, Z., Jiang, L., Qu, W., Hu, Q., Wang, G., Weng, Q., Yan, H., Zhang, X., Qiu, X., Lin, D., Wen, Y., & Jin, X. (2026). Efficient training of large language models on distributed infrastructures : a survey. Vicinagearth, 3(9). https://doi.org/10.1007/s44336-026-00038-z

Falk, S., Corrêa, N. K., Luccioni, S., & Biber-freudenberger, L. (2026). From computation to environmental cost the resource burden of arti fi cial intelligence. Communications Earth & Environment, 7, 1–15. https://doi.org/10.1038/s43247-026-03537-5

Fan, Z., Yan, Z., & Wen, S. (2023). Deep Learning and Artificial Intelligence in Sustainability : A Review of SDGs , Renewable Energy , and Environmental Health. Sustainability, 15(18). https://doi.org/10.3390/SU151813493

Hassan, M. U., Al-awady, A. A., Ali, A., Iqbal, M. M., Akram, M., & Jamil, H. (2024). Smart Resource Allocation in Mobile Cloud Next-Generation Network ( NGN ) Orchestration with Context-Aware Data and Machine Learning for the Cost Optimization of Microservice Applications. Sensors, 24(3). https://doi.org/10.3390/S24030865

He, Y., Wang, Y., Lin, Q., & Li, J. (2022). Meta-Hierarchical Reinforcement Learning (MHRL)-Based Dynamic Resource Allocation for Dynamic Vehicular Networks. IEEE Transactions on Vehicular Technology, 71(4), 3495–3506. https://doi.org/10.1109/TVT.2022.3146439

Kumar, M., Kaur, G., & Rana, P. S. (2025). Performance, portability, productivity, and security in HPC cloud: a systematic literature review. The Journal of Supercomputing, 81(11). https://doi.org/10.1007/S11227-025-07685-X

Liu, Y., He, Z., Xie, X., Liu, A., Li, Z., & Deng, Q. (2026). Data Orchestration Service Placement and Resource Allocation Scheme for Cloud-Edge System. IEEE Transactions on Services Computing, 19(2). https://doi.org/10.1109/TSC.2026.3660225

Lu, J., Postigo-boix, M., Guillén, A. B., De, L. J., & Llopis, C. (2026). FedCAMO : Federated Learning Carbon-Aware Multi-Objective Client Selection. Computer Networks, 286(June), 112486. https://doi.org/10.1016/j.comnet.2026.112486

Ma, L., Cheng, N., Zhou, C., Wang, X., Lu, N., & Zhang, N. (2024). Dynamic Neural Network-Based Resource Management for Mobile Edge Computing in 6G Networks. IEEE Transactions on Cognitive Communications and Networking, 10(3), 953–967. https://doi.org/10.1109/TCCN.2023.3346824

Machado, L. R., Rossato, G. de M., Beck, A. C. S., Jordan, M. G., & Rutzig, M. B. (2026). Energy ‑ aware DVFS ‑ driven workload provisioning in heterogeneous cloud FaaS architectures. The Journal of Supercomputing, 82. https://doi.org/10.1007/s11227-025-08171-0

Manhary, F. N., Mohamed, M. H., & Farouk, M. (2025). A scalable machine learning strategy for resource allocation in database. Scientific Reports, 15(1), 1–17. https://doi.org/10.1038/s41598-025-14962-5

Marmouzi, O., & Oumaira, I. (2026). A Systematic Review of Green and Sustainable AI : Taxonomy , Metrics , Challenges , and Open Research Directions. Sustainability, 18(8). https://doi.org/10.3390/su18084115

Munshi, S., & Fernandez, L. (2026). Carbon-Aware Training Schedules for Machine Learning Models : An Energy-Efficient Green AI Approach. International Journal of Engineering and Information Management, 2(1), 38–51. https://doi.org/10.52756/ijeim.2026.v02.i01.003

Nakaegawa, T. (2022). High-Performance Computing in Meteorology under a Context of an Era of Graphical Processing Units. Computers, 11(7). https://doi.org/10.3390/computers11070114

Narsimhulu, B., & Kumar, T. S. (2026). A hybrid RL – GA – LSTM – AE framework for energy-aware and SLA-driven task scheduling in cloud computing environments. Scientific Reports, 16(1), 1–33. https://doi.org/10.1038/s41598-026-43108-4

Niazmand, V., & Ye, Q. (2025). Joint Task Offloading, DNN Pruning, and Computing Resource Allocation for Fault Detection With Dynamic Constraints in Industrial IoT. IEEE Transactions on Cognitive Communications and Networking, 11(5), 3486–3501. https://doi.org/10.1109/TCCN.2025.3529688

Peykani, P., Emrouznejad, A., Ghanidel, S., & Seyedali, I. J. (2026). Green Artificial Intelligence : A Comprehensive Review of Metrics , Tools , Challenges , Trends , and Future Prospects. Archives of Computational Methods in Engineering. https://doi.org/10.1007/s11831-026-10546-2

Prasetya, A., Herdianto, R., Nur, A., Bella, A., Utama, P., Andika, F., & Drezewski, R. (2026). AI without borders : The rise of cross-disciplinary machine learning. Telematics and Informatics Reports, 21, 100294. https://doi.org/10.1016/j.teler.2026.100294

Rashid, A. Bin, & Kausik, A. K. (2024). AI revolutionizing industries worldwide : A comprehensive overview of its diverse applications. Hybrid Advances, 7, 100277. https://doi.org/10.1016/j.hybadv.2024.100277

Rithani, M., Kumar, R. P., & Doss, S. (2023). A review on big data based on deep neural network approaches. Artificial Intelligence Review, 56, 14765–14801. https://doi.org/10.1007/S10462-023-10512-5

Santos, S., Ottoni, A. L. C., Borgo, R., & Ferreira, D. (2026). A systematic review of Green Machine Learning : practices and challenges for sustainability. Artificial Intelligence Review, 59(132). https://doi.org/10.1007/s10462-026-11515-8 A

Shingne, H., Ghusse, D., Dadiyala, C., & Welekar, R. (2026). Federated deep learning-driven decentralized and cost-aware cloud resource management for load balancing and SLA optimizations. Discover Computing, 29. https://doi.org/10.1007/s10791-026-09988-w

Wang, K., Lu, J., Liu, A., Zhang, G., & Xiong, L. (2023). Evolving Gradient Boost: A Pruning Scheme Based on Loss Improvement Ratio for Learning Under Concept Drift. IEEE Transactions on Cybernetics, 53(4), 2110–2123. https://doi.org/10.1109/TCYB.2021.3109796

Wang, P., Cheng, Y., Peng, Q., Wang, J., & Li, S. (2023). Learning Dynamic Computing Resource Allocation in Convolutional Neural Networks for Wireless Interference Identification. IEEE Transactions on Vehicular Technology, 72(7), 8770–8782. https://doi.org/10.1109/TVT.2023.3244560

Xu, H., & Zhang, N. (2021). Implications of Data Anonymization on the Statistical Evidence of Disparity. Management Science, 68(4). https://doi.org/10.1287/MNSC.2021.4028

Xu, M., Cai, D., Yin, W., Wang, S., Jin, X. I. N., & Liu, X. (2025). Resource-efficient Algorithms and Systems of Foundation Models : A Survey. ACM Computing Surveys, 57(5), 1–39. https://doi.org/10.1145/3706418

Xu, Y., Liu, X., & Cao, X. (2021). Artificial intelligence: A powerful paradigm for scientific research. Innovation, 2(4). https://doi.org/10.1016/j.xinn.2021.100179

Yildirim, E., Hussein, M., Titov, M., & Kilic, O. O. (2026). Predicting runtime and resource utilization of jobs on integrated cloud and HPC systems. Future Generation Computer Systems, 176. https://doi.org/10.1016/J.FUTURE.2025.108230

Zerouali, B., Bailek, N., Tariq, A., Kuriqi, A., Guermoui, M., Alharbi, A. H., Khafaga, D. S., Sayed, E., & El, M. (2024). Enhancing deep learning ‑ based slope stability classification using a novel metaheuristic optimization algorithm for feature selection. Scientific Reports, 14, 1–19. https://doi.org/10.1038/s41598-024-72588-5

Zoller, M.-A., & Huber, M. F. (2021). Benchmark and Survey of Automated Machine Learning Frameworks. Journal of Artificial Intelligence Research, 70, 409–472. https://doi.org/10.1613/JAIR.1.11854