Adaptive Cloud-Based Big Data Analytics Model for Sustainable Supply Chain Management
Main Article Content
Abstract
Freight decisions are simultaneously cost decisions, service decisions and disclosable emission decisions, yet in almost every deployed analytics stack the optimiser sees only the first two. We present a cloud-native architecture that moves the emission calculation onto the operational data path, so that a carbon figure can shape a fulfilment decision rather than document one after the fact. The design combines a streaming emission operator conformant with ISO 14083, a leakage-pruned predictive layer, an online assignment rule whose carbon budget is enforced by a Lagrangian multiplier, and an autoscaler whose set point derives from a queueing-stability condition. We evaluate on the DataCo Global supply chain release: 180,501 usable order lines, geocoded to city level for 99.12% of records, with modal emission factors taken from the EPA GHG Emission Factors Hub. Three findings stand out. Cost and carbon turn out to be aligned rather than opposed in this network, because surface transport is both the cheapest and the cleanest option, so the binding trade-off is service level against emissions rather than money against emissions. Against a service-oriented baseline the controller cuts transport emissions by 35.0% for 1.54 days of additional mean lead time, holding SLA attainment at 100%, and the reduction is sign-stable across every declared assumption we sweep. The platform’s own carbon is negligible at this scale, 0.047% of logistics emissions, with the incremental cost of running the controller amounting to 0.0039% of what it saves. Finally, and independently of the framework, models trained on the unpruned DataCo frame achieve perfect separation because four columns are realised only after delivery; removing them lowers PR-AUC by 0.171 (95% CI 0.166 to 0.176).
Article Details
Issue
Section

This work is licensed under a Creative Commons Attribution 4.0 International License.
Deprecated: json_decode(): Passing null to parameter #1 ($json) of type string is deprecated in /home/u273879158/domains/mesopotamian.press/public_html/journals/plugins/generic/citations/CitationsPlugin.php on line 68
How to Cite
References
[1] D. Ivanov and A. Dolgui, “A digital supply chain twin for managing the disruption risks and resilience in the era of Industry 4.0,” Production Planning & Control, vol. 32, no. 9, pp. 775–788, 2021, doi: 10.1080/09537287.2020.1768450.
[2] D. Ivanov, “Intelligent digital twin (iDT) for supply chain stress-testing, resilience, and viability,” International Journal of Production Economics, vol. 263, Art. no. 108938, 2023, doi: 10.1016/j.ijpe.2023.108938.
[3] J. Mageto, “Big data analytics in sustainable supply chain management: A focus on manufacturing supply chains,” Sustainability, vol. 13, no. 13, Art. no. 7101, 2021, doi: 10.3390/su13137101.
[4] World Resources Institute and World Business Council for Sustainable Development, Corporate Value Chain (Scope 3) Accounting and Reporting Standard. Washington, DC, USA, 2011.
[5] International Organization for Standardization, ISO 14083:2023—Greenhouse Gases: Quantification and Reporting of Greenhouse Gas Emissions Arising From Transport Chain Operations. Geneva, Switzerland, 2023.
[6] Smart Freight Centre, Global Logistics Emissions Council (GLEC) Framework for Logistics Emissions Accounting and Reporting, Version 3.0. Amsterdam, The Netherlands, 2023.
[7] A. Singh, S. Kumari, H. Malekpoor, and N. Mishra, “Big data cloud computing framework for low carbon supplier selection in the beef supply chain,” Journal of Cleaner Production, vol. 202, pp. 139–149, 2018, doi: 10.1016/j.jclepro.2018.07.236.
[8] E. Masanet, A. Shehabi, N. Lei, S. Smith, and J. Koomey, “Recalibrating global data center energy-use estimates,” Science, vol. 367, no. 6481, pp. 984–986, 2020, doi: 10.1126/science.aba3758.
[9] A. Radovanović, R. Koningstein, I. Schneider, B. Chen, A. Duarte, B. Roy, D. Xiao, M. Haridasan, P. Hung, N. Care, S. Talukdar, E. Mullen, K. Smith, M. Cottman, and W. Cirne, “Carbon-aware computing for datacenters,” IEEE Transactions on Power Systems, vol. 38, no. 2, pp. 1270–1280, Mar. 2023, doi: 10.1109/TPWRS.2022.3173250.
[10] Z. Cao, X. Zhou, H. Hu, Z. Wang, and Y. Wen, “Toward a systematic survey for carbon neutral data centers,” IEEE Communications Surveys & Tutorials, vol. 24, no. 2, pp. 895–936, 2022, doi: 10.1109/COMST.2022.3161275.
[11] F. Constante, F. Silva, and A. Pereira, “DataCo SMART SUPPLY CHAIN FOR BIG DATA ANALYSIS,” Mendeley Data, V5, 2019, doi: 10.17632/8gx2fvg2k6.5.
[12] U.S. Environmental Protection Agency, Emission Factors for Greenhouse Gas Inventories (GHG Emission Factors Hub). Washington, DC, USA, Jan. 2025. [Online]. Available: https://www.epa.gov/climateleadership/ghg-emission-factors-hub
[13] J. Dean and S. Ghemawat, “MapReduce: Simplified data processing on large clusters,” Communications of the ACM, vol. 51, no. 1, pp. 107–113, 2008, doi: 10.1145/1327452.1327492.
[14] M. Zaharia, R. S. Xin, P. Wendell, T. Das, M. Armbrust, A. Dave, X. Meng, J. Rosen, S. Venkataraman, M. J. Franklin, A. Ghodsi, J. Gonzalez, S. Shenker, and I. Stoica, “Apache Spark: A unified engine for big data processing,” Communications of the ACM, vol. 59, no. 11, pp. 56–65, 2016, doi: 10.1145/2934664.
[15] J. Kreps, N. Narkhede, and J. Rao, “Kafka: A distributed messaging system for log processing,” in Proc. NetDB Workshop, Athens, Greece, 2011, pp. 1–7.
[16] M. Armbrust, T. Das, J. Torres, B. Yavuz, S. Zhu, R. Xin, A. Ghodsi, I. Stoica, and M. Zaharia, “Structured Streaming: A declarative API for real-time applications in Apache Spark,” in Proc. ACM SIGMOD Int. Conf. Management of Data, Houston, TX, USA, 2018, pp. 601–613, doi: 10.1145/3183713.3190664.
[17] P. Carbone, A. Katsifodimos, S. Ewen, V. Markl, S. Haridi, and K. Tzoumas, “Apache Flink: Stream and batch processing in a single engine,” IEEE Data Engineering Bulletin, vol. 38, no. 4, pp. 28–38, 2015.
[18] N. R. Herbst, S. Kounev, and R. Reussner, “Elasticity in cloud computing: What it is, and what it is not,” in Proc. 10th Int. Conf. Autonomic Computing (ICAC), San Jose, CA, USA, 2013, pp. 23–27.
[19] T. Lorido-Botrán, J. Miguel-Alonso, and J. A. Lozano, “A review of auto-scaling techniques for elastic applications in cloud environments,” Journal of Grid Computing, vol. 12, no. 4, pp. 559–592, 2014, doi: 10.1007/s10723-014-9314-7.
[20] L. A. Barroso and U. Hölzle, “The case for energy-proportional computing,” Computer, vol. 40, no. 12, pp. 33–37, 2007, doi: 10.1109/MC.2007.443.
[21] B. Acun, B. Lee, F. Kazhamiaka, K. Maeng, U. Gupta, M. Chakkaravarthy, D. Brooks, and C.-J. Wu, “Carbon Explorer: A holistic framework for designing carbon aware datacenters,” in Proc. 28th ACM Int. Conf. Architectural Support for Programming Languages and Operating Systems (ASPLOS), vol. 2, Vancouver, BC, Canada, 2023, pp. 118–132, doi: 10.1145/3575693.3575754.
[22] Y. Ran, H. Hu, X. Zhou, and Y. Wen, “DeepEE: Joint optimization of job scheduling and cooling control for data center energy efficiency using deep reinforcement learning,” in Proc. IEEE 39th Int. Conf. Distributed Computing Systems (ICDCS), Dallas, TX, USA, 2019, doi: 10.1109/ICDCS.2019.00070.
[23] L. Breiman, “Random forests,” Machine Learning, vol. 45, no. 1, pp. 5–32, 2001, doi: 10.1023/A:1010933404324.
[24] T. Chen and C. Guestrin, “XGBoost: A scalable tree boosting system,” in Proc. 22nd ACM SIGKDD Int. Conf. Knowledge Discovery and Data Mining, San Francisco, CA, USA, 2016, pp. 785–794, doi: 10.1145/2939672.2939785.
[25] G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu, “LightGBM: A highly efficient gradient boosting decision tree,” in Advances in Neural Information Processing Systems 30, Long Beach, CA, USA, 2017, pp. 3146–3154.
[26] S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural Computation, vol. 9, no. 8, pp. 1735–1780, 1997, doi: 10.1162/neco.1997.9.8.1735.
[27] S. M. Lundberg and S.-I. Lee, “A unified approach to interpreting model predictions,” in Advances in Neural Information Processing Systems 30, Long Beach, CA, USA, 2017, pp. 4765–4774.
[28] R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction, 2nd ed. Cambridge, MA, USA: MIT Press, 2018.
[29] W. B. Powell, Approximate Dynamic Programming: Solving the Curses of Dimensionality. Hoboken, NJ, USA: Wiley, 2007.
[30] V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, A. A. Rusu, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature, vol. 518, no. 7540, pp. 529–533, 2015, doi: 10.1038/nature14236.
[31] J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Proximal policy optimization algorithms,” arXiv:1707.06347, 2017.
[32] J. Gijsbrechts, R. N. Boute, J. A. Van Mieghem, and D. J. Zhang, “Can deep reinforcement learning improve inventory management? Performance on lost sales, dual-sourcing, and multi-echelon problems,” Manufacturing & Service Operations Management, vol. 24, 2022, doi: 10.1287/msom.2021.1064.
[33] A. Oroojlooyjadid, M. Nazari, L. V. Snyder, and M. Takáč, “A deep Q-network for the beer game: Deep reinforcement learning for inventory optimization,” Manufacturing & Service Operations Management, vol. 24, no. 1, pp. 285–304, 2022, doi: 10.1287/msom.2020.0939.
[34] Z. Kegenbekov and I. Jackson, “Adaptive supply chain: Demand–supply synchronization using deep reinforcement learning,” Algorithms, vol. 14, no. 8, Art. no. 240, 2021, doi: 10.3390/a14080240.
[35] R. Tian, M. Lu, H. P. Wang, B. Wang, and Q. X. Tang, “IACPPO: A deep reinforcement learning-based model for warehouse inventory replenishment,” Computers & Industrial Engineering, vol. 187, Art. no. 109829, 2024, doi: 10.1016/j.cie.2023.109829.
[36] K. Wang, C. Long, D. Ong, J. Zhang, and X.-M. Yuan, “Single-site perishable inventory management under uncertainties: A deep reinforcement learning approach,” IEEE Transactions on Knowledge and Data Engineering, vol. 35, no. 10, pp. 10807–10813, 2023, doi: 10.1109/TKDE.2023.3241087.
[37] J. Zhang, Y. Zhao, W. Xue, and J. Li, “Vehicle routing problem with fuel consumption and carbon emission,” International Journal of Production Economics, vol. 170, pp. 234–242, 2015, doi: 10.1016/j.ijpe.2015.09.031.
[38] K. Deb, A. Pratap, S. Agarwal, and T. Meyarivan, “A fast and elitist multiobjective genetic algorithm: NSGA-II,” IEEE Transactions on Evolutionary Computation, vol. 6, no. 2, pp. 182–197, Apr. 2002, doi: 10.1109/4235.996017.
[39] E. Zitzler and L. Thiele, “Multiobjective evolutionary algorithms: A comparative case study and the strength Pareto approach,” IEEE Transactions on Evolutionary Computation, vol. 3, no. 4, pp. 257–271, Nov. 1999, doi: 10.1109/4235.797969.
[40] D. Bertsimas and M. Sim, “The price of robustness,” Operations Research, vol. 52, no. 1, pp. 35–53, 2004, doi: 10.1287/opre.1030.0065.
[41] Y. Yang, W. Ingwersen, T. Hawkins, M. Srocka, and D. Meyer, “USEEIO: A new and transparent United States environmentally-extended input-output model,” Journal of Cleaner Production, vol. 158, pp. 308–318, 2017, doi: 10.1016/j.jclepro.2017.04.150.
[42] W. W. Ingwersen, M. Li, B. Young, J. Vendries, and C. Birney, “USEEIO v2.0, the US environmentally-extended input-output model v2.0,” Scientific Data, vol. 9, Art. no. 194, 2022, doi: 10.1038/s41597-022-01293-7.
[43] W. Ingwersen, “Supply Chain Greenhouse Gas Emission Factors v1.3 by NAICS-6,” U.S. Environmental Protection Agency, Washington, DC, USA, 2024. [Online]. Available: https://catalog.data.gov/dataset/supply-chain-greenhouse-gas-emission-factors-v1-3-by-naics-6
[44] J. Heiss, T. Oegel, M. Shakeri, and S. Tai, “Verifiable carbon accounting in supply chains,” IEEE Transactions on Services Computing, 2023, doi: 10.1109/TSC.2023.3332831.
[45] T. K. Agrawal, V. Kumar, R. Pal, L. Wang, and Y. Chen, “Blockchain-based framework for supply chain traceability: A case example of textile and clothing industry,” Computers & Industrial Engineering, vol. 154, Art. no. 107130, 2021, doi: 10.1016/j.cie.2021.107130.
[46] H. B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. Agüera y Arcas, “Communication-efficient learning of deep networks from decentralized data,” in Proc. 20th Int. Conf. Artificial Intelligence and Statistics (AISTATS), Fort Lauderdale, FL, USA, 2017, pp. 1273–1282.
[47] Q. Li, Y. Cui, T. Song, and L. Zheng, “Federated multiagent actor–critic learning task offloading in intelligent logistics,” IEEE Internet of Things Journal, vol. 10, no. 13, pp. 11696–11707, 2023, doi: 10.1109/JIOT.2023.3244557.
[48] J. D. C. Little, “A proof for the queuing formula: L = λW,” Operations Research, vol. 9, no. 3, pp. 383–387, 1961, doi: 10.1287/opre.9.3.383.
[49] U.S. Environmental Protection Agency, Emissions & Generation Resource Integrated Database (eGRID) With 2023 Data, Summary Tables, Revision 2. Washington, DC, USA, Jun. 2025. [Online]. Available: https://www.epa.gov/egrid/summary-data
[50] K. R. Ahmed, M. E. Ansari, M. N. Ahsan, A. Rohan, M. B. Uddin, and M. A. H. Rivin, “Deep learning framework for interpretable supply chain forecasting using SOM ANN and SHAP,” Scientific Reports, vol. 15, Art. no. 26355, 2025, doi: 10.1038/s41598-025-11510-z.