PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 24, 20260 citationsOpen Access

Data-Driven Process Optimization for US Supply Chain Resilience Using Machine Learning and SQL

View Full Paper
BSBabul Chandra SarkerKMKamana Parvej MishuMAMohammad Tahmid Ahmed

Key Points

  • The aim is to optimize US supply chain processes using a data-driven framework that integrates machine learning and SQL.
  • Developed an integrated data pipeline using SQL for efficient ETL of multi-modal data.
  • Trained ML models, including Gradient Boosting Regressor and Random Forest classifier on the consolidated dataset.
  • Implemented real-time querying and monitoring of key resilience indicators.
  • Achieved a 23% reduction in Mean Absolute Percentage Error (MAPE) for forecasting accuracy.
  • Risk classification model reached an F1-score of 0.89 for predicting disruptions.
  • Enabled proactive identification of potential logistics disruptions.

Abstract

The resilience of the United States supply chain has been critically tested by recent global disruptions, revealing systemic vulnerabilities in forecasting, logistics, and inventory management. This research proposes a robust, data-driven framework that leverages Machine Learning (ML) and Structured Query Language (SQL) to enhance supply chain process optimization and bolster resilience. We developed an integrated data pipeline where SQL was utilized for the efficient extraction, transformation, and loading (ETL) of large-scale, multi-modal data from disparate sources, including ERP systems, IoT sensors, and logistics feeds. Subsequently, ML models, including a Gradient Boosting Regressor for demand forecasting and a Random Forest classifier for risk prediction, were trained on this consolidated dataset. The results demonstrate a significant improvement in forecasting accuracy, with a 23% reduction in Mean Absolute Percentage Error (MAPE) compared to traditional statistical methods. Furthermore, the risk classification model achieved an F1-score of 0.89, enabling proactive identification of potential disruptions in the logistics network. The SQL-driven data infrastructure allowed for real-time querying and monitoring of key resilience indicators, such as inventory turnover and supplier lead time variability. The discussion highlights how this synergistic use of ML for predictive analytics and SQL for scalable data management creates a closed-loop system for continuous process improvement. We conclude that the adoption of such a data-centric approach is imperative for building agile, transparent, and resilient supply chains capable of withstanding future shocks.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Sarker et al. (2025) studied this question.

synapsesocial.com/papers/69746149bb9d90c67120b22chttps://doi.org/10.5281/zenodo.18335827
Ask AI
Helpful
Bookmark
Share
View Full Paper